AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
4 May 2026

BWLA: Breaking the Barrier of W1AX Post-Training Quantization for LLMs

ApplicationsDGX agent

arXiv:2605.00422v1 Announce Type: new Abstract: Large language models (LLMs) have driven major progress in NLP, yet their substantial memory and compute demands still hinder practical deployment. Bina

Copula-enhanced Vision Transformer for high myopia diagnosis through OU UWF fundus images

ResearchDGX agent

arXiv:2501.06540v2 Announce Type: replace Abstract: The advancement of AI-assisted myopia screening necessitates the joint diagnosis of both-eye (OU) high myopia (HM) status and the prediction of axia

Decouple before Integration: Test-time Synthesis of SFT and RLVR Task Vectors

ResearchDGX agent

arXiv:2605.00610v1 Announce Type: new Abstract: SFT and RLVR represent two fundamental yet distinct paradigms for LLM post-training, each excelling in distinct dimensions. SFT expands knowledge breadt

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Deep Kernel Learning for Stratifying Glaucoma Trajectories

ResearchDGX agent

arXiv:2605.00708v1 Announce Type: new Abstract: Effectively stratifying patient risk in chronic diseases like glaucoma is a major clinical challenge. Clinicians need tools to identify patients at high

DeGenTWeb: A First Look at LLM-dominant Websites

TutorialsDGX agent

arXiv:2605.00087v1 Announce Type: cross Abstract: Many recent news reports have claimed that content generated by large language models (LLMs) is taking over the web. However, these claims are typical

EASE: Federated Multimodal Unlearning via Entanglement-Aware Anchor Closure

ResearchDGX agent

arXiv:2605.00733v1 Announce Type: cross Abstract: Federated Multimodal Learning (FML) trains multimodal models across decentralized clients while keeping their image-text pairs private. However, joint

EGREFINE: An Execution-Grounded Optimization Framework for Text-to-SQL Schema Refinement

Local AiDGX agent

arXiv:2605.00628v1 Announce Type: cross Abstract: Text-to-SQL enables non-expert users to query databases in natural language, yet real-world schemas often suffer from ambiguous, abbreviated, or incon

Even our toughest critics come around eventually

ResearchDGX agent

Nous Research likely discusses how their AI models or research have gained acceptance even among skeptical observers, suggesting that rigorous development and demonstrated capabilities eventually conv

Excited to partner with @pinecone!

ApplicationsDGX agent

Excited to partner with @pinecone! Introducing Pinecone Nexus. A knowledge engine for agents. The bottleneck for production agents isn't the model. It's the per-query work of searching, stitching, par

GCGNet: Graph-Consistent Generative Network for Time Series Forecasting with Exogenous Variables

ApplicationsDGX agent

arXiv:2603.08032v2 Announce Type: replace Abstract: Exogenous variables offer valuable supplementary information for predicting future endogenous variables. Forecasting with exogenous variables needs

Graph Rewiring in GNNs to Mitigate Over-Squashing and Over-Smoothing: A Survey

ResearchDGX agent

arXiv:2411.17429v2 Announce Type: replace Abstract: Graph Neural Networks are powerful models for learning from graph-structured data, yet their effectiveness is often limited by two critical challeng

Hierarchical Abstract Tree for Cross-Document Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2605.00529v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models with external knowledge, and tree-based RAG organizes documents into hierarchical in

High-Speed Vision Improves Zero-Shot Semantic Understanding of Human Actions

ResearchDGX agent

arXiv:2605.00496v1 Announce Type: new Abstract: Understanding human actions from visual observations is essential for human--robot interaction, particularly when semantic interpretation of unfamiliar

I think I can say Runway is among the platforms with the most user-friendly interface and navigation. Pretty sure I can say that, it's my op…

TutorialsDGX agent

I think I can say Runway is among the platforms with the most user-friendly interface and navigation. Pretty sure I can say that, it's my opinion. But, Runway interface and combined models really do e

I tried running the same 'Generate an SVG of a pelican riding a bicycle' prompt against 21 different quantized variants of the same IBM Gran…

ToolsDGX agent

I tried running the same 'Generate an SVG of a pelican riding a bicycle' prompt against 21 different quantized variants of the same IBM Granite 4.1 3B model - the results weren't as interesting as I h

Information-geometric adaptive sampling for graph diffusion

ResearchDGX agent

arXiv:2605.00250v1 Announce Type: cross Abstract: Standard diffusion models for graph generation typically rely on uniform time-stepping, an approach that overlooks the non-homogeneous dynamics of dis

Introducing agent quality optimization in AgentCore, now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

Introducing the agent performance loop: AgentCore Optimization now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

Introducing the agent quality loop: AgentCore Optimization now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

@LangChain is doing cool stuff.

AgentsDGX agent

LangChain, a framework for developing applications with large language models, is advancing its capabilities and features according to insights from Harrison Chase, one of its key figures. This post l

Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory

SafetyDGX agent

arXiv:2605.00702v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term user memory for consistent personalization, but limited context windows hinder tracking evolving pre

Learning the Helmholtz equation operator with DeepONet for non-parametric 2D geometries

Local AiDGX agent

arXiv:2605.00760v1 Announce Type: new Abstract: This paper deals with solving the 2D Helmholtz equation on non-parametric domains, leveraging a physics-informed neural operator network based on the De

llamacpp on Apple Silicon, once configured correctly is really rock solid! You can throw anything at it and it will answer. Impressive!

HardwareDGX agent

Llamacpp, when properly configured on Apple Silicon hardware, demonstrates robust performance and reliability for running language models. The tool can handle varied input requests effectively, making

Memory in the LLM Era: Modular Architectures and Strategies in a Unified Framework

AgentsDGX agent

arXiv:2604.01707v2 Announce Type: replace Abstract: Memory emerges as the core module in the large language model (LLM)-based agents for long-horizon complex tasks (e.g., multi-turn dialogue, game pla

MIT's virtual violin offers luthiers a new design tool

IndustryDGX agent

MIT engineers have developed a 'computational violin' that uses physics-based simulation to realistically produce violin sound by modeling how the instrument and its vibrating strings interact with su

Network Digital Untwinning: Towards Backward Optimization of Digital Twins

ApplicationsDGX agent

arXiv:2605.00169v1 Announce Type: cross Abstract: Network digital twins (NDTs) are transforming network management by offering precise virtual replicas of physical network systems. However, their reli

One of the great parts of nous portal is stuff like this

AgentsDGX agent

One of the great parts of nous portal is stuff like this Trinity-Large-Thinking, @arcee_ai's latest model, is now free on Nous Portal for the next week Sign up for Nous Portal to use it in your Hermes

Optimize Supply Chain Decision Systems Using NVIDIA cuOpt Agent Skills

HardwareDGX agent

NVIDIA cuOpt Agent Skills integrate LLM reasoning with GPU-accelerated solvers to enable AI agents to translate natural language supply chain problems into optimized mathematical models, with skills e

PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation

ResearchDGX agent

arXiv:2605.00517v1 Announce Type: new Abstract: Despite substantial progress in text-driven 3D human motion synthesis, generating realistic multi-person interaction sequences remains challenging. Nota

Privacy Amplification in Differentially Private Zeroth-Order Optimization with Hidden States

ResearchDGX agent

arXiv:2506.00158v2 Announce Type: replace Abstract: Zeroth-order optimization has emerged as a promising approach for fine-tuning large language models under differential privacy (DP) and memory const

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference

ResearchDGX agent

arXiv:2602.18196v3 Announce Type: replace Abstract: Structured dilated attention has an appealing inference-time efficiency knob: it reduces the FLOPs of attention and the KV cache size by a factor of

Recovering Hidden Reward in Diffusion-Based Policies

SafetyDGX agent

arXiv:2605.00623v1 Announce Type: new Abstract: This paper introduces EnergyFlow, a framework that unifies generative action modeling with inverse reinforcement learning by parameterizing a scalar ene

Reinforcement Learning for LLM Post-Training: A Survey

SafetyDGX agent

arXiv:2407.16216v3 Announce Type: replace Abstract: Large language models (LLMs) trained via pretraining and supervised fine-tuning (SFT) can still produce harmful and misaligned outputs, or struggle

ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning

SafetyDGX agent

arXiv:2605.00380v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) enhances reasoning of Large Language Models (LLMs) but usually exhibits limited generation diver

RT @LangChain: Excited to partner with @pinecone!

ToolsDGX agent

LangChain announced a partnership with Pinecone, a vector database platform, to integrate vector search capabilities with LangChain's framework for building applications with large language models. Th

RunAgent: Interpreting Natural-Language Plans with Constraint-Guided Execution

AgentsDGX agent

arXiv:2605.00798v1 Announce Type: cross Abstract: Humans solve problems by executing targeted plans, yet large language models (LLMs) remain unreliable for structured workflow execution. We propose Ru

Scale-Aware Adversarial Analysis: A Diagnostic for Generative AI in Multiscale Complex Systems

ResearchDGX agent

arXiv:2605.00510v1 Announce Type: cross Abstract: Complex physical systems, from supersonic turbulence to the macroscopic structure of the universe, are governed by continuous multiscale dynamics. Whi

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.00444v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) advance vision language understanding but face inherent limitations in long-video tasks due to bounded percept

Smart Ensemble Learning Framework for Predicting Groundwater Heavy Metal Pollution

ResearchDGX agent

arXiv:2605.00056v1 Announce Type: new Abstract: Groundwater in the Densu Basin is increasingly threatened by heavy metal contamination, but conventional methods fail to capture the statistical complex

Soft Graph Diffusion Transformer for MIMO Detection

SafetyDGX agent

arXiv:2605.00449v1 Announce Type: cross Abstract: Learning-based MIMO detection has shown strong empirical performance, yet existing methods typically rely on fixed-depth architectures without explici

Spiking Sequence Machines and Transformers

ResearchDGX agent

arXiv:2605.00662v1 Announce Type: cross Abstract: Sequence learning reduces to similarity-based retrieval over a temporally indexed representation space, a constraint on any sequence model, not a prop

SPLICE: Latent Diffusion over JEPA Embeddings for Conformal Time-Series Inpainting

ResearchDGX agent

arXiv:2605.00126v1 Announce Type: new Abstract: Generative models for time-series imputation achieve strong reconstruction accuracy, yet provide no finite-sample reliability guarantees, a critical lim

SynQuE: Estimating Synthetic Dataset Quality Without Annotations

ApplicationsDGX agent

arXiv:2511.03928v5 Announce Type: replace Abstract: We introduce and formalize the Synthetic Dataset Quality Estimation (SynQuE) problem: ranking synthetic datasets by their expected real-world task p

TimeRFT: Stimulating Generalizable Time Series Forecasting for TSFMs via Reinforcement Finetuning

ApplicationsDGX agent

arXiv:2605.00015v1 Announce Type: cross Abstract: Time Series Foundation Models (TSFMs) advance generalization and data efficiency in time series forecasting by unified large-scale pretraining. But TS

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection

ResearchDGX agent

arXiv:2602.03216v2 Announce Type: replace Abstract: The quadratic complexity of attention remains the central bottleneck in long-context inference for large language models. Prior acceleration methods

VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding

ResearchDGX agent

arXiv:2603.22285v2 Announce Type: replace Abstract: Long video understanding remains challenging for multimodal large language models (MLLMs) due to limited context windows, which necessitate identify

'What Are You Really Trying to Do?': Co-Creating Life Goals from Everyday Computer Use

ResearchDGX agent

arXiv:2605.00497v1 Announce Type: cross Abstract: Recent advances in user modeling make it feasible to conduct open-ended inference over a person's everyday computer use. Despite longstanding visions

What if your AI could review its own work before you even see it? @ListenLabs Co-Founder & CTO @florian_jue explained how this works on the …

AgentsDGX agent

This post discusses self-review capabilities in AI systems, where an AI model can evaluate and refine its own outputs before presenting them to users. According to Florian Jue from Listen Labs, this f

When Structure Doesn't Help: LLMs Do Not Read Text-Attributed Graphs as Effectively as We Expected

ResearchDGX agent

arXiv:2511.16767v2 Announce Type: replace Abstract: Graphs provide a unified representation of semantic content and relational structure, making them a natural fit for domains such as molecular modeli

3 May 2026

Comfy developers pushing important updates to fix broken workflows

Local AiDGX agent

ComfyUI developers released important updates addressing workflow compatibility issues, including fixes for model compatibility problems like HunYuan 3D 2.0 support and EasyCache input/output channel

FastSDCPU release v1.0.0-beta.301

Local AiDGX agent

FastSDCPU is an optimized fork of Stable Diffusion designed to run efficiently on CPUs and devices without dedicated GPUs by leveraging Latent Consistency Models and Adversarial Diffusion Distillation

Wiki Lint Report — 2026-05-03

SynthesesDGX agent

Automated lint: 45 errors, 11 warnings, 3 info

Releasing a skill to help build LLM Wikis.

ResearchDGX agent

DAIR.AI released a skill or tool designed to assist in building wikis powered by large language models (LLMs), likely enabling users to create structured knowledge bases with AI capabilities. The reso

The Top AI Papers of the Week (April 26 - May 3) - Latent Agents - RecursiveMAS - OneManCompany - AgenticQwen-30B-A3B - Agentic World Modeli…

AgentsDGX agent

The Top AI Papers of the Week (April 26 - May 3) - Latent Agents - RecursiveMAS - OneManCompany - AgenticQwen-30B-A3B - Agentic World Modeling - Agentic Harness Engineering - From Skill Text to Skill

2 May 2026

Article: https://venturebeat.com/infrastructure/the-ai-scaffolding-layer-is-collapsing-llamaindexs-ceo-explains-what-survives Podcast: https…

AgentsDGX agent

The article discusses how the AI scaffolding layer—infrastructure tools and frameworks built around large language models—is experiencing consolidation and collapse, with LlamaIndex CEO explaining whi

b9008

Local AiDGX agent

B9008 is a build release of llama.cpp from May 2, 2026. Llama.cpp is a C/C++ implementation of LLM inference designed to enable large language model inference with minimal setup and high performance a

Composer 2 is 50% off in the SDK this weekend. Enjoy!

ToolsDGX agent

Composer 2 is 50% off in the SDK this weekend. Enjoy! We’re introducing the Cursor SDK so you can build agents with the same runtime, harness, and models that power Cursor. Run agents from CI/CD pipel

Generally, I would say X is not real life, but I am surprised about how often I get asked by executives about which AI lab is winning or wha…

ApplicationsDGX agent

Generally, I would say X is not real life, but I am surprised about how often I get asked by executives about which AI lab is winning or what is up with a particular model in ways that indicate that t

I need testers. Ollama Cloud Chat android app

Local AiDGX agent

A developer is seeking beta testers for 'Ollama Cloud Chat,' an Android application that integrates Ollama's cloud models with a mobile chat interface. The post likely discusses features, how to parti

@xai @grok this really puts the efficiency into focus

ToolsDGX agent

This post from Swyx discusses Grok (xAI's AI assistant) in the context of efficiency, likely highlighting performance metrics, speed improvements, or resource optimization of the model compared to alt

← Previous
1…785786787788789…1017
Next →