AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,046 results
Agents

From Spark to Fire: Modeling and Mitigating Error Cascades in LLM-Based Multi-Agent Collaboration

DGX agent

arXiv:2603.04474v2 Announce Type: replace-cross Abstract: Large Language Model-based Multi-Agent Systems (LLM-MAS) are increasingly applied to complex collaborative scenarios. However, their collabora

agentsarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Gate-and-Merge: Zero-shot Compositional Personalization of Vision Language Models

DGX agent

arXiv:2605.08702v1 Announce Type: cross Abstract: This paper tackles compositional personalization of vision-language models (VLMs). In this problem, multiple user-defined concepts must be recognized

researcharxiv-cs-ai
12 May 2026
Model Releases

Geometry-Aware Discretization Error of Diffusion Models

DGX agent

arXiv:2605.08392v1 Announce Type: new Abstract: Practical diffusion sampling is a numerical approximation problem: under a fixed inference budget, one must simulate a reverse-time ODE or SDE using onl

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model

DGX agent

arXiv:2605.10739v1 Announce Type: cross Abstract: We introduce SMART-HC-VQA, a Sentinel-2-based visual question answering dataset derived from the IARPA SMART Heavy Construction dataset, designed for

model-releasesarxiv-cs-ai
12 May 2026
Agents

i like that there are models called bert and ernie, but in all seriousness, this update looks impressive

DGX agent

i like that there are models called bert and ernie, but in all seriousness, this update looks impressive ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total pa

agentsyohei-nakajima--x
12 May 2026
Research

In KAME, a fast speech model starts replying instantly, while a backend LLM runs in parallel to inject deep knowledge on the fly. It’s a com…

DGX agent

In KAME, a fast speech model starts replying instantly, while a backend LLM runs in parallel to inject deep knowledge on the fly. It’s a completely different way to approach conversational AI, making

researchdavid-ha--x
12 May 2026
Safety

Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning

DGX agent

arXiv:2605.09009v1 Announce Type: cross Abstract: Large language models (LLMs) have shown remarkable in-context learning (ICL) capabilities, yet their potential for sequential decision-making remains

safetyarxiv-cs-ai
12 May 2026
Model Releases

Learning Less Is More: Premature Upper-Layer Attention Specialization Hurts Language Model Pretraining

DGX agent

arXiv:2605.10504v1 Announce Type: new Abstract: A causal-decoder block is hierarchical: lower layers build the residual basis that upper layers attend over. We identify a failure mode in GPT pretraini

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

LLM Agents Already Know When to Call Tools -- Even Without Reasoning

DGX agent

arXiv:2605.09252v1 Announce Type: new Abstract: Tool-augmented LLM agents tend to call tools indiscriminately, even when the model can answer directly. Each unnecessary call wastes API fees and latenc

model-releasesarxiv-cs-cl
12 May 2026
Safety

Metropolis-Adjusted Diffusion Models

DGX agent

arXiv:2605.09654v1 Announce Type: cross Abstract: Sampling from score-based diffusion models incurs bias due to both time discretisation and the approximation of the score function. A common strategy

safetyarxiv-cs-lg
12 May 2026
Safety

Mismatch-Aware Adaptive Constraint Tightening for Bicycle-Model Trajectory Optimization

DGX agent

arXiv:2605.09376v1 Announce Type: new Abstract: Trajectory optimization for autonomous vehicles usually relies on the kinematic bicycle model because of its computational simplicity. However, when the

safetyarxiv-cs-ro
12 May 2026
Research

Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models

DGX agent

arXiv:2602.01698v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have recently achieved strong mathematical and code reasoning performance through Reinforcement Learning (RL) post-tra

researcharxiv-cs-cl
12 May 2026
Research

Rethinking Event-Based Object Dtection through Representation-Level Temporal Aggregation and Model-Level Hypergraph Reasoning

DGX agent

arXiv:2605.08825v1 Announce Type: new Abstract: Event cameras provide microsecond-level temporal resolution, low latency, and high dynamic range, offering potential for perception under fast motion an

researcharxiv-cs-cv
12 May 2026
Research

Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models

DGX agent

arXiv:2605.08145v1 Announce Type: cross Abstract: Current vision language models face hallucination and robustness issues against ambiguous or corrupted modalities. We hypothesize that these issues ca

researcharxiv-cs-ai
12 May 2026
Industry

Talked to a friend at a top AI lab. Their whole team is former journalists, training models on poems, summaries, and creative writing. I use…

DGX agent

Talked to a friend at a top AI lab. Their whole team is former journalists, training models on poems, summaries, and creative writing. I use AI every day and can see that it tends to flatten my writin

industryallie-k--miller--x
12 May 2026
Research

TARO: Temporal Adversarial Rectification Optimization Using Diffusion Models as Purifiers

DGX agent

arXiv:2605.08440v1 Announce Type: cross Abstract: Adversarial purification with diffusion models seeks to project adversarial examples back toward the data manifold, but balancing semantic preservatio

researcharxiv-cs-cv
12 May 2026
Model Releases

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

DGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

model-releasesarxiv-cs-ai
12 May 2026
Applications

Though the smartness comes with a cost: all of the prompts that were written for the old realtime voice model now need to be revised for a m…

DGX agent

Ethan Mollick discusses a tradeoff in OpenAI's newer realtime voice model, where improved capabilities require developers to revise prompts that were written for the previous version. The post highlig

applicationsethan-mollick--x
12 May 2026
Agents

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

DGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

agentsarxiv-cs-ro
12 May 2026
Safety

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

DGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

safetyarxiv-cs-ai
12 May 2026
Research

UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing

DGX agent

arXiv:2601.08321v3 Announce Type: replace Abstract: With the rapid advancement of image generation, visual text editing using natural language instructions has received increasing attention. The main

researcharxiv-cs-cv
12 May 2026
Applications

Unlocking air traffic flow prediction through microscopic aircraft-state modeling

DGX agent

arXiv:2605.10083v1 Announce Type: new Abstract: Short-term air traffic flow prediction in terminal airspace is essential for proactive air traffic management. Existing approaches predominantly model t

applicationsarxiv-cs-lg
12 May 2026
Research

UxSID: Semantic-Aware User Interests Modeling for Ultra-Long Sequence

DGX agent

arXiv:2605.09040v1 Announce Type: new Abstract: Modeling ultra-long user sequences involves a difficult trade-off between efficiency and effectiveness. While current paradigms rely on either item-spec

researcharxiv-cs-ai
12 May 2026
Agents

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models

DGX agent

arXiv:2605.10106v1 Announce Type: cross Abstract: Recent advances in Multi-modal Large Language Models (MLLMs) target 3D spatial intelligence, yet the progress has been largely driven by post-training

agentsarxiv-cs-ai
12 May 2026
Research

ViSurf: Visual Supervised-and-Reinforcement Fine-Tuning for Large Vision-and-Language Models

DGX agent

arXiv:2510.10606v4 Announce Type: replace Abstract: Post-training Large Vision-and-Language Models (LVLMs) typically involves Supervised Fine-Tuning (SFT) for knowledge injection or Reinforcement Lear

researcharxiv-cs-cv
12 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Safety

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

DGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

safetyarxiv-cs-ai
12 May 2026
Safety

When More Parameters Hurt: Foundation Model Priors Amplify Worst-Client Disparity Under Extreme Federated Heterogeneity

DGX agent

arXiv:2605.08992v1 Announce Type: new Abstract: Federated learning (FL) is increasingly used to fine-tune foundation models (FMs) on distributed private data. The community largely assumes that large-

safetyarxiv-cs-lg
12 May 2026
Safety

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph

DGX agent

arXiv:2605.08037v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) aligns language models using pairwise preference comparisons, offering a simple and effective alternative to Rein

safetyarxiv-cs-ai
11 May 2026
Agents

CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment

DGX agent

arXiv:2605.06702v1 Announce Type: new Abstract: Large language models (LLMs) have become a central foundation of modern artificial intelligence, yet their lifecycle remains constrained by a rigid sepa

agentsarxiv-cs-ai
11 May 2026
Research

Code Generation and Conic Constraints for Model-Predictive Control on Microcontrollers with Conic-TinyMPC

DGX agent

arXiv:2403.18149v3 Announce Type: replace Abstract: Model-predictive control (MPC) is a state-of-the-art control method for constrained robotic systems, yet deployment on resource-limited hardware rem

researcharxiv-cs-ro
11 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Agents

Computer use with any model Hermes Agent × @trycua

DGX agent

This post from Nous Research discusses the integration of computer use capabilities with Hermes Agent models in collaboration with Claude (CUA), enabling AI agents to interact with computer interfaces

agentsnous-research--x
11 May 2026
Research

Conservative Flows: A New Paradigm of Generative Models

DGX agent

arXiv:2605.06905v1 Announce Type: new Abstract: Modern generative modeling is dominated by transport from a noise prior to data. We propose an alternative paradigm in which generation is performed by

researcharxiv-cs-lg
11 May 2026
Model Releases

Dataset Watermarking for Closed LLMs with Provable Detection

DGX agent

arXiv:2605.06865v1 Announce Type: new Abstract: Large language models (LLMs) are pre-trained and post-trained on vast amounts of loosely curated data, raising the possibility that these models may hav

model-releasesarxiv-cs-lg
11 May 2026
Applications

Emergence of Distortions in High-Dimensional Guided Diffusion Models

DGX agent

arXiv:2602.00716v4 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is the de facto standard for conditional sampling in diffusion models, yet it often reduces sample diversity. U

applicationsarxiv-cs-lg
11 May 2026
Research

Equivalence of Coarse and Fine-Grained Models for Learning with Distribution Shift

DGX agent

arXiv:2605.07005v1 Announce Type: cross Abstract: Recent work on provably efficient algorithms for learning with distribution shift has focused on two models: PQ learning (Goldwasser et al. (2020)) an

researcharxiv-cs-lg
11 May 2026
Local Ai

FLAM: Evaluating Model Performance with Aggregatable Measures in Federated Learning

DGX agent

arXiv:2605.07962v1 Announce Type: new Abstract: Performance evaluation is essential for assessing the quality of machine learning (ML) models and guiding deployment decisions. In federated learning (F

local-aiarxiv-cs-lg
11 May 2026
Tools

For browser-use AI agents, every task is dozens of model calls in a tight loop. The inference layer isn’t background infrastructure. It’s wh…

DGX agent

For browser-use AI agents, every task is dozens of model calls in a tight loop. The inference layer isn’t background infrastructure. It’s what the product runs on. @yutori_ai runs Scouts, Delegate, an

toolstogether-ai--x
11 May 2026
Research

From Average Sensitivity to Small-Loss Regret Bounds under Random-Order Model

DGX agent

arXiv:2602.09457v2 Announce Type: replace-cross Abstract: We study online learning in the random-order model, where the multiset of loss functions is chosen adversarially but revealed in a uniformly r

researcharxiv-cs-lg
11 May 2026
Hardware

GATO: GPU-Accelerated and Batched Trajectory Optimization for Scalable Edge Model Predictive Control

DGX agent

arXiv:2510.07625v2 Announce Type: replace Abstract: While Model Predictive Control (MPC) delivers strong performance across robotics applications, solving the underlying (batches of) nonlinear traject

hardwarearxiv-cs-ro
11 May 2026
Safety

HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models

DGX agent

arXiv:2605.07973v1 Announce Type: new Abstract: Text-to-image diffusion models can generate visually stunning images, yet, controlling what appears and how it appears, remains surprisingly difficult,

safetyarxiv-cs-cv
11 May 2026
Research

In modern ML accelerators, FLOPS have absolutely exploded. Often though, the bottleneck is not FLOPS but memory bandwidth. Similarly, model …

DGX agent

In modern ML accelerators, FLOPS have absolutely exploded. Often though, the bottleneck is not FLOPS but memory bandwidth. Similarly, model intelligence has exploded, causing the bottleneck to be huma

researchsoumith-chintala--x
11 May 2026
Research

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

DGX agent

arXiv:2605.07721v1 Announce Type: cross Abstract: Recurrent LLM architectures have emerged as a promising approach for improving reasoning, as they enable multi-step computation in the embedding space

researcharxiv-cs-ai
11 May 2026
Research

Mixture of Masters: Sparse Chess Language Models with Player Routing

DGX agent

arXiv:2602.04447v2 Announce Type: replace-cross Abstract: Modern chess language models are dense transformers trained on millions of games played by thousands of high-rated individuals. However, these

researcharxiv-cs-ai
11 May 2026
Safety

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models

DGX agent

arXiv:2602.07026v2 Announce Type: replace-cross Abstract: Despite the success of multimodal contrastive learning in aligning visual and linguistic representations, a persistent geometric anomaly, the

safetyarxiv-cs-ai
11 May 2026
Applications

Neural CDEs as Correctors for Learned Time Series Models

DGX agent

arXiv:2512.12116v3 Announce Type: replace Abstract: Learned time-series models, whether continuous or discrete, are widely used for forecasting the states of dynamical systems but suffer from error ac

applicationsarxiv-cs-lg
11 May 2026
Safety

One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy

DGX agent

arXiv:2605.07931v1 Announce Type: cross Abstract: Vision-language-action (VLA) models increasingly rely on auxiliary world modules to plan over long horizons, yet how such modules should be parameteri

safetyarxiv-cs-ai
11 May 2026
← Previous
1…236237238239240…1272
Next →