AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
1 Jun 2026

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2605.30621v1 Announce Type: new Abstract: LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task

Memory by Design: Probabilistic Sequence Layers

Model ReleasesDGX agent

arXiv:2605.31163v1 Announce Type: cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writ

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

AgentsDGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Not All Synthetic Data Is Yours to Learn From

Model ReleasesDGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

NumLeak: Public Numeric Benchmarks as Latent Labels in Foundation Models

ApplicationsDGX agent

arXiv:2605.30393v1 Announce Type: cross Abstract: Public numeric benchmarks appear in pretraining, so an evaluation that conditions on a date may be measuring memorized recall rather than out-of-sampl

Protocol for evaluating ChatGPT in biomedical association generation and verification using a RAG-enabled, cross-model majority voting workflow

ResearchDGX agent

arXiv:2605.30400v1 Announce Type: new Abstract: We present a protocol to evaluate ChatGPT's ability to generate disease-centric biomedical associations. It outlines how we generate the associations, v

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs

Model ReleasesDGX agent

arXiv:2605.30646v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in clinical applications. However, their behavior remains highly sensitive to subtle linguistic var

SubsurfaceGen: Procedural Generation of Field-Scale Earth Models and Seismic Data

HardwareDGX agent

arXiv:2605.30541v1 Announce Type: new Abstract: Full waveform inversion (FWI) is the gold standard for subsurface imaging, with applications from carbon sequestration to energy and mineral exploration

Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Transformers

Model ReleasesDGX agent

arXiv:2603.09453v3 Announce Type: replace-cross Abstract: Foundation models are increasingly being deployed in contexts where understanding the uncertainty of their outputs is critical to ensuring res

When Are Multimodal Predictions Biologically Supported? A Diagnostic Evaluation Framework

ResearchDGX agent

arXiv:2605.31504v1 Announce Type: new Abstract: Multimodal models in oncology can produce accurate predictions, but accurate prediction does not reveal whether the model has learned biology that is sh

29 May 2026

Comparing Post-Hoc Explainable AI Methods for Interpreting Black-Box EEG Models in Depression Detection

ResearchDGX agent

arXiv:2605.28977v1 Announce Type: cross Abstract: Recent advances in deep learning have enabled increasingly accurate electroencephalography (EEG)-based classification of Major Depressive Disorder (MD

Conf-Gen: Conformal Uncertainty Quantification for Generative Models

ResearchDGX agent

arXiv:2605.28920v1 Announce Type: cross Abstract: Conformal prediction (CP) and its extension, conformal risk control (CRC), are established frameworks for quantifying uncertainty in supervised machin

Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop

SafetyDGX agent

arXiv:2601.17670v2 Announce Type: replace-cross Abstract: Mathematical programming is widely employed across various sectors - such as logistics, energy, and workforce planning - to model and solve in

GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human

Model ReleasesDGX agent

arXiv:2605.28882v1 Announce Type: cross Abstract: With the rapid advancement of large language models, evaluating human-likeness in open-ended conversation has become increasingly important. However,

OmniCustom: Sync Audio-Video Customization Via Joint Audio-Video Generation Model

ResearchDGX agent

arXiv:2602.12304v4 Announce Type: replace-cross Abstract: Existing mainstream video customization methods focus on generating identity-consistent videos based on given reference images and textual pro

RadioFormer3D: Weakly Supervised 3D Radio Map Estimation in Low-Altitude Airspace via Generative Modeling

Local AiDGX agent

arXiv:2605.29538v1 Announce Type: new Abstract: With the emergence of wireless applications in three-dimensional environments, such as the low-altitude airspace and 3D heterogeneous networks, radio ma

RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models

AgentsDGX agent

arXiv:2603.18859v2 Announce Type: replace Abstract: Reinforcement learning (RL) shows promise for enhancing LLM agentic reasoning, yet sparse terminal rewards hinder fine-grained optimization. Process

SalsaAgent: A multimodal embodied language model for interactive dance generation

ResearchDGX agent

arXiv:2605.29219v1 Announce Type: new Abstract: Interaction between humanoids involves bidirectional and nonverbal reactivity, coordination and synchrony. Toward socially aware robots and interactive

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

Model ReleasesDGX agent

arXiv:2605.29358v1 Announce Type: new Abstract: We demonstrate that sparse autoencoders can extract interpretable features from Claude 3 Sonnet, a production-scale language model, addressing the open

SLAD : Shared LoRA Adapters for Task Specific Distillation

Model ReleasesDGX agent

arXiv:2605.29726v1 Announce Type: new Abstract: In the context of resource-constrained environments such as embedded systems, adapting reduced-size foundation models to downstream tasks has become inc

Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for Human Cognitive Diversity and World-Model Alignment

SafetyDGX agent

arXiv:2605.29930v1 Announce Type: new Abstract: Mutual misunderstanding in contemporary society does not arise merely because people hold different opinions or values. Even under the same observations

Towards Localized and Disentangled Knowledge Editing for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.29826v1 Announce Type: cross Abstract: Existing methods in Multimodal Knowledge Editing (MKE) have advanced the ability to correct outdated or inaccurate knowledge in Multimodal Large Langu

28 May 2026

A Multiscale Kinetic Framework for Image Segmentation: From Particle Systems to Continuum Models

ResearchDGX agent

arXiv:2605.28619v1 Announce Type: new Abstract: In this work, we present a multiscale kinetic framework for consensus-based image segmentation. By interpreting an image as a system of interacting part

A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space Operations

ResearchDGX agent

arXiv:2605.27444v1 Announce Type: cross Abstract: The rapid expansion of space activities has led to an unprecedented accumulation of technical documentation, operational guidelines, and scientific li

Bandwidth-Efficient and Privacy-Preserving Edge-Cloud Many-to-Many Speech Translation

Model ReleasesDGX agent

arXiv:2605.28642v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated significant potential for speech-to-text translation (S2TT). However, existing deployment par

Constrained Auto-Bidding via Generative Response Modeling

ResearchDGX agent

arXiv:2605.27811v1 Announce Type: new Abstract: Auto-bidding systems aim to maximize advertiser value over long horizons under budget constraints and ratio targets such as cost-per-acquisition, yet fu

Diffusion Large Language Models for Visual Speech Recognition

SafetyDGX agent

arXiv:2605.28456v1 Announce Type: new Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually

Fine-Tuning Vision-Language Models for Understanding Current Damage and Scoring Priority with Quality Guard Agent

AgentsDGX agent

arXiv:2605.27452v1 Announce Type: new Abstract: Bridge inspection in Japan requires mandatory visual assessments every five years, yet qualitative damage ratings (levels a-e) assigned by different eng

Gradient Step Plug-and-Play Model for Dental Cone-Beam CT Reconstruction

ResearchDGX agent

arXiv:2605.28124v1 Announce Type: new Abstract: The goal of this work is to reduce the effect of photon noise in dental cone-beam CT reconstruction. We consider an inverse problem formulation and deve

How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving

Model ReleasesDGX agent

arXiv:2605.28302v1 Announce Type: cross Abstract: Modern large language model (LLM) inference has progressively disaggregated to keep pace with growing model sizes and tight TTFT and TPOT service-leve

Large Language Models as Automatic Annotators and Annotation Adjudicators for Fine-Grained Opinion Analysis

AgentsDGX agent

arXiv:2601.16800v3 Announce Type: replace Abstract: Fine-grained opinion analysis of text provides a detailed understanding of expressed sentiments, including the addressed entity. Although this level

Mag-VLA: Vision-Language-Action Model for Bimanual Magnetically Actuated Microrobot Manipulation

SafetyDGX agent

arXiv:2605.28486v1 Announce Type: new Abstract: Magnetically actuated microrobots have been used as wireless, non-contact manipulation tools at microscales, making them promising for minimally invasiv

PocketGS: On-Device Training of 3D Gaussian Splatting for High Perceptual Modeling

Local AiDGX agent

arXiv:2601.17354v5 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) enables real-time rendering, its training demands workstation-level compute and memory, making mobile deployment

Skillful high-resolution weather forecasting independent of physical models

ResearchDGX agent

arXiv:2605.28153v1 Announce Type: cross Abstract: Accurate and timely weather forecasts are critical for high-impact decisions in modern society. Machine-learning-based weather prediction is emerging

Stage-wise Distortion-Perception Traversal in Zero-shot Inverse Problems with Diffusion Models

ApplicationsDGX agent

arXiv:2605.28711v1 Announce Type: new Abstract: The distortion-perception (D-P) tradeoff is a fundamental phenomenon of Bayesian inverse problems, which characterizes the inherent tension between dist

The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic

Model ReleasesDGX agent

arXiv:2605.28700v1 Announce Type: new Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

Model ReleasesDGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

Thinking as Compression: Your Reasoning Model is Secretly a Context Compressor

ResearchDGX agent

arXiv:2605.28713v1 Announce Type: new Abstract: Context compression aims to shorten long context inputs with minimal information loss for LLM inference acceleration. While existing methods have shown

Why LLMs Fail at Causal Discovery and How Interventional Agents Escape

Model ReleasesDGX agent

arXiv:2605.27567v1 Announce Type: new Abstract: Causal discovery is a cornerstone of scientific reasoning, yet whether large language models can perform it reliably remains an open question. Recent be

27 May 2026

Adversarial Water-Filling: Theory, Algorithms and Foundation Model

TutorialsDGX agent

arXiv:2605.26163v1 Announce Type: cross Abstract: Competitive resource allocation problems over frequency and space can be formulated as minimax interaction between transmit power and worst-case inter

ARBITER: Reasoning Trajectory Basins and Majority Vote Failures in Test-Time Sampling

Model ReleasesDGX agent

arXiv:2605.26172v1 Announce Type: new Abstract: When language models use test-time sampling, they generate multiple reasoning trajectories and select an answer by majority vote. We show that these tra

BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma

TutorialsDGX agent

arXiv:2605.26376v1 Announce Type: cross Abstract: Hepatocellular carcinoma (HCC) is biologically heterogeneous, shaped by the interplay between hepatic functional reserve and tumor-related oncologic f

Dense2MoE: Pushing the Pareto Frontier of On-Device LLMs via Unified Pruning and Upcycling

Model ReleasesDGX agent

arXiv:2605.26496v1 Announce Type: cross Abstract: The Mixture of Experts MoE architecture is highly promising for resource constrained on device deployments yet training these models from scratch incu

Generative Animations: A Multi-Model Pipeline for Prompt-Driven Motion Synthesis

ApplicationsDGX agent

arXiv:2605.27203v1 Announce Type: cross Abstract: Animation elevates digital documents into immersive experiences, yet creating custom motion paths remains cumbersome, requiring designers to manually

METATR: A Multilingual, Evolving Benchmark for Automatic Text Recognition

Model ReleasesDGX agent

arXiv:2605.26712v1 Announce Type: new Abstract: Benchmarks that reflect the diversity and complexity of real-world documents are essential for accurately evaluating Automatic Text Recognition (ATR) sy

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

Personalized Generative Models for Contextual Debiasing

TutorialsDGX agent

arXiv:2605.26353v1 Announce Type: cross Abstract: Different visual patterns appear with different frequencies in the world: e.g., beach balls appear on sand more often than they do on a road. These st

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

AgentsDGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

Probabilistic Recurrent Intention Switching Model

ResearchDGX agent

arXiv:2605.26998v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) recovers reward functions from observed behavior, yet traditional methods assume a single stationary reward that ca

Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuitive Scenes

Model ReleasesDGX agent

arXiv:2601.07737v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable performance in mainstream visual understanding tasks, but their ability

Self-Cascaded Diffusion Models for Arbitrary-Scale Image Super-Resolution

ResearchDGX agent

arXiv:2506.07813v2 Announce Type: replace-cross Abstract: Arbitrary-scale image super-resolution aims to upsample images to any desired resolution, offering greater flexibility than traditional fixed-

SL-BiLEM: Structured Learnable Behavior-in-the-Loop Epidemic Modeling for Forecasting and Policy Evaluation

SafetyDGX agent

arXiv:2605.26704v1 Announce Type: cross Abstract: Epidemic forecasting faces a fundamental challenge: human behavior dynamically responds to disease spread, creating feedback loops that induce distrib

The Need for an External Observer Formalizing the Sufficiency Gap: A Mathematical Extension of Mixture Identifiability and Contextual Grounding in Sequence Models

AgentsDGX agent

arXiv:2605.26711v1 Announce Type: new Abstract: We construct a binary mixed-regime process with one deterministic textual regime and one random regime governed by an unobserved latent state. Even an i

Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets

Model ReleasesDGX agent

arXiv:2605.26165v1 Announce Type: cross Abstract: Agentic RAG systems that equip language models with dozens to hundreds of tool definitions face a critical resource conflict: tool schemas consume the

Transfer Learning using 66 Diseases for Disease Forecasting Applications

ResearchDGX agent

arXiv:2605.27269v1 Announce Type: new Abstract: Disease forecasting models typically rely on a single data stream, making models brittle when histories are short or noisy. Recent top-performing models

UltraCUA: A Foundation Model for Computer Use Agents with Hybrid Action

ApplicationsDGX agent

arXiv:2510.17790v3 Announce Type: replace-cross Abstract: Computer-use agents face a fundamental limitation. They rely exclusively on primitive GUI actions (click, type, scroll), creating brittle exec

Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems

Model ReleasesDGX agent

arXiv:2605.26302v1 Announce Type: new Abstract: Long-lived AI agents are increasingly deployed as persistent operational systems, yet they are still evaluated like freshly initialized models. Day-one

26 May 2026

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

AgentsDGX agent

arXiv:2605.24823v1 Announce Type: new Abstract: Manufacturing has passed through four widely recognized paradigms - mechanization, electrification, programmable automation, and Smart Manufacturing - e

Approximating Safety Feedback Without a Safety Oracle via Model Predictive Control

SafetyDGX agent

arXiv:2510.20955v2 Announce Type: replace Abstract: Safe decision-making algorithms for control of mobile robots often require the existence of feedback to verify the safety of proposed actions. This

Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models

AgentsDGX agent

arXiv:2605.24755v1 Announce Type: new Abstract: Speech monologues recorded in naturalistic settings provide opportunities to characterize mental illness phenomenology and detect symptom exacerbation.

← Previous
1…234235236237238…1018
Next →