AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Research

Amortized Factor Inference Networks for Posterior Inference

DGX agent

arXiv:2605.26419v1 Announce Type: new Abstract: Amortized inference promises fast test-time Bayesian inference, but existing methods are inherently tied to fixed models. Extending amortization to unse

researcharxiv-cs-lg
27 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

An End-to-End Learning Approach for Solving Capacitated Location-Routing Problems

DGX agent

arXiv:2511.02525v2 Announce Type: replace-cross Abstract: The capacitated location-routing problems (CLRPs) are classical problems in combinatorial optimization, which require simultaneously making lo

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Anchor: Mitigating Artifact Drift in Agent Benchmark Generation

DGX agent

arXiv:2605.26321v1 Announce Type: new Abstract: AI agents are beginning to complete valuable, long-horizon business operations tasks, but training and evaluation environments for enterprise work still

model-releasesarxiv-cs-ai
27 May 2026
Agents

APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2603.13853v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) connects large language models (LLMs) to external knowledge, but single-round retrieval is often insuffic

agentsarxiv-cs-ai
27 May 2026
Applications

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework …

DGX agent

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework now supports hierarchical KV cache optimization for SWA. Pro

applicationsjeremy-howard--x
27 May 2026
Model Releases

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation

DGX agent

arXiv:2601.08146v3 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting their use on diverse natural text. We adapt Co

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

BhashaSetu: A Data-Centric Approach to Low-Resource Machine Translation

DGX agent

arXiv:2605.27050v1 Announce Type: new Abstract: We present BhashaSetu, a linguistically enriched English--Marathi parallel dataset addressing persistent data limitations in low-resource neural machine

model-releasesarxiv-cs-cl
27 May 2026
Agents

Building self-improving tax agents with Codex

DGX agent

This article describes how OpenAI's Codex model can be used to build autonomous tax agents capable of self-improvement through code generation and execution. The work demonstrates using large language

agentsopenai
27 May 2026
Model Releases

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

DGX agent

arXiv:2605.27000v1 Announce Type: cross Abstract: Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@K as the canonical metric. Yet the

model-releasesarxiv-cs-ai
27 May 2026
Research

Conceptual Steganography

DGX agent

arXiv:2605.26537v1 Announce Type: new Abstract: Language Models (LMs) emit Chains-of-Thought (CoTs) that drive much of their capability. However, the same sequence that carries useful reasoning can al

researcharxiv-cs-cl
27 May 2026
Safety

Counteraction-Aware Multi-Teacher On-Policy Distillation for General Capability Recovery with Domain Preservation

DGX agent

arXiv:2605.27115v1 Announce Type: new Abstract: Domain specialization can improve LLM behavior in vertical domains, but often weakens the general capabilities inherited from the original model. Recent

safetyarxiv-cs-ai
27 May 2026
Safety

Curriculum Learning for Safety Alignment

DGX agent

arXiv:2605.26315v1 Announce Type: cross Abstract: Direct Preference Optimisation (DPO) is widely used for safety alignment in large language models. However, prior work shows it is brittle and exhibit

safetyarxiv-cs-ai
27 May 2026
Model Releases

Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% — For months,

model-releasestechmeme
27 May 2026
Model Releases

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

DGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

model-releasesarxiv-cs-ai
27 May 2026
Applications

Dynamic Link Prediction with Temporally Enhanced Signed Graph Neural Networks

DGX agent

arXiv:2605.26290v1 Announce Type: new Abstract: Temporal signed networks (TSNs) model the time evolution of cooperative and adversarial relationships that arise in applications such as social media an

applicationsarxiv-cs-lg
27 May 2026
Tutorials

DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding

DGX agent

arXiv:2605.26680v1 Announce Type: cross Abstract: Recent video multimodal large language models (MLLMs) increasingly couple step-by-step reasoning with on-demand visual evidence retrieval, allowing mo

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

ECSEL: Explainable Classification via Signomial Equation Learning

DGX agent

arXiv:2601.21789v2 Announce Type: replace-cross Abstract: We introduce ECSEL, an explainable classification method that learns formal expressions in the form of signomial equations, motivated by the o

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Efficient Prediction of SO(3)-Equivariant Hamiltonian Matrices via SO(2) Local Frames

DGX agent

arXiv:2506.09398v3 Announce Type: replace Abstract: We consider the task of predicting Hamiltonian matrices to accelerate electronic structure calculations, which plays an important role in physics, c

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

DGX agent

arXiv:2605.24456v2 Announce Type: replace Abstract: Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life.

model-releasesarxiv-cs-cv
27 May 2026
Safety

Elias in the Lighthouse, Again? Diagnosing Low Diversity in LLM Stories

DGX agent

arXiv:2605.26492v1 Announce Type: cross Abstract: LLM-generated stories are a popular use case, but they show very low variability. We sample 20,000 total stories from four current models using five p

safetyarxiv-cs-ai
27 May 2026
Model Releases

Enhancing Autonomous Online Intrusion Detection for IoT with Balanced Learning, Reliable Pseudo-Labels, and Lightweight Architectures

DGX agent

arXiv:2605.26166v1 Announce Type: cross Abstract: The rapid proliferation of Internet of Things (IoT) devices has created an urgent demand for adaptive, resource-efficient Intrusion Detection Systems

model-releasesarxiv-cs-ai
27 May 2026
Safety

Ethical Fairness without Demographics in Human-Centered AI

DGX agent

arXiv:2603.13373v3 Announce Type: replace-cross Abstract: In ubiquitous and mobile health systems, computational models infer human states from wearable, behavioral, and physiological sensing data. In

safetyarxiv-cs-ai
27 May 2026
Research

Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination

DGX agent

arXiv:2605.27016v1 Announce Type: cross Abstract: Large language models (LLMs) are prone to hallucinations, i.e., statements unsupported by the input or training data, hindering reliable deployment. I

researcharxiv-cs-ai
27 May 2026
Applications

Function-Valued Causal Influence in Nonlinear Time Series

DGX agent

arXiv:2605.26408v1 Announce Type: new Abstract: Causal discovery in time series is increasingly performed using nonlinear machine-learning models, yet the resulting causal relationships are almost alw

applicationsarxiv-cs-lg
27 May 2026
Model Releases

GEM: Geometric Entropy Mixing for Optimal LLM Data Curation

DGX agent

arXiv:2605.26121v1 Announce Type: cross Abstract: LLM pre-training efficacy increasingly depends on data composition rather than sheer volume. Yet, optimal mixing is hindered by categorization flaws:

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Grammar of the Wave: Towards Explainable Multivariate Time Series Event Detection via Neuro-Symbolic VLM Agents

DGX agent

arXiv:2603.11479v2 Announce Type: replace-cross Abstract: Time Series Event Detection (TSED) aims to localize semantically meaningful events in time series data, with critical applications in high-sta

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

JobBench: Aligning Agent Work With Human Will

DGX agent

arXiv:2605.26329v1 Announce Type: new Abstract: Current benchmarks for occupational AI agents are scoped primarily by economic values, telling a replacement story. We introduce JobBench, which evaluat

model-releasesarxiv-cs-ai
27 May 2026
Safety

LAD-VF: LLM-Automatic Differentiation Enables Fine-Tuning-Free Robot Planning from Formal Methods Feedback

DGX agent

arXiv:2509.18384v2 Announce Type: replace Abstract: Large language models (LLMs) can translate natural language instructions into executable action plans for robotics, autonomous driving, and other do

safetyarxiv-cs-ro
27 May 2026
Model Releases

LaRe: Latent Refocusing for Multimodal Reasoning

DGX agent

arXiv:2511.02360v4 Announce Type: replace-cross Abstract: Chain of Thought (CoT) reasoning enhances logical performance by decomposing complex tasks, yet its multimodal extension faces a trade-off. Th

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior

DGX agent

arXiv:2605.26797v1 Announce Type: cross Abstract: We study Latent Recurrent Transformer (LRT), a lightweight augmentation of autoregressive transformers that reuses a high-level source-layer hidden st

model-releasesarxiv-cs-cl
27 May 2026
Agents

Learning GUI Grounding with Spatial Reasoning from Visual Feedback

DGX agent

arXiv:2509.21552v2 Announce Type: replace-cross Abstract: Graphical User Interface (GUI) grounding is commonly framed as a coordinate prediction task -- given a natural language instruction, generate

agentsarxiv-cs-cl
27 May 2026
Agents

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments

DGX agent

arXiv:2605.27209v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have facilitated the widespread deployment of LLMs as interactive agents capable of reasoning, planning,

agentsarxiv-cs-ai
27 May 2026
Model Releases

LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring

DGX agent

arXiv:2605.27088v1 Announce Type: new Abstract: Aligning LLMs for math tutoring typically requires RL-based training with multi-GPU infrastructure. We investigate whether training-free prompt optimiza

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

DGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV

DGX agent

arXiv:2605.26244v1 Announce Type: new Abstract: Audio-visual generation is rapidly advancing from short clips to minute-long content, while existing evaluation protocols remain largely confined to sho

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

LongCat-Video-Avatar 1.5 Technical Report

DGX agent

arXiv:2605.26486v1 Announce Type: new Abstract: Despite advances in audio-driven video generation, achieving commercial-grade stability remains challenging. We present LongCat-Video-Avatar 1.5, an upg

model-releasesarxiv-cs-cv
27 May 2026
Research

MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning

DGX agent

arXiv:2601.18904v2 Announce Type: replace-cross Abstract: Auditory Large Language Models (LLMs) have demonstrated strong performance across a wide range of speech and audio understanding tasks. Nevert

researcharxiv-cs-ai
27 May 2026
Safety

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

DGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

safetyarxiv-cs-ai
27 May 2026
Model Releases

Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

DGX agent

arXiv:2605.26844v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level teacher supervision. Recent selective OPD methods exploit the non-uni

model-releasesarxiv-cs-lg
27 May 2026
Research

Physics Informed Neural Networks for damped harmonic oscillator and Burger's Equation (with extrapolation analysis) [P]

DGX agent

Physics-informed neural networks (PINNs) are machine learning models that incorporate physical laws and equations as constraints during training to solve differential equations. This post likely discu

researchr-machinelearning
27 May 2026
Model Releases

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

DGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

model-releasesarxiv-cs-ai
27 May 2026
Research

PinPoint: Prompting with Informative Interior Points

DGX agent

arXiv:2605.26689v1 Announce Type: cross Abstract: Modern referring image segmentation pipelines couple a vision-language model (VLM) for grounding with a promptable segmenter such as the Segment Anyth

researcharxiv-cs-cl
27 May 2026
Local Ai

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2510.01833v2 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local de

local-aiarxiv-cs-ai
27 May 2026
Model Releases

Provably Communication-Efficient and Privacy-Preserving Federated Graph Neural Networks

DGX agent

arXiv:2605.26243v1 Announce Type: new Abstract: Graph neural networks (GNNs) achieve strong performance on relational data, but real-world graphs are often distributed across organizations that cannot

model-releasesarxiv-cs-lg
27 May 2026
Applications

Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks

DGX agent

arXiv:2605.26934v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to post-training reasoning models, yet a key limitation of existing studies i

applicationsarxiv-cs-ai
27 May 2026
Model Releases

Receipt Replay OOD: A Small Benchmark for Screen Replay Detection Under Domain Shift

DGX agent

arXiv:2605.26855v1 Announce Type: new Abstract: Public datasets such as DLC-2021, SynID, and KID34K have significantly contributed to research on presentation attack detection for identity documents,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Resolving Ambiguity in Composed Image Retrieval via Calibrated Interaction

DGX agent

arXiv:2605.24634v2 Announce Type: replace Abstract: Composed image retrieval (CIR) searches a corpus with a reference image and a text describing how to modify it. Despite rapid progress from triplet-

model-releasesarxiv-cs-cv
27 May 2026
← Previous
1…710711712713714…1371
Next →