AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,598Total entries
1Added by human
91,597Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,253 results
Model Releases

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

DGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

model-releasesarxiv-cs-ai
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems

DGX agent

arXiv:2512.01556v3 Announce Type: replace Abstract: Foundation models often generate unreliable answers, while heuristic uncertainty estimators fail to fully distinguish correct from incorrect outputs

researcharxiv-cs-ai
27 May 2026
Research

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)

DGX agent

arXiv:2605.27268v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) are often criticized for producing repetitive and homogeneous text, despite possessing vast latent vocabularies. W

researcharxiv-cs-ai
27 May 2026
Model Releases

MerLean-Prover: A Recursive Looping Harness for End-to-End Lean 4 Theorem Proving

DGX agent

arXiv:2605.26959v1 Announce Type: cross Abstract: MerLean-Prover is an end-to-end Lean4 theorem prover that replaces sorry declarations with kernel-checkable proofs. It is built from three agent types

model-releasesarxiv-cs-cl
27 May 2026
Research

Object Pose and Shape Estimation for Grasping: Does it Work?

DGX agent

arXiv:2605.26944v1 Announce Type: cross Abstract: The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models

researcharxiv-cs-cv
27 May 2026
Model Releases

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

DGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes

DGX agent

arXiv:2605.26619v1 Announce Type: new Abstract: Reconstructing continuous state trajectories of chaotic dynamical systems from sparse, noisy observations remains a fundamental open problem in nonlinea

model-releasesarxiv-cs-lg
27 May 2026
Research

Prospective evaluation of multimodal respiratory failure prediction: Do chest X-rays improve performance beyond EHR signals?

DGX agent

arXiv:2605.26255v1 Announce Type: cross Abstract: Early prediction of respiratory failure is critical for timely clinical intervention in intensive care units. Existing electronic health record (EHR)-

researcharxiv-cs-ai
27 May 2026
Research

Recursive Flow Matching

DGX agent

arXiv:2605.26535v1 Announce Type: cross Abstract: Generative models have emerged as a powerful paradigm for solving physics systems and modeling complex spatiotemporal dynamics. However, achieving hig

researcharxiv-cs-ai
27 May 2026
Model Releases

RLVR Datasets and Where to Find Them: Tracing Data Lineage for Better Training Data

DGX agent

arXiv:2605.26971v1 Announce Type: new Abstract: The proliferation of Reinforcement Learning from Verifiable Rewards (RLVR) datasets has exacerbated provenance collapse due to unclear lineage among exi

model-releasesarxiv-cs-lg
27 May 2026
Research

RoMo: A Large-Scale, Richly Organized Dataset and Semantic Taxonomy for Human Motion Generation

DGX agent

arXiv:2605.26241v1 Announce Type: new Abstract: Success in generative modeling across language, image, and video demonstrates that large, well-curated datasets are the key driver for building capable

researcharxiv-cs-cv
27 May 2026
Research

Searching the Internet for Challenging Benchmarks at Scale

DGX agent

arXiv:2509.26619v3 Announce Type: replace-cross Abstract: Many static benchmarks are beginning to saturate: as models rapidly improve, they achieve near-perfect scores on fixed test sets, leaving litt

researcharxiv-cs-ai
27 May 2026
Tools

Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL

DGX agent

Delta Weight Sync (DWS) is a technique for efficiently distributing and synchronizing large language models with trillions of parameters across distributed systems using Hugging Face's TRL (Transforme

toolshugging-face
27 May 2026
Model Releases

Shopping Companion: A Memory-Augmented LLM Agent for Real-World E-Commerce Tasks

DGX agent

arXiv:2603.14864v2 Announce Type: replace Abstract: In e-commerce, LLM agents show promise for shopping tasks such as recommendations, budget management, and bundle deals, where accurately capturing u

model-releasesarxiv-cs-cl
27 May 2026
Research

Tracing Computation Density in LLMs

DGX agent

arXiv:2605.27033v1 Announce Type: cross Abstract: Transformer-based large language models (LLMs) are comprised of billions of parameters arranged in deep and wide computational graphs, but it is not c

researcharxiv-cs-ai
27 May 2026
Model Releases

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering.…

DGX agent

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering. This new research introduces AgingBench, a longitudinal rel

model-releasesdair-ai--x
27 May 2026
Research

A Dynamical Framework for Cognitive Processes Based on Transformations and Semantic Equivalence

DGX agent

arXiv:2605.23942v1 Announce Type: new Abstract: This paper proposes a structural and dynamical framework for modeling cognitive processes within a cybernetic perspective. Cognitive states are represen

researcharxiv-cs-ai
26 May 2026
Model Releases

A Matched Spectral Benchmark of Quantum Inspired Feature Maps

DGX agent

arXiv:2605.24324v1 Announce Type: cross Abstract: Quantum machine learning is often motivated by the idea that quantum systems can expose useful high-dimensional structure that is difficult to access

model-releasesarxiv-cs-lg
26 May 2026
Safety

Adaptive Preference Optimization with Uncertainty-aware Utility Anchor

DGX agent

arXiv:2509.10515v1 Announce Type: cross Abstract: Offline preference optimization methods are efficient for large language models (LLMs) alignment. Direct Preference optimization (DPO)-like learning,

safetyarxiv-cs-cl
26 May 2026
Model Releases

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

DGX agent

arXiv:2505.24876v2 Announce Type: replace-cross Abstract: Deep reasoning is fundamental for solving complex tasks, especially in vision-centric scenarios that demand sequential, multimodal understandi

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

AI-Associated Lexical Shifts Across 34 Languages: Cross-Lingual Convergence and Diachronic Uptake in News Writing

DGX agent

arXiv:2605.25358v1 Announce Type: cross Abstract: AI-associated lexical shifts have been documented mainly in Scientific English. We extend this work to 34 languages in the WMT News Crawl corpus, refi

model-releasesarxiv-cs-ai
26 May 2026
Research

All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting

DGX agent

arXiv:2602.17234v2 Announce Type: replace Abstract: Backtesting LLMs on resolved events assumes models reason only from pre-cutoff knowledge, yet pretrained models inevitably leak post-cutoff knowledg

researcharxiv-cs-ai
26 May 2026
Model Releases

Autoregression-Free Neural Operators for Time-Dependent PDEs

DGX agent

arXiv:2605.25413v1 Announce Type: cross Abstract: Neural operators learn mappings from function-dependent inputs to solutions, providing an effective framework for solving partial differential equatio

model-releasesarxiv-cs-ai
26 May 2026
Research

BigMac: Breaking the Pareto Frontier of Compute and Memory in Multimodal LLM Training

DGX agent

arXiv:2605.25451v1 Announce Type: new Abstract: Training multimodal large language models (MLLMs) is challenged by both model and data heterogeneity. Existing systems redesign the training pipeline to

researcharxiv-cs-lg
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Conformalised imprecise inference for robust extrapolation under limited data

DGX agent

arXiv:2605.25882v1 Announce Type: new Abstract: Recent advances in uncertainty quantification increasingly emphasise the distinction between aleatory and epistemic uncertainty in machine learning, mot

model-releasesarxiv-cs-lg
26 May 2026
Research

Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory

DGX agent

arXiv:2605.24602v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) frequently suffer from object hallucinations, yet the visual perceptual mechanism underlying this failure rem

researcharxiv-cs-ai
26 May 2026
Model Releases

Courtroom Analogy: New Perspective on Uncertainty-Aware Classification

DGX agent

arXiv:2605.25616v1 Announce Type: new Abstract: Single-pass uncertainty quantification (UQ) methods for classification represent uncertainty by predicting a tractable distribution over the class proba

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Equation-Free Coarse Control of Distributed Parameter Systems via Local Neural Operators

DGX agent

arXiv:2509.23975v2 Announce Type: replace-cross Abstract: The control of high-dimensional distributed parameter systems (DPS) remains a challenge when explicit coarse-grained equations are unavailable

model-releasesarxiv-cs-lg
26 May 2026
Research

Fundamental Limitation in Explaining AI

DGX agent

arXiv:2605.24727v1 Announce Type: new Abstract: While large-scale models such as LLMs and diffusion models have achieved practical success, public institutions have emphasized the importance of explai

researcharxiv-cs-ai
26 May 2026
Research

GEESE: Genotype-aware End-to-End Spatio-temporal Embedding for Behavioral Phenotyping

DGX agent

arXiv:2605.24370v1 Announce Type: new Abstract: Behavioral phenotyping of genetic animal models currently requires labor-intensive manual feature engineering that limits reproducibility and scalabilit

researcharxiv-cs-lg
26 May 2026
Model Releases

GL-LFGNN:A Global-Local Dual-branch Causal Graph Neural Network Based on Liang-Kleeman Information Flow for EEG Emotion Recognition

DGX agent

arXiv:2605.25061v1 Announce Type: cross Abstract: EEG-based emotion recognition holds significant promise for objective diagnosis of mood disorders. Graph neural networks (GNNs) have emerged as the do

model-releasesarxiv-cs-ai
26 May 2026
Safety

Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation

DGX agent

arXiv:2605.25170v1 Announce Type: cross Abstract: Training data for olfaction is scattered through disparate, non-standardized datasets that limit the ability to build representative world models. Olf

safetyarxiv-cs-ai
26 May 2026
Local Ai

How to ship a local LLM that matches frontier LLMs with evals and prompt engineering

DGX agent

Most production AI features don't need a frontier model. Here's how capability evals and prompt engineering can help ship a local SLM that matches frontier-model quality with lower latency and cost. T

local-aiarize-ai
26 May 2026
Safety

Improved Scaling Laws via Weak-to-Strong Generalization in Random Feature Ridge Regression

DGX agent

arXiv:2603.05691v2 Announce Type: replace Abstract: It is increasingly common in machine learning to use learned models to label data and then employ such data to train more capable models. The phenom

safetyarxiv-cs-lg
26 May 2026
Safety

Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization

DGX agent

arXiv:2605.24960v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) faithfulness, i.e., whether CoTs genuinely reflect large language models' (LLM) underlying behavior, is typically evaluated und

safetyarxiv-cs-ai
26 May 2026
Safety

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection

DGX agent

arXiv:2509.13608v2 Announce Type: replace Abstract: As Large Multimodal Models (LMMs) become integral to daily digital life, understanding their safety architectures is a critical problem for AI Align

safetyarxiv-cs-lg
26 May 2026
Model Releases

Learning Fine-grained Parameter Sharing via Sparse Tensor Decomposition

DGX agent

arXiv:2411.09816v4 Announce Type: replace Abstract: Large neural networks achieve state-of-the-art performance on many tasks, yet their sheer size hinders deployment on resource-constrained devices. A

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries

DGX agent

arXiv:2508.15760v2 Announce Type: replace-cross Abstract: Tool calling has emerged as a critical capability for AI agents. In contrast to conventional tool calling frameworks that rely on static, prov

model-releasesarxiv-cs-ai
26 May 2026
Applications

Manifold-Constrained MPPI: Real-Time Sampling-Based Control Under Hard Constraints

DGX agent

arXiv:2605.24813v1 Announce Type: new Abstract: Sampling-based model predictive control methods, such as Model Predictive Path Integral (MPPI), offer derivative-free optimization and robustness in com

applicationsarxiv-cs-ro
26 May 2026
Safety

Measuring the Depth of LLM Unlearning via Activation Patching

DGX agent

arXiv:2605.24614v1 Announce Type: cross Abstract: Large language model (LLM) unlearning has emerged as a crucial post-hoc mechanism for privacy protection and AI safety, yet auditing whether target kn

safetyarxiv-cs-ai
26 May 2026
Model Releases

MEDAL: Manifold Embedding Distillation via Autoencoder Learning

DGX agent

arXiv:2605.24244v1 Announce Type: cross Abstract: Low-dimensional embeddings are widely used as visual summaries of high-dimensional data and to enable downstream scientific discoveries. Yet, popular

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

MuJoCoUni:Persistent Batched Runtime Primitives for MuJoCo

DGX agent

arXiv:2605.24922v1 Announce Type: new Abstract: We present MuJoCoUni, a downstream MuJoCo distribution for online robot learning and batched physics evaluation. Alongside the open-loop batched traject

model-releasesarxiv-cs-ro
26 May 2026
Research

Multimodality Stacking with Blockwise missing values and application to the PIONeeR biomarkers study for prediction of resistance to immunotherapy

DGX agent

arXiv:2605.25050v1 Announce Type: cross Abstract: Integrating multimodal datasets in clinical oncology is frequently hindered by high dimensionality and blockwise missingness, where entire data source

researcharxiv-cs-lg
26 May 2026
Local Ai

Multiscale Real-Time Object Detection in the NMS-Free Era: A Comparative Performance Evaluation of YOLOv8 and YOLO26

DGX agent

arXiv:2605.24831v1 Announce Type: cross Abstract: Non-Maximum Suppression (NMS) remains a key post-processing step in many real-time object detection pipelines, but it can introduce latency variation

local-aiarxiv-cs-ai
26 May 2026
Model Releases

Neural Integral Operators for Inverse Problems: An Operator-Learning Framework for Small-Sample Spectroscopic Classification

DGX agent

arXiv:2505.03677v3 Announce Type: replace Abstract: Learning maps between function spaces with a strong inductive bias is a central challenge in soft computing, especially when training data are scarc

model-releasesarxiv-cs-lg
26 May 2026
Research

NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding

DGX agent

arXiv:2605.24993v1 Announce Type: new Abstract: Current fMRI decoders face a performance-fidelity trade-off where efficient ID encoders outperform geometrically faithful surface-based models. We argue

researcharxiv-cs-ai
26 May 2026
Safety

OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation

DGX agent

arXiv:2605.25829v1 Announce Type: cross Abstract: Recent vision-language-action (VLA) models and world action models (WAMs) advance robotic manipulation by enriching intermediate representations with

safetyarxiv-cs-ai
26 May 2026
← Previous
1…613614615616617…1381
Next →