AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise

DGX agent

arXiv:2605.18832v1 Announce Type: cross Abstract: The Transformer is the foundational building block of modern AI, yet offers no principled handling of uncertainty, which is prevalent in real applicat

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PRISM: A Benchmark for Programmatic Spatial-Temporal Reasoning

DGX agent

arXiv:2605.19382v1 Announce Type: new Abstract: Programmatic video generation through code offers geometric precision and temporal coherence beyond pixel-level diffusion models, yet rigorously evaluat

model-releasesarxiv-cs-ai
20 May 2026
Research

Robustness and Regularization in Hierarchical Re-Basin

DGX agent

arXiv:2510.09174v3 Announce Type: replace Abstract: This paper takes a closer look at Git Re-Basin, an interesting new approach to merge trained models. We propose a hierarchical model merging scheme

researcharxiv-cs-lg
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions

DGX agent

arXiv:2605.19895v1 Announce Type: new Abstract: Constraint programming practitioners accelerate hard problems through a layered set of techniques applied in order of risk. Standard hardening (symmetry

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding

DGX agent

arXiv:2605.19866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) parse documents end-to-end but frequently break down on layouts unlike those seen in training. We attribute this to a two-

model-releasesarxiv-cs-cv
20 May 2026
Hardware

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

DGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

hardwarearxiv-cs-cl
20 May 2026
Model Releases

Toto 2.0: Time Series Forecasting Enters the Scaling Era

DGX agent

arXiv:2605.20119v1 Announce Type: cross Abstract: We show that time series foundation models scale: a single training recipe produces reliable forecast-quality improvements from 4M to 2.5B parameters.

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs

DGX agent

arXiv:2605.19528v1 Announce Type: new Abstract: 3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera i

model-releasesarxiv-cs-cv
20 May 2026
Safety

Towards Distillation Guarantees under Algorithmic Alignment for Combinatorial Optimization

DGX agent

arXiv:2605.20074v1 Announce Type: new Abstract: Distillation transfers knowledge from a large model trained on broad data to a smaller, more efficient model suitable for deployment. In structured pred

safetyarxiv-cs-lg
20 May 2026
Model Releases

Trust or Abstain? A Self-Aware RAG Approach

DGX agent

arXiv:2605.18792v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves large language models (LLMs) by incorporating external evidence, but it also introduces knowledge confli

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins

DGX agent

arXiv:2603.06740v2 Announce Type: replace-cross Abstract: Protein language models (pLMs) have shown strong potential for zero-shot prediction of missense variant effects, yet systematic benchmarking o

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

WARC-Bench: Web Archive Based Benchmark for GUI Subtask Executions

DGX agent

arXiv:2510.09872v2 Announce Type: replace-cross Abstract: Training web agents to navigate complex, real-world websites requires them to master extit{subtasks} - short-horizon interactions on multiple

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Makes a Representation Good for Single-Cell Perturbation Prediction?

DGX agent

arXiv:2605.19343v1 Announce Type: new Abstract: Single-cell perturbation modeling is fundamental for understanding and predicting cellular responses to genetic perturbations. However, existing approac

model-releasesarxiv-cs-lg
20 May 2026
Local Ai

Your Neighbors Know: Leveraging Local Neighborhoods for Backdoor Detection in Decentralized Learning

DGX agent

arXiv:2605.19969v1 Announce Type: new Abstract: Decentralized learning (DL) is an emerging machine learning paradigm where nodes collaboratively train models without a central server. However, the col

local-aiarxiv-cs-lg
20 May 2026
Model Releases

A Machine with Short-Term, Episodic, and Semantic Memory Systems

DGX agent

arXiv:2212.02098v5 Announce Type: replace Abstract: Inspired by the cognitive science theory of the explicit human memory systems, we have modeled an agent with short-term, episodic, and semantic memo

model-releasesarxiv-cs-ai
19 May 2026
Research

A More Word-like Image Tokenization for MLLMs

DGX agent

arXiv:2605.17954v1 Announce Type: cross Abstract: Modern multimodal large language models (MLLMs) typically keep the language model fixed and train a visual projector that maps the pixels into a seque

researcharxiv-cs-ai
19 May 2026
Model Releases

Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring

DGX agent

arXiv:2605.16386v1 Announce Type: new Abstract: Multimodal large language models (LLMs) are increasingly explored as automated evaluators in clinical settings, yet their scoring behavior on ordinal cl

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

BESplit: Bias-Compensated Split Federated Learning with Evidential Aggregation

DGX agent

arXiv:2605.17508v1 Announce Type: cross Abstract: Split Federated Learning (SFL) enables privacy-preserving collaborative training by partitioning models between clients and a server. However, under n

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Point-Wise Matching: Structural Representation Alignment for Accelerating Diffusion Transformers

DGX agent

arXiv:2605.16949v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) demonstrate that aligning noisy latent states with well-trained semantic features-as pioneered by Repre

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Boundedly Rational Meta-Learning in Sequential Consumer Choice

DGX agent

arXiv:2605.16532v1 Announce Type: new Abstract: Many consumer decisions are repeated choices under uncertainty. Standard models capture these decisions using Bayesian learning and dynamic programming:

model-releasesarxiv-cs-lg
19 May 2026
Research

CADS: Conformal Adaptive Decision System for Cost-Efficient Image Classification

DGX agent

arXiv:2605.16401v1 Announce Type: new Abstract: While high-capacity AI models have advanced state-of-the-art performance, their practical deployment is often hindered by high inference costs, environm

researcharxiv-cs-cv
19 May 2026
Model Releases

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean

DGX agent

arXiv:2605.17255v1 Announce Type: new Abstract: Formal theorem-proving benchmarks enable mechanically verifiable evaluation of mathematical reasoning in large language models. However, existing benchm

model-releasesarxiv-cs-ai
19 May 2026
Research

Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks

DGX agent

arXiv:2510.01782v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) should refuse to answer questions beyond their knowledge. This capability, which we term knowledge-aware refusal,

researcharxiv-cs-ai
19 May 2026
Model Releases

CasualSynth: Generating Structurally Sound Synthetic Data

DGX agent

arXiv:2605.17528v1 Announce Type: cross Abstract: Large Language Models (LLMs) generate realistic synthetic data but offer no guarantee that their outputs respect the causal mechanisms governing the t

model-releasesarxiv-cs-ai
19 May 2026
Safety

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks

DGX agent

arXiv:2605.17458v1 Announce Type: new Abstract: Text classification models are typically trained via supervised fine-tuning (SFT). However, SFT essentially performs behavior cloning from instance-wise

safetyarxiv-cs-lg
19 May 2026
Model Releases

Closing the Gap at CRAC 2026: Two-Stage Adaptation for LLM-Based Multilingual Coreference Resolution

DGX agent

arXiv:2605.16984v1 Announce Type: new Abstract: We present our submission to the LLM track of the 2026 Computational Models of Reference, Anaphora and Coreference (CRAC 2026) shared task. With an aver

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

ContractBench: Can LLM Agents Preserve Observation Contracts?

DGX agent

arXiv:2605.17281v1 Announce Type: cross Abstract: Tool-augmented LLM agents call APIs whose intermediate outputs, such as presigned URLs, session tokens, and OAuth state parameters, are observation co

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning

DGX agent

arXiv:2602.02979v2 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong potential in complex reasoning, yet their progress remains fundamentally constrained by reliance

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

DBES: A Systematic Benchmark and Metric Suite for Evaluating Expert Specialization in Large-Scale MoEs

DGX agent

arXiv:2605.18498v1 Announce Type: cross Abstract: Expert specialization in Mixture-of-Experts (MoE) models remains poorly understood, with traditional evaluations conflating architectural load-balanci

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Distributed Perceptron under Bounded Staleness, Partial Participation, and Noisy Communication

DGX agent

arXiv:2601.10705v3 Announce Type: replace Abstract: We study a semi-asynchronous client-server perceptron trained via iterative parameter mixing (IPM-style averaging): clients run local perceptron upd

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

DriveSafer: End-to-End Autonomous Driving with Safety Guidance

DGX agent

arXiv:2605.16737v1 Announce Type: cross Abstract: End-to-End (E2E) autonomous driving models have shown growing capability in recent years, with performance improving on increasingly challenging bench

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

DynMuon: A Dynamic Spectral Shaping View of Muon

DGX agent

arXiv:2605.17109v1 Announce Type: cross Abstract: In recent years, Muon has emerged as the dominant method for training large language models, and transformers more broadly. The essential difference,

model-releasesarxiv-cs-ai
19 May 2026
Research

Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks

DGX agent

arXiv:2603.11689v2 Announce Type: replace Abstract: Frontier Multimodal Large Language Models (MLLMs) exhibit remarkable capabilities in Visual-Language Comprehension (VLC) tasks. However, they are of

researcharxiv-cs-ai
19 May 2026
Local Ai

Federated Learning by Utility-Constrained Stochastic Aggregation for Improving Rational Participation

DGX agent

arXiv:2605.18020v1 Announce Type: new Abstract: Federated Learning (FL) algorithms implicitly assume that clients passively comply with server-side orchestration by sharing local model updates upon se

local-aiarxiv-cs-lg
19 May 2026
Local Ai

FedSDR: Federated Self-Distillation with Rectification

DGX agent

arXiv:2605.18028v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models faces severe statistical heterogeneity. However, existing model-level defenses often overlook the root

local-aiarxiv-cs-ai
19 May 2026
Research

Haptic Rendering of Fractional-Order Viscoelasticity: Passivity and Rendering Fidelity

DGX agent

arXiv:2605.16389v1 Announce Type: cross Abstract: Haptic rendering of viscoelastic materials that exhibit creep and stress relaxation is crucial for many applications, such as medical training with re

researcharxiv-cs-ai
19 May 2026
Model Releases

High-dimensional ridge regression with random features for non-identically distributed data with a variance profile

DGX agent

arXiv:2504.03035v2 Announce Type: replace-cross Abstract: Random feature ridge regression is often analyzed in the high-dimensional regime under the homogeneous sampling model x_i=Sigma^{1/2}x_i', whe

model-releasesarxiv-cs-lg
19 May 2026
Local Ai

In-context learning enables continental-scale subsurface temperature prediction from sparse local observations

DGX agent

arXiv:2605.16665v1 Announce Type: new Abstract: Continental-scale knowledge of subsurface temperature is limited by the cost and sparsity of borehole measurements, but such information is essential fo

local-aiarxiv-cs-lg
19 May 2026
Model Releases

Joint Parameter and State-Space Bayesian Optimization: Using Process Expertise to Accelerate Manufacturing Optimization

DGX agent

arXiv:2602.17679v2 Announce Type: replace Abstract: Bayesian optimization (BO) is a powerful method for optimizing black-box manufacturing processes, but its performance is often limited when dealing

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Latency-Aware Deep Learning Benchmark for Real-Time Cyber-Physical Attack and Fault Classification in Inverter-Dominated Power Grids

DGX agent

arXiv:2605.17256v1 Announce Type: cross Abstract: This work introduces a latency-aware benchmarking framework for evaluating deep learning models in power system anomaly detection using high-fidelity,

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LEAF: A Living Benchmark for Event-Augmented Forecasting

DGX agent

arXiv:2605.16358v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to forecasting. To evaluate this capability while mitigating pre-training data contamination, se

model-releasesarxiv-cs-ai
19 May 2026
Research

Learning more physically realistic dynamics in machine-learning based weather forecasting with latent-space constraints

DGX agent

arXiv:2510.04006v2 Announce Type: replace Abstract: Data-driven machine learning (ML) models are reshaping weather forecasting and have shown the potential to accelerate and surpass traditional physic

researcharxiv-cs-lg
19 May 2026
Research

Learning Quantifiable Visual Explanations Without Ground-Truth

DGX agent

arXiv:2605.18681v1 Announce Type: new Abstract: Explainable AI (XAI) techniques are increasingly important for the validation and responsible use of modern deep learning models, but are difficult to e

researcharxiv-cs-ai
19 May 2026
Model Releases

LoopQ: Quantization for Recursive Transformers

DGX agent

arXiv:2605.16343v1 Announce Type: cross Abstract: Looped language models (LoopLMs) improve parameter efficiency by recursively reusing Transformer blocks, enabling deeper computation under a fixed mod

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

DGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

model-releasesarxiv-cs-ai
19 May 2026
Safety

Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics

DGX agent

arXiv:2605.18549v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) introduce new opportunities for safety monitoring through their Chain of Thought (CoT) reasoning. However, CoT is not alwa

safetyarxiv-cs-cl
19 May 2026
Model Releases

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

DGX agent

arXiv:2605.17596v1 Announce Type: new Abstract: We present NeuSymMS, an adaptive memory system that enables large language model (LLM) agents to learn, remember, and reason about users across sessions

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…367368369370371…1074
Next →