AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
Model Releases

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

DGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

model-releasesarxiv-cs-ai
20 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

DGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Search Self-play: Pushing the Frontier of Agent Capability without Supervision

DGX agent

arXiv:2510.18821v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become the mainstream technique for training LLM agents. However, RLVR highly depends on w

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Self-Filtered Distillation with LLMs-generated Trust Indicators for Reliable Patent Classification

DGX agent

arXiv:2510.05431v4 Announce Type: replace Abstract: Organizing large-scale patent corpora according to classification schemes is a core information management task that determines the accuracy and eff

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Self-improving AI is a big deal! As a first step, I've been exploring how much of the post-training can be automated. Here is a first post o…

DGX agent

Self-improving AI is a big deal! As a first step, I've been exploring how much of the post-training can be automated. Here is a first post on how I am using @FireworksAI_HQ Agent to automate LLM fine-

model-releasesdair-ai--x
20 May 2026
Model Releases

Semantic-Enriched Latent Visual Reasoning

DGX agent

arXiv:2605.19342v1 Announce Type: new Abstract: Multimodal latent-space reasoning aims to replace explicit thinking with images by performing visual reasoning directly in a compact latent space. Howev

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Sequential Consensus for Multi-Agent LLM Debates: A Wald-SPRT compute governor with calibration-based failure detection

DGX agent

arXiv:2605.19193v1 Announce Type: new Abstract: Multi-agent LLM debate improves factuality and reasoning, but most recipes pick a fixed round count, over-spending on easy items and under-spending on h

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Sharper Bounds for Chebyshev Moment Matching, with Applications

DGX agent

arXiv:2408.12385v3 Announce Type: replace-cross Abstract: We study the problem of approximately recovering a probability distribution given noisy measurements of its Chebyshev polynomial moments. This

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Simply Stabilizing the Loop via Fully Looped Transformer

DGX agent

arXiv:2605.18797v1 Announce Type: cross Abstract: Scaling model performance typically requires increasing model size. Looped Transformer offers a compelling alternative by iteratively reusing the same

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

DGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Smooth Piecewise Cutting for Neural Operator to Handle Discontinuities and Sharp Transitions

DGX agent

arXiv:2605.19823v1 Announce Type: cross Abstract: Neural operators have achieved strong performance in learning solution operators of partial differential equations (PDEs), but their inherently contin

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Sonar-TS: Search-Then-Verify Natural Language Querying for Time Series Databases

DGX agent

arXiv:2602.17001v2 Announce Type: replace Abstract: Natural Language Querying for Time Series Databases (NLQ4TSDB) aims to assist non-expert users retrieve meaningful events, intervals, and summaries

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

SpecX: A Large-Scale Benchmark for Multi-Modal Spectroscopy and Cross-Paradigm Evaluation

DGX agent

arXiv:2605.18791v1 Announce Type: cross Abstract: Existing spectral benchmarks are limited in scale, modality alignment, and evaluation scope, and typically focus on either specialized models or multi

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

STAR: Semantic-Tuned and Tail-Adaptive Retriever for Graph-Augmented Generation

DGX agent

arXiv:2605.18765v1 Announce Type: cross Abstract: To augment Large Language Models (LLMs) for multi-hop question answering, a mainstream solution within Graph Retrieval Augmented Generation (GraphRAG)

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Stochastic Gradient Variational Inference with Price's Gradient Estimator from Bures-Wasserstein to Parameter Space

DGX agent

arXiv:2602.18718v2 Announce Type: replace-cross Abstract: For approximating a target distribution given only its unnormalized log-density, stochastic gradient-based variational inference (VI) algorith

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions

DGX agent

arXiv:2605.19895v1 Announce Type: new Abstract: Constraint programming practitioners accelerate hard problems through a layered set of techniques applied in order of risk. Standard hardening (symmetry

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding

DGX agent

arXiv:2605.19866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) parse documents end-to-end but frequently break down on layouts unlike those seen in training. We attribute this to a two-

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Subagents running locally and simultaneously on MacBook Pro M5 with Codex CLI + @lmstudio to review code and find bugs using Qwen 3.6 Powere…

DGX agent

Subagents running locally and simultaneously on MacBook Pro M5 with Codex CLI + @lmstudio to review code and find bugs using Qwen 3.6 Powered by the updated MLX engine with batching in beta in the app

model-releaseslm-studio--x
20 May 2026
Model Releases

SynGR: Unleashing the Potential of Cross-Modal Synergy for Generative Recommendation

DGX agent

arXiv:2605.18920v1 Announce Type: cross Abstract: Generative Recommendation (GR) has emerged as a promising paradigm by formulating item recommendation as a sequence-to-sequence generation task over i

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Synthesis and Evaluation of Long-term History-aware Medical Dialogue

DGX agent

arXiv:2605.19766v1 Announce Type: cross Abstract: An effective healthcare agent must be able to recall and reason over a patient's longitudinal medical history. However, the absence of datasets with r

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Synthetic Data Generation for Brain-Computer Interfaces: Overview, Benchmarking, and Future Directions

DGX agent

arXiv:2603.12296v2 Announce Type: replace-cross Abstract: Deep learning has achieved transformative performance across diverse domains, largely driven by large-scale and high-quality training data. In

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

TADA! Tuning Audio Diffusion Models through Activation Steering

DGX agent

arXiv:2602.11910v2 Announce Type: replace-cross Abstract: Audio diffusion models can synthesize high-fidelity music from text, yet achieving fine-grained control over specific musical attributes remai

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Tail Annealing for Heavy-Tailed Flow Matching

DGX agent

arXiv:2605.20068v1 Announce Type: cross Abstract: Standard generative models struggle with heavy-tailed data: Lipschitz architectures cannot produce power-law tails from Gaussian noise, and interpolat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

DGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Target-Aligned Reinforcement Learning

DGX agent

arXiv:2603.29501v2 Announce Type: replace-cross Abstract: Many value-based deep reinforcement learning algorithms rely on target networks - lagged copies of the online network - to stabilize training.

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Targeted Downstream-Agnostic Attack

DGX agent

arXiv:2605.19446v1 Announce Type: cross Abstract: Recently, pre-trained encoders have gained widespread use due to their strong capability in representation extraction. However, they are vulnerable to

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards

DGX agent

arXiv:2605.19320v1 Announce Type: new Abstract: Faithful text rendering remains a persistent weakness of large text-to-image generative models, as it requires both semantic instruction following and f

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

DGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

The Evaluation Game: Beyond Static LLM Benchmarking

DGX agent

arXiv:2605.19377v1 Announce Type: cross Abstract: As jailbreaks, adversarially crafted inputs that bypass safety constraints, continue to be discovered in Large Language Models, practitioners increasi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

The frontier is still jagged though (here is Gemini 3.5 Flash messing up counting letters in words) https://x.com/RRiscio37389/status/205719…

DGX agent

The frontier is still jagged though (here is Gemini 3.5 Flash messing up counting letters in words) https://x.com/RRiscio37389/status/2057193260745883962?s=20 @emollick Proof: https://gemini.google.co

model-releasesethan-mollick--x
20 May 2026
Model Releases

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measur…

DGX agent

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measuring every biomarker, or @sytses openly sharing and analyzing

model-releasesclem-delangue--x
20 May 2026
Model Releases

The Growing Pains of Frontier Models: When Leaderboards Stop Separating and What to Measure Next

DGX agent

arXiv:2605.18840v1 Announce Type: cross Abstract: Leaderboards rank frontier models on independent axes but do not reveal whether capabilities reinforce or trade off across releases -- and at the fron

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular,…

DGX agent

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular, and represents an important milestone for the math and AI c

model-releasesopenai--x
20 May 2026
Model Releases

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

DGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

The World Won't Stay Still: Programmable Evolution for Agent Benchmarks

DGX agent

arXiv:2603.05910v2 Announce Type: replace Abstract: LLM-powered tool-calling agents fulfill user requests by interacting with environments, querying data, and invoking tools in a multi-turn process. Y

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Theory-optimal Quantization Based on Flatness

DGX agent

arXiv:2605.18800v1 Announce Type: cross Abstract: Post-training quantization has emerged as a widely adopted technique for compressing and accelerating the inference of Large Language Models (LLMs). T

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting …

DGX agent

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting ideas across distant fields, and surfacing paths researchers

model-releasesopenai--x
20 May 2026
Model Releases

TideGS: Scalable Training of Over One Billion 3D Gaussian Splatting Primitives via Out-of-Core Optimization

DGX agent

arXiv:2605.20150v1 Announce Type: new Abstract: Training 3D Gaussian Splatting (3DGS) at billion-primitive scale is fundamentally memory-bound: each Gaussian primitive carries a large attribute vector

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Time-optimal neural feedback control of nilpotent systems as a binary classification problem

DGX agent

arXiv:2503.17581v2 Announce Type: replace-cross Abstract: A computational method for the synthesis of time-optimal feedback control laws for linear nilpotent systems is proposed. The method is based o

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

DGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents

DGX agent

arXiv:2605.18882v1 Announce Type: cross Abstract: LLM agents exhibit a consistent tendency to over-call, invoking tools even in situations where none is needed. On the When2Call benchmark, six models

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul Erdős in 1946. For nearly 80 …

DGX agent

Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul Erdős in 1946. For nearly 80 years, mathematicians believed the best possible solutions l

model-releasesopenai--x
20 May 2026
Model Releases

Toto 2.0: Time Series Forecasting Enters the Scaling Era

DGX agent

arXiv:2605.20119v1 Announce Type: cross Abstract: We show that time series foundation models scale: a single training recipe produces reliable forecast-quality improvements from 4M to 2.5B parameters.

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs

DGX agent

arXiv:2605.19528v1 Announce Type: new Abstract: 3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera i

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs

DGX agent

arXiv:2605.19798v1 Announce Type: new Abstract: As Socially Interactive Agents (SIAs) become increasingly integrated into daily life, the ability to calibrate user trust to an agent's actual capabilit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Training Neural Networks with Optimal Double-Bayesian Learning

DGX agent

arXiv:2605.20009v1 Announce Type: cross Abstract: Backpropagation with gradient descent is a common optimization strategy employed by most neural network architectures in machine learning. However, fi

model-releasesarxiv-cs-ai
20 May 2026
← Previous
1…288289290291292…471
Next →