AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,508 results
Local Ai

Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping

DGX agent

arXiv:2605.10937v1 Announce Type: new Abstract: Recently, post-training methods based on reinforcement learning, with a particular focus on Group Relative Policy Optimization (GRPO), have emerged as t

local-aiarxiv-cs-cv
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.10925v1 Announce Type: new Abstract: Large-scale pretraining has made Vision-Language-Action (VLA) models promising foundations for generalist robot manipulation, yet adapting them to downs

tutorialsarxiv-cs-ro
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Safety

Relative Score Policy Optimization for Diffusion Language Models

DGX agent

arXiv:2605.10218v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a promising route to parallel and efficient text generation, but improving their reasoning ability require

safetyarxiv-cs-cl
12 May 2026
Safety

RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.09410v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models remain brittle in long-horizon, contact-rich manipulation because success-only imitation provides little supervisi

safetyarxiv-cs-ai
12 May 2026
Research

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models

DGX agent

arXiv:2602.11824v2 Announce Type: replace Abstract: Despite the advanced capabilities of Large Vision-Language Models (LVLMs), they frequently suffer from object hallucination. One reason is that visu

researcharxiv-cs-ai
12 May 2026
Research

SynerDiff: Synergetic Continuous Batching for Fast and Parallel Diffusion Model Inference

DGX agent

arXiv:2605.08835v1 Announce Type: new Abstract: The expansion of Artificial Intelligence-generated content service requires diffusion model serving to simultaneously achieve high throughput and low ta

researcharxiv-cs-ai
12 May 2026
Research

The Astonishing Ability of Large Language Models to Parse Jabberwockified Language

DGX agent

arXiv:2602.23928v2 Announce Type: replace Abstract: We show that large language models (LLMs) have an astonishing ability to recover meaning from severely degraded English texts. Texts in which conten

researcharxiv-cs-cl
12 May 2026
Agents

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

DGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

agentsclem-delangue--x
12 May 2026
Applications

The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies (Connie Loizos/TechCrunch)

DGX agent

Connie Loizos / TechCrunch: The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies — Neil Batlivala has

applicationstechmeme
12 May 2026
Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

DGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

safetyyohei-nakajima--x
12 May 2026
Tutorials

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm

DGX agent

arXiv:2605.10640v1 Announce Type: cross Abstract: Continual Pre-Training (CPT) is essential for enabling Language Models (LMs) to integrate new knowledge without erasing old. While classical CPT techn

tutorialsarxiv-cs-ai
12 May 2026
Applications

TrajDLM: Topology-Aware Block Diffusion Language Model for Trajectory Generation

DGX agent

arXiv:2605.10020v1 Announce Type: new Abstract: Generating high-fidelity synthetic GPS trajectories is increasingly important for applications in transportation, urban planning, and what-if scenario s

applicationsarxiv-cs-lg
12 May 2026
Industry

We've just hit 1M open datasets on the Hugging Face Hub 🎉 Open models need open data. Today we hit that milestone, together with the most i…

DGX agent

We've just hit 1M open datasets on the Hugging Face Hub 🎉 Open models need open data. Today we hit that milestone, together with the most incredible community in AI! 🤗 Onwards to the next million 🚀 Me

industryclem-delangue--x
12 May 2026
Research

Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits

DGX agent

arXiv:2605.08200v1 Announce Type: new Abstract: A pervasive intuition holds that vision-language models (VLMs) are most trustworthy when their attention maps look sharp: concentrated attention on the

researcharxiv-cs-ai
12 May 2026
Tutorials

World Models: 10 Things That Matter in AI Right Now

DGX agent

World models recently made our list of 10 Things That Matter in AI Right Now. Watch executive editor Niall Firth explain why this emerging area of AI is gaining so much attention. Join MIT Technology

tutorialsmit-tech-review
12 May 2026
Research

A Behavioral Framework for Data-Driven Modeling of Nonlinear Systems in Vector-Valued Reproducing Kernel Hilbert Spaces

DGX agent

arXiv:2605.07052v1 Announce Type: cross Abstract: We generalize Jan Willems' behavioral approach to a class of discrete-time nonlinear systems in a vector-valued reproducing kernel Hilbert space (RKHS

researcharxiv-cs-lg
11 May 2026
Research

A Rod Flow Model for Adam at the Edge of Stability

DGX agent

arXiv:2605.06821v1 Announce Type: cross Abstract: Cohen et al. (arXiv:2207.14484) observed that adaptive gradient methods such as Adam operate at the edge of stability. While there has been significan

researcharxiv-cs-ai
11 May 2026
Applications

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models

DGX agent

arXiv:2605.07308v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have significantly advanced the capabilities of robotic agents in executing diverse tasks; however, they still face

applicationsarxiv-cs-ro
11 May 2026
Safety

Better Protein Function Prediction by Modeling Survivorship Bias

DGX agent

arXiv:2605.06879v1 Announce Type: new Abstract: Protein sequence data from nature exhibits survivorship bias: we only observe data from those organisms that survive and reproduce, while non-functional

safetyarxiv-cs-lg
11 May 2026
Tutorials

Bifurcation Models: Learning Set-Valued Solution Maps with Weight-Tied Dynamics

DGX agent

arXiv:2605.07277v1 Announce Type: cross Abstract: Many scientific and combinatorial problems admit multiple correct solutions, not a single label. Standard supervised learning resolves this ambiguity

tutorialsarxiv-cs-ai
11 May 2026
Research

Black-box model classification under the discriminative factorization

DGX agent

arXiv:2605.07878v1 Announce Type: new Abstract: Access to modern generative systems is often restricted to querying an API (the ``black-box' setting) and many properties of the system are unknown to t

researcharxiv-cs-lg
11 May 2026
Research

DIMoE-Adapters: Dynamic Expert Evolution for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07494v1 Announce Type: new Abstract: Continual learning enables vision-language models to accumulate knowledge and adapt to evolving tasks without retraining from scratch. However, in multi

researcharxiv-cs-cv
11 May 2026
Research

Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport

DGX agent

arXiv:2605.06785v1 Announce Type: cross Abstract: Inference-time scaling methods rely on Process Reward Models (PRMs), which are often poorly calibrated and overestimate success probabilities. We prop

researcharxiv-cs-ai
11 May 2026
Research

EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting

DGX agent

arXiv:2605.07642v1 Announce Type: new Abstract: Forecasting future 3D hand pose sequences from egocentric video is essential for understanding human intention and enabling embodied applications such a

researcharxiv-cs-cv
11 May 2026
Safety

Emergent Symbolic Structure in Health Foundation Models: Extraction, Alignment, and Cross-Modal Transfer

DGX agent

arXiv:2605.07407v1 Announce Type: new Abstract: Health foundation models (FMs) learn useful representations from wearable sensors, but interpreting what they encode and transferring that knowledge acr

safetyarxiv-cs-lg
11 May 2026
Research

How Do Language Models Compose Functions?

DGX agent

arXiv:2510.01685v2 Announce Type: replace-cross Abstract: While large language models (LLMs) appear to be increasingly capable of solving compositional tasks, it is an open question whether they do so

researcharxiv-cs-ai
11 May 2026
Local Ai

🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally with any compatible GGUF/MLX model > shipped native…

DGX agent

🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally with any compatible GGUF/MLX model > shipped native traces support for Hermes Agent: visualize your Hermes traces

local-aiclem-delangue--x
11 May 2026
Agents

I have a new job! Excited to announce that I will be working with Hugging Face to make local models work great in OpenClaw and other open ag…

DGX agent

I have a new job! Excited to announce that I will be working with Hugging Face to make local models work great in OpenClaw and other open agent harnesses! I will be building in public and documenting

agentsclem-delangue--x
11 May 2026
Research

ImplantMamba: Long-range Sequential Modeling Mamba For Dental Implant Position Prediction

DGX agent

arXiv:2605.07082v1 Announce Type: new Abstract: In the design of surgical guides for implant placement, determining the precise implant position is a critical step. However, the implant region itself

researcharxiv-cs-cv
11 May 2026
Research

Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

DGX agent

arXiv:2602.01166v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models benefit from chain-of-thought (CoT) reasoning, but existing approaches incur high inference overhead and rely on

researcharxiv-cs-ro
11 May 2026
Local Ai

Learned Lagrangian Models of PDEs via Euler-Lagrange Residual Minimization

DGX agent

arXiv:2605.07157v1 Announce Type: new Abstract: We present the first method to directly use a learned continuous Lagrangian to forecast the dynamics of systems governed by partial differential equatio

local-aiarxiv-cs-lg
11 May 2026
Research

Linear Response Estimators for Singular Statistical Models

DGX agent

arXiv:2605.07970v1 Announce Type: cross Abstract: We define susceptibilities as a measure of the response of an observable quantity of a parameterized statistical model to a perturbation of the data f

researcharxiv-cs-lg
11 May 2026
Research

MAST: A Multi-fidelity Augmented Surrogate model via Spatial Trust-weighting

DGX agent

arXiv:2602.20974v2 Announce Type: replace Abstract: In engineering design and scientific computing, computational cost and predictive accuracy are intrinsically coupled. High-fidelity simulations prov

researcharxiv-cs-lg
11 May 2026
Safety

Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models

DGX agent

arXiv:2605.08031v1 Announce Type: new Abstract: Vision-language models (VLMs) raise growing concerns about privacy, copyright, and bias, motivating machine unlearning to remove sensitive knowledge. Ho

safetyarxiv-cs-cv
11 May 2026
Research

Pre-trained Tabular Foundation Models as Versatile Summary Networks for Neural Posterior Estimation

DGX agent

arXiv:2605.07765v1 Announce Type: new Abstract: In this work, we study TabPFN as a training-free, modular summary network for simulation-based Bayesian inference (SBI). Tabular foundation models such

researcharxiv-cs-lg
11 May 2026
Research

Pretraining a Foundation Model for Small-Molecule Natural Products

DGX agent

arXiv:2503.17656v4 Announce Type: replace-cross Abstract: Natural products, as metabolites from microorganisms, animals, or plants, exhibit diverse biological activities, making them crucial for drug

researcharxiv-cs-ai
11 May 2026
Tutorials

Rethinking State Tracking in Recurrent Models Through Error Control Dynamics

DGX agent

arXiv:2605.07755v1 Announce Type: cross Abstract: The theory of state tracking in recurrent architectures has predominantly focused on expressive capacity: whether a fixed architecture can theoretical

tutorialsarxiv-cs-cl
11 May 2026
Research

Saliency-Aware Regularized Quantization Calibration for Large Language Models

DGX agent

arXiv:2605.05693v2 Announce Type: replace Abstract: Post-training quantization (PTQ) is an effective approach for deploying large language models (LLMs) under memory and latency constraints. Most exis

researcharxiv-cs-ai
11 May 2026
Industry

speaking of things that have gotten over a threshold for me, the combo of the new ChatGPT model, personality, and personalization feels like…

DGX agent

Sam Altman comments on OpenAI's new ChatGPT model, noting that the combination of improved capabilities, personality features, and personalization options has crossed an important threshold of functio

industrysam-altman--x
11 May 2026
Research

StreamPhy: Streaming Inference of High-Dimensional Physical Dynamics via State Space Models

DGX agent

arXiv:2605.07384v1 Announce Type: new Abstract: Inferring the evolution of high-dimensional and multi-modal (e.g., spatio-temporal) physical fields from irregular sparse measurements in real time is a

researcharxiv-cs-lg
11 May 2026
Agents

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, tel…

DGX agent

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, teleoperators, and real consequences. Shrey’s piece on simulati

agentsyohei-nakajima--x
11 May 2026
Tools

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building t…

DGX agent

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building toward this for a while. Visual intelligence is becoming real

toolsswyx--x
11 May 2026
Research

Three-in-One World Model: Energy-Based Consistency, Prediction, and Counterfactual Inference for Marketing Intervention

DGX agent

arXiv:2605.07199v1 Announce Type: new Abstract: Marketing decisions reflect the interaction of latent consumer heterogeneity, time-varying internal states, and explicit interventions, a structure that

researcharxiv-cs-ai
11 May 2026
Safety

Toward Better Geometric Representations for Molecule Generative Models

DGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

safetyarxiv-cs-lg
11 May 2026
Research

TTF: Temporal Token Fusion for Efficient Video-Language Model

DGX agent

arXiv:2605.07355v1 Announce Type: cross Abstract: Video-language models (VLMs) face rapid inference costs as visual token counts scale with video length. For example, 32 frames at 448{imes}448 resolut

researcharxiv-cs-ai
11 May 2026
Research

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

DGX agent

arXiv:2605.07969v1 Announce Type: new Abstract: Diffusion models perform remarkably well on high-dimensional data such as images, often using only a modest number of reverse-time steps. Despite this p

researcharxiv-cs-lg
11 May 2026
Tutorials

With the model's simultaneous speech capability, Horace has gotten a lot easier to work with recently.

DGX agent

The post discusses improvements in working with Horace (likely a tool or system) due to recent implementation of simultaneous speech capability in its underlying model. This advancement has made the i

tutorialsjeremy-howard--x
11 May 2026
← Previous
1…281282283284285…1303
Next →