AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,137 results
Research

Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training

DGX agent

arXiv:2605.09994v1 Announce Type: cross Abstract: Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evo

researcharxiv-cs-lg
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

DGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

safetyarxiv-cs-lg
12 May 2026
Local Ai

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

DGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

local-aiarxiv-cs-lg
12 May 2026
Research

Learning predictive models for combinations of heterogeneous proteomic data sources

DGX agent

arXiv:2605.08958v1 Announce Type: new Abstract: Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease

researcharxiv-cs-lg
12 May 2026
Safety

LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models

DGX agent

arXiv:2605.09948v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models typically treat the deepest representation of a vision-language backbone as universally optimal for action p

safetyarxiv-cs-ai
12 May 2026
Research

Majority Bit-Aware Watermarking For Large Language Models

DGX agent

arXiv:2508.03829v2 Announce Type: replace Abstract: The growing deployment of Large Language Models (LLMs) has raised concerns about their misuse in generating harmful or deceptive content. To address

researcharxiv-cs-cl
12 May 2026
Hardware

mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters

DGX agent

arXiv:2605.08300v1 Announce Type: cross Abstract: Manifold-Constrained Hyper-Connections (mHC) introduce a stability-motivated variant of multi stream residual mixing by constraining residual stream m

hardwarearxiv-cs-ai
12 May 2026
Safety

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

DGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

safetyarxiv-cs-lg
12 May 2026
Research

Model-Reference Adaptive Flight Control of the 95-mg Bee++

DGX agent

arXiv:2605.08525v1 Announce Type: new Abstract: We introduce a model-reference adaptive control (MRAC) architecture for high-performance positional tracking of the Bee++, a 95-mg insect-scale flapping

researcharxiv-cs-ro
12 May 2026
Local Ai

MolWorld: Molecule World Models for Actionable Molecular Optimization

DGX agent

arXiv:2605.08954v1 Announce Type: cross Abstract: Molecular optimization in drug discovery aims to discover molecules with improved target properties, but practical lead optimization often requires mo

local-aiarxiv-cs-ai
12 May 2026
Safety

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

DGX agent

arXiv:2605.09258v1 Announce Type: cross Abstract: Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of mo

safetyarxiv-cs-ai
12 May 2026
Research

Not-So-Strange Love: Language Models and Generative Linguistic Theories are More Compatible than They Appear

DGX agent

arXiv:2605.10061v1 Announce Type: cross Abstract: Futrell and Mahowald (2025) frame the success of neural language models (LMs) as supporting gradient, usage-based linguistic theories. I argue that LM

researcharxiv-cs-ai
12 May 2026
Research

Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens and paid tools toget…

DGX agent

Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens and paid tools together for hassle-free setup and simple billing. http://portal.

researchnous-research--x
12 May 2026
Research

On the global convergence of gradient descent for wide shallow models with bounded nonlinearities

DGX agent

arXiv:2605.10775v1 Announce Type: cross Abstract: A surprising phenomenon in the training of neural networks is the ability of gradient descent to find global minimizers of the training loss despite i

researcharxiv-cs-lg
12 May 2026
Local Ai

Path-Dependent Denoising: A Non-Conservative Field Perspective on Order Collapse in Diffusion Language Models

DGX agent

arXiv:2605.09303v1 Announce Type: new Abstract: Diffusion language models (DLMs) offer a structural alternative to autoregressive generation: denoising can update tokens in arbitrary orders or in para

local-aiarxiv-cs-lg
12 May 2026
Local Ai

Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping

DGX agent

arXiv:2605.10937v1 Announce Type: new Abstract: Recently, post-training methods based on reinforcement learning, with a particular focus on Group Relative Policy Optimization (GRPO), have emerged as t

local-aiarxiv-cs-cv
12 May 2026
Tutorials

PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.10925v1 Announce Type: new Abstract: Large-scale pretraining has made Vision-Language-Action (VLA) models promising foundations for generalist robot manipulation, yet adapting them to downs

tutorialsarxiv-cs-ro
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Safety

Relative Score Policy Optimization for Diffusion Language Models

DGX agent

arXiv:2605.10218v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a promising route to parallel and efficient text generation, but improving their reasoning ability require

safetyarxiv-cs-cl
12 May 2026
Safety

RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.09410v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models remain brittle in long-horizon, contact-rich manipulation because success-only imitation provides little supervisi

safetyarxiv-cs-ai
12 May 2026
Research

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models

DGX agent

arXiv:2602.11824v2 Announce Type: replace Abstract: Despite the advanced capabilities of Large Vision-Language Models (LVLMs), they frequently suffer from object hallucination. One reason is that visu

researcharxiv-cs-ai
12 May 2026
Research

SynerDiff: Synergetic Continuous Batching for Fast and Parallel Diffusion Model Inference

DGX agent

arXiv:2605.08835v1 Announce Type: new Abstract: The expansion of Artificial Intelligence-generated content service requires diffusion model serving to simultaneously achieve high throughput and low ta

researcharxiv-cs-ai
12 May 2026
Research

The Astonishing Ability of Large Language Models to Parse Jabberwockified Language

DGX agent

arXiv:2602.23928v2 Announce Type: replace Abstract: We show that large language models (LLMs) have an astonishing ability to recover meaning from severely degraded English texts. Texts in which conten

researcharxiv-cs-cl
12 May 2026
Agents

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

DGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

agentsclem-delangue--x
12 May 2026
Applications

The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies (Connie Loizos/TechCrunch)

DGX agent

Connie Loizos / TechCrunch: The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies — Neil Batlivala has

applicationstechmeme
12 May 2026
Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

DGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

safetyyohei-nakajima--x
12 May 2026
Tutorials

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm

DGX agent

arXiv:2605.10640v1 Announce Type: cross Abstract: Continual Pre-Training (CPT) is essential for enabling Language Models (LMs) to integrate new knowledge without erasing old. While classical CPT techn

tutorialsarxiv-cs-ai
12 May 2026
Applications

TrajDLM: Topology-Aware Block Diffusion Language Model for Trajectory Generation

DGX agent

arXiv:2605.10020v1 Announce Type: new Abstract: Generating high-fidelity synthetic GPS trajectories is increasingly important for applications in transportation, urban planning, and what-if scenario s

applicationsarxiv-cs-lg
12 May 2026
Industry

We've just hit 1M open datasets on the Hugging Face Hub 🎉 Open models need open data. Today we hit that milestone, together with the most i…

DGX agent

We've just hit 1M open datasets on the Hugging Face Hub 🎉 Open models need open data. Today we hit that milestone, together with the most incredible community in AI! 🤗 Onwards to the next million 🚀 Me

industryclem-delangue--x
12 May 2026
Research

Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits

DGX agent

arXiv:2605.08200v1 Announce Type: new Abstract: A pervasive intuition holds that vision-language models (VLMs) are most trustworthy when their attention maps look sharp: concentrated attention on the

researcharxiv-cs-ai
12 May 2026
Tutorials

World Models: 10 Things That Matter in AI Right Now

DGX agent

World models recently made our list of 10 Things That Matter in AI Right Now. Watch executive editor Niall Firth explain why this emerging area of AI is gaining so much attention. Join MIT Technology

tutorialsmit-tech-review
12 May 2026
Research

A Behavioral Framework for Data-Driven Modeling of Nonlinear Systems in Vector-Valued Reproducing Kernel Hilbert Spaces

DGX agent

arXiv:2605.07052v1 Announce Type: cross Abstract: We generalize Jan Willems' behavioral approach to a class of discrete-time nonlinear systems in a vector-valued reproducing kernel Hilbert space (RKHS

researcharxiv-cs-lg
11 May 2026
Research

A Rod Flow Model for Adam at the Edge of Stability

DGX agent

arXiv:2605.06821v1 Announce Type: cross Abstract: Cohen et al. (arXiv:2207.14484) observed that adaptive gradient methods such as Adam operate at the edge of stability. While there has been significan

researcharxiv-cs-ai
11 May 2026
Applications

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models

DGX agent

arXiv:2605.07308v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have significantly advanced the capabilities of robotic agents in executing diverse tasks; however, they still face

applicationsarxiv-cs-ro
11 May 2026
Safety

Better Protein Function Prediction by Modeling Survivorship Bias

DGX agent

arXiv:2605.06879v1 Announce Type: new Abstract: Protein sequence data from nature exhibits survivorship bias: we only observe data from those organisms that survive and reproduce, while non-functional

safetyarxiv-cs-lg
11 May 2026
Tutorials

Bifurcation Models: Learning Set-Valued Solution Maps with Weight-Tied Dynamics

DGX agent

arXiv:2605.07277v1 Announce Type: cross Abstract: Many scientific and combinatorial problems admit multiple correct solutions, not a single label. Standard supervised learning resolves this ambiguity

tutorialsarxiv-cs-ai
11 May 2026
Research

Black-box model classification under the discriminative factorization

DGX agent

arXiv:2605.07878v1 Announce Type: new Abstract: Access to modern generative systems is often restricted to querying an API (the ``black-box' setting) and many properties of the system are unknown to t

researcharxiv-cs-lg
11 May 2026
Research

DIMoE-Adapters: Dynamic Expert Evolution for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07494v1 Announce Type: new Abstract: Continual learning enables vision-language models to accumulate knowledge and adapt to evolving tasks without retraining from scratch. However, in multi

researcharxiv-cs-cv
11 May 2026
Research

Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport

DGX agent

arXiv:2605.06785v1 Announce Type: cross Abstract: Inference-time scaling methods rely on Process Reward Models (PRMs), which are often poorly calibrated and overestimate success probabilities. We prop

researcharxiv-cs-ai
11 May 2026
Research

EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting

DGX agent

arXiv:2605.07642v1 Announce Type: new Abstract: Forecasting future 3D hand pose sequences from egocentric video is essential for understanding human intention and enabling embodied applications such a

researcharxiv-cs-cv
11 May 2026
Safety

Emergent Symbolic Structure in Health Foundation Models: Extraction, Alignment, and Cross-Modal Transfer

DGX agent

arXiv:2605.07407v1 Announce Type: new Abstract: Health foundation models (FMs) learn useful representations from wearable sensors, but interpreting what they encode and transferring that knowledge acr

safetyarxiv-cs-lg
11 May 2026
Research

How Do Language Models Compose Functions?

DGX agent

arXiv:2510.01685v2 Announce Type: replace-cross Abstract: While large language models (LLMs) appear to be increasingly capable of solving compositional tasks, it is an open question whether they do so

researcharxiv-cs-ai
11 May 2026
Local Ai

🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally with any compatible GGUF/MLX model > shipped native…

DGX agent

🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally with any compatible GGUF/MLX model > shipped native traces support for Hermes Agent: visualize your Hermes traces

local-aiclem-delangue--x
11 May 2026
Agents

I have a new job! Excited to announce that I will be working with Hugging Face to make local models work great in OpenClaw and other open ag…

DGX agent

I have a new job! Excited to announce that I will be working with Hugging Face to make local models work great in OpenClaw and other open agent harnesses! I will be building in public and documenting

agentsclem-delangue--x
11 May 2026
Research

ImplantMamba: Long-range Sequential Modeling Mamba For Dental Implant Position Prediction

DGX agent

arXiv:2605.07082v1 Announce Type: new Abstract: In the design of surgical guides for implant placement, determining the precise implant position is a critical step. However, the implant region itself

researcharxiv-cs-cv
11 May 2026
Research

Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

DGX agent

arXiv:2602.01166v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models benefit from chain-of-thought (CoT) reasoning, but existing approaches incur high inference overhead and rely on

researcharxiv-cs-ro
11 May 2026
Local Ai

Learned Lagrangian Models of PDEs via Euler-Lagrange Residual Minimization

DGX agent

arXiv:2605.07157v1 Announce Type: new Abstract: We present the first method to directly use a learned continuous Lagrangian to forecast the dynamics of systems governed by partial differential equatio

local-aiarxiv-cs-lg
11 May 2026
Research

Linear Response Estimators for Singular Statistical Models

DGX agent

arXiv:2605.07970v1 Announce Type: cross Abstract: We define susceptibilities as a measure of the response of an observable quantity of a parameterized statistical model to a perturbation of the data f

researcharxiv-cs-lg
11 May 2026
← Previous
1…284285286287288…1316
Next →