AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlog
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
Research

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embe…

DGX agent

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embeddings averaged within each bag, and predicts the next bag o

researchnous-research--x
13 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Industry

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates…

DGX agent

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates more risk than openness. 'Six, seven years ago, at the time

industryclem-delangue--x
13 May 2026
Tutorials

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

DGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

tutorialsarxiv-cs-ai
12 May 2026
Research

ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models

DGX agent

arXiv:2605.08612v1 Announce Type: new Abstract: Addressing the escalating security vulnerabilities in Vision-Language-Action (VLA) models, this study investigates backdoor attacks targeting the visual

researcharxiv-cs-ro
12 May 2026
Agents

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

DGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

agentsarxiv-cs-ai
12 May 2026
Applications

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

DGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

applicationsarxiv-cs-ai
12 May 2026
Research

Can Muon Fine-tune Adam-Pretrained Models?

DGX agent

arXiv:2605.10468v1 Announce Type: new Abstract: Muon has emerged as an efficient alternative to Adam for pretraining, yet remains underused for fine-tuning. A key obstacle is that most open models are

researcharxiv-cs-lg
12 May 2026
Safety

Composing Policy Gradients and Prompt Optimization for Language Model Programs

DGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

safetyarxiv-cs-cl
12 May 2026
Research

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd …

DGX agent

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd - CfP and details on the website: https://wm-booth.org - @yl

researchyann-lecun--x
12 May 2026
Local Ai

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding

DGX agent

arXiv:2512.06673v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are rapidly expanding from general video understanding to finer-grained understanding such as spatio-tempor

local-aiarxiv-cs-cv
12 May 2026
Local Ai

DP-LAC: Lightweight Adaptive Clipping for Differentially Private Federated Fine-tuning of Language Models

DGX agent

arXiv:2605.10272v1 Announce Type: cross Abstract: Federated learning (FL) enables the collaborative training of large-scale language models (LLMs) across edge devices while keeping user data on-device

local-aiarxiv-cs-ai
12 May 2026
Safety

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

DGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

safetyarxiv-cs-ai
12 May 2026
Research

FERA: Uncertainty-Aware Federated Reasoning for Large Language Models

DGX agent

arXiv:2605.10082v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong reasoning capabilities when guided by high-quality demonstrations, yet such data is often distributed across

researcharxiv-cs-cl
12 May 2026
Applications

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

DGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

applicationsarxiv-cs-ai
12 May 2026
Local Ai

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

DGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

local-aiarxiv-cs-lg
12 May 2026
Research

Functional Subspace, where language models can use vector algebra to solve problems

DGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

researcharxiv-cs-ai
12 May 2026
Agents

GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents

DGX agent

arXiv:2510.13896v2 Announce Type: replace-cross Abstract: Cellular image segmentation is essential for quantitative biology yet remains difficult due to heterogeneous modalities, morphological variabi

agentsarxiv-cs-ai
12 May 2026
Research

Generative Giants, Retrieval Weaklings: Why do Multimodal Large Language Models Fail at Multimodal Retrieval?

DGX agent

arXiv:2512.19115v2 Announce Type: replace Abstract: Despite the remarkable success of multimodal large language models (MLLMs) in generative tasks, we observe that they exhibit a counterintuitive defi

researcharxiv-cs-cv
12 May 2026
Safety

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

DGX agent

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

safetyarxiv-cs-ai
12 May 2026
Tools

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can …

DGX agent

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can now find the right voice for their app faster by describing

toolstogether-ai--x
12 May 2026
Agents

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory ha…

DGX agent

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory has a lot of gravity.' - @hwchase17, Co-Founder & CEO, @LangCh

agentsharrison-chase--x
12 May 2026
Research

Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training

DGX agent

arXiv:2605.09994v1 Announce Type: cross Abstract: Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evo

researcharxiv-cs-lg
12 May 2026
Safety

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

DGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

safetyarxiv-cs-lg
12 May 2026
Local Ai

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

DGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

local-aiarxiv-cs-lg
12 May 2026
Research

Learning predictive models for combinations of heterogeneous proteomic data sources

DGX agent

arXiv:2605.08958v1 Announce Type: new Abstract: Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease

researcharxiv-cs-lg
12 May 2026
Safety

LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models

DGX agent

arXiv:2605.09948v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models typically treat the deepest representation of a vision-language backbone as universally optimal for action p

safetyarxiv-cs-ai
12 May 2026
Research

Majority Bit-Aware Watermarking For Large Language Models

DGX agent

arXiv:2508.03829v2 Announce Type: replace Abstract: The growing deployment of Large Language Models (LLMs) has raised concerns about their misuse in generating harmful or deceptive content. To address

researcharxiv-cs-cl
12 May 2026
Hardware

mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters

DGX agent

arXiv:2605.08300v1 Announce Type: cross Abstract: Manifold-Constrained Hyper-Connections (mHC) introduce a stability-motivated variant of multi stream residual mixing by constraining residual stream m

hardwarearxiv-cs-ai
12 May 2026
Safety

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

DGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

safetyarxiv-cs-lg
12 May 2026
Research

Model-Reference Adaptive Flight Control of the 95-mg Bee++

DGX agent

arXiv:2605.08525v1 Announce Type: new Abstract: We introduce a model-reference adaptive control (MRAC) architecture for high-performance positional tracking of the Bee++, a 95-mg insect-scale flapping

researcharxiv-cs-ro
12 May 2026
Local Ai

MolWorld: Molecule World Models for Actionable Molecular Optimization

DGX agent

arXiv:2605.08954v1 Announce Type: cross Abstract: Molecular optimization in drug discovery aims to discover molecules with improved target properties, but practical lead optimization often requires mo

local-aiarxiv-cs-ai
12 May 2026
Safety

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

DGX agent

arXiv:2605.09258v1 Announce Type: cross Abstract: Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of mo

safetyarxiv-cs-ai
12 May 2026
Research

Not-So-Strange Love: Language Models and Generative Linguistic Theories are More Compatible than They Appear

DGX agent

arXiv:2605.10061v1 Announce Type: cross Abstract: Futrell and Mahowald (2025) frame the success of neural language models (LMs) as supporting gradient, usage-based linguistic theories. I argue that LM

researcharxiv-cs-ai
12 May 2026
Research

Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens and paid tools toget…

DGX agent

Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens and paid tools together for hassle-free setup and simple billing. http://portal.

researchnous-research--x
12 May 2026
Research

On the global convergence of gradient descent for wide shallow models with bounded nonlinearities

DGX agent

arXiv:2605.10775v1 Announce Type: cross Abstract: A surprising phenomenon in the training of neural networks is the ability of gradient descent to find global minimizers of the training loss despite i

researcharxiv-cs-lg
12 May 2026
Local Ai

Path-Dependent Denoising: A Non-Conservative Field Perspective on Order Collapse in Diffusion Language Models

DGX agent

arXiv:2605.09303v1 Announce Type: new Abstract: Diffusion language models (DLMs) offer a structural alternative to autoregressive generation: denoising can update tokens in arbitrary orders or in para

local-aiarxiv-cs-lg
12 May 2026
Local Ai

Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping

DGX agent

arXiv:2605.10937v1 Announce Type: new Abstract: Recently, post-training methods based on reinforcement learning, with a particular focus on Group Relative Policy Optimization (GRPO), have emerged as t

local-aiarxiv-cs-cv
12 May 2026
Tutorials

PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.10925v1 Announce Type: new Abstract: Large-scale pretraining has made Vision-Language-Action (VLA) models promising foundations for generalist robot manipulation, yet adapting them to downs

tutorialsarxiv-cs-ro
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Safety

Relative Score Policy Optimization for Diffusion Language Models

DGX agent

arXiv:2605.10218v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a promising route to parallel and efficient text generation, but improving their reasoning ability require

safetyarxiv-cs-cl
12 May 2026
Safety

RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.09410v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models remain brittle in long-horizon, contact-rich manipulation because success-only imitation provides little supervisi

safetyarxiv-cs-ai
12 May 2026
Research

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models

DGX agent

arXiv:2602.11824v2 Announce Type: replace Abstract: Despite the advanced capabilities of Large Vision-Language Models (LVLMs), they frequently suffer from object hallucination. One reason is that visu

researcharxiv-cs-ai
12 May 2026
Research

SynerDiff: Synergetic Continuous Batching for Fast and Parallel Diffusion Model Inference

DGX agent

arXiv:2605.08835v1 Announce Type: new Abstract: The expansion of Artificial Intelligence-generated content service requires diffusion model serving to simultaneously achieve high throughput and low ta

researcharxiv-cs-ai
12 May 2026
Research

The Astonishing Ability of Large Language Models to Parse Jabberwockified Language

DGX agent

arXiv:2602.23928v2 Announce Type: replace Abstract: We show that large language models (LLMs) have an astonishing ability to recover meaning from severely degraded English texts. Texts in which conten

researcharxiv-cs-cl
12 May 2026
Agents

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

DGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

agentsclem-delangue--x
12 May 2026
Applications

The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies (Connie Loizos/TechCrunch)

DGX agent

Connie Loizos / TechCrunch: The US' Centers for Medicare & Medicaid Services is testing ACCESS, an outcome-based payment model for AI-driven medical care, with 150 tech companies — Neil Batlivala has

applicationstechmeme
12 May 2026
Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

DGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

safetyyohei-nakajima--x
12 May 2026
Tutorials

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm

DGX agent

arXiv:2605.10640v1 Announce Type: cross Abstract: Continual Pre-Training (CPT) is essential for enabling Language Models (LMs) to integrate new knowledge without erasing old. While classical CPT techn

tutorialsarxiv-cs-ai
12 May 2026
← Previous
1…278279280281282…1294
Next →