AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlog
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
Agents

announcing deepagents v0.6, our biggest release yet! it’s all about performance: at the model layer w harness profiles, agent layer w code i…

DGX agent

announcing deepagents v0.6, our biggest release yet! it’s all about performance: at the model layer w harness profiles, agent layer w code interpreter, and at scale w streaming and delta channels cont

agentsharrison-chase--x
18 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models

DGX agent

arXiv:2605.15755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can produce fluent artwork emotion explanations, but they often suffer from attribute flooding: they enumerate

agentsarxiv-cs-cv
18 May 2026
Research

BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models

DGX agent

arXiv:2602.04163v2 Announce Type: replace Abstract: Large language model inference is often bounded by memory footprint and bandwidth in resource-constrained deployments, making quantization fundament

researcharxiv-cs-lg
18 May 2026
Tutorials

Constrained latent state modeling: A unifying perspective on representation learning under competing constraints

DGX agent

arXiv:2605.15995v1 Announce Type: cross Abstract: Learning latent representations from complex data is central to modern machine learning, spanning temporal, multimodal, and partially observed systems

tutorialsarxiv-cs-ai
18 May 2026
Safety

DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation

DGX agent

arXiv:2605.16113v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved unprecedented success due to their exceptional generative capabilities. However, because they depend on kno

safetyarxiv-cs-ai
18 May 2026
Agents

Differentiable Mixture-of-Agents Incentivizes Swarm Intelligence of Large Language Models

DGX agent

arXiv:2605.15706v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have catalyzed the development of multi-agent systems (MAS) for complex reasoning tasks. However, existi

agentsarxiv-cs-lg
18 May 2026
Safety

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?

DGX agent

arXiv:2605.15855v1 Announce Type: new Abstract: Despite strong image-generation performance, diffusion models' reconstruction objectives limit alignment with human preferences. RL enables such alignme

safetyarxiv-cs-cv
18 May 2026
Safety

Embedding-perturbed Exploration Preference Optimization for Flow Models

DGX agent

arXiv:2605.15803v1 Announce Type: new Abstract: Recent advancements have established Reinforcement Learning (RL) as a pivotal paradigm for aligning generative models with human intent. However, group-

safetyarxiv-cs-cv
18 May 2026
Safety

From Model Design to Organizational Design: Complexity Redistribution and Trade-Offs in Generative AI

DGX agent

arXiv:2506.22440v2 Announce Type: replace-cross Abstract: This paper introduces the Generality-Accuracy-Simplicity (GAS) framework to analyze how large language models (LLMs) are reshaping organizatio

safetyarxiv-cs-lg
18 May 2026
Local Ai

I believe on-prem and local AI - based on @huggingface open-source models - will be an important answer to the GPU shortages this year (beca…

DGX agent

I believe on-prem and local AI - based on @huggingface open-source models - will be an important answer to the GPU shortages this year (because they are cheaper, faster, safer than cloud APIs)! Great

local-aiclem-delangue--x
18 May 2026
Research

Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds

DGX agent

arXiv:2605.15822v1 Announce Type: new Abstract: Score-based generative models are trained in high-dimensional ambient spaces, yet many data distributions are supported on low-dimensional nonlinear str

researcharxiv-cs-lg
18 May 2026
Agents

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in …

DGX agent

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in real-time. Try our playable research preview today, with Ago

agentsyohei-nakajima--x
18 May 2026
Agents

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say…

DGX agent

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say a neutral observability layer still wins. http://venturebea

agentsharrison-chase--x
18 May 2026
Research

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation

DGX agent

arXiv:2605.15831v1 Announce Type: cross Abstract: Autoregressive music generation depends strongly on the audio tokenizer. Existing high-fidelity codecs often use residual multi-codebook quantization,

researcharxiv-cs-ai
18 May 2026
Local Ai

Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection

DGX agent

arXiv:2605.16107v1 Announce Type: new Abstract: Machine-generated texts (MGTs) pose risks such as disinformation and phishing, underscoring the need for reliable detection. Metric-based methods, which

local-aiarxiv-cs-cl
18 May 2026
Tutorials

okay maybe it's a good time? We have a small colbert model trained at pplx, it is a continue-training of pplx-embed-0.6b, so native multilin…

DGX agent

okay maybe it's a good time? We have a small colbert model trained at pplx, it is a continue-training of pplx-embed-0.6b, so native multilingual, just made it open and added a section how to use MaxSi

tutorialsclem-delangue--x
18 May 2026
Agents

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

DGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

agentsarxiv-cs-ai
18 May 2026
Industry

To protect passengers or cargo, the powered rear seats & trunk in Model Y will automatically pop back up if detecting an obstruction while f…

DGX agent

Tesla Model Y's powered rear seats and trunk are equipped with automatic obstruction detection that causes them to automatically reverse and pop back up if an obstruction is detected during operation,

industryelon-musk--x
18 May 2026
Industry

Developers say Chinese AI labs lead US rivals in video generation, as ByteDance and Kuaishou train models on vast short-form video libraries from their own apps (Eleanor Olcott/Financial Times)

DGX agent

Eleanor Olcott / Financial Times: Developers say Chinese AI labs lead US rivals in video generation, as ByteDance and Kuaishou train models on vast short-form video libraries from their own apps — Chi

industrytechmeme
17 May 2026
Agents

Publicis agrees to acquire LiveRamp, which allows companies to share and build new data sets and models that can power agentic frameworks, for $2.2B in cash (Alison Weissbrot/Adweek)

DGX agent

Alison Weissbrot / Adweek: Publicis agrees to acquire LiveRamp, which allows companies to share and build new data sets and models that can power agentic frameworks, for 2.2B in cash — Publicis Groupe

agentstechmeme
17 May 2026
Hardware

A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 — Every major A

hardwaretechmeme
15 May 2026
Industry

AI teams shouldn’t have to choose between expensive object storage and painful git workflows. @huggingface Storage is built for model weight…

DGX agent

AI teams shouldn’t have to choose between expensive object storage and painful git workflows. @huggingface Storage is built for model weights, datasets, checkpoints and artifacts: - simple per-TB pric

industryclem-delangue--x
15 May 2026
Research

Beyond What to Select: A Plug-and-play Oscillatory Data-Volume Scheduling for Efficient Model Training

DGX agent

arXiv:2605.14773v1 Announce Type: cross Abstract: Data selection accelerates training by identifying representative training data while preserving model performance. However, existing methods mainly f

researcharxiv-cs-ai
15 May 2026
Agents

BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: O…

DGX agent

BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: Opus 4.7 by @AnthropicAI 2nd: Opus 4.7 (Thinking) by @Anthrop

agentszhipu-ai--x
15 May 2026
Agents

Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

DGX agent

arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U

agentsarxiv-cs-ai
15 May 2026
Applications

DT-Transformer: A Foundation Model for Disease Trajectory Prediction on a Real-world Health System

DGX agent

arXiv:2605.14227v1 Announce Type: cross Abstract: Accurate disease trajectory prediction is critical for early intervention, resource allocation, and improving long-term outcomes. While electronic hea

applicationsarxiv-cs-cl
15 May 2026
Safety

Exploring Geographic Relative Space in Large Language Models through Activation Patching

DGX agent

arXiv:2605.14535v1 Announce Type: new Abstract: The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide ra

safetyarxiv-cs-lg
15 May 2026
Agents

Graph of States: Solving Abductive Tasks with Large Language Models

DGX agent

arXiv:2603.21250v2 Announce Type: replace Abstract: Logical reasoning encompasses deduction, induction, and abduction. However, while Large Language Models (LLMs) have effectively mastered the former

agentsarxiv-cs-ai
15 May 2026
Research

Image Restoration via Diffusion Models with Dynamic Resolution

DGX agent

arXiv:2605.14267v1 Announce Type: cross Abstract: Diffusion models (DMs) have exhibited remarkable efficacy in various image restoration tasks. However, existing approaches typically operate within th

researcharxiv-cs-ai
15 May 2026
Industry

In a viral X post that parodies the old Mac vs. PC commercials, General Catalyst posted a 'VC vs GC' video, with the VC apparently modeled after Marc Andreessen (Julie Bort/TechCrunch)

DGX agent

Julie Bort / TechCrunch: In a viral X post that parodies the old Mac vs. PC commercials, General Catalyst posted a “VC vs GC” video, with the VC apparently modeled after Marc Andreessen — One of the m

industrytechmeme
15 May 2026
Safety

Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control

DGX agent

arXiv:2605.14935v1 Announce Type: new Abstract: We present MSCoT, a multi-scale, coarse-to-fine model for test-time human motion synthesis and control. Unlike recent approaches that rely on multiple i

safetyarxiv-cs-cv
15 May 2026
Applications

MultiMat: Multimodal Program Synthesis for Procedural Materials using Large Multimodal Models

DGX agent

arXiv:2509.22151v3 Announce Type: replace Abstract: Material node graphs are programs that generate the 2D channels of procedural materials, including geometry such as roughness and displacement maps,

applicationsarxiv-cs-cv
15 May 2026
Research

SeaVis: Modeling and Control of a Remotely Operated Towed Vehicle for Seabed Visualization and Mapping

DGX agent

arXiv:2605.14683v1 Announce Type: new Abstract: High-resolution seafloor mapping necessitates stable and precise positioning for underwater robots. This paper introduces a novel mathematical model for

researcharxiv-cs-ro
15 May 2026
Local Ai

TRIO: Token Reduction via Inference-Objective Guidance for Efficient Vision-Language Models

DGX agent

arXiv:2602.04657v3 Announce Type: replace Abstract: Recently, reducing redundant visual tokens in vision-language models (VLMs) to accelerate VLM inference has emerged as a hot topic. However, most ex

local-aiarxiv-cs-cv
15 May 2026
Industry

We’re releasing a 30B-A3B reasoning model that reaches gold-medal level across both physics and math Olympiad evaluations: IPhO directly, an…

DGX agent

We’re releasing a 30B-A3B reasoning model that reaches gold-medal level across both physics and math Olympiad evaluations: IPhO directly, and IMO/USAMO with test-time self-verification and refinement.

industryclem-delangue--x
15 May 2026
Research

A Hierarchical Language Model with Predictable Scaling Laws and Provable Benefits of Reasoning

DGX agent

arXiv:2605.13687v1 Announce Type: cross Abstract: We introduce a family of synthetic languages with hierarchical structure -- generated by a broadcast process on trees -- for which the role of context

researcharxiv-cs-ai
14 May 2026
Industry

Are scaling laws finally working for time series foundation models? Today, @datadoghq is releasing Toto 2.0 weights in Apache 2.0 on @huggin…

DGX agent

Are scaling laws finally working for time series foundation models? Today, @datadoghq is releasing Toto 2.0 weights in Apache 2.0 on @huggingface. It's a family of open-weights TSFMs from 4M to 2.5B p

industryclem-delangue--x
14 May 2026
Research

Assessing the Creativity of Large Language Models: Testing, Limits, and New Frontiers

DGX agent

arXiv:2605.13450v1 Announce Type: new Abstract: Measuring the creativity of large language models (LLMs) is essential for designing methods that can improve creativity and for enhancing our scientific

researcharxiv-cs-ai
14 May 2026
Local Ai

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements t…

DGX agent

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements to caching for faster inference overall. Turn on Developer Mod

local-ailm-studio--x
14 May 2026
Research

CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models

DGX agent

arXiv:2605.13178v1 Announce Type: cross Abstract: In large vision-language models, visual tokens typically constitute the majority of input tokens, leading to substantial computational overhead. To ad

researcharxiv-cs-ai
14 May 2026
Safety

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

DGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

safetyarxiv-cs-lg
14 May 2026
Industry

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me …

DGX agent

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me flexibility. Give me delight. Give me color. Please create t

industryallie-k--miller--x
14 May 2026
Safety

DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models

DGX agent

arXiv:2605.12702v1 Announce Type: new Abstract: General-purpose safety benchmarks for large language models do not adequately evaluate disability-related harms. We introduce DisaBench: a taxonomy of t

safetyarxiv-cs-ai
14 May 2026
Research

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models

DGX agent

arXiv:2605.13156v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in bridging visual perception and natural language understanding, enabling a wid

researcharxiv-cs-cv
14 May 2026
Research

Generative Modeling by Minimizing the Wasserstein-2 Loss

DGX agent

arXiv:2406.13619v4 Announce Type: replace-cross Abstract: This paper develops a generative model by minimizing the second-order Wasserstein loss (the W_2 loss) through a distribution-dependent ordinar

researcharxiv-cs-lg
14 May 2026
Local Ai

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

local-aiarxiv-cs-ai
14 May 2026
Local Ai

Learning to See What You Need: Gaze Attention for Multimodal Large Language Models

DGX agent

arXiv:2605.13080v1 Announce Type: new Abstract: When humans describe a visual scene, they do not process the entire image uniformly; instead, they selectively fixate on regions relevant to their inten

local-aiarxiv-cs-cv
14 May 2026
Applications

MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling

DGX agent

arXiv:2605.13711v1 Announce Type: new Abstract: Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels.

applicationsarxiv-cs-lg
14 May 2026
← Previous
1…276277278279280…1293
Next →