AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
18 May 2026

An overview of macro tech trends: capex explosion, unprecedented surge in chip demand, supply chain bottlenecks, model commoditization, AI automation, and more (Benedict Evans)

IndustryDGX agent

Benedict Evans: An overview of macro tech trends: capex explosion, unprecedented surge in chip demand, supply chain bottlenecks, model commoditization, AI automation, and more — Twice a year, I produc

announcing deepagents v0.6, our biggest release yet! it’s all about performance: at the model layer w harness profiles, agent layer w code i…

AgentsDGX agent

announcing deepagents v0.6, our biggest release yet! it’s all about performance: at the model layer w harness profiles, agent layer w code interpreter, and at scale w streaming and delta channels cont

Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents
DGX agent

arXiv:2605.15755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can produce fluent artwork emotion explanations, but they often suffer from attribute flooding: they enumerate

BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models

ResearchDGX agent

arXiv:2602.04163v2 Announce Type: replace Abstract: Large language model inference is often bounded by memory footprint and bandwidth in resource-constrained deployments, making quantization fundament

Constrained latent state modeling: A unifying perspective on representation learning under competing constraints

TutorialsDGX agent

arXiv:2605.15995v1 Announce Type: cross Abstract: Learning latent representations from complex data is central to modern machine learning, spanning temporal, multimodal, and partially observed systems

DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2605.16113v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved unprecedented success due to their exceptional generative capabilities. However, because they depend on kno

Differentiable Mixture-of-Agents Incentivizes Swarm Intelligence of Large Language Models

AgentsDGX agent

arXiv:2605.15706v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have catalyzed the development of multi-agent systems (MAS) for complex reasoning tasks. However, existi

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?

SafetyDGX agent

arXiv:2605.15855v1 Announce Type: new Abstract: Despite strong image-generation performance, diffusion models' reconstruction objectives limit alignment with human preferences. RL enables such alignme

Embedding-perturbed Exploration Preference Optimization for Flow Models

SafetyDGX agent

arXiv:2605.15803v1 Announce Type: new Abstract: Recent advancements have established Reinforcement Learning (RL) as a pivotal paradigm for aligning generative models with human intent. However, group-

From Model Design to Organizational Design: Complexity Redistribution and Trade-Offs in Generative AI

SafetyDGX agent

arXiv:2506.22440v2 Announce Type: replace-cross Abstract: This paper introduces the Generality-Accuracy-Simplicity (GAS) framework to analyze how large language models (LLMs) are reshaping organizatio

I believe on-prem and local AI - based on @huggingface open-source models - will be an important answer to the GPU shortages this year (beca…

Local AiDGX agent

I believe on-prem and local AI - based on @huggingface open-source models - will be an important answer to the GPU shortages this year (because they are cheaper, faster, safer than cloud APIs)! Great

Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds

ResearchDGX agent

arXiv:2605.15822v1 Announce Type: new Abstract: Score-based generative models are trained in high-dimensional ambient spaces, yet many data distributions are supported on low-dimensional nonlinear str

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in …

AgentsDGX agent

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in real-time. Try our playable research preview today, with Ago

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say…

AgentsDGX agent

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say a neutral observability layer still wins. http://venturebea

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation

ResearchDGX agent

arXiv:2605.15831v1 Announce Type: cross Abstract: Autoregressive music generation depends strongly on the audio tokenizer. Existing high-fidelity codecs often use residual multi-codebook quantization,

Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection

Local AiDGX agent

arXiv:2605.16107v1 Announce Type: new Abstract: Machine-generated texts (MGTs) pose risks such as disinformation and phishing, underscoring the need for reliable detection. Metric-based methods, which

okay maybe it's a good time? We have a small colbert model trained at pplx, it is a continue-training of pplx-embed-0.6b, so native multilin…

TutorialsDGX agent

okay maybe it's a good time? We have a small colbert model trained at pplx, it is a continue-training of pplx-embed-0.6b, so native multilingual, just made it open and added a section how to use MaxSi

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

AgentsDGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

To protect passengers or cargo, the powered rear seats & trunk in Model Y will automatically pop back up if detecting an obstruction while f…

IndustryDGX agent

Tesla Model Y's powered rear seats and trunk are equipped with automatic obstruction detection that causes them to automatically reverse and pop back up if an obstruction is detected during operation,

17 May 2026

Developers say Chinese AI labs lead US rivals in video generation, as ByteDance and Kuaishou train models on vast short-form video libraries from their own apps (Eleanor Olcott/Financial Times)

IndustryDGX agent

Eleanor Olcott / Financial Times: Developers say Chinese AI labs lead US rivals in video generation, as ByteDance and Kuaishou train models on vast short-form video libraries from their own apps — Chi

Publicis agrees to acquire LiveRamp, which allows companies to share and build new data sets and models that can power agentic frameworks, for $2.2B in cash (Alison Weissbrot/Adweek)

AgentsDGX agent

Alison Weissbrot / Adweek: Publicis agrees to acquire LiveRamp, which allows companies to share and build new data sets and models that can power agentic frameworks, for 2.2B in cash — Publicis Groupe

15 May 2026

A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 (Rebecca Bellan/TechCrunch)

HardwareDGX agent

Rebecca Bellan / TechCrunch: A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 — Every major A

AI teams shouldn’t have to choose between expensive object storage and painful git workflows. @huggingface Storage is built for model weight…

IndustryDGX agent

AI teams shouldn’t have to choose between expensive object storage and painful git workflows. @huggingface Storage is built for model weights, datasets, checkpoints and artifacts: - simple per-TB pric

Beyond What to Select: A Plug-and-play Oscillatory Data-Volume Scheduling for Efficient Model Training

ResearchDGX agent

arXiv:2605.14773v1 Announce Type: cross Abstract: Data selection accelerates training by identifying representative training data while preserving model performance. However, existing methods mainly f

BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: O…

AgentsDGX agent

BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: Opus 4.7 by @AnthropicAI 2nd: Opus 4.7 (Thinking) by @Anthrop

Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

AgentsDGX agent

arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U

DT-Transformer: A Foundation Model for Disease Trajectory Prediction on a Real-world Health System

ApplicationsDGX agent

arXiv:2605.14227v1 Announce Type: cross Abstract: Accurate disease trajectory prediction is critical for early intervention, resource allocation, and improving long-term outcomes. While electronic hea

Exploring Geographic Relative Space in Large Language Models through Activation Patching

SafetyDGX agent

arXiv:2605.14535v1 Announce Type: new Abstract: The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide ra

Graph of States: Solving Abductive Tasks with Large Language Models

AgentsDGX agent

arXiv:2603.21250v2 Announce Type: replace Abstract: Logical reasoning encompasses deduction, induction, and abduction. However, while Large Language Models (LLMs) have effectively mastered the former

Image Restoration via Diffusion Models with Dynamic Resolution

ResearchDGX agent

arXiv:2605.14267v1 Announce Type: cross Abstract: Diffusion models (DMs) have exhibited remarkable efficacy in various image restoration tasks. However, existing approaches typically operate within th

In a viral X post that parodies the old Mac vs. PC commercials, General Catalyst posted a 'VC vs GC' video, with the VC apparently modeled after Marc Andreessen (Julie Bort/TechCrunch)

IndustryDGX agent

Julie Bort / TechCrunch: In a viral X post that parodies the old Mac vs. PC commercials, General Catalyst posted a “VC vs GC” video, with the VC apparently modeled after Marc Andreessen — One of the m

Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control

SafetyDGX agent

arXiv:2605.14935v1 Announce Type: new Abstract: We present MSCoT, a multi-scale, coarse-to-fine model for test-time human motion synthesis and control. Unlike recent approaches that rely on multiple i

MultiMat: Multimodal Program Synthesis for Procedural Materials using Large Multimodal Models

ApplicationsDGX agent

arXiv:2509.22151v3 Announce Type: replace Abstract: Material node graphs are programs that generate the 2D channels of procedural materials, including geometry such as roughness and displacement maps,

SeaVis: Modeling and Control of a Remotely Operated Towed Vehicle for Seabed Visualization and Mapping

ResearchDGX agent

arXiv:2605.14683v1 Announce Type: new Abstract: High-resolution seafloor mapping necessitates stable and precise positioning for underwater robots. This paper introduces a novel mathematical model for

TRIO: Token Reduction via Inference-Objective Guidance for Efficient Vision-Language Models

Local AiDGX agent

arXiv:2602.04657v3 Announce Type: replace Abstract: Recently, reducing redundant visual tokens in vision-language models (VLMs) to accelerate VLM inference has emerged as a hot topic. However, most ex

We’re releasing a 30B-A3B reasoning model that reaches gold-medal level across both physics and math Olympiad evaluations: IPhO directly, an…

IndustryDGX agent

We’re releasing a 30B-A3B reasoning model that reaches gold-medal level across both physics and math Olympiad evaluations: IPhO directly, and IMO/USAMO with test-time self-verification and refinement.

14 May 2026

A Hierarchical Language Model with Predictable Scaling Laws and Provable Benefits of Reasoning

ResearchDGX agent

arXiv:2605.13687v1 Announce Type: cross Abstract: We introduce a family of synthetic languages with hierarchical structure -- generated by a broadcast process on trees -- for which the role of context

Are scaling laws finally working for time series foundation models? Today, @datadoghq is releasing Toto 2.0 weights in Apache 2.0 on @huggin…

IndustryDGX agent

Are scaling laws finally working for time series foundation models? Today, @datadoghq is releasing Toto 2.0 weights in Apache 2.0 on @huggingface. It's a family of open-weights TSFMs from 4M to 2.5B p

Assessing the Creativity of Large Language Models: Testing, Limits, and New Frontiers

ResearchDGX agent

arXiv:2605.13450v1 Announce Type: new Abstract: Measuring the creativity of large language models (LLMs) is essential for designing methods that can improve creativity and for enhancing our scientific

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements t…

Local AiDGX agent

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements to caching for faster inference overall. Turn on Developer Mod

CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models

ResearchDGX agent

arXiv:2605.13178v1 Announce Type: cross Abstract: In large vision-language models, visual tokens typically constitute the majority of input tokens, leading to substantial computational overhead. To ad

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

SafetyDGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me …

IndustryDGX agent

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me flexibility. Give me delight. Give me color. Please create t

DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models

SafetyDGX agent

arXiv:2605.12702v1 Announce Type: new Abstract: General-purpose safety benchmarks for large language models do not adequately evaluate disability-related harms. We introduce DisaBench: a taxonomy of t

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models

ResearchDGX agent

arXiv:2605.13156v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in bridging visual perception and natural language understanding, enabling a wid

Generative Modeling by Minimizing the Wasserstein-2 Loss

ResearchDGX agent

arXiv:2406.13619v4 Announce Type: replace-cross Abstract: This paper develops a generative model by minimizing the second-order Wasserstein loss (the W_2 loss) through a distribution-dependent ordinar

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

Local AiDGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

Learning to See What You Need: Gaze Attention for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.13080v1 Announce Type: new Abstract: When humans describe a visual scene, they do not process the entire image uniformly; instead, they selectively fixate on regions relevant to their inten

MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling

ApplicationsDGX agent

arXiv:2605.13711v1 Announce Type: new Abstract: Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels.

Model. Harness. Context. The 3 main components of agents. As you build more agents, context increasingly lives AGENTS.md, skills, policies, …

AgentsDGX agent

Model. Harness. Context. The 3 main components of agents. As you build more agents, context increasingly lives AGENTS.md, skills, policies, examples, + generated research files. Context needs its own

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governan…

TutorialsDGX agent

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governance. Fireworks AI on @Azure AI Foundry provides the inference

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy

ResearchDGX agent

arXiv:2304.11193v2 Announce Type: replace-cross Abstract: Predicting the outcomes of robotic actions, often referred to as learning a world model, in complex environments remains a fundamental challen

Multitask Multimodal Fusion with Tabular Foundation Models for Peak and Durability Prediction of Pertussis Booster Response

ResearchDGX agent

arXiv:2605.12852v1 Announce Type: new Abstract: Pertussis booster vaccination produces immune responses that vary widely across individuals in both peak magnitude and long-term durability. These two p

On the Limits of Latent Reuse in Diffusion Models

ResearchDGX agent

arXiv:2605.13448v1 Announce Type: cross Abstract: Diffusion models are often trained in low-dimensional latent spaces, which are then reused for related but shifted datasets. In this work, we study wh

Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP

SafetyDGX agent

arXiv:2605.13436v1 Announce Type: cross Abstract: Subword regularization methods such as BPE dropout are typically applied only during fine-tuning, while pretraining is usually done with deterministic

SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models

ResearchDGX agent

arXiv:2605.13667v1 Announce Type: new Abstract: Scene graph generation provides a compact structured representation for visual perception, but accurate and fast graph prediction from images and videos

TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models

ResearchDGX agent

arXiv:2605.12518v1 Announce Type: cross Abstract: The proliferation of online news poses a challenge to extracting structured timelines from unstructured content. While recent studies have shown that

Toto 2.0 is here: Datadog AI's 5 open-weights forecasting models (4m-2.5B params) finally make scaling work for time series forecasting! #1 …

IndustryDGX agent

Toto 2.0 is here: Datadog AI's 5 open-weights forecasting models (4m-2.5B params) finally make scaling work for time series forecasting! #1 on BOOM, GIFT-Eval, and TIME. Weights/code Apache 2.0. 🔗 Rea

Try Rime Mist v3 in voice finder directly: https://findtherightvoice.com/rime-labs--rime-mist-v3 Model pages: http://www.together.ai/models/…

ToolsDGX agent

Try Rime Mist v3 in voice finder directly: https://findtherightvoice.com/rime-labs--rime-mist-v3 Model pages: http://www.together.ai/models/rime-mist-v3 http://www.together.ai/models/rime-mist-v3-omni

What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models

SafetyDGX agent

arXiv:2605.13105v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning has shown promise for Vision-Language-Action (VLA) models in robotic manipulation, but deployment-time visual sh

← Previous
1…217218219220221…1017
Next →