AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation

DGX agent

arXiv:2606.02320v1 Announce Type: new Abstract: Deep Research Agents have shown strong capability in multi-step information retrieval, reasoning, and long-form report generation, but existing benchmar

model-releasesarxiv-cs-cl
2 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool 'to responsibly encourage agentic AI adoption' (Natalie Lung/Bloomberg)

DGX agent

Natalie Lung / Bloomberg: Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool “to responsibly encourage agentic AI adoption” — Uber Technologies Inc. has set

model-releasestechmeme
2 Jun 2026
Model Releases

Uncovering Competency Gaps in Large Language Models and Their Benchmarks

DGX agent

arXiv:2512.20638v2 Announce Type: replace-cross Abstract: The evaluation of large language models relies heavily on standardized benchmarks. These benchmarks provide useful aggregated metrics, but can

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Understanding Identity Continuity in Thermal Video through Scene-Level Consistency

DGX agent

arXiv:2606.01694v1 Announce Type: cross Abstract: Thermal pedestrian MOT remains challenging because weak appearance cues and frequent detection interruptions cause severe trajectory fragmentation. We

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Understanding LLM Behavior in Multi-Target Cross-Lingual Summarization

DGX agent

arXiv:2606.01252v1 Announce Type: cross Abstract: Multi-target cross-lingual text summarization (MTXLS), which summarizes a source document into multiple target languages, is increasingly important as

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

UniD^3: A Knowledge Graph-Enhanced RAG Framework for Drug-Disease Discovery and Reasoning

DGX agent

arXiv:2606.01394v1 Announce Type: new Abstract: Systematic characterization of drug-disease relationships is essential for drug discovery and repurposing, yet is hindered by the heterogeneity and rapi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Universal Quantum Transformer

DGX agent

arXiv:2606.00045v1 Announce Type: new Abstract: Classical continuous-space neural networks fundamentally struggle to lock into exact mathematical symmetries, such as modular arithmetic and non-commuta

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Unlocking the Black Box of Latent Reasoning: An Interpretability-Guided Approach to Intervention

DGX agent

arXiv:2606.01243v1 Announce Type: new Abstract: Latent reasoning enables Large Language Models (LLMs) to perform multi-step inference within continuous hidden states, offering efficiency gains over ex

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets

DGX agent

arXiv:2603.14010v2 Announce Type: replace Abstract: Articulated objects are fundamental for robotics, simulation of physics, and interactive virtual environments. However, recovering them from visual

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound

DGX agent

arXiv:2509.25773v3 Announce Type: replace-cross Abstract: AI models capable of comprehending humor hold real-world promise -- for example, enhancing engagement in human-machine interactions. To gauge

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

v0.30.1: llm: ignore llama-server SSE ping comments (#16443)

DGX agent

Ollama v0.30.1 addresses an issue where the LLM component now ignores Server-Sent Events (SSE) ping comments from llama-server, resolving problem #16443. This fix improves the stability and reliabilit

model-releasesollama-releases
2 Jun 2026
Model Releases

Value Flows

DGX agent

arXiv:2510.07650v4 Announce Type: replace-cross Abstract: While most reinforcement learning methods today flatten the distribution of future returns to a single scalar value, distributional RL methods

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VESTA: Visual Exploration with Statistical Tool Agents

DGX agent

arXiv:2606.00384v1 Announce Type: new Abstract: Fitting quantitative models to data is a central step in scientific workflows, yet it remains one of the least automated. Recent agent-based systems lev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Vision-language Models for Driver Monitoring Systems: A Driver Activity Description Dataset

DGX agent

arXiv:2606.02273v1 Announce Type: new Abstract: Understanding subtle driver actions is essential for building reliable driver monitoring systems. Existing visionlanguage models (VLMs) are trained on g

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning

DGX agent

arXiv:2606.00105v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress on vision-language tasks, but they may also memorize and expose sensitive o

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting

DGX agent

arXiv:2606.02138v1 Announce Type: cross Abstract: Out of distribution (OOD) events in multivariate time series forecasting are rare but often dominate real world risk, making average case forecasting

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio

DGX agent

arXiv:2512.10120v2 Announce Type: replace-cross Abstract: General-purpose audio representations aim to map acoustically variable instances of the same event to nearby points, resolving content identit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models

DGX agent

arXiv:2510.22276v3 Announce Type: replace-cross Abstract: Contrastive vision-language models have achieved remarkable progress through large-scale pretraining. Recent work has shown that removing Engl

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1.…

DGX agent

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1. MSA (MiniMax Sparse Attention) is the star ⭐️. Unlike CSA/HC

model-releasestogether-ai--x
2 Jun 2026
Model Releases

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours a…

DGX agent

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours again. Small enough that it's inexpensive to run, big enough

model-releasescohere--x
2 Jun 2026
Model Releases

What Do LLMs Know About Alzheimer's Disease? Multi-loss Fine-Tuning and Probing for AD Detection

DGX agent

arXiv:2602.11177v2 Announce Type: replace-cross Abstract: Reliable early detection of Alzheimer's disease (AD) is challenging, particularly due to the limited availability of labeled data. While large

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

What to Format and How: A Benchmark and Workflow Approach for Document Formatting

DGX agent

arXiv:2606.01936v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have opened up new possibilities for automated document formatting. However, real-world formatting often

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation

DGX agent

arXiv:2602.16763v2 Announce Type: replace Abstract: Artificial intelligence benchmarks are an important mechanism for measuring model progress and guiding deployment decisions. However, benchmarks qui

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

DGX agent

arXiv:2602.08236v2 Announce Type: replace-cross Abstract: Despite rapid progress in MLLMs, visual spatial reasoning remains unreliable when correct answers depend on how a scene would appear under uns

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Jokes Cross the Line: Analyzing Regular Humor and Dark Humor in YouTube Shorts

DGX agent

arXiv:2606.00046v1 Announce Type: cross Abstract: Video platforms such as YouTube have reshaped how users engage with entertainment and information, emphasizing brief, highly engaging content such as

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

DGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Single Answer Is Not Enough: Rethinking Single-Step Retrosynthesis Benchmarks for LLMs

DGX agent

arXiv:2602.03554v2 Announce Type: replace-cross Abstract: Recent progress has expanded the use of large language models (LLMs) in drug discovery, including synthesis planning. However, objective evalu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

DGX agent

arXiv:2606.02060v1 Announce Type: new Abstract: Deep-research agents solve tasks through long trajectories of search, tool use, evidence inspection, and answer synthesis. Evaluation based on final ans

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?

DGX agent

arXiv:2606.01247v1 Announce Type: new Abstract: Humans can reproduce the viewpoint specified by a target image through active head and body motion, yet spatial intelligence in foundation models has la

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Which Leakage Types Matter? A Quantitative Landscape Across 2,047 Benchmark Datasets

DGX agent

arXiv:2604.04199v2 Announce Type: replace Abstract: Twenty-eight within-subject counterfactual experiments across 2,047 iid tabular datasets, plus a boundary experiment on 129 temporal datasets, measu

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Why Not Hyperparameter-Friendly Optimisation? A Monotonic Adaptive Norm Rescaling Approach For Long-Tailed Recognition

DGX agent

arXiv:2606.02526v1 Announce Type: cross Abstract: Long-tailed recognition poses a significant challenge for deep learning. The two-stage decoupling paradigm, which separates representation learning fr

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WildCat: Near-Linear Attention in Theory and Practice

DGX agent

arXiv:2602.10056v2 Announce Type: replace Abstract: We introduce WildCat, a high-accuracy, low-cost approach to compressing the attention mechanism in neural networks. While attention is a staple of m

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Wordle 1,808 5/6 ⬛⬛⬛⬛⬛ 🟨🟨⬛⬛⬛ ⬛🟨🟩🟨⬛ 🟩🟩🟩🟩⬛ 🟩🟩🟩🟩🟩

DGX agent

This post shows the solution path for Wordle puzzle #1,808, solved in 5 guesses, with a visual representation of each guess's results using colored tiles indicating correct letters (green), misplaced

model-releasesanthropic--x
2 Jun 2026
Model Releases

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out b…

DGX agent

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out best practices, examples and more. I’m particularly excited a

model-releasesthariq--x
2 Jun 2026
Model Releases

WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching

DGX agent

arXiv:2603.06331v2 Announce Type: replace Abstract: Diffusion-based world models have shown strong potential for unified world simulation, but the iterative denoising remains too costly for interactiv

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis

DGX agent

arXiv:2606.01869v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly asked not only to write static interfaces, but to construct executable interactive worlds from natural lan

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World

DGX agent

arXiv:2512.10958v2 Announce Type: replace Abstract: Generative world models are reshaping embodied AI, enabling agents to synthesize realistic 4D driving environments that look convincing but often fa

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Worlds Within Words: Translating Culture in Ancient Chinese Texts with Multi-Agent Coordination

DGX agent

arXiv:2606.01276v1 Announce Type: new Abstract: Large language model (LLM)-based machine translation has advanced cross-cultural communication, yet it still struggles with culture-loaded words (CLWs)

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

DGX agent

arXiv:2512.00956v3 Announce Type: replace Abstract: Quantizing LLM weights and activations is a standard approach for efficient deployment, but a few extreme outliers can stretch the dynamic range and

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding

DGX agent

arXiv:2606.02482v1 Announce Type: new Abstract: While video streaming understanding has made significant strides, real-world applications, such as live sports broadcasting, autonomous driving, and mul

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

XAI-SOH-FL: Enhancing SOH-FL with Adaptive Aggregation and Explainable AI for Intrusion Detection in Heterogeneous IoT

DGX agent

arXiv:2606.00134v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDS) in Internet of Things (IoT) environments face significant challenges due to data heterogeneity, lack of labeled data

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

You Can Learn Tokenization End-to-End with Reinforcement Learning

DGX agent

arXiv:2602.13940v2 Announce Type: replace-cross Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large Language Models (LLMs), despite a general trend t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Zamba2-VL Technical Report

DGX agent

arXiv:2606.00390v1 Announce Type: cross Abstract: We present Zamba2-VL, a suite of vision-language models built on Zamba2, a hybrid language-model architecture combining Mamba2 state-space layers with

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Zero-Shot Off-Policy Learning

DGX agent

arXiv:2602.01962v2 Announce Type: replace-cross Abstract: Off-policy learning methods seek to derive an optimal policy directly from a fixed dataset of prior interactions. This objective presents sign

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

3DAE: Binaural Quality Assessment for Audio Novel View Synthesis with Spatial Maps and Benchmark

DGX agent

arXiv:2605.30469v1 Announce Type: cross Abstract: 3D audio and novel-view acoustic synthesis models are usually evaluated with global metrics.However, global metrics often hide where and why binaural

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

A Kinetic Energy Perspective of Flow Matching

DGX agent

arXiv:2602.07928v2 Announce Type: replace-cross Abstract: Flow-based generative models can be viewed through a physics lens: sampling transports a particle from noise to data by integrating a learned

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

A Lightweight Ensemble-Based Face Image Quality Assessment Method with Correlation-Aware Loss

DGX agent

arXiv:2509.10114v2 Announce Type: replace Abstract: Face image quality assessment (FIQA) plays a critical role in face recognition and verification systems, especially in uncontrolled, real-world envi

model-releasesarxiv-cs-cv
1 Jun 2026
← Previous
1…236237238239240…475
Next →