AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
Model Releases

ToolFailBench: Diagnosing Tool-Use Failures in LLM Agents

DGX agent

arXiv:2607.04686v1 Announce Type: cross Abstract: Tool calling is central to modern language model agents, but aggregate benchmark scores often hide where tool use fails. A model that never calls a ne

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Topology-Driven Transferability Estimation for 3D Medical Vision Foundation Models

DGX agent

arXiv:2607.04199v1 Announce Type: new Abstract: The growing number of medical vision foundation models highlights the need for effective model selection. However, mainstream selection methods rely on

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Toward Efficient Agents: Memory, Tool learning, and Planning

DGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Toward Trustworthy Large Language Model Agents in Healthcare

DGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker

DGX agent

arXiv:2605.25706v2 Announce Type: replace Abstract: Referring expression comprehension (REC) aims to localize a target object within an image based on a given expression. Although recent advances in v

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Realistic Remote Sensing Dataset Distillation with Discriminative Prototype-guided Diffusion

DGX agent

arXiv:2601.15829v2 Announce Type: replace Abstract: Recent years have witnessed the remarkable success of deep learning in remote sensing image interpretation, driven by the availability of large-scal

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation

DGX agent

arXiv:2603.17673v2 Announce Type: replace-cross Abstract: LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to repro

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Standardized Light Field Quality Assessment: Hybrid Subjective Benchmarking and Objective Metric Evaluation

DGX agent

arXiv:2607.03494v1 Announce Type: new Abstract: Benchmarking immersive media coding solutions, especially in the standardization context, requires reliable and reproducible subjective quality assessme

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards transferable lightweight neuromorphic computing through a model-free temporal-switch framework

DGX agent

arXiv:2607.02608v1 Announce Type: cross Abstract: Lightweight neuromorphic computing offers a promising route to efficient AI, with particular benefits for resource-constrained edge deployments. Howev

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TRACE: Capability-Targeted Agentic Training

DGX agent

arXiv:2604.05336v2 Announce Type: replace Abstract: Models often fail to complete agentic tasks because they lack core capabilities required by the target environment. However, mainstream approaches f

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Training-Free Model Selection and Domain-Aware Score Calibration for First-Shot Anomalous Sound Detection

DGX agent

arXiv:2607.04526v1 Announce Type: cross Abstract: First-shot anomalous sound detection in DCASE Challenge Task 2 must flag anomalies of unseen machine types with a single threshold, without knowing wh

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Training Hybrid Block Diffusion Language Models with Partial Bidirectionality

DGX agent

arXiv:2607.02805v1 Announce Type: cross Abstract: High-throughput long-context generation is one of the central challenges for large language models. Generation is typically memory-bandwidth-bound rat

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Transcribe Arabic is built to bring frontier capabilities to the millions of Arabic speakers in business and developer communities. Built to…

DGX agent

Transcribe Arabic is built to bring frontier capabilities to the millions of Arabic speakers in business and developer communities. Built to handle code-switching, multiple dialects, and Arabic-accent

model-releasescohere--x
7 Jul 2026
Model Releases

Transformers with Physics-Informed Encodings and Simulation-Based Inference for Robust Detection of Eccentric Binary Black Holes in Pulsar Timing Array Data

DGX agent

arXiv:2607.03904v1 Announce Type: new Abstract: Pulsar timing arrays (PTAs) provide a unique window into nanohertz gravitational waves (GWs), but extracting astrophysical parameters from noisy, long-b

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TREK: Distill to Explore, Reinforce to Refine

DGX agent

arXiv:2607.05339v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) is effective when the current policy already samples useful reasoning trajectories, but it stalls on hard pr

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking

DGX agent

arXiv:2410.15135v5 Announce Type: replace Abstract: With the surge of online misinformation, Large Language Models (LLMs) and Reasoning Large Language Models (RLMs) serving as Automatic Fact-Checking

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Triple-Phase Multimodal Knowledge Aggregation Framework for Microbial Keratitis Subtype Diagnosis on Slit-Lamp Photography

DGX agent

arXiv:2607.03740v1 Announce Type: cross Abstract: Microbial keratitis requires rapid pathogen identification to guide treatment, but culture- and PCR-based diagnostics are slow and resource-intensive.

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TSP with Predictions: Heatmap to Tour with Provable Guarantees

DGX agent

arXiv:2607.03791v1 Announce Type: cross Abstract: The Traveling Salesperson Problem (TSP) has long served as a benchmark for evaluating the strength of optimization techniques in the classical theory

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

U-Joint CAAMS: Experimental Evaluation of a Universal-Joint Continuum Manipulator for Aerial Manipulation

DGX agent

arXiv:2607.03321v1 Announce Type: new Abstract: Continuum manipulators mounted on multi-rotor UAVs enable compliant aerial manipulation, but payloads and propeller downwash amplify out-of-plane bendin

model-releasesarxiv-cs-ro
7 Jul 2026
Model Releases

Unbiased Alignment for Large Language Models with Noisy Preferences

DGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Uncertainty-aware damage identification in short-span bridges via physics-informed variational autoencoder

DGX agent

arXiv:2607.05025v1 Announce Type: new Abstract: Vibration-based damage identification in civil infrastructure is a challenging, ill-posed inverse problem due to measurement noise, sparse sensor arrays

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Unified Audio Intelligence Without Regressing on Text Intelligence

DGX agent

arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy

DGX agent

arXiv:2603.24690v2 Announce Type: replace Abstract: In-context learning (ICL) enables fast task adaptation from demonstrations without per-task parameter updates but remains highly sensitive to exampl

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

UniVideo: Unified Understanding, Generation, and Editing for Videos

DGX agent

arXiv:2510.08377v4 Announce Type: replace Abstract: Unified multimodal models have shown promising results in multimodal content generation and editing but remain largely limited to the image domain.

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching

DGX agent

arXiv:2603.27044v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is widely recognized as sample-inefficient, a limitation attributable in part to the high dimensionality and

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

DGX agent

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Re

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

DGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

DGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

DGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

DGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs

DGX agent

arXiv:2511.20272v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) have become adept at recognizing objects, they often lack the intuitive, human-like understanding of

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

DGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

DGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

We're extending access to Claude Fable 5 on all paid plans through July 12.

DGX agent

Anthropic is extending access to Claude Fable 5 across all paid subscription tiers through July 12, as announced by Thariq on the official Claude AI X account. This indicates a temporary or promotiona

model-releasesthariq--x
7 Jul 2026
Model Releases

What’s New in Microsoft Foundry | June 2026

DGX agent

Claude is now generally available in Microsoft Foundry. Here's everything else that shipped between Build 2026 and the end of June — autopilot agents, expanded Toolboxes and Routines, Agent Optimizer'

model-releasesmicrosoft-foundry
7 Jul 2026
Model Releases

When Aggregate Alignment Misleads: Auditing Policy Repair Without Per-State Expert Actions

DGX agent

arXiv:2607.03386v1 Announce Type: new Abstract: Agentic AI systems are increasingly used to edit, refine, and repair decision policies, but evaluating these edits is difficult when per-state expert ac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

DGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters

DGX agent

arXiv:2607.04919v1 Announce Type: new Abstract: Deploying a time series foundation model requires GPU infrastructure, engineering overhead, and carries no guarantee of improvement over XGBoost. We pro

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

When Does High-CFG Diffusion Inversion Fail? A Controlled Study of Prompt--Latent Interactions

DGX agent

arXiv:2607.04731v1 Announce Type: new Abstract: Text-guided diffusion inversion is central to image editing, where an image is mapped to an initial latent and then edited by replaying the denoising pr

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

When is a System Discoverable from Data? Discovery Requires Chaos

DGX agent

arXiv:2511.08860v2 Announce Type: replace-cross Abstract: The deep learning revolution has spurred a rise in advances of using AI in sciences. Within physical sciences the main focus has been on disco

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Rubrics Fail: Error Enumeration as Reward in Reference-Free RL Post-Training for Virtual Try-On

DGX agent

arXiv:2603.05659v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and Rubrics as Rewards (RaR) have driven strong gains in domains with clear correctness

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Simpler Is Better: Evaluating Translation Pipelines for Medieval Latin Manuscripts

DGX agent

arXiv:2607.03836v1 Announce Type: cross Abstract: Despite remarkable progress in machine translation, Vision Language Models (VLMs) struggle on historical manuscripts, a domain that stresses core Natu

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Users Are Happy but Agents Are Wrong: Multi-Dimensional Evaluation of Tool-Augmented Dialogue

DGX agent

arXiv:2510.19186v3 Announce Type: replace Abstract: Evaluating conversational AI systems that use external tools is challenging, as errors can arise from complex interactions among user, agent, and to

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Which Algorithm Specification Formats Help Language Models Implement Machine Learning Algorithms?

DGX agent

arXiv:2607.03158v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to implement algorithms from research manuscripts, but papers often leave implementation choices im

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Why Open Data Matters | Nemotron Labs https://x.com/i/broadcasts/1DxLddoaAOoxm

DGX agent

This broadcast likely discusses the importance and benefits of open data in AI development and machine learning, covering topics such as accessibility, transparency, democratization of AI technology,

model-releasesclem-delangue--x
7 Jul 2026
Model Releases

Wordle 1,843 6/6 ⬛🟨⬛⬛⬛ ⬛⬛⬛⬛🟨 🟩🟩⬛⬛⬛ 🟩🟩⬛⬛🟩 🟩🟩⬛⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This is a Wordle game result shared by Anthropic on X (Twitter), showing a player who solved puzzle #1,843 on their sixth and final attempt. The colored emoji squares document the progression of guess

model-releasesanthropic--x
7 Jul 2026
Model Releases

WorldBagel: Uncovering the Power of Unified Multimodal Models for Vision-Language-Action-World Modeling

DGX agent

arXiv:2607.03461v1 Announce Type: new Abstract: World models aim to capture environment dynamics in ways that support perception, reasoning, and action, and have recently become a central direction in

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Wrong Before Right: Late Rescue and Interface Failure in Aligned Language Models

DGX agent

arXiv:2607.04640v1 Announce Type: new Abstract: We study how correctness is assembled inside aligned language models, not only whether the final answer is right. Using layer-wise difference-in-differe

model-releasesarxiv-cs-cl
7 Jul 2026
← Previous
1…129130131132133…471
Next →