AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

UNSPECIFIC: General Constraint Synthesis for Breaking Copy-and-Paste Shortcut in LLM Instruction Following

DGX agent

arXiv:2608.09154v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly expected to follow long lists of constraints in complex instructions, and synthesizing instructions from a

model-releasesarxiv-cs-cl
11 Aug 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Unsupervised Point Cloud Registration with Self-Distillation

DGX agent

arXiv:2409.07558v2 Announce Type: replace Abstract: Rigid point cloud registration is a fundamental problem and highly relevant in robotics and autonomous driving. Nowadays deep learning methods can b

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

v0.32.8

DGX agent

Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such

model-releasesollama-releases
11 Aug 2026
Model Releases

v0.32.9

DGX agent

NVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for that execution layer of always-on agents. It is designed f

model-releasesollama-releases
11 Aug 2026
Model Releases

VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging

DGX agent

arXiv:2511.18121v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel on benchmarks, their processing paradigm differs from the human ability to integrate visu

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

DGX agent

arXiv:2608.08477v1 Announce Type: new Abstract: We present VectraYX-Vision-1B, a sub-2B vision-language model (VLM) for Spanish/LATAM cybersecurity imagery, coupling a frozen SigLIP-so400m encoder to

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting

DGX agent

arXiv:2608.09286v1 Announce Type: cross Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Verication-driven closed-loop multi-agent large language modelframework for code-compliant structural design

DGX agent

arXiv:2608.07978v1 Announce Type: cross Abstract: Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

VideoVIBE: A Video-Grounded Diagnostic Benchmark for One-Shot Interactive Website Generation

DGX agent

arXiv:2608.09573v1 Announce Type: new Abstract: Natural-language-driven 'vibe coding' enables the one-shot generation of visually rich and interactive web applications, yet reliable assessment of thei

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VIGIL: Tackling Hallucination Detection in Image Recontextualization

DGX agent

arXiv:2602.14633v2 Announce Type: replace Abstract: We introduce VIGIL (Visual Inconsistency & Generative In-context Lucidity), a benchmark dataset and framework that provides a fine-grained categoriz

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VLZip: Unified Visual and Textual Compression for Interleaved Long-Context Modeling

DGX agent

arXiv:2608.08630v1 Announce Type: new Abstract: Vision Language Models (VLMs) face significant challenges with ultra-long, interleaved image-text sequences due to the quadratic complexity of self-atte

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VTO: Visual Tool Orchestration for Video Anomaly Detection

DGX agent

arXiv:2608.08219v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is a critical yet challenging task due to the complex and diverse nature of real-world scenarios. Traditional deep learn

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8× RTX 5090

DGX agent

We converted the model from the original safetensors and found two issues. The first one made our quantization fail several times, the second one does not fail at all, it just quietly ruins the base 1

model-releasesr-localllama
11 Aug 2026
Model Releases

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap!

DGX agent

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap! We ran DeepSeek V4 Flash through 4 more agent harnesses (Hermes Agent, Pi Agent, Pri

model-releasesharrison-chase--x
11 Aug 2026
Model Releases

Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems

DGX agent

arXiv:2401.04013v2 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of in

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks

DGX agent

arXiv:2506.01952v2 Announce Type: replace-cross Abstract: Powered by large language models (LLMs), web browsing agents operate graphical user interfaces in a human-like manner, offering a transparent

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

What to Edit Next: Visually Aligned Image-Editing Follow-Up Suggestions in Conversational Systems

DGX agent

arXiv:2608.07565v1 Announce Type: cross Abstract: Conversational assistants increasingly recommend follow-up edits to help users continue a task. Existing systems primarily target text-only interactio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

What Would Fix This RAG Failure? Auditing Counterfactual Response with Paired Evidence Interventions

DGX agent

arXiv:2608.08944v1 Announce Type: cross Abstract: A failed retrieval-augmented generation (RAG) answer can be consistent with several unseen responses to evidence repair. We introduce Pair-ID, an offl

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

DGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition

DGX agent

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predi

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Does An Extra View Help? Adapting Single-View 3D Reconstruction with Extra Imagery

DGX agent

arXiv:2608.08132v1 Announce Type: new Abstract: Reconstruction of 3D objects from a single image is a challenging research problem in computer vision. The key challenge is the lack of critical informa

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

DGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Is Benchmark Contamination Detectable? Information Limits and Power-Calibrated Audits

DGX agent

arXiv:2608.07914v1 Announce Type: new Abstract: Behavioral contamination detectors can return 'no evidence' either because a benchmark is clean or because the audit has little power. We formalize this

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

DGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When should we trust the annotation? Selective prediction for molecular structure retrieval from mass spectra

DGX agent

arXiv:2603.10950v2 Announce Type: replace Abstract: Machine learning methods for identifying molecular structures from tandem mass spectra (MS/MS) have advanced rapidly, yet current approaches still e

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs

DGX agent

arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, c

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

DGX agent

arXiv:2608.07813v1 Announce Type: new Abstract: An LLM judge deployed inside a reasoning pipeline does not merely measure quality, it decides which answer ships. We show that the cost of that decision

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer m…

DGX agent

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer model is a 30B multimodal reasoning model with a Gemma-like a

model-releasessebastian-raschka--x
11 Aug 2026
Model Releases

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

DGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Wiener Representation Filtering for VLM Hallucination Suppression

DGX agent

arXiv:2608.08167v1 Announce Type: new Abstract: Vision-language models (VLMs) excel at open-ended captioning and visual QA but often describe objects, attributes, or relations absent from the image, a

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Wix launches Symphony, a new standalone multi-agent system built for business operations

DGX agent

Cloud-based website builder Wix Ltd. today announced the launch of Symphony, a new standalone agentic artificial intelligence platform that proactively learns business values, interests, needs, practi

model-releasessiliconangle
11 Aug 2026
Model Releases

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

DGX agent

arXiv:2608.09298v1 Announce Type: cross Abstract: Action-conditioned world models (ACWMs) promise to provide embodied AI with scalable predictive simulators for planning, policy evaluation, and data g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management

DGX agent

arXiv:2608.07529v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation

DGX agent

arXiv:2505.11146v3 Announce Type: replace-cross Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition challenge due to the significant

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

XFeat Revisited: Reproducibility and Evaluation of a Lightweight Image Matcher

DGX agent

arXiv:2608.09519v1 Announce Type: new Abstract: We present a reproducibility study of XFeat, a lightweight local feature extractor and matcher designed to identify corresponding points across images e

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

XPolicyLab: A Unified Standard and Open Ecosystem for Robot Policy Evaluation and Deployment

DGX agent

arXiv:2608.09892v1 Announce Type: new Abstract: Robot policy evaluation and deployment remain fragmented by model-specific software dependencies, data representations, and runtime interfaces, so that

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

You Don't Need To Stay in The Loop: An Agentic Robotics Loop for Robot-Policy Improvement

DGX agent

arXiv:2608.07555v1 Announce Type: new Abstract: Coding agents such as Claude Code and Codex close the software loop: a main agent manages the loop, subagents analyze and execute, tools do the work. We

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Your Prompt Is Not the Only Prompt: How Much Do LLMs Weight Structured-Output Schema Descriptions?

DGX agent

arXiv:2608.08254v1 Announce Type: new Abstract: Structured output, where an LLM populates a predefined JSON schema, has become a default mechanism for data labeling and information extraction, but it

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Your VLM Already Knows When: Training-Free Temporal Grounding by Asking Yes or No

DGX agent

arXiv:2608.08315v1 Announce Type: new Abstract: Multimodal LLMs that recognise events reliably still fail to say when they happen. Prompted for timestamps, strong VLMs reach as little as 3.8% R@0.5 on

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Zero-shot 2D Grounding with Novel Affordance Types

DGX agent

arXiv:2608.08929v1 Announce Type: new Abstract: 2D affordance grounding aims to locate the region of an object that a human can interact with. Existing research focuses on recognizing affordance types

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Zero-Shot Traffic Accident Detection via a Coarse-to-Fine VLM-Tracking Pipeline

DGX agent

arXiv:2608.08867v1 Announce Type: new Abstract: Traffic surveillance cameras capture accidents continuously, yet converting raw CCTV footage into structured event records that pinpoint when, where, an

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

DGX agent

arXiv:2608.07925v1 Announce Type: new Abstract: EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

1M context with 17 GB model in 24 GB VRAM: 'for the first time I was able to load a context of almost 1M tokens and extract 7 needles from various parts of the text'

DGX agent

https://preview.redd.it/xxjh11f38jih1.png?width=1852&format=png&auto=webp&s=76850ed51e29a8bc86c2ca718d4320075eed4363 Just wanted to share a user report that I found to be very interesting. Some person

model-releasesr-localllama
10 Aug 2026
Model Releases

8 days ago, while jogging, I asked Claude to solve the Riemann Hypothesis It didn’t. 1.5 days later, it proved >= 67% of the zeros are on th…

DGX agent

8 days ago, while jogging, I asked Claude to solve the Riemann Hypothesis It didn’t. 1.5 days later, it proved >= 67% of the zeros are on the line (prev: 41.6%) Still not sure what that means, but som

model-releasesboris-cherny--x
10 Aug 2026
Model Releases

A downside with VLM-based parsing is that they’re generally slower than text-based heuristic approaches. As a result they add latency to any…

DGX agent

A downside with VLM-based parsing is that they’re generally slower than text-based heuristic approaches. As a result they add latency to any ad-hoc file processing *in-the agent loop* (e.g. if you upl

model-releasesjerry-liu--x
10 Aug 2026
Model Releases

A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers

DGX agent

arXiv:2608.06694v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics extends polymer simulation beyond the scales accessible to all-atom (AA) methods, but bottom-up CG modeling is la

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

DGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

model-releasesarxiv-cs-ai
10 Aug 2026
← Previous
1…1718192021…464
Next →