AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models

DGX agent

arXiv:2608.05948v1 Announce Type: new Abstract: Physics engines facilitate large-scale training and evaluation for embodied intelligence, while generative video world models are emerging as implicit s

model-releasesarxiv-cs-ai
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning

DGX agent

arXiv:2604.02721v3 Announce Type: replace Abstract: Competitive programming remains one of the last few human strongholds in coding against AI. The best AI system to date still underperforms the best

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

GROM: Gradient-Free Rapid One-Shot Machine Unlearning

DGX agent

arXiv:2608.05783v1 Announce Type: cross Abstract: Machine unlearning has become a critical capability for safely removing specific, sensitive knowledge from large language models (LLMs). Current state

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Grounded Well-Condition Anomaly Detection on the Volve Field: Constructed Labels, a Baseline, and a Dual-Head Model

DGX agent

arXiv:2608.05685v1 Announce Type: new Abstract: Most public benchmarks for machine-condition monitoring come from test rigs, where faults are induced on purpose and every event is known. Real producti

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

DGX agent

arXiv:2608.06130v1 Announce Type: cross Abstract: AI agents performing cryptographic operations (signing Git commits, authenticating API calls, issuing certificates) currently store private keys in so

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

DGX agent

arXiv:2608.06301v1 Announce Type: new Abstract: As LLMs are increasingly deployed within agentic systems, their capabilities depend not only on the model weights but also on the harness: the prompts,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

DGX agent

arXiv:2608.06012v1 Announce Type: new Abstract: Search-agent rewards mix answer quality, citation grounding, tool cost, and anti-hacking terms; a high score therefore need not imply that cited evidenc

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots

DGX agent

arXiv:2608.05715v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as planners in robotic systems, where they translate natural-language commands into executable

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

How Much Reconstruction Does Quantum Machine Learning Need? Late Fusion of Independently Trained Quantum Subcircuits

DGX agent

arXiv:2608.05595v1 Announce Type: cross Abstract: Circuit cutting lets a large quantum neural network (QNN) run as independent subcircuits on small devices, but rebuilding its outputs by reconstructio

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Human-Like Anaphor Resolution in Large Language Models

DGX agent

arXiv:2608.05630v1 Announce Type: new Abstract: Anaphors are expressions that refer to other expressions, called antecedents. The process of connecting the two is called resolution. Cognitive science

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Hyper-ES: Effective Evolution Strategies for LLM Reasoning via Descent Direction Merging

DGX agent

arXiv:2608.05541v1 Announce Type: new Abstract: Evolution Strategy (ES) is a promising alternative to gradient-based fine-tuning for resource-constrained Large Language Model (LLM) reasoning. However,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Improving the Realism of Synthetic Clinical Benchmarks Under Utility Constraints

DGX agent

arXiv:2608.06265v1 Announce Type: new Abstract: Synthetic clinical benchmarks for enterprise AI agents can pass existing utility checks and still remain structurally unrealistic, especially in privacy

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Innocent Panels, Hateful Stories: Evaluating and Detecting Hateful Intent in Multi-Turn Visual Story Generation

DGX agent

arXiv:2608.05210v1 Announce Type: cross Abstract: Picture books and comics have long been used to disseminate hateful narratives because they are easily understood even by children, as exemplified by

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Invariant Representation Learning for Source-Free Time Series Forecasting with LLM-Centric Proxy Denoising

DGX agent

arXiv:2510.05589v3 Announce Type: replace-cross Abstract: Effective time series forecasting enables various real-world applications, benefiting from the proliferation of mobile devices. However, the v

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

IPV-Bench: Benchmarking Image Protection Methods under Diverse Image-to-Video Generation Scenarios

DGX agent

arXiv:2603.26154v2 Announce Type: replace Abstract: Image-to-video (I2V) generation models can be misused to animate a single image into a convincing fake video, motivating perturbation-based image pr

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Iterate or Widen? When Test-Time Refinement Helps LiDAR Scene Completion: A Controlled Study of Evidence Geometry, Training Coverage, and Compute

DGX agent

arXiv:2608.06014v1 Announce Type: new Abstract: Should a completion model spend extra test-time compute by iterating, or spend a similar parameter budget on a wider one-shot predictor? The answer is e

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

JoyAI-RA 0.5: Scaling Robot Manipulation Learning via Dual Action Alignment

DGX agent

arXiv:2608.05674v1 Announce Type: new Abstract: Robot data is scarce, so generalist policies need to learn from heterogeneous sources, including human egocentric video, simulation, and real robots, wh

model-releasesarxiv-cs-ro
7 Aug 2026
Model Releases

Kastor: An efficient fine-tuning strategy for generative emulation of PDE simulations

DGX agent

arXiv:2608.06107v1 Announce Type: new Abstract: Machine learning offers a promising avenue to accelerate physical simulations by replacing computationally expensive traditional Partial Differential Eq

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

KILVO: Kinematic-Inertial-LiDAR-Visual Odometry with Robust Multimodal Adaptation for Humanoid Robots

DGX agent

arXiv:2608.05647v1 Announce Type: new Abstract: This article presents a kinematic-inertial-LiDAR-visual odometry for humanoid robots, called KILVO. Tailored to the platform features, requirements, and

model-releasesarxiv-cs-ro
7 Aug 2026
Model Releases

KV-Skill: Forging Expertise in the Model's Native Language

DGX agent

arXiv:2608.05475v1 Announce Type: new Abstract: Task knowledge is commonly stored either as text in the prompt or as an update to model weights. Text is modular but must be interpreted on every use, w

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

LangChoiceBench: Measuring and Explaining Programming-Language Choice in LLMs

DGX agent

arXiv:2608.06041v1 Announce Type: cross Abstract: Large language models (LLMs) have been shown to exhibit strong Python preferences when generating project-level code, but there is currently no system

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Layer-wise Positional Bias in Short-Context Language Modeling

DGX agent

arXiv:2601.04098v2 Announce Type: replace-cross Abstract: Transformer language models systematically prefer tokens at specific input positions regardless of semantic relevance---a phenomenon known as

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Learning When to Trust via Selective Context Preference Optimization

DGX agent

arXiv:2608.06377v1 Announce Type: cross Abstract: Language models increasingly condition their answers on external signals, and a single misleading one can turn a correct answer wrong. The obvious rem

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

LLM Inference Under Bursty Workload Distribution: Modifying the WAIT Algorithm

DGX agent

arXiv:2608.06135v1 Announce Type: new Abstract: Large Language Models (LLMs) such as ChatGPT and Claude are widely used for information retrieval and problem-solving. Recent work has focused on improv

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

LoDA: A Level of Detection Aware Method and a Multimodal Sensing Benchmark for Object Level Change Detection

DGX agent

arXiv:2608.05356v1 Announce Type: new Abstract: High-definition 3D LiDAR maps are important for autonomous driving and smart-city services, which require reliable detection of object-level changes in

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Look Twice: Training-Free Evidence Highlighting for Knowledge-based Visual Question Answering

DGX agent

arXiv:2604.01280v2 Announce Type: replace-cross Abstract: Knowledge-based Visual Question Answering (KB-VQA) requires Multimodal Large Language Models (MLLMs) to identify and combine fine-grained visu

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs

DGX agent

arXiv:2608.05246v1 Announce Type: new Abstract: Existing personalized LLM benchmarks primarily rely on textual personas or isolated behavioral signals, providing limited evaluation of cross-domain beh

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

M^3R-Bench: A Unified Benchmark for Evidence-Grounded Multimodal Metaphor Understanding

DGX agent

arXiv:2608.05817v1 Announce Type: new Abstract: Metaphor enables the understanding of abstract concepts through cross-domain mappings while conveying affective attitudes. In multimodal scenarios, visu

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

MAC 2026: Advancing Micro-Action Analysis Towards Fine-Grained Understanding

DGX agent

arXiv:2607.16284v2 Announce Type: replace Abstract: Micro-Actions (MAs) are subtle and spontaneous human behaviors that provide important non-verbal cues in social interaction and affective communicat

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

MameLoshnLM: Yiddish Language Model and Evaluation Benchmark

DGX agent

arXiv:2608.05850v1 Announce Type: cross Abstract: We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, it

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

MASS: Multiplayer World Models with Authoritative Shared State

DGX agent

arXiv:2608.06257v1 Announce Type: new Abstract: Current video world models struggle in multiplayer environments because they entangle world state with view-dependent visual latents, leading to redunda

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Matching Matters: A Fair Quality-Efficiency Benchmark for Command-Line Agents

DGX agent

arXiv:2606.21140v2 Announce Type: replace-cross Abstract: Rapid advances in large language models have improved the task-solving capabilities of command-line-interface (CLI)-based agents, whose CLIs d

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Matrix Zonotopic Attention: A Context-Adaptive Value Projection for Set Transformers

DGX agent

arXiv:2608.05472v1 Announce Type: cross Abstract: Multi-head attention combines an input-dependent softmax routing with an input-independent linear value projection, so the per-sample operator mapping

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

MAVISEG: Manifold Propagation and Visual Prototypes for Zero-Shot Open-Vocabulary Segmentation in Diffusion Transformers

DGX agent

arXiv:2608.05878v1 Announce Type: new Abstract: Text-to-image diffusion transformers learn about objects and scenes by learning to generate them, making them strong candidates for training-free zero-s

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

MetaboLLM: a metabolomics-specialized large language model for biochemical knowledge integration and predictive metabolite graph construction

DGX agent

arXiv:2608.06253v1 Announce Type: new Abstract: Metabolomics knowledge is distributed across heterogeneous resources and remains difficult to translate into predictive representations. We developed Me

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Mind the Gaps: Mixture-of-Minds for Human Simulation

DGX agent

arXiv:2608.06115v1 Announce Type: new Abstract: Predicting how a population will answer a new question is a long-standing goal. Statistical methods succeed at the level of the mass but falter at the l

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

MoCA: Implicit Social Context Analysis

DGX agent

arXiv:2608.05825v1 Announce Type: new Abstract: Human social communication, such as affection and intent, is often conveyed in highly implicit ways, where underlying meanings are expressed through ind

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification

DGX agent

arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of bet

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Multi-Representation Geometric Hierarchy Fusion: An Implicit-Submap Driven Framework for Resilient 3D Place Recognition

DGX agent

arXiv:2506.14243v4 Announce Type: replace Abstract: LiDAR-based place recognition is critical for long-term autonomous driving without GPS. Existing handcrafted feature methods face dual limitations.

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

NavTrust: Benchmarking Trustworthiness for Embodied Navigation

DGX agent

arXiv:2603.19229v2 Announce Type: replace-cross Abstract: There are two major categories of embodied navigation: Vision-Language Navigation (VLN), where agents navigate by following natural language i

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

DGX agent

arXiv:2608.06292v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. H

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Neuro-Symbolic Closed-Loop Control of Laser Powder Bed Fusion with an In-Loop Ontology

DGX agent

arXiv:2608.05773v1 Announce Type: new Abstract: A geometry-conditioned, neuro-symbolic closed-loop architecture is proposed for laser powder bed fusion, in which a standards-aligned ontology operates

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

DGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

One Leak Away: How Pretrained Model Exposure Amplifies Jailbreak Risks in Finetuned LLMs

DGX agent

arXiv:2512.14751v3 Announce Type: replace-cross Abstract: Finetuning pretrained large language models (LLMs) has become the standard paradigm for developing downstream applications. However, its secur

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Operating Multi-Node Full Fine-Tuning on NVIDIA B300: A Field Report on Telemetry-Based Triage, Negative Results, and Operational Hardening

DGX agent

arXiv:2608.05944v1 Announce Type: cross Abstract: We report operational experience full-fine-tuning a 32.76B-parameter dense model (Qwen3-32B) on 16 x NVIDIA B300 (two nodes, FSDP / ZeRO-3) -- among t

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

DGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Otter: A Time-Aware, History-Conditioned Human Chess AI

DGX agent

arXiv:2608.05206v1 Announce Type: new Abstract: Otter is a 15.3M-parameter human chess AI that predicts human move selection by modeling play as a time-aware, sequential process rather than treating e

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…1920212223…357
Next →