AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

DGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

model-releasesarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Trajectory inference via Acceleration Matching

DGX agent

arXiv:2608.03916v1 Announce Type: new Abstract: Trajectory inference is a fundamental problem in many scientific domains: given a collection of unpaired snapshots of observations at discrete time poin

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

DGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection

DGX agent

arXiv:2608.03627v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for automated fact-checking, yet their susceptibility to gender bias in this context remains underexp

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

DGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

UrbanAgent: A Tool-Augmented Agent for Cross-System Urban Tasks

DGX agent

arXiv:2608.03018v1 Announce Type: new Abstract: Modern cities rely on an increasing number of digital services to operate, but residents' daily needs are still difficult to meet. Services are fragment

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Verifier-Guided Model Discovery for Physical Dynamical Systems with Pretrained Symbolic Transformers

DGX agent

arXiv:2608.02662v1 Announce Type: cross Abstract: Reliable forecasting of nonlinear physical systems underpins scientific discovery and engineering decision-making. Yet high-fidelity simulations are p

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

DGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

DGX agent

arXiv:2608.03810v1 Announce Type: cross Abstract: Large language models routinely describe socially salient targets, including political figures, countries, religions, organizations, historical events

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

VIBE: Vector Index Benchmark for Embeddings

DGX agent

arXiv:2505.17810v2 Announce Type: replace Abstract: Approximate nearest neighbor (ANN) search is a performance-critical component of many machine learning pipelines, and rigorous benchmarking is essen

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

DGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP

DGX agent

arXiv:2608.03095v1 Announce Type: new Abstract: We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurati

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Vulnerabilities, Secrets and Misconfiguration in the Highest-Exposure Docker Hub Images

DGX agent

arXiv:2608.02669v1 Announce Type: cross Abstract: Docker Hub is the registry underneath most container deployments, and a flaw in a widely reused base image is inherited by every image built on it. Pr

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks

DGX agent

arXiv:2608.03499v1 Announce Type: new Abstract: Recent advances in persistent personal-agent frameworks are making human-centered agent networks realistic deployment targets: each user can be served b

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

DGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

When Attention Goes Blind: Numerical Failure in ALiBi Positional Encodings

DGX agent

arXiv:2608.03994v1 Announce Type: new Abstract: We identify a previously overlooked failure mode of ALiBi positional encoding: its linear bias scaling underflows floating-point precision, which zeroes

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

When Classes Evolve: A Benchmark and Framework for Stage-Aware Class-Incremental Learning

DGX agent

arXiv:2602.00573v2 Announce Type: replace-cross Abstract: Class-Incremental Learning (CIL) aims to sequentially learn new classes while mitigating catastrophic forgetting of previously learned knowled

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

When Correct Solutions Repeat: Rarity-Aware Credit Redistribution for GRPO

DGX agent

arXiv:2608.03467v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) com- monly optimizes each correct completion as an independent learning signal. In GRPO, this comp

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Do Fewer Visual Tokens Accelerate Multimodal Inference? A Break-Even Study Across Decision Locations and Hardware

DGX agent

arXiv:2608.03649v1 Announce Type: new Abstract: Fewer visual tokens do not guarantee lower end-to-end latency. We evaluate break-even with a reproducible protocol that accounts for decision overhead,

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs

DGX agent

arXiv:2608.03506v1 Announce Type: new Abstract: Self-consistency assumes the most frequent answer among sampled reasoning traces is the most reliable, but this can fail in causal reasoning: samples of

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Memory Becomes Authority: Benchmarking Authority Collapse at the Memory Consolidation Boundary

DGX agent

arXiv:2608.01679v2 Announce Type: replace Abstract: Persistent memory allows (self-evolving) LLM agents to adapt across tasks by consolidating heterogeneous interaction histories into reusable facts,

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Outputs Disperse, Does Epistemic Revision Follow? A Black-Box Coupling Diagnostic for Machine Collectives

DGX agent

arXiv:2608.03722v1 Announce Type: new Abstract: Collective intelligence research treats disagreement as evidence of epistemic diversity: if agents express different views, the group should retain capa

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models

DGX agent

arXiv:2608.03201v1 Announce Type: new Abstract: Safety guards are widely used to filter harmful content and are typically trained via supervised fine-tuning on labeled prompt-response pairs. We audit

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking

DGX agent

arXiv:2606.15673v2 Announce Type: replace Abstract: Web agents act through long interaction sequences, yet existing benchmarks evaluate only terminal success, discarding all process information and of

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't

DGX agent

arXiv:2608.02829v1 Announce Type: new Abstract: Model families train every size from scratch. Can a pretrained large model be converted into a smaller sibling? We characterize the 1.4B->410M conversio

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament

DGX agent

arXiv:2608.04008v1 Announce Type: new Abstract: Benchmarks that measure the forecasting ability of large language models are almost always retrospective: the event has happened, the answer is somewher

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

DGX agent

arXiv:2608.02657v1 Announce Type: cross Abstract: Agentic LLMs are vulnerable to indirect prompt injection (IPI) attacks, e.g., malicious side-tasks hidden in external tool results. While many efforts

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

A Benchmark Dataset for MLLM-Generated Image Detection: GPT Image2 & Nano Banana2

DGX agent

arXiv:2608.01258v1 Announce Type: new Abstract: The realism of images generated by multimodal large language models (MLLMs), such as GPT Image2 and Nano Banana2, has improved rapidly in recent years.

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

A Comparative Analysis of MLP and Kolmogorov-Arnold Networks (KAN) for Faster-than-Nyquist (FTN) Signaling Detection

DGX agent

arXiv:2608.02062v1 Announce Type: cross Abstract: Faster-than-Nyquist signaling improves spectral ef- ficiency by deliberately introducing inter-symbol interference. Classical sequence detectors such

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

A Few Neurons Reveal When LLMs Misuse Tools: Sparse Detection and Selective Steering for Reliable Tool Use

DGX agent

arXiv:2608.00218v1 Announce Type: new Abstract: Agentic LLMs exhibit three consequential tool-use failures: invalid arguments (validity), unnecessary calls (over-calling), and omitted calls when tools

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

A General-Purpose VLM Can Teach an Astronomy Foundation Model to Better Recognize Galaxy Morphology

DGX agent

arXiv:2608.02300v1 Announce Type: new Abstract: Existing astronomy foundation models provide strong galaxy representations, but adapting them to new survey conditions and survey-specific morphology re

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

A Large-Scale Multi-Dimensional Empirical Study of LLMs for Conversation Summarization

DGX agent

arXiv:2606.15974v2 Announce Type: replace Abstract: Despite the significant advancement of LLMs in conversation summarization, their evaluation remains limited by insufficient scenarios, input lengths

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

A Simple Approximation to the Distribution of the Ridge Regression Estimator

DGX agent

arXiv:2608.02539v1 Announce Type: cross Abstract: We present a simple Gaussian approximation to the finite-sample distribution of the classical ridge regression estimator. Our approximation captures t

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

A Triple-Robustness Analysis of Retrieval-Augmented Generation for Multi-Hop Requirements Traceability

DGX agent

arXiv:2608.00705v1 Announce Type: cross Abstract: Reported verdicts on GraphRAG versus vector RAG disagree, and the evidence is typically tied to a single corpus, embedder, and judge -- and, we show,

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Abduction Without a Body? Representational Grounding and the Abduction Loop for Scientific Hypothesis Generation

DGX agent

arXiv:2608.02505v1 Announce Type: cross Abstract: Can scientific abduction occur without continuous sensorimotor embodiment? Recent arguments in AI and philosophy of science hold that genuine hypothes

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Accuracy Does Not Guarantee Human-Likeness: Cross-Domain Human-Centered Benchmark in Monocular Depth Estimation

DGX agent

arXiv:2512.08163v2 Announce Type: replace Abstract: Deep neural networks (DNNs) are increasingly used as functional models of human vision, yet standard monocular depth estimation (MDE) benchmarks lar

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

AdaMTP: An Adaptive Training Paradigm for Multi-Token Prediction

DGX agent

arXiv:2608.00434v1 Announce Type: new Abstract: Multi-Token Prediction (MTP) has emerged as an effective paradigm that augments a shared Large Language Model backbone with auxiliary heads, training th

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Adaptive Quantum Physics-Informed Neural Networks for Differential Equations with Applications to Fluid Dynamics

DGX agent

arXiv:2608.00850v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a versatile approach for solving nonlinear partial differential equations (PDEs), yet achieving

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Adaptive Reconstruction of Bosonic Quantum States

DGX agent

arXiv:2608.02049v1 Announce Type: cross Abstract: Bosonic quantum systems provide a hardware-efficient platform for quantum information processing but remain challenging to characterise due to their l

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification

DGX agent

arXiv:2509.24560v2 Announce Type: replace Abstract: Extended Chain-of-Thought (CoT) reasoning has significantly bolstered the capabilities of medical large language models (LLMs). However, current mod

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

AdvPlan-Bench: Adversarial Evaluation of Structured Plan-Generation Agents

DGX agent

arXiv:2608.00832v1 Announce Type: new Abstract: Structured plan-generation agents are often evaluated as if a plan has quality in isolation, yet many realistic planning tasks require asking how a cand

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AeroLLE: Constrained Pseudo-Supervision for Nighttime Aerial Image Enhancement with the AeroNight-1.5K Benchmark

DGX agent

arXiv:2608.00702v1 Announce Type: new Abstract: Nighttime aerial image enhancement is challenged by spatially nonuniform exposure, mixed illumination, and weak structural evidence, while registered no

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

DGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AgentMemBench: A Systematic Benchmark for Evaluating Long-Term Memory Management Strategies in Conversational AI Agents

DGX agent

arXiv:2608.00009v1 Announce Type: new Abstract: Long-term memory remains a critical bottleneck for conversational AI agents, whose finite context windows cannot support coherent recall across thousand

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Amortized Inference of Multi-Modal Posteriors using Likelihood-Weighted Normalizing Flows

DGX agent

arXiv:2512.04954v3 Announce Type: replace Abstract: We present a novel technique for amortized posterior estimation using Normalizing Flows trained with likelihood-weighted importance sampling. This a

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

An Accessible Solution for Deformable Image Registration Compared with Learning-Based Approaches

DGX agent

arXiv:2608.02248v1 Announce Type: cross Abstract: Deformable image registration (DIR) is a core problem in medical image analysis; but, unlike labeling decision problems such as classification and seg

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

An Embedded RISC-V Evaluation of Kolmogorov--Arnold Networks in Hard-Constrained Recurrent Physics-Informed Models

DGX agent

arXiv:2608.00737v1 Announce Type: new Abstract: Hard-constrained recurrent physics-informed networks (HRPINNs) embed known dynamics inside a recurrent numerical integrator and restrict a neural branch

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…2930313233…357
Next →