AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,587 results
Model Releases

Who Checks the Citations? Benchmarking Legal Hallucination Detection

DGX agent

arXiv:2606.21155v2 Announce Type: replace Abstract: Attorneys, judges, and pro se filers increasingly use AI to draft legal documents, yet these tools frequently fabricate citations. Despite predictio

model-releasesarxiv-cs-cl
7 Aug 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Mechanistic Analysis of Transformers for Dynamical Systems

DGX agent

arXiv:2512.21113v2 Announce Type: replace Abstract: Transformers are increasingly adopted for modeling and forecasting time-series, yet their internal mechanisms remain poorly understood from a dynami

researcharxiv-cs-lg
6 Aug 2026
Model Releases

A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents

DGX agent

arXiv:2602.06052v4 Announce Type: replace-cross Abstract: Research in artificial intelligence is shifting from model innovations and benchmark scores towards problem definition and rigorous real-world

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

DGX agent

arXiv:2608.04682v1 Announce Type: cross Abstract: Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific b

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Adaptive Finite-Budget Training for CVaR Risk-Aware Q-Learning

DGX agent

arXiv:2608.04305v1 Announce Type: new Abstract: Risk-aware Q-learning (RaQL) provides a model-free, two-timescale estimator for dynamic risk objectives, but its finite-budget behavior remains fragile:

model-releasesarxiv-cs-lg
6 Aug 2026
Safety

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

DGX agent

arXiv:2608.05042v1 Announce Type: new Abstract: Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as a promising paradigm for 3D robot m

safetyarxiv-cs-ro
6 Aug 2026
Research

Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference

DGX agent

arXiv:2608.04428v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have emerged as a key component in embodied AI. Among existing approaches, diffusion-based VLA models achieve supe

researcharxiv-cs-lg
6 Aug 2026
Model Releases

EgoAfford: Task-Oriented Affordance Grounding via Egocentric Referring Segmentation

DGX agent

arXiv:2608.04533v1 Announce Type: new Abstract: Part-level affordance grounding has advanced the localization of functional object regions associated with elemental actions. Extending this capability

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

LaPrune: Controllable Differentiable Sparsity at Million Scale

DGX agent

arXiv:2608.04057v1 Announce Type: cross Abstract: Top-k selection determines which components of a sparse model remain active. Hard selection blocks gradients, while continuous relaxations often coupl

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning

DGX agent

arXiv:2510.21084v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong potential for clinical decision support through their advanced language understanding and reaso

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

DGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

DGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

model-releasesarxiv-cs-ai
6 Aug 2026
Research

PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior

DGX agent

arXiv:2512.18367v2 Announce Type: replace-cross Abstract: Diffusion models are highly expressive image priors for Bayesian inverse problems. However, most diffusion models cannot operate on large-scal

researcharxiv-cs-lg
6 Aug 2026
Model Releases

RESPClinBench: Benchmarking Multimodal Clinical Decision-Making and Longitudinal Disease Management in Respiratory Specialty Care

DGX agent

arXiv:2608.04514v1 Announce Type: new Abstract: Background: Respiratory specialty care requires multimodal interpretation, longitudinal risk assessment, guideline-concordant intervention, and whole-co

model-releasesarxiv-cs-cl
6 Aug 2026
Research

REZE: Recognition-Based Zero-Shot Extraction for Video Temporal Grounding

DGX agent

arXiv:2608.04480v1 Announce Type: new Abstract: Video temporal grounding (VTG) refers to the task of identifying the time interval in a video that corresponds to a given natural-language query. A comm

researcharxiv-cs-cv
6 Aug 2026
Model Releases

Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load

DGX agent

arXiv:2608.05018v1 Announce Type: new Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Tactus: Open-Vocabulary Object Recognition from Low-Cost Pressure Arrays

DGX agent

arXiv:2608.04043v1 Announce Type: new Abstract: Resistive pressure arrays are the cheapest and most widely shipped tactile sensors, yet tactile representation learning has concentrated on optical sens

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

The Loss Does Not See the Basis, but Adam Does

DGX agent

arXiv:2608.05136v1 Announce Type: new Abstract: Gradient descent on a factored model W = UV^op is implicitly biased toward low-rank solutions, while Adam, starting from the same small initialization,

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Transfer Learning for Named Entity Recognition of Classical Latin through LLM Prompting

DGX agent

arXiv:2608.04015v1 Announce Type: new Abstract: With the increase in digitized resources of Classical Latin texts and modern breakthroughs of Large Language Models (LLMs), I contribute to ancient lang

model-releasesarxiv-cs-cl
6 Aug 2026
Safety

A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues

DGX agent

arXiv:2608.03927v1 Announce Type: new Abstract: Engineered Skeletal Muscle Tissues (ESMs) have become a key structure for biomedical disease modeling and pharmacological screening, yet their functiona

safetyarxiv-cs-lg
5 Aug 2026
Model Releases

Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems

DGX agent

arXiv:2608.03744v1 Announce Type: new Abstract: Clinical decision support is moving toward committees of language-model agents deliberating on a shared workspace. We ask whether such committees can be

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

DGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety

DGX agent

arXiv:2601.17003v2 Announce Type: replace-cross Abstract: Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual dive

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

DGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

CARE-Bench: Benchmarking Patient-Facing LLM Triage

DGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ConlangBench: Exploring Language Knowledge and Learning in LLMs through Diverse Constructed Languages

DGX agent

arXiv:2608.03505v1 Announce Type: new Abstract: Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity. Despite their potential for s

model-releasesarxiv-cs-cl
5 Aug 2026
Research

EmbodiedVAE: Disentangled Video VAE for Efficient and Controllable Embodied Manipulation

DGX agent

arXiv:2608.02990v1 Announce Type: new Abstract: Latent diffusion models (LDMs) have recently significantly advanced embodied learning in constructing powerful embodied manipulation world models. Howev

researcharxiv-cs-ro
5 Aug 2026
Research

Every Wrong Answer Counts: Option-Level Psychometrics for LLM Multiple-Choice Benchmarks

DGX agent

arXiv:2608.02966v1 Announce Type: new Abstract: Most multiple-choice question (MCQ) benchmarks evaluate Large Language Models (LLMs) only by whether they select the correct answers. This binary scorin

researcharxiv-cs-cl
5 Aug 2026
Model Releases

HyVIC: A Metric-Driven Spatio-Spectral Hyperspectral Image Compression Architecture Based on Variational Autoencoders

DGX agent

arXiv:2603.26468v2 Announce Type: replace Abstract: The rapid growth of hyperspectral data archives in remote sensing (RS) necessitates effective compression methods for storage and transmission. Rece

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization

DGX agent

arXiv:2608.03210v1 Announce Type: new Abstract: Foundation models have achieved remarkable success across diverse tasks, but they remain vulnerable. To investigate such vulnerabilities, semantic-shift

safetyarxiv-cs-cl
5 Aug 2026
Model Releases

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation

DGX agent

arXiv:2608.02639v1 Announce Type: cross Abstract: Production prompts rarely carry a single instruction. One system message may require valid JSON, a word limit, three citations, and a fixed tone at th

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Inverted Detection and Control in Steering Vectors

DGX agent

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption un

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Knowing the Form, Not the Function: Automatically Auditing Answer--Authority Decoupling in Legal Benchmarks

DGX agent

arXiv:2608.02621v1 Announce Type: cross Abstract: Legal benchmarks typically score final answers even when models also state legal authority. We test whether answer correctness can serve as a proxy fo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Scale

DGX agent

arXiv:2608.02613v1 Announce Type: cross Abstract: Edge-deployed personal memory assistants must handle private interpersonal conversations on-device with open-weight models. Yet, existing memory bench

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MoEGen: Mixture-of-Experts for Instance-Adaptive LoRA Generation

DGX agent

arXiv:2608.03275v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) enables efficient adaptation of large language models, but existing MoE-based PEFT methods typically improve capa

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification

DGX agent

arXiv:2608.03474v1 Announce Type: new Abstract: Recent advances in Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities in web UI generation. However, existing benchmarks pre

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding

DGX agent

arXiv:2608.03708v1 Announce Type: new Abstract: Text-to-image diffusion models enable personalization of specific visual concepts from a small number of reference images. However, generating a single

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

MuRA: Multi-Rank Adaptation for Efficient and Effective Test-Time Vision-Language Generalization

DGX agent

arXiv:2608.03885v1 Announce Type: new Abstract: Vision-language models exhibit remarkable zero-shot capabilities but suffer significant performance degradation under distribution shifts. While test-ti

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Particle-based Generalised Stochastic Optimisation

DGX agent

arXiv:2608.02844v1 Announce Type: cross Abstract: We develop a class of diffusion-based stochastic particle optimisation methods for loss functions with intractable gradients. Specifically, we conside

model-releasesarxiv-cs-lg
5 Aug 2026
Safety

PhyAI: Real-Time Physical AI at the Edge, Scalable Rollouts in the Cloud

DGX agent

arXiv:2608.03682v1 Announce Type: new Abstract: Physical AI policies require inference throughout their lifecycle, including model evaluation, cloud reinforcement learning rollout, edge GPU serving, a

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Rectify Then Diffuse: Disentangling Concepts Before Denoising Trajectory Unfolds

DGX agent

arXiv:2608.03135v1 Announce Type: cross Abstract: Text-to-image diffusion models can generate individual concepts well, but they often omit or merge concepts incorrectly with multiple concepts. We tra

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

S^3: Improving Agent Safety through Multi-Stage Defense

DGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

DGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA

DGX agent

arXiv:2509.25459v2 Announce Type: replace Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. How

model-releasesarxiv-cs-cl
5 Aug 2026
Research

The Geometric Nature and a Free Proxy for Flow-Matching Uncertainty

DGX agent

arXiv:2607.27933v2 Announce Type: replace Abstract: Flow matching (FM) has become a popular action head paradigm for modern embodied models. However, as a conditional generative model, it does not exp

researcharxiv-cs-ai
5 Aug 2026
Model Releases

The Ignition Is Real, and It Lives at the Readout: Latent composition, difficulty-clocked ignition, and the interface-constituted commit in a recurrent-depth reasoner

DGX agent

arXiv:2608.03263v1 Announce Type: cross Abstract: We test whether the 'compositional ignition' reported in latent-reasoning models is real computation, an instrument artifact, or inherited from verbal

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Traceable Multi-Agent System for Knowledge-Based Forecasting

DGX agent

arXiv:2608.03339v1 Announce Type: new Abstract: Enterprise forecasting increasingly relies on autonomous agents that interpret documents, search for data, generate code, and revise models. While this

agentsarxiv-cs-ai
5 Aug 2026
← Previous
1…394395396397398…1075
Next →