AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Model Releases

MiraMind: Benchmarking Reliable Mental Health Reasoning beyond Answer Accuracy

DGX agent

arXiv:2512.09636v3 Announce Type: replace Abstract: Mental-health reasoning with large language models (LLMs) is an evidence-constrained judgment problem: models must transform limited, subjective, an

model-releasesarxiv-cs-cl
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling

DGX agent

arXiv:2608.08553v1 Announce Type: new Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from

safetyarxiv-cs-cv
11 Aug 2026
Research

MRI super-resolution in ten sampling steps using a diffusion bridge model

DGX agent

arXiv:2608.08819v1 Announce Type: new Abstract: Objective. MRI provides excellent soft-tissue contrast, but long acquisition times can cause patient discomfort and lead to motion artifacts, forcing a

researcharxiv-cs-cv
11 Aug 2026
Model Releases

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

DGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

SwissCrop25: A National Multi-Year Benchmark for Operational Crop Mapping

DGX agent

arXiv:2608.09497v1 Announce Type: new Abstract: Operational crop mapping requires models that generalise across years, resolve fine-grained crop taxonomies, and distinguish cropland from surrounding l

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?

DGX agent

arXiv:2608.07899v1 Announce Type: new Abstract: Agent systems increasingly expose execution traces, yet telemetry that reveals a failure may still be inadequate for identifying where that failure orig

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Research

A foundation-model approach to pediatric headache classification from rs-fMRI

DGX agent

arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona

researcharxiv-cs-lg
10 Aug 2026
Research

A Transferable Autologistic Model for Predicting Rare Failures in Heterogeneous Equipment

DGX agent

arXiv:2608.06695v1 Announce Type: new Abstract: Predicting failures before they occur remains a major challenge in predictive maintenance, particularly when failures are rare, when equipment of the sa

researcharxiv-cs-lg
10 Aug 2026
Safety

Bootstrap-Conditioned Action Selection with Tabular Foundation Models

DGX agent

arXiv:2608.06559v1 Announce Type: new Abstract: Contextual bandits offer a natural framework for sample-efficient personalization, but practical deployment remains difficult under sparse, biased inter

safetyarxiv-cs-lg
10 Aug 2026
Model Releases

Cascading Through the Hierarchy: Regularizer-Induced Feature Detection as Phase Transitions in Deep Linear Neural Networks

DGX agent

arXiv:2608.06597v1 Announce Type: cross Abstract: A scientific theory of deep learning, comprising learning dynamics and statistical properties of learned models, is rapidly gaining attention. One of

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers

DGX agent

arXiv:2608.06867v1 Announce Type: new Abstract: No single large language model (LLM) is optimal across all queries and budget constraints, making model routing essential for cost-effective deployment.

model-releasesarxiv-cs-cl
10 Aug 2026
Research

Measurements Automatically Extracted from Zero Echo Time MRI Using Deep Learning Image Segmentation and Geometric Modeling Agree with Expert Manual Readings

DGX agent

arXiv:2608.07368v1 Announce Type: cross Abstract: Computed tomography (CT) remains the reference for 3D osseous morphometry in femoroacetabular impingement (FAI) but requires ionizing radiation and ma

researcharxiv-cs-ai
10 Aug 2026
Model Releases

SABRE: Scalable and Automated Benchmarking of VLMs under Stress

DGX agent

arXiv:2608.07435v1 Announce Type: cross Abstract: Vision-language models (VLMs) are improving rapidly, but benchmark development lags behind, making weaknesses hard to identify. Building stress tests

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Seeking SOTA: Time-Series Forecasting Must Adopt Taxonomy-Specific Evaluation to Dispel Illusory Gains

DGX agent

arXiv:2603.15506v2 Announce Type: replace-cross Abstract: We argue that the current practice of evaluating AI/ML time-series forecasting models, predominantly on benchmarks characterized by strong, pe

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Toward a Causal Data Management Ecosystem for Decision Making and Agentic AI

DGX agent

arXiv:2608.07214v1 Announce Type: cross Abstract: Modern AI is no longer a single model but an ecosystem: classical ML predictors, deep and multimodal models, large language models, and agents, each t

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

A Six-Dimensional Taxonomy of Post-Training Adaptation Techniques with Applications in AI Governance

DGX agent

arXiv:2608.06246v1 Announce Type: new Abstract: Post-training adaptation has become central to modern machine learning practice and includes techniques such as retraining, fine-tuning, parameter-effic

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features

DGX agent

arXiv:2608.06008v1 Announce Type: new Abstract: Large video diffusion models provide rich spatiotemporal priors for autonomous driving, but existing world-action models often inherit the cost of itera

model-releasesarxiv-cs-ro
7 Aug 2026
Tutorials

ALTER: Modeling Longitudinal Changes via Regional Differencing for 3D CT Report Generation

DGX agent

arXiv:2608.05615v1 Announce Type: new Abstract: Computed tomography (CT) is widely used for clinical diagnosis and longitudinal follow-up, yet automatically generating accurate and complete radiology

tutorialsarxiv-cs-cv
7 Aug 2026
Safety

ARGUS: Aligning Robot Scene Geometry Under Shifting Views with Large 3D Vision Models

DGX agent

arXiv:2608.05579v1 Announce Type: new Abstract: Large-scale visuomotor policies have demonstrated impressive performance across a wide range of robot manipulation tasks. However, despite this success,

safetyarxiv-cs-ro
7 Aug 2026
Research

Controllable Clothing: Precise Labels and Generation for Virtual Try-On with Latent Diffusion Models

DGX agent

arXiv:2608.05834v1 Announce Type: new Abstract: In this technical report, I present a new method for guiding image generation in the context of Virtual- Try-On (VITON). The proposed method leverages n

researcharxiv-cs-cv
7 Aug 2026
Research

Counterfactual Analysis via Large Language Models

DGX agent

arXiv:2608.05367v1 Announce Type: new Abstract: Counterfactual analysis aims to predict potential outcomes under hypothetical scenarios, offering valuable insights for decision-making. This paper inve

researcharxiv-cs-ai
7 Aug 2026
Research

HyTBE: Hyperbolic Target-Background Expert Model for Cross-Domain Infrared Small Target Detection

DGX agent

arXiv:2608.05771v1 Announce Type: cross Abstract: Infrared small target detection (IRSTD) has achieved substantial progress under domain-consistent evaluation, yet detector performance often degrades

researcharxiv-cs-ai
7 Aug 2026
Safety

Large Language Models Threaten Double-blind Review

DGX agent

arXiv:2608.05157v1 Announce Type: cross Abstract: Double blind peer review serves as the scientific community primary defense against status and affiliation bias. Its effectiveness rests on the assump

safetyarxiv-cs-ai
7 Aug 2026
Tutorials

RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction

DGX agent

arXiv:2608.06259v1 Announce Type: new Abstract: Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, l

tutorialsarxiv-cs-lg
7 Aug 2026
Research

ViSR-KGC: Visual Subgraph Reasoning with Vision-Language Models for Multimodal Knowledge Graph Completion

DGX agent

arXiv:2608.05833v1 Announce Type: new Abstract: Knowledge graph completion (KGC) aims to infer missing entities or relations from incomplete graph structures, and has evolved into multimodal knowledge

researcharxiv-cs-ai
7 Aug 2026
Tutorials

VLAff: Vision-Language-Affordance Model for Unified Actionable Affordances

DGX agent

arXiv:2608.05215v1 Announce Type: cross Abstract: Learning manipulation skills from human videos is promising for scalable robot learning. However, the embodiment mismatch between humans and robots ma

tutorialsarxiv-cs-cv
7 Aug 2026
Research

When Drafts Evolve: Speculative Decoding Meets Online Learning

DGX agent

arXiv:2603.12617v2 Announce Type: replace-cross Abstract: Speculative decoding has emerged as a widely adopted paradigm for accelerating large language model inference, where a lightweight draft model

researcharxiv-cs-ai
7 Aug 2026
Applications

Benchmarking Deep Learning Models for Dense Event Classification of Offshore Wind Infrastructure in Sentinel-1 Time Series

DGX agent

arXiv:2608.04706v1 Announce Type: new Abstract: Monitoring of offshore wind energy infrastructure life cycles, especially during the deployment phase, is an important contribution for stakeholders to

applicationsarxiv-cs-lg
6 Aug 2026
Model Releases

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings

DGX agent

arXiv:2608.04735v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is increasingly treated as an important safety layer for frontier reasoning models. Most monitorability evaluations st

model-releasesarxiv-cs-ai
6 Aug 2026
Research

CSGen: A Multi-Domain Curvilinear Structure Generation Model via Hierarchical Multimodal Diffusion

DGX agent

arXiv:2608.04655v1 Announce Type: cross Abstract: Curvilinear structure analysis is an important and fundamental task in multimedia. However, the controllable generation of images with precise curvili

researcharxiv-cs-ai
6 Aug 2026
Model Releases

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

DGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

model-releasesarxiv-cs-cl
6 Aug 2026
Research

Do Language Models Know Their Slang? Queer Slang Understanding in User-Generated Content

DGX agent

arXiv:2608.04847v1 Announce Type: new Abstract: Despite its cultural relevance and diffusion, queer slang remains underrepresented in Natural Language Processing research. Towards addressing this gap,

researcharxiv-cs-cl
6 Aug 2026
Model Releases

Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark

DGX agent

arXiv:2608.04670v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed computational linguistics and achieved remarkable performance across numerous natural language processin

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

Embedding Large Language Models into Flow Controls: An Agentic Framework for Adaptive and Trustworthy Automated Cooking

DGX agent

arXiv:2608.04768v1 Announce Type: new Abstract: Automated cooking robots have traditionally relied on predefined procedures and rule-based control, ensuring stable execution but offering limited perso

agentsarxiv-cs-cv
6 Aug 2026
Model Releases

GEB-Bench: Abstract Structures Told in Many Voices

DGX agent

arXiv:2608.04111v1 Announce Type: cross Abstract: Can a model look at a river delta and a lightning bolt and see that they share a structure? We introduce GEB-Bench, a benchmark whose unit is an abstr

model-releasesarxiv-cs-cl
6 Aug 2026
Research

Group-Equivariant Diffusion Models for Lattice Field Theory

DGX agent

arXiv:2510.26081v2 Announce Type: replace-cross Abstract: Near the critical point, Markov Chain Monte Carlo (MCMC) simulations of lattice quantum field theories (LQFT) become increasingly inefficient

researcharxiv-cs-lg
6 Aug 2026
Model Releases

MathDebugger: Detecting and Diagnosing Errors in Synthetic Mathematical Data

DGX agent

arXiv:2502.19058v2 Announce Type: replace Abstract: Synthetic mathematical data has become an important resource for scaling the reasoning capabilities of large language models, yet errors in generate

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

DGX agent

arXiv:2608.04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the iss

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

Spoken Function Calling: A New Perspective on Spoken Language Understanding for Large Audio Language Models

DGX agent

arXiv:2608.05126v1 Announce Type: new Abstract: Spoken Language Understanding (SLU) is the core component of task-oriented dialogue systems and a pivotal link in achieving seamless human-agent interac

agentsarxiv-cs-cl
6 Aug 2026
Model Releases

The Calibration Floor: Format Repair Can Masquerade as Self-Correction at Small-to-Mid Scale

DGX agent

arXiv:2608.04355v1 Announce Type: new Abstract: Accuracy changes after language-model self-revision are usually interpreted as changes in reasoning. We show this can fail at the answer-extraction boun

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

A Unified 2D Framework for DeepLesion Detection, Segmentation and Short Report Generation

DGX agent

arXiv:2608.02805v1 Announce Type: cross Abstract: In previous work, we integrated large language models (LLMs) into the lesion segmentation model based on the ULS23 DeepLesion dataset, using short-for

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Calibrating Semantic Uncertainty from Observable Language-Model Probabilities

DGX agent

arXiv:2607.17447v2 Announce Type: replace-cross Abstract: As generative artificial intelligence enters scientific and professional work, its uncertainty must be defined on the states that matter for i

researcharxiv-cs-cl
5 Aug 2026
Research

Conditionally Identifiable Latent-Environment Modeling for Out-of-Distribution Recommendation

DGX agent

arXiv:2608.03647v1 Announce Type: cross Abstract: Out-of-distribution (OOD) recommendation is vulnerable to preference shifts induced by a latent environment. Existing methods can infer latent states

researcharxiv-cs-lg
5 Aug 2026
Model Releases

Does Forgetting Transfer Across Modalities? A Real-World Benchmark for Cross-Modal Knowledge Unlearning Evaluation

DGX agent

arXiv:2608.03791v1 Announce Type: new Abstract: Vision-Language Models (VLMs), like Large Language Models (LLMs), may memorize sensitive, copyrighted, or harmful knowledge from their pretraining corpo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Efficient Multilingual Neural Machine Translation via Corpus-Driven Vocabulary Pruning: An English-Arabic Case Study

DGX agent

arXiv:2608.03480v1 Announce Type: new Abstract: The adoption of large pre-trained multilingual models for neural machine translation (MNMT) faces a major challenge: excessive memory and computational

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Evaluating LLM Trade-offs for Enterprise Automation: Lessons from Workflow Generation in a Production Enterprise Platform

DGX agent

arXiv:2608.03311v1 Announce Type: cross Abstract: Enterprise compliance management requires rapid adaptation to evolving regulatory frameworks (e.g., DORA, AI RMF, FedRAMP) and tight remediation SLAs.

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…225226227228229…1038
Next →