AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering

DGX agent

arXiv:2411.19504v2 Announce Type: replace Abstract: The advance of large language models (LLMs) has unlocked great opportunities in complex multi-modal data management tasks, particularly in question

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

TRACER: Token ReAssignment for Concept ERasure in Generative Recommendation

DGX agent

arXiv:2606.07688v1 Announce Type: cross Abstract: Generative recommendation formulates next-item prediction as autoregressive generation over semantic ID (SID) sequences derived from users' historical

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model

DGX agent

arXiv:2603.25184v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for post-training large language models (LLMs) in reasoning tasks. While scaling rollouts can

tutorialsarxiv-cs-ai
9 Jun 2026
Research

Training-Free Intelligibility-Guided Observation Addition for Noisy ASR

DGX agent

arXiv:2602.20967v2 Announce Type: replace-cross Abstract: Automatic speech recognition (ASR) degrades severely in noisy environments. Although speech enhancement (SE) front-ends effectively suppress b

researcharxiv-cs-ai
9 Jun 2026
Safety

Training-Inference Kernel Contracts: Bounding Divergence in Post-Training and Deployment

DGX agent

arXiv:2606.07581v1 Announce Type: cross Abstract: A modern post-training pipeline often writes one symbol for its policy, pi_theta, while evaluating it through two different programs: a training kerne

safetyarxiv-cs-ai
9 Jun 2026
Safety

Trait-space Monitoring for Emergent Misalignment During Supervised Finetuning

DGX agent

arXiv:2606.07631v1 Announce Type: cross Abstract: Emergent misalignment (EM) occurs when narrow finetuning causes a model to behave dangerously outside the finetuning task. Standard training signals c

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Transforming Police-Car Swerving for Mitigating Isolated Stop-and-Go Traffic Waves: A Practice-Oriented Jam-Absorption Driving Strategy

DGX agent

arXiv:2602.10234v3 Announce Type: replace-cross Abstract: Stop-and-go traffic waves, a major form of freeway congestion, impose severe and persistent adverse impacts, including reduced traffic efficie

safetyarxiv-cs-ai
9 Jun 2026
Research

Transition-Based Digital Twin Modelling for Alzheimer's Disease under Sparse Longitudinal Data

DGX agent

arXiv:2606.09671v1 Announce Type: cross Abstract: Alzheimer's disease (AD) progression is highly heterogeneous and is typically observed through sparse and irregular longitudinal data, posing challeng

researcharxiv-cs-ai
9 Jun 2026
Agents

Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing

DGX agent

arXiv:2606.08256v1 Announce Type: new Abstract: Verifiability, attribution, and reproducibility are foundational requirements of scientific knowledge, yet current publishing infrastructure does not en

agentsarxiv-cs-ai
9 Jun 2026
Research

TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs

DGX agent

arXiv:2606.09030v1 Announce Type: cross Abstract: Clinical early warning systems built on electronic health records, in which clinical observations are recorded as irregularly sampled medical time ser

researcharxiv-cs-ai
9 Jun 2026
Model Releases

TRL-Bench: Standardizing Cross-Paradigm Representation-Level Evaluation of Tabular Encoders

DGX agent

arXiv:2606.09323v1 Announce Type: new Abstract: Tabular encoders are usually evaluated inside task-specific end-to-end pipelines, so models from different training paradigms are difficult to compare d

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Trustworthy Smart Fabs via Professional Proxies: Scaling Safe and Sustainable by Design (SSbD) through Industrial Data Spaces

DGX agent

arXiv:2606.09227v1 Announce Type: cross Abstract: The convergence of the 2026 European Union Safe and Sustainable by Design (SSbD) framework, Corporate Sustainability Due Diligence Directive (CSDDD),

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

TT-DAC-PS: Twin-Target Deterministic Actor-Critic with Policy Smoothing for Optimal Trade Execution

DGX agent

arXiv:2606.08379v1 Announce Type: new Abstract: This study addresses the optimal execution of large stock sell programs by introducing TT-DAC-PS (Twin-Target Deterministic Actor-Critic with Policy Smo

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Tyan-WP: A Wind Power Foundation Model for Ultra-Short-Term Probabilistic Forecasting

DGX agent

arXiv:2606.08630v1 Announce Type: cross Abstract: Global wind power capacity, especially in China, is booming, with new farms spanning diverse terrains and climates. The industry urgently needs accura

researcharxiv-cs-ai
9 Jun 2026
Applications

UA-DCM: Uncertainty-aware Causal Decision Making via Effect Bound Decomposition

DGX agent

arXiv:2601.22736v2 Announce Type: replace-cross Abstract: Causal inference from observational data can provide strong evidence for finding the best action in a decision-making scenario without having

applicationsarxiv-cs-ai
9 Jun 2026
Research

Unambiguous Representations in Neural Networks: An Information-Theoretic Approach to Intentionality

DGX agent

arXiv:2512.11000v2 Announce Type: replace-cross Abstract: Representations pervade our daily experience, from letters representing sounds to bit strings encoding digital files. While such representatio

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Understanding Benchmark Language Under Weakened Formal Semantics

DGX agent

arXiv:2509.17455v2 Announce Type: replace-cross Abstract: State-of-the-art NLP benchmarks require interpretation of natural language that specifies conditions, procedures, and exceptions, often relyin

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Understanding Quantization-Aware Training: Gradients at Quantized Weights Bias to the Low-Loss Basin

DGX agent

arXiv:2606.09012v1 Announce Type: cross Abstract: Post-training quantization (PTQ) converts a trained full-precision model into low-bit weights without task-level retraining, while quantization-aware

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

Unification of Closed-Open Industrial Detection Scenarios: New Large-Scale Benchmarks,Challenges and Baselines

DGX agent

arXiv:2606.07953v1 Announce Type: new Abstract: Large-scale Visual-Language Models (LVLMs) have achieved remarkable success in natural visual tasks, yet their application to industrial defect detectio

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Unified Energy for Invariant and Independent Decoding in Diffusion Language Models

DGX agent

arXiv:2606.09159v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) enable parallel text generation by iteratively denoising a full sequence, offering attractive flexibility compared to

researcharxiv-cs-ai
9 Jun 2026
Safety

Unifying Object-Centric World Models and Diffusion Policy: A Hierarchical Framework for Multi-Stage Robotic Tasks

DGX agent

arXiv:2606.08775v1 Announce Type: cross Abstract: Visual world models have shown great potential in learning complex system dynamics. Recent advancements leverage these models as transition functions

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

UniQL: Towards Dialect-Universal Benchmarking for Text-to-SQL

DGX agent

arXiv:2606.08018v1 Announce Type: new Abstract: Existing text-to-SQL benchmarks are largely centered on SQLite, making it difficult to evaluate whether models can generalize across heterogeneous SQL d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Unsupervised Partner Design Enables Robust Ad-hoc Teamwork

DGX agent

arXiv:2508.06336v2 Announce Type: replace-cross Abstract: We introduce Unsupervised Partner Design (UPD), a population-free multi-agent reinforcement learning method for robust ad-hoc teamwork. UPD ge

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges

DGX agent

arXiv:2606.09125v1 Announce Type: cross Abstract: Privacy risks in text-only Large Language Models (LLMs) are well studied, particularly their tendency to memorize and leak sensitive information. Howe

researcharxiv-cs-ai
9 Jun 2026
Research

UnWeaving the knots of GraphRAG -- turns out VectorRAG is almost enough

DGX agent

arXiv:2603.29875v3 Announce Type: replace-cross Abstract: One of the key problems in Retrieval-augmented generation (RAG) systems is that chunk-based retrieval pipelines represent the source chunks as

researcharxiv-cs-ai
9 Jun 2026
Research

Variational Speculative Decoding: Rethinking Draft Training from Token Likelihood to Sequence Acceptance

DGX agent

arXiv:2602.05774v4 Announce Type: replace-cross Abstract: Speculative decoding accelerates inference for (M)LLMs, yet a training-decoding discrepancy persists: while existing methods optimize single g

researcharxiv-cs-ai
9 Jun 2026
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

DGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

safetyarxiv-cs-ai
9 Jun 2026
Research

VFEM: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion

DGX agent

arXiv:2510.03244v2 Announce Type: replace-cross Abstract: Large time series foundation models often adopt channel-independent architectures to handle varying data dimensions, but this design ignores c

researcharxiv-cs-ai
9 Jun 2026
Safety

Video Understanding by Design: How Datasets Shape Video Models

DGX agent

arXiv:2509.09151v2 Announce Type: replace-cross Abstract: Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While exi

safetyarxiv-cs-ai
9 Jun 2026
Agents

ViMax: Agentic Video Generation

DGX agent

arXiv:2606.07649v1 Announce Type: cross Abstract: Long-form video generation requires systematic narrative planning and visual consistency that current short-clip methods cannot provide. Existing meth

agentsarxiv-cs-ai
9 Jun 2026
Research

Vision-Based Early Fault Diagnosis and Self-Recovery for Strawberry Harvesting Robots

DGX agent

arXiv:2601.02085v3 Announce Type: replace-cross Abstract: Strawberry-harvesting robots faced challenges such as poor visual perception, gripper misalignment, empty grasp/misgrasp, and slippage, which

researcharxiv-cs-ai
9 Jun 2026
Tutorials

Vision Language Model Helps Private Information De-Identification in Vision Data

DGX agent

arXiv:2606.09132v1 Announce Type: new Abstract: Visual Language Models (VLMs) have gained significant popularity due to their remarkable ability. While various methods exist to enhance privacy in text

tutorialsarxiv-cs-ai
9 Jun 2026
Applications

Visual Prompting Meets Feature Reconstruction-Based Anomaly Detection with Dual-Teacher Supervision

DGX agent

arXiv:2606.09670v1 Announce Type: cross Abstract: Recent Anomaly Detection methods achieve perfect detection and segmentation scores on well-established datasets, such as MVTec. However, many of these

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

DGX agent

arXiv:2606.07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking extern

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Voting Protocols as Coordination Mechanisms for Role-Constrained Multi-Agent Tutoring Systems

DGX agent

arXiv:2606.08030v1 Announce Type: cross Abstract: Agentic tutoring systems introduce a coordination challenge: multiple agents may propose different but reasonable interventions, yet only one response

agentsarxiv-cs-ai
9 Jun 2026
Tutorials

Weak-Driven Learning: How Weak Agents make Strong Agents Stronger

DGX agent

arXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces

DGX agent

arXiv:2606.09426v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly operate in runtimes that combine visual desktop control, command-line execution, code editing, browsers, and ext

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Web Agents Should Use Typed Actions Instead of Click-Based Browsing

DGX agent

arXiv:2602.17245v2 Announce Type: replace Abstract: This position paper argues that building a reliable agentic Web requires shifting from low-level interaction primitives to typed actions supported b

safetyarxiv-cs-ai
9 Jun 2026
Safety

What Makes a Desired Graph for Relational Deep Learning?

DGX agent

arXiv:2606.08491v1 Announce Type: new Abstract: Relational deep learning (RDL) converts relational databases (RDBs) into heterogeneous graphs, but graphs derived directly from database schemas are oft

safetyarxiv-cs-ai
9 Jun 2026
Research

What Makes Video World Model Latents Action-Relevant: Prediction over Reconstruction

DGX agent

arXiv:2606.07687v1 Announce Type: cross Abstract: Video world models are increasingly used to provide predictive visual representations, yet it remains unclear which pretraining signals induce action-

researcharxiv-cs-ai
9 Jun 2026
Research

What's the Point? Spatial Grammar & Index Resolution for Sign Language Processing

DGX agent

arXiv:2606.08056v1 Announce Type: cross Abstract: Sign language models are predominantly trained with gloss-sequence or text supervision, thereby under-modeling non-lexical and productive construction

researcharxiv-cs-ai
9 Jun 2026
Model Releases

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective

DGX agent

arXiv:2606.08044v1 Announce Type: cross Abstract: Large Language Model (LLM) safety has often been evaluated at the behavior level, which provides limited evidence of internal robustness, as these eva

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

DGX agent

arXiv:2602.08235v2 Announce Type: replace-cross Abstract: Although computer-use agents (CUAs) hold significant potential to automate increasingly complex OS workflows, they can demonstrate unsafe unin

model-releasesarxiv-cs-ai
9 Jun 2026
Research

When Does Delegation Beat Majority? A Delegation-Based Aggregator for Multi-Sample LLM Inference

DGX agent

arXiv:2606.08098v1 Announce Type: new Abstract: Majority voting over sampled answers is the dominant unsupervised aggregator for multi-sample LLM inference. We show that piping the signals every sampl

researcharxiv-cs-ai
9 Jun 2026
Research

When No Answer Is Correct: Diagnosing Absent Answer Detection for MLLMs in Video Understanding

DGX agent

arXiv:2606.08239v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made substantial advancements in video understanding, yet the reliability of their responses remains under

researcharxiv-cs-ai
9 Jun 2026
← Previous
1…187188189190191…448
Next →