AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Research

Variational Speculative Decoding: Rethinking Draft Training from Token Likelihood to Sequence Acceptance

DGX agent

arXiv:2602.05774v4 Announce Type: replace-cross Abstract: Speculative decoding accelerates inference for (M)LLMs, yet a training-decoding discrepancy persists: while existing methods optimize single g

researcharxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

DGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

safetyarxiv-cs-ai
9 Jun 2026
Research

VFEM: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion

DGX agent

arXiv:2510.03244v2 Announce Type: replace-cross Abstract: Large time series foundation models often adopt channel-independent architectures to handle varying data dimensions, but this design ignores c

researcharxiv-cs-ai
9 Jun 2026
Safety

Video Understanding by Design: How Datasets Shape Video Models

DGX agent

arXiv:2509.09151v2 Announce Type: replace-cross Abstract: Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While exi

safetyarxiv-cs-ai
9 Jun 2026
Agents

ViMax: Agentic Video Generation

DGX agent

arXiv:2606.07649v1 Announce Type: cross Abstract: Long-form video generation requires systematic narrative planning and visual consistency that current short-clip methods cannot provide. Existing meth

agentsarxiv-cs-ai
9 Jun 2026
Research

Vision-Based Early Fault Diagnosis and Self-Recovery for Strawberry Harvesting Robots

DGX agent

arXiv:2601.02085v3 Announce Type: replace-cross Abstract: Strawberry-harvesting robots faced challenges such as poor visual perception, gripper misalignment, empty grasp/misgrasp, and slippage, which

researcharxiv-cs-ai
9 Jun 2026
Tutorials

Vision Language Model Helps Private Information De-Identification in Vision Data

DGX agent

arXiv:2606.09132v1 Announce Type: new Abstract: Visual Language Models (VLMs) have gained significant popularity due to their remarkable ability. While various methods exist to enhance privacy in text

tutorialsarxiv-cs-ai
9 Jun 2026
Applications

Visual Prompting Meets Feature Reconstruction-Based Anomaly Detection with Dual-Teacher Supervision

DGX agent

arXiv:2606.09670v1 Announce Type: cross Abstract: Recent Anomaly Detection methods achieve perfect detection and segmentation scores on well-established datasets, such as MVTec. However, many of these

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

DGX agent

arXiv:2606.07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking extern

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Voting Protocols as Coordination Mechanisms for Role-Constrained Multi-Agent Tutoring Systems

DGX agent

arXiv:2606.08030v1 Announce Type: cross Abstract: Agentic tutoring systems introduce a coordination challenge: multiple agents may propose different but reasonable interventions, yet only one response

agentsarxiv-cs-ai
9 Jun 2026
Tutorials

Weak-Driven Learning: How Weak Agents make Strong Agents Stronger

DGX agent

arXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces

DGX agent

arXiv:2606.09426v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly operate in runtimes that combine visual desktop control, command-line execution, code editing, browsers, and ext

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Web Agents Should Use Typed Actions Instead of Click-Based Browsing

DGX agent

arXiv:2602.17245v2 Announce Type: replace Abstract: This position paper argues that building a reliable agentic Web requires shifting from low-level interaction primitives to typed actions supported b

safetyarxiv-cs-ai
9 Jun 2026
Safety

What Makes a Desired Graph for Relational Deep Learning?

DGX agent

arXiv:2606.08491v1 Announce Type: new Abstract: Relational deep learning (RDL) converts relational databases (RDBs) into heterogeneous graphs, but graphs derived directly from database schemas are oft

safetyarxiv-cs-ai
9 Jun 2026
Research

What Makes Video World Model Latents Action-Relevant: Prediction over Reconstruction

DGX agent

arXiv:2606.07687v1 Announce Type: cross Abstract: Video world models are increasingly used to provide predictive visual representations, yet it remains unclear which pretraining signals induce action-

researcharxiv-cs-ai
9 Jun 2026
Research

What's the Point? Spatial Grammar & Index Resolution for Sign Language Processing

DGX agent

arXiv:2606.08056v1 Announce Type: cross Abstract: Sign language models are predominantly trained with gloss-sequence or text supervision, thereby under-modeling non-lexical and productive construction

researcharxiv-cs-ai
9 Jun 2026
Model Releases

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective

DGX agent

arXiv:2606.08044v1 Announce Type: cross Abstract: Large Language Model (LLM) safety has often been evaluated at the behavior level, which provides limited evidence of internal robustness, as these eva

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

DGX agent

arXiv:2602.08235v2 Announce Type: replace-cross Abstract: Although computer-use agents (CUAs) hold significant potential to automate increasingly complex OS workflows, they can demonstrate unsafe unin

model-releasesarxiv-cs-ai
9 Jun 2026
Research

When Does Delegation Beat Majority? A Delegation-Based Aggregator for Multi-Sample LLM Inference

DGX agent

arXiv:2606.08098v1 Announce Type: new Abstract: Majority voting over sampled answers is the dominant unsupervised aggregator for multi-sample LLM inference. We show that piping the signals every sampl

researcharxiv-cs-ai
9 Jun 2026
Research

When No Answer Is Correct: Diagnosing Absent Answer Detection for MLLMs in Video Understanding

DGX agent

arXiv:2606.08239v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made substantial advancements in video understanding, yet the reliability of their responses remains under

researcharxiv-cs-ai
9 Jun 2026
Agents

When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA

DGX agent

arXiv:2606.08542v1 Announce Type: cross Abstract: Exploratory manipulation often turns an apparent failed attempt into the key evidence for what to do next. For example, a robot pulls a locked cabinet

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing

DGX agent

arXiv:2606.07710v1 Announce Type: cross Abstract: The autoregressive nature of large language models (LLMs) remains a significant bottleneck for inference, particularly in complex agentic workloads. W

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Who Earns the Safety? Intervention-Aware Quantum Predictive Control with Safety Attribution

DGX agent

arXiv:2606.09778v1 Announce Type: cross Abstract: Hard safety filters are increasingly placed downstream of learned controllers to guarantee constraint satisfaction at run time. Yet a filtered control

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Why Limit the Residual Stream to Layers and Not Tokens? Persistent Memory for Continuous Latent Reasoning

DGX agent

arXiv:2606.07720v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable reasoning abilities on mathematical and multi-hop planning tasks. The CoCoNuT (Chain of Contin

researcharxiv-cs-ai
9 Jun 2026
Tutorials

XAInomaly: Explainable and Interpretable Deep Contractive Autoencoder for O-RAN Traffic Anomaly Detection

DGX agent

arXiv:2502.09194v1 Announce Type: cross Abstract: Generative Artificial Intelligence (AI) techniques have become integral part in advancing next generation wireless communication systems by enabling s

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

XCR-Bench: Benchmarking Cross-Cultural Reasoning in LLMs via Culture-Specific Items and Hall's Triad

DGX agent

arXiv:2601.14063v2 Announce Type: replace-cross Abstract: Cross-cultural competence in large language models (LLMs) requires understanding and adapting Culture-Specific Items (CSIs) across varying cul

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

DGX agent

arXiv:2606.07965v1 Announce Type: new Abstract: Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ZIPP:Zero-shot Image Personalization from Personas

DGX agent

arXiv:2606.08841v1 Announce Type: new Abstract: Text-to-image diffusion models are increasingly deployed in open-ended creative contexts, yet their outputs remain impersonal, optimized for aggregate a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

DGX agent

arXiv:2606.07410v1 Announce Type: cross Abstract: The emergence of 'Aha moments' in large language models, particularly DeepSeek-R1-0120, has raised the question of whether these systems genuinely rea

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

A Geometric Account of Activation Steering through Angle-Norm Decomposition

DGX agent

arXiv:2606.06735v1 Announce Type: new Abstract: Linear activation steering has gained popularity as a simple and empirically effective way to control language model behavior. More recently, spherical

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

A Geometric Gaussian Mixture Representation of Plane Curves

DGX agent

arXiv:2606.06505v1 Announce Type: cross Abstract: We introduce a user defined probabilistic polygonal representation for plane curves. Given a curve, we select vertices on the curve and connect consec

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders

DGX agent

arXiv:2606.07007v1 Announce Type: cross Abstract: We propose a unified mathematical framework for a geometric understanding of concept learning and neuron interpretation in sparse autoencoders (SAEs).

safetyarxiv-cs-ai
8 Jun 2026
Tutorials

A Mechanism-Coupled Split Window Network for Medium- to High-Resolution Land Surface Temperature Retrieval

DGX agent

arXiv:2509.04991v2 Announce Type: replace-cross Abstract: Land surface temperature (LST) is a fundamental physical variable in land-atmosphere interactions, surface energy budgets, and climate process

tutorialsarxiv-cs-ai
8 Jun 2026
Tutorials

A robust PPG foundation model using multimodal physiological supervision

DGX agent

arXiv:2606.07365v1 Announce Type: cross Abstract: Photoplethysmography (PPG), a non-invasive measure of changes in blood volume, is widely used in both wearable devices and clinical settings. Recent P

tutorialsarxiv-cs-ai
8 Jun 2026
Research

A Study of Parallel Continuous Local Search

DGX agent

arXiv:2606.06656v1 Announce Type: new Abstract: We study parallel Continuous Local Search (CLS) as a solution approach for Boolean satisfiability problems with symmetric pseudo-Boolean (PB) constraint

researcharxiv-cs-ai
8 Jun 2026
Local Ai

A Temporal Spatial Minimax Rate for Smoothly-Varying Distributions in Wasserstein Space

DGX agent

arXiv:2606.07325v1 Announce Type: cross Abstract: We study the minimax rate of estimating a future value mu_{t_n+h} of a curve tmapstomu_t in the 2-Wasserstein space P_2(R^d) from finitely many noisy

local-aiarxiv-cs-ai
8 Jun 2026
Hardware

Accelerated Fourier SAT (AFSAT): Fully Realising a GPU-based Symmetric Pseudo-Boolean SAT Solver

DGX agent

arXiv:2606.06641v1 Announce Type: new Abstract: We present Accelerated Fourier SAT (AFSAT), a GPU-accelerated solver for pseudo-Boolean satisfiability based on continuous local search (CLS). AFSAT rea

hardwarearxiv-cs-ai
8 Jun 2026
Safety

Accounting for Context: Shaping Moral Credences for Value Alignment

DGX agent

arXiv:2606.06972v1 Announce Type: new Abstract: Ensuring that agent behaviours are aligned with human moral values inevitably raises the problem of how to account for the plurality of moral perspectiv

safetyarxiv-cs-ai
8 Jun 2026
Safety

Acoustic Cue Alignment in Audio Language Models for Speech Emotion Recognition

DGX agent

arXiv:2606.07309v1 Announce Type: cross Abstract: Instruction-following audio language models (ALMs) can be augmented with explicit acoustic cues, yet it remains unclear whether such cues are used in

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

DGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

AdMem: Advanced Memory for Task-solving Agents

DGX agent

arXiv:2606.06787v1 Announce Type: new Abstract: Large Language Models (LLMs) show promise as tool-using agents but remain limited in long-horizon tasks that require remembering, organizing, and reusin

agentsarxiv-cs-ai
8 Jun 2026
Safety

AEGIS: A Backup Reflex for Physical AI

DGX agent

arXiv:2606.06660v1 Announce Type: new Abstract: Long-horizon robot manipulation tends to fail gradually: one bad step degrades the state, and the policy spirals into a basin from which it cannot recov

safetyarxiv-cs-ai
8 Jun 2026
Agents

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

DGX agent

arXiv:2606.06525v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text genera

agentsarxiv-cs-ai
8 Jun 2026
Research

AI-Driven Test Case Generation from Natural Language Requirements: A Survey of Techniques and Research Gaps

DGX agent

arXiv:2606.06563v1 Announce Type: cross Abstract: Software testing is critical for verifying that systems meet specified requirements, yet remains among the most time-consuming and expensive activitie

researcharxiv-cs-ai
8 Jun 2026
Research

AI Sovereignty: A Qualitative Model of Strategic Competition as AI Becomes an Instrument of National Power

DGX agent

arXiv:2606.07245v1 Announce Type: cross Abstract: AI sovereignty is the extent to which a nation independently controls its artificial intelligence (AI) technologies. The race toward ever-more-sophist

researcharxiv-cs-ai
8 Jun 2026
← Previous
1…192193194195196…452
Next →