AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

UxSID: Semantic-Aware User Interests Modeling for Ultra-Long Sequence

ResearchDGX agent

arXiv:2605.09040v1 Announce Type: new Abstract: Modeling ultra-long user sequences involves a difficult trade-off between efficiency and effectiveness. While current paradigms rely on either item-spec

Value-Decomposed Reinforcement Learning Framework for Taxiway Routing with Hierarchical Conflict-Aware Observations

SafetyDGX agent

arXiv:2605.08754v1 Announce Type: new Abstract: Taxiway routing and on-surface conflict avoidance are coupled safety-critical decision problems in airport surface operations. Existing planning and opt

Variational Inference for Levy Process-Driven SDEs via Neural Tilting

SafetyDGX agent

arXiv:2605.10934v1 Announce Type: cross Abstract: Modelling extreme events and heavy-tailed phenomena is central to building reliable predictive systems in domains such as finance, climate science, an


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models

Model ReleasesDGX agent

arXiv:2603.18113v2 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly shape content generation, interaction, and decision-making across the Web, aligning them with hum

VECTOR-Drive: Tightly Coupled Vision-Language and Trajectory Expert Routing for End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2605.08830v1 Announce Type: cross Abstract: End-to-end autonomous driving requires models to understand traffic scenes, infer driving intent, and generate executable motion plans. Recent vision-

VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation

Model ReleasesDGX agent

arXiv:2605.08553v1 Announce Type: cross Abstract: Large language models can generate useful code from natural language, but their outputs come without correctness guarantees. Verifiable code generatio

Verifiable Process Rewards for Agentic Reasoning

Local AiDGX agent

arXiv:2605.10325v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has improved the reasoning abilities of large language models (LLMs), but most existing approaches

Verifier-Free RL for LLMs via Intrinsic Gradient-Norm Reward

SafetyDGX agent

arXiv:2605.09920v1 Announce Type: cross Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has recently emerged as a promising post-training paradigm for Large Language Models (LLMs

Virtual Personas for Language Models via an Anthology of Backstories

ResearchDGX agent

arXiv:2407.06576v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are trained from vast repositories of text authored by millions of distinct authors, reflecting an enormous diver

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models

AgentsDGX agent

arXiv:2605.10106v1 Announce Type: cross Abstract: Recent advances in Multi-modal Large Language Models (MLLMs) target 3D spatial intelligence, yet the progress has been largely driven by post-training

Visual-ERM: Reward Modeling for Visual Equivalence

Model ReleasesDGX agent

arXiv:2603.13224v2 Announce Type: replace-cross Abstract: Vision-to-code tasks require models to reconstruct structured visual inputs, such as charts, tables, and SVGs, into executable or structured r

VLADriver-RAG: Retrieval-Augmented Vision-Language-Action Models for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.08133v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for end-to-end autonomous driving, yet their reliance on implicit parametric

Voice Biomarkers for Depression and Anxiety

TutorialsDGX agent

arXiv:2605.09908v1 Announce Type: cross Abstract: Current approaches to detecting depression and anxiety from speech primarily rely on machine learning techniques that utilize hand-engineered paraling

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

Model ReleasesDGX agent

arXiv:2605.08146v1 Announce Type: cross Abstract: Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domai

VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection

TutorialsDGX agent

arXiv:2605.09461v1 Announce Type: new Abstract: Automated vulnerability detection is a fundamental task in software security, yet existing learning-based methods still struggle to capture the structur

WATCH: Wide-Area Archaeological Site Tracking for Change Detection

Model ReleasesDGX agent

arXiv:2605.08160v1 Announce Type: cross Abstract: Monitoring archaeological sites at scale is vital for protecting cultural heritage, yet pinpointing when disturbances occur remains difficult because

Watermarking Graph Neural Networks via Explanations for Ownership Protection

ResearchDGX agent

arXiv:2501.05614v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) are widely deployed in industry, making their intellectual property valuable. However, protecting GNNs from unaut

WavesFM: Hierarchical Representation Learning for Longitudinal Wearable Sensor Waveforms

Local AiDGX agent

arXiv:2605.09173v1 Announce Type: cross Abstract: Wearable sensors enable the continuous acquisition of high-resolution physiological waveforms, such as photoplethysmography and accelerometry, under f

Weakly Supervised Concept Learning for Object-centric Visual Reasoning

ApplicationsDGX agent

arXiv:2605.08201v1 Announce Type: cross Abstract: Neurosymbolic systems promise to combine deep neural network's (DNN) processing of raw sensor inputs with few-shot performance of symbolic artificial

WebTrap: Stealthy Mid-Task Hijacking of Browser Agents During Navigation

AgentsDGX agent

arXiv:2605.08310v1 Announce Type: cross Abstract: Browser agents are increasingly deployed in long-horizon tasks, which require executing extended action chains to accomplish user goals. However, this

Weight Pruning Amplifies Bias: A Multi-Method Study of Compressed LLMs for Edge AI

Model ReleasesDGX agent

arXiv:2605.08137v1 Announce Type: cross Abstract: Weight pruning is widely advocated for deploying Large Language Models on resource-constrained IoT and edge devices, yet its impact on model fairness

Weighted Rules under the Stable Model Semantics

ResearchDGX agent

arXiv:2605.09519v1 Announce Type: new Abstract: We introduce the concept of weighted rules under the stable model semantics following the log-linear models of Markov Logic. This provides versatile met

What Cohort INRs Encode and Where to Freeze Them

TutorialsDGX agent

arXiv:2605.08298v1 Announce Type: cross Abstract: Reusing the early layers of cohort-trained INRs as initialization for new signals has been shown to accelerate and improve signal fitting, yet it rema

What Does Flow Matching Bring To TD Learning?

ResearchDGX agent

arXiv:2603.04333v2 Announce Type: replace-cross Abstract: Recent work shows that flow matching can be effective for scalar Q-value function estimation in reinforcement learning (RL), but it remains un

What If We Let Forecasting Forget? A Sparse Bottleneck for Cross-Variable Dependencies

ApplicationsDGX agent

arXiv:2605.08289v1 Announce Type: cross Abstract: Multivariate time series forecasting is critical in many real-world systems, and thus modeling cross-channel dependencies is essential. Although exist

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

AgentsDGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

What Structural Inductive Bias Helps Transformers Reason Over Knowledge Graphs? A Study with Tabula RASA

SafetyDGX agent

arXiv:2602.02834v3 Announce Type: replace-cross Abstract: What structural inductive bias helps transformers reason over knowledge graphs? Through controlled ablations of a minimal transformer modifica

What Will Happen Next: Large Models-Driven Deduction for Emergency Instances

Model ReleasesDGX agent

arXiv:2605.08599v1 Announce Type: new Abstract: Traditional simulation methods reproduce occurred emergency instances through presetting to assist people in risk assessment and emergency decision-maki

What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering

Model ReleasesDGX agent

arXiv:2601.20164v2 Announce Type: replace-cross Abstract: Prior work suggests that language models, while trained on next token prediction, show implicit planning behavior: they may select the next to

When a Robot is More Capable than a Human: Learning from Constrained Demonstrators

SafetyDGX agent

arXiv:2510.09096v3 Announce Type: replace-cross Abstract: Learning from demonstrations enables experts to teach robots complex tasks using interfaces such as kinesthetic teaching, joystick control, an

When Agents Overtrust Environmental Evidence: An Extensible Agentic Framework for Benchmarking Evidence-Grounding Defects in LLM Agents

SafetyDGX agent

arXiv:2605.08828v1 Announce Type: new Abstract: Large language model agents increasingly operate through environment-facing scaffolds that expose files, web pages, APIs, and logs. These observations i

When Agents Say One Thing and Do Another: Validating Elicited Beliefs from LLMs

AgentsDGX agent

arXiv:2602.06286v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings where good decisions require forming beliefs over the probability of

When AI Meets Science: Research Diversity, Interdisciplinarity, Visibility, and Retractions across Disciplines in a Global Surge

ResearchDGX agent

arXiv:2605.06033v2 Announce Type: replace-cross Abstract: The extent to which Artificial Intelligence (AI) can trigger generalized paradigm shifts in science is unclear. Although some of these technol

When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.09109v1 Announce Type: new Abstract: Many continuous-control problems ship with a competent but suboptimal controller (a tuned PID, a hand-designed gait). A growing family of methods uses s

When Attention Beats Fourier: Multi-Scale Transformers for PDE Solving on Irregular Domains

Model ReleasesDGX agent

arXiv:2605.08318v1 Announce Type: cross Abstract: We study the problem of architecture selection for deep learning models trained to solve partial differential equations (PDEs), asking when transforme

When Can Digital Personas Reliably Approximate Human Survey Findings?

SafetyDGX agent

arXiv:2605.10659v1 Announce Type: cross Abstract: Digital personas powered by Large Language Models (LLMs) are increasingly proposed as substitutes for human survey respondents, yet it remains unclear

When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees

ResearchDGX agent

arXiv:2605.08710v1 Announce Type: new Abstract: Human-AI teams fail to outperform their best member in 70% of studies, yet no theory specifies when complementarity is achievable. We derive tight bound

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

AgentsDGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

When Does Non-Uniform Replay Matter in Reinforcement Learning?

Model ReleasesDGX agent

arXiv:2605.10236v1 Announce Type: cross Abstract: Modern off-policy reinforcement learning algorithms often rely on simple uniform replay sampling and it remains unclear when and why non-uniform repla

When Does Value-Aware KV Eviction Help? A Fixed-Contract Diagnostic for Non-Monotone Cache Compression

ResearchDGX agent

arXiv:2605.08234v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by the memory and bandwidth cost of reading large KV caches during decoding. KV compression reduces this co

When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation

ResearchDGX agent

arXiv:2605.09460v1 Announce Type: cross Abstract: Identity-preserved image generation is typically built on many-step diffusion backbones, making personalized generation expensive at deployment time.

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

SafetyDGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

When Normality Shifts: Risk-Aware Test-Time Adaptation for Unsupervised Tabular Anomaly Detection

TutorialsDGX agent

arXiv:2605.10242v1 Announce Type: cross Abstract: Unsupervised tabular anomaly detection methods typically learn feature patterns from normal samples during training and subsequently identify samples

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

Model ReleasesDGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

Model ReleasesDGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

When Tables Leak: Attacking String Memorization in LLM-Based Tabular Data Generation

ResearchDGX agent

arXiv:2512.08875v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently demonstrated remarkable performance in generating high-quality tabular synthetic data. In practice,

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

Model ReleasesDGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

When to Trust Imagination: Adaptive Action Execution for World Action Models

Model ReleasesDGX agent

arXiv:2605.06222v2 Announce Type: replace-cross Abstract: World Action Models (WAMs) have recently emerged as a promising paradigm for robotic manipulation by jointly predicting future visual observat

Where Do Flow Semantics Reside? A Protocol-Native Tabular Pretraining Paradigm for Encrypted Traffic Classification

SafetyDGX agent

arXiv:2603.10051v2 Announce Type: replace-cross Abstract: Self-supervised masked modeling shows promise for encrypted traffic classification by masking and reconstructing raw bytes. Yet recent work re

Where Do Reasoning Models Refuse?

ResearchDGX agent

arXiv:2507.03167v3 Announce Type: replace-cross Abstract: Chat models without chain-of-thought (CoT) reasoning must decide whether to refuse a harmful request before generating their first response to

Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits

ResearchDGX agent

arXiv:2605.08200v1 Announce Type: new Abstract: A pervasive intuition holds that vision-language models (VLMs) are most trustworthy when their attention maps look sharp: concentrated attention on the

Why Adam Works Better with eta_1 = eta_2: The Missing Gradient Scale Invariance Principle

SafetyDGX agent

arXiv:2601.21739v2 Announce Type: replace-cross Abstract: Adam has been at the core of large-scale training for almost a decade, yet a simple empirical fact remains unaccounted for: both validation sc

Why Do Aligned LLMs Remain Jailbreakable: Refusal-Escape Directions, Operator-Level Sources, and Safety-Utility Trade-off

Local AiDGX agent

arXiv:2605.08878v1 Announce Type: cross Abstract: Aligned large language models (LLMs) remain vulnerable to jailbreak attacks. Recent mechanistic studies have identified latent features and representa

Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space

SafetyDGX agent

arXiv:2605.08250v1 Announce Type: cross Abstract: Recent advances in diffusion transformers (DiTs) have enabled promising single-turn image editing capabilities. However, multi-turn editing often lead

Why Low-Resource NLP Needs More Than Cross-Lingual Transfer: Lessons Learned from Luxembourgish

ResearchDGX agent

arXiv:2605.10714v1 Announce Type: cross Abstract: Cross-lingual transfer has become a central paradigm for extending natural language processing (NLP) technologies to low-resource languages. By levera

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

Model ReleasesDGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

Willful Disobedience: Automatically Detecting Failures in Agentic Traces

AgentsDGX agent

arXiv:2603.23806v2 Announce Type: replace-cross Abstract: AI agents are increasingly embedded in real software systems, where they execute multi-step workflows through multi-turn dialogue, tool invoca

WindINR: Latent-State INR for Fast Local Wind Query and Correction in Complex Terrain

Model ReleasesDGX agent

arXiv:2605.09511v1 Announce Type: new Abstract: Many downstream decisions in complex terrain require fast wind estimates at a small number of user-specified locations and heights for a given forecast

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

SafetyDGX agent

arXiv:2605.09765v1 Announce Type: cross Abstract: Representation learning in electronic health records (EHR) has largely followed paradigms inherited from natural language processing, relying on seque

Workspace Optimization: How to Train Your Agent

AgentsDGX agent

arXiv:2605.09650v1 Announce Type: new Abstract: Modern agents built on frontier language models often cannot adapt their weights. What, then, remains trainable? We argue it is the agent's workspace, t

← Previous
1…276277278279280…358
Next →