AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

Text as Partial Constraint: Core-Residual Alignment for Robust Vision-Language Learning

SafetyDGX agent

arXiv:2607.03143v1 Announce Type: cross Abstract: Vision-language alignment powers open-vocabulary recognition, retrieval, and LVLM grounding, yet natural captions are often underspecified, making sim

Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation

SafetyDGX agent

arXiv:2606.22726v2 Announce Type: replace Abstract: Choreographic motion generation poses unique challenges for AI, demanding precise semantic control over complex, temporally structured, and expressi

The agent creates, we validate: A Lightweight Framework for Agentic Artifact Generation

SafetyDGX agent

arXiv:2607.02615v1 Announce Type: cross Abstract: Generating structured artifacts with Large Language Models - e.g. database queries, threat framework mappings, entity schemas - is relatively straight


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Anatomy of Uncertainty in LLMs

ResearchDGX agent

arXiv:2603.24967v2 Announce Type: replace Abstract: Understanding why a large language model (LLM) is uncertain about the response is important for their reliable deployment. Current approaches, which

The Changing Role of Symbolic Methods in Artificial Intelligence

AgentsDGX agent

arXiv:2607.05168v1 Announce Type: new Abstract: Why do intelligent systems need to perform explicit symbolic reasoning? Computer science has traditionally regarded symbolic reasoning as a defining com

The Foreign Policy AI Evaluation Gap

SafetyDGX agent

arXiv:2607.02955v1 Announce Type: cross Abstract: We argue that AI systems used in conducting foreign policy tasks - broadly enacting 'statecraft' - should be a priority test case for technical AI gov

The Hidden Water Geography of U.S. Hyperscale Data Centers in the AI Era

ResearchDGX agent

arXiv:2607.02531v1 Announce Type: cross Abstract: Water use by data centers is routinely reported as a single footprint, but water is consumed through two physically distinct pathways: at the site for

The 'I Don't Know' Filter: Enhancing Agentic Reliability in Function Calling

AgentsDGX agent

arXiv:2607.04034v1 Announce Type: cross Abstract: The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the traini

The Language of Bargaining: Linguistic Effects in LLM Negotiations

AgentsDGX agent

arXiv:2601.04387v2 Announce Type: replace Abstract: Negotiation is a core component of social intelligence, requiring agents to balance strategic reasoning, cooperation, and social norms. Recent work

The Map Behind the Flow: Finite-Step Gradient Descent as a Dynamical System

Model ReleasesDGX agent

arXiv:2607.04993v1 Announce Type: cross Abstract: Many phenomena of deep learning are dynamical: they concern not only which minima exist, but how gradient descent reaches, avoids, or selects among th

The Remarkable Effectiveness of Providing AI Agents with Natural Language Tools: A Replication Study Validating NLT Performance Across 14 Models

Model ReleasesDGX agent

arXiv:2607.03953v1 Announce Type: cross Abstract: This study independently replicates and extends the Natural Language Tools (NLT) framework of Johnson et al.~(2025), which questions the use of struct

The Role of Prompt Language and Translation-Theory-Driven Prompts in Large Language Models: A Case Study on Spanish-Chinese Journalistic Translation

Model ReleasesDGX agent

arXiv:2607.03160v1 Announce Type: cross Abstract: This study examines how prompt language and translation theory-driven prompt design influence the quality of Spanish-Chinese journalistic translations

The Role of Rigor in Artificial Intelligence

ResearchDGX agent

arXiv:2607.03634v1 Announce Type: new Abstract: Artificial intelligence (AI) has achieved extraordinary capabilities despite lacking many of the conceptual and scientific foundations associated with m

The S-ICDF Dataset: Sionna-Simulated Dynamic Interference Characterization and Direction Finding

HardwareDGX agent

arXiv:2607.03411v1 Announce Type: cross Abstract: Jamming and spoofing threaten wireless and satellite navigation by disrupting or manipulating radio frequency (RF) signals, undermining availability,

The Three Regimes of Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2510.01460v4 Announce Type: replace-cross Abstract: Offline-to-online reinforcement learning (RL) has emerged as a practical paradigm that leverages offline datasets for pretraining and online i

They Infer What You Meant: Models Represent Communicative Intent More Reliably Than They Act On It

ResearchDGX agent

arXiv:2607.03598v1 Announce Type: cross Abstract: When a person shares something with a language model, the model often answers the surface of the message rather than what the sender was doing by send

Three-Phase Evaluation of AI-Assisted Software Development Life Cycle

ResearchDGX agent

arXiv:2607.05125v1 Announce Type: cross Abstract: This paper presents an exploratory evaluation of how increasing levels of AI autonomy affect software development productivity, requirement adherence,

TIER: Trajectory-Invariant Explanation Regularization for Membership Privacy

ResearchDGX agent

arXiv:2607.02903v1 Announce Type: cross Abstract: Explainability is central to building trustworthy AI, yet explanation interfaces can inadvertently provide adversaries with an expanded privacy-relate

TokAN: Accent Normalization Using Self-Supervised Speech Tokens

SafetyDGX agent

arXiv:2607.03928v1 Announce Type: cross Abstract: Accent normalization (AN) seeks to convert non-native (L2) accented speech into standard (L1) speech while preserving speaker identity. The current te

Token-Based Affordance Grounding with Large Vision-Language Models

ResearchDGX agent

arXiv:2607.03595v1 Announce Type: cross Abstract: Affordance grounding aims to localize image regions that support a specific action, serving as a core capability for physical intelligence and embodie

ToolFailBench: Diagnosing Tool-Use Failures in LLM Agents

Model ReleasesDGX agent

arXiv:2607.04686v1 Announce Type: cross Abstract: Tool calling is central to modern language model agents, but aggregate benchmark scores often hide where tool use fails. A model that never calls a ne

Topological Shape Representation for Aneurysm -- Bifurcation Detection

ResearchDGX agent

arXiv:2607.05317v1 Announce Type: cross Abstract: Automated detection of intracranial aneurysms (IAs) from CT angiography (CTA) is severely hindered by high false-positive rates. Convolutional neural

TORINO: Token Reduction via Interpretable Concept Overlap in Vision-Language Models

ResearchDGX agent

arXiv:2607.04593v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated impressive capabilities across different tasks, but their computational cost is dominated by the large

Toward Efficient Agents: Memory, Tool learning, and Planning

Model ReleasesDGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

Toward Trustworthy Large Language Model Agents in Healthcare

Model ReleasesDGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

Towards Diverse and Comprehensive Benchmarks for Mutual Information Estimation

ApplicationsDGX agent

arXiv:2607.03487v1 Announce Type: cross Abstract: Mutual information (MI) estimation is a central problem in machine learning and statistics; however, existing benchmarks typically evaluate estimators

Towards Mitigation of Hallucination for LLM-empowered Agents: Progressive Generalization Bound Exploration and Watchdog Monitor

AgentsDGX agent

arXiv:2507.15903v2 Announce Type: replace-cross Abstract: Empowered by large language models (LLMs), intelligent agents have become a popular paradigm for interacting with open environments to facilit

Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation

Model ReleasesDGX agent

arXiv:2603.17673v2 Announce Type: replace-cross Abstract: LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to repro

Towards Understanding Deep Learning Model in Image Recognition via Coverage Test

ResearchDGX agent

arXiv:2505.08814v3 Announce Type: replace-cross Abstract: Deep neural networks (DNNs) play a crucial role in the field of artificial intelligence, and their security-related testing has been a promine

TRACE: Capability-Targeted Agentic Training

Model ReleasesDGX agent

arXiv:2604.05336v2 Announce Type: replace Abstract: Models often fail to complete agentic tasks because they lack core capabilities required by the target environment. However, mainstream approaches f

Tracing 3D Anatomy in 2D Strokes: A Multi-Stage Projection Driven Approach to Cervical Spine Fracture Identification

ResearchDGX agent

arXiv:2601.15235v4 Announce Type: replace-cross Abstract: Cervical spine fractures require rapid and accurate diagnosis, yet automatic CT interpretation remains challenging as subtle injuries must be

Training-Free Model Selection and Domain-Aware Score Calibration for First-Shot Anomalous Sound Detection

Model ReleasesDGX agent

arXiv:2607.04526v1 Announce Type: cross Abstract: First-shot anomalous sound detection in DCASE Challenge Task 2 must flag anomalies of unseen machine types with a single threshold, without knowing wh

Training Hybrid Block Diffusion Language Models with Partial Bidirectionality

Model ReleasesDGX agent

arXiv:2607.02805v1 Announce Type: cross Abstract: High-throughput long-context generation is one of the central challenges for large language models. Generation is typically memory-bandwidth-bound rat

Transferability Between Understanding and Generation in Unified Multimodal Models

ResearchDGX agent

arXiv:2607.04423v1 Announce Type: cross Abstract: Unified Multimodal Models (UMMs) integrate image understanding and generation within a single architecture, yet how the two tasks interact remains und

Transition Information Density: Morphological Trajectories, Synesthetic Perception, and Structured Interpolation in Neural Training (or: The Synesthetic AI)

ResearchDGX agent

arXiv:2607.03210v1 Announce Type: cross Abstract: Standard machine learning training presents data as discrete endpoint pairs, omitting the structure of the space between them. This paper introduces T

TransitNet: A Compact Attention-Augmented Deep Learning Framework for Low-SNR Transit Blind Searches

ResearchDGX agent

arXiv:2606.18932v2 Announce Type: replace-cross Abstract: Motivated by the observational incompleteness of intermediate-to-long-period Earth-size planets, we present TransitNet, a compact attention-au

Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5

ApplicationsDGX agent

arXiv:2607.04510v1 Announce Type: cross Abstract: Emergent misalignment (EM) -- the broad misbehaviour a language model acquires after fine-tuning on narrow harmful data -- is mediated in Qwen2.5 mode

TREK: Distill to Explore, Reinforce to Refine

Model ReleasesDGX agent

arXiv:2607.05339v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) is effective when the current policy already samples useful reasoning trajectories, but it stalls on hard pr

TRIAGE: Trustworthy Retrieval Instrumentation And Graph Evaluation

ResearchDGX agent

arXiv:2607.03447v1 Announce Type: cross Abstract: Knowledge graphs (KGs) that underpin Graph-based Retrieval-Augmented Generation (Graph-RAG) are increasingly built automatically by LLM-driven extract

Trust-Region Noise Search for Black-Box Alignment of Diffusion and Flow Models

Local AiDGX agent

arXiv:2603.14504v2 Announce Type: replace-cross Abstract: Optimizing the noise samples of diffusion and flow models is an increasingly popular approach to align these models to target rewards at infer

Trust Region Policy Distillation

SafetyDGX agent

arXiv:2607.04751v1 Announce Type: cross Abstract: Big goals are hard to achieve all at once; breaking them into small steps is wiser. We present Trust Region Policy Distillation (TOP-D), which transfo

Turbo-Muon: Almost-Orthogonal Pre-Conditioning for Fast Muon Updates

ResearchDGX agent

arXiv:2512.04632v2 Announce Type: replace Abstract: Orthogonality-based optimizers, such as Muon, have recently shown strong performance across large-scale training and community-driven efficiency cha

Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment

SafetyDGX agent

arXiv:2607.04728v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of 'rollout then update', which inevitably res

Two Black Boxes, One Solver: Encoder Probing and Decoder Attribution for Neural Multi-Attribute VRP under Hard-Mask and Recourse Decoders

SafetyDGX agent

arXiv:2607.04487v1 Announce Type: cross Abstract: Neural autoregressive solvers for the Multi-Attribute Vehicle Routing Problem (MAVRP) reach competitive cost but offer no per-step justification, a pr

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

SafetyDGX agent

arXiv:2607.04425v1 Announce Type: cross Abstract: Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task execution toward cross-platform int

Unbiased Alignment for Large Language Models with Noisy Preferences

Model ReleasesDGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

UNDREAM: Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks

SafetyDGX agent

arXiv:2510.16923v3 Announce Type: replace-cross Abstract: Deep learning models deployed in safety critical applications like autonomous driving use simulations to test their robustness against adversa

Unified Audio Intelligence Without Regressing on Text Intelligence

Model ReleasesDGX agent

arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A

Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching

Model ReleasesDGX agent

arXiv:2603.27044v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is widely recognized as sample-inefficient, a limitation attributable in part to the high dimensionality and

Unsupervised Features Mining via Activation Geometry

ResearchDGX agent

arXiv:2607.04222v1 Announce Type: new Abstract: Interpretability methods aim to reveal the features represented inside large language models (LLMs). Many existing methods begin with labeled examples o

Unveiling the Unborn: Advancing Fetal Health Classification through Machine Learning

ApplicationsDGX agent

arXiv:2310.00505v3 Announce Type: replace-cross Abstract: Fetal health classification is a critical task in obstetrics, enabling early identification and management of potential health problems. Howev

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

Model ReleasesDGX agent

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Re

Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models

ResearchDGX agent

arXiv:2503.06269v3 Announce Type: replace-cross Abstract: Traditional white-box methods for creating adversarial perturbations against LLMs typically rely only on gradient computation from the targete

Verifier-free Test-Time Sampling for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2510.05681v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) have demonstrated remarkable performance in robot control. However, they remain fundamentally limited in

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

Model ReleasesDGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

ResearchDGX agent

arXiv:2607.03657v1 Announce Type: cross Abstract: Gloss-free Sign Language Translation (SLT) translates sign language videos into spoken-language sentences without gloss annotations, avoiding costly l

Vision Token Manipulation Attacks on Cloud-Edge Inference of Large Vision-Language Models

ResearchDGX agent

arXiv:2607.02819v1 Announce Type: cross Abstract: Cloud-edge Large Vision-Language Model (LVLM) inference enables efficient deployment by splitting computation between edge devices and cloud servers.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

SafetyDGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models

SafetyDGX agent

arXiv:2508.08521v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to

← Previous
1…9293949596…358
Next →