AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
26 Jun 2026

Accelerating Skill Assessment in Chess: A Drift-Diffusion-Enhanced Elo Rating System

SafetyDGX agent

arXiv:2606.26267v1 Announce Type: new Abstract: Rating systems such as Elo serve as the gold standard for matchmaking in competitive chess. However, they inherently suffer from response lag due to the

Active Adversarial Perturbation-driven Associative Memory Retrieval for RGB-Event Visual Object Tracking

Model ReleasesDGX agent

arXiv:2606.26455v1 Announce Type: cross Abstract: RGB-Event tracking improves localization robustness by fusing RGB appearance textures and dense temporal motion cues from event sensors. While this mu

Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.26479v1 Announce Type: cross Abstract: Recent work (2024 to 2026) has converged on a strategy for defending tool-using LLM agents against indirect prompt injection: rather than training the

Adaptive Utility driven Resource Orchestration for Resilient AI (AURORA-AI)

SafetyDGX agent

arXiv:2606.27005v1 Announce Type: new Abstract: Modern AI systems are increasingly deployed under non-stationary computational, demographic, and operational conditions in which static resource allocat

Advancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical Autonomy

AgentsDGX agent

arXiv:2606.27251v1 Announce Type: cross Abstract: Building persistent embodied agents in unstructured environments demands unified orchestration of heterogeneous tools spanning both cyber (APIs, IoT)

Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols

SafetyDGX agent

arXiv:2606.26203v1 Announce Type: new Abstract: As AI agent protocols proliferate, the governance structures shaping their interoperability standards remain empirically underexamined. We introduce an

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents

Model ReleasesDGX agent

arXiv:2606.26627v1 Announce Type: cross Abstract: Large language model agents increasingly query databases, search document collections, call external APIs, remember past interactions, and act on a us

AgentX: Towards Agent-Driven Self-Iteration of Industrial Recommender Systems

SafetyDGX agent

arXiv:2606.26859v1 Announce Type: new Abstract: Recommendation algorithm iteration is moving from an artisanal, engineer-bound process toward an industrialized research loop, but this transition remai

AI Healthcare Chatbots as Information Infrastructure: A Large-Scale Study of User-Reported Breakdowns

ApplicationsDGX agent

arXiv:2606.27302v1 Announce Type: cross Abstract: AI healthcare chatbots are increasingly used to support health information seeking and self-management, yet their performance and impact on users rema

AIGP: An LLM-Based Framework for Long-Term Value Alignment in E-Commerce Pricing

SafetyDGX agent

arXiv:2606.26787v1 Announce Type: cross Abstract: Traditional dynamic pricing models in large-scale e-commerce suffer from limited interpretability, poor utilization of unstructured information, and m

AlgoEvolve: LLM-driven Meta-evolution of Algorithmic Trading Programs

AgentsDGX agent

arXiv:2606.26173v1 Announce Type: new Abstract: Recent work shows that Large Language Models (LLMs) can act as semantic mutation operators for the evolutionary discovery of programs and proofs. Most c

Algorithmic Foundations of Deep Learning: Complexity-Theoretic Rates and a Characterization of Universal Approximation

Model ReleasesDGX agent

arXiv:2606.26705v1 Announce Type: cross Abstract: Feedforward neural network (NN) expressivity is typically studied by emulating optimal basis-expansion schemes. While powerful, this perspective is in

An Empirical Study of LLM-Generated Specifications for VeriFast

Model ReleasesDGX agent

arXiv:2606.26490v1 Announce Type: cross Abstract: Static verification tools can assure industrial scale software, but require significant human labor to write specifications. This is particularly true

Anatomy-Guided Residual Motion Diffusion for Controllable 4D Cardiac MRI Synthesis

ResearchDGX agent

arXiv:2606.26764v1 Announce Type: cross Abstract: Developing robust artificial intelligence models for 4D (3D + time) medical imaging is constrained by limited annotated data, inter-device domain shif

Application of LLMs to Threat Assessment of Foreign Peacekeeping Missions

ResearchDGX agent

arXiv:2606.27106v1 Announce Type: cross Abstract: We present a novel approach for applying Large Language Models (LLMs) to threat assessment in the context of foreign peacekeeping missions. Building o

Ask, Don't Judge: Binary Questions for Interpretable LLM Evaluation and Self-Improvement

ResearchDGX agent

arXiv:2606.27226v1 Announce Type: new Abstract: Evaluating LLM outputs remains a major bottleneck in NLP: human evaluation is expensive and slow, lexical metrics correlate poorly with human judgments

Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfare

Model ReleasesDGX agent

arXiv:2606.26104v1 Announce Type: cross Abstract: Animal-welfare advocates produce a lot of writing, and increasingly that writing trains the language models that millions of people then ask about ani

Auditing Framing-Sensitive Behavioral Instability in Large Language Models for Mental Health Interactions

ResearchDGX agent

arXiv:2606.26982v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly being integrated into mental health support tools and other psychologically sensitive conversational app

Augmentation techniques for video surveillance in the visible and thermal spectral range

TutorialsDGX agent

arXiv:2606.13042v2 Announce Type: replace Abstract: In intelligent video surveillance, cameras record image sequences during day and night. Commonly, this demands different sensors. To achieve a bette

auto-psych: Automating the science of mind using agent-driven theory discovery and experimentation

Model ReleasesDGX agent

arXiv:2606.26460v1 Announce Type: new Abstract: AI-based scientific automation is increasingly possible by using agents to generate hypotheses, design experiments, and analyze data. Data collection is

Autoformalization of Agent Instructions into Policy-as-Code

Model ReleasesDGX agent

arXiv:2606.26649v1 Announce Type: new Abstract: Agent safety in high-stakes domains requires formal policy enforcement, but most existing approaches either rely on probabilistic guardrails (fine-tuned

Automated reproducibility assessments in the social and behavioral sciences using large language models

ResearchDGX agent

arXiv:2606.13670v2 Announce Type: replace Abstract: Reproducibility in the social and behavioral sciences is typically evaluated by independent researchers who reanalyze the original data to assess wh

Automating Potential-based Reward Shaping with Vision Language Model Guidance

SafetyDGX agent

arXiv:2606.27180v1 Announce Type: cross Abstract: Sparse rewards are inherently challenging for reinforcement learning agents as they lack intermediate feedback to guide exploration and to correctly a

Autoregressive Boltzmann Generators

Model ReleasesDGX agent

arXiv:2606.27361v1 Announce Type: cross Abstract: Efficient sampling of molecular systems at thermodynamic equilibrium is a hallmark challenge in statistical physics. This challenge has driven the dev

AXLE: A Cloud Infrastructure for Lean 4 Theorem Proving Utilities

AgentsDGX agent

arXiv:2606.26442v1 Announce Type: cross Abstract: We present AXLE (Axiom Lean Engine), a cloud service for Lean 4 proof manipulation, extraction, and verification. Recent progress in AI for mathematic

Benchmarking Open-Weight Foundation Models for Global AI Technical Governance

Model ReleasesDGX agent

arXiv:2606.26099v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in artificial intelligence (AI) governance analysis across national and international organisat

Beyond Feedforward Networks: Reentry Neural Systems as the Fundamental Basis of Subjecthood and Intrinsic Safety of Next-Generation AGI

SafetyDGX agent

arXiv:2606.26406v1 Announce Type: cross Abstract: We propose a complete architectural blueprint for safe artificial general intelligence based on a closed reentry loop (D I cycle). In contrast to feed

Beyond Global Divergences: A Local-Mass Perspective on Bayesian Inference

Model ReleasesDGX agent

arXiv:2606.27090v1 Announce Type: cross Abstract: Global objectives, such as KL divergence and ELBO, are widely used in Bayesian inference for measuring distributional discrepancy. This paper studies

Beyond Logical Forms: LLM-Extracted Patterns for Fallacy Classification

ResearchDGX agent

arXiv:2606.26698v1 Announce Type: cross Abstract: In today's fast-paced information era, logical fallacies, defined as defective patterns of reasoning, inevitably contribute to the growth of informati

Beyond the Hard Budget: Sparsity Regularizers for More Interpretable Top-k Sparse Autoencoders

ResearchDGX agent

arXiv:2606.27321v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a leading tool for interpreting the representations of vision foundation models, decomposing their polysemantic

Boundary-Aware Context Grounding for A Low-Channel EEG Agent

Model ReleasesDGX agent

arXiv:2606.26519v1 Announce Type: new Abstract: Large language models (LLMs) can make scientific software easier to use. However, a general model does not automatically know which measurements a parti

Bridging Talk and Thought: Understanding Dialogue Dynamics Across Collaborative Problem-Solving Contexts

AgentsDGX agent

arXiv:2606.27233v1 Announce Type: cross Abstract: We present a conceptual framework for analyzing dialogue in collaborative problem-solving contexts, with an emphasis on the emerging dynamics of human

Bridging Vision and Language Concepts through Optimal Transport Semantic Flow

Local AiDGX agent

arXiv:2606.26891v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) promise transparent reasoning by predicting through human-interpretable concepts, yet their effectiveness fundamental

Byzantine-Robust Aggregation for Securing Decentralized Federated Learning

Local AiDGX agent

arXiv:2409.17754v2 Announce Type: replace-cross Abstract: Federated Learning (FL) emerges as a distributed machine learning approach that addresses privacy concerns by training AI models locally on de

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention

HardwareDGX agent

arXiv:2606.27229v1 Announce Type: cross Abstract: Recurrent models must forget in order to remember, yet the state of the art decides what to erase without consulting what is stored -- the gate sees o

CascadeFormer: Depth-Tapered Transformers Motivated by Gradient Fan-in Asymmetry

Model ReleasesDGX agent

arXiv:2606.26538v1 Announce Type: cross Abstract: Deep Transformers are composed of uniformly stacked residual blocks, yet their deepest layers often add little value. We present two efficiency method

CAT-Q: Cost-efficient and Accurate Ternary Quantization for LLMs

HardwareDGX agent

arXiv:2606.26650v1 Announce Type: cross Abstract: In this paper, we present CAT-Q, Cost-efficient and Accurate Ternary Quantization, for compressing and accelerating LLMs. Unlike existing state-of-the

Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities

SafetyDGX agent

arXiv:2606.26933v1 Announce Type: cross Abstract: AI-assisted vulnerability discovery has proven effective for bug classes like memory safety, where instrumentation confirms memory violations and effi

Charting the Growth of Social-Physical HRI (spHRI): A Systematic Review Pipeline Augmented by Small Language Models

ResearchDGX agent

arXiv:2606.26382v1 Announce Type: cross Abstract: Social-physical human-robot interaction (spHRI) has grown rapidly across robotics, human-computer interaction, human-robot interaction, and haptics. Y

Clinical Harness for Governable Medical AI Skill Ecosystems

ResearchDGX agent

arXiv:2606.26494v1 Announce Type: new Abstract: Medical AI remains organized around isolated models, whereas clinical care requires accountable capabilities that persist across time. We propose clinic

Closing the Loop to Discover Psychological Theories with an Automated Cognitive Scientist

AgentsDGX agent

arXiv:2606.26448v1 Announce Type: cross Abstract: Across the sciences, autonomous systems are increasingly being used in closed-loop discovery, proposing new theories and designing and running experim

Computational Analysis of Heart Rate Variability in Healthy Adults

ResearchDGX agent

arXiv:2606.26816v1 Announce Type: new Abstract: Heart Rate Variability (HRV) analysis is a key indicator of cardiac physiological state and aids in disease diagnosis. However, research on HRV paramete

Confidence-Aware Tool Orchestration for Robust Video Understanding

Model ReleasesDGX agent

arXiv:2606.26904v1 Announce Type: cross Abstract: Video reasoning language models implicitly assume that every input frame is equally reliable. This leads to what we term the Blind Trust Problem: unde

ConflictScore: Identifying and Measuring How Language Models Handle Conflicting Evidence

Model ReleasesDGX agent

arXiv:2606.26437v1 Announce Type: cross Abstract: Existing metrics for factuality and faithfulness evaluate whether an answer is supported or contradicted by its grounding documents, but they fail to

Content-Based Smart E-Mail Dispatcher Using Large Language Models

AgentsDGX agent

arXiv:2606.26593v1 Announce Type: new Abstract: Email communication has become an integral part of personal and professional life, but handling its vast volume is still a significant issue for large o

Context-Aware Synthesis of Optimization Pipelines for Warehouse Optimization

Model ReleasesDGX agent

arXiv:2606.26852v1 Announce Type: new Abstract: Order fulfillment in manual picker-to-goods warehouses involves interconnected decisions such as item assignment, order batching, and picker routing. Wh

Context Recycling for Long-Horizon LLM Inference

Model ReleasesDGX agent

arXiv:2606.26105v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong capabilities in short-context reasoning but degrade in performance over long conversational horizons due t

COrigami: An AI Pipeline for Co-Designing Flat-Foldable Visually Recognisable Origami

AgentsDGX agent

arXiv:2606.26299v1 Announce Type: new Abstract: While generative AI has achieved remarkable success in solving problems with verifiable solutions, generating physical art that satisfies both strict ge

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation

SafetyDGX agent

arXiv:2606.26423v1 Announce Type: cross Abstract: Long-horizon, contact-rich complex manipulation tasks, such as seating a GPU into a PCIe slot, demand both millimeter high precision and out-of-the-bo

CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?

Model ReleasesDGX agent

arXiv:2606.26216v1 Announce Type: cross Abstract: We present CyberChainBench, a benchmark for evaluating LLM-based agents on smart contract security across three complementary tasks: vulnerability det

Data-driven Machine Learning Cannot Reach Symbolic-level Logical Reasoning -- The Limit of the Scaling Law

Model ReleasesDGX agent

arXiv:2606.26454v1 Announce Type: new Abstract: Sphere neural networks have achieved symbolic level syllogistic reasoning without training data, raising the question of where the limit of the scaling

Data-Free Reservoir Features for Efficient Long-Horizon Cold-Start Continual Learning

SafetyDGX agent

arXiv:2606.27095v1 Announce Type: cross Abstract: Cold-start exemplar-free class-incremental learning requires learning a growing set of classes without replay, external pretraining, or a large initia

Decision-Aligned Evaluation of Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2606.26990v1 Announce Type: cross Abstract: Uncertainty estimates in machine learning are typically evaluated using generic metrics such as the negative log-likelihood and expected calibration e

Delegation and Verification Under AI

ResearchDGX agent

arXiv:2603.02961v2 Announce Type: replace-cross Abstract: As AI systems enter institutional workflows, workers must decide whether to delegate task execution to AI and how much effort to invest in ver

Detecting and Controlling Sycophancy with Cascading Linear Features

ResearchDGX agent

arXiv:2606.26155v1 Announce Type: new Abstract: Interpreting and controlling model behaviors through activation steering methods requires many pairs of contrastive samples that clearly exhibit desired

Deterministic Pareto-Optimal Policy Synthesis for Multi-Objective Reinforcement Learning

SafetyDGX agent

arXiv:2606.26397v1 Announce Type: cross Abstract: Real-world decision-making often requires balancing multiple conflicting objectives, a challenge that standard Reinforcement Learning (RL) frequently

Diagnosing Task Insensitivity in Language Agents

SafetyDGX agent

arXiv:2606.26918v1 Announce Type: new Abstract: Large language models can serve as capable long-horizon agents, but their out-of-distribution (OOD) generalization remains weak. We identify a key sourc

Digital Twin-Driven Communication-Efficient Federated Anomaly Detection for Industrial IoT

Model ReleasesDGX agent

arXiv:2601.01701v2 Announce Type: replace-cross Abstract: Anomaly detection is increasingly becoming crucial for maintaining the safety, reliability, and efficiency of industrial systems. Recently, wi

Disco-LoRA: Disentangled Composition of Content, Style, and Motion for Multi-concept Video Customization

Model ReleasesDGX agent

arXiv:2606.26668v1 Announce Type: cross Abstract: Video customization based on Text-to-Video (T2V) models aims to learn specific features from reference data to generate controllable videos. While sig

Discovering Millions of Interpretable Features with Sparse Autoencoders

ApplicationsDGX agent

arXiv:2606.26620v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a powerful tool for decomposing superposed language model representations into sparse and interpretable fea

← Previous
1…120121122123124…358
Next →