AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 May 2026

Deepfake Audio Detection Using Self-supervised Fusion Representations

ResearchDGX agent

arXiv:2605.03420v1 Announce Type: cross Abstract: This paper describes a submission to the Environment-Aware Speech and Sound Deepfake Detection Challenge (ESDD2) 2026, which addresses component-level

Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks

ApplicationsDGX agent

arXiv:2601.17644v3 Announce Type: replace-cross Abstract: The growing adoption of multimodal Retrieval-Augmented Generation (mRAG) pipelines for vision-centric tasks (e.g., visual QA) introduces impor

Enhancing Agent Safety Judgment: Controlled Benchmark Rewriting and Analogical Reasoning for Deceptive Out-of-Distribution Scenarios

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.03242v1 Announce Type: new Abstract: Tool-using agent systems powered by large language models (LLMs) are increasingly deployed across web, app, operating-system, and transactional environm

Evaluating Prompting and Execution-Based Methods for Deterministic Computation in LLMs

ResearchDGX agent

arXiv:2605.03227v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in natural language understanding and reasoning. However, their ability to perform ex

Evaluating Semantic Fragility in Text-to-Audio Generation Systems Under Controlled Prompt Perturbations

SafetyDGX agent

arXiv:2603.13824v2 Announce Type: replace-cross Abstract: Recent advances in text-to-audio generation enable models to translate natural-language descriptions into diverse musical output. However, the

EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics

Model ReleasesDGX agent

arXiv:2605.03871v1 Announce Type: new Abstract: Language models encode substantial evaluative knowledge from pretraining, yet current post-training methods rely on external supervision (human annotati

Forget BIT, It is All about TOKEN: Towards Semantic Information Theory for LLMs

TutorialsDGX agent

arXiv:2511.01202v3 Announce Type: replace-cross Abstract: Despite the unprecedented empirical triumphs of LLMs across diverse real-world applications, the prevailing research paradigm remains overwhel

From Barrier to Bridge: The Case for AI Data Center/Power Grid Co-Design

ResearchDGX agent

arXiv:2605.03090v1 Announce Type: cross Abstract: For over a century, the electric grid has relied on a single statistical assumption: load diversity, the principle that the uncorrelated demands of mi

From Intent to Execution: Composing Agentic Workflows with Agent Recommendation

Local AiDGX agent

arXiv:2605.03986v1 Announce Type: new Abstract: Multi-Agent Systems (MAS) built using AI agents fulfill a variety of user intents that may be used to design and build a family of related applications.

From Knowledge to Action: Outcomes of the 2025 Large Language Model (LLM) Hackathon for Applications in Materials Science and Chemistry

AgentsDGX agent

arXiv:2605.03205v1 Announce Type: cross Abstract: Large language models (LLMs) are rapidly changing how researchers in materials science and chemistry discover, organize, and act on scientific knowled

From Passive Feeds to Guided Discovery: AI-Initiated Interaction for Vague Intent in Content Exploration

ResearchDGX agent

arXiv:2605.02902v1 Announce Type: cross Abstract: Recommendation feeds work well when people are simply browsing, and search works well when they can formulate a query. Between these two cases is a co

Generalization Bounds of Spiking Neural Networks via Rademacher Complexity

Model ReleasesDGX agent

arXiv:2605.02927v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have garnered increasing attention as one of bio-inspired models due to their great potential in neuromorphic computing

GeoDecider: A Coarse-to-Fine Agentic Workflow for Explainable Lithology Classification

AgentsDGX agent

arXiv:2605.03383v1 Announce Type: new Abstract: Lithology classification aims to infer subsurface rock types from well-logging signals, supporting downstream applications like reservoir characterizati

Geometry over Density: Few-Shot Cross-Domain OOD Detection

ApplicationsDGX agent

arXiv:2605.03410v2 Announce Type: new Abstract: Out-of-distribution (OOD) detection identifies test samples that fall outside a model's training distribution, a capability critical for safe deployment

Human-Provenance Verification should be Treated as Labor Infrastructure in AI-Saturated Markets

AgentsDGX agent

arXiv:2605.03210v1 Announce Type: cross Abstract: We argue that AI-saturated markets are likely to create Veblen-good premiums, which we term human-provenance premiums, for verified human presence, an

Inconsistent Databases and Argumentation Frameworks with Collective Attacks

ResearchDGX agent

arXiv:2605.03954v1 Announce Type: cross Abstract: The connection between subset-maximal repairs for inconsistent databases involving various integrity constraints and acceptable sets of arguments with

Intelligent Knowledge Mining Framework: Bridging AI Analysis and Trustworthy Preservation

ResearchDGX agent

arXiv:2512.17795v2 Announce Type: replace-cross Abstract: The unprecedented proliferation of digital data presents significant challenges in access, integration, and value creation across all data-int

Keyword spotting using convolutional neural network for speech recognition in Hindi

Local AiDGX agent

arXiv:2605.02928v1 Announce Type: cross Abstract: In this study, we investigate the application of keyword spotting (KWS) in the domain of Hindi speech recognition, utilizing a dataset comprising 40,0

Learning Correct Behavior from Examples: Validating Sequential Execution in Autonomous Agents

AgentsDGX agent

arXiv:2605.03159v1 Announce Type: new Abstract: As autonomous agents become increasingly sophisticated, validating their sequential behavior presents a significant challenge. Traditional testing appro

Magic-Informed Quantum Architecture Search

Model ReleasesDGX agent

arXiv:2605.03932v1 Announce Type: cross Abstract: Nonstabilizerness, commonly referred to as magic, is a fundamental resource underpinning quantum advantage. In this paper, we propose a magic-informed

Making the Invisible Visible: Understanding the Mismatch Between Organizational Goals and Worker Experiences in AI Adoption

ApplicationsDGX agent

arXiv:2605.03078v1 Announce Type: new Abstract: While AI is often introduced into organizations to drive innovation and efficiency, many adoption efforts fail as workers resist and struggle to integra

Mechanical Conscience: A Mathematical Framework for Dependability of Machine Intelligenc

SafetyDGX agent

arXiv:2605.03847v1 Announce Type: new Abstract: Distributed collaborative intelligence (DCI), encompassing edge-to-edge architectures, federated learning, transfer learning, and swarm systems, creates

MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2605.03675v1 Announce Type: new Abstract: Long-running autonomous AI agents suffer from a well-documented memory coherence problem: tool-execution success rates degrade 14 percentage points over

MenuNet: A Strategy-Proof Mechanism for Matching Markets

SafetyDGX agent

arXiv:2605.03216v1 Announce Type: cross Abstract: Strategy-proofness is a fundamental desideratum in mechanism design, ensuring truthful reporting and robust participation. Stability is another centra

MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents

Model ReleasesDGX agent

arXiv:2605.03952v1 Announce Type: cross Abstract: Coding agents often pass per-prompt safety review yet ship exploitable code when their tasks are decomposed into routine engineering tickets. The chal

Multi-Agent Strategic Games with LLMs

AgentsDGX agent

arXiv:2605.03604v1 Announce Type: cross Abstract: This paper asks whether large language models (LLMs) can be used to study the strategic foundations of conflict and cooperation. I introduce LLMs as e

Multi Language Models for On-the-Fly Syntax Highlighting

TutorialsDGX agent

arXiv:2510.04166v2 Announce Type: replace-cross Abstract: Syntax highlighting is a critical feature in modern software development environments, enhancing code readability and developer productivity.

On the evolutionary cognitive pressure for experiential awareness: do machines need it?

AgentsDGX agent

arXiv:2510.20839v2 Announce Type: replace-cross Abstract: The consciousness standing for artificial intelligence divides opinions across epistemological positions. Whether or not machines can be consc

OptiLookUp: An Optical ROM-Based Loop up Table Engine for Photonic Accelerators

HardwareDGX agent

arXiv:2605.03241v1 Announce Type: cross Abstract: Read-only memory (ROM) provides deterministic access to predefined data mappings. Extending ROM concepts to the optical domain enables high-bandwidth,

OracleProto: A Reproducible Framework for Benchmarking LLM Native Forecasting via Knowledge Cutoff and Temporal Masking

SafetyDGX agent

arXiv:2605.03762v1 Announce Type: new Abstract: Large language models are moving from static text generators toward real-world decision-support systems, where forecasting is a composite capability tha

Pact: A Choreographic Language for Agentic Ecosystems

AgentsDGX agent

arXiv:2605.03143v1 Announce Type: cross Abstract: Recent advances in large language models have led to the rise of software systems (i.e. agents) that execute with increasing autonomy on behalf of use

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal

SafetyDGX agent

arXiv:2603.22844v4 Announce Type: replace Abstract: Surgical smoke severely degrades intraoperative video quality, obscuring anatomical structures and limiting surgical perception. Existing learning-b

Physics-Grounded Multi-Agent Architecture for Traceable, Risk-Aware Human-AI Decision Support in Manufacturing

Model ReleasesDGX agent

arXiv:2605.04003v1 Announce Type: cross Abstract: High-precision CNC machining of free-form aerospace components requires bounded compensations informed by inspection, simulation, and process knowledg

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

AgentsDGX agent

arXiv:2506.00886v3 Announce Type: replace Abstract: As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Exi

ProgramBench: Can Language Models Rebuild Programs From Scratch?

AgentsDGX agent

arXiv:2605.03546v1 Announce Type: cross Abstract: Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain,

Programmatic Context Augmentation for LLM-based Symbolic Regression

ResearchDGX agent

arXiv:2605.03101v1 Announce Type: new Abstract: Symbolic regression (SR), the task of discovering mathematical expressions that best describe a given dataset, remains a fundamental challenge in scient

QKVShare: Quantized KV-Cache Handoff for Multi-Agent On-Device LLMs

Model ReleasesDGX agent

arXiv:2605.03884v1 Announce Type: new Abstract: Multi-agent LLM systems on edge devices need to hand off latent context efficiently, but the practical choices today are expensive re-prefill or full-pr

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents

SafetyDGX agent

arXiv:2604.03976v2 Announce Type: replace Abstract: Prior work on trustworthy AI emphasizes model-internal properties such as bias mitigation, adversarial robustness, and interpretability. As AI syste

RAMoEA-QA: Hierarchical Specialization for Robust Respiratory Audio Question Answering

ApplicationsDGX agent

arXiv:2603.06542v2 Announce Type: replace-cross Abstract: Conversational generative AI is increasingly explored in healthcare, where models must integrate heterogeneous patient signals and support div

Real-Time Evaluation of Autonomous Systems under Adversarial Attacks

Model ReleasesDGX agent

arXiv:2605.03491v1 Announce Type: new Abstract: Most evaluations of autonomous driving policies under adversarial conditions are conducted in simulation, due to cost efficiency and the absence of phys

ReasonAudio: A Benchmark for Evaluating Reasoning Beyond Matching in Text-Audio Retrieval

Model ReleasesDGX agent

arXiv:2605.03361v2 Announce Type: new Abstract: As multimodal content continues to expand at a rapid pace, audio retrieval has emerged as a key enabling technology for media search, content organizati

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours

Model ReleasesDGX agent

arXiv:2605.04019v1 Announce Type: new Abstract: AI systems are entering critical domains like healthcare, finance, and defense, yet remain vulnerable to adversarial attacks. While AI red teaming is a

Replacing Parameters with Preferences: Federated Alignment of Heterogeneous Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.03426v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have broad potential in privacy-sensitive domains such as healthcare and finance, yet strict data-sharing constraints rend

Revisiting the Travel Planning Capabilities of Large Language Models

AgentsDGX agent

arXiv:2605.03308v1 Announce Type: new Abstract: Travel planning serves as a critical task for long-horizon reasoning, exposing significant deficits in LLMs. However, existing benchmarks and evaluation

Robust Agent Compensation (RAC): Teaching AI Agents to Compensate

SafetyDGX agent

arXiv:2605.03409v1 Announce Type: new Abstract: We present Robust Agent Compensation (RAC), a log-based recovery paradigm (providing a safety net) implemented through an architectural extension that c

Safety Must Precede the Deployment of Open-Ended AI

SafetyDGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

Same Voice, Different Lab: On the Homogenization of Frontier LLM Personalities

ResearchDGX agent

arXiv:2605.02897v1 Announce Type: cross Abstract: LLM assistant personalities play a critical role in user experience and perceived response quality. We present a large-scale experiment of frontier LL

ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting

Local AiDGX agent

arXiv:2605.03804v1 Announce Type: new Abstract: Long-term personalized memory for LLM agents is challenging on resource-limited edge devices due to high storage costs and multimodal complexity. To add

Self-Improvement for Fast, High-Quality Plan Generation

TutorialsDGX agent

arXiv:2605.03625v1 Announce Type: new Abstract: Generative models trained on synthetic plan data are a promising approach to generalized planning. Recent work has focused on finding any valid plan, ra

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

Model ReleasesDGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

Smart Passive Acoustic Monitoring: Embedding a Classifier on AudioMoth Microcontroller

AgentsDGX agent

arXiv:2605.03412v1 Announce Type: cross Abstract: Passive Acoustic Monitoring (PAM) is an efficient and non-invasive method for surveying ecosystems at a reduced cost. Typically, autonomous recorders

Stable Agentic Control: Tool-Mediated LLM Architecture for Autonomous Cyber Defense

Model ReleasesDGX agent

arXiv:2605.03034v1 Announce Type: new Abstract: Agentic systems involved in high-stake decision-making under adversarial pressure need formal guarantees not offered by existing approaches. Motivated b

Stage Light is Sequence^2: Multi-Light Control via Imitation Learning

Model ReleasesDGX agent

arXiv:2605.03660v1 Announce Type: cross Abstract: Music-inspired Automatic Stage Lighting Control (ASLC) has gained increasing attention in recent years due to the substantial time and financial costs

Stop Automating Peer Review Without Rigorous Evaluation

ResearchDGX agent

arXiv:2605.03202v1 Announce Type: new Abstract: Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used

SymptomAI: Towards a Conversational AI Agent for Everyday Symptom Assessment

AgentsDGX agent

arXiv:2605.04012v1 Announce Type: new Abstract: Language models excel at diagnostic assessments on currated medical case-studies and vignettes, performing on par with, or better than, clinical profess

Tailored Prompts, Targeted Protection: Vulnerability-Specific LLM Analysis for Smart Contracts

ApplicationsDGX agent

arXiv:2605.03697v1 Announce Type: cross Abstract: Smart contracts on blockchains are prone to diverse security vulnerabilities that can lead to significant financial losses due to their immutable natu

TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference

Model ReleasesDGX agent

arXiv:2603.26498v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) power platforms like ChatGPT, Gemini, and Copilot, enabling richer interactions with text, images, an

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

Model ReleasesDGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

The Hive Mind is a Single Reinforcement Learning Agent

Local AiDGX agent

arXiv:2410.17517v5 Announce Type: replace-cross Abstract: Decision-making is an essential attribute of any intelligent agent or group. Natural systems are known to converge to effective strategies thr

Towards Open World Sound Event Detection

TutorialsDGX agent

arXiv:2605.03934v1 Announce Type: cross Abstract: Sound Event Detection (SED) plays a vital role in audio understanding, with applications in surveillance, smart cities, healthcare, and multimedia ind

← Previous
1…285286287288289…358
Next →