AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents

DGX agent

arXiv:2605.03952v1 Announce Type: cross Abstract: Coding agents often pass per-prompt safety review yet ship exploitable code when their tasks are decomposed into routine engineering tickets. The chal

model-releasesarxiv-cs-ai
7 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-Agent Strategic Games with LLMs

DGX agent

arXiv:2605.03604v1 Announce Type: cross Abstract: This paper asks whether large language models (LLMs) can be used to study the strategic foundations of conflict and cooperation. I introduce LLMs as e

agentsarxiv-cs-ai
7 May 2026
Tutorials

Multi Language Models for On-the-Fly Syntax Highlighting

DGX agent

arXiv:2510.04166v2 Announce Type: replace-cross Abstract: Syntax highlighting is a critical feature in modern software development environments, enhancing code readability and developer productivity.

tutorialsarxiv-cs-ai
7 May 2026
Agents

On the evolutionary cognitive pressure for experiential awareness: do machines need it?

DGX agent

arXiv:2510.20839v2 Announce Type: replace-cross Abstract: The consciousness standing for artificial intelligence divides opinions across epistemological positions. Whether or not machines can be consc

agentsarxiv-cs-ai
7 May 2026
Hardware

OptiLookUp: An Optical ROM-Based Loop up Table Engine for Photonic Accelerators

DGX agent

arXiv:2605.03241v1 Announce Type: cross Abstract: Read-only memory (ROM) provides deterministic access to predefined data mappings. Extending ROM concepts to the optical domain enables high-bandwidth,

hardwarearxiv-cs-ai
7 May 2026
Safety

OracleProto: A Reproducible Framework for Benchmarking LLM Native Forecasting via Knowledge Cutoff and Temporal Masking

DGX agent

arXiv:2605.03762v1 Announce Type: new Abstract: Large language models are moving from static text generators toward real-world decision-support systems, where forecasting is a composite capability tha

safetyarxiv-cs-ai
7 May 2026
Agents

Pact: A Choreographic Language for Agentic Ecosystems

DGX agent

arXiv:2605.03143v1 Announce Type: cross Abstract: Recent advances in large language models have led to the rise of software systems (i.e. agents) that execute with increasing autonomy on behalf of use

agentsarxiv-cs-ai
7 May 2026
Safety

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal

DGX agent

arXiv:2603.22844v4 Announce Type: replace Abstract: Surgical smoke severely degrades intraoperative video quality, obscuring anatomical structures and limiting surgical perception. Existing learning-b

safetyarxiv-cs-ai
7 May 2026
Model Releases

Physics-Grounded Multi-Agent Architecture for Traceable, Risk-Aware Human-AI Decision Support in Manufacturing

DGX agent

arXiv:2605.04003v1 Announce Type: cross Abstract: High-precision CNC machining of free-form aerospace components requires bounded compensations informed by inspection, simulation, and process knowledg

model-releasesarxiv-cs-ai
7 May 2026
Agents

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

DGX agent

arXiv:2506.00886v3 Announce Type: replace Abstract: As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Exi

agentsarxiv-cs-ai
7 May 2026
Agents

ProgramBench: Can Language Models Rebuild Programs From Scratch?

DGX agent

arXiv:2605.03546v1 Announce Type: cross Abstract: Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain,

agentsarxiv-cs-ai
7 May 2026
Research

Programmatic Context Augmentation for LLM-based Symbolic Regression

DGX agent

arXiv:2605.03101v1 Announce Type: new Abstract: Symbolic regression (SR), the task of discovering mathematical expressions that best describe a given dataset, remains a fundamental challenge in scient

researcharxiv-cs-ai
7 May 2026
Model Releases

QKVShare: Quantized KV-Cache Handoff for Multi-Agent On-Device LLMs

DGX agent

arXiv:2605.03884v1 Announce Type: new Abstract: Multi-agent LLM systems on edge devices need to hand off latent context efficiently, but the practical choices today are expensive re-prefill or full-pr

model-releasesarxiv-cs-ai
7 May 2026
Safety

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents

DGX agent

arXiv:2604.03976v2 Announce Type: replace Abstract: Prior work on trustworthy AI emphasizes model-internal properties such as bias mitigation, adversarial robustness, and interpretability. As AI syste

safetyarxiv-cs-ai
7 May 2026
Applications

RAMoEA-QA: Hierarchical Specialization for Robust Respiratory Audio Question Answering

DGX agent

arXiv:2603.06542v2 Announce Type: replace-cross Abstract: Conversational generative AI is increasingly explored in healthcare, where models must integrate heterogeneous patient signals and support div

applicationsarxiv-cs-ai
7 May 2026
Model Releases

Real-Time Evaluation of Autonomous Systems under Adversarial Attacks

DGX agent

arXiv:2605.03491v1 Announce Type: new Abstract: Most evaluations of autonomous driving policies under adversarial conditions are conducted in simulation, due to cost efficiency and the absence of phys

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

ReasonAudio: A Benchmark for Evaluating Reasoning Beyond Matching in Text-Audio Retrieval

DGX agent

arXiv:2605.03361v2 Announce Type: new Abstract: As multimodal content continues to expand at a rapid pace, audio retrieval has emerged as a key enabling technology for media search, content organizati

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours

DGX agent

arXiv:2605.04019v1 Announce Type: new Abstract: AI systems are entering critical domains like healthcare, finance, and defense, yet remain vulnerable to adversarial attacks. While AI red teaming is a

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Replacing Parameters with Preferences: Federated Alignment of Heterogeneous Vision-Language Models

DGX agent

arXiv:2605.03426v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have broad potential in privacy-sensitive domains such as healthcare and finance, yet strict data-sharing constraints rend

model-releasesarxiv-cs-ai
7 May 2026
Agents

Revisiting the Travel Planning Capabilities of Large Language Models

DGX agent

arXiv:2605.03308v1 Announce Type: new Abstract: Travel planning serves as a critical task for long-horizon reasoning, exposing significant deficits in LLMs. However, existing benchmarks and evaluation

agentsarxiv-cs-ai
7 May 2026
Safety

Robust Agent Compensation (RAC): Teaching AI Agents to Compensate

DGX agent

arXiv:2605.03409v1 Announce Type: new Abstract: We present Robust Agent Compensation (RAC), a log-based recovery paradigm (providing a safety net) implemented through an architectural extension that c

safetyarxiv-cs-ai
7 May 2026
Safety

Safety Must Precede the Deployment of Open-Ended AI

DGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

safetyarxiv-cs-ai
7 May 2026
Research

Same Voice, Different Lab: On the Homogenization of Frontier LLM Personalities

DGX agent

arXiv:2605.02897v1 Announce Type: cross Abstract: LLM assistant personalities play a critical role in user experience and perceived response quality. We present a large-scale experiment of frontier LL

researcharxiv-cs-ai
7 May 2026
Local Ai

ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting

DGX agent

arXiv:2605.03804v1 Announce Type: new Abstract: Long-term personalized memory for LLM agents is challenging on resource-limited edge devices due to high storage costs and multimodal complexity. To add

local-aiarxiv-cs-ai
7 May 2026
Tutorials

Self-Improvement for Fast, High-Quality Plan Generation

DGX agent

arXiv:2605.03625v1 Announce Type: new Abstract: Generative models trained on synthetic plan data are a promising approach to generalized planning. Recent work has focused on finding any valid plan, ra

tutorialsarxiv-cs-ai
7 May 2026
Model Releases

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

DGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

model-releasesarxiv-cs-ai
7 May 2026
Agents

Smart Passive Acoustic Monitoring: Embedding a Classifier on AudioMoth Microcontroller

DGX agent

arXiv:2605.03412v1 Announce Type: cross Abstract: Passive Acoustic Monitoring (PAM) is an efficient and non-invasive method for surveying ecosystems at a reduced cost. Typically, autonomous recorders

agentsarxiv-cs-ai
7 May 2026
Model Releases

Stable Agentic Control: Tool-Mediated LLM Architecture for Autonomous Cyber Defense

DGX agent

arXiv:2605.03034v1 Announce Type: new Abstract: Agentic systems involved in high-stake decision-making under adversarial pressure need formal guarantees not offered by existing approaches. Motivated b

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Stage Light is Sequence^2: Multi-Light Control via Imitation Learning

DGX agent

arXiv:2605.03660v1 Announce Type: cross Abstract: Music-inspired Automatic Stage Lighting Control (ASLC) has gained increasing attention in recent years due to the substantial time and financial costs

model-releasesarxiv-cs-ai
7 May 2026
Research

Stop Automating Peer Review Without Rigorous Evaluation

DGX agent

arXiv:2605.03202v1 Announce Type: new Abstract: Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used

researcharxiv-cs-ai
7 May 2026
Agents

SymptomAI: Towards a Conversational AI Agent for Everyday Symptom Assessment

DGX agent

arXiv:2605.04012v1 Announce Type: new Abstract: Language models excel at diagnostic assessments on currated medical case-studies and vignettes, performing on par with, or better than, clinical profess

agentsarxiv-cs-ai
7 May 2026
Applications

Tailored Prompts, Targeted Protection: Vulnerability-Specific LLM Analysis for Smart Contracts

DGX agent

arXiv:2605.03697v1 Announce Type: cross Abstract: Smart contracts on blockchains are prone to diverse security vulnerabilities that can lead to significant financial losses due to their immutable natu

applicationsarxiv-cs-ai
7 May 2026
Model Releases

TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference

DGX agent

arXiv:2603.26498v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) power platforms like ChatGPT, Gemini, and Copilot, enabling richer interactions with text, images, an

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

DGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

model-releasesarxiv-cs-ai
7 May 2026
Local Ai

The Hive Mind is a Single Reinforcement Learning Agent

DGX agent

arXiv:2410.17517v5 Announce Type: replace-cross Abstract: Decision-making is an essential attribute of any intelligent agent or group. Natural systems are known to converge to effective strategies thr

local-aiarxiv-cs-ai
7 May 2026
Tutorials

Towards Open World Sound Event Detection

DGX agent

arXiv:2605.03934v1 Announce Type: cross Abstract: Sound Event Detection (SED) plays a vital role in audio understanding, with applications in surveillance, smart cities, healthcare, and multimedia ind

tutorialsarxiv-cs-ai
7 May 2026
Safety

Unifying Dynamical Systems and Graph Theory to Mechanistically Understand Computation in Neural Networks

DGX agent

arXiv:2605.03598v2 Announce Type: cross Abstract: Understanding how biological and artificial neural networks implement computation from connectivity is a central problem in neuroscience and machine l

safetyarxiv-cs-ai
7 May 2026
Model Releases

VCBench: Benchmarking LLMs in Venture Capital

DGX agent

arXiv:2509.14448v2 Announce Type: replace Abstract: Benchmarks such as SWE-bench and ARC-AGI demonstrate how shared datasets accelerate progress toward artificial general intelligence (AGI). We introd

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

DGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne

model-releasesarxiv-cs-ai
7 May 2026
Agents

What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity

DGX agent

arXiv:2605.03782v1 Announce Type: new Abstract: To navigate partially observable visual environments, recent VLM agents increasingly internalize world modeling capabilities into their policies via exp

agentsarxiv-cs-ai
7 May 2026
Hardware

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

DGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

hardwarearxiv-cs-ai
7 May 2026
Model Releases

12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation

DGX agent

arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence

DGX agent

arXiv:2605.01546v1 Announce Type: cross Abstract: Sixth-generation (6G) networks are increasingly envisioned as AI-native infrastructures integrating communication, sensing, and computing into a unifi

model-releasesarxiv-cs-ai
6 May 2026
Local Ai

A Cellular Doctrine of Morality: Intrinsic Active Precision and the Mind-Reality Overload Dilemma

DGX agent

arXiv:2605.01376v1 Announce Type: new Abstract: Current AI systems, grounded in oversimplified neuroscience, risk eroding the distinction between truth and falsehood. They maximize reward by amplifyin

local-aiarxiv-cs-ai
6 May 2026
Agents

A Compound AI Agent for Conversational Grant Discovery

DGX agent

arXiv:2605.02366v1 Announce Type: new Abstract: Research funding discovery remains fundamentally fragmented: researchers navigate disparate agency portals (e.g., in the United States, NSF, NIH, DARPA,

agentsarxiv-cs-ai
6 May 2026
Safety

A Knowledge-Driven LLM-Based Decision-Support System for Explainable Defect Analysis and Mitigation Guidance in Laser Powder Bed Fusion

DGX agent

arXiv:2605.01100v1 Announce Type: new Abstract: This work presents a knowledge-driven decision-support system that integrates structured defect knowledge with LLM-based reasoning to provide explainabl

safetyarxiv-cs-ai
6 May 2026
Agents

A Low-Latency Fraud Detection Layer for Detecting Adversarial Interaction Patterns in LLM-Powered Agents

DGX agent

arXiv:2605.01143v1 Announce Type: new Abstract: Large Language Model (LLM)-powered agents demonstrate strong capabilities in autonomous task execution, tool use, and multi-step reasoning. However, the

agentsarxiv-cs-ai
6 May 2026
Applications

A Neuro-Symbolic Framework for Accountability in Public-Sector AI

DGX agent

arXiv:2512.12109v3 Announce Type: replace-cross Abstract: Automated eligibility systems increasingly determine access to essential public benefits, but the explanations they generate often fail to ref

applicationsarxiv-cs-ai
6 May 2026
← Previous
1…357358359360361…448
Next →