AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
17 Apr 2026

FRAGATA: Semantic Retrieval of HPC Support Tickets via Hybrid RAG over 20 Years of Request Tracker History

TutorialsDGX agent

arXiv:2604.13721v1 Announce Type: cross Abstract: The technical support team of a supercomputing centre accumulates, over the course of decades, a large volume of resolved incidents that constitute cr

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability

HardwareDGX agent

arXiv:2604.13048v1 Announce Type: cross Abstract: Modern cloud-native platforms expose thousands of time series metrics through systems like Prometheus, yet formulating correct queries in domain-speci

GeoAgentBench: A Dynamic Execution Benchmark for Tool-Augmented Agents in Spatial Analysis

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.13888v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into Geographic Information Systems (GIS) marks a paradigm shift toward autonomous spatial analysis. How

GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning

Model ReleasesDGX agent

arXiv:2603.01410v2 Announce Type: replace Abstract: Knowledge graphs provide structured and reliable information for many real-world applications, motivating increasing interest in combining large lan

Hijacking online reviews: sparse manipulation and behavioral buffering in popularity-biased rating systems

AgentsDGX agent

arXiv:2604.13049v1 Announce Type: cross Abstract: Online reviews and recommendation systems help users navigate overwhelming choice, but they are vulnerable to self-reinforcing distortions. This paper

ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08064v2 Announce Type: replace Abstract: Existing memory benchmarks for LLM agents evaluate explicit recall of facts, yet overlook implicit memory where experience becomes automated behavio

In-Context Autonomous Network Incident Response: An End-to-End Large Language Model Agent Approach

AgentsDGX agent

arXiv:2602.13156v2 Announce Type: replace-cross Abstract: Rapidly evolving cyberattacks demand incident response systems that can autonomously learn and adapt to changing threats. Prior work has exten

Inclusive Kitchen Design for Older Adults: Generative AI Visualizations to Support Mild Cognitive Impairment

SafetyDGX agent

arXiv:2604.13203v1 Announce Type: cross Abstract: Mild Cognitive Impairment (MCI) affects 15-20% of adults aged 65 and older, often making kitchen navigation and independent living difficult, particul

Integration of Deep Reinforcement Learning and Agent-based Simulation to Explore Strategies Counteracting Information Disorder

AgentsDGX agent

arXiv:2604.13047v1 Announce Type: cross Abstract: In recent years, the spread of fake news has triggered a growing interest in Information Disorders (ID) on social media, a phenomenon that has become

Large Language Models to Enhance Business Process Modeling: Past, Present, and Future Trends

ApplicationsDGX agent

arXiv:2604.14034v1 Announce Type: cross Abstract: Recent advances in Generative Artificial Intelligence, particularly Large Language Models (LLMs), have stimulated growing interest in automating or as

Lazy or Efficient? Towards Accessible Eye-Tracking Event Detection Using LLMs

ResearchDGX agent

arXiv:2604.13243v1 Announce Type: cross Abstract: Gaze event detection is fundamental to vision science, human-computer interaction, and applied analytics. However, current workflows often require spe

Listening Alone, Understanding Together: Collaborative Context Recovery for Privacy-Aware AI

ResearchDGX agent

arXiv:2604.13348v1 Announce Type: new Abstract: We introduce CONCORD, a privacy-aware asynchronous assistant-to-assistant (A2A) framework that leverages collaboration between proactive speech-based AI

MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2509.06477v2 Announce Type: replace Abstract: Shortcuts such as APIs and deep-links have emerged as efficient complements to flexible GUI operations, fostering a promising hybrid paradigm for ML

MCPThreatHive: Automated Threat Intelligence for Model Context Protocol Ecosystems

AgentsDGX agent

arXiv:2604.13849v1 Announce Type: cross Abstract: The rapid proliferation of Model Context Protocol (MCP)-based agentic systems has introduced a new category of security threats that existing framewor

MIND: AI Co-Scientist for Material Research

AgentsDGX agent

arXiv:2604.13699v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning witho

Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities

SafetyDGX agent

arXiv:2509.06921v2 Announce Type: replace-cross Abstract: Cybersecurity demands both rapid pattern recognition and deliberative reasoning, yet purely neural or purely symbolic approaches each address

On the Creativity of AI Agents

AgentsDGX agent

arXiv:2604.13242v1 Announce Type: cross Abstract: Large language models (LLMs), particularly when integrated into agentic systems, have demonstrated human- and even superhuman-level performance across

On the Use of Evolutionary Optimization for the Dynamic Chance Constrained Open-Pit Mine Scheduling Problem

ApplicationsDGX agent

arXiv:2604.13385v1 Announce Type: cross Abstract: Open-pit mine scheduling is a complex real world optimization problem that involves uncertain economic values and dynamically changing resource capaci

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

Model ReleasesDGX agent

arXiv:2506.03610v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game b

OVT-MLCS: An Online Visual Tool for MLCS Mining from Long or Big Sequences

ResearchDGX agent

arXiv:2604.13037v1 Announce Type: cross Abstract: Mining multiple longest common subsequences (extit{MLCS}) from a set of sequences of three or more over a finite alphabet Sigma (a classical NP-hard p

ProRe: A Proactive Reward System for GUI Agents via Reasoner-Actor Collaboration

SafetyDGX agent

arXiv:2509.21823v2 Announce Type: replace Abstract: Reward is critical to the evaluation and training of large language models (LLMs). However, existing rule-based or model-based reward methods strugg

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization

TutorialsDGX agent

arXiv:2508.10164v2 Announce Type: replace Abstract: Recent advances in Large Reasoning Models (LRMs) have demonstrated strong performance on complex tasks through long Chain-of-Thought (CoT) reasoning

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

SafetyDGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

ReSS: Learning Reasoning Models for Tabular Data Prediction via Symbolic Scaffold

TutorialsDGX agent

arXiv:2604.13392v1 Announce Type: new Abstract: Tabular data remains prevalent in high-stakes domains such as healthcare and finance, where predictive models are expected to provide both high accuracy

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

Local AiDGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

Model ReleasesDGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

Sandwich: Joint Configuration Search and Hot-Switching for Efficient CPU LLM Serving

ResearchDGX agent

arXiv:2507.18454v2 Announce Type: replace-cross Abstract: CPUs are critical for LLM serving due to their availability, cost efficiency, and edge applicability. However, efficient CPU serving is hinder

SAQ: Stabilizer-Aware Quantum Error Correction Decoder

Model ReleasesDGX agent

arXiv:2512.08914v2 Announce Type: replace-cross Abstract: Quantum Error Correction (QEC) decoding faces a fundamental accuracy-efficiency tradeoff. Classical methods like Minimum Weight Perfect Matchi

SciFi: A Safe, Lightweight, User-Friendly, and Fully Autonomous Agentic AI Workflow for Scientific Applications

AgentsDGX agent

arXiv:2604.13180v1 Announce Type: new Abstract: Recent advances in agentic AI have enabled increasingly autonomous workflows, but existing systems still face substantial challenges in achieving reliab

Secure and Privacy-Preserving Vertical Federated Learning

Model ReleasesDGX agent

arXiv:2604.13474v1 Announce Type: cross Abstract: We propose a novel end-to-end privacy-preserving framework, instantiated by three efficient protocols for different deployment scenarios, covering bot

SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits

Model ReleasesDGX agent

arXiv:2604.01473v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are powerful tools for answering user queries, yet they remain highly vulnerable to jailbreak attacks. Existing g

Sentiment analysis for software engineering: How far can zero-shot learning (ZSL) go?

ResearchDGX agent

arXiv:2604.13826v1 Announce Type: cross Abstract: Sentiment analysis in software engineering focuses on understanding emotions expressed in software artifacts. Previous research highlighted the limita

Strategic Response of News Publishers to Generative AI

ApplicationsDGX agent

arXiv:2512.24968v4 Announce Type: replace-cross Abstract: Generative AI can adversely impact news publishers by lowering consumer demand. It can also reduce demand for newsroom employees, and increase

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

AgentsDGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

The Code Whisperer: LLM and Graph-Based AI for Smell and Vulnerability Resolution

ResearchDGX agent

arXiv:2604.13114v1 Announce Type: cross Abstract: Code smells and software vulnerabilities both increase maintenance cost, yet they are often handled by separate tools that miss structural context and

The Cognitive Circuit Breaker: A Systems Engineering Framework for Intrinsic AI Reliability

ResearchDGX agent

arXiv:2604.13417v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in mission-critical software systems, detecting hallucinations and ``faked truthfulness'' ha

The Mirror Design Pattern: Strict Data Geometry over Model Scale for Prompt Injection Detection

Model ReleasesDGX agent

arXiv:2603.11875v2 Announce Type: replace-cross Abstract: Prompt injection defenses are often framed as semantic understanding problems and delegated to increasingly large neural detectors. For the fi

TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds

ResearchDGX agent

arXiv:2604.13737v1 Announce Type: cross Abstract: Recommender systems have historically developed along two largely independent paradigms: feature interaction models for modeling correlations among mu

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

SafetyDGX agent

arXiv:2604.13715v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) enable general audio understanding and demonstrate remarkable performance across various audio tasks. However, the

Towards Scalable Lightweight GUI Agents via Multi-role Orchestration

SafetyDGX agent

arXiv:2604.13488v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) enable digital automation on end-user devices. Whil

Trust and Reliance on AI in Education: AI Literacy and Need for Cognition as Moderators

ApplicationsDGX agent

arXiv:2604.01114v2 Announce Type: replace-cross Abstract: As generative AI systems are integrated into educational settings, students often encounter AI-generated output while working through learning

Variance Computation for Weighted Model Counting with Knowledge Compilation Approach

ApplicationsDGX agent

arXiv:2601.03523v2 Announce Type: replace Abstract: One of the most important queries in knowledge compilation is weighted model counting (WMC), which has been applied to probabilistic inference on va

VeruSAGE: A Study of Agent-Based Verification for Rust Systems

Model ReleasesDGX agent

arXiv:2512.18436v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive capability to understand and develop code. However, their capability to rigorously reason a

Weight Patching: Toward Source-Level Mechanistic Localization in LLMs

Model ReleasesDGX agent

arXiv:2604.13694v1 Announce Type: new Abstract: Mechanistic interpretability seeks to localize model behavior to the internal components that causally realize it. Prior work has advanced activation-sp

WybeCoder: Verified Imperative Code Generation

AgentsDGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

Young people's perceptions and recommendations for conversational generative artificial intelligence in youth mental health

AgentsDGX agent

arXiv:2604.13381v1 Announce Type: cross Abstract: Conversational generative artificial intelligence agents (or genAI chatbots) could benefit youth mental health, yet young people's perspectives remain

15 Apr 2026

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

A document is worth a structured record: Principled inductive bias design for document recognition

SafetyDGX agent

arXiv:2507.08458v2 Announce Type: replace-cross Abstract: Many document types use intrinsic, convention-driven structures that serve to encode precise and structured information, such as the conventio

A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators

Model ReleasesDGX agent

arXiv:2603.27557v2 Announce Type: replace-cross Abstract: In this paper, we analyze two main factors of Bonafide Resource (BR) or AI-based Generator (AG) which affect the performance and the generalit

A hierarchical spatial-aware algorithm with efficient reinforcement learning for human-robot task planning and allocation in production

AgentsDGX agent

arXiv:2604.12669v1 Announce Type: new Abstract: In advanced manufacturing systems, humans and robots collaborate to conduct the production process. Effective task planning and allocation (TPA) is cruc

A Layer-wise Analysis of Supervised Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.11838v1 Announce Type: cross Abstract: While critical for alignment, Supervised Fine-Tuning (SFT) incurs the risk of catastrophic forgetting, yet the layer-wise emergence of instruction-fol

A longitudinal health agent framework

SafetyDGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

A Scoping Review of Large Language Model-Based Pedagogical Agents

AgentsDGX agent

arXiv:2604.12253v1 Announce Type: new Abstract: This scoping review examines the emerging field of Large Language Model (LLM)-based pedagogical agents in educational settings. While traditional pedago

A Survey of Multimodal Mathematical Reasoning: From Perception, Alignment to Reasoning

SafetyDGX agent

arXiv:2603.08291v3 Announce Type: replace Abstract: Multimodal Mathematical Reasoning (MMR) has recently attracted increasing attention for its capability to solve mathematical problems involving both

A Two-Stage LLM Framework for Accessible and Verified XAI Explanations

ApplicationsDGX agent

arXiv:2604.12543v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to translate the technical outputs of eXplainable Artificial Intelligence (XAI) methods into accessib

A2-DIDM: Privacy-preserving Accumulator-enabled Auditing for Distributed Identity of DNN Model

ResearchDGX agent

arXiv:2405.04108v2 Announce Type: replace-cross Abstract: Recent booming development of Generative Artificial Intelligence (GenAI) has facilitated model commercialization to reinforce the model perfor

AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought

ResearchDGX agent

arXiv:2501.16154v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive multilingual capabilities through pretraining on diverse corpora. Although these models sho

Adaptive Domain Models: Bayesian Evolution, Warm Rotation, and Principled Training for Geometric and Neuromorphic AI

ResearchDGX agent

arXiv:2603.18104v3 Announce Type: replace Abstract: Prevailing AI training infrastructure assumes reverse-mode automatic differentiation over IEEE-754 arithmetic. The memory overhead of training relat

Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents

AgentsDGX agent

arXiv:2604.12129v1 Announce Type: new Abstract: The transition from stateless model inference to stateful agentic execution is reshaping the systems assumptions underlying modern AI infrastructure. Wh

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

Model ReleasesDGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

← Previous
1…325326327328329…354
Next →