AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
10 Apr 2026

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

SafetyDGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

ApplicationsDGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2501.16150v3 Announce Type: replace Abstract: Agents for computer use (ACUs) are an emerging class of systems capable of executing complex tasks on digital devices -- such as desktops, mobile ph

A First Guess is Rarely the Final Answer: Learning to Search in the Travelling Salesperson Problem

SafetyDGX agent

arXiv:2604.06940v1 Announce Type: cross Abstract: Most neural solvers for the Traveling Salesperson Problem (TSP) are trained to output a single solution, even though practitioners rarely stop there:

A Goal-Oriented Chatbot for Engaging the Elderly Through Family Photo Conversations

ResearchDGX agent

arXiv:2604.06184v1 Announce Type: cross Abstract: We propose a personalized chatbot designed for elderly individuals. The chatbot initiates discussions based on family photos, encouraging users to int

A Graph-Enhanced Defense Framework for Explainable Fake News Detection with LLM

ResearchDGX agent

arXiv:2604.06666v1 Announce Type: cross Abstract: Explainable fake news detection aims to assess the veracity of news claims while providing human-friendly explanations. Existing methods incorporating

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

HardwareDGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

A-MBER: Affective Memory Benchmark for Emotion Recognition

Model ReleasesDGX agent

arXiv:2604.07017v1 Announce Type: new Abstract: AI assistants that interact with users over time need to interpret the user's current emotional state in order to respond appropriately and personally.

A Novel Automatic Framework for Speaker Drift Detection in Synthesized Speech

Model ReleasesDGX agent

arXiv:2604.06327v1 Announce Type: cross Abstract: Recent diffusion-based text-to-speech (TTS) models achieve high naturalness and expressiveness, yet often suffer from speaker drift, a subtle, gradual

A Parameter-Efficient Transfer Learning Approach through Multitask Prompt Distillation and Decomposition for Clinical NLP

Model ReleasesDGX agent

arXiv:2604.06650v1 Announce Type: cross Abstract: Existing prompt-based fine-tuning methods typically learn task-specific prompts independently, imposing significant computing and storage overhead at

A Severity-Based Curriculum Learning Strategy for Arabic Medical Text Generation

TutorialsDGX agent

arXiv:2604.06365v1 Announce Type: cross Abstract: Arabic medical text generation is increasingly needed to help users interpret symptoms and access general health guidance in their native language. Ne

A Study of LLMs' Preferences for Libraries and Programming Languages

ResearchDGX agent

arXiv:2503.17181v3 Announce Type: replace-cross Abstract: Despite the rapid progress of large language models (LLMs) in code generation, existing evaluations focus on functional correctness or syntact

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

Model ReleasesDGX agent

arXiv:2604.07274v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from

AdaProb: Efficient Machine Unlearning via Adaptive Probability

ResearchDGX agent

arXiv:2411.02622v3 Announce Type: replace-cross Abstract: Machine unlearning, enabling a trained model to forget specific data, is crucial for addressing erroneous data and adhering to privacy regulat

Adaptive Differential Privacy for Federated Medical Image Segmentation Across Diverse Modalities

ApplicationsDGX agent

arXiv:2604.06518v1 Announce Type: cross Abstract: Large volumes of medical data remain underutilized because centralizing distributed data is often infeasible due to strict privacy regulations and ins

Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2512.10510v2 Announce Type: replace-cross Abstract: Offline-to-Online Reinforcement Learning (O2O RL) faces a critical dilemma in balancing the use of a fixed offline dataset with newly collecte

AEROS: A Single-Agent Operating Architecture with Embodied Capability Modules

SafetyDGX agent

arXiv:2604.07039v1 Announce Type: cross Abstract: Robotic systems lack a principled abstraction for organizing intelligence, capabilities, and execution in a unified manner. Existing approaches either

AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power

SafetyDGX agent

arXiv:2604.07007v1 Announce Type: cross Abstract: Autonomous AI agents are beginning to operate across organizational boundaries on the open internet -- discovering, transacting with, and delegating t

AgentGate: A Lightweight Structured Routing Engine for the Internet of Agents

Model ReleasesDGX agent

arXiv:2604.06696v1 Announce Type: new Abstract: The rapid development of AI agent systems is leading to an emerging Internet of Agents, where specialized agents operate across local devices, edge node

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

Model ReleasesDGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

AI-Driven Research for Databases

SafetyDGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

An Automated Survey of Generative Artificial Intelligence: Large Language Models, Architectures, Protocols, and Applications

Model ReleasesDGX agent

arXiv:2306.02781v4 Announce Type: replace-cross Abstract: Generative artificial intelligence, and large language models in particular, have emerged as one of the most transformative paradigms in moder

An empirical study of LoRA-based fine-tuning of large language models for automated test case generation

Model ReleasesDGX agent

arXiv:2604.06946v1 Announce Type: cross Abstract: Automated test case generation from natural language requirements remains a challenging problem in software engineering due to the ambiguity of requir

Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes

TutorialsDGX agent

arXiv:2410.22177v2 Announce Type: replace-cross Abstract: As more applications of large language models (LLMs) for 3D content for immersive environments emerge, it is crucial to study user behaviour t

Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions

SafetyDGX agent

arXiv:2604.07277v1 Announce Type: cross Abstract: Online reinforcement learning (RL) serves as an effective method for enhancing the capabilities of Android agents. However, guiding agents to learn th

Asking like Socrates: Socrates helps VLMs understand remote sensing images

Model ReleasesDGX agent

arXiv:2511.22396v2 Announce Type: replace-cross Abstract: Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (

Assessing the Added Value of Onboard Earth Observation Processing with the IRIDE HEO Service Segment

AgentsDGX agent

arXiv:2604.07120v1 Announce Type: cross Abstract: Current operational Earth Observation (EO) services, including the Copernicus Emergency Management Service (CEMS), the European Forest Fire Informatio

ATANT: An Evaluation Framework for AI Continuity

Model ReleasesDGX agent

arXiv:2604.06710v1 Announce Type: new Abstract: We present ATANT (Automated Test for Acceptance of Narrative Truth), an open evaluation framework for measuring continuity in AI systems: the ability to

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis

Model ReleasesDGX agent

arXiv:2604.02022v2 Announce Type: replace Abstract: Evaluating the safety of LLM-based agents is increasingly important because risks in realistic deployments often emerge over multi-step interactions

Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries

SafetyDGX agent

arXiv:2604.06416v1 Announce Type: cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We ev

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

TutorialsDGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

SafetyDGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

Automating Database-Native Function Code Synthesis with LLMs

Model ReleasesDGX agent

arXiv:2604.06231v1 Announce Type: cross Abstract: Database systems incorporate an ever-growing number of functions in their kernels (a.k.a., database native functions) for scenarios like new applicati

AutoReproduce: Automatic AI Experiment Reproduction with Paper Lineage

Model ReleasesDGX agent

arXiv:2505.20662v3 Announce Type: replace Abstract: Efficient reproduction of research papers is pivotal to accelerating scientific progress. However, the increasing complexity of proposed methods oft

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views

Model ReleasesDGX agent

arXiv:2604.07041v1 Announce Type: cross Abstract: Text-to-SQL is the task of translating natural language queries into executable SQL for a given database, enabling non-expert users to access structur

BadImplant: Injection-based Multi-Targeted Graph Backdoor Attack

ResearchDGX agent

arXiv:2601.15474v2 Announce Type: replace-cross Abstract: Graph neural network (GNN) have demonstrated exceptional performance in solving critical problems across diverse domains yet remain susceptibl

BDI-Kit Demo: A Toolkit for Programmable and Conversational Data Harmonization

ResearchDGX agent

arXiv:2604.06405v1 Announce Type: new Abstract: Data harmonization remains a major bottleneck for integrative analysis due to heterogeneity in schemas, value representations, and domain-specific conve

Before Humans Join the Team: Diagnosing Coordination Failures in Healthcare Robot Team Simulation

SafetyDGX agent

arXiv:2508.04691v2 Announce Type: replace-cross Abstract: As humans move toward collaborating with coordinated robot teams, understanding how these teams coordinate and fail is essential for building

Before We Trust Them: Decision-Making Failures in Navigation of Foundation Models

Model ReleasesDGX agent

arXiv:2601.05529v5 Announce Type: replace Abstract: High success rates on navigation-related tasks do not necessarily translate into reliable decision making by foundation models. To examine this gap,

Benchmarking LLM Tool-Use in the Wild

Model ReleasesDGX agent

arXiv:2604.06185v1 Announce Type: cross Abstract: Fulfilling user needs through Large Language Model multi-turn, multi-step tool-use is rarely a straightforward process. Real user interactions are inh

Between Century and Poet: Graph-Based Lexical Semantic Change in Persian Poetry

ResearchDGX agent

arXiv:2604.06674v1 Announce Type: cross Abstract: Meaning in Persian poetry is both historical and relational. Words persist through literary tradition while shifting their force through changing cons

Beyond Case Law: Evaluating Structure-Aware Retrieval and Safety in Statute-Centric Legal QA

Model ReleasesDGX agent

arXiv:2604.06173v1 Announce Type: cross Abstract: Legal QA benchmarks have predominantly focused on case law, overlooking the unique challenges of statute-centric regulatory reasoning. In statutory do

Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06201v1 Announce Type: cross Abstract: While most reading comprehension benchmarks for LLMs focus on factual information that can be answered by localizing specific textual evidence, many r

Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects

AgentsDGX agent

arXiv:2604.06373v1 Announce Type: cross Abstract: New generation of AI coding tools, including AI-powered IDEs equipped with agentic capabilities, can generate code within the context of the project.

Beyond Surface Judgments: Human-Grounded Risk Evaluation of LLM-Generated Disinformation

SafetyDGX agent

arXiv:2604.06820v1 Announce Type: new Abstract: Large language models (LLMs) can generate persuasive narratives at scale, raising concerns about their potential use in disinformation campaigns. Assess

Bi-Level Optimization for Single Domain Generalization

ResearchDGX agent

arXiv:2604.06349v1 Announce Type: cross Abstract: Generalizing from a single labeled source domain to unseen target domains, without access to any target data during training, remains a fundamental ch

BiScale-GTR: Fragment-Aware Graph Transformers for Multi-Scale Molecular Representation Learning

Model ReleasesDGX agent

arXiv:2604.06336v1 Announce Type: cross Abstract: Graph Transformers have recently attracted attention for molecular property prediction by combining the inductive biases of graph neural networks (GNN

Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses

Model ReleasesDGX agent

arXiv:2604.06216v1 Announce Type: cross Abstract: As LLM-powered chatbots are increasingly deployed in mental health services, detecting hallucinations and omissions has become critical for user safet

Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules

Model ReleasesDGX agent

arXiv:2604.06233v1 Announce Type: new Abstract: Safety-trained language models routinely refuse requests for help circumventing rules. But not all rules deserve compliance. When users ask for help eva

Blockchain and AI: Securing Intelligent Networks for the Future

AgentsDGX agent

arXiv:2604.06323v2 Announce Type: cross Abstract: Blockchain and artificial intelligence (AI) are increasingly proposed together for securing intelligent networks, but the literature remains fragmente

Bridging MRI and PET physiology: Untangling complementarity through orthogonal representations

ResearchDGX agent

arXiv:2604.07154v1 Announce Type: cross Abstract: Multimodal imaging analysis often relies on joint latent representations, yet these approaches rarely define what information is shared versus modalit

Bridging Natural Language and Microgrid Dynamics: A Context-Aware Simulator and Dataset

ApplicationsDGX agent

arXiv:2604.05429v2 Announce Type: replace-cross Abstract: Addressing the critical need for intelligent, context-aware energy management in renewable systems, we introduce the OpenCEM Simulator and Dat

Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code

Model ReleasesDGX agent

arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u

CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models

ResearchDGX agent

arXiv:2604.06987v1 Announce Type: cross Abstract: Palmprint recognition is deployed in security-critical applications, including access control and palm-based payment, due to its contactless acquisiti

CADENCE: Context-Adaptive Depth Estimation for Navigation and Computational Efficiency

Model ReleasesDGX agent

arXiv:2604.07286v1 Announce Type: cross Abstract: Autonomous vehicles deployed in remote environments typically rely on embedded processors, compact batteries, and lightweight sensors. These hardware

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

SafetyDGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning

AgentsDGX agent

arXiv:2510.18034v2 Announce Type: replace-cross Abstract: Autonomous driving systems remain critically vulnerable to the long-tail of rare, out-of-distribution semantic anomalies. While VLMs have emer

Chatbot-Based Assessment of Code Understanding in Automated Programming Assessment Systems

Local AiDGX agent

arXiv:2604.07304v1 Announce Type: cross Abstract: Large Language Models (LLMs) challenge conventional automated programming assessment because students can now produce functionally correct code withou

ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understanding

ApplicationsDGX agent

arXiv:2604.06685v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have demonstrated significant potential in chemical visual understanding, current models are predominantly optimiz

ChopGrad: Pixel-Wise Losses for Latent Video Diffusion via Truncated Backpropagation

ResearchDGX agent

arXiv:2603.17812v2 Announce Type: replace-cross Abstract: Recent video diffusion models achieve high-quality generation through recurrent frame processing where each frame generation depends on previo

← Previous
1…343344345346347…350
Next →