AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
23 Apr 2026

A Vision-Language-Action Model for Adaptive Ultrasound-Guided Needle Insertion and Needle Tracking

Model ReleasesDGX agent

arXiv:2604.20347v1 Announce Type: cross Abstract: Ultrasound (US)-guided needle insertion is a critical yet challenging procedure due to dynamic imaging conditions and difficulties in needle visualiza

AAC: Admissible-by-Architecture Differentiable Landmark Compression for ALT

Model ReleasesDGX agent

arXiv:2604.20744v1 Announce Type: new Abstract: We introduce extbf{AAC} (Architecturally Admissible Compressor), a differentiable landmark-selection module for ALT (A*, Landmarks, and Triangle inequal

Accelerating PayPal's Commerce Agent with Speculative Decoding: An Empirical Study on EAGLE3 with Fine-Tuned Nemotron Models

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.19767v1 Announce Type: cross Abstract: We evaluate speculative decoding with EAGLE3 as an inference-time optimization for PayPal's Commerce Agent, powered by a fine-tuned llama3.1-nemotron-

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

Model ReleasesDGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

Adaptive Conformal Anomaly Detection with Time Series Foundation Models for Signal Monitoring

ApplicationsDGX agent

arXiv:2604.20122v1 Announce Type: cross Abstract: We propose a post-hoc adaptive conformal anomaly detection method for monitoring time series that leverages predictions from pre-trained foundation mo

AgentLens: Adaptive Visual Modalities for Human-Agent Interaction in Mobile GUI Agents

AgentsDGX agent

arXiv:2604.20279v1 Announce Type: cross Abstract: Mobile GUI agents can automate smartphone tasks by interacting directly with app interfaces, but how they should communicate with users during executi

AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation

SafetyDGX agent

arXiv:2604.20134v1 Announce Type: cross Abstract: Security Operations Centers (SOCs) increasingly encounter difficulties in correlating heterogeneous alerts, interpreting multi-stage attack progressio

Agnostic Language Identification and Generation

ResearchDGX agent

arXiv:2601.23258v2 Announce Type: replace-cross Abstract: Recent works on language identification and generation have established tight statistical rates at which these tasks can be achieved. These wo

AI models of unstable flow exhibit hallucination

SafetyDGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains

Model ReleasesDGX agent

arXiv:2604.19751v1 Announce Type: new Abstract: Generative AI is entering research, education, and professional work faster than current governance frameworks can specify how AI-assisted outputs shoul

Algorithm Selection with Zero Domain Knowledge via Text Embeddings

ResearchDGX agent

arXiv:2604.19753v1 Announce Type: new Abstract: We propose a feature-free approach to algorithm selection that replaces hand-crafted instance features with pretrained text embeddings. Our method, Zero

Analyzing Shapley Additive Explanations to Understand Anomaly Detection Algorithm Behaviors and Their Complementarity

ResearchDGX agent

arXiv:2602.00208v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection is a challenging problem due to the diversity of data distributions and the lack of labels. Ensemble methods ar

AnatomicalNets: A Multi-Structure Segmentation and Contour-Based Distance Estimation Pipeline for Clinically Grounded Lung Cancer T-Staging

ResearchDGX agent

arXiv:2511.19367v2 Announce Type: replace-cross Abstract: Accurate tumor staging in lung cancer is crucial for prognosis and treatment planning and is governed by explicit anatomical criteria under fi

Anchor-and-Resume Concession Under Dynamic Pricing for LLM-Augmented Freight Negotiation

Model ReleasesDGX agent

arXiv:2604.20732v1 Announce Type: cross Abstract: Freight brokerages negotiate thousands of carrier rates daily under dynamic pricing conditions where models frequently revise targets mid-conversation

AROMA: Augmented Reasoning Over a Multimodal Architecture for Virtual Cell Genetic Perturbation Modeling

ResearchDGX agent

arXiv:2604.20263v1 Announce Type: cross Abstract: Virtual cell modeling predicts molecular state changes under genetic perturbations in silico, which is essential for biological mechanism studies. How

Assessing the Robustness of Climate Foundation Models under No-Analog Distribution Shifts

Model ReleasesDGX agent

arXiv:2603.23043v2 Announce Type: replace-cross Abstract: The accelerating pace of climate change introduces profound non-stationarities that challenge the ability of Machine Learning based climate em

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

AgentsDGX agent

arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and

ATIR: Towards Audio-Text Interleaved Contextual Retrieval

Model ReleasesDGX agent

arXiv:2604.20267v1 Announce Type: cross Abstract: Audio carries richer information than text, including emotion, speaker traits, and environmental context, while also enabling lower-latency processing

Atomic Decision Boundaries: A Structural Requirement for Guaranteeing Execution-Time Admissibility in Autonomous Systems

SafetyDGX agent

arXiv:2604.17511v2 Announce Type: replace-cross Abstract: Autonomous systems increasingly execute actions that directly modify shared state, creating an urgent need for precise control over which tran

Auditing and Controlling AI Agent Actions in Spreadsheets

AgentsDGX agent

arXiv:2604.20070v1 Announce Type: cross Abstract: Advances in AI agent capabilities have outpaced users' ability to meaningfully oversee their execution. AI agents can perform sophisticated, multi-ste

Auto-Unrolled Proximal Gradient Descent: An AutoML Approach to Interpretable Waveform Optimization

ResearchDGX agent

arXiv:2603.17478v2 Announce Type: replace-cross Abstract: This study explores the combination of automated machine learning (AutoML) with model-based deep unfolding (DU) for optimizing wireless beamfo

AutoGraphAD: Unsupervised network anomaly detection using Variational Graph Autoencoders

ResearchDGX agent

arXiv:2511.17113v2 Announce Type: replace-cross Abstract: Network Intrusion Detection Systems (NIDS) are essential tools for detecting network attacks and intrusions. While extensive research has expl

Automated Detection of Dosing Errors in Clinical Trial Narratives: A Multi-Modal Feature Engineering Approach with LightGBM

Model ReleasesDGX agent

arXiv:2604.19759v1 Announce Type: new Abstract: Clinical trials require strict adherence to medication protocols, yet dosing errors remain a persistent challenge affecting patient safety and trial int

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

Model ReleasesDGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

AVISE: Framework for Evaluating the Security of AI Systems

Model ReleasesDGX agent

arXiv:2604.20833v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across critical domains, their security vulnerabilities pose growing risks of high-p

BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching

HardwareDGX agent

arXiv:2412.03594v3 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly play an important role in a wide range of information processing and management tasks in industry. M

Behavioral Transfer in AI Agents: Evidence and Privacy Implications

AgentsDGX agent

arXiv:2604.19925v1 Announce Type: cross Abstract: AI agents powered by large language models are increasingly acting on behalf of humans in social and economic environments. Prior research has focused

Benefits of Low-Cost Bio-Inspiration in the Age of Overparametrization

Model ReleasesDGX agent

arXiv:2604.20365v1 Announce Type: cross Abstract: While Central Pattern Generators (CPGs) and Multi-Layer Perceptrons (MLP) are widely used paradigms in robot control, few systematic studies have been

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

Model ReleasesDGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

Beyond ZOH: Advanced Discretization Strategies for Vision Mamba

ResearchDGX agent

arXiv:2604.20606v1 Announce Type: cross Abstract: Vision Mamba, as a state space model (SSM), employs a zero-order hold (ZOH) discretization, which assumes that input signals remain constant between s

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

SafetyDGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

Model ReleasesDGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

Can 'AI' Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

Model ReleasesDGX agent

arXiv:2604.20791v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in healthcare, yet their communicative alignment with clinical standards remains insufficiently

Can LLMs Infer Conversational Agent Users' Personality Traits from Chat History?

AgentsDGX agent

arXiv:2604.19785v1 Announce Type: cross Abstract: Sensitive information, such as knowledge about an individual's personality, can be can be misused to influence behavior (e.g., via personalized messag

Can We Locate and Prevent Stereotypes in LLMs?

Model ReleasesDGX agent

arXiv:2604.19764v1 Announce Type: cross Abstract: Stereotypes in large language models (LLMs) can perpetuate harmful societal biases. Despite the widespread use of models, little is known about where

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

Model ReleasesDGX agent

arXiv:2603.28032v2 Announce Type: replace-cross Abstract: The convergence of low-altitude economies, embodied intelligence, and air-ground cooperative systems creates growing demand for simulation inf

CAST: Achieving Stable LLM-based Text Analysis for Data Analytics

SafetyDGX agent

arXiv:2602.15861v2 Announce Type: replace-cross Abstract: Text analysis of tabular data relies on two core operations: summarization for corpus-level theme extraction and tagging for row-level labelin

Catalyzing Informed Residential Energy Retrofit Decisions via Domain-Specific LLM

Model ReleasesDGX agent

arXiv:2602.20181v2 Announce Type: replace-cross Abstract: Residential energy retrofit initiation is often stalled by an expertise gap, where homeowners lack the technical literacy required for structu

Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?

ResearchDGX agent

arXiv:2502.07963v4 Announce Type: replace-cross Abstract: Medical research faces well-documented challenges in translating novel treatments into clinical practice. Publishing incentives encourage rese

CEDAR: Context Engineering for Agentic Data Science

Local AiDGX agent

arXiv:2601.06606v2 Announce Type: replace-cross Abstract: We demonstrate CEDAR, an application for automating data science (DS) tasks with an agentic setup. Solving DS problems with LLMs is an underex

Centering Ecological Goals in Automated Identification of Individual Animals

ResearchDGX agent

arXiv:2604.20626v1 Announce Type: cross Abstract: Recognizing individual animals over time is central to many ecological and conservation questions, including estimating abundance, survival, movement,

CHASM: Unveiling Covert Advertisements on Chinese Social Media

ApplicationsDGX agent

arXiv:2604.20511v1 Announce Type: cross Abstract: Current benchmarks for evaluating large language models (LLMs) in social media moderation completely overlook a serious threat: covert advertisements,

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

SafetyDGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

CHORUS: An Agentic Framework for Generating Realistic Deliberation Data

AgentsDGX agent

arXiv:2604.20651v1 Announce Type: new Abstract: Understanding the intricate dynamics of online discourse depends on large-scale deliberation data, a resource that remains scarce across interactive web

Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation

ResearchDGX agent

arXiv:2604.19826v1 Announce Type: cross Abstract: AI coding assistants increasingly generate code alongside tests. How developers structure test code, whether inline with the implementation or in sepa

CoAuthorAI: A Human in the Loop System For Scientific Book Writing

ResearchDGX agent

arXiv:2604.19772v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in scientific writing but struggle with book-length tasks, often producing inconsistent structure a

CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment

SafetyDGX agent

arXiv:2510.18471v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) excel at code generation by learning from vast code corpora, a fundamental semantic gap remains between the

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

Model ReleasesDGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

Cognis: Context-Aware Memory for Conversational AI Agents

ApplicationsDGX agent

arXiv:2604.19771v1 Announce Type: cross Abstract: LLM agents lack persistent memory, causing conversations to reset each session and preventing personalization over time. We present Lyzr Cognis, a uni

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

SafetyDGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

Model ReleasesDGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

Combo-Gait: Unified Transformer Framework for Multi-Modal Gait Recognition and Attribute Analysis

TutorialsDGX agent

arXiv:2510.10417v2 Announce Type: replace-cross Abstract: Gait recognition is an important biometric for human identification at a distance, particularly under low-resolution or unconstrained environm

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

Model ReleasesDGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

Computing the Reachability Value of Posterior-Deterministic POMDPs

ResearchDGX agent

arXiv:2602.07473v2 Announce Type: replace Abstract: Partially observable Markov decision processes (POMDPs) are a fundamental model for sequential decision-making under uncertainty. However, many veri

Context Attribution with Multi-Armed Bandit Optimization

ResearchDGX agent

arXiv:2506.19977v2 Announce Type: replace Abstract: Understanding which parts of the retrieved context contribute to a large language model's generated answer is essential for building interpretable a

Convergent Evolution: How Different Language Models Learn Similar Number Representations

TutorialsDGX agent

arXiv:2604.20817v1 Announce Type: cross Abstract: Language models trained on natural text learn to represent numbers using periodic features with dominant periods at T=2, 5, 10. In this paper, we iden

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment

ApplicationsDGX agent

arXiv:2604.20246v1 Announce Type: cross Abstract: Industrial robotic manipulation demands reliable long-horizon execution across embodiments, tasks, and changing object distributions. While Vision-Lan

Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation

SafetyDGX agent

arXiv:2604.20763v1 Announce Type: cross Abstract: Retrieval quality is the primary bottleneck for accuracy and robustness in retrieval-augmented generation (RAG). Current evaluation relies on heuristi

CreativeGame:Toward Mechanic-Aware Creative Game Generation

AgentsDGX agent

arXiv:2604.19926v1 Announce Type: new Abstract: Large language models can generate plausible game code, but turning this capability into iterative creative improvement remains difficult. In practice,

Cross-Modal Taxonomic Generalization in (Vision-) Language Models

ApplicationsDGX agent

arXiv:2603.07474v2 Announce Type: replace-cross Abstract: What is the interplay between semantic representations learned by language models (LM) from surface form alone to those learned from more grou

← Previous
1…311312313314315…354
Next →