AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
14 Apr 2026

Attention-Guided Flow-Matching for Sparse 3D Geological Generation

ResearchDGX agent

arXiv:2604.09700v1 Announce Type: cross Abstract: Constructing high-resolution 3D geological models from sparse 1D borehole and 2D surface data is a highly ill-posed inverse problem. Traditional heuri

Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music

Model ReleasesDGX agent

arXiv:2604.10905v1 Announce Type: cross Abstract: We present Audio Flamingo Next (AF-Next), the next-generation and most capable large audio-language model in the Audio Flamingo series, designed to ad

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

Auto-regressive transformation for image alignment

SafetyDGX agent

arXiv:2505.04864v2 Announce Type: replace-cross Abstract: Existing methods for image alignment struggle in cases involving feature-sparse regions, extreme scale and field-of-view differences, and larg

Automating Structural Analysis Across Multiple Software Platforms Using Large Language Models

AgentsDGX agent

arXiv:2604.09866v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have shown the promise to significantly accelerate the workflow by automating structural modeling and

AutoMS: Multi-Agent Evolutionary Search for Cross-Physics Inverse Microstructure Design

Model ReleasesDGX agent

arXiv:2603.27195v2 Announce Type: replace Abstract: Designing microstructures with coupled cross-physics objectives is a fundamental challenge where traditional topology optimization is often computat

AXIL: Exact Instance Attribution for Gradient Boosting

ResearchDGX agent

arXiv:2301.01864v2 Announce Type: replace-cross Abstract: We derive an exact, prediction-specific instance-attribution method for fitted gradient boosting machines (GBMs) trained with squared-error lo

Back to the Barn with LLAMAs: Evolving Pretrained LLM Backbones in Finetuning Vision Language Models

Model ReleasesDGX agent

arXiv:2604.10985v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have rapidly advanced by leveraging powerful pre-trained Large Language Models (LLMs) as core reasoning backbones. As new

Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward

SafetyDGX agent

arXiv:2604.09748v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an emerging paradigm that significantly boosts a Large Language Model's (LLM's) reasoning abi

bacpipe: a Python package to make bioacoustic deep learning models accessible

ResearchDGX agent

arXiv:2604.11560v1 Announce Type: cross Abstract: 1. Natural sounds have been recorded for millions of hours over the previous decades using passive acoustic monitoring. Improvements in deep learning

BankerToolBench: Evaluating AI Agents in End-to-End Investment Banking Workflows

Model ReleasesDGX agent

arXiv:2604.11304v1 Announce Type: new Abstract: Existing AI benchmarks lack the fidelity to assess economically meaningful progress on professional workflows. To evaluate frontier AI agents in a high-

Belief-Aware VLM Model for Human-like Reasoning

SafetyDGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

Beyond A Fixed Seal: Adaptive Stealing Watermark in Large Language Models

ApplicationsDGX agent

arXiv:2604.10893v1 Announce Type: cross Abstract: Watermarking provides a critical safeguard for large language model (LLM) services by facilitating the detection of LLM-generated text. Correspondingl

Beyond Compliance: A Resistance-Informed Motivation Reasoning Framework for Challenging Psychological Client Simulation

SafetyDGX agent

arXiv:2604.10507v1 Announce Type: new Abstract: Psychological client simulators have emerged as a scalable solution for training and evaluating counselor trainees and psychological LLMs. Yet existing

Beyond Fluency: Toward Reliable Trajectories in Agentic IR

AgentsDGX agent

arXiv:2604.04269v2 Announce Type: replace Abstract: Information Retrieval is shifting from passive document ranking toward autonomous agentic workflows that operate in multi-step Reason-Act-Observe lo

Beyond LLMs, Sparse Distributed Memory, and Neuromorphics <A Hyper-Dimensional SRAM-CAM 'VaCoAl' for Ultra-High Speed, Ultra-Low Power, and Low Cost>

ResearchDGX agent

arXiv:2604.11665v1 Announce Type: cross Abstract: This paper reports an unexpected finding: in a deterministic hyperdimensional computing (HDC) architecture based on Galois-field algebra, a path-depen

Beyond Matching to Tiles: Bridging Unaligned Aerial and Satellite Views for Vision-Only UAV Navigation

Model ReleasesDGX agent

arXiv:2603.22153v3 Announce Type: replace-cross Abstract: Recent advances in cross-view geo-localization (CVGL) methods have shown strong potential for supporting unmanned aerial vehicle (UAV) navigat

Beyond Message Passing: A Semantic View of Agent Communication Protocols

SafetyDGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

SafetyDGX agent

arXiv:2604.10367v1 Announce Type: new Abstract: Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation

Beyond Offline A/B Testing: Context-Aware Agent Simulation for Recommender System Evaluation

AgentsDGX agent

arXiv:2604.09549v1 Announce Type: cross Abstract: Recommender systems are central to online services, enabling users to navigate through massive amounts of content across various domains. However, the

Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation

AgentsDGX agent

arXiv:2602.02007v3 Announce Type: replace-cross Abstract: Agent memory systems often adopt the standard Retrieval-Augmented Generation (RAG) pipeline, yet its underlying assumptions differ in this set

Beyond RAG for Cyber Threat Intelligence: A Systematic Evaluation of Graph-Based and Agentic Retrieval

AgentsDGX agent

arXiv:2604.11419v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) analysts must answer complex questions over large collections of narrative security reports. Retrieval-augmented generat

Beyond Statistical Co-occurrence: Unlocking Intrinsic Semantics for Tabular Data Clustering

Model ReleasesDGX agent

arXiv:2604.10865v1 Announce Type: new Abstract: Deep Clustering (DC) has emerged as a powerful tool for tabular data analysis in real-world domains like finance and healthcare. However, most existing

Beyond Theory of Mind in Robotics

ResearchDGX agent

arXiv:2604.09612v1 Announce Type: new Abstract: Theory of Mind, the capacity to explain and predict behavior by inferring hidden mental states, has become the dominant paradigm for social interaction

BiCLIP: Domain Canonicalization via Structured Geometric Transformation

Model ReleasesDGX agent

arXiv:2603.08942v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities, yet adapting these models to specialized

Bottleneck Tokens for Unified Multimodal Retrieval

ResearchDGX agent

arXiv:2604.11095v1 Announce Type: cross Abstract: Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rel

BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning

ResearchDGX agent

arXiv:2604.11136v1 Announce Type: cross Abstract: Object-level spatial-temporal understanding is essential for video question answering, yet existing multimodal large language models (MLLMs) encode fr

BridgeSim: Unveiling the OL-CL Gap in End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2604.10856v1 Announce Type: cross Abstract: Open-loop (OL) to closed-loop (CL) gap (OL-CL gap) exists when OL-pretrained policies scoring high in OL evaluations fail to transfer effectively in c

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

SafetyDGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

Brief2Design: A Multi-phased, Compositional Approach to Prompt-based Graphic Design

ResearchDGX agent

arXiv:2604.11019v1 Announce Type: cross Abstract: Professional designers work from client briefs that specify goals and constraints but often lack concrete design details. Translating these abstract r

Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning

ResearchDGX agent

arXiv:2604.10701v1 Announce Type: cross Abstract: Credit assignment is a central challenge in reinforcement learning (RL). Classical actor-critic methods address this challenge through fine-grained ad

Budget-Aware Uncertainty for Radiotherapy Segmentation QA Using nnU-Net

SafetyDGX agent

arXiv:2604.11798v1 Announce Type: cross Abstract: Accurate delineation of the Clinical Target Volume (CTV) is essential for radiotherapy planning, yet remains time-consuming and difficult to assess, e

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

Model ReleasesDGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

C2F-Thinker: Coarse-to-Fine Reasoning with Hint-Guided Reinforcement Learning for Multimodal Sentiment Analysis

SafetyDGX agent

arXiv:2604.00013v2 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to integrate textual, acoustic, and visual information for deep emotional understanding. Despite the progre

CAGE: Bridging the Accuracy-Aesthetics Gap in Educational Diagrams via Code-Anchored Generative Enhancement

ResearchDGX agent

arXiv:2604.09691v1 Announce Type: cross Abstract: Educational diagrams -- labeled illustrations of biological processes, chemical structures, physical systems, and mathematical concepts -- are essenti

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

SafetyDGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

Camyla: Scaling Autonomous Research in Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

Can Large Language Models Infer Causal Relationships from Real-World Text?

Model ReleasesDGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

HardwareDGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

TutorialsDGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

Model ReleasesDGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

CASK: Core-Aware Selective KV Compression for Reasoning Traces

HardwareDGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

Causally Sufficient and Necessary Feature Expansion for Class-Incremental Learning

TutorialsDGX agent

arXiv:2603.09145v2 Announce Type: replace-cross Abstract: Current expansion-based methods for Class Incremental Learning (CIL) effectively mitigate catastrophic forgetting by freezing old features. Ho

CFMS: A Coarse-to-Fine Multimodal Synthesis Framework for Enhanced Tabular Reasoning

TutorialsDGX agent

arXiv:2604.10973v1 Announce Type: new Abstract: Reasoning over tabular data is a crucial capability for tasks like question answering and fact verification, as it requires models to comprehend both fr

CHAIRO: Contextual Hierarchical Analogical Induction and Reasoning Optimization for LLMs

ApplicationsDGX agent

arXiv:2604.10502v1 Announce Type: new Abstract: Content moderation in online platforms faces persistent challenges due to the evolving complexity of user-generated content and the limitations of tradi

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

HardwareDGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

Model ReleasesDGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

Model ReleasesDGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

CID-TKG: Collaborative Historical Invariance and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

SafetyDGX agent

arXiv:2604.09600v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to infer future facts at unseen timestamps from temporally evolving entities and relations. Despite recent

CircuitSynth: Reliable Synthetic Data Generation

ResearchDGX agent

arXiv:2604.10114v1 Announce Type: cross Abstract: The generation of high-fidelity synthetic data is a cornerstone of modern machine learning, yet Large Language Models (LLMs) frequently suffer from ha

Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems

Model ReleasesDGX agent

arXiv:2604.10305v1 Announce Type: cross Abstract: Cooperative perception allows connected vehicles and roadside infrastructure to share sensor observations, creating a fused scene representation beyon

ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection

Local AiDGX agent

arXiv:2604.11790v1 Announce Type: cross Abstract: Tool-augmented Large Language Model (LLM) agents have demonstrated impressive capabilities in automating complex, multi-step real-world tasks, yet rem

ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents

AgentsDGX agent

arXiv:2604.11784v1 Announce Type: cross Abstract: GUI agents drive applications through their visual interfaces instead of programmatic APIs, interacting with arbitrary software via taps, swipes, and

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

CLAY: Conditional Visual Similarity Modulation in Vision-Language Embedding Space

ResearchDGX agent

arXiv:2604.11539v1 Announce Type: cross Abstract: Human perception of visual similarity is inherently adaptive and subjective, depending on the users' interests and focus. However, most image retrieva

Closed-Form Concept Erasure via Double Projections

SafetyDGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

CocoaBench: Evaluating Unified Digital Agents in the Wild

Model ReleasesDGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

CodaRAG: Connecting the Dots with Associativity Inspired by Complementary Learning

ResearchDGX agent

arXiv:2604.10426v1 Announce Type: cross Abstract: Large Language Models (LLMs) struggle with knowledge-intensive tasks due to hallucinations and fragmented reasoning over dispersed information. While

CodeTracer: Towards Traceable Agent States

AgentsDGX agent

arXiv:2604.11641v1 Announce Type: cross Abstract: Code agents are advancing rapidly, but debugging them is becoming increasingly difficult. As frameworks orchestrate parallel tool calls and multi-stag

CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

Model ReleasesDGX agent

arXiv:2604.01687v2 Announce Type: replace Abstract: Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool

← Previous
1…332333334335336…354
Next →