AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

From Passive Reuse to Active Reasoning: Grounding Large Language Models for Neuro-Symbolic Experience Replay

SafetyDGX agent

arXiv:2605.09419v1 Announce Type: new Abstract: While experience replay is essential for data efficiency in reinforcement learning (RL), standard methods treat the replay buffer as a passive memory sy

From Single-Step Edit Response to Multi-Step Molecular Optimization

Local AiDGX agent

arXiv:2605.10035v1 Announce Type: new Abstract: Conditional molecular optimization aims to edit a molecule to realize a specified property shift. In practice, structurally similar molecule data is sca

From Spark to Fire: Modeling and Mitigating Error Cascades in LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2603.04474v2 Announce Type: replace-cross Abstract: Large Language Model-based Multi-Agent Systems (LLM-MAS) are increasingly applied to complex collaborative scenarios. However, their collabora


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Traditional Taggers to LLMs: A Comparative Study of POS Tagging for Medieval Romance Languages

Model ReleasesDGX agent

arXiv:2605.09147v1 Announce Type: cross Abstract: Part-of-speech (POS) tagging for Medieval Romance languages remains challenging due to orthographic variation, morphological complexity, and limited a

Functional Stable Model Semantics and Answer Set Programming Modulo Theories

ResearchDGX agent

arXiv:2605.09524v1 Announce Type: new Abstract: Recently there has been an increasing interest in incorporating ``intensional'' functions in answer set programming. Intensional functions are those who

Functional Subspace, where language models can use vector algebra to solve problems

ResearchDGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

G-Zero: Self-Play for Open-Ended Generation from Zero Data

AgentsDGX agent

arXiv:2605.09959v1 Announce Type: cross Abstract: Self-evolving LLMs excel in verifiable domains but struggle in open-ended tasks, where reliance on proxy LLM judges introduces capability bottlenecks

Gate-and-Merge: Zero-shot Compositional Personalization of Vision Language Models

ResearchDGX agent

arXiv:2605.08702v1 Announce Type: cross Abstract: This paper tackles compositional personalization of vision-language models (VLMs). In this problem, multiple user-defined concepts must be recognized

GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents

AgentsDGX agent

arXiv:2510.13896v2 Announce Type: replace-cross Abstract: Cellular image segmentation is essential for quantitative biology yet remains difficult due to heterogeneous modalities, morphological variabi

Gender Fairness in Audio Deepfake Detection: Performance and Disparity Analysis

SafetyDGX agent

arXiv:2603.09007v2 Announce Type: replace-cross Abstract: Audio deepfake detection aims to detect real human voices from those generated by Artificial Intelligence (AI) and has emerged as a significan

General Agent Evaluation

Model ReleasesDGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

Generalization Bounds of Emergent Communications for Agentic AI Networking

AgentsDGX agent

arXiv:2605.08613v1 Announce Type: new Abstract: The evolution of 6G networking toward agentic AI networking (AgentNet) systems requires a shift from traditional data pipelines to task-aware, agentic A

Generalized Category Discovery in Federated Graph Learning

SafetyDGX agent

arXiv:2605.08178v1 Announce Type: cross Abstract: Federated Graph Learning (FGL) enables collaborative learning over distributed graph data, yet existing approaches largely rely on a closed-world assu

Generating Leakage-Free Benchmarks for Robust RAG Evaluation

Model ReleasesDGX agent

arXiv:2605.08838v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) is widely used to augment large language models (LLMs) with external knowledge. However, many benchmark datasets,

Generative AI Fuels Solo Entrepreneurship, but Teams Still Lead at the Top

Model ReleasesDGX agent

arXiv:2605.10291v1 Announce Type: cross Abstract: Recent advances in generative artificial intelligence (AI) are reshaping who enters entrepreneurship, but not who reaches the top of the quality distr

Generative Experiences for Digital Mental Health Interventions: Evidence from a Randomized Study

TutorialsDGX agent

arXiv:2604.07558v2 Announce Type: replace-cross Abstract: Digital mental health (DMH) tools have extensively explored personalization of interventions to users' needs and contexts. However, this perso

Geometric 4D Stitching for Grounded 4D Generation

HardwareDGX agent

arXiv:2605.09984v1 Announce Type: cross Abstract: Recent 4D generation methods complete scene-level missing information using generative models and reconstruct the scene into radiance-based representa

Geometrically Constrained Stenosis Editing in Coronary Angiography via Entropic Optimal Transport

Model ReleasesDGX agent

arXiv:2605.08851v1 Announce Type: cross Abstract: The scarcity of high-quality imaging data for coronary angiography (CAG) stenosis limits the clinical translation of automated stenosis detection. Syn

Geometry Guided Self-Consistency for Physical AI

ResearchDGX agent

arXiv:2605.08638v1 Announce Type: cross Abstract: State-of-the-art physical AI models generate a chunk of actions per inference through diffusion or flow matching, iteratively refining an initial nois

Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model

Model ReleasesDGX agent

arXiv:2605.10739v1 Announce Type: cross Abstract: We introduce SMART-HC-VQA, a Sentinel-2-based visual question answering dataset derived from the IARPA SMART Heavy Construction dataset, designed for

GESR: A Genetic Programming-Based Symbolic Regression Method with Gene Editing

TutorialsDGX agent

arXiv:2605.10685v1 Announce Type: new Abstract: Mathematical formulas serve as a language through which humans communicate with nature. Discovering mathematical laws from scientific data to describe n

GLAI: GreenLightningAI for Accelerated Training through Knowledge Decoupling

ResearchDGX agent

arXiv:2510.00883v2 Announce Type: replace-cross Abstract: In this work we introduce GreenLightningAI (GLAI), a new architectural block designed as an alternative to conventional MLPs. The central idea

GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction

Model ReleasesDGX agent

arXiv:2605.09973v1 Announce Type: cross Abstract: Reliable detection of personally identifiable information (PII) is increasingly important across modern data-processing systems, yet the task remains

GNN for Structural Displacement Prediction

SafetyDGX agent

arXiv:2605.08303v1 Announce Type: cross Abstract: Accurate prediction of structural displacements under external loading is fundamental to structural health monitoring and seismic safety assessment. A

Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect

SafetyDGX agent

arXiv:2605.05248v2 Announce Type: replace-cross Abstract: AI systems increasingly synthesize executable structure at runtime: LLMs generate programs, agents construct workflows,self-improving systems

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

SafetyDGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

Graph Computation Meets Circuit Algebra: A Task-Aligned Analysis of Graph Neural Networks for Electronic Design Automation

ResearchDGX agent

arXiv:2605.08291v1 Announce Type: cross Abstract: EDA problems are graph-structured, but not all graph-structured problems call for the same GNN computation. We argue that successful GNN-for-EDA metho

GraphBench: Next-generation graph learning benchmarking

Model ReleasesDGX agent

arXiv:2512.04475v5 Announce Type: replace-cross Abstract: Machine learning on graphs has made substantial progress across domains such as molecular property prediction and chip design. Yet benchmarkin

GridProbe: Posterior-Probing for Adaptive Test-Time Compute in Long-Video VLMs

ResearchDGX agent

arXiv:2605.10762v1 Announce Type: cross Abstract: Long-video understanding in VLMs is bottlenecked by a single monolithic forward pass over thousands of frames at quadratic attention cost. A common mi

GRIT: Teaching MLLMs to Think with Images

ResearchDGX agent

arXiv:2505.15879v2 Announce Type: replace-cross Abstract: Recent studies have demonstrated the efficacy of using Reinforcement Learning (RL) in building reasoning models that articulate chains of thou

Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models

Model ReleasesDGX agent

arXiv:2603.16253v2 Announce Type: replace-cross Abstract: Vision-language process reward models (VL-PRMs) are increasingly used to score intermediate reasoning steps and rerank candidates under test-t

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing

ResearchDGX agent

arXiv:2605.10582v1 Announce Type: cross Abstract: This paper proposes a guaranteed defense method for large language models (LLMs) to safeguard against jailbreaking attacks. Drawing inspiration from t

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs

Model ReleasesDGX agent

arXiv:2508.20325v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly integral to various domains, their potential to generate harmful responses has prompted si

GuardAD: Safeguarding Autonomous Driving MLLMs via Markovian Safety Logic

SafetyDGX agent

arXiv:2605.10386v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly integrated into autonomous driving (AD) systems; however, they remain vulnerable to diverse sa

Guided Streaming Stochastic Interpolant Policy

SafetyDGX agent

arXiv:2605.10051v1 Announce Type: cross Abstract: Inference-time guidance is essential for steering generative robot policies toward dynamic objectives without retraining, yet existing methods are lar

HAGE: Harnessing Agentic Memory via RL-Driven Weighted Graph Evolution

AgentsDGX agent

arXiv:2605.09942v1 Announce Type: new Abstract: Memory retrieval in agentic large language model (LLM) systems is often treated as a static lookup problem, relying on flat vector search or fixed binar

HAMLET: A Hierarchical and Adaptive Multi-Agent Framework for Live Embodied Theatrics

AgentsDGX agent

arXiv:2507.15518v5 Announce Type: replace Abstract: Creating an immersive and interactive theatrical experience is a long-term goal in the field of interactive narrative. The emergence of large langua

HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation

SafetyDGX agent

arXiv:2605.09971v1 Announce Type: cross Abstract: Text-to-vibration generation converts natural language into haptic feedback, enabling vibration-effect designers to get scenarios-fitted vibrations mo

HeteroGenManip: Generalizable Manipulation For Heterogeneous Object Interactions

Local AiDGX agent

arXiv:2605.10201v1 Announce Type: cross Abstract: Generalizable manipulation involving cross-type object interactions is a critical yet challenging capability in robotics. To reliably accomplish such

HH-SAE: Discovering and Steering Hierarchical Knowledge of Complex Manifolds

ResearchDGX agent

arXiv:2605.10536v1 Announce Type: cross Abstract: Rare semantic innovations in high-dimensional, mission-critical domains are often obscured by dense background contexts, a challenge we define as exti

Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal

Model ReleasesDGX agent

arXiv:2605.09502v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting assumes that generated reasoning reflects a model's internal computation. We show this assumption is wrong in a speci

Hidden Heroes and Gradient Bloats: Layer-Wise Redundancy Inverts Attribution in Transformers

ResearchDGX agent

arXiv:2602.01442v3 Announce Type: replace-cross Abstract: Gradient-based attribution is the workhorse of mechanistic interpretability, yet whether it reliably tracks causal importance at the component

Hierarchical Attention-based Graph Neural Network with Relevance-driven Pruning

ResearchDGX agent

arXiv:2605.09308v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) excel at relational reasoning but face two persistent challenges: the lack of interpretable attribution for heterogeneous

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

SafetyDGX agent

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

Hierarchical Mixture-of-Experts with Two-Stage Optimization

ResearchDGX agent

arXiv:2605.08292v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) models scale capacity by routing each token to a small subset of experts. However, their routers exhibit a fundamental

HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities

Model ReleasesDGX agent

arXiv:2605.09348v1 Announce Type: cross Abstract: Large Language Models (LLMs) provide flexible natural language processing capabilities, while knowledge graphs (KGs) offer explicit and structured kno

HoReN: Normalized Hopfield Retrieval for Large-Scale Sequential Model Editing

Model ReleasesDGX agent

arXiv:2605.08143v1 Announce Type: cross Abstract: Large language models encode vast factual knowledge that inevitably becomes outdated or incorrect after deployment, yet retraining is costly prohibiti

How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients

ResearchDGX agent

arXiv:2504.10766v2 Announce Type: replace-cross Abstract: As the post-training of large language models (LLMs) advances from instruction-following to complex reasoning tasks, understanding how differe

How LLMs Are Persuaded: A Few Attention Heads, Rerouted

SafetyDGX agent

arXiv:2605.09314v1 Announce Type: new Abstract: Language models can be persuaded to abandon factual knowledge. This vulnerability is central to AI safety, but its internal mechanism remains poorly und

How Mobile World Model Guides GUI Agents?

ResearchDGX agent

arXiv:2605.10347v1 Announce Type: new Abstract: Recent advances in vision-language models have enabled mobile GUI agents to perceive visual interfaces and execute user instructions, but reliable predi

How Much is Brain Data Worth for Machine Learning?

SafetyDGX agent

arXiv:2605.09243v1 Announce Type: new Abstract: If a person can solve a task, can measuring their brain make it easier to train a model to solve that task too? Recent NeuroAI work suggests that supple

How You Begin is How You Reason: Driving Exploration in RLVR via Prefix-Tuned Priors

ResearchDGX agent

arXiv:2605.08817v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) recently thrives in large language model (LLM) reasoning tasks. However, the reward sparsity and t

HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control

SafetyDGX agent

arXiv:2605.08283v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a pivotal technique for enhancing the reasoning capabilities of Large Language Mo

Human-Inspired Memory Architecture for LLM Agents

Model ReleasesDGX agent

arXiv:2605.08538v1 Announce Type: new Abstract: Current LLM agents lack principled mechanisms for managing persistent memory across long interaction horizons. We present a biologically-grounded memory

Human-LLM Dialogue Improves Diagnostic Accuracy in Emergency Care

Model ReleasesDGX agent

arXiv:2605.08533v1 Announce Type: new Abstract: Clinical decision-making in emergency medicine demands rapid, accurate diagnoses under uncertainty. Despite benchmark progress, evidence for LLMs as int

HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding

SafetyDGX agent

arXiv:2605.08158v1 Announce Type: cross Abstract: Long-video understanding with multimodal language models suffers from three compounding bottlenecks: heavy decode cost to obtain dense RGB frames, qua

HyPER: Bridging Exploration and Exploitation for Scalable LLM Reasoning with Hypothesis Path Expansion and Reduction

SafetyDGX agent

arXiv:2602.06527v2 Announce Type: replace Abstract: Scaling test-time compute with multi-path chain-of-thought improves reasoning accuracy, but its effectiveness depends critically on the exploration-

Hyperbolic Distillation: Geometry-Guided Cross-Modal Transfer for Robust 3D Object Detection

ResearchDGX agent

arXiv:2605.09899v1 Announce Type: cross Abstract: Cross-modal knowledge distillation has emerged as an effective strategy for integrating point cloud and image features in 3D perception tasks. However

HYPERPOSE: Hyperbolic Kinematic Phase-Space Attention for 3D Human Pose Estimation

ResearchDGX agent

arXiv:2605.10100v1 Announce Type: cross Abstract: We introduce HYPERPOSE, a novel 3D human pose estimation framework that performs spatio-temporal reasoning entirely within the Lorentz model of hyperb

HyperSpace: A Generalized Framework for Spatial Encoding in Hyperdimensional Representations

Model ReleasesDGX agent

arXiv:2604.15113v2 Announce Type: replace Abstract: Vector Symbolic Architectures (VSAs) provide a well-defined algebraic framework for compositional representations in hyperdimensional spaces. We int

← Previous
1…267268269270271…358
Next →