AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
26 May 2026

MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research

AgentsDGX agent

arXiv:2605.26114v1 Announce Type: new Abstract: We present MobileGym, a browser-hosted, lightweight, fully controllable environment for everyday mobile use, targeting interaction fidelity without repl

MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM

ResearchDGX agent

arXiv:2602.20191v2 Announce Type: replace-cross Abstract: Dynamic runtime latency and memory constraints necessitate flexible large language model (LLM) deployment, where an LLM can be inferred with v

Mode-as-Sequence: Translating Multimodal Motion Prediction into Unified Sequential Mode Modeling

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.24037v1 Announce Type: cross Abstract: Multimodal motion forecasting is inherently under-supervised: each training scene provides only one realized future, yet multiple plausible futures ex

Momentum Streams for Optimizer-Inspired Transformers

ResearchDGX agent

arXiv:2605.24425v1 Announce Type: cross Abstract: The residual update of a pre-norm Transformer layer admits an interpretation as one step of a first-order optimizer acting on a surrogate token energy

More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries

AgentsDGX agent

arXiv:2605.24050v1 Announce Type: cross Abstract: Skill libraries allow LLM agents to load task-specific instructions on demand, letting non-expert users solve domain-specific tasks through natural la

Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending

Model ReleasesDGX agent

arXiv:2605.25574v1 Announce Type: cross Abstract: Concept erasure has emerged as a key research direction for ensuring safe and ethical image synthesis in Text-to-Image (T2I) models. While existing st

Motion-Compensated Weight Compression

SafetyDGX agent

arXiv:2605.24754v1 Announce Type: cross Abstract: Neural network weights are increasingly a bottleneck for deployment, yet most compression pipelines treat layers independently and overlook cross-laye

MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning

Model ReleasesDGX agent

arXiv:2605.25842v1 Announce Type: new Abstract: Vision-language models (VLMs) increasingly rely on chain-of-thought (CoT) reasoning to solve complex multimodal tasks, but their large parameter sizes m

Multi-Agent Coordination Adaptation via Structure-Guided Orchestration

SafetyDGX agent

arXiv:2605.25746v1 Announce Type: cross Abstract: As large language model (LLM)-based multi-agent systems scale to handle increasingly complex tasks, balancing structural stability and dynamic adaptab

Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations

AgentsDGX agent

arXiv:2605.25101v1 Announce Type: cross Abstract: In many industrial domains, the Functional Mock-up Interface (FMI) is used to exchange simulation models as Functional Mock-up Units (FMUs) across dif

Multi-market value-stacking: Battery control for combined imbalance participation and non-uniform FCR bidding

AgentsDGX agent

arXiv:2605.23964v1 Announce Type: cross Abstract: The growing share of Renewable Energy Sources (RES) in modern power systems increases both grid imbalances and frequency deviations, reinforcing the n

Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval

SafetyDGX agent

arXiv:2209.11572v3 Announce Type: replace-cross Abstract: As an increasingly popular task in multimedia information retrieval, video moment retrieval (VMR) aims to localize the target moment from an u

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning

SafetyDGX agent

arXiv:2605.25210v1 Announce Type: cross Abstract: Diffusion models are increasingly used as powerful conditional generators, yet real deployments often involve multiple target distributions arising fr

Multimodal Alignment and Preference Optimization for Zero-Shot Conditional RNA Generation

SafetyDGX agent

arXiv:2605.23961v1 Announce Type: cross Abstract: The design of RNA molecules that interact with specific proteins is a critical challenge in experimental and computational biology. Despite recent pro

Multimodal Functional Maximum Correlation for Emotion Recognition

SafetyDGX agent

arXiv:2512.23076v2 Announce Type: replace-cross Abstract: Emotional states manifest as coordinated yet heterogeneous physiological responses across central and autonomic systems, posing a fundamental

MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection

SafetyDGX agent

arXiv:2505.23803v2 Announce Type: replace-cross Abstract: Phishing email detection faces significant challenges due to evolving adversarial tactics and heterogeneous attack patterns. Traditional appro

MultiPUFFIN: A Multimodal Domain-Constrained Foundation Model for Molecular Property Prediction of Small Molecules

TutorialsDGX agent

arXiv:2603.00857v2 Announce Type: replace-cross Abstract: MultiPUFFIN is a domain-informed multimodal foundation model for predicting thermophysical properties of small molecules, addressing a critica

Multiscale Real-Time Object Detection in the NMS-Free Era: A Comparative Performance Evaluation of YOLOv8 and YOLO26

Local AiDGX agent

arXiv:2605.24831v1 Announce Type: cross Abstract: Non-Maximum Suppression (NMS) remains a key post-processing step in many real-time object detection pipelines, but it can introduce latency variation

MuNet: A Mutualistic Network for Joint 3D Human Mesh Recovery and 3D Clothed Human Reconstruction from Single Images

Model ReleasesDGX agent

arXiv:2605.25861v1 Announce Type: cross Abstract: 3D human mesh recovery and 3D clothed human reconstruction are inherently related, yet they have long been studied in isolation, thereby overlooking t

MX-SAFE: Versatile Inference- and Training-Proof Microscaling Format with On-the-Fly Exponent and Mantissa Bit Allocation

Local AiDGX agent

arXiv:2605.24391v1 Announce Type: cross Abstract: As the demand for deep learning grows, cost reduction through quantization has become essential for both training and inference. In 2022, the Open Com

Nano World Models: A Minimalist Implementation of Future Video Prediction

ResearchDGX agent

arXiv:2605.23993v1 Announce Type: cross Abstract: World models have become a central paradigm for learning predictive simulators that support generation, planning, and decision-making. Yet, despite ra

Neural Scalable Symbolic Search Framework for Complex Logical Queries with Multiple Free Variables

Model ReleasesDGX agent

arXiv:2605.25985v1 Announce Type: new Abstract: Complex Query Answering (CQA) is a fundamental knowledge representation and reasoning task over incomplete knowledge graphs (KGs). Answering existential

NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding

ResearchDGX agent

arXiv:2605.24993v1 Announce Type: new Abstract: Current fMRI decoders face a performance-fidelity trade-off where efficient ID encoders outperform geometrically faithful surface-based models. We argue

Neuro-Inspired Inverse Learning for Planning and Control

SafetyDGX agent

arXiv:2605.24152v1 Announce Type: new Abstract: We present a neuro-inspired framework for embodied planning and control. Building on three principles that enable fast and highly effective goal-directe

Neuromorphic LiDAR-based Bird's Eye View Object Detection using Energy-efficient Spiking Neural Networks

Model ReleasesDGX agent

arXiv:2605.25293v1 Announce Type: cross Abstract: Autonomous driving perception demands accurate and efficient processing of three-dimensional sensor data under strict power constraints. Traditional c

Neuronal Stochastic Attention Circuit (NSAC) for Probabilistic Representation Learning

AgentsDGX agent

arXiv:2605.26061v1 Announce Type: cross Abstract: Reliable quantification of uncertainty estimates in continuous-time (CT) representation learning remains nascent, particularly within CT attention arc

Noise-Robust Financial Numerical Entity Attribute Tagging

Model ReleasesDGX agent

arXiv:2605.24910v1 Announce Type: new Abstract: Financial Numerical Entity (FNE) understanding aims to recover the meaning of numerical mentions in financial reports. Existing studies primarily focus

Non-Invasive Reconstruction of Intracranial EEG Across the Deep Temporal Lobe from Scalp EEG based on Conditional Normalizing Flow

Local AiDGX agent

arXiv:2603.03354v3 Announce Type: replace-cross Abstract: Although obtaining deep brain activity from non-invasive scalp electroencephalography (sEEG) is crucial for neuroscience and clinical diagnosi

Not All Transitions Matter: Evidence from PPO

SafetyDGX agent

arXiv:2605.24071v1 Announce Type: cross Abstract: Training a reinforcement learning agent on-policy means collecting fresh experience at every update, and that experience comes with a hidden problem.

NPSolver: Neural Poisson Solver with Iterative Physics Supervision

ResearchDGX agent

arXiv:2605.25786v1 Announce Type: cross Abstract: Efficiently solving Poisson equations on complex, irregular domains remains a fundamental challenge in scientific computing, as classical iterative so

NSR-Boost: A Neuro-Symbolic Residual Boosting Framework for Industrial Legacy Models

ApplicationsDGX agent

arXiv:2601.10457v3 Announce Type: replace Abstract: Although the Gradient Boosted Decision Trees (GBDTs) dominate industrial tabular applications, upgrading legacy models in high-concurrency productio

OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation

SafetyDGX agent

arXiv:2605.25829v1 Announce Type: cross Abstract: Recent vision-language-action (VLA) models and world action models (WAMs) advance robotic manipulation by enriching intermediate representations with

OmniSapiens: A Foundation Model for Social Behavior Processing via Heterogeneity-Aware Relative Policy Optimization

SafetyDGX agent

arXiv:2602.10635v2 Announce Type: replace Abstract: Socially intelligent AI systems must entail reasoning across diverse human behavioral tasks, and generalization to new contexts. However, AI has yet

On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits

SafetyDGX agent

arXiv:2605.25789v1 Announce Type: cross Abstract: We study a stochastic multi-armed bandit problem where an agent is granted a free exploration budget before regret accumulates, a setting not captured

On the Epistemic Uncertainty of Overparametrized Neural Networks

Model ReleasesDGX agent

arXiv:2605.25234v1 Announce Type: cross Abstract: Epistemic uncertainty is often viewed as a reducible uncertainty that vanishes with increasing data. This perspective implicitly assumes parameter ide

On the Impact of Class Imbalance on the Learning Dynamics of Deep Neural Networks:An Intuitive Insight

ResearchDGX agent

arXiv:2605.24908v1 Announce Type: cross Abstract: Class imbalance in deep neural networks (DNNs) has witnessed a rapid increase in research attention in recent years. However, the varying accounts of

On the Stability and Realizability of Recurrent Polynomial Surrogate Ternary Logic Gate Networks

SafetyDGX agent

arXiv:2605.24649v1 Announce Type: cross Abstract: Recurrent Neural Networks (RNNs) can learn to predict Signal Temporal Logic (STL) verdicts online from partial trajectories, but deploying them as run

Operationalizing Reconstructive Authority: Runtime Construction, Dependency Resolution, and Execution Gating in Autonomous Agent Systems

SafetyDGX agent

arXiv:2605.23935v1 Announce Type: new Abstract: Autonomous agent systems fail not only due to incorrect decisions, but due to executing decisions whose authority no longer holds at runtime. Prior work

Optimizing Sensor Placement for Flow Reconstruction in Urban Drainage Networks: A Digital Twin-Based Sparse Sensing Approach

TutorialsDGX agent

arXiv:2511.04556v2 Announce Type: replace Abstract: Urban flooding triggered by intense rainfall is becoming increasingly frequent and widespread. While flood prediction and monitoring in high spatio-

OrpQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Transformer Quantization

Model ReleasesDGX agent

arXiv:2605.26092v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) and Vision Transformers (ViTs) on edge devices is significantly constrained by memory limitations and t

OSDTW: Optimal Shared Depth and Task Weighting for Long-Tailed Recognition

Model ReleasesDGX agent

arXiv:2605.24969v1 Announce Type: cross Abstract: Long-tailed recognition suffers from a persistent head--tail trade-off: improving tail performance often degrades head accuracy and can increase train

Overcoming 'Physics Shock' in Earth Observation A Heteroscedastic Uncertainty Framework for PINN-based Flood Inference

ApplicationsDGX agent

arXiv:2605.24106v1 Announce Type: cross Abstract: Rapid and accurate flood extent mapping from Remote Sensing data, such as Synthetic Aperture Radar (SAR), is critical for operational disaster respons

PACZero: PAC-Private Fine-Tuning of Language Models via Sign Quantization

Model ReleasesDGX agent

arXiv:2605.06505v2 Announce Type: replace-cross Abstract: We introduce PACZero, a family of PAC-private zeroth-order mechanisms for fine-tuning large language models that delivers usable utility at I(

PageLLM: A Multi-Grained Reward Framework for Whole-Page Optimization with Large Language Models

SafetyDGX agent

arXiv:2506.09084v2 Announce Type: replace-cross Abstract: Whole-page optimization (WPO) decides how search and recommendation results are surfaced to users, and large language models (LLMs) open a new

Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs

Model ReleasesDGX agent

arXiv:2605.24154v1 Announce Type: new Abstract: Current safety alignment of foundation models largely follows a one-size-fits-all paradigm, applying the same refusal policy across users and contexts.

PALoRA: Projection-Adaptive LoRA for Preserving Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.24549v1 Announce Type: new Abstract: Efficiently updating Large Language Models (LLMs) with new or evolving factual knowledge remains a central challenge, as even parameter-efficient adapta

PANDO: Efficient Multimodal AI Agents via Online Skill Distillation

AgentsDGX agent

arXiv:2605.24785v1 Announce Type: new Abstract: Recent advances in multimodal web agents often rely on increased inference-time computation, including rollout search, verifier passes, offline skill di

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

HardwareDGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

Parameter-Efficient CT Reconstruction via Deep Graph Laplacian Regularization

Model ReleasesDGX agent

arXiv:2605.25348v1 Announce Type: cross Abstract: Low-dose computed tomography (LDCT) reconstruction faces a critical tradeoff between reconstruction quality and resource requirements. While recent de

Parameter Efficient Multi-Class Intelligent Scheduling for Multimodal Online Distributed Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.23984v1 Announce Type: cross Abstract: Industrial anomaly detection has attracted significant attention as a fundamental challenge in industrial systems. The rapid advancement of heterogene

Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.24792v1 Announce Type: cross Abstract: The major limitations of gastrointestinal (GI) endoscopy AI systems arise from a shortage of annotated data, strict privacy policies, and significant

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

PASA: A Principled Embedding-Space Watermarking Approach for LLM-Generated Text under Semantic-Invariant Attacks

ResearchDGX agent

arXiv:2605.10977v2 Announce Type: replace-cross Abstract: Watermarking for large language models (LLMs) is a promising approach for detecting LLM-generated text and enabling responsible deployment. Ho

PathMem: Toward Cognition-Aligned Memory Transformation for Pathology MLLMs

ResearchDGX agent

arXiv:2603.09943v2 Announce Type: replace Abstract: Computational pathology demands both visual pattern recognition and dynamic integration of structured domain knowledge, including taxonomy, grading

PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs

SafetyDGX agent

arXiv:2601.20539v3 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled automated heuristic design (AHD) for combinatorial optimization problems (COPs), but existing frameworks'

PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning

AgentsDGX agent

arXiv:2502.10906v2 Announce Type: replace Abstract: Reward design plays a pivotal role in the training of game AIs, requiring substantial domain-specific knowledge and human effort. In recent years, s

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

Model ReleasesDGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation

Model ReleasesDGX agent

arXiv:2605.25572v1 Announce Type: cross Abstract: The growing complexity of quantum programming frameworks has exposed a critical limitation in existing large language model (LLM)-based code assistant

Performance Comparison of Classical and Neural Sampling Algorithms for Robotic Navigation

AgentsDGX agent

arXiv:2605.25010v1 Announce Type: cross Abstract: Integrating artificial intelligence (AI) into sampling-based motion planning provides new possibilities for improving autonomous navigation efficiency

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

Model ReleasesDGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

← Previous
1…216217218219220…358
Next →