AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs

TutorialsDGX agent

arXiv:2512.04746v2 Announce Type: replace-cross Abstract: Extremely low-bit quantization is critical for efficiently deploying Large Language Models (LLMs), yet it often leads to severe performance de

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment

SafetyDGX agent

arXiv:2605.17341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable success, yet their reliance on massive datasets and unintended memorization of training data ra

SIPO: Stabilized and Improved Preference Optimization for Aligning Diffusion Models

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2505.21893v3 Announce Type: replace-cross Abstract: Preference learning has garnered extensive attention as an effective technique for aligning diffusion models with human preferences in visual

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models

Local AiDGX agent

arXiv:2605.16842v1 Announce Type: new Abstract: Diffusion Multi-Modal Large Language Models (dMLLMs) are powerful for image generation, but optimizing them through reinforcement learning (RL) remains

SKG-Eval: Stateful Evaluation of Multi-Turn Dialogue via Incremental Semantic Knowledge Graphs

Local AiDGX agent

arXiv:2605.16650v1 Announce Type: cross Abstract: Evaluating multi-turn dialogue systems remains challenging because response quality depends not only on the current prompt, but also on previously est

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

Model ReleasesDGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

AgentsDGX agent

arXiv:2602.14211v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to extend LLM agents with task-specific instructions, executable scripts, and auxiliary resources. While im

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

Model ReleasesDGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

Model ReleasesDGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

Skim: Speculative Execution for Fast and Efficient Web Agents

AgentsDGX agent

arXiv:2605.16565v1 Announce Type: new Abstract: Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is

SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

Model ReleasesDGX agent

arXiv:2605.16626v1 Announce Type: cross Abstract: Since autonomous coding agents generate complex behaviors at high-volume, we may want to use other LLMs to monitor actions to reduce the risk from dan

Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules

Model ReleasesDGX agent

arXiv:2604.08432v2 Announce Type: replace-cross Abstract: Photonic neural networks promise ultrafast inference, yet most architectures rely on linear optical meshes with electronic nonlinearities, rei

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

Model ReleasesDGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

SomaliWeb v1: A Quality-Filtered Somali Web Corpus with a Matched Tokenizer and a Public Language-Identification Benchmark

Model ReleasesDGX agent

arXiv:2605.18232v1 Announce Type: cross Abstract: Somali is a Cushitic language of the Horn of Africa with ~25 million speakers, yet no documented dedicated Somali pretraining corpus with a companion

Some[Body] Must Receive That Pain for Agent Accountability

AgentsDGX agent

arXiv:2605.16872v1 Announce Type: cross Abstract: AI agents increasingly act consequentially in the real world. This creates a problem we call consequence reception: harm occurs, the producing system

SonarSweep: Fusing Sonar and Vision for Robust 3D Reconstruction via Plane Sweeping

ApplicationsDGX agent

arXiv:2511.00392v2 Announce Type: replace-cross Abstract: Accurate 3D reconstruction in visually-degraded underwater environments remains a formidable challenge. Single-modality approaches are insuffi

SparseSAM: Structured Sparsification of Activations in Segment Anything Models

ResearchDGX agent

arXiv:2605.17633v1 Announce Type: cross Abstract: The Segment Anything Model (SAM) achieves strong open-vocabulary segmentation, but its ViT-based image encoders dominate inference latency and memory.

Spatial Blindness in Whole-Slide Multiple Instance Learning

ResearchDGX agent

arXiv:2605.17449v1 Announce Type: cross Abstract: Whole-slide MIL models are often called context-aware once graphs, Transform ers, or state-space modules are placed above patch embeddings. We show th

Spatially Aware Linear Transformer (SAL-T) for Particle Jet Tagging

Local AiDGX agent

arXiv:2510.23641v2 Announce Type: replace-cross Abstract: Transformers are very effective in capturing both global and local correlations within high-energy particle collisions, but they present deplo

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

Spatiotemporal Robustness of Temporal Logic Tasks using Multi-Objective Reasoning

AgentsDGX agent

arXiv:2603.29868v2 Announce Type: replace Abstract: The reliability of autonomous systems depends on their robustness, i.e., their ability to meet their objectives under uncertainty. In this paper, we

Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception

AgentsDGX agent

arXiv:2601.09413v2 Announce Type: replace-cross Abstract: We introduce a voice-agentic framework that learns one critical omni-understanding skill: knowing when to trust itself versus when to consult

Spherical VAE with Cluster-Aware Feasible Regions: Guaranteed Prevention of Posterior Collapse

Model ReleasesDGX agent

arXiv:2603.10935v4 Announce Type: replace-cross Abstract: Variational autoencoders (VAEs) frequently suffer from posterior collapse, where the latent variables become uninformative as the approximate

Spiker-LL: An Energy-Efficient FPGA Accelerator Enabling Adaptive Local Learning in Spiking Neural Networks

Local AiDGX agent

arXiv:2605.18003v1 Announce Type: cross Abstract: Deploying adaptive intelligence at the edge remains challenging due to the high computational and energy cost of training neural models. Spiking Neura

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

SafetyDGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

Stabilizing Temporal Inference Dynamics for Online Surgical Phase Recognition

ResearchDGX agent

arXiv:2605.16387v1 Announce Type: cross Abstract: Online Surgical Phase Recognition (SPR) models can reach high frame-wise accuracy, yet their predictions often lack temporal stability, fragmenting wo

Stable Audio 3

HardwareDGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video

SafetyDGX agent

arXiv:2605.18553v1 Announce Type: cross Abstract: Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, whe

STAG-CN: Spatio-Temporal Apiary Graph Convolutional Network for Disease Onset Prediction in Beehive Sensor Networks

ResearchDGX agent

arXiv:2603.14462v2 Announce Type: replace-cross Abstract: Honey bee colony losses threaten global pollination services, yet current monitoring systems treat each hive as an isolated unit, ignoring the

State Contamination in Memory-Augmented LLM Agents

SafetyDGX agent

arXiv:2605.16746v1 Announce Type: new Abstract: LLM agents increasingly rely on persistent state, including transcripts, summaries, retrieved context, and memory buffers, to support long-horizon inter

State-of-the-Art Claims Require State-of-the-Art Evidence

Model ReleasesDGX agent

arXiv:2605.17273v1 Announce Type: cross Abstract: State-of-the-Art (SOTA) claims pervade Artificial Intelligence (AI) and Machine Learning (ML) research. These claims rest on benchmark evaluations, wh

Statistical Limits and Efficient Algorithms for Differentially Private Federated Learning

Model ReleasesDGX agent

arXiv:2605.18656v1 Announce Type: cross Abstract: Federated Learning is a leading framework for training ML and AI models collaboratively across numerous user devices or databases. We study the trade-

Stochastic Penalty-Barrier Methods for Constrained Machine Learning

SafetyDGX agent

arXiv:2605.18618v1 Announce Type: cross Abstract: Constrained machine learning enables fairness-aware training, physics-informed neural networks, and integration of symbolic domain knowledge into stat

Strategic Over-Parameterization for Generalizable Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2605.16470v1 Announce Type: cross Abstract: Adapting large language models (LLMs) to downstream tasks via full fine-tuning is increasingly impractical due to its computational and memory demands

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video

Model ReleasesDGX agent

arXiv:2605.16381v1 Announce Type: cross Abstract: Proactive streaming video understanding requires models to continuously process video streams and decide when to respond, rather than merely what to r

STRIDE: A Self-Reflective Agent Framework for Reliable Automatic Equation Discovery

AgentsDGX agent

arXiv:2605.17790v1 Announce Type: new Abstract: LLM-based equation discovery offers a promising route to recovering symbolic laws from data, but many systems still rely on generation-centered loops th

STRIDE-AI: A Threat Modeling Framework for Generative AI Security Assessment

ApplicationsDGX agent

arXiv:2605.17163v1 Announce Type: cross Abstract: Traditional cybersecurity methodologies target deterministic systems and fail to address the probabilistic nature of AI, leaving systems vulnerable to

StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs

Model ReleasesDGX agent

arXiv:2605.16353v1 Announce Type: cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models to incrementally acquire new abilities. However, existing CVIT met

StructLens: A Structural Lens for Language Models via Maximum Spanning Trees

ResearchDGX agent

arXiv:2603.03328v2 Announce Type: replace-cross Abstract: Language exhibits inherent structures, a property that explains both language acquisition and language change. Given this characteristic, we e

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics

Model ReleasesDGX agent

arXiv:2605.18548v1 Announce Type: cross Abstract: Large language models (LLMs) deployed in real-world agentic applications must be capable of replanning and adapting when mid-task disruptions invalida

StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting

Model ReleasesDGX agent

arXiv:2605.17309v1 Announce Type: cross Abstract: We present StyleText, a large-scale dataset and benchmark for localized scene-text inpainting with style preservation. StyleText contains 28,518 image

Supervising the search process produces reliable and generalizable information-seeking agents

Model ReleasesDGX agent

arXiv:2502.13957v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming web search by shifting from document ranking to synthesizing answers, and are increasingly deplo

Support-Safe Variational Hybrid Filtering for Contact-Mode and Sparse-Law Recovery

ResearchDGX agent

arXiv:2605.16398v1 Announce Type: cross Abstract: Contact-rich robot dynamics are hybrid: a single observation can match several latent states and contact regimes (free, impact, stick--slip). A standa

SuReNav: Superpixel Graph-based Constraint Relaxation for Navigation in Over-constrained Environments

SafetyDGX agent

arXiv:2602.06807v2 Announce Type: replace-cross Abstract: We address the over-constrained planning problem in semi-static environments. The planning objective is to find a best-effort solution that av

Surface-Form Neural Sparse Retrieval: Robust Fuzzy Matching for Industrial Music Search

TutorialsDGX agent

arXiv:2605.17762v1 Announce Type: new Abstract: Music search at the scale of Amazon Music presents a unique challenge: queries frequently deviate from indexed metadata due to misspellings, transpositi

Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention

SafetyDGX agent

arXiv:2603.01683v2 Announce Type: replace-cross Abstract: Injecting new reasoning knowledge into Large Language Models (LLMs) via post-training often induces catastrophic forgetting. Recent studies em

Sustainability via LLM Right-sizing

Model ReleasesDGX agent

arXiv:2504.13217v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have become increasingly embedded in organizational workflows. This has raised concerns over their energy consump

Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents

Local AiDGX agent

arXiv:2605.16671v1 Announce Type: new Abstract: Rapid biodiversity loss underscore the urgency of effective monitoring, yet manual surveys remain resource-intensive. While on-device AI offers a scalab

SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space

SafetyDGX agent

arXiv:2603.26720v2 Announce Type: replace-cross Abstract: Predicting surgical needle trajectories from endoscopic video is critical for robot-assisted suturing, enabling anticipatory planning, real-ti

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain

Model ReleasesDGX agent

arXiv:2605.17946v1 Announce Type: new Abstract: Multimodal large language models are increasingly used as agent backbones that understand multimodal inputs, plan retrieval actions, invoke external too

SwordBench: Evaluating Orthogonality of Steering Image Representations

Model ReleasesDGX agent

arXiv:2605.16372v1 Announce Type: cross Abstract: Steering or intervening on model representations at inference time to correct predictions is essential for AI interpretability and safety, yet existin

Symmetry-Compatible Principle for Optimizer Design: Embeddings, LM Heads, SwiGLU MLPs, and MoE Routers

Model ReleasesDGX agent

arXiv:2605.18106v1 Announce Type: cross Abstract: A striking geometric disparity has long persisted in the practice of deep learning. While modern neural network architectures naturally exhibit rich s

Symphony for Speech-to-Text: Supporting Real-Time Medical Voice Interfaces

Model ReleasesDGX agent

arXiv:2605.16545v1 Announce Type: cross Abstract: After decades of use in dictation and, more recently, ambient documentation, speech is emerging as a primary modality for interacting with technology

SynVA: A Modular Toolkit for Vessel Generation and Aneurysm Editing

ResearchDGX agent

arXiv:2605.17620v1 Announce Type: cross Abstract: Intracranial aneurysms (IAs), characterized by unpredictable growth and risk of rupture, are a major cause of stroke and can lead to life-threatening

Systematic Evaluation of the Quality of Synthetic Clinical Notes Rephrased by LLMs at Million-Note Scale

ResearchDGX agent

arXiv:2605.17775v1 Announce Type: cross Abstract: Large language models (LLMs) can generate or synthesize clinical text for a wide range of applications, from improving clinical documentation to augme

Systematic Evaluation of Vision Transformers for Automated Cervical Cancer Classification: Optimization, Statistical Validation, and Clinical Interpretability

ResearchDGX agent

arXiv:2605.17236v1 Announce Type: cross Abstract: Manual Pap smear analysis for cervical cancer screening is limited by inter-observer variability, time constraints, and restricted expert availability

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

HardwareDGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

TailedTS: Benchmark Dataset for Heavy-Tailed Time Series Prediction and Periodicity Quantification

Model ReleasesDGX agent

arXiv:2605.16361v1 Announce Type: cross Abstract: We present TailedTS, a large-scale benchmark dataset derived from Wikipedia hourly page view observations throughout 2024, specifically designed to te

Task Abstention for Large Language Models in Code Generation

Model ReleasesDGX agent

arXiv:2605.17029v1 Announce Type: cross Abstract: Large language models (LLMs) have revolutionized automated code generation. One serious concern, however, is the so-called ``hallucination'', i.e., LL

← Previous
1…242243244245246…358
Next →