AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

DGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

safetyarxiv-cs-ai
19 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Rethinking GNNs and Missing Features: Challenges, Evaluation and a Robust Solution

DGX agent

arXiv:2601.04855v2 Announce Type: replace-cross Abstract: Handling missing node features is a key challenge for deploying Graph Neural Networks (GNNs) in real-world domains such as healthcare and sens

model-releasesarxiv-cs-ai
19 May 2026
Safety

Retrieval and competition: how a protein foundation model starts a protein

DGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

safetyarxiv-cs-ai
19 May 2026
Agents

Reversa: A Reverse Documentation Engineering Framework for Converting Legacy Software into Operational Specifications for AI Agents

DGX agent

arXiv:2605.18684v1 Announce Type: cross Abstract: Legacy systems concentrate business rules, architectural decisions, and operational exceptions that often remain implicit in code, data, configuration

agentsarxiv-cs-ai
19 May 2026
Model Releases

Reverse-Engineering Model Editing on Language Models

DGX agent

arXiv:2602.10134v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are pretrained on corpora containing trillions of tokens and, therefore, inevitably memorize sensitive informatio

model-releasesarxiv-cs-ai
19 May 2026
Applications

Revisiting Long-term Time Series Forecasting: An Investigation on Linear Mapping

DGX agent

arXiv:2305.10721v2 Announce Type: replace-cross Abstract: Introduction: Long-term time series forecasting (LTSF) has gained significant attention in recent years. While various specialized designs exi

applicationsarxiv-cs-ai
19 May 2026
Research

RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots

DGX agent

arXiv:2605.18197v1 Announce Type: cross Abstract: Current approaches to 3D scene graph generation rely on dedicated depth sensors, such as LiDAR or RGB-D cameras, for metric 3D reconstruction. This li

researcharxiv-cs-ai
19 May 2026
Model Releases

RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards

DGX agent

arXiv:2509.21319v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Human Feedback (RLHF) and Reinforcement Learning with Verifiable Rewards (RLVR) are the main RL paradigms used in

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies

DGX agent

arXiv:2603.04639v2 Announce Type: replace-cross Abstract: Memory is critical for long-horizon and history-dependent robotic manipulation. Such tasks often involve counting repeated actions or manipula

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

Rover: Context-aware Conflict Resolution with LLM

DGX agent

arXiv:2605.17279v1 Announce Type: cross Abstract: Code merging is a significant challenge, particularly in large-scale projects. Existing solutions, including program analysis and machine learning, sh

tutorialsarxiv-cs-ai
19 May 2026
Agents

S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination

DGX agent

arXiv:2605.17076v1 Announce Type: cross Abstract: Concurrent LLM agents sharing mutable natural-language state produce Structural Race Conditions (SRCs): write-write and cross-shard stale-read conflic

agentsarxiv-cs-ai
19 May 2026
Model Releases

SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering

DGX agent

arXiv:2605.17526v1 Announce Type: cross Abstract: As autonomous coding agents become capable of handling increasingly long-horizon tasks, they have gradually demonstrated the potential to complete end

model-releasesarxiv-cs-ai
19 May 2026
Research

SAFE-SVD: Sensitivity-Aware Fidelity-Enforcing SVD for Physics Foundation Models

DGX agent

arXiv:2605.17985v1 Announce Type: cross Abstract: We propose a new method for compressing physics foundation models (PFMs) which is a new trend in AI for Science. While model compression is essential

researcharxiv-cs-ai
19 May 2026
Safety

Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction

DGX agent

arXiv:2605.18104v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) often fail to transfer safety capabilities learned in the text modality to semantically equivalent non-text inp

safetyarxiv-cs-ai
19 May 2026
Model Releases

SAME: A Semantically-Aligned Music Autoencoder

DGX agent

arXiv:2605.18613v1 Announce Type: cross Abstract: Latent representations are at the heart of the majority of modern generative models. In the audio domain they are typically produced by a neural-audio

model-releasesarxiv-cs-ai
19 May 2026
Agents

Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents

DGX agent

arXiv:2605.18332v1 Announce Type: cross Abstract: Behavioral studies of LLM-based software engineering agents extract operational rules about which trajectory shapes correlate with higher resolution r

agentsarxiv-cs-ai
19 May 2026
Safety

SAPO: Step-Aligned Policy Optimization for Reasoning-Based Generative Recommendation

DGX agent

arXiv:2605.17648v1 Announce Type: new Abstract: Generative recommendation treats next-item prediction as autoregressive item-identifier generation. Specifically, items are encoded as semantic identifi

safetyarxiv-cs-ai
19 May 2026
Research

SAS: Semantic-aware Sampling for Generative Dataset Distillation

DGX agent

arXiv:2605.18012v1 Announce Type: cross Abstract: Deep neural networks have achieved impressive performance across a wide range of tasks, but this success often comes with substantial computational an

researcharxiv-cs-ai
19 May 2026
Research

Scalable Environments Drive Generalizable Agents

DGX agent

arXiv:2605.18181v1 Announce Type: new Abstract: Generalizable agents should adapt to diverse tasks and unseen environments beyond their training distribution. This position paper argues that such gene

researcharxiv-cs-ai
19 May 2026
Applications

Scalable Uncertainty Reasoning in Knowledge Graphs

DGX agent

arXiv:2605.16568v1 Announce Type: new Abstract: Knowledge Graphs are pivotal for semantic data integration. The real-world data they model is often inherently uncertain. Within knowledge graphs, uncer

applicationsarxiv-cs-ai
19 May 2026
Model Releases

Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings

DGX agent

arXiv:2510.26384v2 Announce Type: replace Abstract: The prohibitive cost of evaluating large language models (LLMs) on comprehensive benchmarks necessitates the creation of small yet representative da

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Scheduling That Speaks: An Interpretable Programmatic Reinforcement Learning Framework

DGX agent

arXiv:2605.18454v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) has recently emerged as a promising approach to solve combinatorial optimization problems such as job shop schedulin

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

DGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

model-releasesarxiv-cs-ai
19 May 2026
Research

Scientific Logicality Enriched Methodology for LLM Reasoning: A Practice in Physics

DGX agent

arXiv:2605.17104v1 Announce Type: new Abstract: With the continuous advancement of reasoning abilities in Large Language Models (LLMs), their application to scientific reasoning tasks has gained signi

researcharxiv-cs-ai
19 May 2026
Safety

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

DGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

safetyarxiv-cs-ai
19 May 2026
Local Ai

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding

DGX agent

arXiv:2605.18018v1 Announce Type: cross Abstract: We present SWIM (See What I Mean), a novel training strategy that aligns vision and language representations to enable fine-grained object understandi

local-aiarxiv-cs-ai
19 May 2026
Safety

Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency

DGX agent

arXiv:2605.18162v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have made striking progress, yet their spatial reasoning remains fragile: models that answer an original input correctly

safetyarxiv-cs-ai
19 May 2026
Model Releases

Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain

DGX agent

arXiv:2603.02218v2 Announce Type: replace-cross Abstract: Large language models (LLMs) make it plausible to build systems that improve through self-evolving loops, but many existing proposals are bett

model-releasesarxiv-cs-ai
19 May 2026
Agents

Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning

DGX agent

arXiv:2602.08167v2 Announce Type: replace-cross Abstract: Embodied Chain-of-Thought (CoT) reasoning has significantly enhanced Vision-Language-Action (VLA) models, yet current methods rely on rigid te

agentsarxiv-cs-ai
19 May 2026
Model Releases

Self-supervised Hierarchical Visual Reasoning with World Model

DGX agent

arXiv:2605.17537v1 Announce Type: new Abstract: 3D open-world environments with adversarial opponents remain a core challenge for reinforcement learning due to their vast state spaces. Effective reaso

model-releasesarxiv-cs-ai
19 May 2026
Agents

SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning

DGX agent

arXiv:2605.17101v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) is widely employed to mitigate risks such as hallucinations and knowledge obsolescence in medical question answer

agentsarxiv-cs-ai
19 May 2026
Research

Semantic Generative Tuning for Unified Multimodal Models

DGX agent

arXiv:2605.18714v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) strive to consolidate visual understanding and visual generation within a single architecture. However, prevailing tr

researcharxiv-cs-ai
19 May 2026
Safety

Semantic Smoothing via Novel View Synthesis for Robust SAR Image Classification

DGX agent

arXiv:2605.16440v1 Announce Type: cross Abstract: Deep neural networks are vulnerable to adversarial perturbations, limiting deployment in safety-critical applications such as synthetic aperture radar

safetyarxiv-cs-ai
19 May 2026
Research

SENSE: Satellite-based ENergy Synthesis for Sustainable Environment

DGX agent

arXiv:2605.18101v1 Announce Type: cross Abstract: Urban Building Energy Modeling plays a critical role in achieving the United Nations' Sustainable Development Goals 7 and 11. Although existing studie

researcharxiv-cs-ai
19 May 2026
Model Releases

ShareChat: A Dataset of Chatbot Conversations in the Wild

DGX agent

arXiv:2512.17843v4 Announce Type: replace-cross Abstract: By evaluating Large Language Models (LLMs) through uniform, text-only interfaces, current academic benchmarks obscure how the unique designs a

model-releasesarxiv-cs-ai
19 May 2026
Safety

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

DGX agent

arXiv:2605.17999v1 Announce Type: new Abstract: This paper proposes a Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) algorithm. By sharing the base module between the Actor and Cri

safetyarxiv-cs-ai
19 May 2026
Tutorials

SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs

DGX agent

arXiv:2512.04746v2 Announce Type: replace-cross Abstract: Extremely low-bit quantization is critical for efficiently deploying Large Language Models (LLMs), yet it often leads to severe performance de

tutorialsarxiv-cs-ai
19 May 2026
Safety

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment

DGX agent

arXiv:2605.17341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable success, yet their reliance on massive datasets and unintended memorization of training data ra

safetyarxiv-cs-ai
19 May 2026
Model Releases

SIPO: Stabilized and Improved Preference Optimization for Aligning Diffusion Models

DGX agent

arXiv:2505.21893v3 Announce Type: replace-cross Abstract: Preference learning has garnered extensive attention as an effective technique for aligning diffusion models with human preferences in visual

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models

DGX agent

arXiv:2605.16842v1 Announce Type: new Abstract: Diffusion Multi-Modal Large Language Models (dMLLMs) are powerful for image generation, but optimizing them through reinforcement learning (RL) remains

local-aiarxiv-cs-ai
19 May 2026
Local Ai

SKG-Eval: Stateful Evaluation of Multi-Turn Dialogue via Incremental Semantic Knowledge Graphs

DGX agent

arXiv:2605.16650v1 Announce Type: cross Abstract: Evaluating multi-turn dialogue systems remains challenging because response quality depends not only on the current prompt, but also on previously est

local-aiarxiv-cs-ai
19 May 2026
Model Releases

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

DGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

model-releasesarxiv-cs-ai
19 May 2026
Agents

SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

DGX agent

arXiv:2602.14211v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to extend LLM agents with task-specific instructions, executable scripts, and auxiliary resources. While im

agentsarxiv-cs-ai
19 May 2026
Model Releases

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

DGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

DGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

model-releasesarxiv-cs-ai
19 May 2026
Agents

Skim: Speculative Execution for Fast and Efficient Web Agents

DGX agent

arXiv:2605.16565v1 Announce Type: new Abstract: Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is

agentsarxiv-cs-ai
19 May 2026
Model Releases

SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

DGX agent

arXiv:2605.16626v1 Announce Type: cross Abstract: Since autonomous coding agents generate complex behaviors at high-volume, we may want to use other LLMs to monitor actions to reduce the risk from dan

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules

DGX agent

arXiv:2604.08432v2 Announce Type: replace-cross Abstract: Photonic neural networks promise ultrafast inference, yet most architectures rely on linear optical meshes with electronic nonlinearities, rei

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…302303304305306…448
Next →