AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding

DGX agent

arXiv:2605.07141v1 Announce Type: cross Abstract: Open-world referring segmentation requires grounding unconstrained language expressions to precise pixel-level regions. Existing multimodal large lang

model-releasesarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes

DGX agent

arXiv:2601.20599v2 Announce Type: replace-cross Abstract: Gradient temporal-difference (GTD) learning algorithms are widely used for off-policy policy evaluation with function approximation. However,

safetyarxiv-cs-ai
11 May 2026
Research

R^3L: Reasoning 3D Layouts from Relative Spatial Relations

DGX agent

arXiv:2605.06758v1 Announce Type: cross Abstract: Relative spatial relations provide a compact representation of spatial structure and are fundamental to relative spatial reasoning in 3D layout genera

researcharxiv-cs-ai
11 May 2026
Model Releases

Randomness is sometimes necessary for coordination

DGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

model-releasesarxiv-cs-ai
11 May 2026
Safety

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

DGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

safetyarxiv-cs-ai
11 May 2026
Model Releases

ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning

DGX agent

arXiv:2605.06483v2 Announce Type: replace Abstract: Signal Temporal Logic (STL) is an expressive formal language for specifying spatio-temporal requirements over real-valued, real-time signals. It has

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference

DGX agent

arXiv:2605.07234v1 Announce Type: cross Abstract: Large language models (LLMs) support long-context inference but suffer from substantial memory and runtime overhead due to Key-Value (KV) Cache growth

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Region4Web: Rethinking Observation Space Granularity for Web Agents

DGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

model-releasesarxiv-cs-ai
11 May 2026
Research

Regulating Branch Parallelism in LLM Serving

DGX agent

arXiv:2605.06914v1 Announce Type: cross Abstract: Recent methods expose intra-request parallelism in LLM outputs, allowing independent branches to decode concurrently. Existing serving systems execute

researcharxiv-cs-ai
11 May 2026
Safety

RELO: Reinforcement Learning to Localize for Visual Object Tracking

DGX agent

arXiv:2605.07379v1 Announce Type: cross Abstract: Conventional visual object trackers localize targets using handcrafted spatial priors, often in the form of heatmaps. Such priors provide only surroga

safetyarxiv-cs-ai
11 May 2026
Hardware

Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement

DGX agent

arXiv:2605.06298v2 Announce Type: replace-cross Abstract: Training world models on vast quantities of unlabelled videos is a critical step toward fully autonomous intelligence. However, the prevailing

hardwarearxiv-cs-ai
11 May 2026
Model Releases

Rep2Text: Decoding Full Text from a Single LLM Token Representation

DGX agent

arXiv:2511.06571v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress across diverse tasks, yet their internal mechanisms remain largely opaque. In t

model-releasesarxiv-cs-ai
11 May 2026
Safety

Repeated Deceptive Path Planning against Learnable Observer

DGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

safetyarxiv-cs-ai
11 May 2026
Research

Replicating Human Motivated Reasoning Studies with LLMs

DGX agent

arXiv:2601.16130v2 Announce Type: replace-cross Abstract: Motivated reasoning - the idea that individuals processing information may be motivated to either arrive at accurate beliefs or arrive at desi

researcharxiv-cs-ai
11 May 2026
Safety

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

DGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

safetyarxiv-cs-ai
11 May 2026
Safety

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

DGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

safetyarxiv-cs-ai
11 May 2026
Safety

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

DGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

safetyarxiv-cs-ai
11 May 2026
Model Releases

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

DGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

model-releasesarxiv-cs-ai
11 May 2026
Research

Revisiting Adam for Streaming Reinforcement Learning

DGX agent

arXiv:2605.06764v1 Announce Type: cross Abstract: Learning from a sequence of interactions, as soon as observations are perceived and acted upon, without explicitly storing them, holds the promise of

researcharxiv-cs-ai
11 May 2026
Model Releases

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

DGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Safety

Rubric-based On-policy Distillation

DGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

safetyarxiv-cs-ai
11 May 2026
Model Releases

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

DGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Safety

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

DGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

safetyarxiv-cs-ai
11 May 2026
Model Releases

Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents

DGX agent

arXiv:2605.07630v1 Announce Type: cross Abstract: When a phone-use agent avoids harm, does that show safety, or simply inability to act? Existing evaluations often cannot tell. A harmful outcome may b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Safety Anchor: Defending Harmful Fine-tuning via Geometric Bottlenecks

DGX agent

arXiv:2605.05995v2 Announce Type: replace-cross Abstract: The safety alignment of Large Language Models (LLMs) remains vulnerable to Harmful Fine-tuning (HFT). While existing defenses impose constrain

model-releasesarxiv-cs-ai
11 May 2026
Research

Saliency-Aware Regularized Quantization Calibration for Large Language Models

DGX agent

arXiv:2605.05693v2 Announce Type: replace Abstract: Post-training quantization (PTQ) is an effective approach for deploying large language models (LLMs) under memory and latency constraints. Most exis

researcharxiv-cs-ai
11 May 2026
Research

SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild

DGX agent

arXiv:2605.07604v1 Announce Type: cross Abstract: 3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scene

researcharxiv-cs-ai
11 May 2026
Tutorials

Same Brain, Different Prediction: How Preprocessing Choices Undermine EEG Decoding Reliability

DGX agent

arXiv:2605.07212v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a cornerstone of brain-computer interfaces and clinical neuroscience, yet deep learning models are typically trained a

tutorialsarxiv-cs-ai
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Safety

SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints

DGX agent

arXiv:2512.23770v3 Announce Type: replace-cross Abstract: In safety-critical domains, reinforcement learning (RL) agents must often satisfy strict, zero-cost safety constraints while accomplishing tas

safetyarxiv-cs-ai
11 May 2026
Research

Scalable Option Learning in High-Throughput Environments

DGX agent

arXiv:2509.00338v3 Announce Type: replace-cross Abstract: Hierarchical reinforcement learning (RL) has the potential to enable effective decision-making over long timescales. Existing approaches, whil

researcharxiv-cs-ai
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

DGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

model-releasesarxiv-cs-ai
11 May 2026
Applications

Script Sensitivity: Benchmarking Language Models on Unicode, Romanized and Mixed-Script Sinhala

DGX agent

arXiv:2601.14958v3 Announce Type: replace-cross Abstract: The performance of Language Models (LMs) on low-resource, morphologically rich languages like Sinhala remains largely unexplored, particularly

applicationsarxiv-cs-ai
11 May 2026
Agents

Searching for Privacy Risks in LLM Agents via Simulation

DGX agent

arXiv:2508.10880v3 Announce Type: replace-cross Abstract: The widespread deployment of LLM-based agents is likely to introduce a critical privacy threat: malicious agents that proactively engage other

agentsarxiv-cs-ai
11 May 2026
Safety

Self-Programmed Execution for Language-Model Agents

DGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

safetyarxiv-cs-ai
11 May 2026
Hardware

Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding

DGX agent

arXiv:2605.07897v1 Announce Type: cross Abstract: Online streaming video understanding requires models to process continuous visual inputs and respond to user queries in real time, where the unbounded

hardwarearxiv-cs-ai
11 May 2026
Model Releases

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression

DGX agent

arXiv:2502.01941v3 Announce Type: replace-cross Abstract: While Key-Value (KV) cache compression is essential for efficient LLM inference, current evaluations disproportionately focus on sparse retrie

model-releasesarxiv-cs-ai
11 May 2026
Research

Shaping the Future of Mathematics in the Age of AI

DGX agent

arXiv:2603.24914v2 Announce Type: replace-cross Abstract: Artificial intelligence is transforming mathematics at a speed and scale that demand active engagement from the mathematical community. We exa

researcharxiv-cs-ai
11 May 2026
Research

SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion

DGX agent

arXiv:2605.07482v1 Announce Type: cross Abstract: Machine unlearning for large language models (LLMs) aims to selectively remove memorized content such as private data, copyrighted text, or hazardous

researcharxiv-cs-ai
11 May 2026
Local Ai

Signal Reshaping for GRPO in Weak-Feedback Agentic Code Repair

DGX agent

arXiv:2605.07276v1 Announce Type: new Abstract: Code-agent RL often receives weak feedback: rollout-time signals are reliable and executable, but capture only necessary or surface conditions for task

local-aiarxiv-cs-ai
11 May 2026
Safety

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

DGX agent

arXiv:2605.06130v2 Announce Type: replace Abstract: A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupl

safetyarxiv-cs-ai
11 May 2026
Research

Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models

DGX agent

arXiv:2509.25584v2 Announce Type: replace Abstract: Vision-language models achieve incredible performance across a wide range of tasks, but their large size makes inference costly. Recent work has sho

researcharxiv-cs-ai
11 May 2026
Model Releases

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

DGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

model-releasesarxiv-cs-ai
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Agents

SOM: Structured Opponent Modeling for LLM-based Agents via Structural Causal Model

DGX agent

arXiv:2605.07301v1 Announce Type: new Abstract: Accurately predicting opponents' behavior from interactions is a fundamental capability for large language model (LLM)-based agents in multi-agent and g

agentsarxiv-cs-ai
11 May 2026
← Previous
1…353354355356357…448
Next →