AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

RKSC: Reasoning-Aware KV Cache Sharing and Confident Early Exit for Multi-Step LLM Inference

DGX agent

arXiv:2606.09937v1 Announce Type: cross Abstract: We introduce RKSC (Reasoning-Aware KV Cache Sharing), a training-free inference framework that eliminates two structural redundancies in multi-branch

researcharxiv-cs-ai
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

DGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning

DGX agent

arXiv:2606.11092v1 Announce Type: cross Abstract: Elite humanoid soccer shooting requires whole-body stability, high-impulse whole-body interactions, and accuracy to targets. Motion tracking-driven re

safetyarxiv-cs-ai
10 Jun 2026
Agents

Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey

DGX agent

arXiv:2403.00420v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is a subfield of machine learning for training autonomous agents that take sequential actions across complex

agentsarxiv-cs-ai
10 Jun 2026
Safety

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

DGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

DGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Routing-Aware Expert Calibration for Machine Unlearning in Mixture-of-Experts Language Models

DGX agent

arXiv:2606.10338v1 Announce Type: cross Abstract: Machine unlearning is increasingly important for large language models, yet unlearning in Mixture-of-Experts (MoE) architectures remains underexplored

researcharxiv-cs-ai
10 Jun 2026
Model Releases

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

DGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

DGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SCOPE: Sequential Causal Optimization of Process Interventions

DGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

SD-GRPO: Verifiable Segment Decomposition for Long-Form Vision-Language Generation

DGX agent

arXiv:2606.09871v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) and its variants, originally developed for Large Language Models (LLMs), have recently been applied to Multi

safetyarxiv-cs-ai
10 Jun 2026
Safety

Self-Distillation Policy Optimization via Visual Feedback: Bridging Code and Visual Artifacts

DGX agent

arXiv:2606.10334v1 Announce Type: new Abstract: Code-generating large language models (LLMs) increasingly produce visual artifacts such as charts, web pages, and slides by writing programs that are ex

safetyarxiv-cs-ai
10 Jun 2026
Safety

Self-EmoQ: Plutchik-Guided Value-based Planning to Drive Streaming Emotional TTS

DGX agent

arXiv:2606.09837v1 Announce Type: cross Abstract: Emotional interaction is increasingly crucial for conversational AI, yet current systems lack a self-emotion determination mechanism to drive the stre

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPE: Coalition-Aware Expert Pruning for Sparse Mixture-of-Experts LLMs

DGX agent

arXiv:2606.09886v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) large language models achieve strong quality with low per-token compute, yet their deployment is often limited by the

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

DGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

DGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

DGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

DGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

SocraticPO: Policy Optimization via Interactive Guidance

DGX agent

arXiv:2606.09887v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewar

safetyarxiv-cs-ai
10 Jun 2026
Research

Soul Computing: A Theoretical Framework and Technical Architecture for Intelligent Agents with Independent Consciousness

DGX agent

arXiv:2606.10413v1 Announce Type: new Abstract: Breakthroughs in large language models and multimodal generation technologies have propelled the digital reconstruction of human mental traits, emotiona

researcharxiv-cs-ai
10 Jun 2026
Model Releases

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

DGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

model-releasesarxiv-cs-ai
10 Jun 2026
Local Ai

Spatial-Omni: Spatial Audio Understanding Integration in Multimodal LLMs via FOA Encoding

DGX agent

arXiv:2606.10738v1 Announce Type: cross Abstract: Recent multimodal large language models mainly process audio as monaural signals, thereby discarding the spatial cues contained in spatial audio for s

local-aiarxiv-cs-ai
10 Jun 2026
Research

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation

DGX agent

arXiv:2606.10368v1 Announce Type: cross Abstract: Speech-to-text (S2T) systems for recognition (ASR) and translation (S2TT) typically generate discrete text tokens. In contrast, continuous-target lang

researcharxiv-cs-ai
10 Jun 2026
Model Releases

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

DGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Stop Early, Spend Less: Hidden-State Probes as a Practical Recipe for Streaming Moderation of LLM Outputs

DGX agent

arXiv:2606.10487v1 Announce Type: cross Abstract: Deploying large language models in user-facing systems requires efficient output safety filtering. Existing approaches typically rely on a separate mo

safetyarxiv-cs-ai
10 Jun 2026
Research

STORM: Stepwise Token Optimization with Reward-Guided Beam Search

DGX agent

arXiv:2606.10621v1 Announce Type: cross Abstract: Modern retrieval increasingly relies on dense and learned-sparse neural models that are effective but require encoding the entire corpus into a specia

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning

DGX agent

arXiv:2606.11015v1 Announce Type: new Abstract: Tuning controllers for strongly coupled multi-input multi-output (MIMO) industrial processes is hard: decentralized classical auto-tuning ignores loop i

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Structure-Preserving Learning Improves Geometry Generalization in Neural PDEs

DGX agent

arXiv:2602.02788v2 Announce Type: replace-cross Abstract: We aim to develop physics foundation models for science and engineering that provide real-time solutions to Partial Differential Equations (PD

safetyarxiv-cs-ai
10 Jun 2026
Research

Superficial Beliefs in LLM Decision-Making

DGX agent

arXiv:2606.11016v1 Announce Type: new Abstract: We ask whether large language models (LLMs) merely imitate rationales when choosing between two options, or whether their choices reflect a systematic u

researcharxiv-cs-ai
10 Jun 2026
Applications

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction

DGX agent

arXiv:2606.10279v1 Announce Type: new Abstract: Supervised fine-tuning with synthetic rationale data is widely assumed to improve language model performance on clinical prediction tasks by teaching mo

applicationsarxiv-cs-ai
10 Jun 2026
Safety

Support sufficiency as action-sufficient compression: a single-cycle rate-regret formulation

DGX agent

arXiv:2606.09858v1 Announce Type: cross Abstract: Robust decision-making requires compression. A system that forms a rich support state cannot usually preserve its full structure at the point of actio

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving

DGX agent

arXiv:2602.23499v4 Announce Type: replace-cross Abstract: Collecting a high-quality dataset is a critical task that demands meticulous attention to detail, as overlooking certain aspects can render th

agentsarxiv-cs-ai
10 Jun 2026
Safety

TD-Grokking: Learning from Zero-Reward Problems by Training-Time Decomposition

DGX agent

arXiv:2606.09883v1 Announce Type: cross Abstract: Large language models (LLMs) have made remarkable progress in reasoning tasks, largely driven by post-training paradigms, especially reinforcement lea

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Temporal Context Conditioning for Seasonality-Aware Precipitation Nowcasting of High-Intensity Rainfall

DGX agent

arXiv:2606.09959v1 Announce Type: cross Abstract: Precipitation nowcasting is increasingly being approached with deep learning models that learn directly from recent radar observations. Although such

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Temporal Sheaf Neural Networks with Dynamic Orthogonal Transport

DGX agent

arXiv:2606.10071v1 Announce Type: cross Abstract: We introduce Temporal Sheaf Neural Networks (TSNN), a temporal link prediction framework that equips each node with a time-varying orthogonal frame an

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Test-time Adversarial Takeover: A Real-time Hijacking Interface against Robotic Diffusion Policies

DGX agent

arXiv:2606.10371v1 Announce Type: cross Abstract: Diffusion-based action generation has become a foundational component of embodied AI, but its reliance on visual conditioning leaves deployed visuomot

safetyarxiv-cs-ai
10 Jun 2026
Safety

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning

DGX agent

arXiv:2606.11087v1 Announce Type: cross Abstract: Expressive continuous control policies, such as diffusion and flow models, form the backbone of recent advances in scaling imitation learning for simu

safetyarxiv-cs-ai
10 Jun 2026
Agents

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

DGX agent

arXiv:2606.10747v1 Announce Type: new Abstract: As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiati

agentsarxiv-cs-ai
10 Jun 2026
Research

The Bioelectrical Information Theory: Investigating the theoretical compression limit of bioelectrical signals under artificial intelligence

DGX agent

arXiv:2606.09922v1 Announce Type: cross Abstract: Bioelectrical signals are increasingly acquired at scales that challenge the bandwidth of brain-computer interfaces. However, their compression is sti

researcharxiv-cs-ai
10 Jun 2026
Agents

The Confident Liar: Diagnosing Multi-Agent Debate with Log-Probabilities and LLM-as-Judge

DGX agent

arXiv:2606.10296v1 Announce Type: cross Abstract: Multi-agent debate systems are typically evaluated only on whether the final answer is correct, overlooking the quality of the intermediate reasoning

agentsarxiv-cs-ai
10 Jun 2026
Agents

The Distributed Detectability Band Against Marginal-Preserving Attacks

DGX agent

arXiv:2606.10456v1 Announce Type: cross Abstract: AI-control monitors score individual agent actions to detect misbehavior, but real harm can be distributed across many benign-looking steps, each indi

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

The Role of Feedback Alignment in Self-Distillation

DGX agent

arXiv:2606.11173v1 Announce Type: new Abstract: Conditioning a language model on additional context, such as feedback on a previous attempt, typically improves its response. Self-distillation trains t

safetyarxiv-cs-ai
10 Jun 2026
Safety

The Whale That Outswam Evolution: Swarm Intelligence Maximises Memory in Connectome Reservoirs

DGX agent

arXiv:2606.09902v1 Announce Type: cross Abstract: Reservoir computing exploits the fixed dynamics of a recurrent network for temporal processing, requiring only a trained linear readout. Biological ne

safetyarxiv-cs-ai
10 Jun 2026
Research

Time Series as Language: A Universal Tokenizer for General-Purpose Time Series Foundation Models

DGX agent

arXiv:2606.09861v1 Announce Type: cross Abstract: While Next-Token Prediction (NTP) has unified LLM pretraining, its adaptation to unbounded, continuous time series (TS) remains open. To bridge the ga

researcharxiv-cs-ai
10 Jun 2026
Hardware

torch-sla: Differentiable Sparse Linear Algebra with Adjoint Solvers and Sparse Tensor Parallelism for PyTorch

DGX agent

arXiv:2601.13994v3 Announce Type: replace-cross Abstract: Differentiable sparse linear algebra is foundational for scientific machine learning, yet PyTorch lacks a unified library for it: torch.sparse

hardwarearxiv-cs-ai
10 Jun 2026
Agents

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

DGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

agentsarxiv-cs-ai
10 Jun 2026
← Previous
1…173174175176177…448
Next →