AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading

DGX agent

arXiv:2606.27032v1 Announce Type: cross Abstract: Energy trading decisions depend not only on current market prices, but also on expected future market conditions, and operational constraints. This ma

safetyarxiv-cs-ai
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

DGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

model-releasesarxiv-cs-lg
25 Jun 2026
Tutorials

Lifelong In-Context Learning with Transformers Requires Parametric Forms of Attention

DGX agent

arXiv:2606.25342v1 Announce Type: new Abstract: Lifelong continual learning remains an obstacle on the path to human-like intelligence. Modern transformers show sparks of intelligence with in-context

tutorialsarxiv-cs-lg
25 Jun 2026
Safety

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

DGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

safetyarxiv-cs-lg
25 Jun 2026
Local Ai

Swarm-Inspired Generation of Collective Behaviors in Graph Dynamical Systems

DGX agent

arXiv:2606.24958v1 Announce Type: new Abstract: Collective behavior arises when locally interacting units produce coordinated global organization, from synchronization in dynamical systems to task-rel

local-aiarxiv-cs-lg
25 Jun 2026
Safety

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

DGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

safetyarxiv-cs-ai
25 Jun 2026
Model Releases

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

DGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

An Introduction to Causal Reinforcement Learning

DGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

DGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

DGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

safetyarxiv-cs-lg
23 Jun 2026
Research

Leaderless Collective Motion in Affine Formation Control over the Complex Plane

DGX agent

arXiv:2604.05648v2 Announce Type: replace Abstract: We propose a method for the collective maneuvering of affine formations in the plane by modifying the original weights of the Laplacian matrix used

researcharxiv-cs-ro
23 Jun 2026
Model Releases

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

DGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Scaling Self-Play for End-to-End Driving

DGX agent

arXiv:2606.19641v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving models are typically trained on offline human-demonstration datasets that provide limited state coverage and oft

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Select-to-Act: Hierarchical Reinforcement Learning via Adaptive Language Guidance

DGX agent

arXiv:2606.22350v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been widely applied to sequential decision-making, yet it often suffers from poor sample efficiency due to costly intera

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Two-Bridge: Exclusive Objectives and Extended Horizon StarCraft II Benchmark

DGX agent

arXiv:2603.06608v2 Announce Type: replace-cross Abstract: The research community lacks a middle ground between StarCraft II full game and its mini-games. The full-game's sprawling state-action space r

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct

DGX agent

arXiv:2606.23543v1 Announce Type: cross Abstract: Scaling reinforcement learning for visual mathematical reasoning requires more than generating harder questions: as data volume grows, the reward labe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

From Content to Knowledge: Lightning Fast Long-Video Understanding with Neural Knowledge Representations

DGX agent

arXiv:2606.11913v1 Announce Type: new Abstract: We propose a new paradigm for long video understanding by treating a long video as a Neural Knowledge Representation (NKR). NKR represents video content

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

MPC-Patch-Bench: Security-Aware LLM Code Patch for Multi-Party Computation

DGX agent

arXiv:2606.11416v1 Announce Type: cross Abstract: Repository-level benchmarks for evaluating Large Language Model (LLM) code repair on Secure Multi-Party Computation (MPC) software do not yet exist, a

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

DGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

AgenticNav: Zero-Shot Vision-and-Language Navigation as a Tool-Calling Harness

DGX agent

arXiv:2606.10577v1 Announce Type: new Abstract: Zero-shot vision-and-language navigation in continuous environments (VLN-CE) has recently become feasible with large vision-language models (VLMs). Howe

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Evaluating Research-Level Math Proofs via Strict Step-Level Verification

DGX agent

arXiv:2606.10799v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle to rigorously verify complex mathematical proofs. Standard global evaluation approaches suffer from 'context poiso

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Event-Driven Reinforcement Learning Enables Long-Horizon Control in Semiconductor Fabrication

DGX agent

arXiv:2606.10705v1 Announce Type: cross Abstract: Reinforcement learning promises to optimize sequential decisions in large-scale systems. Semiconductor manufacturing systems are stochastic and highly

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

FailureScope: Cross-Regime Behavioral Diagnosis of Language Model Weaknesses

DGX agent

arXiv:2606.09878v1 Announce Type: new Abstract: Standard benchmarks report aggregate accuracy, but practitioners need to know which specific capabilities a model lacks. We introduce FailureScope, a be

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Geometry-Aware Reinforcement Learning for 2D Irregular Nesting

DGX agent

arXiv:2606.10611v1 Announce Type: cross Abstract: Traditional heuristic solvers for the 2D irregular nesting problem share a fundamental limitation: they are blind to polygon geometry, relying on guid

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake

DGX agent

arXiv:2606.10460v1 Announce Type: cross Abstract: Recent large language models (LLMs) have shown rapid progress in reading-based question answering (QA), where evidence is explicitly provided or can b

model-releasesarxiv-cs-ai
10 Jun 2026
Local Ai

Multi-task LLMs for Bug Classification: Efficient Inference with Auxiliary Decoding Heads

DGX agent

arXiv:2606.09956v1 Announce Type: cross Abstract: The rapid adoption of LLM-powered code generation has dramatically accelerated software development, yet effective verification methods remain severel

local-aiarxiv-cs-lg
10 Jun 2026
Model Releases

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

DGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents

DGX agent

arXiv:2606.08590v1 Announce Type: cross Abstract: Kubernetes incidents are diagnosed reliably only when a root-cause system's reported gains come from incident evidence rather than scenario-specific s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ComplexConstraints and Beyond: Expert Rubrics for RLVR

DGX agent

arXiv:2606.09118v1 Announce Type: new Abstract: As LLM capabilities advance rapidly, the evaluation methods used to assess them increasingly lag behind. Traditional benchmarks relied on programmatic v

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Decentralized End-to-End Multi-AAV Pursuit Using Predictive Spatio-Temporal Observation via Deep Reinforcement Learning

DGX agent

arXiv:2603.24238v2 Announce Type: replace Abstract: Decentralized cooperative pursuit in cluttered environments is challenging for autonomous aerial swarms, especially under partial and noisy percepti

safetyarxiv-cs-ro
9 Jun 2026
Safety

Decoupling Semantics and Logic: A Training-Free Coarse-to-Fine Pipeline for Video Retrieval-Augmented Generation

DGX agent

arXiv:2606.07924v1 Announce Type: cross Abstract: This paper presents our system description for the 2nd Workshop on Multimodal Augmented Generation via MultimodAl Retrieval (MAGMaR). Addressing the c

safetyarxiv-cs-ai
9 Jun 2026
Safety

Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration

DGX agent

arXiv:2606.08596v1 Announce Type: new Abstract: Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of w

safetyarxiv-cs-ai
9 Jun 2026
Local Ai

Dynamic Distributed Constraint Optimization and Metareasoning for Continual, Large-Scale Satellite Operations

DGX agent

arXiv:2601.06188v3 Announce Type: replace Abstract: As Earth-observing satellite constellations grow in size and capability, distributed onboard control offers a pathway to novel responses and time-se

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

End-to-End Context Compression at Scale

DGX agent

arXiv:2606.09659v1 Announce Type: cross Abstract: Long-context language model inference is bottlenecked by memory, as the KV cache grows with context length. Recent techniques to compress the KV cache

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Implementing Grassroots Logic Programs with Multiagent Transition Systems and AI (Full Version)

DGX agent

arXiv:2602.06934v4 Announce Type: replace-cross Abstract: Grassroots Logic Programs (GLP) is a concurrent logic programming language in which logic variables are partitioned into paired readers and wr

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Instrumental convergence and power-seeking

DGX agent

arXiv:2606.08832v1 Announce Type: new Abstract: Recent years have seen increasing concern that artificial intelligence may soon pose an existential risk to humanity. One leading ground for concern is

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

PEDRA: Evaluating the Realism of Pedestrian Dynamics in Video Generation

DGX agent

arXiv:2510.20182v2 Announce Type: replace Abstract: Pedestrian simulation traditionally relies on expert-tuned, hand-crafted models that limit scalability and generalization. Meanwhile, large-scale vi

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

DGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

DGX agent

arXiv:2606.08481v1 Announce Type: cross Abstract: Enterprise property graphs vary widely in schema structure, internal terminology, domain assumptions, governance constraints, and user interaction pat

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation

DGX agent

arXiv:2606.07723v1 Announce Type: new Abstract: Open-vocabulary long-horizon manipulation requires robots to reason over flexible instructions and complex multi-object scenes while adaptively planning

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

Learn to Match: Two-Sided Matching with Temporally Extended Feedback

DGX agent

arXiv:2606.06744v1 Announce Type: new Abstract: Two-sided matching markets often involve information that unfolds over time through interviews, repeated interaction, learning, and separation. Existing

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Uncertainty-Aware LLM-Guided Policy Shaping for Sparse-Reward Reinforcement Learning

DGX agent

arXiv:2606.06673v1 Announce Type: new Abstract: Sparse rewards and heterogeneous task sequences remain persistent challenges in Reinforcement Learning (RL), often resulting in slow convergence, weak g

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Learning Adaptive Parallel Execution for Efficient Code Localization

DGX agent

arXiv:2601.19568v2 Announce Type: replace Abstract: Code localization constitutes a key bottleneck in automated software development pipelines. While concurrent tool execution can enhance discovery sp

researcharxiv-cs-ai
6 Jun 2026
Model Releases

WorldFly: A World-Model-Based Vision-Language-Action Model for UAV Navigation

DGX agent

arXiv:2606.06147v1 Announce Type: new Abstract: End-to-end Vision-Language-Action (VLA) models have shown promise in UAV navigation. However, existing approaches typically rely on historical observati

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional Writing

DGX agent

arXiv:2602.23845v2 Announce Type: replace Abstract: Chinese text correction has traditionally focused on spelling and grammar, while factual error correction is usually treated separately. However, in

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments

DGX agent

arXiv:2606.05661v1 Announce Type: cross Abstract: Continual learning, the ability of AI systems to improve through sequential experience, has attracted substantial interest, but no high-quality benchm

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset

DGX agent

arXiv:2606.06338v1 Announce Type: new Abstract: Video question answering (VideoQA) aims to answer questions about given videos. While existing approaches excel on factoid VideoQA, they struggle with d

model-releasesarxiv-cs-cv
5 Jun 2026
← Previous
1…190191192193194…233
Next →