AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge

DGX agent

arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete parametric knowledge. Retrieval-augmented genera

model-releasesarxiv-cs-ai
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Green Energy Management for Sustainable Data Centers Using Deep Reinforcement Learning

DGX agent

arXiv:2507.21153v2 Announce Type: replace Abstract: The exponential growth of digital services has positioned data centers among the most energy-intensive infrastructures in the modern economy, raisin

safetyarxiv-cs-lg
5 May 2026
Safety

High entropy leads to symmetry equivariant policies in Dec-POMDPs

DGX agent

arXiv:2511.22581v3 Announce Type: replace Abstract: We prove that in any Dec-POMDP, sufficiently high entropy regularization ensures that the policy gradient flow with tabular softmax parametrization

safetyarxiv-cs-lg
5 May 2026
Safety

Hybrid Quantum Reinforcement Learning with QAOA for Improved Vehicle Routing Optimization

DGX agent

arXiv:2605.01574v1 Announce Type: new Abstract: Vehicle Routing Problem (VRP) is one of the most complex NP-hard combinatorial optimization problem in transportation and logistics that requires a dyna

safetyarxiv-cs-lg
5 May 2026
Safety

Knowledge-Based Design Requirements for Generative Social Robots in Higher Education

DGX agent

arXiv:2602.12873v4 Announce Type: replace-cross Abstract: Generative social robots (GSRs) powered by large language models enable adaptive, conversational tutoring but also introduce risks such as mis

safetyarxiv-cs-ai
5 May 2026
Model Releases

MolViBench: Evaluating LLMs on Molecular Vibe Coding

DGX agent

arXiv:2605.02351v1 Announce Type: new Abstract: Molecular Vibe Coding, a paradigm where chemists interact with LLMs to generate executable programs for molecular tasks, has emerged as a flexible alter

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

OpenAI GPT-5 System Card

DGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation

DGX agent

arXiv:2605.02537v1 Announce Type: new Abstract: Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological p

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

PACE: Parameter Change for Unsupervised Environment Design

DGX agent

arXiv:2605.01358v1 Announce Type: new Abstract: Unsupervised Environment Design (UED) offers a promising paradigm for improving reinforcement learning generalization by adaptively shaping training env

model-releasesarxiv-cs-lg
5 May 2026
Applications

RAST-MoE-RL: A Regime-Aware Spatio-Temporal MoE Framework for Deep Reinforcement Learning in Ride-Hailing

DGX agent

arXiv:2512.13727v2 Announce Type: replace Abstract: Ride-hailing platforms face the challenge of balancing passenger waiting times with overall system efficiency under highly uncertain supply-demand c

applicationsarxiv-cs-lg
5 May 2026
Safety

Reliability-Oriented Multilingual Orthopedic Diagnosis: A Domain-Adaptive Modeling and a Conceptual Validation Framework

DGX agent

arXiv:2605.02266v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly proposed for clinical decision support including multilingual diagnosis in low-resource settings. However,

safetyarxiv-cs-cl
5 May 2026
Hardware

SplitZip: Ultra Fast Lossless KV Compression for Disaggregated LLM Serving

DGX agent

arXiv:2605.01708v1 Announce Type: cross Abstract: Contemporary systems serving large language models (LLMs) have adopted prefill-decode disaggregation to better load-balance between the compute-bound

hardwarearxiv-cs-lg
5 May 2026
Safety

STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

DGX agent

arXiv:2605.02122v1 Announce Type: new Abstract: Human evaluation remains the primary standard for assessing modern AI systems, yet annotator disagreement, bias, and variability make system rankings fr

safetyarxiv-cs-lg
5 May 2026
Model Releases

STAGE: A Full-Screenplay Benchmark for Reasoning over Evolving Storie

DGX agent

arXiv:2601.08510v3 Announce Type: replace Abstract: Movie screenplays are rich long-form narratives that interleave complex character relationships, temporally ordered events, and dialogue-driven inte

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation

DGX agent

arXiv:2508.15658v5 Announce Type: replace Abstract: The rapid growth of academic literature makes the manual creation of scientific surveys increasingly infeasible. While large language models show pr

model-releasesarxiv-cs-cl
5 May 2026
Research

Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments

DGX agent

arXiv:2505.09901v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate or automate human behavior in complex sequential decision-making settings. A na

researcharxiv-cs-cl
4 May 2026
Safety

Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations

DGX agent

arXiv:2512.20260v5 Announce Type: replace Abstract: Weakly-Supervised Camouflaged Object Detection (WSCOD) aims to locate and segment objects that are visually concealed within their surrounding scene

safetyarxiv-cs-cv
4 May 2026
Safety

Decentralized Proximal Stochastic Gradient Langevin Dynamics

DGX agent

arXiv:2605.00723v1 Announce Type: cross Abstract: We propose Decentralized Proximal Stochastic Gradient Langevin Dynamics (DE-PSGLD), a decentralized Markov chain Monte Carlo (MCMC) algorithm for samp

safetyarxiv-cs-lg
4 May 2026
Model Releases

FollowTable: A Benchmark for Instruction-Following Table Retrieval

DGX agent

arXiv:2605.00400v1 Announce Type: cross Abstract: Table Retrieval (TR) has traditionally been formulated as an ad-hoc retrieval problem, where relevance is primarily determined by topical semantic sim

model-releasesarxiv-cs-cl
4 May 2026
Safety

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding

DGX agent

arXiv:2605.00642v1 Announce Type: cross Abstract: Graphical User Interface (GUI) grounding maps natural language instructions to the visual coordinates of target elements and serves as a core capabili

safetyarxiv-cs-cv
4 May 2026
Safety

Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values

DGX agent

arXiv:2605.00762v1 Announce Type: new Abstract: We propose a new framework for meritocratic fairness in budgeted combinatorial multi-armed bandits with full-bandit feedback (BCMAB-FBF). Unlike semi-ba

safetyarxiv-cs-lg
4 May 2026
Model Releases

Minimizing Human Intervention in Online Classification

DGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

model-releasesarxiv-cs-lg
4 May 2026
Safety

Optimizing Resource-Constrained Non-Pharmaceutical Interventions for Multi-Cluster Outbreak Control Using Hierarchical Reinforcement Learning

DGX agent

arXiv:2603.19397v2 Announce Type: replace Abstract: Non-pharmaceutical interventions (NPIs), such as diagnostic testing and quarantine, are crucial for controlling infectious disease outbreaks but are

safetyarxiv-cs-lg
4 May 2026
Safety

Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization

DGX agent

arXiv:2604.07669v2 Announce Type: replace Abstract: Lead optimization in drug discovery requires improving therapeutic properties while ensuring that molecular modifications correspond to feasible syn

safetyarxiv-cs-lg
4 May 2026
Safety

ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning

DGX agent

arXiv:2605.00380v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) enhances reasoning of Large Language Models (LLMs) but usually exhibits limited generation diver

safetyarxiv-cs-cl
4 May 2026
Model Releases

Semantic Level of Detail for Knowledge Graphs: Discovering Abstraction Boundaries via Spectral Heat Diffusion

DGX agent

arXiv:2603.08965v2 Announce Type: replace Abstract: Graph-structured knowledge systems -- from knowledge graphs to GraphRAG pipelines -- organize information into hierarchical communities, yet lack a

model-releasesarxiv-cs-lg
4 May 2026
Safety

World Model for Robot Learning: A Comprehensive Survey

DGX agent

arXiv:2605.00080v1 Announce Type: cross Abstract: World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They s

safetyarxiv-cs-cv
4 May 2026
Model Releases

AppTek Call-Center Dialogues: A Multi-Accent Long-Form Benchmark for English ASR

DGX agent

arXiv:2604.27543v1 Announce Type: new Abstract: Evaluating English ASR systems for conversational AI applications remains difficult, as many publicly available corpora are either pre-segmented into sh

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

DGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

model-releasesarxiv-cs-lg
1 May 2026
Safety

Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations

DGX agent

arXiv:2604.27372v1 Announce Type: cross Abstract: This paper investigates the continuous-time counterpart of the Q-function for entropy-regularized mean-field control (MFC) with controlled common nois

safetyarxiv-cs-lg
1 May 2026
Model Releases

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction

DGX agent

arXiv:2604.27906v1 Announce Type: new Abstract: Persistent AI memory is often reduced to a retrieval problem: store prior interactions as text, embed them, and ask the model to recover relevant contex

model-releasesarxiv-cs-ai
1 May 2026
Applications

IACDM: Interactive Adversarial Convergence Development Methodology -- A Structured Framework for AI-Assisted Software Development

DGX agent

arXiv:2604.16399v2 Announce Type: replace-cross Abstract: The widespread adoption of AI-assisted development tools in 2025 -- and the emergence of vibe coding, a practice of generating complete applic

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

DGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

DGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

model-releasesarxiv-cs-ai
1 May 2026
Safety

Knowledge Graph Representations for LLM-Based Policy Compliance Reasoning

DGX agent

arXiv:2604.27713v1 Announce Type: new Abstract: The risks posed by AI features are increasing as they are rapidly integrated into software applications. In response, regulations and standards for safe

safetyarxiv-cs-ai
1 May 2026
Model Releases

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

DGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

model-releasesarxiv-cs-ai
1 May 2026
Hardware

Predictive Multi-Tier Memory Management for KV Cache in Large-Scale GPU Inference

DGX agent

arXiv:2604.26968v1 Announce Type: cross Abstract: Key-value (KV) cache memory management is the primary bottleneck limiting throughput and cost-efficiency in large-scale GPU inference serving. Current

hardwarearxiv-cs-ai
1 May 2026
Model Releases

TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question Answering

DGX agent

arXiv:2604.28076v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced Table Question Answering, where most queries can be answered by extracting information or simple aggregatio

model-releasesarxiv-cs-ai
1 May 2026
Safety

A Scaled Three-Vehicle Platooning Platform

DGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

safetyarxiv-cs-ro
30 Apr 2026
Safety

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

DGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

safetyarxiv-cs-ai
30 Apr 2026
Hardware

AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving

DGX agent

arXiv:2604.26103v1 Announce Type: cross Abstract: All current LLM serving systems place the GPU at the center, from production-level attention-FFN disaggregation to NVIDIA's Rubin GPU-LPU heterogeneou

hardwarearxiv-cs-ai
30 Apr 2026
Model Releases

Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI

DGX agent

arXiv:2604.26382v1 Announce Type: cross Abstract: Most enterprise document AI today is a pipeline. Parse, index, retrieve, generate. Each of those stages has been studied to death on its own -- what's

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Decide less, communicate more: On the construct validity of end-to-end fact-checking in medicine

DGX agent

arXiv:2506.20876v4 Announce Type: replace Abstract: Technological progress has led to concrete advancements in tasks that were regarded as challenging, such as automatic fact-checking. Interest in ado

researcharxiv-cs-cl
30 Apr 2026
Model Releases

Entropy Centroids as Intrinsic Rewards for Test-Time Scaling

DGX agent

arXiv:2604.26173v1 Announce Type: cross Abstract: An effective way to scale up test-time compute of large language models is to sample multiple responses and then select the best one, as in Grok Heavy

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Evergreen: Efficient Claim Verification for Semantic Aggregates

DGX agent

arXiv:2604.26180v1 Announce Type: cross Abstract: With recent semantic query processing engines, semantic aggregation has become a primitive operator, enabling the reduction of a relation into a natur

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

FlowS: One-Step Motion Prediction via Local Transport Conditioning

DGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

model-releasesarxiv-cs-ro
30 Apr 2026
Safety

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

DGX agent

arXiv:2510.08547v2 Announce Type: replace-cross Abstract: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to wor

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

DGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

model-releasesarxiv-cs-cv
30 Apr 2026
← Previous
1…220221222223224…230
Next →