AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

DGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

safetyarxiv-cs-cl
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

DGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

safetyarxiv-cs-ai
26 May 2026
Model Releases

Decision-Making with Lightweight Confidence-Aware Language Model for Autonomous Driving

DGX agent

arXiv:2605.25393v1 Announce Type: new Abstract: Large Language Models (LLMs) and Multimodal LLMs (MLLMs) have demonstrated immense potential in autonomous driving (AD) by offering human-like reasoning

model-releasesarxiv-cs-ro
26 May 2026
Safety

Eureka: Intelligent Feature Engineering for Enterprise AI Cloud Resource Demand Prediction

DGX agent

arXiv:2605.25297v1 Announce Type: cross Abstract: Effective features are crucial for predictive model performance, but creating them often requires domain expertise, limiting scalability across applic

safetyarxiv-cs-ai
26 May 2026
Model Releases

FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization

DGX agent

arXiv:2605.25246v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimizat

model-releasesarxiv-cs-ai
26 May 2026
Safety

Hide to Guide: Learning via Semantic Masking

DGX agent

arXiv:2605.25198v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a powerful paradigm for improving language models on reasoning-intensive tasks, but i

safetyarxiv-cs-ai
26 May 2026
Model Releases

How Well Do Models Follow Their Constitutions?

DGX agent

arXiv:2605.24229v1 Announce Type: new Abstract: Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Inference Time Context Sparsity: Illusion or Opportunity?

DGX agent

arXiv:2605.24168v1 Announce Type: new Abstract: Sparsity has long been a central theme in LLM efficiency, but its role in context processing remains unresolved. As LLM workloads shift toward longer co

hardwarearxiv-cs-ai
26 May 2026
Safety

Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol

DGX agent

arXiv:2605.24538v1 Announce Type: cross Abstract: Every major framework for governing artificial intelligence presupposes an identifiable entity -- a developer, deployer, or operator -- who can be hel

safetyarxiv-cs-ai
26 May 2026
Safety

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible

DGX agent

arXiv:2605.25739v1 Announce Type: new Abstract: We prove that no reinforcement learning policy with confidence-gated autonomy can simultaneously achieve maximum helpfulness, optimal calibration, and f

safetyarxiv-cs-lg
26 May 2026
Model Releases

BOHM: Zero-Cost Hierarchical Attribution for Compound AI Systems

DGX agent

arXiv:2605.22866v1 Announce Type: new Abstract: Compound AI systems route tasks through hierarchies of specialised components. Attribution is dominated by Shapley-based methods (SHAP), which decompose

model-releasesarxiv-cs-ai
25 May 2026
Safety

Classical State Preparation for Variational Quantum Algorithms via Reinforcement Learning

DGX agent

arXiv:2605.23138v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) potentially offer a pathway to practical quantum advantage, but their optimization is heavily hindered by barren

safetyarxiv-cs-ai
25 May 2026
Safety

Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models

DGX agent

arXiv:2509.06858v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly used to simulate human opinion dynamics, yet the effect of genuine interaction is often obscured by sys

safetyarxiv-cs-ai
25 May 2026
Model Releases

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

DGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

model-releasesarxiv-cs-ai
25 May 2026
Safety

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

DGX agent

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

safetyarxiv-cs-ro
25 May 2026
Safety

LACO: Adaptive Latent Communication for Collaborative Driving

DGX agent

arXiv:2605.22504v1 Announce Type: cross Abstract: Collaborative driving aims to improve safety and efficiency by enabling connected vehicles to coordinate under partial observability. Recent approache

safetyarxiv-cs-cv
22 May 2026
Model Releases

Open-World Evaluations for Measuring Frontier AI Capabilities

DGX agent

arXiv:2605.20520v1 Announce Type: new Abstract: Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it

model-releasesarxiv-cs-ai
22 May 2026
Hardware

Frontier: Towards Comprehensive and Accurate LLM Inference Simulation

DGX agent

arXiv:2605.21312v1 Announce Type: cross Abstract: Modern LLM serving is no longer homogeneous or monolithic. Production systems now combine disaggregated execution, complex parallelism, runtime optimi

hardwarearxiv-cs-lg
21 May 2026
Model Releases

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

DGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

model-releasesarxiv-cs-cv
21 May 2026
Safety

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

DGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

safetyarxiv-cs-lg
21 May 2026
Safety

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

DGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

safetyarxiv-cs-lg
21 May 2026
Local Ai

Mechanisms of Misgeneralization in Physical Sequence Modeling

DGX agent

arXiv:2605.20299v1 Announce Type: new Abstract: Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to

local-aiarxiv-cs-lg
21 May 2026
Safety

Tutor-Student Reinforcement Learning: A Dynamic Curriculum for Robust Deepfake Detection

DGX agent

arXiv:2603.24139v2 Announce Type: replace Abstract: Standard supervised training for deepfake detection treats all samples with uniform importance, which can be suboptimal for learning robust and gene

safetyarxiv-cs-cv
21 May 2026
Safety

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails

DGX agent

arXiv:2510.13727v2 Announce Type: replace Abstract: Generative AI systems are increasingly assisting and acting on behalf of end users in practical settings, from digital shopping assistants to next-g

safetyarxiv-cs-ai
20 May 2026
Safety

Guiding Neuro-Symbolic Scenario Generation with Spatio-Temporal Logic

DGX agent

arXiv:2605.19038v1 Announce Type: cross Abstract: The rapid advancement of autonomous driving (AD) technologies has outpaced the development of robust safety evaluation methods. Conventional testing r

safetyarxiv-cs-lg
20 May 2026
Safety

When Critics Disagree: Adaptive Reward Poisoning Attacks in RIS-Aided Wireless Control System

DGX agent

arXiv:2605.20037v1 Announce Type: cross Abstract: Reward-poisoning attacks present a significant risk to learning-based wireless control systems. Given this, we propose a Disagreement-Guided Reward Po

safetyarxiv-cs-ai
20 May 2026
Model Releases

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback

DGX agent

arXiv:2605.18073v1 Announce Type: cross Abstract: Large Language Models (LLMs) demonstrate strong potential for automated code generation, yet their ability to iteratively refine solutions using execu

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting

DGX agent

arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers and limited scalability. While

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

DGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

LEAF: A Living Benchmark for Event-Augmented Forecasting

DGX agent

arXiv:2605.16358v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to forecasting. To evaluate this capability while mitigating pre-training data contamination, se

model-releasesarxiv-cs-ai
19 May 2026
Safety

LISTEN to Your Preferences: An LLM Framework for Multi-Objective Selection

DGX agent

arXiv:2510.25799v2 Announce Type: replace Abstract: Human experts often struggle to select the best option from a large set of items with multiple competing objectives, a process bottlenecked by the d

safetyarxiv-cs-cl
19 May 2026
Safety

oldsymbol{f}-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control

DGX agent

arXiv:2605.17862v1 Announce Type: cross Abstract: Scaling on-policy distillation (OPD) for large language models (LLMs) confronts a fundamental tension: asynchronous execution is necessary for system

safetyarxiv-cs-ai
19 May 2026
Safety

QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning

DGX agent

arXiv:2605.16429v1 Announce Type: cross Abstract: We introduce extbf{QuantFPFlow}, a reinforcement learning framework that integrates quantum amplitude estimation into the Fokker--Planck~(FP) formulat

safetyarxiv-cs-ai
19 May 2026
Safety

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

DGX agent

arXiv:2605.17999v1 Announce Type: new Abstract: This paper proposes a Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) algorithm. By sharing the base module between the Actor and Cri

safetyarxiv-cs-ai
19 May 2026
Model Releases

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

DGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation

DGX agent

arXiv:2605.15585v1 Announce Type: new Abstract: Large language models can generate executable code for educational animations, but the resulting renders often exhibit visual defects, including element

local-aiarxiv-cs-ai
18 May 2026
Applications

FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale

DGX agent

arXiv:2605.14445v1 Announce Type: new Abstract: Many real-world coding challenges are open-ended and admit no known optimal solution. Yet, recent progress in LLM coding has focused on well-defined tas

applicationsarxiv-cs-lg
15 May 2026
Safety

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

DGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

safetyarxiv-cs-ai
14 May 2026
Safety

Flow Matching for Offline Reinforcement Learning with Discrete Actions

DGX agent

arXiv:2602.06138v2 Announce Type: replace Abstract: Generative policies based on diffusion models and flow matching have shown strong promise for offline reinforcement learning (RL), but their applica

safetyarxiv-cs-lg
14 May 2026
Safety

In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores

DGX agent

arXiv:2605.12530v1 Announce Type: cross Abstract: LLM fairness should be evaluated through in-situ conversational behavior rather than standardized-test Q&A benchmarks. We show that the standardized-t

safetyarxiv-cs-ai
14 May 2026
Safety

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

DGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

safetyarxiv-cs-ai
14 May 2026
Safety

Leveraging RAG for Training-Free Alignment of LLMs

DGX agent

arXiv:2605.11217v1 Announce Type: new Abstract: Large language model (LLM) alignment algorithms typically consist of post-training over preference pairs. While such algorithms are widely used to enabl

safetyarxiv-cs-lg
13 May 2026
Model Releases

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

DGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

DGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

model-releasesarxiv-cs-lg
12 May 2026
Safety

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

DGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

safetyarxiv-cs-lg
12 May 2026
Research

Communicating Sound Through Natural Language

DGX agent

arXiv:2605.08750v1 Announce Type: cross Abstract: Natural language is widely used to describe, prompt, and control audio systems, but rarely serves as the representation carrying audio itself. We intr

researcharxiv-cs-ai
12 May 2026
Tutorials

EFGCL: Learning Dynamic Motion through Spotting-Inspired External Force Guided Curriculum Learning

DGX agent

arXiv:2605.10063v1 Announce Type: new Abstract: Learning dynamic whole-body motions for legged robots through reinforcement learning (RL) remains challenging due to the high risk of failure, which mak

tutorialsarxiv-cs-ro
12 May 2026
← Previous
1…192193194195196…233
Next →