AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

DGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

model-releasesarxiv-cs-cl
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

Chain-of-thought obfuscation learned from output supervision can generalise to unseen tasks

DGX agent

arXiv:2601.23086v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning provides a significant performance uplift to LLMs by enabling planning, exploration, and deliberation of their acti

tutorialsarxiv-cs-ai
22 May 2026
Model Releases

ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning

DGX agent

arXiv:2605.22734v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) treat disease associations as static facts, but temporal information is crucial for clinical reasoning, e.g., a sympto

model-releasesarxiv-cs-cl
22 May 2026
Research

Mind the Gaps: Multi-Robot Feedback-Driven Ergodic Coverage in Unknown Environments

DGX agent

arXiv:2605.21719v1 Announce Type: new Abstract: In this work, we address the problem of multi-robot adaptive coverage, where teams of robots perform dynamic sampling by continuously adjusting their po

researcharxiv-cs-ro
22 May 2026
Model Releases

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

DGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

DGX agent

arXiv:2605.22109v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in human-facing roles where personality perception is critical, yet existing benchm

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Putnam 2025 Problems in Rocq using Opus 4.6 and Rocq-MCP

DGX agent

arXiv:2603.20405v2 Announce Type: replace-cross Abstract: We report on an experiment in which Claude Opus~4.6, equipped with a suite of Model Context Protocol (MCP) tools for the Rocq proof assistant,

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Residual Skill Optimization for Text-to-SQL Ensembles

DGX agent

arXiv:2605.21792v1 Announce Type: new Abstract: Text-to-SQL ensembles improve over single-candidate generation by drawing multiple SQL candidates and selecting one, but their effectiveness is bounded

model-releasesarxiv-cs-cl
22 May 2026
Safety

ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driving

DGX agent

arXiv:2605.21168v1 Announce Type: new Abstract: Safety-critical scenarios are central to evaluating autonomous driving systems, yet their rarity in naturalistic logs makes simulation-based stress test

safetyarxiv-cs-ai
22 May 2026
Safety

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

DGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

safetyarxiv-cs-ro
22 May 2026
Model Releases

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

DGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Causal Path Alignment: Anchoring the Optimization Trajectory for Controllable In-Parameter Knowledge Editing

DGX agent

arXiv:2506.04042v2 Announce Type: replace Abstract: Knowledge editing is pivotal for efficiently updating the parametric memory of Large Language Models (LLMs), enabling them to function as evolving a

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

FedCritic: Serverless Federated Critic Learning-based Resource Allocation for Multi-Cell OFDMA in 6G

DGX agent

arXiv:2605.21418v1 Announce Type: cross Abstract: In sixth-generation (6G) ultra-dense networks, aggressive frequency reuse amplifies inter-cell interference (ICI), making multi-cell orthogonal freque

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

DGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

DGX agent

arXiv:2605.20729v1 Announce Type: new Abstract: Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational

model-releasesarxiv-cs-cl
21 May 2026
Safety

Multi-Head Attention as Ensemble Nadaraya-Watson Estimation: Variance Reduction, Decorrelation, and Optimal Head Diversity

DGX agent

arXiv:2605.20271v1 Announce Type: cross Abstract: We develop a rigorous statistical theory of multi-head attention (MHA) as an ensemble of Nadaraya-Watson (NW) kernel regression estimators. Building o

safetyarxiv-cs-lg
21 May 2026
Model Releases

On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists

DGX agent

arXiv:2605.20668v1 Announce Type: new Abstract: With the advancement of AI capabilities, AI reviewers are beginning to be deployed in scientific peer review, yet their capability and credibility remai

model-releasesarxiv-cs-cl
21 May 2026
Safety

Q-SpiRL: Quantum Spiking Reinforcement Learning for Adaptive Robot Navigation

DGX agent

arXiv:2605.20801v1 Announce Type: new Abstract: Adaptive robot navigation in dynamic environments requires policies that can reach the target reliably while producing efficient and stable trajectories

safetyarxiv-cs-ro
21 May 2026
Research

Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes

DGX agent

arXiv:2605.21211v1 Announce Type: cross Abstract: In this work we present an efficient and practically implementable approach for the application of reinforcement learning (RL)-based control in chemic

researcharxiv-cs-lg
21 May 2026
Safety

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces

DGX agent

arXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov

safetyarxiv-cs-lg
21 May 2026
Model Releases

roto 2.0: The Robot Tactile Olympiad

DGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

model-releasesarxiv-cs-lg
21 May 2026
Safety

The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study

DGX agent

arXiv:2605.20767v1 Announce Type: new Abstract: Large language models (LLMs) show potential as simulators of human behavior, offering a scalable way to study responses to interventions. However, becau

safetyarxiv-cs-cl
21 May 2026
Applications

Validating Navmesh using Geometry: Voxel-Based Analysis with Prioritized Exploration

DGX agent

arXiv:2605.21397v1 Announce Type: cross Abstract: Navigation mesh (Navmesh) inconsistencies affect the player experience by directly impacting the navigation systems used by non-playable characters (N

applicationsarxiv-cs-ro
21 May 2026
Model Releases

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

DGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

model-releasesarxiv-cs-cv
20 May 2026
Safety

Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay

DGX agent

arXiv:2605.19352v1 Announce Type: cross Abstract: Understanding how humans and artificial intelligence systems predict and plan by interacting with their environment is a fundamental challenge at the

safetyarxiv-cs-ai
20 May 2026
Model Releases

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

DGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models

DGX agent

arXiv:2605.18824v1 Announce Type: cross Abstract: Evaluation of foundation models often rely on aggregate scores from benchmarks that lack comprehensive coverage and metadata for a fine-grained evalua

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models

DGX agent

arXiv:2605.19341v1 Announce Type: cross Abstract: Hallucination remains a central failure mode of large language models, but existing benchmarks operationalize it inconsistently across summarization,

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

DGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

optimize_anything: A Universal API for Optimizing any Text Parameter

DGX agent

arXiv:2605.19633v1 Announce Type: cross Abstract: Can a single LLM-based optimization system match specialized tools across fundamentally different domains? We show that when optimization problems are

model-releasesarxiv-cs-ai
20 May 2026
Research

Projecting Latent RL Actions: Towards Generalizable and Scalable Graph Combinatorial Optimization

DGX agent

arXiv:2605.19721v1 Announce Type: new Abstract: Graph combinatorial optimization (GCO) has attracted growing interest, as many NP-hard problems naturally admit graph formulations, yet their combinator

researcharxiv-cs-ai
20 May 2026
Model Releases

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

DGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

model-releasesarxiv-cs-ai
20 May 2026
Safety

Safe Continual Reinforcement Learning under Nonstationarity via Adaptive Safety Constraints

DGX agent

arXiv:2605.18842v1 Announce Type: new Abstract: Safe reinforcement learning in nonstationary environments requires safety mechanisms that adapt as environmental conditions change. Standard safe reinfo

safetyarxiv-cs-lg
20 May 2026
Safety

SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects

DGX agent

arXiv:2605.19587v1 Announce Type: new Abstract: Indoor scene synthesis underpins embodied AI, robotic manipulation, and simulation-based policy evaluation, where a useful scene must specify not only w

safetyarxiv-cs-ai
20 May 2026
Safety

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making

DGX agent

arXiv:2602.07008v2 Announce Type: replace Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typical

safetyarxiv-cs-cv
20 May 2026
Model Releases

A Production-Ready RL Framework for Personalized Utility Tuning with Pareto Sweeping in Pinterest Recommender Systems

DGX agent

arXiv:2605.16344v1 Announce Type: cross Abstract: Large-scale recommenders encode multi-objective trade-offs by combining multiple predicted outcomes into a single utility score. Although this utility

model-releasesarxiv-cs-lg
19 May 2026
Safety

A Visual Reinforcement Learning-Based Separate Primitive Policy for Peg-in-Hole Tasks

DGX agent

arXiv:2504.14820v2 Announce Type: replace Abstract: For peg-in-hole tasks, humans rely on binocular visual perception to locate the peg above the hole surface and then proceed with insertion. This pap

safetyarxiv-cs-ro
19 May 2026
Model Releases

AI4BayesCode: From Natural Language Descriptions to Validated Modular Stateful Bayesian Samplers

DGX agent

arXiv:2605.18476v1 Announce Type: cross Abstract: Coding and computation remain major bottlenecks in Markov chain Monte Carlo (MCMC) workflows, especially as modern sampling algorithms have become inc

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

ALIGN: A Vision-Language Framework for High-Accuracy Accident Location Inference through Geo-Spatial Neural Reasoning

DGX agent

arXiv:2511.06316v3 Announce Type: replace Abstract: In low- and middle-income countries, public safety and urban planning initiatives frequently face a critical shortage of accurate, location-specific

local-aiarxiv-cs-ai
19 May 2026
Research

Are Researchers Being Replaced by Artificial Intelligence?

DGX agent

arXiv:2605.16294v1 Announce Type: cross Abstract: A Nature survey from 2023 involving 1,600 researchers shows that scientists are ``concerned, as well as excited, by the increasing use of artificial-i

researcharxiv-cs-ai
19 May 2026
Model Releases

ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery

DGX agent

arXiv:2605.16902v1 Announce Type: new Abstract: Scientific artifacts such as models and datasets are foundations for research. With the rapid growth of platforms like HuggingFace, researchers now have

model-releasesarxiv-cs-lg
19 May 2026
Safety

Automatic Generation of High-Performance RL Environments

DGX agent

arXiv:2603.12145v2 Announce Type: replace-cross Abstract: Translating complex reinforcement learning (RL) environments into high-performance implementations has traditionally required months of specia

safetyarxiv-cs-ai
19 May 2026
Safety

Beyond Safety Filtering: Control Barrier Function-Informed Reinforcement Learning for Connected and Automated Vehicles

DGX agent

arXiv:2605.16894v1 Announce Type: new Abstract: Reinforcement Learning (RL) uses rewards to guide learning, yet reward design is typically hand-crafted using heuristics that can be difficult to tune.

safetyarxiv-cs-ro
19 May 2026
Model Releases

CAM-Bench: A Benchmark for Computational and Applied Mathematics in Lean

DGX agent

arXiv:2605.17255v1 Announce Type: new Abstract: Formal theorem-proving benchmarks enable mechanically verifiable evaluation of mathematical reasoning in large language models. However, existing benchm

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench

DGX agent

arXiv:2605.17079v1 Announce Type: cross Abstract: LLMs are increasingly used as ``digital consumers'' to simulate public opinion, pre-test marketing decisions, and anticipate audience response. Howeve

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

CMAG: Concept-Scaffolded Retrieval for Marketplace Avatar Generation

DGX agent

arXiv:2605.18680v1 Announce Type: new Abstract: Metaverse platforms rely on creator-driven marketplaces where avatars are assembled from discrete, taxonomy-labeled 3D assets (e.g., tops, bottoms, shoe

local-aiarxiv-cs-cv
19 May 2026
← Previous
1…215216217218219…230
Next →