AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
Human
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Research

Reachability and asymptotics of Gaussian Transformer dynamics

DGX agent

arXiv:2606.07600v1 Announce Type: cross Abstract: We formulate data propagation through the Transformer, the machine learning architecture powering large language models, as a nonlinear control system

researcharxiv-cs-ai
9 Jun 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

REACT 2026: The Fourth Multiple Appropriate Facial Reaction Generation Challenge: Personalised MAFRG and Appropriate EEG Reaction Prediction

DGX agent

arXiv:2606.07935v1 Announce Type: new Abstract: In dyadic interactions, various human facial reactions could be appropriate for responding to each human speaker behaviour. Following the successful org

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Readable Yet Unpredictable: Rotated-Outcome Prediction in Vision-Language Models

DGX agent

arXiv:2606.07641v1 Announce Type: new Abstract: Can vision-language models predict what a 180{eg} rotation would reveal from the original image alone? We study this ability through Rotated-Outcome Pre

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Real-IKEA: Physical Fidelity is the Prerequisite for Robust Manipulation

DGX agent

arXiv:2606.08564v1 Announce Type: new Abstract: Robotic manipulation robustness often founders on the physics gap between simplified simulations and the resistance-laden real world. In this work, we e

model-releasesarxiv-cs-ro
9 Jun 2026
Applications

Real-Time and Accurate Collision-Free Teleoperation via Differentiable Constraint-Based Trajectory Planning

DGX agent

arXiv:2606.08725v1 Announce Type: new Abstract: In teleoperation, the human operator typically controls only the end-effector pose, which often leads to self-collisions of the manipulator and collisio

applicationsarxiv-cs-ro
9 Jun 2026
Model Releases

Real-time body pose non-verbal communication with a consistency-based reliability measure

DGX agent

arXiv:2606.09390v1 Announce Type: cross Abstract: Body movement communicates intent at distances and in conditions where neither the face, nor speech can be captured. We study the recognition of commu

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning

DGX agent

arXiv:2606.09303v1 Announce Type: new Abstract: The rapid development of pretrained foundation models has enabled more general image segmentation. Multimodal large language models (MLLMs) have been wi

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

DGX agent

arXiv:2606.09380v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a leading paradigm for improving the reasoning ability of large language models throu

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing and forecasting disease trajectories of patients with Alzheimer's disease using routine data in resource-constrained settings

DGX agent

arXiv:2606.07798v1 Announce Type: new Abstract: Alzheimer's disease is a progressive neurodegenerative disorder, and its progression varies substantially across patients. Existing work aims to forecas

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing Synthetic SDO/AIA 193 A EUV Images from He I 10830 A Observations with Diffusion Model Translator

DGX agent

arXiv:2606.08652v1 Announce Type: cross Abstract: Routine full-disk EUV imaging has been available only since the modern era, such as SOHO and SDO. To extend EUV coronal context into earlier periods,

researcharxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

DGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

REFINE: Super-efficient 3D Gaussian Splatting Pruning via Rendering-Free Primitive Importance

DGX agent

arXiv:2606.09074v1 Announce Type: new Abstract: Existing pruning methods for 3D Gaussian splatting (3DGS) suffer from either severe quality degradation or prohibitive computational overhead. In this p

model-releasesarxiv-cs-cv
9 Jun 2026
Agents

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

DGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

agentsarxiv-cs-ai
9 Jun 2026
Agents

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

DGX agent

arXiv:2603.18388v2 Announce Type: replace Abstract: Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective A

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Reformulate LLM Reinforcement Learning for Efficient Training under Black-box Discrepancy

DGX agent

arXiv:2606.08779v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a pivotal post-training paradigm, yet it frequently suffers from unpredictable sub-optimum performance or eve

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

DGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

safetyarxiv-cs-ro
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement learning in linear embedding space unlocks generalizable control across soft robot configurations

DGX agent

arXiv:2606.08104v1 Announce Type: new Abstract: Soft-bodied organisms such as octopuses and elephant trunks exhibit remarkable morphological adaptability, dynamically reconfiguring body shape and stif

safetyarxiv-cs-ro
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Local Ai

Relational Epipolar Graphs for Robust Relative Camera Pose Estimation

DGX agent

arXiv:2604.04554v2 Announce Type: replace Abstract: A key component of Visual Simultaneous Localization and Mapping (VSLAM) is estimating relative camera poses using matched keypoints. Accurate estima

local-aiarxiv-cs-cv
9 Jun 2026
Safety

Reliable to Expressive: A Curriculum for Rubric-Following Safety Judges

DGX agent

arXiv:2606.09165v1 Announce Type: new Abstract: Safety judges are increasingly deployed to evaluate model outputs against evolving criteria, yet recent meta-evaluation work shows they remain brittle u

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Remember with Confidence: Uncertainty Quantification for Spatio-temporal Memory with Probabilistic Guarantees

DGX agent

arXiv:2606.08277v1 Announce Type: new Abstract: Long-horizon robot operation requires spatio-temporal memory to record the environment state and recall it for downstream reasoning. Scene graphs and re

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them

DGX agent

arXiv:2606.07597v1 Announce Type: cross Abstract: Pre-training data mixtures are commonly tuned by running small-scale experiments and extrapolating to the target training budget. When high-quality da

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

RepoLaunch: Automating Build and Management of Code Repositories across Languages and Platforms

DGX agent

arXiv:2603.05026v2 Announce Type: replace-cross Abstract: Language model (LM) agents have driven substantial progress in automated software engineering (SWE), yet building and testing software reposit

agentsarxiv-cs-lg
9 Jun 2026
Research

Report on CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS)

DGX agent

arXiv:2606.08936v1 Announce Type: cross Abstract: This report summarizes the CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS), which examined how GenAI is reshaping academic search sy

researcharxiv-cs-ai
9 Jun 2026
Research

Report the Floor: A Training-Free Conformal Interval Is a Mandatory Baseline for Probabilistic Time-Series Forecasting

DGX agent

arXiv:2606.09473v1 Announce Type: cross Abstract: Probabilistic forecasters are increasingly learned, yet the baselines they are compared against are often weak or omitted. We show that the simplest p

researcharxiv-cs-lg
9 Jun 2026
Model Releases

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

DGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

Resource-aware Computation-Communication Overlap for multi-GPU ML Workloads

DGX agent

arXiv:2606.09200v1 Announce Type: cross Abstract: The rapid growth of large-scale machine learning (ML) has made distributed training across multiple GPUs a fundamental component of modern ML systems.

hardwarearxiv-cs-ai
9 Jun 2026
Tutorials

ReTabSyn: Realistic Tabular Data Synthesis via Reinforcement Learning

DGX agent

arXiv:2603.10823v2 Announce Type: replace-cross Abstract: Deep generative models can help with data scarcity and privacy by producing synthetic training data, but they struggle in low-data, imbalanced

tutorialsarxiv-cs-lg
9 Jun 2026
Research

Rethinking 3D Shape Generation: Diffusion over Superquadrics

DGX agent

arXiv:2606.08957v1 Announce Type: new Abstract: Diffusion models have advanced 3D shape generation, yet most methods still denoise in high-cardinality spaces (e.g., voxel/SDF grids, meshes, or point c

researcharxiv-cs-cv
9 Jun 2026
Safety

Rethinking the Divergence Regularization in LLM RL

DGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

safetyarxiv-cs-lg
9 Jun 2026
Applications

Retrieval Augmented Generation Framework for the Nepali Legal Domain Question Answering

DGX agent

arXiv:2606.07523v1 Announce Type: cross Abstract: Legal domains in high-resource languages like English have widely adopted artificial intelligence for legal question answering. However, data scarcity

applicationsarxiv-cs-ai
9 Jun 2026
Research

RetroReasoner: A Reasoning LLM for Strategic Retrosynthesis Prediction

DGX agent

arXiv:2603.12666v2 Announce Type: replace-cross Abstract: Retrosynthesis prediction aims to identify reactants that can synthesize a given product molecule. Although molecular large language models (L

researcharxiv-cs-ai
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Safety

Revisiting the shutdown problem

DGX agent

arXiv:2606.08296v1 Announce Type: new Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut d

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Reward Evolution with Graph-of-Thoughts: A Bi-Level Language Model Framework for Reinforcement Learning

DGX agent

arXiv:2509.16136v5 Announce Type: replace Abstract: Designing effective reward functions remains a major challenge in reinforcement learning (RL), often requiring considerable human expertise and iter

agentsarxiv-cs-ro
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Tutorials

RGB-S: Image-Aligned Tactile Saliency for Robust Dexterous Manipulation

DGX agent

arXiv:2606.08765v1 Announce Type: cross Abstract: Effective visuo-tactile integration is critical for robotic dexterous manipulation, especially when visual observations are unreliable or occluded. Ho

tutorialsarxiv-cs-cv
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

safetyarxiv-cs-lg
9 Jun 2026
Safety

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

DGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

safetyarxiv-cs-ro
9 Jun 2026
Research

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks

DGX agent

arXiv:2506.06891v3 Announce Type: replace Abstract: We study the corruption-robustness of in-context reinforcement learning (ICRL), focusing on the Decision-Pretrained Transformer (DPT, Lee et al., 20

researcharxiv-cs-lg
9 Jun 2026
← Previous
1…598599600601602…1344
Next →