AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Research

Reachability and asymptotics of Gaussian Transformer dynamics

DGX agent

arXiv:2606.07600v1 Announce Type: cross Abstract: We formulate data propagation through the Transformer, the machine learning architecture powering large language models, as a nonlinear control system

researcharxiv-cs-ai
9 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

REACT 2026: The Fourth Multiple Appropriate Facial Reaction Generation Challenge: Personalised MAFRG and Appropriate EEG Reaction Prediction

DGX agent

arXiv:2606.07935v1 Announce Type: new Abstract: In dyadic interactions, various human facial reactions could be appropriate for responding to each human speaker behaviour. Following the successful org

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Readable Yet Unpredictable: Rotated-Outcome Prediction in Vision-Language Models

DGX agent

arXiv:2606.07641v1 Announce Type: new Abstract: Can vision-language models predict what a 180{eg} rotation would reveal from the original image alone? We study this ability through Rotated-Outcome Pre

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Real-IKEA: Physical Fidelity is the Prerequisite for Robust Manipulation

DGX agent

arXiv:2606.08564v1 Announce Type: new Abstract: Robotic manipulation robustness often founders on the physics gap between simplified simulations and the resistance-laden real world. In this work, we e

model-releasesarxiv-cs-ro
9 Jun 2026
Applications

Real-Time and Accurate Collision-Free Teleoperation via Differentiable Constraint-Based Trajectory Planning

DGX agent

arXiv:2606.08725v1 Announce Type: new Abstract: In teleoperation, the human operator typically controls only the end-effector pose, which often leads to self-collisions of the manipulator and collisio

applicationsarxiv-cs-ro
9 Jun 2026
Model Releases

Real-time body pose non-verbal communication with a consistency-based reliability measure

DGX agent

arXiv:2606.09390v1 Announce Type: cross Abstract: Body movement communicates intent at distances and in conditions where neither the face, nor speech can be captured. We study the recognition of commu

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning

DGX agent

arXiv:2606.09303v1 Announce Type: new Abstract: The rapid development of pretrained foundation models has enabled more general image segmentation. Multimodal large language models (MLLMs) have been wi

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

DGX agent

arXiv:2606.09380v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a leading paradigm for improving the reasoning ability of large language models throu

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing and forecasting disease trajectories of patients with Alzheimer's disease using routine data in resource-constrained settings

DGX agent

arXiv:2606.07798v1 Announce Type: new Abstract: Alzheimer's disease is a progressive neurodegenerative disorder, and its progression varies substantially across patients. Existing work aims to forecas

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing Synthetic SDO/AIA 193 A EUV Images from He I 10830 A Observations with Diffusion Model Translator

DGX agent

arXiv:2606.08652v1 Announce Type: cross Abstract: Routine full-disk EUV imaging has been available only since the modern era, such as SOHO and SDO. To extend EUV coronal context into earlier periods,

researcharxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

DGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

REFINE: Super-efficient 3D Gaussian Splatting Pruning via Rendering-Free Primitive Importance

DGX agent

arXiv:2606.09074v1 Announce Type: new Abstract: Existing pruning methods for 3D Gaussian splatting (3DGS) suffer from either severe quality degradation or prohibitive computational overhead. In this p

model-releasesarxiv-cs-cv
9 Jun 2026
Agents

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

DGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

agentsarxiv-cs-ai
9 Jun 2026
Agents

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

DGX agent

arXiv:2603.18388v2 Announce Type: replace Abstract: Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective A

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Reformulate LLM Reinforcement Learning for Efficient Training under Black-box Discrepancy

DGX agent

arXiv:2606.08779v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a pivotal post-training paradigm, yet it frequently suffers from unpredictable sub-optimum performance or eve

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

DGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

safetyarxiv-cs-ro
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement learning in linear embedding space unlocks generalizable control across soft robot configurations

DGX agent

arXiv:2606.08104v1 Announce Type: new Abstract: Soft-bodied organisms such as octopuses and elephant trunks exhibit remarkable morphological adaptability, dynamically reconfiguring body shape and stif

safetyarxiv-cs-ro
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Local Ai

Relational Epipolar Graphs for Robust Relative Camera Pose Estimation

DGX agent

arXiv:2604.04554v2 Announce Type: replace Abstract: A key component of Visual Simultaneous Localization and Mapping (VSLAM) is estimating relative camera poses using matched keypoints. Accurate estima

local-aiarxiv-cs-cv
9 Jun 2026
Safety

Reliable to Expressive: A Curriculum for Rubric-Following Safety Judges

DGX agent

arXiv:2606.09165v1 Announce Type: new Abstract: Safety judges are increasingly deployed to evaluate model outputs against evolving criteria, yet recent meta-evaluation work shows they remain brittle u

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Remember with Confidence: Uncertainty Quantification for Spatio-temporal Memory with Probabilistic Guarantees

DGX agent

arXiv:2606.08277v1 Announce Type: new Abstract: Long-horizon robot operation requires spatio-temporal memory to record the environment state and recall it for downstream reasoning. Scene graphs and re

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them

DGX agent

arXiv:2606.07597v1 Announce Type: cross Abstract: Pre-training data mixtures are commonly tuned by running small-scale experiments and extrapolating to the target training budget. When high-quality da

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

RepoLaunch: Automating Build and Management of Code Repositories across Languages and Platforms

DGX agent

arXiv:2603.05026v2 Announce Type: replace-cross Abstract: Language model (LM) agents have driven substantial progress in automated software engineering (SWE), yet building and testing software reposit

agentsarxiv-cs-lg
9 Jun 2026
Research

Report on CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS)

DGX agent

arXiv:2606.08936v1 Announce Type: cross Abstract: This report summarizes the CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS), which examined how GenAI is reshaping academic search sy

researcharxiv-cs-ai
9 Jun 2026
Research

Report the Floor: A Training-Free Conformal Interval Is a Mandatory Baseline for Probabilistic Time-Series Forecasting

DGX agent

arXiv:2606.09473v1 Announce Type: cross Abstract: Probabilistic forecasters are increasingly learned, yet the baselines they are compared against are often weak or omitted. We show that the simplest p

researcharxiv-cs-lg
9 Jun 2026
Model Releases

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

DGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

Resource-aware Computation-Communication Overlap for multi-GPU ML Workloads

DGX agent

arXiv:2606.09200v1 Announce Type: cross Abstract: The rapid growth of large-scale machine learning (ML) has made distributed training across multiple GPUs a fundamental component of modern ML systems.

hardwarearxiv-cs-ai
9 Jun 2026
Tutorials

ReTabSyn: Realistic Tabular Data Synthesis via Reinforcement Learning

DGX agent

arXiv:2603.10823v2 Announce Type: replace-cross Abstract: Deep generative models can help with data scarcity and privacy by producing synthetic training data, but they struggle in low-data, imbalanced

tutorialsarxiv-cs-lg
9 Jun 2026
Research

Rethinking 3D Shape Generation: Diffusion over Superquadrics

DGX agent

arXiv:2606.08957v1 Announce Type: new Abstract: Diffusion models have advanced 3D shape generation, yet most methods still denoise in high-cardinality spaces (e.g., voxel/SDF grids, meshes, or point c

researcharxiv-cs-cv
9 Jun 2026
Safety

Rethinking the Divergence Regularization in LLM RL

DGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

safetyarxiv-cs-lg
9 Jun 2026
Applications

Retrieval Augmented Generation Framework for the Nepali Legal Domain Question Answering

DGX agent

arXiv:2606.07523v1 Announce Type: cross Abstract: Legal domains in high-resource languages like English have widely adopted artificial intelligence for legal question answering. However, data scarcity

applicationsarxiv-cs-ai
9 Jun 2026
Research

RetroReasoner: A Reasoning LLM for Strategic Retrosynthesis Prediction

DGX agent

arXiv:2603.12666v2 Announce Type: replace-cross Abstract: Retrosynthesis prediction aims to identify reactants that can synthesize a given product molecule. Although molecular large language models (L

researcharxiv-cs-ai
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Safety

Revisiting the shutdown problem

DGX agent

arXiv:2606.08296v1 Announce Type: new Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut d

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Reward Evolution with Graph-of-Thoughts: A Bi-Level Language Model Framework for Reinforcement Learning

DGX agent

arXiv:2509.16136v5 Announce Type: replace Abstract: Designing effective reward functions remains a major challenge in reinforcement learning (RL), often requiring considerable human expertise and iter

agentsarxiv-cs-ro
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Tutorials

RGB-S: Image-Aligned Tactile Saliency for Robust Dexterous Manipulation

DGX agent

arXiv:2606.08765v1 Announce Type: cross Abstract: Effective visuo-tactile integration is critical for robotic dexterous manipulation, especially when visual observations are unreliable or occluded. Ho

tutorialsarxiv-cs-cv
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

safetyarxiv-cs-lg
9 Jun 2026
Safety

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

DGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

safetyarxiv-cs-ro
9 Jun 2026
Research

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks

DGX agent

arXiv:2506.06891v3 Announce Type: replace Abstract: We study the corruption-robustness of in-context reinforcement learning (ICRL), focusing on the Decision-Pretrained Transformer (DPT, Lee et al., 20

researcharxiv-cs-lg
9 Jun 2026
← Previous
1…598599600601602…1344
Next →