AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,270 results
Research

Raymoval: Raycasting-based Dynamic Object Removal for Static 3D Mapping

DGX agent

arXiv:2605.08937v1 Announce Type: new Abstract: Static mapping is fundamental to robot navigation, providing a persistent geometric prior and a consistent reference for long-term autonomy. However, dy

researcharxiv-cs-ro
12 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RDEx-CASK: Cauchy Mutation, Archive, and Stagnation Kick for RDEx-CSOP

DGX agent

arXiv:2605.09652v1 Announce Type: cross Abstract: We extend RDEx-CSOP with 3 changes that target stagnation & late-stage variance, plus minor parameter tuning. The second scale factor in the standard

model-releasesarxiv-cs-ai
12 May 2026
Research

RDKV: Rate-Distortion Bit Allocation for Joint Eviction and Quantization of the KV Cache

DGX agent

arXiv:2605.08317v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong performance across diverse tasks, but their inference with long input contexts is bottlenecked by memor

researcharxiv-cs-ai
12 May 2026
Safety

Re-Triggering Safeguards within LLMs for Jailbreak Detection

DGX agent

arXiv:2605.10611v1 Announce Type: cross Abstract: This paper proposes a jailbreaking prompt detection method for large language models (LLMs) to defend against jailbreak attacks. Although recent LLMs

safetyarxiv-cs-ai
12 May 2026
Model Releases

Re^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

DGX agent

arXiv:2605.09012v1 Announce Type: new Abstract: Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the

model-releasesarxiv-cs-ai
12 May 2026
Industry

Reachy is mad, but RAM costs + tariffs are forcing our hand. Prices will go up on June 1st! Still at the early bird price until then though …

DGX agent

Reachy is mad, but RAM costs + tariffs are forcing our hand. Prices will go up on June 1st! Still at the early bird price until then though if you were looking for an excuse to get one now: http://hf.

industryclem-delangue--x
12 May 2026
Model Releases

Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation

DGX agent

arXiv:2605.10430v1 Announce Type: cross Abstract: Estimating heterogeneous treatment effects with machine learning has attracted substantial attention in both academic research and industrial practice

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Really amazing dissection. LLMs hitting rarely. Tools list is amazing.

DGX agent

Really amazing dissection. LLMs hitting rarely. Tools list is amazing. 🤩🤯🤩 Claude Code (still not AGI but biggest advance since GPT-4) is the most neurosymbolic thing I have ever seen in my life. 53 s

model-releasesgary-marcus--x
12 May 2026
Model Releases

ReaMOT: A Benchmark and Framework for Reasoning-based Multi-Object Tracking

DGX agent

arXiv:2505.20381v4 Announce Type: replace Abstract: Referring Multi-Object Tracking (RMOT) aims to track targets specified by language instructions. However, existing RMOT paradigms heavily rely on ex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage

DGX agent

arXiv:2604.01527v3 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fi

model-releasesarxiv-cs-ai
12 May 2026
Agents

REAP: Reinforcement-Learning End-to-End Autonomous Parking with Gaussian Splatting Simulator for Real2Sim2Real Transfer

DGX agent

arXiv:2605.08713v1 Announce Type: cross Abstract: In recent years, autonomous parking has made significant advances, yet parking tasks still face challenges in extreme scenarios such as mechanical and

agentsarxiv-cs-ai
12 May 2026
Research

Reasoning-Aware Training for Time Series Forecasting

DGX agent

arXiv:2605.08625v1 Announce Type: cross Abstract: Time Series Foundation Models (TSFMs) excel at numerical forecasting but operate as black boxes lacking qualitative reasoning. Conversely, applying LL

researcharxiv-cs-ai
12 May 2026
Safety

Reasoning Compression with Mixed-Policy Distillation

DGX agent

arXiv:2605.08776v1 Announce Type: new Abstract: Reasoning-centric large language models (LLMs) achieve strong performance by generating intermediate reasoning trajectories, but often incur excessive t

safetyarxiv-cs-ai
12 May 2026
Model Releases

Reasoning emerges from constrained inference manifolds in large language models

DGX agent

arXiv:2605.08142v1 Announce Type: cross Abstract: Reasoning in large language models is predominantly evaluated through labeled benchmarks, conflating task performance with the quality of internal inf

model-releasesarxiv-cs-cl
12 May 2026
Safety

Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge

DGX agent

arXiv:2605.10805v1 Announce Type: new Abstract: Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings r

safetyarxiv-cs-ai
12 May 2026
Tutorials

Reasoning Trajectories for Socratic Debugging of Student Code: From Misconceptions to Contradictions and Updated Beliefs

DGX agent

arXiv:2511.00371v2 Announce Type: replace Abstract: In Socratic debugging, instructors guide students towards identifying and fixing a bug on their own, instead of providing the bug fix directly. Most

tutorialsarxiv-cs-cl
12 May 2026
Research

Rebellious Student: Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR

DGX agent

arXiv:2605.10781v1 Announce Type: cross Abstract: Self-distillation has emerged as a powerful framework for post-training LLMs, where a teacher conditioned on extra information guides a student withou

researcharxiv-cs-cl
12 May 2026
Research

Reconciling Consistency-Based Diagnosis with Actual-Causality-Based Explanations

DGX agent

arXiv:2605.08688v1 Announce Type: new Abstract: We establish, from the point of view of Explainable AI (XAI), connections between Consistency-Based Diagnosis (CBD), on one side, and Actual Causality a

researcharxiv-cs-ai
12 May 2026
Research

Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science

DGX agent

arXiv:2605.10612v1 Announce Type: cross Abstract: Graph neural networks are increasingly adopted in trigger systems for collider experiments, where strict latency and throughput constraints render dep

researcharxiv-cs-lg
12 May 2026
Model Releases

Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency

DGX agent

arXiv:2605.08454v1 Announce Type: cross Abstract: Recovering continuous-time dynamics from discrete observations is difficult because local supervision (e.g., pointwise regression targets, derivative

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Recursive Language Models

DGX agent

arXiv:2512.24601v3 Announce Type: replace Abstract: We study allowing large language models (LLMs) to process arbitrarily long prompts through the lens of inference-time scaling. We propose Recursive

model-releasesarxiv-cs-ai
12 May 2026
Agents

Red Hat expands agentic AI strategy with new inference, automation and sovereignty capabilities

DGX agent

IBM Corp. subsidiary Red Hat today is unveiling a broad set of product and partnership announcements aimed at helping enterprises put artificial intelligence into operation, modernize infrastructure a

agentssiliconangle
12 May 2026
Research

Reducing Annotation Burden for Femoral Cartilage Segmentation in Knee MRI via Cross-Sequence Transfer Learning

DGX agent

arXiv:2605.09067v1 Announce Type: new Abstract: Purpose: To develop and evaluate cross-sequence transfer learning for automatic femoral cartilage segmentation, testing bidirectional transfer between d

researcharxiv-cs-cv
12 May 2026
Safety

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning

DGX agent

arXiv:2605.09614v1 Announce Type: new Abstract: Long chain-of-thought (CoT) reasoning improves large vision--language models, but visual information often fades during generation, limiting long-horizo

safetyarxiv-cs-cv
12 May 2026
Safety

Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias

DGX agent

arXiv:2605.08315v1 Announce Type: new Abstract: Existing LLM-based policy optimizers see only scalar rewards: that a policy scored 0.45, but not whether the agent got stuck in a loop, fell into a hole

safetyarxiv-cs-lg
12 May 2026
Agents

Regime-Calibrated Fleet Repositioning with a Spatial Queue-Regret Decomposition

DGX agent

arXiv:2604.03883v2 Announce Type: replace-cross Abstract: Ride-hailing and autonomous mobility-on-demand operators reposition idle supply before future demand is fully observed. We study a retrieval-c

agentsarxiv-cs-ai
12 May 2026
Research

Region Seeding via Pre-Activation Regularization: A Geometric View of Piecewise Affine Neural Networks

DGX agent

arXiv:2605.06300v2 Announce Type: replace Abstract: Deep networks with continuous piecewise affine activations induce polyhedral partitions of the input space, making the number of realized affine reg

researcharxiv-cs-lg
12 May 2026
Research

Regret Analysis of Guided Diffusion for Black-Box Optimization over Structured Inputs

DGX agent

arXiv:2605.10385v1 Announce Type: cross Abstract: Guided-diffusion black-box optimization (BO) has shown strong empirical performance on structured design problems such as molecules and crystals, but

researcharxiv-cs-lg
12 May 2026
Research

Regret Minimization in Bilateral Trade With Perturbed Markets

DGX agent

arXiv:2605.10475v1 Announce Type: cross Abstract: We address the problem of maximizing Gain from Trade (GFT) in repeated buyer-seller exchanges subject to global budget balance constraints. While this

researcharxiv-cs-lg
12 May 2026
Model Releases

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

DGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

model-releasesarxiv-cs-ai
12 May 2026
Research

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models

DGX agent

arXiv:2605.10759v1 Announce Type: cross Abstract: Diffusion and flow-matching models scale because pretraining is supervised regression: a clean sample is noised analytically, and a model regresses ag

researcharxiv-cs-cv
12 May 2026
Safety

Reinforcement learning for inverse structural design and rapid laser cutting of kirigami prototypes

DGX agent

arXiv:2605.08098v1 Announce Type: new Abstract: Kirigami is an increasingly useful fabrication method to produce shape-programmable metamaterial structures. However, inverse design remains difficult b

safetyarxiv-cs-lg
12 May 2026
Safety

Reinforcement Learning for Scalable and Trustworthy Intelligent Systems

DGX agent

arXiv:2605.08378v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for improving the capability of intelligent systems, but its practical deployment faces two cent

safetyarxiv-cs-ai
12 May 2026
Model Releases

Reinforcement Learning Measurement Model

DGX agent

arXiv:2605.09305v1 Announce Type: cross Abstract: Interactive assessments generate sequential process data that are not well handled by conventional item response models. Existing MDP-based measuremen

model-releasesarxiv-cs-lg
12 May 2026
Safety

Reinforcement Learning with Action Chunking

DGX agent

arXiv:2507.07969v4 Announce Type: replace-cross Abstract: We present Q-chunking, a simple yet effective recipe for improving reinforcement learning (RL) algorithms for long-horizon, sparse-reward task

safetyarxiv-cs-ai
12 May 2026
Safety

Reinforcing Multimodal Reasoning Against Visual Degradation

DGX agent

arXiv:2605.09262v1 Announce Type: cross Abstract: Reinforcement Learning has significantly advanced the reasoning capabilities of Multimodal Large Language Models (MLLMs), yet the resulting policies r

safetyarxiv-cs-cl
12 May 2026
Safety

Relational reasoning and inductive bias in transformers and large language models

DGX agent

arXiv:2506.04289v3 Announce Type: replace Abstract: Transformer-based models have demonstrated remarkable reasoning abilities, but the mechanisms underlying relational reasoning remain poorly understo

safetyarxiv-cs-lg
12 May 2026
Safety

Relational Retrieval: Leveraging Known-Novel Interactions for Generalized Category Discovery

DGX agent

arXiv:2605.09420v1 Announce Type: cross Abstract: In this study, we tackle Generalized Category Discovery (GCD) via a Relational Retrieval perspective, explicitly coupling labeled and unlabeled data t

safetyarxiv-cs-ai
12 May 2026
Safety

Relations Are Channels: Knowledge Graph Embedding via Kraus Decompositions

DGX agent

arXiv:2605.10317v1 Announce Type: cross Abstract: Knowledge graph embedding (KGE) models typically represent each relation as an operator on entity embeddings. In this work, we identify three structur

safetyarxiv-cs-ai
12 May 2026
Model Releases

Relative Kinetic Utility for Reasoning-Aware Structural Pruning in Large Language Models

DGX agent

arXiv:2605.09008v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) prompting symbolized a huge improvement of reasoning capabilities of Large Language Models (LLMs). However, scaling up test-tim

model-releasesarxiv-cs-cl
12 May 2026
Safety

Relative Score Policy Optimization for Diffusion Language Models

DGX agent

arXiv:2605.10218v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a promising route to parallel and efficient text generation, but improving their reasoning ability require

safetyarxiv-cs-cl
12 May 2026
Model Releases

RelBench v2: A Large-Scale Benchmark and Repository for Relational Data

DGX agent

arXiv:2602.12606v2 Announce Type: replace Abstract: Relational deep learning (RDL) has emerged as a powerful paradigm for learning directly on relational databases by modeling entities and their relat

model-releasesarxiv-cs-lg
12 May 2026
Research

RelFlexformer: Efficient Attention 3D-Transformers for Integrable Relative Positional Encodings

DGX agent

arXiv:2605.10706v1 Announce Type: new Abstract: We present a new class of efficient attention mechanisms applying universal 3D Relative Positional Encoding (RPE) methods given by arbitrary integrable

researcharxiv-cs-lg
12 May 2026
Model Releases

Reliable LLM-Based Edge-Cloud-Expert Cascades for Telecom Knowledge Systems

DGX agent

arXiv:2512.20012v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are emerging as key enablers of automation in domains such as telecommunications, assisting with tasks including

model-releasesarxiv-cs-lg
12 May 2026
Research

ReLibra: Routing-Replay-Guided Load Balancing for MoE Training in Reinforcement Learning

DGX agent

arXiv:2605.08639v1 Announce Type: new Abstract: Load imbalance is a long-standing challenge in Mixture-of-Experts (MoE) training and is exacerbated in reinforcement learning (RL) for LLMs, where hot e

researcharxiv-cs-lg
12 May 2026
Local Ai

Relightable Gaussian Splatting for Virtual Production Using Image-Based Illumination

DGX agent

arXiv:2605.09024v1 Announce Type: new Abstract: Virtual production (VP) use LED walls to provide both background imagery and image-based lighting. While this enables on-set compositing, it couples lig

local-aiarxiv-cs-cv
12 May 2026
Agents

Remember the Decision, Not the Description: A Rate-Distortion Framework for Agent Memory

DGX agent

arXiv:2605.10870v1 Announce Type: new Abstract: Long-horizon language agents must operate under limited runtime memory, yet existing memory mechanisms often organize experience around descriptive crit

agentsarxiv-cs-ai
12 May 2026
Safety

Remember to Forget: Gated Adaptive Positional Encoding

DGX agent

arXiv:2605.10414v1 Announce Type: new Abstract: Rotary Positional Encoding (RoPE) is widely used in modern large language models. However, when sequences are extended beyond the range seen during trai

safetyarxiv-cs-lg
12 May 2026
← Previous
1…12951296129712981299…1839
Next →