AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
Tutorials

Reading Recognition in the Wild

DGX agent

arXiv:2505.24848v4 Announce Type: replace Abstract: To enable egocentric contextual AI in always-on smart glasses, it is crucial to be able to keep a record of the user's interactions with the world,

tutorialsarxiv-cs-cv
10 Apr 2026
Safety

Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.07165v1 Announce Type: new Abstract: Reinforcement learning for Large Language Model agents is often hindered by sparse rewards in multi-step reasoning tasks. Existing approaches like Group

safetyarxiv-cs-ai
10 Apr 2026
Safety

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

DGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

safetyarxiv-cs-cv
10 Apr 2026
Applications

Reasoning-Based Refinement of Unsupervised Text Clusters with LLMs

DGX agent

arXiv:2604.07562v1 Announce Type: new Abstract: Unsupervised methods are widely used to induce latent semantic structure from large text collections, yet their outputs often contain incoherent, redund

applicationsarxiv-cs-cl
10 Apr 2026
Research

Reasoning Fails Where Step Flow Breaks

DGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

researcharxiv-cs-ai
10 Apr 2026
Agents

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

DGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

agentsarxiv-cs-cl
10 Apr 2026
Safety

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

DGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

safetyarxiv-cs-cl
10 Apr 2026
Research

ReCellTy: Domain-Specific Knowledge Graph Retrieval-Augmented LLMs Reasoning Workflow for Single-Cell Annotation

DGX agent

arXiv:2505.00017v2 Announce Type: replace Abstract: With the rapid development of large language models (LLMs), their application to cell type annotation has drawn increasing attention. However, gener

researcharxiv-cs-cl
10 Apr 2026
Research

ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video

DGX agent

arXiv:2604.07882v1 Announce Type: new Abstract: Reconstructing non-rigid objects with physical plausibility remains a significant challenge. Existing approaches leverage differentiable rendering for p

researcharxiv-cs-cv
10 Apr 2026
Research

RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification

DGX agent

arXiv:2503.02537v4 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress across various visual generation tasks. However, their performance significantly declines when ge

researcharxiv-cs-cv
10 Apr 2026
Research

Rectifying LLM Thought from Lens of Optimization

DGX agent

arXiv:2512.01925v2 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have been driven by their emergent reasoning capabilities, particularly through long chain

researcharxiv-cs-ai
10 Apr 2026
Agents

ReDAct: Uncertainty-Aware Deferral for LLM Agents

DGX agent

arXiv:2604.07036v1 Announce Type: cross Abstract: Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However,

agentsarxiv-cs-lg
10 Apr 2026
Safety

Reflection-Based Task Adaptation for Self-Improving VLA

DGX agent

arXiv:2510.12710v3 Announce Type: replace Abstract: Pre-trained Vision-Language-Action (VLA) models represent a major leap towards general-purpose robots, yet efficiently adapting them to novel, speci

safetyarxiv-cs-ro
10 Apr 2026
Safety

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

DGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

safetyarxiv-cs-cl
10 Apr 2026
Local Ai

Region-Graph Optimal Transport Routing for Mixture-of-Experts Whole-Slide Image Classification

DGX agent

arXiv:2604.07298v1 Announce Type: cross Abstract: Multiple Instance Learning (MIL) is the dominant framework for gigapixel whole-slide image (WSI) classification in computational pathology. However, c

local-aiarxiv-cs-ai
10 Apr 2026
Safety

Region-R1: Reinforcing Query-Side Region Cropping for Multi-Modal Re-Ranking

DGX agent

arXiv:2604.05268v2 Announce Type: replace-cross Abstract: Multi-modal retrieval-augmented generation (MM-RAG) relies heavily on re-rankers to surface the most relevant evidence for image-question quer

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Reinforcement-Guided Synthetic Data Generation for Privacy-Sensitive Identity Recognition

DGX agent

arXiv:2604.07884v1 Announce Type: new Abstract: High-fidelity generative models are increasingly needed in privacy-sensitive scenarios, where access to data is severely restricted due to regulatory an

model-releasesarxiv-cs-cv
10 Apr 2026
Agents

RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs

DGX agent

arXiv:2604.07765v1 Announce Type: new Abstract: Earth Observation (EO) systems are essentially designed to support domain experts who often express their requirements through vague natural language ra

agentsarxiv-cs-cv
10 Apr 2026
Model Releases

Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning

DGX agent

arXiv:2601.04268v2 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

Reset-Free Reinforcement Learning for Real-World Agile Driving: An Empirical Study

DGX agent

arXiv:2604.07672v1 Announce Type: new Abstract: This paper presents an empirical study of reset-free reinforcement learning (RL) for real-world agile driving, in which a physical 1/10-scale vehicle le

safetyarxiv-cs-ro
10 Apr 2026
Research

Resistance Distance and Linearized Optimal Transport on Graphs

DGX agent

arXiv:2404.15261v4 Announce Type: replace-cross Abstract: We study the linearization of a discrete transportation distance between probability distributions on finite weighted graphs originally due to

researcharxiv-cs-lg
10 Apr 2026
Research

Resource-constrained Amazons chess decision framework integrating large language models and graph attention

DGX agent

arXiv:2603.10512v2 Announce Type: replace Abstract: Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for deci

researcharxiv-cs-ai
10 Apr 2026
Model Releases

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

DGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Rethinking Data Mixing from the Perspective of Large Language Models

DGX agent

arXiv:2604.07963v1 Announce Type: new Abstract: Data mixing strategy is essential for large language model (LLM) training. Empirical evidence shows that inappropriate strategies can significantly redu

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

DGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability

DGX agent

arXiv:2604.06628v1 Announce Type: new Abstract: A prevailing narrative in LLM post-training holds that supervised finetuning (SFT) memorizes while reinforcement learning (RL) generalizes. We revisit t

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

DGX agent

arXiv:2511.23158v2 Announce Type: replace-cross Abstract: The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Revisiting Fairness Impossibility with Endogenous Behavior

DGX agent

arXiv:2604.06378v1 Announce Type: cross Abstract: In many real-world settings, institutions can and do adjust the consequences attached to algorithmic classification decisions, such as the size of fin

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Revisiting Radar Perception With Spectral Point Clouds

DGX agent

arXiv:2604.08282v1 Announce Type: new Abstract: Radar perception models are trained with different inputs, from range-Doppler spectra to sparse point clouds. Dense spectra are assumed to outperform sp

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

RewardFlow: Generate Images by Optimizing What You Reward

DGX agent

arXiv:2604.08536v1 Announce Type: new Abstract: We introduce RewardFlow, an inversion-free framework that steers pretrained diffusion and flow-matching models at inference time through multi-reward La

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Riemann-Bench: A Benchmark for Moonshot Mathematics

DGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

model-releasesarxiv-cs-ai
10 Apr 2026
Hardware

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

DGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

hardwarearxiv-cs-lg
10 Apr 2026
Safety

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

DGX agent

arXiv:2604.07774v1 Announce Type: cross Abstract: This paper focuses on embodied task planning, where an agent acquires visual observations from the environment and executes atomic actions to accompli

safetyarxiv-cs-cv
10 Apr 2026
Agents

Robust Multi-Agent Target Tracking in Intermittent Communication Environments via Analytical Belief Merging

DGX agent

arXiv:2604.07575v1 Announce Type: new Abstract: Autonomous multi-agent target tracking in GPS-denied and communication-restricted environments (e.g., underwater exploration, subterranean search and re

agentsarxiv-cs-ro
10 Apr 2026
Model Releases

Robust support vector model based on bounded asymmetric elastic net loss for binary classification

DGX agent

arXiv:2603.06257v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel bounded asymmetric elastic net (L_{baen}) loss function and combine it with the support vector machine (SV

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Robustness Risk of Conversational Retrieval: Identifying and Mitigating Noise Sensitivity in Qwen3-Embedding Model

DGX agent

arXiv:2604.06176v1 Announce Type: cross Abstract: We present an empirical study of embedding-based retrieval under realistic conversational settings, where queries are short, dialogue-like, and weakly

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

RoSHI: A Versatile Robot-oriented Suit for Human Data In-the-Wild

DGX agent

arXiv:2604.07331v1 Announce Type: cross Abstract: Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting

safetyarxiv-cs-ai
10 Apr 2026
Safety

Rotation Equivariant Convolutions in Deformable Registration of Brain MRI

DGX agent

arXiv:2604.08034v1 Announce Type: new Abstract: Image registration is a fundamental task that aligns anatomical structures between images. While CNNs perform well, they lack rotation equivariance - a

safetyarxiv-cs-cv
10 Apr 2026
Tutorials

RPM-Net Reciprocal Point MLP Network for Unknown Network Security Threat Detection

DGX agent

arXiv:2604.06638v1 Announce Type: cross Abstract: Effective detection of unknown network security threats in multi-class imbalanced environments is critical for maintaining cyberspace security. Curren

tutorialsarxiv-cs-ai
10 Apr 2026
Agents

RQR3D: Reparametrizing the regression targets for BEV-based 3D object detection

DGX agent

arXiv:2505.17732v2 Announce Type: replace Abstract: Accurate, fast, and reliable 3D perception is essential for autonomous driving. Recently, bird's-eye view (BEV)-based perception approaches have eme

agentsarxiv-cs-cv
10 Apr 2026
Research

$S^3$: Stratified Scaling Search for Test-Time in Diffusion Language Models

DGX agent

arXiv:2604.06260v1 Announce Type: cross Abstract: Test-time scaling investigates whether a fixed diffusion language model (DLM) can generate better outputs when given more inference compute, without a

researcharxiv-cs-ai
10 Apr 2026
Safety

Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU

DGX agent

arXiv:2604.07644v1 Announce Type: new Abstract: We present GPU-SLS, a GPU-parallelized framework for safe, robust nonlinear model predictive control (MPC) that scales to high-dimensional uncertain rob

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

SALLIE: Safeguarding Against Latent Language & Image Exploits

DGX agent

arXiv:2604.06247v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) remain highly vulnerable to textual and visual jailbreaks, as well as prompt injections

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Sampling-Aware 3D Spatial Analysis in Multiplexed Imaging

DGX agent

arXiv:2604.07890v1 Announce Type: new Abstract: Highly multiplexed microscopy enables rich spatial characterization of tissues at single-cell resolution, yet most analyses rely on two-dimensional sect

researcharxiv-cs-cv
10 Apr 2026
Model Releases

SANDO: Safe Autonomous Trajectory Planning for Dynamic Unknown Environments

DGX agent

arXiv:2604.07599v1 Announce Type: new Abstract: SANDO is a safe trajectory planner for 3D dynamic unknown environments, where obstacle locations and motions are unknown a priori and a collision-free p

model-releasesarxiv-cs-ro
10 Apr 2026
Research

SAT: Balancing Reasoning Accuracy and Efficiency with Stepwise Adaptive Thinking

DGX agent

arXiv:2604.07922v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have revolutionized complex problem-solving, yet they exhibit a pervasive 'overthinking', generating unnecessarily long

researcharxiv-cs-cl
10 Apr 2026
Local Ai

SAT: Selective Aggregation Transformer for Image Super-Resolution

DGX agent

arXiv:2604.07994v1 Announce Type: new Abstract: Transformer-based approaches have revolutionized image super-resolution by modeling long-range dependencies. However, the quadratic computational comple

local-aiarxiv-cs-cv
10 Apr 2026
Research

Say Something Else: Rethinking Contextual Privacy as Information Sufficiency

DGX agent

arXiv:2604.06409v1 Announce Type: cross Abstract: LLM agents increasingly draft messages on behalf of users, yet users routinely overshare sensitive information and disagree on what counts as private.

researcharxiv-cs-ai
10 Apr 2026
← Previous
1…12831284128512861287…1290
Next →