AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
4 Aug 2026

RAP: KV-Cache Compression via RoPE-Aligned Pruning

Model ReleasesDGX agent

arXiv:2602.02599v4 Announce Type: replace Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the memory and compute of the key-value (KV) cache. Structured pruning is

Rapid Embodiment Adaptation for Quadrupedal Locomotion

SafetyDGX agent

arXiv:2608.01506v1 Announce Type: cross Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learning-based robot policies often break whe

ReACT-CLIP: Response-Aware Test-Time Defense for Vision--Language Models

ResearchDGX agent

arXiv:2608.01067v1 Announce Type: new Abstract: Training-free test-time defenses offer a practical way to improve the adversarial robustness of CLIP-style vision--language models without modifying the

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Real-Time Detection and Repair of LLM Agent Failures

Model ReleasesDGX agent

arXiv:2608.02464v1 Announce Type: cross Abstract: LLM agents fail mid-episode -- they loop, cascade tool errors, drift off goal, fabricate results, or silently absorb corrupted content -- and the stan

Real-Time Visual Obstruction Detection in Surgical Augmented Reality

Model ReleasesDGX agent

arXiv:2608.00232v1 Announce Type: new Abstract: Surgical augmented reality (AR) can provide contextual guidance by overlaying virtual annotations, tool cues, and procedural information onto the surgic

ReasonCast: Towards Explainable Time Series Forecasting with Reasoning

Model ReleasesDGX agent

arXiv:2608.01875v1 Announce Type: cross Abstract: Most time series (TS) models are specialized for a single task, either understanding (i.e., returning text answers about a TS) or generation (i.e., re

Reassessing the Feasibility of PPG-Based Non-Invasive Blood Glucose Level Estimation

ApplicationsDGX agent

arXiv:2608.01820v1 Announce Type: cross Abstract: Non-invasive blood glucose level (BGL) estimation from photoplethysmography (PPG) holds great promise for wearable health monitoring, but results acro

ReBRAC-v2: The Return of the King

ResearchDGX agent

arXiv:2608.01205v1 Announce Type: new Abstract: Recent offline reinforcement learning methods increasingly rely on expressive generative policies and specialized value-guidance mechanisms. We ask whet

Recompute or Reuse? Diagnosing and Mitigating Textual Shortcuts in VLM Self-Reflection

ResearchDGX agent

arXiv:2608.01930v1 Announce Type: new Abstract: Vision-language models (VLMs) are expected to revise their reasoning when visual evidence changes. Failures to do so are often attributed to insufficien

Reconstruction-Shift Discrimination via Mask-Guided Latent Diffusion for Medical Anomaly Detection

Local AiDGX agent

arXiv:2608.00444v1 Announce Type: new Abstract: Unsupervised medical anomaly detection learns normal anatomical patterns from healthy training images and identifies deviations at test time. Reconstruc

Recursive Gaussian Processes and the Bayesian Brain

ResearchDGX agent

arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobio

Recursive Vision Language Models for General Symbolic Reasoning

Model ReleasesDGX agent

arXiv:2608.01534v1 Announce Type: new Abstract: Hard symbolic-reasoning tasks such as Sudoku, maze pathfinding, and ARC remain challenging for LLMs due to their fixed-depth autoregressive reasoning, w

Refine Drugs, Don't Complete Them: Uniform-Source Discrete Flows for Fragment-Based Drug Discovery

Model ReleasesDGX agent

arXiv:2509.26405v2 Announce Type: replace Abstract: We introduce InVirtuoGen, a discrete flow generative model for fragmented SMILES for de novo and fragment-constrained generation, and target-propert

REFLEX: Rethinking MoE Inference as Refinement-Aware Compute Allocation in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2608.01784v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) models increase parameter capacity by activating only a small subset of experts for each token. This conditional-computation

ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection

ResearchDGX agent

arXiv:2608.01793v1 Announce Type: new Abstract: Unified anomaly detection requires modeling highly heterogeneous normal data without access to anomalous samples. While foundation models like DINOv2 pr

Regularized Schrodinger Bridge via Distortion-Perception Perturbation for High-Fidelity Speech Enhancement

SafetyDGX agent

arXiv:2511.11686v4 Announce Type: replace Abstract: Speech enhancement (SE) requires high-fidelity reconstruction of clean speech that preserves linguistic and paralinguistic cues while maintaining hi

Relative Parameter Importance in Task-Agnostic Replay-Free Continual Learning

Model ReleasesDGX agent

arXiv:2608.00630v1 Announce Type: new Abstract: Achieving continual learning (CL) with deep neural networks requires balancing stability and plasticity while enabling knowledge transfer. In this work,

Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00574v1 Announce Type: new Abstract: Vision-language MoE batches contain different numbers of image and text tokens. Image resolution, image count, tiling, and prompt length all change this

Remember-R1: Mitigating Long-Context Visual Forgetting through Reinforcement Learning

ResearchDGX agent

arXiv:2608.01314v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly rely on long chain-of-thought reasoning for complex tasks. However, as reasoning sequences lengthe

ReMiX-MAE: Learning Missing-Channel Cross-Modal Representations from RGB-Only Clinical Facial Videos for Sympathetic-Mediated Pain Assessment

ApplicationsDGX agent

arXiv:2608.02561v1 Announce Type: new Abstract: Automated pain assessment in real clinics is limited by scarce clinically grounded facial video data with weak labels (often sequence-level self-report)

Representation Transfer of Foundation Models for Ultra-Widefield Retinal Imaging

ResearchDGX agent

arXiv:2608.00586v1 Announce Type: new Abstract: Despite the widespread adoption of foundation models as feature extractors for medical imaging, relatively little is understood about how different pret

Residual-Based Adaptive Kalman Filtering for Legged Robot State Estimation

Model ReleasesDGX agent

arXiv:2608.02316v1 Announce Type: new Abstract: State estimation is a key component in model-based control of walking robots and, more broadly, applicable wherever hidden variables must be inferred. T

Response Magnitude as a Dominant Signal for Held-Out CRISPRi Perturbation Effect Prediction

Model ReleasesDGX agent

arXiv:2608.00152v1 Announce Type: new Abstract: Predicting the magnitude of a CRISPRi perturbation's transcriptomic effect on held-out target genes is an important open problem in single-cell biology.

Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance

TutorialsDGX agent

arXiv:2510.21590v3 Announce Type: replace Abstract: Current image super-resolution methods show strong performance on natural images but distort text, creating a fundamental trade-off between image qu

RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

Model ReleasesDGX agent

arXiv:2608.01247v1 Announce Type: new Abstract: Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future queries, but performance can collapse und

Rethinking and formalising the state across languages: a unified computational learning theory account

ResearchDGX agent

arXiv:2608.00523v1 Announce Type: new Abstract: The linguistic notion of state has traditionally been restricted to the construct (annexation) state of Afroasiatic languages and treated as a language-

Rethinking Federated Graph Foundation Models: A Graph-Language Alignment-based Approach

Model ReleasesDGX agent

arXiv:2601.21369v2 Announce Type: replace Abstract: Recent studies of federated graph foundational models (FedGFMs) break the idealized and untenable assumption of having centralized data storage to t

Rethinking IRSTD: Single-Point Supervision Guided Encoder-only Framework is Enough for Infrared Small Target Detection

ResearchDGX agent

arXiv:2604.05363v2 Announce Type: replace Abstract: Infrared small target detection (IRSTD) aims to separate small targets from clutter backgrounds. Extensive research is dedicated to the pixel-level

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning

Local AiDGX agent

arXiv:2608.01556v1 Announce Type: new Abstract: Large language models are increasingly aligned to human preferences via reward modeling, but user preference data are sensitive and often cannot be cent

Rethinking PPG-based Sleep Staging: Datasets, Metrics, and Benchmarks

ResearchDGX agent

arXiv:2608.00943v1 Announce Type: cross Abstract: Automated sleep staging assigns discrete stage labels to successive time epochs throughout an overnight recording; conventionally each window spans at

Rethinking Pretraining for Specialized Design Data: Evidence from the JONES-19 Cultural Design Dataset

ResearchDGX agent

arXiv:2608.00135v1 Announce Type: cross Abstract: Design and architectural archives encode expert human knowledge in graphical formats, providing a critical testbed for design-inspired Machine Learnin

Rethinking Total Absorption Gamma Spectroscopy Deconvolution: Supervised Machine Learning vs Response-Matrix Methods

ResearchDGX agent

arXiv:2608.00090v1 Announce Type: cross Abstract: The extraction of eta-feeding distributions in Total Absorption gamma-ray Spectroscopy constitutes a challenging inverse problem, particularly in nucl

Rethinking Video Token Compression with a Global Codebook: Learning Once, Compressing Everywhere

ResearchDGX agent

arXiv:2608.01271v1 Announce Type: new Abstract: Video large language models (Video-LLMs) represent videos as dense sequences of visual tokens, whose length grows with the temporal and spatial extent o

ReTouch: Empowering Contact-Rich Dexterous Manipulation with Online-Refined Tactile Prediction

ApplicationsDGX agent

arXiv:2608.01824v1 Announce Type: new Abstract: Fusing tactile signals has proven effective for contact-rich manipulation, enabling robots to perceive contact states and adapt to rapidly changing phys

Retrieval Augmented Biomedical Question Answering with Weak Question Recovery and Neural Reranking for BioASQ Task 14b

ResearchDGX agent

arXiv:2608.01468v1 Announce Type: new Abstract: This work presents DS@GT ARC BioASQ team's work for a biomedical question answering pipeline, integrating multi-source query expansion, neural reranking

Retrieval-Augmented Interpretable Learning: Towards Task-Specific Zero-Shot Models in Healthcare

ApplicationsDGX agent

arXiv:2607.17508v2 Announce Type: replace Abstract: We introduce Retrieval-Augmented Interpretable Learning (RAIL), a probabilistic meta-learning framework for zero-shot generation of task-specific in

Retrieval-Based Cross-Domain Generalization in Optical Networks via Global Features

ResearchDGX agent

arXiv:2608.00044v1 Announce Type: cross Abstract: We propose a retrieval-based framework for crossdomain quality-of-transmission (QoT) estimation that leverages transferable feature representations wh

Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforcement Learning

Model ReleasesDGX agent

arXiv:2608.01418v1 Announce Type: cross Abstract: Autoregressive rollout generation is a major computational cost in reinforcement learning for large language models. Reusing each rollout batch for ad

Revisiting Generalization Across Difficulty Levels: It's Not So Easy

ResearchDGX agent

arXiv:2511.21692v2 Announce Type: replace Abstract: We investigate how well large language models (LLMs) generalize across different task difficulties, a key question for effective data curation and e

RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals

ApplicationsDGX agent

arXiv:2608.00289v1 Announce Type: cross Abstract: Recognizing Human-Object Interactions (HOI) is essential for intelligent systems, underpinning applications in virtual and augmented reality, embodied

RH-RAG: Trustworthy Long-Form Generation for Privacy-Constrained Settings

Local AiDGX agent

arXiv:2608.01311v1 Announce Type: new Abstract: Generating long-form content from extensive internal reports remains challenging for organizations operating under strict privacy and security constrain

RHEA: Reliability-Harmonized Reconstruction and Assignment for Robust Multimodal-Attributed Graph Clustering

ResearchDGX agent

arXiv:2608.00621v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), whose nodes carry heterogeneous attributes such as text and images over a relational structure, have become a funda

Rhythm of the Deep: Two-Tier Combinatorial Structure in Sperm Whale Codas Revealed by Acoustic Unit Induction

ResearchDGX agent

arXiv:2606.16084v2 Announce Type: replace-cross Abstract: Sperm-whale codas are conventionally described as recurring click-count and timing patterns. We show instead that their waveforms contain a tw

Riemannian Attention Mechanisms for Transformers: A Theoretical Framework and Architecture Design

Model ReleasesDGX agent

arXiv:2608.01283v1 Announce Type: new Abstract: All Transformer-based large language models compute attention via the Euclidean inner product, an architectural choice that Dong et al. (2021) proved ca

Right Answer, Wrong Method: Shortcut Hacking Misleads the Evaluation of LLM Reasoning on Frontier Science Benchmarks

Model ReleasesDGX agent

arXiv:2608.02442v1 Announce Type: cross Abstract: Scientific reasoning benchmarks typically evaluate large language models (LLMs) using final-answer accuracy. However, a correct answer does not necess

RING: Retrieval-Internalized Generation for Continual Large-Scale Knowledge Injection

Model ReleasesDGX agent

arXiv:2608.01630v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves factuality but adds latency and engineering overhead at serving time. We propose RING (Retrieval-Internali

RIT*: Riemannian Informed Trees for Cost-Adaptive Optimal Motion Planning

Model ReleasesDGX agent

arXiv:2608.00822v1 Announce Type: new Abstract: We present Riemannian Informed Trees (RIT*), a planning framework that replaces Euclidean primitives in batch-informed search with their Riemannian coun

RL Bootstrapping of OpenVLA-OFT for a Novel Robot Embodiment

SafetyDGX agent

arXiv:2608.01013v1 Announce Type: new Abstract: Adapting a pretrained vision-language-action (VLA) policy to a new robot usually assumes embodiment-specific demonstrations. This assumption is especial

RMSWeb: Reflection, Failure-Mode Mining, and Salvage-DS for Web Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.00335v1 Announce Type: cross Abstract: Compact web agents can reduce deployment cost, but training them poses challenges in both data collection and post-SFT reinforcement learning (RL). Su

RobotDancing: Residual-Action Reinforcement Learning Enables Robust Long-Horizon Humanoid Motion Tracking

SafetyDGX agent

arXiv:2509.20717v2 Announce Type: replace Abstract: Long-horizon, high-dynamic motion tracking on humanoids remains brittle: retargeted reference motions are typically kinematically plausible but dyna

Robust Bayesian Optimization via Tempered Posteriors

ResearchDGX agent

arXiv:2601.07094v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquis

Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Stealthy Backdoor

Model ReleasesDGX agent

arXiv:2608.00543v1 Announce Type: cross Abstract: Although semantic watermarking is considered a promising safeguard for images generated by Latent Diffusion Models (LDMs), the reliance of the waterma

Role Steering of Language Models for Social Simulations

Model ReleasesDGX agent

arXiv:2608.00023v1 Announce Type: new Abstract: Social simulations built from language-model agents need role-conditioned behavior that can be checked before agents are placed into a simulated populat

Rolling Shutter Camera Self-Calibration

ResearchDGX agent

arXiv:2608.01509v1 Announce Type: new Abstract: Rolling shutter (RS) cameras are widely used in consumer devices, but their row-wise exposure causes distortions under motion, making geometric 3D visio

Romanized Arabic Across Dialects: Views, Usage Patterns, and Linguistic Variation

SafetyDGX agent

arXiv:2608.02555v1 Announce Type: new Abstract: Arabizi refers to Arabic written in Latin script. Although previous studies have shown that the prevalence and usage of Arabizi vary by factors such as

RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States

AgentsDGX agent

arXiv:2608.02508v1 Announce Type: cross Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the inte

Roomer: Reflective Object-Grounded Model Editing and Repair for 3D Indoor Layout Synthesis

Local AiDGX agent

arXiv:2608.01973v1 Announce Type: cross Abstract: Existing indoor layout generators produce globally plausible layouts yet may retain local violations such as collisions, out-of-bounds placements, obs

Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors

Model ReleasesDGX agent

arXiv:2608.00675v1 Announce Type: cross Abstract: Autoregressive models accumulate error over long rollouts, yet at deployment there is no ground truth to measure it against. We train a single conditi

RPL-UIE: Reliable Prior Learning for Underwater Image Enhancement

ApplicationsDGX agent

arXiv:2608.00137v1 Announce Type: cross Abstract: Underwater image enhancement (UIE) aims to recover clear images from observations affected by wavelength-dependent absorption, scattering, and spatial

RSC-GestureNet: Reliability-Aware Selective Causal Recognition of Chinese Traffic Police Gestures

Model ReleasesDGX agent

arXiv:2608.02200v1 Announce Type: new Abstract: Traffic police gestures are safety-critical perception cues for autonomous driving. A deployable recognizer must infer commands causally from continuous

← Previous
1…100101102103104…998
Next →