AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
Human
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
21 May 2026

Reducing Object Hallucination in LVLMs via Emphasizing Image-negative Tokens

TutorialsDGX agent

arXiv:2605.21300v1 Announce Type: new Abstract: Object hallucination is a significant challenge that hinders the application of large vision-language models (LVLMs) in practice. We hypothesize that on

Refining and Reusing Annotation Guidelines for LLM Annotation

Model ReleasesDGX agent

arXiv:2605.20809v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable performance on zero-shot annotation tasks, they often struggle with the specialized convention

REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak

SafetyDGX agent

arXiv:2605.20654v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable capabilities, they remain susceptible to sophisticated, multi-step jailbreak attacks that circ

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis

SafetyDGX agent

arXiv:2605.20277v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have rapidly advanced as general-purpose multimodal assistants, yet their deployment in 3D Computed Tomography (CT

Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes

ResearchDGX agent

arXiv:2605.21211v1 Announce Type: cross Abstract: In this work we present an efficient and practically implementable approach for the application of reinforcement learning (RL)-based control in chemic

Reinforcement Learning for Risk Adaptation via Differentiable CVaR Barrier Functions

SafetyDGX agent

arXiv:2605.21257v1 Announce Type: new Abstract: Planning through crowded environments under uncertain obstacle motions remains difficult, as stochastic interactions often induce overly conservative be

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces

SafetyDGX agent

arXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov

Reinforcing Human Behavior Simulation via Verbal Feedback

Model ReleasesDGX agent

arXiv:2605.20506v1 Announce Type: cross Abstract: Humans learn social norms and behaviors from verbal feedback (e.g., a parent saying 'that was rude' or a friend explaining 'here's why that hurt'). Ye

Reliable Automated Triage in Spanish Clinical Notes: A Hybrid Framework for Risk-Aware HIV Suspicion Identification

ResearchDGX agent

arXiv:2605.21256v1 Announce Type: new Abstract: Standard clinical Natural Language Processing (NLP) benchmarks often yield inflated metrics by forcing deterministic classification on ambiguous instanc

RelWitness: Open-Vocabulary 3D Scene Graph Generation with Visual-Geometric Relation Witnesses

ResearchDGX agent

arXiv:2605.20823v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph generation seeks to describe object instances and their relations with flexible natural-language predicates. The central

ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes

ResearchDGX agent

arXiv:2605.21440v1 Announce Type: new Abstract: Atmospheric turbulence severely degrades video quality by introducing distortions such as geometric warping, blur, and temporal flickering, posing signi

RePCM: Region-Specific and Phenotype-Adaptive Bi-Ventricular Cardiac Motion Synthesis

Local AiDGX agent

arXiv:2605.21237v1 Announce Type: new Abstract: Cardiac motion over a cardiac cycle is crucial for quantifying regional function and is strongly affected by cardiovascular diseases. Since temporally d

rePIRL: Learn PRM with Inverse RL for LLM Reasoning

SafetyDGX agent

arXiv:2602.07832v2 Announce Type: replace Abstract: Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In

Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing

Model ReleasesDGX agent

arXiv:2605.20262v1 Announce Type: new Abstract: We study selective refusal editing as a three-way control problem: induce non-refusal on designated edit prompts while preserving benign behavior and ha

ResNet-50 with Class Reweighting and Anatomy-Guided Temporal Decoding for Gastrointestinal Video Analysis

ResearchDGX agent

arXiv:2603.17784v2 Announce Type: replace Abstract: We developed a multi-label gastrointestinal video analysis pipeline based on a ResNet-50 frame classifier followed by anatomy-guided temporal event

Resolving Long-Tail Ambiguity in Unsupervised 3D Point Cloud Segmentation with Language Priors

Model ReleasesDGX agent

arXiv:2605.20737v1 Announce Type: new Abstract: Existing approaches for unsupervised 3D point cloud segmentation predominantly rely on a purely visual similarity-based learning-by-clustering paradigm,

Rethinking Cross-Layer Information Routing in Diffusion Transformers

SafetyDGX agent

arXiv:2605.20708v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have become a de facto backbone of modern visual generation, and nearly every major axis of their design -- tokenization,

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

AgentsDGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

Retrieval-Augmented Long-Context Translation for Cultural Image Captioning: Gators submission for AmericasNLP 2026 shared task

Model ReleasesDGX agent

arXiv:2605.20626v1 Announce Type: new Abstract: We present the University of Florida Gators submission to the AmericasNLP 2026 shared task on cultural image captioning for Indigenous languages. Our tw

Retrospective Sparse Attention for Efficient Long-Context Generation

ResearchDGX agent

arXiv:2508.09001v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly deployed in long-context tasks such as reasoning, code generation, and multi-turn dialogue. However, i

ReversedQ: Opportunities for Faster Q-Learning in Episodic Online Reinforcement Learning

ResearchDGX agent

arXiv:2605.20592v1 Announce Type: new Abstract: We study model-free Q-learning in finite-horizon episodic Markov Decision Processes (MDPs) with stationary dynamics across episodes. We identify a centr

Reviving Error Correction in Modern Deep Time-Series Forecasting

ResearchDGX agent

arXiv:2605.21088v1 Announce Type: new Abstract: Modern deep-learning models have achieved remarkable success in time-series forecasting. Yet, their performance degrades in long-term prediction due to

Riemannian MeanFlow for One-Step Generation on Manifolds

ResearchDGX agent

arXiv:2603.10718v2 Announce Type: replace Abstract: Flow Matching enables simulation-free training of generative models on Riemannian manifolds, yet sampling typically still relies on numerically inte

RISE: Reliable Improvement in Self-Evolving Vision-Language Models

ResearchDGX agent

arXiv:2605.20914v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong multimodal reasoning capabilities, but further improving them still relies heavily on large-scale hum

RoadTones: Tone Controllable Text Generation from Road Event Videos

ResearchDGX agent

arXiv:2605.21411v1 Announce Type: new Abstract: Existing video-language models can generate factual descriptions of road events but lack control over how these events are expressed: their tone, urgenc

ROAR-3D: Routing Arbitrary Views for High-Fidelity 3D Generation

ResearchDGX agent

arXiv:2605.21121v1 Announce Type: new Abstract: Single-image-to-3D generative models can now produce high-quality geometry, yet conditioning on a single view inevitably introduces ambiguity about unse

Robust Personalized Recommendation under Hidden Confounding in MNAR

Model ReleasesDGX agent

arXiv:2605.21066v1 Announce Type: new Abstract: Recommender systems often rely on observational user--item interaction data, which is prone to selection bias due to users' selective interactions with

Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework

SafetyDGX agent

arXiv:2605.20721v1 Announce Type: new Abstract: Learning from implicit feedback in recommender systems is fundamentally challenged by pervasive label noise. While conventional denoising approaches oft

Robust Subspace-Constrained Quadratic Models for Low-Dimensional Structure Learning

ResearchDGX agent

arXiv:2605.20300v1 Announce Type: new Abstract: In this paper, we propose a robust subspace-constrained quadratic model (SCQM) for learning low-dimensional structure from high-dimensional data. Buildi

RoPeSLR: 3D RoPE-driven Sparse-LowRank Attention for Efficient Diffusion Transformers

ResearchDGX agent

arXiv:2605.20659v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have revolutionized high-fidelity video generation, yet their O(L^2) attention complexity poses a formidable bottleneck fo

roto 2.0: The Robot Tactile Olympiad

Model ReleasesDGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

Runtime-Certified Bounded-Error Quantized Attention

Model ReleasesDGX agent

arXiv:2605.20868v1 Announce Type: new Abstract: KV cache quantization reduces the memory cost of long-context LLM inference, but introduces approximation error that is typically validated only empiric

Safety-Critical Control for Smoothed Implicit Contact Dynamics

Model ReleasesDGX agent

arXiv:2605.21138v1 Announce Type: new Abstract: Smoothed implicit contact dynamics enables gradient-based planning and control for contact-rich tasks without predefined mode sequences. However, safety

SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection

SafetyDGX agent

arXiv:2605.21186v1 Announce Type: new Abstract: Interpretability in object detection provides crucial confidence support for clinical auxiliary diagnosis. However, in tiny bacteria detection, traditio

Same Target, Different Basins: Hard vs. Soft Labels for Annotator Distributions

ResearchDGX agent

arXiv:2605.20642v1 Announce Type: new Abstract: When annotators disagree, that disagreement can reflect epistemic uncertainty rather than simple label noise. We study hard-label delivery as an alterna

Sample Complexity of Transfer Learning: An Optimal Transport Approach

ResearchDGX agent

arXiv:2605.20545v1 Announce Type: cross Abstract: Transfer learning is an essential technique for many machine learning/AI models of complex structures such as large language models and generative AI.

SAVER: Selective As-Needed Vision Evidence for Multimodal Information Extraction

ResearchDGX agent

arXiv:2605.20713v1 Announce Type: new Abstract: Multimodal IE in social media is difficult because a post may attach multiple images that are weakly related, redundant, or even misleading with respect

Scalable Multi-robot Motion Planning via Hierarchical Subproblem Expansion and Workspace Decomposition Refinement

ResearchDGX agent

arXiv:2605.20395v1 Announce Type: new Abstract: A fundamental challenge in multi-robot motion planning is achieving sufficient coordination to avoid inter-robot conflicts without incurring the large c

Scale-Calibrated Median-of-Means for Robust Distributed Principal Component Analysis

Local AiDGX agent

arXiv:2605.20681v1 Announce Type: cross Abstract: Distributed principal component analysis (PCA) produces node-level estimates of both a mean vector and a principal subspace. Robustly aggregating thes

Score-Based Causal Discovery of Latent Variable Causal Models

ResearchDGX agent

arXiv:2605.20396v1 Announce Type: new Abstract: Identifying latent variables and the causal structure involving them is essential across various scientific fields. While many existing works fall under

SCRIBE: Diagnostic Evaluation and Rich Transcription Models for Indic ASR

SafetyDGX agent

arXiv:2605.20712v1 Announce Type: new Abstract: Automatic speech recognition replaces typing only when correction costs less than manual entry, a threshold determined by error types, not counts: fixin

SDM: A Powerful Tool for Evaluating Model Robustness

ResearchDGX agent

arXiv:2605.20308v1 Announce Type: new Abstract: Gradient-based attacks are important methods for evaluating model robustness. However, since the proposal of APGD, it has been difficult for such method

Secure, Verifiable, and Scalable Multi-Client Data Sharing via Consensus-Based Privacy-Preserving Data Distribution

SafetyDGX agent

arXiv:2601.00418v2 Announce Type: replace-cross Abstract: We propose the Consensus-Based Privacy-Preserving Data Distribution (CPPDD) framework, a lightweight and post-setup autonomous protocol for se

Seeing Through Fog: Towards Fog-Invariant Action Recognition

Model ReleasesDGX agent

arXiv:2605.20645v1 Announce Type: new Abstract: Foggy conditions are commonly encountered in real-world applications; however, existing action recognition approaches typically assume favorable weather

Self-Improving Skill Learning for Robust Skill-based Meta-Reinforcement Learning

ResearchDGX agent

arXiv:2502.03752v5 Announce Type: replace Abstract: Meta-reinforcement learning (Meta-RL) facilitates rapid adaptation to unseen tasks but faces challenges in long-horizon environments. Skill-based ap

Self-Refining Video Sampling

SafetyDGX agent

arXiv:2601.18577v2 Announce Type: replace Abstract: Modern video generators still struggle with complex physical dynamics, often falling short of physical realism. Existing approaches address this usi

Self-Training Doesn't Flatten Language -- It Restructures It: Surface Markers Amplify While Deep Syntax Dies

ResearchDGX agent

arXiv:2605.20602v1 Announce Type: new Abstract: Successive self-training on a language model's own outputs is widely characterized as a process of flattening: diversity drops, distributions narrow, an

Semantic Granularity Navigation in Image Editing

Local AiDGX agent

arXiv:2605.21190v1 Announce Type: new Abstract: Despite the generative capabilities of diffusion and flow models, real-image editing remains constrained by a persistent trade-off between semantic edit

Semiparametric Efficient Bilevel Gradient Estimation

Model ReleasesDGX agent

arXiv:2605.21341v1 Announce Type: cross Abstract: Functional bilevel methods estimate a lower-level function and plug it into a hypergradient, but this plug-in gradient can retain first-order bias whe

Sequential Data Augmentation for Generative Recommendation

Model ReleasesDGX agent

arXiv:2509.13648v3 Announce Type: replace Abstract: Generative recommendation plays a crucial role in personalized systems, predicting users' future interactions from their historical behavior sequenc

ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society

Model ReleasesDGX agent

arXiv:2605.20510v1 Announce Type: new Abstract: Urban heat exposure is becoming an increasingly critical challenge due to the intensifying urban heat island effect. Fine-grained shade patterns, especi

ShapeBench: A Scalable Benchmark and Diagnostic Suite for Standardized Evaluation in Aerodynamic Shape Optimization

Model ReleasesDGX agent

arXiv:2605.20763v1 Announce Type: new Abstract: Rapid progress in aerodynamic shape optimization (ASO) has outpaced currently-available standardized evaluation frameworks. Fair comparison requires a u

SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass

Model ReleasesDGX agent

arXiv:2602.06358v2 Announce Type: replace Abstract: We propose SHINE (Scalable Hyper In-context NEtwork), a scalable hypernetwork that can map diverse meaningful contexts into high-quality LoRA adapte

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

SafetyDGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

ShowMak3r: Compositional TV Show Reconstruction

ApplicationsDGX agent

arXiv:2504.19584v3 Announce Type: replace Abstract: Reconstructing dynamic radiance fields from video clips is challenging, especially when entertainment videos like TV shows are given. Many challenge

Single-Pass, Depth-Selective Reading for Multi-Aspect Sentiment Analysis

ResearchDGX agent

arXiv:2605.20998v1 Announce Type: new Abstract: Aspect-Term Sentiment Analysis (ATSA) in multi-aspect sentences faces a fundamental tradeoff between efficiency and expressiveness. Existing models eith

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches

ResearchDGX agent

arXiv:2605.20733v1 Announce Type: new Abstract: Converting hand-drawn sketches into structured 3D geometries remains challenging due to the difficulty of representing non-Euclidean surfaces and mainta

SMA-DP: Spectral Memory-Aware Differential Privacy for Deep Learning

SafetyDGX agent

arXiv:2605.20450v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) enables private deep learning through per-example clipping and calibrated Gaussian noise, bu

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

AgentsDGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

Smarter edits? Post-editing with error highlights and translation suggestions

ResearchDGX agent

arXiv:2605.21135v1 Announce Type: new Abstract: As MT quality increases, interest in enhanced post-editing features such as QE-derived error highlights is growing, yet evidence for their usefulness re

← Previous
1…631632633634635…1032
Next →