AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
14 Apr 2026

Rethinking the Diffusion Model from a Langevin Perspective

ResearchDGX agent

arXiv:2604.10465v1 Announce Type: cross Abstract: Diffusion models are often introduced from multiple perspectives, such as VAEs, score matching, or flow matching, accompanied by dense and technically

Rethinking Token-Level Credit Assignment in RLVR: A Polarity-Entropy Analysis

SafetyDGX agent

arXiv:2604.11056v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has substantially improved the reasoning ability of Large Language Models (LLMs). However, its s

Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation

Model ReleasesDGX agent

arXiv:2604.10397v1 Announce Type: cross Abstract: Video-based human-object interaction (HOI) understanding requires both detecting ongoing interactions and anticipating their future evolution. However

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Retinal Cyst Detection from Optical Coherence Tomography Images

ResearchDGX agent

arXiv:2604.10843v1 Announce Type: cross Abstract: Retinal Cysts are formed by leakage and accumulation of fluid in the retina due to the incompetence of retinal vasculature. These cystic spaces have s

Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning

TutorialsDGX agent

arXiv:2604.11407v1 Announce Type: cross Abstract: We revisit retrieval-augmented generation (RAG) by embedding retrieval control directly into generation. Instead of treating retrieval as an external

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

Model ReleasesDGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure

ResearchDGX agent

arXiv:2604.11759v1 Announce Type: new Abstract: Organizational knowledge used by AI agents typically lacks epistemic structure: retrieval systems surface semantically relevant content without distingu

Retrieving to Recover: Towards Incomplete Audio-Visual Question Answering via Semantic-consistent Purification

ApplicationsDGX agent

arXiv:2604.10695v1 Announce Type: new Abstract: Recent Audio-Visual Question Answering (AVQA) methods have advanced significantly. However, most AVQA methods lack effective mechanisms for handling mis

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

SafetyDGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

Revisiting the Scale Loss Function and Gaussian-Shape Convolution for Infrared Small Target Detection

Model ReleasesDGX agent

arXiv:2604.09991v1 Announce Type: new Abstract: Infrared small target detection still faces two persistent challenges: training instability from non-monotonic scale loss functions, and inadequate spat

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

Model ReleasesDGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

RF-LEGO: Modularized Signal Processing-Deep Learning Co-Design for RF Sensing via Deep Unrolling

ApplicationsDGX agent

arXiv:2604.10183v1 Announce Type: cross Abstract: Wireless sensing, traditionally relying on signal processing (SP) techniques, has recently shifted toward data-driven deep learning (DL) to achieve pe

Rhizome OS-1: Rhizome's Semi-Autonomous Operating System for Small Molecule Drug Discovery

Model ReleasesDGX agent

arXiv:2604.07512v2 Announce Type: replace Abstract: We present Rhizome OS-1, a semi-autonomous operating system for small molecule drug discovery in which multi-modal AI agents operate as a full multi

Riemannian Zeroth-Order Gradient Estimation with Structure-Preserving Metrics for Geodesically Incomplete Manifolds

ResearchDGX agent

arXiv:2601.08039v2 Announce Type: replace Abstract: In this paper, we study Riemannian zeroth-order optimization in settings where the underlying Riemannian metric g is geodesically incomplete, and th

RISK: A Framework for GUI Agents in E-commerce Risk Management

Model ReleasesDGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

SafetyDGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

Model ReleasesDGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

Model ReleasesDGX agent

arXiv:2604.03768v2 Announce Type: replace Abstract: Unsustainable land-use practices in ecologically sensitive regions threaten biodiversity, water resources, and the livelihoods of millions. This pap

RL makes MLLMs see better than SFT

Model ReleasesDGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

Ro-SLM: Onboard Small Language Models for Robot Task Planning and Operation Code Generation

TutorialsDGX agent

arXiv:2604.10929v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) provide robots with contextual reasoning abilities to comprehend human instructions. Yet, current LLM-en

RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation

ResearchDGX agent

arXiv:2511.17441v3 Announce Type: replace Abstract: Despite the critical role of bimanual manipulation in endowing robots with human-like dexterity, large-scale and diverse datasets remain scarce due

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

Model ReleasesDGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

SafetyDGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

Robust Adversarial Policy Optimization Under Dynamics Uncertainty

Model ReleasesDGX agent

arXiv:2604.10974v1 Announce Type: new Abstract: Reinforcement learning (RL) policies often fail under dynamics that differ from training, a gap not fully addressed by domain randomization or existing

Robust Fair Disease Diagnosis in CT Images

Model ReleasesDGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

Robust Real-Time Coordination of CAVs: A Distributed Optimization Framework under Uncertainty

SafetyDGX agent

arXiv:2508.21322v2 Announce Type: replace Abstract: Achieving both safety guarantees and real-time performance in cooperative vehicle coordination remains a fundamental challenge, particularly in dyna

RobustMedSAM: Degradation-Resilient Medical Image Segmentation via Robust Foundation Model Adaptation

Model ReleasesDGX agent

arXiv:2604.09814v1 Announce Type: new Abstract: Medical image segmentation models built on Segment Anything Model (SAM) achieve strong performance on clean benchmarks, yet their reliability often degr

RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo

Model ReleasesDGX agent

arXiv:2505.09368v2 Announce Type: replace Abstract: Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image cor

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

Local AiDGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

rPPG-VQA: A Video Quality Assessment Framework for Unsupervised rPPG Training

ResearchDGX agent

arXiv:2604.11156v1 Announce Type: new Abstract: Unsupervised remote photoplethysmography (rPPG) promises to leverage unlabeled video data, but its potential is hindered by a critical challenge: traini

RTMC: Step-Level Credit Assignment via Rollout Trees

AgentsDGX agent

arXiv:2604.11037v1 Announce Type: cross Abstract: Multi-step agentic reinforcement learning benefits from fine-grained credit assignment, yet existing approaches offer limited options: critic-free met

RUMLEM: A Dictionary-Based Lemmatizer for Romansh

ResearchDGX agent

arXiv:2604.11233v1 Announce Type: new Abstract: Lemmatization -- the task of mapping an inflected word form to its dictionary form -- is a crucial component of many NLP applications. In this paper, we

S^3: Structured Sparsity Specification

ResearchDGX agent

arXiv:2604.11315v1 Announce Type: cross Abstract: We introduce the Structured Sparsity Specification (S^3), an algebraic framework for defining, composing, and implementing structured sparse patterns.

S4M: 4-points to Segment Anything

ResearchDGX agent

arXiv:2503.05534v3 Announce Type: replace Abstract: Purpose: The Segment Anything Model (SAM) promises to ease the annotation bottleneck in medical segmentation, but overlapping anatomy and blurred bo

Saar-Voice: A Multi-Speaker Saarbrucken Dialect Speech Corpus

Local AiDGX agent

arXiv:2604.11803v1 Announce Type: new Abstract: Natural language processing (NLP) and speech technologies have made significant progress in recent years; however, they remain largely focused on standa

Safe Human-to-Humanoid Motion Imitation Using Control Barrier Functions

SafetyDGX agent

arXiv:2604.11447v1 Announce Type: new Abstract: Ensuring operational safety is critical for human-to-humanoid motion imitation. This paper presents a vision-based framework that enables a humanoid rob

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

SafetyDGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

Safety Guarantees in Zero-Shot Reinforcement Learning for Cascade Dynamical Systems

SafetyDGX agent

arXiv:2604.10429v1 Announce Type: new Abstract: This paper considers the problem of zero-shot safety guarantees for cascade dynamical systems. These are systems where a subset of the states (the inner

Sanity Checks for Agentic Data Science

AgentsDGX agent

arXiv:2604.11003v1 Announce Type: new Abstract: Agentic data science (ADS) pipelines have grown rapidly in both capability and adoption, with systems such as OpenAI Codex now able to directly analyze

Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping

ResearchDGX agent

arXiv:2505.13777v2 Announce Type: replace-cross Abstract: We present Sat2Sound, a unified multimodal framework for geospatial soundscape understanding, designed to predict and map the distribution of

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

HardwareDGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

SBAMP: Sampling Based Adaptive Motion Planning

AgentsDGX agent

arXiv:2511.12022v2 Announce Type: replace Abstract: Autonomous robots operating in dynamic environments must balance global path optimality with real-time responsiveness to disturbances. This requires

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight

Model ReleasesDGX agent

arXiv:2512.19691v3 Announce Type: replace Abstract: Reference labels for machine-learning benchmarks are increasingly synthesized with LLM assistance, but their reliability remains underexamined. We a

Scaling Up AI-Generated Image Detection with Generator-Aware Prototypes

ResearchDGX agent

arXiv:2512.12982v2 Announce Type: replace Abstract: The pursuit of a universal AI-generated image (AIGI) detector often relies on aggregating data from numerous generators to improve generalization. H

Scene Change Detection with Vision-Language Representation Learning

ApplicationsDGX agent

arXiv:2604.11402v1 Announce Type: new Abstract: Scene change detection (SCD) is crucial for urban monitoring and navigation but remains challenging in real-world environments due to lighting variation

SciPostLayoutTree: A Dataset for Structural Analysis of Scientific Posters

ResearchDGX agent

arXiv:2511.18329v3 Announce Type: replace Abstract: Scientific posters play a vital role in academic communication by presenting ideas through visual summaries. Analyzing reading order and parent-chil

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

Model ReleasesDGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

Model ReleasesDGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

HardwareDGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

Model ReleasesDGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

SafetyDGX agent

arXiv:2604.10688v1 Announce Type: cross Abstract: On-policy reinforcement learning has become the dominant paradigm for reasoning alignment in large language models, yet its sparse, outcome-level rewa

ScoRe-Flow: Complete Distributional Control via Score-Based Reinforcement Learning for Flow Matching

ResearchDGX agent

arXiv:2604.10962v1 Announce Type: new Abstract: Flow Matching (FM) policies have emerged as an efficient backbone for robotic control, offering fast and expressive action generation that underpins rec

Score-matching-based Structure Learning for Temporal Data on Networks

ApplicationsDGX agent

arXiv:2412.07469v3 Announce Type: replace-cross Abstract: Causal discovery is a crucial initial step in establishing causality from empirical data and background knowledge. Numerous algorithms have be

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding

Model ReleasesDGX agent

arXiv:2604.11244v1 Announce Type: new Abstract: Advances in Multimodal Large Language Models (MLLMs) are transforming video captioning from a descriptive endpoint into a semantic interface for both vi

Search-MIND: Training-Free Multi-Modal Medical Image Registration

Local AiDGX agent

arXiv:2604.09743v1 Announce Type: cross Abstract: Multi-modal image registration plays a critical role in precision medicine but faces challenges from non-linear intensity relationships and local opti

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

SafetyDGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

Model ReleasesDGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment

SafetyDGX agent

arXiv:2604.09749v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) frequently hallucinate objects that are absent from the visual input, often because attention during decoding i

← Previous
1…954955956957958…989
Next →