AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

Bridging the Ex-Vivo to In-Vivo Gap: Synthetic Priors for Monocular Depth Estimation in Specular Surgical Environments

AgentsDGX agent

arXiv:2512.23786v2 Announce Type: replace Abstract: Accurate Monocular Depth Estimation (MDE) is critical for autonomous robotic surgery. However, existing self-supervised methods often exhibit a seve

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

Model ReleasesDGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.11038v2 Announce Type: replace Abstract: We study the reasoning behavior of large language models (LLMs) under limited computation budgets. In such settings, producing useful partial soluti

C-GenReg: Training-Free 3D Point Cloud Registration by Multi-View-Consistent Geometry-to-Image Generation with Probabilistic Modalities Fusion

SafetyDGX agent

arXiv:2604.16680v1 Announce Type: new Abstract: We introduce C-GenReg, a training-free framework for 3D point cloud registration that leverages the complementary strengths of world-scale generative pr

CAARL: In-Context Learning for Interpretable Co-Evolving Time Series Forecasting

TutorialsDGX agent

arXiv:2604.18305v1 Announce Type: new Abstract: In this paper we investigate forecasting coevolving time series that feature intricate dependencies and nonstationary dynamics by using an LLM Large Lan

Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA

ApplicationsDGX agent

arXiv:2604.17316v1 Announce Type: new Abstract: Safe clinical deployment of Large Language Models (LLMs) requires not only high accuracy but also robust uncertainty calibration to ensure models defer

Calibrating Model-Based Evaluation Metrics for Summarization

ResearchDGX agent

arXiv:2604.17200v1 Announce Type: new Abstract: Recent advances in summary evaluation are based on model-based metrics to assess quality dimensions, such as completeness, conciseness, and faithfulness

CAM3DNet: Comprehensively mining the multi-scale features for 3D Object Detection with Multi-View Cameras

Model ReleasesDGX agent

arXiv:2604.17024v1 Announce Type: new Abstract: Query-based 3D object detection methods using multi-view images often struggle to efficiently leverage dynamic multi-scale information, e.g., the relati

Camo-M3FD: A New Benchmark Dataset for Cross-Spectral Camouflaged Pedestrian Detection

Model ReleasesDGX agent

arXiv:2604.16582v1 Announce Type: new Abstract: Pedestrian detection is fundamental to autonomous driving, robotics, and surveillance. Despite progress in deep learning, reliable identification remain

CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion

ResearchDGX agent

arXiv:2509.19979v2 Announce Type: replace Abstract: Recently, camera-controlled video generation has seen rapid development, offering more precise control over video generation. However, existing meth

Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study

SafetyDGX agent

arXiv:2604.17896v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models map multimodal inputs directly to robot actions and are typically trained through large-scale imitation learning. Wh

Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?

Model ReleasesDGX agent

arXiv:2604.18134v1 Announce Type: new Abstract: Recent advancements in self-supervised learning have led to powerful surgical vision encoders capable of spatiotemporal understanding. However, extendin

Can we generate portable representations for clinical time series data using LLMs?

ApplicationsDGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition

ApplicationsDGX agent

arXiv:2604.18184v1 Announce Type: new Abstract: Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

SafetyDGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

CAPO: Counterfactual Credit Assignment in Sequential Cooperative Teams

SafetyDGX agent

arXiv:2604.17693v1 Announce Type: new Abstract: In cooperative teams where agents act in a fixed order and share a single team reward, it is hard to know how much each agent contributed, and harder st

Capture Timing-Attention of Events in Clinical Time Series

Model ReleasesDGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction

Model ReleasesDGX agent

arXiv:2512.11988v3 Announce Type: replace Abstract: Accurate capture of human-object interaction from ubiquitous sensors like RGB cameras is important for applications in human understanding, gaming,

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

Model ReleasesDGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

Cat-DPO: Category-Adaptive Safety Alignment

SafetyDGX agent

arXiv:2604.17299v1 Announce Type: new Abstract: Aligning large language models with human preferences must balance two competing goals: responding helpfully to legitimate requests and reliably refusin

CATP: Confidence-Aware Token Pruning for Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2604.16854v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to segment targets that share extreme textural and structural similarities with their complex environments. Leve

CaTS-Bench: Can Language Models Describe Time Series?

Model ReleasesDGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

Causally-Constrained Probabilistic Forecasting for Time-Series Anomaly Detection

Local AiDGX agent

arXiv:2604.17998v1 Announce Type: new Abstract: Anomaly detection in multivariate time series is a central challenge in industrial monitoring, as failures frequently arise from complex temporal dynami

CBR-to-SQL: Rethinking Retrieval-based Text-to-SQL using Case-based Reasoning in the Healthcare Domain

ApplicationsDGX agent

arXiv:2603.05569v2 Announce Type: replace-cross Abstract: Extracting insights from Electronic Health Record (EHR) databases often requires SQL expertise, creating a barrier for clinical decision-makin

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

Model ReleasesDGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

CCAR: Intrinsic Robustness as an Emergent Geometric Property

SafetyDGX agent

arXiv:2604.16861v1 Announce Type: cross Abstract: Standard supervised learning optimizes for predictive accuracy but remains agnostic to the internal geometry of learned features, often yielding repre

CDSA-Net:Collaborative Decoupling of Vascular Structure and Background for High-Fidelity Coronary Digital Subtraction Angiography

Model ReleasesDGX agent

arXiv:2604.17208v1 Announce Type: new Abstract: Digital subtraction angiography (DSA) in coronary imaging is fundamentally challenged by physiological motion, forcing reliance on raw angiograms clutte

Central Limit Theorems for Asynchronous Averaged Q-Learning

ResearchDGX agent

arXiv:2509.18964v3 Announce Type: replace Abstract: This paper establishes central limit theorems for Polyak-Ruppert averaged Q-learning under asynchronous updates. We prove a non-asymptotic central l

Centre manifold theorem for maps along manifolds of fixed points

ResearchDGX agent

arXiv:2604.18202v1 Announce Type: cross Abstract: We prove a centre manifold theorem for a map along a manifold-with-boundary of fixed points, and provide an application to the study of gradient desce

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

Model ReleasesDGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

CFSR: Geometry-Conditioned Shadow Removal via Physical Disentanglement

Local AiDGX agent

arXiv:2604.18032v1 Announce Type: new Abstract: Traditional shadow removal networks often treat image restoration as an unconstrained mapping, lacking the physical interpretability required to balance

CGCMA: Conditionally-Gated Cross-Modal Attention for Event-Conditioned Asynchronous Fusion

SafetyDGX agent

arXiv:2604.16411v1 Announce Type: new Abstract: We study asynchronous alignment, a first-class multimodal learning setting in which a dense primary stream must be fused with sporadic external context

Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction

Model ReleasesDGX agent

arXiv:2604.16886v1 Announce Type: new Abstract: Generalist embodied agents must perform interactive, causally-dependent reasoning, continually interacting with the environment, acquiring information,

Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation

Model ReleasesDGX agent

arXiv:2604.16806v1 Announce Type: new Abstract: Referring image segmentation (RIS) requires accurate segmentation of target regions in images according to language descriptions, which is a cross-modal

Chaos-Enhanced Prototypical Networks for Few-Shot Medical Image Classification

ResearchDGX agent

arXiv:2604.17300v1 Announce Type: cross Abstract: The scarcity of labeled clinical data in oncology makes Few-Shot Learning (FSL) a critical framework for Computer Aided Diagnostics, but we observed t

Characterizing Model-Native Skills

SafetyDGX agent

arXiv:2604.17614v1 Announce Type: cross Abstract: Skills are a natural unit for describing what a language model can do and how its behavior can be changed. However, existing characterizations rely on

Chasing Ghosts: A Simulation-to-Real Olfactory Navigation Stack with Optional Vision Augmentation

SafetyDGX agent

arXiv:2602.19577v2 Announce Type: replace Abstract: Autonomous odor source localization remains a challenging problem for aerial robots due to turbulent airflow, sparse and delayed sensory signals, an

Chatting about Conditional Trajectory Prediction

AgentsDGX agent

arXiv:2604.18126v1 Announce Type: cross Abstract: Human behavior has the nature of mutual dependencies, which requires human-robot interactive systems to predict surrounding agents trajectories by mod

Chatting about Upper-Body Expressive Human Pose and Shape Estimation

Model ReleasesDGX agent

arXiv:2604.17959v1 Announce Type: new Abstract: Expressive Human Pose and Shape Estimation (EHPS) plays a crucial role in various AR/VR applications and has witnessed significant progress in recent ye

CHIMERA: A Knowledge Base of Scientific Idea Recombinations for Research Analysis and Ideation

ResearchDGX agent

arXiv:2505.20779v5 Announce Type: replace Abstract: A hallmark of human innovation is recombination -- the creation of novel ideas by integrating elements from existing concepts and mechanisms. In thi

Chronax: A Jax Library for Univariate Statistical Forecasting and Conformal Inference

ApplicationsDGX agent

arXiv:2604.16719v1 Announce Type: new Abstract: Time-series forecasting is central to many scientific and industrial domains, such as energy systems, climate modeling, finance, and retail. While forec

CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents

AgentsDGX agent

arXiv:2603.15421v2 Announce Type: replace Abstract: Large language model agents heavily rely on external memory to support knowledge reuse and complex reasoning tasks. Yet most memory systems store ex

CLASP: Training-Free LLM-Assisted Source Code Watermarking via Semantic-Preserving Transformations

Local AiDGX agent

arXiv:2510.11251v2 Announce Type: replace-cross Abstract: The proliferation of open-source code and large language models (LLMs) for code generation has amplified the risks of unauthorized reuse and i

Class-specific diffusion models improve military object detection in a low-data domain

ResearchDGX agent

arXiv:2604.18076v1 Announce Type: new Abstract: Diffusion-based image synthesis has emerged as a promising source of synthetic training data for AI-based object detection and classification. In this w

Classification of systolic murmurs in heart sounds using multiresolution complex Gabor dictionary and vision transformer

ResearchDGX agent

arXiv:2604.16563v1 Announce Type: new Abstract: Systolic murmurs are extra heart sounds that occur during the contraction phase of the cardiac cycle, often indicating heart abnormalities caused by tur

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

Model ReleasesDGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

Clinical Note Bloat Reduction for Efficient LLM Use

ResearchDGX agent

arXiv:2604.16364v1 Announce Type: cross Abstract: Health systems are rapidly deploying large language models (LLMs) that use clinical notes for clinical decision support applications. However, modern

Closing the Modality Reasoning Gap for Speech Large Language Models

SafetyDGX agent

arXiv:2601.05543v2 Announce Type: replace Abstract: Although Speech Large Language Models have achieved notable progress, a substantial modality reasoning gap remains: their reasoning performance on s

Clusterability-Based Assessment of Potentially Noisy Views for Multi-View Clustering

ApplicationsDGX agent

arXiv:2604.18024v1 Announce Type: new Abstract: In multi-view clustering, the quality of different views may vary substantially, and low-quality or degraded views can impair overall clustering perform

Co-generation of Layout and Shape from Text via Autoregressive 3D Diffusion

ResearchDGX agent

arXiv:2604.16552v1 Announce Type: new Abstract: Recent text-to-scene generation approaches largely reduced the manual efforts required to create 3D scenes. However, their focus is either to generate a

CoAct: Co-Active LLM Preference Learning with Human-AI Synergy

ResearchDGX agent

arXiv:2604.17501v1 Announce Type: new Abstract: Learning from preference-based feedback has become an effective approach for aligning LLMs across diverse tasks. However, high-quality human-annotated p

CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

Model ReleasesDGX agent

arXiv:2604.18027v1 Announce Type: cross Abstract: Transpilation, or code translation, aims to convert source code from one programming language (PL) to another. It is beneficial for many downstream ap

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

Model ReleasesDGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

Coevolving Representations in Joint Image-Feature Diffusion

ResearchDGX agent

arXiv:2604.17492v1 Announce Type: new Abstract: Joint image-feature generative modeling has recently emerged as an effective strategy for improving diffusion training by coupling low-level VAE latents

COFFAIL: A Dataset of Successful and Anomalous Robot Skill Executions in the Context of Coffee Preparation

SafetyDGX agent

arXiv:2604.18236v1 Announce Type: new Abstract: In the context of robot learning for manipulation, curated datasets are an important resource for advancing the state of the art; however, available dat

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

AgentsDGX agent

arXiv:2509.00789v2 Announce Type: replace Abstract: The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs):

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

Model ReleasesDGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

Cognitive Policy-Driven LLM for Diagnosis and Intervention of Cognitive Distortions in Emotional Support Conversation

SafetyDGX agent

arXiv:2604.17178v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) plays a critical role in mental health assistance by providing accessible psychological support in real-world appli

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering

TutorialsDGX agent

arXiv:2604.16930v1 Announce Type: new Abstract: Visual Question Answering (VQA) requires models to identify the correct answer options based on both visual and textual evidence. Recent Mixture-of-Expe

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

Model ReleasesDGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

← Previous
1…883884885886887…998
Next →