AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,414 results
Model Releases

Learning Robustness at Test-Time from a Non-Robust Teacher

DGX agent

arXiv:2604.11590v1 Announce Type: new Abstract: Nowadays, pretrained models are increasingly used as general-purpose backbones and adapted at test-time to downstream environments where target data are

model-releasesarxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Learning to Assist: Physics-Grounded Human-Human Control via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2603.11346v2 Announce Type: replace Abstract: Humanoid robotics has strong potential to transform daily service and caregiving applications. Although recent advances in general motion tracking w

agentsarxiv-cs-cv
14 Apr 2026
Research

Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video

DGX agent

arXiv:2511.18322v3 Announce Type: replace-cross Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpretability or rely on prior assumpt

researcharxiv-cs-cv
14 Apr 2026
Model Releases

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

DGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

DGX agent

arXiv:2604.10677v1 Announce Type: cross Abstract: Scaling up robot learning is hindered by the scarcity of robotic demonstrations, whereas human videos offer a vast, untapped source of interaction dat

safetyarxiv-cs-cv
14 Apr 2026
Research

LiveGesture Streamable Co-Speech Gesture Generation Model

DGX agent

arXiv:2604.10927v1 Announce Type: new Abstract: We propose LiveGesture, the first fully streamable, speech-driven full-body gesture generation framework that operates with zero look-ahead and supports

researcharxiv-cs-cv
14 Apr 2026
Research

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation

DGX agent

arXiv:2604.11789v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable progress in general-purpose vision--language understanding, yet they remain limited in tasks req

researcharxiv-cs-cv
14 Apr 2026
Research

LogitDynamics: Reliable ViT Error Detection from Layerwise Logit Trajectories

DGX agent

arXiv:2604.10643v1 Announce Type: new Abstract: Reliable confidence estimation is critical when deploying vision models. We study error prediction: determining whether an image classifier's output is

researcharxiv-cs-cv
14 Apr 2026
Model Releases

LoGo-MR: Screening Breast MRI for Cancer Risk Prediction by Efficient Omni-Slice Modeling

DGX agent

arXiv:2604.11348v1 Announce Type: new Abstract: Efficient and explainable breast cancer (BC) risk prediction is critical for large-scale population-based screening. Breast MRI provides functional info

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation

DGX agent

arXiv:2604.10103v1 Announce Type: new Abstract: Streaming video generation (SVG) distills a pretrained bidirectional video diffusion model into an autoregressive model equipped with sliding window att

local-aiarxiv-cs-cv
14 Apr 2026
Model Releases

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

DGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

DGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment: Methods and Results

DGX agent

arXiv:2604.11207v1 Announce Type: new Abstract: This paper reviews the LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment. This challenge aims to raise a new direction, i.e., how

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

DGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LumiMotion: Improving Gaussian Relighting with Scene Dynamics

DGX agent

arXiv:2604.10994v1 Announce Type: new Abstract: In 3D reconstruction, the problem of inverse rendering, namely recovering the illumination of the scene and the material properties, is fundamental. Exi

model-releasesarxiv-cs-cv
14 Apr 2026
Research

M^{2}SNet: Multi-scale in Multi-scale Subtraction Network for Medical Image Segmentation

DGX agent

arXiv:2303.10894v3 Announce Type: replace Abstract: Accurate medical image segmentation is critical for early medical diagnosis. Most existing methods are based on U-shape structure and use element-wi

researcharxiv-cs-cv
14 Apr 2026
Agents

MapATM: Enhancing HD Map Construction through Actor Trajectory Modeling

DGX agent

arXiv:2604.11081v1 Announce Type: new Abstract: High-definition (HD) mapping tasks, which perform lane detections and predictions, are extremely challenging due to non-ideal conditions such as view oc

agentsarxiv-cs-cv
14 Apr 2026
Applications

Masked Training for Robust Arrhythmia Detection from Digitalized Multiple Layout ECG Images

DGX agent

arXiv:2508.09165v3 Announce Type: replace-cross Abstract: Background: Electrocardiograms are indispensable for diagnosing cardiovascular diseases, yet in many settings they exist only as paper printou

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

DGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration

DGX agent

arXiv:2604.11197v1 Announce Type: new Abstract: Contrastive Language-Image Pre-training (CLIP) has demonstrated outstanding performance in global image understanding and zero-shot transfer through lar

local-aiarxiv-cs-cv
14 Apr 2026
Model Releases

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

DGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

MetroGS: Efficient and Stable Reconstruction of Geometrically Accurate High-Fidelity Large-Scale Scenes

DGX agent

arXiv:2511.19172v4 Announce Type: replace Abstract: Recently, 3D Gaussian Splatting and its derivatives have achieved significant breakthroughs in large-scale scene reconstruction. However, how to eff

tutorialsarxiv-cs-cv
14 Apr 2026
Research

Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation Models

DGX agent

arXiv:2604.10095v1 Announce Type: new Abstract: With the emergence of 3D foundation models, there is growing interest in fine-tuning them for downstream tasks, where LoRA is the dominant fine-tuning p

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Mirai: Autoregressive Visual Generation Needs Foresight

DGX agent

arXiv:2601.14671v2 Announce Type: replace Abstract: Autoregressive (AR) visual generators model images as sequences of discrete tokens and are trained with a next-token likelihood objective. This stri

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MLLM-as-a-Judge Exhibits Model Preference Bias

DGX agent

arXiv:2604.11589v1 Announce Type: new Abstract: Automatic evaluation using multimodal large language models (MLLMs), commonly referred to as MLLM-as-a-Judge, has been widely used to measure model perf

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark

DGX agent

arXiv:2604.10755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced clinical tasks for common conditions, but their performance on rare diseases remains largely unte

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

DGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

MosaicMRI: A Diverse Dataset and Benchmark for Raw Musculoskeletal MRI

DGX agent

arXiv:2604.11762v1 Announce Type: new Abstract: Deep learning underpins a wide range of applications in MRI, including reconstruction, artifact removal, and segmentation. However, progress has been dr

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

DGX agent

arXiv:2505.23606v4 Announce Type: replace-cross Abstract: Unified generation models aim to handle diverse tasks across modalities -- such as text generation, image generation, and vision-language reas

researcharxiv-cs-cv
14 Apr 2026
Safety

Multi-Granularity Reasoning for Image Quality Assessment via Attribute-Aware Reinforcement Learning to Rank

DGX agent

arXiv:2604.09704v1 Announce Type: new Abstract: Recent advances in reasoning-induced image quality assessment (IQA) have demonstrated the power of reinforcement learning to rank (RL2R) for training vi

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Multi-Head Attention based interaction-aware architecture for Bangla Handwritten Character Recognition: Introducing a Primary Dataset

DGX agent

arXiv:2604.09717v1 Announce Type: new Abstract: Character recognition is the fundamental part of an optical character recognition (OCR) system. Word recognition, sentence transcription, document digit

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi

DGX agent

arXiv:2604.10347v1 Announce Type: new Abstract: Vision foundation models have been shown to be effective at processing satellite imagery into representations fit for downstream tasks, however, creatin

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

MuPPet: Multi-person 2D-to-3D Pose Lifting

DGX agent

arXiv:2604.09715v1 Announce Type: new Abstract: Multi-person social interactions are inherently built on coherence and relationships among all individuals within the group, making multi-person localiz

local-aiarxiv-cs-cv
14 Apr 2026
Model Releases

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

DGX agent

arXiv:2512.00336v2 Announce Type: replace Abstract: The rapid advancement of AI-generated multimodal video-audio content has raised significant concerns regarding information security and content auth

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Naka-GS: A Bionics-inspired Dual-Branch Naka Correction and Progressive Point Pruning for Low-Light 3DGS

DGX agent

arXiv:2604.11142v1 Announce Type: new Abstract: Low-light conditions severely hinder 3D restoration and reconstruction by degrading image visibility, introducing color distortions, and contaminating g

safetyarxiv-cs-cv
14 Apr 2026
Research

Near OOD Detection for Vision-Language Prompt Learning with Contrastive Logit Score

DGX agent

arXiv:2405.16091v2 Announce Type: replace Abstract: Prompt learning has emerged as an efficient and effective method for fine-tuning vision-language models such as CLIP. While many studies have explor

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Neural Stochastic Processes for Satellite Precipitation Refinement

DGX agent

arXiv:2604.10414v1 Announce Type: new Abstract: Accurate precipitation estimation is critical for flood forecasting, water resource management, and disaster preparedness. Satellite products provide gl

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Neural Surface Reconstruction from Sparse Views Using Epipolar Geometry

DGX agent

arXiv:2406.04301v3 Announce Type: replace Abstract: Reconstructing accurate surfaces from sparse multi-view images remains challenging due to severe geometric ambiguity and occlusions. Existing genera

tutorialsarxiv-cs-cv
14 Apr 2026
Research

NeuVolEx: Implicit Neural Features for Volume Exploration

DGX agent

arXiv:2604.11172v1 Announce Type: cross Abstract: Direct volume rendering (DVR) aims to help users identify and examine regions of interest (ROIs) within volumetric data, and feature representations t

researcharxiv-cs-cv
14 Apr 2026
Applications

Ninja Codes: Neurally Generated Fiducial Markers for Stealthy 6-DoF Tracking

DGX agent

arXiv:2510.18976v2 Announce Type: replace Abstract: In this paper we describe Ninja Codes, neurally generated fiducial markers that can be made to naturally blend into various real-world environments.

applicationsarxiv-cs-cv
14 Apr 2026
Applications

NTIRE 2026 Challenge on Robust AI-Generated Image Detection in the Wild

DGX agent

arXiv:2604.11487v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Challenge on Robust AI-Generated Image Detection in the Wild, held in conjunction with the NTIRE works

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results

DGX agent

arXiv:2604.10551v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models. This challenge utili

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

NTIRE 2026 Challenge on Single Image Reflection Removal in the Wild: Datasets, Results, and Methods

DGX agent

arXiv:2604.10321v1 Announce Type: new Abstract: In this paper, we review the NTIRE 2026 challenge on single-image reflection removal (SIRR) in the Wild. SIRR is a fundamental task in image restoration

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: AI Flash Portrait (Track 3)

DGX agent

arXiv:2604.11230v1 Announce Type: new Abstract: In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 The Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results

DGX agent

arXiv:2604.10634v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images. Building upon the success

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Observe Less, Understand More: Cost-aware Cross-scale Observation for Remote Sensing Understanding

DGX agent

arXiv:2604.11415v1 Announce Type: new Abstract: Remote sensing understanding inherently requires multi-resolution observation, since different targets and application tasks demand different levels of

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video

DGX agent

arXiv:2604.11102v1 Announce Type: new Abstract: Current multimodal large language models (MLLMs) have demonstrated remarkable capabilities in short-form video understanding, yet translating long-form

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation

DGX agent

arXiv:2604.11804v1 Announce Type: new Abstract: In this work, we study Human-Object Interaction Video Generation (HOIVG), which aims to synthesize high-quality human-object interaction videos conditio

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…246247248249250…259
Next →