AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

Context Unrolling in Omni Models

DGX agent

arXiv:2604.21921v1 Announce Type: new Abstract: We present Omni, a unified multimodal model natively trained on diverse modalities, including text, images, videos, 3D geometry, and hidden representati

researcharxiv-cs-cv
24 Apr 2026
Research

Cross-Distribution Diffusion Priors-Driven Iterative Reconstruction for Sparse-View CT

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.13576v2 Announce Type: replace-cross Abstract: Sparse-View CT (SVCT) reconstruction enhances temporal resolution and reduces radiation dose, yet its clinical use is hindered by artifacts du

researcharxiv-cs-cv
24 Apr 2026
Model Releases

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

DGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

model-releasesarxiv-cs-cv
24 Apr 2026
Research

DCMorph: Face Morphing via Dual-Stream Cross-Attention Diffusion

DGX agent

arXiv:2604.21627v1 Announce Type: new Abstract: Advancing face morphing attack techniques is crucial to anticipate evolving threats and develop robust defensive mechanisms for identity verification sy

researcharxiv-cs-cv
24 Apr 2026
Research

Deep kernel video approximation for unsupervised action segmentation

DGX agent

arXiv:2604.21572v1 Announce Type: new Abstract: This work focuses on per-video unsupervised action segmentation, which is of interest to applications where storing large datasets is either not possibl

researcharxiv-cs-cv
24 Apr 2026
Safety

Demystifying Action Space Design for Robotic Manipulation Policies

DGX agent

arXiv:2602.23408v2 Announce Type: replace-cross Abstract: The specification of the action space plays a pivotal role in imitation-based robotic manipulation policy learning, fundamentally shaping the

safetyarxiv-cs-cv
24 Apr 2026
Safety

DepthMaster: Taming Diffusion Models for Monocular Depth Estimation

DGX agent

arXiv:2501.02576v2 Announce Type: replace Abstract: Monocular depth estimation within the diffusion-denoising paradigm demonstrates impressive generalization ability but suffers from low inference spe

safetyarxiv-cs-cv
24 Apr 2026
Research

DiffNR: Diffusion-Enhanced Neural Representation Optimization for Sparse-View 3D Tomographic Reconstruction

DGX agent

arXiv:2604.21518v1 Announce Type: cross Abstract: Neural representations (NRs), such as neural fields and 3D Gaussians, effectively model volumetric data in computed tomography (CT) but suffer from se

researcharxiv-cs-cv
24 Apr 2026
Safety

Directional Confusions Reveal Divergent Inductive Biases Through Rate-Distortion Geometry in Human and Machine Vision

DGX agent

arXiv:2604.21909v1 Announce Type: new Abstract: Humans and modern vision models can reach similar classification accuracy while making systematically different kinds of mistakes - differing not in how

safetyarxiv-cs-cv
24 Apr 2026
Applications

Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery

DGX agent

arXiv:2604.21712v1 Announce Type: new Abstract: 3D human mesh recovery from monocular RGB images aims to estimate anatomically plausible 3D human models for downstream applications, but remains challe

applicationsarxiv-cs-cv
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Tutorials

DualSplat: Robust 3D Gaussian Splatting via Pseudo-Mask Bootstrapping from Reconstruction Failures

DGX agent

arXiv:2604.21631v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) achieves real-time photorealistic rendering, its performance degrades significantly when training images contain tran

tutorialsarxiv-cs-cv
24 Apr 2026
Local Ai

EdgeFormer: local patch-based edge detection transformer on point clouds

DGX agent

arXiv:2604.21387v1 Announce Type: new Abstract: Edge points on 3D point clouds can clearly convey 3D geometry and surface characteristics, therefore, edge detection is widely used in many vision appli

local-aiarxiv-cs-cv
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

Encoder-Free Human Motion Understanding via Structured Motion Descriptions

DGX agent

arXiv:2604.21668v1 Announce Type: new Abstract: The world knowledge and reasoning capabilities of text-based large language models (LLMs) are advancing rapidly, yet current approaches to human motion

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

DGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

model-releasesarxiv-cs-cv
24 Apr 2026
Research

Foveated Reasoning: Stateful, Action-based Visual Focusing for Vision-Language Models

DGX agent

arXiv:2604.21079v1 Announce Type: new Abstract: Vision-language models benefit from high-resolution images, but the increase in visual-token count incurs high compute overhead. Humans resolve this ten

researcharxiv-cs-cv
24 Apr 2026
Model Releases

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

DGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

model-releasesarxiv-cs-cv
24 Apr 2026
Local Ai

Frozen LLMs as Map-Aware Spatio-Temporal Reasoners for Vehicle Trajectory Prediction

DGX agent

arXiv:2604.21479v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong reasoning capabilities and attracted increasing research attention in the field of autono

local-aiarxiv-cs-cv
24 Apr 2026
Safety

FryNet: Dual-Stream Adversarial Fusion for Non-Destructive Frying Oil Oxidation Assessment

DGX agent

arXiv:2604.21321v1 Announce Type: new Abstract: Monitoring frying oil degradation is critical for food safety, yet current practice relies on destructive wet-chemistry assays that provide no spatial i

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure

DGX agent

arXiv:2512.22274v2 Announce Type: replace Abstract: We introduce GeCo, a geometry-grounded metric for jointly detecting geometric deformation and occlusion-inconsistency artifacts in static scenes. By

model-releasesarxiv-cs-cv
24 Apr 2026
Local Ai

Geometry-aided Vision-based Localization of Future Mars Helicopters in Challenging Illumination Conditions

DGX agent

arXiv:2502.09795v3 Announce Type: replace Abstract: Planetary exploration using aerial assets has the potential for unprecedented scientific discoveries on Mars. While NASA's Mars helicopter Ingenuity

local-aiarxiv-cs-cv
24 Apr 2026
Local Ai

Gmd: Gaussian mixture descriptor for pair matching of 3D fragments

DGX agent

arXiv:2604.21519v1 Announce Type: new Abstract: In the automatic reassembly of fragments acquired using laser scanners to reconstruct objects, a crucial step is the matching of fractured surfaces. In

local-aiarxiv-cs-cv
24 Apr 2026
Hardware

GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA

DGX agent

arXiv:2604.21290v1 Announce Type: new Abstract: Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fix

hardwarearxiv-cs-cv
24 Apr 2026
Model Releases

Grounding Video Reasoning in Physical Signals

DGX agent

arXiv:2604.21873v1 Announce Type: new Abstract: Physical video understanding requires more than naming an event correctly. A model can answer a question about pouring, sliding, or collision from textu

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping

DGX agent

arXiv:2604.21127v1 Announce Type: new Abstract: The NASA PACE mission provides unprecedented hyperspectral observations of ocean color, aerosols, and clouds, offering new insights into how these compo

model-releasesarxiv-cs-cv
24 Apr 2026
Research

ID-Eraser: Proactive Defense Against Face Swapping via Identity Perturbation

DGX agent

arXiv:2604.21465v1 Announce Type: new Abstract: Deepfake technologies have rapidly advanced with modern generative AI, and face swapping in particular poses serious threats to privacy and digital secu

researcharxiv-cs-cv
24 Apr 2026
Local Ai

ImageHD: Energy-Efficient On-Device Continual Learning of Visual Representations via Hyperdimensional Computing

DGX agent

arXiv:2604.21280v1 Announce Type: new Abstract: On-device continual learning (CL) is critical for edge AI systems operating on non-stationary data streams, but most existing methods rely on backpropag

local-aiarxiv-cs-cv
24 Apr 2026
Tutorials

Information Bottleneck-Guided Heterogeneous Graph Learning for Interpretable Neurodevelopmental Disorder Diagnosis

DGX agent

arXiv:2502.20769v3 Announce Type: replace Abstract: Developing interpretable models for neurodevelopmental disorders (NDDs) diagnosis presents significant challenges in effectively encoding, decoding,

tutorialsarxiv-cs-cv
24 Apr 2026
Applications

Instance-level Visual Active Tracking with Occlusion-Aware Planning

DGX agent

arXiv:2604.21453v1 Announce Type: new Abstract: Visual Active Tracking (VAT) aims to control cameras to follow a target in 3D space, which is critical for applications like drone navigation and securi

applicationsarxiv-cs-cv
24 Apr 2026
Model Releases

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

DGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

KD-CVG: A Knowledge-Driven Approach for Creative Video Generation

DGX agent

arXiv:2604.21362v1 Announce Type: new Abstract: Creative Generation (CG) leverages generative models to automatically produce advertising content that highlights product features, and it has been a si

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

Latent Denoising Improves Visual Alignment in Large Multimodal Models

DGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

model-releasesarxiv-cs-cv
24 Apr 2026
Research

LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation

DGX agent

arXiv:2604.21279v1 Announce Type: new Abstract: Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control

researcharxiv-cs-cv
24 Apr 2026
Research

Linear Image Generation by Synthesizing Exposure Brackets

DGX agent

arXiv:2604.21008v1 Announce Type: new Abstract: The life of a photo begins with photons striking the sensor, whose signals are passed through a sophisticated image signal processing (ISP) pipeline to

researcharxiv-cs-cv
24 Apr 2026
Model Releases

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

DGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

Local Neighborhood Instability in Parametric Projections: Quantitative and Visual Analysis

DGX agent

arXiv:2604.21617v1 Announce Type: new Abstract: Parametric projections let analysts embed new points in real time, but input variations from measurement noise or data drift can produce unpredictable s

safetyarxiv-cs-cv
24 Apr 2026
Research

LRDUN: A Low-Rank Deep Unfolding Network for Efficient Spectral Compressive Imaging

DGX agent

arXiv:2511.18513v2 Announce Type: replace Abstract: Deep unfolding networks (DUNs) have achieved remarkable success and become the mainstream paradigm for spectral compressive imaging (SCI) reconstruc

researcharxiv-cs-cv
24 Apr 2026
Model Releases

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

DGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

model-releasesarxiv-cs-cv
24 Apr 2026
Research

Micro-DualNet: Dual-Path Spatio-Temporal Network for Micro-Action Recognition

DGX agent

arXiv:2604.21011v1 Announce Type: new Abstract: Micro-actions are subtle, localized movements lasting 1-3 seconds such as scratching one's head or tapping fingers. Such subtle actions are essential fo

researcharxiv-cs-cv
24 Apr 2026
Safety

Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational Adaptation

DGX agent

arXiv:2603.12845v2 Announce Type: replace Abstract: Predicting enzyme kinetic parameters quantifies how efficiently an enzyme catalyzes a specific substrate under defined biochemical conditions. Canon

safetyarxiv-cs-cv
24 Apr 2026
Research

Multiscale Super Resolution without Image Priors

DGX agent

arXiv:2604.21810v1 Announce Type: new Abstract: We address the ambiguities in the super-resolution problem under translation. We demonstrate that combinations of low-resolution images at different sca

researcharxiv-cs-cv
24 Apr 2026
Research

Neuro-Symbolic Manipulation Understanding with Enriched Semantic Event Chains

DGX agent

arXiv:2604.21053v1 Announce Type: cross Abstract: Robotic systems operating in human environments must reason about how object interactions evolve over time, which actions are currently being performe

researcharxiv-cs-cv
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Applications

Optimizing Diffusion Priors with a Single Observation

DGX agent

arXiv:2604.21066v1 Announce Type: new Abstract: While diffusion priors generate high-quality posterior samples across many inverse problems, they are often trained on limited training sets or purely s

applicationsarxiv-cs-cv
24 Apr 2026
Research

PanGuide3D: Cohort-Robust Pancreas Tumor Segmentation via Probabilistic Pancreas Conditioning and a Transformer Bottleneck

DGX agent

arXiv:2604.20981v1 Announce Type: cross Abstract: Pancreatic tumor segmentation in contrast-enhanced computed tomography (CT) is clinically important yet technically challenging: lesions are often sma

researcharxiv-cs-cv
24 Apr 2026
Research

PAT3D: Physics-Augmented Text-to-3D Scene Generation

DGX agent

arXiv:2511.21978v2 Announce Type: replace Abstract: We introduce PAT3D, the first physics-augmented text-to-3D scene generation framework that integrates vision-language models with physics-based simu

researcharxiv-cs-cv
24 Apr 2026
Research

PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & Editing

DGX agent

arXiv:2511.02777v2 Announce Type: replace Abstract: We present PercHead, a model for single-image 3D head reconstruction and disentangled 3D editing - two tasks that are inherently challenging due to

researcharxiv-cs-cv
24 Apr 2026
← Previous
1…217218219220221…261
Next →