AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

UltraStar: Semantic-Aware Star Graph Modeling for Echocardiography Navigation

DGX agent

arXiv:2603.01461v2 Announce Type: replace Abstract: Echocardiography is critical for diagnosing cardiovascular diseases, yet the shortage of skilled sonographers hinders timely patient care, due to hi

researcharxiv-cs-cv
26 Jun 2026
Agents
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

UniFlow: Zero-Shot LiDAR Scene Flow for Autonomous Vehicles

DGX agent

arXiv:2511.18254v3 Announce Type: replace Abstract: LiDAR scene flow is the task of estimating per-point 3D motion between consecutive point clouds. Recent methods achieve centimeter-level accuracy on

agentsarxiv-cs-cv
26 Jun 2026
Model Releases

Unison: Benchmarking Unified Multimodal Models via Synergistic Understanding and Generation

DGX agent

arXiv:2606.26984v1 Announce Type: new Abstract: Unified multimodal models capable of both understanding and generation have achieved remarkable strides. However, despite their unified designs, existin

model-releasesarxiv-cs-cv
26 Jun 2026
Research

ViQ: Text-Aligned Visual Quantized Representations at Any Resolution

DGX agent

arXiv:2606.27313v1 Announce Type: new Abstract: A unified representation for text and vision is a natural pursuit, as it enables simpler multimodal modeling and more efficient training. However, repre

researcharxiv-cs-cv
26 Jun 2026
Safety

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning

DGX agent

arXiv:2606.18974v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) interleave generated ''visual thoughts'' (VTs) with text reasoning to improve spatial tasks. This incurs roughly an

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations

DGX agent

arXiv:2606.26384v1 Announce Type: new Abstract: As deepfake generators approach perceptual indistinguishability, reliable detection becomes critical. Yet, detectors that score well on benchmarks routi

model-releasesarxiv-cs-cv
26 Jun 2026
Safety

World Action Models Enable Continual Imitation Learning with Recurrent Generative Replays

DGX agent

arXiv:2606.27374v1 Announce Type: cross Abstract: Going beyond predicting robot actions, World Action Models (WAMs) can also generate future visual observations. We build on this generative capability

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

1000 Rallies: An Event-Camera Dataset and Real-Time Learned Ball-State Estimation for Robotic Table Tennis

DGX agent

arXiv:2606.25620v1 Announce Type: cross Abstract: Robotic table tennis has emerged as a compelling benchmark for real-time robotic perception due to its fast ball dynamics and stringent timing require

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

2K Retrofit: Entropy-Guided Efficient Sparse Refinement for High-Resolution 3D Geometry Prediction

DGX agent

arXiv:2603.19964v3 Announce Type: replace Abstract: High-resolution geometric prediction is essential for robust perception in autonomous driving, robotics, and AR/MR, but current foundation models ar

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

A Benchmark for Heterogeneous Stereo Deblurring with Physically- and Epipolar-constrained Cross Attention

DGX agent

arXiv:2606.25962v1 Announce Type: new Abstract: Modern stereo-capable smartphones enable immersive XR content capture. However, hardware heterogeneity across camera modules often causes severe asymmet

model-releasesarxiv-cs-cv
25 Jun 2026
Research

A cross-process welding penetration status prediction algorithm based on unsupervised domain adaptation in laser and TIG welding

DGX agent

arXiv:2606.26078v1 Announce Type: new Abstract: Supervised deep learning has been widely used for weld penetration state classification; however, its performance often degrades significantly under dom

researcharxiv-cs-cv
25 Jun 2026
Model Releases

A Leakage-Aware Comparative Benchmark of Machine Learning, Deep Learning, and Transformer Models for Reliable Leukemia Detection

DGX agent

arXiv:2606.24944v1 Announce Type: cross Abstract: Automated classification of acute lymphoblastic leukemia (ALL) from peripheral blood smear images has often reported near-perfect performance on the C

model-releasesarxiv-cs-cv
25 Jun 2026
Research

A welding penetration prediction model for laser welding process based on self-supervised learning using physics-informed neural networks

DGX agent

arXiv:2606.26059v1 Announce Type: new Abstract: The laser welding full-penetration is of critical importance, as it constitutes one of the fundamental factors in achieving defect-free welded joints. A

researcharxiv-cs-cv
25 Jun 2026
Applications

ADM-Fusion: Adaptive Deep Multi-Sensor Fusion for Robust Ego-Motion Estimation in Diverse Conditions

DGX agent

arXiv:2606.25111v1 Announce Type: cross Abstract: Robust multi-sensor fusion is essential for reliable autonomy in diverse and degraded environments, where sensor reliability can fluctuate rapidly. Be

applicationsarxiv-cs-cv
25 Jun 2026
Model Releases

AISPO: Enhancing Depth Reliability for Robotic Manipulation of Non-Lambertian Objects via Affine-Invariant Shape Prior

DGX agent

arXiv:2606.25503v1 Announce Type: cross Abstract: Reliable depth perception is critical for robotic manipulation, especially for non-Lambertian objects such as transparent or highly specular surfaces,

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs

DGX agent

arXiv:2601.17037v2 Announce Type: replace Abstract: We investigate visual reasoning limitations of both multimodal large language models (MLLMs) and image generation models (IGMs) by creating a novel

model-releasesarxiv-cs-cv
25 Jun 2026
Research

An Improved Variational Method for Image Denoising

DGX agent

arXiv:2410.02587v2 Announce Type: replace Abstract: The total variation (TV) method is an image denoising technique that aims to reduce noise by minimizing the total variation of the image, which meas

researcharxiv-cs-cv
25 Jun 2026
Tutorials

An Integrated Hardware-Software Design for Low-Data Spatial Defect Detection in Robotic Visual Inspection with Hybrid Optoelectronic Neural Networks

DGX agent

arXiv:2606.25277v1 Announce Type: cross Abstract: To address data overload and inefficient shape-level annotation in robotic visual inspection, this paper proposes a hardware-software integrated optoe

tutorialsarxiv-cs-cv
25 Jun 2026
Model Releases

An iterative energy-based multimodal transformer for joint retrieval of wheat soil moisture, leaf area index, and plant height from Sentinel-1 and Sentinel-2 time series

DGX agent

arXiv:2606.25174v1 Announce Type: cross Abstract: Field-scale retrieval of surface soil moisture (SM), leaf area index (LAI), and plant height (PH) is essential for precision agriculture, yet it remai

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Anatomically-conditioned Latent Diffusion Model for Data-Efficient Few-Shot Cross-Domain 3D Glioma MRI Synthesis

DGX agent

arXiv:2606.25390v1 Announce Type: new Abstract: Accurate classification of diffuse gliomas is often hindered by domain shifts across centers and a lack of large, annotated datasets. We propose the Ana

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Are We There Yet? Exploring the Capabilities of MLLMs in Assistive AI Applications

DGX agent

arXiv:2606.25084v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have redefined visual understanding by combining vision encoders with large-scale language models. This unified

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

ArteryX: A Reliable End-to-End Toolbox for Standardized Intracranial Artery Feature Extraction from 3D TOF-MRA

DGX agent

arXiv:2507.07920v2 Announce Type: replace-cross Abstract: Cerebrovascular research heavily relies on quantitative analysis of intracranial arteries from time-of-flight magnetic resonance angiography,

safetyarxiv-cs-cv
25 Jun 2026
Applications

Articulat3D: Reconstructing Articulated Digital Twins From Monocular Videos with Geometric and Motion Constraints

DGX agent

arXiv:2603.11606v2 Announce Type: replace Abstract: Building high-fidelity digital twins of articulated objects from visual data remains a central challenge. Existing approaches depend on multi-view c

applicationsarxiv-cs-cv
25 Jun 2026
Agents

ASSCG: Just-Right Gating over Chattering for Fast-Slow LLM Planning in Autonomous Driving

DGX agent

arXiv:2606.25509v1 Announce Type: cross Abstract: Large language models (LLMs) can improve autonomous driving planning but are costly to query online, and existing fast-slow planners often rely on han

agentsarxiv-cs-cv
25 Jun 2026
Model Releases

Auto-Labelling-Based Domain Transfer for 3D Object Detection on a Bicycle-Mounted LiDAR Platform

DGX agent

arXiv:2606.25652v1 Announce Type: new Abstract: Reliable 3D perception of vulnerable road users (VRUs) such as cyclists and pedestrians is essential for their safety in urban traffic and a core requir

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Benchmarking Deep Learning Models for Laryngeal Cancer Staging Using the LaryngealCT Dataset

DGX agent

arXiv:2510.11047v2 Announce Type: replace Abstract: Laryngeal cancer imaging research lacks standardised public datasets to enable reproducible deep learning (DL) model development. We present Larynge

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Benchmarking the Alignment of Data-Quality Metrics, Human Judgment and Land-Cover Segmentation Performance for Earth Observation

DGX agent

arXiv:2606.25128v1 Announce Type: cross Abstract: Volume and quality of datasets are crucial for deep learning model training, yet they are often constrained by availability and data acquisition costs

safetyarxiv-cs-cv
25 Jun 2026
Model Releases

Beyond Visual Forensics: Auditing Multimodal Robustness for Synthetic Medical Image Detection

DGX agent

arXiv:2606.25375v1 Announce Type: new Abstract: With the rapid adoption of generative AI, synthetic medical images pose growing risks, including diagnostic deception and insurance fraud. Although prio

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

BOFA: Bridge-Layer Orthogonal Low-Rank Fusion for CLIP-Based Class-Incremental Learning

DGX agent

arXiv:2511.11421v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) aims to continually learn new categories without forgetting previously acquired knowledge. Vision-language models s

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation

DGX agent

arXiv:2606.25432v1 Announce Type: cross Abstract: Inference efficiency is typically pursued by shrinking the model: distillation, pruning, quantization, and sparse routing each lower per-token cost wh

researcharxiv-cs-cv
25 Jun 2026
Local Ai

C2RM-Seg: Causal Counterfactual Reasoning with Structural-Semantic Priors for Weakly Supervised Histopathological Tissue Segmentation

DGX agent

arXiv:2606.25508v1 Announce Type: new Abstract: Histopathological tissue segmentation is essential for computer-aided diagnosis, yet weakly supervised methods often suffer from noisy pseudo-labels gen

local-aiarxiv-cs-cv
25 Jun 2026
Model Releases

C3-Bench: A Context-Aware Change Captioning Benchmark

DGX agent

arXiv:2606.25445v1 Announce Type: new Abstract: While Change Captioning systems have garnered substantial attention to respond to our evolving world, their true performance on diverse real-world chang

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Cage-based Texture Transfer with Geometric Filtering

DGX agent

arXiv:2606.25220v1 Announce Type: new Abstract: Real-time texture transfer expands the creative horizon for interactive applications, enabling seamless detail projection in scenarios that range from d

researcharxiv-cs-cv
25 Jun 2026
Research

Calousel: Extrinsic Calibration of Non-overlapping Multi-camera Systems from Pure Rotation

DGX agent

arXiv:2606.25646v1 Announce Type: cross Abstract: Extrinsic calibration of multi-camera systems with non-overlapping FOVs has been a challenging problem in the robotics literature. Conventional target

researcharxiv-cs-cv
25 Jun 2026
Safety

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

DGX agent

arXiv:2606.25473v1 Announce Type: new Abstract: Autoregressive video diffusion with causal diffusion transformers has emerged as a major paradigm for real-time streaming video generation and action-co

safetyarxiv-cs-cv
25 Jun 2026
Research

Chorus II: Cross-Request Sparsity Reuse for Efficient Image-to-Video Generation

DGX agent

arXiv:2606.25040v1 Announce Type: new Abstract: Serving diffusion models for image-to-video generation is computationally expensive, posing significant challenges for large-scale deployment. Real I2V

researcharxiv-cs-cv
25 Jun 2026
Research

CoGeoAD: Hierarchical Color-Geometric Fusion with Multi-View Attention for Zero-Shot 3D Anomaly Detection

DGX agent

arXiv:2606.25273v1 Announce Type: new Abstract: Zero-shot 3D anomaly detection is essential for industrial quality inspection, where labeled anomaly samples are scarce. Meanwhile, existing methods lac

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos

DGX agent

arXiv:2603.25645v2 Announce Type: replace-cross Abstract: Early screening via colonoscopy is critical for colon cancer prevention, yet developing robust AI systems for this domain is hindered by the l

model-releasesarxiv-cs-cv
25 Jun 2026
Research

Color Matters: Trigger Color Affects Success in Federated Backdoor Attacks

DGX agent

arXiv:2606.25858v1 Announce Type: cross Abstract: Federated learning is vulnerable to backdoor attacks in which malicious clients inject poisoned updates while preserving benign-task performance. In t

researcharxiv-cs-cv
25 Jun 2026
Research

Concept Removal for Frontier Image Generative Models

DGX agent

arXiv:2606.25548v1 Announce Type: new Abstract: Image generative models are trained on massive, largely uncurated internet-scale datasets that contain undesirable visual concepts. Efficiently removing

researcharxiv-cs-cv
25 Jun 2026
Safety

Contrastive Conditional-Unconditional Alignment for Long-tailed Diffusion Model

DGX agent

arXiv:2507.09052v3 Announce Type: replace Abstract: Training data for class-conditional image synthesis often exhibit a long-tailed distribution with limited amount of images for tail classes. Such an

safetyarxiv-cs-cv
25 Jun 2026
Research

Counterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering

DGX agent

arXiv:2512.04554v2 Announce Type: replace Abstract: Document Visual Question Answering (DocVQA) enables end-to-end reasoning grounded on information present in a document input. While recent models ha

researcharxiv-cs-cv
25 Jun 2026
Research

Cross-Attention Multimodal Learning for Predicting Response to Neoadjuvant Imatinib in Gastrointestinal Stromal Tumors: A Multicenter Retrospective Study

DGX agent

arXiv:2606.25579v1 Announce Type: cross Abstract: Background: Response to neoadjuvant imatinib in gastrointestinal stromal tumors (GISTs) is highly variable and cannot be reliably predicted using curr

researcharxiv-cs-cv
25 Jun 2026
Tutorials

Cross-Modality Structural Guidance in 3D Latent Diffusion for Robust FLAIR Super-Resolution

DGX agent

arXiv:2606.25255v1 Announce Type: new Abstract: High-resolution (HR) MRI acquisition is often hampered by scan time constraints, resulting in anisotropic or low-resolution scans (e.g., thick-slice FLA

tutorialsarxiv-cs-cv
25 Jun 2026
Safety

Cross-View Variance Correlation in Path-Traced Stereo:A Hidden Shortcut in Synthetic Training Data

DGX agent

arXiv:2606.25483v1 Announce Type: new Abstract: Path-traced synthetic stereo data underlie a large fraction of modern disparity-estimation training pipelines. We report a previously unrecognised prope

safetyarxiv-cs-cv
25 Jun 2026
Model Releases

Curvature-Guided Mixing for MLLM Adaptation

DGX agent

arXiv:2606.24963v1 Announce Type: new Abstract: Fine-tuning Multimodal Large Language Models (MLLMs) on specialized tasks often leads to catastrophic forgetting of their general capabilities. Existing

model-releasesarxiv-cs-cv
25 Jun 2026
Research

CustomX: Unified Character, Action, and Scene Customization in Video World Models

DGX agent

arXiv:2512.17796v2 Announce Type: replace Abstract: Recent advances in world models have greatly enhanced interactive environment simulation. Existing methods mainly fall into two categories: (1) stat

researcharxiv-cs-cv
25 Jun 2026
Local Ai

Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability

DGX agent

arXiv:2512.05394v2 Announce Type: replace Abstract: Latent diffusion models pair VAEs with diffusion backbones, and the structure of VAE latents strongly influences the difficulty of diffusion trainin

local-aiarxiv-cs-cv
25 Jun 2026
← Previous
1…8889909192…263
Next →