AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,414 results
Model Releases

JUMP-lite: Compact, reproducible benchmarking of cell representations

DGX agent

arXiv:2608.07632v1 Announce Type: cross Abstract: Image-based profiling captures rich phenotypic signatures for drug discovery and functional genomics. Large public datasets like JUMP Cell Painting no

model-releasesarxiv-cs-cv
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

LAD-COD: Language-Aligned Dense Perception for Camouflaged Object Detection

DGX agent

arXiv:2608.07941v1 Announce Type: new Abstract: Camouflaged object detection (COD) aims to segment objects that exhibit high visual similarity to their surroundings, which reduces foreground-backgroun

tutorialsarxiv-cs-cv
11 Aug 2026
Safety

LASA: Language-and-Source-Anchored Alignment for Domain Generalized Semantic Segmentation

DGX agent

arXiv:2608.08805v1 Announce Type: new Abstract: Domain Generalization Semantic Segmentation (DGSS) focuses on generalizing knowledge from labeled source domains to unseen target domains where data is

safetyarxiv-cs-cv
11 Aug 2026
Agents

LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents

DGX agent

arXiv:2608.07585v1 Announce Type: new Abstract: Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video too

agentsarxiv-cs-cv
11 Aug 2026
Safety

Learning Deep Modality-Shared Self-Expressiveness for Image Clustering with Textual Information

DGX agent

arXiv:2608.08418v1 Announce Type: new Abstract: Leveraging textual information for image clustering has emerged as a promising direction, largely owing to the powerful representations learned by Visio

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

Learning How the World Evolves: Extrapolative Video World Models via Latent Dynamics Reasoning

DGX agent

arXiv:2608.09926v1 Announce Type: new Abstract: The world evolves following its dynamics, i.e., its laws of motion. However, leading video diffusion models largely fit the pixels without modeling how

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Learning human joint torques from pixels

DGX agent

arXiv:2608.09083v1 Announce Type: new Abstract: Estimating human joint torques from visual observations is a key step toward bringing biomechanical analysis from controlled laboratories to real-world

model-releasesarxiv-cs-cv
11 Aug 2026
Research

Learning Physical Interaction: A Survey of Tactile- and Force-aware Robot Learning

DGX agent

arXiv:2608.07558v1 Announce Type: cross Abstract: Physically grounded robot intelligence requires robots to perceive, reason about, and regulate their interactions with the physical world. This capabi

researcharxiv-cs-cv
11 Aug 2026
Model Releases

Learning Structural Illumination for Unsupervised Low-light Enhancement

DGX agent

arXiv:2608.08153v1 Announce Type: new Abstract: Existing unsupervised low-light image enhancement (LLIE) methods often estimate illumination directly from the entire low-light input, without separatin

model-releasesarxiv-cs-cv
11 Aug 2026
Tutorials

Let Geometry GUIDE: Layer-wise Unrolling of Geometric Priors in Multimodal LLMs

DGX agent

arXiv:2604.05695v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in 2D visual tasks but still struggle to understand physical space in rea

tutorialsarxiv-cs-cv
11 Aug 2026
Research

LHSDet: High-Resolution AI-Generated Image Detection via Visual Question Answering

DGX agent

arXiv:2608.07863v1 Announce Type: new Abstract: Driven by advances in diffusion models and autoregressive models, the fidelity and resolution of AI-generated images now rival those of real images. How

researcharxiv-cs-cv
11 Aug 2026
Model Releases

LIBAD: A Multimodal Anomaly Detection Benchmark for Li-Ion Battery Electrode Manufacturing

DGX agent

arXiv:2608.07958v1 Announce Type: new Abstract: Multimodal industrial anomaly detection has largely focused on discrete products using strongly correlated RGB and 3D observations, leaving continuous p

model-releasesarxiv-cs-cv
11 Aug 2026
Research

LightAIR: Lightweight Action Inversion and Riemannian Rectification for Text-based Person Anomaly Search

DGX agent

arXiv:2608.09152v1 Announce Type: new Abstract: Traditional Text-based Person Search (TPS) is typically limited to matching static appearance attributes, severely neglecting dynamic action information

researcharxiv-cs-cv
11 Aug 2026
Applications

LineGraph2Road: Structural Graph Reasoning on Line Graphs for Road Network Extraction

DGX agent

arXiv:2602.23290v2 Announce Type: replace Abstract: Extracting routable road networks from satellite imagery requires accurate topology recovery beyond pixel-level segmentation. Recent methods decompo

applicationsarxiv-cs-cv
11 Aug 2026
Safety

Linguistically-Aligned and Visually-Grounded Preference Optimization for Clinically-Augmented Medical Report Generation

DGX agent

arXiv:2608.08494v1 Announce Type: new Abstract: Despite significant advances in Medical Report Generation (MRG), the reliability remains constrained by the prevalence of factual errors. While Direct P

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding

DGX agent

arXiv:2608.07596v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models transform representations from pretrained vision-language models (VLMs) into robot actions, yet the interface that

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

LogiShot: Logically Coherent Cross-Shot Video Generation

DGX agent

arXiv:2608.08820v1 Announce Type: new Abstract: Generating cross-shot videos that are logically connected is essential for content creation. Currently, most cross-shot video-generation workflows, such

model-releasesarxiv-cs-cv
11 Aug 2026
Research

LookAgain: Closed-Loop GUI Grounding with Visually Grounded Reflection

DGX agent

arXiv:2608.09723v1 Announce Type: new Abstract: Recent graphical user interface (GUI) grounders have significantly advanced single-shot accuracy on standard benchmarks, yet their performance degrades

researcharxiv-cs-cv
11 Aug 2026
Research

LoRA-based Adaptation Alone Is Not Enough: Understanding the Limits of Foundation Models for Face Presentation Attack Detection

DGX agent

arXiv:2608.09633v1 Announce Type: new Abstract: Face presentation attack detection (PAD) aims to reliably detect a wide range of presentation attacks. While PAD methods achieve strong performance with

researcharxiv-cs-cv
11 Aug 2026
Research

Lying mirror using structured surfaces

DGX agent

arXiv:2410.15521v2 Announce Type: replace-cross Abstract: We introduce an all-optical system, termed the 'lying mirror', to hide input information by transforming it into misleading, ordinary-looking

researcharxiv-cs-cv
11 Aug 2026
Safety

MAGIC-SSCIL: Manifold Anchoring and Geometric Incremental Calibration for Semi-Supervised Class Incremental Learning

DGX agent

arXiv:2608.07586v1 Announce Type: new Abstract: Semi-supervised Class Incremental Learning (SSCIL) is a severe challenge for neural networks, and it is hardest in the exemplar-free setting where no pa

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

Marrying Optimal Transport and ODEs for Unified Continuous-Time 4D Reconstruction and Tracking

DGX agent

arXiv:2608.09613v1 Announce Type: new Abstract: Existing unified 4D reconstruction and point tracking approaches typically rely on heuristic interpolations or just predict at integer timestamps, lacki

model-releasesarxiv-cs-cv
11 Aug 2026
Applications

Mask-aware inference with State-Space Models

DGX agent

arXiv:2603.04568v2 Announce Type: replace Abstract: Many real-world computer vision tasks, such as depth completion, must handle inputs with arbitrarily shaped regions of missing or invalid data. For

applicationsarxiv-cs-cv
11 Aug 2026
Applications

MeanSR: Restoration Trajectory Learning for One-Step Perceptual Super-Resolution

DGX agent

arXiv:2608.09405v1 Announce Type: new Abstract: Diffusion-based super-resolution (SR) achieves strong perceptual quality but requires costly iterative denoising. Existing one-step distillation methods

applicationsarxiv-cs-cv
11 Aug 2026
Model Releases

Mechanistic Interpretability-Guided Selective Fine-Tuning of Vision-Language Models for Centimeter-Level Flood Depth Estimation

DGX agent

arXiv:2608.07562v1 Announce Type: new Abstract: Urban flooding poses an escalating threat to transportation infrastructure, yet no operational system provides real-time, street-level flood-depth estim

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

DGX agent

arXiv:2608.09316v1 Announce Type: new Abstract: Offline context optimization improves an agent by revising its instructions and examples while keeping the model frozen. This approach learns from rollo

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Model the Edit, Not the Image: Visual Autoregressive Editing from a Source-Centric Perspective

DGX agent

arXiv:2608.09057v1 Announce Type: new Abstract: Next-scale visual autoregressive models (VARs) have emerged as a powerful generative paradigm, producing high-quality images through efficient coarse-to

safetyarxiv-cs-cv
11 Aug 2026
Applications

Motion-Aware Animatable Gaussian Avatars Deblurring

DGX agent

arXiv:2411.16758v4 Announce Type: replace Abstract: The creation of 3D human avatars from multi-view videos is a significant yet challenging task in computer vision. However, existing techniques rely

applicationsarxiv-cs-cv
11 Aug 2026
Safety

MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling

DGX agent

arXiv:2608.08553v1 Announce Type: new Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

MPISuperRes-PnP: A Super-Resolution Zero-Shot Plug-and-Play Reconstruction Algorithm for Magnetic Particle Imaging

DGX agent

arXiv:2608.09672v1 Announce Type: new Abstract: Magnetic Particle Imaging (MPI) is an emerging medical imaging modality. MPI is based on the non-linear response of magnetic nanoparticles to an applied

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MRBench: A Comprehensive Benchmark for Human Motion-Text Retrieval

DGX agent

arXiv:2608.07993v1 Announce Type: new Abstract: Human motion-text retrieval provides a rigorous means of assessing cross-modal alignment. Prevailing benchmarks are dominated by homogeneous indoor moti

model-releasesarxiv-cs-cv
11 Aug 2026
Research

MRI super-resolution in ten sampling steps using a diffusion bridge model

DGX agent

arXiv:2608.08819v1 Announce Type: new Abstract: Objective. MRI provides excellent soft-tissue contrast, but long acquisition times can cause patient discomfort and lead to motion artifacts, forcing a

researcharxiv-cs-cv
11 Aug 2026
Model Releases

MSP-Net: Manifold-Guided Spectral Prompt Network for Hyperspectral Object Tracking

DGX agent

arXiv:2608.09575v1 Announce Type: new Abstract: Hyperspectral object tracking leverages abundant spectral information to provide unique advantages for target discrimination in complex scenes. However,

model-releasesarxiv-cs-cv
11 Aug 2026
Tutorials

Multi-Relational Knowledge Graph Enhanced Embedding for Trajectory-User Linking

DGX agent

arXiv:2608.08646v1 Announce Type: cross Abstract: Trajectory-User Linking (TUL) aims to identify the owner of an anonymous trajectory from a set of candidate users, providing a basis for user mobility

tutorialsarxiv-cs-cv
11 Aug 2026
Local Ai

Multi-Submap Implicit Neural SLAM with Local-to-Global Loop Closure for Large-Scale Scene Reconstruction

DGX agent

arXiv:2608.09146v1 Announce Type: new Abstract: Neural Radiance Fields (NeRF)-based SLAM has demonstrated impressive results in small-scale scene reconstruction, yet scaling these methods to extensive

local-aiarxiv-cs-cv
11 Aug 2026
Research

Multimodal Skin Lesion Classification with Swin Transformer and Clinical Metadata Fusion

DGX agent

arXiv:2608.07574v1 Announce Type: new Abstract: Skin lesion classification plays an important role in supporting the early diagnosis of skin cancer. However, automated analysis remains challenging due

researcharxiv-cs-cv
11 Aug 2026
Safety

MultiShadow: Multi-Object Shadow Generation for Image Compositing via Diffusion Model

DGX agent

arXiv:2603.02743v4 Announce Type: replace Abstract: Realistic shadow generation is crucial for achieving seamless image compositing, yet existing methods primarily focus on single-object insertion and

safetyarxiv-cs-cv
11 Aug 2026
Research

MVMD: A Multi-View Approach for Enhanced Mirror Detection

DGX agent

arXiv:2608.07559v1 Announce Type: new Abstract: In 3D reconstruction, mirrors introduce significant challenges by creating distorted and fragmented spaces, resulting in inaccurate and unreliable 3D mo

researcharxiv-cs-cv
11 Aug 2026
Model Releases

NBA_Streaming: A Large-Scale Benchmark for Fine-Grained Basketball Commentary Generation in Continuous Streams

DGX agent

arXiv:2608.09200v1 Announce Type: new Abstract: Live basketball commentary generation requires determining when an event is sufficiently observable and describing it before subsequent events unfold. H

model-releasesarxiv-cs-cv
11 Aug 2026
Research

NeuralDMD: Interpretable Neural Representation of Dynamics from Sparse and Noisy Measurements

DGX agent

arXiv:2507.03094v2 Announce Type: replace Abstract: Many challenges in scientific imaging involve solving ill-posed inverse problems, where the goal is to recover spatio-temporal fields from indirect,

researcharxiv-cs-cv
11 Aug 2026
Model Releases

NeuroGuard: Neural Gradient Update Aware of Representation Damage

DGX agent

arXiv:2608.08068v1 Announce Type: new Abstract: Long-tailed class-incremental learning (LT-CIL) must learn new classes from imbalanced streams while retaining old classes. Existing methods mainly chan

model-releasesarxiv-cs-cv
11 Aug 2026
Research

NewtonGS: Physics-Structured Object-Level Neural Newtonian Dynamics for Gaussian Scene Animation

DGX agent

arXiv:2608.07598v1 Announce Type: new Abstract: Animating objects in a static 3D Gaussian scene requires an explicit object-level dynamic state and a controllable model of object motion. Existing dyna

researcharxiv-cs-cv
11 Aug 2026
Research

Not All Frames Deserve Full Computation: Accelerating Autoregressive Video Generation via Selective Computation and Predictive Extrapolation

DGX agent

arXiv:2604.02979v2 Announce Type: replace Abstract: Autoregressive (AR) video diffusion models enable long-form video generation but remain expensive due to repeated multi-step denoising. Existing tra

researcharxiv-cs-cv
11 Aug 2026
Applications

NTIRE 2026 Low-light Enhancement: Twilight Cowboy Challenge

DGX agent

arXiv:2608.09782v1 Announce Type: new Abstract: This paper presents a review of the NTIRE 2026 Low-light Enhancement: Twilight Cowboy Challenge. The objective of the competition was to merge a set of

applicationsarxiv-cs-cv
11 Aug 2026
Research

OccAnyScene: Towards Unified Indoor-Outdoor 3D Occupancy Predictio

DGX agent

arXiv:2608.08696v1 Announce Type: new Abstract: 3D occupancy prediction is fundamental to scene understanding, yet existing 3D semantic occupancy methods are typically specialized to fixed scene types

researcharxiv-cs-cv
11 Aug 2026
Model Releases

OGG-FR: Orthogonal Gradient Gaming and Frequency Rectification for Unmanned Aerial Vehicle Infrared Image Super-Resolution

DGX agent

arXiv:2608.09150v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) infrared image super-resolution aims to recover weak thermal structures for deployment on resource-constrained platforms;

model-releasesarxiv-cs-cv
11 Aug 2026
Research

One Model to Magnify Them All: Efficient Scale-Invariant Histopathology via Conditional Normalization and Continuous Magnification Training

DGX agent

arXiv:2608.09403v1 Announce Type: new Abstract: Whole slide images (WSIs) in digital histopathology are acquired at discrete magnification levels encoding complementary diagnostic information from glo

researcharxiv-cs-cv
11 Aug 2026
Local Ai

One-Time Training for All Grains: Open-Set Grain Recognition and Quantitative Analysis

DGX agent

arXiv:2608.09345v1 Announce Type: new Abstract: Advances in crop breeding have introduced an increasing number of grain varieties, creating a growing demand for efficient variety recognition and quant

local-aiarxiv-cs-cv
11 Aug 2026
← Previous
1…45678…259
Next →