AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators

DGX agent

arXiv:2506.19348v2 Announce Type: replace Abstract: Training-free motion customization imposes motion patterns from reference videos onto video generators through test-time computation. Most existing

researcharxiv-cs-cv
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Why Fake ? Unveiling the Semantic Vocabulary of Deepfake Detectors

DGX agent

arXiv:2607.07216v1 Announce Type: new Abstract: Deepfake (DF) technology poses a significant threat to information integrity, driving the need for robust detection methods. Most DF detectors only cons

local-aiarxiv-cs-cv
9 Jul 2026
Research

Widest-Path Reachability Fields for Connectivity-Preserving Slender Structure Segmentation

DGX agent

arXiv:2607.07123v1 Announce Type: new Abstract: Segmenting slender curvilinear structures such as retinal vessels, cracks, and roads demands topological correctness, as even a single-pixel discontinui

researcharxiv-cs-cv
9 Jul 2026
Agents

WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence

DGX agent

arXiv:2607.06838v1 Announce Type: new Abstract: Humans can navigate an unfamiliar city and gradually form a coherent spatial mental map spanning tens of square kilometers. Can AI build spatial represe

agentsarxiv-cs-cv
9 Jul 2026
Safety

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer

DGX agent

arXiv:2510.24108v2 Announce Type: replace-cross Abstract: Human demonstrations are widely considered the cornerstone of end-to-end (E2E) autonomous driving despite human demonstration's scarcity for l

safetyarxiv-cs-cv
9 Jul 2026
Research

A Task-Driven Evaluation of UAV Detection and Tracking under Synthetic Fog

DGX agent

arXiv:2607.05467v1 Announce Type: new Abstract: Fog severely degrades the visibility of small unmanned aerial vehicles (UAVs) in skydominant, long-range imagery, reducing the reliability of downstream

researcharxiv-cs-cv
8 Jul 2026
Model Releases

A VLM-Enhanced Framework for Comprehensive Traffic Sign Condition Assessment Integrating Daytime Visual Performance and Nighttime Retroreflectivity Evaluation

DGX agent

arXiv:2607.06478v1 Announce Type: new Abstract: Traffic signs are crucial components of road safety, serving as visual tools under all lighting conditions. The Manual on Uniform Traffic Control Device

model-releasesarxiv-cs-cv
8 Jul 2026
Applications

Abductive Corroboration of Probabilistic AI Models for Forensic Synthetic Media Detection

DGX agent

arXiv:2607.05434v1 Announce Type: cross Abstract: Artificial Intelligence (AI) models, at their core, apply general learnings from broad datasets to individual circumstances using probabilistic behavi

applicationsarxiv-cs-cv
8 Jul 2026
Local Ai

AEGIS: A Mechanism-Guided Defense against Visual Synonym Jailbreaks in Text-to-Image Models

DGX agent

arXiv:2607.06120v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved high visual fidelity and broad adoption, but remain vulnerable to safety violations when adversaries exploi

local-aiarxiv-cs-cv
8 Jul 2026
Applications

AlayaWorld: Long-Horizon and Playable Video World Generation

DGX agent

arXiv:2607.06291v1 Announce Type: new Abstract: Game worlds have traditionally been built through labor-intensive production pipelines, making them costly to develop, difficult to customization, and e

applicationsarxiv-cs-cv
8 Jul 2026
Research

Andha-Dhun: A First Look at Audio Descriptions in Hindi

DGX agent

arXiv:2607.06457v1 Announce Type: new Abstract: Audio Descriptions (ADs) narrate visual content for Blind and Low Vision (BLV) audiences during gaps in audiovisual media. There is growing momentum aro

researcharxiv-cs-cv
8 Jul 2026
Research

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation

DGX agent

arXiv:2607.01499v2 Announce Type: replace Abstract: Recent advances in Image-to-Video generation allow a single image to be animated into a convincing video under text guidance, raising serious copyri

researcharxiv-cs-cv
8 Jul 2026
Safety

ARMS: Anchor-Relational Motion Streaming for Seamless Solo-Social Motion Transitions

DGX agent

arXiv:2607.05733v1 Announce Type: new Abstract: Generating temporally continuous and socially coherent human motion from text remains a fundamental challenge, particularly in realistic streams where p

safetyarxiv-cs-cv
8 Jul 2026
Agents

Assessing the Operational Impact of Poisoning Attacks over Augmented 3D Point Cloud Public Datasets for Connected and Autonomous Vehicles

DGX agent

arXiv:2607.06484v1 Announce Type: cross Abstract: Poisoning attacks against public datasets lead to major concerns, such as (i) misclassification of perceived objects when the poisoned data is used fo

agentsarxiv-cs-cv
8 Jul 2026
Research

Association Restoration Test: Revealing Restorable Shortcuts after Unlearning

DGX agent

arXiv:2607.05726v1 Announce Type: new Abstract: Association unlearning aims to disable learned label-attribute shortcuts while preserving task performance. Existing evaluations mainly measure output-l

researcharxiv-cs-cv
8 Jul 2026
Local Ai

AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring

DGX agent

arXiv:2607.05859v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are promising for construction-site monitoring, and recent construction-tailored VLMs have primarily adapted pretrained VL

local-aiarxiv-cs-cv
8 Jul 2026
Model Releases

Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective

DGX agent

arXiv:2607.05783v1 Announce Type: new Abstract: Environmental illusions (eg., shadows, reflections, and tire marks) are naturally existing yet overlooked phenomena in real-world driving environments.

model-releasesarxiv-cs-cv
8 Jul 2026
Research

BitFair: A 12nm Bit-Serial CNN Accelerator with Learnable Early Termination and Adaptive Bit Ordering for Ultra-Low-Power XR Vision

DGX agent

arXiv:2607.05445v1 Announce Type: cross Abstract: Extended Reality (XR) wearables require always-on perception within tight power envelopes of a few watts and motion-to-photon latency budgets below 20

researcharxiv-cs-cv
8 Jul 2026
Tutorials

Blind Quality Enhancement of Compressed Video via Fine-Grained Degradation-Guided Sequential Inference

DGX agent

arXiv:2511.16137v2 Announce Type: replace Abstract: Existing studies on quality enhancement for compressed video (QECV) predominantly rely on known quantization parameters (QPs), training separate enh

tutorialsarxiv-cs-cv
8 Jul 2026
Tutorials

Breaking Spurious Correlations via Generative Randomization and Cross-Variant Self-Supervised Learning

DGX agent

arXiv:2607.05850v1 Announce Type: new Abstract: Deep neural networks trained with Empirical Risk Minimization (ERM) often fail under distribution shifts because they exploit spurious correlations betw

tutorialsarxiv-cs-cv
8 Jul 2026
Safety

Bridging Diffusion Pruning and Step Distillation with Teacher-Aligned Repair

DGX agent

arXiv:2607.06335v1 Announce Type: new Abstract: Diffusion models generate high-quality images, but their inference cost comes from two sources: large denoising networks and repeated denoising steps. E

safetyarxiv-cs-cv
8 Jul 2026
Model Releases

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models

DGX agent

arXiv:2607.06534v1 Announce Type: new Abstract: Existing 3D scene-grounded Large Language Models (3D-LLMs) focus on answering questions grounded in simplified single-room 3D scenes, lacking the abilit

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Clustered Codebook Quantization for 2D Gaussian-based Image Compression

DGX agent

arXiv:2607.05667v1 Announce Type: new Abstract: Gaussian-based image representations effectively model image content using compact parametric primitives while preserving high visual fidelity, yet stor

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Conformal Prediction Sets for Instance Segmentation

DGX agent

arXiv:2602.10045v2 Announce Type: replace Abstract: Current instance segmentation models achieve high performance on average predictions, but lack principled uncertainty quantification: their outputs

model-releasesarxiv-cs-cv
8 Jul 2026
Safety

Cross-Contextual Vision-Language Adaptation with LoRA for Personalized Severe Adverse Event Detection in Clinical Wound Monitoring

DGX agent

arXiv:2607.05625v1 Announce Type: new Abstract: Wound monitoring is a critical yet underserved clinical challenge, where timely identification of severe adverse events (SAEs) such as infection, tissue

safetyarxiv-cs-cv
8 Jul 2026
Safety

DeSeG: Decoupling Semantic Intent and Geometric Constraints for Physically Plausible Human-Scene Interaction

DGX agent

arXiv:2607.05787v1 Announce Type: new Abstract: Synthesizing physically plausible human-scene interactions (HSI) remains a critical challenge in computer vision and the development of human avatars. A

safetyarxiv-cs-cv
8 Jul 2026
Applications

EeveeDark: A Binary Neural Framework for Low-Light Video Enhancement via Event-Guided Sensor-Level Fusion

DGX agent

arXiv:2607.06217v1 Announce Type: new Abstract: Enhancing videos under extreme low-light conditions remains challenging due to the difficulty of balancing restoration quality and computational efficie

applicationsarxiv-cs-cv
8 Jul 2026
Model Releases

EgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera Footage

DGX agent

arXiv:2607.06468v1 Announce Type: new Abstract: We introduce EgoPolice, a carefully curated dataset of real, egocentric police-civilian interactions, sourced from publicly available body-worn camera v

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Enhanced Seam Segmentation for Automated Welding Robot in Construction Through Transfer Learning: Addressing Limitations of Bilateral Segmentation Network

DGX agent

arXiv:2607.06150v1 Announce Type: new Abstract: Reliable seam segmentation is essential for autonomous robotic welding in construction, where harsh illumination, specular reflections, and thin weld ge

model-releasesarxiv-cs-cv
8 Jul 2026
Safety

FADRA: Frequency-Aware Diffusion with Residual Adaptation for Video Face Restoration

DGX agent

arXiv:2607.06389v1 Announce Type: new Abstract: Video face restoration (VFR) aims to recover high-quality and temporally consistent facial details from severely degraded video sequences; however, exis

safetyarxiv-cs-cv
8 Jul 2026
Applications

FedDAF: Federated Domain Adaptation Using Model Functional Distance

DGX agent

arXiv:2509.11819v2 Announce Type: replace-cross Abstract: Federated Domain Adaptation (FDA) is a federated learning (FL) approach that improves model performance at the target client by collaborating

applicationsarxiv-cs-cv
8 Jul 2026
Tutorials

FGAA-FPN: Foreground-Guided Angle-Aware Feature Pyramid Network for Oriented Object Detection

DGX agent

arXiv:2602.10710v2 Announce Type: replace Abstract: With the increasing availability of high-resolution remote sensing and aerial imagery, oriented object detection has become a key capability for geo

tutorialsarxiv-cs-cv
8 Jul 2026
Research

FIELDS: Face reconstruction with accurate Inference of Expression using Learning with Direct Supervision

DGX agent

arXiv:2511.21245v3 Announce Type: replace Abstract: Monocular 3D face reconstruction estimates a 3D morphable model (3DMM) representation from a single image, providing geometry-aware expression codes

researcharxiv-cs-cv
8 Jul 2026
Model Releases

FourTune: Towards Fully 4-Bit Efficient Post-Training for Diffusion Models

DGX agent

arXiv:2607.05711v1 Announce Type: cross Abstract: Diffusion models have become a dominant paradigm for high-quality generative modeling, while post-training is essential for adapting them to diverse d

model-releasesarxiv-cs-cv
8 Jul 2026
Local Ai

Freqformer: Image-Demoireing Transformer via Effective Frequency Decomposition

DGX agent

arXiv:2505.19120v2 Announce Type: replace Abstract: Image demoireing remains a challenging task due to the complex interplay between texture corruption and color distortions caused by moire patterns.

local-aiarxiv-cs-cv
8 Jul 2026
Safety

From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications

DGX agent

arXiv:2308.16041v2 Announce Type: replace Abstract: Talking head generation has progressed rapidly from landmark- and GAN-based facial animation to diffusion models, neural rendering, 3D-aware avatars

safetyarxiv-cs-cv
8 Jul 2026
Research

From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models

DGX agent

arXiv:2607.06553v1 Announce Type: new Abstract: Large-scale text-to-image models are attractive backbones for dense prediction because RGB generation pretraining learns rich semantic, structural, and

researcharxiv-cs-cv
8 Jul 2026
Research

FUSE: A Flow-based Mapping Between Shapes

DGX agent

arXiv:2511.13431v2 Announce Type: replace Abstract: We introduce a novel neural representation for maps between 3D shapes based on flow-matching models, which is computationally efficient and supports

researcharxiv-cs-cv
8 Jul 2026
Safety

GaussFusion: Towards Multimodal 3D Gaussian Pretraining

DGX agent

arXiv:2607.05906v1 Announce Type: new Abstract: 3D Gaussian Splatting provides an explicit representation that jointly models geometry and appearance, serving as a scalable foundation for 3D represent

safetyarxiv-cs-cv
8 Jul 2026
Model Releases

GEM-Occ: From Visual Geometry Evidence to Embodied Semantic Occupancy Memory

DGX agent

arXiv:2607.05543v1 Announce Type: cross Abstract: Semantic occupancy provides a structured spatial memory for embodied indoor agents by jointly representing occupied regions, observed free space, unkn

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Generalized Synthetic Image Detection with Enhanced RGB-Noise Representation Learning

DGX agent

arXiv:2607.06354v1 Announce Type: new Abstract: The rapid advancement of large-scale generative models has accelerated the spread of highly deceptive AI-generated images, making generalized synthetic

model-releasesarxiv-cs-cv
8 Jul 2026
Safety

GraspIT: A Dataset Bridging the Sim-to-Real gap and back for Validated Grasping SE(3) Pose Generation

DGX agent

arXiv:2607.05869v1 Announce Type: cross Abstract: Robust robotic grasping of novel objects requires datasets that simultaneously provide photorealistic RGB-D observations, physically validated grasp q

safetyarxiv-cs-cv
8 Jul 2026
Applications

Ground3D-LMM: Fine-Grained 3D Point Grounding and Spatial Reasoning with LMM

DGX agent

arXiv:2607.05493v1 Announce Type: new Abstract: Natural-language queries about 3D environments become actionable when responses are verifiable and metric. Verifiability requires explicit grounding to

applicationsarxiv-cs-cv
8 Jul 2026
Local Ai

High-Resolution Artwork Outpainting with Global Blueprint Guidance and Layout Control

DGX agent

arXiv:2607.06162v1 Announce Type: new Abstract: Image outpainting extends an image beyond its original borders, requiring seamless style integration and globally coherent scene completion. Building on

local-aiarxiv-cs-cv
8 Jul 2026
Model Releases

HoloCount: A Holistic Visual Counting Benchmark for MLLMs

DGX agent

arXiv:2607.06420v1 Announce Type: new Abstract: Visual counting is a fundamental pillar of multimodal intelligence, requiring a seamless integration of fine-grained grounding and spatial reasoning. Wh

model-releasesarxiv-cs-cv
8 Jul 2026
Applications

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

DGX agent

arXiv:2607.05765v1 Announce Type: new Abstract: Embodied navigation aims to build agents that interpret multimodal goals, reason in 3D space, and reach target destinations reliably in the real world.

applicationsarxiv-cs-cv
8 Jul 2026
Local Ai

Imbalance-Robust and Sampling-Efficient Continuous Conditional GANs via Adaptive Vicinal Learning and Auxiliary Regularization

DGX agent

arXiv:2508.01725v5 Announce Type: replace-cross Abstract: Recent advances in continuous conditional generative modeling, including Continuous conditional Generative Adversarial Network (CcGAN) and Con

local-aiarxiv-cs-cv
8 Jul 2026
Safety

KOAL: Knowledge-Driven Prostate Cancer Grading with Ordinal-Aware Learning

DGX agent

arXiv:2607.06019v1 Announce Type: new Abstract: Non-invasive prediction of Gleason Grade Group (GGG) in prostate cancer using multiparametric MRI (mpMRI) is clinically vital for reducing unnecessary b

safetyarxiv-cs-cv
8 Jul 2026
← Previous
1…5758596061…261
Next →