AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Model Releases

MoCRA: Mixture of Compositional Rank-1 Atoms for 4K All-in-One Video Restoration

DGX agent

arXiv:2608.01829v1 Announce Type: new Abstract: Real-world video arrives hazy, rainy, dark, or noisy, and a deployable restorer faces three demands at once: no degradation label, native 4K output, and

model-releasesarxiv-cs-cv
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

DGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

MonitorVLM-v2: A Deployed Vision-Language Framework for Real-Time Safety Violation Detection

DGX agent

arXiv:2608.00975v1 Announce Type: new Abstract: Large vision--language models (VLMs) can reason step by step about complex visual scenes, but this open-ended, autoregressive chain-of-thought (CoT) app

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

DGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

DGX agent

arXiv:2608.01628v1 Announce Type: new Abstract: Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural corr

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

DGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

DGX agent

arXiv:2603.14686v2 Announce Type: replace Abstract: Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preservi

safetyarxiv-cs-cv
4 Aug 2026
Research

Neural Born Series Operator for Biomedical Ultrasound Computed Tomography

DGX agent

arXiv:2312.15575v2 Announce Type: replace-cross Abstract: Ultrasound Computed Tomography (USCT) provides a radiation-free option for high-resolution clinical imaging. Despite its potential, the comput

researcharxiv-cs-cv
4 Aug 2026
Model Releases

New York Smells: A Large Multimodal Dataset for Olfaction

DGX agent

arXiv:2511.20544v2 Announce Type: replace Abstract: While olfaction is central to how animals perceive the world, this rich chemical sensory modality remains largely inaccessible to machines. One key

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

NISF++: Geometrically-grounded implicit representations of 3D+time cardiac function from 2D short- and long-axis MR views

DGX agent

arXiv:2608.00752v1 Announce Type: new Abstract: Clinical acquisition in cardiac magnetic resonance (CMR) imaging involves obtaining cross-sectional planes of the heart along the radial and longitudina

safetyarxiv-cs-cv
4 Aug 2026
Tutorials

Noise-Robust Conditional Flow Matching: Generating Clean Samples from Noisy Datasets

DGX agent

arXiv:2608.00064v1 Announce Type: new Abstract: Generative models learn the statistical properties of their training data, so high-quality generation depends on clean and representative datasets. In s

tutorialsarxiv-cs-cv
4 Aug 2026
Model Releases

On the Viability of Semi-Supervised Segmentation Methods for Statistical Shape Modeling

DGX agent

arXiv:2407.15260v3 Announce Type: replace Abstract: Statistical Shape Models (SSMs) excel at identifying population level anatomical variations, which is at the core of various clinical and biomedical

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Onboard Satellite Image Classification for Earth Observation: A Comparative Study of ViT Models

DGX agent

arXiv:2409.03901v4 Announce Type: replace Abstract: Remote sensing (RS) image classification is central to Earth observation, but onboard deployment requires models that are accurate, efficient, and r

researcharxiv-cs-cv
4 Aug 2026
Model Releases

One Query, Many Scales: Sparse Mixture-of-Experts for Efficient Hierarchical Cross-View Geo-Localization

DGX agent

arXiv:2608.01060v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) retrieves geo-tagged satellite imagery for a ground-view query. Most systems exhaustively search a flat, fixed-resolu

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

One-Sided Quantile Coupling for Flow Matching

DGX agent

arXiv:2608.00978v1 Announce Type: cross Abstract: Flow Matching trains continuous-time generative models by regressing the velocity field of a probability path between a simple source distribution and

safetyarxiv-cs-cv
4 Aug 2026
Local Ai

Open-Set Visual Text Forensics via Sparse-Constraint Rectified Flow

DGX agent

arXiv:2608.02258v1 Announce Type: new Abstract: Rapidly evolving Generative AI enables sophisticated visual text manipulations that increasingly evade current forensic detectors. Existing discriminati

local-aiarxiv-cs-cv
4 Aug 2026
Applications

Optical Flow from Photons

DGX agent

arXiv:2608.00499v1 Announce Type: new Abstract: Optical flow remains challenging in high-speed and low-light scenes, where the limited frame rate and sensitivity of conventional cameras lead to motion

applicationsarxiv-cs-cv
4 Aug 2026
Model Releases

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression

DGX agent

arXiv:2608.00345v1 Announce Type: new Abstract: A 3D CT scan entering a vision-language model produces a long sequence of visual tokens, often thousands to tens of thousands per volume, and this seque

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

ORCESTRA: VLM-driven Visual Robot programming in Mixed Reality

DGX agent

arXiv:2608.00775v1 Announce Type: cross Abstract: ORCESTRA is a mixed-reality system for programming robot digital twins through no-code waypoint teaching and language-guided control. In a passthrough

safetyarxiv-cs-cv
4 Aug 2026
Safety

OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs

DGX agent

arXiv:2603.11804v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) adapted to remote sensing rely heavily on domain-specific image-text supervision, yet high-quality annotations for sat

safetyarxiv-cs-cv
4 Aug 2026
Research

OSSDD - a New Open Dataset for Sentinel-1 Ship Detection

DGX agent

arXiv:2608.01963v1 Announce Type: new Abstract: Ship detection in Synthetic Aperture Radar (SAR) images plays an important role for maritime situational awareness, especially with respect to different

researcharxiv-cs-cv
4 Aug 2026
Model Releases

PackingGPT: 3D Packing Agent for Real Furniture in Last-Mile Delivery

DGX agent

arXiv:2608.01427v1 Announce Type: new Abstract: 3D bin packing rectangular items into standardised containers to maximise space utilisation under geometric shipping automation. Loading a furniture pur

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Parameter-Dynamic Adaptive Fusion and Calibration Network for RGBT Tracking

DGX agent

arXiv:2608.01807v1 Announce Type: new Abstract: Existing RGBT trackers typically employ fusion functions with fixed parameters across different targets and scenarios. Although dynamic-architecture met

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Parameter-Efficient CLIP Adaptation for 3D Understanding via Unified Tokenization

DGX agent

arXiv:2505.18819v2 Announce Type: replace Abstract: Vision-language models, such as CLIP, encode rich semantic knowledge through large-scale image-text pretraining. Reusing these models for 3D underst

model-releasesarxiv-cs-cv
4 Aug 2026
Hardware

Partial FC: Training 10 Million Identities on a Single Machine

DGX agent

arXiv:2010.05222v3 Announce Type: replace Abstract: Training face recognition models with millions of identities is challenging because classifier storage, logit memory, and computation grow linearly

hardwarearxiv-cs-cv
4 Aug 2026
Tutorials

PartMat: Material-Aware 3D Part Decomposition with a Single Global Latent

DGX agent

arXiv:2608.01825v1 Announce Type: new Abstract: Part-level 3D generation has recently attracted increasing attention for producing structured and editable 3D assets. However, existing methods typicall

tutorialsarxiv-cs-cv
4 Aug 2026
Local Ai

PatchAlign3D: Local Feature Alignment for Dense 3D Shape Understanding

DGX agent

arXiv:2601.02457v2 Announce Type: replace Abstract: Current foundation models for 3D shapes excel at global tasks (retrieval, classification) but transfer poorly to local part-level reasoning. Recent

local-aiarxiv-cs-cv
4 Aug 2026
Applications

PeCA: Palette Context Assisted Inference for Test-Time Paint-Bucket Colourisation on Animation Videos

DGX agent

arXiv:2608.00903v1 Announce Type: new Abstract: In animation production, paint-bucket colourisation for hand-drawn animation is a labour-intensive procedure that assigns each enclosed region in line s

applicationsarxiv-cs-cv
4 Aug 2026
Research

PhenoStitch: Training-Free Panoptic Crop Mapping from Satellite Image Time Series

DGX agent

arXiv:2608.00870v1 Announce Type: new Abstract: Panoptic crop mapping requires both delineating individual agricultural parcels and assigning a crop type to each parcel from satellite image time serie

researcharxiv-cs-cv
4 Aug 2026
Applications

PhotoHOI: Synthesizing 3D Hand-Object Interactions from a Single RGB Photograph

DGX agent

arXiv:2608.01905v1 Announce Type: new Abstract: Hand-object interaction (HOI) is a fundamental human behavior with broad applications in AR/VR, digital humans, and embodied interaction. Existing metho

applicationsarxiv-cs-cv
4 Aug 2026
Research

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs

DGX agent

arXiv:2608.02150v1 Announce Type: new Abstract: Embodied intelligence and world models require video understanding systems to go beyond recognizing objects and actions and develop an understanding of

researcharxiv-cs-cv
4 Aug 2026
Model Releases

PhysAgent: A Multi-Agent Framework for Reliable Remote Heart Rate Estimation

DGX agent

arXiv:2608.00066v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact heart-rate estimation from facial videos, but its weak physiological signal is easily corrupted b

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Pixel Ignores, Superpixel Sees: Adverse Weather Image Restoration via Semantic-Center SSM

DGX agent

arXiv:2608.01760v1 Announce Type: new Abstract: Adverse weather image restoration aims to recover clear visibility from degraded images in complex weather conditions. Existing works attempt to address

researcharxiv-cs-cv
4 Aug 2026
Local Ai

PixelSR: Efficient Screen Content Super-Resolution via Pixel Classification

DGX agent

arXiv:2608.00646v1 Announce Type: new Abstract: Screen content images are generally composed of texts and graphics. Compared to natural images, these man-made images contain a large quantity of sharp

local-aiarxiv-cs-cv
4 Aug 2026
Tutorials

PixVL: Self-Supervised Training of Pixel-Level MLLMs via a Unified Mask--Text Consistency Cycle

DGX agent

arXiv:2608.01354v1 Announce Type: new Abstract: Recent studies develop pixel-level multimodal large language models (MLLMs) that support both Region Segmentation and Region Understanding, extending mu

tutorialsarxiv-cs-cv
4 Aug 2026
Safety

PlantRig - From Bones to Branches: Adaptation of Autoregressive Rigging Models for Plant Skeletal Reconstruction

DGX agent

arXiv:2608.01072v1 Announce Type: new Abstract: Autoregressive rigging models such as UniRig and SkinTokens perform well on articulated characters, but their ability to generalize to plant structures

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

PNEC-Mamba: Prototype-Guided Positive-Negative Evidence Calibration for Hyperspectral Image Classification

DGX agent

arXiv:2608.01910v1 Announce Type: new Abstract: In real-world hyperspectral scenes, pixel representations are often ambiguous due to factors such as spectral similarity, mixed pixels, and local contex

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis

DGX agent

arXiv:2608.00440v1 Announce Type: new Abstract: Recent image generators can synthesize convincing human-centric images, yet producing a useful collection remains different from producing a single succ

researcharxiv-cs-cv
4 Aug 2026
Safety

Practical Noise Modeling for SPAD Intensity Imaging

DGX agent

arXiv:2608.00489v1 Announce Type: new Abstract: Single-photon avalanche diode (SPAD) cameras are promising for low-light and high-dynamic-range intensity imaging, but their practical use is limited by

safetyarxiv-cs-cv
4 Aug 2026
Agents

PRISM: Privileged Probabilistic Latent Supervision for End-to-End Autonomous Driving Motion Planning

DGX agent

arXiv:2608.01201v1 Announce Type: cross Abstract: End-to-end autonomous driving (E2E AD) systems integrate perception, prediction, and planning into a single differentiable architecture. While these m

agentsarxiv-cs-cv
4 Aug 2026
Tutorials

Probing the 3D Object-Level Understanding of Pre-Trained Detection Transformers

DGX agent

arXiv:2608.01495v1 Announce Type: new Abstract: Detection transformer models, including DETR and its extensions, learn to output a set of object-level embeddings that can be simultaneously decoded int

tutorialsarxiv-cs-cv
4 Aug 2026
Research

Prompt-Driven Simulation with Feature Perturbation for Cross-Domain Few-Shot Object Detection

DGX agent

arXiv:2608.01348v1 Announce Type: new Abstract: Data augmentation, which simulates diverse visual variations to expand the source distribution and induce synthetic domain shifts, is a simple yet effec

researcharxiv-cs-cv
4 Aug 2026
Research

PromptPath: Prompt-Adaptive Computational Pathways for In-Context Learning

DGX agent

arXiv:2608.02129v1 Announce Type: new Abstract: In-context learning (ICL) has attracted increasing attention for enabling models to perform new tasks using only a few ``input--output'' prompt examples

researcharxiv-cs-cv
4 Aug 2026
Safety

Proteus: A Truncation-Robust Entropy Model for Progressive LiDAR Compression

DGX agent

arXiv:2608.00687v1 Announce Type: new Abstract: LiDAR point clouds provide explicit, deterministic physical boundaries critical for collaborative safety-critical perception. However, wireless channels

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

Protocol generalisation for brain tissue microstructure estimation via hypernetwork-controlled geometric deep learning

DGX agent

arXiv:2608.02053v1 Announce Type: cross Abstract: Brain tissue microstructure estimation with machine learning provides higher computational efficiency than conventional fitting. However, machine lear

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation

DGX agent

arXiv:2608.01978v1 Announce Type: new Abstract: Audio-driven portrait animation has advanced rapidly with diffusion-based generative models, yet real-time one-shot generation with expressive emotion c

researcharxiv-cs-cv
4 Aug 2026
Research

Quaternion Tensor Modeling for Joint Color-Polarization Demosaicking

DGX agent

arXiv:2608.02144v1 Announce Type: new Abstract: Division-of-focal-plane (DoFP) color polarization cameras enable snapshot acquisition of color polarization mosaic images, but the inherently sparse sam

researcharxiv-cs-cv
4 Aug 2026
Model Releases

QuerySplat: Decoupling Geometry and Appearance Representations in 3DGS Prediction

DGX agent

arXiv:2608.01186v1 Announce Type: new Abstract: While feed-forward 3D Gaussian Splatting (3DGS) enables efficient 3D reconstruction, achieving high-fidelity rendering remains challenging. Existing pix

model-releasesarxiv-cs-cv
4 Aug 2026
← Previous
1…2324252627…261
Next →