AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Research

DeGS: A Scalable 3DGS Architecture via Decoupled Workload Parsing and Reorganization

DGX agent

arXiv:2608.02099v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading technique for real-time novel view synthesis, yet existing 3DGS accelerators suffer from poor ar

researcharxiv-cs-cv
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Deja Cue: Localizing States in Object Histories via Vocabulary-Relative Coordinates

DGX agent

arXiv:2608.02044v1 Announce Type: new Abstract: Tracking links observations of the same object through visual change, yet cannot by itself determine when the object is empty or filled, intact or cut.

researcharxiv-cs-cv
4 Aug 2026
Agents

DerainSplat: Feed-Forward Clean 3D Gaussian Splatting from Sparse Rainy Views

DGX agent

arXiv:2608.02191v1 Announce Type: new Abstract: Although image deraining has advanced substantially, existing methods mainly focus on 2D image restoration. As spatial intelligence applications such as

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

Detail Continuation over a Trustworthy Coarse Scale for Autoregressive Super-Resolution

DGX agent

arXiv:2608.01823v1 Announce Type: new Abstract: Hallucination remains a persistent challenge in generative super-resolution (GSR), where reconstructed results may contain visually plausible yet weakly

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

Device-First Feedback: Toward Mobile-Native LLM-Driven Neural Architecture Search

DGX agent

arXiv:2608.00078v1 Announce Type: new Abstract: Deploying convolutional neural networks generated by large language models (LLMs) on real mobile hardware requires more than GPU validation accuracy: IN

local-aiarxiv-cs-cv
4 Aug 2026
Research

DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation

DGX agent

arXiv:2608.01343v1 Announce Type: new Abstract: The emergence of transformer-based deep learning models has brought unprecedented performance across various domains, particularly in natural language p

researcharxiv-cs-cv
4 Aug 2026
Tutorials

DexMani: Human-Derived Manipulability Guidance for Dexterous Rotation

DGX agent

arXiv:2608.00554v1 Announce Type: cross Abstract: Dexterous object rotation is a sequential contact problem: each support, release, and re-contact decision must both produce the desired object motion,

tutorialsarxiv-cs-cv
4 Aug 2026
Agents

DF^3: World Modeling via Decoder-Free Feature Forecasting in Autonomous Navigation

DGX agent

arXiv:2608.02428v1 Announce Type: new Abstract: Forecasting future states from video sequences is a critical challenge for autonomous robotic systems and a fundamental objective of world modeling. Pri

agentsarxiv-cs-cv
4 Aug 2026
Research

Diagnosing Under-Development of Irreversible Processes in Video Generation

DGX agent

arXiv:2608.00617v1 Announce Type: new Abstract: Many physical attributes are irreversible: ice melts but does not re-freeze, paper chars but does not un-burn. Do video generators respect this? We show

researcharxiv-cs-cv
4 Aug 2026
Agents

DiffPhysCam: Differentiable Physics-Based Camera Simulation for Inverse Rendering and Embodied AI

DGX agent

arXiv:2508.08831v2 Announce Type: replace-cross Abstract: Generating synthetic images that closely mimic those from real cameras is instrumental in training visual models and enabling end-to-end visuo

agentsarxiv-cs-cv
4 Aug 2026
Tutorials

DiffPrune: differentiable information throttling for token pruning in vision-language models

DGX agent

arXiv:2608.01985v1 Announce Type: new Abstract: Visual token pruning reduces the computational cost of Vision-Language Models (VLMs) by removing redundant visual tokens. The key is to learn a score th

tutorialsarxiv-cs-cv
4 Aug 2026
Safety

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

DGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

safetyarxiv-cs-cv
4 Aug 2026
Agents

Direct and Adaptable Mesh-Gaussian Scene Reconstruction from Multi-View Images

DGX agent

arXiv:2405.06945v4 Announce Type: replace Abstract: Jointly recovering explicit surface geometry and high-quality appearance from multi-view images remains challenging. This capability is essential fo

agentsarxiv-cs-cv
4 Aug 2026
Research

Distill What RGB Can Recover: Privileged 3D Evidence for RGB-Only Vision-Language Models

DGX agent

arXiv:2608.00110v1 Announce Type: new Abstract: 3D scene understanding requires reasoning about entity existence, spatial layout, and object relations, yet RGB images alone often provide insufficient

researcharxiv-cs-cv
4 Aug 2026
Research

Distributional Matching for Vector Quantization: A Unified Theoretical and Empirical Framework

DGX agent

arXiv:2607.15933v2 Announce Type: replace Abstract: The effectiveness of modern visual representation learning and autoregressive models critically depends on vector quantization (VQ), which discretiz

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Do Maps Still Matter for Machines: Revisiting the Role of Choropleth Maps in Foundation Model Spatial Understanding

DGX agent

arXiv:2607.17999v2 Announce Type: replace-cross Abstract: Spatial understanding is crucial for foundation models (FMs), and maps have long helped humans organize and reason about geographic informatio

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards

DGX agent

arXiv:2608.00536v1 Announce Type: new Abstract: Reinforcement learning (RL) for document parsing often relies on reference-based rewards rooted in edit distance (e.g., tree edit distance), yet it rema

safetyarxiv-cs-cv
4 Aug 2026
Research

DODA: A Database of Datasets for Aesthetics Research

DGX agent

arXiv:2608.00089v1 Announce Type: new Abstract: With rapid growth in the fields of empirical and computational aesthetics we have seen a vast increase in large image datasets annotated for aesthetics.

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Does Explainability Transfer? A Controlled Benchmark of Attribution Methods on Vision Transformers and CNNs

DGX agent

arXiv:2608.02396v1 Announce Type: new Abstract: Most evidence on the effectiveness of explainable artificial intelligence (XAI) attribution methods has been established on convolutional neural network

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

DrawAI: Agentic Benchmark and Workflow for Making Raster Images Editable

DGX agent

arXiv:2608.00548v1 Announce Type: new Abstract: Recent image-generation models and multimodal agents can produce high-quality visuals for increasingly complex visual communication tasks. Yet their ras

model-releasesarxiv-cs-cv
4 Aug 2026
Research

DreamTraj: Generating 6-DoF Object Trajectories by Reading Unrendered Video Diffusion Latents

DGX agent

arXiv:2608.00486v1 Announce Type: new Abstract: Accurate prediction of object trajectories during manipulation is essential for closing the perception-action loop. Progress is limited on two fronts: a

researcharxiv-cs-cv
4 Aug 2026
Agents

DriveCode: Domain Specific Numerical Encoding for LLM-Based Autonomous Driving

DGX agent

arXiv:2603.00919v3 Announce Type: replace Abstract: Large language models (LLMs) have shown great promise for autonomous driving. However, discretizing numbers into tokens limits precise numerical rea

agentsarxiv-cs-cv
4 Aug 2026
Safety

Driver2Map: Imitating Human Driving for Online High-Definition Map Construction

DGX agent

arXiv:2608.01338v1 Announce Type: new Abstract: High-definition (HD) maps are essential for autonomous driving systems. In constructing such maps, onboard multi-view camera images, standard-definition

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

DS@GT ARC at MEDIQA-CORE-Task-1 2026: Trimodal Model Fusion with Task-Specific Gates for Brain Tumor Subtype Classification

DGX agent

arXiv:2608.00086v1 Announce Type: new Abstract: Brain tumor diagnosis is a time-sensitive process in which patients may wait weeks for a finalized pathology report. This problem motivates automated sy

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

DyFrDet: Towards Accurate Small Object Detection via Dynamic Frequency Suppression with Label Disambiguation

DGX agent

arXiv:2608.02495v1 Announce Type: new Abstract: Despite the remarkable progress over the past decades, accurately identifying small objects remains challenging because of their insufficient visual cue

local-aiarxiv-cs-cv
4 Aug 2026
Agents

DynActiveGS: Active Gaussian Splatting for Dynamic Scene Reconstruction

DGX agent

arXiv:2608.01178v1 Announce Type: new Abstract: We present DynActiveGS, a dynamic-aware active reconstruction framework based on 3D Gaussian Splatting (3DGS) for autonomous exploration in dynamic envi

agentsarxiv-cs-cv
4 Aug 2026
Local Ai

Dynamic Resolution Routing for Efficient Egocentric Grounding

DGX agent

arXiv:2608.01638v1 Announce Type: new Abstract: Egocentric visual grounding requires high-resolution inputs to localize small objects. However, scaling Multimodal Large Language Models to this domain

local-aiarxiv-cs-cv
4 Aug 2026
Model Releases

DynamicManip: Enabling Dynamic Manipulation from a Single Static Demonstration

DGX agent

arXiv:2608.01452v1 Announce Type: cross Abstract: Dynamic manipulation is a critical capability for robots operating in complex and dynamic environments, where robots must interact with objects that a

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

E2Pano: Learning Event-to-Panorama Image Reconstruction

DGX agent

arXiv:2608.00694v1 Announce Type: new Abstract: Event cameras offer microsecond-level temporal resolution and high dynamic range, potentially facilitating motion-blur-free panoramic imaging from fast

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation

DGX agent

arXiv:2608.02474v1 Announce Type: new Abstract: Audio-driven video generation (A2V) has achieved promising progress in synthesizing temporally coherent and audio-visually aligned videos, yet its infer

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models

DGX agent

arXiv:2601.17883v3 Announce Type: replace-cross Abstract: Electroencephalography (EEG) foundation models (FMs) have recently emerged as a promising paradigm for brain-computer interfaces, aiming to le

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next

DGX agent

arXiv:2603.12147v2 Announce Type: replace Abstract: Egocentric video provides a natural modality for studying human behavior, but conventional visual understanding captures mainly observable scenes, o

model-releasesarxiv-cs-cv
4 Aug 2026
Research

ELECTRIC: Evidential Learning-Enhanced CT Reconstruction via Iterative Correction

DGX agent

arXiv:2608.00060v1 Announce Type: new Abstract: Here we introduce ELECTRIC (Evidential Learning-Enhanced CT Reconstruction via Iterative Correction), a physics-guided Bayesian formulation. An evidenti

researcharxiv-cs-cv
4 Aug 2026
Safety

Element-Aware Group Learning for E-Commerce Image Generation

DGX agent

arXiv:2608.00584v1 Announce Type: new Abstract: Recent advances in image generation and editing have made prompt quality a key bottleneck for e-commerce creatives. Vision-language models (VLMs) can ge

safetyarxiv-cs-cv
4 Aug 2026
Research

EmoScene: A Dual-space Dataset for Controllable Affective Image Generation

DGX agent

arXiv:2604.00933v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve high visual fidelity, yet fine-grained affective control remains difficult because textual emotion cues often

researcharxiv-cs-cv
4 Aug 2026
Research

Empirical investigation of 3D CT Foundation Models and Unsupervised Adaptation for Head and Neck Cancer Recurrence Prediction

DGX agent

arXiv:2608.00071v1 Announce Type: new Abstract: The rapid emergence of 3D CT foundation models has opened new avenues for predictive modeling from CT imaging, offering a compelling alternative to trad

researcharxiv-cs-cv
4 Aug 2026
Agents

Enhancing Visual Perception in Foggy Conditions via Multiclass Fog Density Modeling

DGX agent

arXiv:2608.01572v1 Announce Type: new Abstract: Autonomous driving (AD) systems have advanced rapidly over the past decade; however, robust perception under adverse weather conditions remains a major

agentsarxiv-cs-cv
4 Aug 2026
Safety

Entity-Aware Sequence Transduction for Player-Centric Ball Action Spotting

DGX agent

arXiv:2608.01696v1 Announce Type: new Abstract: Player-centric ball action spotting requires temporally precise event detection together with actor attribution in crowded, partially observed multi-age

safetyarxiv-cs-cv
4 Aug 2026
Research

EOVSAM: Efficient Open-Vocabulary Segmentation with SAM 3 in One Pass

DGX agent

arXiv:2608.02284v1 Announce Type: new Abstract: Open-vocabulary segmentation identifies and segments objects from arbitrary textual descriptions. SAM 3 supports noun-phrase-guided segmentation and ach

researcharxiv-cs-cv
4 Aug 2026
Research

Estimating SSIM from MSE for DCT-Based Compressed Images

DGX agent

arXiv:2608.02549v1 Announce Type: cross Abstract: Efficient and perceptually meaningful quality assessment is a fundamental requirement for image and video processing, compression, and streaming syste

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Event ActivityNet: A Large-Scale Simulated-Event Benchmark for Untrimmed Action Understanding

DGX agent

arXiv:2608.01948v1 Announce Type: new Abstract: Long-horizon event-based action understanding remains underexplored because existing datasets largely comprise short, trimmed clips, while collecting na

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Explainable Multimodal AI for Adaptive Calibration of Archaeological Sensing Workflows

DGX agent

arXiv:2608.00074v1 Announce Type: new Abstract: This paper presents a multimodal machine-learning framework for calibration monitoring, quality assessment, and adaptive acquisition support in archaeol

researcharxiv-cs-cv
4 Aug 2026
Research

Extended Field of View Analysis for VideoGAN-based Trajectory Generation

DGX agent

arXiv:2608.02289v1 Announce Type: new Abstract: Realistic and diverse trajectory generation is central to enabling higher levels of vehicle automation. While rule-based and classical learning-based me

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Extended KAFR: A kinematic-adaptive paradigm for the efficient analysis of surgical video

DGX agent

arXiv:2608.01058v1 Announce Type: new Abstract: Artificial Intelligence is increasingly applied to surgical video analysis for phase segmentation, skill assessment, and workflow optimization. A key ch

model-releasesarxiv-cs-cv
4 Aug 2026
Agents

FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

DGX agent

arXiv:2608.01049v1 Announce Type: cross Abstract: World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this e

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

FairForensics: Seeing Expressions and Parsing Demographics via Vision-Language Modeling for Generalizable Fair Deepfake Detection

DGX agent

arXiv:2608.01661v1 Announce Type: new Abstract: The challenge of fair deepfake detection (FDD) has attracted increasing attention. Existing fairness-enhanced detectors often suffer from suboptimal gen

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

FAST-GS: Frequency Aware Space-time Gaussian Splatting for Photorealistic Dynamic Novel View Synthesis

DGX agent

arXiv:2608.01958v1 Announce Type: new Abstract: 4D Gaussian Splatting (4DGS) excels in dynamic 3D reconstruction and real-time novel view synthesis via efficient 4D Gaussian representations and parall

local-aiarxiv-cs-cv
4 Aug 2026
Research

Fast Trainable Multilinear Bases for Image Compression

DGX agent

arXiv:2608.00053v1 Announce Type: cross Abstract: The Discrete Fourier Transform, the Discrete Cosine Transform, and their block-wise variants underpin most deployed image and video codecs. Their effe

researcharxiv-cs-cv
4 Aug 2026
← Previous
1…2021222324…261
Next →