AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Research

Edge-directed geometric partitioning for versatile video coding

DGX agent

arXiv:2606.01641v1 Announce Type: new Abstract: To improve the coding performance, geometric partition (GEO) was proposed for the upcoming VVC standard. GEO provides 140 partition candidates. The inde

researcharxiv-cs-cv
2 Jun 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Edge Prediction for Roof Wireframe Reconstruction with Transformers

DGX agent

arXiv:2606.02406v1 Announce Type: new Abstract: This paper presents a competitive solution to the S23DR Challenge 2026, which aims to reconstruct 3D house roof wireframe models from sparse SfM point c

researcharxiv-cs-cv
2 Jun 2026
Research

Effective Multi-sensor Conditioning for Street-view Novel-view Synthesis

DGX agent

arXiv:2606.01590v1 Announce Type: new Abstract: Modern vehicle platforms are equipped with a rich sensor suite, including LiDAR, calibrated multi-camera rigs, and accurate ego-motion, that in principl

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark

DGX agent

arXiv:2606.02246v1 Announce Type: new Abstract: To operate in the physical world, embodied agents must perceive their environment in an 'always-on' fashion, selectively accessing the most informative

model-releasesarxiv-cs-cv
2 Jun 2026
Research

EIVE: End-to-End Instance-Specific Visual Explanations for Detection Transformers

DGX agent

arXiv:2606.01601v1 Announce Type: new Abstract: Visual explainability for object detection remains challenging due to the multi-instance nature of detection. Existing approaches predominantly adopt po

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Enhancing Blind Source Separation with Dissociative Principal Component Analysis

DGX agent

arXiv:2411.12321v2 Announce Type: replace Abstract: Principal component analysis (PCA) and its sparse variants (sPCA) are widely used as a precursor to independent component analysis (ICA) for blind s

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Entropy Minimization without Model Collapse: Mitigating Prediction Bias in Medical Imaging

DGX agent

arXiv:2606.02339v1 Announce Type: cross Abstract: Entropy minimization (EM) is the dominant objective for test-time adaptation, yet its failure mode, model collapse, remains poorly understood. In this

safetyarxiv-cs-cv
2 Jun 2026
Safety

Equilibrated Diffusion: Frequency-aware Textual Embedding for Equilibrated Image Customization

DGX agent

arXiv:2606.02129v1 Announce Type: new Abstract: Image customization learns target subjects from reference concept images and generates conditioned images per text prompts, mainly modifying styles or b

safetyarxiv-cs-cv
2 Jun 2026
Research

ETC: Extreme Token Compression via Task-aware Visual Information Distillation in VLMs

DGX agent

arXiv:2606.00543v1 Announce Type: new Abstract: In Vision-Language Models (VLMs), high-resolution images produce a large number of visual tokens, resulting in high computational costs and KV-cache ove

researcharxiv-cs-cv
2 Jun 2026
Research

Event-Based Vision in Space: Applications, Trends, and Future Directions

DGX agent

arXiv:2606.01280v1 Announce Type: new Abstract: Earth Observation (EO) is undergoing a significant transformation driven by the deployment of novel sensing technologies. Traditional frame-based optica

researcharxiv-cs-cv
2 Jun 2026
Research

EvoCut: Multi-Layer Evolution-Aware Visual Token Compression for Efficient Large Vision-Language Models

DGX agent

arXiv:2606.01756v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong performance on image and video understanding tasks, but their inference efficiency is constrained by

researcharxiv-cs-cv
2 Jun 2026
Applications

Evolving to the Aesthetics of a Vision-Language Model

DGX agent

arXiv:2606.00112v1 Announce Type: cross Abstract: Evolutionary systems have demonstrated remarkable results in creative domains, with recent applications in generative typography, design, and music. H

applicationsarxiv-cs-cv
2 Jun 2026
Safety

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs

DGX agent

arXiv:2606.01072v1 Announce Type: cross Abstract: Imitation learning enables robots to learn how to execute tasks via observation. However, real-world environments like homes and offices are often sev

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos

DGX agent

arXiv:2606.02402v1 Announce Type: new Abstract: The rapid advancement of AI-driven video generation has transformed content creation, while simultaneously increasing the risk of misinformation through

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Exploiting In-Sensor Computing for Energy-Efficient Earth Observation

DGX agent

arXiv:2606.01271v1 Announce Type: new Abstract: The rapid growth of the satellite industry has driven a significant increase in geospatial data acquisition, highlighting a critical bottleneck: the sev

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression

DGX agent

arXiv:2606.01608v1 Announce Type: new Abstract: Most existing extreme compression methods fail to achieve an optimal rate-distortion-perception trade-off, as they typically prioritize perceptual fidel

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays

DGX agent

arXiv:2509.15234v2 Announce Type: replace Abstract: Multimodal learning from paired medical images and clinical text is a central challenge in medical data-driven informatics, where effective cross-mo

model-releasesarxiv-cs-cv
2 Jun 2026
Research

ext{VG}^2GT: Voxel-Gaussian Splatting Visual Geometry Grounded Transformer

DGX agent

arXiv:2606.01573v1 Announce Type: new Abstract: Gaussian splatting has shown strong potential for 3D reconstruction and novel view synthesis. However, most existing methods require accurate camera par

researcharxiv-cs-cv
2 Jun 2026
Model Releases

FACT: A Simple and Efficient Framework for Active Finetuning

DGX agent

arXiv:2606.02079v1 Announce Type: new Abstract: The main goal of active finetuning is to improve a pretrained model's performance on a specific task or domain by finetuning it with carefully selected

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Family Matters: A Systematic Study of Spatial vs. Frequency Masking for Continual Test-Time Adaptation

DGX agent

arXiv:2512.08048v3 Announce Type: replace Abstract: Recent continual test-time adaptation (CTTA) methods adopt masked image modeling to stabilize learning under distribution shift, yet each treats its

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

Fast-SAM3D: 3Dfy Anything in Images but Faster

DGX agent

arXiv:2602.05293v2 Announce Type: replace Abstract: SAM3D enables scalable, open-world 3D reconstruction from complex scenes, yet its deployment is hindered by prohibitive inference latency. In this w

model-releasesarxiv-cs-cv
2 Jun 2026
Local Ai

FDIO: Frequency Decomposed Inertial Odometry

DGX agent

arXiv:2511.15645v3 Announce Type: replace Abstract: Pedestrian inertial odometry (PIO) estimates autonomous pedestrian motion using only acceleration and angular velocity measurements collected by an

local-aiarxiv-cs-cv
2 Jun 2026
Safety

Feature Alignment Determines Fusion Strategy: A Comparative Study of Cross-Attention and Concatenation in Multimodal Learning

DGX agent

arXiv:2606.01207v1 Announce Type: new Abstract: The choice between cross-attention and concatenation for multimodal fusion remains governed by practitioner intuition rather than principled understandi

safetyarxiv-cs-cv
2 Jun 2026
Tutorials

FiSeR: Fine-Grained Source Representations for Cross-Domain AI Image Detection

DGX agent

arXiv:2606.00606v1 Announce Type: new Abstract: Real-world synthetic image detectors often generalize poorly under domain shift despite strong in-domain performance. Using unsupervised UMAP projection

tutorialsarxiv-cs-cv
2 Jun 2026
Model Releases

FLAME: Physics-Guided Neural Operators for Onboard Satellite Methane Detection in Hyperspectral Imagery

DGX agent

arXiv:2606.01577v1 Announce Type: new Abstract: Methane is a major driver of near-term climate change, and rapidly identifying its emission sources is a critical climate intervention. Spaceborne hyper

model-releasesarxiv-cs-cv
2 Jun 2026
Research

FlatVPR: Plug-and-play Geo-linear Residual Adapter for Geometric Rectification of Foundation Model Feature Manifolds

DGX agent

arXiv:2606.01734v1 Announce Type: new Abstract: This paper proposes ``FlatVPR,'' a novel geometric rectification paradigm that effectively bridges the trade-off between map lightweightness and localiz

researcharxiv-cs-cv
2 Jun 2026
Safety

Flexible Control of 3D CT Generation via Text and Semantically-Defined Segmentation Prompts

DGX agent

arXiv:2606.00967v1 Announce Type: new Abstract: Generative models for volumetric medical images have found many applications in medical imaging, ranging from data augmentation to serving as priors for

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

FlowIt: Global Matching via Hierarchical Transformers and Optimal Transport for Optical Flow

DGX agent

arXiv:2603.28759v2 Announce Type: replace Abstract: We present FlowIt, a novel architecture for optical flow estimation that combines global matching with confidence and occlusion-guided refinement. A

model-releasesarxiv-cs-cv
2 Jun 2026
Research

FlowNar: Scalable Streaming Narration for Long-Form Videos

DGX agent

arXiv:2606.00620v1 Announce Type: new Abstract: Recent Large Multimodal Models (LMMs), primarily designed for offline settings, are ill-suited for the dynamic requirements of streaming video. While re

researcharxiv-cs-cv
2 Jun 2026
Model Releases

FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection

DGX agent

arXiv:2606.00782v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has achieved remarkable progress through large-scale vision-language pre-training. Existing methods, however, typ

model-releasesarxiv-cs-cv
2 Jun 2026
Research

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation

DGX agent

arXiv:2606.02090v1 Announce Type: new Abstract: Diffusion transformer (DiT) has been widely adopted in the generative diffusion field, advancing the denoising of query tokens through attention and Fee

researcharxiv-cs-cv
2 Jun 2026
Local Ai

ForestMamba: Sparse Mamba with Geometry-guided Queries for 3D Forest Point Cloud Segmentation

DGX agent

arXiv:2606.01549v1 Announce Type: new Abstract: AI-based semantic and instance segmentation of terrestrial and drone LiDAR point clouds is emerging as a transformative approach for converting the comp

local-aiarxiv-cs-cv
2 Jun 2026
Research

FOVI: A biologically-inspired foveated interface for deep vision models

DGX agent

arXiv:2602.03766v2 Announce Type: replace Abstract: Human vision is foveated, with variable resolution peaking at the center of a large field of view; this reflects an efficient trade-off for active s

researcharxiv-cs-cv
2 Jun 2026
Research

From Extrinsic to Intrinsic: Geodesic-Guided Representation Learning for 3D Geometric Data

DGX agent

arXiv:2606.02268v1 Announce Type: new Abstract: Geometric analysis fundamentally distinguishes between extit{extrinsic} and extit{intrinsic} perspectives. The dominant paradigm in current 3D represent

researcharxiv-cs-cv
2 Jun 2026
Research

From Zero to Hero: Training-Free Custom Concept Spawning in World Models

DGX agent

arXiv:2606.02575v1 Announce Type: new Abstract: Autoregressive world models have emerged as a powerful paradigm for interactive video generation, allowing users to navigate dynamically generated envir

researcharxiv-cs-cv
2 Jun 2026
Safety

FROST-STA: Frozen Dense Features for the Ego4D Short-Term Object Interaction Anticipation

DGX agent

arXiv:2606.00694v1 Announce Type: new Abstract: Short-term anticipation in egocentric video requires more than recognizing the current scene: a system must infer which object the camera wearer will co

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

GABI: Geometry-Aware Boundary Integration for Spacecraft Segmentation

DGX agent

arXiv:2606.00886v1 Announce Type: new Abstract: Accurate segmentation is crucial for autonomous spacecraft, as it directly affects downstream tasks related to 3D situational awareness. The harsh illum

model-releasesarxiv-cs-cv
2 Jun 2026
Research

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling

DGX agent

arXiv:2606.00110v1 Announce Type: new Abstract: Achieving robust generalization from limited data is a central challenge in embodied intelligence. Prevailing methods fail by regressing absolute coordi

researcharxiv-cs-cv
2 Jun 2026
Research

Generalization Limits in Vehicle Re-Identification

DGX agent

arXiv:2606.01981v1 Announce Type: new Abstract: Vehicle re-identification focuses on retrieving images of the same vehicle from a gallery given a query image. Upon closer inspection of commonly used d

researcharxiv-cs-cv
2 Jun 2026
Research

Generate in Reconstruction Space, Match in Semantic Space: Transport Geometry for One-Step Generation

DGX agent

arXiv:2606.00514v1 Announce Type: cross Abstract: Generative modeling and self-supervised representation learning (SSL) optimize structurally different objectives: generative training rewards distribu

researcharxiv-cs-cv
2 Jun 2026
Tutorials

Generative Diffusion Priors for 3D Mapping of the Dark Universe

DGX agent

arXiv:2606.00803v1 Announce Type: cross Abstract: Reconstructing the three-dimensional distribution of dark matter from weak-lensing observations is a central but highly ill-posed inverse problem in c

tutorialsarxiv-cs-cv
2 Jun 2026
Model Releases

Geometry-Aware Implicit Memory for Video World Models

DGX agent

arXiv:2606.02436v1 Announce Type: new Abstract: Video world models aim to simulate controllable visual environments, but long-horizon rollouts depend on what the model remembers after observations lea

model-releasesarxiv-cs-cv
2 Jun 2026
Research

GloResNet: A lightweight 3D CNN with global topological features for preterm brain injury prediction

DGX agent

arXiv:2606.02498v1 Announce Type: new Abstract: This study introduces an automated deep learning framework for predicting brain injury (BI) in preterm infants from T2-weighted MRI (dHCP dataset). We p

researcharxiv-cs-cv
2 Jun 2026
Agents

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation

DGX agent

arXiv:2606.01621v1 Announce Type: new Abstract: Vision-language models (VLMs) have become a common foundation for vision-and-language navigation in continuous environments (VLN-CE). Yet most VLM-based

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

HakushoBench: A Japanese Chart and Table VQA Benchmark from Governmental White Papers

DGX agent

arXiv:2606.01132v1 Announce Type: new Abstract: Understanding chart and table images is essential for applying vision-language models (VLMs) to real-world document understanding. While English benchma

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Hallucination-Aware Diffusion Sampling for Inverse Problems via Robust Prior Updates

DGX agent

arXiv:2606.02331v1 Announce Type: new Abstract: Diffusion-based inverse problem solvers can produce realistic reconstructions, but realism alone does not ensure that the recovered details are supporte

researcharxiv-cs-cv
2 Jun 2026
Safety

Hard Labels In! Rethinking the Role of Hard Labels in Mitigating Local Semantic Drift

DGX agent

arXiv:2512.15647v3 Announce Type: replace Abstract: Soft labels from teacher models are a de facto practice for knowledge transfer and large-scale dataset distillation (e.g., SRe2L, LPLD). However, wh

safetyarxiv-cs-cv
2 Jun 2026
Research

Head-Pose-Aware Visual Speech Recognition with FiLM Modulation

DGX agent

arXiv:2606.00751v1 Announce Type: new Abstract: Visual Speech Recognition (VSR) aims to recognize speech from visual cues such as lip movements, but its performance is fundamentally limited by viseme

researcharxiv-cs-cv
2 Jun 2026
← Previous
1…126127128129130…263
Next →