AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,618 results
Model Releases

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

DGX agent

arXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i

model-releasesarxiv-cs-cv
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Time-Conditioned and Multi-Time Survival Prediction from 2D PET/CT Projections in Lung Cancer

DGX agent

arXiv:2606.12140v1 Announce Type: new Abstract: Accurate prediction of overall survival (OS) from positron emission tomography/computed tomography (PET/CT) can support personalized treatment and follo

researcharxiv-cs-cv
11 Jun 2026
Tutorials

TopoCap: Learning Topology-Agnostic Motion Priors for Monocular Video-to-Animation

DGX agent

arXiv:2606.12153v1 Announce Type: new Abstract: The explosion of generative 3D assets has created a massive demand for animation, yet current motion capture methods remain brittle, restricted to speci

tutorialsarxiv-cs-cv
11 Jun 2026
Safety

Towards Conditional Feature Alignment for Cross-Domain Counting

DGX agent

arXiv:2506.17137v3 Announce Type: replace Abstract: Object counting models often degrade under cross-domain deployment because density composition varies across domains and is itself task-relevant. St

safetyarxiv-cs-cv
11 Jun 2026
Safety

Traits Run Deeper: Trait-Specific Asymmetric Fusion for Personality Assessment

DGX agent

arXiv:2606.11269v1 Announce Type: new Abstract: Personality assessment aims to infer stable personality traits from dynamic behaviors across language, voice, and facial cues. Since different personali

safetyarxiv-cs-cv
11 Jun 2026
Applications

TRON: Tracing Rays to Orchestrate a Neural Renderer for 3D Gaussian Reconstructions

DGX agent

arXiv:2606.11314v1 Announce Type: new Abstract: We introduce TRON, a rendering framework that combines 3D Gaussian ray tracing with neural rendering to enable realistic and controllable rendering of r

applicationsarxiv-cs-cv
11 Jun 2026
Research

Understanding Cross-Sensor Feature Variations for Generalizable 3D Perception

DGX agent

arXiv:2606.11573v1 Announce Type: new Abstract: Radar-camera BEV perception often suffers from degraded performance when evaluated across datasets, as changes in driving scenes, sensor configurations,

researcharxiv-cs-cv
11 Jun 2026
Research

Vision Transformers for Face Recognition Need More Registers

DGX agent

arXiv:2606.12036v1 Announce Type: new Abstract: Recent advances in Vision Transformers (ViTs) for face recognition (FR) have moved beyond the standard CLS-token paradigm. In this paradigm, a special c

researcharxiv-cs-cv
11 Jun 2026
Safety

ViT-FREE: Efficient Face Recognition via Early Exiting and Synthetic Adaptation

DGX agent

arXiv:2606.12023v1 Announce Type: new Abstract: Vision Transformers (ViTs) have gained significant attention in computer vision and shown strong potential for face recognition (FR). However, their hig

safetyarxiv-cs-cv
11 Jun 2026
Model Releases

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

DGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

model-releasesarxiv-cs-cv
11 Jun 2026
Safety

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving

DGX agent

arXiv:2606.12396v1 Announce Type: new Abstract: Vision-language-action (VLA) models can describe scenes and reason about them in language, yet still struggle to ground their actions in the dense 3D wo

safetyarxiv-cs-cv
11 Jun 2026
Research

VOID: Defeating Unauthorized Mimicry in Latent Diffusion Models

DGX agent

arXiv:2606.12263v1 Announce Type: new Abstract: While Latent Diffusion Models (LDMs) have revolutionized visual synthesis, they are increasingly exploited for unauthorized mimicry of individuals. Exis

researcharxiv-cs-cv
11 Jun 2026
Applications

Wild3R: Feed-Forward 3D Gaussian Splatting from Unconstrained Sparse Photo Collection

DGX agent

arXiv:2606.11894v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) removes the need for time-consuming per-scene optimization required by traditional 3DGS. However, existing fee

applicationsarxiv-cs-cv
11 Jun 2026
Model Releases

World Model Self-Distillation: Training World Models to Solve General Tasks

DGX agent

arXiv:2606.12072v1 Announce Type: new Abstract: Pretrained video generators are promising visual world models that exhibit emergent task-solving abilities; however, their reliance on detailed textual

model-releasesarxiv-cs-cv
11 Jun 2026
Research

XPR: An Extensible Cross-Platform Point-Based Differentiable Renderer

DGX agent

arXiv:2606.11529v1 Announce Type: cross Abstract: Point-based differentiable rendering underpins modern 3D reconstruction, novel-view synthesis, and learning-based graphics pipelines, but developing n

researcharxiv-cs-cv
11 Jun 2026
Agents

3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis

DGX agent

arXiv:2606.10478v1 Announce Type: new Abstract: Most recent 3D reconstruction and editing systems operate on implicit and explicit representations such as NeRF, point clouds, or meshes. While these re

agentsarxiv-cs-cv
10 Jun 2026
Model Releases

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

A fine-grained attention and geometric correspondence model for musculoskeletal risk classification in athletes using multimodal visual and skeletal features

DGX agent

arXiv:2509.05913v3 Announce Type: replace Abstract: Musculoskeletal disorders pose significant risks to athletes, and early risk assessment is essential for prevention. However, most existing methods

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

A Large Scale Open-Source Image and Video Dataset for Robust Wildfire Detection and Classification

DGX agent

arXiv:2606.10174v1 Announce Type: new Abstract: Wildfire detection and monitoring are critical for mitigating fire spread and reducing environmental and infrastructural damage. In this work, we introd

model-releasesarxiv-cs-cv
10 Jun 2026
Research

A Multimodal RGB and Events Dataset for Hand Detection in First-Person View

DGX agent

arXiv:2606.10790v1 Announce Type: new Abstract: Existing hand detection algorithms work on images and the detection rate is restricted by the frame rate of the camera. In hand detection applications f

researcharxiv-cs-cv
10 Jun 2026
Applications

ABot-Earth 0.5: Generative 3D Earth Model

DGX agent

arXiv:2606.09967v1 Announce Type: new Abstract: We present ABot-Earth 0.5, a generative 3D framework designed to synthesize vast, seamless 3D environments from ubiquitous, geospatially referenced sate

applicationsarxiv-cs-cv
10 Jun 2026
Research

Advancing Wood Identification in the Philippines: Utilizing the Xylorix Platform for Efficient AI Model Development and Deployment for Five Key Species

DGX agent

arXiv:2606.10876v1 Announce Type: new Abstract: Illegal logging and timber trade continue to pose significant challenges in the Philippines, where accurate wood species identification is essential for

researcharxiv-cs-cv
10 Jun 2026
Research

An Uncertainty Estimation Framework for Dose Accumulation in Adaptive Radiotherapy: Application to CBCT-Guided Radiotherapy for Cervical Cancer

DGX agent

arXiv:2606.11012v1 Announce Type: new Abstract: Background and purpose: oART enables daily plan adaptation to interfraction anatomical variations, but cumulative dose estimation remains limited by DIR

researcharxiv-cs-cv
10 Jun 2026
Applications

Analyzing Training-Free Corruption Detection for Object Detection Datasets

DGX agent

arXiv:2606.10666v1 Announce Type: new Abstract: Annotation errors are widespread in computer vision datasets and can significantly degrade the performance of systems trained on them, particularly in c

applicationsarxiv-cs-cv
10 Jun 2026
Safety

AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects

DGX agent

arXiv:2606.10988v1 Announce Type: new Abstract: While recent advancements in generative AI have substantially accelerated static 3D model creation workflows, the synthesis of category-agnostic 3D anim

safetyarxiv-cs-cv
10 Jun 2026
Applications

AnyMod-LLVE: Low-Light Video Enhancement with Modality-Agnostic Inference

DGX agent

arXiv:2606.11186v1 Announce Type: new Abstract: Low-light video enhancement (LLVE) remains a challenging task due to severe information degradation under low-illumination conditions. Recent multimodal

applicationsarxiv-cs-cv
10 Jun 2026
Safety

ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations

DGX agent

arXiv:2606.11188v1 Announce Type: new Abstract: This paper introduces ARM, a discrete representation-based AutoRegressive Model that unifies image understanding, generation, and editing within a next-

safetyarxiv-cs-cv
10 Jun 2026
Research

Audio-Visual Exchange-Aware Token Pruning for Efficient Audio-Visual Captioning

DGX agent

arXiv:2606.10533v1 Announce Type: new Abstract: Audio-visual captioning generates natural language descriptions from video and audio content. Multimodal LLMs have advanced this task, but both modaliti

researcharxiv-cs-cv
10 Jun 2026
Safety

Automatic Labelling for Low-Light Pedestrian Detection

DGX agent

arXiv:2507.02513v4 Announce Type: replace Abstract: Pedestrian detection in RGB images is a key task in pedestrian safety, as the most common sensor in autonomous vehicles and advanced driver assistan

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

Benchmarking stereo reconstruction for 3D printable Martian terrain models

DGX agent

arXiv:2606.10364v1 Announce Type: new Abstract: Reconstructing printable 3D models from Mars rover imagery is challenging because Martian terrain is low-texture, irregular, and partially observed. We

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Beyond Model Size: Probing the Gaps in Visual in-Context Learning by Training a Tiny Model

DGX agent

arXiv:2606.10905v1 Announce Type: new Abstract: Visual in-Context Learning (VICL) aims at making progress towards adaptive vision models, that can -- based on a few examples -- adapt to a new task at

model-releasesarxiv-cs-cv
10 Jun 2026
Tutorials

Breaking the Curse of Dimensionality: Diffusion Models Efficiently Learn Low-Dimensional Distributions

DGX agent

arXiv:2409.02426v5 Announce Type: replace-cross Abstract: Despite their empirical success across a wide range of generative tasks, the fundamental principles underlying the ability of diffusion models

tutorialsarxiv-cs-cv
10 Jun 2026
Local Ai

CapStARE: Capsule-based Sequential Architecture for Robust and Efficient Gaze Estimation

DGX agent

arXiv:2509.19936v2 Announce Type: replace Abstract: Human gaze estimation is essential for applications such as human-computer interaction, social robotics, and assistive systems. However, achieving a

local-aiarxiv-cs-cv
10 Jun 2026
Model Releases

ChartLens: A Dual-Branch Framework for Chart Data Correction and Factual Summary Refinement

DGX agent

arXiv:2606.10640v1 Announce Type: new Abstract: In this report, we present our champion solution for the DataMFM Challenge Track 2: Chart Understanding. This track requires models to recover structure

model-releasesarxiv-cs-cv
10 Jun 2026
Research

ClinReadNet: A clinical reading-inspired network for low-dose abdominal CT image quality assessment

DGX agent

arXiv:2606.10372v1 Announce Type: new Abstract: In abdominal CT imaging, developing a low-dose, no-reference image quality assessment (No-reference IQA) model that mimics doctors' reading habits for e

researcharxiv-cs-cv
10 Jun 2026
Model Releases

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

DGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Continuous Neural Reparameterization as a Deep Geometric Prior for Robust Fixed-Chart UV Repair

DGX agent

arXiv:2606.10050v1 Announce Type: cross Abstract: Traditional UV unwrapping relies on direct optimization of geometric distortion energies and can fail through invalid initialization, local minima, or

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP

DGX agent

arXiv:2601.19210v2 Announce Type: replace Abstract: Vision-language models (VLMs) such as CLIP have demonstrated remarkable zero-shot generalization, yet remain highly vulnerable to adversarial exampl

safetyarxiv-cs-cv
10 Jun 2026
Research

Cost-Aware Routing for Efficient Text-To-Image Generation

DGX agent

arXiv:2506.14753v3 Announce Type: replace Abstract: Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfo

researcharxiv-cs-cv
10 Jun 2026
Model Releases

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

DGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

DGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

model-releasesarxiv-cs-cv
10 Jun 2026
Research

DD-INR: Dynamics-Driven Implicit Neural Representation for Accelerated Whole-Brain Functional MRI Reconstruction

DGX agent

arXiv:2606.10756v1 Announce Type: new Abstract: Accelerated acquisition of fMRI enables enhanced detection of neurovascular (BOLD) activity in the brain, but image reconstruction becomes challenging w

researcharxiv-cs-cv
10 Jun 2026
Research

Deep learning for echo sounder data

DGX agent

arXiv:2606.10811v1 Announce Type: new Abstract: There is no doubt that over the last decade, techniques from the field of machine learning have revolutionized how we process and interpret data, especi

researcharxiv-cs-cv
10 Jun 2026
Safety

Dexterous Point Policy: Learning Point-based Dexterous Hand Policies from Human Demonstrations

DGX agent

arXiv:2606.10614v1 Announce Type: cross Abstract: Robotic foundation models pre-trained on human demonstration videos have shown promise, but a significant embodiment gap remains when the resulting po

safetyarxiv-cs-cv
10 Jun 2026
Safety

Dissect and Prune: Enhancing Robustness in AI-Generated Image Detection

DGX agent

arXiv:2606.10309v1 Announce Type: new Abstract: While existing AI-generated image detectors report high performance, we identify that this is largely driven by a critical prediction asymmetry: a bias

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

Don't waste SAM

DGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

model-releasesarxiv-cs-cv
10 Jun 2026
Local Ai

Dual-stream attention-guided learning for weakly supervised whole slide image classification

DGX agent

arXiv:2505.23341v3 Announce Type: replace Abstract: Whole slide images (WSIs) play a crucial role in cancer diagnosis due to their ultra-high resolution and rich morphological information, and multipl

local-aiarxiv-cs-cv
10 Jun 2026
Model Releases

Efficient RWKV-based Representation Learning for 3D Point Clouds

DGX agent

arXiv:2606.10395v1 Announce Type: new Abstract: The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic

model-releasesarxiv-cs-cv
10 Jun 2026
← Previous
1…106107108109110…263
Next →