AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlog
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,414 results
Research

HQ-DM: Single Hadamard Transformation-Based Quantization-Aware Training for Low-Bit Diffusion Models

DGX agent

arXiv:2512.05746v3 Announce Type: replace Abstract: Diffusion models have demonstrated significant applications in the field of image generation. However, their high computational and memory costs pos

researcharxiv-cs-cv
12 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

HUI360: A 360{eg} Egocentric Dataset and Baselines for Human-Robot Interaction Anticipation

DGX agent

arXiv:2608.11051v1 Announce Type: new Abstract: As robots increasingly operate in human-populated environments, anticipating human intentions is essential for enabling proactive and socially aware beh

model-releasesarxiv-cs-cv
12 Aug 2026
Tutorials

Human versus Computer Vision

DGX agent

arXiv:2608.10181v1 Announce Type: new Abstract: Computer vision saliency models predict where people will look, one map per image, and a billion-dollar predicted-attention industry sells those maps in

tutorialsarxiv-cs-cv
12 Aug 2026
Model Releases

Implicit representations are dead. Long live explicit primitives!

DGX agent

arXiv:2608.10001v1 Announce Type: cross Abstract: Continuous parameterization of medical data has emerged as a powerful paradigm for resolution-independent image representation. While Implicit Neural

model-releasesarxiv-cs-cv
12 Aug 2026
Research

InterPruner: Interactive Structured Pruning via Taylor-Implicit Criterion and Language-Prior Modulator for Multimodal Object Detection

DGX agent

arXiv:2608.10724v1 Announce Type: new Abstract: Multimodal object detection proves effective in remote sensing, especially the RGB-Infrared paradigm. The parallel feature extractors provide rich multi

researcharxiv-cs-cv
12 Aug 2026
Model Releases

Introspective Attention Modulation for Safe Text-to-Image Generation

DGX agent

arXiv:2607.14945v2 Announce Type: replace Abstract: State-of-the-art flow based text-to-image (T2I) models exhibit remarkable generative abilities but remain vulnerable to producing unsafe content. Pr

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Is There Really a Camouflaged Object? Towards Realistic Camouflaged Object Detection

DGX agent

arXiv:2608.11135v1 Announce Type: new Abstract: Camouflaged object detection (COD) aims to segment objects that are visually concealed in their surroundings and has attracted increasing attention in r

model-releasesarxiv-cs-cv
12 Aug 2026
Research

Iterative Erasure Count Is Not an Affine-Invariant Concept Dimension

DGX agent

arXiv:2608.10566v1 Announce Type: cross Abstract: How many directions does a neural representation use to encode a concept? A common answer repeatedly erases probe directions and reports the stopping

researcharxiv-cs-cv
12 Aug 2026
Local Ai

Learning Gaussian Structure: Intervention-Guided Density Control for Feed-Forward Driving Reconstruction

DGX agent

arXiv:2608.11077v1 Announce Type: new Abstract: Feed-forward Gaussian reconstruction has recently emerged as an efficient approach for driving scene reconstruction. However, prevailing LiDAR-based met

local-aiarxiv-cs-cv
12 Aug 2026
Safety

Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen)

DGX agent

arXiv:2509.06191v2 Announce Type: replace-cross Abstract: Recent 3D generative models, which are capable of generating full object shapes from just a few images, now open up new opportunities in robot

safetyarxiv-cs-cv
12 Aug 2026
Research

LEGO: Leveled Language Gaussian Splatting

DGX agent

arXiv:2608.10057v1 Announce Type: new Abstract: We introduce LEGO for advanced open-vocabulary scene understanding. Beyond basic concept recognition, its core innovation lies in capturing the intrinsi

researcharxiv-cs-cv
12 Aug 2026
Local Ai

Lesion-Aware Adaptive Fourier Neural Operator for CT-to-PSMA PET Synthesis in Prostate Cancer

DGX agent

arXiv:2608.10429v1 Announce Type: new Abstract: Deep learning models that synthesize PET from CT or MRI can reduce patient dose and scanner demand, but are typically optimized with global losses such

local-aiarxiv-cs-cv
12 Aug 2026
Research

Logit Lens Supervision for Patch-Level Explanations in Vision-Language Models

DGX agent

arXiv:2602.01530v2 Announce Type: replace Abstract: Modern autoregressive Vision-Language Models (VLMs) can generate fluent answers while their visual-token representations become weakly tied to the i

researcharxiv-cs-cv
12 Aug 2026
Research

Longitudinal 3D Foundation Modeling for Neoadjuvant Breast Cancer Response Prediction from Serial DCE-MRI

DGX agent

arXiv:2608.09991v1 Announce Type: cross Abstract: Pathologic complete response (pCR) is an important endpoint in neoadjuvant chemotherapy (NAC) for breast cancer, and predicting pCR from imaging durin

researcharxiv-cs-cv
12 Aug 2026
Safety

LoRCA: LoRA Cycle Adaptation for Histology to HiP-CT Translation with DINOv3

DGX agent

arXiv:2608.10002v1 Announce Type: cross Abstract: Hierarchical Phase-Contrast Tomography (HiP-CT) is a synchrotron based X-ray imaging technique that enables non-destructive, volumetric imaging of int

safetyarxiv-cs-cv
12 Aug 2026
Model Releases

MAD-HOI: Masked Autoregressive Diffusion for Generating Articulated Hand Object Interactions from Text

DGX agent

arXiv:2608.10162v1 Announce Type: new Abstract: Methods for text-based generation of hand-object interaction (HOI) sequences primarily focus on producing smooth, physically plausible trajectories. A t

model-releasesarxiv-cs-cv
12 Aug 2026
Applications

MammoMix: Leveraging Mixture of Experts for Robust Mammogram Breast Detection

DGX agent

arXiv:2608.10437v1 Announce Type: new Abstract: Breast lesion detection in mammography remains a challenging task due to variations in image quality, lesion appearance, and population demographics acr

applicationsarxiv-cs-cv
12 Aug 2026
Research

Mixture-of-Experts-based Entropy Model for Learned Image Compression

DGX agent

arXiv:2608.10947v1 Announce Type: new Abstract: Learned image compression has seen significant progress in recent years with the development of end-to-end learned models that achieve better compressio

researcharxiv-cs-cv
12 Aug 2026
Research

MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understanding

DGX agent

arXiv:2608.10706v1 Announce Type: new Abstract: Recent vision-language models demonstrate impressive general visual understanding, yet their art interpretation remains shallow: they describe surface c

researcharxiv-cs-cv
12 Aug 2026
Model Releases

More Accurate, Less Human: Gestalt Grouping in Vision Models

DGX agent

arXiv:2608.10195v1 Announce Type: new Abstract: Human vision organizes what it sees into wholes: same-colored points group into series, similar marks cohere into categories, and shapes complete into r

model-releasesarxiv-cs-cv
12 Aug 2026
Research

Motion Artifact-Aware Self-Supervised Representation Learning for 3D Brain MRI Motion Artifact Reduction

DGX agent

arXiv:2608.10170v1 Announce Type: new Abstract: Patient motion remains a source of image degradation in brain MRI, leading to signal loss, blurring, and geometric distortion that compromise quantitati

researcharxiv-cs-cv
12 Aug 2026
Research

Multi-Level Evidence Aggregation for Robust Facial Phenotype Retrieval in Rare Genetic Disorder Prioritization

DGX agent

arXiv:2608.11037v1 Announce Type: new Abstract: AI-assisted facial phenotyping supports rare genetic disorder prioritization by retrieving visually similar diagnosed cases from facial image reference

researcharxiv-cs-cv
12 Aug 2026
Safety

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

DGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

safetyarxiv-cs-cv
12 Aug 2026
Research

Multimodal Ambivalence and Hesitancy Recognition via Cross-Attention and Gated Fusion

DGX agent

arXiv:2607.15779v2 Announce Type: replace Abstract: We present a multimodal framework for Ambivalence/Hesitancy (A/H) recognition in video, developed for the ABAW11 challenge at ECCV 2026. The propose

researcharxiv-cs-cv
12 Aug 2026
Research

Multiple Scale Latents for Learned Image Compression

DGX agent

arXiv:2608.10952v1 Announce Type: new Abstract: Most learned image compression systems rely on a single latent representation combined with a hyperprior, which limits their ability to efficiently capt

researcharxiv-cs-cv
12 Aug 2026
Model Releases

Neural Introspection Gating for Adaptive KV-Cache Reuse in Vision-Language-Action Models

DGX agent

arXiv:2608.10824v1 Announce Type: cross Abstract: Vision-Language-Action(VLA) models map camera images and language instructions directly to motor commands through a single autoregressive transformer.

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

NullEdit: Stealthy Image Protection via VLM Condition Redirection

DGX agent

arXiv:2608.10870v1 Announce Type: new Abstract: Modern image editors combine vision-language models (VLMs) with diffusion transformer backbones to modify a single reference image according to instruct

model-releasesarxiv-cs-cv
12 Aug 2026
Tutorials

Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMs

DGX agent

arXiv:2608.10959v1 Announce Type: new Abstract: Existing vision-language model (VLM) backdoors are usually treated as static vulnerabilities: one-to-one and N-to-N attacks bind one or more triggers to

tutorialsarxiv-cs-cv
12 Aug 2026
Local Ai

P3CA: Encoder-Agnostic Interpretation of Vision Foundation Model Embeddings via Spatial Probing

DGX agent

arXiv:2608.10131v1 Announce Type: new Abstract: Vision foundation models are increasingly used as reusable encoders in medical image computing, yet their high-dimensional spatial embeddings are diffic

local-aiarxiv-cs-cv
12 Aug 2026
Model Releases

PEAK: Precise and Persistent Concept Erasure via k-Sparse Autoencoders

DGX agent

arXiv:2608.10985v1 Announce Type: new Abstract: Erasing concepts from large-scale text-to-image (T2I) diffusion models has become increasingly crucial due to the growing concerns over copyright infrin

model-releasesarxiv-cs-cv
12 Aug 2026
Applications

PolyLayout: Hierarchical VLM-Guided Layout Generation Beyond Rectangular Rooms

DGX agent

arXiv:2608.10838v1 Announce Type: new Abstract: Generating physically plausible 3D room layouts is essential for home furnishing retail, enabling customers to visualize products in their own homes and

applicationsarxiv-cs-cv
12 Aug 2026
Research

PolypVision: A Three-Stage Hierarchical Deep Learning Framework for Classification and Segmentation of Colorectal Polyps

DGX agent

arXiv:2608.10649v1 Announce Type: new Abstract: Colorectal cancer (CRC) remains one of the leading causes of cancer-related mortality worldwide, predominantly arising from precancerous polyps. Accurat

researcharxiv-cs-cv
12 Aug 2026
Research

Pre- to Post-Contrast Synthesis of Breast DCE-MRI using Latent Bridge Matching

DGX agent

arXiv:2608.10000v1 Announce Type: cross Abstract: Dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI) is central to breast cancer imaging, but gadolinium administration increases scan burde

researcharxiv-cs-cv
12 Aug 2026
Applications

Precise Top-Layer Fabric Segmentation for Fabric Destacking with Edge- and Shape-Aware Deep Networks

DGX agent

arXiv:2608.10648v1 Announce Type: new Abstract: Fabric destacking requires precise segmentation of the topmost fabric layer, a task complicated by subtle fabric boundaries and high visual similarity b

applicationsarxiv-cs-cv
12 Aug 2026
Model Releases

PRMU: A Corpus-Free Benchmark for Person-Centric Knowledge Unlearning in Multimodal Large Language Models

DGX agent

arXiv:2608.11149v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in storing and recalling rich person-related knowledge, raising incre

model-releasesarxiv-cs-cv
12 Aug 2026
Agents

Protection Levels for Vision-Based Pose Estimation

DGX agent

arXiv:2608.10023v1 Announce Type: cross Abstract: Vision-based navigation complements Global Navigation Satellite Systems, but certification demands integrity guarantees that account for faulty measur

agentsarxiv-cs-cv
12 Aug 2026
Safety

Rethinking Data Efficiency in Industrial Dense Prediction: Pretraining Coherence, Not Inductive Bias, Determines ViTs Low-Data Advantage

DGX agent

arXiv:2608.10590v1 Announce Type: new Abstract: Vision Transformers (ViTs) are widely believed to require more labeled data than CNNs for industrial dense prediction. Through controlled experiments on

safetyarxiv-cs-cv
12 Aug 2026
Model Releases

Rethinking LLM Verification: Evidence Structure, Uncertainty, and Selective Refinement

DGX agent

arXiv:2608.10725v1 Announce Type: new Abstract: Large language models (LLMs) often rely on shortcuts rather than systematic reasoning, raising safety concerns in medical applications. Allowing models

model-releasesarxiv-cs-cv
12 Aug 2026
Applications

Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets

DGX agent

arXiv:2608.10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing sin

applicationsarxiv-cs-cv
12 Aug 2026
Research

Robustness of transferability estimation metrics for medical imaging

DGX agent

arXiv:2608.09999v1 Announce Type: cross Abstract: In transfer learning, the choice of source model largely influences the performance on a target dataset. Still, selecting a fitting source remains a c

researcharxiv-cs-cv
12 Aug 2026
Local Ai

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense

DGX agent

arXiv:2608.10933v1 Announce Type: new Abstract: Text-to-Video (T2V) generative models are vulnerable to jailbreak attacks in real-world deployment, leading them to produce harmful or inappropriate con

local-aiarxiv-cs-cv
12 Aug 2026
Safety

SapiensID 2.0: Aligning Human Recognition Foundation Models with Human Perception

DGX agent

arXiv:2608.10497v1 Announce Type: new Abstract: While foundation models have significantly advanced human recognition across diverse modalities, they predominantly rely on static, geometric feature ex

safetyarxiv-cs-cv
12 Aug 2026
Model Releases

SAR2Agri: Learning SAR Intensity Representations for Agricultural Monitoring

DGX agent

arXiv:2608.11142v1 Announce Type: new Abstract: Agricultural monitoring faces unique challenges, arising from the landscape's complex temporal, phenological, and climate dynamics, yet monitoring them

model-releasesarxiv-cs-cv
12 Aug 2026
Tutorials

SceneNAT: Masked Generative Modeling for Language-Guided Indoor Scene Synthesis

DGX agent

arXiv:2601.07218v2 Announce Type: replace Abstract: We present SceneNAT, a masked non-autoregressive Transformer for 3D indoor scene synthesis from natural language instructions. It generates complete

tutorialsarxiv-cs-cv
12 Aug 2026
Safety

SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural Networks

DGX agent

arXiv:2608.10289v1 Announce Type: new Abstract: Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-r

safetyarxiv-cs-cv
12 Aug 2026
Research

Self-Geometry: GT-Free and Plug-and-Play Test-Time Adaptation for Geometrically Consistent 3D Vision Foundation Models

DGX agent

arXiv:2608.10708v1 Announce Type: new Abstract: Recent Vision Foundation Models (VFMs) predict depth, camera pose, and pointmap in a single forward pass without per-scene optimization, achieving stron

researcharxiv-cs-cv
12 Aug 2026
Local Ai

Sensor-Informed Per-Point Covariance for Structured-Light 3D Imaging

DGX agent

arXiv:2608.10888v1 Announce Type: new Abstract: Per-point uncertainty models are important in structured-light 3D reconstruction for probabilistic registration, fusion, and quality assessment. In prac

local-aiarxiv-cs-cv
12 Aug 2026
Model Releases

Significance and Stability Analysis of Gene-Environment Interaction using GxEStat

DGX agent

arXiv:2604.03337v2 Announce Type: replace Abstract: Genotype-environment (GxE) interactions can influence the performance of genotypes across diverse environments, limiting the reliability of genotype

model-releasesarxiv-cs-cv
12 Aug 2026
← Previous
1234…259
Next →