AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cv”

GridTimelineEvolution
12,515 results
Safety

Improving Model Safety by Targeted Error Correction

DGX agent

arXiv:2605.02544v1 Announce Type: cross Abstract: The widespread adoption of machine learning in critical applications demands techniques to mitigate high-consequence errors. Our method utilizes a dua

safetyarxiv-cs-cv
5 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

InfiltrNet: Dual-Branch CNN-Transformer Architecture for Brain Tumor Infiltration Risk Prediction

DGX agent

arXiv:2605.02230v1 Announce Type: new Abstract: Gliomas are aggressive brain tumors that infiltrate surrounding tissue beyond the visible tumor margins observed on Magnetic Resonance Imaging (MRI). Pr

researcharxiv-cs-cv
5 May 2026
Hardware

InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation

DGX agent

arXiv:2512.08309v4 Announce Type: replace Abstract: For decades, procedural worlds have been built on procedural noise functions such as Perlin noise, which are fast and infinite, yet fundamentally li

hardwarearxiv-cs-cv
5 May 2026
Model Releases

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation

DGX agent

arXiv:2512.21788v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning of Diffusion Transformers (DiTs) for diverse, multi-conditional tasks often suffers from task interference when usin

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Interactive Multi-Turn Retrieval for Health Videos

DGX agent

arXiv:2605.01409v1 Announce Type: cross Abstract: The growing availability of health-related instructional videos creates new opportunities for clinical training, patient rehabilitation, and health ed

model-releasesarxiv-cs-cv
5 May 2026
Research

Interlaced R2D2 DNN Series for Scalable Non-Cartesian MRI with Sensitivity Self-calibration

DGX agent

arXiv:2503.09559v3 Announce Type: replace-cross Abstract: We introduce interlaced R2D2 (iR2D2), a DNN series paradigm for scalable image reconstruction from accelerated non-Cartesian k-space acquisiti

researcharxiv-cs-cv
5 May 2026
Model Releases

InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene

DGX agent

arXiv:2605.01036v1 Announce Type: new Abstract: This paper tackles the problem of physics-aware human motion synthesis in a dynamic scene. Unlike existing works which mainly tend to generate physicall

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

Intervention-Based Self-Supervised Learning: A Causal Probe Paradigm for Remote Photoplethysmography

DGX agent

arXiv:2605.00882v1 Announce Type: new Abstract: Remote Photoplethysmography (rPPG) enables convenient non-contact physiological measurement. Existing Self-Supervised Learning (SSL) methods commonly fa

tutorialsarxiv-cs-cv
5 May 2026
Safety

Investigating Anthropometric Fidelity in SAM 3D Body

DGX agent

arXiv:2601.06035v2 Announce Type: replace-cross Abstract: The release of SAM 3D Body is a recent development in human mesh recovery, demonstrating improved performance in producing clean, topologicall

safetyarxiv-cs-cv
5 May 2026
Model Releases

Joint Architecture-Token-Bitwidth Multi-Axis Optimization of Vision Transformers for Semiconductor IC Packaging

DGX agent

arXiv:2605.01742v1 Announce Type: new Abstract: Vision Transformers (ViTs) have achieved strong performance in visual recognition, yet their deployment in resource-constrained industrial environments

model-releasesarxiv-cs-cv
5 May 2026
Research

Know Yourself Better: Diverse Object-Related Features Improve Open Set Recognition

DGX agent

arXiv:2404.10370v3 Announce Type: replace Abstract: Open set recognition (OSR) is a critical aspect of machine learning, addressing the challenge of detecting novel classes during inference. Within th

researcharxiv-cs-cv
5 May 2026
Model Releases

LabBuilder: Protocol-Grounded 3D Layout Generation for Interactable and Safe Laboratory

DGX agent

arXiv:2605.02288v1 Announce Type: new Abstract: Automated laboratories hold the promise of accelerating scientific discovery, yet their deployment is bottlenecked by the difficulty of designing safe a

model-releasesarxiv-cs-cv
5 May 2026
Applications

Laplacian Frequency Interaction Network for Rural Thematic Road Extraction

DGX agent

arXiv:2605.02866v1 Announce Type: new Abstract: Rural thematic road network construction aims to extract topological road structures from movement trajectory images of agricultural machinery. However,

applicationsarxiv-cs-cv
5 May 2026
Research

Latent Space Probing for Adult Content Detection in Video Generative Models

DGX agent

arXiv:2605.00874v1 Announce Type: new Abstract: The rapid proliferation of AI-powered video generation systems has introduced significant challenges in content moderation, particularly with respect to

researcharxiv-cs-cv
5 May 2026
Model Releases

LatentDiff: Scaling Semantic Dataset Comparison to Millions of Images

DGX agent

arXiv:2605.00899v1 Announce Type: new Abstract: We present LatentDiff, a scalable framework for semantic dataset comparison that operates directly in the latent space of pretrained vision encoders. By

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

Learning Equivariant Neural-Augmented Object Dynamics From Few Interactions

DGX agent

arXiv:2605.02699v1 Announce Type: cross Abstract: Learning data-efficient object dynamics models for robotic manipulation remains challenging, especially for deformable objects. A popular approach is

tutorialsarxiv-cs-cv
5 May 2026
Research

Learning to Place Objects with Programs and Iterative Self Training

DGX agent

arXiv:2503.04496v2 Announce Type: replace-cross Abstract: In this work we study indoor scene object placement. Given a 3D indoor scene and an object, the task is to predict placement locations within

researcharxiv-cs-cv
5 May 2026
Safety

Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy

DGX agent

arXiv:2509.21173v5 Announce Type: replace Abstract: Vision-Language Models (VLMs) such as CLIP have revolutionized zero-shot classification and safety-critical tasks, including Out-of-Distribution (OO

safetyarxiv-cs-cv
5 May 2026
Model Releases

Leveraging Imperfect Medical Data: A Manifold-Consistent Spatio-Temporal Network for Sensor-based Human Activity Recognition

DGX agent

arXiv:2605.00913v1 Announce Type: new Abstract: Sensor-based Human Activity Recognition (HAR) has attracted increasing attention in medical and healthcare monitoring, particularly with the growth of I

model-releasesarxiv-cs-cv
5 May 2026
Research

LGDWT-GS: Local and Global Discrete Wavelet-Regularized 3D Gaussian Splatting for Sparse-View Scene Reconstruction

DGX agent

arXiv:2601.17185v2 Announce Type: replace Abstract: We propose a new method for few-shot 3D reconstruction that integrates global and local frequency regularization to stabilize geometry and preserve

researcharxiv-cs-cv
5 May 2026
Agents

LIE: LiDAR-only HD Map Construction with Intensity Enhancement via Online Knowledge Distillation

DGX agent

arXiv:2605.01478v1 Announce Type: new Abstract: Online High-Definition (HD) map construction is a key component of autonomous driving. Recent methods rely on multi-view camera images for cost-effectiv

agentsarxiv-cs-cv
5 May 2026
Tutorials

Limited-Angle Tomography Reconstruction via Projector Guided 3D Diffusion

DGX agent

arXiv:2510.06516v2 Announce Type: replace Abstract: Limited-angle electron tomography aims to reconstruct 3D shapes from 2D projections of Transmission Electron Microscopy (TEM) within a restricted ra

tutorialsarxiv-cs-cv
5 May 2026
Model Releases

Linear-Time Global Visual Modeling without Explicit Attention

DGX agent

arXiv:2605.01711v1 Announce Type: new Abstract: Existing research largely attributes the global sequence modeling capability of Transformers to the explicit computation of attention weights, a process

model-releasesarxiv-cs-cv
5 May 2026
Local Ai

Linearizing Vision Transformer with Test-Time Training

DGX agent

arXiv:2605.02772v1 Announce Type: new Abstract: While linear-complexity attention mechanisms offer a promising alternative to Softmax attention for overcoming the quadratic bottleneck, training such m

local-aiarxiv-cs-cv
5 May 2026
Safety

Linking spatial biology and clinical histology via Haiku

DGX agent

arXiv:2605.00925v1 Announce Type: cross Abstract: Integrating molecular, morphological, and clinical data is essential for basic and translational biomedical research, yet systematic frameworks for jo

safetyarxiv-cs-cv
5 May 2026
Research

LinMU: Multimodal Understanding Made Linear

DGX agent

arXiv:2601.01322v2 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) achieve impressive performance but are limited by the quadratic complexity of self-attention, which prevents th

researcharxiv-cs-cv
5 May 2026
Model Releases

LiteVLA-H: Dual-Rate Vision-Language-Action Inference for Onboard Aerial Guidance and Semantic Perception

DGX agent

arXiv:2605.00884v1 Announce Type: new Abstract: Vision-language-action (VLA) models have shown strong semantic grounding and task generalization in manipulation, but aerial deployment remains difficul

model-releasesarxiv-cs-cv
5 May 2026
Research

Low-Latency Embedded Driver Monitoring System with a Multi-Task Neural Network

DGX agent

arXiv:2605.02563v1 Announce Type: new Abstract: Road traffic accidents remain a significant global concern, with the majority attributed to human factors such as driver distraction and fatigue. This s

researcharxiv-cs-cv
5 May 2026
Safety

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance

DGX agent

arXiv:2410.18717v2 Announce Type: replace Abstract: Recent advancements in artificial intelligence hold ample potential for monitoring applications using surveillance cameras. However, concerns about

safetyarxiv-cs-cv
5 May 2026
Safety

LVLM-Aided Alignment of Task-Specific Vision Models

DGX agent

arXiv:2512.21985v2 Announce Type: replace Abstract: In high-stakes domains, small task-specific vision models are crucial due to their low computational requirements and the availability of numerous m

safetyarxiv-cs-cv
5 May 2026
Model Releases

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE

DGX agent

arXiv:2605.02641v1 Announce Type: new Abstract: We present Mamoda2.5, a unified AR-Diffusion framework that seamlessly integrates multimodal understanding and generation within a single architecture.

model-releasesarxiv-cs-cv
5 May 2026
Research

Manifold-Aligned Guided Integrated Gradients for Reliable Feature Attribution

DGX agent

arXiv:2605.02167v1 Announce Type: cross Abstract: Feature attribution is central to diagnosing and trusting deep neural networks, and Integrated Gradients (IG) is widely used due to its axiomatic prop

researcharxiv-cs-cv
5 May 2026
Agents

MapRF: Weakly Supervised Online HD Map Construction via NeRF-Guided Self-Training

DGX agent

arXiv:2511.19527v2 Announce Type: replace Abstract: Autonomous driving systems benefit from high-definition (HD) maps that provide critical information about road infrastructure. The online constructi

agentsarxiv-cs-cv
5 May 2026
Agents

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

DGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

agentsarxiv-cs-cv
5 May 2026
Applications

MER-DG: Modality-Entropy Regularization for Multimodal Domain Generalization

DGX agent

arXiv:2605.01967v1 Announce Type: cross Abstract: Deploying multimodal models in real-world scenarios requires generalization to new environments where recording conditions differ from training, a cha

applicationsarxiv-cs-cv
5 May 2026
Model Releases

Metric Unreliability in Multimodal Machine Unlearning: A Systematic Analysis and Principled Unified Score

DGX agent

arXiv:2605.02206v1 Announce Type: new Abstract: Machine unlearning in Vision-Language Models (VLMs) is required for compliance with the General Data Protection Regulation (GDPR), yet current evaluatio

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Mextsuperscript{4}Fuse: Lightweight State-Space MoE with a Cross-Scale Gating Bridge for Brain Tumor Segmentation

DGX agent

arXiv:2605.02444v1 Announce Type: new Abstract: Encoder-decoder imbalance and the reliance on large input volumes make many 3D brain tumor segmentation models both compute-heavy and brittle. We presen

model-releasesarxiv-cs-cv
5 May 2026
Research

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

DGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

researcharxiv-cs-cv
5 May 2026
Research

Mixture Prototype Flow Matching for Open-Set Supervised Anomaly Detection

DGX agent

arXiv:2605.02438v1 Announce Type: new Abstract: Open-set supervised anomaly detection (OSAD) aims to identify unseen anomalies using limited anomalous supervision. However, existing prototype-based me

researcharxiv-cs-cv
5 May 2026
Safety

MOC-3D: Manifold-Order Consistency for Text-to-3D Generation

DGX agent

arXiv:2605.01743v1 Announce Type: new Abstract: With the burgeoning development of fields such as the Metaverse, Virtual Reality (VR), and Digital Twins, text-to-3D generation has emerged as a researc

safetyarxiv-cs-cv
5 May 2026
Model Releases

MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation

DGX agent

arXiv:2506.05952v4 Announce Type: replace Abstract: Recent advances in transformer-based text-to-motion generation have led to impressive progress in synthesizing high-quality human motion. Neverthele

model-releasesarxiv-cs-cv
5 May 2026
Safety

Momentum-Anchored Multi-Scale Fusion Model for Long-Tailed Chest X-Ray Classification

DGX agent

arXiv:2605.02292v1 Announce Type: new Abstract: Chest X-ray classification suffers from severe class imbalance where gradient updates bias toward majority classes, causing feature drift and poor perfo

safetyarxiv-cs-cv
5 May 2026
Research

MooD: An Efficient VA-Driven Affective Image Editing Framework via Fine-Grained Semantic Control

DGX agent

arXiv:2605.02521v1 Announce Type: new Abstract: Affective image editing (AIE) aims to edit visual content to evoke target emotions. However, existing methods often overlook inference efficiency and pr

researcharxiv-cs-cv
5 May 2026
Research

Motion-Aware Caching for Efficient Autoregressive Video Generation

DGX agent

arXiv:2605.01725v1 Announce Type: new Abstract: Autoregressive video generation paradigms offer theoretical promise for long video synthesis, yet their practical deployment is hindered by the computat

researcharxiv-cs-cv
5 May 2026
Research

Multi-Branch Non-Homogeneous Image Dehazing via Concentration Partitioning and Image Fusion

DGX agent

arXiv:2605.00885v1 Announce Type: new Abstract: Existing single image dehazing methods have demonstrated satisfactory performance on homogeneous thin-haze images; however, they often struggle with non

researcharxiv-cs-cv
5 May 2026
Research

Multi-Dataset Cross-Domain Knowledge Distillation for Unified Medical Image Segmentation, Classification, and Detection

DGX agent

arXiv:2605.01563v1 Announce Type: new Abstract: We propose a unified cross-domain transfer learning framework that leverages knowledge from multiple heterogeneous medical imaging datasets to improve p

researcharxiv-cs-cv
5 May 2026
Research

Multi-Rater Calibrated Segmentation Models

DGX agent

arXiv:2605.02437v1 Announce Type: new Abstract: Objective: Accurate probability estimates are essential for the safe deployment of medical image segmentation models in clinical decision-making. Howeve

researcharxiv-cs-cv
5 May 2026
Safety

Multi-Scale Gaussian-Language Map for Zero-shot Embodied Navigation and Reasoning

DGX agent

arXiv:2605.01736v1 Announce Type: new Abstract: Understanding the geometric and semantic structure of environments is essential for embodied navigation and reasoning. Existing semantic mapping methods

safetyarxiv-cs-cv
5 May 2026
← Previous
1…196197198199200…261
Next →