AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,737 results
Research

BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training

DGX agent

arXiv:2604.09022v1 Announce Type: new Abstract: With the rapid adoption of diffusion models, synthetic data generation has emerged as a promising approach for addressing the growing demand for large-s

researcharxiv-cs-cv
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CatalogStitch: Dimension-Aware and Occlusion-Preserving Object Compositing for Catalog Image Generation

DGX agent

arXiv:2604.08836v1 Announce Type: new Abstract: Generative object compositing methods have shown remarkable ability to seamlessly insert objects into scenes. However, when applied to real-world catalo

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Dual Mamba for Node-Specific Representation Learning: Tackling Over-Smoothing with Selective State Space Modeling

DGX agent

arXiv:2511.06756v3 Announce Type: replace Abstract: Over-smoothing remains a fundamental challenge in deep Graph Neural Networks (GNNs), where repeated message passing causes node representations to b

researcharxiv-cs-lg
13 Apr 2026
Model Releases

ELT: Elastic Looped Transformers for Visual Generation

DGX agent

arXiv:2604.09168v1 Announce Type: new Abstract: We introduce Elastic Looped Transformers (ELT), a highly parameter-efficient class of visual generative models based on a recurrent transformer architec

model-releasesarxiv-cs-cv
13 Apr 2026
Agents

Gen-n-Val: Agentic Image Data Generation and Validation

DGX agent

arXiv:2506.04676v2 Announce Type: replace-cross Abstract: The data scarcity, label noise, and long-tailed category imbalance remain important and unresolved challenges in many computer vision tasks, s

agentsarxiv-cs-ai
13 Apr 2026
Safety

Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation

DGX agent

arXiv:2604.09231v1 Announce Type: new Abstract: Although recent advances have improved the quality of 3D texture generation, existing methods still struggle with incomplete texture coverage, cross-vie

safetyarxiv-cs-cv
13 Apr 2026
Research

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

DGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

researcharxiv-cs-cv
13 Apr 2026
Tutorials

OmniPrism: Learning Disentangled Visual Concept for Image Generation

DGX agent

arXiv:2412.12242v2 Announce Type: replace-cross Abstract: Creative visual concept generation often draws inspiration from specific concepts in a reference image to produce relevant outcomes. However,

tutorialsarxiv-cs-ai
13 Apr 2026
Research

Overhang Tower: Resource-Rational Adaptation in Sequential Physical Planning

DGX agent

arXiv:2604.09072v1 Announce Type: new Abstract: Humans effortlessly navigate the physical world by predicting how objects behave under gravity and contact forces, yet how such judgments support sequen

researcharxiv-cs-ai
13 Apr 2026
Agents

RIRF: Reasoning Image Restoration Framework

DGX agent

arXiv:2604.09511v1 Announce Type: new Abstract: Universal image restoration (UIR) aims to recover clean images from diverse and unknown degradations using a unified model. Existing UIR methods primari

agentsarxiv-cs-cv
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Agents

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

DGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

agentsarxiv-cs-ro
13 Apr 2026
Model Releases

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

DGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Balanced Diffusion-Guided Fusion for Multimodal Remote Sensing Classification

DGX agent

arXiv:2509.23310v3 Announce Type: replace Abstract: Deep learning-based techniques for the analysis of multimodal remote sensing data have become popular due to their ability to effectively integrate

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

DGX agent

arXiv:2604.08287v1 Announce Type: new Abstract: Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings.

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Distilling Specialized Orders for Visual Generation

DGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

researcharxiv-cs-cv
10 Apr 2026
Research

DMin: Scalable Training Data Influence Estimation for Diffusion Models

DGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

researcharxiv-cs-cv
10 Apr 2026
Research

DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction

DGX agent

arXiv:2604.07986v1 Announce Type: new Abstract: Egocentric video is crucial for next-generation 4D scene reconstruction, with applications in AR/VR and embodied AI. However, reconstructing dynamic fir

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Evaluating Low-Light Image Enhancement Across Multiple Intensity Levels

DGX agent

arXiv:2511.15496v2 Announce Type: replace Abstract: Imaging in low-light environments is challenging due to reduced scene radiance, which leads to elevated sensor noise and reduced color saturation. M

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration

DGX agent

arXiv:2603.16570v2 Announce Type: replace Abstract: Recent advances in image restoration have enabled high-fidelity recovery of faces from degraded inputs using reference-based face restoration models

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

DGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement

DGX agent

arXiv:2507.08390v4 Announce Type: replace Abstract: Discrete diffusion models have recently emerged as strong alternatives to autoregressive language models, matching their performance through large-s

tutorialsarxiv-cs-lg
10 Apr 2026
Research

LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models

DGX agent

arXiv:2512.17489v2 Announce Type: replace Abstract: Text-to-image (T2I) models have demonstrated remarkable progress in creative image generation, yet they still lack precise control over scene illumi

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

DGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

model-releasesarxiv-cs-lg
10 Apr 2026
Local Ai

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models

DGX agent

arXiv:2601.04068v3 Announce Type: replace Abstract: Aligning text-to-video diffusion models with human preferences is crucial for generating high-quality videos. Existing Direct Preference Otimization

local-aiarxiv-cs-cv
10 Apr 2026
Tutorials

MoRight: Motion Control Done Right

DGX agent

arXiv:2604.07348v1 Announce Type: cross Abstract: Generating motion-controlled videos--where user-specified actions drive physically plausible scene dynamics under freely chosen viewpoints--demands tw

tutorialsarxiv-cs-ai
10 Apr 2026
Applications

MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation

DGX agent

arXiv:2603.11633v2 Announce Type: replace Abstract: Recent unified 3D generation models have made remarkable progress in producing high-quality 3D assets from a single image. Notably, layout-aware app

applicationsarxiv-cs-cv
10 Apr 2026
Model Releases

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

DGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Physical Knot Classification Beyond Accuracy: A Benchmark and Diagnostic Study

DGX agent

arXiv:2603.23286v3 Announce Type: replace Abstract: Physical knot classification is a fine-grained task in which the intended cue is rope crossing structure, but high accuracy may still come from appe

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

DGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards

DGX agent

arXiv:2512.01236v2 Announce Type: replace Abstract: Personalized generation models for a single subject have demonstrated remarkable effectiveness, highlighting their significant potential. However, w

model-releasesarxiv-cs-cv
10 Apr 2026
Research

ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video

DGX agent

arXiv:2604.07882v1 Announce Type: new Abstract: Reconstructing non-rigid objects with physical plausibility remains a significant challenge. Existing approaches leverage differentiable rendering for p

researcharxiv-cs-cv
10 Apr 2026
Model Releases

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

DGX agent

arXiv:2604.08211v1 Announce Type: new Abstract: Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Self-Improving 4D Perception via Self-Distillation

DGX agent

arXiv:2604.08532v1 Announce Type: new Abstract: Large-scale multi-view reconstruction models have made remarkable progress, but most existing approaches still rely on fully supervised training with gr

researcharxiv-cs-cv
10 Apr 2026
Research

SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning

DGX agent

arXiv:2604.06636v1 Announce Type: cross Abstract: Process supervision has emerged as a promising approach for enhancing LLM reasoning, yet existing methods fail to distinguish meaningful progress from

researcharxiv-cs-ai
10 Apr 2026
Research

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

DGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

researcharxiv-cs-cl
10 Apr 2026
Research

A Generalized Theory of Load Distribution in Redundantly-actuated Robotic Systems

DGX agent

arXiv:2603.11431v2 Announce Type: replace Abstract: This paper presents a generalized theory which describes how applied loads are distributed within rigid bodies handled by redundantly-actuated robot

researcharxiv-cs-ro
13 Aug 2026
Model Releases

Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence

DGX agent

arXiv:2608.12290v1 Announce Type: cross Abstract: Modern black-box Image-to-Video (I2V) models offer powerful capabilities in automated content creation, yet their lack of fine-grained control and rel

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

D3D-GEN: Robot-Aware Domain-Grounded Interactive 3D World Generation for Social Robotics

DGX agent

arXiv:2608.11876v1 Announce Type: new Abstract: Training and validation of Embodied AI for social navigation critically depends on realistic simulation environments, yet many current approaches fail t

agentsarxiv-cs-ro
13 Aug 2026
Agents

Diffusion Probe: Generated Image Result Prediction Using CNN Probes

DGX agent

arXiv:2602.23783v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models lack an efficient mechanism for early quality assessment, leading to costly trial-and-error in multi-generation

agentsarxiv-cs-cv
13 Aug 2026
Model Releases

Explainability in Practice: A Survey of Explainable NLP Across Various Domains

DGX agent

arXiv:2502.00837v3 Announce Type: replace-cross Abstract: Natural Language Processing (NLP) is now embedded in critical sectors including healthcare, finance, and customer relationship management, whe

model-releasesarxiv-cs-ai
13 Aug 2026
Research

First-order friction models with bristle dynamics: lumped and distributed formulations

DGX agent

arXiv:2602.09429v3 Announce Type: replace-cross Abstract: Dynamic models, particularly rate-dependent models, have proven effective in capturing the key phenomenological features of frictional process

researcharxiv-cs-ro
13 Aug 2026
Research

From Monolithic to Modular: Segment-level Automatic Prompt Optimization

DGX agent

arXiv:2608.11219v1 Announce Type: new Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others. We present SAPO, a seg

researcharxiv-cs-ai
13 Aug 2026
Research

Generation as Auxiliary Supervision: Enhancing Visual Understanding at Zero Inference Overhead via Decoupled Embedding Prediction

DGX agent

arXiv:2608.12209v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have achieved remarkable progress, visual understanding and generation are typically treated as divergent

researcharxiv-cs-cv
13 Aug 2026
Agents

GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Settings

DGX agent

arXiv:2608.12133v1 Announce Type: new Abstract: Enterprise guideline documents are heterogeneous and multimodal, combining narrative text, complex tables, and embedded images. Existing LLM and VLM sys

agentsarxiv-cs-ai
13 Aug 2026
Tutorials

How Organizations Use AI: Evidence from ChatGPT

DGX agent

arXiv:2608.12236v1 Announce Type: cross Abstract: We study how organizations use frontier generative AI by linking ChatGPT Enterprise account records to usage, worker roles, task classifications, and

tutorialsarxiv-cs-ai
13 Aug 2026
Model Releases

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

DGX agent

arXiv:2608.11616v1 Announce Type: new Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Multi-Agent Target-Existence Verification and Learned Mask Geometry Refinement: Winning Report of the MeViS-Text Track at the 8th LSVOS Challenge 2026

DGX agent

arXiv:2608.11458v1 Announce Type: new Abstract: We present the first-place solution to the MeViS-Text track of the 8th Large-scale Video Object Segmentation (LSVOS) Challenge 2026: referring video obj

agentsarxiv-cs-cv
13 Aug 2026
← Previous
1…2627282930…58
Next →