AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,844 results
Safety

Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation

DGX agent

arXiv:2604.09231v1 Announce Type: new Abstract: Although recent advances have improved the quality of 3D texture generation, existing methods still struggle with incomplete texture coverage, cross-vie

safetyarxiv-cs-cv
13 Apr 2026
Tutorials
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All o…

DGX agent

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All of them are variations of GRPO and all of them are incrementa

tutorialsjeremy-howard--x
13 Apr 2026
Research

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

DGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

researcharxiv-cs-cv
13 Apr 2026
Tutorials

OmniPrism: Learning Disentangled Visual Concept for Image Generation

DGX agent

arXiv:2412.12242v2 Announce Type: replace-cross Abstract: Creative visual concept generation often draws inspiration from specific concepts in a reference image to produce relevant outcomes. However,

tutorialsarxiv-cs-ai
13 Apr 2026
Research

Overhang Tower: Resource-Rational Adaptation in Sequential Physical Planning

DGX agent

arXiv:2604.09072v1 Announce Type: new Abstract: Humans effortlessly navigate the physical world by predicting how objects behave under gravity and contact forces, yet how such judgments support sequen

researcharxiv-cs-ai
13 Apr 2026
Agents

RIRF: Reasoning Image Restoration Framework

DGX agent

arXiv:2604.09511v1 Announce Type: new Abstract: Universal image restoration (UIR) aims to recover clean images from diverse and unknown degradations using a unified model. Existing UIR methods primari

agentsarxiv-cs-cv
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Agents

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

DGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

agentsarxiv-cs-ro
13 Apr 2026
Model Releases

struggling choosing one edit model from klein 9b or qwen 2511.

DGX agent

This r/StableDiffusion thread discusses the community debate around choosing between FLUX.2 [klein] 9B and Qwen Image Edit 2511 as an image editing model, two strong open-source contenders in the spac

model-releasesr-stablediffusion
11 Apr 2026
Tutorials

The one thing I still don't know how to do: TTS/singing a specific song but with a specific voice

DGX agent

The specific Reddit post could not be retrieved from the search results. However, based on the context of the URL and related results, I can provide the following best-effort summary based on what ...

tutorialsr-stablediffusion
11 Apr 2026
Model Releases

Advanced inpaint/edit Klein/Qwen workflows

DGX agent

A Reddit post on r/StableDiffusion discussing advanced ComfyUI workflows that combine the FLUX Klein and Qwen Image Edit models for precision inpainting and image editing tasks. FLUX Klein offers ...

model-releasesr-stablediffusion
10 Apr 2026
Model Releases

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

DGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Balanced Diffusion-Guided Fusion for Multimodal Remote Sensing Classification

DGX agent

arXiv:2509.23310v3 Announce Type: replace Abstract: Deep learning-based techniques for the analysis of multimodal remote sensing data have become popular due to their ability to effectively integrate

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

DGX agent

arXiv:2604.08287v1 Announce Type: new Abstract: Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings.

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Distilling Specialized Orders for Visual Generation

DGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

researcharxiv-cs-cv
10 Apr 2026
Research

DMin: Scalable Training Data Influence Estimation for Diffusion Models

DGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

researcharxiv-cs-cv
10 Apr 2026
Research

DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction

DGX agent

arXiv:2604.07986v1 Announce Type: new Abstract: Egocentric video is crucial for next-generation 4D scene reconstruction, with applications in AR/VR and embodied AI. However, reconstructing dynamic fir

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Evaluating Low-Light Image Enhancement Across Multiple Intensity Levels

DGX agent

arXiv:2511.15496v2 Announce Type: replace Abstract: Imaging in low-light environments is challenging due to reduced scene radiance, which leads to elevated sensor noise and reduced color saturation. M

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration

DGX agent

arXiv:2603.16570v2 Announce Type: replace Abstract: Recent advances in image restoration have enabled high-fidelity recovery of faces from degraded inputs using reference-based face restoration models

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

DGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement

DGX agent

arXiv:2507.08390v4 Announce Type: replace Abstract: Discrete diffusion models have recently emerged as strong alternatives to autoregressive language models, matching their performance through large-s

tutorialsarxiv-cs-lg
10 Apr 2026
Research

LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models

DGX agent

arXiv:2512.17489v2 Announce Type: replace Abstract: Text-to-image (T2I) models have demonstrated remarkable progress in creative image generation, yet they still lack precise control over scene illumi

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

DGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

model-releasesarxiv-cs-lg
10 Apr 2026
Local Ai

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models

DGX agent

arXiv:2601.04068v3 Announce Type: replace Abstract: Aligning text-to-video diffusion models with human preferences is crucial for generating high-quality videos. Existing Direct Preference Otimization

local-aiarxiv-cs-cv
10 Apr 2026
Tutorials

MoRight: Motion Control Done Right

DGX agent

arXiv:2604.07348v1 Announce Type: cross Abstract: Generating motion-controlled videos--where user-specified actions drive physically plausible scene dynamics under freely chosen viewpoints--demands tw

tutorialsarxiv-cs-ai
10 Apr 2026
Applications

MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation

DGX agent

arXiv:2603.11633v2 Announce Type: replace Abstract: Recent unified 3D generation models have made remarkable progress in producing high-quality 3D assets from a single image. Notably, layout-aware app

applicationsarxiv-cs-cv
10 Apr 2026
Model Releases

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

DGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Physical Knot Classification Beyond Accuracy: A Benchmark and Diagnostic Study

DGX agent

arXiv:2603.23286v3 Announce Type: replace Abstract: Physical knot classification is a fine-grained task in which the intended cue is rope crossing structure, but high accuracy may still come from appe

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

DGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards

DGX agent

arXiv:2512.01236v2 Announce Type: replace Abstract: Personalized generation models for a single subject have demonstrated remarkable effectiveness, highlighting their significant potential. However, w

model-releasesarxiv-cs-cv
10 Apr 2026
Research

ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video

DGX agent

arXiv:2604.07882v1 Announce Type: new Abstract: Reconstructing non-rigid objects with physical plausibility remains a significant challenge. Existing approaches leverage differentiable rendering for p

researcharxiv-cs-cv
10 Apr 2026
Model Releases

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

DGX agent

arXiv:2604.08211v1 Announce Type: new Abstract: Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Self-Improving 4D Perception via Self-Distillation

DGX agent

arXiv:2604.08532v1 Announce Type: new Abstract: Large-scale multi-view reconstruction models have made remarkable progress, but most existing approaches still rely on fully supervised training with gr

researcharxiv-cs-cv
10 Apr 2026
Research

SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning

DGX agent

arXiv:2604.06636v1 Announce Type: cross Abstract: Process supervision has emerged as a promising approach for enhancing LLM reasoning, yet existing methods fail to distinguish meaningful progress from

researcharxiv-cs-ai
10 Apr 2026
Research

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

DGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

researcharxiv-cs-cl
10 Apr 2026
Agents

Studying Sutton and Barto's RL book and its connections to RL for LLMs (e.g., tool use, math reasoning, agents, and so on)? [D]

DGX agent

A Reddit discussion thread on r/MachineLearning in which practitioners explore how foundational concepts from Sutton and Barto's *Reinforcement Learning: An Introduction* — including MDPs, policy g...

agentsr-machinelearning
9 Apr 2026
Model Releases

PS: I finally got around to trying out @randal_olson 's Tufte Test tool to prettify the benchmark plot. Great tool 👌! https://www.goodeyela…

DGX agent

Sebastian Raschka (rasbt) used Randal Olson's Tufte Test tool, developed by Goodeye Labs, to improve the visual quality of a machine learning benchmark plot. The Tufte Test encodes seven of Tufte'...

model-releasessebastian-raschka--x
8 Apr 2026
Research

A Generalized Theory of Load Distribution in Redundantly-actuated Robotic Systems

DGX agent

arXiv:2603.11431v2 Announce Type: replace Abstract: This paper presents a generalized theory which describes how applied loads are distributed within rigid bodies handled by redundantly-actuated robot

researcharxiv-cs-ro
13 Aug 2026
Model Releases

Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence

DGX agent

arXiv:2608.12290v1 Announce Type: cross Abstract: Modern black-box Image-to-Video (I2V) models offer powerful capabilities in automated content creation, yet their lack of fine-grained control and rel

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

D3D-GEN: Robot-Aware Domain-Grounded Interactive 3D World Generation for Social Robotics

DGX agent

arXiv:2608.11876v1 Announce Type: new Abstract: Training and validation of Embodied AI for social navigation critically depends on realistic simulation environments, yet many current approaches fail t

agentsarxiv-cs-ro
13 Aug 2026
Agents

Diffusion Probe: Generated Image Result Prediction Using CNN Probes

DGX agent

arXiv:2602.23783v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models lack an efficient mechanism for early quality assessment, leading to costly trial-and-error in multi-generation

agentsarxiv-cs-cv
13 Aug 2026
Model Releases

Explainability in Practice: A Survey of Explainable NLP Across Various Domains

DGX agent

arXiv:2502.00837v3 Announce Type: replace-cross Abstract: Natural Language Processing (NLP) is now embedded in critical sectors including healthcare, finance, and customer relationship management, whe

model-releasesarxiv-cs-ai
13 Aug 2026
Research

First-order friction models with bristle dynamics: lumped and distributed formulations

DGX agent

arXiv:2602.09429v3 Announce Type: replace-cross Abstract: Dynamic models, particularly rate-dependent models, have proven effective in capturing the key phenomenological features of frictional process

researcharxiv-cs-ro
13 Aug 2026
Research

From Monolithic to Modular: Segment-level Automatic Prompt Optimization

DGX agent

arXiv:2608.11219v1 Announce Type: new Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others. We present SAPO, a seg

researcharxiv-cs-ai
13 Aug 2026
Research

Generation as Auxiliary Supervision: Enhancing Visual Understanding at Zero Inference Overhead via Decoupled Embedding Prediction

DGX agent

arXiv:2608.12209v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have achieved remarkable progress, visual understanding and generation are typically treated as divergent

researcharxiv-cs-cv
13 Aug 2026
Agents

GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Settings

DGX agent

arXiv:2608.12133v1 Announce Type: new Abstract: Enterprise guideline documents are heterogeneous and multimodal, combining narrative text, complex tables, and embedded images. Existing LLM and VLM sys

agentsarxiv-cs-ai
13 Aug 2026
Tutorials

How Organizations Use AI: Evidence from ChatGPT

DGX agent

arXiv:2608.12236v1 Announce Type: cross Abstract: We study how organizations use frontier generative AI by linking ChatGPT Enterprise account records to usage, worker roles, task classifications, and

tutorialsarxiv-cs-ai
13 Aug 2026
Model Releases

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

DGX agent

arXiv:2608.11616v1 Announce Type: new Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to

model-releasesarxiv-cs-ai
13 Aug 2026
← Previous
1…2728293031…60
Next →