AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,844 results
Applications

ReaDy-Go: Real-to-Sim Dynamic 3D Gaussian Splatting Simulation for Environment-Specific Visual Navigation with Moving Obstacles

DGX agent

arXiv:2602.11575v3 Announce Type: replace-cross Abstract: Visual navigation models often struggle in real-world dynamic environments due to limited robustness to the sim-to-real gap and the difficulty

applicationsarxiv-cs-cv
25 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SSMNBench: Diagnosing Image-based Cross-View Human-Object Understanding via Single-View Sufficiency and Multi-View Necessity

DGX agent

arXiv:2606.25634v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in single-image perception, yet their ability to reason about complex cross-view

model-releasesarxiv-cs-cv
25 Jun 2026
Research

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms

DGX agent

arXiv:2606.25450v1 Announce Type: cross Abstract: Traditional evaluations measure a learning algorithm's final performance on an i.i.d. test set, reducing learning to a single aggregate score. This ap

researcharxiv-cs-cl
25 Jun 2026
Model Releases

WOLF-VLA: Whole-Body Humanoid Optimal Locomotion Framework for Vision-Language-Action Learning

DGX agent

arXiv:2606.25591v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently demonstrated strong generalization in robotic manipulation, yet their applicability to whole-body, con

model-releasesarxiv-cs-ro
25 Jun 2026
Safety

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

DGX agent

arXiv:2606.25034v1 Announce Type: new Abstract: General-purpose models often struggle to reliably identify and understand real-world multimodal risks, largely due to the inherent multimodal adversaria

safetyarxiv-cs-cv
25 Jun 2026
Research

Advancing WordArt-Oriented Scene Text Recognition: Datasets and Methods

DGX agent

arXiv:2606.24484v1 Announce Type: new Abstract: WordArt (artistic text) features highly customized fonts, textures, and layouts, making WordArt-oriented scene TExt Recognition (WATER) substantially mo

researcharxiv-cs-cv
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming

DGX agent

arXiv:2606.24245v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly automate complex tasks by integrating language models with external tools and environments. However, th

safetyarxiv-cs-ai
24 Jun 2026
Safety

Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation

DGX agent

arXiv:2606.24464v1 Announce Type: new Abstract: Text-driven Referring Video Object Segmentation (RVOS) aims to locate and segment target objects in videos given natural language. However, existing mod

safetyarxiv-cs-cv
24 Jun 2026
Research

Configurable Holography: Towards Display and Scene Adaptation

DGX agent

arXiv:2405.01558v4 Announce Type: replace Abstract: Rendering holograms for holographic displays is often an iterative and computationally costly process. Emerging learned holography methods have alle

researcharxiv-cs-cv
24 Jun 2026
Agents

Debate2Create: Robot Co-design via Multi-Agent LLM Debate

DGX agent

arXiv:2510.25850v3 Announce Type: replace-cross Abstract: We introduce Debate2Create (D2C), a multi-agent LLM framework that formulates robot co-design as structured, iterative debate grounded in phys

agentsarxiv-cs-lg
24 Jun 2026
Model Releases

DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects

DGX agent

arXiv:2606.24779v1 Announce Type: cross Abstract: Birth defects are a major cause of fetal loss, neonatal morbidity and long-term disability. In the subset with suspected genetic etiologies, exome and

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

DramaDirector: Geometry-Guided Short Drama Generation

DGX agent

arXiv:2606.24107v1 Announce Type: cross Abstract: Short dramas, with their rapid shot rhythms, dialogue-driven focus shifts, and demanding cinematographic grounding, pose challenges that prompt-level

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

FlowPipe: LLM-Enhanced Conditional Generative Flow Networks for Data Preparation Pipeline Construction

DGX agent

arXiv:2606.24679v1 Announce Type: cross Abstract: Data preparation pipelines improve data quality in machine learning by transforming raw tables into learning-ready data through sequential cleaning an

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Grounding Generative Policies in Physics: Optimization-Guided Diffusion for Robot Control

DGX agent

arXiv:2606.24208v1 Announce Type: new Abstract: Diffusion models sample effectively from high-dimensional, multimodal distributions, but their outputs may violate deployment constraints. For task-spac

researcharxiv-cs-ro
24 Jun 2026
Hardware

Hardware-Oriented Inference Complexity of Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.03345v2 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) have recently emerged as a powerful architecture for various machine learning applications. However, their unique

hardwarearxiv-cs-lg
24 Jun 2026
Research

High-Fidelity Synthetic Transmission Electron Microscopy Image Generation Using Diffusion Probabilistic Models for Data-Limited Semiconductor Metrology

DGX agent

arXiv:2606.24817v1 Announce Type: new Abstract: Advanced semiconductor nodes drastically increased demand for Transmission Electron Microscopy (TEM), yet destructive sample preparation, slow imaging a

researcharxiv-cs-cv
24 Jun 2026
Model Releases

MortarBench: Evaluating Mortgage Loan Origination Agents

DGX agent

arXiv:2606.19416v2 Announce Type: replace Abstract: Loan origination is the process by which a lender creates a new loan, from application and underwriting through approval and funding. This process s

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Navigating User Behavior toward Personalized Multimodal Generation

DGX agent

arXiv:2606.24196v1 Announce Type: new Abstract: Modern AIGC pipelines deliver high-fidelity images and videos but presuppose a well-formed creation instruction, while end users rarely articulate visua

model-releasesarxiv-cs-ai
24 Jun 2026
Local Ai

Neural Particle Automata: Learning Self-Organizing Particle Dynamics

DGX agent

arXiv:2601.16096v2 Announce Type: replace-cross Abstract: We introduce Neural Particle Automata (NPA), a Lagrangian generalization of Neural Cellular Automata (NCA) from static lattices to dynamic par

local-aiarxiv-cs-cv
24 Jun 2026
Research

S1-Omni-Image: A Unified Model for Scientific Image Understanding, Generation, and Editing

DGX agent

arXiv:2606.24441v1 Announce Type: new Abstract: We present S1-Omni-Image, an open-weight unified multimodal model for scientific image understanding, generation, and editing. Unlike general-purpose im

researcharxiv-cs-cv
24 Jun 2026
Applications

Sat2City v2: Native 3D City Asset Generation from a Single Satellite Image

DGX agent

arXiv:2606.24138v1 Announce Type: new Abstract: Generating explicit 3D city assets from a single satellite image is important for digital twins, urban simulation, and geospatial intelligence. Unlike s

applicationsarxiv-cs-cv
24 Jun 2026
Applications

SLEEPING-DISCO 9M: A large-scale pre-training dataset for generative music modeling

DGX agent

arXiv:2506.14293v4 Announce Type: replace-cross Abstract: We present Sleeping-DISCO 9M, a large-scale pre-training dataset for music and song. To the best of our knowledge, there are no open-source hi

applicationsarxiv-cs-lg
24 Jun 2026
Research

2D Versus 3D Diffusion for In Silico Training of Interventional X-ray AI Models

DGX agent

arXiv:2606.21414v1 Announce Type: cross Abstract: The ability to synthesize realistic X-ray images has catalyzed the development of AI models for X-ray image-guided procedures, which otherwise suffer

researcharxiv-cs-cv
23 Jun 2026
Agents

3D Vessel Reconstruction from Sparse-View Dynamic DSA Images via Vessel Probability Guided Attenuation Learning

DGX agent

arXiv:2405.10705v3 Announce Type: replace-cross Abstract: Digital Subtraction Angiography (DSA) is one of the gold standards for vascular disease diagnosis. With the help of a contrast agent, time-res

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

ACE-GS: Acing the Trade-off with Accurate, Compact and Efficient 3D Gaussian Splatting

DGX agent

arXiv:2606.21244v1 Announce Type: new Abstract: 3D Gaussian Splatting achieves exceptional real-time rendering, but its substantial computational and storage demands hinder widespread deployment. Exis

model-releasesarxiv-cs-cv
23 Jun 2026
Hardware

ACEsplat: Accelerated 3D Gaussian Scene Regression via RGB and Poses Only

DGX agent

arXiv:2606.22091v1 Announce Type: new Abstract: Per-scene 3D Gaussian Splatting (3DGS) enables high-fidelity rendering, but practical robotic and AR scene capture pipelines often depend on external ge

hardwarearxiv-cs-ro
23 Jun 2026
Applications

Beyond a Single Light: A Large-Scale Aerial Dataset for Urban Scene Reconstruction Under Varying Illumination

DGX agent

arXiv:2512.14200v2 Announce Type: replace Abstract: Recent advances in Neural Radiance Fields and 3D Gaussian Splatting have demonstrated strong potential for large-scale UAV-based 3D reconstruction t

applicationsarxiv-cs-cv
23 Jun 2026
Safety

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

DGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

DGX agent

arXiv:2606.21337v1 Announce Type: new Abstract: Massive unstructured multimodal streams suffer from high 'data entropy,' impeding both efficient human knowledge acquisition and high-quality AI post-tr

model-releasesarxiv-cs-lg
23 Jun 2026
Research

ECGFlowCMR: Pretraining with ECG-Generated Cine CMR Helps Cardiac Disease Classification and Phenotype Prediction

DGX agent

arXiv:2601.20904v3 Announce Type: replace-cross Abstract: Cardiac Magnetic Resonance (CMR) imaging provides a comprehensive assessment of cardiac structure and function but remains constrained by high

researcharxiv-cs-lg
23 Jun 2026
Model Releases

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion

DGX agent

arXiv:2606.20924v1 Announce Type: new Abstract: In multi-object text-to-image (T2I) diffusion, ensuring semantic consistency between textual prompts and generated visual content is crucial for image s

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Enhancing Creativity in 3D Generative Design via a TRIZ-Inspired Text-to-CAD Framework

DGX agent

arXiv:2606.21378v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated significant potential in supporting engineering design tasks, including computer-aided

agentsarxiv-cs-lg
23 Jun 2026
Research

Iterative Diffusion-Refined Neural Attenuation Fields for Multi-Source Stationary CT Reconstruction: NAF Meets Diffusion Model

DGX agent

arXiv:2511.14310v2 Announce Type: replace Abstract: Multi-source stationary computed tomography (CT) has recently attracted attention for its ability to achieve rapid image reconstruction, making it s

researcharxiv-cs-cv
23 Jun 2026
Model Releases

LIBERO-Safety: A Comprehensive Benchmark for Physical and Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2606.23686v1 Announce Type: new Abstract: Despite the impressive manipulation capabilities of Vision-Language-Action (VLA) models, their operational safety under strict constraints remains large

model-releasesarxiv-cs-ro
23 Jun 2026
Research

LVQAC: Lattice Vector Quantization Coupled with Spatially Adaptive Companding for Efficient Learned Image Compression

DGX agent

arXiv:2304.12319v2 Announce Type: replace-cross Abstract: Recently, numerous end-to-end optimized image compression neural networks have been developed and proved themselves as leaders in rate-distort

researcharxiv-cs-cv
23 Jun 2026
Model Releases

MammoExpert: Benchmarking Chain-of-Thought Reasoning in Mammography Diagnosis

DGX agent

arXiv:2606.21119v1 Announce Type: new Abstract: Mammography is an essential tool for breast cancer detection, with millions of examinations conducted annually. However, publicly available high-quality

model-releasesarxiv-cs-cv
23 Jun 2026
Research

MeGAS: Thermomechanical Dynamic Gaussian Splatting for Thermophysical Scene Editing

DGX agent

arXiv:2606.23455v1 Announce Type: new Abstract: Recent advances integrate physically grounded Newtonian dynamics with neural rendering frameworks, narrowing the gap between photorealistic scene recons

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Open Annotations and Synthetic Data for Field Localisation in Indian Bank Cheques

DGX agent

arXiv:2606.20682v1 Announce Type: new Abstract: Automated cheque processing requires localising key fields (date, legal amount, IFSC code, account number, signature, and payee name) before any recogni

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

DGX agent

arXiv:2601.03357v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a standard approach to reconstruct and render photorealistic 3D head avatars. A major challenge is to religh

safetyarxiv-cs-cv
23 Jun 2026
Research

Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation

DGX agent

arXiv:2603.08305v2 Announce Type: replace Abstract: Text-conditioned generative models for volumetric medical imaging provide semantic control but lack explicit anatomical guidance, often resulting in

researcharxiv-cs-cv
23 Jun 2026
Research

S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix

DGX agent

arXiv:2508.08048v2 Announce Type: replace Abstract: While video generation models excel at producing high-quality monocular videos, generating 3D stereoscopic and spatial videos for immersive applicat

researcharxiv-cs-cv
23 Jun 2026
Local Ai

SteerVTE: Seamless Video Text Editing with Style and Glyph Control

DGX agent

arXiv:2606.23254v1 Announce Type: new Abstract: Visual text editing aims to precisely modify text in images and videos while preserving stylistic consistency and visual realism. Despite significant ad

local-aiarxiv-cs-cv
23 Jun 2026
Research

Synthetic Network Packet Generation through Statistical Learning and Genetic Algorithms

DGX agent

arXiv:2606.20864v1 Announce Type: cross Abstract: Developing robust intrusion detection systems (IDS) for IoT environments requires large, labeled datasets capturing realistic traffic distributions ac

researcharxiv-cs-lg
23 Jun 2026
Model Releases

TeleStyle V2: Beyond Content-Preserving Style Transfer with Self-Distillation and Distribution-Matching-Distillation

DGX agent

arXiv:2606.20709v1 Announce Type: new Abstract: Given a content reference and a style reference, content-preserving style transfer requires the model to generate stylized outputs with content and styl

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

Training Diffusion Policies via Prior-Mapping Co-Evolution

DGX agent

arXiv:2512.02581v3 Announce Type: replace Abstract: Reinforcement learning (RL) faces a persistent tension: policies that are stable to optimize (e.g., Gaussians) are often too simple to represent the

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Another new idea to push the state of AI architectures forward. Sakana released a model that effectively uses a mixture of models to get wor…

DGX agent

Another new idea to push the state of AI architectures forward. Sakana released a model that effectively uses a mixture of models to get work done. You get a single API but then the work gets farmed o

model-releasesdavid-ha--x
22 Jun 2026
Agents

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

DGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

agentsdavid-ha--x
22 Jun 2026
← Previous
1…3738394041…60
Next →