AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
7 Jul 2026

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Model ReleasesDGX agent

arXiv:2511.07403v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress in vision-language tasks, but continue to struggle with spatial rea

TexTailor: Inference-Time Textual Guidance Tailoring for Multimodal Diffusion Transformers

Model ReleasesDGX agent

arXiv:2601.02211v2 Announce Type: replace Abstract: Recent breakthroughs of transformer-based diffusion models, particularly with Multimodal Diffusion Transformers (MMDiT) driven models like FLUX and

TokAN: Accent Normalization Using Self-Supervised Speech Tokens

SafetyDGX agent

arXiv:2607.03928v1 Announce Type: cross Abstract: Accent normalization (AN) seeks to convert non-native (L2) accented speech into standard (L1) speech while preserving speaker identity. The current te


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language

ResearchDGX agent

arXiv:2607.04515v1 Announce Type: new Abstract: Efik, a tonal language spoken by about 3 million second language speakers and 1.5 million native speakers in Southeastern Nigeria, remains underrepresen

UniVideo: Unified Understanding, Generation, and Editing for Videos

Model ReleasesDGX agent

arXiv:2510.08377v4 Announce Type: replace Abstract: Unified multimodal models have shown promising results in multimodal content generation and editing but remain largely limited to the image domain.

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

ResearchDGX agent

The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection attempts remain limited to offline trajectories fo

XPlainVerse: A Million-Scale Benchmark for Explainable Deepfake Detection

Model ReleasesDGX agent

arXiv:2607.03562v1 Announce Type: new Abstract: As deepfake detection models increasingly produce natural language explanations, their reasoning often remains weakly grounded in visual artifacts, limi

3 Jul 2026

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

HardwareDGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

SafetyDGX agent

arXiv:2603.22435v2 Announce Type: replace-cross Abstract: 'Code-as-Policy' considers how executable code can complement data-intensive Vision-Language-Action (VLA) methods, yet their effectiveness as

DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving

SafetyDGX agent

arXiv:2603.18315v2 Announce Type: replace-cross Abstract: Traditional reinforcement learning (RL) methods rely on manually engineered rewards or sparse collision signals, which fail to capture the ric

Hawk: Harnessing Hardware-Aware Knowledge for High-Performance NPU Kernel Generation

ApplicationsDGX agent

arXiv:2607.01590v1 Announce Type: new Abstract: Developing high-performance kernels for Neural Processing Units (NPUs) is a critical industry bottleneck, requiring developers to manually navigate impl

How Indian Dermatologists are Utilizing Artificial Intelligence for Clinical Practice and Workflow Management: A Nationwide Survey with a Special Focus on atopic dermatitis

ResearchDGX agent

arXiv:2607.01252v1 Announce Type: cross Abstract: Background: Dermatology AI has mainly focused on image-based diagnosis, while chronic disease workflows have received less attention. We surveyed Indi

LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models

Model ReleasesDGX agent

arXiv:2504.02327v2 Announce Type: replace Abstract: Natural Language to SQL (NL2SQL) aims to translate natural language queries into executable SQL statements, offering non-expert users intuitive acce

Learning 3D-Gaussian Simulators from RGB Videos

TutorialsDGX agent

arXiv:2503.24009v3 Announce Type: replace-cross Abstract: Realistic simulation is critical for applications ranging from robotics to animation. Learned simulators have emerged as a possibility to capt

LEFT: Learnable Fusion of Tri-view Tokens for Unsupervised Time Series Anomaly Detection

TutorialsDGX agent

arXiv:2602.08638v2 Announce Type: replace-cross Abstract: As a fundamental data mining task, unsupervised time series anomaly detection (TSAD) aims to build a model for identifying abnormal timestamps

Office Comprehension Benchmark

Model ReleasesDGX agent

arXiv:2607.01245v1 Announce Type: cross Abstract: We introduce Office Comprehension Bench (OCB), the first public benchmark to jointly evaluate LLM systems on Word, Excel, and PowerPoint comprehension

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

Model ReleasesDGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

Traceable Fault Diagnosis for Battery Energy Storage Systems via Retrieval-Augmented Multi-Agent O&M Assistant

AgentsDGX agent

arXiv:2607.01992v1 Announce Type: new Abstract: Large-scale battery energy storage systems (BESSs) require O&M decisions that combine alarms, cell-level measurements, device topology, diagnostic table

VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object Retrieval

Model ReleasesDGX agent

arXiv:2607.02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, rec

2 Jul 2026

An LLM-Based Framework for Intent-Driven Network Topology Design

Model ReleasesDGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

ASPIRE: Agentic /Skills Discovery for Robotics

SafetyDGX agent

arXiv:2607.00272v1 Announce Type: cross Abstract: Traditional robot programming is challenging: it requires orchestrating multimodal perception, managing physical contact dynamics, and handling divers

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

SafetyDGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

HardwareDGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

Model ReleasesDGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

HardwareDGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

Graded strength of comparative illusions is explained by Bayesian inference

ResearchDGX agent

arXiv:2511.14642v2 Announce Type: replace Abstract: Like visual processing, language processing is susceptible to illusions in which people systematically misperceive stimuli. In one such case--the co

Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination

SafetyDGX agent

arXiv:2607.00924v1 Announce Type: new Abstract: Accelerating materials discovery requires AI systems that can generate scientifically valid hypotheses through multi-step, domain-grounded reasoning. St

Human-Machine Collaboration on Generative Meta-Learning: Model and Algorithm

AgentsDGX agent

arXiv:2607.00926v1 Announce Type: cross Abstract: Generalizing machine learning models to environments that differ from their training distribution remains a critical hurdle, particularly when data fr

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models

ResearchDGX agent

arXiv:2607.01222v1 Announce Type: new Abstract: Recent 3D generative models can synthesize high-quality geometry but often struggle to reproduce intricate textures from reference images, largely due t

Knowdit: Agentic Smart Contract Vulnerability Detection with Auditing Knowledge Summarization

AgentsDGX agent

arXiv:2603.26270v2 Announce Type: replace-cross Abstract: Smart contracts govern billions of dollars in decentralized finance (DeFi), yet automated vulnerability detection remains challenging because

Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads

Model ReleasesDGX agent

arXiv:2607.01002v1 Announce Type: cross Abstract: In long-context use, large language models frequently synthesize answers from the meaning of a relevant context span rather than literally copy-pastin

Measuring the Gap Between Human and LLM Research Ideas

ResearchDGX agent

arXiv:2607.01233v1 Announce Type: cross Abstract: LLMs are increasingly used to brainstorm research ideas, but existing evaluations mostly judge individual ideas by novelty, feasibility, or expert pre

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

SafetyDGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

Optimal Resource Utilization for Autonomous Laboratory Orchestrators

AgentsDGX agent

arXiv:2607.01188v1 Announce Type: new Abstract: In autonomous laboratories, AI agents suggest the next batch of experiments to do. However, planning and executing those tasks taking full advantage of

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

Model ReleasesDGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

Prompt2Effect: Training-Free Image-to-Video Model Specialization via LoRA Generation

SafetyDGX agent

arXiv:2606.13971v2 Announce Type: replace Abstract: While personalizing Image-to-Video (I2V) diffusion models with specific visual effects is increasingly demanded for high-end generation, current pra

Radial Interaction Tomography: Recognizing Non-Transitive Evolutionary Games from One Range-Expansion Image

Model ReleasesDGX agent

arXiv:2607.00378v1 Announce Type: new Abstract: Colored sectors in a microbial range expansion encode more than lineage survival counts. We formulate a computer-vision inverse problem: from one endpoi

Unleashing More Actions via Action Compositional Training for VLA Models

SafetyDGX agent

arXiv:2607.00351v1 Announce Type: new Abstract: Vision-Language-Action models excel at robotic manipulation, driven by the scale and diversity of demonstration data. However, standard training paradig

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video

ResearchDGX agent

arXiv:2607.01202v1 Announce Type: cross Abstract: We present World from Motion, a method for generating freely renderable dynamic 3D Gaussian representations from monocular videos. Our approach condit

1 Jul 2026

A Scalable Whole-body Motion Transfer via Implicit Kinodynamic Motion Retargeting

SafetyDGX agent

arXiv:2509.15443v2 Announce Type: replace-cross Abstract: Human-to-humanoid imitation learning presents a promising pathway to address the severe data scarcity bottleneck in robotics by utilizing abun

AnyBokeh: Physics-Guided Any-to-Any Bokeh Editing with Optical Fingerprint Transfer

ApplicationsDGX agent

arXiv:2606.31959v1 Announce Type: new Abstract: Depth-of-field control is a fundamental tool in photography, yet post-capture bokeh editing from a single image remains challenging. A practical editor

Cross-Resolution Distribution Matching for Diffusion Distillation

ResearchDGX agent

arXiv:2603.06136v2 Announce Type: replace Abstract: Diffusion distillation is central to accelerating image and video generation, yet existing methods are fundamentally limited by the denoising proces

DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation

SafetyDGX agent

arXiv:2606.31537v1 Announce Type: new Abstract: Text-rich image generation is one of the most challenging settings in image generation, since models must simultaneously produce visually realistic imag

DPPE: Rethinking Camera-Based Positional Encoding for Scaling Multi-View Transformers

ResearchDGX agent

arXiv:2606.31585v1 Announce Type: cross Abstract: The remarkable scalability of Transformers has expanded their application to 3D computer vision, where camera-aware positional encoding is crucial for

ELEVATE: Designing Human-Centered GenAI Virtual Tutors for Scalable and Inclusive Education

Local AiDGX agent

arXiv:2606.30662v1 Announce Type: cross Abstract: The advent of Generative Artificial Intelligence (GenAI), and in particular Large Language Models (LLMs), is reshaping educational practice, while int

Few-Shot Synthetic Image Attribution: Identifying Unseen Generators with Limited Samples

ApplicationsDGX agent

arXiv:2509.25682v2 Announce Type: replace Abstract: AI-generated image (AIGI) attribution presents a pressing challenge that goes beyond mere AIGI detection, aiming to identify the source model or tec

From Grasps to Dexterity: Large-Scale Grasp Pretraining for Dexterous Manipulation

Model ReleasesDGX agent

arXiv:2606.30749v1 Announce Type: new Abstract: Large-scale dexterous grasp datasets encode rich priors over hand-object interaction, but their use has largely been confined to grasp generation and pi

From Materials Database to Materials Bank: Assetizing Data for AI Driven Materials Innovation

ApplicationsDGX agent

arXiv:2606.31366v1 Announce Type: cross Abstract: Driven by high-throughput experimentation, computational modeling, and artificial intelligence (AI), materials data has expanded at an unprecedented r

HistoriQA-ThirdRepublic: Multi-Hop Question Answering Corpus for Historical Research, Parliamentary Debates from the French Third Republic (1870-1940)

SafetyDGX agent

arXiv:2606.31325v1 Announce Type: new Abstract: We present HistoriQA-ThirdRepublic: a French-language dataset of multi-hop historical questions derived from parliamentary debates and newspapers of the

@hwchase17 s “wiki memory” really resonates with how we’ve been thinking about Enterprise Knowledgebases. Agent memory is still early, but o…

AgentsDGX agent

@hwchase17 s “wiki memory” really resonates with how we’ve been thinking about Enterprise Knowledgebases. Agent memory is still early, but one pattern we observe consistently: enterprises don’t need a

InfiniVerse: Occupancy Guided Unbounded Scene Generation for Autonomous Driving

SafetyDGX agent

arXiv:2606.31109v1 Announce Type: new Abstract: Generating realistic, controllable, and temporally coherent urban environments is a critical yet unresolved challenge in the autonomous driving communit

IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing

SafetyDGX agent

arXiv:2606.13368v2 Announce Type: replace Abstract: Computer-Aided Design is pivotal in modern manufacturing, yet existing automated methods predominantly rely on open-loop, one-shot generation, creat

Loc2Repair: A Framework for Evaluating the Impact of File-Level Issue Localization in Repo-Level LLM Repair

Local AiDGX agent

arXiv:2606.30963v1 Announce Type: cross Abstract: Repository-grounded automated repair is often reported as a single end-to-end capability, which hides distinct failure modes such as poor file targeti

Motion Planning in Compressed Representation Spaces

AgentsDGX agent

arXiv:2606.30940v1 Announce Type: cross Abstract: Deep learning methods have vastly expanded the capabilities of motion planning in robotics applications, as learning priors from large-scale data has

On the pod: 'Constrained Adaptive Rejection Sampling' with @ucsd_cse professor @lorisdanto. Hear how symbolic AI experts have navigated the …

ResearchDGX agent

On the pod: 'Constrained Adaptive Rejection Sampling' with @ucsd_cse professor @lorisdanto. Hear how symbolic AI experts have navigated the LLM era and why the future of AI code generation depends on

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation

Model ReleasesDGX agent

arXiv:2606.30673v1 Announce Type: cross Abstract: Autoregressive Transformers dominate high-quality mesh generation by producing artist-worthy topologies, yet their inherent sequential decoding induce

ShardNet: Training Neural Controllers with Hard, Non-Convex Constraints

SafetyDGX agent

arXiv:2606.30935v1 Announce Type: cross Abstract: While neural network control policies are powerful, their deployment on safety critical systems depends on ensuring that they obey strict constraints.

SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views

ResearchDGX agent

arXiv:2509.17246v2 Announce Type: replace Abstract: We introduce SPFSplatV2, an efficient feed-forward framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth pose

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

SafetyDGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

SafetyDGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

← Previous
1…2728293031…47
Next →