AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
Model Releases

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

DGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

model-releasesarxiv-cs-cl
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

DGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

model-releasesarxiv-cs-ai
19 May 2026
Research

GaussianZoom: Progressive Zoom-in Generative 3D Gaussian Splatting with Geometric and Semantic Guidance

DGX agent

arXiv:2605.18252v1 Announce Type: new Abstract: We introduce GaussianZoom, a generative zoom-in 3D reconstruction system with an iterative progressive framework that combines geometry-consistent scene

researcharxiv-cs-cv
19 May 2026
Tutorials

Generating Pretraining Tokens from Organic Data for Data-Bound Scaling

DGX agent

arXiv:2605.17849v1 Announce Type: cross Abstract: LLM pretraining is shifting from a compute-bound to a data-bound regime, where available human (organic) text falls far short of scaling demands. Howe

tutorialsarxiv-cs-ai
19 May 2026
Research

Generative 3D Gaussians with Learned Density Control

DGX agent

arXiv:2605.16355v1 Announce Type: cross Abstract: We present Density-Sampled Gaussians (DeG), a novel 3D representation designed to bridge the gap between adaptive rendering primitives and scalable ge

researcharxiv-cs-cv
19 May 2026
Research

Geospatial-Reasoning-Driven Vocabulary-Agnostic Remote Sensing Semantic Segmentation

DGX agent

arXiv:2602.08206v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation has become an important direction in remote sensing, as it enables recognition beyond predefined land-cover ca

researcharxiv-cs-cv
19 May 2026
Research

HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos

DGX agent

arXiv:2605.17543v1 Announce Type: new Abstract: Video outpainting generates plausible visual content beyond the original spatial extent of a video, playing a key role in adapting videos to diverse dis

researcharxiv-cs-cv
19 May 2026
Research

Human-Certified Module Repositories for the AI Age

DGX agent

arXiv:2603.02512v4 Announce Type: replace-cross Abstract: Human-Certified Module Repositories (HCMRs) are introduced in this work as a new architectural model for constructing trustworthy software in

researcharxiv-cs-ai
19 May 2026
Local Ai

IdGlow: Dynamic Identity Modulation for Multi-Subject Generation

DGX agent

arXiv:2603.00607v2 Announce Type: replace-cross Abstract: Multi-subject image generation requires seamlessly harmonizing multiple reference identities within a coherent scene. However, existing method

local-aiarxiv-cs-ai
19 May 2026
Research

InstructAV2AV: Instruction-Guided Audio-Video Joint Editing

DGX agent

arXiv:2605.18467v1 Announce Type: new Abstract: Recent diffusion-based methods have achieved impressive progress in video content manipulation. However, they typically ignore the accompanying audio, l

researcharxiv-cs-cv
19 May 2026
Safety

Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems

DGX agent

arXiv:2605.16278v1 Announce Type: cross Abstract: The use of Artificial Intelligence (AI) in high-risk, decision-making scenarios presents technical, safety, and normative challenges; problems that ma

safetyarxiv-cs-ai
19 May 2026
Research

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

DGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

researcharxiv-cs-cv
19 May 2026
Model Releases

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

DGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

model-releasesarxiv-cs-ai
19 May 2026
Agents

OProver: A Unified Framework for Agentic Formal Theorem Proving

DGX agent

arXiv:2605.17283v1 Announce Type: cross Abstract: Recent progress in formal theorem proving has benefited from large-scale proof generation and verifier-aware training, but agentic proving is rarely i

agentsarxiv-cs-ai
19 May 2026
Research

PIXLRelight: Controllable Relighting via Intrinsic Conditioning

DGX agent

arXiv:2605.18735v1 Announce Type: new Abstract: We present PIXLRelight, a feed-forward approach for physically controllable single-image relighting. Existing methods either provide limited lighting co

researcharxiv-cs-cv
19 May 2026
Research

Position: Weight Space Should Be a First-Class Generative AI Modality

DGX agent

arXiv:2605.18632v1 Announce Type: cross Abstract: Neural network checkpoints have quietly become a large-scale data resource: millions of trained weight vectors now exist, each encoding task-, domain-

researcharxiv-cs-ai
19 May 2026
Model Releases

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

DGX agent

arXiv:2605.16612v1 Announce Type: new Abstract: Rapid identification of candidate materials with target properties has become a key task in materials science. Machine learning has emerged as an altern

model-releasesarxiv-cs-ai
19 May 2026
Agents

Real-time Multi-instrument Autonomous Discovery of Novel Phase-change Memory Materials

DGX agent

arXiv:2605.18033v1 Announce Type: cross Abstract: Autonomous labs enable the integration of automated experiment execution, data analysis and decision making. The main challenge remains the integratio

agentsarxiv-cs-lg
19 May 2026
Applications

RHINO: Reconstructing Human Interactions with Novel Objects from Monocular Videos

DGX agent

arXiv:2605.17014v1 Announce Type: new Abstract: Reconstructing people, objects, and their interactions in 3D is a long-standing goal for intelligent systems. Often the input is RGB video from a moving

applicationsarxiv-cs-cv
19 May 2026
Model Releases

Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback

DGX agent

arXiv:2605.17448v1 Announce Type: cross Abstract: Computer-aided design (CAD) is the backbone of modern industrial design, yet learned CAD generators still fall short of real engineering pipelines: th

model-releasesarxiv-cs-cl
19 May 2026
Safety

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

DGX agent

arXiv:2605.17423v1 Announce Type: new Abstract: We study series-level cinematic remaking, a long-horizon video-to-video generation problem that localizes full episodes or films via stylization or acto

safetyarxiv-cs-cv
19 May 2026
Research

StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model

DGX agent

arXiv:2511.14223v3 Announce Type: replace Abstract: This paper focuses on the task of speech-driven 3D facial animation, which aims to generate realistic and synchronized facial motions driven by spee

researcharxiv-cs-cv
19 May 2026
Research

Systematic Evaluation of the Quality of Synthetic Clinical Notes Rephrased by LLMs at Million-Note Scale

DGX agent

arXiv:2605.17775v1 Announce Type: cross Abstract: Large language models (LLMs) can generate or synthesize clinical text for a wide range of applications, from improving clinical documentation to augme

researcharxiv-cs-ai
19 May 2026
Hardware

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

DGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

hardwarearxiv-cs-ai
19 May 2026
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
Research

The IsalProgram Programming Language

DGX agent

arXiv:2605.17008v1 Announce Type: cross Abstract: We introduce IsalProgram (Instruction Set and Language for Programming), a novel assembly-like programming language with three distinctive theoretical

researcharxiv-cs-ai
19 May 2026
Model Releases

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

DGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

model-releasesarxiv-cs-ai
19 May 2026
Agents

Tongyi DeepResearch Technical Report

DGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

agentsarxiv-cs-ai
19 May 2026
Safety

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

DGX agent

arXiv:2605.16420v1 Announce Type: new Abstract: This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structure

safetyarxiv-cs-cv
19 May 2026
Model Releases

WavFlow: Audio Generation in Waveform Space

DGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

model-releasesarxiv-cs-cv
19 May 2026
Research

Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road

DGX agent

arXiv:2605.17026v1 Announce Type: new Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through spec

researcharxiv-cs-lg
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Agents

Xiaomi EV World Model: A Joint World Model Integrating Reconstruction and Generation for Autonomous Driving

DGX agent

arXiv:2605.18137v1 Announce Type: new Abstract: This report presents a unified technical system addressing the two core capabilities of world models for autonomous driving: world representation and wo

agentsarxiv-cs-cv
19 May 2026
Model Releases

A3D: Agentic AI flow for autonomous Accelerator Design

DGX agent

arXiv:2605.15237v1 Announce Type: cross Abstract: Accelerating applications through the design of hardware accelerators can significantly enhance system performance and energy efficiency. Despite adva

model-releasesarxiv-cs-ai
18 May 2026
Local Ai

ChangeFlow -- Latent Rectified Flow for Change Detection in Remote Sensing

DGX agent

arXiv:2605.15375v1 Announce Type: cross Abstract: Remote sensing change detection (RSCD) aims to localise changes between two images of the same geographic region. In practice, change masks often foll

local-aiarxiv-cs-ai
18 May 2026
Safety

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

DGX agent

arXiv:2605.15532v1 Announce Type: cross Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically

safetyarxiv-cs-ai
18 May 2026
Model Releases

Diagonal Adaptive Non-local Observables on Quantum Neural Networks

DGX agent

arXiv:2605.15410v1 Announce Type: cross Abstract: Adaptive Non-local Observables (ANOs) have shown that making quantum observables dynamic can substantially enlarge the function space of Variational Q

model-releasesarxiv-cs-ai
18 May 2026
Safety

Differentially Private Motif-Preserving Multi-modal Hashing

DGX agent

arXiv:2605.15460v1 Announce Type: cross Abstract: Cross-modal hashing enables efficient retrieval by encoding images and text into compact binary codes. State-of-the-art methods rely on semantic simil

safetyarxiv-cs-ai
18 May 2026
Local Ai

DreamSR: Towards Ultra-High-Resolution Image Super-Resolution via a Receptive-Field Enhanced Diffusion Transformer

DGX agent

arXiv:2605.15682v1 Announce Type: new Abstract: Large-scale pre-trained diffusion models have been extensively adopted for real-world image Super-Resolution because of their powerful generative priors

local-aiarxiv-cs-cv
18 May 2026
Research

IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation

DGX agent

arXiv:2605.16258v1 Announce Type: cross Abstract: Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most ex

researcharxiv-cs-ai
18 May 2026
Model Releases

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

DGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

DGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

model-releasesarxiv-cs-ro
18 May 2026
Research

PanoWorld: Geometry-Consistent Panoramic Video World Modeling

DGX agent

arXiv:2605.15391v1 Announce Type: cross Abstract: We present PanoWorld, a panoramic video world model that generates geometry-consistent 360egree video from a single image and a caption. Existing pano

researcharxiv-cs-ai
18 May 2026
Safety

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

DGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

safetyarxiv-cs-lg
18 May 2026
Agents

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

DGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

agentsarxiv-cs-ai
18 May 2026
Local Ai

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

DGX agent

arXiv:2605.16138v1 Announce Type: cross Abstract: Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on

local-aiarxiv-cs-ai
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation

DGX agent

arXiv:2605.16079v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown significant progress in video understanding, yet they face substantial challenges in tasks requiring p

model-releasesarxiv-cs-ai
18 May 2026
← Previous
1…4445464748…59
Next →