AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,711 results
Model Releases

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

DGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

The IsalProgram Programming Language

DGX agent

arXiv:2605.17008v1 Announce Type: cross Abstract: We introduce IsalProgram (Instruction Set and Language for Programming), a novel assembly-like programming language with three distinctive theoretical

researcharxiv-cs-ai
19 May 2026
Model Releases

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

DGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

model-releasesarxiv-cs-ai
19 May 2026
Agents

Tongyi DeepResearch Technical Report

DGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

agentsarxiv-cs-ai
19 May 2026
Safety

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

DGX agent

arXiv:2605.16420v1 Announce Type: new Abstract: This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structure

safetyarxiv-cs-cv
19 May 2026
Model Releases

WavFlow: Audio Generation in Waveform Space

DGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

model-releasesarxiv-cs-cv
19 May 2026
Research

Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road

DGX agent

arXiv:2605.17026v1 Announce Type: new Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through spec

researcharxiv-cs-lg
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Agents

Xiaomi EV World Model: A Joint World Model Integrating Reconstruction and Generation for Autonomous Driving

DGX agent

arXiv:2605.18137v1 Announce Type: new Abstract: This report presents a unified technical system addressing the two core capabilities of world models for autonomous driving: world representation and wo

agentsarxiv-cs-cv
19 May 2026
Model Releases

A3D: Agentic AI flow for autonomous Accelerator Design

DGX agent

arXiv:2605.15237v1 Announce Type: cross Abstract: Accelerating applications through the design of hardware accelerators can significantly enhance system performance and energy efficiency. Despite adva

model-releasesarxiv-cs-ai
18 May 2026
Local Ai

ChangeFlow -- Latent Rectified Flow for Change Detection in Remote Sensing

DGX agent

arXiv:2605.15375v1 Announce Type: cross Abstract: Remote sensing change detection (RSCD) aims to localise changes between two images of the same geographic region. In practice, change masks often foll

local-aiarxiv-cs-ai
18 May 2026
Safety

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

DGX agent

arXiv:2605.15532v1 Announce Type: cross Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically

safetyarxiv-cs-ai
18 May 2026
Model Releases

Diagonal Adaptive Non-local Observables on Quantum Neural Networks

DGX agent

arXiv:2605.15410v1 Announce Type: cross Abstract: Adaptive Non-local Observables (ANOs) have shown that making quantum observables dynamic can substantially enlarge the function space of Variational Q

model-releasesarxiv-cs-ai
18 May 2026
Safety

Differentially Private Motif-Preserving Multi-modal Hashing

DGX agent

arXiv:2605.15460v1 Announce Type: cross Abstract: Cross-modal hashing enables efficient retrieval by encoding images and text into compact binary codes. State-of-the-art methods rely on semantic simil

safetyarxiv-cs-ai
18 May 2026
Local Ai

DreamSR: Towards Ultra-High-Resolution Image Super-Resolution via a Receptive-Field Enhanced Diffusion Transformer

DGX agent

arXiv:2605.15682v1 Announce Type: new Abstract: Large-scale pre-trained diffusion models have been extensively adopted for real-world image Super-Resolution because of their powerful generative priors

local-aiarxiv-cs-cv
18 May 2026
Research

IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation

DGX agent

arXiv:2605.16258v1 Announce Type: cross Abstract: Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most ex

researcharxiv-cs-ai
18 May 2026
Model Releases

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

DGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

DGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

model-releasesarxiv-cs-ro
18 May 2026
Research

PanoWorld: Geometry-Consistent Panoramic Video World Modeling

DGX agent

arXiv:2605.15391v1 Announce Type: cross Abstract: We present PanoWorld, a panoramic video world model that generates geometry-consistent 360egree video from a single image and a caption. Existing pano

researcharxiv-cs-ai
18 May 2026
Safety

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

DGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

safetyarxiv-cs-lg
18 May 2026
Agents

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

DGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

agentsarxiv-cs-ai
18 May 2026
Local Ai

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

DGX agent

arXiv:2605.16138v1 Announce Type: cross Abstract: Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on

local-aiarxiv-cs-ai
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation

DGX agent

arXiv:2605.16079v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown significant progress in video understanding, yet they face substantial challenges in tasks requiring p

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Cognitive-Uncertainty Guided Knowledge Distillation for Accurate Classification of Student Misconceptions

DGX agent

arXiv:2605.14752v1 Announce Type: cross Abstract: Accurately identifying student misconceptions is crucial for personalized education but faces three challenges: (1) data scarcity with long-tail distr

model-releasesarxiv-cs-ai
15 May 2026
Agents

DriveCtrl: Conditioned Sim-to-Real Driving Video Generation

DGX agent

arXiv:2605.15116v1 Announce Type: new Abstract: Large-scale labelled driving video data is essential for training autonomous driving systems. Although simulation offers scalable and fully annotated da

agentsarxiv-cs-cv
15 May 2026
Model Releases

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations

DGX agent

arXiv:2605.14498v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly serve as personal assistants and workplace collaborators, where their utility depends on memory systems t

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

HDRFace: Rethinking Face Restoration with High-Dimensional Representation

DGX agent

arXiv:2605.14821v1 Announce Type: new Abstract: Face restoration under complex degradations still remains an ill-posed inverse problem due to severe information loss. Although diffusion models benefit

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory

DGX agent

arXiv:2605.15128v1 Announce Type: cross Abstract: Long-term agent memory is increasingly multimodal, yet existing evaluations rarely test whether agents preserve the visual evidence needed for later r

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

PolitNuggets: Benchmarking Agentic Discovery of Long-Tail Political Facts

DGX agent

arXiv:2605.14002v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) embedded in agentic frameworks have transformed information retrieval from static, long context question answering into op

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Sat3DGen: Comprehensive Street-Level 3D Scene Generation from Single Satellite Image

DGX agent

arXiv:2605.14984v1 Announce Type: cross Abstract: Generating a street-level 3D scene from a single satellite image is a crucial yet challenging task. Current methods present a stark trade-off: geometr

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades

DGX agent

arXiv:2605.14415v1 Announce Type: cross Abstract: Coding agents powered by large language models are increasingly expected to perform realistic software maintenance tasks beyond isolated issue resolut

model-releasesarxiv-cs-ai
15 May 2026
Local Ai

Venus-DeFakerOne: Unified Fake Image Detection & Localization

DGX agent

arXiv:2605.14091v1 Announce Type: new Abstract: In recent years, the rapid evolution of generative AI has fundamentally reshaped the paradigm of image forgery, breaking the traditional boundaries betw

local-aiarxiv-cs-cv
15 May 2026
Research

WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition

DGX agent

arXiv:2603.09921v3 Announce Type: replace Abstract: Open-domain visual entity recognition (VER) seeks to associate images with entities in encyclopedic knowledge bases such as Wikipedia. Recent genera

researcharxiv-cs-cv
15 May 2026
Research

Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models

DGX agent

arXiv:2605.12517v1 Announce Type: cross Abstract: Vision-language models (VLMs) are often deployed on text-only inputs, although they are trained with images. We find that removing the vision modality

researcharxiv-cs-ai
14 May 2026
Model Releases

ConRetroBert: EMA Stabilized Dual Encoders for Template-Based Single-Step Retrosynthesis

DGX agent

arXiv:2605.12736v1 Announce Type: new Abstract: Template based single step retrosynthesis predicts reactants by selecting and applying an explicit reaction template, making each prediction traceable t

model-releasesarxiv-cs-lg
14 May 2026
Research

Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation

DGX agent

arXiv:2605.13122v1 Announce Type: new Abstract: Instruction-based image editing (IIE) models have recently demonstrated strong capability in modifying specific image regions according to natural langu

researcharxiv-cs-cv
14 May 2026
Model Releases

EcoGEO: Trajectory-Aware Evidence Ecosystems for Web-Enabled LLM Search Agents

DGX agent

arXiv:2605.12887v1 Announce Type: cross Abstract: Web-enabled LLM agents are changing how online information influences search outcomes. Existing Generative Engine Optimization (GEO) studies mainly fo

model-releasesarxiv-cs-ai
14 May 2026
Research

GAAMA: Graph Augmented Associative Memory for Agents

DGX agent

arXiv:2603.27910v2 Announce Type: replace Abstract: AI agents that interact with users across multiple sessions require persistent long-term memory to maintain coherent, personalized behavior. Current

researcharxiv-cs-ai
14 May 2026
Safety

Generative Texture Diversification of 3D Pedestrians for Robust Autonomous Driving Perception

DGX agent

arXiv:2605.13755v1 Announce Type: new Abstract: In recent years, autonomous driving has significantly in creased the demand for high-quality data to train 2D and 3D perception models for safety-critic

safetyarxiv-cs-cv
14 May 2026
Model Releases

GeomHair: Reconstruction of Hair Strands from Colorless 3D Scans

DGX agent

arXiv:2505.05376v3 Announce Type: replace Abstract: We propose a novel method that reconstructs hair strands directly from colorless 3D scans by leveraging multi-modal hair orientation extraction. Hai

model-releasesarxiv-cs-cv
14 May 2026
Research

HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation

DGX agent

arXiv:2605.13586v1 Announce Type: cross Abstract: Generating controllable and physically plausible indoor scenes is a pivotal prerequisite for constructing high-fidelity simulation environments for em

researcharxiv-cs-ai
14 May 2026
Safety

HIR-ALIGN: Enhancing Hyperspectral Image Restoration via Diffusion-Based Data Generation

DGX agent

arXiv:2605.13581v1 Announce Type: new Abstract: Hyperspectral image (HSI) restoration is crucial for reliable analysis, as real HSIs suffer from degradations like noise, blur, and resolution loss. How

safetyarxiv-cs-cv
14 May 2026
Safety

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

DGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

safetyarxiv-cs-ai
14 May 2026
Model Releases

OmniLiDAR: A Unified Diffusion Framework for Multi-Domain 3D LiDAR Generation

DGX agent

arXiv:2605.13815v1 Announce Type: new Abstract: LiDAR scene generation is increasingly important for scalable simulation and synthetic data creation, especially under diverse sensing conditions that a

model-releasesarxiv-cs-cv
14 May 2026
Safety

PRA-PoE: Robust Alzheimer's Diagnosis with Arbitrary Missing Modalities

DGX agent

arXiv:2605.13081v1 Announce Type: new Abstract: Missing modalities are prevalent in real-world Alzheimer's disease (AD) assessment and pose a significant challenge to multimodal learning, particularly

safetyarxiv-cs-cv
14 May 2026
Safety

Real2Sim: A Physics-driven and Editable Gaussian Splatting Framework for Autonomous Driving Scenes

DGX agent

arXiv:2605.13591v1 Announce Type: new Abstract: Reliable autonomous driving relies on large-scale, well-labeled data and robust models. However, manual data collection is resource-intensive, and tradi

safetyarxiv-cs-cv
14 May 2026
Model Releases

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning

DGX agent

arXiv:2507.01908v3 Announce Type: replace Abstract: Instruction-based image editing (IIE) has advanced rapidly with the success of diffusion models. However, existing efforts primarily focus on simple

model-releasesarxiv-cs-cv
14 May 2026
← Previous
1…4344454647…57
Next →