AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
27 May 2026

Object Pose and Shape Estimation for Grasping: Does it Work?

ResearchDGX agent

arXiv:2605.26944v1 Announce Type: cross Abstract: The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models

Verilog-Evolve: Feedback-Driven and Skill-Evolving Verilog Generation

ResearchDGX agent

arXiv:2605.26498v1 Announce Type: new Abstract: Large language models (LLMs) have improved Verilog generation from natural-language specifications, but most pipelines still treat generation as isolate

26 May 2026

Auditing medical multi-agent AI reveals risks of false consensus

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.10185v2 Announce Type: replace-cross Abstract: Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through sp

Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches

SafetyDGX agent

arXiv:2605.23910v1 Announce Type: cross Abstract: Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (mult

Hypothesis Generation and Inductive Inference in Children and Language Models

ApplicationsDGX agent

arXiv:2605.24528v1 Announce Type: new Abstract: Real world decision-making requires constructing mental models under uncertainty over evidence, over the underlying causal rules, and over the state of

25 May 2026

USIM and U0: A Vision-Language-Action Dataset and Model for General Underwater Robots

ResearchDGX agent

arXiv:2510.07869v4 Announce Type: replace Abstract: Underwater environments pose unique challenges for robotic navigation and manipulation. While existing research has primarily focused on task-specif

21 May 2026

AI-Assisted Scientific Assessment: A Case Study on Climate Change

Model ReleasesDGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

20 May 2026

HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance

SafetyDGX agent

arXiv:2506.07209v2 Announce Type: replace-cross Abstract: We present HOI-PAGE, a new approach that prioritizes part-level affordance reasoning to generate high-fidelity 4D human-object interactions (H

LiFT: Lifted Inter-slice Feature Trajectories for 3D Image Generation from 2D Generators

ResearchDGX agent

arXiv:2605.19060v1 Announce Type: cross Abstract: High-resolution 3D medical image generation remains challenging because fully volumetric models are computationally expensive, while efficient 2D slic

Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions

Model ReleasesDGX agent

arXiv:2605.19895v1 Announce Type: new Abstract: Constraint programming practitioners accelerate hard problems through a layered set of techniques applied in order of risk. Standard hardening (symmetry

19 May 2026

Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models

AgentsDGX agent

arXiv:2510.07799v2 Announce Type: replace-cross Abstract: The efficiency of multi-agent systems driven by large language models (LLMs) largely hinges on their communication topology. However, designin

HAD: Hallucination-Aware Diffusion Priors for 3D Reconstruction

ResearchDGX agent

arXiv:2605.16873v1 Announce Type: new Abstract: Diffusion priors have recently demonstrated strong capability in enhancing the quality of sparse-view 3D reconstruction by augmenting training views at

Knowledge-to-Verification: Exploring RLVR for LLMs in Knowledge-Intensive Domains

ResearchDGX agent

arXiv:2605.18261v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has demonstrated promising potential to enhance the reasoning capabilities of large language model

On Applicability of Synthetic Datasets for Facial Expression Recognition

ResearchDGX agent

arXiv:2605.17483v1 Announce Type: new Abstract: Facial Expression Recognition faces two core challenges. The first is class imbalance in public datasets, which skews the learning process and weakens g

SynVA: A Modular Toolkit for Vessel Generation and Aneurysm Editing

ResearchDGX agent

arXiv:2605.17620v1 Announce Type: cross Abstract: Intracranial aneurysms (IAs), characterized by unpredictable growth and risk of rupture, are a major cause of stroke and can lead to life-threatening

UniPPTBench: A Unified Benchmark for Presentation Generation Across Diverse Input Settings

Model ReleasesDGX agent

arXiv:2605.17356v1 Announce Type: new Abstract: Existing works typically focus on presentation generation under isolated input settings, whereas real-world use cases span diverse scenarios, including

Vector RAG vs LLM-Compiled Wiki: A Preregistered Comparison on a Small Multi-Domain Research

ResearchDGX agent

arXiv:2605.18490v1 Announce Type: new Abstract: We preregistered a comparison of two ways to help an LLM answer questions over a small research corpus: a single-round Vector RAG system and an LLM-comp

15 May 2026

Generating HDR Video from SDR Video

ResearchDGX agent

arXiv:2605.14703v1 Announce Type: new Abstract: The high dynamic range (HDR) video ecosystem is approaching maturity, but the problem of upconverting legacy standard dynamic range (SDR) videos persist

14 May 2026

Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

SafetyDGX agent

arXiv:2605.12729v1 Announce Type: cross Abstract: Large language models are increasingly being used to support network operations (NetOps) and artificial intelligence for IT operations (AIOps), includ

Phy-CoSF: Physics-Guided Continuous Spectral Fields Reconstruction and Super-Resolution for Snapshot Compressive Imaging

ResearchDGX agent

arXiv:2605.13583v1 Announce Type: new Abstract: Recent advances have demonstrated that coded aperture snapshot spectral imaging (CASSI) systems show great potential for capturing 3D hyperspectral imag

13 May 2026

Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling

ResearchDGX agent

arXiv:2605.11666v1 Announce Type: new Abstract: The reasoning frontier of Large Language Models (LLMs) has advanced significantly through modern post-training paradigms (e.g., Reinforcement Learning f

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

SafetyDGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

12 May 2026

Engineering Robustness into Personal Agents with the AI Workflow Store

AgentsDGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

HardwareDGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding

Model ReleasesDGX agent

arXiv:2605.08412v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made rapid progress in single-video understanding, yet their ability to reason across multiple independent

Towards Generative Predictive Display for Vision-Based Teleoperation: A Zero-Shot Benchmark of Off-the-Shelf Video Models

Model ReleasesDGX agent

arXiv:2605.09670v1 Announce Type: cross Abstract: Teleoperation systems are fundamentally limited by communication latency, which degrades situational awareness and control performance. Predictive dis

11 May 2026

A^2RD: Agentic Autoregressive Diffusion for Long Video Consistency

Model ReleasesDGX agent

arXiv:2605.06924v1 Announce Type: cross Abstract: Synthesizing consistent and coherent long video remains a fundamental challenge. Existing methods suffer from semantic drift and narrative collapse ov

VDCook:DIY video data cook your MLLMs

AgentsDGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

VDEGaussian: Video Diffusion Enhanced 4D Gaussian Splatting for Dynamic Urban Scenes Modeling

SafetyDGX agent

arXiv:2508.02129v2 Announce Type: replace Abstract: Dynamic urban scene modeling is a rapidly evolving area with broad applications. While current approaches leveraging neural radiance fields or Gauss

7 May 2026

Agentic publications: redesigning scientific publishing in the age of thinking large language models

AgentsDGX agent

arXiv:2505.13246v2 Announce Type: replace Abstract: Purpose: This paper introduces the concept of 'Agentic Publication,' a novel LLM-driven framework designed to complement traditional scientific publ

RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos

Model ReleasesDGX agent

arXiv:2412.03077v2 Announce Type: replace Abstract: 4D reconstruction from casually captured monocular videos is challenging due to inherent ambiguity in reconstructing dynamic 3D geometry. To address

UniPCB: A Generation-Assisted Detection Framework for PCB Defect Inspection

ResearchDGX agent

arXiv:2605.04635v1 Announce Type: new Abstract: Printed Circuit Board (PCB) defect inspection faces two compounding challenges: scarce and imbalanced defect samples that limit model training, and insu

5 May 2026

Combining Trained Models in Reinforcement Learning

SafetyDGX agent

arXiv:2605.02159v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) has delivered strong results in domains such as Atari and Go, but it still suffers from high sample cost and weak tran

Neural Cellular Automata: From Cells to Pixels

Local AiDGX agent

arXiv:2506.22899v3 Announce Type: replace Abstract: Neural Cellular Automata (NCAs) are bio-inspired dynamical systems in which identical cells iteratively apply a learned local update rule to self-or

StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.01939v1 Announce Type: new Abstract: Static benchmarks for LLMs are increasingly compromised by contamination and overfitting especially on knowledge intensive reasoning tasks While recent

SVGS: Enhancing Gaussian Splatting Using Primitives with Spatially Varying Colors

ApplicationsDGX agent

arXiv:2411.18966v2 Announce Type: replace Abstract: Gaussian Splatting demonstrates impressive results in multi-view reconstruction based on Gaussian explicit representations. However, the current Gau

1 May 2026

AutoSurfer -- Teaching Web Agents through Comprehensive Surfing, Learning, and Modeling

Model ReleasesDGX agent

arXiv:2604.27253v1 Announce Type: new Abstract: Recent advances in multimodal large language models (LLMs) have revolutionized web agents that can automate complex tasks on websites. However, their ac

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

Model ReleasesDGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.27629v1 Announce Type: new Abstract: We present WaferSAGE, a framework for wafer defect visual question answering using small vision-language models. To address data scarcity in semiconduct

29 Apr 2026

Accurate and Robust Generative Approach for Overcoming Data Sparsity and Imbalance in Landslide Modeling with A Tabular Foundation Model

TutorialsDGX agent

arXiv:2604.25159v1 Announce Type: new Abstract: Landslide investigation relies on sufficient and well-balanced observational data influenced by geological, hydrological, and anthropogenic factors. Ava

C3G: Learning Compact 3D Representations with 2K Gaussians

TutorialsDGX agent

arXiv:2512.04021v2 Announce Type: replace Abstract: Reconstructing and understanding 3D scenes from unposed sparse views in a feed-forward manner remains as a challenging task in 3D computer vision. R

Contrastive Image-Metadata Pre-Training for Materials Transmission Electron Microscopy

TutorialsDGX agent

arXiv:2604.24909v1 Announce Type: new Abstract: The vast majority of transmission electron microscopy (TEM) data never gets published and ends up on a backup drive until deleted to free up space. Thes

Korean aegyo speech shows systematic F1 increase to signal childlike qualities

ResearchDGX agent

arXiv:2604.25133v1 Announce Type: new Abstract: Korean aegyo is a socially recognized childlike speaking style used predominantly in romantic interactions among adults. This study examined vowel space

LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model

ApplicationsDGX agent

arXiv:2604.25297v1 Announce Type: new Abstract: In recent years, the rapid proliferation of open-source large language models (LLMs) has spurred efforts to turn general-purpose models into domain spec

MiMo-Embodied: X-Embodied Foundation Model Technical Report

AgentsDGX agent

arXiv:2511.16518v2 Announce Type: replace-cross Abstract: We open-source MiMo-Embodied, the first cross-embodied foundation model to successfully integrate and achieve state-of-the-art performance in

Novel 3D Binary Indexed Tree for Volume Computation of 3D Reconstructed Models from Volumetric Data

ResearchDGX agent

arXiv:2412.10441v2 Announce Type: replace-cross Abstract: In the burgeoning field of medical imaging, precise computation of 3D volume holds a significant importance for subsequent qualitative analysi

28 Apr 2026

AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance

SafetyDGX agent

arXiv:2604.23909v1 Announce Type: new Abstract: Navigational aids for blind and low vision individuals struggle conveying dynamic real-world environments, leading to cognitive overload from continuous

BERT-APC: A Reference-free Framework for Automatic Pitch Correction via Musical Context Inference

ResearchDGX agent

arXiv:2511.20006v2 Announce Type: replace-cross Abstract: Automatic Pitch Correction (APC) enhances vocal recordings by aligning pitch deviations with intended musical notes. However, existing APC sys

Coarse-to-Real: Generative Rendering for Populated Dynamic Scenes

ResearchDGX agent

arXiv:2601.22301v2 Announce Type: replace Abstract: Traditional rendering pipelines rely on complex assets, accurate materials and lighting, and substantial computational resources to produce realisti

Cooptimizing Safety and Performance Using Safety Value-Constrained Model Predictive Control

SafetyDGX agent

arXiv:2604.23863v1 Announce Type: new Abstract: Autonomous systems are increasingly deployed in real-world environments, where they must achieve high performance while maintaining safety under state a

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

Model ReleasesDGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata

AgentsDGX agent

arXiv:2510.18289v2 Announce Type: replace Abstract: Food insecurity remains a persistent public health emergency in the United States, tightly interwoven with chronic disease, mental illness, and opio

From Equations to Algorithms and Data: Transforming Microwave Engineering and Education with Machine Learning

ApplicationsDGX agent

arXiv:2604.22792v1 Announce Type: cross Abstract: Conventional microwave engineering education relies heavily on analytical methods, canonical circuit topologies, and intuition-driven design, which ha

Guiding Vector Field Generation via Score-based Diffusion Model

ResearchDGX agent

arXiv:2604.24487v1 Announce Type: new Abstract: Guiding Vector Fields (GVFs) are a powerful tool for robotic path following. However, classical methods assume smooth, ordered curves and fail when path

Infrastructure-Guided Connectivity-Enhanced Road Crack Detection and Estimation

ResearchDGX agent

arXiv:2604.24616v1 Announce Type: new Abstract: In this paper, we report the world's first infrastructure-guided communication-enhanced road crack detection pipeline that is effective and implementabl

LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation

ApplicationsDGX agent

arXiv:2505.12009v2 Announce Type: replace Abstract: Expressive human pose and shape estimation (EHPS) plays a central role in digital human generation, particularly in live-streaming applications. How

Learning to Decipher from Pixels -- A Case Study of Copiale

ApplicationsDGX agent

arXiv:2604.23683v1 Announce Type: new Abstract: Historical encrypted manuscripts require both paleographic interpretation of cipher symbols and cryptanalytic recovery of plaintext. Most existing compu

MuSc-V2: Zero-Shot Multimodal Industrial Anomaly Classification and Segmentation with Mutual Scoring of Unlabeled Samples

ResearchDGX agent

arXiv:2511.10047v2 Announce Type: replace Abstract: Zero-shot anomaly classification (AC) and segmentation (AS) methods aim to identify and outline defects without using any labeled samples. In this p

Point-MF: One-step Point Cloud Generation from a Single Image via Mean Flows

TutorialsDGX agent

arXiv:2604.24586v1 Announce Type: new Abstract: Single-image point cloud reconstruction must infer complete 3D geometry, including occluded parts, from a single RGB image. While diffusion-based recons

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

ResearchDGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

← Previous
1…1819202122…47
Next →