AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
1 Jun 2026

Beyond Tokens: Enhancing RTL Quality Estimation via Structural Graph Learning

ResearchDGX agent

arXiv:2508.18730v2 Announce Type: replace Abstract: Estimating the quality of register transfer level (RTL) designs is crucial in the electronic design automation (EDA) workflow, as it enables instant

Chatterbox-Flash: Prior-Calibrated Block Diffusion for Streaming Zero-Shot TTS

ResearchDGX agent

arXiv:2605.30748v1 Announce Type: cross Abstract: We present Chatterbox-Flash, a zero-shot text-to-speech model obtained by fine-tuning a pretrained autoregressive TTS decoder into a block-diffusion d

Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery

SafetyDGX agent

arXiv:2509.20941v2 Announce Type: replace Abstract: As surgical AI transitions from pixel-level detection to complex reasoning, Scene Graphs (SGs) offer the structured, relational representations nece


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Developing a Culturally Grounded, AI-Augmented UX Research Point of View (POV): An Exemplar Case Study from Telemedicine Dementia Care

TutorialsDGX agent

arXiv:2605.31147v1 Announce Type: cross Abstract: User Experience Research (UXR) Points of View (POVs) distil complex and often fragmented research evidence into actionable perspectives that guide how

Developing a UXR Point of View for Cognitive Accessibility in Mobile Learning with Generative AI

ResearchDGX agent

arXiv:2605.31149v1 Announce Type: cross Abstract: This study investigates how UX research (UXR) principles, combined with Large Language Model (LLM)-supported analysis, can be used to improve the qual

DTBench: A Synthetic Benchmark for Document-to-Table Extraction

Model ReleasesDGX agent

arXiv:2602.13812v3 Announce Type: replace-cross Abstract: Document-to-table (Doc2Table) extraction derives structured tables from unstructured documents under a target schema, enabling reliable and ve

Effective Biological Representation Learning by Masking Gene Expression

ResearchDGX agent

arXiv:2605.31562v1 Announce Type: new Abstract: RNA sequencing produces rich and diverse datasets of gene expression, offering compelling insights into cellular state and function that have many appli

Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship

Model ReleasesDGX agent

arXiv:2605.30947v1 Announce Type: new Abstract: LLM-based research agents have advanced rapidly in science and engineering, where research is organized around executable experiments, code, and quantit

Extending the UXR Point of View Pyramid: A Generative AI-Augmented Methodology for Human-Centred AI Systems

SafetyDGX agent

arXiv:2605.31143v1 Announce Type: cross Abstract: Rising household debt and cost-of-living pressures in the United Kingdom have intensified the role of AI-driven financial technologies in mediating cr

Fighting Numerical Hallucinations via Data-centric Compilation for Online Financial QA

AgentsDGX agent

arXiv:2605.31064v1 Announce Type: cross Abstract: Large Language Models (LLMs) have significantly advanced online data services, particularly in the domain of financial question answering (FinQA). How

Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report Generation

Model ReleasesDGX agent

arXiv:2605.30984v1 Announce Type: cross Abstract: Modern 3D medical vision-language models (VLMs) can generate fluent radiology-style text while exhibit critically low pathology detection and output d

GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration

ApplicationsDGX agent

arXiv:2605.31039v1 Announce Type: new Abstract: Real-world image restoration (IR) is bottlenecked by the scarcity of high-quality paired training data. Synthetic datasets are abundant but often fail t

Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response

SafetyDGX agent

arXiv:2605.30680v1 Announce Type: new Abstract: Healthcare mechanisms are inseparable from the strategic provider response they induce: existing healthcare AI benchmarks hold this response fixed and s

Learning Global Motion with Compact Gaussians for Feed-Forward 4D Reconstruction

ResearchDGX agent

arXiv:2605.31595v1 Announce Type: new Abstract: Dynamic scene reconstruction from monocular video remains a fundamental challenge in computer vision. Existing feed-forward methods predict 3D Gaussians

MultiAct: Text-to-Motion Generation from Composite Text via Tailored Attention Guidance

ResearchDGX agent

arXiv:2605.30925v1 Announce Type: new Abstract: Text-to-motion generation has progressed rapidly in recent years, offering an expressive interface for animation and human-computer interaction. However

Optimizing Rank for High-Fidelity Implicit Neural Representations

SafetyDGX agent

arXiv:2512.14366v2 Announce Type: replace Abstract: Implicit Neural Representations (INRs) based on vanilla Multi-Layer Perceptrons (MLPs) are widely believed to be incapable of representing high-freq

Planner-Centric Reinforcement Learning for Deep Research with Structure-Aware Reward

ResearchDGX agent

arXiv:2605.30824v1 Announce Type: new Abstract: Deep research tasks require LLMs to plan what to investigate, retrieve evidence, and synthesize long-form answers across multiple branches of inquiry. E

SCOPE: Self-Play via Co-Evolving Policies for Open-Ended Tasks

ResearchDGX agent

arXiv:2605.31433v1 Announce Type: new Abstract: Self-play can train language models without external supervision. However, existing methods require rule-checkable answers, leaving open-ended tasks dep

SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation

ResearchDGX agent

arXiv:2605.31033v1 Announce Type: new Abstract: Streaming video generation models typically rely on temporal-centric memory, which organizes historical context as raw frames, chunk segments, or unclus

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

Model ReleasesDGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

Symbolic Intermediaries as a Linguistic-Numerical Interface for LLM-Driven Geometric Reasoning

Model ReleasesDGX agent

arXiv:2505.17607v3 Announce Type: replace Abstract: Large Language Models (LLMs) display reasoning capabilities over linguistic and symbolic objects but have limited capabilities to directly interpret

TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens

ResearchDGX agent

arXiv:2605.31294v1 Announce Type: new Abstract: Recent advances in Audio-LLMs like GPT-4o have ushered in an era of conversational interaction with language models. Conversational avatars however, sti

Triangle Splatting SLAM

ResearchDGX agent

arXiv:2605.31419v1 Announce Type: new Abstract: We present a dense RGB-D SLAM system using differentiable triangles as the 3D map representation. While 3D Gaussian Splatting has emerged as the leading

UniMedVL: Unifying Medical Multimodal Understanding and Generation through Observation-Knowledge-Analysis

ResearchDGX agent

arXiv:2510.15710v3 Announce Type: replace Abstract: Medical workflows routinely combine reading images with producing visual and textual outputs, making both image understanding and generation central

UXR PoV for Neuroinclusive Emotion Regulation

SafetyDGX agent

arXiv:2605.31131v1 Announce Type: cross Abstract: Attention-deficit/hyperactivity disorder (ADHD) is a psychiatric disorder which presents itself in individuals through patterns of developmentally ina

Value Functions as Supermartingale Certificates

SafetyDGX agent

arXiv:2605.31524v1 Announce Type: new Abstract: Certification methods for stochastic systems provide sufficient proof rules, based on real-valued supermartingale certificates, to determine the almost-

30 May 2026

Open source : Turning vocal imitations into sound effects. (New UX for sound generation)

Local AiDGX agent

An open-source system that converts vocal imitations—human-made sound recreations—into synthesized sound effects for creative applications. The technology produces sound effects from vocal imitations

29 May 2026

Bridging Functional and Representational Similarity via Usable Information

ResearchDGX agent

arXiv:2601.21568v2 Announce Type: replace Abstract: We present a unified framework for quantifying the similarity between representations through the lens of extit{usable} information, offering a rigo

Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation

Model ReleasesDGX agent

arXiv:2605.30000v1 Announce Type: new Abstract: Front-end web code has become a core product surface for every frontier LLM release, yet evaluating these interactive applications at development speed

DeepTool: Scaling Interleaved Deliberation in Tool-Integrated Reasoning via Process-Supervised Reinforcement Learning

ResearchDGX agent

arXiv:2605.29568v1 Announce Type: new Abstract: Tool-Integrated Reasoning (TIR) extends LLM capabilities by leveraging external environments. However, existing methods lack the deliberation during seq

Evolutionary Refinement of Generative Graph Topologies: A Hybrid WGAN-GA Approach

SafetyDGX agent

arXiv:2605.29161v1 Announce Type: cross Abstract: Generating realistic graph-structured data is challenging due to discrete connectivity, varying graph sizes, and class-specific structural patterns. R

F-RNG: Feed-Forward Relightable Neural Gaussians

ApplicationsDGX agent

arXiv:2605.25975v2 Announce Type: replace-cross Abstract: Capturing relightable 3D assets from real-world objects is a widely researched problem. Several per-scene optimization-based methods, based on

Knowing What to Solve Before How: Preplan Empowered LLM Mathematical Reasoning

TutorialsDGX agent

arXiv:2605.30245v1 Announce Type: new Abstract: Current plan-based reasoning methods improve large language models (LLMs) by inserting a planning stage before execution, giving rise to the question ri

LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents

Model ReleasesDGX agent

arXiv:2605.29559v1 Announce Type: new Abstract: Mastering terminal environments requires language agents capable of multi-step planning, feedback-grounded execution, and dynamic state adaptation. Howe

LiveSVG: Zero-Shot SVG Animation via Video Generation

Model ReleasesDGX agent

arXiv:2605.30174v1 Announce Type: new Abstract: We introduce LiveSVG, a zero-shot approach for generating Scalable Vector Graphics (SVG) animations using video diffusion models. Current SVG animation

On the Geometry of Games and their Solvers

SafetyDGX agent

arXiv:2605.29919v1 Announce Type: new Abstract: A central challenge in game theory and learning systems such as GANs is understanding which algorithms can efficiently compute equilibria across the het

Singularity-free dynamical invariants-based quantum control

ResearchDGX agent

arXiv:2510.15340v2 Announce Type: replace-cross Abstract: State preparation is a cornerstone of quantum technologies, underpinning applications in computation, communication, and sensing. Its importan

Supercharging Thermal Gaussian Splatting with Depth Estimation

AgentsDGX agent

arXiv:2605.30328v1 Announce Type: new Abstract: Efficient and robust 3D scene representation is crucial in autonomous driving, robotics, and related fields. While RGB images provide valuable content f

Taming Data Challenges in ML-based Security Tasks Using Generative AI

ResearchDGX agent

arXiv:2507.06092v4 Announce Type: replace-cross Abstract: Machine learning-based supervised classifiers are widely used for security tasks, and their improvement has been largely focused on algorithmi

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

AgentsDGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

TraceCodec: A Compiler-Backed Neural Codec for Stateful Multi-Flow Network Traffic Traces

SafetyDGX agent

arXiv:2605.29941v1 Announce Type: cross Abstract: Critical networking workflows require high-fidelity packet captures (PCAPs) for testing, security analysis, and protocol validation, not just statisti

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

AgentsDGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis

AgentsDGX agent

arXiv:2605.28978v1 Announce Type: new Abstract: Finite Element Analysis (FEA) serves as the cornerstone of modern engineering design. However, its workflow is inherently complex and relies heavily on

28 May 2026

A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

Model ReleasesDGX agent

arXiv:2605.28556v1 Announce Type: new Abstract: As agent capabilities advance, existing benchmarks, such as au^2-Bench, are becoming increasingly saturated. Yet constructing new benchmark tasks remain

Accelerating Diffusion Sampling via Exploiting Local Transition Coherence

ResearchDGX agent

arXiv:2503.09675v3 Announce Type: replace Abstract: Text-based diffusion models have made significant breakthroughs in generating high-quality images and videos from textual descriptions. However, the

Asynchronous Remote Sensing Time-Series Fusion for Cloud Removal and Anytime Reconstruction

Model ReleasesDGX agent

arXiv:2605.27726v1 Announce Type: new Abstract: Frequent cloud cover severely limits the usability of Sentinel-2 (S2) optical time series for Earth surface monitoring. Sentinel-1 (S1) SAR provides all

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning

ApplicationsDGX agent

arXiv:2601.18116v2 Announce Type: replace Abstract: We argue that multi-document reasoning is constrained not only by how much text a model can read, but also by how limited query-time evidence budget

CircuitLM: A Multi-Agent LLM-Aided Design Framework for Generating Circuit Schematics from Natural Language Prompts

AgentsDGX agent

arXiv:2601.04505v3 Announce Type: replace Abstract: Generating accurate circuit schematics from high-level natural language descriptions remains a persistent challenge in electronic design automation

CubePart: An Open-Vocabulary Part-Controllable 3D Generator

ResearchDGX agent

arXiv:2605.28763v1 Announce Type: new Abstract: Interactive 3D assets used in games and simulation are typically decomposed into specific semantic parts to support animation, physics, and scripted beh

D^2Turb: Depth-Aware Simulation and Decoupled Learning for Single-Frame Atmospheric Turbulence Mitigation

TutorialsDGX agent

arXiv:2605.27460v1 Announce Type: new Abstract: Single-frame atmospheric turbulence mitigation is inherently ill-posed due to spatially varying blur coupled with non-rigid geometric distortion. Existi

Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox

Model ReleasesDGX agent

arXiv:2605.27772v1 Announce Type: cross Abstract: Audio large language models (Audio LLMs) demonstrate strong performance on speech understanding tasks, yet their ability to understand paralinguistic

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence

Model ReleasesDGX agent

arXiv:2605.28371v1 Announce Type: new Abstract: Industrial Prognostics and Health Management (PHM) provides a representative case study for a broader challenge in applied machine learning: translating

GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection

AgentsDGX agent

arXiv:2605.28534v1 Announce Type: new Abstract: Despite the rapid progress of multimodal large language models in building Graphical User Interface (GUI) agents, their real-world task completion is fu

IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction

Model ReleasesDGX agent

arXiv:2510.06928v2 Announce Type: replace Abstract: Autoregressive models have emerged as a powerful paradigm for visual content creation, but often overlook the intrinsic structural properties of vis

Integrating Inductive Biases in Transformers via Distillation for Financial Time Series Forecasting

Local AiDGX agent

arXiv:2603.16985v2 Announce Type: replace Abstract: Transformer-based models have been widely adopted for time-series forecasting due to their high representational capacity and architectural flexibil

Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents

AgentsDGX agent

arXiv:2605.28775v1 Announce Type: cross Abstract: Computer-use agents (CUAs) have recently made substantial progress, but deploying a separate large expert for each software domain remains expensive.

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration

Model ReleasesDGX agent

arXiv:2602.20497v3 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in image and video generation tasks. However, the high computational demands of Diffusion Tr

MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation

Model ReleasesDGX agent

arXiv:2605.28173v1 Announce Type: new Abstract: End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page la

MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents

Model ReleasesDGX agent

arXiv:2605.27853v1 Announce Type: new Abstract: We present MolLingo, a multi-agent system that emulates the reasoning process of a chemist to automate molecular design. Existing LLM-based approaches e

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

ApplicationsDGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

← Previous
1…3233343536…47
Next →