AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
21 Apr 2026

MESA: A Training-Free Multi-Exemplar Deep Framework for Restoring Ancient Inscription Textures

SafetyDGX agent

arXiv:2604.17390v1 Announce Type: new Abstract: Ancient inscriptions frequently suffer missing or corrupted regions from fragmentation, erosion, or other damage, hindering reading, and analysis. We re

NullFace: Training-Free Localized Face Anonymization

ApplicationsDGX agent

arXiv:2503.08478v2 Announce Type: replace Abstract: Privacy concerns around ever increasing number of cameras are increasing in today's digital age. Although existing anonymization methods are able to

Personalizing Student-Agent Interactions Using Log-Contextualized Retrieval-Augmented Generation (RAG)

AgentsDGX agent

arXiv:2505.17238v3 Announce Type: replace Abstract: Collaborative dialogue offers rich insights into students' learning and critical thinking, which is essential for personalizing pedagogical agent in


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards

ResearchDGX agent

arXiv:2604.17957v1 Announce Type: new Abstract: Process Reward Models (PRMs) have emerged as a powerful tool for providing step-level feedback when evaluating the reasoning of Large Language Models (L

Same prompts as before, but now in GPT image-generator 2, page excerpts from: 'Eldritch Horrors as Pets: A Guide' 'How Womblenauts Work' 'Ph…

TutorialsDGX agent

Same prompts as before, but now in GPT image-generator 2, page excerpts from: 'Eldritch Horrors as Pets: A Guide' 'How Womblenauts Work' 'Photographs of the People of New York Who Look Like Birds' 'Ca

Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding

AgentsDGX agent

arXiv:2510.15253v3 Announce Type: replace Abstract: Document understanding is critical for applications from financial analysis to scientific discovery. Current approaches, whether OCR-based pipelines

Source-Free Domain Adaptation with Vision-Language Prior

SafetyDGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

Model ReleasesDGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges

ResearchDGX agent

arXiv:2511.05152v2 Announce Type: replace Abstract: Deformable Gaussian Splatting (GS) accomplishes photorealistic dynamic 3-D reconstruction from dense multi-view video (MVV) by learning to deform a

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models

ResearchDGX agent

arXiv:2604.17565v1 Announce Type: new Abstract: Camera-controllable image editing aims to synthesize novel views of a given scene under varying camera poses while strictly preserving cross-view geomet

Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation

Local AiDGX agent

arXiv:2604.17656v1 Announce Type: cross Abstract: Video-to-music (V2M) is the fundamental task of creating background music for an input video. Recent V2M models achieve audiovisual alignment by typic

20 Apr 2026

ATTNPO: Attention-Guided Process Supervision for Efficient Reasoning

ResearchDGX agent

arXiv:2602.09953v2 Announce Type: replace Abstract: Large reasoning models trained with reinforcement learning and verifiable rewards (RLVR) achieve strong performance on complex reasoning tasks, yet

Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards

ResearchDGX agent

arXiv:2510.07774v3 Announce Type: replace Abstract: In this paper, we observe that current models are susceptible to reward hacking, leading to a substantial overestimation of a model's reasoning abil

DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AI

AgentsDGX agent

arXiv:2604.15456v1 Announce Type: new Abstract: Trustworthiness and transparency are essential for the clinical adoption of artificial intelligence (AI) in healthcare and biomedical research. Recent d

Diffusion Autoencoder for Unsupervised Artifact Restoration in Handheld Fundus Images

TutorialsDGX agent

arXiv:2604.15723v1 Announce Type: cross Abstract: The advent of handheld fundus imaging devices has made ophthalmologic diagnosis and disease screening more accessible, efficient, and cost-effective.

FD-NL2SQL: Feedback-Driven Clinical NL2SQL that Improves with Use

ResearchDGX agent

arXiv:2604.15646v1 Announce Type: new Abstract: Clinicians exploring oncology trial repositories often need ad-hoc, multi-constraint queries over biomarkers, endpoints, interventions, and time, yet wr

NK-GAD: Neighbor Knowledge-Enhanced Unsupervised Graph Anomaly Detection

ApplicationsDGX agent

arXiv:2604.15668v1 Announce Type: new Abstract: Graph anomaly detection aims to identify irregular patterns in graph-structured data. Most unsupervised GNN-based methods rely on the homophily assumpti

Reward Modeling for Scientific Writing Evaluation

ResearchDGX agent

arXiv:2601.11374v2 Announce Type: replace Abstract: Scientific writing is an expert-domain task that demands deep domain knowledge, task-specific requirements and reasoning capabilities that leverage

Verification Modulo Tested Library Contracts

ResearchDGX agent

arXiv:2604.15533v1 Announce Type: cross Abstract: We consider the problem of verification modulo tested library contracts as a step towards automating the verification of client programs that use comp

19 Apr 2026

Video of my opening keynote for the 2026 World Modeling Workshop at Mila, Quebec AI Institute: simple but powerful ways of using world model…

ResearchDGX agent

Video of my opening keynote for the 2026 World Modeling Workshop at Mila, Quebec AI Institute: simple but powerful ways of using world models and their latent space (sorry for sound problem 0:20-0:43)

17 Apr 2026

Attention to Mamba: A Recipe for Cross-Architecture Distillation

TutorialsDGX agent

arXiv:2604.14191v1 Announce Type: new Abstract: State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher thro

C2W-Tune: Cavity-to -Wall Transfer Learning for Thin Atrial Wall Segmentation in 3D Late Gadolinium-enhanced Magnetic Resonance

ResearchDGX agent

arXiv:2603.24992v2 Announce Type: replace Abstract: Accurate segmentation of the left atrial (LA) wall in 3D late gadolinium-enhanced MRI (LGE-MRI) is essential for wall thickness mapping and fibrosis

EchoAgent: Towards Reliable Echocardiography Interpretation with 'Eyes','Hands' and 'Minds'

AgentsDGX agent

arXiv:2604.05541v2 Announce Type: replace Abstract: Reliable interpretation of echocardiography (Echo) is crucial for assessing cardiac function, which demands clinicians to synchronously orchestrate

Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks

ResearchDGX agent

arXiv:2601.03448v2 Announce Type: replace Abstract: Language models (LMs) are pre-trained on raw text datasets to generate text sequences token-by-token. While this approach facilitates the learning o

Few-Shot Left Atrial Wall Segmentation in 3D LGE MRI via Meta-Learning

Local AiDGX agent

arXiv:2603.24985v2 Announce Type: replace Abstract: Segmenting the left atrial wall from late gadolinium enhancement magnetic resonance images (MRI) is challenging due to the wall's thin geometry, low

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution

TutorialsDGX agent

arXiv:2604.15097v1 Announce Type: cross Abstract: This beta technical report asks how reusable experience should be represented so that it can function as effective test-time control and as a substrat

Geometrically Consistent Multi-View Scene Generation from Freehand Sketches

ResearchDGX agent

arXiv:2604.14302v1 Announce Type: new Abstract: We tackle a new problem: generating geometrically consistent multi-view scenes from a single freehand sketch. Freehand sketches are the most geometrical

MapSR: Prompt-Driven Land Cover Map Super-Resolution via Vision Foundation Models

ResearchDGX agent

arXiv:2604.14582v1 Announce Type: new Abstract: High-resolution (HR) land-cover mapping is often constrained by the high cost of dense HR annotations. We revisit this problem from the perspective of m

QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment

SafetyDGX agent

arXiv:2604.14175v1 Announce Type: new Abstract: We present a unified system addressing both Subtask 3 (answer generation) and Subtask 4 (evidence sentence alignment) of the ArchEHR-QA Shared Task. For

ReSS: Learning Reasoning Models for Tabular Data Prediction via Symbolic Scaffold

TutorialsDGX agent

arXiv:2604.13392v1 Announce Type: new Abstract: Tabular data remains prevalent in high-stakes domains such as healthcare and finance, where predictive models are expected to provide both high accuracy

16 Apr 2026

C2: Scalable Rubric-Augmented Reward Modeling from Binary Preferences

ResearchDGX agent

arXiv:2604.13618v1 Announce Type: new Abstract: Rubric-augmented verification guides reward models with explicit evaluation criteria, yielding more reliable judgments than single-model verification. H

Designing synthetic datasets for the real world: Mechanism design and reasoning from first principles

ApplicationsDGX agent

This Google Research work presents guidelines for synthetic data mechanism design and provides insights into generating and evaluating synthetic data at scale. The research introduces a reasoning-driv

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

Model ReleasesDGX agent

arXiv:2604.14032v1 Announce Type: cross Abstract: Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its de

Kwame 2.0: Human-in-the-Loop Generative AI Teaching Assistant for Large Scale Online Coding Education in Africa

TutorialsDGX agent

arXiv:2603.29159v2 Announce Type: replace Abstract: Providing timely and accurate learning support in large-scale online coding courses is challenging, particularly in resource-constrained contexts. W

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging

Model ReleasesDGX agent

arXiv:2604.13756v1 Announce Type: new Abstract: The potential of Multimodal Large Language Models (MLLMs) in domain of medical imaging raise the demands of systematic and rigorous evaluation framework

Power Transform Revisited: Numerically Stable, and Federated

ApplicationsDGX agent

arXiv:2510.04995v3 Announce Type: replace Abstract: Power transforms are popular parametric methods for making data more Gaussian-like, and are widely used as preprocessing steps in statistical analys

RobotPan: A 360^irc Surround-View Robotic Vision System for Embodied Perception

ResearchDGX agent

arXiv:2604.13476v1 Announce Type: cross Abstract: Surround-view perception is increasingly important for robotic navigation and loco-manipulation, especially in human-in-the-loop settings such as tele

Scale-Invariant Sampling in Multi-Arm Bandit Motion Planning for Object Extraction

ResearchDGX agent

arXiv:2604.14026v1 Announce Type: new Abstract: Object extraction tasks often occur in disassembly problems, where bolts, screws, or pins have to be removed from tight, narrow spaces. In such problems

SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation

ResearchDGX agent

arXiv:2604.13941v1 Announce Type: new Abstract: Local feature matching plays a critical role in understanding the correspondence between cross-view images. However, traditional methods are constrained

UNBOX: Unveiling Black-box visual models with Natural-language

SafetyDGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

Local AiDGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators

ResearchDGX agent

arXiv:2604.13309v1 Announce Type: new Abstract: In this paper we present a novel visual servoing framework to control a robotic manipulator in the configuration space by using purely natural visual fe

Why Your Agents Can’t Read Enterprise Documents — and How to Fix It

TutorialsDGX agent

AI agents struggle to effectively process and extract information from complex enterprise documents due to limitations in context windows, reasoning capabilities, and handling of unstructured data for

15 Apr 2026

ART-VITON: Measurement-Guided Latent Diffusion for Artifact-Free Virtual Try-On

SafetyDGX agent

arXiv:2509.25749v2 Announce Type: cross Abstract: Virtual try-on (VITON) aims to generate realistic images of a person wearing a target garment, requiring precise garment alignment in try-on regions a

JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence

ResearchDGX agent

arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate

Mid-training is essential for LLM reasoning, IBM study shows

ResearchDGX agent

IBM Research has demonstrated that mid-training — a dedicated training phase between initial pre-training and fine-tuning — is essential for improving reasoning capabilities in large language models (

One Model for All: Unified Try-On and Try-Off in Any Pose via LLM-Inspired Bidirectional Tweedie Diffusion

ApplicationsDGX agent

arXiv:2508.04559v3 Announce Type: replace Abstract: Recent diffusion-based approaches have made significant advances in image-based virtual try-on, enabling more realistic and end-to-end garment synth

Siamese Foundation Models for Crystal Structure Prediction

TutorialsDGX agent

arXiv:2503.10471v2 Announce Type: replace-cross Abstract: Predicting crystal structures from chemical compositions is a fundamental challenge in materials discovery, complicated by complex 3D geometri

Using Learning Progressions to Guide AI Feedback for Science Learning

TutorialsDGX agent

arXiv:2603.03249v2 Announce Type: replace Abstract: Generative artificial intelligence (AI) offers scalable support for formative feedback, yet most AI-generated feedback relies on task-specific rubri

14 Apr 2026

A Diffusion-Contrastive Graph Neural Network with Virtual Nodes for Wind Nowcasting in Unobserved Regions

TutorialsDGX agent

arXiv:2604.10328v1 Announce Type: cross Abstract: Accurate weather nowcasting remains one of the central challenges in atmospheric science, with critical implications for climate resilience, energy se

Beyond Reconstruction: Reconstruction-to-Vector Diffusion for Hyperspectral Anomaly Detection

SafetyDGX agent

arXiv:2604.11390v1 Announce Type: new Abstract: While Hyperspectral Anomaly Detection (HAD) excels at identifying sparse targets in complex scenes, existing models remain trapped in a scalar 'reconstr

Dual-Control Frequency-Aware Diffusion Model for Depth-Dependent Optical Microrobot Microscopy Image Generation

AgentsDGX agent

arXiv:2604.11680v1 Announce Type: new Abstract: Optical microrobots actuated by optical tweezers (OT) are important for cell manipulation and microscale assembly, but their autonomous operation depend

EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model

ResearchDGX agent

arXiv:2604.10268v1 Announce Type: new Abstract: We propose EditCrafter, a high-resolution image editing method that operates without tuning, leveraging pretrained text-to-image (T2I) diffusion models

F3G-Avatar : Face Focused Full-body Gaussian Avatar

ResearchDGX agent

arXiv:2604.09835v1 Announce Type: cross Abstract: Existing full-body Gaussian avatar methods primarily optimize global reconstruction quality and often fail to preserve fine-grained facial geometry an

Hierarchical Textual Knowledge for Enhanced Image Clustering

TutorialsDGX agent

arXiv:2604.11144v1 Announce Type: cross Abstract: Image clustering aims to group images in an unsupervised fashion. Traditional methods focus on knowledge from visual space, making it difficult to dis

How to Bridge the Sim-to-Real Gap in Digital Twin-Aided Telecommunication Networks

TutorialsDGX agent

arXiv:2507.07067v3 Announce Type: replace-cross Abstract: Training effective artificial intelligence models for telecommunications is challenging due to the scarcity of deployment-specific data. Real

Iterative Inference-time Scaling with Adaptive Frequency Steering for Image Super-Resolution

ResearchDGX agent

arXiv:2512.23532v2 Announce Type: replace Abstract: Diffusion models have become a leading paradigm for image super-resolution (SR), but existing methods struggle to guarantee both the high-frequency

Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System

Model ReleasesDGX agent

arXiv:2602.09667v2 Announce Type: replace Abstract: Integrating domain knowledge into neural networks is a central challenge in scientific machine learning. Three paradigms have emerged -- data-driven

Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images

ResearchDGX agent

arXiv:2604.10573v1 Announce Type: new Abstract: Robust 3D representation learning forms the perceptual foundation of spatial intelligence, enabling downstream tasks in scene understanding and embodied

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

ResearchDGX agent

arXiv:2505.23606v4 Announce Type: replace-cross Abstract: Unified generation models aim to handle diverse tasks across modalities -- such as text generation, image generation, and vision-language reas

← Previous
1…2021222324…47
Next →