AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,818 results
14 May 2026

Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation

ResearchDGX agent

arXiv:2605.13122v1 Announce Type: new Abstract: Instruction-based image editing (IIE) models have recently demonstrated strong capability in modifying specific image regions according to natural langu

EcoGEO: Trajectory-Aware Evidence Ecosystems for Web-Enabled LLM Search Agents

Model ReleasesDGX agent

arXiv:2605.12887v1 Announce Type: cross Abstract: Web-enabled LLM agents are changing how online information influences search outcomes. Existing Generative Engine Optimization (GEO) studies mainly fo

GAAMA: Graph Augmented Associative Memory for Agents

ResearchDGX agent

arXiv:2603.27910v2 Announce Type: replace Abstract: AI agents that interact with users across multiple sessions require persistent long-term memory to maintain coherent, personalized behavior. Current


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generative Texture Diversification of 3D Pedestrians for Robust Autonomous Driving Perception

SafetyDGX agent

arXiv:2605.13755v1 Announce Type: new Abstract: In recent years, autonomous driving has significantly in creased the demand for high-quality data to train 2D and 3D perception models for safety-critic

GeomHair: Reconstruction of Hair Strands from Colorless 3D Scans

Model ReleasesDGX agent

arXiv:2505.05376v3 Announce Type: replace Abstract: We propose a novel method that reconstructs hair strands directly from colorless 3D scans by leveraging multi-modal hair orientation extraction. Hai

HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation

ResearchDGX agent

arXiv:2605.13586v1 Announce Type: cross Abstract: Generating controllable and physically plausible indoor scenes is a pivotal prerequisite for constructing high-fidelity simulation environments for em

HIR-ALIGN: Enhancing Hyperspectral Image Restoration via Diffusion-Based Data Generation

SafetyDGX agent

arXiv:2605.13581v1 Announce Type: new Abstract: Hyperspectral image (HSI) restoration is crucial for reliable analysis, as real HSIs suffer from degradations like noise, blur, and resolution loss. How

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

SafetyDGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

🌟Introducing🎻Violin — an Open-source Video Translation Skill. 📹Video is the dominant medium on the internet, yet most high-quality conten…

AgentsDGX agent

🌟Introducing🎻Violin — an Open-source Video Translation Skill. 📹Video is the dominant medium on the internet, yet most high-quality content (lecture, talk, podcast) is locked behind a single language,

OmniLiDAR: A Unified Diffusion Framework for Multi-Domain 3D LiDAR Generation

Model ReleasesDGX agent

arXiv:2605.13815v1 Announce Type: new Abstract: LiDAR scene generation is increasingly important for scalable simulation and synthetic data creation, especially under diverse sensing conditions that a

PRA-PoE: Robust Alzheimer's Diagnosis with Arbitrary Missing Modalities

SafetyDGX agent

arXiv:2605.13081v1 Announce Type: new Abstract: Missing modalities are prevalent in real-world Alzheimer's disease (AD) assessment and pose a significant challenge to multimodal learning, particularly

Real2Sim: A Physics-driven and Editable Gaussian Splatting Framework for Autonomous Driving Scenes

SafetyDGX agent

arXiv:2605.13591v1 Announce Type: new Abstract: Reliable autonomous driving relies on large-scale, well-labeled data and robust models. However, manual data collection is resource-intensive, and tradi

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning

Model ReleasesDGX agent

arXiv:2507.01908v3 Announce Type: replace Abstract: Instruction-based image editing (IIE) has advanced rapidly with the success of diffusion models. However, existing efforts primarily focus on simple

Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation

TutorialsDGX agent

arXiv:2605.12975v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a standard approach for knowledge-intensive question answering, but existing systems remain brittle on m

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data

SafetyDGX agent

arXiv:2605.13775v1 Announce Type: cross Abstract: The scalability of robotic manipulation is fundamentally bottlenecked by the scarcity of task-aligned physical interaction data. While vision-language

Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism

Model ReleasesDGX agent

arXiv:2605.12524v1 Announce Type: cross Abstract: We introduce ProofGrid, a benchmark suite for evaluating LLM reasoning through machine-checkable proofs rather than final answers alone. ProofGrid con

Structural Diversity Drives Disruptive Scientific Innovation

SafetyDGX agent

arXiv:2605.12514v1 Announce Type: cross Abstract: Scientific innovation increasingly depends on collaboration, yet the organizational structure that fosters breakthrough ideas remains poorly understoo

Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents

SafetyDGX agent

arXiv:2605.12620v1 Announce Type: new Abstract: Building generalist embodied agents capable of solving complex real-world tasks remains a fundamental challenge in AI. Multimodal Large Language Models

When Diffusion Breaks Constraints: Sequential Autoregressive Generation with RL and MCTS

ResearchDGX agent

arXiv:2512.01242v3 Announce Type: replace-cross Abstract: Data-driven generative models excel in language and vision, but diffusion models often fail in constrained planning and design tasks, exhibiti

13 May 2026

3D-Belief: Embodied Belief Inference via Generative 3D World Modeling

Model ReleasesDGX agent

arXiv:2605.11367v1 Announce Type: new Abstract: Recent advances in visual generative models have highlighted the promise of learning generative world models. However, most existing approaches frame wo

A Survey of On-Policy Distillation for Large Language Models

SafetyDGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

CAST: Collapse-Aware multi-Scale Topology Fusion for Multimodal Coreset Selection

ResearchDGX agent

arXiv:2605.11705v1 Announce Type: new Abstract: The training of large multimodal models fundamentally relies on massive image-text datasets, which inevitably incur prohibitive computational overhead.

Correcting Selection Bias in Sparse User Feedback for Large Language Model Quality Estimation: A Multi-Agent Hierarchical Bayesian Approach

Model ReleasesDGX agent

arXiv:2605.12177v1 Announce Type: new Abstract: [Abridged] Production LLM deployments receive feedback from a non-random fraction of users: thumbs sit mostly in the tails of the satisfaction distribut

DarkQA: Benchmarking Vision-Language Models on Visual-Primitive Question Answering in Low-Light Indoor Scenes

Model ReleasesDGX agent

arXiv:2512.24985v4 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly adopted as central reasoning modules for embodied agents. Existing benchmarks evaluate their capabili

ECHO: Continuous Hierarchical Memory for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2605.10993v1 Announce Type: new Abstract: Memory capacity is a critical factor determining the performance of Vision-Language-Action (VLA) models in long-horizon manipulation tasks. Existing mem

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

SafetyDGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR

Model ReleasesDGX agent

arXiv:2605.11115v1 Announce Type: new Abstract: High Dynamic Range (HDR) generation remains challenging for generative models, which are largely limited to low dynamic range outputs. Recent diffusionb

MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2605.12361v1 Announce Type: new Abstract: Evaluating large language models (LLMs) in the biomedical domain requires benchmarks that can distinguish reasoning from pattern matching and remain dis

PoseCompass: Intelligent Synthetic Pose Selection for Visual Localization

Local AiDGX agent

arXiv:2605.12144v1 Announce Type: new Abstract: In visual localization, Absolute Pose Regression (APR) enables real-time 6-DoF camera pose inference from single images, yet critically depends on fine-

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

Model ReleasesDGX agent

arXiv:2605.11887v1 Announce Type: new Abstract: Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, li

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

AgentsDGX agent

arXiv:2605.12500v1 Announce Type: new Abstract: Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as disti

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images

ResearchDGX agent

arXiv:2605.11462v1 Announce Type: new Abstract: Recent advancements in Large Vision-Language Models (VLMs) have demonstrated exceptional semantic understanding, yet these models consistently struggle

12 May 2026

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

Model ReleasesDGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

Model ReleasesDGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

Model ReleasesDGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care

SafetyDGX agent

arXiv:2605.08480v1 Announce Type: new Abstract: Individuals with Alzheimer's disease (AD) and Alzheimer's disease-related dementia (ADRD) experience memory and thinking changes that impact their abili

AllocMV: Optimal Resource Allocation for Music Video Generation via Structured Persistent State

ResearchDGX agent

arXiv:2605.10723v1 Announce Type: cross Abstract: Generating long-horizon music videos (MVs) is frequently constrained by prohibitive computational costs and difficulty maintaining cross-shot consiste

Anchor-guided Hypergraph Condensation with Dual-level Discrimination

ResearchDGX agent

arXiv:2605.10001v1 Announce Type: new Abstract: The increasing prevalence of large-scale hypergraphs poses significant computational challenges for hypergraph neural network (HNN) training. To address

Attribution-based Explanations for Markov Decision Processes

ResearchDGX agent

arXiv:2605.09780v1 Announce Type: new Abstract: Attribution techniques explain the outcome of an AI model by assigning a numerical score to its inputs. So far, these techniques have mainly focused on

Automated Approach for Solving Infinite-state Polynomial Reachability Games

Model ReleasesDGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD

Model ReleasesDGX agent

arXiv:2605.10865v1 Announce Type: new Abstract: Industrial Computer-Aided Design (CAD) code generation requires models to produce executable parametric programs from visual or textual inputs. Beyond r

Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenizatio

ResearchDGX agent

arXiv:2605.10780v1 Announce Type: cross Abstract: Representation autoencoders that reuse frozen pretrained vision encoders as visual tokenizers have achieved strong reconstruction and generation quali

Communicating Sound Through Natural Language

ResearchDGX agent

arXiv:2605.08750v1 Announce Type: cross Abstract: Natural language is widely used to describe, prompt, and control audio systems, but rarely serves as the representation carrying audio itself. We intr

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision

SafetyDGX agent

arXiv:2509.14234v3 Announce Type: replace Abstract: Where do learning signals come from when there is no ground truth in post-training? We show that inference compute itself can serve as supervision.

ConFixGS: Learning to Fix Feedforward 3D Gaussian Splatting with Confidence-Aware Diffusion Priors in Driving Scenes

TutorialsDGX agent

arXiv:2605.09688v1 Announce Type: new Abstract: Feedforward 3D Gaussian Splatting (3DGS) often struggles in trajectory-based sparse-view driving scenes. Existing Gaussian repair methods mainly target

Containment Verification: AI Safety Guarantees Independent of Alignment

SafetyDGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

Count Anything at Any Granularity

ApplicationsDGX agent

arXiv:2605.10887v1 Announce Type: new Abstract: Open-world object counting remains brittle: despite rapid advances in vision-language models (VLMs), reliably counting the objects a user intends is far

DeformMaster: An Interactive Physics-Neural World Model for Deformable Objects from Videos

Model ReleasesDGX agent

arXiv:2605.09586v1 Announce Type: new Abstract: World models for deformable objects should recover not only geometry and appearance, but also underlying physical dynamics, interaction grounding, and m

EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents

SafetyDGX agent

arXiv:2605.10366v1 Announce Type: new Abstract: Graph reasoning agents operating from natural-language inputs must solve a coupled problem: they must reconstruct a structured graph instance from text,

HH-SAE: Discovering and Steering Hierarchical Knowledge of Complex Manifolds

ResearchDGX agent

arXiv:2605.10536v1 Announce Type: cross Abstract: Rare semantic innovations in high-dimensional, mission-critical domains are often obscured by dense background contexts, a challenge we define as exti

How Sapu Indexed 28 Million PubMed Abstracts to Accelerate Cancer Research with Qdrant

ApplicationsDGX agent

Sapu is an early-stage biopharmaceutical company developing treatments for hard-to-treat cancers. From its San Diego facility, the team is pioneering a nanomedicine pipeline that takes existing FDA-ap

Introducing voice finder — a new tool to quickly find the right voice for your app from over 600+ voices

ToolsDGX agent

Together AI launched Voice Finder, a tool designed to help developers quickly select appropriate voices for their applications from a library of over 600 voice options. The tool streamlines the voice

Learning More from Less: Exploiting Counterfactuals for Data-Efficient Chart Understanding

ResearchDGX agent

arXiv:2605.10855v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in chart understanding, largely driven by supervised fine-tuning (SFT) on increasing

MedMeta: A Benchmark for LLMs in Synthesizing Meta-Analysis Conclusion from Medical Studies

Model ReleasesDGX agent

arXiv:2605.09661v1 Announce Type: cross Abstract: Large language models (LLMs) have saturated standard medical benchmarks that test factual recall, yet their ability to perform higher-order reasoning,

NARRA-Gym for Evaluating Interactive Narrative Agents

Model ReleasesDGX agent

arXiv:2605.08503v1 Announce Type: new Abstract: Interactive narrative tasks require LLMs to sustain a coherent, evolving story while adapting to a user over multiple turns. However, suitable benchmark

ProactBench: Beyond What The User Asked For

Model ReleasesDGX agent

arXiv:2605.09228v1 Announce Type: cross Abstract: Most LLM benchmarks score how well a model responds to explicit requests. They leave unmeasured a different conversational ability: noticing and actin

ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation

SafetyDGX agent

arXiv:2605.08774v1 Announce Type: cross Abstract: Long-horizon robotic manipulation requires dense feedback that reflects how a task advances through its procedural stages, not merely whether the fina

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents

Local AiDGX agent

arXiv:2605.08468v1 Announce Type: cross Abstract: Local LLM-based coding agents increasingly work in settings where correctness is earned through execution feedback, persistent state, and bounded repa

RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards

SafetyDGX agent

arXiv:2605.10899v1 Announce Type: new Abstract: Training deep research agents, namely systems that plan, search, evaluate evidence, and synthesize long-form reports, pushes reinforcement learning beyo

SciIntegrity-Bench: A Benchmark for Evaluating Academic Integrity in AI Scientist Systems

Model ReleasesDGX agent

arXiv:2605.10246v1 Announce Type: new Abstract: AI scientist systems are increasingly deployed for autonomous research, yet their academic integrity has never been systematically evaluated. We introdu

← Previous
1…3637383940…47
Next →