AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,844 results
4 Aug 2026

Learning to Tessellate: Point Cloud Generation via Recursive Spectral Partitioning

ResearchDGX agent

arXiv:2608.02432v1 Announce Type: new Abstract: Autoregressive models have emerged as an effective paradigm for point cloud generation. However, most existing approaches rely on heuristic tokenization

Lethe: How Hard Is It to Forget? A Benchmark for Federated Unlearning in Medical Imaging

Model ReleasesDGX agent

arXiv:2608.01094v1 Announce Type: new Abstract: Federated learning enables medical-imaging models to be trained across hospitals, and privacy law, most explicitly the GDPR ``right to be forgotten'', t

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

AgentsDGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Model ReleasesDGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

Optimizing Minimax Regret in Uncertain MDPs with Small Sets of Policies

SafetyDGX agent

arXiv:2608.02509v1 Announce Type: cross Abstract: Sequential decision-making in real-world applications often involves uncertainty about the environment's model. Uncertain Markov decision processes (U

Practical Noise Modeling for SPAD Intensity Imaging

SafetyDGX agent

arXiv:2608.00489v1 Announce Type: new Abstract: Single-photon avalanche diode (SPAD) cameras are promising for low-light and high-dynamic-range intensity imaging, but their practical use is limited by

Prompt-Driven Simulation with Feature Perturbation for Cross-Domain Few-Shot Object Detection

ResearchDGX agent

arXiv:2608.01348v1 Announce Type: new Abstract: Data augmentation, which simulates diverse visual variations to expand the source distribution and induce synthetic domain shifts, is a simple yet effec

QuerySplat: Decoupling Geometry and Appearance Representations in 3DGS Prediction

Model ReleasesDGX agent

arXiv:2608.01186v1 Announce Type: new Abstract: While feed-forward 3D Gaussian Splatting (3DGS) enables efficient 3D reconstruction, achieving high-fidelity rendering remains challenging. Existing pix

Recursive Gaussian Processes and the Bayesian Brain

ResearchDGX agent

arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobio

RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals

ApplicationsDGX agent

arXiv:2608.00289v1 Announce Type: cross Abstract: Recognizing Human-Object Interactions (HOI) is essential for intelligent systems, underpinning applications in virtual and augmented reality, embodied

SG-Layout: Structured Scene Graph-Guided Layout Generation with LLMs

SafetyDGX agent

arXiv:2608.01106v1 Announce Type: new Abstract: Understanding and generating spatially coherent layouts from natural language remains a fundamental yet challenging task for large language models (LLMs

TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity Understanding

Local AiDGX agent

arXiv:2608.00200v1 Announce Type: cross Abstract: Wearable sensors capture fine-grained motion patterns that support rich behavioral understanding, yet most existing methods reduce these signals to ac

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

SafetyDGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

VC-Tooler: Learning Compositional and Adaptive Visual Tool Use

AgentsDGX agent

arXiv:2608.02217v1 Announce Type: new Abstract: Agentic multimodal reasoning extends passive image understanding by allowing VLMs to actively acquire and refine visual evidence through visual tool int

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills

SafetyDGX agent

arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that w

WorldMirror: Universal 3D World Reconstruction with Any-Prior Prompting

ResearchDGX agent

arXiv:2510.10726v2 Announce Type: replace Abstract: We present WorldMirror, a unified feed-forward model for comprehensive 3D geometric prediction tasks. Unlike existing methods constrained to image-o

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding

Model ReleasesDGX agent

arXiv:2608.00036v1 Announce Type: new Abstract: Real-world document tasks often ask professionals to answer questions from annual reports, regulations, clinical guidelines, and technical manuals that

3 Aug 2026

AgenticRepair: Multi-Faceted Program Context Engineering for Agentic Vulnerability Repair

AgentsDGX agent

arXiv:2607.29422v1 Announce Type: cross Abstract: Automated vulnerability repair aims to reduce the time and effort required to patch security flaws from a vulnerability triage report. Recent agentic

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements

Model ReleasesDGX agent

arXiv:2607.28661v1 Announce Type: new Abstract: Do Large Language Models (LLMs) possess genuine structural reasoning, or merely rely on surface-level pattern matching? The financial domain, demanding

Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review

Model ReleasesDGX agent

arXiv:2607.28631v1 Announce Type: new Abstract: AI Scientist systems capable of autonomous research have the potential to significantly accelerate scientific discovery. However, evaluating and compari

CLIFT: Turning Gemini Robotics On-Device into Humanoid Specialists via Non-Invasive Closed-Loop Iterative Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.29172v1 Announce Type: cross Abstract: While robot foundation models are growing increasingly capable, the strongest models are typically trained on proprietary data and remain closed-sourc

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

Model ReleasesDGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation

ResearchDGX agent

arXiv:2603.23413v2 Announce Type: replace Abstract: Despite remarkable progress in video generation, maintaining long-term scene consistency upon revisiting previously explored areas remains challengi

Inference-time Trajectory Optimization for Structure-Preserving Manga Image Editing

Model ReleasesDGX agent

arXiv:2603.27790v2 Announce Type: replace Abstract: We present a lightweight, training-free trajectory correction method that adapts a pretrained image editing model to each input manga image using on

Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

ResearchDGX agent

arXiv:2607.28684v1 Announce Type: new Abstract: Existing benchmarks for scientific equation discovery are largely composed of well-known equations available in the public domain, making it difficult t

Meshy T2: Fast Native Mesh Generation with Flow Matching

ResearchDGX agent

arXiv:2607.28675v1 Announce Type: cross Abstract: Polygonal meshes are the standard surface representation of modern 3D pipelines, and generating high-quality meshes with artist-style topology is esse

Mirror Learning

SafetyDGX agent

arXiv:2607.28737v1 Announce Type: cross Abstract: We investigate imitation learning through the lens of third-person observation and propose a framework for mirror learning: acquiring actionable polic

MoRoute: Dynamic Routing for In-Context Multimodal Video Generation

Model ReleasesDGX agent

arXiv:2607.29545v1 Announce Type: new Abstract: Multimodal video generation aims to generate and edit videos conditioned on arbitrary combinations of text, images, and videos within a single model, al

RTLCurator: Label-Efficient Data Curation for RTL Generation

SafetyDGX agent

arXiv:2607.29283v1 Announce Type: cross Abstract: Training large language models (LLMs) to write register-transfer level (RTL) requires large corpora of paired specifications and code, and such data i

Running gpt-oss:20b locally and grading it head to head against a frontier model on real tasks. It held up better than I expected

Local AiDGX agent

I serve a free local model on my Mac Mini and route real agent work to it. To check I was not fooling myself, I set up a blind grader that replays frontier tasks locally and scores both. https://previ

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

Model ReleasesDGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens

ResearchDGX agent

arXiv:2607.29363v1 Announce Type: cross Abstract: Balancing sequence length, representational capacity, and long-horizon stability is a central problem in autoregressive (AR) speech and audio generati

TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking

SafetyDGX agent

arXiv:2604.01207v2 Announce Type: replace Abstract: Existing 3D Gaussian Splatting (3DGS) editing methods primarily focus on appearance modification and often struggle to support flexible geometry edi

2 Aug 2026

New research from Google DeepMind. (bookmark it) SkillSmith treats model weights as an additional modality the LLM reads natively. The augme…

ResearchDGX agent

New research from Google DeepMind. (bookmark it) SkillSmith treats model weights as an additional modality the LLM reads natively. The augmented model ingests existing prefix weights alongside rich te

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 month…

ResearchDGX agent

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 months later in December 2024 and has now become completely ubiqu

31 Jul 2026

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans

ResearchDGX agent

arXiv:2607.27634v1 Announce Type: new Abstract: Generating high-quality 360-degree dynamic human assets from text prompts is challenging. Existing methods usually synthesize monocular or multi-view vi

AgenticCANN: Automated Ascend C Operator Generation via Knowledge-Augmented Agentic Evolution

HardwareDGX agent

arXiv:2607.26661v1 Announce Type: new Abstract: Ascend C operator optimization is critical for NPU (Neural Processing Unit) inference performance but requires deep hardware expertise.While large langu

Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments

AgentsDGX agent

arXiv:2607.28591v1 Announce Type: cross Abstract: Scaling coding agents requires a continuing supply of executable data for training, benchmarking, and continuous evaluation. Each task must couple a r

ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science

Model ReleasesDGX agent

arXiv:2607.26155v1 Announce Type: new Abstract: Clinical data-science agents must transform heterogeneous longitudinal records into auditable analyses, yet existing benchmarks largely isolate medical

CRMWeaver: Building Powerful Business Agent via Agentic RL and Shared Memories

AgentsDGX agent

arXiv:2510.25333v2 Announce Type: replace Abstract: Recent years have witnessed the rapid development of LLM-based agents, which shed light on using language agents to solve complex real-world problem

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

Model ReleasesDGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

ICLE++: Modeling Fine-Grained Traits for Holistic Essay Scoring

ResearchDGX agent

arXiv:2607.27671v1 Announce Type: new Abstract: The majority of the recently-developed models for automated essay scoring (AES) are evaluated solely on the ASAP corpus. However, ASAP is not without it

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

SafetyDGX agent

arXiv:2607.28565v1 Announce Type: new Abstract: Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typical

MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation

SafetyDGX agent

arXiv:2607.26698v1 Announce Type: cross Abstract: Cover song generation (CSG) should preserve the melodic and linguistic content of a reference song while recreating the remaining musical components.

Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges

AgentsDGX agent

arXiv:2607.26212v1 Announce Type: cross Abstract: Multi-Agent Debate (MAD) is a promising paradigm for improving the accuracy and robustness of Large Language Model (LLM)-based agentic systems. It ena

Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork

AgentsDGX agent

arXiv:2607.27177v1 Announce Type: new Abstract: Effective collaboration with novel and diverse partners is a crucial skill for autonomous agents. Most current ad-hoc teamwork (AHT) approaches assume t

PrintAnything: Learning an Intermediate Representation for 3D printing G-code Generation

ResearchDGX agent

arXiv:2607.27729v1 Announce Type: new Abstract: Point clouds are one of the most fundamental and widely used 3D representations, serving as the most basic geometric representation of 3D shapes. Nevert

Scaling medical imaging report generation with multimodal reinforcement learning

Model ReleasesDGX agent

arXiv:2601.17151v2 Announce Type: replace-cross Abstract: Frontier models have demonstrated remarkable capabilities in understanding and reasoning with natural-language text, but they still exhibit ma

SciDataSailor: Deep Scientific Data Exploring

AgentsDGX agent

arXiv:2607.28098v1 Announce Type: cross Abstract: Scientific datasets are commonly organized as hierarchical repositories containing heterogeneous and interdependent files, making their inspection, in

SkillSmith: Learning to Compose Parametric Skills and Textual Knowledge

AgentsDGX agent

arXiv:2607.27497v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) regularly feature two key mechanisms to autonomously solve complex problems: synthesizing text-ba

ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

AgentsDGX agent

arXiv:2607.27528v1 Announce Type: cross Abstract: Threat modeling is essential for secure software development, yet manual analysis of cloud-native architectures is slow and demands scarce security ex

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

Model ReleasesDGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

VETO: Towards Protecting Images From Frontier AI Editing

ResearchDGX agent

arXiv:2607.27292v1 Announce Type: new Abstract: The rise of powerful, accessible image-editing models such as FLUX.2 has brought high-fidelity editing within broad reach. Their capabilities now extend

30 Jul 2026

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering

SafetyDGX agent

arXiv:2607.26411v1 Announce Type: new Abstract: Unified multimodal models (UMMs) aim to integrate understanding and generation within a single architecture, yet it remains unclear whether these capabi

HERMES: A Hybrid Ensemble for Head-and-Neck Tumor Segmentation, TN Staging, and Recurrence-Free Survival on PET/CT

ResearchDGX agent

arXiv:2607.26498v1 Announce Type: new Abstract: We present HERMES (Hybrid Ensemble for Radiotherapy-target segmentation, Malignancy staging, and Event-free Survival), a single containerized algorithm

InkShield: Writing Style Protection Against Unauthorized Handwriting Mimicry

ResearchDGX agent

arXiv:2607.26976v1 Announce Type: cross Abstract: Recent handwritten text generators can reproduce a writer's style from publicly available references, posing risks of document forgery and identity mi

TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning

Model ReleasesDGX agent

arXiv:2607.26977v1 Announce Type: new Abstract: Travel planning is a demanding stress test for tool-using LLM agents: a usable itinerary is a single artifact that must be right along many axes at once

29 Jul 2026

Adversarial Deepfake Generation and an Investigation of Purification-Based Adversarial Detection

ResearchDGX agent

arXiv:2607.25842v1 Announce Type: new Abstract: This paper describes the participation of team 'Go To Germany' in the ImageCLEF 2026 Deepfake Detection and Generation Task. For the image generation ta

Cinematic Compositing Using Character-Environment-Harmonized Video Generation Models

ResearchDGX agent

arXiv:2606.20233v2 Announce Type: replace Abstract: Cinematic compositing aims to integrate green-screen characters into novel environments while maintaining physical and photometric realism. Previous

CORF-GS: Real-Time Wireless Radiance Field Reconstruction via Coupled Optical-RF Gaussian Splatting

ResearchDGX agent

arXiv:2607.25569v1 Announce Type: cross Abstract: Recent advances in 3D Gaussian Splatting (3DGS)-based wireless radiance field (WRF) reconstruction provide an efficient solution for wireless channel

← Previous
1…2425262728…48
Next →