AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “synthesis”

GridTimelineEvolution
2,844 results
Tutorials

Style-Aware Gloss Control for Generative Non-Photorealistic Rendering

DGX agent

arXiv:2602.16611v3 Announce Type: replace-cross Abstract: Humans can infer material characteristics of objects from their visual appearance, and this ability extends to artistic depictions, where simi

tutorialsarxiv-cs-cv
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Towards a new paradigm of scientific discovery with socialized artificial intelligence

DGX agent

arXiv:2608.02775v1 Announce Type: new Abstract: Scientific discovery has advanced through successive transformations in the organization of knowledge. Observation and experimentation established the e

researcharxiv-cs-ai
5 Aug 2026
Model Releases

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

DGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

DeepVoyager-VL: Incentivizing Vision-in-the-Loop Search for Long-Horizon Multimodal Agents

DGX agent

arXiv:2608.01827v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced visual understanding and reasoning, yet their static parametric knowledge limits their ability to

agentsarxiv-cs-cv
4 Aug 2026
Agents

DerainSplat: Feed-Forward Clean 3D Gaussian Splatting from Sparse Rainy Views

DGX agent

arXiv:2608.02191v1 Announce Type: new Abstract: Although image deraining has advanced substantially, existing methods mainly focus on 2D image restoration. As spatial intelligence applications such as

agentsarxiv-cs-cv
4 Aug 2026
Agents

DiffPhysCam: Differentiable Physics-Based Camera Simulation for Inverse Rendering and Embodied AI

DGX agent

arXiv:2508.08831v2 Announce Type: replace-cross Abstract: Generating synthetic images that closely mimic those from real cameras is instrumental in training visual models and enabling end-to-end visuo

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

DiffusionGemma Technical Report

DGX agent

arXiv:2608.00146v1 Announce Type: new Abstract: We introduce DiffusionGemma, an experimental open-weight language model that uses discrete diffusion to generate text at exceptionally high speed. Rathe

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study

DGX agent

arXiv:2608.02235v1 Announce Type: new Abstract: Recent advances in neural text-to-speech (TTS) systems have substantially improved speech naturalness and intelligibility across many languages. However

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Enhancing Large Language Model Reasoning with Reward Models: An Analytical Survey

DGX agent

arXiv:2510.01925v3 Announce Type: replace Abstract: Reward models (RMs) play a critical role in enhancing the reasoning performance of LLMs. For example, they can provide training signals to finetune

researcharxiv-cs-cl
4 Aug 2026
Safety

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning

DGX agent

arXiv:2608.00613v1 Announce Type: new Abstract: Physical-world interaction is inherently dynamic, as environments can evolve during execution, requiring agents to adapt their plans under non-stationar

safetyarxiv-cs-ro
4 Aug 2026
Safety

Human-LLM Alignment in Language Attitudes Toward Non-Native Japanese

DGX agent

arXiv:2608.01629v1 Announce Type: new Abstract: Large language models (LLMs) increasingly evaluate human writing in high-stakes domains such as hiring and academic assessment, putting non-native speak

safetyarxiv-cs-cl
4 Aug 2026
Research

InstancePin: Instance-Addressable Layout-to-Image Diffusion via Coordinate Pinning

DGX agent

arXiv:2608.00588v1 Announce Type: new Abstract: Layout-to-image diffusion models have achieved impressive semantic controllability by conditioning generation on category-level segmentation maps. Howev

researcharxiv-cs-cv
4 Aug 2026
Research

Learning to Tessellate: Point Cloud Generation via Recursive Spectral Partitioning

DGX agent

arXiv:2608.02432v1 Announce Type: new Abstract: Autoregressive models have emerged as an effective paradigm for point cloud generation. However, most existing approaches rely on heuristic tokenization

researcharxiv-cs-cv
4 Aug 2026
Model Releases

Lethe: How Hard Is It to Forget? A Benchmark for Federated Unlearning in Medical Imaging

DGX agent

arXiv:2608.01094v1 Announce Type: new Abstract: Federated learning enables medical-imaging models to be trained across hospitals, and privacy law, most explicitly the GDPR ``right to be forgotten'', t

model-releasesarxiv-cs-cv
4 Aug 2026
Agents

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

DGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

DGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

Optimizing Minimax Regret in Uncertain MDPs with Small Sets of Policies

DGX agent

arXiv:2608.02509v1 Announce Type: cross Abstract: Sequential decision-making in real-world applications often involves uncertainty about the environment's model. Uncertain Markov decision processes (U

safetyarxiv-cs-lg
4 Aug 2026
Safety

Practical Noise Modeling for SPAD Intensity Imaging

DGX agent

arXiv:2608.00489v1 Announce Type: new Abstract: Single-photon avalanche diode (SPAD) cameras are promising for low-light and high-dynamic-range intensity imaging, but their practical use is limited by

safetyarxiv-cs-cv
4 Aug 2026
Research

Prompt-Driven Simulation with Feature Perturbation for Cross-Domain Few-Shot Object Detection

DGX agent

arXiv:2608.01348v1 Announce Type: new Abstract: Data augmentation, which simulates diverse visual variations to expand the source distribution and induce synthetic domain shifts, is a simple yet effec

researcharxiv-cs-cv
4 Aug 2026
Model Releases

QuerySplat: Decoupling Geometry and Appearance Representations in 3DGS Prediction

DGX agent

arXiv:2608.01186v1 Announce Type: new Abstract: While feed-forward 3D Gaussian Splatting (3DGS) enables efficient 3D reconstruction, achieving high-fidelity rendering remains challenging. Existing pix

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Recursive Gaussian Processes and the Bayesian Brain

DGX agent

arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobio

researcharxiv-cs-lg
4 Aug 2026
Applications

RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals

DGX agent

arXiv:2608.00289v1 Announce Type: cross Abstract: Recognizing Human-Object Interactions (HOI) is essential for intelligent systems, underpinning applications in virtual and augmented reality, embodied

applicationsarxiv-cs-ro
4 Aug 2026
Safety

SG-Layout: Structured Scene Graph-Guided Layout Generation with LLMs

DGX agent

arXiv:2608.01106v1 Announce Type: new Abstract: Understanding and generating spatially coherent layouts from natural language remains a fundamental yet challenging task for large language models (LLMs

safetyarxiv-cs-cv
4 Aug 2026
Local Ai

TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity Understanding

DGX agent

arXiv:2608.00200v1 Announce Type: cross Abstract: Wearable sensors capture fine-grained motion patterns that support rich behavioral understanding, yet most existing methods reduce these signals to ac

local-aiarxiv-cs-cl
4 Aug 2026
Safety

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

DGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

safetyarxiv-cs-lg
4 Aug 2026
Agents

VC-Tooler: Learning Compositional and Adaptive Visual Tool Use

DGX agent

arXiv:2608.02217v1 Announce Type: new Abstract: Agentic multimodal reasoning extends passive image understanding by allowing VLMs to actively acquire and refine visual evidence through visual tool int

agentsarxiv-cs-cv
4 Aug 2026
Safety

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills

DGX agent

arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that w

safetyarxiv-cs-ro
4 Aug 2026
Research

WorldMirror: Universal 3D World Reconstruction with Any-Prior Prompting

DGX agent

arXiv:2510.10726v2 Announce Type: replace Abstract: We present WorldMirror, a unified feed-forward model for comprehensive 3D geometric prediction tasks. Unlike existing methods constrained to image-o

researcharxiv-cs-cv
4 Aug 2026
Model Releases

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding

DGX agent

arXiv:2608.00036v1 Announce Type: new Abstract: Real-world document tasks often ask professionals to answer questions from annual reports, regulations, clinical guidelines, and technical manuals that

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

AgenticRepair: Multi-Faceted Program Context Engineering for Agentic Vulnerability Repair

DGX agent

arXiv:2607.29422v1 Announce Type: cross Abstract: Automated vulnerability repair aims to reduce the time and effort required to patch security flaws from a vulnerability triage report. Recent agentic

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements

DGX agent

arXiv:2607.28661v1 Announce Type: new Abstract: Do Large Language Models (LLMs) possess genuine structural reasoning, or merely rely on surface-level pattern matching? The financial domain, demanding

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review

DGX agent

arXiv:2607.28631v1 Announce Type: new Abstract: AI Scientist systems capable of autonomous research have the potential to significantly accelerate scientific discovery. However, evaluating and compari

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

CLIFT: Turning Gemini Robotics On-Device into Humanoid Specialists via Non-Invasive Closed-Loop Iterative Fine-Tuning

DGX agent

arXiv:2607.29172v1 Announce Type: cross Abstract: While robot foundation models are growing increasingly capable, the strongest models are typically trained on proprietary data and remain closed-sourc

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

DGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

model-releasesarxiv-cs-cl
3 Aug 2026
Research

I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation

DGX agent

arXiv:2603.23413v2 Announce Type: replace Abstract: Despite remarkable progress in video generation, maintaining long-term scene consistency upon revisiting previously explored areas remains challengi

researcharxiv-cs-cv
3 Aug 2026
Model Releases

Inference-time Trajectory Optimization for Structure-Preserving Manga Image Editing

DGX agent

arXiv:2603.27790v2 Announce Type: replace Abstract: We present a lightweight, training-free trajectory correction method that adapts a pretrained image editing model to each input manga image using on

model-releasesarxiv-cs-cv
3 Aug 2026
Research

Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

DGX agent

arXiv:2607.28684v1 Announce Type: new Abstract: Existing benchmarks for scientific equation discovery are largely composed of well-known equations available in the public domain, making it difficult t

researcharxiv-cs-ai
3 Aug 2026
Research

Meshy T2: Fast Native Mesh Generation with Flow Matching

DGX agent

arXiv:2607.28675v1 Announce Type: cross Abstract: Polygonal meshes are the standard surface representation of modern 3D pipelines, and generating high-quality meshes with artist-style topology is esse

researcharxiv-cs-cv
3 Aug 2026
Safety

Mirror Learning

DGX agent

arXiv:2607.28737v1 Announce Type: cross Abstract: We investigate imitation learning through the lens of third-person observation and propose a framework for mirror learning: acquiring actionable polic

safetyarxiv-cs-cv
3 Aug 2026
Model Releases

MoRoute: Dynamic Routing for In-Context Multimodal Video Generation

DGX agent

arXiv:2607.29545v1 Announce Type: new Abstract: Multimodal video generation aims to generate and edit videos conditioned on arbitrary combinations of text, images, and videos within a single model, al

model-releasesarxiv-cs-cv
3 Aug 2026
Safety

RTLCurator: Label-Efficient Data Curation for RTL Generation

DGX agent

arXiv:2607.29283v1 Announce Type: cross Abstract: Training large language models (LLMs) to write register-transfer level (RTL) requires large corpora of paired specifications and code, and such data i

safetyarxiv-cs-lg
3 Aug 2026
Local Ai

Running gpt-oss:20b locally and grading it head to head against a frontier model on real tasks. It held up better than I expected

DGX agent

I serve a free local model on my Mac Mini and route real agent work to it. To check I was not fooling myself, I set up a blind grader that replays frontier tasks locally and scores both. https://previ

local-air-ollama
3 Aug 2026
Model Releases

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

DGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

model-releasesarxiv-cs-lg
3 Aug 2026
Research

Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens

DGX agent

arXiv:2607.29363v1 Announce Type: cross Abstract: Balancing sequence length, representational capacity, and long-horizon stability is a central problem in autoregressive (AR) speech and audio generati

researcharxiv-cs-ai
3 Aug 2026
Safety

TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking

DGX agent

arXiv:2604.01207v2 Announce Type: replace Abstract: Existing 3D Gaussian Splatting (3DGS) editing methods primarily focus on appearance modification and often struggle to support flexible geometry edi

safetyarxiv-cs-cv
3 Aug 2026
Research

New research from Google DeepMind. (bookmark it) SkillSmith treats model weights as an additional modality the LLM reads natively. The augme…

DGX agent

New research from Google DeepMind. (bookmark it) SkillSmith treats model weights as an additional modality the LLM reads natively. The augmented model ingests existing prefix weights alongside rich te

researchdair-ai--x
2 Aug 2026
Research

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 month…

DGX agent

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 months later in December 2024 and has now become completely ubiqu

researchfrancois-chollet--x
2 Aug 2026
Research

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans

DGX agent

arXiv:2607.27634v1 Announce Type: new Abstract: Generating high-quality 360-degree dynamic human assets from text prompts is challenging. Existing methods usually synthesize monocular or multi-view vi

researcharxiv-cs-cv
31 Jul 2026
← Previous
1…3031323334…60
Next →