AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlog
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,457 results
Model Releases

Verifier-Guided Code Translation via Meta-Step Decoding

DGX agent

arXiv:2605.17626v1 Announce Type: new Abstract: Test-time scaling is an important mechanism for improving large language models, especially on tasks with deterministic verifiers. Code translation is a

model-releasesarxiv-cs-lg
19 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub

Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study

DGX agent

arXiv:2605.17998v1 Announce Type: cross Abstract: As multi-agent systems move from short interactions to tool-using workflows with specialized roles and persistent state, completion becomes a runtime-

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

DGX agent

arXiv:2605.17467v1 Announce Type: new Abstract: Large language model-driven multi-agent systems (LLM-MAS) excel at complex tasks, yet unreliable agents remain a key bottleneck to system-level reliabil

model-releasesarxiv-cs-cl
19 May 2026
Research

VeriHGN: Heterogeneous Graph-Based Congestion Prediction for Chip Layout Verification

DGX agent

arXiv:2603.11075v2 Announce Type: replace-cross Abstract: As Very Large Scale Integration (VLSI) designs continue to scale in size and complexity, layout verification has become a central challenge in

researcharxiv-cs-ai
19 May 2026
Model Releases

VGGT-CD: Training-Free Robust Registration for 3D Change Detection

DGX agent

arXiv:2605.16859v1 Announce Type: cross Abstract: 3D change detection from multi-view images is essential for urban monitoring, disaster assessment, and autonomous driving. However, existing methods p

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction

DGX agent

arXiv:2605.16911v1 Announce Type: new Abstract: 3D semantic occupancy prediction requires accurate 2D-to-3D feature lifting, yet current methods restrict camera geometry to initial projections. Subseq

model-releasesarxiv-cs-cv
19 May 2026
Safety

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

DGX agent

arXiv:2605.16420v1 Announce Type: new Abstract: This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structure

safetyarxiv-cs-cv
19 May 2026
Model Releases

[video] why we need a new continuity layer for long-running agents (claude did this video! all except the voice which was @elevenlabs)

DGX agent

This video discusses the architectural need for a continuity layer in long-running AI agents, explaining how agents require persistent memory and state management mechanisms to maintain coherence acro

model-releasesyohei-nakajima--x
19 May 2026
Research

VideoNeuMat: Neural Material Extraction from Generative Video Models

DGX agent

arXiv:2602.07272v2 Announce Type: replace Abstract: Creating photorealistic materials for 3D rendering requires exceptional artistic skill. Generative models for materials could help, but are currentl

researcharxiv-cs-cv
19 May 2026
Research

Vidya: An AI-Driven Modular Pipeline for Archival Automation and Semantic Metadata Enrichment

DGX agent

arXiv:2605.16338v1 Announce Type: cross Abstract: The large-scale digitization of historical archives has created a paradox: 'dark data'-digital objects lacking metadata for retrieval. Manual archival

researcharxiv-cs-cl
19 May 2026
Safety

View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification

DGX agent

arXiv:2605.18192v1 Announce Type: new Abstract: Aerial-Ground Person Re-Identification (AGPReID) remains highly challenging due to drastic viewpoint variations between drones and fixed cameras. Existi

safetyarxiv-cs-cv
19 May 2026
Research

Virtual Nodes Guided Dynamic Graph Neural Network for Brain Tumor Segmentation with Missing Modalities

DGX agent

arXiv:2605.16880v1 Announce Type: new Abstract: Multimodal magnetic resonance imaging (MRI) is crucial for brain tumor segmentation, with many methods leveraging its four key modalities to capture com

researcharxiv-cs-ai
19 May 2026
Research

Virtues of Ordered Chaos: Planning with Topple Actions in Tabletop Stack Rearrangement

DGX agent

arXiv:2605.17815v1 Announce Type: cross Abstract: Efficient object manipulation strategies have significant impact in automation applications. In this work, the stack rearrangement in tabletop setting

researcharxiv-cs-ai
19 May 2026
Applications

VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation

DGX agent

arXiv:2605.18547v1 Announce Type: new Abstract: Emotion Recognition in Conversation (ERC) is essential for effective human-machine interaction, aiming to identify speakers' emotional states in multi-t

applicationsarxiv-cs-ai
19 May 2026
Research

Vision Foundation Models as Generalist Tokenizers for Image Generation

DGX agent

arXiv:2605.18390v1 Announce Type: new Abstract: In this work, we explore the largely unexplored direction of building a generalist image tokenizer directly on top of a frozen vision foundation model (

researcharxiv-cs-cv
19 May 2026
Model Releases

Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models

DGX agent

arXiv:2605.18160v1 Announce Type: cross Abstract: In recent years, multimodal large language models (MLLMs) have achieved remarkable progress, primarily attributed to effective paradigms for integrati

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

DGX agent

arXiv:2605.18740v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) still struggle with fine-grained visual understanding, where answers often depend on small but decisive evide

local-aiarxiv-cs-ai
19 May 2026
Safety

Vision Transformer-Conditioned UNet for Domain-Adaptive Semantic Segmentation

DGX agent

arXiv:2605.16393v1 Announce Type: cross Abstract: Semantic segmentation is essential for analysing anatomical features in biomedical research, yet a performance gap remains for Vision Transformers (Vi

safetyarxiv-cs-ai
19 May 2026
Model Releases

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

DGX agent

arXiv:2602.04802v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have achieved impressive performance in cross-modal understanding across textual and visual inputs, yet existing bench

model-releasesarxiv-cs-cv
19 May 2026
Research

VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers

DGX agent

arXiv:2605.17312v1 Announce Type: new Abstract: Video style transfer aims to render videos in a target artistic style while preserving content, structure, and motion. While image stylization has advan

researcharxiv-cs-cv
19 May 2026
Local Ai

VISTA: Variance-Gated Inter-Sequence Test-Time Adaptation for Multi-Sequence MRI Segmentation

DGX agent

arXiv:2605.17433v1 Announce Type: new Abstract: Deploying multi-sequence magnetic resonance imaging (MRI) segmentation models to new clinical environments is challenging due to variations in scanners

local-aiarxiv-cs-cv
19 May 2026
Model Releases

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

DGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

model-releasesarxiv-cs-ai
19 May 2026
Safety

Visual Sculpting: Visually-Aligned Planning Representations for Long-Horizon Robot Clay Sculpting

DGX agent

arXiv:2605.17556v1 Announce Type: cross Abstract: Clay sculpting is a nuanced, artistic task involving dexterous manipulation with long-horizon planning to achieve high-level goals. As a robotics prob

safetyarxiv-cs-ai
19 May 2026
Research

Visual Search Patterns in 3D Pancreatic Imaging: An Eye Tracking Study

DGX agent

arXiv:2605.16408v1 Announce Type: new Abstract: Eye tracking has emerged as a powerful tool for examining visual perception and search strategies in various domains, including medicine. While it is re

researcharxiv-cs-cv
19 May 2026
Research

Visual Timelines of Police Encounters in Body-Worn Camera Footage: Operational Context and Activity Cataloging for Training and Analysis in OpenBWC

DGX agent

arXiv:2605.17095v1 Announce Type: cross Abstract: Law enforcement agencies are accumulating vast amounts of body-worn camera (BWC) footage. However, this remains operationally opaque. That is, analyst

researcharxiv-cs-ai
19 May 2026
Model Releases

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

DGX agent

arXiv:2605.18172v1 Announce Type: new Abstract: Leveraging the universal representations of pre-trained LLMs and MLLMs offers a promising path toward brain foundation models. However, visually-evoked

model-releasesarxiv-cs-ai
19 May 2026
Safety

VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events

DGX agent

arXiv:2603.18178v2 Announce Type: replace-cross Abstract: The rapid growth of ego-centric dashcam footage presents a major challenge for detecting safety-critical events such as collisions and near-co

safetyarxiv-cs-ai
19 May 2026
Research

Voice ''Cloning'' is Style Transfer

DGX agent

arXiv:2605.16578v1 Announce Type: cross Abstract: Artificially generated speech is increasingly embedded in everyday life. Voice cloning in particular enables applications where identity preservation

researcharxiv-cs-ai
19 May 2026
Tools

Voice 'cloning' is style transfer. Across three widely used systems — ElevenLabs V3, Coqui-XTTS, Chatterbox — clones don't just copy speaker…

DGX agent

Voice 'cloning' is style transfer. Across three widely used systems — ElevenLabs V3, Coqui-XTTS, Chatterbox — clones don't just copy speakers, they reshape them to be warmer, more authoritative, more

toolstogether-ai--x
19 May 2026
Safety

Voices in the Loop: Mapping Participatory AI

DGX agent

arXiv:2605.16827v1 Announce Type: new Abstract: Participatory approaches to artificial intelligence are increasingly documented across public, civic, and humanitarian settings, but evidence about how

safetyarxiv-cs-ai
19 May 2026
Agents

Voker raises $2.2M to help teams understand how AI agents perform in the wild

DGX agent

Voker, an agent analytics platform for artificial intelligence product teams, today announced it has raised 2.2 million in pre-seed funding from Y Combinator and FundersClub. As more companies push AI

agentssiliconangle
19 May 2026
Safety

VolTA-3D: Self-Supervised Learning for Brain MRI using 3D Volumetric Token Alignment

DGX agent

arXiv:2605.16775v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has advanced medical image analysis be enabling learning form large unlabelled data. However, in brain magnetic resonan

safetyarxiv-cs-ai
19 May 2026
Tools

volunteer here https://ai.engineer/cfp !

DGX agent

This post links to a Call for Proposals (CFP) for volunteering opportunities at AI.Engineer, likely inviting community members to submit talk proposals, workshop ideas, or volunteer roles for an AI en

toolsswyx--x
19 May 2026
Research

VoxScene: Anchor-Conditioned Voxel Diffusion for Indoor Scene Arrangement

DGX agent

arXiv:2605.17102v1 Announce Type: cross Abstract: We present VoxScene, a novel anchor-conditioned voxel diffusion framework tailored for 3D scene synthesis. Current data-driven layout generation techn

researcharxiv-cs-cv
19 May 2026
Research

VoxShield: Protecting 3D Medical Datasets from Unauthorized Training via Frequency-Aware Inter-Slice Disruption

DGX agent

arXiv:2605.17345v1 Announce Type: new Abstract: The release of public 3D medical image segmentation (MIS) datasets accelerates clinical research but simultaneously heightens risks of unauthorized AI m

researcharxiv-cs-cv
19 May 2026
Hardware

Vultr Announces Milan, Italy, as 33rd Cloud Data Center Region

DGX agent

Vultr expanded its global cloud infrastructure by opening a new data center region in Milan, Italy, marking the company's 33rd cloud data center location worldwide. This expansion provides European cu

hardwarevultr
19 May 2026
Applications

VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos

DGX agent

arXiv:2605.17584v1 Announce Type: new Abstract: Unsupervised pixel-level video understanding remains challenging in real-world scenarios, where motion blur, occlusion, and fast object dynamics often c

applicationsarxiv-cs-cv
19 May 2026
Research

WASIL: In-the-Wild Arabic Spoken Interactions with LLMs

DGX agent

arXiv:2605.16364v1 Announce Type: cross Abstract: Large Language Models (LLMs) voice assistants are commonly built as cascaded Automatic Speech recognition (ASR) to LLM systems, where recognition erro

researcharxiv-cs-ai
19 May 2026
Research

Wasserstein bounds for denoising diffusion probabilistic models via the Follmer process

DGX agent

arXiv:2605.18069v1 Announce Type: cross Abstract: This paper studies sampling error bounds for denoising diffusion probabilistic models (DDPMs) in the 2-Wasserstein distance. Our contributions are thr

researcharxiv-cs-lg
19 May 2026
Model Releases

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

DGX agent

arXiv:2605.18313v1 Announce Type: cross Abstract: Small vision-language models (2-8B) are well-suited for clin- ical deployment due to privacy constraints, limited connectivity, and low-latency requir

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning

DGX agent

arXiv:2601.06943v2 Announce Type: replace-cross Abstract: In real-world video question answering scenarios, videos often provide only localized visual cues, while verifiable answers are distributed ac

model-releasesarxiv-cs-ai
19 May 2026
Research

Watermarks Attack Watermarks: Re-Watermarking as a Generic Removal Strategy

DGX agent

arXiv:2605.16796v1 Announce Type: cross Abstract: Watermarking combines an imperceptible change to an input image that will trigger a detector, to assert provenance and protect intellectual property.

researcharxiv-cs-cv
19 May 2026
Research

Wavelet Flow Matching for Multi-Scale Physics Emulation

DGX agent

arXiv:2605.16573v1 Announce Type: cross Abstract: Accurate emulation of multi-scale physical systems governed by PDEs demands models that remain stable over long autoregressive rollouts while preservi

researcharxiv-cs-ai
19 May 2026
Model Releases

WavFlow: Audio Generation in Waveform Space

DGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

model-releasesarxiv-cs-cv
19 May 2026
Research

We are aware of a @Railway outage impacting Nous Portal users and have contacted their team for more information on service restoration ETA.

DGX agent

Nous Research reported an outage affecting Railway that impacted access to Nous Portal, and the company stated they had contacted Railway's team to obtain information about when service would be resto

researchnous-research--x
19 May 2026
Hardware

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human geno…

DGX agent

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human genome on a single GPU in <2 days. Here are the tricks we used:

hardwareclem-delangue--x
19 May 2026
Industry

We just added Grok's new imagine model in Paper so you can explore images even faster. Here's what we found: - Super fast generations for th…

DGX agent

We just added Grok's new imagine model in Paper so you can explore images even faster. Here's what we found: - Super fast generations for the quality - Saves details when editing - Perfect for 30 rapi

industryelon-musk--x
19 May 2026
Model Releases

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong

DGX agent

arXiv:2509.22510v3 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) is the ability to satisfy desired objectives during generation, which is critical for trustworthy deployme

model-releasesarxiv-cs-cl
19 May 2026
← Previous
1…11671168116911701171…1843
Next →