AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,356 results
30 Jun 2026

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

Model ReleasesDGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

TUGS: Physics-based Compact Representation of Underwater Scenes by Tensorized Gaussian

ApplicationsDGX agent

arXiv:2505.08811v3 Announce Type: replace Abstract: Underwater 3D scene reconstruction is crucial for multimedia applications in adverse environments, such as underwater robotic perception and navigat

Tumor-aware augmentation with task-guided attention analysis improves rectal cancer segmentation from magnetic resonance images

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2605.05522v2 Announce Type: replace-cross Abstract: Although self-supervised pretraining is expected to learn broadly transferable representations, its effectiveness across imaging modalities su

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

ResearchDGX agent

arXiv:2606.28548v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have become a useful tool for extracting interpretable features in language models. However, standard SAE architectures opera

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for t…

Model ReleasesDGX agent

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for the naturalness/latency Building voice agents can come with t

Two kinds of robustness are not the same: disentangling fault tolerance and low-SNR robustness in multi-domain event detection on real data

Model ReleasesDGX agent

arXiv:2606.29339v1 Announce Type: cross Abstract: Reliable event detection underpins induced-seismicity monitoring for Carbon dioxide Capture and Storage (CCS) and geothermal operations, distributed a

Two-Stage Prompt Optimization for Few-Shot Relation Extraction: From Reasoning-Guided Search to Gradient-Guided Refinement

ResearchDGX agent

arXiv:2606.29639v1 Announce Type: cross Abstract: Automatic prompt optimization is still underexplored for episodic few-shot relation extraction with smaller language models. We propose a two-stage fr

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models

ApplicationsDGX agent

arXiv:2602.22960v2 Announce Type: replace Abstract: World models based on video generation demonstrate remarkable potential for simulating interactive environments yet suffer from persistent difficult

UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation

Local AiDGX agent

arXiv:2606.29502v1 Announce Type: new Abstract: Skill memories can improve agentic reinforcement learning by reusing past experience as textual guidance, but retrieved skills are not oracular: they ma

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention

Local AiDGX agent

arXiv:2510.16325v3 Announce Type: replace Abstract: Ultra-high-resolution text-to-image generation is increasingly vital for applications requiring fine-grained textures and global structural fidelity

Uncertainty-Aware Generation and Decision-Making Under Ambiguity

ApplicationsDGX agent

arXiv:2606.30578v1 Announce Type: new Abstract: With rapidly improving capabilities, Large Language Models (LLMs) are increasingly used in many complex real-world tasks. Beyond requiring in-depth know

Uncertainty Estimation in Pathology Foundation Models via Deep Mutual Learning

ResearchDGX agent

arXiv:2606.30020v1 Announce Type: new Abstract: Pathology foundation models (PFMs) offer generalizable representations for whole-slide image (WSI) analysis, yet their clinical adoption remains limited

Uncovering Salience-Driven Dynamics in Consumer Confidence with Generative Social Simulation

SafetyDGX agent

arXiv:2606.30395v1 Announce Type: cross Abstract: Consumer confidence is typically modeled as a persistent macroeconomic index, yet its movements arise from households that interpret economic informat

Understanding Evaluation Illusion in Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.29228v1 Announce Type: new Abstract: Despite the capability of parallel decoding, diffusion large language models (dLLMs) require many denoising steps to maintain generation quality, motiva

Understanding LLM Intervention Explanations in Multi-Party Human-Robot Interaction

ResearchDGX agent

arXiv:2606.29460v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly embedded in social robots to support natural group interactions, yet their role in complex multi-party set

UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image

AgentsDGX agent

arXiv:2606.30608v1 Announce Type: new Abstract: Articulated 3D objects are essential for interactive environments in embodied AI, robotics, and virtual reality, but reconstructing their structure and

UniCA: Bi-directional Cross-Attention with Positive Similarity Loss for Robust Multi-Modal Retrieval

Model ReleasesDGX agent

arXiv:2606.28350v1 Announce Type: cross Abstract: Multi-modal retrieval has become increasingly critical for handling the growing volume of integrated visual-textual data in real-world applications, b

Unified Complex-valued Neural Network: A Magnitude-Phase Computational Model for Event-Driven Neuromorphic Learning

Local AiDGX agent

arXiv:2606.29099v1 Announce Type: cross Abstract: Artificial neural networks (ANN) provide accurate continuous-valued representation, whereas spiking neural networks (SNN) offer event-driven temporal

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

Model ReleasesDGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

SafetyDGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation

SafetyDGX agent

arXiv:2603.22282v2 Announce Type: replace-cross Abstract: We present UniMotion, to our knowledge the first unified framework for simultaneous understanding and generation of human motion, natural lang

UniPR-3D: Towards Universal Visual Place Recognition with Visual Geometry Grounded Transformer

ResearchDGX agent

arXiv:2512.21078v3 Announce Type: replace Abstract: Visual Place Recognition (VPR) has been traditionally formulated as a single-image retrieval task. Using multiple views offers clear advantages, yet

UniTriSplat: A Unified 3D Gaussian Splatting Framework with Uniform Spherical Rasterization for Universal Cameras

ResearchDGX agent

arXiv:2606.29794v1 Announce Type: new Abstract: Existing 3D Gaussian Splatting (3DGS) frameworks rely on camera-specific rasterization, suffering from inconsistent solid-angle sampling and degraded pe

UniVAD v2: Unified Visual Anomaly Detection via Support-Conditioned Boundary Construction

ResearchDGX agent

arXiv:2606.29714v1 Announce Type: new Abstract: Unified visual anomaly detection seeks to train a single detector that can be deployed across categories, domains, and application scenarios. In the few

Universality of empirical risk minimization

ResearchDGX agent

arXiv:2202.08832v3 Announce Type: replace-cross Abstract: We study a general class of optimization problems with decision variable oldsymbol{Theta} in R^{p imes k} and cost function which is the sum o

Unlocking the Visual Record of Materials Science: A Large-Scale Multimodal Dataset from Scientific Literature

Model ReleasesDGX agent

arXiv:2606.29667v1 Announce Type: cross Abstract: The materials science literature encodes decades of experimental knowledge in figures, yet this visual record remains locked away and inaccessible to

Unveiling Novelty Evolution in the field of Library and Information Science in China

ResearchDGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

UrbanCDNet: Appearance-Robust and Boundary-Aware Bitemporal Change Detection for Korean Urban Building Monitoring

Model ReleasesDGX agent

arXiv:2606.29781v1 Announce Type: new Abstract: Urban building change detection from bi-temporal aerial imagery is important for redevelopment monitoring, infrastructure management, and unauthorized-c

Using Large Language Models as Low-Cost Statistical Estimators for Human-Response Data

SafetyDGX agent

arXiv:2606.30372v1 Announce Type: new Abstract: Quantitative research across the social and behavioral sciences depends on human subject experiments that are expensive, slow, and subject to sampling b

Value-Action Alignment in Large Language Models under Privacy-Prosocial Conflict

SafetyDGX agent

arXiv:2601.03546v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate decision-making tasks involving personal data sharing, where privacy concerns a

Variance Reduction for Stochastic Gradient Generalized Non-reversible Langevin Monte Carlo Algorithms

ResearchDGX agent

arXiv:2606.28808v1 Announce Type: cross Abstract: We study the leading-order fluctuation of stochastic gradient Euler-Maruyama estimators for generalized non-reversible Langevin dynamics. Under struct

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

Model ReleasesDGX agent

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

VCS-SLAM: Geometry-Validated Semantic Evidence Fusion for 3D Gaussian SLAM

ApplicationsDGX agent

arXiv:2606.29494v1 Announce Type: new Abstract: Visual SLAM performance often deteriorates in complex real-world applications. Semantic 3D Gaussian SLAM commonly fuses 2D semantic priors into a persis

Vercel Agent has updated pricing

AgentsDGX agent

Vercel Agent pricing has been updated, likely reflecting changes to the cost structure for using Vercel's AI agent capabilities. The specific details of the new pricing tiers, feature inclusions, and

Vercel and Shopify are rebuilding Hydrogen

ToolsDGX agent

Vercel and Shopify collaborated to rebuild Hydrogen, Shopify's React-based framework for building custom storefronts. The rebuild likely focused on improving developer experience, performance, and int

Vercel Private Blob is now generally available

ToolsDGX agent

Vercel Private Blob, a storage solution for secure file management, has reached general availability status. This feature enables developers to store and manage private files within Vercel deployments

Vercel Sandbox now support Custom Images

ToolsDGX agent

Vercel Sandbox has expanded its capabilities to support custom images, allowing developers to use personalized or pre-configured container images for their sandbox environments. This feature enables g

Vercel Services gives your full stack app atomic deployment and rollback, a single preview URL, and a private network between services. One …

ToolsDGX agent

Vercel Services gives your full stack app atomic deployment and rollback, a single preview URL, and a private network between services. One Vercel project for your entire application ↓ https://vercel.

Vercel Services: Run full stack on Vercel

ToolsDGX agent

Vercel Services enables developers to run complete full-stack applications on the Vercel platform, extending beyond static site hosting to include backend capabilities. This feature allows seamless in

very successful first poster sessions day at AIEWF, thanks to @heathercmiller and @ACM_President for the support!! tomorrow: Poaster session…

ToolsDGX agent

very successful first poster sessions day at AIEWF, thanks to @heathercmiller and @ACM_President for the support!! tomorrow: Poaster sessions! submit your hottest tweets to @vibhuuuus for printing! sp

VIB-AVSR: Variational Information Bottleneck for Noise-Robust LLM-Based Audio-Visual Speech Recognition

ResearchDGX agent

arXiv:2606.29632v1 Announce Type: cross Abstract: Audio-Visual Speech Recognition takes two input modalities, acoustic and visual streams, where visual information from lip movements aids recognition

VibES: Induced Vibration for Persistent Event-Based Sensing

ApplicationsDGX agent

arXiv:2508.19094v3 Announce Type: replace Abstract: Event cameras are a bio-inspired class of sensors that asynchronously measure per-pixel intensity changes. Under fixed illumination conditions in st

ViewSplat: View-Adaptive 3D Gaussian Splatting for Feed-Forward Synthesis

ResearchDGX agent

arXiv:2603.25265v2 Announce Type: replace Abstract: We present ViewSplat, a view-adaptive 3D Gaussian splatting network for novel view synthesis from unposed images. While recent feed-forward 3D Gauss

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

Model ReleasesDGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

Vimeo owner Bending Spoons raised 1.68B by selling 58M shares at 29 each, valuing it at ~$18.4B, in one of the largest US IPOs by a European company in 2026 (Subrat Patnaik/Bloomberg)

IndustryDGX agent

Subrat Patnaik / Bloomberg: Vimeo owner Bending Spoons raised 1.68B by selling 58M shares at 29 each, valuing it at ~18.4B, in one of the largest US IPOs by a European company in 2026 — Bending Spoons

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

Virtual Ring Try-On

ResearchDGX agent

arXiv:2606.28792v1 Announce Type: new Abstract: This paper presents an innovative approach that enables the users to capture their hand and try the jewel ring on their hand. The user captures the imag

Visa, Mastercard, Stripe, BlackRock, Coinbase, and 140+ companies join Open Standard to launch Open USD, a stablecoin that shares earnings from its reserves (Kyle Baird/The Block)

IndustryDGX agent

Kyle Baird / The Block: Visa, Mastercard, Stripe, BlackRock, Coinbase, and 140+ companies join Open Standard to launch Open USD, a stablecoin that shares earnings from its reserves — Quick Take — Paym

Visa, Stripe and 140 others back new Open USD stablecoin to challenge Tether

IndustryDGX agent

A consortium of more than 140 financial, payments and technology companies, including Visa Inc., Stripe Inc. and BlackRock Inc., is backing a new stablecoin called Open USD, taking direct aim at the m

Vision-driven Preference Synthesis for Mitigating Hallucinations in VLMs

SafetyDGX agent

arXiv:2606.28401v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong performance in visual understanding, yet they still suffer from hallucinations, generating content that

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

SafetyDGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

Vision-Language Models for Deployable Social Robot Navigation: Bridging Semantic Reasoning and Low-Level Control

SafetyDGX agent

arXiv:2606.28760v1 Announce Type: new Abstract: Social robot navigation (SRN) requires more than geometric path planning; it demands understanding human intentions, social norms, and contextual cues t

VisReflect: Latent Visual Reflection for Fine-Grained Perception in Long Visual Context

Local AiDGX agent

arXiv:2606.30288v1 Announce Type: new Abstract: Large Vision Language Models (LVLMs) have achieved remarkable success on vision-language tasks, yet fine-grained perception over high-resolution images

VISTA-DZ: Visual Semantic Trajectory Adaptation for Personalized Dilemma Zone Prediction

SafetyDGX agent

arXiv:2606.29548v1 Announce Type: cross Abstract: Driver decision making in the dilemma zone at signalized intersections is safety critical, as vehicles approaching a yellow signal must decide whether

Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration

SafetyDGX agent

arXiv:2508.14483v4 Announce Type: replace Abstract: We present Vivid-VR, a DiT-based generative video restoration method built upon an advanced T2V foundation model, where ControlNet is leveraged to c

VLK: Learning Humanoid Loco-Manipulation from Synthetic Interactions in Reconstructed Scenes

SafetyDGX agent

arXiv:2606.30645v1 Announce Type: cross Abstract: Perception-based humanoid loco-manipulation requires connecting egocentric observations and task instructions to whole-body motion. Learning this mapp

VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors

SafetyDGX agent

arXiv:2510.00458v3 Announce Type: replace Abstract: Vision-language object detectors (VLODs) such as YOLO-World and Grounding DINO exhibit strong zero-shot generalization, but their performance degrad

Voice AI is the core interface for future devices but it’s still not mainstream. And that is crazy to me. The quality is good enough to dict…

IndustryDGX agent

Voice AI is the core interface for future devices but it’s still not mainstream. And that is crazy to me. The quality is good enough to dictate everything (I’ve only had to add 20 words to my Wispr fl

Vorlon debuts Guardian to block risky AI agent actions before they complete

AgentsDGX agent

Agentic ecosystem security startup Vorlon Inc. today launched Guardian, a real-time enforcement gateway that aims to block risky actions by artificial intelligence agents before a transaction complete

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

Model ReleasesDGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

← Previous
1…461462463464465…1473
Next →