AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,832 results
Research

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation

DGX agent

arXiv:2604.02467v2 Announce Type: replace-cross Abstract: Cinematic camera control relies on a tight feedback loop between director and cinematographer, where camera motion and framing are continuousl

researcharxiv-cs-ai
14 Apr 2026
Research

Vestibular reservoir computing

DGX agent
X Post
Paper
YouTube
Reddit
GitHub

arXiv:2604.09943v1 Announce Type: new Abstract: Reservoir computing (RC) is a computational framework known for its training efficiency, making it ideal for physical hardware implementations. However,

researcharxiv-cs-lg
14 Apr 2026
Model Releases

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

DGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

VGGT-HPE: Reframing Head Pose Estimation as Relative Pose Prediction

DGX agent

arXiv:2604.10106v1 Announce Type: new Abstract: Monocular head pose estimation is traditionally formulated as direct regression from a single image to an absolute pose. This paradigm forces the networ

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Vibe-driven model-based engineering

DGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

researcharxiv-cs-ai
14 Apr 2026
Model Releases

VidAudio-Bench: Benchmarking V2A and VT2A Generation across Four Audio Categories

DGX agent

arXiv:2604.10542v1 Announce Type: cross Abstract: Video-to-Audio (V2A) generation is essential for immersive multimedia experiences, yet its evaluation remains underexplored. Existing benchmarks typic

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Video-based Heart Rate Estimation with Angle-guided ROI Optimization and Graph Signal Denoising

DGX agent

arXiv:2604.11395v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact heart rate measurement from facial videos, but its performance is significantly degraded by facia

researcharxiv-cs-cv
14 Apr 2026
Safety

VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG

DGX agent

arXiv:2604.05418v2 Announce Type: replace-cross Abstract: Scaling multimodal large language models (MLLMs) to long videos is constrained by limited context windows. While retrieval-augmented generatio

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Virtual Smart Metering in District Heating Networks via Heterogeneous Spatial-Temporal Graph Neural Networks

DGX agent

arXiv:2604.10166v1 Announce Type: cross Abstract: Intelligent operation of thermal energy networks aims to improve energy efficiency, reliability, and operational flexibility through data-driven contr

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

ViserDex: Visual Sim-to-Real for Robust Dexterous In-hand Reorientation

DGX agent

arXiv:2604.11138v1 Announce Type: cross Abstract: In-hand object reorientation requires precise estimation of the object pose to handle complex task dynamics. While RGB sensing offers rich semantic cu

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Vision-Language-Action Model, Robustness, Multi-modal Learning, Robot Manipulation

DGX agent

arXiv:2604.10055v1 Announce Type: new Abstract: Despite their strong performance in embodied tasks, recent Vision-Language-Action (VLA) models remain highly fragile under multimodal perturbations, whe

model-releasesarxiv-cs-ro
14 Apr 2026
Research

VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites

DGX agent

arXiv:2506.14629v3 Announce Type: replace-cross Abstract: Mosquito-borne diseases pose a major global health risk, requiring early detection and proactive control of breeding sites to prevent outbreak

researcharxiv-cs-cl
14 Apr 2026
Safety

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning

DGX agent

arXiv:2604.10500v1 Announce Type: new Abstract: Multimodal latent reasoning has emerged as a promising paradigm that replaces explicit Chain-of-Thought (CoT) decoding with implicit feature propagation

safetyarxiv-cs-cv
14 Apr 2026
Research

Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval

DGX agent

arXiv:2604.10167v1 Announce Type: cross Abstract: Multi-vector models dominate Visual Document Retrieval (VDR) due to their fine-grained matching capabilities, but their high storage and computational

researcharxiv-cs-cl
14 Apr 2026
Safety

VLMaterial: Vision-Language Model-Based Camera-Radar Fusion for Physics-Grounded Material Identification

DGX agent

arXiv:2604.11671v1 Announce Type: cross Abstract: Accurate material recognition is a fundamental capability for intelligent perception systems to interact safely and effectively with the physical worl

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions

DGX agent

arXiv:2604.10533v1 Announce Type: cross Abstract: Conventional Vision-and-Language Navigation (VLN) benchmarks assume instructions are feasible and the referenced target exists, leaving agents ill-equ

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Volumetric Ergodic Control

DGX agent

arXiv:2511.11533v3 Announce Type: replace-cross Abstract: Ergodic control synthesizes optimal coverage behaviors over spatial distributions for nonlinear systems. However, existing formulations model

researcharxiv-cs-ai
14 Apr 2026
Model Releases

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

DGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments

DGX agent

arXiv:2506.02387v3 Announce Type: replace Abstract: Recent advancements in Vision Language Models (VLMs) have expanded their capabilities to interactive agent tasks, yet existing benchmarks remain lim

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

VTC: DNN Compilation with Virtual Tensors for Data Movement Elimination

DGX agent

arXiv:2604.09558v1 Announce Type: cross Abstract: With the widening gap between compute and memory operation latencies, data movement optimizations have become increasingly important for DNN compilati

hardwarearxiv-cs-lg
14 Apr 2026
Safety

Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration

DGX agent

arXiv:2604.10245v1 Announce Type: new Abstract: Registration between preoperative CT and intraoperative laparoscopic video plays a crucial role in augmented reality (AR) guidance for minimally invasiv

safetyarxiv-cs-cv
14 Apr 2026
Safety

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations

DGX agent

arXiv:2604.10809v1 Announce Type: new Abstract: Recent advancements in learning from human demonstration have shown promising results in addressing the scalability and high cost of data collection req

safetyarxiv-cs-ro
14 Apr 2026
Agents

WaterAdmin: Orchestrating Community Water Distribution Optimization via AI Agents

DGX agent

arXiv:2604.10343v1 Announce Type: new Abstract: We study the operation of community water systems, where pumps and valves must be scheduled to reliably meet water demands while minimizing energy consu

agentsarxiv-cs-lg
14 Apr 2026
Model Releases

WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting

DGX agent

arXiv:2604.10544v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have recently achieved remarkable success in universal forecasting by leveraging large-scale pretraining on dive

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

DGX agent

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

local-air-stablediffusion
14 Apr 2026
Model Releases

WBCBench 2026: A Challenge for Robust White Blood Cell Classification Under Class Imbalance

DGX agent

arXiv:2604.10797v1 Announce Type: new Abstract: We present WBCBench 2026, an ISBI challenge and benchmark for automated WBC classification designed to stress-test algorithms under three key difficulti

model-releasesarxiv-cs-cv
14 Apr 2026
Tools

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model…

DGX agent

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model, in a new preprint with @MayoClinic. We're now releasing an

toolslinus-lee--x
14 Apr 2026
Research

We are aware of issues impacting model availability via the Nous Portal and are working to restore service ASAP

DGX agent

Nous Research posted a status update on X acknowledging service disruptions affecting model availability through their Nous Portal. The announcement indicated the team was actively working to restore

researchnous-research--x
14 Apr 2026
Research

We benchmarked TranslateGemma against 5 other LLMs on subtitle translation across 6 languages. At first glance the numbers told a clean story, but then human QA added a chapter. [D]

DGX agent

This r/MachineLearning discussion post details a hands-on benchmark study in which TranslateGemma — Google's open translation model suite built on Gemma 3, available in 4B, 12B, and 27B sizes and cove

researchr-machinelearning
14 Apr 2026
Industry

We don’t have the budget for that. We don’t have the time to do something like that. It’s too hard and will take too long. You have to be re…

DGX agent

We don’t have the budget for that. We don’t have the time to do something like that. It’s too hard and will take too long. You have to be reasonable. The client doesn’t have budget for an idea like th

industrycristobal-valenzuela--x
14 Apr 2026
Local Ai

We may have a new SOTA open-source model: ERNIE-Image Comparisons

DGX agent

ERNIE-Image is an open-weight text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

local-air-stablediffusion
14 Apr 2026
Agents

We see this as further validation that multi-agent architectures excel at novel problems outside training data distribution. These technique…

DGX agent

Cursor AI shared observations on X validating that multi-agent architectures demonstrate superior performance when tackling novel problems that fall outside the distribution of training data. The post

agentscursor--x
14 Apr 2026
Industry

We strictly prohibit users from generating non-consensual explicit deepfakes and from using our tools to undress real people. xAI has extens…

DGX agent

We strictly prohibit users from generating non-consensual explicit deepfakes and from using our tools to undress real people. xAI has extensive safeguards in place to prevent such misuse, such as cont

industryelon-musk--x
14 Apr 2026
Research

We will read histories of today’s US in a few decades and wonder how something like this could happen. Slowly allowing Polio and Measles to …

DGX agent

Yann LeCun shared or engaged with a post on X expressing concern about the erosion of public health infrastructure in the United States, specifically referencing the potential resurgence of eradicated

researchyann-lecun--x
14 Apr 2026
Model Releases

WearBCI Dataset: Understanding and Benchmarking Real-World Wearable Brain-Computer Interfaces Signals

DGX agent

arXiv:2604.09649v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) have opened new platforms for human-computer interaction, medical diagnostics, and neurorehabilitation. Wearable BCI

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

WebForge: Breaking the Realism-Reproducibility-Scalability Trilemma in Browser Agent Benchmark

DGX agent

arXiv:2604.10988v1 Announce Type: new Abstract: Existing browser agent benchmarks face a fundamental trilemma: real-website benchmarks lack reproducibility due to content drift, controlled environment

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

WebLLM: A High-Performance In-Browser LLM Inference Engine

DGX agent

arXiv:2412.15803v2 Announce Type: replace-cross Abstract: Advancements in large language models (LLMs) have unlocked remarkable capabilities. While deploying these models typically requires server-gra

local-aiarxiv-cs-ai
14 Apr 2026
Safety

Weird Generalization is Weirdly Brittle

DGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

We’re expanding Trusted Access for Cyber with additional tiers for authenticated cybersecurity defenders. Customers in the highest tiers can…

DGX agent

We’re expanding Trusted Access for Cyber with additional tiers for authenticated cybersecurity defenders. Customers in the highest tiers can request access to GPT-5.4-Cyber, a version of GPT-5.4 fine-

model-releasesopenai--x
14 Apr 2026
Tools

We're Transferring the Stripe Sync Engine to Stripe

DGX agent

Supabase originally built the open-source Stripe Sync Engine to keep its own Postgres database of billing data current with Stripe, using webhooks to update records for customers, invoices, and paymen

toolssupabase-blog
14 Apr 2026
Tools

We've been building something that doesn't fit in a wave. Coming soon

DGX agent

Windsurf, the AI-powered coding platform, has teased an upcoming product or feature announcement that is described as something that doesn't fit within their existing 'wave' release framework. The cry

toolswindsurf--x
14 Apr 2026
Hardware

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to ap…

DGX agent

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to apply it to optimizing CUDA kernels. In 3 weeks, it delivered

hardwarecursor--x
14 Apr 2026
Model Releases

We've been working on this for a while. Can't wait to hear what you think

DGX agent

We've been working on this for a while. Can't wait to hear what you think We've redesigned Claude Code on desktop. You can now run multiple Claude sessions side by side from one window, with a new sid

model-releasesboris-cherny--x
14 Apr 2026
Industry

We’ve entered the era of AI

DGX agent

We’ve entered the era of AI MPA boss Charles Rivkin says AI can “bolster the art of storytelling” and “improve the fan experience”: “We’ve entered the era of AI,” Rivkin told theater operators at #Cin

industrycristobal-valenzuela--x
14 Apr 2026
Model Releases

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cuttin…

DGX agent

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cutting it by 2-5x would be transformative. Last year, OSS models

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

DGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

What are some good env versions for speed and compatibility?

DGX agent

This Reddit thread from r/StableDiffusion discusses community recommendations for optimal Python, CUDA, PyTorch, and related dependency versions when setting up a Stable Diffusion environment that bal

local-air-stablediffusion
14 Apr 2026
Applications

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

DGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

applicationsarxiv-cs-cl
14 Apr 2026
← Previous
1…16671668166916701671…1747
Next →