AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,375 results
30 Jun 2026

VIB-AVSR: Variational Information Bottleneck for Noise-Robust LLM-Based Audio-Visual Speech Recognition

ResearchDGX agent

arXiv:2606.29632v1 Announce Type: cross Abstract: Audio-Visual Speech Recognition takes two input modalities, acoustic and visual streams, where visual information from lip movements aids recognition

VibES: Induced Vibration for Persistent Event-Based Sensing

ApplicationsDGX agent

arXiv:2508.19094v3 Announce Type: replace Abstract: Event cameras are a bio-inspired class of sensors that asynchronously measure per-pixel intensity changes. Under fixed illumination conditions in st

ViewSplat: View-Adaptive 3D Gaussian Splatting for Feed-Forward Synthesis

ResearchDGX agent

arXiv:2603.25265v2 Announce Type: replace Abstract: We present ViewSplat, a view-adaptive 3D Gaussian splatting network for novel view synthesis from unposed images. While recent feed-forward 3D Gauss

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

Model ReleasesDGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

Vimeo owner Bending Spoons raised 1.68B by selling 58M shares at 29 each, valuing it at ~$18.4B, in one of the largest US IPOs by a European company in 2026 (Subrat Patnaik/Bloomberg)

IndustryDGX agent

Subrat Patnaik / Bloomberg: Vimeo owner Bending Spoons raised 1.68B by selling 58M shares at 29 each, valuing it at ~18.4B, in one of the largest US IPOs by a European company in 2026 — Bending Spoons

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

Virtual Ring Try-On

ResearchDGX agent

arXiv:2606.28792v1 Announce Type: new Abstract: This paper presents an innovative approach that enables the users to capture their hand and try the jewel ring on their hand. The user captures the imag

Visa, Mastercard, Stripe, BlackRock, Coinbase, and 140+ companies join Open Standard to launch Open USD, a stablecoin that shares earnings from its reserves (Kyle Baird/The Block)

IndustryDGX agent

Kyle Baird / The Block: Visa, Mastercard, Stripe, BlackRock, Coinbase, and 140+ companies join Open Standard to launch Open USD, a stablecoin that shares earnings from its reserves — Quick Take — Paym

Visa, Stripe and 140 others back new Open USD stablecoin to challenge Tether

IndustryDGX agent

A consortium of more than 140 financial, payments and technology companies, including Visa Inc., Stripe Inc. and BlackRock Inc., is backing a new stablecoin called Open USD, taking direct aim at the m

Vision-driven Preference Synthesis for Mitigating Hallucinations in VLMs

SafetyDGX agent

arXiv:2606.28401v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong performance in visual understanding, yet they still suffer from hallucinations, generating content that

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

SafetyDGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

Vision-Language Models for Deployable Social Robot Navigation: Bridging Semantic Reasoning and Low-Level Control

SafetyDGX agent

arXiv:2606.28760v1 Announce Type: new Abstract: Social robot navigation (SRN) requires more than geometric path planning; it demands understanding human intentions, social norms, and contextual cues t

VisReflect: Latent Visual Reflection for Fine-Grained Perception in Long Visual Context

Local AiDGX agent

arXiv:2606.30288v1 Announce Type: new Abstract: Large Vision Language Models (LVLMs) have achieved remarkable success on vision-language tasks, yet fine-grained perception over high-resolution images

VISTA-DZ: Visual Semantic Trajectory Adaptation for Personalized Dilemma Zone Prediction

SafetyDGX agent

arXiv:2606.29548v1 Announce Type: cross Abstract: Driver decision making in the dilemma zone at signalized intersections is safety critical, as vehicles approaching a yellow signal must decide whether

Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration

SafetyDGX agent

arXiv:2508.14483v4 Announce Type: replace Abstract: We present Vivid-VR, a DiT-based generative video restoration method built upon an advanced T2V foundation model, where ControlNet is leveraged to c

VLK: Learning Humanoid Loco-Manipulation from Synthetic Interactions in Reconstructed Scenes

SafetyDGX agent

arXiv:2606.30645v1 Announce Type: cross Abstract: Perception-based humanoid loco-manipulation requires connecting egocentric observations and task instructions to whole-body motion. Learning this mapp

VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors

SafetyDGX agent

arXiv:2510.00458v3 Announce Type: replace Abstract: Vision-language object detectors (VLODs) such as YOLO-World and Grounding DINO exhibit strong zero-shot generalization, but their performance degrad

Voice AI is the core interface for future devices but it’s still not mainstream. And that is crazy to me. The quality is good enough to dict…

IndustryDGX agent

Voice AI is the core interface for future devices but it’s still not mainstream. And that is crazy to me. The quality is good enough to dictate everything (I’ve only had to add 20 words to my Wispr fl

Vorlon debuts Guardian to block risky AI agent actions before they complete

AgentsDGX agent

Agentic ecosystem security startup Vorlon Inc. today launched Guardian, a real-time enforcement gateway that aims to block risky actions by artificial intelligence agents before a transaction complete

VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On

Model ReleasesDGX agent

arXiv:2603.11734v2 Announce Type: replace Abstract: As virtual try-on (VTON) continues to advance, a growing number of real-world scenarios have emerged, pushing beyond the ability of the existing spe

W4A4 Quantization for Inference on Wan2.2-I2V-A14B

ResearchDGX agent

arXiv:2606.29337v1 Announce Type: new Abstract: We summarize our submission to Sub-Challenge 1: W4A4 Quantization for Inference (HiF4 / MXFP4) of the ICME 2026 Low-Bit-width Large-Model Quantization C

Walking in the Implicit: Interactive World Exploration via Neural Scene Representation

Local AiDGX agent

arXiv:2606.30045v1 Announce Type: new Abstract: Interactive video generation systems for camera-controlled world exploration roll out growing sequences of latent video frames, entangling state transit

Warm-Starting Iterative Gaussian Processes for Faster Sequential Inference

ResearchDGX agent

arXiv:2511.16340v2 Announce Type: replace Abstract: Efficient Gaussian process (GP) inference is critical for sequential decision-making tasks such as active learning, online prediction, and Bayesian

WARP: Whole-Body Retargeting for Learning from Offline Human Demonstrations

ApplicationsDGX agent

arXiv:2606.29940v1 Announce Type: new Abstract: Direct transfer from human demonstration to learnable robot action is a crucial step towards scalable whole-body mobile manipulation. While human data s

Wasserstein Distributionally Robust Regret Optimization

ResearchDGX agent

arXiv:2504.10796v4 Announce Type: replace-cross Abstract: Distributionally robust optimization (DRO) is widely used for decision-making under uncertainty, but its adversarial focus on worst-case loss

Watch Devin Fusion work: The sidekick model handles fetching information and skimming sources. The frontier model makes a plan for implement…

AgentsDGX agent

Watch Devin Fusion work: The sidekick model handles fetching information and skimming sources. The frontier model makes a plan for implementation, then delegates well-scoped sub-tasks to the sidekick

wav2VOT: Automatic estimation of voice onset time, closure duration, and burst realisation with wav2vec2

ResearchDGX agent

arXiv:2606.28857v1 Announce Type: cross Abstract: While automatic tools for speech annotation are now commonplace within phonetic research pipelines, many tasks require substantial manual correction o

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, …

Model ReleasesDGX agent

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, who are exploring to host multiple open source models and mo

We are rebuilding Hydrogen from the ground up with @Shopify. It's agent-first, runtime-agnostic, and runs anywhere JavaScript does. Learn mo…

AgentsDGX agent

We are rebuilding Hydrogen from the ground up with @Shopify. It's agent-first, runtime-agnostic, and runs anywhere JavaScript does. Learn more and try the Next.js developer preview ↓ https://vercel.co

we are talking loops today ✅

ToolsDGX agent

This post discusses loop constructs and control flow mechanisms in programming, likely covering concepts such as for loops, while loops, and their practical applications. Based on Swyx's typical conte

We build differently than other tech firms. No superintelligence. No competition to spend the most money. We're creating empowering, efficie…

Model ReleasesDGX agent

We build differently than other tech firms. No superintelligence. No competition to spend the most money. We're creating empowering, efficient AI to enhance human potential, not replace it. With high

We can finally say AI isn't killing jobs. A new paper from me, @tryramp, and @RevelioLabs uses firm-level spend and workforce data across 21…

ResearchDGX agent

We can finally say AI isn't killing jobs. A new paper from me, @tryramp, and @RevelioLabs uses firm-level spend and workforce data across 21K U.S. businesses to measure AI's impact on jobs. Firms that

We heard your feedback. You want to go faster. Introducing GLM 5.2 Fast The same model and quality as GLM 5.2 standard, now at 140 tok/s Fli…

ToolsDGX agent

Fireworks AI announced GLM 5.2 Fast, an optimized version of their GLM 5.2 model that maintains the same quality and capabilities as the standard version while delivering significantly faster inferenc

We just made it a lot easier for AI agents to work with your documents. LlamaParse MCP now does more than parse or classify files: it can pu…

Model ReleasesDGX agent

We just made it a lot easier for AI agents to work with your documents. LlamaParse MCP now does more than parse or classify files: it can pull structured data out of contracts, invoices, and reports a

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instea…

AgentsDGX agent

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instead of chained tool calls Code interpreters don't need a sandb

We tax cigarettes to reduce smoking. We tax alcohol to reduce drinking. We tax fuel to reduce driving. What do you think happens when you ta…

IndustryDGX agent

This post presents an argument about the unintended consequences of taxation policies, drawing a parallel between sin taxes on cigarettes, alcohol, and fuel—which aim to reduce consumption—and implyin

Weak Dominant Balance for Robust Identification of Dynamically Consistent Fluid Flow Structure

ResearchDGX agent

arXiv:2606.29047v1 Announce Type: cross Abstract: Extracting interpretable, localized physical mechanisms from complex spatiotemporal data is a foundational challenge across physics, biology, and engi

wearing my new babyagi/activegraph tee

AgentsDGX agent

Yohei Nakajima shared a post about wearing a new t-shirt related to BabyAGI and ActiveGraph, likely promoting or celebrating these AI projects. BabyAGI is a simplified autonomous AI agent framework, w

Weighted Contrastive Learning for Anomaly-Aware Time-Series Forecasting

ResearchDGX agent

arXiv:2512.07569v2 Announce Type: replace-cross Abstract: Reliable forecasting of multivariate time series under anomalous conditions is crucial in applications such as ATM cash logistics, where sudde

We're coming out of stealth. We've built our first racks after a successful A0 tapeout, 1B+ in customer contracts, and 800m raised. Early …

ResearchDGX agent

We're coming out of stealth. We've built our first racks after a successful A0 tapeout, 1B+ in customer contracts, and 800m raised. Early customer tests show us achieving SOTA throughput, latency, and

We're excited to be sponsoring and co-organizing IOL-AI 2026, a new open challenge with the International Linguistics Olympiad @IOLing_offic…

Model ReleasesDGX agent

We're excited to be sponsoring and co-organizing IOL-AI 2026, a new open challenge with the International Linguistics Olympiad @IOLing_official.🌎💬 🌐 Open-science 🗓️ One-month competition 🎯 Targeting A

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological …

Model ReleasesDGX agent

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment call

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then in…

Model ReleasesDGX agent

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then instantly animate them with the other—all at a fraction of the

We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring acces…

Model ReleasesDGX agent

We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring access tomorrow, and will share an update soon. We’re grateful to

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs f…

Model ReleasesDGX agent

What an honor to emcee the first day of @aiDotEngineer and introduce the Software Factories Track Thank you @swyx & team, and @KeycardLabs for the support. “A year ago @GeoffreyHuntley released the Ra

What Capable Agents Must Know: Selection Theorems for Robust Decision-Making under Uncertainty

AgentsDGX agent

arXiv:2603.02491v3 Announce Type: replace-cross Abstract: As artificial agents become increasingly capable, what internal structure is necessary for an agent to act competently under uncertainty? Clas

What Color is the Sky (for a non-human) ?

ResearchDGX agent

arXiv:2606.28912v1 Announce Type: new Abstract: The light of the daytime sky contains a mixture of many colors yet is perceived as blue by human observers. This is largely due to the particular respon

What Drives the Inlier-Memorization Effect? A Theory of Outlier Detection via Early Training Dynamics

Model ReleasesDGX agent

arXiv:2606.29791v1 Announce Type: cross Abstract: Outlier detection (OD) aims to identify anomalous instances by learning the underlying structure of normal data (inliers), and is particularly challen

What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs

TutorialsDGX agent

arXiv:2606.28615v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, where free-text explanations such as chain-of-thought and post-hoc rati

What's new in Claude Sonnet 5

Model ReleasesDGX agent

What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official

When AI Reviews Its Own Code: Recursive Self-Training Collapse in Code LLMs

Model ReleasesDGX agent

arXiv:2606.28438v1 Announce Type: cross Abstract: Recursive self-training can degrade neural generative models when generated data is reused without fresh human data or external quality control. We st

When Can Conformal Risk Control Certify LLM Outputs? Bounds, Impossibility, and Adaptation for Structured Generation

ResearchDGX agent

arXiv:2606.29054v1 Announce Type: new Abstract: Large language models (LLMs) deployed for structured generation (NER, JSON extraction, QA, and classification) lack formal reliability guarantees, and s

When Does Online Imitation Learning Help in LLM Post-Training? The Role of (Non-)Realizability Beyond Horizon

SafetyDGX agent

arXiv:2606.30445v1 Announce Type: new Abstract: Online imitation learning (IL), particularly on-policy distillation, has emerged as a strong LLM post-training approach, often outperforming offline sup

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

When Does Sparsity Mitigate the Curse of Depth in LLMs

ResearchDGX agent

arXiv:2603.15389v2 Announce Type: replace Abstract: Recent work has demonstrated the curse of depth in large language models (LLMs), where later layers contribute less to learning and representation t

When Does Synthetic CT Transfer? A Label-Free Donor/Host Diagnostic for Medical Vision-Language Model Routing on Real Lung CT

ResearchDGX agent

arXiv:2606.29232v1 Announce Type: new Abstract: A synthetic measurement of model competence is useful only if it survives the move to real data, yet the real labels that would verify it are exactly wh

When, How Long and How Much? Interpretable Neural Networks for Time Series Regression by Learning to Mask and Aggregate

ApplicationsDGX agent

arXiv:2512.03578v3 Announce Type: replace-cross Abstract: Time series extrinsic regression (TSER) refers to the task of predicting a continuous target variable from an input time series. It appears in

When Is a Draft Accepted? A Theory of Acceptance in Speculative Decoding

Local AiDGX agent

arXiv:2606.30265v1 Announce Type: cross Abstract: Speculative decoding accelerates language model inference by using a fast drafter to propose candidate tokens that are then verified by a larger targe

When LLMs Develop Languages: Symbolic Communication for Efficient Multi-Agent Reasoning

AgentsDGX agent

arXiv:2606.29354v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves large language models (LLMs) on difficult reasoning tasks, but it often incurs long natural-language rationales that are

When May I Help You? On The Effect of Proactivity on Group Human-Robot Collaboration

ResearchDGX agent

arXiv:2606.28469v1 Announce Type: new Abstract: Robot initiative is a central challenge in multi-party human-robot collaboration. A robot that contributes without being addressed may provide timely su

← Previous
1…462463464465466…1473
Next →