AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,619 results
22 Apr 2026

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

Model ReleasesDGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

Model ReleasesDGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2604.19635v1 Announce Type: cross Abstract: While generative models have set new benchmarks for Target Speaker Extraction (TSE), their inherent reliance on global context precludes deployment in

Towards Understanding the Robustness of Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

Trainability Beyond Linearity in Variational Quantum Objectives

ResearchDGX agent

arXiv:2604.18846v1 Announce Type: cross Abstract: Barren-plateau results have established exponential gradient suppression as a widely cited obstacle to the scalability of variational quantum algorith

TransSplat: Unbalanced Semantic Transport for Language-Driven 3DGS Editing

TutorialsDGX agent

arXiv:2604.19571v1 Announce Type: new Abstract: Language-driven 3D Gaussian Splatting (3DGS) editing provides a more convenient approach for modifying complex scenes in VR/AR. Standard pipelines typic

Treehub launches with Tim Draper and Anne Wojcicki to back the next wave of AI health founders

Model ReleasesDGX agent

Treehub, a new Stanford University-adjacent residency program backed by the AI Health Fund, launched today, with billionaire investor Tim Draper and 23andMe Holding Co. founder Anne Wojcicki among the

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

ResearchDGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

SafetyDGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

SafetyDGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages…

Model ReleasesDGX agent

Truly sorry for any confusion or frustration caused by unclear, misleading, or inappropriate rules in our moderation system and on our pages. OpenClaw, Hermes, and SillyTavern are now explicitly marke

Try Kimi K2.6 now on the AI Native Cloud: http://www.together.ai/models/kimi-k26#

ToolsDGX agent

Kimi K2.6 is now available for use on Together AI's cloud platform, which offers AI model deployment and inference services. This announcement indicates the model has been added to Together AI's roste

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation

ResearchDGX agent

arXiv:2604.19473v1 Announce Type: new Abstract: Generating high-quality videos from complex temporal descriptions that contain multiple sequential actions is a key unsolved problem. Existing methods a

Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items

Model ReleasesDGX agent

arXiv:2604.19748v1 Announce Type: new Abstract: Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet compl

TurboEvolve: Towards Fast and Robust LLM-Driven Program Evolution

ResearchDGX agent

arXiv:2604.18607v1 Announce Type: cross Abstract: LLM-driven program evolution can discover high-quality programs, but its cost and run-to-run variance hinder reliable progress. We propose TurboEvolve

... turns out if you dig around in the browser network inspector enough you CAN find the prompt - here's the prompt it used for this image

ToolsDGX agent

This post documents a method for extracting the text prompt used by an image generation AI by examining network traffic in a web browser's developer tools. Simon Willison demonstrates that prompts can

Two-dimensional early exit optimisation of LLM inference

Model ReleasesDGX agent

arXiv:2604.18592v1 Announce Type: cross Abstract: We introduce a two-dimensional (2D) early exit strategy that coordinates layer-wise and sentence-wise exiting for classification tasks in large langua

Two new TPUs to power the next wave of AI training and inference at Google

HardwareDGX agent

Google LLC introduced two new custom silicon chips for artificial intelligence today at Google Cloud Next 2026, unveiling two distinct Tensor Processor Unit architectures built for training and infere

UAF: A Unified Audio Front-end LLM for Full-Duplex Speech Interaction

ApplicationsDGX agent

arXiv:2604.19221v1 Announce Type: new Abstract: Full-duplex speech interaction, as the most natural and intuitive mode of human communication, is driving artificial intelligence toward more human-like

Ultrametric OGP - parametric RDT symmetric binary perceptron connection

ResearchDGX agent

arXiv:2604.19712v1 Announce Type: new Abstract: In [97,99,100], an fl-RDT framework is introduced to characterize statistical computational gaps (SCGs). Studying symmetric binary perceptrons (SBPs), [

Uncertainty Quantification in Detection Transformers: Object-Level Calibration and Image-Level Reliability

Local AiDGX agent

arXiv:2412.01782v4 Announce Type: replace-cross Abstract: DETR and its variants have emerged as promising architectures for object detection, offering an end-to-end prediction pipeline. In practice, h

Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length

ResearchDGX agent

arXiv:2603.22608v2 Announce Type: replace Abstract: Users often rely on Large Language Models (LLMs) for processing multiple documents or performing analysis over a number of instances. For example, a

Unifying Controller Design for Stabilizing Nonlinear Systems with Norm-Bounded Control Inputs

ResearchDGX agent

arXiv:2403.03030v2 Announce Type: replace-cross Abstract: This paper revisits a classical challenge in the design of stabilizing controllers for nonlinear systems with a norm-bounded input constraint.

UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

Model ReleasesDGX agent

arXiv:2604.19734v1 Announce Type: cross Abstract: Scaling humanoid foundation models is bottlenecked by the scarcity of robotic data. While massive egocentric human data offers a scalable alternative,

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

Model ReleasesDGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

Unposed-to-3D: Learning Simulation-Ready Vehicles from Real-World Images

AgentsDGX agent

arXiv:2604.19257v1 Announce Type: new Abstract: Creating realistic and simulation-ready 3D assets is crucial for autonomous driving research and virtual environment construction. However, existing 3D

Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation

ResearchDGX agent

arXiv:2604.19444v1 Announce Type: new Abstract: Reasoning language models can solve increasingly complex tasks, but struggle to produce the calibrated confidence estimates necessary for reliable deplo

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

Model ReleasesDGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

URoPE: Universal Relative Position Embedding across Geometric Spaces

Model ReleasesDGX agent

arXiv:2604.18747v1 Announce Type: new Abstract: Relative position embedding has become a standard mechanism for encoding positional information in Transformers. However, existing formulations are typi

User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation

SafetyDGX agent

arXiv:2501.04410v2 Announce Type: replace Abstract: User simulation is an emerging interdisciplinary topic with multiple critical applications in the era of Generative AI. It involves creating an inte

Vast Data raises 1B at 30B valuation as AI infrastructure demand accelerates

IndustryDGX agent

Vast Data Inc. today said it has raised about 1 billion in a late-stage Series F funding round that values the company at 30 billion amid growing demand for infrastructure to support artificial intell

Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation (Kai Nicol-Schwarz/CNBC)

HardwareDGX agent

Kai Nicol-Schwarz / CNBC: Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation — Vast Data announc

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

Model ReleasesDGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

VDPP: Video Depth Post-Processing for Speed and Scalability

Model ReleasesDGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

Model ReleasesDGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. Th…

AgentsDGX agent

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. They watched 13 of them code in the wild and surveyed 99 more.

VideoAgent: Personalized Synthesis of Scientific Videos

Model ReleasesDGX agent

arXiv:2509.11253v2 Announce Type: replace Abstract: The technical complexity of research papers often limits their reach, necessitating more accessible formats like scientific videos to disseminate ke

ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios

Model ReleasesDGX agent

arXiv:2601.08620v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) pipelines must address challenges beyond simple single-document retrieval, such as interpreting visual elements

VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers

SafetyDGX agent

arXiv:2604.03261v2 Announce Type: replace Abstract: The rise of generative AI is posing increasing risks to online information integrity and civic discourse. Most concretely, such risks can materialis

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

SafetyDGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

Virtual boundary integral neural network for three-dimensional exterior acoustic problems

ResearchDGX agent

arXiv:2604.18636v1 Announce Type: cross Abstract: This paper presents a virtual boundary integral neural network (VBINN) for exterior acoustic problems in three dimensions. The method introduces a vir

Vision-Based Human Awareness Estimation for Enhanced Safety and Efficiency of AMRs in Industrial Warehouses

SafetyDGX agent

arXiv:2604.18627v1 Announce Type: new Abstract: Ensuring human safety is of paramount importance in warehouse environments that feature mixed traffic of human workers and autonomous mobile robots (AMR

VISTA: Verification In Sequential Turn-based Assessment

ResearchDGX agent

arXiv:2510.27052v5 Announce Type: replace Abstract: Hallucination--defined here as generating statements unsupported or contradicted by available evidence or conversational context--remains a major ob

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

SafetyDGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

Model ReleasesDGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images

Model ReleasesDGX agent

arXiv:2509.07966v2 Announce Type: replace-cross Abstract: Visual reasoning over structured data such as tables is a critical capability for modern vision-language models (VLMs), yet current benchmarks

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified chec…

Model ReleasesDGX agent

VLM Performance:Qwen3.6-27B is natively multimodal, supporting both vision-language thinking and non-thinking modes in a single unified checkpoint — the same as Qwen3.6-35B-A3B. It handles images and

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

Model ReleasesDGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding

ResearchDGX agent

arXiv:2604.19609v1 Announce Type: new Abstract: Transformers have become a common foundation across deep learning, yet 3D scene understanding still relies on specialized backbones with strong domain p

VoteGCL: Enhancing Graph-based Recommendations with Majority-Voting LLM-Rerank Augmentation

SafetyDGX agent

arXiv:2507.21563v4 Announce Type: replace-cross Abstract: Recommendation systems often suffer from data sparsity caused by limited user-item interactions, which degrade their performance and amplify p

Warmth and Competence in the Swarm: Designing Effective Human-Robot Teams

AgentsDGX agent

arXiv:2604.19270v1 Announce Type: new Abstract: As groups of robots increasingly collaborate with humans, understanding how humans perceive them is critical for designing effective human-robot teams.

Watch Sony’s elite ping-pong robot beat top-ranked players

IndustryDGX agent

Humans have been building ping-pong playing robots for decades, such as Omron's FOREPHUS that challenged amateur competitors at CES 2017. What sets Ace apart from the rest is that the robot, which was

Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs

Model ReleasesDGX agent

arXiv:2508.00161v3 Announce Type: replace-cross Abstract: The releases of powerful open-weight large language models (LLMs) are often not accompanied by access to their full training data. Existing in

We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress…

AgentsDGX agent

VideoGameBench is a new benchmark launched on Antim Labs, created by a1zhang, Thomas L. Griffiths, Karthik R. N, and Ofir Press. The benchmark likely evaluates AI model performance on video game-relat

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across …

Model ReleasesDGX agent

We just hit #1 on the @huggingface BrowseComp-Plus leaderboard. Best accuracy: 92.53%. Best recall: 88.79%. Lowest calibration error across all submissions. Built with @AI21Labs Maestro. https://huggi

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17

Model ReleasesDGX agent

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17 People are misreading the SpaceX/Cursor deal as an M&A story. It’s actually a b

We want you to have a lot of AI!

IndustryDGX agent

We want you to have a lot of AI! I don't know what they are doing over there, but Codex will continue to be available both in the FREE and PLUS ($20) plans. We have the compute and efficient models to

Weakly supervised framework for wildlife detection and counting in challenging Arctic environments: a case study on caribou (Rangifer tarandus)

SafetyDGX agent

arXiv:2601.18891v3 Announce Type: replace Abstract: Caribou across the Arctic has declined in recent decades, motivating scalable and accurate monitoring approaches to guide evidence-based conservatio

WebUncertainty: Dual-Level Uncertainty Driven Planning and Reasoning For Autonomous Web Agent

AgentsDGX agent

arXiv:2604.17821v2 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have empowered autonomous web agents to execute natural language instructions directly on real-w

← Previous
1…12181219122012211222…1411
Next →