AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,020 results
2 Jun 2026

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out b…

Model ReleasesDGX agent

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out best practices, examples and more. I’m particularly excited a

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications

SafetyDGX agent

arXiv:2606.00133v1 Announce Type: new Abstract: World models, internal simulators that learn the structure and dynamics of an environment, have emerged as a central paradigm in the pursuit of artifici

World Models for Robotic Manipulation: A Survey

SafetyDGX agent

arXiv:2606.00113v1 Announce Type: new Abstract: Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world mode

Content type
AllBlogX PostPaperYouTubeRedditGitHub

World-Task Factorization for Robot Learning

SafetyDGX agent

arXiv:2606.02027v1 Announce Type: cross Abstract: Robot learning must produce policies that generalize to new combinations of constraints, teammates, and environments. To achieve this, we must structu

WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching

Model ReleasesDGX agent

arXiv:2603.06331v2 Announce Type: replace Abstract: Diffusion-based world models have shown strong potential for unified world simulation, but the iterative denoising remains too costly for interactiv

WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis

Model ReleasesDGX agent

arXiv:2606.01869v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly asked not only to write static interfaces, but to construct executable interactive worlds from natural lan

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World

Model ReleasesDGX agent

arXiv:2512.10958v2 Announce Type: replace Abstract: Generative world models are reshaping embodied AI, enabling agents to synthesize realistic 4D driving environments that look convincing but often fa

Worlds Within Words: Translating Culture in Ancient Chinese Texts with Multi-Agent Coordination

Model ReleasesDGX agent

arXiv:2606.01276v1 Announce Type: new Abstract: Large language model (LLM)-based machine translation has advanced cross-cultural communication, yet it still struggles with culture-loaded words (CLWs)

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

Model ReleasesDGX agent

arXiv:2512.00956v3 Announce Type: replace Abstract: Quantizing LLM weights and activations is a standard approach for efficient deployment, but a few extreme outliers can stretch the dynamic range and

X-Foresight: A Joint Vision-Action Causal Forecasting Network via Predictive World Modeling

SafetyDGX agent

arXiv:2605.24892v2 Announce Type: replace Abstract: Physical world knowledge resides mainly in videos. Equipping Vision-Language-Action (VLA) models with such knowledge is fundamental for safe and gen

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding

Model ReleasesDGX agent

arXiv:2606.02482v1 Announce Type: new Abstract: While video streaming understanding has made significant strides, real-world applications, such as live sports broadcasting, autonomous driving, and mul

@xai @imagine see more about flipbook from @zan2434 and @eddiejiao_obj ! https://x.com/eddiejiao_obj/status/2046975783324004732?s=20

ToolsDGX agent

@xai @imagine see more about flipbook from @zan2434 and @eddiejiao_obj ! https://x.com/eddiejiao_obj/status/2046975783324004732?s=20 What if your whole computer were just pixels streamed to you from a

XAI-SOH-FL: Enhancing SOH-FL with Adaptive Aggregation and Explainable AI for Intrusion Detection in Heterogeneous IoT

Model ReleasesDGX agent

arXiv:2606.00134v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDS) in Internet of Things (IoT) environments face significant challenges due to data heterogeneity, lack of labeled data

XD-RCDepth: Lightweight Radar-Camera Depth Estimation with Explainability-Aligned and Distribution-Aware Distillation

AgentsDGX agent

arXiv:2510.13565v2 Announce Type: replace Abstract: Depth estimation remains central to autonomous driving, and radar-camera fusion offers robustness in adverse conditions by providing complementary g

'You are not likely to see Henry Nowak’s words stenciled on a mural. No corporation will change its logo. The same establishment that made a…

IndustryDGX agent

'You are not likely to see Henry Nowak’s words stenciled on a mural. No corporation will change its logo. The same establishment that made a few words immortal when spoken by a black man in Minneapoli

You Can Learn Tokenization End-to-End with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.13940v2 Announce Type: replace-cross Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large Language Models (LLMs), despite a general trend t

You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models

TutorialsDGX agent

arXiv:2603.00133v2 Announce Type: replace-cross Abstract: Generative models have been shown to 'memorize' certain training data, leading to verbatim or near-verbatim generating images, which may cause

you have to give @MicrosoftAI props for training all these in-house from scratch and getting ALL of them to near-SOTA. Mustafa built a full …

ToolsDGX agent

you have to give @MicrosoftAI props for training all these in-house from scratch and getting ALL of them to near-SOTA. Mustafa built a full fledged neolab inside Microsoft in 2 years, that now MS full

Zamba2-VL Technical Report

Model ReleasesDGX agent

arXiv:2606.00390v1 Announce Type: cross Abstract: We present Zamba2-VL, a suite of vision-language models built on Zamba2, a hybrid language-model architecture combining Mamba2 state-space layers with

Zero-Shot Multi-Animal Tracking in the Wild

ResearchDGX agent

arXiv:2511.02591v2 Announce Type: replace Abstract: Multi-animal tracking is crucial for understanding animal ecology and behavior, yet remains challenging due to variations in habitat, motion pattern

Zero-Shot Off-Policy Learning

Model ReleasesDGX agent

arXiv:2602.01962v2 Announce Type: replace-cross Abstract: Off-policy learning methods seek to derive an optimal policy directly from a fixed dataset of prior interactions. This objective presents sign

Zhipu AI says it plans to apply for a listing in Shanghai; Zhipu's Hong Kong-listed shares are up over 10x since its January IPO, giving it an $83B market cap (Reuters)

IndustryDGX agent

Reuters: Zhipu AI says it plans to apply for a listing in Shanghai; Zhipu's Hong Kong-listed shares are up over 10x since its January IPO, giving it an $83B market cap — Knowledge Atlas Technology JSC

1 Jun 2026

$200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀

HardwareDGX agent

200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀 ANOTHER big AI Lab is massively investing in London. This time it is @runwayml

3DAE: Binaural Quality Assessment for Audio Novel View Synthesis with Spatial Maps and Benchmark

Model ReleasesDGX agent

arXiv:2605.30469v1 Announce Type: cross Abstract: 3D audio and novel-view acoustic synthesis models are usually evaluated with global metrics.However, global metrics often hide where and why binaural

3ViewSense: Spatial and Mental Perspective Reasoning from Orthographic Views in Vision-Language Models

ResearchDGX agent

arXiv:2603.07751v2 Announce Type: replace-cross Abstract: Current Large Language Models have achieved Olympiad-level logic, yet Vision-Language Models paradoxically falter on elementary spatial tasks

6 months ago I said I won't stop until I have Kimi at home, after 10+ botched REAPs I finally have it Needs benchmarking of course. - 45 tok…

TutorialsDGX agent

6 months ago I said I won't stop until I have Kimi at home, after 10+ botched REAPs I finally have it Needs benchmarking of course. - 45 tok/s decode - 954 tok/s prefill no cache - 95k+ tok/s cached p

本日6月1日より、Sakana AI @SakanaAILabs に Applied Research Engineer としてジョインしました!🐟 防衛・インテリジェンス領域のAI開発に取り組みます。国の安全保障に関わる重要な分野において、最先端技術を通じた実社会へのインパク…

ResearchDGX agent

David Ha announced joining Sakana AI as an Applied Research Engineer starting June 1st, focusing on AI development in defense and intelligence sectors. His work will apply cutting-edge technology to c

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents

AgentsDGX agent

arXiv:2602.08964v2 Announce Type: replace-cross Abstract: Understanding an agent's goals helps explain and predict its behaviour, yet there is no established methodology for reliably attributing goals

A Context-Aware Middleware for Medical Image Based Reports: An approach based on image feature extraction and association rules

ResearchDGX agent

arXiv:2605.30699v1 Announce Type: cross Abstract: This work proposes a context-aware middleware for medical workflow organization and efficiency improvement. In hospitals, laboratories and teleradiolo

A Google image search of “white mother” returns images of white women with only black children. Google is peak woke.

IndustryDGX agent

I can't verify this claim without access to current Google search results. Google's image search results are algorithmic and based on factors like image metadata, user engagement patterns, and relevan

A hitchhiker's guide to Poisson gradient estimation

SafetyDGX agent

arXiv:2602.03896v2 Announce Type: replace-cross Abstract: Poisson-distributed latent variable models are widely used in computational neuroscience, but differentiating through discrete stochastic samp

A holomorphic neural network framework for 3D boundary value problems governed by harmonic potentials

ResearchDGX agent

arXiv:2605.31231v1 Announce Type: cross Abstract: We present a neural-network-based framework for the solution of three-dimensional boundary value problems where the solution is expressible in terms o

A Kinetic Energy Perspective of Flow Matching

Model ReleasesDGX agent

arXiv:2602.07928v2 Announce Type: replace-cross Abstract: Flow-based generative models can be viewed through a physics lens: sampling transports a particle from noise to data by integrating a learned

A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models

SafetyDGX agent

arXiv:2605.30843v1 Announce Type: new Abstract: In the forward reinforcement-learning problem, the reward is fixed and known; the learner is asked to find a good policy or value function. Here we turn

A Lightweight Ensemble-Based Face Image Quality Assessment Method with Correlation-Aware Loss

Model ReleasesDGX agent

arXiv:2509.10114v2 Announce Type: replace Abstract: Face image quality assessment (FIQA) plays a critical role in face recognition and verification systems, especially in uncontrolled, real-world envi

A local to the police as soon as they arrive: 'He has a mouth full of blood' Henry, in and out of consciousness : 'I can't breathe, I've bee…

IndustryDGX agent

A local to the police as soon as they arrive: 'He has a mouth full of blood' Henry, in and out of consciousness : 'I can't breathe, I've been stabbed' …WHAT was 'complex' about that Hampshire police?

A look at the Seckinger school cluster in Georgia, including the US' 'first AI-themed educational institution', as parents say AI integration is often sparse (New York Times)

IndustryDGX agent

New York Times: A look at the Seckinger school cluster in Georgia, including the US' “first AI-themed educational institution”, as parents say AI integration is often sparse — It was 9 a.m. on a Thurs

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition

HardwareDGX agent

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition Introducing the Cosmos Coalition A new global initiative with NVIDIA and leading AI labs to build and open-source

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent.

HardwareDGX agent

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent. This is the NVIDIA RTX Spark Superchip. A new beginning for personal computers. Designed for creators,

A Novel Computer Vision Approach for Assessing Fish Responses to Intrusive Objects in Aquaculture

ApplicationsDGX agent

arXiv:2605.30399v1 Announce Type: cross Abstract: The aquaculture industry needs to address several challenges to secure sustainable seafood production that can serve an increasing global demand. One

A Novel Evaluation Metric for Unsupervised Learning in AIS-Based Maritime Anomaly Detection: MADQI

ResearchDGX agent

arXiv:2605.30388v1 Announce Type: new Abstract: This paper introduces a new systematic framework for detecting anomalies in maritime Automatic Identification System (AIS) datasets. These anomalies inc

A Novel Global Context-aware Deep Neural Network for Enhanced Brain Tumor Segmentation using Magnetic Resonance Images

Model ReleasesDGX agent

arXiv:2605.30510v1 Announce Type: cross Abstract: Brain cancer's severity necessitates precise brain tumor segmentation, which is crucial for effective brain tumor diagnosis. Manual identification, bu

A Padding Method for Enhanced Encoding of Inorganic Structures with Varying Chemical Compositions

ResearchDGX agent

arXiv:2605.30743v1 Announce Type: cross Abstract: Designing novel inorganic materials through generative models remains an important challenge for material science, driven by the complexity and divers

a parallel experiment building a coding agent on top of @activegraphai. you can see everything flattened down to a single event log trace

AgentsDGX agent

This post documents a parallel experiment where Yohei Nakajima built a coding agent using ActiveGraphAI, showcasing the system's architecture through a flattened event log trace that makes all operati

A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI

SafetyDGX agent

arXiv:2605.31021v1 Announce Type: new Abstract: Current alignment paradigms for generative artificial intelligence rely predominantly on monolithic benchmarking frameworks that reduce the plurality of

A Perturbation Approach to Unconstrained Linear Bandits

ResearchDGX agent

arXiv:2603.28201v2 Announce Type: replace Abstract: We revisit the standard perturbation-based approach of Abernethy et al. (2008) in the context of unconstrained Bandit Linear Optimization (uBLO). We

A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models

ResearchDGX agent

arXiv:2605.31080v1 Announce Type: cross Abstract: Blind and low-vision (BLV) audiences remain underserved by visual art descriptions, particularly across languages and in museum settings where privacy

A prompt can cost a million times more than an HTTP request, so token theft is a high-margin business for attackers. How we protect our AI e…

ToolsDGX agent

A prompt can cost a million times more than an HTTP request, so token theft is a high-margin business for attackers. How we protect our AI endpoints ↓ https://vercel.com/blog/protecting-against-token-

A study on a Real-Time VR-Based Teleoperation Framework for Manipulator in Dynamic Environment

SafetyDGX agent

arXiv:2605.30989v1 Announce Type: new Abstract: Robot teleoperation enables safe, non-contact task execution in hazardous environments where direct human access is difficult, and its application has e

A Survey on Semantic Communication for Vision: Categories, Frameworks, Enabling Techniques, and Applications

ResearchDGX agent

arXiv:2601.22202v2 Announce Type: replace-cross Abstract: Semantic communication (SemCom) emerges as a transformative paradigm for traffic-intensive visual data transmission, shifting focus from raw d

A Tight Theory of Error Feedback Algorithms in Distributed Optimization

AgentsDGX agent

arXiv:2605.31594v1 Announce Type: new Abstract: Communication costs are a major bottleneck in distributed learning and first-order optimization. A common approach to alleviate this issue is to compres

A Unified and Reproducible Experimentation Framework for Speech Understanding

AgentsDGX agent

arXiv:2605.30899v1 Announce Type: cross Abstract: Speech foundation models and Speech LLMs have advanced speech understanding, yet deployment-oriented model selection is hindered by non-comparable eva

A Unified Framework for Gradient Aggregation in Multi-Objective Optimization

SafetyDGX agent

arXiv:2605.30452v1 Announce Type: cross Abstract: Many machine learning problems involve multiple inherent trade-offs that are best addressed by gradient-based multi-objective optimization (MOO) algor

A Unifying View of Anchoring via Operator-Side Tikhonov Regularization

ResearchDGX agent

arXiv:2605.30905v1 Announce Type: cross Abstract: Anchored fixed point and monotone equation methods, including Halpern iteration, extra anchored gradient, and their relatives, add a vanishing pull to

A Unifying View of Variational Generative Wasserstein Flows

ResearchDGX agent

arXiv:2605.31369v1 Announce Type: cross Abstract: Many modern generative models can be viewed as minimizing divergences between probability distributions, yet they rely on different algorithmic and ge

A Visually Impaired Assistance Benchmark for VLM-as-a-Judge Evaluation

Model ReleasesDGX agent

arXiv:2605.31351v1 Announce Type: new Abstract: AI-based Visually Impaired Assistance (VIA) remains challenging, largely due to the high cost of human evaluation. The VLM-as-a-Judge paradigm may offer

AbstainGNN: Teaching Graph Neural Networks to Abstain for Graph Classification

Model ReleasesDGX agent

arXiv:2605.30786v1 Announce Type: new Abstract: Graph classification is a core task in graph data mining with widespread real-world applications. Recent advances in graph neural networks (GNNs) have l

Accelerate LLM model loading and increase context windows with GPUDirect on Amazon FSx for Lustre and TurboQuant

HardwareDGX agent

If you’re iterating on deploying large language models (LLMs) on AWS GPU instances, you’ve probably noticed the larger the model to be loaded into GPU High Bandwidth Memory (HBM), the longer the painf

Accelerated Multiple Wasserstein Gradient Flows for Multi-objective Distributional Optimization

ResearchDGX agent

arXiv:2601.19220v2 Announce Type: replace Abstract: We study multi-objective optimization over probability distributions in Wasserstein space. Recently, Nguyen et al. (2025) introduced Multiple Wasser

Active Timepoint Selection for Learning Measure-Valued Trajectories

SafetyDGX agent

arXiv:2605.30625v1 Announce Type: cross Abstract: Inferring continuous probability paths from sparse snapshots is a fundamental challenge in domains like single-cell biology, where high-fidelity data

← Previous
1…763764765766767…1517
Next →