AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
1 Jul 2026

Scaling LLM Inference: Multi-Node KV Cache Offloading with GKE & Managed Lustre

Model ReleasesDGX agent

Significant contributors to this article include Sneha Aradhey, Software Engineer, Google Kubernetes Engine, and Michael MacDonald, Sr Software Engineer, Google Cloud Managed Lustre. Enterprise produc

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

SafetyDGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

Stage-Transition Dense Reward Modeling for Reinforcement Learning

ResearchDGX agent

arXiv:2606.31377v1 Announce Type: cross Abstract: Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping si

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Truth or Sophistry? LoFa: A Benchmark for LLM Robustness Against Logical Fallacies

Model ReleasesDGX agent

arXiv:2606.31039v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong semantic capabilities, yet their resilience to manipulative linguistic patterns such as logical fallacies re

What If We Allocate Test-Time Compute Adaptively?

Model ReleasesDGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

Model ReleasesDGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

30 Jun 2026

A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming

ResearchDGX agent

arXiv:2606.29221v1 Announce Type: new Abstract: We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots

A Machine-Verified Proof of a Quantum-Optimization Conjecture

Model ReleasesDGX agent

arXiv:2606.29687v1 Announce Type: cross Abstract: We report a machine-verified resolution of a problem open for over a decade in quantum optimization: the Farhi, Goldstone and Gutmann (FGG) conjecture

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

Model ReleasesDGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

Cross-Session 3D LiDAR and Camera Fusion for Robust Localization of Unmanned Aerial Vehicles in GPS-Denied Environments

ResearchDGX agent

arXiv:2606.28951v1 Announce Type: new Abstract: Accurate localization of unmanned aerial vehicles (UAVs) is essential for applications such as structural health monitoring, especially in environments

Cybersecurity is the True Frontier for Generative AI Success or Failure

ResearchDGX agent

arXiv:2606.28929v1 Announce Type: cross Abstract: Cybersecurity is a real-life test-bed for many machine learning problems at once, especially when considering modern strides in using Large Language M

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

Model ReleasesDGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

Embodiment Meets Environment: Toward Context-Aware, Safe Physical Caregiving Robots

ApplicationsDGX agent

arXiv:2606.28592v1 Announce Type: new Abstract: Physical caregiving robots need to assist different users with different tasks in diverse environments, and they come in many embodiments. While substan

Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

HardwareDGX agent

arXiv:2606.29082v1 Announce Type: new Abstract: Would experience designing faster GPU kernels also help close in on a long-standing open mathematical conjecture? Large Language Models (LLMs) integrate

Field Order Should Not Matter: Permutation-Invariant Embedding Model Fine-Tuning for Structured Metadata Retrieval

Model ReleasesDGX agent

arXiv:2606.30473v1 Announce Type: cross Abstract: We study retrieval over catalogs of structured metadata, where each record is a small schema whose fields answer different kinds of query. Embedding a

Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data

SafetyDGX agent

arXiv:2606.29793v1 Announce Type: new Abstract: Demand for personalized financial advising is growing, but consistent advisor expertise is difficult to obtain, scale, and encode in LLM systems. Simple

Google's new Nano Banana 2 Lite image model is its fastest and cheapest yet

IndustryDGX agent

Google introduced Nano Banana 2 Lite, its fastest and most cost-efficient image model built for high throughput, speed and scale. The model generates text-to-image outputs in about four seconds and co

Grounding LLM Reasoning under Incomplete Graph Evidence

TutorialsDGX agent

arXiv:2606.30247v1 Announce Type: new Abstract: Knowledge graphs can guide large language models (LLMs) reasoning, but the graph seen by a system is usually a retrieved, linked, temporally scoped, and

HEARTS: Benchmarking LLM Reasoning on Health Time Series

Model ReleasesDGX agent

arXiv:2603.06638v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has shifted time series analysis from narrow analytics to general-purpose reasoning. Yet, existing be

How much of an LLM-generated clinical corpus is actually new? A production-scale measurement of content redundancy for provenance classification

Model ReleasesDGX agent

arXiv:2606.29605v1 Announce Type: new Abstract: Clinical machine learning increasingly relies on training corpora generated by large language models (LLMs) rather than annotated by clinicians, and suc

Introducing Claude Sonnet 5 on AWS: Anthropic’s most capable Sonnet model

Model ReleasesDGX agent

Today, we’re excited to announce the availability of Anthropic’s most advanced Sonnet model, Claude Sonnet 5, on Amazon Bedrock and Claude Platform on AWS. Claude Sonnet 5 is the first Sonnet model of

Learning to Distributedly Estimate under Partially Known Dynamics: A Covariance-Agnostic Neural Kalman Consensus Filter

ResearchDGX agent

arXiv:2606.28441v1 Announce Type: cross Abstract: Online latent state estimation constitutes a fundamental challenge within the artificial intelligence field, serving as a foundational tool for divers

LLM Semantic Signaling Game and Mechanism Design: Systematic Blindness, Awareness Shaping, and Mindset Dynamics

SafetyDGX agent

arXiv:2606.29113v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate strategic interactions through natural language, making semantic control a critical element of commu

Parametric Skills

Model ReleasesDGX agent

arXiv:2606.30015v1 Announce Type: new Abstract: Since intelligence fundamentally relies on efficient skill acquisition (Chollet, 2019), the ability to leverage skills is critical. For LLMs, skills, ma

Physics Models for Sim-to-Real Transfer in Professional-Level Robot Table Tennis

SafetyDGX agent

arXiv:2606.28805v1 Announce Type: new Abstract: At competitive speeds and spins, a table tennis ball follows complex, counterintuitive trajectories that a robot must track and precisely counter within

Predictive Objectives Discard Exogenous Control-Relevant Features: A Controlled Mechanistic Study

SafetyDGX agent

arXiv:2606.30068v1 Announce Type: new Abstract: Joint-embedding predictive (JEPA-style) objectives learn representations by predicting future latents. In doing so they can discard features that are ex

Privacy-Preserving Decentralized Cooperative Localization with Range-Only Measurements: A Convex Optimization Based Approach

ResearchDGX agent

arXiv:2606.29673v1 Announce Type: new Abstract: Cooperative localization using range-based measurements is critical for multi-robot systems operating in GPS-denied and unstructured environments. Howev

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

Model ReleasesDGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

Model ReleasesDGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

Showing how to display dynamic subagents was tricky - this is what we aligned on

TutorialsDGX agent

This post likely discusses technical decisions and design patterns for dynamically displaying subagents within an AI system, documenting the alignment the team reached after encountering implementatio

Staged Hybridisation for Visual Quantum Reinforcement Learning via Knowledge Distillation

SafetyDGX agent

arXiv:2606.30520v1 Announce Type: cross Abstract: Visual environments are a demanding setting for quantum reinforcement learning (QRL): high-dimensional observations, unstable RL optimisation, and con

STEMGym: Benchmarking Sequential Decision-Making under Dose Budgets in Autonomous Electron Microscopy

Model ReleasesDGX agent

arXiv:2606.29592v1 Announce Type: new Abstract: A central premise of autonomous scientific imaging is that smarter navigation, whether Bayesian, RL-based, or otherwise adaptive, is the principal lever

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios

Model ReleasesDGX agent

arXiv:2511.17649v4 Announce Type: replace-cross Abstract: Tangible control interfaces (TCIs), such as appliance panels, remotes, elevators, and embedded GUIs, are a fundamental component of everyday h

TERC: A Transfer Entropy Redundancy Criterion for State Variable Selection in Reinforcement Learning

SafetyDGX agent

arXiv:2401.11512v2 Announce Type: replace-cross Abstract: Identifying the most suitable variables to represent the state is a fundamental challenge in Reinforcement Learning (RL). These variables must

The CRISTAL Method: Neurosymbolic analysis from AI-synthesized world models

Model ReleasesDGX agent

arXiv:2606.29799v1 Announce Type: new Abstract: This project introduces the CRISTAL Method (Coherent Reliable Intentional Synthesis of Truthful Analysis Logic), a neurosymbolic framework for automatin

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by h…

Model ReleasesDGX agent

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by how strongly GLM-5.2 is pushing long-horizon performance .. not just

The registrar's function in a hybrid society. AI value chain,smart data and the concept of property

ApplicationsDGX agent

arXiv:2606.28789v1 Announce Type: cross Abstract: Artificial intelligence reaches the land registry not as another tool but as a value chain that turns data into intelligence and intelligence into eco

The Verbose Context Problem in Medical Records

Model ReleasesDGX agent

arXiv:2606.29503v1 Announce Type: cross Abstract: The verbose context problem occurs when structured concepts have token-inefficient textual representations. This bottleneck is acute in population hea

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then in…

Model ReleasesDGX agent

We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then instantly animate them with the other—all at a fraction of the

When Stopping Fails: Rethinking Minimal Risk Conditions through Human-Interactive Autonomous Driving for Safe Transportation Systems

SafetyDGX agent

arXiv:2606.29115v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) are increasingly deployed in urban environments, yet their safety frameworks remain primarily designed around collision avoi

Zero-Label Driving Scenario Complexity Detection via Joint Embedding Predictive Architecture

SafetyDGX agent

arXiv:2606.28383v1 Announce Type: new Abstract: Identifying complex and safety-critical driving scenarios in large unlabelled datasets is an important but expensive problem. Existing approaches rely o

29 Jun 2026

AirGroundBench: Probing Spatial Intelligence in Multimodal Large Models under Heterogeneous Multi-View Embodied Collaboration

Model ReleasesDGX agent

arXiv:2606.28049v1 Announce Type: new Abstract: In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometric

because i'm not a design engineer myself, this track is one of the harder ones I struggle to curate. very fortunate to befriend Geoff who ha…

ToolsDGX agent

because i'm not a design engineer myself, this track is one of the harder ones I struggle to curate. very fortunate to befriend Geoff who has lent a hand to the past 2 years of AI UX meetups, and now

Benchmarking Multi-Modal Graph-based Social Media Popularity Prediction

Model ReleasesDGX agent

arXiv:2606.27539v1 Announce Type: cross Abstract: Social media popularity prediction aims to forecast the future reach or influence of online content from early-stage observations. Accurate prediction

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for…

Model ReleasesDGX agent

Chess engines tell you the best move. But grandmasters are human, they don’t always play it. So I built 'Kibitz': a human move predictor for chess broadcasts. I trained this model on my Nvidia RTX 508

LocalNav: Distilling Frontier VLMs and Embodied RL for On-Device Object Goal Navigation

Model ReleasesDGX agent

arXiv:2606.27871v1 Announce Type: new Abstract: Vision Language Models (VLMs) have emerged in the robotic domain as a powerful tool that enables environmental perception with language context, serving

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemin…

Model ReleasesDGX agent

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemini / Opus 4.6 level 35tr tokens trained entirely on 50k Chine

NormAct: A Benchmark for Hidden Social Norm Compliance in Embodied Planning

Model ReleasesDGX agent

arXiv:2606.27826v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly deployed as embodied planners in egocentric environments, where task success requires not only

Not All Relations Rotate Alike: Transformation-Aware Decoupling for Viewpoint-Robust 3D Scene Graph Generation

Model ReleasesDGX agent

arXiv:2606.27412v1 Announce Type: cross Abstract: 3D Scene Graph Generation (3DSGG) represents 3D scenes as structured object-relation-object graphs, providing a compact relational abstraction for spa

OverFlowLight: Real-Time Gridlock Prevention and Traffic Signal Optimization for Urban Intersections

SafetyDGX agent

arXiv:2606.27381v1 Announce Type: cross Abstract: Queue overflow, a severe consequence of urban traffic congestion, occurs when vehicle queues exceed intersection capacity, obstructing upstream traffi

Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs

ResearchDGX agent

arXiv:2603.20239v2 Announce Type: replace-cross Abstract: 3D Scene Graphs (3DSGs) provide hierarchical, multi-resolution abstractions that encode the geometric and semantic structure of an environment

Scaling Network Analysis for Fraud Prevention with BigQuery Graph

IndustryDGX agent

Based in the UK, Curve are building a financial super-app, a smart wallet that consolidates all your debit and credit cards into a single app and card, simplifying how millions of users spend, send an

Synthesize the big picture and analyze trends with BigQuery's AI.AGG function

Model ReleasesDGX agent

We recently announced the preview of the BigQuery AI.AGG() function. With AI.AGG(), you can use natural-language instructions within a single line of SQL to summarize or synthesize information over mi

Tandem Reinforcement Learning with Verifiable Rewards

ResearchDGX agent

arXiv:2606.28166v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly improved the reasoning capability of large language models, reaching expert or e

Towards Automating Scientific Review with Google's Paper Assistant Tool

Model ReleasesDGX agent

arXiv:2606.28277v1 Announce Type: cross Abstract: Artificial intelligence is driving a revolution in scientific discovery, accelerating everything from hypothesis generation to mathematical theorem pr

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs

Model ReleasesDGX agent

arXiv:2606.27909v1 Announce Type: cross Abstract: Theory-of-mind evaluations of large language models typically use dyadic social-deduction games, where every observable cue points to a single hidden

v0.30.12

Local AiDGX agent

v0.30.12 is a release candidate (rc0) that fixes a gemma4:12b floating point exception crash on x86, CUDA, Linux, and Windows systems. It includes improvements to ollama launch for Hermes Desktop, all

28 Jun 2026

Après avoir vu Elon répondre au Programme alimentaire mondial de l'ONU qui lui réclamait 6 milliards pour 'résoudre la faim dans le monde', …

IndustryDGX agent

Après avoir vu Elon répondre au Programme alimentaire mondial de l'ONU qui lui réclamait 6 milliards pour 'résoudre la faim dans le monde', j'ai compris quelque chose que je vais essayer de prouver ic

← Previous
1…270271272273274…296
Next →