AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
Industry

Open-ended coding training data may no longer be the bottleneck: AI can scale open-ended tasks—and even outperform human-expert curation. Fr…

DGX agent

Open-ended coding training data may no longer be the bottleneck: AI can scale open-ended tasks—and even outperform human-expert curation. FrontierCS team is releasing FrontierSmith: a system for synth

industryclem-delangue--x
15 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Peng's Q(lambda) for Conservative Value Estimation in Offline Reinforcement Learning

DGX agent

arXiv:2605.14779v1 Announce Type: new Abstract: We propose a model-free offline multi-step reinforcement learning (RL) algorithm, Conservative Peng's Q(lambda) (CPQL). Our algorithm adapts the Peng's

model-releasesarxiv-cs-lg
15 May 2026
Safety

Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands

DGX agent

arXiv:2605.15164v1 Announce Type: cross Abstract: This position paper argues that behavioural assurance, even when carefully designed, is being asked to carry safety claims it cannot verify. AI govern

safetyarxiv-cs-ai
15 May 2026
Model Releases

QOuLiPo: What a quantum computer sees when it reads a book

DGX agent

arXiv:2605.14188v1 Announce Type: cross Abstract: What does a book look like to a quantum computer? This paper takes eight classical works of the Renaissance and its late-antique inheritance -- from A

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Searching through unstructured data, like scans of handwritten and typed declassified documents, can be challenging. But with Cohere Compass…

DGX agent

Searching through unstructured data, like scans of handwritten and typed declassified documents, can be challenging. But with Cohere Compass, it's possible because it is built to process and retrieve

model-releasescohere--x
15 May 2026
Model Releases

SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection

DGX agent

arXiv:2605.14110v1 Announce Type: new Abstract: Vision Transformers (ViTs) enable strong multi-view 3D detection but are limited by high inference latency from dense token and query processing across

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding

DGX agent

arXiv:2510.13016v3 Announce Type: replace Abstract: A truly capable AI system must do more than detect objects or recognize activities in isolation. It must form unified, grounded representations of w

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Test-Time Learning with an Evolving Library

DGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily …

DGX agent

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily for all the other models. Claude Code: ollama launch claude

model-releasesollama--x
15 May 2026
Tools

A conversation with @sirupsen on scaling Shopify, building turbopuffer, and the future of databases. 0:00 - Scaling Shopify through flash sa…

DGX agent

A conversation with @sirupsen on scaling Shopify, building turbopuffer, and the future of databases. 0:00 - Scaling Shopify through flash sales and outages 8:13 - How top infrastructure teams collabor

toolscursor--x
14 May 2026
Safety

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

DGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

safetyarxiv-cs-ai
14 May 2026
Safety

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

DGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

safetyarxiv-cs-lg
14 May 2026
Tools

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brough…

DGX agent

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brought SF to SG. great showings from @daytonaio @usetusk @arizeai

toolsswyx--x
14 May 2026
Safety

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

DGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

safetyarxiv-cs-lg
14 May 2026
Safety

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

DGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

safetyarxiv-cs-ai
14 May 2026
Model Releases

BEAVER: An Enterprise Benchmark for Text-to-SQL

DGX agent

arXiv:2409.02038v3 Announce Type: replace-cross Abstract: Existing text-to-SQL benchmarks have largely been constructed from public databases with well-structured schemas and simplistic question-SQL p

model-releasesarxiv-cs-ai
14 May 2026
Safety

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

DGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

safetyarxiv-cs-ai
14 May 2026
Model Releases

CodeClash: Benchmarking Goal-Oriented Software Engineering

DGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

model-releasesarxiv-cs-ai
14 May 2026
Local Ai

Contextual Bandits for Resource-Constrained Devices using Probabilistic Learning

DGX agent

arXiv:2605.13346v1 Announce Type: new Abstract: Contextual bandits (CB) are online sequential decision-making problems under partial feedback that underpin many adaptive services. There is a growing d

local-aiarxiv-cs-lg
14 May 2026
Model Releases

D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models

DGX agent

arXiv:2605.13276v1 Announce Type: new Abstract: The rapid evolution of Embodied AI has enabled Vision-Language-Action (VLA) models to excel in multimodal perception and task execution. However, applyi

model-releasesarxiv-cs-ai
14 May 2026
Safety

Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration

DGX agent

arXiv:2603.22273v4 Announce Type: replace Abstract: The process of discovery requires active exploration -- the act of collecting new and informative data. However, efficient autonomous exploration re

safetyarxiv-cs-lg
14 May 2026
Local Ai

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

local-aiarxiv-cs-ai
14 May 2026
Model Releases

HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation

DGX agent

arXiv:2605.13321v1 Announce Type: new Abstract: VLN has achieved remarkable progress by scaling data and model capacity. However, the assumption of a static environment breaks down in real-world indoo

model-releasesarxiv-cs-ro
14 May 2026
Applications

Introducing Rime Mist v3 on Together AI, a production TTS family built for deterministic pronunciation and controllable voice output. AI nat…

DGX agent

Introducing Rime Mist v3 on Together AI, a production TTS family built for deterministic pronunciation and controllable voice output. AI natives can now deploy @rimelabs Mist v3 on Together AI dedicat

applicationstogether-ai--x
14 May 2026
Model Releases

Large Language Models Lack Temporal Awareness of Medical Knowledge

DGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

model-releasesarxiv-cs-lg
14 May 2026
Research

Limits of Personalizing Differential Privacy Budgets

DGX agent

arXiv:2605.13503v1 Announce Type: cross Abstract: A key technical difficulty in differential privacy is selecting a privacy budget that satisfies privacy requirements while maximizing utility. A natur

researcharxiv-cs-lg
14 May 2026
Research

Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems

DGX agent

arXiv:2512.08411v2 Announce Type: replace Abstract: Model-based planning in robotic domains is challenged by the hybrid nature of physical dynamics, where continuous motion is punctuated by discrete e

researcharxiv-cs-ai
14 May 2026
Local Ai

PROMETHEUS: Automating Deep Causal Research Integrating Text, Data and Models

DGX agent

arXiv:2605.12835v1 Announce Type: new Abstract: Large language models can extract local causal claims from text, but those claims become more useful when organized as persistent, navigable world model

local-aiarxiv-cs-ai
14 May 2026
Research

Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation

DGX agent

arXiv:2605.13129v1 Announce Type: cross Abstract: Recent 3D generative models can synthesize high-quality assets, but their outputs are typically static: they lack the skeletal rigs, joint hierarchies

researcharxiv-cs-cv
14 May 2026
Model Releases

scShapeBench: Discovering geometry from high dimensional scRNAseq data

DGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs

DGX agent

arXiv:2605.13737v1 Announce Type: new Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what it actually sees or hears, does the failure lie in perc

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SupChain-Bench: Benchmarking Large Language Models for Real-World Supply Chain Management

DGX agent

arXiv:2602.07342v2 Announce Type: replace Abstract: Large language models (LLMs) have shown promise in complex reasoning and tool-based decision making, motivating their application to real-world supp

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

TiCo: Time-Controllable Spoken Dialogue Model

DGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

model-releasesarxiv-cs-ai
14 May 2026
Hardware

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

DGX agent

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

hardwaretogether-ai--x
14 May 2026
Safety

Unweighted ranking for value-based decision making with uncertainty

DGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

safetyarxiv-cs-ai
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Model Releases

我們開源了這顆星球🌎上速度最快的低成本 bm25 引擎。

DGX agent

我們開源了這顆星球🌎上速度最快的低成本 bm25 引擎。 so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the

model-releasesemad-mostaque--x
13 May 2026
Safety

Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing

DGX agent

arXiv:2505.05665v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated success in decision-making tasks including planning, control, and prediction, but thei

safetyarxiv-cs-cl
13 May 2026
Safety

Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control

DGX agent

arXiv:2605.11775v1 Announce Type: cross Abstract: Policy entropy has emerged as a fundamental measure for understanding and controlling exploration in reinforcement learning with verifiable rewards (R

safetyarxiv-cs-cl
13 May 2026
Local Ai

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

DGX agent

arXiv:2605.11509v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control an

local-aiarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Safety

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

DGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

safetyarxiv-cs-lg
13 May 2026
Model Releases

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

DGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

model-releasesarxiv-cs-cl
13 May 2026
Safety

Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making

DGX agent

arXiv:2602.07668v2 Announce Type: replace Abstract: The looking-in-looking-out (LILO) framework has enabled intelligent vehicle applications that understand both the outside scene and the driver state

safetyarxiv-cs-cv
13 May 2026
Tutorials

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work eit…

DGX agent

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work either inflates context or retrains the model. This paper shows

tutorialsdair-ai--x
13 May 2026
Applications

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

DGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

applicationsarxiv-cs-lg
13 May 2026
Safety

Rainbow Deep Q-Learning with Kinematics-Aware Design for Cooperative Delta and 3-RRS Parallel Robot Insertion

DGX agent

arXiv:2605.11697v1 Announce Type: new Abstract: This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipula

safetyarxiv-cs-ro
13 May 2026
Safety

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

DGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

safetyarxiv-cs-ro
13 May 2026
← Previous
1…348349350351352…367
Next →