AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,723 results
Model Releases

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment

DGX agent

arXiv:2606.04569v1 Announce Type: new Abstract: Underground mines present extreme conditions for autonomous robot navigation: GPS is denied, lighting is degraded, and tunnel topology is loop-rich and

model-releasesarxiv-cs-ro
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

NVIDIA Nemotron 3 Ultra now available on Amazon SageMaker JumpStart

DGX agent

The search results primarily discuss the Nemotron 3 Nano Omni model rather than Nemotron 3 Ultra. However, I found a recent NVIDIA blog post reference that indicates Nemotron 3 Ultra is available thro

model-releasesaws-ml-blog
4 Jun 2026
Model Releases

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC init…

DGX agent

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC initiative. We are scaling up to build Japan’s first 1T paramete

model-releasesdavid-ha--x
4 Jun 2026
Research

Reasoning Shift: How Context Silently Shortens LLM Reasoning

DGX agent

arXiv:2604.01161v2 Announce Type: replace Abstract: Large language models (LLMs) exhibiting test-time scaling behavior, such as extended reasoning traces and self-verification, have demonstrated remar

researcharxiv-cs-lg
4 Jun 2026
Safety

Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.04923v1 Announce Type: cross Abstract: Rubric-based reinforcement learning (RL) uses an LLM-as-a-Judge (LaaJ) to score model outputs according to rubrics as rewards. However, policy models

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Self-Evolving Deep Research via Joint Generation and Evaluation

DGX agent

arXiv:2606.04507v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly adopted in daily applications, with deep research standing out as a particularly important capab

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Trace-Mediated Peak Bias: Bridging Temporal Credit Assignment and Cognitive Heuristics in Deep Reinforcement Learning

DGX agent

arXiv:2606.04735v1 Announce Type: cross Abstract: Temporal credit assignment is central to both biological and artificial intelligence, yet its interaction with non-linear function approximation is po

safetyarxiv-cs-ai
4 Jun 2026
Applications

Trusted healthcare AI hinges on data foundations, not models alone

DGX agent

Healthcare AI is moving out of the pilot phase and into production environments, where the gap between a compelling demo and a clinically trustworthy output has never been more consequential. As AI ag

applicationssiliconangle
4 Jun 2026
Model Releases

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k p…

DGX agent

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages of real-world enterprise documents ✅ It has comprehensi

model-releasesjerry-liu--x
4 Jun 2026
Tutorials

what makes building on replit different? it all happens in one place. → describe your idea in plain english, get working software → generate…

DGX agent

what makes building on replit different? it all happens in one place. → describe your idea in plain english, get working software → generate UI, add auth + databases, deploy → collaborate with your te

tutorialsreplit--x
4 Jun 2026
Model Releases

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

DGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

DGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Cosmos 3: Omnimodal World Models for Physical AI

DGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Google's new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM

DGX agent

Google DeepMind released Gemma 4 12B, an open AI model that brings multimodal capabilities to everyday laptops by processing text, images, and audio natively without separate encoders. Small enough to

model-releasesars-technica
3 Jun 2026
Applications

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure…

DGX agent

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure Opus 4.7 on quality and costs less. This is the multi-model

applicationsclem-delangue--x
3 Jun 2026
Model Releases

Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition

DGX agent

arXiv:2511.21731v2 Announce Type: replace-cross Abstract: We present the results of cognitive tests on conceptual combinations, performed using specific Large Language Models (LLMs) as test subjects.

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Learning to Bet for Horizon-Aware Anytime-Valid Testing

DGX agent

arXiv:2603.19551v2 Announce Type: replace-cross Abstract: We develop horizon-aware anytime-valid tests and confidence sequences for bounded means under a strict deadline N. Using the betting/e-process

safetyarxiv-cs-lg
3 Jun 2026
Industry

Microsoft and OpenAI broke up — now they’re ready to fight

DGX agent

At Microsoft's annual Build conference on Tuesday, the company announced a slew of new or expanded AI initiatives, including a super app, in-house reasoning models, a cybersecurity tool, and OpenClaw-

industrythe-verge-ai
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Tutorials

Nice primer on post-training reasoning data. (bookmark it) This is one of the first primers to pull the scattered post-training reasoning-da…

DGX agent

Nice primer on post-training reasoning data. (bookmark it) This is one of the first primers to pull the scattered post-training reasoning-data literature into one place, synthesizing over 150 public s

tutorialsdair-ai--x
3 Jun 2026
Safety

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

DGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Proof-Refactor: Refactoring Generated Formal Proofs into Modular Artifacts

DGX agent

arXiv:2606.03743v1 Announce Type: new Abstract: While Large Language Models (LLMs) have shown strong performance in generating formal proofs, their outputs often remain less readable, modular, maintai

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as m…

DGX agent

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as most companies are currently learning (in addition to giving

applicationsclem-delangue--x
3 Jun 2026
Model Releases

SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series

DGX agent

arXiv:2606.03301v1 Announce Type: new Abstract: We introduce SagaQA, a long-form video benchmark for multi-hop reasoning over full-length TV series. Existing video reasoning benchmarks often emphasize

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Spike-Aware C++ INT8 Inference for Sparse Spiking Language Models on Commodity CPUs

DGX agent

arXiv:2606.03026v1 Announce Type: cross Abstract: Spiking language models expose activation sparsity that dense Transformer runtimes do not directly exploit. This paper studies that property from a sy

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

Two bits here I think we're going to see a lot more: 1. custom harnesses / finetunes on smaller open source models to beat frontier models a…

DGX agent

Two bits here I think we're going to see a lot more: 1. custom harnesses / finetunes on smaller open source models to beat frontier models at specific skill or task 2. using frontier models as critics

applicationsclem-delangue--x
3 Jun 2026
Safety

Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning

DGX agent

arXiv:2606.03962v1 Announce Type: cross Abstract: Classical reinforcement learning (RL) typically seeks a deterministic policy that maximizes the expected sum of a scalar reward. Yet, modern applicati

safetyarxiv-cs-ai
3 Jun 2026
Hardware

What’s new in serverless Managed Service for Apache Spark

DGX agent

Whether you use it for data preparation, real-time interactive queries, AI model training, or something entirely different, running Apache Spark at scale is demanding — you shouldn’t have to manage th

hardwaregoogle-cloud-ai
3 Jun 2026
Research

A Theoretical Framework for Self-Play Theorem Proving Algorithms

DGX agent

arXiv:2606.01861v1 Announce Type: new Abstract: Self-play, a type of training algorithm that enables a model to self-improve, has recently shown promising empirical results in the context of formal th

researcharxiv-cs-lg
2 Jun 2026
Applications

Algorithmic algorithm development with LLMs: A Case Study on LLM-Usage for Contraction Order Optimization in Tensor Networks

DGX agent

arXiv:2606.01975v1 Announce Type: new Abstract: We consider LLM-based algorithm development through a case study on contractionorder optimisation for tensor networks with OpenEvolve. We pay particular

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation

DGX agent

arXiv:2606.00987v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown strong visual understanding and language-guided grounding abilities, yet their capacity for multi-temp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Announcing Spanner Graph algorithms: Google-grade intelligence for connected data

DGX agent

At Google Cloud Next, we announced the preview of graph algorithms with Spanner Graph, bringing Google Research’s state-of-the-art graph mining capabilities natively to your database. These graph inte

model-releasesgoogle-cloud-ai
2 Jun 2026
Model Releases

ChartArena: Benchmarking Chart Parsing across Languages, Scenarios, and Formats

DGX agent

arXiv:2606.01348v1 Announce Type: new Abstract: Charts are a primary medium for conveying quantitative and relational information, yet systematically evaluating chart parsing models remains difficult.

model-releasesarxiv-cs-cv
2 Jun 2026
Industry

Cisco’s new cloud platform aimed at securing AI infrastructure

DGX agent

Cisco Systems Inc. today introduced a sweeping set of products and services designed to help enterprises manage, secure and automate increasingly complex information technology environments as artific

industrysiliconangle
2 Jun 2026
Safety

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence

DGX agent

arXiv:2606.00235v1 Announce Type: cross Abstract: We argue that governance must transition from a normative discipline to an engineering discipline, and we develop a formal framework, inspired by the

safetyarxiv-cs-ai
2 Jun 2026
Safety

Control of a Twin Rotor using Twin Delayed Deep Deterministic Policy Gradient (TD3)

DGX agent

arXiv:2512.13356v2 Announce Type: replace-cross Abstract: This paper proposes a reinforcement learning (RL) framework for controlling and stabilizing the Twin Rotor Aerodynamic System (TRAS) at specif

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation

DGX agent

arXiv:2606.01212v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are widely deployed and increasingly influential, but their reliance on external corpora exposes new secu

local-aiarxiv-cs-ai
2 Jun 2026
Applications

Distributed GNEP Algorithms without Multiplier Sharing and Applications to Multi-Robot Coordination and Contextual Bandit-Based Active Learning

DGX agent

arXiv:2606.00759v1 Announce Type: new Abstract: Recent advances in artificial intelligence have expanded the focus from classical optimization to include equilibrium analysis in noncooperative games.

applicationsarxiv-cs-lg
2 Jun 2026
Safety

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs

DGX agent

arXiv:2606.01072v1 Announce Type: cross Abstract: Imitation learning enables robots to learn how to execute tasks via observation. However, real-world environments like homes and offices are often sev

safetyarxiv-cs-cv
2 Jun 2026
Safety

From Cues to Horizons: Dynamic Risk Horizon Profiling for Trajectory Prediction

DGX agent

arXiv:2606.00857v1 Announce Type: cross Abstract: Accurate and reliable vehicle trajectory prediction is essential for safe autonomous driving. Recent studies have incorporated safety risk into trajec

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

HAIM: Human-AI Music Datasets for AI Music Production Tracking Benchmark

DGX agent

arXiv:2606.01686v1 Announce Type: cross Abstract: As generative platforms such as Suno and Udio reach human-grade audio quality, the scope of AI's utility has expanded across the entire music producti

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning

DGX agent

arXiv:2509.12263v3 Announce Type: replace Abstract: Large multimodal models (LMMs) encode physical laws observed during training, such as momentum conservation, as parametric knowledge. It allows LMMs

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

InstructSAM: Segment Any Instance with Any Instructions

DGX agent

arXiv:2605.26102v2 Announce Type: replace Abstract: In this paper, we introduce InstructSAM, a unified and streamlined framework designed for multi-instance segmentation under arbitrary instructions.

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Interpretable Policy Distillation for Power Grid Topology Control

DGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions

DGX agent

arXiv:2606.01703v1 Announce Type: cross Abstract: We address the challenge of generating high-fidelity, long-form soundtracks that remain coherent across scene transitions. Existing AI music systems a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Measuring and Mitigating Bias in Code Generated by Large Language Models

DGX agent

arXiv:2606.00049v1 Announce Type: cross Abstract: Large language models (LLMs) are widely recognised for their applications in natural language generation and are increasingly used for code generation

model-releasesarxiv-cs-ai
2 Jun 2026
Hardware

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major…

DGX agent

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse attention, paged MSA decode, optimized index scoring

hardwaretogether-ai--x
2 Jun 2026
Research

MOSS-Audio Technical Report

DGX agent

arXiv:2606.01802v1 Announce Type: cross Abstract: MOSS-Audio is a unified audio-language model for speech, environmental sound, and music understanding, supporting audio captioning, time-aware questio

researcharxiv-cs-ai
2 Jun 2026
← Previous
1…343344345346347…370
Next →