AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Safety

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

DGX agent

On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental

safetyapple-ml-research
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma

DGX agent

arXiv:2607.06987v1 Announce Type: new Abstract: Reinforcement learning (RL) has become the standard paradigm for enhancing the complex reasoning capabilities of large language models (LLMs). To achiev

safetyarxiv-cs-lg
9 Jul 2026
Model Releases

VCDP: Variation-Conditioned Distributional Proxy Learning for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2607.07416v1 Announce Type: new Abstract: Semi-supervised 3D medical image segmentation reduces the need for dense voxel-level annotations by exploiting unlabeled volumes. Although existing meth

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool cal…

DGX agent

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool call, and subagent run shows up as an inspectable trace. Setup

model-releasesharrison-chase--x
9 Jul 2026
Tools

We measured cost per task against GLM 5.2 as the baseline. On WANDR, GLM 5.2 + advisor runs at 2.1x versus Opus at 6.1x, averaging roughly h…

DGX agent

Perplexity conducted a cost efficiency comparison of language models, using GLM 5.2 as the baseline metric for cost per task. Results showed GLM 5.2 with an advisor achieved 2.1x cost efficiency on th

toolsperplexity--x
9 Jul 2026
Research

When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs

DGX agent

arXiv:2607.07395v1 Announce Type: cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-sh

researcharxiv-cs-ai
9 Jul 2026
Model Releases

Wordle 1,845 5/6 ⬛⬛⬛⬛🟨 ⬛⬛🟨⬛🟨 🟨⬛🟨🟨⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result (puzzle #1,845) played by Anthropic, showing the progression of guesses through color-coded tile feedback until reaching the correct five-letter word solution

model-releasesanthropic--x
9 Jul 2026
Model Releases

1/3 Best-of-N leaves $$ on the table by not accounting for variance in task difficulty. We built budget-aware execution: turn the dial on co…

DGX agent

1/3 Best-of-N leaves $$ on the table by not accounting for variance in task difficulty. We built budget-aware execution: turn the dial on compute or speed, while keeping quality constant, to save cost

model-releasesai21-labs--x
8 Jul 2026
Model Releases

2/3 By building a reliable early stopping mechanism, we could apply cascading (save up to 44% compute by not running unnecessary rollouts ) …

DGX agent

2/3 By building a reliable early stopping mechanism, we could apply cascading (save up to 44% compute by not running unnecessary rollouts ) or parallel execution (up to 25% faster by sparing wait time

model-releasesai21-labs--x
8 Jul 2026
Agents

A Three-Layer Framework for AI in Scientific Discovery

DGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Arkenstone Defense launches with $35M to help startups sell to the Pentagon

DGX agent

Federal contracting software startup Arkenstone Defense Inc. formally launched today with 35 million in new funding to take on the compliance, security and workforce setup that keeps many commercial t

model-releasessiliconangle
8 Jul 2026
Model Releases

Artificial Analysis assessment

DGX agent

Artificial Analysis assessment SpaceXAI just released Grok 4.5, and it ranks #4 on GDPval-AA v2 with an Elo of 1543 - behind only the latest Claude releases from Anthropic on real-world agentic knowle

model-releaseselon-musk--x
8 Jul 2026
Local Ai

b9908

DGX agent

b9908 is a build-tagged release from llama.cpp , the open-source C/C++ inference engine for large language models. llama.cpp is an open-source software library that performs inference on various large

local-aillama-cpp-releases
8 Jul 2026
Local Ai

b9932

DGX agent

B9932 is a continuous build-tagged release from the llama.cpp project , an open-source C/C++ inference engine for running large language models locally. llama.cpp performs inference on various large l

local-aillama-cpp-releases
8 Jul 2026
Agents

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

DGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Box Agent uses @LangChain’s Deep Agents harness to bring specialized agents into the enterprise content platform.w @NVIDIA + LangChain’s wor…

DGX agent

Box Agent uses @LangChain’s Deep Agents harness to bring specialized agents into the enterprise content platform.w @NVIDIA + LangChain’s work with Nemotron 3 Ultra reinforces where AI is headed: open,

model-releasesharrison-chase--x
8 Jul 2026
Model Releases

Breaking Structural Isolation: Scalable Graph Clustering via Community-Aware Sampling and Structural Entropy

DGX agent

arXiv:2607.05469v1 Announce Type: cross Abstract: Unsupervised graph clustering is a fundamental technique for uncovering underlying semantic patterns in large-scale networks. Although Graph Contrasti

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Bridging Diffusion Pruning and Step Distillation with Teacher-Aligned Repair

DGX agent

arXiv:2607.06335v1 Announce Type: new Abstract: Diffusion models generate high-quality images, but their inference cost comes from two sources: large denoising networks and repeated denoising steps. E

safetyarxiv-cs-cv
8 Jul 2026
Model Releases

Building and connecting a production-ready ecommerce MCP server using Amazon Bedrock AgentCore and Mistral AI Studio

DGX agent

In this post, you build and connect that server end to end. You will implement MCP tools, set up two-layer JSON Web Token (JWT) authentication, deploy with AWS Cloud Development Kit (AWS CDK), and con

model-releasesaws-ml-blog
8 Jul 2026
Model Releases

Classification of Financial Data Using Quantum Support Vector Machine

DGX agent

arXiv:2412.10860v2 Announce Type: replace-cross Abstract: Quantum Support Vector Machine is a kernel-based approach to classification problems. We study the applicability of quantum kernels to financi

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP!

DGX agent

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP! Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with

model-releasesharrison-chase--x
8 Jul 2026
Research

Decision-Focused Scenario Generation and Selection for Efficient and Robust Grid Dispatch

DGX agent

arXiv:2607.05830v1 Announce Type: cross Abstract: The increasing uncertainty from flexible demand and renewable generation has made distributionally robust optimization (DRO) an important tool for rob

researcharxiv-cs-ai
8 Jul 2026
Model Releases

Deep Neural Variation Spaces: A Unifying Perspective on Depth and Complexity

DGX agent

arXiv:2607.05546v1 Announce Type: cross Abstract: We develop a unified function space theory of deep fully connected neural networks. Functions in our spaces are defined recursively as ell^1-bounded l

model-releasesarxiv-cs-lg
8 Jul 2026
Hardware

DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

DGX agent

arXiv:2607.06523v1 Announce Type: new Abstract: Long-context language model inference is increasingly limited by the memory bandwidth and capacity required to store key-value caches, yet existing comp

hardwarearxiv-cs-ai
8 Jul 2026
Model Releases

Didn't even have time to write up thoughts on my early access to the new GPT voice, but it is really good & much closer to the science ficti…

DGX agent

Didn't even have time to write up thoughts on my early access to the new GPT voice, but it is really good & much closer to the science fiction experience of talking to AI (in part because it is much s

model-releasesethan-mollick--x
8 Jul 2026
Model Releases

Discovering Frequent Closed Embedded Sub-DAGs in Spatio-Temporal Event Data

DGX agent

arXiv:2607.05995v1 Announce Type: cross Abstract: We propose a novel approach to mine patterns in spatio-temporal event data based on discovering frequent closed embedded sub-Directed Acyclic Graphs (

model-releasesarxiv-cs-lg
8 Jul 2026
Safety

Domain-Adaptive Climate Downscaling Under Temporal Distribution Shift

DGX agent

arXiv:2607.05645v1 Announce Type: new Abstract: Deep-learning-based climate downscaling aims to learn relationships from historical low-resolution (LR) and high-resolution (HR) climate data to generat

safetyarxiv-cs-lg
8 Jul 2026
Applications

Drift Happens: An Empirical Study of Neural Architecture Robustness to Temporal Distribution Shift

DGX agent

arXiv:2607.05908v1 Announce Type: new Abstract: Real-world data distributions evolve over time, inducing temporal distribution shift that can substantially degrade the reliability of deployed machine

applicationsarxiv-cs-lg
8 Jul 2026
Research

Early Language Learning via Spreading Activation and Category Exploration in Complex Networks

DGX agent

arXiv:2607.06258v1 Announce Type: new Abstract: Is word acquisition in children uneven with respect to semantic and lexical categories? To answer this question, we model early language learning as a s

researcharxiv-cs-cl
8 Jul 2026
Local Ai

Energy-Efficient GPU DVFS for Fine-Tuning of SLMs on Resource-constrained Embedded Devices

DGX agent

arXiv:2607.05933v1 Announce Type: cross Abstract: Dynamic Voltage Frequency Scaling (DVFS) on resource-constrained embedded GPU platforms is essential for energy-efficient small language model (SLM) f

local-aiarxiv-cs-lg
8 Jul 2026
Model Releases

Enhanced Seam Segmentation for Automated Welding Robot in Construction Through Transfer Learning: Addressing Limitations of Bilateral Segmentation Network

DGX agent

arXiv:2607.06150v1 Announce Type: new Abstract: Reliable seam segmentation is essential for autonomous robotic welding in construction, where harsh illumination, specular reflections, and thin weld ge

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the i…

DGX agent

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the idea behind our work with LangChain. Read more: https://nvda.

model-releasessonya-huang--x
8 Jul 2026
Research

Estimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs

DGX agent

arXiv:2607.06327v1 Announce Type: cross Abstract: Uncertainty estimation (UE) enables LLM-powered systems to recognize when to abstain, yet existing research has predominantly focused on English. We p

researcharxiv-cs-ai
8 Jul 2026
Model Releases

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @D…

DGX agent

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @DeepInfra @togethercompute Introducing the NemoClaw Deep Agen

model-releasesharrison-chase--x
8 Jul 2026
Model Releases

Fable feels very different than Opus. GPT-5.6 feels like a part of the GPT-5 family. I developed a very complex set of heuristics about when…

DGX agent

Fable feels very different than Opus. GPT-5.6 feels like a part of the GPT-5 family. I developed a very complex set of heuristics about when to use which. Fable was often “smarter” but was also too se

model-releasesethan-mollick--x
8 Jul 2026
Model Releases

Federated Physics-Grounded Reinforcement Learning for Distributed Stability Control in Smart Grids

DGX agent

arXiv:2607.05553v1 Announce Type: new Abstract: Transient stability control in smart grids requires rapid post-fault damping of generator frequency and rotor angle deviations to prevent cascading fail

model-releasesarxiv-cs-lg
8 Jul 2026
Research

Few-Medoids: An Embarrassingly Simple Coreset Selection Method for Few-Shot Knowledge Distillation

DGX agent

arXiv:2607.05891v1 Announce Type: cross Abstract: Coreset selection aims to identify a small and highly representative subset of a massive dataset for efficient model training. The problem remains cha

researcharxiv-cs-ai
8 Jul 2026
Model Releases

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games

DGX agent

arXiv:2607.06514v1 Announce Type: new Abstract: We present FootsiesGym, an open-source environment for learning in a non-trivial two-player, zero-sum, imperfect-information game. Built on HiFight's mi

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

FORGE: Towards Functional Tool-Use Generalization via Keypoint Trajectory Reasoning

DGX agent

arXiv:2607.05780v1 Announce Type: cross Abstract: While humans readily repurpose a book, a stone, or a shoe to drive a nail, robots trained on specific tools fail to transfer the same function to nove

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia (Radhika Rajkumar/ZDNET)

DGX agent

Radhika Rajkumar / ZDNET: Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia — ZDNET's key

model-releasestechmeme
8 Jul 2026
Safety

From Application-Layer Simulation to Native Meta-Architecture: Structural Tension as an Endogenous Driver for Heterogeneous AI Evolution

DGX agent

arXiv:2607.06269v1 Announce Type: new Abstract: Current large language models (LLMs) are fundamentally stateless: their behavior is fully determined by input at inference time, and any higher-order co

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

GEM-Occ: From Visual Geometry Evidence to Embodied Semantic Occupancy Memory

DGX agent

arXiv:2607.05543v1 Announce Type: cross Abstract: Semantic occupancy provides a structured spatial memory for embodied indoor agents by jointly representing occupied regions, observed free space, unkn

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

GLM-5.2 on Ollama's cloud just got more capacity in US & Europe! Ollama's cloud for GLM 5.2 consistently delivers between 80 to 120 output t…

DGX agent

GLM-5.2 on Ollama's cloud just got more capacity in US & Europe! Ollama's cloud for GLM 5.2 consistently delivers between 80 to 120 output tokens per second, even during peak hours, compared to 30 to

model-releasesollama--x
8 Jul 2026
Model Releases

gpt-5.6 sol isnt the only thinking launching thursday we're also releasing a big update to OpenWiki (auto create wikis of code bases... and …

DGX agent

gpt-5.6 sol isnt the only thinking launching thursday we're also releasing a big update to OpenWiki (auto create wikis of code bases... and more?) we're also going live with a webinar to talk all thin

model-releasesharrison-chase--x
8 Jul 2026
Model Releases

GPT-live (next-generation voice) launches today in ChatGPT. it feels magical and 'real'. i have always preferred typing to talking to an AI,…

DGX agent

OpenAI launched GPT-Live, a next-generation voice feature for ChatGPT that enables more natural, real-time voice interactions with the AI. The feature represents an advancement in conversational AI ca

model-releasessam-altman--x
8 Jul 2026
Tutorials

Graph Convolutional Attention: A Spectral Perspective on Graph Denoising and Diffusion

DGX agent

arXiv:2607.06546v1 Announce Type: cross Abstract: Denoising graphs is a fundamental problem in graph learning and the core operation of graph diffusion models. Attention-based architectures like graph

tutorialsarxiv-cs-ai
8 Jul 2026
Model Releases

GraphAllocBench: A Flexible Benchmark for Preference-Conditioned Multi-Objective Policy Learning

DGX agent

arXiv:2601.20753v4 Announce Type: replace Abstract: Preference-Conditioned Policy Learning (PCPL) in Multi-Objective Reinforcement Learning (MORL) approximates diverse Pareto-optimal solutions by cond

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benc…

DGX agent

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benchmark scores keep climbing, yet the same agent failures resu

model-releasesdair-ai--x
8 Jul 2026
← Previous
1…789790791792793…1326
Next →