AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
9 Jul 2026

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

SafetyDGX agent

On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental

UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma

SafetyDGX agent

arXiv:2607.06987v1 Announce Type: new Abstract: Reinforcement learning (RL) has become the standard paradigm for enhancing the complex reasoning capabilities of large language models (LLMs). To achiev

VCDP: Variation-Conditioned Distributional Proxy Learning for Semi-Supervised Medical Image Segmentation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.07416v1 Announce Type: new Abstract: Semi-supervised 3D medical image segmentation reduces the need for dense voxel-level annotations by exploiting unlabeled volumes. Although existing meth

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool cal…

Model ReleasesDGX agent

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool call, and subagent run shows up as an inspectable trace. Setup

We measured cost per task against GLM 5.2 as the baseline. On WANDR, GLM 5.2 + advisor runs at 2.1x versus Opus at 6.1x, averaging roughly h…

ToolsDGX agent

Perplexity conducted a cost efficiency comparison of language models, using GLM 5.2 as the baseline metric for cost per task. Results showed GLM 5.2 with an advisor achieved 2.1x cost efficiency on th

When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs

ResearchDGX agent

arXiv:2607.07395v1 Announce Type: cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-sh

Wordle 1,845 5/6 ⬛⬛⬛⬛🟨 ⬛⬛🟨⬛🟨 🟨⬛🟨🟨⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result (puzzle #1,845) played by Anthropic, showing the progression of guesses through color-coded tile feedback until reaching the correct five-letter word solution

8 Jul 2026

1/3 Best-of-N leaves $$ on the table by not accounting for variance in task difficulty. We built budget-aware execution: turn the dial on co…

Model ReleasesDGX agent

1/3 Best-of-N leaves $$ on the table by not accounting for variance in task difficulty. We built budget-aware execution: turn the dial on compute or speed, while keeping quality constant, to save cost

2/3 By building a reliable early stopping mechanism, we could apply cascading (save up to 44% compute by not running unnecessary rollouts ) …

Model ReleasesDGX agent

2/3 By building a reliable early stopping mechanism, we could apply cascading (save up to 44% compute by not running unnecessary rollouts ) or parallel execution (up to 25% faster by sparing wait time

A Three-Layer Framework for AI in Scientific Discovery

AgentsDGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

Arkenstone Defense launches with $35M to help startups sell to the Pentagon

Model ReleasesDGX agent

Federal contracting software startup Arkenstone Defense Inc. formally launched today with 35 million in new funding to take on the compliance, security and workforce setup that keeps many commercial t

Artificial Analysis assessment

Model ReleasesDGX agent

Artificial Analysis assessment SpaceXAI just released Grok 4.5, and it ranks #4 on GDPval-AA v2 with an Elo of 1543 - behind only the latest Claude releases from Anthropic on real-world agentic knowle

b9908

Local AiDGX agent

b9908 is a build-tagged release from llama.cpp , the open-source C/C++ inference engine for large language models. llama.cpp is an open-source software library that performs inference on various large

b9932

Local AiDGX agent

B9932 is a continuous build-tagged release from the llama.cpp project , an open-source C/C++ inference engine for running large language models locally. llama.cpp performs inference on various large l

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

Box Agent uses @LangChain’s Deep Agents harness to bring specialized agents into the enterprise content platform.w @NVIDIA + LangChain’s wor…

Model ReleasesDGX agent

Box Agent uses @LangChain’s Deep Agents harness to bring specialized agents into the enterprise content platform.w @NVIDIA + LangChain’s work with Nemotron 3 Ultra reinforces where AI is headed: open,

Breaking Structural Isolation: Scalable Graph Clustering via Community-Aware Sampling and Structural Entropy

Model ReleasesDGX agent

arXiv:2607.05469v1 Announce Type: cross Abstract: Unsupervised graph clustering is a fundamental technique for uncovering underlying semantic patterns in large-scale networks. Although Graph Contrasti

Bridging Diffusion Pruning and Step Distillation with Teacher-Aligned Repair

SafetyDGX agent

arXiv:2607.06335v1 Announce Type: new Abstract: Diffusion models generate high-quality images, but their inference cost comes from two sources: large denoising networks and repeated denoising steps. E

Building and connecting a production-ready ecommerce MCP server using Amazon Bedrock AgentCore and Mistral AI Studio

Model ReleasesDGX agent

In this post, you build and connect that server end to end. You will implement MCP tools, set up two-layer JSON Web Token (JWT) authentication, deploy with AWS Cloud Development Kit (AWS CDK), and con

Classification of Financial Data Using Quantum Support Vector Machine

Model ReleasesDGX agent

arXiv:2412.10860v2 Announce Type: replace-cross Abstract: Quantum Support Vector Machine is a kernel-based approach to classification problems. We study the applicability of quantum kernels to financi

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP!

Model ReleasesDGX agent

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP! Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with

Decision-Focused Scenario Generation and Selection for Efficient and Robust Grid Dispatch

ResearchDGX agent

arXiv:2607.05830v1 Announce Type: cross Abstract: The increasing uncertainty from flexible demand and renewable generation has made distributionally robust optimization (DRO) an important tool for rob

Deep Neural Variation Spaces: A Unifying Perspective on Depth and Complexity

Model ReleasesDGX agent

arXiv:2607.05546v1 Announce Type: cross Abstract: We develop a unified function space theory of deep fully connected neural networks. Functions in our spaces are defined recursively as ell^1-bounded l

DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

HardwareDGX agent

arXiv:2607.06523v1 Announce Type: new Abstract: Long-context language model inference is increasingly limited by the memory bandwidth and capacity required to store key-value caches, yet existing comp

Didn't even have time to write up thoughts on my early access to the new GPT voice, but it is really good & much closer to the science ficti…

Model ReleasesDGX agent

Didn't even have time to write up thoughts on my early access to the new GPT voice, but it is really good & much closer to the science fiction experience of talking to AI (in part because it is much s

Discovering Frequent Closed Embedded Sub-DAGs in Spatio-Temporal Event Data

Model ReleasesDGX agent

arXiv:2607.05995v1 Announce Type: cross Abstract: We propose a novel approach to mine patterns in spatio-temporal event data based on discovering frequent closed embedded sub-Directed Acyclic Graphs (

Domain-Adaptive Climate Downscaling Under Temporal Distribution Shift

SafetyDGX agent

arXiv:2607.05645v1 Announce Type: new Abstract: Deep-learning-based climate downscaling aims to learn relationships from historical low-resolution (LR) and high-resolution (HR) climate data to generat

Drift Happens: An Empirical Study of Neural Architecture Robustness to Temporal Distribution Shift

ApplicationsDGX agent

arXiv:2607.05908v1 Announce Type: new Abstract: Real-world data distributions evolve over time, inducing temporal distribution shift that can substantially degrade the reliability of deployed machine

Early Language Learning via Spreading Activation and Category Exploration in Complex Networks

ResearchDGX agent

arXiv:2607.06258v1 Announce Type: new Abstract: Is word acquisition in children uneven with respect to semantic and lexical categories? To answer this question, we model early language learning as a s

Energy-Efficient GPU DVFS for Fine-Tuning of SLMs on Resource-constrained Embedded Devices

Local AiDGX agent

arXiv:2607.05933v1 Announce Type: cross Abstract: Dynamic Voltage Frequency Scaling (DVFS) on resource-constrained embedded GPU platforms is essential for energy-efficient small language model (SLM) f

Enhanced Seam Segmentation for Automated Welding Robot in Construction Through Transfer Learning: Addressing Limitations of Bilateral Segmentation Network

Model ReleasesDGX agent

arXiv:2607.06150v1 Announce Type: new Abstract: Reliable seam segmentation is essential for autonomous robotic welding in construction, where harsh illumination, specular reflections, and thin weld ge

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the i…

Model ReleasesDGX agent

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the idea behind our work with LangChain. Read more: https://nvda.

Estimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs

ResearchDGX agent

arXiv:2607.06327v1 Announce Type: cross Abstract: Uncertainty estimation (UE) enables LLM-powered systems to recognize when to abstain, yet existing research has predominantly focused on English. We p

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @D…

Model ReleasesDGX agent

Excited to team up with partners across the AI infrastructure + enterprise ecosystem. @EY_US @baseten @FireworksAI_HQ @nebiusai @CrusoeAI @DeepInfra @togethercompute Introducing the NemoClaw Deep Agen

Fable feels very different than Opus. GPT-5.6 feels like a part of the GPT-5 family. I developed a very complex set of heuristics about when…

Model ReleasesDGX agent

Fable feels very different than Opus. GPT-5.6 feels like a part of the GPT-5 family. I developed a very complex set of heuristics about when to use which. Fable was often “smarter” but was also too se

Federated Physics-Grounded Reinforcement Learning for Distributed Stability Control in Smart Grids

Model ReleasesDGX agent

arXiv:2607.05553v1 Announce Type: new Abstract: Transient stability control in smart grids requires rapid post-fault damping of generator frequency and rotor angle deviations to prevent cascading fail

Few-Medoids: An Embarrassingly Simple Coreset Selection Method for Few-Shot Knowledge Distillation

ResearchDGX agent

arXiv:2607.05891v1 Announce Type: cross Abstract: Coreset selection aims to identify a small and highly representative subset of a massive dataset for efficient model training. The problem remains cha

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games

Model ReleasesDGX agent

arXiv:2607.06514v1 Announce Type: new Abstract: We present FootsiesGym, an open-source environment for learning in a non-trivial two-player, zero-sum, imperfect-information game. Built on HiFight's mi

FORGE: Towards Functional Tool-Use Generalization via Keypoint Trajectory Reasoning

Model ReleasesDGX agent

arXiv:2607.05780v1 Announce Type: cross Abstract: While humans readily repurpose a book, a stone, or a shoe to drive a nail, robots trained on specific tools fail to transfer the same function to nove

Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia (Radhika Rajkumar/ZDNET)

Model ReleasesDGX agent

Radhika Rajkumar / ZDNET: Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia — ZDNET's key

From Application-Layer Simulation to Native Meta-Architecture: Structural Tension as an Endogenous Driver for Heterogeneous AI Evolution

SafetyDGX agent

arXiv:2607.06269v1 Announce Type: new Abstract: Current large language models (LLMs) are fundamentally stateless: their behavior is fully determined by input at inference time, and any higher-order co

GEM-Occ: From Visual Geometry Evidence to Embodied Semantic Occupancy Memory

Model ReleasesDGX agent

arXiv:2607.05543v1 Announce Type: cross Abstract: Semantic occupancy provides a structured spatial memory for embodied indoor agents by jointly representing occupied regions, observed free space, unkn

GLM-5.2 on Ollama's cloud just got more capacity in US & Europe! Ollama's cloud for GLM 5.2 consistently delivers between 80 to 120 output t…

Model ReleasesDGX agent

GLM-5.2 on Ollama's cloud just got more capacity in US & Europe! Ollama's cloud for GLM 5.2 consistently delivers between 80 to 120 output tokens per second, even during peak hours, compared to 30 to

gpt-5.6 sol isnt the only thinking launching thursday we're also releasing a big update to OpenWiki (auto create wikis of code bases... and …

Model ReleasesDGX agent

gpt-5.6 sol isnt the only thinking launching thursday we're also releasing a big update to OpenWiki (auto create wikis of code bases... and more?) we're also going live with a webinar to talk all thin

GPT-live (next-generation voice) launches today in ChatGPT. it feels magical and 'real'. i have always preferred typing to talking to an AI,…

Model ReleasesDGX agent

OpenAI launched GPT-Live, a next-generation voice feature for ChatGPT that enables more natural, real-time voice interactions with the AI. The feature represents an advancement in conversational AI ca

Graph Convolutional Attention: A Spectral Perspective on Graph Denoising and Diffusion

TutorialsDGX agent

arXiv:2607.06546v1 Announce Type: cross Abstract: Denoising graphs is a fundamental problem in graph learning and the core operation of graph diffusion models. Attention-based architectures like graph

GraphAllocBench: A Flexible Benchmark for Preference-Conditioned Multi-Objective Policy Learning

Model ReleasesDGX agent

arXiv:2601.20753v4 Announce Type: replace Abstract: Preference-Conditioned Policy Learning (PCPL) in Multi-Objective Reinforcement Learning (MORL) approximates diverse Pareto-optimal solutions by cond

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benc…

Model ReleasesDGX agent

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benchmark scores keep climbing, yet the same agent failures resu

Grok 4.5 brings frontier performance across coding and knowledge work

Model ReleasesDGX agent

Grok 4.5 brings frontier performance across coding and knowledge work SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and O

Grok 4.5 context window will upgrade to 1M probably by next week

Model ReleasesDGX agent

Grok 4.5 context window will upgrade to 1M probably by next week SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and Opus 4

Grok 4.5 is built for real-world engineering. It excels in large codebases and handles long-running tasks that span multiple repositories, h…

ApplicationsDGX agent

Grok 4.5 is an AI model designed for real-world engineering tasks, with particular strengths in handling large codebases and managing complex, long-running tasks that span multiple repositories. The m

Grok 4.5 is now officially available in Grok Build......try it now

IndustryDGX agent

Grok 4.5, an updated version of xAI's conversational AI model, has been officially released and is now accessible through Grok Build, xAI's development platform. The announcement encourages users to t

Grok has always been very strong on law

Model ReleasesDGX agent

Elon Musk stated that Grok, the AI assistant developed by xAI, has demonstrated strong capabilities in understanding and applying legal concepts and law-related reasoning. The statement suggests Grok

Guess what we added someone else to this webinar: @jeffreyhuber, ceo and cofounder of chroma What does a vector database think about the con…

Model ReleasesDGX agent

Guess what we added someone else to this webinar: @jeffreyhuber, ceo and cofounder of chroma What does a vector database think about the concept of “wikis”? Come find out! It’s turning into a party gp

Heckman-Corrected Epistemic Uncertainty: Selection on Unobservables Defeats Importance Weighting

SafetyDGX agent

arXiv:2607.05806v1 Announce Type: new Abstract: Training data for machine learning is routinely collected by a selection process the model never sees: loans are observed only when granted, outcomes on

Hilti-Trimble-Oxford Dataset: 360 Visual-Inertial Benchmark with Floor Plan Priors for SLAM and Localization

Model ReleasesDGX agent

arXiv:2607.06464v1 Announce Type: new Abstract: Automated progress monitoring on construction sites is an active area of research and development. Robot and human-carried mapping systems have been dev

I am more excited to try this pattern instead. Same Executor-Advisor setup but with GPT-5.6 as the executor and Fable 5 as the advisor. It a…

Model ReleasesDGX agent

I am more excited to try this pattern instead. Same Executor-Advisor setup but with GPT-5.6 as the executor and Fable 5 as the advisor. It already works wonderfully using GPT-5.5, so I think 5.6 shoul

I cannot put into words how stoked I am on this It's not an understatement to say that this blueprint could represent the future of enterpri…

Model ReleasesDGX agent

I cannot put into words how stoked I am on this It's not an understatement to say that this blueprint could represent the future of enterprise inference Introducing the NemoClaw Deep Agents Blueprint,

Imbalance-Robust and Sampling-Efficient Continuous Conditional GANs via Adaptive Vicinal Learning and Auxiliary Regularization

Local AiDGX agent

arXiv:2508.01725v5 Announce Type: replace-cross Abstract: Recent advances in continuous conditional generative modeling, including Continuous conditional Generative Adversarial Network (CcGAN) and Con

IMR: Iterative Mode-World Weighted Regression for Multi-Agent Trajectory Prediction

Model ReleasesDGX agent

arXiv:2607.05705v1 Announce Type: cross Abstract: Multi-agent motion prediction is essential for automated vehicles to understand the intentions of surrounding vehicles. However, previous prediction-b

← Previous
1…629630631632633…1059
Next →