AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,413 results
11 Aug 2026

💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to c…

Model ReleasesDGX agent

💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to choose the right model for the right task. As part of this, we

Tools to Explain Neural Networks for Power System Dynamics

ResearchDGX agent

arXiv:2608.08048v1 Announce Type: cross Abstract: This paper presents, for the first time in power systems literature to our knowledge, analytical tools to explain the training performance of machine

v0.32.8

Model ReleasesDGX agent

Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

We built the Agentic World Cup - LLMs that compete in 1v1 Soccer. [P]

AgentsDGX agent

Hey everyone - we've been building something particularly relevant to ML at large - The Agentic World Cup - a platform where Agents compete in sports. As you know, today's Agents can code, do math, an

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

Local AiDGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

10 Aug 2026

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

Local AiDGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

Model ReleasesDGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

ArchEGraph: A Large-Scale Graph Dataset for Geometry-Topology-Physics Aligned Building Energy Modeling

Model ReleasesDGX agent

arXiv:2608.06772v1 Announce Type: new Abstract: Accurate estimation of building energy use is essential for achieving carbon neutral and sustainable buildings. To better understand the influence of de

Best Local LLMs - August 2026

Local AiDGX agent

Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the closed frontier, Opus level models on non-insane hardwa

Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation

Model ReleasesDGX agent

arXiv:2608.06751v1 Announce Type: cross Abstract: Artist-grounded image generation requires more than appending an artist name to a prompt. Image models often respond to artist names through canonical

Comparing how Cline, Kilo, and Qwen Code handle long-task context/state (and why context loops keep happening)

Model ReleasesDGX agent

I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside the conversation that gets reinjected on a c

Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education

ApplicationsDGX agent

arXiv:2608.07364v1 Announce Type: new Abstract: Contribution: This paper presents a six-phase AI-assisted instructional design architecture based on the Curriculum as Code paradigm, integrating Genera

Density-Functional Excited-State Gradients and Nonadiabatic Couplings on a Consumer GPU from a Contraction-DAG

SafetyDGX agent

arXiv:2608.06536v1 Announce Type: cross Abstract: Nonadiabatic dynamics needs an excited-state gradient and an interstate nonadiabatic coupling matrix element (NACME) at every nuclear geometry, and a

Dual-Node NVIDIA DGX Spark over Tailscale: A Remote-Access Testbed for Distributed LLM Training and Cyber-Threat-Intelligence Fine-Tuning

Local AiDGX agent

arXiv:2608.07226v1 Announce Type: cross Abstract: Compact AI systems make local language-model experimentation increasingly accessible, yet practical evidence for multi-node training on desktop-class

DynaCrys: Crystal Generation with Dynamic Space-Group Diffusion

ResearchDGX agent

arXiv:2608.07401v1 Announce Type: cross Abstract: The search for new crystalline materials spans an enormous compositional and structural space. Generating candidates in this space requires jointly mo

Evaluating Useful Surrogate Models for Configuration Tuning Beyond Accuracy: A Fitness Landscape Analysis Perspective

ResearchDGX agent

arXiv:2509.21945v2 Announce Type: replace-cross Abstract: To efficiently tune configuration for better software system performance (e.g., latency) at the deployment and maintenance stage, many tuners

Georeferencing Non-Gazetteered Place Names using Biological Specimen Records

Model ReleasesDGX agent

arXiv:2608.06884v1 Announce Type: cross Abstract: Biological specimen records collected by natural history institutions constitute a rich source of temporal geographic knowledge, capturing biodiversit

Stockmark-Nemotron-3-Nano-Omni-JapanDocReader: Structured Document Parsing via Capability Injection and Forgetting Control

Model ReleasesDGX agent

arXiv:2608.06758v1 Announce Type: new Abstract: We present Stockmark-Nemotron-3-Nano-Omni-JapanDocReader, a Japanese document understanding model built from Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

Model ReleasesDGX agent

arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and simulation by synthesizing realistic ins

Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools

ApplicationsDGX agent

arXiv:2608.07446v1 Announce Type: cross Abstract: Rapid adoption of large language models (LLMs) in enterprise settings has introduced operational, security, and governance risks. As generative AI app

Ultra-High Interactivity on NVIDIA GPUs? - TileRT InferenceX

HardwareDGX agent

The TileRT InferenceX article (Aug 10 2026) examines whether the TileRT software stack on NVIDIA GPUs can compete with dedicated inference systems such as Cerebras, Groq LPUs and SambaNova for ultra‑h

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

Model ReleasesDGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

9 Aug 2026

I Turned My Underused Gaming Laptop Into a Local AI Workstation

Local AiDGX agent

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

not wrong!

AgentsDGX agent

In a recent Twitter exchange, Luca Ambrogioni stated that large‑language models (LLMs) likely possess fundamental limitations that may be obscured by progress in reasoning and agentic pipelines, thoug

The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotten 3x more …

Model ReleasesDGX agent

The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotten 3x more expensive while flatlining on visual recognition across comp

We compared how far the same budget goes with DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. Two DeepSeek V4 Flash attempts solved MORE task…

Model ReleasesDGX agent

We compared how far the same budget goes with DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. Two DeepSeek V4 Flash attempts solved MORE tasks than one Luna attempt at roughly one-third the cost. Media

8 Aug 2026

Anthropic says auto mode will be the default in Claude Code for Pro, Max, Team plans, starting on Aug. 14, claiming it's good enough at catching harmful actions (Simon Willison/Simon Willison's Weblog)

Model ReleasesDGX agent

Simon Willison / Simon Willison's Weblog: Anthropic says auto mode will be the default in Claude Code for Pro, Max, Team plans, starting on Aug. 14, claiming it's good enough at catching harmful actio

Auto mode is now the default in Claude Code for Pro, Max, and Team plans

Model ReleasesDGX agent

Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses

Showoff Saturday: Local 4x 6000 Pro (multi-year progression)

Model ReleasesDGX agent

Not the biggest or shiniest, but it's mine From gaming machine inference on the original llama models, to a 4x RTX 6000 Pro Max Q + 4x 3090s local AI cluster. Pictures are in reverse chronological ord

We analyzed DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. A DeepSeek-first cascade with test-suite verification solved MORE tasks than Luna…

Model ReleasesDGX agent

Researchers from TogetherAI analyzed DeepSeek V4 Flash and GPT‑5.6 Luna on the DeepSWE benchmark. The study found that employing a DeepSeek‑first cascade with test‑suite verification solved more tasks

7 Aug 2026

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

Model ReleasesDGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals

SafetyDGX agent

arXiv:2608.05255v1 Announce Type: cross Abstract: Retail investors lack access to the kind of personalized, tax-aware portfolio management that institutional clients take for granted -- existing robo-

Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux (Marcus Mendes/9to5Mac)

Model ReleasesDGX agent

Marcus Mendes / 9to5Mac: Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux — Users running

Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with Negative Feedback

SafetyDGX agent

arXiv:2509.03206v2 Announce Type: replace-cross Abstract: Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that

Basically every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be signficantly higher with a bett…

Model ReleasesDGX agent

On August 7, 2026 Ethan Mollick tweeted that “every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be significantly higher with a better harness.” The commen

Computationally Efficient Collaborative Communication Via Regularity-Based Coarsening

AgentsDGX agent

arXiv:2608.05327v1 Announce Type: cross Abstract: Our results show that the existence of a short high-utility protocol already suffices for efficient communication. In particular, in a game with n pos

Counterfactual Analysis via Large Language Models

ResearchDGX agent

arXiv:2608.05367v1 Announce Type: new Abstract: Counterfactual analysis aims to predict potential outcomes under hypothetical scenarios, offering valuable insights for decision-making. This paper inve

EqDeepRx: Learning a Scalable and Interference Mitigating MIMO Receiver

ResearchDGX agent

arXiv:2602.11834v2 Announce Type: replace-cross Abstract: While machine learning (ML)-based receiver algorithms have received a great deal of attention in the recent literature, they often suffer from

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

FI-TW: An Open Train-Weather Dataset for Railway Delay Analysis in Finland

SafetyDGX agent

arXiv:2601.16592v2 Announce Type: replace-cross Abstract: Train delays result from complex interactions between operational, technical, and environmental factors. While weather impacts railway reliabi

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

SafetyDGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

Grounded Well-Condition Anomaly Detection on the Volve Field: Constructed Labels, a Baseline, and a Dual-Head Model

Model ReleasesDGX agent

arXiv:2608.05685v1 Announce Type: new Abstract: Most public benchmarks for machine-condition monitoring come from test rigs, where faults are induced on purpose and every event is known. Real producti

How Google Cloud detects, contains, and protects against emerging threats

Model ReleasesDGX agent

At Google Cloud, securing your data and business systems is our foundational commitment. We empower our customers with the tools, governance, and infrastructure needed to securely deploy workloads and

IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation

HardwareDGX agent

arXiv:2608.06088v1 Announce Type: new Abstract: Robotics simulators serve as a foundational infrastructure for embodied AI, facilitating safe and scalable robotic system development. NVIDIA Isaac Sim

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use

SafetyDGX agent

arXiv:2608.05738v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become the dominant recipe for generalist manipulation, yet they are almost universally trained by behavior clo

Learning Globally Reusable Skills for Coding Agents

AgentsDGX agent

arXiv:2608.06153v1 Announce Type: cross Abstract: Automated skill evolution enables Large Language Model (LLM) agents to continuously improve without expensive retraining. However, existing approaches

MIRA: A Modular Open-Source Micro-UAV for Indoor Research

SafetyDGX agent

arXiv:2607.11785v2 Announce Type: replace Abstract: Indoor robotics research increasingly uses micro-UAV platforms whose airframes, electronics, and control software are open to modification. Off-the-

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

Model ReleasesDGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

Ordered Diffusion for 3D Human Registration

SafetyDGX agent

arXiv:2608.05804v1 Announce Type: new Abstract: 3D human registration has historically been treated as a regression task, assuming a unique ground-truth alignment exists between the template and an in

Posture and Sustainment Optimization Under Adversarial Uncertainty

ResearchDGX agent

arXiv:2608.05256v1 Announce Type: new Abstract: Pre-commitment posture, the assignment of military assets to theater locations before conflict scenarios resolve, is a critical and formally unsolved pr

Predicting Social Media User Actions: A Hybrid Approach for Common and Rare Behavior Prediction on Bluesky

ResearchDGX agent

arXiv:2511.17241v2 Announce Type: replace Abstract: Understanding and predicting user behavior on social media platforms is crucial for content recommendation and platform design. While existing appro

Reinforcing Action Policies by Prophesying

TutorialsDGX agent

arXiv:2511.20633v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) policies excel in aligning language, perception, and robot control. However, most VLAs are trained purely by imitation,

TAU-Bench: From Anomaly Instance Tracking to Fine-Grained Video Anomaly Understanding

Model ReleasesDGX agent

arXiv:2608.05699v1 Announce Type: new Abstract: Humans understand anomalous events through a coherent perceptual process in which they identify the focal instance, follow its behavior as the event unf

Training a Conditioned Video Game Agent on a VLM Annotated Dataset

SafetyDGX agent

arXiv:2608.05954v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the gam

6 Aug 2026

A Unified Model for Cross-Domain Clone Detection via Model Merging

Model ReleasesDGX agent

arXiv:2608.04215v1 Announce Type: cross Abstract: The growing diversity of code clone types, from syntactic copies to cross-language semantic clones to AI-generated duplicates, has created a fragmenta

Agent Skills for Automated Reasoning policies in Amazon Bedrock

SafetyDGX agent

Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custo

AMA: MiniMax H3 Team — Ask us anything about our open video generation model, training, and future plans

Local AiDGX agent

https://preview.redd.it/kihat320ashh1.png?width=1672&format=png&auto=webp&s=a7ccc40ba3fb229ac7ebf57e8e6a314e0ee45646 Hi r/StableDiffusion! u/New-Requirement1419 -> dacongya (Head of H3 Researcher) u/A

An accurate characterization of the arc of AI is that it is shaped by two trends: 1. Moving more and more logic to a neural model for tasks …

AgentsDGX agent

An accurate characterization of the arc of AI is that it is shaped by two trends: 1. Moving more and more logic to a neural model for tasks where training data can be densely sampled (e.g. the shift f

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

HardwareDGX agent

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classific

Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning

Model ReleasesDGX agent

arXiv:2608.05144v1 Announce Type: new Abstract: Long-horizon reasoning requires an agentic runtime that can persist when evidence supports its current approach and pivot when measurements reveal failu

← Previous
1…4647484950…91
Next →