AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,420 results
Agents

OmniReasoner: Thinking with Long Audio-Video via Native Tool Use

DGX agent

arXiv:2607.19339v1 Announce Type: new Abstract: Long audio-video reasoning is difficult for omnimodal LLMs because the decisive evidence is often sparse, cross-modal, and too expensive to preserve wit

agentsarxiv-cs-cv
23 Jul 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Prober.ai: Gated Inquiry-Based Feedback via LLM-Constrained Personas for Argumentative Writing Development

DGX agent

arXiv:2605.05598v2 Announce Type: replace Abstract: The proliferation of large language models (LLMs) in educational settings has paradoxically undermined the cognitive processes they purport to suppo

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference

DGX agent

arXiv:2607.20327v1 Announce Type: new Abstract: Large language models (LLMs) provide strong reasoning capabilities but are expensive to serve at scale, whereas small language models (SLMs) are cheaper

safetyarxiv-cs-cl
23 Jul 2026
Model Releases

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

DGX agent

arXiv:2607.20058v1 Announce Type: new Abstract: Large language models can answer scientific questions, yet a correct output does not reveal whether the model represents or uses the governing physics.

model-releasesarxiv-cs-ai
23 Jul 2026
Local Ai

The Giant Hippocampus: From Structural Monoculture to a System of Systems

DGX agent

arXiv:2607.19973v1 Announce Type: new Abstract: AI researchers describe state-of-the-art models as one thing repeated at scale: the Transformer, wired identically for text, pixels, or speech. Neurosci

local-aiarxiv-cs-ai
23 Jul 2026
Research

Time Series Network Utilization KPI Forecasting Using Advanced AI/ML Models

DGX agent

arXiv:2607.19974v1 Announce Type: cross Abstract: The rapid proliferation of data-intensive applications, cloud infrastructure, and IoT ecosystems has made proactive resource provisioning critical for

researcharxiv-cs-ai
23 Jul 2026
Research

Total Variation Distance Estimation in Autoregressive Models

DGX agent

arXiv:2607.19510v1 Announce Type: new Abstract: Modern LLM deployments use a number of implementation choices and inference optimizations (e.g., batching, custom kernels, and quantization) on top of f

researcharxiv-cs-lg
23 Jul 2026
Model Releases

Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patches

DGX agent

TL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA

model-releasesr-stablediffusion
23 Jul 2026
Research

Understanding Developer Pain Points in Federated Learning: Insights from Stack Overflow and GitHub

DGX agent

arXiv:2607.19621v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training without centralizing raw data, but building and operating FL systems remains difficult du

researcharxiv-cs-ai
23 Jul 2026
Safety

Wish Dario had taken this bet.

DGX agent

Wish Dario had taken this bet. 🎺 I am hereby publicly offering to bet @darioamodei $1,000,000 that AI in 2027 will NOT be “smarter than Nobel Prize winners across most fields in science and engineerin

safetygary-marcus--x
23 Jul 2026
Model Releases

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

DGX agent

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke

model-releasessimon-willison
22 Jul 2026
Model Releases

Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

DGX agent

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

model-releasesr-ollama
22 Jul 2026
Agents

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dens…

DGX agent

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dense tables, footnoted adjustments, and the details buried in t

agentsjerry-liu--x
22 Jul 2026
Hardware

Inside NVIDIA Rubin GPU Architecture: Powering the Era of Agentic AI

DGX agent

NVIDIA’s Rubin GPU, the core of the Vera Rubin platform, delivers up to 10× the agentic inference throughput per watt compared with previous generations, using 336 billion transistors, 224 SMs, 896 Te

hardwarenvidia-developer
21 Jul 2026
Model Releases

Now in preview: Find and fix software vulnerabilities with CodeMender

DGX agent

As adversarial AI threats accelerate attacks on code, security teams must counter them with machine-speed defenses that can automate code remediation and fight AI with AI. CodeMender is our managed co

model-releasesgoogle-cloud-ai
21 Jul 2026
Model Releases

This is a neat feature. I wrote an article a few weeks back about how I built this into my agent orchestrator: https://x.com/omarsar0/status…

DGX agent

This is a neat feature. I wrote an article a few weeks back about how I built this into my agent orchestrator: https://x.com/omarsar0/status/2073404610501329247?s=20 But I made it multimodal from the

model-releasesdair-ai--x
21 Jul 2026
Safety

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by r…

DGX agent

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by removing their biggest bottleneck: AI compute fragmentation.

safetyclem-delangue--x
21 Jul 2026
Model Releases

We talked about Claude Code, Claude Tag, Fable, coding agent security, evals, tool design, and how Anthropic use these tools themselves Clau…

DGX agent

We talked about Claude Code, Claude Tag, Fable, coding agent security, evals, tool design, and how Anthropic use these tools themselves Claude Tag (Claude Code via Slack) is already landing 65% of the

model-releasessimon-willison--x
21 Jul 2026
Hardware

Accelerating automotive innovation with C4A-metal and Panasonic Automotive vSkipGen

DGX agent

As the automotive landscape accelerates toward software-defined vehicles, Cockpit Domain Controllers (CDCs) are becoming the core of next-generation in-cabin experiences. The ability to rapidly develo

hardwaregoogle-cloud-ai
20 Jul 2026
Hardware

At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI

DGX agent

At SIGGRAPH 2026, NVIDIA unveiled a suite of AI‑driven graphics and simulation advances, highlighting neural rendering, agentic and physical AI world models, and real‑time simulation methods. Key rele

hardwarenvidia-blog
20 Jul 2026
Model Releases

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without s…

DGX agent

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without sacrificing performance. Model routing will become a core par

model-releasesdair-ai--x
20 Jul 2026
Safety

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

DGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

safetyarxiv-cs-ro
16 Jul 2026
Model Releases

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

DGX agent

arXiv:2607.13705v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, the need for unified evaluation infrastructure becomes critical. However, current evaluat

model-releasesarxiv-cs-ai
16 Jul 2026
Research

AI advice suppresses people's willingness to say 'I don't know', even when the advice is wrong and accuracy is incentivized

DGX agent

arXiv:2607.13562v1 Announce Type: new Abstract: Knowing when to say 'I don't know' is fundamental to human judgment, yet AI assistants offer a fluent answer to almost any question. In five experiments

researcharxiv-cs-ai
16 Jul 2026
Research

Analyzing Curricular Pattern Complexity Using AI to Improve On-Time Graduation Rates

DGX agent

arXiv:2607.13094v1 Announce Type: cross Abstract: The rise of Artificial Intelligence (AI) enables automatic analysis of large amounts of data. Previously time-consuming and labor-intensive tasks can

researcharxiv-cs-ai
16 Jul 2026
Model Releases

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

DGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

DGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

safetyarxiv-cs-ai
16 Jul 2026
Hardware

Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit

DGX agent

arXiv:2607.13095v1 Announce Type: cross Abstract: We present a full-pipeline inference optimization for the MiMo-V2.5 model family, which combines Hybrid Sliding Window Attention (Hybrid SWA), sparse

hardwarearxiv-cs-ai
16 Jul 2026
Research

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

DGX agent

arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encour

researcharxiv-cs-lg
16 Jul 2026
Model Releases

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

DGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

DGX agent

arXiv:2604.00830v3 Announce Type: replace-cross Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

DGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

model-releasesarxiv-cs-ai
16 Jul 2026
Applications

LPM: Industrial-Scale Generative Video Restoration

DGX agent

arXiv:2607.13460v1 Announce Type: new Abstract: We present the Large Processing Model (LPM), a diffusion-based generative framework for photorealistic video restoration under complex, in-the-wild degr

applicationsarxiv-cs-cv
16 Jul 2026
Model Releases

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B on 2x3090s

DGX agent

I managed to get this model working on 2x 3090s with full 262k ctx and N=4, if anyone is interested to try it, thanks to this quant: https://huggingface.co/danielrmay/NVIDIA-Nemotron-Labs-3-Puzzle-75B

model-releasesr-localllama
16 Jul 2026
Model Releases

OvisOCR2 Technical Report

DGX agent

arXiv:2607.13639v1 Announce Type: cross Abstract: We introduce OvisOCR2, a 0.8B document parsing model. OvisOCR2 is designed as an end-to-end parser: given a document page image, it generates a Markdo

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

DGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset

DGX agent

arXiv:2603.11811v2 Announce Type: replace-cross Abstract: The acquisition of large-scale physical interaction data, a critical prerequisite for modern robot learning, is severely bottlenecked by the p

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

DGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

The Cafe in Amsterdam: When the Incumbent Becomes the Oracle

DGX agent

arXiv:2607.13393v1 Announce Type: cross Abstract: A field can reformulate its computations freely exactly where its demand is stated independently of any incumbent implementation, and finds itself una

safetyarxiv-cs-ai
16 Jul 2026
Local Ai

A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study

DGX agent

arXiv:2607.12886v1 Announce Type: new Abstract: Clinical notes contain many of the signs and symptoms that bring patients to care, yet this information rarely reaches structured fields. Existing extra

local-aiarxiv-cs-ai
15 Jul 2026
Model Releases

A Shared Subcircuit Lets LLMs Count Down Across Tasks

DGX agent

arXiv:2607.12279v1 Announce Type: new Abstract: Writing a sentence of exactly twelve words; ending a DNA sequence at the right codon; formatting an ASCII table. These are all tasks that language model

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

DGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

model-releasesr-localllama
15 Jul 2026
Research

An Omnilingual-ASR-Based Speech-LLM System for the 2nd MLC-SLM Challenge

DGX agent

arXiv:2607.12468v1 Announce Type: cross Abstract: We describe our submission to Task 1 of the 2nd MLCSLM Challenge: a cascaded diarization-then-recognition system that combines DiariZen-Large-s80 (Wav

researcharxiv-cs-ai
15 Jul 2026
Model Releases

Bonsai-27B & Ternary-Bonsai-27B - Updates (on PRs)

DGX agent

Below Upstream Status sections are from https://github.com/PrismML-Eng/Bonsai-demo Upstream Status for Binary Q1_0 is supported out of the box in upstream llama.cpp across many backends: CPU (generic,

model-releasesr-localllama
15 Jul 2026
Hardware

Cadence unveils AuraStack AI Super Agent, an AI platform for PCB and advanced chip packaging design, with Nvidia, TSMC, and Schneider Electric among early users (Marco Chiappetta/Forbes)

DGX agent

Marco Chiappetta / Forbes: Cadence unveils AuraStack AI Super Agent, an AI platform for PCB and advanced chip packaging design, with Nvidia, TSMC, and Schneider Electric among early users — As systems

hardwaretechmeme
15 Jul 2026
Safety

Calibration-First Reward-Component Auditing for Reinforcement Learning Control in Smart Greenhouses

DGX agent

arXiv:2607.11959v1 Announce Type: new Abstract: Greenhouse reinforcement learning can test climate-control ideas at a speed and scale that is difficult to achieve with crop experiments alone. For smar

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

Deep4ge: DNN Training Trajectories for Fault Detection and Diagnosis

DGX agent

arXiv:2607.12868v1 Announce Type: cross Abstract: Deep learning systems often fail due to subtle implementation faults that alter training behavior. Recent work has studied how to detect and diagnose

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

Designing Agent-Ready Websites for AI Web Agents: A Framework for Machine Readability, Actionability, and Decision Reliability

DGX agent

arXiv:2607.12056v1 Announce Type: new Abstract: Online shopping is increasingly shifting toward a model in which AI agents independently search for products, compare options, evaluate constraints, and

model-releasesarxiv-cs-ai
15 Jul 2026
← Previous
1…6566676869…113
Next →