AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,412 results
20 Jul 2026

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without s…

Model ReleasesDGX agent

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without sacrificing performance. Model routing will become a core par

16 Jul 2026

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

SafetyDGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2607.13705v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, the need for unified evaluation infrastructure becomes critical. However, current evaluat

AI advice suppresses people's willingness to say 'I don't know', even when the advice is wrong and accuracy is incentivized

ResearchDGX agent

arXiv:2607.13562v1 Announce Type: new Abstract: Knowing when to say 'I don't know' is fundamental to human judgment, yet AI assistants offer a fluent answer to almost any question. In five experiments

Analyzing Curricular Pattern Complexity Using AI to Improve On-Time Graduation Rates

ResearchDGX agent

arXiv:2607.13094v1 Announce Type: cross Abstract: The rise of Artificial Intelligence (AI) enables automatic analysis of large amounts of data. Previously time-consuming and labor-intensive tasks can

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

Model ReleasesDGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

SafetyDGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit

HardwareDGX agent

arXiv:2607.13095v1 Announce Type: cross Abstract: We present a full-pipeline inference optimization for the MiMo-V2.5 model family, which combines Hybrid Sliding Window Attention (Hybrid SWA), sparse

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

ResearchDGX agent

arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encour

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

Model ReleasesDGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

SafetyDGX agent

arXiv:2604.00830v3 Announce Type: replace-cross Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

Model ReleasesDGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

LPM: Industrial-Scale Generative Video Restoration

ApplicationsDGX agent

arXiv:2607.13460v1 Announce Type: new Abstract: We present the Large Processing Model (LPM), a diffusion-based generative framework for photorealistic video restoration under complex, in-the-wild degr

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B on 2x3090s

Model ReleasesDGX agent

I managed to get this model working on 2x 3090s with full 262k ctx and N=4, if anyone is interested to try it, thanks to this quant: https://huggingface.co/danielrmay/NVIDIA-Nemotron-Labs-3-Puzzle-75B

OvisOCR2 Technical Report

Model ReleasesDGX agent

arXiv:2607.13639v1 Announce Type: cross Abstract: We introduce OvisOCR2, a 0.8B document parsing model. OvisOCR2 is designed as an end-to-end parser: given a document page image, it generates a Markdo

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

Model ReleasesDGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset

SafetyDGX agent

arXiv:2603.11811v2 Announce Type: replace-cross Abstract: The acquisition of large-scale physical interaction data, a critical prerequisite for modern robot learning, is severely bottlenecked by the p

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

Model ReleasesDGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

The Cafe in Amsterdam: When the Incumbent Becomes the Oracle

SafetyDGX agent

arXiv:2607.13393v1 Announce Type: cross Abstract: A field can reformulate its computations freely exactly where its demand is stated independently of any incumbent implementation, and finds itself una

15 Jul 2026

A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study

Local AiDGX agent

arXiv:2607.12886v1 Announce Type: new Abstract: Clinical notes contain many of the signs and symptoms that bring patients to care, yet this information rarely reaches structured fields. Existing extra

A Shared Subcircuit Lets LLMs Count Down Across Tasks

Model ReleasesDGX agent

arXiv:2607.12279v1 Announce Type: new Abstract: Writing a sentence of exactly twelve words; ending a DNA sequence at the right codon; formatting an ASCII table. These are all tasks that language model

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

Model ReleasesDGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

An Omnilingual-ASR-Based Speech-LLM System for the 2nd MLC-SLM Challenge

ResearchDGX agent

arXiv:2607.12468v1 Announce Type: cross Abstract: We describe our submission to Task 1 of the 2nd MLCSLM Challenge: a cascaded diarization-then-recognition system that combines DiariZen-Large-s80 (Wav

Bonsai-27B & Ternary-Bonsai-27B - Updates (on PRs)

Model ReleasesDGX agent

Below Upstream Status sections are from https://github.com/PrismML-Eng/Bonsai-demo Upstream Status for Binary Q1_0 is supported out of the box in upstream llama.cpp across many backends: CPU (generic,

Cadence unveils AuraStack AI Super Agent, an AI platform for PCB and advanced chip packaging design, with Nvidia, TSMC, and Schneider Electric among early users (Marco Chiappetta/Forbes)

HardwareDGX agent

Marco Chiappetta / Forbes: Cadence unveils AuraStack AI Super Agent, an AI platform for PCB and advanced chip packaging design, with Nvidia, TSMC, and Schneider Electric among early users — As systems

Calibration-First Reward-Component Auditing for Reinforcement Learning Control in Smart Greenhouses

SafetyDGX agent

arXiv:2607.11959v1 Announce Type: new Abstract: Greenhouse reinforcement learning can test climate-control ideas at a speed and scale that is difficult to achieve with crop experiments alone. For smar

Deep4ge: DNN Training Trajectories for Fault Detection and Diagnosis

Model ReleasesDGX agent

arXiv:2607.12868v1 Announce Type: cross Abstract: Deep learning systems often fail due to subtle implementation faults that alter training behavior. Recent work has studied how to detect and diagnose

Designing Agent-Ready Websites for AI Web Agents: A Framework for Machine Readability, Actionability, and Decision Reliability

Model ReleasesDGX agent

arXiv:2607.12056v1 Announce Type: new Abstract: Online shopping is increasingly shifting toward a model in which AI agents independently search for products, compare options, evaluate constraints, and

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

ResearchDGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

FormalAnalyticGeo: A Neural-Symbolic Based Framework for Multimodal Analytic Geometry Problem Generation

Model ReleasesDGX agent

arXiv:2607.12982v1 Announce Type: new Abstract: Math reasoning has achieved significant progress with the rapid advancement of Multimodal Large Language Models (MLLMs), however analytic geometry remai

FoundationGeo: Learning Spatial Pixel-Wise Fields for Monocular Metric Geometry

Local AiDGX agent

arXiv:2607.11588v2 Announce Type: replace Abstract: We present FoundationGeo, a two-stage framework that explicitly bridges relative and metric prediction via spatial calibration and principled data d

GRID: Grammar-Railed Decoding for Enterprise SQL Generation

Model ReleasesDGX agent

arXiv:2607.11951v1 Announce Type: new Abstract: Large language models can write SQL, but enterprise deployment demands more than plausible text: outputs must be syntactically valid, must respect per-r

How Inference Compute Shapes Frontier LLM Evaluation

Model ReleasesDGX agent

arXiv:2606.17930v2 Announce Type: replace Abstract: AI evaluations are shifting toward harder tasks that benefit from longer trajectories involving tool use and iterative problem solving. As a result,

Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement

AgentsDGX agent

arXiv:2606.19387v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success in software development. However, they are susceptible to hallucinations, meanin

Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions

SafetyDGX agent

arXiv:2607.12406v1 Announce Type: new Abstract: The capability of LLM agents to function as the ``brain'' of a system fundamentally expands the scope of analysis beyond a standalone model. Consequentl

Jetson-PI: Towards Onboard Real-Time Robot Control via Foresight-Aligned Asynchronous Inference

Model ReleasesDGX agent

arXiv:2607.12659v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved impressive performance on diverse embodied tasks. However, deploying VLA models on low-power onboard

Learning the Graphical Nature of Symmetries

ResearchDGX agent

arXiv:2607.12026v1 Announce Type: cross Abstract: Finite groups are rigid algebraic objects, whose Cayley graphs expose a rich network geometry through which group-theoretic structure can be measured,

MaxSAT-Based Feedback for Guiding Vision-Language Models in Sudoku

TutorialsDGX agent

arXiv:2607.12711v1 Announce Type: new Abstract: Vision--Language Models (VLMs) have recently demonstrated promising performance on structured visual reasoning tasks, including grid-based puzzles. Howe

Optimization Is Not All You Need

Model ReleasesDGX agent

arXiv:2607.11977v1 Announce Type: new Abstract: In 2019, OpenAI released two million GPT-2 outputs-ungrammatical, half broken-to aid the detection of machine-generated text. The alignment that produce

PFAdapter: Hierarchical LoRA Decomposition for Personalized Federated MLLMs

Model ReleasesDGX agent

arXiv:2607.12111v1 Announce Type: cross Abstract: Agentic AI systems are reshaping communications and networking by deploying autonomous intelligent agents capable of collaborative learning while main

PolarBM: Complex-valued Boltzmann Machine for Modeling Audio Signals in Polar and Log-polar Coordinates

ResearchDGX agent

arXiv:2607.12417v1 Announce Type: new Abstract: Although vast amounts of data, such as audio signal spectra, are naturally represented using complex numbers, conventional machine learning methods ofte

QwenPaw-Data: Bridging Facts, Methodology, and Execution for Autonomous Enterprise Data Analytics

AgentsDGX agent

arXiv:2607.11019v2 Announce Type: replace Abstract: Enterprise data analysis is emerging as a distinct frontier for autonomous agents. Compared with general-purpose interaction and software engineerin

RepTran: Search-Based Repair of Transformer Models

ResearchDGX agent

arXiv:2607.11193v2 Announce Type: replace-cross Abstract: To ensure the overall quality of AI-enabled software, not only traditional software components but also AI components need to be tested and re

Self-Regulated Reading with AI Support: An Eight-Week Study with Students

TutorialsDGX agent

arXiv:2602.09907v2 Announce Type: replace-cross Abstract: College students increasingly use AI chatbots to support academic reading, yet we lack granular understanding of how these interactions shape

Sophos launches Fusion, an AI-native ‘defense system’ to unify security tools

Model ReleasesDGX agent

Cybersecurity firm Sophos Ltd. today launched Sophos Fusion, a single platform that ties together its security operations, endpoint, network, identity, email and cloud protection. Sophos calls it the

TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

Model ReleasesDGX agent

arXiv:2607.13028v1 Announce Type: cross Abstract: Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground beh

The Risk of Exposed Cloud Functions and How to Harden

Model ReleasesDGX agent

Written by: Corné de Jong Introduction Mandiant security assessments frequently identify publicly exposed serverless applications that lack authentication, often as a result of specific business requi

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

Model ReleasesDGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

Win by Silence: Deletion Non-Monotonicity, Autonomous Exploitation, and Typed-State Gating in LLM Plan Evaluation

AgentsDGX agent

arXiv:2607.12986v1 Announce Type: new Abstract: Plan evaluators can reward a strategic plan for becoming less explicit. This paper studies that failure in a staged expected-value scorer for LLM-genera

14 Jul 2026

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But …

ApplicationsDGX agent

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But that speed allows you to make little tweaks to your system s

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told yo…

Model ReleasesDGX agent

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told you so!' - The mass majority of researchers and academics in d

Lessons From the Leaderboard: What 5,000+ Kagglers Taught Us About Improving AI Reasoning

Model ReleasesDGX agent

The NVIDIA Nemotron Model Reasoning Challenge on Kaggle attracted over 5,000 participants who all began from the same open model, benchmark, and infrastructure. The strongest solutions treated reasoni

Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills

HardwareDGX agent

NVIDIA Cosmos 3 was post‑trained in under a day using TAO agent skills and LoRA adapters, raising accuracy on the Woven Traffic Safety video QA dataset from 54.41 % to 93.35 %. The mixture‑of‑transfor

simonw/pedalican

Model ReleasesDGX agent

simonw/pedalican Clearly I wasn't paying attention when these were first announced back in May, but today I accidentally activated a 'pet' in Codex Desktop - a little animated robot, reminiscent of Cl

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

AgentsDGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

13 Jul 2026

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

Model ReleasesDGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

10 Jul 2026

A safety-oriented hypothetico-deductive framework for AI-assisted differential diagnosis

Model ReleasesDGX agent

arXiv:2607.08038v1 Announce Type: new Abstract: Diagnostic error is a major threat to patient safety, yet current large language model (LLM) systems often treat diagnosis as a one-shot prediction task

And then what is the point of Work? Just a dumbed-down version of Codex that secretly does coding but hides it? (I can understand, maybe, if…

ApplicationsDGX agent

And then what is the point of Work? Just a dumbed-down version of Codex that secretly does coding but hides it? (I can understand, maybe, if the release was a 1st step towards something but there is n

b9956

Local AiDGX agent

Release b9956 is a build version of llama.cpp, an open-source project for LLM inference in C/C++ . The project releases frequently, with multiple releases published in a single day , and b9956 represe

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

SafetyDGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

← Previous
1…5253545556…91
Next →