AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
28 Jul 2026

A Motion-Aware Vector Quantization Framework with Centroid Reuse for Efficient VLA Inference

HardwareDGX agent

arXiv:2607.24148v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated strong potential for embodied AI, yet their high inference latency on GPUs limits real-time deploy

AMD calls its shot, but the real race is engineering velocity

HardwareDGX agent

At AMD’s Advancing AI event, Lisa Su called her shot – just like Babe Ruth. The question now is whether AMD has built the engineering machine to hit it. Last week, we argued that AMD’s next reinventio

Anthropic and Nvidia come out against blanket bans on open-weight AI models

HardwareDGX agent

As the United States government debates new artificial intelligence rules and regulations, Anthropic PBC and Nvidia Corp. are drawing a line against blanket bans on open-weight models, urging regulato


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Anthropic faces backlash from Silicon Valley partners, founders, and researchers for competitive tactics, guardrails, and lack of support for open-weight models (Wall Street Journal)

HardwareDGX agent

Wall Street Journal: Anthropic faces backlash from Silicon Valley partners, founders, and researchers for competitive tactics, guardrails, and lack of support for open-weight models — The AI pioneer f

Apple overtakes Nvidia to become the world's most valuable publicly traded company, two years after it lost the crown, closing at a $4.9T market cap on Monday (Kalley Huang/New York Times)

HardwareDGX agent

Kalley Huang / New York Times: Apple overtakes Nvidia to become the world's most valuable publicly traded company, two years after it lost the crown, closing at a $4.9T market cap on Monday — The comp

Are single GPU research still published in ML/DL and its applications nowadays? Which are the most notable recent ones? [D]

HardwareDGX agent

ML research is progressing at breakneck speed where frontier labs in both academia and industry have access to considerably large computes (GPUs). Where do small labs or independent researchers go in

Beyond Block Boundaries: Multi-Block Editing for Diffusion Large Language Models

HardwareDGX agent

arXiv:2607.22663v1 Announce Type: new Abstract: Block diffusion has emerged as the dominant paradigm for scaling discrete diffusion language models (dLLMs), because decoding text in fixed-size blocks

Causal-TS: A Python Library for Causal Discovery in High-Dimensional and Nonstationary Time Series

HardwareDGX agent

arXiv:2607.24673v1 Announce Type: new Abstract: We describe Causal-TS, an open-source Python library for causal discovery in high-dimensional and nonstationary multivariate time series. Causal-TS prov

Certified Parallel-in-Time Sinkhorn for Dynamic Entropic Optimal Transport

HardwareDGX agent

arXiv:2607.24741v1 Announce Type: cross Abstract: Dynamic applications, including optimal-transport Flow Matching, repeatedly solve related entropic optimal transport problems, yet conventional distri

CHS-SQL: A Text-to-SQL approach based on Confidence-Guided Heuristic Search Schema Linking process

HardwareDGX agent

arXiv:2607.22624v1 Announce Type: new Abstract: Recently, there have been several works in the Text-to-SQL domain that utilize Small Language Models (SLMs) for training. These approaches achieve perfo

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference

HardwareDGX agent

arXiv:2511.21702v2 Announce Type: replace-cross Abstract: Large language models face significant computational bottlenecks during inference due to the expensive output layer computation over large voc

Developing Healthcare Robotics with GPU-Native Medical Physics Simulation

HardwareDGX agent

Healthcare robotics struggles with data scarcity, limited generalization to rare clinical scenarios, and slow prototyping due to the need for annotated demonstrations and costly experimental setups. N

DynaResize: Runtime GPU Reallocation for Disaggregated LLM Post-Training

HardwareDGX agent

arXiv:2607.22614v1 Announce Type: new Abstract: RL-based LLM post-training increasingly disaggregates Rollout and Training across separate GPU resources, but static GPU partitioning suffers from sever

Energy Constrained Hierarchical Underwater Monitoring via Local Multi-Agent RAG

HardwareDGX agent

arXiv:2607.24313v1 Announce Type: cross Abstract: Marine life monitoring is limited by strict energy constraints, poor underwater connectivity, and the high cost of transmitting raw multimodal data fr

Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models

HardwareDGX agent

arXiv:2507.21438v2 Announce Type: replace Abstract: Ontologies and knowledge graphs require continuous evolution to remain comprehensive and accurate, but manual curation is labor intensive. Large Lan

Hybrid Semantic and Spectral Ensemble for Robust Synthetic Image Source Attribution

HardwareDGX agent

arXiv:2607.22808v1 Announce Type: cross Abstract: The rapid advancement of text-to-image (T2I) models has necessitated robust Synthetic Image Source Attribution (SIA) methodologies. A critical challen

InnerGS: Internal Scenes Reconstruction and Segmentation via Factorized 3D Gaussian Splatting

HardwareDGX agent

arXiv:2508.13287v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has recently gained popularity for efficient scene rendering by representing scenes as explicit sets of anisotrop

Kalypso: Relational LLM Serving

HardwareDGX agent

arXiv:2607.23815v1 Announce Type: cross Abstract: Large language models are increasingly used as semantic operators for filtering, extracting, ranking, joining, and transforming unstructured data. Exi

Multi-primitive in-memory computing for Monte Carlo tree search

HardwareDGX agent

arXiv:2607.22869v1 Announce Type: cross Abstract: Monte Carlo tree search (MCTS) enables artificial intelligence (AI) decision-making, but requires 55-300 W on conventional processors, limiting edge d

netflix’s stack

HardwareDGX agent

netflix’s stack In-House LLM Serving at Netflix Netflix built an in-house LLM serving platform using vLLM and NVIDIA Triton, integrating self-hosted models into its existing production infrastructure

Omni-Prune: Query-Aware Unified Token Pruning for Efficient Omnimodal Large Language Models

HardwareDGX agent

arXiv:2607.23445v1 Announce Type: cross Abstract: Omnimodal large language models (OmniLLMs) are rapidly extending multimodal reasoning to cover synchronized audio and video. However, the resulting au

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models

HardwareDGX agent

arXiv:2607.23193v1 Announce Type: new Abstract: Existing token compression methods for omnimodal large language models typically rely on one modality to determine what to retain in the other. We show

OpenWiki supports 12 LLM providers: - @OpenAI API & @ChatGPT sub login - @AnthropicAI - @GeminiApp AI Studio & Enterprise Vertex - @awscloud…

HardwareDGX agent

OpenWiki supports 12 LLM providers: - @OpenAI API & @ChatGPT sub login - @AnthropicAI - @GeminiApp AI Studio & Enterprise Vertex - @awscloud Bedrock - @OpenRouter - @FireworksAI_HQ - @baseten - @nebiu

Powerful Compute So Compact, It’s Clutch — Build AI Anywhere With NVIDIA Jetson

HardwareDGX agent

As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent video, Guo, founder of

Real-Time Semantic Segmentation with Optimized RetinaNet Architectures for Embedded Automotive Systems

HardwareDGX agent

arXiv:2607.22714v1 Announce Type: cross Abstract: Real-time perception is a foundational requirement for advanced driver assistance systems (ADAS) and autonomous vehicles, yet embedded automotive plat

Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline

HardwareDGX agent

arXiv:2607.22997v1 Announce Type: cross Abstract: Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the real world -- has emerged as the next

Rendering on Real Silicon: GPU Render-Timing as a Passive, AI-Resistant CAPTCHA Signal

HardwareDGX agent

arXiv:2607.23389v1 Announce Type: cross Abstract: Conventional CAPTCHAs pose puzzles that modern AI systems increasingly solve, while behavioral and cryptographic-attestation defenses carry privacy or

Small-Pollinator Detection in Cluttered Field Video

HardwareDGX agent

arXiv:2607.22913v1 Announce Type: new Abstract: Detecting pollinators in field video is challenging: targets are small, visually similar, and observed against cluttered vegetation under blur and occlu

Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach

HardwareDGX agent

arXiv:2607.22584v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation pipelines rank retrieved documents by semantic similarity alone, without accounting for source provenance or cre

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

HardwareDGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

🦔The cost of insuring Big Tech debt against default just hit record highs. Credit default swaps are insurance policies investors buy when t…

HardwareDGX agent

🦔The cost of insuring Big Tech debt against default just hit record highs. Credit default swaps are insurance policies investors buy when they think a company might not pay back what it borrowed. When

We recently joined NVIDIA, Microsoft, and others in supporting open-weight AI. Today, we’re putting that belief into the product with model …

HardwareDGX agent

We recently joined NVIDIA, Microsoft, and others in supporting open-weight AI. Today, we’re putting that belief into the product with model choice on Replit, starting with Kimi K3. The future is the r

X-Stage: An Overlooked Pipeline Stage for Communication-Computation Overlap in DiT Inference

HardwareDGX agent

arXiv:2607.23264v1 Announce Type: cross Abstract: Fine-grained, device-initiated communication lets persistent GPU kernels in distributed diffusion transformer (DiT) inference issue remote stores and

27 Jul 2026

Advancing Semiconductor Innovation Across Materials Engineering and Manufacturing

HardwareDGX agent

Applied Materials and NVIDIA have created an end‑to‑end digital development platform that unifies GPU‑accelerated tools (Ginestra, cuDSS, cuEST, PhysicsNeMo, Omniverso) with CUDA‑X libraries to accele

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, …

HardwareDGX agent

AI security improves when organizations share research, tools and real-world experience. We’re joining industry leaders, including @NVIDIA, in the Open Secure AI Alliance to help organizations identif

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. …

HardwareDGX agent

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. During the Hugging Face incident, closed AI blocked essentia

Built & Trained a Transformer from Scratch in Pure PyTorch for English-to-Tamil Machine Translation [Math + Code Breakdown] [P]

HardwareDGX agent

Hi everyone! 👋 I built and trained the complete Transformer architecture from scratch using pure PyTorch (`torch.nn` primitives) based on the original 'Attention Is All You Need' paper. I trained the

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open mo…

HardwareDGX agent

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open models and harnesses for defense. Lets stop believing the PR t

If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to …

HardwareDGX agent

If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to IlyAGI We are announcing a long-term strategic partnership w

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can …

HardwareDGX agent

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can take the best open model in the world and efficiently tune i

Nvidia could reportedly backstop $250B loan for new OpenAI data center campus

HardwareDGX agent

Nvidia Corp. is reportedly in talks to backstop a 250 billion loan for OpenAI Group PBC. The Wall Street Journal on Sunday cited sources as saying that the financing would cover the cost of a new data

NVIDIA Harnesses Vera CPU to Speed Up Design of Next-Generation CPUs and GPUs

HardwareDGX agent

The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDIA is collaborating w

Nvidia is putting its Vera CPUs to work alongside AI agents to speed up chip design

HardwareDGX agent

Nvidia Corp. said tonight it’s now running the chip design software its engineers are using to design the next generation of its graphics processing units, on its own silicon. It’s partnering with Cad

NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning

HardwareDGX agent

NVIDIA Ising Calibration 1.5 is a 31‑billion‑parameter vision‑language model that diagnoses and tunes quantum processors, delivering state‑of‑the‑art zero‑shot and in‑context learning on the QCalEval

Nvidia says it made a 'substantial' investment in Safe Superintelligence, which will get access to Nvidia GPUs to up computing power by an 'order of magnitude' (Keach Hagey/Wall Street Journal)

HardwareDGX agent

Keach Hagey / Wall Street Journal: Nvidia says it made a “substantial” investment in Safe Superintelligence, which will get access to Nvidia GPUs to up computing power by an “order of magnitude” — Par

On theCUBE Pod: AMD nips at Nvidia’s heels, and IBM undergoes quiet transformation

HardwareDGX agent

The artificial intelligence competition is no longer between large language models, but rather between AI systems. Advanced Micro Devices Inc. has shifted from being primarily a chip business to a sec

Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to acceleratin…

HardwareDGX agent

Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to accelerating the adoption and trust of open models and harnesses throug

Proud to join @nvidia as a founding member of the Open Secure AI Alliance. Defenders need open tools, collaboration, and transparency to sec…

HardwareDGX agent

Proud to join @nvidia as a founding member of the Open Secure AI Alliance. Defenders need open tools, collaboration, and transparency to secure AI agents and systems. And customers want the ability to

RIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention

HardwareDGX agent

arXiv:2607.21927v1 Announce Type: new Abstract: Full self-attention in large language models scales as O(N^2), which limits long-context document analysis to 65,536 tokens and requires costly GPU clus

Singular value soft-thresholding via the polar decomposition

HardwareDGX agent

arXiv:2607.22484v1 Announce Type: cross Abstract: Singular value soft-thresholding can be computed via a reduction to the matrix polar decomposition, which allows one to exploit GPU-friendly algorithm

Six Agent Harness Capabilities for Higher Model Performance

HardwareDGX agent

The performance of AI agents depends not only on the underlying models but also heavily on their “harness”—the surrounding architecture that supplies context, state management, action execution, and t

Sources: Nvidia has committed to invest 5B in Ilya Sutskever's SSI; the startup has previously raised about 3B in funding and was valued at $32B last year (Shirin Ghaffary/Bloomberg)

HardwareDGX agent

Shirin Ghaffary / Bloomberg: Sources: Nvidia has committed to invest 5B in Ilya Sutskever's SSI; the startup has previously raised about 3B in funding and was valued at 32B last year — Nvidia Corp. ha

TensorWave targets focused AI cloud strategy to deliver better customer experience

HardwareDGX agent

As AI infrastructure matures, AI cloud strategy is increasingly defined by reliability and open ecosystems rather than raw GPU performance. That evolution is creating new opportunities for specialized

Unified Static-Dynamic Pruning for Efficient LLM Inference

HardwareDGX agent

arXiv:2607.21985v1 Announce Type: cross Abstract: The increasing deployment of large language models (LLMs) has magnified the computational and memory bottlenecks of autoregressive decoding, where low

We're proud to join @NVIDIA and the Open Secure AI Alliance. To support open source models, we're contributing our research on measuring the…

HardwareDGX agent

We're proud to join @NVIDIA and the Open Secure AI Alliance. To support open source models, we're contributing our research on measuring the trustworthiness and security of open source models. Closing

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔

HardwareDGX agent

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔 JUST IN : NVIDIA IN TALKS TO PROVIDE 250 BILLION FINANCIAL BACKSTOP FOR OPENAI DATA CENTER IN O

26 Jul 2026

It continues to boggle the mind how many people, who otherwise seem to have a functioning intellect, appear to lose all cognitive capacity w…

HardwareDGX agent

It continues to boggle the mind how many people, who otherwise seem to have a functioning intellect, appear to lose all cognitive capacity when it comes to thinking about actions that impact the compa

Sources: Nvidia is in talks to provide a ~$250B backstop for OpenAI as part of a 10 GW data center project that SoftBank is developing in Ohio (Wall Street Journal)

HardwareDGX agent

Wall Street Journal: Sources: Nvidia is in talks to provide a ~$250B backstop for OpenAI as part of a 10 GW data center project that SoftBank is developing in Ohio — Project would be one of the larges

Understanding GPU Inference Workloads [D]

HardwareDGX agent

Hey everyone, I have been looking into how people source compute for their Inference workloads (and in general). I wanted to understand some specific pain points here. If you've used online services l

25 Jul 2026

AMD moves beyond challenger status in the race for AI platform leadership

HardwareDGX agent

AI hardware competition entered a sharper phase this week as Advanced Micro Devices Inc. used its flagship AI event to argue it isn’t merely chasing Nvidia Corp. — it intends to lead the market outrig

← Previous
123456…29
Next →