AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,477 results
5 Aug 2026

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

4 Aug 2026

Adaptive Reconstruction of Bosonic Quantum States

Model ReleasesDGX agent

arXiv:2608.02049v1 Announce Type: cross Abstract: Bosonic quantum systems provide a hardware-efficient platform for quantum information processing but remain challenging to characterise due to their l

Device-First Feedback: Toward Mobile-Native LLM-Driven Neural Architecture Search

Local Ai
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.00078v1 Announce Type: new Abstract: Deploying convolutional neural networks generated by large language models (LLMs) on real mobile hardware requires more than GPU validation accuracy: IN

Loggia dei Lanzi: AI Thermography Enhancement Comparisons through 3D Photogrammetry

Model ReleasesDGX agent

arXiv:2608.02404v1 Announce Type: new Abstract: The Loggia dei Lanzi in the Piazza della Signoria is one of Florence's most prominent structures visited by millions every year. Its construction histor

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

Model ReleasesDGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

On the Limits of Machine-Learned Ranking for Modern Microarchitectural Policies

ResearchDGX agent

arXiv:2608.01041v1 Announce Type: cross Abstract: Machine-learning predictors estimate processor performance far faster than cycle-level simulation. For design-space exploration, however, the valuable

Open-DiffLoco: Open-Source Differentiable Learning for Deployable Blind Quadruped Locomotion

Local AiDGX agent

arXiv:2608.02069v1 Announce Type: cross Abstract: Developing deployable locomotion policies through conventional reinforcement learning often requires complex reward engineering and expensive training

Pseudorandom Streams within Diffusion Models Act as Learnable Inputs That Affect Generation Quality

Local AiDGX agent

arXiv:2608.02575v1 Announce Type: new Abstract: Diffusion models rely on stochastic inputs, yet on finite-precision hardware, the 'randomness' they consume is realized as deterministic numerical orbit

Rapid Embodiment Adaptation for Quadrupedal Locomotion

SafetyDGX agent

arXiv:2608.01506v1 Announce Type: cross Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learning-based robot policies often break whe

SoniSpeech: A Large-Scale Open-Vocabulary Tri-Modal Dataset for Wearable Silent Speech Interfaces

Model ReleasesDGX agent

arXiv:2608.00803v1 Announce Type: cross Abstract: Wearable silent speech interfaces (SSIs) are limited to small, closed vocabularies. Approaches achieving larger vocabularies require obtrusive hardwar

The Elements of Differentiable Programming

ResearchDGX agent

arXiv:2403.14606v4 Announce Type: replace Abstract: Artificial intelligence has recently experienced remarkable advances, fueled by large models, vast datasets, accelerated hardware, and, last but not

3 Aug 2026

BWM: A Low-Cost High-Fidelity World Simulator for Robot Learning

Model ReleasesDGX agent

arXiv:2607.29302v1 Announce Type: cross Abstract: Reliable robot learning requires a world simulator that can predict action consequences before execution on physical hardware, including risky and fai

Netflix co-founder backs $312M round for optical inference appliance maker Olix

IndustryDGX agent

Artificial intelligence hardware startup Olix Computing Ltd. today announced that it has raised 312 million in funding. The Series C round included contributions from Arm Holding plc, Netflix Inc. co-

2 Aug 2026

Encrypted Clouds?

Model ReleasesDGX agent

I love the progress happening on open models but I feel like it is kind of getting clear that hardware to run good sized models is completely unaffordable for me right now. I know that you all love Qw

What’s the community’s favorite benchmark to validate performance?

Model ReleasesDGX agent

Built my 1st inference machine and have been tweaking models trying to get the most out of my modest hardware. I think I’m at a good place but I’m testing with my own prompts. I’ve looked into some of

31 Jul 2026

Compact Task-Aligned Imitation Learning for Laboratory Automation

Local AiDGX agent

arXiv:2603.01110v2 Announce Type: replace Abstract: Robotic laboratory automation has traditionally relied on carefully engineered motion pipelines and task-specific hardware interfaces, resulting in

Experience sharing: How do you use your local models and for what kind of tasks?

Model ReleasesDGX agent

Here is my experience, which I would like to share with you and I also would like to hear your thoughts and valuable tips&tricks. Hardware: Mac Mini M4 (32GB Unified Memory) Model Server: Ollama Orche

Has anyone actually benchmarked where the 'big-model orchestrator + local-model worker' split breaks down?

Model ReleasesDGX agent

I keep seeing the 'use a big model via API as the architect, run local small/mid models as workers' pattern recommended for people with modest local hardware. I've been running it myself (orchestrator

We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

Model ReleasesDGX agent

I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B. On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is

30 Jul 2026

LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving

Model ReleasesDGX agent

arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and t

MLVC: Multi-platform Learned Video Codec for Real-World Deployment [P]

ApplicationsDGX agent

I've always found it a little strange that AI is everywhere, but the codecs we use in practice are the traditional hand-engineered systems like h.264, h.265, av1. Alexnet started the wave of neural ne

P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Model ReleasesDGX agent

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

Shot-based quantum encoding: a data-loading paradigm for quantum neural networks

ResearchDGX agent

arXiv:2604.06135v2 Announce Type: replace-cross Abstract: Efficient data loading remains a bottleneck for near-term quantum machine learning. Existing schemes (angle, amplitude, and basis encoding) ei

Under the Hood: Serving Kimi K3

Model ReleasesDGX agent

DigitalOcean launched Kimi K3 on day 0. It’s already one of the most popular models on the platform and across the market: second most likes on Hugging Face, sixth most traffic on OpenCode. Getting a

29 Jul 2026

Characterizing and Mitigating the Effects of Device Temperature on RF Fingerprinting Accuracy

ApplicationsDGX agent

arXiv:2607.25070v1 Announce Type: cross Abstract: Radio Frequency Fingerprinting (RFFP) has emerged as a promising approach for device authentication by exploiting hardware-specific impairments embedd

28 Jul 2026

AI model compression startup Multiverse raises 570M at 1.7B valuation

IndustryDGX agent

Multiverse Computing SL, a startup working on technology that compresses artificial intelligence models so that they run more efficiently on less hardware, announced Monday it raised 570 million in Se

ArmnetBench v0.1: Parallel Real-World Evaluation of Manipulation Policies on a Low-Cost Arm Farm

Model ReleasesDGX agent

arXiv:2607.24481v1 Announce Type: new Abstract: Real-world evaluation is a bottleneck in developing generalist robot manipulation policies. Each rollout requires physical hardware and an operator to s

Bigger or Cheaper? Scale and Quantization Effects on Uncertainty Signals in Vision-Language Models Under Image Degradation

ResearchDGX agent

arXiv:2607.24440v1 Announce Type: cross Abstract: Vision-language models (VLMs) deployed on consumer hardware must decide when to answer and when to defer, and that decision depends on having a confid

ChatGPT literally saved me money this weekend

IndustryDGX agent

I was never a huge AI guy but I was starting to wonder if my ISP was really providing the speeds I was paying for, and it turns out they were, but I had a bottleneck somewhere in my hardware chain. I

Formally Verified Synthesizable Floating-Point Data Types in ARCH HDL

ResearchDGX agent

arXiv:2607.23715v1 Announce Type: new Abstract: We report the design and end-to-end verification of first-class IEEE-754 binary32 (FP32) and bfloat16 (BF16) arithmetic for ARCH, a hardware description

Occlusion-Point Reuse for Ray-Traced Ambient Occlusion and Shadow

ResearchDGX agent

arXiv:2607.23122v1 Announce Type: cross Abstract: Ambient occlusion (AO) and soft shadows are critical visibility cues for spatial perception in real-time rendering. Hardware ray tracing provides a di

27 Jul 2026

Entanglement geometry separates circuit cutting, classical hardness, and trainability

ResearchDGX agent

arXiv:2607.17872v2 Announce Type: replace-cross Abstract: Circuit cutting promises to scale quantum computations beyond current hardware, but variational quantum advantage also requires low cutting ov

Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation

ResearchDGX agent

arXiv:2607.22034v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on consumer hardware where input images are degraded by compression, camera shake, and poor li

24 Jul 2026

AMD targets AI PCs to curb agentic AI costs as enterprises rethink cloud token economics

Local AiDGX agent

As AI moves beyond chatbots toward autonomous agents, attention is shifting enterprise AI PCs as a new layer of AI infrastructure. That transition is driving demand for hardware and software designed

Beyond Independent Optimization: Compression, MoE Routing, and Quantization Interactions in Multimodal Edge Intelligence

ResearchDGX agent

arXiv:2607.20981v1 Announce Type: new Abstract: Efficient multimodal inference is increasingly constrained not only by model quality or FLOP count, but also by the cost of preserving, moving, routing,

Identifying Good Rules for Efficient SAT Encodings of Single-Constant Multiplication Using Machine Learning

ResearchDGX agent

arXiv:2607.21188v1 Announce Type: new Abstract: The Single Constant Multiplication problem is a fundamental NP-hard optimization task in hardware design, which seeks to decompose a fixed constant usin

MKEvolve: A Modular Multi-Agent Framework for Kernel Code Generation

AgentsDGX agent

arXiv:2607.20501v1 Announce Type: new Abstract: Despite rapid progress in LLM-based code generation, writing correct and performant kernels for hardware accelerators remains a key bottleneck in scalin

Neural Guided Sampling for Quantum Circuit Optimization

ResearchDGX agent

arXiv:2510.12430v2 Announce Type: replace-cross Abstract: Translating a general quantum circuit on a specific hardware topology with a reduced set of available gates, also known as transpilation, come

Stokes-Informed Diffusion for Robust Linear Polarization Estimation

SafetyDGX agent

arXiv:2607.21239v1 Announce Type: new Abstract: Polarization cues benefit applications such as material detection and de-reflection, yet acquiring them typically requires dedicated hardware. This moti

Towards an Automated Test of LLM Security Knowledge

Model ReleasesDGX agent

arXiv:2607.18496v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for a range of software, hardware and human-centered security tasks. Consequently, LLM perf

23 Jul 2026

AI infrastructure demand is outrunning even the boldest supply chain playbooks

AgentsDGX agent

AI infrastructure buildouts are moving so fast that plans made just months ago are already obsolete, forcing hardware makers to rewrite how they design, source and ship the systems powering the next g

20 Jul 2026

What are the current best local models to run on 48GB VRAM?

Local AiDGX agent

I have a 48GB M5 Pro and have far too many development projects going that just don't need the power of Anthropic to churn through so have started looking into running local models and while it certai

16 Jul 2026

Efficient and Privacy Aware Edge Cloud Collaborative Inference for Large Language Models

Local AiDGX agent

arXiv:2607.13093v1 Announce Type: cross Abstract: On-device LLM inference faces a trilemma of response latency, limited hardware resources and user privacy. Full cloud inference delivers strong comput

ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation

SafetyDGX agent

arXiv:2607.13124v1 Announce Type: cross Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compre

15 Jul 2026

Physics-Informed Structure Anchoring With Capture-Aware Prototype Calibration for Cross-Environment RF Fingerprinting

Model ReleasesDGX agent

arXiv:2607.09760v2 Announce Type: replace-cross Abstract: Radio frequency fingerprint identification (RFFI) exploits transmitter-specific hardware imperfections as physicallayer identity cues for Inte

10 Jul 2026

breaking: company built on stolen IP and lies allegedly steals more IP

SafetyDGX agent

breaking: company built on stolen IP and lies allegedly steals more IP “OpenAI’s nascent hardware business now rests on the shakiest of foundations, rotten to its core by its illegal reliance on misap

Sam Altman should be in prison. But if we can’t have that, I’ll settle for him being a broke pariah.

TutorialsDGX agent

Sam Altman should be in prison. But if we can’t have that, I’ll settle for him being a broke pariah. OpenAI hardware chief Tang Tan allegedly helped coach recruits on how to evade Apple’s data securit

Valve's new Steam Machine verification system is silent on these Steam Deck-busters

IndustryDGX agent

Valve has introduced new verification systems for its upcoming Steam Machine and Steam Frame hardware, detailed at GDC 2026 . Steam Machine Verified functions as a higher-performance extension of Stea

9 Jul 2026

Fast token generation emerges as the key differentiator as heterogeneous inference takes hold

Model ReleasesDGX agent

The race for fast token generation has moved from benchmark sheets into production data centers, and the hardware blueprint for winning it is no longer a GPU-only story. As agentic AI use cases multip

Quantum simulation of real-world nonlinear dynamics via Koopman method

ApplicationsDGX agent

arXiv:2607.07338v1 Announce Type: cross Abstract: Nonlinear dynamics is ubiquitous in nature, ranging from chemical pattern formation to ocean circulation, yet its simulation on quantum computers is f

Solve harder problems with AlphaEvolve, now available to everyone on Google Cloud

Model ReleasesDGX agent

Many of the most challenging and valuable problems in the world are related to optimization. Now, AI is now making these problems tractable. If you've ever tried to design a microchip, plan a delivery

8 Jul 2026

b9910

Local AiDGX agent

Based on available information, b9910 is a release tag from the llama.cpp project, an open-source C/C++ implementation for running large language model inference locally on consumer hardware. Llama.cp

b9923

Local AiDGX agent

b9923 is a release build of llama.cpp, an open-source C/C++ project for LLM inference with minimal setup and state-of-the-art performance on various hardware . The specific b9923 build includes binary

Binocular Gaze Estimation with Single Camera and Single Light Source

ResearchDGX agent

arXiv:2607.05473v1 Announce Type: cross Abstract: According to commonly consented theories, the minimum hardware requirement for gaze tracker is one camera and two light sources to realize gaze estima

Google Cloud named Leader in the 2026 Gartner® Magic Quadrant™ for AI Infrastructure

Model ReleasesDGX agent

In the agentic era, AI is evolving from answering questions to reasoning and taking action. Companies who want to lead in this next phase of AI need computing infrastructure that’s designed and optimi

Quantum startup Oratomic banks $300M to race straight to fault-tolerance

IndustryDGX agent

Oratomic Inc., a quantum computing startup, has raised 300 million in early-stage funding to scale its quantum hardware and accelerate its path to fault tolerance and error correction, creating the ne

What's new at IBM Quantum Q2 2026

ResearchDGX agent

IBM Quantum's Q2 2026 updates likely showcase advancements in quantum computing hardware, software, or services, including new processor capabilities, improved error correction, expanded cloud access,

7 Jul 2026

COMET: Combinatorial Optimization for Multiplex Editing Targets Via Constraint-Preserving QAOA

TutorialsDGX agent

arXiv:2607.02622v1 Announce Type: cross Abstract: Multiplex CRISPR-Cas9 gene editing requires selecting one guide RNA per target gene subject to cross-gene interactions: a constrained combinatorial pr

GelNeuro: A Sensing-Computing Integrated Neuromorphic Tactile System for Texture Recognition

Model ReleasesDGX agent

arXiv:2607.05241v1 Announce Type: new Abstract: Neuromorphic visuo-tactile sensing offers a promising paradigm for low-latency and low-power robotic perception. However, existing systems still rely he

LACE-SVD: Loss-Aware SVD with Cumulative Error Correction for LLM Compression

Model ReleasesDGX agent

arXiv:2607.03057v1 Announce Type: cross Abstract: The rapid growth in the parameter scale of large language models (LLMs) has created a strong demand for efficient compression techniques. As a hardwar

← Previous
1…3536373839…75
Next →