AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
4 May 2026

Quantum Gradient-Based Approach for Edge and Corner Detection Using Sobel Kernels

ResearchDGX agent

arXiv:2605.00744v1 Announce Type: new Abstract: Edge detection refers to identifying points in a digital image where intensity changes sharply, indicating object boundaries or structural features. Cor

Tempus: A Temporally Scalable Resource-Invariant GEMM Streaming Framework for Versal AI Edge

Local AiDGX agent

arXiv:2605.00536v1 Announce Type: cross Abstract: Scaling laws for Large Language Models (LLMs) establish that model quality improves with computational scale, yet edge deployment imposes strict const

Would a 2nd hand custom built 2080 Ti 22GB vram be worth it? How usable would it be with ComfyUI? Or maybe even 2 pcs with NVLink? Can large models, like Wan2.2 be split with NVLink like it's 44GB, or will it always be 22GB for 1 model, and 22GB for another model (encoders, CLIP, anything else)?

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local AiDGX agent

This Reddit post discusses the viability of using second-hand custom-built RTX 2080 Ti GPUs with 22GB VRAM for running Stable Diffusion models in ComfyUI, including whether two cards could be linked v

3 May 2026

b9012

Local AiDGX agent

b9012 is a build version from the llama.cpp GitHub repository, which is an LLM inference implementation in C/C++ . The release likely contains updates, bug fixes, and improvements to the llama.cpp inf

Best Local Vision-Language Models?

Local AiDGX agent

This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques

Reachy-Mini under assembly. ⁦@huggingface⁩

IndustryDGX agent

Reachy-Mini is a compact robotic arm under development or assembly, likely a smaller version of the Reachy robot platform. The post from Hugging Face's leadership suggests this is part of their roboti

2 May 2026

b9004

Local AiDGX agent

llama.cpp is an LLM inference framework in C/C++ that enables efficient local execution of large language models. Build b9004 is a recent release with optimizations and improvements for supporting var

b9008

Local AiDGX agent

B9008 is a build release of llama.cpp from May 2, 2026. Llama.cpp is a C/C++ implementation of LLM inference designed to enable large language model inference with minimal setup and high performance a

b9009

Local AiDGX agent

Build b9009 is an incremental release of llama.cpp from May 2026, continuing the rapid development cycle that characterized April 2026's updates with tensor parallelism, 1-bit quantization, and expand

b9010

Local AiDGX agent

b9010 is a build release of llama.cpp, an open-source project for LLM inference in C/C++ . As a numbered build tag in the llama.cpp release system, it represents a specific development build containin

1 May 2026

Adaptive Nonlinear MPC for Trajectory Tracking of An Overactuated Tiltrotor Hexacopter

ResearchDGX agent

arXiv:2211.06762v2 Announce Type: replace Abstract: Omnidirectional micro aerial vehicles (OMAVs) are more capable of doing environmentally interactive tasks due to their ability to exert full wrenche

AutoSP: Unlocking Long-Context LLM Training Via Compiler-Based Sequence Parallelism

Model ReleasesDGX agent

arXiv:2604.27089v1 Announce Type: new Abstract: Large-language-models (LLMs) demonstrate enormous utility in long-context tasks which require processing prompts that consist of tens to hundreds of tho

End-to-end autonomous scientific discovery on a real optical platform

AgentsDGX agent

arXiv:2604.27092v1 Announce Type: new Abstract: Scientific research has long been human-led, driving new knowledge and transformative technologies through the continual revision of questions, methods

Focus Session: Autonomous Systems Dependability in the era of AI: Design Challenges in Safety, Security, Reliability and Certification

SafetyDGX agent

arXiv:2604.27807v1 Announce Type: new Abstract: The design of embedded safety-critical systems such as those used in next-generation automotive and autonomous platforms, is increasingly challenged by

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

Model ReleasesDGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

HAVEN: Hybrid Automated Verification ENgine for UVM Testbench Synthesis with LLMs

ResearchDGX agent

arXiv:2604.27643v1 Announce Type: cross Abstract: Integrated Circuit (IC) verification consumes nearly 70% of the IC development cycle, and recent research leverages Large Language Models (LLMs) to au

HQ-UNet: A Hybrid Quantum-Classical U-Net with a Quantum Bottleneck for Remote Sensing Image Segmentation

Model ReleasesDGX agent

arXiv:2604.27206v1 Announce Type: new Abstract: Semantic segmentation in remote sensing is commonly addressed using classical deep learning architectures such as U-Net, which require a large number of

LLM-Guided Runtime Parameter Optimization for Energy-Efficient Model Inference

Model ReleasesDGX agent

arXiv:2604.27032v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of many real-world workflows. However, LLMs consume a lot of energy, which becomes a large c

Multibit neural inference in a N-ary crossbar architecture

ResearchDGX agent

arXiv:2604.26979v1 Announce Type: cross Abstract: In-memory computing (IMC) enables energy-efficient neural network inference by computing analog matrix-vector multiplications (MVM) in memory crossbar

my car just warned me that my eyes were closed when they weren’t so now i’m a little offended but whatever

AgentsDGX agent

A user reported experiencing a false positive from their vehicle's driver monitoring system, which incorrectly detected their eyes as closed when they were actually open, prompting a humorous reaction

OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction

SafetyDGX agent

arXiv:2604.28197v1 Announce Type: cross Abstract: Human-robot collaboration has been studied primarily in dyadic or sequential settings. However, real homes require multiadic collaboration, where mult

On the Predictive Skill of Artificial Intelligence-based Weather Models for Extreme Events using Uncertainty Quantification

ResearchDGX agent

arXiv:2511.17176v2 Announce Type: replace-cross Abstract: Accurate prediction of extreme weather events remains a major challenge for artificial intelligence-based weather prediction systems. While de

ORFS-agent: Tool-Using Agents for Chip Design Optimization

Model ReleasesDGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

Physically-Informed Fuzzy Clustering of Vertical Sounding Ionograms

ResearchDGX agent

arXiv:2604.27721v1 Announce Type: cross Abstract: This paper presents a physically-informed fuzzy clustering of vertical sounding ionograms for automatically separating the ionogram into tracks suitab

Real-Time GPU-Accelerated Monte Carlo Evaluation of Safety-Critical AEB Systems Under Uncertainty

SafetyDGX agent

arXiv:2604.27193v1 Announce Type: new Abstract: Automatic Emergency Braking (AEB) systems represent a safety-critical national interest, with the National Highway Traffic Safety Administration (NHTSA)

Seasoned dev but new to local LLMs: help me pick the right Apple product for hosting model in the 27B - 36B size

Local AiDGX agent

A seasoned developer seeking advice on selecting an Apple product to locally host large language models in the 27-36 billion parameter range, discussing the trade-offs and specifications of different

Transformer-Empowered Actor-Critic Reinforcement Learning for Sequence-Aware Service Function Chain Partitioning

ResearchDGX agent

arXiv:2504.18902v2 Announce Type: replace-cross Abstract: In the forthcoming era of 6G networks, characterized by unprecedented data rates, ultra-low latency, and ubiquitous connectivity, effective ma

30 Apr 2026

b8982

Local AiDGX agent

The search did not return specific details about release b8982. Based on the llama.cpp project context from the search results, b8982 is a build/release version from the llama.cpp repository that foll

b8983

Local AiDGX agent

b8983 is a release of llama.cpp, a C/C++ implementation for LLM inference. The release represents a specific build/commit version in the active development of the llama.cpp project, which supports run

b8984

Local AiDGX agent

B8984 is a build/release version of llama.cpp, an open-source C/C++ library for large language model inference. Llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a

b8987

Local AiDGX agent

Release b8987 of llama.cpp updates cpp-httplib to version 0.43.2 . This is a minor maintenance release from the llama.cpp project, which provides C/C++ inference for large language models. The release

b8992

Local AiDGX agent

b8992 is a build release from the llama.cpp project, which publishes multiple releases in a single day . llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a wide r

Cloud CISO Perspectives: At Next ‘26, why we’re multicloud and multi-AI

Model ReleasesDGX agent

Welcome to the second Cloud CISO Perspectives for April 2026. Today, Francis deSouza, COO Google Cloud and President, Security Products, explains why Google is multicloud and multi-AI, straight from N

Color-Encoded Illumination for High-Speed Volumetric Scene Reconstruction

ApplicationsDGX agent

arXiv:2604.26920v1 Announce Type: new Abstract: The task of capturing and rendering 3D dynamic scenes from 2D images has become increasingly popular in recent years. However, most conventional cameras

FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility

SafetyDGX agent

arXiv:2510.26841v2 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training without data sharing, yet participants face a fundamental challenge, e.g., simult

Folding Tensor and Sequence Parallelism for Memory-Efficient Transformer Training & Inference

Model ReleasesDGX agent

arXiv:2604.26294v1 Announce Type: new Abstract: We present tensor and sequence parallelism (TSP), a parallel execution strategy that folds tensor parallelism and sequence parallelism onto a single dev

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

Model ReleasesDGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

Star-Fusion: A Multi-modal Transformer Architecture for Discrete Celestial Orientation via Spherical Topology

AgentsDGX agent

arXiv:2604.26582v1 Announce Type: cross Abstract: Reliable celestial attitude determination is a critical requirement for autonomous spacecraft navigation, yet traditional 'Lost-in-Space' (LIS) algori

The Serial Scaling Hypothesis

ResearchDGX agent

arXiv:2507.12549v4 Announce Type: replace Abstract: While machine learning has advanced through massive parallelization, we identify a critical blind spot: some problems are fundamentally sequential.

These folks are trying to ban open source. They're looking to take away your freedom to choose. They're also looking to take away the rights…

ResearchDGX agent

These folks are trying to ban open source. They're looking to take away your freedom to choose. They're also looking to take away the rights of businesses like Cursor to fine tune and make their produ

Vertex Features for Neural Global Illumination

ResearchDGX agent

arXiv:2508.07852v2 Announce Type: replace-cross Abstract: Recent research on learnable neural representations has been widely adopted in the field of 3D scene reconstruction and neural rendering appli

29 Apr 2026

b8969

Local AiDGX agent

b8969 is a release build of llama.cpp (commit bdc9c74), published on April 29, 2026. Llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art perfor

b8972

Local AiDGX agent

b8972 is a release of llama.cpp, a C/C++ implementation for LLM inference . The release identifier follows llama.cpp's development versioning scheme with frequent builds published to the GitHub releas

b8973

Local AiDGX agent

b8973 is a release of llama.cpp that added SVE tuned code for the gemm_q8_0_4x8_q8_0() kernel and changed arrays to static const in repack.cpp . The llama.cpp project enables LLM inference with minima

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

Model ReleasesDGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

Model ReleasesDGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

Building the compute infrastructure for the Intelligence Age

Model ReleasesDGX agent

OpenAI outlines the computational infrastructure requirements and strategies necessary to support advanced AI systems in the emerging Intelligence Age. The piece likely discusses scaling challenges, h

Data-Driven Hamiltonian Reduction for Superconducting Qubits via Meta-Learning

Model ReleasesDGX agent

arXiv:2604.24912v1 Announce Type: cross Abstract: We introduce HAML (Hamiltonian Adaptation via Meta-Learning), a framework for fast online adaptation of effective Hamiltonian models of superconductin

Feasible-First Exploration for Constrained ML Deployment Optimization in Crash-Prone Hierarchical Search Spaces

Model ReleasesDGX agent

arXiv:2604.25073v1 Announce Type: new Abstract: Deploying machine learning models under production constraints requires joint optimization over model family, quantization scheme, runtime backend, and

Filing: TSMC sold its remaining 1.11M Arm shares at 207.65 per share for a total of ~231M this week; TSMC invested ~100M in Arm's 2023 IPO at 51 per share (Wen-Yee Lee/Reuters)

ApplicationsDGX agent

Wen-Yee Lee / Reuters: Filing: TSMC sold its remaining 1.11M Arm shares at 207.65 per share for a total of ~231M this week; TSMC invested ~100M in Arm's 2023 IPO at 51 per share — Taiwan Semiconductor

Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver

Model ReleasesDGX agent

arXiv:2604.25067v1 Announce Type: cross Abstract: Forecasting when AI systems will become capable of meaningfully accelerating AI research is a central challenge for AI safety. Existing benchmarks mea

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

Model ReleasesDGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpos…

Model ReleasesDGX agent

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpose-built for agentic AI on your personal devices. 💡Key insights

MiMo-V2.5-GGUF (preview available)

Local AiDGX agent

MiMo-V2.5 is Xiaomi's multimodal AI model with native visual and audio understanding that supports up to 1 million tokens of context. The GGUF format refers to quantized versions of the model optimize

NimbleReg: A light-weight deep-learning framework for diffeomorphic image registration

SafetyDGX agent

arXiv:2503.07768v2 Announce Type: replace Abstract: This paper presents NimbleReg, a light-weight deep-learning (DL) framework for diffeomorphic image registration leveraging surface representation of

Our cofounder @the_bunny_chen joined @GregorVand to talk about open models, production inference, and how RFT is unlocking model customizati…

ApplicationsDGX agent

Our cofounder @the_bunny_chen joined @GregorVand to talk about open models, production inference, and how RFT is unlocking model customization for teams without a dedicated ML org. They covered custom

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

Model ReleasesDGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

QB-LIF: Learnable-Scale Quantized Burst Neurons for Efficient SNNs

Model ReleasesDGX agent

arXiv:2604.25688v1 Announce Type: new Abstract: Binary spike coding enables sparse and event-driven computation in spiking neural networks (SNNs), yet its 1-bit-per-timestep representation fundamental

Qwen3.6 27B on dual RTX 5060 Ti 16GB with vLLM: ~60 tok/s, 204k context working

Local AiDGX agent

Qwen3.6 27B is a 27-billion parameter language model that can achieve approximately 60 tokens per second throughput when running on dual RTX 5060 Ti GPUs with 16GB memory each, using the vLLM inferenc

← Previous
1…6667686970…75
Next →