AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
Model Releases

b10189

DGX agent

Remove custom cpu op from the M3 graph, express with stock ops (#26297) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS I

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10192

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10194

DGX agent

ggml-cuda: Allow transpose-free gemmv computation (#26171) When matrix's weights are shaped 1xK is leverage a transpose-free computation to use mat_mul_vec_f. Website: https://llama.app macOS/iOS: mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10195

DGX agent

tests : avoid building get-model.cpp many times (#26317) tests : remove get-model.cpp tests : fix quant type selection Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Sil

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10196

DGX agent

llama-context : sync pending async copies before clearing embd_seq (#25676) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10197

DGX agent

Test support for alternative conv layout (#25617) add bool cwhn = true to conv_2d test cases add layout check at graph building time extend layout checks for conv2d.cu kernel in CPU back-end kernel ne

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10198

DGX agent

vulkan: Support quantized concat (#25684) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Lin

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10199

DGX agent

server: support inp embd to generate next token (#26313) server: support embd for sampled token fix ~server_batch() Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silico

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

Batten Down Your Packages: Mitigation Guidance for Supply Chain Compromise

DGX agent

Written by: Kelli Vanderlee, Stuart Carrera For years, the cybersecurity industry's understanding of software supply chain compromise has been anchored by a few watershed events, including Russian cyb

model-releasesgoogle-cloud-ai
30 Jul 2026
Model Releases

BayesAME: Bayesian Active Model Evaluation

DGX agent

arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Benchmarked: MindControl for Llama.cpp

DGX agent

I recently shared the original MindControl PoC (and on github) - sampler-level guided reasoning budgets for llama.cpp, nudging the model with self-aware statements about its own thinking budget instea

model-releasesr-localllama
30 Jul 2026
Model Releases

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

DGX agent

arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization

DGX agent

arXiv:2607.26232v1 Announce Type: new Abstract: Background manipulation is a practical but under-specified image-forensics setting: the manipulated evidence can sit outside the salient foreground obje

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation

DGX agent

arXiv:2606.19651v2 Announce Type: replace-cross Abstract: Three-dimensional (3D) brain MRI is central to clinical neurology and neuro-oncology, where generative models could augment under-represented

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility

DGX agent

arXiv:2607.26828v1 Announce Type: new Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adapti

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Calibri: Enhancing Diffusion Transformers via Parameter-Efficient Calibration

DGX agent

arXiv:2603.24800v2 Announce Type: replace Abstract: In this paper, we uncover the hidden potential of Diffusion Transformers (DiTs) to significantly enhance generative tasks. Through an in-depth analy

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Choosing Where and How to Moderate: End-to-End Trade-offs in Filter Placement and Response Rewriting

DGX agent

arXiv:2607.26200v1 Announce Type: new Abstract: Content-moderation classifiers are usually evaluated in isolation, but deployment requires choosing where to intervene and what follows a flag. We evalu

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

CMT-RAG: Complementary Memory Traces for Multi-turn Multi-hop RAG

DGX agent

arXiv:2607.26470v1 Announce Type: new Abstract: Multi-turn information-seeking conversations require both multi-hop reasoning and long-range dependency tracking across turns. However, existing RAG sys

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

CoCaRS: Correlation Calibration-Based Redundancy Suppression for Heterogeneous Knowledge Distillation

DGX agent

arXiv:2607.27054v1 Announce Type: new Abstract: Knowledge distillation (KD) enables a compact student model to learn from a powerful teacher and has become an effective paradigm for model compression.

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyU…

DGX agent

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyUI - combining Florence 2, SAM2, and WAN Video into a single

model-releasescomfyui--x
30 Jul 2026
Model Releases

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success …

DGX agent

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success rate. - Unoptimized? - Buggy? - Deliberate (coz that helps i

model-releasessebastian-raschka--x
30 Jul 2026
Model Releases

Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark

DGX agent

arXiv:2607.27143v1 Announce Type: new Abstract: High-stakes decision systems in credit scoring, fraud detection, healthcare, and industrial safety require reliable uncertainty quantification under sev

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Credit Cards, Confusion, Computation, and Consequences: What Can We Uncover About Language Model Reasoning?

DGX agent

arXiv:2607.26952v1 Announce Type: new Abstract: We introduce CreditCardQA, the first financial literacy benchmark for numerical reasoning derived from real credit card agreements. The dataset contains

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Crossing-Free Probabilistic K-Line Forecasts Without Retraining

DGX agent

arXiv:2607.26792v1 Announce Type: cross Abstract: Probabilistic K-line forecasting describes uncertainty in four complementary prices, namely open--high--low--close (OHLC). However, it introduces two

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Defending Against Backdoor Attacks via Alignment Checking in Model-Contrastive Federated Learning

DGX agent

arXiv:2607.26933v1 Announce Type: cross Abstract: Federated Learning (FL) is vulnerable to backdoor attacks because of its distributed nature in edge computing scenarios. Existing defense methods show

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

DenseOn with the LateOn: Fully Open Dense and Late-Interaction Models for Multilingual, Long-Context, and Code Search

DGX agent

arXiv:2607.27178v1 Announce Type: new Abstract: State-of-the-art retrieval models increasingly rely on closed training data, creating a reproducibility gap. We present an open end-to-end recipe for tr

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Diagnosing Fine-Grained Inconsistency Classification in Financial Disclosure Text

DGX agent

arXiv:2607.26368v1 Announce Type: new Abstract: Financial disclosures contain numerical claims, temporal statements, entity references, policy commitments, and risk descriptions that may conflict in q

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English

DGX agent

arXiv:2601.22888v4 Announce Type: replace Abstract: More than 80% of the 1.6B English speakers do not use Standard American English (SAE), yet LLMs often fail to correctly identify non-SAE dialects an

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Does MTP head get loaded in VRAM by default?

DGX agent

I ran into a doubt when using the following command. It seems that the System RAM usage keeps increasing even though there is >10GB of space left in VRAM while using the MTP mode. Does the MTP head lo

model-releasesr-localllama
30 Jul 2026
Model Releases

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding

DGX agent

arXiv:2607.26518v1 Announce Type: new Abstract: Reliable visual safety understanding in real-world scenarios demands more than just object recognition; it requires causal reasoning under epistemic unc

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Enhancing Automated Machine Learning via Homogeneous Train-Test Splitting Methods

DGX agent

arXiv:2607.26625v1 Announce Type: new Abstract: Accurate model evaluation in machine learning depends critically on how datasets are split into training and testing subsets. Standard random splitting

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Enhancing Generative Information Extraction with Two-step Validation: A Product Attribute Use Case

DGX agent

arXiv:2607.26780v1 Announce Type: new Abstract: The ability of large language models (LLMs) to process and generate text has introduced potential for applications in information extraction (IE). While

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces

DGX agent

arXiv:2505.16035v3 Announce Type: replace Abstract: We introduce Equivariant Neural Eikonal Solvers, a novel framework that integrates Equivariant Neural Fields (ENFs) with Neural Eikonal Solvers. Our

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Evaluating Prompt Scope and Demonstration Similarity in Local LLM Machine Translation

DGX agent

arXiv:2607.26286v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as general-purpose translation systems, but their behavior is usually evaluated under a single prompt

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks

DGX agent

arXiv:2607.26490v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

exttt{AMEND++}: Benchmarking Eligibility Criteria Amendments in Clinical Trials

DGX agent

arXiv:2601.06300v2 Announce Type: replace Abstract: Clinical trial amendments frequently introduce delays, increased costs, and administrative burden, with eligibility criteria being the most commonly

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

FAS-R1: A Unified Multi-Task MLLM for Reasoning Face Anti-Spoofing

DGX agent

arXiv:2607.26432v1 Announce Type: new Abstract: Face anti-spoofing (FAS) is increasingly expected to provide not only bona fide/spoof decisions, but also attack semantics and image-grounded evidence f

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

FedTopo: Relation-Level Topology Sharing for Model-Heterogeneous Federated Learning

DGX agent

arXiv:2607.26801v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative learning over decentralized data silos without centralizing raw data. However, heterogeneous local archite

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

DGX agent

arXiv:2607.26618v1 Announce Type: cross Abstract: Federated PEFT enables LLMs to collaboratively adapt to decentralized private data without sharing raw examples. However, task heterogeneity across cl

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Financial Volatility and Risk Forecasting Incorporating a Larger Number of Realized Measures

DGX agent

arXiv:2411.17136v2 Announce Type: replace-cross Abstract: Realised volatility has become increasingly prominent in volatility forecasting due to its ability to capture intraday price fluctuations. Wit

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Flow Map Learning via Nongradient Vector Flow

DGX agent

arXiv:2607.26398v1 Announce Type: new Abstract: Diffusion and flow-based models benefit from simple regression losses, but inference incurs significant overhead because sampling requires integration.

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making th…

DGX agent

For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @Goog

model-releasesgoogle-ai--x
30 Jul 2026
Model Releases

ForgetBench: Benchmarking Forgetting Dynamics of Long-Term Parametric Memory in Language Models

DGX agent

arXiv:2607.26455v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong capabilities in knowledge acquisition and reasoning, yet their ability to retain previously acquir

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Foundation Models for Face Presentation Attack Detection: A Unified Linear-Probing Benchmark

DGX agent

arXiv:2607.26993v1 Announce Type: new Abstract: Face presentation attack detection (PAD) remains challenging under cross-dataset evaluation, where domain shift degrades models trained on a single data

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

From Tokens to Watt-hours: Analytical Energy Estimation for LLM Inference on Modern GPUs

DGX agent

arXiv:2607.26571v1 Announce Type: new Abstract: The operational energy consumption of large language model (LLM) inference is becoming an increasingly important component of the environmental footprin

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Gated Adaptation for Continual Learning in Human Activity Recognition

DGX agent

arXiv:2603.10046v2 Announce Type: replace Abstract: Wearable sensors in Internet of Things (IoT) ecosystems increasingly support applications such as remote health monitoring, elderly care, and smart

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration

DGX agent

Gemini Robotics ER 2 helps robots reason, collaborate, and solve real-world tasks. It represents a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic app

model-releasesgoogle-deepmind
30 Jul 2026
← Previous
1…6162636465…466
Next →