AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
Model Releases

APEX-Accounting

DGX agent

arXiv:2607.27189v1 Announce Type: new Abstract: We introduce APEX-Accounting, a benchmark built by Mercor in partnership with Ramp, to assess whether frontier models can do the real work of accountant

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

ARC-Encoder: learning compressed text representations for large language models

DGX agent

arXiv:2510.20535v2 Announce Type: replace Abstract: Recent techniques such as retrieval-augmented generation or chain-of-thought reasoning have led to longer contexts and increased inference costs. Co

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Archetypes or ability? Clustering for modelling student mathematical competence

DGX agent

arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672

DGX agent

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672 We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are

model-releasesswyx--x
30 Jul 2026
Model Releases

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

DGX agent

arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last b

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

b10184

DGX agent

mimo2: address MTP review feedback (#26228) Co-authored-by: tnhnyc 115956684+tnhnyc@users.noreply.github.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm6

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10186

DGX agent

ggml : Fix issue with kleidiai ci and stringop overflow warning (#26277) Signed-off-by: Jonathan Clohessy Jonathan.Clohessy@arm.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) ma

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10188

DGX agent

metal: fix memory unwire if model is freed without any GPU operations (#26082) metal: fix memory leak if model is freed without any GPU operations metal: run dummy work only if residency sets are used

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10189

DGX agent

Remove custom cpu op from the M3 graph, express with stock ops (#26297) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS I

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10192

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10194

DGX agent

ggml-cuda: Allow transpose-free gemmv computation (#26171) When matrix's weights are shaped 1xK is leverage a transpose-free computation to use mat_mul_vec_f. Website: https://llama.app macOS/iOS: mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10195

DGX agent

tests : avoid building get-model.cpp many times (#26317) tests : remove get-model.cpp tests : fix quant type selection Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Sil

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10196

DGX agent

llama-context : sync pending async copies before clearing embd_seq (#25676) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10197

DGX agent

Test support for alternative conv layout (#25617) add bool cwhn = true to conv_2d test cases add layout check at graph building time extend layout checks for conv2d.cu kernel in CPU back-end kernel ne

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10198

DGX agent

vulkan: Support quantized concat (#25684) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Lin

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10199

DGX agent

server: support inp embd to generate next token (#26313) server: support embd for sampled token fix ~server_batch() Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silico

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

Batten Down Your Packages: Mitigation Guidance for Supply Chain Compromise

DGX agent

Written by: Kelli Vanderlee, Stuart Carrera For years, the cybersecurity industry's understanding of software supply chain compromise has been anchored by a few watershed events, including Russian cyb

model-releasesgoogle-cloud-ai
30 Jul 2026
Model Releases

BayesAME: Bayesian Active Model Evaluation

DGX agent

arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Benchmarked: MindControl for Llama.cpp

DGX agent

I recently shared the original MindControl PoC (and on github) - sampler-level guided reasoning budgets for llama.cpp, nudging the model with self-aware statements about its own thinking budget instea

model-releasesr-localllama
30 Jul 2026
Model Releases

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

DGX agent

arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization

DGX agent

arXiv:2607.26232v1 Announce Type: new Abstract: Background manipulation is a practical but under-specified image-forensics setting: the manipulated evidence can sit outside the salient foreground obje

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation

DGX agent

arXiv:2606.19651v2 Announce Type: replace-cross Abstract: Three-dimensional (3D) brain MRI is central to clinical neurology and neuro-oncology, where generative models could augment under-represented

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility

DGX agent

arXiv:2607.26828v1 Announce Type: new Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adapti

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Calibri: Enhancing Diffusion Transformers via Parameter-Efficient Calibration

DGX agent

arXiv:2603.24800v2 Announce Type: replace Abstract: In this paper, we uncover the hidden potential of Diffusion Transformers (DiTs) to significantly enhance generative tasks. Through an in-depth analy

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Choosing Where and How to Moderate: End-to-End Trade-offs in Filter Placement and Response Rewriting

DGX agent

arXiv:2607.26200v1 Announce Type: new Abstract: Content-moderation classifiers are usually evaluated in isolation, but deployment requires choosing where to intervene and what follows a flag. We evalu

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

CMT-RAG: Complementary Memory Traces for Multi-turn Multi-hop RAG

DGX agent

arXiv:2607.26470v1 Announce Type: new Abstract: Multi-turn information-seeking conversations require both multi-hop reasoning and long-range dependency tracking across turns. However, existing RAG sys

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

CoCaRS: Correlation Calibration-Based Redundancy Suppression for Heterogeneous Knowledge Distillation

DGX agent

arXiv:2607.27054v1 Announce Type: new Abstract: Knowledge distillation (KD) enables a compact student model to learn from a powerful teacher and has become an effective paradigm for model compression.

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyU…

DGX agent

ComfyUI Face Swap Tutorial: Fast, Clean Results VFX artist Doug Hogan walks through a fully automated face swap workflow built inside ComfyUI - combining Florence 2, SAM2, and WAN Video into a single

model-releasescomfyui--x
30 Jul 2026
Model Releases

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success …

DGX agent

Consistent with what I found with Qwen3.6 a while back: Claude Code uses 2-3x as many tokens than (many) other harnesses at similar success rate. - Unoptimized? - Buggy? - Deliberate (coz that helps i

model-releasessebastian-raschka--x
30 Jul 2026
Model Releases

Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark

DGX agent

arXiv:2607.27143v1 Announce Type: new Abstract: High-stakes decision systems in credit scoring, fraud detection, healthcare, and industrial safety require reliable uncertainty quantification under sev

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Credit Cards, Confusion, Computation, and Consequences: What Can We Uncover About Language Model Reasoning?

DGX agent

arXiv:2607.26952v1 Announce Type: new Abstract: We introduce CreditCardQA, the first financial literacy benchmark for numerical reasoning derived from real credit card agreements. The dataset contains

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Crossing-Free Probabilistic K-Line Forecasts Without Retraining

DGX agent

arXiv:2607.26792v1 Announce Type: cross Abstract: Probabilistic K-line forecasting describes uncertainty in four complementary prices, namely open--high--low--close (OHLC). However, it introduces two

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Defending Against Backdoor Attacks via Alignment Checking in Model-Contrastive Federated Learning

DGX agent

arXiv:2607.26933v1 Announce Type: cross Abstract: Federated Learning (FL) is vulnerable to backdoor attacks because of its distributed nature in edge computing scenarios. Existing defense methods show

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

DenseOn with the LateOn: Fully Open Dense and Late-Interaction Models for Multilingual, Long-Context, and Code Search

DGX agent

arXiv:2607.27178v1 Announce Type: new Abstract: State-of-the-art retrieval models increasingly rely on closed training data, creating a reproducibility gap. We present an open end-to-end recipe for tr

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Diagnosing Fine-Grained Inconsistency Classification in Financial Disclosure Text

DGX agent

arXiv:2607.26368v1 Announce Type: new Abstract: Financial disclosures contain numerical claims, temporal statements, entity references, policy commitments, and risk descriptions that may conflict in q

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English

DGX agent

arXiv:2601.22888v4 Announce Type: replace Abstract: More than 80% of the 1.6B English speakers do not use Standard American English (SAE), yet LLMs often fail to correctly identify non-SAE dialects an

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Does MTP head get loaded in VRAM by default?

DGX agent

I ran into a doubt when using the following command. It seems that the System RAM usage keeps increasing even though there is >10GB of space left in VRAM while using the MTP mode. Does the MTP head lo

model-releasesr-localllama
30 Jul 2026
Model Releases

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding

DGX agent

arXiv:2607.26518v1 Announce Type: new Abstract: Reliable visual safety understanding in real-world scenarios demands more than just object recognition; it requires causal reasoning under epistemic unc

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Enhancing Automated Machine Learning via Homogeneous Train-Test Splitting Methods

DGX agent

arXiv:2607.26625v1 Announce Type: new Abstract: Accurate model evaluation in machine learning depends critically on how datasets are split into training and testing subsets. Standard random splitting

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Enhancing Generative Information Extraction with Two-step Validation: A Product Attribute Use Case

DGX agent

arXiv:2607.26780v1 Announce Type: new Abstract: The ability of large language models (LLMs) to process and generate text has introduced potential for applications in information extraction (IE). While

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces

DGX agent

arXiv:2505.16035v3 Announce Type: replace Abstract: We introduce Equivariant Neural Eikonal Solvers, a novel framework that integrates Equivariant Neural Fields (ENFs) with Neural Eikonal Solvers. Our

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Evaluating Prompt Scope and Demonstration Similarity in Local LLM Machine Translation

DGX agent

arXiv:2607.26286v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as general-purpose translation systems, but their behavior is usually evaluated under a single prompt

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks

DGX agent

arXiv:2607.26490v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

exttt{AMEND++}: Benchmarking Eligibility Criteria Amendments in Clinical Trials

DGX agent

arXiv:2601.06300v2 Announce Type: replace Abstract: Clinical trial amendments frequently introduce delays, increased costs, and administrative burden, with eligibility criteria being the most commonly

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

FAS-R1: A Unified Multi-Task MLLM for Reasoning Face Anti-Spoofing

DGX agent

arXiv:2607.26432v1 Announce Type: new Abstract: Face anti-spoofing (FAS) is increasingly expected to provide not only bona fide/spoof decisions, but also attack semantics and image-grounded evidence f

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

FedTopo: Relation-Level Topology Sharing for Model-Heterogeneous Federated Learning

DGX agent

arXiv:2607.26801v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative learning over decentralized data silos without centralizing raw data. However, heterogeneous local archite

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

DGX agent

arXiv:2607.26618v1 Announce Type: cross Abstract: Federated PEFT enables LLMs to collaboratively adapt to decentralized private data without sharing raw examples. However, task heterogeneity across cl

model-releasesarxiv-cs-cl
30 Jul 2026
← Previous
1…6263646566…466
Next →