AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
All
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages a…

DGX agent

We created a document OCR router that can estimate the complexity of every single page and parse it with the relevant mode 💫 * Some pages are full of native text, which can be directly handled with Li

model-releasesjerry-liu--x
31 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

DGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

model-releasesjerry-liu--x
31 Jul 2026
Model Releases

We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

DGX agent

I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B. On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is

model-releasesr-localllama
31 Jul 2026
Model Releases

What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation

DGX agent

arXiv:2607.26935v1 Announce Type: new Abstract: Bot detectors deployed at scale treat traffic as binary: human or bot. This assumption breaks when AI agents browse the web through browser automation,

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

DGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

What’s new in AI infrastructure and orchestration this month

DGX agent

At Google, AI is a soup-to-nuts endeavor. Obviously, we make leading AI models like Gemini and Nano Banana. We incorporate AI into the tools you use every day (think Gmail, BigQuery, AlloyDB, Google C

model-releasesgoogle-cloud-ai
31 Jul 2026
Model Releases

What's your local AI coding setup on a MacBook Pro M4?

DGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

model-releasesr-ollama
31 Jul 2026
Model Releases

When Does Muon Help Agentic Reinforcement Learning?

DGX agent

arXiv:2607.16169v3 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its operating regime in reinforcement-learning post-training remains unclear. We map

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

WhisperRec: Latent Reasoning for Efficient Foundation Recommendation Models

DGX agent

arXiv:2607.26621v2 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong reasoning capabilities, motivating their adoption as backbones for foundation recommendation mod

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Why are AI model tests always the same generic prompts?

DGX agent

Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are 'make a car game,' 'make a website,' or something equally generic, usually from a prompt that's barely a lin

model-releasesr-localllama
31 Jul 2026
Model Releases

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

DGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Will ollama upgrade Deepseek V4 Flash on cloud?

DGX agent

https://preview.redd.it/vxl4zslewigh1.png?width=1435&format=png&auto=webp&s=5a419870ca0cb13076be6c9ff4ef33d8177d5eda New version is 25% better than previous one and is near GLM-5.2 quality submitted b

model-releasesr-ollama
31 Jul 2026
Model Releases

With release of Deepseek V4 I wanted see how the model sizes are trending over time. The trend is that by this time next year, we probably will have Opus 4.5 level models on consumer grade laptops!

DGX agent

I was surprised to see that Deepseek V4 Flash is extremely smart and small enough to fit in setup that can be built with < $50,000. Expensive, but not a datacenter. So I wanted to see the trend over t

model-releasesr-localllama
31 Jul 2026
Model Releases

Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning

DGX agent

arXiv:2607.28478v1 Announce Type: new Abstract: As large language models (LLMs) continue to advance in complex reasoning tasks, they have learned to heavily prioritize explicit conditions provided in

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

ZUNA1.1: A more flexible EEG foundation model for Denoising and Super-resolution

DGX agent

arXiv:2607.27308v1 Announce Type: new Abstract: We introduce ZUNA1.1, a 380M-parameter diffusion autoencoder for flexible EEG signal reconstruction. ZUNA1.1 is capable of reconstructing variable lengt

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

4090 + 5060 Ti + 64GB RAM: 206 t/s on a 35B-A3B, and a 122B at 37 t/s

DGX agent

I've been benchmarking a two-card box for a few weeks and I still can't quite get over some of these numbers, so I'm dumping them here. Box: RTX 4090 (24GB) + RTX 5060 Ti (16GB), i9-13900K, 64GB DDR5.

model-releasesr-localllama
30 Jul 2026
Model Releases

AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing

DGX agent

arXiv:2601.11007v2 Announce Type: replace-cross Abstract: LLM role-playing aims to portray arbitrary characters in interactive narratives, yet existing systems often suffer from limited immersion and

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Advancing the price-performance frontier with GPT‑5.6

DGX agent

Advancing the price-performance frontier with GPT‑5.6 Huge price drop from OpenAI today: GPT-5.6 Terra got a 20% reduction, and GPT-5.6 Luna got a massive 80% drop. OpenAI credit 5.6 Sol with enabling

model-releasessimon-willison
30 Jul 2026
Model Releases

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

DGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

model-releasesperplexity--x
30 Jul 2026
Model Releases

Aligning LLM-Simulated and Human Examinees for Psychometric Calibration: A Cognitive Diagnostic Profiling Approach

DGX agent

arXiv:2607.26317v1 Announce Type: cross Abstract: Psychometric calibration for educational tests typically requires costly human response data. Large language models (LLMs) simulated examinees offer a

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Amortized Moment Matching for Visual Generation

DGX agent

arXiv:2607.26860v1 Announce Type: new Abstract: We propose amortized moment matching, utilizing neural networks to learn data moments as distributional training signals. By casting diffusion denoisers

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

And Grok 4.6 comes out in a week

DGX agent

And Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl

model-releaseselon-musk--x
30 Jul 2026
Model Releases

Anthropic says it discovered three of its models had breached three organizations after launching a review in response to the OpenAI-Hugging Face incident (Anthropic)

DGX agent

Anthropic: Anthropic says it discovered three of its models had breached three organizations after launching a review in response to the OpenAI-Hugging Face incident — In a review of our cybersecurity

model-releasestechmeme
30 Jul 2026
Model Releases

APEX-Accounting

DGX agent

arXiv:2607.27189v1 Announce Type: new Abstract: We introduce APEX-Accounting, a benchmark built by Mercor in partnership with Ramp, to assess whether frontier models can do the real work of accountant

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

ARC-Encoder: learning compressed text representations for large language models

DGX agent

arXiv:2510.20535v2 Announce Type: replace Abstract: Recent techniques such as retrieval-augmented generation or chain-of-thought reasoning have led to longer contexts and increased inference costs. Co

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Archetypes or ability? Clustering for modelling student mathematical competence

DGX agent

arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672

DGX agent

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672 We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are

model-releasesswyx--x
30 Jul 2026
Model Releases

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

DGX agent

arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last b

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

b10184

DGX agent

mimo2: address MTP review feedback (#26228) Co-authored-by: tnhnyc 115956684+tnhnyc@users.noreply.github.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm6

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10186

DGX agent

ggml : Fix issue with kleidiai ci and stringop overflow warning (#26277) Signed-off-by: Jonathan Clohessy Jonathan.Clohessy@arm.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) ma

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10188

DGX agent

metal: fix memory unwire if model is freed without any GPU operations (#26082) metal: fix memory leak if model is freed without any GPU operations metal: run dummy work only if residency sets are used

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10189

DGX agent

Remove custom cpu op from the M3 graph, express with stock ops (#26297) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS I

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10192

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10194

DGX agent

ggml-cuda: Allow transpose-free gemmv computation (#26171) When matrix's weights are shaped 1xK is leverage a transpose-free computation to use mat_mul_vec_f. Website: https://llama.app macOS/iOS: mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10195

DGX agent

tests : avoid building get-model.cpp many times (#26317) tests : remove get-model.cpp tests : fix quant type selection Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Sil

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10196

DGX agent

llama-context : sync pending async copies before clearing embd_seq (#25676) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED mac

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10197

DGX agent

Test support for alternative conv layout (#25617) add bool cwhn = true to conv_2d test cases add layout check at graph building time extend layout checks for conv2d.cu kernel in CPU back-end kernel ne

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10198

DGX agent

vulkan: Support quantized concat (#25684) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Lin

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

b10199

DGX agent

server: support inp embd to generate next token (#26313) server: support embd for sampled token fix ~server_batch() Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silico

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

Batten Down Your Packages: Mitigation Guidance for Supply Chain Compromise

DGX agent

Written by: Kelli Vanderlee, Stuart Carrera For years, the cybersecurity industry's understanding of software supply chain compromise has been anchored by a few watershed events, including Russian cyb

model-releasesgoogle-cloud-ai
30 Jul 2026
Model Releases

BayesAME: Bayesian Active Model Evaluation

DGX agent

arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Benchmarked: MindControl for Llama.cpp

DGX agent

I recently shared the original MindControl PoC (and on github) - sampler-level guided reasoning budgets for llama.cpp, nudging the model with self-aware statements about its own thinking budget instea

model-releasesr-localllama
30 Jul 2026
Model Releases

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

DGX agent

arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization

DGX agent

arXiv:2607.26232v1 Announce Type: new Abstract: Background manipulation is a practical but under-specified image-forensics setting: the manipulated evidence can sit outside the salient foreground obje

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation

DGX agent

arXiv:2606.19651v2 Announce Type: replace-cross Abstract: Three-dimensional (3D) brain MRI is central to clinical neurology and neuro-oncology, where generative models could augment under-represented

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Budget-Aware LLM Discovery via Cost-Calibrated Frontier Utility

DGX agent

arXiv:2607.26828v1 Announce Type: new Abstract: Large language models increasingly support scientific and algorithmic discovery through inference-time search over evaluated candidates. Existing adapti

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Calibri: Enhancing Diffusion Transformers via Parameter-Efficient Calibration

DGX agent

arXiv:2603.24800v2 Announce Type: replace Abstract: In this paper, we uncover the hidden potential of Diffusion Transformers (DiTs) to significantly enhance generative tasks. Through an in-depth analy

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Choosing Where and How to Moderate: End-to-End Trade-offs in Filter Placement and Response Rewriting

DGX agent

arXiv:2607.26200v1 Announce Type: new Abstract: Content-moderation classifiers are usually evaluated in isolation, but deployment requires choosing where to intervene and what follows a flag. We evalu

model-releasesarxiv-cs-cl
30 Jul 2026
← Previous
1…6566676869…469
Next →