AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
49+ results
Research

Failure to Reproduce Modern Paper Claims [D]

DGX agent

This r/MachineLearning discussion thread addresses the widespread challenge of reproducing results claimed in modern ML research papers, a topic of significant concern in the field. Community members

researchr-machinelearning
15 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NeurIPS 2026 Reviewer: AI-Generated Rebuttals (and Paper) [D]

DGX agent

One of the papers I reviewed has what seems to be entirely LLM-generated rebuttals, and the original paper is also clearly LLM-generated, with Claude-speak everywhere. While the authors acknowledge LL

model-releasesr-machinelearning
28 Jul 2026
Research

Tri-Net v2: Open-source implementation of our Scientific Reports paper on unified skin lesion and symptom-based monkeypox detection [R]

DGX agent

Hi everyone, We've open-sourced Tri-Net v2, the official implementation accompanying our recently published Scientific Reports (Nature Portfolio) paper: 'Tri-Net: Unified Deep Learning for Skin Lesion

researchr-machinelearning
21 Jul 2026
Research

Call for Papers - Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]

DGX agent

This call for papers announces a workshop focused on the growing need for efficient and effective techniques for editing trained models, especially large generative models. The workshop solicits paper

researchr-machinelearning
25 May 2026
Research

What is the criteria for a ML paper to be published?[D]

DGX agent

This r/MachineLearning discussion thread explores the standards and expectations reviewers and program committees use when evaluating ML papers for acceptance at top venues such as NeurIPS, ICML, and

researchr-machinelearning
15 Apr 2026
Model Releases

[Paper] Statistically-Lossless Quantization of Large Language Models

DGX agent

Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as GPTQ and AWQ achieve practical compression but

model-releasesr-localllama
24 Jul 2026
Research

[ICML 2026] Scores for Position papers post discussion? [D]

DGX agent

This r/MachineLearning discussion thread focuses on the post-discussion reviewer scores for ICML 2026's dedicated Position Paper Track, where authors and community members share and compare their revi

researchr-machinelearning
13 Apr 2026
Local Ai

Built a fully-local paper-RAG across 2× 1080 Ti + a 3090. Three Ollama gotchas that each cost me a day.

DGX agent

A developer documented their experience building a fully-local paper Retrieval-Augmented Generation (RAG) system using two NVIDIA GTX 1080 Ti GPUs and one RTX 3090, sharing three significant challenge

local-air-ollama
6 Jun 2026
Model Releases

DeepSeek V4 paper full version is out, FP4 QAT details and stability tricks [D]

DGX agent

DeepSeek released the full technical report for DeepSeek-V4 on April 24, 2026, titled 'DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence.' The paper details FP4 quantization-awa

model-releasesr-machinelearning
9 May 2026
Research

Was looking at a ICLR 2025 Oral paper and I am shocked it got oral [D]

DGX agent

This Reddit thread from r/MachineLearning reflects community skepticism about the peer review standards at top ML conferences, with a user expressing surprise that a particular paper received an oral

researchr-machinelearning
15 Apr 2026
Model Releases

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

DGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

model-releasesr-localllama
28 Jul 2026
Model Releases

How many people in this sub try to train their own AI from scratch on their systems just for fun and to test out techniques from research papers?

DGX agent

As for me, I own a system with an RTX 5090, Ryzen 9 9950X3D2, and 64 GB of DDR5. Every time I see research come out with a new way to train AI, I immediately think to try it on my system to see the re

model-releasesr-localllama
6 Aug 2026
Local Ai

[Paper] RecGPT-V3 Technical Report

DGX agent

Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent that drives it. RecGPT-V1 pioneered this

local-air-localllama
26 Jul 2026
Model Releases

[Paper] SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including severe memory pressure, non-overlappe

model-releasesr-localllama
23 Jul 2026
Safety

PNAS: Over Half of All Academic Articles Now Show LLM Influence—7.3M-Paper Study [R]

DGX agent

Largest empirical study of AI penetration in academic publishing ever conducted—51%-by-2025 is the most authoritative quantitative marker yet of how thoroughly LLMs have reshaped scientific writing, a

safetyr-machinelearning
28 Jul 2026
Research

Paper from Kimi: Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter [R]

DGX agent

Mooncake is the serving platform for Kimi developed by Moonshot AI, featuring a KVCache-centric disaggregated architecture that separates prefill and decoding clusters while leveraging underutilized C

researchr-machinelearning
18 Apr 2026
Model Releases

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

DGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

model-releasesr-localllama
23 Jul 2026
Research

When are ICML openreviews made public? [R]

DGX agent

Reviews and discussions for all accepted papers at ICML are made public on OpenReview after the reviewing period concludes. Authors of rejected papers may also opt-in to have their reviews and discuss

researchr-machinelearning
31 May 2026
Research

Mandatory In-Person Presentation in CVPR 2026 [D]

DGX agent

This Reddit thread on r/MachineLearning discusses CVPR 2026's policy requiring accepted papers to be registered under an in-person author registration, with virtual attendance still permitted if circu

researchr-machinelearning
13 Apr 2026
Research

[ECCV2026] Workshop notification of reject/accept[D]

DGX agent

This Reddit thread on r/MachineLearning discusses the workshop paper accept/reject notifications for ECCV 2026, the European Conference on Computer Vision (ECCV), a biennial premier research conferenc

researchr-machinelearning
13 Apr 2026
Research

[P] citracer: a small CLI tool to trace where a concept comes from in a citation graph

DGX agent

`citracer` is a small Python CLI tool available on PyPI that traces citation chains for any keyword across research papers, helping researchers identify where a concept originates in a citation gra...

researchr-machinelearning
8 Apr 2026
Model Releases

Nanbeige4.2-3B: I'm not impressed

DGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

model-releasesr-localllama
30 Jul 2026
Research

So Confused about Polarizing ICML Reviews [D]

DGX agent

This r/MachineLearning discussion thread addresses the common and frustrating experience of receiving highly polarizing reviewer scores for ICML paper submissions — for example, one reviewer rating a

researchr-machinelearning
12 Apr 2026
Model Releases

New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :)

DGX agent

Hey Folks, I've been making quants for a while - recently I took a short break to get into hardcore research (submitted my first EMNLP paper during it!). Along the way, I built up a little arsenal of

model-releasesr-localllama
12 Aug 2026
Safety

Github repo to learn the OPD/OPSD and how they perform compared to GRPO, on a consumer grade GPU [P]

DGX agent

I am trying to learn concepts like On Policy Distillation (OPD), On Policy Self Distillation (OPSD) and how do they compare to RL algorithms like GRPO. There are a lot of papers on this, but because o

safetyr-machinelearning
1 Aug 2026
Research

KDD 2026 Cycle 2 Results [D]

DGX agent

KDD 2026 Cycle 2 Results covers the announcement of review decisions and outcomes from the second submission cycle of KDD 2026, which accepts papers across multiple tracks including Research, Applied

researchr-machinelearning
16 May 2026
Research

[D] What Happened to Neurips Creative AI Track? [R]

DGX agent

The NeurIPS Creative AI track became part of the main conference proceedings for 2025, with papers presented as posters during the conference , marking a change from 2024 when the track was not part o

researchr-machinelearning
4 May 2026
Research

How much harder is it these days to get into a PhD program without having a high ranking degree for UG? [D]

DGX agent

This Reddit discussion thread from r/MachineLearning explores the growing challenges faced by applicants from non-elite undergraduate institutions when applying to PhD programs in machine learning and

researchr-machinelearning
15 Apr 2026
Research

Ijcai 2026 rebuttal doubt [D]

DGX agent

This r/MachineLearning discussion thread addresses questions and concerns from researchers navigating the IJCAI 2026 author rebuttal phase, during which approximately 70% of submitted papers remained

researchr-machinelearning
12 Apr 2026
Model Releases

Hidden Reasoning from Claude and GPT are Decoded, and it is interesting

DGX agent

Yesteday a paper showed a gap that allows to see 100% of the reasoning tokens form ALL Claude and GPT models Stealing Reasoning Traces from Proprietary LLM APIs. check it out, they have published lots

model-releasesr-localllama
12 Aug 2026
Model Releases

I built a weird, low-power llama.cpp server using an Intel N100 + RTX 5060Ti

DGX agent

Everything started with the sudden death of my old ASRock J1900. While looking for the perfect ITX replacement, I stumbled upon the Chinese CW-NAS-ADLN-K motherboard, which looked perfect on paper: In

model-releasesr-localllama
11 Aug 2026
Hardware

Open Model: Google Weather Next 2

DGX agent

I am not a meteorologist, but I just read a very interesting article: https://arstechnica.com/science/2026/08/deepminds-hurricane-model-bought-forecasters-an-extra-day/ In a paper published on Thursda

hardwarer-localllama
9 Aug 2026
Hardware

Built & Trained a Transformer from Scratch in Pure PyTorch for English-to-Tamil Machine Translation [Math + Code Breakdown] [P]

DGX agent

Hi everyone! 👋 I built and trained the complete Transformer architecture from scratch using pure PyTorch (`torch.nn` primitives) based on the original 'Attention Is All You Need' paper. I trained the

hardwarer-machinelearning
27 Jul 2026
Model Releases

GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]

DGX agent

The interesting finding from a new [arXiv paper](https://arxiv.org/abs/2607.16165) isn't that a frontier vision model failed a new benchmark, that happens weekly, but the specific shape of the failure

model-releasesr-machinelearning
23 Jul 2026
Model Releases

SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]

DGX agent

Paper:https://arxiv.org/abs/2607.19058 Code (GitHub):https://github.com/nuemaan/skewadam Hi everyone, I just published a preprint on a new optimizer designed to tackle the massive VRAM bottleneck in M

model-releasesr-machinelearning
22 Jul 2026
Local Ai

I just wanted a small WebUI with an admin panel… it escalated into a full open-source agent framework runs fully local with Ollama

DGX agent

Let me try to explain this clearly, simply, and neatly. Originally, I just wanted to build a small WebUI adapter with an admin panel, but things escalated over the last few months. At first, I faced t

local-air-ollama
21 Jul 2026
Research

Follow the Mean: Reference-Guided Flow Matching [R]

DGX agent

This paper presents a method for controllable generation in flow matching models by adapting them through reference examples. The key insight is that the velocity field in flow matching is determined

researchr-machinelearning
14 May 2026
Research

CHI PLAY reviews [R]

DGX agent

This Reddit post on r/MachineLearning, tagged [R] (Research), is a community discussion thread sharing or reviewing peer feedback related to CHI PLAY — the international and interdisciplinary ACM SIGC

researchr-machinelearning
15 Apr 2026
Research

Thinking Deeper, Not Longer: Depth-Recurrent Transformers for Compositional Generalization [R]

DGX agent

This research paper explores a depth-recurrent transformer architecture designed to improve compositional generalization in language models by increasing computational depth rather than sequence lengt

researchr-machinelearning
13 Apr 2026
Research

What if your HNSW index stored 3-bit embeddings instead of float32? [R]

DGX agent

A research paper (arXiv:2601.11557) proposes replacing the dominant 'HNSW + float32 + cosine similarity' vector database stack with an information-theoretic alternative that uses Maximally Informat...

researchr-machinelearning
11 Apr 2026
Local Ai

Why Speculative Decoding went mature in 2026?

DGX agent

Spec-dec has been a thing for a while, in fact, it's wasn't an idea that was born for LLM inference. E.g. Uber's https://github.com/uber/submitqueue applied it to a merge queue. Apple & GDM had been r

local-air-localllama
10 Aug 2026
Model Releases

[2606.05682] Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillation

DGX agent

Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost constrained production environments. Quantization

model-releasesr-localllama
9 Aug 2026
Model Releases

KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding.

DGX agent

First of all, I'm not a lab, this was a solo summer research project that finally culminated into the github repo and the writeup. The repo includes a much deeper dive with methods, findings about qua

model-releasesr-localllama
9 Aug 2026
Model Releases

[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms

DGX agent

TL;DR: I spent 9 days developing a new quantization method for MLX models and measured 18 variants against each other on a single M5 Max MacBook Pro (128 GB). The result is the best-measuring MLX quan

model-releasesr-localllama
2 Aug 2026
Model Releases

Vacuum 16T

DGX agent

https://huggingface.co/tsfrm/vacuum-16t A 16.5-trillion-parameter model that contains nothing. This model is just a ████ you to the labs and companies who say that 'haha I have the biggest model out t

model-releasesr-localllama
2 Aug 2026
Local Ai

Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast

DGX agent

Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders

local-air-localllama
29 Jul 2026
Model Releases

We compared different LLMs on IMO 2026 [R]

DGX agent

There are a few reasons why problems from International Mathematical Olympiad function as a good benchmark for LLMs: - The problems are new, not included in the training data of any model - Hard math

model-releasesr-machinelearning
26 Jul 2026
Safety

Fizgig Krea 2 training features update

DGX agent

https://github.com/shootthesound/Fizgig Intelligent trainer - Per-image loss tracking with self-adapting training runs — every image gets its own verdict (easy / suspect / stuck / exhausted) and its o

safetyr-stablediffusion
24 Jul 2026
← Previous
1
Next →
86 results
← Previous
12
Next →