AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “r-machinelearning”

GridTimelineEvolution
173 results
16 May 2026

KDD 2026 Cycle 2 Results [D]

ResearchDGX agent

KDD 2026 Cycle 2 Results covers the announcement of review decisions and outcomes from the second submission cycle of KDD 2026, which accepts papers across multiple tracks including Research, Applied

14 May 2026

Continual Harness: Online Adaptation for Self-Improving Foundation Agents [R]

ResearchDGX agent

Continual Harness proposes an approach to online adaptation for foundation agents that moves beyond traditional gradient-based retraining by introducing a dual-agent architecture (Teacher/Student) wit

Follow the Mean: Reference-Guided Flow Matching [R]

ResearchDGX agent

This paper presents a method for controllable generation in flow matching models by adapting them through reference examples. The key insight is that the velocity field in flow matching is determined


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Integrating 3D Heat Equation into a PINN for Real-Time Aerospace Simulation (C++ WASM Engine)[P]

ResearchDGX agent

This post discusses implementing Physics-Informed Neural Networks (PINNs) to solve the 3D heat equation within a C++ WebAssembly engine for real-time aerospace applications. The work combines machine

12 May 2026

TabPFN-3 just released: a pre-trained tabular foundation model for up to 1M rows [R][N]

Model ReleasesDGX agent

TabPFN-3 is a pre-trained tabular foundation model that supports datasets up to 1,000,000 rows × 200 features , representing a significant scaling improvement for the TabPFN family. The model delivers

11 May 2026

Interactive Jensen–Shannon Divergence Visualisation [P]

ResearchDGX agent

Jensen-Shannon divergence is a method of measuring the similarity between two probability distributions. It is based on the Kullback-Leibler divergence, with the notable difference that it is symmetri

9 May 2026

DeepSeek V4 paper full version is out, FP4 QAT details and stability tricks [D]

Model ReleasesDGX agent

DeepSeek released the full technical report for DeepSeek-V4 on April 24, 2026, titled 'DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence.' The paper details FP4 quantization-awa

8 May 2026

Disillusionment with mechanistic interpretability research [D]

ResearchDGX agent

Mechanistic interpretability research aims to uncover specific neurons and circuits in neural networks responsible for tasks, but over a decade of efforts suggests these findings may not translate int

7 May 2026

Diffusion for generating/editing ASTs? [D]

ResearchDGX agent

Recent advances in diffusion-based language models enable controllable sequence generation, but applying them to structured code remains challenging, prompting exploration of syntax-aware diffusion fr

Quantization and Fast Inference (MEAP) - How much performance are you actually getting from quantization in production? [D]

ApplicationsDGX agent

This discussion examines the practical performance gains of quantization techniques when deployed in production environments, particularly addressing whether theoretical speedups translate to real-wor

ROCm Status in mid 2026 [D]

ResearchDGX agent

ROCm has reached production-ready status for PyTorch and vLLM workloads in 2026. PyTorch runs well and MI300X benchmarks are competitive, though the ecosystem gap with CUDA remains real. AMD's MI355X

5 May 2026

Charting the AI Perception Gap: Across 71 scenarios, AI experts (N=119) and the public (N=1100) have differing views on the risks, benefits, and value of AI. More importantly, AI experts discount the influence of risks stronger than the public does when forming their value judgments [R]

ResearchDGX agent

A study examining 71 AI scenarios found that AI experts (N=119) and the public (N=1100) hold differing views on AI risks, benefits, and value, with experts discounting risk influences more heavily tha

4 May 2026

[D] What Happened to Neurips Creative AI Track? [R]

ResearchDGX agent

The NeurIPS Creative AI track became part of the main conference proceedings for 2025, with papers presented as posters during the conference , marking a change from 2024 when the track was not part o

[P] QLoRA Fine-Tuning of Qwen2.5-1.5B for CEFR English Proficiency Classification (A1–C2) [P]

ResearchDGX agent

This post likely describes a machine learning project implementing QLoRA (Quantized Low-Rank Adaptation) fine-tuning on the Qwen2.5-1.5B model to classify English language proficiency levels according

3 May 2026

Anyone submit ML articles to ACM journals (eg. TOPML or TIST)? [D]

ResearchDGX agent

A discussion post on r/MachineLearning where users share experiences and advice about submitting machine learning research articles to ACM journals, with focus on publications like ACM Transactions on

Public interpretability dataset and benchmark library for a novel transformer architecture [R]

Model ReleasesDGX agent

This work presents an explainability library for transformer models that provides tools for understanding transformer behavior through attributions and concept-based explanations . The resource likely

Struggling with Chebyshev Filter Integration in CNN — Any Advice? [R]

ResearchDGX agent

This Reddit discussion covers technical challenges encountered when attempting to integrate Chebyshev filters—mathematical filters with equiripple characteristics used in signal processing—into convol

torch-nvenc-compress: GPU NVENC silicon as a PCIe bandwidth multiplier — PCA + pure-ctypes Video Codec SDK wrapper. Parallel-path overlap measured at 67% of theoretical max on a real GEMM + encode workload. [P]

HardwareDGX agent

torch-nvenc-compress is a Python library that leverages GPU NVENC (NVIDIA's hardware video encoding) to optimize PCIe bandwidth utilization by compressing data during transfer. The project implements

1 May 2026

[D] Simple Questions Thread

ResearchDGX agent

This is a discussion thread from r/MachineLearning where community members ask and answer beginner-level and straightforward questions about machine learning concepts, techniques, and practical implem

What benchmark would you build for “reply quality” in SDR generation? [D]

Model ReleasesDGX agent

Based on the Reddit discussion title, this likely discusses how to design and establish benchmarks for evaluating the quality of AI-generated responses in Sales Development Representative (SDR) system

30 Apr 2026

[R] Joint Embedding Variational Bayes (TMLR ’26)

ResearchDGX agent

Variational Joint Embedding (VJE) is a framework that synthesizes joint embedding and variational inference to enable self-supervised learning of probabilistic representations in a reconstruction-free

29 Apr 2026

Free Registration & $20K Prize Pool: 2nd MLC-SLM Challenge 2026 on Multilingual Speech LLMs [N]

ResearchDGX agent

The 2nd Multilingual Conversational Speech Language Model (MLC-SLM) Challenge is an open research competition inviting teams worldwide to participate , featuring free registration with a $20K prize po

What are people using for low-latency autocomplete in production? [P]

ApplicationsDGX agent

Production low-latency autocomplete implementations employ diverse strategies including inference server optimization (tools like vLLM, llama.cpp, NVIDIA Triton), deployment choices (cloud APIs, on-pr

28 Apr 2026

ACL ARR March 2026 Cycle [D]

ResearchDGX agent

The ACL Rolling Review (ARR) operates on a two-monthly review cycle , and the March 2026 cycle refers to one of these periodic submission and peer review rounds for computational linguistics research.

Karpathy dropped a 200-line GPT, so I used the math to turn pandas DataFrames into searchable context windows and open sourced it (and automated my stats pipeline). [P]

ResearchDGX agent

This post describes a project where the author leveraged Andrej Karpathy's compact GPT implementation to create a tool for converting pandas DataFrames into searchable context windows, which they then

23 Apr 2026

First time fine-tuning, need a sanity check — 3B or 7B for multi-task reasoning? [D]

ResearchDGX agent

This Reddit discussion post addresses a beginner's question about choosing between 3B and 7B parameter models for fine-tuning on multi-task reasoning problems. The post likely contains advice from exp

22 Apr 2026

GPU Compass – open-source, real-time GPU pricing across 20+ clouds [P]

HardwareDGX agent

GPU Compass is an open-source tool that tracks and displays real-time GPU pricing information across over 20 cloud providers. The platform likely helps machine learning practitioners and researchers c

INT3 compression+fused metal kernels [R]

ResearchDGX agent

INT3 compression with fused Metal kernels enables large language models to compute attention operations directly on compressed (INT3/INT4) key-value cache representations using custom GPU kernels that

19 Apr 2026

Why production systems keep making “correct” decisions that are no longer right [D]

ApplicationsDGX agent

Production machine learning systems can experience performance degradation where predictions appear reasonable and metrics don't change immediately, yet decision quality decays over time in ways diffi

18 Apr 2026

easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P]

SafetyDGX agent

easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization

Paper from Kimi: Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter [R]

ResearchDGX agent

Mooncake is the serving platform for Kimi developed by Moonshot AI, featuring a KVCache-centric disaggregated architecture that separates prefill and decoding clusters while leveraging underutilized C

16 Apr 2026

AI for Materials Science starter kit [D]

ResearchDGX agent

This r/MachineLearning discussion post serves as a community-curated beginner's resource for applying artificial intelligence and machine learning to materials science, likely compiling recommended to

Built an political benchmark for LLMs. KIMI K2 can't answer about Taiwan (Obviously). GPT-5.3 refuses 100% of questions when given an opt-out. [P]

Model ReleasesDGX agent

A researcher on r/MachineLearning built a political benchmark to evaluate how various LLMs handle sensitive geopolitical and politically contentious questions. Key findings include that Kimi K2 (Moons

Why dynamically routing multi-timescale advantages in PPO causes policy collapse (and a simple decoupled fix) [R]

SafetyDGX agent

This Reddit post discusses a known instability in PPO when advantage estimates operating across different temporal scales (e.g., short-horizon and long-horizon returns) are dynamically routed or mixed

15 Apr 2026

Are gamers being used as free labeling labor? The rise of 'Simulators' that look like AI training grounds [D]

ResearchDGX agent

This r/MachineLearning discussion examines a growing concern in the AI and gaming community: that certain video game 'simulators' may be deliberately designed — or repurposed — to harvest player actio

Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]

Model ReleasesDGX agent

A Reddit post on r/MachineLearning sharing an open-source project and accompanying book by Sebastian Raschka that walks through implementing GPT-2, Llama 3, and DeepSeek from scratch using PyTorch, wi

CHI PLAY reviews [R]

ResearchDGX agent

This Reddit post on r/MachineLearning, tagged [R] (Research), is a community discussion thread sharing or reviewing peer feedback related to CHI PLAY — the international and interdisciplinary ACM SIGC

Failure to Reproduce Modern Paper Claims [D]

ResearchDGX agent

This r/MachineLearning discussion thread addresses the widespread challenge of reproducing results claimed in modern ML research papers, a topic of significant concern in the field. Community members

Give me your ideass [N]

ResearchDGX agent

A community thread on r/MachineLearning where users are invited to share research ideas, project concepts, or suggestions related to machine learning. The '[N]' tag indicates it is a discussion post r

Hosting Live session for sub 10ms retrieval by Moss (YC backed) [N]

ResearchDGX agent

Moss is a YC-backed high-performance runtime for real-time semantic search that delivers sub-10ms lookups, instant index updates, and zero infrastructure overhead, running where the agent lives — clou

How much harder is it these days to get into a PhD program without having a high ranking degree for UG? [D]

ResearchDGX agent

This Reddit discussion thread from r/MachineLearning explores the growing challenges faced by applicants from non-elite undergraduate institutions when applying to PhD programs in machine learning and

Jailbreaks as social engineering: 5 case studies suggest LLMs inherit human psychological vulnerabilities from training data [D]

ResearchDGX agent

This r/MachineLearning discussion post examines LLM jailbreaks through the lens of social engineering, arguing that the psychological vulnerabilities found in LLMs are not random artifacts but structu

[N] AMA Reminder: Max Welling

ResearchDGX agent

This Reddit post is a reminder to the r/MachineLearning community about an upcoming Ask Me Anything (AMA) session with Max Welling, Chief AI Officer at CuspAI and Professor of Machine Learning at the

One of the fastest ways to lose trust in a self-hosted LLM: prompt injection compliance [P]

ResearchDGX agent

This r/MachineLearning post discusses how self-hosted LLMs that comply with prompt injection attempts — effectively following malicious or overriding instructions embedded in user input — represent a

[P] Added 8 Indian languages to Chatterbox TTS via LoRA — 1.4% of parameters, no phoneme engineering [P]

ResearchDGX agent

A community researcher shared on r/MachineLearning how they extended Chatterbox TTS — Resemble AI's open-source, 500M-parameter model — to support 8 Indian languages using LoRA (Low-Rank Adaptation),

Seeking Critique on Research Approach to Open Set Recognition (Novelty Detection) [R]

ResearchDGX agent

This is a community discussion post on the r/MachineLearning subreddit where a researcher shares their proposed methodology for tackling Open Set Recognition (OSR) and/or novelty detection, inviting p

Was looking at a ICLR 2025 Oral paper and I am shocked it got oral [D]

ResearchDGX agent

This Reddit thread from r/MachineLearning reflects community skepticism about the peer review standards at top ML conferences, with a user expressing surprise that a particular paper received an oral

What is the criteria for a ML paper to be published?[D]

ResearchDGX agent

This r/MachineLearning discussion thread explores the standards and expectations reviewers and program committees use when evaluating ML papers for acceptance at top venues such as NeurIPS, ICML, and

14 Apr 2026

20M+ Indian legal documents with citation graphs and vector embeddings – potential uses for legal NLP? [D]

ApplicationsDGX agent

A r/MachineLearning discussion thread exploring the potential NLP applications of a large-scale dataset comprising over 20 million Indian legal documents, enriched with citation graphs and pre-compute

ClawBench: Can AI Agents Complete Everyday Online Tasks? 153 tasks, 144 live websites, best model at 33.3% [R]

ResearchDGX agent

ClawBench is a benchmark of 153 everyday web tasks spanning 144 live platforms across 15 categories — from completing purchases and booking appointments to submitting job applications. Unlike existing

'I don't know!': Teaching neural networks to abstain with the HALO-Loss. [R]

ResearchDGX agent

This r/MachineLearning post discusses research on training neural networks to recognize and express uncertainty by introducing a novel HALO-Loss function that gives models an explicit 'abstention' opt

No agent maintained moral reasoning consistency across scenarios. Findings from a structured study with 11 agents on classic ethical dilemmas [R]

AgentsDGX agent

No agent maintained moral reasoning consistency across scenarios. Findings from a structured study with 11 agents on classic ethical dilemmas [R]

We benchmarked TranslateGemma against 5 other LLMs on subtitle translation across 6 languages. At first glance the numbers told a clean story, but then human QA added a chapter. [D]

ResearchDGX agent

This r/MachineLearning discussion post details a hands-on benchmark study in which TranslateGemma — Google's open translation model suite built on Gemma 3, available in 4B, 12B, and 27B sizes and cove

What is the AC guidance for ICML? (Or: ICML qq thread) [D]

ResearchDGX agent

This r/MachineLearning Reddit thread serves as a community Q&A ('qq thread') focused on the Area Chair (AC) process for ICML, where researchers ask and answer questions about AC guidance, responsibili

You can decompose models into a graph database [N]

ResearchDGX agent

This Reddit post from r/MachineLearning discusses the concept of decomposing machine learning models into a graph database representation, treating a model's components — such as layers, weights, and

13 Apr 2026

Built an AI tool that cleans datasets, fills missing values, and predicts unknown fields [P]

ResearchDGX agent

A developer shared a project on r/MachineLearning showcasing an AI-powered tool designed to automate common data preprocessing tasks, including dataset cleaning, intelligent imputation of missing valu

Claude code skill for neurotech/BCI machine learning [P]

Model ReleasesDGX agent

This r/MachineLearning post discusses a Claude Code skill tailored for neurotechnology and brain-computer interface (BCI) machine learning workflows, likely covering domain-specific tasks such as neur

[ECCV2026] Workshop notification of reject/accept[D]

ResearchDGX agent

This Reddit thread on r/MachineLearning discusses the workshop paper accept/reject notifications for ECCV 2026, the European Conference on Computer Vision (ECCV), a biennial premier research conferenc

hands on workshop: context engineering for multi agent systems [D]

AgentsDGX agent

This r/MachineLearning discussion thread centers on a hands-on workshop covering context engineering — the process of designing and optimizing the information an AI agent sends to and receives from a

I scaled a pure Spiking Neural Network (SNN) to 1.088B parameters from scratch. Ran out of budget, but here is what I found [R]

ResearchDGX agent

A researcher on r/MachineLearning documented an independent attempt to scale a pure Spiking Neural Network (SNN) to 1.088 billion parameters, built from scratch, exploring whether SNNs can achieve per

← Previous
123
Next →