AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
Model Releases

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robo…

DGX agent

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robots — Gemini 3.5 Flash-Lite is our fastest, most cost-effecti

model-releasesgoogle-ai--x
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles

DGX agent

arXiv:2607.27670v1 Announce Type: new Abstract: Jigsaw puzzle solving requires jointly reasoning about visual content and geometric constraints, yet existing benchmarks use rectangular cuts that creat

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

K-EXAONE 2.0 released

DGX agent

https://huggingface.co/LGAI-EXAONE/K-EXAONE-2.0-750B-A37B https://huggingface.co/LGAI-EXAONE/K-EXAONE-2.0-750B-A37B-FP8 https://huggingface.co/LGAI-EXAONE/K-EXAONE-2.0-750B-A37B-NVFP4 https://huggingf

model-releasesr-localllama
31 Jul 2026
Model Releases

KAISEN: Reproducible Subgroup Fairness Auditing for Clinical Risk Models

DGX agent

arXiv:2607.28608v1 Announce Type: new Abstract: Clinical risk models routinely achieve strong aggregate performance while producing materially different error rates across patient subgroups. Audit pip

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning

DGX agent

arXiv:2607.27610v1 Announce Type: new Abstract: Reinforcement learning (RL) finetuning significantly enhances the reasoning capabilities of large language models (LLMs), yet its effectiveness critical

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

DGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LayerRAG-Bench: A Cross-Layer Reliability Benchmark for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2607.27353v1 Announce Type: new Abstract: Agentic retrieval-augmented generation systems can produce answers that appear grounded while failing at the evidence, tool-contract, authorization, or

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Learning-Augmented and Randomized Algorithms for Line Aggregation with Delays

DGX agent

arXiv:2607.27807v1 Announce Type: new Abstract: This paper studies learning-augmented and randomized online aggregation with delays on a line metric. We consider advice given as online suggested servi

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Learning Color Grading, No Photo Sharing: Federated Aesthetic Preference Learning for Personalized Image Enhancement

DGX agent

arXiv:2607.27659v1 Announce Type: new Abstract: Personalized image enhancement should reflect individual aesthetic taste, yet learning such preferences commonly depends on private photos and ratings t

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Learning features from Newton's algorithm: a way to accelerate nonlinear parametrized PDE solvers

DGX agent

arXiv:2607.28036v1 Announce Type: new Abstract: It is well known that Newton's method converges faster when the initial guess is closer to a root of a system of nonlinear equations. In this paper, a t

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Learning path to fully understand the Kimi K3 technical report?

DGX agent

Hi everyone, Can anyone suggest a learning path to fully understand the technical report for Kimi K3? My background: • I've taken a graduate-level deep learning course. • I understand the Transformer

model-releasesr-ollama
31 Jul 2026
Model Releases

Learning to Trace Seiberg Dualities

DGX agent

arXiv:2607.28628v1 Announce Type: cross Abstract: Dualities play an important role in establishing both microscopic and emergent phenomena in a wide range of physical systems. In practice, though, it

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Lightning OPD 2.0: Mitigating Style Bias in Cross-Teacher On-Policy Distillation for Large Reasoning Models

DGX agent

arXiv:2607.28449v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense token-level supervision from a teacher, but its effectiveness can depend on teacher consistency, meaning tha

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM-Guided Initialization for Accelerated Hybrid Quantum-Classical Medical Image Classification

DGX agent

arXiv:2607.27262v1 Announce Type: cross Abstract: Variational quantum algorithms often encounter barren plateaus, where cost gradients decay rapidly with increasing circuit depth, undermining the trai

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints

DGX agent

arXiv:2410.06458v2 Announce Type: replace Abstract: Instruction following is a key capability for LLMs. However, recent studies have shown that LLMs often struggle with instructions containing multipl

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM2Vec-Gen: Generative Embeddings from Large Language Models

DGX agent

arXiv:2603.10913v3 Announce Type: replace Abstract: Fine-tuning LLM-based text embedders via contrastive learning maps inputs and outputs into a new representational space, discarding the LLM's output

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLMs struggle to simulate human belief updates in controlled environments

DGX agent

arXiv:2607.28347v1 Announce Type: new Abstract: LLMs are increasingly deployed as proxies for human study participants in social science experiments, yet the fidelity of this practice has rarely been

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LM-GRASP: Instance-Specific Language Models for Combinatorial Construction via Online Imitation Learning

DGX agent

arXiv:2607.28135v1 Announce Type: new Abstract: Machine learning for combinatorial optimization typically relies on neural constructors trained via reinforcement learning on large offline datasets for

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

LoMeVQA: A Comprehensive Benchmark for Longitudinal Medical VQA

DGX agent

arXiv:2607.27806v1 Announce Type: new Abstract: In clinical practice, patients often undergo multiple imaging examinations over successive visits, yielding longitudinal data. Modeling such temporal in

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Looped Transformers with Source-Centered State Evolution

DGX agent

arXiv:2607.27656v1 Announce Type: cross Abstract: Looped Transformers create a useful train- and test-time compute axis by reusing the same Transformer block over recurrent depth, increasing effective

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts

DGX agent

arXiv:2607.27787v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) for mathematical reasoning suffers from a structural blind spot: on 'cliff' prompts-those on which

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

MagicSelector: Joint Optimization for Agent Tool Selection via Counterfactual Decomposition and Progressive Reranking

DGX agent

arXiv:2607.17751v2 Announce Type: cross Abstract: We present MagicSelector, a joint optimization framework integrating Counterfactual task decomposition, Progressive reranking, and Dynamic Top-K, desi

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

MatCreatioNN: Machine learning-guided computational discovery of photocatalysts for environmental applications

DGX agent

arXiv:2607.27295v1 Announce Type: cross Abstract: The rational design of photocatalysts for environmental remediation and CO2 conversion remains limited by the high computational cost and sparse exper

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Measuring Alignment With Reader Highlights Net of Position and Length

DGX agent

arXiv:2607.27739v1 Announce Type: cross Abstract: Context compression discards most of a document before a language model reads it, and is normally evaluated by downstream task accuracy - which makes

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models

DGX agent

arXiv:2502.20780v2 Announce Type: replace-cross Abstract: The increasing use of vision-language models (VLMs) in healthcare applications presents great challenges related to hallucinations, in which t

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Meituan just dropped LongCat-Flash-Lite-Sparse

DGX agent

It’s an MoE with ~3B active params and a 30B n-gram lookup table offloaded to RAM for fast 256k context on a 24GB GPU. Reminds me of Gemma 4’s PLE trick. Initial analysis suggest it wont be replacing

model-releasesr-localllama
31 Jul 2026
Model Releases

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory

DGX agent

arXiv:2607.27919v1 Announce Type: new Abstract: Decoder-only language models entangle long-term memory and reasoning in a single parameter set, making it difficult to scale memory capacity independent

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

MemSecBench: Tracking Agent Memory Poisoning from Persistence to Consequence and Repair

DGX agent

arXiv:2607.27080v1 Announce Type: cross Abstract: Memory systems allow agents to retain and reuse information from past interactions, but they can also let malicious content persist. A malicious instr

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

MiniMax H3 discussion

DGX agent

https://x.com/MiniMax_AI/status/2083008095488516262 (says it's coming in a few days) it can do text-to-image, and image editing everyone seems to be mostly hyped about the video generation part (I am

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

MiniMax H3: Open-weight multimodel video model

DGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

Minimax-H3 video model released, open weights coming in the next few days

DGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

model-releasesr-localllama
31 Jul 2026
Model Releases

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

DGX agent

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageabl

model-releasesr-localllama
31 Jul 2026
Model Releases

MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning

DGX agent

arXiv:2607.27109v2 Announce Type: cross Abstract: With the development of audio large language models (AudioLLMs), audio captioning needs to move from brief descriptions toward open-ended and fine-gra

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

MMHBench: A Multi-Perspective Benchmark for Mental Health Understanding in Long-Form Videos

DGX agent

arXiv:2607.27895v1 Announce Type: cross Abstract: Mental health understanding in long-form videos requires nuanced reasoning over observable behavior, interpersonal context, and latent psychological s

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models

DGX agent

arXiv:2607.27637v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on a wide range of vision-language tasks, but often fail under imperfect or sh

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Models for minimalist RAG: B1ade 335M Embedding and 1B Parameter Small Language Models

DGX agent

arXiv:2607.27506v1 Announce Type: new Abstract: Language and embedding models used in RAG systems are conventionally assumed to require large-scale pretraining and explicit grounding supervision. We p

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding

DGX agent

arXiv:2511.12449v3 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) have significantly advanced e-commerce product understanding. However, they still face three challen

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek

DGX agent

arXiv:2607.28274v1 Announce Type: new Abstract: Modern Greek is a richly inflected language, yet the language models built for it are evaluated mainly on factual knowledge, and no benchmark is dedicat

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

DGX agent

arXiv:2607.27616v1 Announce Type: new Abstract: Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into share

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MSCM-net: A hyperspectral image classiffcation method based on multi-scale convolution and Mamba

DGX agent

arXiv:2607.28277v1 Announce Type: new Abstract: Hyperspectral imaging is widely used in remote sensing and engineering. Therefore, research on its classification methods is crucial. While CNN and Tran

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning

DGX agent

arXiv:2607.26465v1 Announce Type: new Abstract: Multimodal Large Language Models have sparked significant interest due to their potential for social intelligence; however, their ability to perform seq

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them…

DGX agent

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them exchange findings only at phase boundaries, through staged

model-releasesdair-ai--x
31 Jul 2026
Model Releases

Neural Network-Assisted CLEAN for Channel Modeling in Low-SNR Regimes

DGX agent

arXiv:2607.27450v1 Announce Type: new Abstract: Accurate multipath parameter estimation is critical for modern wireless communication systems, particularly in challenging low-SNR environments. Traditi

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk,…

DGX agent

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk, which moved the bottleneck from how many exist to what is i

model-releasesdair-ai--x
31 Jul 2026
Model Releases

Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?

DGX agent

Laguna 2.1 at NVFP4 Deepseek v4 at Q2 Inkling-Small at IQ3 Which models you guys running now ? How it compares to 122b? Upcoming in few days : Ling 3.0 124B (Could be new king) LongCat 69B A3B ( very

model-releasesr-localllama
31 Jul 2026
Model Releases

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA

DGX agent

arXiv:2607.27566v1 Announce Type: new Abstract: Multi-frame medical VQA appears to reward increasingly complex adaptation: controller-style inference, localization-aware reranking, static hard-negativ

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and ge…

DGX agent

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps

model-releasessimon-willison--x
31 Jul 2026
← Previous
1…5960616263…466
Next →