AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

K3 already got in the top 5 most liked models of all time on Hugging Face, just 24 hours after being released! Ahead of Llama 3, Whisper and…

DGX agent

K3 is a language model that entered the top five most‑liked models on Hugging Face merely 24 hours after its release. The achievement surprised many, placing it ahead of prominent models such as Llama

model-releasesclem-delangue--x
28 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

KAYROS: An Anytime and Exact Open-Source Solver for Duration-Minimization Time-Dependent Vehicle Routing. A Technical Report and a Case Study in Human-AI Engineering

DGX agent

arXiv:2607.23116v1 Announce Type: cross Abstract: KAYROS is an open-source solver for duration-minimization time-dependent vehicle routing problems, with or without time windows (TDVRPTW, TDVRP). In t

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory

DGX agent

arXiv:2607.24368v1 Announce Type: new Abstract: Long-term memory systems store what a user says in an external store and retrieve it when a related query arrives. This interface rests on an assumption

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Key-Interval A*: Accelerating Grid Pathfinding via Structural Abstraction

DGX agent

arXiv:2607.23393v1 Announce Type: new Abstract: Existing exact methods for 4-connected grid pathfinding reduce online search, but often either retain fine-grained search states or require substantial

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Kimi K3: Open Frontier Intelligence

DGX agent

arXiv:2607.24653v1 Announce Type: new Abstract: We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories

DGX agent

arXiv:2607.23704v1 Announce Type: cross Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irre

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Language-Routed RAG and Direct Option Scoring for Multilingual Financial QA: DS@GT at FinMMEval

DGX agent

arXiv:2607.22841v1 Announce Type: cross Abstract: We present DS@GT's submission to FinMMEval 2026 Task 1, a multilingual financial exam question answering benchmark spanning English, Spanish, Greek, C

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Language Shapes Instruction Hierarchy Compliance in Multilingual LLMs

DGX agent

arXiv:2607.23545v1 Announce Type: new Abstract: Instruction hierarchy (IH) requires models to prioritize instructions by source, ensuring that higher-priority instructions override lower-priority ones

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Layering Virtual Try-On

DGX agent

arXiv:2607.22924v1 Announce Type: new Abstract: In the real world, fashion is about layering: adding a jacket over a shirt, or a sequence of adding and removing layers, rather than just a single-layer

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

DGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LC-SEPLM: long-range contact-supervised adaptation for sequence-only protein representation learning

DGX agent

arXiv:2607.22777v1 Announce Type: cross Abstract: Protein language models learn transferable sequence representations. However, because they primarily model contextual dependencies along amino-acid se

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks

DGX agent

arXiv:2607.23515v1 Announce Type: new Abstract: Long-horizon manipulation tasks pose significant challenges for reinforcement learning due to sparse reward signals and long horizons. Automatic curricu

model-releasesarxiv-cs-ro
28 Jul 2026
Model Releases

Learning Sampling Parameters for Diffusion Models

DGX agent

arXiv:2607.23488v1 Announce Type: cross Abstract: Text-to-image diffusion models expose many inference-time sampling parameters, including prompts, negative prompts, classifier-free guidance scales, a

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Learning to Access Computation: Accessibility Plasticity as a Principle of Adaptive Intelligence

DGX agent

arXiv:2607.22748v1 Announce Type: cross Abstract: Modern neural networks primarily adapt through parameter modification within predefined computational structures. While recent methods introduce modul

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal LEGO Assembly Assistants

DGX agent

arXiv:2507.05515v4 Announce Type: replace Abstract: Vision-language models (VLMs) are facing the challenges of understanding and following multimodal assembly instructions, particularly when fine-grai

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LFM2.5-Encoders: Fast at Long Context, Even on CPU

DGX agent

LFM2.5-Encoder is a family of multilingual bidirectional encoders built on the LFM2 architecture, available in two sizes: LFM2.5-Encoder-230M — a lightweight encoder for tight latency and memory budge

model-releasesr-localllama
28 Jul 2026
Model Releases

LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models

DGX agent

arXiv:2411.00918v5 Announce Type: replace-cross Abstract: Mixture of experts (MoE) architectures have become a cornerstone for scaling up and are a key component in most large language models such as

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LLM-Assisted Ontology Engineering and Construction of a French Legal Knowledge Graph

DGX agent

arXiv:2607.24551v1 Announce Type: new Abstract: Maintenance regulations are complex legal texts that are difficult to exploit when addressing a specific case and challenging to integrate into operatio

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports

DGX agent

arXiv:2607.24573v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support decisions about uncertain future events, yet evaluating their ability to forecast real-world outcomes

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Looking for Affect in Spontaneous Finnish Speech through Linguistic Interpretability

DGX agent

arXiv:2607.24155v1 Announce Type: new Abstract: Existing research on affect in speech has shown how acoustic surface characteristics and content-related linguistic aspects of speech both relate to per

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers

DGX agent

arXiv:2509.03059v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have shown that their reasoning capabilities can be significantly improved through Reinforceme

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LoRA for Gender-Inclusive Rewriting and Activation Steering for Counter-Narrative Generation

DGX agent

arXiv:2607.23083v1 Announce Type: new Abstract: Gender-inclusive language generation seeks to transform biased text into inclusive alternatives while preserving semantic meaning and contextual coheren

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

DGX agent

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with

model-releasesr-localllama
28 Jul 2026
Model Releases

LoTA-N2N: Local Trace Adaptation for Zero-Shot Self-Supervised Image Denoising

DGX agent

arXiv:2607.24135v1 Announce Type: new Abstract: Single-image self-supervised denoising replaces unavailable clean targets with surrogate targets constructed from noisy observations. Its effectiveness

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Low-Rank Dependence Decomposition via Accelerated Symmetric Non-negative Matrix Factorization

DGX agent

arXiv:2607.24518v1 Announce Type: new Abstract: Symmetric non-negative matrix factorization (SymNMF) recovers latent group structure from a dependence matrix, but its dense, quadratic-memory objective

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

LowAux-RDNet: Low-Pass Residual Supervision with Scene-Balanced Real-World Training for Single-Image Reflection Removal

DGX agent

arXiv:2607.22707v1 Announce Type: new Abstract: Single-image reflection removal aims to recover a clean transmission layer from one image captured through glass. We study an explicit decomposition pip

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

LU-500: A Logo Benchmark for Concept Unlearning

DGX agent

arXiv:2607.24101v1 Announce Type: cross Abstract: Concept unlearning is increasingly used to limit the reproduction of protected or unsafe visual concepts in text-to-image models. Existing evaluations

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation

DGX agent

arXiv:2601.17039v2 Announce Type: replace-cross Abstract: Mangroves are critical for climate-change mitigation, requiring reliable monitoring for effective conservation. While deep learning has emerge

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MATS: A novel multi-modality multi-task learning framework for 3D perception in autonomous driving

DGX agent

arXiv:2607.24224v1 Announce Type: new Abstract: Multi-modality data from different sensors provides rich complementary information for 3D perception, becoming an essential component in reliable autono

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

MAViE: A Multi-scale Adaptive Vision Encoder for Fine-grained Visual Perception and Efficient Multimodal Reasoning

DGX agent

arXiv:2607.24424v1 Announce Type: new Abstract: Vision-language models commonly project all tokens produced by a pretrained vision encoder into a large language model. However, final-layer features ca

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Maximum Satisfiability of Simple Temporal Problems

DGX agent

arXiv:2607.23785v1 Announce Type: cross Abstract: The Simple Temporal Problem (STP) is a core framework for quantitative temporal constraints. As STP data can be inconsistent, we study MAXSTP: compute

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Measuring Negative Campaigning across Languages with Large Language Models: A Study of 18 Million Tweets in 19 Countries

DGX agent

arXiv:2507.17636v2 Announce Type: replace Abstract: Negative campaigning is a defining feature of electoral competition, yet comparative research on its drivers has remained limited by the high cost a

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection

DGX agent

arXiv:2607.15166v2 Announce Type: replace Abstract: Most medical AI benchmarks measure whether a model knows the correct answer. MedFailBench asks a different question: which safety boundary failed? W

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models

DGX agent

arXiv:2607.22566v1 Announce Type: new Abstract: MedLoCoMo is a Medical Long-Context Memory benchmark for patient-specific clinical reasoning over multi-admission medical dialogue. Existing medical QA

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MEGA-CL: A Molecular Foundation Model for Generalizable ADMET Prediction through Graph External Attention and Contrastive Learning

DGX agent

arXiv:2607.24314v1 Announce Type: new Abstract: Predicting the absorption, distribution, metabolism, excretion and toxicity (ADMET) properties of small molecules remains a major challenge in drug disc

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

MegaSlide-DiT: Memory-Centric Adaptation and Deformable Local Attention for Efficient Video Diffusion

DGX agent

arXiv:2607.22696v1 Announce Type: cross Abstract: High-resolution video diffusion models built on Diffusion Transformers (DiTs) deliver strong fidelity but quickly exhaust the memory budget of a singl

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

DGX agent

arXiv:2607.23811v1 Announce Type: cross Abstract: Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple's most powerful

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Meshless Domain Randomization via Explicit Parameter Perturbation of 3D Gaussian Splatting

DGX agent

arXiv:2607.22890v1 Announce Type: cross Abstract: Domain Randomization (DR) is a standard technique for closing the Sim-to-Real gap, yet traditional DR pipelines rely on classical computer graphics re

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

Mage-VL is a codec-native, proactive-streaming multimodal foundation model for image and video understanding, whose visual encoder is trained entirely from scratch at a compact 4B scale. It targets a

model-releasesr-localllama
28 Jul 2026
Model Releases

Might need math+code benchmark for frontier model(LLMs Silently Replace Math)[D]

DGX agent

Hello guys. I found some problems in current frontier models. And want to share. # math_code_hallucination > Record of a failure caused by combining mathematics and code in a single prompt. --- ## Cas

model-releasesr-machinelearning
28 Jul 2026
Model Releases

MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models

DGX agent

arXiv:2607.22556v1 Announce Type: new Abstract: Continual learning (CL) is essential for small language models (SLMs) to adapt to evolving real-world needs in resource-constrained deployments. However

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models

DGX agent

arXiv:2607.23047v1 Announce Type: cross Abstract: Mixed-precision quantization improves the accuracy of post-training quantization by allocating higher bitwidths to sensitive layers, but existing meth

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Mixture-of-Thought-Tokens: Unifying Perception and Reasoning for Free-form Multimodal Grounding

DGX agent

arXiv:2607.24407v1 Announce Type: new Abstract: Multimodal Large Language Models have made great progress in grounding tasks, yet existing methods still struggle to unify precise localization and comp

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

MMOE: Modernizing Diffusion Transformers with Efficient Expert Design

DGX agent

arXiv:2607.24665v1 Announce Type: new Abstract: Modern large language models scale successfully by pairing capacity growth with efficiency, keeping per-token and deployment costs under control as capa

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Modeling Memory-Dependent Reliability of LLMs: A Hidden Markov Model

DGX agent

arXiv:2607.22951v1 Announce Type: cross Abstract: Reliability assessment of large language models (LLMs) seeks to estimate the probability that a model produces correct responses under a specified ope

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

MoE-Based Learned Inertial Odometry for Bicycle Localization

DGX agent

arXiv:2510.17604v2 Announce Type: replace Abstract: GNSS suffers from multipath errors in urban canyons, making reliable bicycle localization difficult. Hand-crafted inertial alternatives, such as cyc

model-releasesarxiv-cs-ro
28 Jul 2026
Model Releases

MoLGE: Mixture of Language Group Experts for Efficient Scaling of Massively Multilingual Speech Recognition

DGX agent

arXiv:2607.24030v1 Announce Type: new Abstract: Massively multilingual automatic speech recognition (ASR) models covering hundreds of languages must maintain robust performance across diverse linguist

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents

DGX agent

arXiv:2607.23870v1 Announce Type: cross Abstract: Smart-city airspace is transforming Uncrewed Aerial Vehicles (UAVs) from passive sensing platforms into cyber-physical decision makers that must follo

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…7879808182…470
Next →