AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
28 Jul 2026

How Context Attribution Handles What the Model Already Knows

Model ReleasesDGX agent

arXiv:2607.23804v1 Announce Type: cross Abstract: Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the ini

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the al…

Model ReleasesDGX agent

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the alternative is locked-down systems that refuse to assist their

https://x.com/Sauers_/status/2082171683645817193

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

New Anthropic research: Discovering cryptographic weaknesses with Claude. Claude Mythos Preview has helped our researchers find weaknesses in cryptographic algorithms—the mathematical methods that are

I built a tool to actually test which weights matter before quantizing, instead of guessing (Qwen3.6-27B, 3 builds: Bedrock/Tightrope/Gambit)

Model ReleasesDGX agent

Most quantization works like this: pick a bit depth, apply it everywhere, maybe let imatrix take a rough guess at what matters, ship it. Most don't check which specific weight groups can take a hit an

I got Kimi-k3 running.....

Model ReleasesDGX agent

Results: prompt eval: 40 tokens / 97.5s → 0.41 tok/s eval: 400 tokens / 1769.9s → 0.23 tok/s total: 440 tokens / 1867s (31 min) Prompt: 'Write a C++ function that reverses a linked list in place. Expl

If I only went off X posts, I'd think Ramp was an AI lab

Model ReleasesDGX agent

If I only went off X posts, I'd think Ramp was an AI lab We’re open-sourcing PorTAL, our framework for shared task representations and cross model LoRA adaptation. It now spans from hybrid attention m

IKS-Instruct: A 24,000-Example Multilingual Dataset for Teaching Language Models Indian Knowledge Systems

Model ReleasesDGX agent

arXiv:2607.23322v1 Announce Type: new Abstract: Instruction tuning has become the standard method for adapting large language models to follow human intent, yet existing instruction datasets are domin

Impute On-Demand: Adaptive Correlated Time Series Imputation for Changing Environments

Model ReleasesDGX agent

arXiv:2607.23503v1 Announce Type: cross Abstract: Internet of Things (IoT) applications generate vast amounts of Correlated Time Series (CTS) data that often contain missing values and require imputat

Indic DiarBench: A Multilingual Joint Diarization and ASR Benchmark for Indian Languages

Model ReleasesDGX agent

arXiv:2607.23808v1 Announce Type: cross Abstract: In this work, we introduce Indic DiarBench, a speaker diarization and ASR benchmark dataset spanning all 22 scheduled languages of India. This corpus

Infinite-Precision Autoregressive Modeling for Vector Graphics and Layouts

Model ReleasesDGX agent

arXiv:2601.05680v2 Announce Type: replace-cross Abstract: While Transformer-based autoregressive models excel in data generation, their token discretization strategy inherently limits their precision

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models

Model ReleasesDGX agent

arXiv:2607.24273v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong potential in financial reasoning, but existing benchmarks often evaluate domain knowledge, numerical reas

Interesting. Grok 4.6 releases around August 7. This will be the 1.5T model with significantly improved SFT & RL. Grok 4.7 will be the 2.1T …

Model ReleasesDGX agent

Interesting. Grok 4.6 releases around August 7. This will be the 1.5T model with significantly improved SFT & RL. Grok 4.7 will be the 2.1T model released a few weeks later. This will be better than 4

Inverse Bayesian Inference for Extracting Lesion Dynamics from Longitudinal Spectral CT

Model ReleasesDGX agent

arXiv:2607.23078v1 Announce Type: new Abstract: Longitudinal medical imaging captures temporal evolution of lesions, yet extracting the underlying dynamical parameters governing this evolution remains

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and …

Model ReleasesDGX agent

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and there needs to be some of severe criminal sanction applied t

JPEG AIC2026: A large-scale dataset for fine-grained assessment of image coding

Model ReleasesDGX agent

arXiv:2607.22783v1 Announce Type: cross Abstract: Recent advances in conventional and learning-based image coding have increased the demand for benchmark datasets that support fine-grained assessment

K-Survival Means

Model ReleasesDGX agent

arXiv:2607.24405v1 Announce Type: new Abstract: In this work, we propose K-SurvMeans, a novel extension of K-Means for clustering survival data. The method explicitly uses the survival outcome in the

K3 already got in the top 5 most liked models of all time on Hugging Face, just 24 hours after being released! Ahead of Llama 3, Whisper and…

Model ReleasesDGX agent

K3 is a language model that entered the top five most‑liked models on Hugging Face merely 24 hours after its release. The achievement surprised many, placing it ahead of prominent models such as Llama

KAYROS: An Anytime and Exact Open-Source Solver for Duration-Minimization Time-Dependent Vehicle Routing. A Technical Report and a Case Study in Human-AI Engineering

Model ReleasesDGX agent

arXiv:2607.23116v1 Announce Type: cross Abstract: KAYROS is an open-source solver for duration-minimization time-dependent vehicle routing problems, with or without time windows (TDVRPTW, TDVRP). In t

Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory

Model ReleasesDGX agent

arXiv:2607.24368v1 Announce Type: new Abstract: Long-term memory systems store what a user says in an external store and retrieve it when a related query arrives. This interface rests on an assumption

Key-Interval A*: Accelerating Grid Pathfinding via Structural Abstraction

Model ReleasesDGX agent

arXiv:2607.23393v1 Announce Type: new Abstract: Existing exact methods for 4-connected grid pathfinding reduce online search, but often either retain fine-grained search states or require substantial

Kimi K3: Open Frontier Intelligence

Model ReleasesDGX agent

arXiv:2607.24653v1 Announce Type: new Abstract: We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token

LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories

Model ReleasesDGX agent

arXiv:2607.23704v1 Announce Type: cross Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irre

Language-Routed RAG and Direct Option Scoring for Multilingual Financial QA: DS@GT at FinMMEval

Model ReleasesDGX agent

arXiv:2607.22841v1 Announce Type: cross Abstract: We present DS@GT's submission to FinMMEval 2026 Task 1, a multilingual financial exam question answering benchmark spanning English, Spanish, Greek, C

Language Shapes Instruction Hierarchy Compliance in Multilingual LLMs

Model ReleasesDGX agent

arXiv:2607.23545v1 Announce Type: new Abstract: Instruction hierarchy (IH) requires models to prioritize instructions by source, ensuring that higher-priority instructions override lower-priority ones

Layering Virtual Try-On

Model ReleasesDGX agent

arXiv:2607.22924v1 Announce Type: new Abstract: In the real world, fashion is about layering: adding a jacket over a shirt, or a sequence of adding and removing layers, rather than just a single-layer

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

LC-SEPLM: long-range contact-supervised adaptation for sequence-only protein representation learning

Model ReleasesDGX agent

arXiv:2607.22777v1 Announce Type: cross Abstract: Protein language models learn transferable sequence representations. However, because they primarily model contextual dependencies along amino-acid se

LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks

Model ReleasesDGX agent

arXiv:2607.23515v1 Announce Type: new Abstract: Long-horizon manipulation tasks pose significant challenges for reinforcement learning due to sparse reward signals and long horizons. Automatic curricu

Learning Sampling Parameters for Diffusion Models

Model ReleasesDGX agent

arXiv:2607.23488v1 Announce Type: cross Abstract: Text-to-image diffusion models expose many inference-time sampling parameters, including prompts, negative prompts, classifier-free guidance scales, a

Learning to Access Computation: Accessibility Plasticity as a Principle of Adaptive Intelligence

Model ReleasesDGX agent

arXiv:2607.22748v1 Announce Type: cross Abstract: Modern neural networks primarily adapt through parameter modification within predefined computational structures. While recent methods introduce modul

LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal LEGO Assembly Assistants

Model ReleasesDGX agent

arXiv:2507.05515v4 Announce Type: replace Abstract: Vision-language models (VLMs) are facing the challenges of understanding and following multimodal assembly instructions, particularly when fine-grai

LFM2.5-Encoders: Fast at Long Context, Even on CPU

Model ReleasesDGX agent

LFM2.5-Encoder is a family of multilingual bidirectional encoders built on the LFM2 architecture, available in two sizes: LFM2.5-Encoder-230M — a lightweight encoder for tight latency and memory budge

LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models

Model ReleasesDGX agent

arXiv:2411.00918v5 Announce Type: replace-cross Abstract: Mixture of experts (MoE) architectures have become a cornerstone for scaling up and are a key component in most large language models such as

LLM-Assisted Ontology Engineering and Construction of a French Legal Knowledge Graph

Model ReleasesDGX agent

arXiv:2607.24551v1 Announce Type: new Abstract: Maintenance regulations are complex legal texts that are difficult to exploit when addressing a specific case and challenging to integrate into operatio

LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports

Model ReleasesDGX agent

arXiv:2607.24573v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support decisions about uncertain future events, yet evaluating their ability to forecast real-world outcomes

Looking for Affect in Spontaneous Finnish Speech through Linguistic Interpretability

Model ReleasesDGX agent

arXiv:2607.24155v1 Announce Type: new Abstract: Existing research on affect in speech has shown how acoustic surface characteristics and content-related linguistic aspects of speech both relate to per

Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers

Model ReleasesDGX agent

arXiv:2509.03059v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have shown that their reasoning capabilities can be significantly improved through Reinforceme

LoRA for Gender-Inclusive Rewriting and Activation Steering for Counter-Narrative Generation

Model ReleasesDGX agent

arXiv:2607.23083v1 Announce Type: new Abstract: Gender-inclusive language generation seeks to transform biased text into inclusive alternatives while preserving semantic meaning and contextual coheren

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

Model ReleasesDGX agent

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with

LoTA-N2N: Local Trace Adaptation for Zero-Shot Self-Supervised Image Denoising

Model ReleasesDGX agent

arXiv:2607.24135v1 Announce Type: new Abstract: Single-image self-supervised denoising replaces unavailable clean targets with surrogate targets constructed from noisy observations. Its effectiveness

Low-Rank Dependence Decomposition via Accelerated Symmetric Non-negative Matrix Factorization

Model ReleasesDGX agent

arXiv:2607.24518v1 Announce Type: new Abstract: Symmetric non-negative matrix factorization (SymNMF) recovers latent group structure from a dependence matrix, but its dense, quadratic-memory objective

LowAux-RDNet: Low-Pass Residual Supervision with Scene-Balanced Real-World Training for Single-Image Reflection Removal

Model ReleasesDGX agent

arXiv:2607.22707v1 Announce Type: new Abstract: Single-image reflection removal aims to recover a clean transmission layer from one image captured through glass. We study an explicit decomposition pip

LU-500: A Logo Benchmark for Concept Unlearning

Model ReleasesDGX agent

arXiv:2607.24101v1 Announce Type: cross Abstract: Concept unlearning is increasingly used to limit the reproduction of protected or unsafe visual concepts in text-to-image models. Existing evaluations

MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation

Model ReleasesDGX agent

arXiv:2601.17039v2 Announce Type: replace-cross Abstract: Mangroves are critical for climate-change mitigation, requiring reliable monitoring for effective conservation. While deep learning has emerge

MATS: A novel multi-modality multi-task learning framework for 3D perception in autonomous driving

Model ReleasesDGX agent

arXiv:2607.24224v1 Announce Type: new Abstract: Multi-modality data from different sensors provides rich complementary information for 3D perception, becoming an essential component in reliable autono

MAViE: A Multi-scale Adaptive Vision Encoder for Fine-grained Visual Perception and Efficient Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2607.24424v1 Announce Type: new Abstract: Vision-language models commonly project all tokens produced by a pretrained vision encoder into a large language model. However, final-layer features ca

Maximum Satisfiability of Simple Temporal Problems

Model ReleasesDGX agent

arXiv:2607.23785v1 Announce Type: cross Abstract: The Simple Temporal Problem (STP) is a core framework for quantitative temporal constraints. As STP data can be inconsistent, we study MAXSTP: compute

Measuring Negative Campaigning across Languages with Large Language Models: A Study of 18 Million Tweets in 19 Countries

Model ReleasesDGX agent

arXiv:2507.17636v2 Announce Type: replace Abstract: Negative campaigning is a defining feature of electoral competition, yet comparative research on its drivers has remained limited by the high cost a

MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection

Model ReleasesDGX agent

arXiv:2607.15166v2 Announce Type: replace Abstract: Most medical AI benchmarks measure whether a model knows the correct answer. MedFailBench asks a different question: which safety boundary failed? W

MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2607.22566v1 Announce Type: new Abstract: MedLoCoMo is a Medical Long-Context Memory benchmark for patient-specific clinical reasoning over multi-admission medical dialogue. Existing medical QA

MEGA-CL: A Molecular Foundation Model for Generalizable ADMET Prediction through Graph External Attention and Contrastive Learning

Model ReleasesDGX agent

arXiv:2607.24314v1 Announce Type: new Abstract: Predicting the absorption, distribution, metabolism, excretion and toxicity (ADMET) properties of small molecules remains a major challenge in drug disc

MegaSlide-DiT: Memory-Centric Adaptation and Deformable Local Attention for Efficient Video Diffusion

Model ReleasesDGX agent

arXiv:2607.22696v1 Announce Type: cross Abstract: High-resolution video diffusion models built on Diffusion Transformers (DiTs) deliver strong fidelity but quickly exhaust the memory budget of a singl

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

Model ReleasesDGX agent

arXiv:2607.23811v1 Announce Type: cross Abstract: Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple's most powerful

Meshless Domain Randomization via Explicit Parameter Perturbation of 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2607.22890v1 Announce Type: cross Abstract: Domain Randomization (DR) is a standard technique for closing the Sim-to-Real gap, yet traditional DR pipelines rely on classical computer graphics re

microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model

Model ReleasesDGX agent

Mage-VL is a codec-native, proactive-streaming multimodal foundation model for image and video understanding, whose visual encoder is trained entirely from scratch at a compact 4B scale. It targets a

Might need math+code benchmark for frontier model(LLMs Silently Replace Math)[D]

Model ReleasesDGX agent

Hello guys. I found some problems in current frontier models. And want to share. # math_code_hallucination > Record of a failure caused by combining mathematics and code in a single prompt. --- ## Cas

MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models

Model ReleasesDGX agent

arXiv:2607.22556v1 Announce Type: new Abstract: Continual learning (CL) is essential for small language models (SLMs) to adapt to evolving real-world needs in resource-constrained deployments. However

MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models

Model ReleasesDGX agent

arXiv:2607.23047v1 Announce Type: cross Abstract: Mixed-precision quantization improves the accuracy of post-training quantization by allocating higher bitwidths to sensitive layers, but existing meth

Mixture-of-Thought-Tokens: Unifying Perception and Reasoning for Free-form Multimodal Grounding

Model ReleasesDGX agent

arXiv:2607.24407v1 Announce Type: new Abstract: Multimodal Large Language Models have made great progress in grounding tasks, yet existing methods still struggle to unify precise localization and comp

MMOE: Modernizing Diffusion Transformers with Efficient Expert Design

Model ReleasesDGX agent

arXiv:2607.24665v1 Announce Type: new Abstract: Modern large language models scale successfully by pairing capacity growth with efficiency, keeping per-token and deployment costs under control as capa

← Previous
1…5859606162…373
Next →