AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
Model Releases

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

DGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

model-releasesarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Guardrails in Logit Space: Safety Token Regularization for LLM Alignment

DGX agent

arXiv:2604.17210v1 Announce Type: new Abstract: Fine-tuning well-aligned large language models (LLMs) on new domains often degrades their safety alignment, even when using benign datasets. Existing sa

model-releasesarxiv-cs-lg
21 Apr 2026
Research

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

DGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

researcharxiv-cs-lg
21 Apr 2026
Model Releases

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

DGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

DGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

model-releasesclem-delangue--x
21 Apr 2026
Model Releases

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

DGX agent

arXiv:2512.04677v5 Announce Type: replace Abstract: Audio-driven avatar interaction demands real-time, streaming, and infinite-length generation -- capabilities fundamentally at odds with the sequenti

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

DGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

DGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

DGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Mitigating Multimodal Hallucination via Phase-wise Self-reward

DGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs

DGX agent

arXiv:2604.18203v1 Announce Type: new Abstract: Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical unde

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

DGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

DGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

DGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

QU-NLP at QIAS 2026: Multi-Stage QLoRA Fine-Tuning for Arabic Islamic Inheritance Reasoning

DGX agent

arXiv:2604.16396v1 Announce Type: new Abstract: Islamic inheritance law (ilm al-mawar{i}th) presents a challenging domain for evaluating large language models' structured reasoning capabilities, requi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Sessa: Selective State Space Attention

DGX agent

arXiv:2604.18580v1 Announce Type: cross Abstract: Modern sequence models are dominated by Transformers, where self-attention mixes information from the visible context in an input-dependent way. Howev

researcharxiv-cs-cl
21 Apr 2026
Model Releases

The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus

DGX agent

arXiv:2604.16913v1 Announce Type: cross Abstract: Decentralized Autonomous Organizations (DAOs) are inclined explore Small Language Models (SLMs) as edge-native constitutional firewalls to vet proposa

model-releasesarxiv-cs-cl
21 Apr 2026
Research

ThinkBrake: Efficient Reasoning via Log-Probability Margin Guided Decoding

DGX agent

arXiv:2510.00546v5 Announce Type: replace Abstract: Large Reasoning Models (LRMs) allocate substantial inference-time compute to Chain-of-Thought (CoT) reasoning, improving performance on mathematics,

researcharxiv-cs-cl
21 Apr 2026
Model Releases

TimeColor: Flexible Reference Colorization via Temporal Concatenation

DGX agent

arXiv:2601.00296v2 Announce Type: replace Abstract: Most colorization models condition only on a single reference, typically the first frame of the scene. However, this approach ignores other sources

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

TinySR: Pruning Diffusion for Real-World Image Super-Resolution

DGX agent

arXiv:2508.17434v2 Announce Type: replace Abstract: Real-world image super-resolution (Real-ISR) focuses on recovering high-quality images from low-resolution inputs that suffer from complex degradati

model-releasesarxiv-cs-cv
21 Apr 2026
Research

TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

DGX agent

arXiv:2603.19684v2 Announce Type: replace Abstract: Automatic tooth segmentation and identification from intra-oral scanned 3D models are fundamental problems in digital dentistry, yet most existing a

researcharxiv-cs-cv
21 Apr 2026
Local Ai

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

DGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

local-aiarxiv-cs-cl
21 Apr 2026
Research

ViT^3: Unlocking Test-Time Training in Vision

DGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

researcharxiv-cs-cv
21 Apr 2026
Research

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

DGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Art3D: Training-Free 3D Generation from Flat-Colored Illustration

DGX agent

arXiv:2504.10466v2 Announce Type: replace Abstract: Large-scale pre-trained image-to-3D generative models have exhibited remarkable capabilities in diverse shape generations. However, most of them str

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

AscendKernelGen: A Systematic Study of LLM-Based Kernel Generation for Neural Processing Units

DGX agent

arXiv:2601.07160v2 Announce Type: replace Abstract: To meet the ever-increasing demand for computational efficiency, Neural Processing Units (NPUs) have become critical in modern AI infrastructure. Ho

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei foo…

DGX agent

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei food. No fluff. April 28 | 5PM | Kaiyō SF Register → http://lum

model-releasesai21-labs--x
20 Apr 2026
Model Releases

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

DGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

model-releasesyann-lecun--x
20 Apr 2026
Model Releases

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

DGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

model-releaseskimi-moonshot--x
20 Apr 2026
Local Ai

kimi k2.6 is now 'available' on ollama cloud

DGX agent

Kimi K2.6 is an open-source model featuring advanced coding, long-horizon execution, and agent swarm capabilities that is now available via Ollama Cloud . The model excels in coding and agentic tools

local-air-ollama
20 Apr 2026
Applications

Modern Structure-Aware Simplicial Spatiotemporal Neural Network

DGX agent

arXiv:2604.15833v1 Announce Type: new Abstract: Spatiotemporal modeling has evolved beyond simple time series analysis to become fundamental in structural time series analysis. While current research

applicationsarxiv-cs-lg
20 Apr 2026
Model Releases

PILOT: A Promptable Interleaved Layout-aware OCR Transformer

DGX agent

arXiv:2504.03621v2 Announce Type: replace Abstract: Classical OCR pipelines decompose document reading into detection, segmentation, and recognition stages, which makes them sensitive to localization

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Reading today's open-closed performance gap

DGX agent

This article analyzes the performance differences between open-source and closed-source AI models in the current landscape, examining factors that influence their relative capabilities and market posi

researchinterconnects
20 Apr 2026
Research

SLE-FNO: Single-Layer Extensions for Task-Agnostic Continual Learning in Fourier Neural Operators

DGX agent

arXiv:2603.20410v2 Announce Type: replace Abstract: Scientific machine learning is increasingly used to build surrogate models, yet most models are trained under a restrictive assumption in which futu

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Was happy with Gemma 4 Cloud, but had to change due to API Errors, GLM 5.1 spends a lot more ressources

DGX agent

A user reported satisfaction with Gemma 4 Cloud but switched to GLM 5.1 due to API errors, noting that the alternative model consumes significantly more resources. The post likely discusses the perfor

model-releasesr-ollama
20 Apr 2026
Local Ai

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it whe…

DGX agent

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it when it got cloned into Jackrongs repo, and it’s going a bit vi

local-aiclem-delangue--x
18 Apr 2026
Local Ai

Deepfake Detection Generalization with Diffusion Noise

DGX agent

arXiv:2604.14570v1 Announce Type: new Abstract: Deepfake detectors face growing challenges in generalization as new image synthesis techniques emerge. In particular, deepfakes generated by diffusion m

local-aiarxiv-cs-cv
17 Apr 2026
Research

ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving

DGX agent

arXiv:2604.14626v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become the dominant architecture for large-scale language models, yet on-premises serving remains fundamentally mem

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)

DGX agent

Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn

model-releasesr-stablediffusion
17 Apr 2026
Research

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

DGX agent

arXiv:2604.14188v1 Announce Type: cross Abstract: Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such mo

researcharxiv-cs-cl
17 Apr 2026
Research

Graph-Based Alternatives to LLMs for Human Simulation

DGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

researcharxiv-cs-cl
17 Apr 2026
Model Releases

IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation

DGX agent

arXiv:2511.01014v3 Announce Type: replace Abstract: Instruction-following is a fundamental ability of Large Language Models (LLMs), requiring their generated outputs to follow multiple constraints imp

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

Learning temporal embeddings from electronic health records of chronic kidney disease patients

DGX agent

arXiv:2601.18675v2 Announce Type: replace Abstract: We investigate whether temporal embedding models trained on longitudinal electronic health records can learn clinically meaningful representations w

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

DGX agent

arXiv:2604.15149v1 Announce Type: new Abstract: As reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for scaling reasoning capabilities in LLMs, a new failure mode

model-releasesarxiv-cs-lg
17 Apr 2026
Agents

LM Studio 0.4.12 is out now - @Alibaba_Qwen Qwen3.6 support! - Nicer PDF exports for chats - MCP servers with OAuth work on Windows - Better…

DGX agent

LM Studio version 0.4.12 introduces support for Alibaba's Qwen 3.6 model, improved PDF export functionality for chat conversations, and fixes for MCP (Model Context Protocol) servers with OAuth authen

agentslm-studio--x
17 Apr 2026
Applications

PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D Perception

DGX agent

arXiv:2510.17568v5 Announce Type: replace Abstract: Recent 3D feed-forward models, such as the Visual Geometry Grounded Transformer (VGGT), have shown strong capability in inferring 3D attributes of s

applicationsarxiv-cs-cv
17 Apr 2026
Research

Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration

DGX agent

arXiv:2503.21970v3 Announce Type: replace Abstract: State-Space Models (SSMs) have attracted considerable attention in Image Restoration (IR) due to their ability to scale linearly sequence length whi

researcharxiv-cs-cv
17 Apr 2026
← Previous
1…409410411412413…1326
Next →