AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
22 Apr 2026

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

Model ReleasesDGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

Model ReleasesDGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

What’s new in GKE at Next ‘26

Model ReleasesDGX agent

This week at Google Cloud Next ‘26, we are sharing the evolution of Google Kubernetes Engine (GKE), delivering leading performance, efficiency, security, and scale for your most demanding and complex

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

Model ReleasesDGX agent

arXiv:2604.02368v4 Announce Type: replace Abstract: As Large Language Models (LLMs) exhibit plateauing performance on conventional benchmarks, a pivotal challenge persists: evaluating their proficienc

21 Apr 2026

An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules

Model ReleasesDGX agent

arXiv:2604.16550v1 Announce Type: new Abstract: Despite the high accuracy of 'black box' deep learning models, drug discovery still relies on protein-ligand interaction principles and heuristics. To i

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning

ResearchDGX agent

arXiv:2510.14738v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have rapidly advanced from perception tasks to complex multi-step reasoning, yet reinforcement learning wit

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

Model ReleasesDGX agent

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs

ResearchDGX agent

arXiv:2604.17629v1 Announce Type: new Abstract: Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-c

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

Model ReleasesDGX agent

arXiv:2507.20879v3 Announce Type: replace Abstract: The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

Model ReleasesDGX agent

arXiv:2509.20360v3 Announce Type: replace Abstract: Recent advances in foundation models highlight a clear trend toward unification and scaling, showing emergent capabilities across diverse domains. W

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

Model ReleasesDGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

'Faithful to What?' On the Limits of Fidelity-Based Explanations

SafetyDGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

FedOBP: Federated Optimal Brain Personalization through Cloud-Edge Element-wise Decoupling

Model ReleasesDGX agent

arXiv:2604.16574v1 Announce Type: new Abstract: Federated Learning (FL) faces challenges from client data heterogeneity and resource-constrained mobile devices, which can degrade model accuracy. Perso

From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation

TutorialsDGX agent

arXiv:2604.17455v1 Announce Type: new Abstract: Visual prompting has emerged as a powerful method for adapting pre-trained models to new domains without updating model parameters. However, existing pr

From Inheritance to Saturation: Disentangling the Evolution of Visual Redundancy for Architecture-Aware MLLM Inference Acceleration

Model ReleasesDGX agent

arXiv:2604.16462v1 Announce Type: new Abstract: High-resolution Multimodal Large Language Models (MLLMs) face prohibitive computational costs during inference due to the explosion of visual tokens. Ex

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

Model ReleasesDGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

Guardrails in Logit Space: Safety Token Regularization for LLM Alignment

Model ReleasesDGX agent

arXiv:2604.17210v1 Announce Type: new Abstract: Fine-tuning well-aligned large language models (LLMs) on new domains often degrades their safety alignment, even when using benign datasets. Existing sa

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

ResearchDGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

Model ReleasesDGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

Model ReleasesDGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Model ReleasesDGX agent

arXiv:2512.04677v5 Announce Type: replace Abstract: Audio-driven avatar interaction demands real-time, streaming, and infinite-length generation -- capabilities fundamentally at odds with the sequenti

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

Model ReleasesDGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

Model ReleasesDGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

Model ReleasesDGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

Model ReleasesDGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

Mitigating Multimodal Hallucination via Phase-wise Self-reward

ResearchDGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs

Model ReleasesDGX agent

arXiv:2604.18203v1 Announce Type: new Abstract: Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical unde

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

Model ReleasesDGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

Model ReleasesDGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

Model ReleasesDGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

QU-NLP at QIAS 2026: Multi-Stage QLoRA Fine-Tuning for Arabic Islamic Inheritance Reasoning

Model ReleasesDGX agent

arXiv:2604.16396v1 Announce Type: new Abstract: Islamic inheritance law (ilm al-mawar{i}th) presents a challenging domain for evaluating large language models' structured reasoning capabilities, requi

Sessa: Selective State Space Attention

ResearchDGX agent

arXiv:2604.18580v1 Announce Type: cross Abstract: Modern sequence models are dominated by Transformers, where self-attention mixes information from the visible context in an input-dependent way. Howev

The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus

Model ReleasesDGX agent

arXiv:2604.16913v1 Announce Type: cross Abstract: Decentralized Autonomous Organizations (DAOs) are inclined explore Small Language Models (SLMs) as edge-native constitutional firewalls to vet proposa

ThinkBrake: Efficient Reasoning via Log-Probability Margin Guided Decoding

ResearchDGX agent

arXiv:2510.00546v5 Announce Type: replace Abstract: Large Reasoning Models (LRMs) allocate substantial inference-time compute to Chain-of-Thought (CoT) reasoning, improving performance on mathematics,

TimeColor: Flexible Reference Colorization via Temporal Concatenation

Model ReleasesDGX agent

arXiv:2601.00296v2 Announce Type: replace Abstract: Most colorization models condition only on a single reference, typically the first frame of the scene. However, this approach ignores other sources

TinySR: Pruning Diffusion for Real-World Image Super-Resolution

Model ReleasesDGX agent

arXiv:2508.17434v2 Announce Type: replace Abstract: Real-world image super-resolution (Real-ISR) focuses on recovering high-quality images from low-resolution inputs that suffer from complex degradati

TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

ResearchDGX agent

arXiv:2603.19684v2 Announce Type: replace Abstract: Automatic tooth segmentation and identification from intra-oral scanned 3D models are fundamental problems in digital dentistry, yet most existing a

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

Local AiDGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

ViT^3: Unlocking Test-Time Training in Vision

ResearchDGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

ResearchDGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

20 Apr 2026

Art3D: Training-Free 3D Generation from Flat-Colored Illustration

Model ReleasesDGX agent

arXiv:2504.10466v2 Announce Type: replace Abstract: Large-scale pre-trained image-to-3D generative models have exhibited remarkable capabilities in diverse shape generations. However, most of them str

AscendKernelGen: A Systematic Study of LLM-Based Kernel Generation for Neural Processing Units

Model ReleasesDGX agent

arXiv:2601.07160v2 Announce Type: replace Abstract: To meet the ever-increasing demand for computational efficiency, Neural Processing Units (NPUs) have become critical in modern AI infrastructure. Ho

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei foo…

Model ReleasesDGX agent

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei food. No fluff. April 28 | 5PM | Kaiyō SF Register → http://lum

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

Model ReleasesDGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

Model ReleasesDGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

kimi k2.6 is now 'available' on ollama cloud

Local AiDGX agent

Kimi K2.6 is an open-source model featuring advanced coding, long-horizon execution, and agent swarm capabilities that is now available via Ollama Cloud . The model excels in coding and agentic tools

Modern Structure-Aware Simplicial Spatiotemporal Neural Network

ApplicationsDGX agent

arXiv:2604.15833v1 Announce Type: new Abstract: Spatiotemporal modeling has evolved beyond simple time series analysis to become fundamental in structural time series analysis. While current research

PILOT: A Promptable Interleaved Layout-aware OCR Transformer

Model ReleasesDGX agent

arXiv:2504.03621v2 Announce Type: replace Abstract: Classical OCR pipelines decompose document reading into detection, segmentation, and recognition stages, which makes them sensitive to localization

Reading today's open-closed performance gap

ResearchDGX agent

This article analyzes the performance differences between open-source and closed-source AI models in the current landscape, examining factors that influence their relative capabilities and market posi

SLE-FNO: Single-Layer Extensions for Task-Agnostic Continual Learning in Fourier Neural Operators

ResearchDGX agent

arXiv:2603.20410v2 Announce Type: replace Abstract: Scientific machine learning is increasingly used to build surrogate models, yet most models are trained under a restrictive assumption in which futu

Was happy with Gemma 4 Cloud, but had to change due to API Errors, GLM 5.1 spends a lot more ressources

Model ReleasesDGX agent

A user reported satisfaction with Gemma 4 Cloud but switched to GLM 5.1 due to API errors, noting that the alternative model consumes significantly more resources. The post likely discusses the perfor

18 Apr 2026

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it whe…

Local AiDGX agent

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it when it got cloned into Jackrongs repo, and it’s going a bit vi

17 Apr 2026

Deepfake Detection Generalization with Diffusion Noise

Local AiDGX agent

arXiv:2604.14570v1 Announce Type: new Abstract: Deepfake detectors face growing challenges in generalization as new image synthesis techniques emerge. In particular, deepfakes generated by diffusion m

ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving

ResearchDGX agent

arXiv:2604.14626v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become the dominant architecture for large-scale language models, yet on-premises serving remains fundamentally mem

Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)

Model ReleasesDGX agent

Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

ResearchDGX agent

arXiv:2604.14188v1 Announce Type: cross Abstract: Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such mo

Graph-Based Alternatives to LLMs for Human Simulation

ResearchDGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation

Model ReleasesDGX agent

arXiv:2511.01014v3 Announce Type: replace Abstract: Instruction-following is a fundamental ability of Large Language Models (LLMs), requiring their generated outputs to follow multiple constraints imp

Learning temporal embeddings from electronic health records of chronic kidney disease patients

TutorialsDGX agent

arXiv:2601.18675v2 Announce Type: replace Abstract: We investigate whether temporal embedding models trained on longitudinal electronic health records can learn clinically meaningful representations w

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

Model ReleasesDGX agent

arXiv:2604.15149v1 Announce Type: new Abstract: As reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for scaling reasoning capabilities in LLMs, a new failure mode

← Previous
1…320321322323324…1042
Next →