AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
28 May 2026

Hurwitz Quaternion Multiplicative Quantization for KV Cache Compression

Model ReleasesDGX agent

arXiv:2605.27646v1 Announce Type: cross Abstract: We propose extbf{Hurwitz Quaternion Multiplicative Quantization (HQMQ)}, a extbf{calibration-free} method for KV cache compression of large language m

Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions

TutorialsDGX agent

arXiv:2605.27523v1 Announce Type: cross Abstract: Deep generative models offer powerful tools for multivariate data analysis, but their black-box architectures are often unidentified and difficult to

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This …

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This is an exciting milestone for Glean, and it's a signal about wh

IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage

Model ReleasesDGX agent

arXiv:2605.28247v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for en- hancing LLM reasoning, yet its data ineffi- ciency remains a

La-Proteina: Atomistic Protein Generation via Partially Latent Flow Matching

ResearchDGX agent

arXiv:2507.09466v2 Announce Type: replace Abstract: Recently, many generative models for de novo protein structure design have emerged. Yet, only few tackle the difficult task of directly generating f

Latent Diffusion for Missing Data

ResearchDGX agent

arXiv:2605.28427v1 Announce Type: new Abstract: Diffusion models have emerged as powerful generative approaches for missing-data imputation, yet most existing methods operate directly in data space an

Learning to Translate from Soft to Hard LLM Prompts

Model ReleasesDGX agent

arXiv:2605.27642v1 Announce Type: new Abstract: Soft prompt tuning is a parameter-efficient method for adapting LLMs to specific tasks, but suffers from a lack of interpretability. Building on recent

LLM Zeroth-Order Fine-Tuning is an Inference Workload

Model ReleasesDGX agent

arXiv:2605.28760v1 Announce Type: new Abstract: Zeroth-order (ZO) fine-tuning is attractive for large language models because it replaces backpropagation with forward objective evaluations. Existing i

MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content

Model ReleasesDGX agent

arXiv:2605.28116v1 Announce Type: cross Abstract: Mobile graphical user interface (GUI) agents driven by vision-language models (VLMs) perceive the screen as rendered pixels and choose actions from wh

NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval

Model ReleasesDGX agent

arXiv:2603.12824v2 Announce Type: replace-cross Abstract: Vision-Language Model (VLM) based retrievers have advanced visual document retrieval (VDR) to impressive quality. They require the same multi-

OccuReward: LLM-Guided Occupant-Centric Reward Shaping for Demographic Equity in Grid-Interactive Buildings

Model ReleasesDGX agent

arXiv:2605.28168v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated promising capability in generating reward functions for deep reinforcement learning (DRL)-based building

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩 https://huggingface.co/nvidia/GLM-5.1-NVFP4

HardwareDGX agent

NVIDIA has released GLM-5.1-NVFP4, a quantized version of the GLM-5.1 model, now available on Hugging Face. The model appears to use NVFP4 (NVIDIA's floating-point 4-bit) quantization format, designed

OGER: A Robust Offline-Guided Exploration Reward for Hybrid Reinforcement Learning

SafetyDGX agent

arXiv:2604.18530v2 Announce Type: replace Abstract: Recent advancements in Reinforcement Learning with Verifiable Rewards (RLVR) have significantly improved Large Language Model (LLM) reasoning, yet m

OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration

Local AiDGX agent

arXiv:2605.28805v1 Announce Type: cross Abstract: Visual outcomes are increasingly central to multimodal large language models, making reliable and fine-grained verification essential for scaling gene

On Compositional Learning Behaviours in Formal Mathematics

Model ReleasesDGX agent

arXiv:2605.28512v1 Announce Type: new Abstract: Self-evolving scientific agents capable of conquering the hard tail of formal mathematics require Compositional Learning Behaviours (CLBs) -- the capaci

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

Model ReleasesDGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

PEAR: Equal Area Weather Forecasting on the Sphere

ResearchDGX agent

arXiv:2505.17720v3 Announce Type: replace Abstract: Artificial intelligence is rapidly reshaping the natural sciences, with weather forecasting emerging as a flagship AI4Science application where mach

PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective

Model ReleasesDGX agent

arXiv:2605.28819v1 Announce Type: cross Abstract: Parameter-efficient finetuning (PEFT) has become the standard approach for adapting large language models, yet evaluations largely emphasize downstrea

PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence

Model ReleasesDGX agent

arXiv:2605.28068v1 Announce Type: new Abstract: Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pr

PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment

Model ReleasesDGX agent

arXiv:2605.28241v1 Announce Type: new Abstract: Point cloud quality plays a critical role in 3D acquisition, reconstruction, rendering, and perception, yet existing point cloud quality assessment (PCQ

PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting

Model ReleasesDGX agent

arXiv:2605.28066v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable efficacy in text embedding, yet current adaptation methods like LoRA face significant bottle

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation

Model ReleasesDGX agent

arXiv:2605.28091v1 Announce Type: new Abstract: Text-to-Image generation has evolved from basic image synthesis into a frequently used core capability in professional creative workflows, where simple

Reflective Dialogue between Teacher and Solver Agents for Video Question Answering

Model ReleasesDGX agent

arXiv:2605.27885v1 Announce Type: new Abstract: Various approaches have been proposed to adapt Vision-Language Models (VLMs) to specialized domains for Video Question Answering, including fine-tuning

ResearchMath-14K: Scaling Research-Level Mathematics via Agents

AgentsDGX agent

arXiv:2605.28003v1 Announce Type: new Abstract: The frontier of mathematics is defined by problems whose solutions are not yet known, yet it remains unclear whether language models can meaningfully en

Resolution-free neural surrogates for geometric parameterization and mapping with spatially varying fields

Model ReleasesDGX agent

arXiv:2605.28551v1 Announce Type: new Abstract: Many imaging problems require computing spatial transformations induced by spatially varying intensity, feature, or density fields. Canonical examples i

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

SEMAGIC: Learning Semantically Consistent Deformable 3D Representations from In-the-Wild Images

SafetyDGX agent

arXiv:2605.27938v1 Announce Type: new Abstract: Learning deformable 3D object models from single-view in-the-wild images has enabled impressive 3D shape reconstruction without supervision. However, it

Singular Vectors of Attention Heads Align with Features

SafetyDGX agent

arXiv:2602.13524v2 Announce Type: replace-cross Abstract: Identifying feature representations in language models is a central task in mechanistic interpretability. Several recent studies have made the

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

Model ReleasesDGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

The Future of Facts: Tracing the Factual Generation-Verification Gap

ResearchDGX agent

arXiv:2605.27564v1 Announce Type: cross Abstract: Language models are becoming the default interface to factual knowledge, yet they often verify outputs more reliably than they generate them. This gen

The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes

SafetyDGX agent

arXiv:2602.15515v2 Announce Type: replace-cross Abstract: Training against white-box deception detectors has been proposed as a way to make AI systems honest. However, such training risks models learn

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

Why We Need Speech to Evaluate Speech Translation

ResearchDGX agent

arXiv:2605.28227v1 Announce Type: new Abstract: Speech translation models are increasingly capable of preserving speech-specific information (e.g., speaker gender, prosody, and emphasis), yet evaluati

27 May 2026

A Dataset of Robot-Patient and Doctor-Patient Medical Dialogues for Spoken Language Processing Tasks

Model ReleasesDGX agent

arXiv:2605.26747v1 Announce Type: new Abstract: Large Language Models (LLMs) have brought huge improvements to Artificial Intelligence (AI), which can be applied to general-purpose tasks. However, the

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

Model ReleasesDGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

Chain Of Thought Compression: A Theoretical Analysis

Model ReleasesDGX agent

arXiv:2601.21576v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has unlocked advanced reasoning abilities of Large Language Models (LLMs) with intermediate steps, yet incurs prohibitive com

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

Model ReleasesDGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning

Model ReleasesDGX agent

arXiv:2605.26789v1 Announce Type: new Abstract: Post-training is routinely evaluated through aggregate benchmark scores that treat multi-hop reasoning as a single capability -- as if a model that answ

CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations

SafetyDGX agent

arXiv:2605.26293v1 Announce Type: cross Abstract: Prior work establishes that controlled contrastiveness between self-generated responses from large language models, set via reward scores, improves do

Developing a Totally Unimodular Linear Program for Optimal Conformance Checking: When and Why It Complements A*

Model ReleasesDGX agent

arXiv:2605.26938v1 Announce Type: new Abstract: Alignment-based conformance checking is the state-of-the-art approach for comparing observed process executions with normative process models. The stand

Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information

Model ReleasesDGX agent

arXiv:2601.03089v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly evaluated with input attribution methods, yet comparing such explanations remains challenging. E

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

Model ReleasesDGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

Furina: Fragmented Uncertainty-Driven Refusal Instability Attack

SafetyDGX agent

arXiv:2605.26158v1 Announce Type: cross Abstract: Safety alignment in large language models (LLMs) and multimodal large language models (MLLMs) is commonly assumed to operate as a near-binary threshol

GeoFaith: A Spatio-Temporal Dual View of Faithful Chain-of-Thought

Model ReleasesDGX agent

arXiv:2605.26893v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has advanced large language models (LLMs), but outcome-based supervision leads to pervasive post-hoc rationalization,

How Chain-of-Thought Works? Tracing Information Flow from Decoding, Projection, and Activation

Model ReleasesDGX agent

arXiv:2507.20758v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting significantly enhances model reasoning, yet its internal mechanisms remain poorly understood. We analyze CoT's oper

InfoQuant: Shaping Activation Distributions for Low-Bit LLM Quantization

Model ReleasesDGX agent

arXiv:2605.26175v1 Announce Type: cross Abstract: Low-bit activation quantization remains a major bottleneck in efficient large language model (LLM) deployment. The difficulty is not only that activat

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and vide…

Model ReleasesDGX agent

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and videos with state-of-the-art models, like Gen-4.5, Seedance 2.0,

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams

Model ReleasesDGX agent

arXiv:2605.27074v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) achieve strong performance on reactive question answering, but real-world streaming assistants require p

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

SafetyDGX agent

arXiv:2605.27288v1 Announce Type: cross Abstract: Large language models (LLMs) are known to abandon their initial stance to conform to user pushback. While prior research largely attributes this behav

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search

HardwareDGX agent

arXiv:2605.26636v1 Announce Type: cross Abstract: We introduce JetViT, a novel family of hybrid-architecture Vision Transformer (ViT) models that match the accuracy of state-of-the-art full-attention

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors

Model ReleasesDGX agent

arXiv:2605.26955v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed to users around the world, they are integrated into everyday tasks across diverse cultural c

L2Rec: Towards Dual-View Understanding of LLMs for Personalized Recommendation

Model ReleasesDGX agent

arXiv:2605.26717v1 Announce Type: cross Abstract: Adapting large language models (LLMs) for personalized recommendation requires aligning their general-purpose capabilities with user-specific preferen

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

Model ReleasesDGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

MemFail: Stress-Testing Failure Modes of LLM Memory Systems

Model ReleasesDGX agent

arXiv:2605.26667v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external memory systems to remain consistent across long-horizon interactions, but little empiric

MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale

Model ReleasesDGX agent

arXiv:2605.27235v1 Announce Type: new Abstract: Layered image generation and editing is a fundamental capability that enables layer-wise reuse, editing, and composition of generated visual content, an

Neural Autoregressive Control Variates for the Quantum Monte Carlo Sign Problem

Model ReleasesDGX agent

arXiv:2605.26814v1 Announce Type: cross Abstract: We train a pair of autoregressive models to construct zero-mean control variates to mitigate the sign problem in quantum Monte Carlo simulations. The

ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection

Model ReleasesDGX agent

arXiv:2508.01253v2 Announce Type: replace Abstract: Existing studies typically investigate domain shift and category shift as independent problems, however, in real-world scenarios, the two types of s

On the Sensitivity of Instruction-tuned LLMs to Harmful Sentences in Long Inputs

Model ReleasesDGX agent

arXiv:2510.05864v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly operate on long inputs, yet their behavior when harmful sentences are sparsely embedded within such inputs

Open-Weight LLM Fine-Tuning Defenses are Susceptible to Simple Attacks

SafetyDGX agent

arXiv:2605.26526v1 Announce Type: new Abstract: Recent defenses for safeguarding open-weight large language models (LLMs) are intended to prevent adversarial usage. Underlying these defenses is an ass

Periodic Topological Deep Learning for Polymer Design and Discovery

Model ReleasesDGX agent

arXiv:2605.26833v1 Announce Type: cross Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging.

← Previous
1…404405406407408…1053
Next →