AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
4 Jun 2026

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

ResearchDGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

ResearchDGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SurvPFN: Towards Foundation Models for Survival Predictions

TutorialsDGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

Test-time reward-guided alignment of language models by importance sampling on pre-logit space

SafetyDGX agent

arXiv:2510.26219v3 Announce Type: replace-cross Abstract: Test-time alignment of large language models (LLMs) attracts attention because fine-tuning of LLMs requires high computational costs. In this

Towards Estimating Normal and Shear Interface Pressures in Prosthetic Sockets via Least Squares and Mechanics Modeling

Local AiDGX agent

arXiv:2606.04222v1 Announce Type: new Abstract: Prosthetic socket fitting remains largely manual and iterative, and objective fit metrics are still limited. Part of the challenge is the lack of long-t

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

SafetyDGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

Who Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have

ResearchDGX agent

arXiv:2606.05107v1 Announce Type: cross Abstract: We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tu

3 Jun 2026

A Graph Foundation Model with Spectral Parsing and Prototype-Guided Spatial Propagation

Local AiDGX agent

arXiv:2606.03315v1 Announce Type: new Abstract: Graph foundation models aim to learn transferable knowledge from diverse graphs for generalization to unseen graphs and tasks. Unlike text and images, g

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

SafetyDGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

SafetyDGX agent

arXiv:2606.03948v1 Announce Type: new Abstract: We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy Alig

Beyond False Stability: High-Noise Drift Gating for Test-Time Adversarial Defenses in Vision-Language Models

ResearchDGX agent

arXiv:2606.03730v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP show strong zero-shot generalization but remain highly vulnerable to adversarial attacks. Adversarial trainin

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

ResearchDGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

Conformal Language Modeling via Posterior Sampling

ResearchDGX agent

arXiv:2606.03731v1 Announce Type: new Abstract: Large Language Models remain plagued by hallucinations. Recent work has sought to tame their prevalence using statistical techniques based on conformal

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study

ApplicationsDGX agent

arXiv:2606.03693v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) are typically evaluated on English radiology visual question answering benchmarks, leaving their robustness under

Exploring Adversarial Robustness and Safety Alignment in Multilingual Multi-Modal Large Language Models

SafetyDGX agent

arXiv:2606.03793v1 Announce Type: new Abstract: Multimodal Large Language Models integrate visual perception into language reasoning, introducing a continuous attack surface susceptible to adversarial

Hybrid Autoregressive-Diffusion Model for Real-Time Sign Language Production

ApplicationsDGX agent

arXiv:2507.09105v4 Announce Type: replace Abstract: Earlier Sign Language Production (SLP) models typically relied on autoregressive decoding, which naturally preserves temporal causality but suffers

Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models

ResearchDGX agent

arXiv:2606.03988v1 Announce Type: new Abstract: Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many s

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

Model ReleasesDGX agent

arXiv:2606.03539v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they main

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

Model ReleasesDGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

SafetyDGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

Multi-component Causal Tracing in Large Language Models

SafetyDGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

SafetyDGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models

Local AiDGX agent

arXiv:2511.19959v2 Announce Type: replace Abstract: Federated learning (FL) has been extensively studied as a privacy-preserving training paradigm. Recently, federated block coordinate descent scheme

Physics-informed diffusion models in spectral space

TutorialsDGX agent

arXiv:2602.09708v2 Announce Type: replace-cross Abstract: We propose physics-informed spectral diffusion (PISD), a methodology that combines generative latent diffusion models with physics-informed ma

Post-Hoc Robustness for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

Privacy-Aware Decoding: Mitigating Privacy Leakage of Large Language Models in Retrieval-Augmented Generation

ApplicationsDGX agent

arXiv:2508.03098v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) enhances the factual accuracy of large language models (LLMs) by conditioning outputs on external knowledge sou

SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory

SafetyDGX agent

arXiv:2511.16275v4 Announce Type: replace-cross Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language models (LLMs) in safety-critical scenarios, as it enables t

Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression

ResearchDGX agent

arXiv:2602.17063v2 Announce Type: replace-cross Abstract: Sub-bit model compression targets storage below one bit per weight; as magnitudes are aggressively compressed, the sign bit becomes a fixed-co

Social Caption: Evaluating Social Understanding in Multimodal Models

ResearchDGX agent

arXiv:2601.14569v2 Announce Type: replace Abstract: Social understanding abilities are crucial for multimodal large language models (MLLMs) to interpret human social interactions. We introduce SOCIAL

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

AgentsDGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

ResearchDGX agent

arXiv:2606.03645v1 Announce Type: cross Abstract: Large Language Models exhibit paradoxical fragility in fundamental arithmetic, implying a disconnect between internal computation and discrete output.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.03127v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models trained on large-scale data have made remarkable progress, but they remain vulnerable to distribution shifts at depl

Your Autoregressive Model Already Reveals the Causal Graph

TutorialsDGX agent

arXiv:2602.01135v3 Announce Type: replace Abstract: Autoregressive models trained via next-token prediction implicitly learn the conditional independence structure of their data-generating process. We

2 Jun 2026

A Monosemantic Attribution Framework for Stable Interpretability in Clinical Neuroscience Transformer-Based Language Models

SafetyDGX agent

arXiv:2601.17952v2 Announce Type: replace-cross Abstract: Interpretability remains a key challenge for deploying language models (LM) in clinical settings such as progression diagnosis of Alzheimer di

AdaptiveK: Complexity-Driven Sparse Autoencoders for Interpretable Language Model Representations

TutorialsDGX agent

arXiv:2508.17320v3 Announce Type: replace Abstract: Understanding the internal representations of large language models (LLMs) remains a central challenge for interpretability research. Sparse autoenc

Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning

ResearchDGX agent

arXiv:2606.01558v1 Announce Type: new Abstract: The effectiveness of Chain-of-Thought (CoT) prompting in Multimodal Large Language Models (MLLMs) remains uncertain: across several visual reasoning ben

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

Model ReleasesDGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

Belief Consistency Between Foundation-Model Evidence and Geometric Perception in Persistent Robotic Maps

AgentsDGX agent

arXiv:2606.00318v1 Announce Type: cross Abstract: Persistent maps used by autonomous robots increasingly fuse a geometric perception stack whose assertions are well-characterized with a foundation-mod

Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Design

ResearchDGX agent

arXiv:2507.15336v3 Announce Type: replace-cross Abstract: Designing high-performance neural networks for new tasks requires balancing optimization quality with search efficiency. Current methods fail

Beyond Pure Sampling: Hybrid Optimization Mechanisms for Non-Convex Model Predictive Control

Local AiDGX agent

arXiv:2606.00737v1 Announce Type: new Abstract: This paper investigates the optimization mechanisms of non-convex Model Predictive Control (MPC) using the Maximum Entropy Differential Dynamic Programm

Bridging Topology and Deep Representation Learning: A TDA-ViT Fusion Model for Four-Class Brain Tumor Classification

ResearchDGX agent

arXiv:2606.00927v1 Announce Type: new Abstract: Accurate brain tumor classification from magnetic resonance imaging (MRI) is a key requirement for early diagnosis and clinical decision-making. Vision

Canonicalized Stable-List Replay for Private Federated Continual Learning over Language-Model Embeddings

ResearchDGX agent

arXiv:2606.00426v1 Announce Type: new Abstract: Federated continual learning (FCL) lets distributed clients adapt language-model heads to evolving NLP tasks without sharing raw text. Under user-level

Challenges in the calibration of tree-based models for imbalanced classification

SafetyDGX agent

arXiv:2412.16209v5 Announce Type: replace Abstract: When using machine learning for imbalanced binary classification problems, it is common to subsample the majority class to create a (more) balanced

Closed-Loop Neural Activation Control in Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.00269v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use

Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards

SafetyDGX agent

arXiv:2606.02194v1 Announce Type: new Abstract: Distilling expert demonstration data into large generative models using behavioral cloning is a scalable approach to learning capable policies for robot

Cross-modal linkage risk in clinical vision-language models

SafetyDGX agent

arXiv:2606.02276v1 Announce Type: cross Abstract: Vision-language models (VLMs) trained on paired chest radiographs and radiology reports learn a shared embedding space that can preserve instance-leve

Detect Before You Leap: Mirage Detection in Vision-Language Models

SafetyDGX agent

arXiv:2606.00435v1 Announce Type: cross Abstract: Vision-language models (VLMs) can produce confident visual answers even when the required visual evidence is missing, blank, or unrelated to the quest

Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games

Model ReleasesDGX agent

arXiv:2606.00103v1 Announce Type: new Abstract: We introduce a multi-turn interactive framework for reasoning evaluation that treats reasoning as active evidence acquisition and belief updating. Where

From Capability Models to Automated Planning: An AAS-Native Approach for Automatic PDDL Generation

ApplicationsDGX agent

arXiv:2606.02167v1 Announce Type: new Abstract: Engineers designing production systems need to verify that a given layout supports all required production sequences. Automated planning techniques can

@garrytan Model, application, and infrastructure layers trying to commoditize each other

IndustryDGX agent

This post discusses how the model, application, and infrastructure layers of AI are competing to commoditize each other, suggesting a dynamic where each layer is attempting to move upstream or downstr

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

AgentsDGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

Hoeffding Concept Bottleneck Models with Applications to Overhead Images

ResearchDGX agent

arXiv:2606.00082v1 Announce Type: cross Abstract: Explainability of deep learning algorithms is critical for computer-vision applications with high-stake decisions. Concept bottleneck models (CBM) hav

Identifiable Markov Switching Models with Instantaneous Effects and Exponential Families

ResearchDGX agent

arXiv:2606.02231v1 Announce Type: cross Abstract: Temporal systems often exhibit non-stationary behaviour, such as seasonal climate variation or glucose fluctuations in patients with type-1 diabetes.

iLRM: An Iterative Large 3D Reconstruction Model

ResearchDGX agent

arXiv:2507.23277v3 Announce Type: replace Abstract: Feed-forward 3D modeling has emerged as a promising approach for rapid and high-quality 3D reconstruction. In particular, directly generating explic

InfoMerge: Information-aware Token Compression for Efficient Video Large Language Models

ApplicationsDGX agent

arXiv:2606.02161v1 Announce Type: cross Abstract: Video Large Language Models (Video-LLMs) achieve strong performance in video understanding, but their excessive visual tokens bring substantial comput

Language-Native Materials Processing Design by Lightly Structured Text Database and Reasoning Large Language Model

Model ReleasesDGX agent

arXiv:2509.06093v4 Announce Type: replace-cross Abstract: Materials synthesis procedures are predominantly documented as narrative text in papers, protocols, and laboratory records, placing them beyon

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

HardwareDGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

Learning Label-Efficient Interpretable Medical Image Diagnosis via Semi-supervised Hypergraph Concept Bottleneck Model

ResearchDGX agent

arXiv:2606.01698v1 Announce Type: new Abstract: Deep learning has revolutionized medical image analysis, delivering exceptional diagnostic accuracy across diverse applications. Yet, the lack of interp

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

SafetyDGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

← Previous
1…181182183184185…1018
Next →