AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
5 Jun 2026

Framing, Judging, Steering: An Assessable Competency Model for Teach-ing Students to Reason With Generative AI

ResearchDGX agent

arXiv:2606.05983v1 Announce Type: cross Abstract: Generative AI makes answers easy and understanding hard, and uncritical use invites cognitive offloading. Schools still measure unaided performance, y

Geometry-Aware Dataset Condensation for Diffusion Model Training

SafetyDGX agent

arXiv:2606.05883v1 Announce Type: new Abstract: Dataset condensation aims to construct compact datasets from real data via synthesis or selection. However, existing approaches are ill-suited for diffu

IR3DE: A Linear Router for Large Language Models

ResearchDGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Large Language Models are Perplexed by some Political Parties

SafetyDGX agent

arXiv:2606.05937v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used, including in political applications, but their political fairness has been little studied. We assess

NAVIRA: Decoupled Stochastic Remasking for Masked Diffusion Language Models

Local AiDGX agent

arXiv:2606.06031v1 Announce Type: new Abstract: Masked diffusion language models generate text by iteratively unmasking many tokens in parallel, but this speed comes with a correction problem: tokens

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions

AgentsDGX agent

arXiv:2602.05843v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has catalyzed the development of autonomous agents capable of navigating complex environments.

PiL-World: A Chunk-Wise World Model for VLA Policy-in-the-Loop Evaluation

SafetyDGX agent

arXiv:2606.05773v1 Announce Type: new Abstract: Vision-language-action (VLA) policies operate in a closed loop in real-world robot tasks: a robot observes the scene, executes an action chunk, and cond

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

ResearchDGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

Model ReleasesDGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

UNIVID: Unified Vision-Language Model for Video Moderation

SafetyDGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

4 Jun 2026

Best Visual Reasoning Model in 2026 (Including APIs) [D]

ResearchDGX agent

Gemini 3.1 Pro and Gemini 3-Pro lead visual reasoning benchmarks , with GPT-5.2, Kimi-K2.5, and GPT-5.2-Pro following . A 2026 evaluation benchmarked 15 leading multimodal models on visual reasoning a

Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

ResearchDGX agent

arXiv:2606.05161v1 Announce Type: cross Abstract: Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the

Causal Multi-fidelity Surrogate Forward and Inverse Models for ICF Implosions

SafetyDGX agent

arXiv:2509.05510v3 Announce Type: replace-cross Abstract: Continued progress in inertial confinement fusion (ICF) requires solving inverse problems relating experimental observations to simulation inp

Covert Influence Between Language Models

SafetyDGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

Efficient Adversarial Attacks on High-dimensional Offline Bandits

SafetyDGX agent

arXiv:2602.01658v2 Announce Type: replace-cross Abstract: Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large

Efficient and Training-Free Single-Image Diffusion Models

ResearchDGX agent

arXiv:2606.04299v1 Announce Type: new Abstract: We consider the problem of generating images whose internal structure -- defined by the distribution of patches across multiple scales -- matches that o

Geospatial Foundation Models to Enable Progress on Sustainable Development Goals

SafetyDGX agent

arXiv:2505.24528v3 Announce Type: replace Abstract: Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and c

Learning What to Learn: Stage-Specific Data Sets for SFT-then-RL in Small Language Model Reasoning

TutorialsDGX agent

arXiv:2606.04466v1 Announce Type: new Abstract: Post-training Small Language Models (SLMs) for reasoning typically follows an SFT-then-RL pipeline, yet existing work rarely considers what data should

Meet LM Studio's mobile app. Your local models, now in your pocket.

Local AiDGX agent

LM Studio has released a mobile app that enables users to run and access local language models on their smartphones. The app extends LM Studio's desktop functionality, allowing users to deploy and int

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

SafetyDGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

Model ReleasesDGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis

ResearchDGX agent

arXiv:2603.09391v2 Announce Type: replace-cross Abstract: Engine sounds originate from sequential exhaust pressure pulses rather than sustained harmonic oscillations. While neural synthesis methods ty

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

ResearchDGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

ResearchDGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

SurvPFN: Towards Foundation Models for Survival Predictions

TutorialsDGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

Test-time reward-guided alignment of language models by importance sampling on pre-logit space

SafetyDGX agent

arXiv:2510.26219v3 Announce Type: replace-cross Abstract: Test-time alignment of large language models (LLMs) attracts attention because fine-tuning of LLMs requires high computational costs. In this

Towards Estimating Normal and Shear Interface Pressures in Prosthetic Sockets via Least Squares and Mechanics Modeling

Local AiDGX agent

arXiv:2606.04222v1 Announce Type: new Abstract: Prosthetic socket fitting remains largely manual and iterative, and objective fit metrics are still limited. Part of the challenge is the lack of long-t

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

SafetyDGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

Who Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have

ResearchDGX agent

arXiv:2606.05107v1 Announce Type: cross Abstract: We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tu

3 Jun 2026

A Graph Foundation Model with Spectral Parsing and Prototype-Guided Spatial Propagation

Local AiDGX agent

arXiv:2606.03315v1 Announce Type: new Abstract: Graph foundation models aim to learn transferable knowledge from diverse graphs for generalization to unseen graphs and tasks. Unlike text and images, g

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

SafetyDGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

SafetyDGX agent

arXiv:2606.03948v1 Announce Type: new Abstract: We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy Alig

Beyond False Stability: High-Noise Drift Gating for Test-Time Adversarial Defenses in Vision-Language Models

ResearchDGX agent

arXiv:2606.03730v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP show strong zero-shot generalization but remain highly vulnerable to adversarial attacks. Adversarial trainin

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

ResearchDGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

Conformal Language Modeling via Posterior Sampling

ResearchDGX agent

arXiv:2606.03731v1 Announce Type: new Abstract: Large Language Models remain plagued by hallucinations. Recent work has sought to tame their prevalence using statistical techniques based on conformal

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study

ApplicationsDGX agent

arXiv:2606.03693v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) are typically evaluated on English radiology visual question answering benchmarks, leaving their robustness under

Exploring Adversarial Robustness and Safety Alignment in Multilingual Multi-Modal Large Language Models

SafetyDGX agent

arXiv:2606.03793v1 Announce Type: new Abstract: Multimodal Large Language Models integrate visual perception into language reasoning, introducing a continuous attack surface susceptible to adversarial

Hybrid Autoregressive-Diffusion Model for Real-Time Sign Language Production

ApplicationsDGX agent

arXiv:2507.09105v4 Announce Type: replace Abstract: Earlier Sign Language Production (SLP) models typically relied on autoregressive decoding, which naturally preserves temporal causality but suffers

Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models

ResearchDGX agent

arXiv:2606.03988v1 Announce Type: new Abstract: Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many s

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

Model ReleasesDGX agent

arXiv:2606.03539v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they main

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

Model ReleasesDGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

SafetyDGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

Multi-component Causal Tracing in Large Language Models

SafetyDGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

SafetyDGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models

Local AiDGX agent

arXiv:2511.19959v2 Announce Type: replace Abstract: Federated learning (FL) has been extensively studied as a privacy-preserving training paradigm. Recently, federated block coordinate descent scheme

Physics-informed diffusion models in spectral space

TutorialsDGX agent

arXiv:2602.09708v2 Announce Type: replace-cross Abstract: We propose physics-informed spectral diffusion (PISD), a methodology that combines generative latent diffusion models with physics-informed ma

Post-Hoc Robustness for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

Privacy-Aware Decoding: Mitigating Privacy Leakage of Large Language Models in Retrieval-Augmented Generation

ApplicationsDGX agent

arXiv:2508.03098v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) enhances the factual accuracy of large language models (LLMs) by conditioning outputs on external knowledge sou

SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory

SafetyDGX agent

arXiv:2511.16275v4 Announce Type: replace-cross Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language models (LLMs) in safety-critical scenarios, as it enables t

Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression

ResearchDGX agent

arXiv:2602.17063v2 Announce Type: replace-cross Abstract: Sub-bit model compression targets storage below one bit per weight; as magnitudes are aggressively compressed, the sign bit becomes a fixed-co

Social Caption: Evaluating Social Understanding in Multimodal Models

ResearchDGX agent

arXiv:2601.14569v2 Announce Type: replace Abstract: Social understanding abilities are crucial for multimodal large language models (MLLMs) to interpret human social interactions. We introduce SOCIAL

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

AgentsDGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

ResearchDGX agent

arXiv:2606.03645v1 Announce Type: cross Abstract: Large Language Models exhibit paradoxical fragility in fundamental arithmetic, implying a disconnect between internal computation and discrete output.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.03127v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models trained on large-scale data have made remarkable progress, but they remain vulnerable to distribution shifts at depl

Your Autoregressive Model Already Reveals the Causal Graph

TutorialsDGX agent

arXiv:2602.01135v3 Announce Type: replace Abstract: Autoregressive models trained via next-token prediction implicitly learn the conditional independence structure of their data-generating process. We

2 Jun 2026

A Monosemantic Attribution Framework for Stable Interpretability in Clinical Neuroscience Transformer-Based Language Models

SafetyDGX agent

arXiv:2601.17952v2 Announce Type: replace-cross Abstract: Interpretability remains a key challenge for deploying language models (LM) in clinical settings such as progression diagnosis of Alzheimer di

AdaptiveK: Complexity-Driven Sparse Autoencoders for Interpretable Language Model Representations

TutorialsDGX agent

arXiv:2508.17320v3 Announce Type: replace Abstract: Understanding the internal representations of large language models (LLMs) remains a central challenge for interpretability research. Sparse autoenc

Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning

ResearchDGX agent

arXiv:2606.01558v1 Announce Type: new Abstract: The effectiveness of Chain-of-Thought (CoT) prompting in Multimodal Large Language Models (MLLMs) remains uncertain: across several visual reasoning ben

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

Model ReleasesDGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

← Previous
1…179180181182183…1010
Next →