AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Agents

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions

DGX agent

arXiv:2602.05843v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has catalyzed the development of autonomous agents capable of navigating complex environments.

agentsarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

PiL-World: A Chunk-Wise World Model for VLA Policy-in-the-Loop Evaluation

DGX agent

arXiv:2606.05773v1 Announce Type: new Abstract: Vision-language-action (VLA) policies operate in a closed loop in real-world robot tasks: a robot observes the scene, executes an action chunk, and cond

safetyarxiv-cs-ro
5 Jun 2026
Research

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

DGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

DGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Research

Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

DGX agent

arXiv:2606.05161v1 Announce Type: cross Abstract: Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the

researcharxiv-cs-cl
4 Jun 2026
Safety

Causal Multi-fidelity Surrogate Forward and Inverse Models for ICF Implosions

DGX agent

arXiv:2509.05510v3 Announce Type: replace-cross Abstract: Continued progress in inertial confinement fusion (ICF) requires solving inverse problems relating experimental observations to simulation inp

safetyarxiv-cs-lg
4 Jun 2026
Safety

Covert Influence Between Language Models

DGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

safetyarxiv-cs-cl
4 Jun 2026
Safety

Efficient Adversarial Attacks on High-dimensional Offline Bandits

DGX agent

arXiv:2602.01658v2 Announce Type: replace-cross Abstract: Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large

safetyarxiv-cs-ai
4 Jun 2026
Research

Efficient and Training-Free Single-Image Diffusion Models

DGX agent

arXiv:2606.04299v1 Announce Type: new Abstract: We consider the problem of generating images whose internal structure -- defined by the distribution of patches across multiple scales -- matches that o

researcharxiv-cs-cv
4 Jun 2026
Safety

Geospatial Foundation Models to Enable Progress on Sustainable Development Goals

DGX agent

arXiv:2505.24528v3 Announce Type: replace Abstract: Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and c

safetyarxiv-cs-cv
4 Jun 2026
Tutorials

Learning What to Learn: Stage-Specific Data Sets for SFT-then-RL in Small Language Model Reasoning

DGX agent

arXiv:2606.04466v1 Announce Type: new Abstract: Post-training Small Language Models (SLMs) for reasoning typically follows an SFT-then-RL pipeline, yet existing work rarely considers what data should

tutorialsarxiv-cs-cl
4 Jun 2026
Safety

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

DGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

safetyarxiv-cs-ai
4 Jun 2026
Research

Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis

DGX agent

arXiv:2603.09391v2 Announce Type: replace-cross Abstract: Engine sounds originate from sequential exhaust pressure pulses rather than sustained harmonic oscillations. While neural synthesis methods ty

researcharxiv-cs-ai
4 Jun 2026
Research

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

DGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

researcharxiv-cs-lg
4 Jun 2026
Research

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

DGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

researcharxiv-cs-cv
4 Jun 2026
Research

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

DGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

researcharxiv-cs-lg
4 Jun 2026
Tutorials

SurvPFN: Towards Foundation Models for Survival Predictions

DGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

tutorialsarxiv-cs-lg
4 Jun 2026
Safety

Test-time reward-guided alignment of language models by importance sampling on pre-logit space

DGX agent

arXiv:2510.26219v3 Announce Type: replace-cross Abstract: Test-time alignment of large language models (LLMs) attracts attention because fine-tuning of LLMs requires high computational costs. In this

safetyarxiv-cs-ai
4 Jun 2026
Local Ai

Towards Estimating Normal and Shear Interface Pressures in Prosthetic Sockets via Least Squares and Mechanics Modeling

DGX agent

arXiv:2606.04222v1 Announce Type: new Abstract: Prosthetic socket fitting remains largely manual and iterative, and objective fit metrics are still limited. Part of the challenge is the lack of long-t

local-aiarxiv-cs-ro
4 Jun 2026
Safety

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

DGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

safetyarxiv-cs-cv
4 Jun 2026
Research

Who Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have

DGX agent

arXiv:2606.05107v1 Announce Type: cross Abstract: We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tu

researcharxiv-cs-ai
4 Jun 2026
Local Ai

A Graph Foundation Model with Spectral Parsing and Prototype-Guided Spatial Propagation

DGX agent

arXiv:2606.03315v1 Announce Type: new Abstract: Graph foundation models aim to learn transferable knowledge from diverse graphs for generalization to unseen graphs and tasks. Unlike text and images, g

local-aiarxiv-cs-lg
3 Jun 2026
Safety

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

DGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

safetyarxiv-cs-ai
3 Jun 2026
Safety

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

DGX agent

arXiv:2606.03948v1 Announce Type: new Abstract: We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy Alig

safetyarxiv-cs-cl
3 Jun 2026
Research

Beyond False Stability: High-Noise Drift Gating for Test-Time Adversarial Defenses in Vision-Language Models

DGX agent

arXiv:2606.03730v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP show strong zero-shot generalization but remain highly vulnerable to adversarial attacks. Adversarial trainin

researcharxiv-cs-cv
3 Jun 2026
Research

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

DGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

researcharxiv-cs-ai
3 Jun 2026
Research

Conformal Language Modeling via Posterior Sampling

DGX agent

arXiv:2606.03731v1 Announce Type: new Abstract: Large Language Models remain plagued by hallucinations. Recent work has sought to tame their prevalence using statistical techniques based on conformal

researcharxiv-cs-lg
3 Jun 2026
Applications

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study

DGX agent

arXiv:2606.03693v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) are typically evaluated on English radiology visual question answering benchmarks, leaving their robustness under

applicationsarxiv-cs-cl
3 Jun 2026
Safety

Exploring Adversarial Robustness and Safety Alignment in Multilingual Multi-Modal Large Language Models

DGX agent

arXiv:2606.03793v1 Announce Type: new Abstract: Multimodal Large Language Models integrate visual perception into language reasoning, introducing a continuous attack surface susceptible to adversarial

safetyarxiv-cs-cl
3 Jun 2026
Applications

Hybrid Autoregressive-Diffusion Model for Real-Time Sign Language Production

DGX agent

arXiv:2507.09105v4 Announce Type: replace Abstract: Earlier Sign Language Production (SLP) models typically relied on autoregressive decoding, which naturally preserves temporal causality but suffers

applicationsarxiv-cs-cv
3 Jun 2026
Research

Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models

DGX agent

arXiv:2606.03988v1 Announce Type: new Abstract: Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many s

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

DGX agent

arXiv:2606.03539v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they main

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

DGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

model-releasesarxiv-cs-cv
3 Jun 2026
Safety

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

DGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

safetyarxiv-cs-ai
3 Jun 2026
Safety

Multi-component Causal Tracing in Large Language Models

DGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

safetyarxiv-cs-cl
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Local Ai

ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models

DGX agent

arXiv:2511.19959v2 Announce Type: replace Abstract: Federated learning (FL) has been extensively studied as a privacy-preserving training paradigm. Recently, federated block coordinate descent scheme

local-aiarxiv-cs-lg
3 Jun 2026
Tutorials

Physics-informed diffusion models in spectral space

DGX agent

arXiv:2602.09708v2 Announce Type: replace-cross Abstract: We propose physics-informed spectral diffusion (PISD), a methodology that combines generative latent diffusion models with physics-informed ma

tutorialsarxiv-cs-ai
3 Jun 2026
Safety

Post-Hoc Robustness for Model-Based Reinforcement Learning

DGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

safetyarxiv-cs-ai
3 Jun 2026
Applications

Privacy-Aware Decoding: Mitigating Privacy Leakage of Large Language Models in Retrieval-Augmented Generation

DGX agent

arXiv:2508.03098v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) enhances the factual accuracy of large language models (LLMs) by conditioning outputs on external knowledge sou

applicationsarxiv-cs-cl
3 Jun 2026
Safety

SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory

DGX agent

arXiv:2511.16275v4 Announce Type: replace-cross Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language models (LLMs) in safety-critical scenarios, as it enables t

safetyarxiv-cs-ai
3 Jun 2026
Research

Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression

DGX agent

arXiv:2602.17063v2 Announce Type: replace-cross Abstract: Sub-bit model compression targets storage below one bit per weight; as magnitudes are aggressively compressed, the sign bit becomes a fixed-co

researcharxiv-cs-ai
3 Jun 2026
Research

Social Caption: Evaluating Social Understanding in Multimodal Models

DGX agent

arXiv:2601.14569v2 Announce Type: replace Abstract: Social understanding abilities are crucial for multimodal large language models (MLLMs) to interpret human social interactions. We introduce SOCIAL

researcharxiv-cs-cl
3 Jun 2026
Agents

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

DGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

agentsarxiv-cs-ai
3 Jun 2026
Research

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

DGX agent

arXiv:2606.03645v1 Announce Type: cross Abstract: Large Language Models exhibit paradoxical fragility in fundamental arithmetic, implying a disconnect between internal computation and discrete output.

researcharxiv-cs-ai
3 Jun 2026
Safety

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

DGX agent

arXiv:2606.03127v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models trained on large-scale data have made remarkable progress, but they remain vulnerable to distribution shifts at depl

safetyarxiv-cs-ro
3 Jun 2026
Tutorials

Your Autoregressive Model Already Reveals the Causal Graph

DGX agent

arXiv:2602.01135v3 Announce Type: replace Abstract: Autoregressive models trained via next-token prediction implicitly learn the conditional independence structure of their data-generating process. We

tutorialsarxiv-cs-lg
3 Jun 2026
← Previous
1…181182183184185…1030
Next →