AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

A Monosemantic Attribution Framework for Stable Interpretability in Clinical Neuroscience Transformer-Based Language Models

DGX agent

arXiv:2601.17952v2 Announce Type: replace-cross Abstract: Interpretability remains a key challenge for deploying language models (LM) in clinical settings such as progression diagnosis of Alzheimer di

safetyarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

AdaptiveK: Complexity-Driven Sparse Autoencoders for Interpretable Language Model Representations

DGX agent

arXiv:2508.17320v3 Announce Type: replace Abstract: Understanding the internal representations of large language models (LLMs) remains a central challenge for interpretability research. Sparse autoenc

tutorialsarxiv-cs-lg
2 Jun 2026
Research

Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning

DGX agent

arXiv:2606.01558v1 Announce Type: new Abstract: The effectiveness of Chain-of-Thought (CoT) prompting in Multimodal Large Language Models (MLLMs) remains uncertain: across several visual reasoning ben

researcharxiv-cs-cv
2 Jun 2026
Model Releases

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

DGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Belief Consistency Between Foundation-Model Evidence and Geometric Perception in Persistent Robotic Maps

DGX agent

arXiv:2606.00318v1 Announce Type: cross Abstract: Persistent maps used by autonomous robots increasingly fuse a geometric perception stack whose assertions are well-characterized with a foundation-mod

agentsarxiv-cs-cv
2 Jun 2026
Research

Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Design

DGX agent

arXiv:2507.15336v3 Announce Type: replace-cross Abstract: Designing high-performance neural networks for new tasks requires balancing optimization quality with search efficiency. Current methods fail

researcharxiv-cs-ai
2 Jun 2026
Local Ai

Beyond Pure Sampling: Hybrid Optimization Mechanisms for Non-Convex Model Predictive Control

DGX agent

arXiv:2606.00737v1 Announce Type: new Abstract: This paper investigates the optimization mechanisms of non-convex Model Predictive Control (MPC) using the Maximum Entropy Differential Dynamic Programm

local-aiarxiv-cs-ro
2 Jun 2026
Research

Bridging Topology and Deep Representation Learning: A TDA-ViT Fusion Model for Four-Class Brain Tumor Classification

DGX agent

arXiv:2606.00927v1 Announce Type: new Abstract: Accurate brain tumor classification from magnetic resonance imaging (MRI) is a key requirement for early diagnosis and clinical decision-making. Vision

researcharxiv-cs-cv
2 Jun 2026
Research

Canonicalized Stable-List Replay for Private Federated Continual Learning over Language-Model Embeddings

DGX agent

arXiv:2606.00426v1 Announce Type: new Abstract: Federated continual learning (FCL) lets distributed clients adapt language-model heads to evolving NLP tasks without sharing raw text. Under user-level

researcharxiv-cs-lg
2 Jun 2026
Safety

Challenges in the calibration of tree-based models for imbalanced classification

DGX agent

arXiv:2412.16209v5 Announce Type: replace Abstract: When using machine learning for imbalanced binary classification problems, it is common to subsample the majority class to create a (more) balanced

safetyarxiv-cs-lg
2 Jun 2026
Safety

Closed-Loop Neural Activation Control in Vision-Language-Action Models

DGX agent

arXiv:2606.00269v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use

safetyarxiv-cs-ai
2 Jun 2026
Safety

Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards

DGX agent

arXiv:2606.02194v1 Announce Type: new Abstract: Distilling expert demonstration data into large generative models using behavioral cloning is a scalable approach to learning capable policies for robot

safetyarxiv-cs-lg
2 Jun 2026
Safety

Cross-modal linkage risk in clinical vision-language models

DGX agent

arXiv:2606.02276v1 Announce Type: cross Abstract: Vision-language models (VLMs) trained on paired chest radiographs and radiology reports learn a shared embedding space that can preserve instance-leve

safetyarxiv-cs-ai
2 Jun 2026
Safety

Detect Before You Leap: Mirage Detection in Vision-Language Models

DGX agent

arXiv:2606.00435v1 Announce Type: cross Abstract: Vision-language models (VLMs) can produce confident visual answers even when the required visual evidence is missing, blank, or unrelated to the quest

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games

DGX agent

arXiv:2606.00103v1 Announce Type: new Abstract: We introduce a multi-turn interactive framework for reasoning evaluation that treats reasoning as active evidence acquisition and belief updating. Where

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

From Capability Models to Automated Planning: An AAS-Native Approach for Automatic PDDL Generation

DGX agent

arXiv:2606.02167v1 Announce Type: new Abstract: Engineers designing production systems need to verify that a given layout supports all required production sequences. Automated planning techniques can

applicationsarxiv-cs-ai
2 Jun 2026
Agents

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

DGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

agentsarxiv-cs-ro
2 Jun 2026
Research

Hoeffding Concept Bottleneck Models with Applications to Overhead Images

DGX agent

arXiv:2606.00082v1 Announce Type: cross Abstract: Explainability of deep learning algorithms is critical for computer-vision applications with high-stake decisions. Concept bottleneck models (CBM) hav

researcharxiv-cs-ai
2 Jun 2026
Research

Identifiable Markov Switching Models with Instantaneous Effects and Exponential Families

DGX agent

arXiv:2606.02231v1 Announce Type: cross Abstract: Temporal systems often exhibit non-stationary behaviour, such as seasonal climate variation or glucose fluctuations in patients with type-1 diabetes.

researcharxiv-cs-lg
2 Jun 2026
Research

iLRM: An Iterative Large 3D Reconstruction Model

DGX agent

arXiv:2507.23277v3 Announce Type: replace Abstract: Feed-forward 3D modeling has emerged as a promising approach for rapid and high-quality 3D reconstruction. In particular, directly generating explic

researcharxiv-cs-cv
2 Jun 2026
Applications

InfoMerge: Information-aware Token Compression for Efficient Video Large Language Models

DGX agent

arXiv:2606.02161v1 Announce Type: cross Abstract: Video Large Language Models (Video-LLMs) achieve strong performance in video understanding, but their excessive visual tokens bring substantial comput

applicationsarxiv-cs-cl
2 Jun 2026
Model Releases

Language-Native Materials Processing Design by Lightly Structured Text Database and Reasoning Large Language Model

DGX agent

arXiv:2509.06093v4 Announce Type: replace-cross Abstract: Materials synthesis procedures are predominantly documented as narrative text in papers, protocols, and laboratory records, placing them beyon

model-releasesarxiv-cs-ai
2 Jun 2026
Hardware

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

DGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

hardwarearxiv-cs-ai
2 Jun 2026
Research

Learning Label-Efficient Interpretable Medical Image Diagnosis via Semi-supervised Hypergraph Concept Bottleneck Model

DGX agent

arXiv:2606.01698v1 Announce Type: new Abstract: Deep learning has revolutionized medical image analysis, delivering exceptional diagnostic accuracy across diverse applications. Yet, the lack of interp

researcharxiv-cs-cv
2 Jun 2026
Safety

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

DGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2603.00171v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) are shifting towards 'Thinking with Images' by actively exploring image details. While effective, lar

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning

DGX agent

arXiv:2606.01301v1 Announce Type: new Abstract: Hallucinations in medical large language models (LLMs) pose serious risks for clinical decision support, particularly when models must reason over compl

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Mitigating Hallucinations in Large Language Models Via Decoder Layer Skipping

DGX agent

arXiv:2606.00819v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across diverse natural language tasks, yet their outputs often suffer from hallucinations

researcharxiv-cs-ai
2 Jun 2026
Research

MLLM-Microscope: Unlocking Hidden Structure Within Multimodal Large Language Models

DGX agent

arXiv:2606.00909v1 Announce Type: cross Abstract: This work presents MLLM-Microscope, a novel system designed for analyzing the hidden representations within Multimodal Large Language Models (MLLMs).

researcharxiv-cs-ai
2 Jun 2026
Agents

Modeling Distinct Human Interaction in Web Agents

DGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

agentsarxiv-cs-cl
2 Jun 2026
Local Ai

Non-vacuous Generalization Bounds for Deep Neural Networks without any modification to the trained models

DGX agent

arXiv:2503.07325v2 Announce Type: replace Abstract: Understanding and certifying the behavior of modern deep neural networks remains a fundamental challenge in reliable machine learning. We introduce

local-aiarxiv-cs-lg
2 Jun 2026
Research

Per-Group Error, Not Total MSE: Fine-Tuning Vision-Language-Action Models for 11-DoF Mobile Manipulation

DGX agent

arXiv:2606.00253v1 Announce Type: cross Abstract: Fine-tuning Vision-Language-Action (VLA) models for mobile manipulators with heterogeneous joint spaces can produce a counterintuitive result: the che

researcharxiv-cs-lg
2 Jun 2026
Model Releases

Profiling Privacy Preservation Against Gradient Inversion Attacks in Tabular Federated Learning

DGX agent

arXiv:2606.00986v1 Announce Type: new Abstract: Federated learning (FL) enables multiple data holders to train machine learning models collaboratively without centralizing raw data, making it useful i

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai

Relational Intervention During Functional Collapse in Large Language Models: A Lexical-Statistical Ablation and a Structure x Register Factorial

DGX agent

arXiv:2606.00935v1 Announce Type: new Abstract: We test whether a relational-style intervention delivered during functional collapse in a small language model produces post-collapse behavior distingui

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

DGX agent

arXiv:2606.00773v1 Announce Type: new Abstract: Vision-language-action (VLA) benchmarks measure whether a policy completes a requested manipulation task, but binary success can hide safety-relevant tr

model-releasesarxiv-cs-ro
2 Jun 2026
Research

Sample Complexity and Decision-Theoretic Guarantees for Bayesian Model Averaging over Decision Trees with Catalan-Exponential Priors

DGX agent

arXiv:2606.01340v1 Announce Type: new Abstract: We ask: when do Bayesian model averaging (BMA) weights over decision trees carry sufficient epistemic information to justify committed exploitation of t

researcharxiv-cs-lg
2 Jun 2026
Research

ShapeLib: Designing a library of programmatic 3D shape abstractions with Large Language Models

DGX agent

arXiv:2502.08884v3 Announce Type: replace-cross Abstract: We present ShapeLib, the first method that uses the priors of Large Language Models (LLMs) to design libraries of programmatic 3D shape abstra

researcharxiv-cs-ai
2 Jun 2026
Safety

SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models

DGX agent

arXiv:2606.00664v1 Announce Type: cross Abstract: Embodied world models have emerged as a promising paradigm in robotics by predicting how robot actions affect the surrounding scene. However, the roll

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

Subliminal Learning is a LoRA Artifact

DGX agent

arXiv:2606.00831v1 Announce Type: new Abstract: Subliminal learning is a phenomenon where language models can transmit behavioral traits to other models through seemingly innocuous data (Cloud et al.,

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Training-free image inversion for one-step diffusion models

DGX agent

arXiv:2606.01380v1 Announce Type: new Abstract: In this work, we introduce a novel training-free inversion (TFinv) framework for one-step diffusion models,addressing key challenges in real image inver

safetyarxiv-cs-cv
2 Jun 2026
Agents

Unified Driving Tokens: Representation- and Geometry-Guided Discrete Tokenizer for Driving World Models and Planning

DGX agent

arXiv:2606.01935v1 Announce Type: new Abstract: Discrete visual tokens should provide a compact representation for both token-based world modeling and planning in autonomous driving. However, most tok

agentsarxiv-cs-cv
2 Jun 2026
Safety

VERA: Variational Inference Framework for Jailbreaking Large Language Models

DGX agent

arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabi

safetyarxiv-cs-cl
2 Jun 2026
Research

Vision Language Models Cannot Reason About Physical Transformation

DGX agent

arXiv:2603.07109v2 Announce Type: replace Abstract: Understanding physical transformations is fundamental for reasoning in dynamic environments. While Vision Language Models (VLMs) show promise in emb

researcharxiv-cs-ai
2 Jun 2026
Safety

Visual Persuasion: What Influences Decisions of Vision-Language Models?

DGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

safetyarxiv-cs-ai
2 Jun 2026
Applications

WALL-WM: Carving World Action Modeling at the Event Joints

DGX agent

arXiv:2606.01955v1 Announce Type: cross Abstract: WALL-WM is a World Action Model that shifts video-action learning from chunk-centric optimization to event-grounded Vision-Language-Action pretraining

applicationsarxiv-cs-cv
2 Jun 2026
Tutorials

When Tabular Foundation Models Transfer Across Modalities: A Systematic Evaluation Across 95 Datasets, 7 Modalities, and Two Regimes

DGX agent

arXiv:2606.02106v1 Announce Type: new Abstract: We present a single classification pipeline that combines an Equiangular Tight Frame (ETF) preprocessing stage with a tabular foundation model for in-co

tutorialsarxiv-cs-lg
2 Jun 2026
Research

Why Do Self-Harm Prediction Models Struggle to Generalise? Lexical and Semantic Variations in Emergency Department Triage Notes

DGX agent

arXiv:2606.01678v1 Announce Type: new Abstract: Self-harm presentations to emergency departments (EDs) are strongly associated with higher suicide risk. NLP models have shown robust performance in det

researcharxiv-cs-cl
2 Jun 2026
Applications

Why Do Time Series Models Need Long Context Windows?

DGX agent

arXiv:2606.01999v1 Announce Type: cross Abstract: Modern deep learning models for forecasting groups of time series rely on increasingly longer observation windows. However, the benefit of increasing

applicationsarxiv-cs-ai
2 Jun 2026
← Previous
1…182183184185186…1030
Next →