AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
2 Jun 2026

Belief Consistency Between Foundation-Model Evidence and Geometric Perception in Persistent Robotic Maps

AgentsDGX agent

arXiv:2606.00318v1 Announce Type: cross Abstract: Persistent maps used by autonomous robots increasingly fuse a geometric perception stack whose assertions are well-characterized with a foundation-mod

Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Design

ResearchDGX agent

arXiv:2507.15336v3 Announce Type: replace-cross Abstract: Designing high-performance neural networks for new tasks requires balancing optimization quality with search efficiency. Current methods fail

Beyond Pure Sampling: Hybrid Optimization Mechanisms for Non-Convex Model Predictive Control

Local AiDGX agent

arXiv:2606.00737v1 Announce Type: new Abstract: This paper investigates the optimization mechanisms of non-convex Model Predictive Control (MPC) using the Maximum Entropy Differential Dynamic Programm

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Bridging Topology and Deep Representation Learning: A TDA-ViT Fusion Model for Four-Class Brain Tumor Classification

ResearchDGX agent

arXiv:2606.00927v1 Announce Type: new Abstract: Accurate brain tumor classification from magnetic resonance imaging (MRI) is a key requirement for early diagnosis and clinical decision-making. Vision

Canonicalized Stable-List Replay for Private Federated Continual Learning over Language-Model Embeddings

ResearchDGX agent

arXiv:2606.00426v1 Announce Type: new Abstract: Federated continual learning (FCL) lets distributed clients adapt language-model heads to evolving NLP tasks without sharing raw text. Under user-level

Challenges in the calibration of tree-based models for imbalanced classification

SafetyDGX agent

arXiv:2412.16209v5 Announce Type: replace Abstract: When using machine learning for imbalanced binary classification problems, it is common to subsample the majority class to create a (more) balanced

Closed-Loop Neural Activation Control in Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.00269v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use

Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards

SafetyDGX agent

arXiv:2606.02194v1 Announce Type: new Abstract: Distilling expert demonstration data into large generative models using behavioral cloning is a scalable approach to learning capable policies for robot

Cross-modal linkage risk in clinical vision-language models

SafetyDGX agent

arXiv:2606.02276v1 Announce Type: cross Abstract: Vision-language models (VLMs) trained on paired chest radiographs and radiology reports learn a shared embedding space that can preserve instance-leve

Detect Before You Leap: Mirage Detection in Vision-Language Models

SafetyDGX agent

arXiv:2606.00435v1 Announce Type: cross Abstract: Vision-language models (VLMs) can produce confident visual answers even when the required visual evidence is missing, blank, or unrelated to the quest

Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games

Model ReleasesDGX agent

arXiv:2606.00103v1 Announce Type: new Abstract: We introduce a multi-turn interactive framework for reasoning evaluation that treats reasoning as active evidence acquisition and belief updating. Where

From Capability Models to Automated Planning: An AAS-Native Approach for Automatic PDDL Generation

ApplicationsDGX agent

arXiv:2606.02167v1 Announce Type: new Abstract: Engineers designing production systems need to verify that a given layout supports all required production sequences. Automated planning techniques can

@garrytan Model, application, and infrastructure layers trying to commoditize each other

IndustryDGX agent

This post discusses how the model, application, and infrastructure layers of AI are competing to commoditize each other, suggesting a dynamic where each layer is attempting to move upstream or downstr

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

AgentsDGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

Hoeffding Concept Bottleneck Models with Applications to Overhead Images

ResearchDGX agent

arXiv:2606.00082v1 Announce Type: cross Abstract: Explainability of deep learning algorithms is critical for computer-vision applications with high-stake decisions. Concept bottleneck models (CBM) hav

Identifiable Markov Switching Models with Instantaneous Effects and Exponential Families

ResearchDGX agent

arXiv:2606.02231v1 Announce Type: cross Abstract: Temporal systems often exhibit non-stationary behaviour, such as seasonal climate variation or glucose fluctuations in patients with type-1 diabetes.

iLRM: An Iterative Large 3D Reconstruction Model

ResearchDGX agent

arXiv:2507.23277v3 Announce Type: replace Abstract: Feed-forward 3D modeling has emerged as a promising approach for rapid and high-quality 3D reconstruction. In particular, directly generating explic

InfoMerge: Information-aware Token Compression for Efficient Video Large Language Models

ApplicationsDGX agent

arXiv:2606.02161v1 Announce Type: cross Abstract: Video Large Language Models (Video-LLMs) achieve strong performance in video understanding, but their excessive visual tokens bring substantial comput

Language-Native Materials Processing Design by Lightly Structured Text Database and Reasoning Large Language Model

Model ReleasesDGX agent

arXiv:2509.06093v4 Announce Type: replace-cross Abstract: Materials synthesis procedures are predominantly documented as narrative text in papers, protocols, and laboratory records, placing them beyon

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

HardwareDGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

Learning Label-Efficient Interpretable Medical Image Diagnosis via Semi-supervised Hypergraph Concept Bottleneck Model

ResearchDGX agent

arXiv:2606.01698v1 Announce Type: new Abstract: Deep learning has revolutionized medical image analysis, delivering exceptional diagnostic accuracy across diverse applications. Yet, the lack of interp

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

SafetyDGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models

Local AiDGX agent

arXiv:2603.00171v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) are shifting towards 'Thinking with Images' by actively exploring image details. While effective, lar

Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning

Model ReleasesDGX agent

arXiv:2606.01301v1 Announce Type: new Abstract: Hallucinations in medical large language models (LLMs) pose serious risks for clinical decision support, particularly when models must reason over compl

Microsoft debuts an expansion of its model families and agentic AI intelligence for developers

AgentsDGX agent

Microsoft Corp. announced an expansion to its artificial intelligence models and agentic AI infrastructure today that brings more data and context into the hands of developers and business users as th

Mitigating Hallucinations in Large Language Models Via Decoder Layer Skipping

ResearchDGX agent

arXiv:2606.00819v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across diverse natural language tasks, yet their outputs often suffer from hallucinations

MLLM-Microscope: Unlocking Hidden Structure Within Multimodal Large Language Models

ResearchDGX agent

arXiv:2606.00909v1 Announce Type: cross Abstract: This work presents MLLM-Microscope, a novel system designed for analyzing the hidden representations within Multimodal Large Language Models (MLLMs).

Modeling Distinct Human Interaction in Web Agents

AgentsDGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

Non-vacuous Generalization Bounds for Deep Neural Networks without any modification to the trained models

Local AiDGX agent

arXiv:2503.07325v2 Announce Type: replace Abstract: Understanding and certifying the behavior of modern deep neural networks remains a fundamental challenge in reliable machine learning. We introduce

Per-Group Error, Not Total MSE: Fine-Tuning Vision-Language-Action Models for 11-DoF Mobile Manipulation

ResearchDGX agent

arXiv:2606.00253v1 Announce Type: cross Abstract: Fine-tuning Vision-Language-Action (VLA) models for mobile manipulators with heterogeneous joint spaces can produce a counterintuitive result: the che

Profiling Privacy Preservation Against Gradient Inversion Attacks in Tabular Federated Learning

Model ReleasesDGX agent

arXiv:2606.00986v1 Announce Type: new Abstract: Federated learning (FL) enables multiple data holders to train machine learning models collaboratively without centralizing raw data, making it useful i

Relational Intervention During Functional Collapse in Large Language Models: A Lexical-Statistical Ablation and a Structure x Register Factorial

Local AiDGX agent

arXiv:2606.00935v1 Announce Type: new Abstract: We test whether a relational-style intervention delivered during functional collapse in a small language model produces post-collapse behavior distingui

SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.00773v1 Announce Type: new Abstract: Vision-language-action (VLA) benchmarks measure whether a policy completes a requested manipulation task, but binary success can hide safety-relevant tr

Sample Complexity and Decision-Theoretic Guarantees for Bayesian Model Averaging over Decision Trees with Catalan-Exponential Priors

ResearchDGX agent

arXiv:2606.01340v1 Announce Type: new Abstract: We ask: when do Bayesian model averaging (BMA) weights over decision trees carry sufficient epistemic information to justify committed exploitation of t

ShapeLib: Designing a library of programmatic 3D shape abstractions with Large Language Models

ResearchDGX agent

arXiv:2502.08884v3 Announce Type: replace-cross Abstract: We present ShapeLib, the first method that uses the priors of Large Language Models (LLMs) to design libraries of programmatic 3D shape abstra

SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models

SafetyDGX agent

arXiv:2606.00664v1 Announce Type: cross Abstract: Embodied world models have emerged as a promising paradigm in robotics by predicting how robot actions affect the surrounding scene. However, the roll

Subliminal Learning is a LoRA Artifact

Model ReleasesDGX agent

arXiv:2606.00831v1 Announce Type: new Abstract: Subliminal learning is a phenomenon where language models can transmit behavioral traits to other models through seemingly innocuous data (Cloud et al.,

Training-free image inversion for one-step diffusion models

SafetyDGX agent

arXiv:2606.01380v1 Announce Type: new Abstract: In this work, we introduce a novel training-free inversion (TFinv) framework for one-step diffusion models,addressing key challenges in real image inver

Unified Driving Tokens: Representation- and Geometry-Guided Discrete Tokenizer for Driving World Models and Planning

AgentsDGX agent

arXiv:2606.01935v1 Announce Type: new Abstract: Discrete visual tokens should provide a compact representation for both token-based world modeling and planning in autonomous driving. However, most tok

VERA: Variational Inference Framework for Jailbreaking Large Language Models

SafetyDGX agent

arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabi

Vision Language Models Cannot Reason About Physical Transformation

ResearchDGX agent

arXiv:2603.07109v2 Announce Type: replace Abstract: Understanding physical transformations is fundamental for reasoning in dynamic environments. While Vision Language Models (VLMs) show promise in emb

Visual Persuasion: What Influences Decisions of Vision-Language Models?

SafetyDGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

WALL-WM: Carving World Action Modeling at the Event Joints

ApplicationsDGX agent

arXiv:2606.01955v1 Announce Type: cross Abstract: WALL-WM is a World Action Model that shifts video-action learning from chunk-centric optimization to event-grounded Vision-Language-Action pretraining

When Tabular Foundation Models Transfer Across Modalities: A Systematic Evaluation Across 95 Datasets, 7 Modalities, and Two Regimes

TutorialsDGX agent

arXiv:2606.02106v1 Announce Type: new Abstract: We present a single classification pipeline that combines an Equiangular Tight Frame (ETF) preprocessing stage with a tabular foundation model for in-co

Why Do Self-Harm Prediction Models Struggle to Generalise? Lexical and Semantic Variations in Emergency Department Triage Notes

ResearchDGX agent

arXiv:2606.01678v1 Announce Type: new Abstract: Self-harm presentations to emergency departments (EDs) are strongly associated with higher suicide risk. NLP models have shown robust performance in det

Why Do Time Series Models Need Long Context Windows?

ApplicationsDGX agent

arXiv:2606.01999v1 Announce Type: cross Abstract: Modern deep learning models for forecasting groups of time series rely on increasingly longer observation windows. However, the benefit of increasing

You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models

TutorialsDGX agent

arXiv:2603.00133v2 Announce Type: replace-cross Abstract: Generative models have been shown to 'memorize' certain training data, leading to verbatim or near-verbatim generating images, which may cause

1 Jun 2026

An OpenAI model solved a famous math problem that stumped humans for 80 years

IndustryDGX agent

An OpenAI model solved the 80-year-old unit distance problem, disproving a major conjecture in discrete geometry and marking a milestone in AI-driven mathematics. The proof came from a new general-pur

Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models

SafetyDGX agent

arXiv:2510.11683v3 Announce Type: replace-cross Abstract: A key challenge in applying reinforcement learning (RL) to diffusion large language models (dLLMs) is the intractability of their likelihood f

Can Aerial VLA Models Cooperate? Evaluating Closed-Loop Air-Ground Coordination with CARLA-Air

SafetyDGX agent

arXiv:2605.31066v1 Announce Type: new Abstract: Recent aerial vision-language-action (VLA) models show promising single-UAV capabilities, such as tracking moving objects and navigating to language-spe

Circuit-Inspired High-Order Neural Networks with Unified Neural Dynamics Modeling for PDE Solving and Visual Perception

ResearchDGX agent

arXiv:2603.23977v2 Announce Type: replace-cross Abstract: Deep networks often rely on architectural heuristics to shape representation evolution, limiting their ability to model data governed by intri

Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models

ResearchDGX agent

arXiv:2605.30631v1 Announce Type: cross Abstract: While automated diagnosis systems have achieved remarkable success in computed tomography (CT)-based lung cancer screening, their development remains

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation

SafetyDGX agent

arXiv:2605.30484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promise for robotic manipulation, yet most existing policies operate reactively by directly regressing ac

Empirical Characterization of Inference-Time Elicited Probability Transformations in Large Language Models

ResearchDGX agent

arXiv:2603.19262v2 Announce Type: replace-cross Abstract: Large language models increasingly rely on inference-time procedures such as chain-of-thought reasoning, self-refinement, retrieval augmentati

Fixed-Point Masked Generative Modeling

Local AiDGX agent

arXiv:2605.31215v1 Announce Type: cross Abstract: Masked Generative Models (MGMs) enable parallel decoding and achieve strong performance across modalities, but require full-sequence bidirectional tra

FlagGAM: Rule-Based Generalized Additive Modeling for Explainable Tabular Prediction

ResearchDGX agent

arXiv:2605.31189v1 Announce Type: new Abstract: Tabular prediction in high-stakes domains requires models that are accurate, transparent, and robust to imperfect inputs. We propose FlagGAM, a rule-def

GitHub Copilot's new pricing model went into effect today, and many noted sticker shock with some saying a few hours of AI usage ate big chunks of monthly caps (Kyle Orland/Ars Technica)

IndustryDGX agent

Kyle Orland / Ars Technica: GitHub Copilot's new pricing model went into effect today, and many noted sticker shock with some saying a few hours of AI usage ate big chunks of monthly caps — In April,

GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring

SafetyDGX agent

arXiv:2605.30865v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) provides a dense view of daily metabolic physiology, yet existing generic time-series and CGM-specific foundation mo

'Intelegi Romaneste?'' A Recipe for Romanian Vision-Language Models

ResearchDGX agent

arXiv:2605.31401v1 Announce Type: new Abstract: Vision-Language Models (VLMs) largely follow the text-only LLM trajectory, excelling on English benchmarks but sharply degrading on low-resource languag

Learning effective models from network dynamics data with multiple initial conditions using weak form SINDy

TutorialsDGX agent

arXiv:2605.30432v1 Announce Type: cross Abstract: Social systems consist of networks of individuals who influence one another through social interactions. Studying how processes evolve on these networ

← Previous
1…180181182183184…1010
Next →