AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Safety

Inverting the Bellman Equation: From Q-Values to World Models

DGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

safetyarxiv-cs-lg
23 Jun 2026
Applications
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

T2MM: An LLM Supported Architecture For Inquiry-Based Modeling

DGX agent

arXiv:2606.11210v1 Announce Type: cross Abstract: Model Construction is a foundational practice in science learning that relies on visualization and interactivity. Large Language Models, increasingly

applicationsarxiv-cs-ai
11 Jun 2026
Applications

FMplex: Model Virtualization for Serving Extensible Foundation Models

DGX agent

arXiv:2606.09643v1 Announce Type: cross Abstract: Foundation models (FMs) are increasingly used as backbones for downstream tasks across language, vision, time-series, and multimodal applications. Yet

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

From Hazard Functions to Language Space: Cox-Supervised Distillation of Survival Risk into a Large Language Model

DGX agent

arXiv:2606.08945v1 Announce Type: new Abstract: We investigate whether information about time-to-event risk estimated by a Cox proportional hazards model can be transferred into a generative large lan

model-releasesarxiv-cs-lg
9 Jun 2026
Research

DSL-Topic: Improving Topic Modeling by Distilling Soft Labelsfrom Language Models

DGX agent

arXiv:2602.17907v2 Announce Type: replace-cross Abstract: Traditional neural topic models are typically optimized by reconstructing the document's Bag-of-Words (BoW) representations, overlooking conte

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models

DGX agent

arXiv:2606.04739v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong potential for automated software vulnerability detection, particularly in retrieval-augmented generatio

model-releasesarxiv-cs-ai
4 Jun 2026
Research

SharedRequest: Privacy-Preserving Model-Agnostic Inference for Large Language Models

DGX agent

arXiv:2606.05004v1 Announce Type: cross Abstract: With the widespread deployment of public large language models (LLMs) such as ChatGPT, protecting user prompt privacy has become an increasingly criti

researcharxiv-cs-ai
4 Jun 2026
Research

How Much of a Model Do We Need? Redundancy and Slimmability in Remote Sensing Foundation Models

DGX agent

arXiv:2601.22841v2 Announce Type: replace Abstract: Large-scale foundation models (FMs) in remote sensing (RS) (denoted as RS FMs) are developed following paradigms established in computer vision (CV)

researcharxiv-cs-cv
3 Jun 2026
Model Releases

Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework

DGX agent

arXiv:2602.18008v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in constructing mechanistic models from data. However, existing evaluations largely focus on s

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Disentangling Similarity and Relatedness in Topic Models

DGX agent

arXiv:2603.10619v2 Announce Type: replace Abstract: The recent success of large pre-trained language models (PLMs) has motivated their integration into topic modeling. However, PLM-augmented topic mod

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

The Case for Model Science: Verify, Explore, Steer, Refine

DGX agent

arXiv:2606.01189v1 Announce Type: new Abstract: We argue that the AI community is now ready to move beyond benchmarking and consolidate scattered efforts in model analysis into a systematic discipline

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models

DGX agent

arXiv:2601.03309v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models, which integrate pretrained large Vision-Language Models (VLM) into their policy backbone, are gaining sig

safetyarxiv-cs-ai
2 Jun 2026
Applications

Diffusion Models Preferentially Memorize Prototypical Examples or: Why Does My Diffusion Model Love Slop?

DGX agent

arXiv:2605.30642v1 Announce Type: new Abstract: Generative models have a persistent limitation: their tendency to memorize training data can create legal liabilities and erode creative diversity. Unde

applicationsarxiv-cs-lg
1 Jun 2026
Model Releases

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models

DGX agent

arXiv:2605.29360v1 Announce Type: new Abstract: Action-conditioned world models are increasingly used as scalable simulators for robot learning, yet current evaluations provide limited evidence that t

model-releasesarxiv-cs-ai
29 May 2026
Safety

Modeling Hierarchical Thinking in Large Reasoning Models

DGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

safetyarxiv-cs-ai
29 May 2026
Model Releases

A Multi-dimensional Framework for Evaluating Generalization in EEG Foundation Models

DGX agent

arXiv:2605.28563v1 Announce Type: cross Abstract: Evaluating foundation models under appropriate adaptation settings is essential for understanding the quality and transferability of the learned repre

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Apple Intelligence Foundation Language Models

DGX agent

arXiv:2407.21075v2 Announce Type: replace Abstract: We present foundation language models developed to power Apple Intelligence features, including a ~3 billion parameter model designed to run efficie

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Revisiting Metafeatures to Explain Model Differences on Tabular Data

DGX agent

arXiv:2605.28418v1 Announce Type: new Abstract: With the rise of tabular foundation models alongside traditional models still performing well on many tasks, choosing the right model for a tabular data

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios

DGX agent

arXiv:2605.27589v1 Announce Type: new Abstract: Video generation models are increasingly used as world simulators for tasks like driving and robotic manipulation. What matters in these settings is not

model-releasesarxiv-cs-cv
28 May 2026
Research

Guess the Unified Model: How Much Can We Recover from Generated Images?

DGX agent

arXiv:2605.25254v1 Announce Type: cross Abstract: With unified model-generated images now widespread online, attributing their model of origin offers a path toward transparency and deeper insight into

researcharxiv-cs-ai
26 May 2026
Research

Understanding the Impact of Geometric Foundation Models on Vision-Language-Action Models

DGX agent

arXiv:2605.24642v1 Announce Type: cross Abstract: Recent work explores new opportunities at the intersection of vision-language-action models (VLAs) and geometric foundation models (GFMs) for 3D recon

researcharxiv-cs-ro
26 May 2026
Model Releases

A Systematic Evaluation of Co-folding Model Representations for Small-Molecule Learning

DGX agent

arXiv:2602.13249v2 Announce Type: replace-cross Abstract: Small-molecule foundation models are typically pretrained on standalone molecular data, unlike vision and language models that often benefit f

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

LLM-driven design of physics-constrained constitutive models: two agents are better than one

DGX agent

arXiv:2605.23754v1 Announce Type: new Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Capability neq Interpretability: Human Interpretability of Vision Foundation Models

DGX agent

arXiv:2605.20337v1 Announce Type: new Abstract: How interpretable are the features of leading vision models? The question is increasingly pressing as these models move from research benchmarks into hi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning

DGX agent

arXiv:2605.20803v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Re

model-releasesarxiv-cs-lg
21 May 2026
Research

Efficient Long-Context Modeling in Diffusion Language Models via Block Approximate Sparse Attention

DGX agent

arXiv:2605.19726v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) enable globally coherent, bidirectional, and controllable text generation, offering advantages over traditional autoreg

researcharxiv-cs-cv
20 May 2026
Model Releases

Can Heterogeneous Language Models Be Fused?

DGX agent

arXiv:2604.01674v2 Announce Type: replace Abstract: Model merging aims to integrate multiple expert models into a single model that inherits their complementary strengths without incurring the inferen

model-releasesarxiv-cs-ai
19 May 2026
Research

Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models

DGX agent

arXiv:2602.01970v2 Announce Type: replace Abstract: Reinforcement learning enhances the reasoning capabilities of large language models but often involves high computational costs due to rollout-inten

researcharxiv-cs-ai
18 May 2026
Local Ai

FeatCal: Feature Calibration for Post-Merging Models

DGX agent

arXiv:2605.13030v1 Announce Type: cross Abstract: Model merging combines task experts into one model and avoids joint training, retraining, or deploying many expert models, but the merged model often

local-aiarxiv-cs-ai
14 May 2026
Model Releases

Reframing preprocessing selection as model-internal calibration in near-infrared spectroscopy: A large-scale benchmark of operator-adaptive PLS and Ridge models

DGX agent

arXiv:2605.13587v1 Announce Type: cross Abstract: Near-infrared spectroscopy (NIRS) is rapid and non-destructive, but reliable calibration still depends heavily on spectral preprocessing. In routine p

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

3D-Belief: Embodied Belief Inference via Generative 3D World Modeling

DGX agent

arXiv:2605.11367v1 Announce Type: new Abstract: Recent advances in visual generative models have highlighted the promise of learning generative world models. However, most existing approaches frame wo

model-releasesarxiv-cs-cv
13 May 2026
Research

Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets

DGX agent

arXiv:2605.10419v1 Announce Type: cross Abstract: This paper investigates the effectiveness of large language models (LLMs) in answering questions over datasets. We examine their performance in two sc

researcharxiv-cs-ai
12 May 2026
Research

Federated Concept-Based Models: Interpretable models with distributed supervision

DGX agent

arXiv:2602.04093v2 Announce Type: replace Abstract: Concept-based Models (CMs) enhance interpretability in deep learning by grounding predictions in human-understandable concepts. However, concept ann

researcharxiv-cs-lg
12 May 2026
Model Releases

MCP-Cosmos: World Model-Augmented Agents for Complex Task Execution in MCP Environments

DGX agent

arXiv:2605.09131v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has unified the interface between Large Language Models (LLMs) and external tools, yet a fundamental gap remains in how

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Sword: Style-Robust World Models as Simulators via Dynamic Latent Bootstrapping for VLA Policy Post-Training

DGX agent

arXiv:2605.07288v1 Announce Type: cross Abstract: The integration of Vision-Language-Action (VLA) models with World Models has gained increasing attention. One representative approach treats learned W

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

CBV: Clean-label Backdoor Attacks on Vision Language Models via Diffusion Models

DGX agent

arXiv:2605.02202v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have achieved remarkable success in tasks such as image captioning and visual question answering (VQA). However, as their

tutorialsarxiv-cs-ai
6 May 2026
Safety

Model Spec Midtraining: Improving How Alignment Training Generalizes

DGX agent

arXiv:2605.02087v1 Announce Type: new Abstract: Some frontier AI developers aim to align language models to a Model Spec or Constitution that describes the intended model behavior. However, standard a

safetyarxiv-cs-ai
6 May 2026
Model Releases

Decoding-Time Debiasing via Process Reward Models: From Controlled Fill-in to Open-Ended Generation

DGX agent

arXiv:2605.02348v1 Announce Type: new Abstract: Large language models pick up social biases from the data they are trained on and carry those biases into downstream applications, often reinforcing ste

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

AgentFloor: How Far Up the tool use Ladder Can Small Open-Weight Models Go?

DGX agent

arXiv:2605.00334v1 Announce Type: cross Abstract: Production agentic systems make many model calls per user request, and most of those calls are short, structured, and routine. This raises a practical

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Physical Foundation Models: Fixed hardware implementations of large-scale neural networks

DGX agent

arXiv:2604.27911v1 Announce Type: new Abstract: Foundation models are deep neural networks (such as GPT-5, Gemini~3, and Opus~4) trained on large datasets that can perform diverse downstream tasks --

model-releasesarxiv-cs-lg
1 May 2026
Research

Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM

DGX agent

arXiv:2604.25119v1 Announce Type: new Abstract: Auditing the fine-tunes of open-weight generative models for harmful specialization has become a new governance challenge for model hosting platforms. T

researcharxiv-cs-lg
29 Apr 2026
Model Releases

Differentiable Faithfulness Alignment for Cross-Model Circuit Transfer

DGX agent

arXiv:2604.24302v1 Announce Type: new Abstract: Mechanistic interpretability has made it possible to localize circuits underlying specific behaviors in language models, but existing methods are expens

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Jailbreaking Frontier Foundation Models Through Intention Deception

DGX agent

arXiv:2604.24082v1 Announce Type: cross Abstract: Large (vision-)language models exhibit remarkable capability but remain highly susceptible to jailbreaking. Existing safety training approaches aim to

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

DGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules

DGX agent

arXiv:2603.29928v2 Announce Type: replace Abstract: Tabular foundation models such as TabPFN and TabICL already produce full predictive distributions, yet prevailing regression benchmarks evaluate the

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Generative Modeling of Neurodegenerative Brain Anatomy with 4D Longitudinal Diffusion Model

DGX agent

arXiv:2604.22700v1 Announce Type: new Abstract: Understanding and predicting the progression of neurodegenerative diseases remains a major challenge in medical AI, with significant implications for ea

researcharxiv-cs-cv
27 Apr 2026
Model Releases

MIRROR: A Hierarchical Benchmark for Metacognitive Calibration in Large Language Models

DGX agent

arXiv:2604.19809v1 Announce Type: new Abstract: We introduce MIRROR, a benchmark comprising eight experiments across four metacognitive levels that evaluates whether large language models can use self

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

On Bayesian Softmax-Gated Mixture-of-Experts Models

DGX agent

arXiv:2604.20551v1 Announce Type: cross Abstract: Mixture-of-experts models provide a flexible framework for learning complex probabilistic input-output relationships by combining multiple expert mode

model-releasesarxiv-cs-lg
23 Apr 2026
← Previous
1…7891011…1012
Next →