AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

The Good, the Bad, and the Brittle: Benchmarking Robustness and Generalisation of Histopathology Foundation Models

DGX agent

arXiv:2607.04401v1 Announce Type: new Abstract: How robust and generalisable are pathology foundation models and have their scaling limites been reached? We benchmarked twelve pathology foundation mod

model-releasesarxiv-cs-cv
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

They Infer What You Meant: Models Represent Communicative Intent More Reliably Than They Act On It

DGX agent

arXiv:2607.03598v1 Announce Type: cross Abstract: When a person shares something with a language model, the model often answers the surface of the message rather than what the sender was doing by send

researcharxiv-cs-ai
7 Jul 2026
Model Releases

TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior

DGX agent

arXiv:2512.20757v2 Announce Type: replace Abstract: Tokenizers provide the fundamental basis through which text is represented and processed by language models (LMs). Despite the importance of tokeniz

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

DGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

DGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters

DGX agent

arXiv:2607.04919v1 Announce Type: new Abstract: Deploying a time series foundation model requires GPU infrastructure, engineering overhead, and carries no guarantee of improvement over XGBoost. We pro

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection

DGX agent

arXiv:2512.10485v2 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiv

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Model Merging as Probabilistic Inference in Fine-Tuning Parameter Space

DGX agent

arXiv:2607.01689v1 Announce Type: cross Abstract: Model merging aims to combine existing single-task solutions into a multi-task solution without additional data-driven fine-tuning.~Most existing appr

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

mupscaling small models: Principled warm starts and hyperparameter transfer

DGX agent

arXiv:2602.10545v2 Announce Type: replace-cross Abstract: Modern large-scale neural networks are often trained and released in multiple sizes to accommodate diverse inference budgets. To improve effic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

DGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge

DGX agent

arXiv:2607.01829v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed for aviation business operations, from documentation and training generation to customer facing a

model-releasesarxiv-cs-ai
3 Jul 2026
Research

CausalMix: Data Mixture as Causal Inference for Language Model Training

DGX agent

arXiv:2607.01104v1 Announce Type: cross Abstract: In Large Language Model (LLM) training, data mixing plays a pivotal role in determining model performance. Recent methods optimize mixture weights via

researcharxiv-cs-ai
2 Jul 2026
Model Releases

CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization

DGX agent

arXiv:2511.05747v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning enhances the problem-solving ability of large language models (LLMs) but leads to substantial inference overhead, l

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Seahorse: A Unified Benchmarking Framework for Spatiotemporal Event Modeling

DGX agent

arXiv:2607.01022v1 Announce Type: new Abstract: Spatiotemporal point processes (STPPs) model event data in continuous time and space, with applications in mobility, epidemiology, and public safety. Re

model-releasesarxiv-cs-lg
2 Jul 2026
Tutorials

Self-conditioned Flow Map Language Models via Fixed-point Flows

DGX agent

arXiv:2607.00714v1 Announce Type: cross Abstract: Self-conditioning is a core technique that enhances continuous flow-based language models, where the model learns to denoise generated text by conditi

tutorialsarxiv-cs-ai
2 Jul 2026
Model Releases

An Empirical Study of Security Calibration in Large Language Models for Code

DGX agent

arXiv:2606.31159v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly transforming software development, yet their use in security-critical contexts raises a key question: do mode

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Calibration, Not Compilation: Detecting and Repairing Misspecified Probabilistic Programs Written by Language Models

DGX agent

arXiv:2606.31630v1 Announce Type: new Abstract: Language models increasingly write probabilistic programs (in NumPyro, Stan, or Pyro), but a program that compiles, runs, and passes every unit test can

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Embodied CAD: Solver-Grounded LLM Agents for Parametric B-Rep Assembly Modeling

DGX agent

arXiv:2606.31252v1 Announce Type: new Abstract: Large language models can write plausible CAD scripts, but reliable industrial CAD modeling requires more than syntactically valid code: every feature,

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

LLM-as-a-judge validity in physics assessment depends more on the task than the model

DGX agent

arXiv:2603.14732v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly considered for automated assessment and feedback, understanding when LLM marking is valid is

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Physics-Constrained Fine-Tuning of Flow-Matching Models for Generation and Inverse Problems

DGX agent

arXiv:2508.09156v3 Announce Type: replace-cross Abstract: We present a framework for fine-tuning flow-matching generative models to enforce physical constraints and solve inverse problems in scientifi

model-releasesarxiv-cs-ai
1 Jul 2026
Research

TotalFM: An Organ-Separated 3D-CT Foundation Model Leveraging Large-Scale Routine Clinical Radiology Data

DGX agent

arXiv:2601.00260v2 Announce Type: replace Abstract: While foundation models in radiology are expected to be applied to various clinical tasks, computational cost constraints remain a major challenge w

researcharxiv-cs-cv
1 Jul 2026
Model Releases

TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios

DGX agent

arXiv:2603.29759v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have accelerated their application to indoor safety hazards assessment. However, existing ben

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

DGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Benchmarking Geospatial Foundation Models for Agriculture Applications

DGX agent

arXiv:2606.29664v1 Announce Type: new Abstract: Geospatial foundation models pretrained on satellite imagery promise broad generalization across remote sensing tasks and regions, but their geographic

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Compressed Sensing for Capability Localization in Large Language Models

DGX agent

arXiv:2603.03335v2 Announce Type: replace Abstract: Large language models (LLMs) exhibit a wide range of capabilities, including mathematical reasoning, code generation, and linguistic behaviors. We s

model-releasesarxiv-cs-cl
30 Jun 2026
Agents

Evaluating Newtonian Mechanics in Video Generative Models with Real Physical Systems

DGX agent

arXiv:2504.02918v3 Announce Type: replace Abstract: Recent advances in image and video generation raise hopes that these models possess world modeling capabilities-the ability to generate realistic, p

agentsarxiv-cs-cv
30 Jun 2026
Research

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement

DGX agent

arXiv:2606.29150v1 Announce Type: new Abstract: Discrete flow models have recently shown promising performance on few-step text generation; however, when naively applied to structured reasoning tasks

researcharxiv-cs-ai
30 Jun 2026
Research

Kriging and neural network models for pressure losses across perforated plates

DGX agent

arXiv:2606.29628v1 Announce Type: cross Abstract: In this paper, two novel data-driven models based on kriging and neural networks (NN) are proposed to predict pressure losses across perforated plates

researcharxiv-cs-lg
30 Jun 2026
Safety

Learning Transferable Dynamics Priors from Action to World Modeling

DGX agent

arXiv:2606.29501v1 Announce Type: new Abstract: We study action-conditioned world modeling as a scalable way to learn transferable dynamics priors for robot learning. By pretraining a model to predict

safetyarxiv-cs-ro
30 Jun 2026
Model Releases

Nonlinear mixture model motivated subspace clustering

DGX agent

arXiv:2606.29261v1 Announce Type: cross Abstract: We derive the linear union-of-subspaces (UoS) model for subspace clustering (SC) from the nonlinear mixture model (NMM) used in blind source separatio

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Obliviate: Erasing Concepts from Autoregressive Image Generation Models

DGX agent

arXiv:2606.28643v1 Announce Type: new Abstract: The widespread adoption of generative AI models has intensified concerns about misuse, including the creation of unsafe or disturbing imagery. To mitiga

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

Pairwise Comparisons without Stochastic Transitivity: Model, Theory and Applications

DGX agent

arXiv:2501.07437v3 Announce Type: replace-cross Abstract: Most statistical models for pairwise comparisons, including the Bradley-Terry (BT) and Thurstone models and many extensions, make a relatively

applicationsarxiv-cs-lg
30 Jun 2026
Hardware

SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces

DGX agent

arXiv:2403.07711v5 Announce Type: replace-cross Abstract: Given the remarkable achievements in image generation through diffusion models, the research community has shown increasing interest in extend

hardwarearxiv-cs-ai
30 Jun 2026
Research

Structural Certification for Reliable Physical Design with Language Models

DGX agent

arXiv:2606.30107v1 Announce Type: new Abstract: An unreliable language model can be made to produce reliable physical designs if the authority to assert is moved out of the model: the model proposes,

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

DGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Diffusion Model Attribution via Spectral Coupling of Denoiser Responses

DGX agent

arXiv:2606.28092v1 Announce Type: new Abstract: Attributing a generated image to its source diffusion model is a fundamental challenge in provenance verification and intellectual property protection.

researcharxiv-cs-cv
29 Jun 2026
Research

From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond

DGX agent

arXiv:2606.28127v1 Announce Type: cross Abstract: The AI community has framed the relationship between large language models (LLMs) and world models as a dichotomy: LLMs predict tokens; world models s

researcharxiv-cs-ai
29 Jun 2026
Research

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models

DGX agent

arXiv:2606.27627v1 Announce Type: cross Abstract: Discrete audio representations have become increasingly popular for building multimodal text-audio systems and integrating audio capabilities into Lar

researcharxiv-cs-ai
29 Jun 2026
Model Releases

Towards Benign Memory Forgetting for Selective Multimodal Large Language Model Unlearning

DGX agent

arXiv:2511.20196v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can inadvertently memorize privacy-sensitive information during training. While existing unlearning methods

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

DGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Do Image Editing Models Understand Lighting?

DGX agent

arXiv:2606.26738v1 Announce Type: new Abstract: While recent advancements in generative image editing models have achieved stunning visual fidelity, it remains an open question whether these systems p

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Unison: Benchmarking Unified Multimodal Models via Synergistic Understanding and Generation

DGX agent

arXiv:2606.26984v1 Announce Type: new Abstract: Unified multimodal models capable of both understanding and generation have achieved remarkable strides. However, despite their unified designs, existin

model-releasesarxiv-cs-cv
26 Jun 2026
Research

Surrogate models for Rock-Fluid Interaction: A Grid-Size-Invariant Approach

DGX agent

arXiv:2602.22188v2 Announce Type: replace Abstract: Modelling rock-fluid interaction requires solving a set of partial differential equations (PDEs) to predict the flow behaviour and the reactions of

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Models and GUI Grounding Datasets

DGX agent

arXiv:2606.25760v1 Announce Type: cross Abstract: Computer-use agents turn vision-language model (VLM) predictions into executable GUI clicks, so reliable uncertainty estimates are essential for rejec

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

BenchX: Benchmarking AI Models for Cancer Detection and Localization with Demographic and Protocol Biases

DGX agent

arXiv:2606.24883v1 Announce Type: new Abstract: Artificial intelligence (AI) has achieved remarkable success in medical imaging, but it is widely recognized that these models often perform inconsisten

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

DriveStack-VLA: Render-Teacher Alignment for BEV-Based DeepStack Vision-Language-Action Model

DGX agent

arXiv:2606.24051v1 Announce Type: new Abstract: Vision-Language-Action driving models convert a pretrained Vision-Language Model into a driving policy, allowing them to use world knowledge and follow

safetyarxiv-cs-cv
24 Jun 2026
Local Ai

Federated Survival Analysis in Healthcare: A Multi-Model Evaluation on Cross-Institutional Heterogeneous Breast Cancer Data

DGX agent

arXiv:2606.23871v1 Announce Type: new Abstract: Survival analysis is central to clinical decision-making, yet reliable time-to-event models require large, diverse cohorts that are rarely available at

local-aiarxiv-cs-lg
24 Jun 2026
Model Releases

Lightweight Transformer Models for On-Device Fault Detection: A Benchmark Study on Resource-Constrained Deployment

DGX agent

arXiv:2606.24173v1 Announce Type: cross Abstract: On-device fault detection enables real-time diagnostics without cloud dependency, but deploying machine learning models on resource-constrained hardwa

model-releasesarxiv-cs-ai
24 Jun 2026
← Previous
1…3334353637…1021
Next →