AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,860 results
15 Jul 2026

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

Model ReleasesDGX agent

arXiv:2607.12756v1 Announce Type: new Abstract: Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated

14 Jul 2026

Nemotron Labs: How Open Models Give Enterprises and Nations AI They Can Trust, Control and Customize

Model ReleasesDGX agent

Enterprises have plenty of powerful models to choose from. The real test is whether the AI an enterprise builds uniquely addresses the needs of the business: improving workflows, tapping into domain k

10 Jul 2026

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Joint Bayesian Parameter and Model Order Estimation for Low-Rank Probability Mass Tensors
Model ReleasesDGX agent

arXiv:2410.06329v4 Announce Type: replace-cross Abstract: Obtaining a reliable estimate of the joint probability mass function (PMF) of a set of random variables from observed data is a significant ob

9 Jul 2026

Checkout the latest version of Zed, adding proper support for llama.cpp as a model provider

Model ReleasesDGX agent

Zed editor has released a new version with integrated support for llama.cpp as a model provider, enabling users to run local language models within the editor. This update allows developers to use lla

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07288v1 Announce Type: new Abstract: Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized st

Meta says its new AI model is ready to compete on coding

Model ReleasesDGX agent

After reentering the AI race with its first in-house Muse Spark model in April, Meta is now opening up the doors to developers with a new model that can plug into AI coding software with the new Meta

Multi-Agent Robotic Control with Onboard Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07403v1 Announce Type: cross Abstract: Vision Language Models (VLMs) and Vision Language Action (VLA) models have shown promise in robotic control. Yet, they face significant challenges reg

Thinking Ahead: Foresight Intelligence in MLLMs and World Model

Model ReleasesDGX agent

arXiv:2511.18735v3 Announce Type: replace-cross Abstract: In this work, we define Foresight Intelligence as the capability to anticipate and interpret future events-an ability essential for applicatio

Yann LeCun claimed on social media [7] that my foundational 1990 paper on Neural World Models [1] 'was never accepted through peer review.' …

AgentsDGX agent

Yann LeCun claimed on social media [7] that my foundational 1990 paper on Neural World Models [1] 'was never accepted through peer review.' This is simply false. The core concepts from the tech report

8 Jul 2026

FedDAF: Federated Domain Adaptation Using Model Functional Distance

ApplicationsDGX agent

arXiv:2509.11819v2 Announce Type: replace-cross Abstract: Federated Domain Adaptation (FDA) is a federated learning (FL) approach that improves model performance at the target client by collaborating

7 Jul 2026

Another day another big public tech co saying they’re using open source AI models extensively

Model ReleasesDGX agent

Another day another big public tech co saying they’re using open source AI models extensively With our internal coding benchmark, we're able to confidently introduce open-weight models into our AI cod

Distribution-free Deviation Bounds and The Role of Domain Knowledge in Learning via Model Selection with Cross-validation Risk Estimation

SafetyDGX agent

arXiv:2303.08777v3 Announce Type: replace-cross Abstract: Cross-validation is one of the most widely used tools for risk estimation and model selection in statistics and machine learning, yet its theo

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

Model ReleasesDGX agent

arXiv:2607.03691v1 Announce Type: cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middlewa

DynaWM: A Base-VLA-Guided World Foundation Model for Moving-Object Manipulation

Model ReleasesDGX agent

arXiv:2607.02604v1 Announce Type: new Abstract: Although vision-language-action (VLA) models have received widespread attention, many challenges remain in manipulating dynamic moving objects. In most

ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models

SafetyDGX agent

arXiv:2506.09740v2 Announce Type: replace-cross Abstract: Diffusion models excel at image generation. Recent studies have shown that these models not only generate high-quality images but also encode

evalci: A Python Library for Statistically Rigorous Comparison of Language Model Evaluations

AgentsDGX agent

arXiv:2607.04429v1 Announce Type: cross Abstract: The dominant practice in language model evaluation is to report a single accuracy number per model and declare the higher one better, without testing

Prior Bias in Vision Language Models on UML Diagram Interpretation

Model ReleasesDGX agent

arXiv:2607.02853v1 Announce Type: new Abstract: Vision Language Models (VLMs) are increasingly applied to software engineering artifacts, especially UML class diagrams whose meaning depends on visual

Punching Above Their Weight: Classification-Head Fine-Tuning of Tiny Language Models (TLMs) for Verifiable Multiple-Choice Tasks

Model ReleasesDGX agent

arXiv:2607.03801v1 Announce Type: cross Abstract: We define Tiny Language Models (TLMs) as models below roughly 3B parameters that fit on mainstream consumer devices. We study how to adapt them for an

Retroactive Chain-of-Thought (RetroCoT): Forensic Reconstruction Prompts as a Safety Diagnostic Across Model Generations

Model ReleasesDGX agent

arXiv:2607.04645v1 Announce Type: cross Abstract: Safety alignment in large language models is typically evaluated against direct, imperative harmful requests. We show that this alignment is highly co

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models

Model ReleasesDGX agent

arXiv:2607.05365v1 Announce Type: cross Abstract: Streaming speech-to-speech language models aim to answer spoken queries directly with synthetic speech. However, standard speech and text benchmarks d

SteelBench: Evaluating Vision-Language Models in Real-World Industrial Environments

Model ReleasesDGX agent

arXiv:2607.05264v1 Announce Type: new Abstract: Existing video benchmarks evaluate action recognition on consumer videos, egocentric recordings, or simulated industrial environments. They do not test

Streaming Model Cascades for Semantic SQL

Model ReleasesDGX agent

arXiv:2604.00660v2 Announce Type: replace-cross Abstract: Modern data warehouses extend SQL with semantic operators that invoke large language models on each qualifying row, making per-row inference o

The Good, the Bad, and the Brittle: Benchmarking Robustness and Generalisation of Histopathology Foundation Models

Model ReleasesDGX agent

arXiv:2607.04401v1 Announce Type: new Abstract: How robust and generalisable are pathology foundation models and have their scaling limites been reached? We benchmarked twelve pathology foundation mod

They Infer What You Meant: Models Represent Communicative Intent More Reliably Than They Act On It

ResearchDGX agent

arXiv:2607.03598v1 Announce Type: cross Abstract: When a person shares something with a language model, the model often answers the surface of the message rather than what the sender was doing by send

TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior

Model ReleasesDGX agent

arXiv:2512.20757v2 Announce Type: replace Abstract: Tokenizers provide the fundamental basis through which text is represented and processed by language models (LMs). Despite the importance of tokeniz

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

SafetyDGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

Model ReleasesDGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters

Model ReleasesDGX agent

arXiv:2607.04919v1 Announce Type: new Abstract: Deploying a time series foundation model requires GPU infrastructure, engineering overhead, and carries no guarantee of improvement over XGBoost. We pro

6 Jul 2026

The Nemotron family just passed 100M downloads! Huge thank you to the community building with us and showing what’s possible with open model…

Model ReleasesDGX agent

The Nemotron model family from NVIDIA has reached 100 million downloads, marking a significant milestone for the open-source AI model community. The achievement reflects growing adoption and collabora

With our internal coding benchmark, we're able to confidently introduce open-weight models into our AI code reviewer w/o degrading code qual…

Model ReleasesDGX agent

With our internal coding benchmark, we're able to confidently introduce open-weight models into our AI code reviewer w/o degrading code quality. Have the frontier model (Fable) to the hardest work, de

3 Jul 2026

From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection

Model ReleasesDGX agent

arXiv:2512.10485v2 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiv

Model Merging as Probabilistic Inference in Fine-Tuning Parameter Space

Model ReleasesDGX agent

arXiv:2607.01689v1 Announce Type: cross Abstract: Model merging aims to combine existing single-task solutions into a multi-task solution without additional data-driven fine-tuning.~Most existing appr

mupscaling small models: Principled warm starts and hyperparameter transfer

Model ReleasesDGX agent

arXiv:2602.10545v2 Announce Type: replace-cross Abstract: Modern large-scale neural networks are often trained and released in multiple sizes to accommodate diverse inference budgets. To improve effic

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

Model ReleasesDGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge

Model ReleasesDGX agent

arXiv:2607.01829v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed for aviation business operations, from documentation and training generation to customer facing a

2 Jul 2026

CausalMix: Data Mixture as Causal Inference for Language Model Training

ResearchDGX agent

arXiv:2607.01104v1 Announce Type: cross Abstract: In Large Language Model (LLM) training, data mixing plays a pivotal role in determining model performance. Recent methods optimize mixture weights via

CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization

Model ReleasesDGX agent

arXiv:2511.05747v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning enhances the problem-solving ability of large language models (LLMs) but leads to substantial inference overhead, l

Dario just declared war on open-source. Anthropic's message is clear: open source could destroy the entire AI business model, and Chinese op…

ApplicationsDGX agent

Dario just declared war on open-source. Anthropic's message is clear: open source could destroy the entire AI business model, and Chinese open-source models are the cause. I sat down with @jasonlk & @

Seahorse: A Unified Benchmarking Framework for Spatiotemporal Event Modeling

Model ReleasesDGX agent

arXiv:2607.01022v1 Announce Type: new Abstract: Spatiotemporal point processes (STPPs) model event data in continuous time and space, with applications in mobility, epidemiology, and public safety. Re

Self-conditioned Flow Map Language Models via Fixed-point Flows

TutorialsDGX agent

arXiv:2607.00714v1 Announce Type: cross Abstract: Self-conditioning is a core technique that enhances continuous flow-based language models, where the model learns to denoise generated text by conditi

1 Jul 2026

An Empirical Study of Security Calibration in Large Language Models for Code

Model ReleasesDGX agent

arXiv:2606.31159v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly transforming software development, yet their use in security-critical contexts raises a key question: do mode

Calibration, Not Compilation: Detecting and Repairing Misspecified Probabilistic Programs Written by Language Models

Model ReleasesDGX agent

arXiv:2606.31630v1 Announce Type: new Abstract: Language models increasingly write probabilistic programs (in NumPyro, Stan, or Pyro), but a program that compiles, runs, and passes every unit test can

Claude Fable 5 is available again in Cursor. It leads all models on CursorBench, but is the most expensive per task.

Model ReleasesDGX agent

Claude Fable 5 has been re-enabled as an available model option within Cursor, where it achieves the highest performance scores on CursorBench among all supported models. However, it comes with the hi

Embodied CAD: Solver-Grounded LLM Agents for Parametric B-Rep Assembly Modeling

Model ReleasesDGX agent

arXiv:2606.31252v1 Announce Type: new Abstract: Large language models can write plausible CAD scripts, but reliable industrial CAD modeling requires more than syntactically valid code: every feature,

LLM-as-a-judge validity in physics assessment depends more on the task than the model

Model ReleasesDGX agent

arXiv:2603.14732v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly considered for automated assessment and feedback, understanding when LLM marking is valid is

Physics-Constrained Fine-Tuning of Flow-Matching Models for Generation and Inverse Problems

Model ReleasesDGX agent

arXiv:2508.09156v3 Announce Type: replace-cross Abstract: We present a framework for fine-tuning flow-matching generative models to enforce physical constraints and solve inverse problems in scientifi

TotalFM: An Organ-Separated 3D-CT Foundation Model Leveraging Large-Scale Routine Clinical Radiology Data

ResearchDGX agent

arXiv:2601.00260v2 Announce Type: replace Abstract: While foundation models in radiology are expected to be applied to various clinical tasks, computational cost constraints remain a major challenge w

TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios

Model ReleasesDGX agent

arXiv:2603.29759v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have accelerated their application to indoor safety hazards assessment. However, existing ben

30 Jun 2026

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

Model ReleasesDGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

Benchmarking Geospatial Foundation Models for Agriculture Applications

Model ReleasesDGX agent

arXiv:2606.29664v1 Announce Type: new Abstract: Geospatial foundation models pretrained on satellite imagery promise broad generalization across remote sensing tasks and regions, but their geographic

Compressed Sensing for Capability Localization in Large Language Models

Model ReleasesDGX agent

arXiv:2603.03335v2 Announce Type: replace Abstract: Large language models (LLMs) exhibit a wide range of capabilities, including mathematical reasoning, code generation, and linguistic behaviors. We s

Evaluating Newtonian Mechanics in Video Generative Models with Real Physical Systems

AgentsDGX agent

arXiv:2504.02918v3 Announce Type: replace Abstract: Recent advances in image and video generation raise hopes that these models possess world modeling capabilities-the ability to generate realistic, p

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement

ResearchDGX agent

arXiv:2606.29150v1 Announce Type: new Abstract: Discrete flow models have recently shown promising performance on few-step text generation; however, when naively applied to structured reasoning tasks

Kriging and neural network models for pressure losses across perforated plates

ResearchDGX agent

arXiv:2606.29628v1 Announce Type: cross Abstract: In this paper, two novel data-driven models based on kriging and neural networks (NN) are proposed to predict pressure losses across perforated plates

Learning Transferable Dynamics Priors from Action to World Modeling

SafetyDGX agent

arXiv:2606.29501v1 Announce Type: new Abstract: We study action-conditioned world modeling as a scalable way to learn transferable dynamics priors for robot learning. By pretraining a model to predict

Nonlinear mixture model motivated subspace clustering

Model ReleasesDGX agent

arXiv:2606.29261v1 Announce Type: cross Abstract: We derive the linear union-of-subspaces (UoS) model for subspace clustering (SC) from the nonlinear mixture model (NMM) used in blind source separatio

Obliviate: Erasing Concepts from Autoregressive Image Generation Models

Model ReleasesDGX agent

arXiv:2606.28643v1 Announce Type: new Abstract: The widespread adoption of generative AI models has intensified concerns about misuse, including the creation of unsafe or disturbing imagery. To mitiga

Pairwise Comparisons without Stochastic Transitivity: Model, Theory and Applications

ApplicationsDGX agent

arXiv:2501.07437v3 Announce Type: replace-cross Abstract: Most statistical models for pairwise comparisons, including the Bradley-Terry (BT) and Thurstone models and many extensions, make a relatively

SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces

HardwareDGX agent

arXiv:2403.07711v5 Announce Type: replace-cross Abstract: Given the remarkable achievements in image generation through diffusion models, the research community has shown increasing interest in extend

Structural Certification for Reliable Physical Design with Language Models

ResearchDGX agent

arXiv:2606.30107v1 Announce Type: new Abstract: An unreliable language model can be made to produce reliable physical designs if the authority to assert is moved out of the model: the model proposes,

← Previous
1…3637383940…998
Next →