AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,515 results
25 May 2026

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up o…

Model ReleasesDGX agent

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up on complex tasks, and it has actual personality. Few rough ed

MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination

AgentsDGX agent

arXiv:2605.22949v1 Announce Type: new Abstract: Foundation model agents increasingly operate in multi-agent deployments where a coordinator must decide which agent's response to trust. The standard ap

MirrorCheck: Efficient Adversarial Defense for Vision-Language Models

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2406.09250v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly susceptible to sophisticated adversarial attacks, including adaptive strategies specifically de

Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling

ResearchDGX agent

arXiv:2602.20210v2 Announce Type: replace-cross Abstract: Crystal modeling spans a family of conditional and unconditional generation tasks, including crystal structure prediction (CSP) and de novo ge

Transform-Invariant Generative Ray Path Sampling for Efficient Radio Propagation Modeling

SafetyDGX agent

arXiv:2603.01655v2 Announce Type: replace Abstract: Ray tracing has become a standard for accurate radio propagation modeling, but suffers from exponential computational complexity, as the number of c

23 May 2026

Conditional Neural Field based Reduced Order Model for Dynamic Ditching Load Prediction

ResearchDGX agent

arXiv:2605.21499v1 Announce Type: cross Abstract: Grid-based neural networks such as convolutional autoencoders are widely used in dimension reduction-based surrogate models for computational fluid dy

Interpreting and Steering State-Space Models via Activation Subspace Bottlenecks

ResearchDGX agent

arXiv:2602.22719v2 Announce Type: replace Abstract: State-space models (SSMs) have emerged as an efficient strategy for building powerful language models, avoiding the quadratic complexity of computin

New to OLLAMA, how to install best model for my Mac?

Local AiDGX agent

Ollama can be installed on Mac by downloading the application and placing it in the Applications folder, after which you use Terminal to run commands that download and launch models . The best model c

SDPM: Survival Diffusion Probabilistic Model for Continuous-Time Survival Analysis

ResearchDGX agent

arXiv:2605.22776v1 Announce Type: new Abstract: Survival analysis aims to estimate a time-to-event distribution from data with censored observations. Many existing methods either impose structural ass

Uniform Diffusion Models Revisited: Leave-One-Out Denoiser and Absorbing State Reformulation

ResearchDGX agent

arXiv:2605.22765v1 Announce Type: new Abstract: Discrete diffusion models are often trained through clean-data prediction, but the prediction can be used in different ways to define the reverse dynami

22 May 2026

Discovering Implicit Large Language Model Alignment Objectives

SafetyDGX agent

arXiv:2602.15338v2 Announce Type: replace-cross Abstract: Large language model (LLM) alignment relies on complex reward signals that often obscure the specific behaviors being incentivized, creating c

Do Factual Recall Mechanisms Carry over from Text to Speech in Multimodal Language Models?

ResearchDGX agent

arXiv:2605.22170v1 Announce Type: new Abstract: In recent years, several Speech Language Models (SLMs) that represent speech and written text jointly have been presented. The question then emerges abo

Evaluating Clinical Competencies of Large Language Models with a General Practice Benchmark

Model ReleasesDGX agent

arXiv:2503.17599v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated considerable potential in general practice. However, existing benchmarks and evaluation frameworks pr

Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly

Model ReleasesDGX agent

arXiv:2605.21625v1 Announce Type: cross Abstract: The emergence of Large Vision-Language Models (LVLMs) has significantly advanced video understanding capabilities. However, existing benchmarks focus

Gemini Flash 3.5 is such a disappointing model. It's intelligence and speed is awesome. Absolutely amazing. But it's been trained to max eva…

Model ReleasesDGX agent

Gemini Flash 3.5 is such a disappointing model. It's intelligence and speed is awesome. Absolutely amazing. But it's been trained to max evals, not to be helpful to humans. It goes off and does random

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic c…

Model ReleasesDGX agent

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic coding, reasoning, and long-horizon autonomy. AI natives can

Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models

Model ReleasesDGX agent

arXiv:2605.21861v1 Announce Type: new Abstract: Multi-modality medical vision (MV) foundation models (FM) are fundamentally challenged by pronounced Non-IID feature statistics across heterogeneous ima

Rethinking Noise-Robust Training for Frozen Vision Foundation Models: A Cross-Dataset Benchmark with a Case Study of Small-Loss Failure

Model ReleasesDGX agent

arXiv:2605.22591v1 Announce Type: new Abstract: Frozen Vision Foundation Models (VFMs) with lightweight classification heads are increasingly used in medical imaging because they offer efficient and r

Towards Selection of Large Multimodal Models as Engines for Burned-in Protected Health Information Detection in Medical Images

Model ReleasesDGX agent

arXiv:2511.02014v2 Announce Type: replace Abstract: The detection of Protected Health Information (PHI) in medical imaging is critical for safeguarding patient privacy and ensuring compliance with reg

Transcription and Recognition of Italian Parliamentary Speeches Using Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.28103v2 Announce Type: replace-cross Abstract: Parliamentary proceedings represent a rich yet challenging resource for computational analysis, particularly when preserved only as scanned hi

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

SafetyDGX agent

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else t…

Model ReleasesDGX agent

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else the same. Available inside our new Edit Studio, you can work

21 May 2026

A Mechanistic Study of Tabular Foundation Models

ResearchDGX agent

arXiv:2605.21288v1 Announce Type: new Abstract: Tabular foundation models with different architectures converge in accuracy across a range of classification and regression tasks. This raises questions

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

Model ReleasesDGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

Can liveness detection models generalise to synthetic media generation techniques they were never trained on? [D]

ResearchDGX agent

Liveness detection and synthetic media detection models often fail to generalize across unseen data and struggle with content from different models. Understanding how factors like data source diversit

DEL: Digit Entropy Loss for Numerical Learning of Large Language Models

Model ReleasesDGX agent

arXiv:2605.20369v1 Announce Type: new Abstract: Number prediction stands as a fundamental capability of large language models (LLMs) in mathematical problem-solving and code generation. The widely ado

Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models

ResearchDGX agent

arXiv:2411.01141v2 Announce Type: replace Abstract: There are two shortages in the current Large Language Models (LLMs) era. The first is short of multilingual models, where most LLMs are English-cent

Efficient training for compact compression models via sequential distillation

ResearchDGX agent

arXiv:2601.05639v2 Announce Type: replace Abstract: Deep learning models for image compression often face practical limitations in hardware-constrained applications. Although these models achieve high

How Open Must Language Models be to Enable Reliable Scientific Inference?

ResearchDGX agent

arXiv:2603.26539v2 Announce Type: replace Abstract: How does the extent to which a model is open or closed impact the scientific inferences that can be drawn from research that involves it? In this pa

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

Model ReleasesDGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

Leveraging Vision-Language Models to Detect Attention in Educational Videos

Model ReleasesDGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

MagenticLite, MagenticBrain, Fara1.5: An agentic experience optimized for small models

AgentsDGX agent

MagenticLite is an agentic system for small models that works across the browser and local file system in a single workflow. It combines specialized models and orchestration to support efficient agent

New inflection point in the accelerating growth of open-source models usage is coming

Model ReleasesDGX agent

New inflection point in the accelerating growth of open-source models usage is coming 🦔Microsoft canceled its internal Claude Code licenses this week after token-based billing made the cost untenable,

Self-Evolving in the Wild:Over the course of ~35 hours of continuous autonomous execution, the model performed 432 kernel evaluations across…

Model ReleasesDGX agent

Self-Evolving in the Wild:Over the course of ~35 hours of continuous autonomous execution, the model performed 432 kernel evaluations across 1,158 tool calls. It wrote, compiled, profiled, and iterati

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefi…

Model ReleasesDGX agent

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefit one at a time inside a 1.7B byte-level pretraining pipelin

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems

ApplicationsDGX agent

arXiv:2603.14392v2 Announce Type: replace Abstract: Trajectory world models play a crucial role in robotic dynamics learning, planning, and control. While recent works have explored trajectory world m

20 May 2026

An OpenAI model has disproved a central conjecture in discrete geometry

Model ReleasesDGX agent

An OpenAI AI model successfully disproved a longstanding conjecture in discrete geometry, a mathematical field studying geometric properties of discrete objects. This achievement demonstrates the pote

Bayesian Joint Model of Multi-Sensor and Failure Event Data for Multi-Mode Failure Prediction

ResearchDGX agent

arXiv:2506.17036v2 Announce Type: replace-cross Abstract: Modern industrial systems are often subject to multiple failure modes, and their conditions are monitored by multiple sensors, generating mult

Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay

SafetyDGX agent

arXiv:2605.19352v1 Announce Type: cross Abstract: Understanding how humans and artificial intelligence systems predict and plan by interacting with their environment is a fundamental challenge at the

Chessformer: A Unified Architecture for Chess Modeling

SafetyDGX agent

arXiv:2605.19091v1 Announce Type: new Abstract: Chess has long served as a canonical testbed for artificial intelligence, but modeling approaches for its central tasks have diverged. Maximizing playin

CPC-VAR:Continual Personalized and Compositional Generation in Visual Autoregressive Models

ResearchDGX agent

arXiv:2605.19750v1 Announce Type: new Abstract: Visual autoregressive (VAR) models have recently emerged as an efficient paradigm for text-to-image generation. Despite their strong generative capabili

D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models

SafetyDGX agent

arXiv:2605.19690v1 Announce Type: new Abstract: Navigation Foundation Models (NFMs) trained on large cross-embodied datasets have demonstrated powerful generalizability in various scenarios. Adopting

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

Model ReleasesDGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

High-quality generation of dynamic game content via small language models: A proof of concept

AgentsDGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.19301v1 Announce Type: new Abstract: Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastroph

Neural Operators for Design-Space Surrogate Modeling of Tendon-Actuated Continuum Robots

ResearchDGX agent

arXiv:2605.19104v1 Announce Type: cross Abstract: Continuum robots enable dexterous manipulation in constrained environments, but require accurate and efficient models for real-time manipulation and c

Probabilistic Tiny Recursive Model

ResearchDGX agent

arXiv:2605.19943v1 Announce Type: new Abstract: Tiny Recursive Models (TRM) solve complex reasoning tasks with a fraction of the parameters of modern large language models (LLMs) by iteratively refini

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

Model ReleasesDGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models

Local AiDGX agent

arXiv:2605.20158v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) show promise in medical applications, but their inability to faithfully ground responses in visual evidence raise

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

Model ReleasesDGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

Shaping the Prior: How Synthetic Task Distributions Determine Tabular Foundation Model Quality

ResearchDGX agent

arXiv:2605.18971v1 Announce Type: cross Abstract: What determines the quality of a tabular foundation model? Unlike language or vision, tabular foundation models acquire their inductive biases almost

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

Model ReleasesDGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

TADA! Tuning Audio Diffusion Models through Activation Steering

Model ReleasesDGX agent

arXiv:2602.11910v2 Announce Type: replace-cross Abstract: Audio diffusion models can synthesize high-fidelity music from text, yet achieving fine-grained control over specific musical attributes remai

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular,…

Model ReleasesDGX agent

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular, and represents an important milestone for the math and AI c

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry p…

IndustryDGX agent

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry problems. Less than 1 year ago frontier AI models were at IMO

Tweedie's Formulae and Diffusion Generative Models Beyond Gaussian

TutorialsDGX agent

arXiv:2605.19391v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success in generating samples from unknown data distributions. Most popular stochastic differential equation

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗

Model ReleasesDGX agent

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗 Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it t

19 May 2026

ARROW: Augmented Replay for RObust World models

SafetyDGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States

SafetyDGX agent

arXiv:2605.17144v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models leverage powerful perceptual priors from web-scale Vision-Language Model (VLM) pre-training, yet they remain surpr

CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark

Model ReleasesDGX agent

arXiv:2605.18621v1 Announce Type: cross Abstract: Spatial intelligence requires multimodal large language models (MLLMs) to move beyond single-view perception and reason consistently about objects, vi

← Previous
1…111112113114115…1009
Next →