AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
Model Releases

Bridging Foundation Models and ASTM Metallurgical Standards for Automated Grain Size Estimation from Microscopy Images

DGX agent

arXiv:2604.18957v1 Announce Type: new Abstract: Extracting standardized metallurgical metrics from microscopy images remains challenging due to complex grain morphology and the data demands of supervi

model-releasesarxiv-cs-cv
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Industry

Here's how anyone can find models that work for your hardware easily. 1. Go to http://huggingface.co and make an account 2. Models tab to fi…

DGX agent

Here's how anyone can find models that work for your hardware easily. 1. Go to http://huggingface.co and make an account 2. Models tab to find weights and all compressions 3. Click on your profile on

industryclem-delangue--x
21 Apr 2026
Model Releases

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

DGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning

DGX agent

arXiv:2503.03480v4 Announce Type: replace Abstract: Vision-language-action models (VLAs) show potential as generalist robot policies. However, these models pose extreme safety challenges during real-w

model-releasesarxiv-cs-ro
21 Apr 2026
Applications

Thermal-GEMs: Generalized Models for Building Thermal Dynamics

DGX agent

arXiv:2604.16443v1 Announce Type: cross Abstract: Data-driven models for building thermal dynamics are a scalable approach for enabling energy-efficient operation through fault detection & diagnosis o

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

Using large language models for embodied planning introduces systematic safety risks

DGX agent

arXiv:2604.18463v1 Announce Type: cross Abstract: Large language models are increasingly used as planners for robotic systems, yet how safely they plan remains an open question. To evaluate safe plann

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Think Multilingual, Not Harder: A Data-Efficient Framework for Teaching Reasoning Models to Code-Switch

DGX agent

arXiv:2604.15490v1 Announce Type: new Abstract: Recent developments in reasoning capabilities have enabled large language models to solve increasingly complex mathematical, symbolic, and logical tasks

researcharxiv-cs-cl
20 Apr 2026
Applications

> grok4.20-beta1 is a much smaller model than opus but is #1 ranked in medicine and healthcare > 4.3 and 4.4 will be much larger models, and…

DGX agent

> grok4.20-beta1 is a much smaller model than opus but is #1 ranked in medicine and healthcare > 4.3 and 4.4 will be much larger models, and likely will have a significant boost in performance on comp

applicationselon-musk--x
18 Apr 2026
Model Releases

Parameter estimation for land-surface models using Neural Physics

DGX agent

arXiv:2505.02979v3 Announce Type: replace-cross Abstract: We propose a novel inverse-modelling approach which estimates the parameters of a simple land-surface model (LSM) by assimilating data into a

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agentic coding on par wi…

DGX agent

⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agentic coding on par with models 10x its active size 📷 Strong multimodal perception an

model-releasesjeremy-howard--x
16 Apr 2026
Model Releases

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?

DGX agent

arXiv:2511.17792v2 Announce Type: replace Abstract: While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and u

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Benchmarking Deflection and Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.12033v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) increasingly rely on retrieval to answer knowledge-intensive multimodal questions. Existing benchmarks overlook c

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

AIM: Intent-Aware Unified world action Modeling with Spatial Value Maps

DGX agent

arXiv:2604.11135v1 Announce Type: cross Abstract: Pretrained video generation models provide strong priors for robot control, but existing unified world action models still struggle to decode reliable

model-releasesarxiv-cs-lg
14 Apr 2026
Agents

Automating Structural Analysis Across Multiple Software Platforms Using Large Language Models

DGX agent

arXiv:2604.09866v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have shown the promise to significantly accelerate the workflow by automating structural modeling and

agentsarxiv-cs-ai
14 Apr 2026
Research

Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning

DGX agent

arXiv:2604.10701v1 Announce Type: cross Abstract: Credit assignment is a central challenge in reinforcement learning (RL). Classical actor-critic methods address this challenge through fine-grained ad

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Given the messy naming scheme used by all the AI companies, I caused a chart to be made showing the gain in GPQA per 0.1 version in model na…

DGX agent

Given the messy naming scheme used by all the AI companies, I caused a chart to be made showing the gain in GPQA per 0.1 version in model names (estimated, since model names skip version numbers). The

model-releasesethan-mollick--x
14 Apr 2026
Research

Lost in Diffusion: Uncovering Hallucination Patterns and Failure Modes in Diffusion Large Language Models

DGX agent

arXiv:2604.10556v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) have emerged as a promising non-autoregressive paradigm comparable to autoregressive (AR) models, their fa

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration

DGX agent

arXiv:2604.11446v1 Announce Type: cross Abstract: Recently, scaling reinforcement learning with verifiable rewards (RLVR) for large language models (LLMs) has emerged as an effective training paradigm

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Pando: Do Interpretability Methods Work When Models Won't Explain Themselves?

DGX agent

arXiv:2604.11061v1 Announce Type: cross Abstract: Mechanistic interpretability is often motivated for alignment auditing, where a model's verbal explanations can be absent, incomplete, or misleading.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass

DGX agent

arXiv:2604.10966v1 Announce Type: cross Abstract: We present a discriminative multimodal reward model that scores all candidate responses in a single forward pass. Conventional discriminative reward m

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

DGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Which model would be the best to generate fictional country flags? SDXL/Qwen/Wan/ZIT/ZIB/Flux Klein/Flux Dev?

DGX agent

This r/StableDiffusion post discusses community recommendations for the best AI image generation model to create fictional country flags, comparing options including SDXL, Qwen, Wan, ZIT, ZIB, Flux Kl

model-releasesr-stablediffusion
12 Apr 2026
Model Releases

AE-ViT: Stable Long-Horizon Parametric Partial Differential Equations Modeling

DGX agent

arXiv:2604.06475v1 Announce Type: new Abstract: Deep Learning Reduced Order Models (ROMs) are becoming increasingly popular as surrogate models for parametric partial differential equations (PDEs) due

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Before We Trust Them: Decision-Making Failures in Navigation of Foundation Models

DGX agent

arXiv:2601.05529v5 Announce Type: replace Abstract: High success rates on navigation-related tasks do not necessarily translate into reliable decision making by foundation models. To examine this gap,

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

People ask me how I choose what model to route queries to. It's simple. Claude gets knowledge work anything below that would be degrading Cl…

DGX agent

I was unable to retrieve the specific tweet content from that URL, as the post requires a logged-in X (Twitter) session to access, and search results did not surface the full text of that specific ...

model-releasesdylan-patel--x
10 Apr 2026
Model Releases

Temporally Phenotyping GLP-1RA Case Reports with Large Language Models: A Textual Time Series Corpus and Risk Modeling

DGX agent

arXiv:2604.06197v1 Announce Type: cross Abstract: Type 2 diabetes case reports describe complex clinical courses, but their timelines are often expressed in language that is difficult to reuse in long

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment

DGX agent

arXiv:2604.06377v1 Announce Type: cross Abstract: We investigate whether post-trained capabilities can be transferred across models without retraining, with a focus on transfer across different model

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Which Way Does Time Flow? A Psychophysics-Grounded Evaluation for Vision-Language Models

DGX agent

arXiv:2510.26241v5 Announce Type: replace-cross Abstract: Modern vision-language models (VLMs) excel at many multimodal tasks, yet their grasp of temporal information in video remains weak and has not

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

if you want claude's managed agents w/o model provider lock in, this is the path! `deepagents deploy` lets you deploy an agent built on our …

DGX agent

if you want claude's managed agents w/o model provider lock in, this is the path! `deepagents deploy` lets you deploy an agent built on our model agnostic, open source harness in minutes you can easil

model-releasesharrison-chase--x
9 Apr 2026
Model Releases

Lots of love for Gemma 4! Team just told me it’s already had 10M+ downloads since last week’s launch. Gemma models have now been downloaded …

DGX agent

Google's Gemma 4, launched on April 2, 2026, surpassed 10 million downloads within its first week, according to Google CEO Sundar Pichai, pushing the cumulative total for all Gemma models past 500 ...

model-releasesollama--x
8 Apr 2026
Model Releases

Mapping and Measuring the Behavioral Evolution of Large Language Models

DGX agent

arXiv:2608.11027v1 Announce Type: cross Abstract: Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across genera

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

Reference-Free Post-Training of Open Large Language Models for Multilingual Machine Translation

DGX agent

arXiv:2608.10812v1 Announce Type: cross Abstract: We study reference-free post-training for multilingual machine translation with open large language models. Starting from the supervised-finetuned MiL

model-releasesarxiv-cs-ai
12 Aug 2026
Research

A Hybrid Neural-Microfacet BRDF Model for Real-Time Rendering

DGX agent

arXiv:2608.09604v1 Announce Type: cross Abstract: Over the past decade, microfacet-based BRDF models have formed the foundation of real-time rendering pipelines. Despite their widespread use, they oft

researcharxiv-cs-cv
11 Aug 2026
Model Releases

ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models

DGX agent

arXiv:2608.09548v1 Announce Type: cross Abstract: Large language models are increasingly deployed in education as tutors, teaching assistants, and content generators. These roles place demands that or

model-releasesarxiv-cs-ai
11 Aug 2026
Research

HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model selection

DGX agent

arXiv:2608.08069v1 Announce Type: cross Abstract: Foundation models are increasingly reused as software components, making model selection a critical software-engineering decision. Current model hubs

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Addressable Memory for Video World Models

DGX agent

arXiv:2608.07408v1 Announce Type: new Abstract: We study visual persistence in interactive video world models. These models rely on a Key-Value (KV) cache as a growing visual memory to carry forward p

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation

DGX agent

arXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Meta releases open-source Muse Glimmer model with 30B parameters

DGX agent

Meta Platforms Inc. today released Muse Glimmer, an open-source language model that can run on personal computers. The company also published a lengthy essay penned by Chief Executive Officer Mark Zuc

model-releasessiliconangle
10 Aug 2026
Agents

Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving

DGX agent

arXiv:2603.06054v2 Announce Type: replace-cross Abstract: The use of Vision-Language Models (VLMs) in automated driving applications is becoming increasingly common, with the aim of leveraging their r

agentsarxiv-cs-ai
10 Aug 2026
Safety

The Fairness Collapse Phenomenon: Bias Amplification in Language Models Trained on Synthetic Data

DGX agent

arXiv:2608.04268v1 Announce Type: new Abstract: Generative models trained on artificially generated data have been shown to exhibit model collapse, resulting in significant performance degradation. As

safetyarxiv-cs-cl
6 Aug 2026
Applications

Conformal risk control for model-form uncertainty in parametric non-intrusive reduced-order models

DGX agent

arXiv:2608.03360v1 Announce Type: cross Abstract: Non-intrusive reduced-order models (NIROMs) have become a standard tool for approximating parametric partial differential equations from computer desi

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

DGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

model-releasesr-localllama
5 Aug 2026
Model Releases

Xiaomi-Robotics-1: New robotics model released

DGX agent

Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipu

model-releasesr-localllama
5 Aug 2026
Model Releases

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

DGX agent

arXiv:2607.29122v1 Announce Type: new Abstract: Pixel-space diffusion models aim to learn an end-to-end generator directly over raw pixels. This is challenging because a single model must capture both

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Alibaba debuts Qwen3.8-Max model with 2.4T parameters

DGX agent

Alibaba Group Holding Ltd. today debuted a new addition to its Qwen series of open-source large language models. Qwen3.8-Max is the Chinese e-commerce giant’s most capable LLM to date. It features 2.4

model-releasessiliconangle
3 Aug 2026
Model Releases

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models

DGX agent

arXiv:2607.27386v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) offer a compelling alternative to autoregressive (AR) generation by enabling bidirectional context and iterative refi

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

MiniMax H3: Open-weight multimodel video model

DGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

model-releasesr-stablediffusion
31 Jul 2026
Research

Predict before you train: Scaling Laws for particle physics foundation models

DGX agent

arXiv:2607.23377v1 Announce Type: cross Abstract: The largest machine learning models in particle physics are also the most expensive to train, yet the return on scaling a given architecture cannot be

researcharxiv-cs-ai
31 Jul 2026
← Previous
1…1617181920…1238
Next →