AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
30 Jun 2026

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios

Model ReleasesDGX agent

arXiv:2511.17649v4 Announce Type: replace-cross Abstract: Tangible control interfaces (TCIs), such as appliance panels, remotes, elevators, and embedded GUIs, are a fundamental component of everyday h

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

29 Jun 2026

AI-Model Network: Concept, Current State and Future

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research
DGX agent

arXiv:2606.27382v1 Announce Type: new Abstract: While the primary function of computers lies in computation and processing, the core value of the Internet is rooted in sharing and collaboration. Compu

Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

Model ReleasesDGX agent

Anthropic’s Claude models in Microsoft Foundry — hosted on Microsoft Azure and running on NVIDIA GB300 Blackwell Ultra GPUs — are now generally available, giving Azure-native enterprises a powerful ne

Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?

ResearchDGX agent

arXiv:2606.27755v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable instruction-driven robotic manipulation, but they inherit oversized language backbones from pretrained VLMs

Grounded Iterative Language Planning: How Parameterized World Models Reduce Hallucination Propagation in LLM Agents

AgentsDGX agent

arXiv:2606.27806v1 Announce Type: new Abstract: World models for language agents come in two useful forms. An agent-based world model calls an LLM API and reasons flexibly in language, but its errors

Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-Based Retrieval Agents Improves Interpretable Claim Verification

Model ReleasesDGX agent

arXiv:2511.03217v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in generating fluent utterances but can lack reliable grounding in verified information. At the same time,

Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning

SafetyDGX agent

arXiv:2606.27483v1 Announce Type: new Abstract: Large language model (LLM) agents have demonstrated strong capability in sequential decision-making, yet they remains fundamentally reactive in long-hor

Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: …

Model ReleasesDGX agent

Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: praxis-spanfinder-3b and praxis-relevance-3b. They swap your

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemin…

Model ReleasesDGX agent

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemini / Opus 4.6 level 35tr tokens trained entirely on 50k Chine

Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron

Model ReleasesDGX agent

Showcasing the importance of open source innovation in American AI, Palantir’s new intelligent engine — introduced today — uses NVIDIA Nemotron open models to serve the needs of U.S. government agenci

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @Fireworks…

AgentsDGX agent

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @FireworksAI_HQ team there’s a mountain of Agent Improvement gold sitt

Sampling the Schwinger Model with Gauge-Equivariant Diffusion

ResearchDGX agent

arXiv:2606.27481v1 Announce Type: cross Abstract: We present a first study of a diffusion-based approach to accelerated sampling of the N_f = 2 lattice Schwinger model. Our work is inspired by recent

SIFT: Self-Imagination Fine-Tuning for Physically Plausible Motion in Video Diffusion Models

SafetyDGX agent

arXiv:2606.27741v1 Announce Type: new Abstract: Recent advances in video diffusion models have greatly improved visual fidelity, yet their generated motions often violate physical plausibility. We obs

The USG launching models on Hugging Face. Go @jgebbia

IndustryDGX agent

The U.S. Government is launching machine learning models on Hugging Face, a popular open-source platform for sharing AI models and datasets. This initiative, highlighted by Hugging Face co-founder Cle

28 Jun 2026

GLM-5.2 is good but it is not GPT-5.5/Opus 4.8, and even further from Mythos. Yet it is solid & it demonstrates that the open models continu…

Model ReleasesDGX agent

GLM-5.2 is good but it is not GPT-5.5/Opus 4.8, and even further from Mythos. Yet it is solid & it demonstrates that the open models continue to chase the frontier What is happening is that open weigh

27 Jun 2026

So this new licensing regime is probably the end of new model vague posting from the Labs. Good night, sweet prince, and flights of angels s…

Model ReleasesDGX agent

Ethan Mollick comments on how a new licensing regime will likely eliminate vague model announcements and promotional posting practices previously used by AI labs. The post uses literary language ('Goo

This is from a popular inference provider GLM-5.2 plus the US banning the most capable new models means open source caught up to SOTA closed…

SafetyDGX agent

This is from a popular inference provider GLM-5.2 plus the US banning the most capable new models means open source caught up to SOTA closed source coding models This could be v problematic for Anthro

26 Jun 2026

A Generalization Theory for JEPA-Based World Models

ResearchDGX agent

arXiv:2606.27014v1 Announce Type: new Abstract: Joint Embedding Predictive Architectures (JEPAs) have recently emerged as a promising paradigm for world modeling by learning predictive dynamics in a l

A Latent ODE Approach to Spatiotemporal Modeling of Cine Cardiac MRI

ResearchDGX agent

arXiv:2606.26718v1 Announce Type: new Abstract: Cardiac magnetic resonance imaging (CMR) captures rich spatiotemporal information about ventricular structure and motion, but conventional risk models u

EO-WM: A Physically Informed World Model for Probabilistic Earth Observation Forecasting

Model ReleasesDGX agent

arXiv:2606.27277v1 Announce Type: new Abstract: Earth Observation (EO) forecasting aims to predict future Earth surface dynamics from satellite observations under changing meteorological conditions. I

Exploring the Intrinsic Geometry of Diffusion Models with Constrained Inverse Kinematics

ResearchDGX agent

arXiv:2606.26408v1 Announce Type: new Abstract: Recent studies suggest that diffusion models can recover geometric structure in the data manifolds they are trained on, yet the supporting evidence has

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as …

Model ReleasesDGX agent

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as both the evolution harness backend and the LLM scorer. Along

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.27079v1 Announce Type: new Abstract: In embodied intelligence, safety is a prerequisite for reliable robot deployment in the physical world. Current vision-language-action (VLA) models cont

GenRecal: Generation after Recalibration from Large to Small Vision-Language Models

ApplicationsDGX agent

arXiv:2506.15681v4 Announce Type: replace Abstract: Recent advancements in vision-language models (VLMs) have leveraged large language models (LLMs) to achieve performance on par with closed-source sy

GPT-5.6 Sol is our most capable model yet for cybersecurity. It shifts the performance-efficiency frontier for long-horizon security tasks i…

Model ReleasesDGX agent

GPT-5.6 Sol represents OpenAI's latest advancement in AI capabilities, specifically optimized for cybersecurity applications. The model demonstrates improved performance-efficiency tradeoffs, particul

in other news, we updated the 5.5 instant model used in chatgpt this week. i like its vibes.

IndustryDGX agent

OpenAI updated the GPT-4o mini model (version 5.5) used in ChatGPT during this period, with Sam Altman expressing satisfaction with the model's performance and characteristics. The update likely inclu

KARLA: Knowledge-base Augmented Retrieval for Language Models

Model ReleasesDGX agent

arXiv:2606.26807v1 Announce Type: new Abstract: We propose a new method that allows an LLM to automatically pull in factual knowledge from a knowledge base during token generation. This means that (1)

MLFFM-SegDiff: A Multi-Level Feature Fusion Diffusion Model for Skin Lesion Segmentation

Model ReleasesDGX agent

arXiv:2606.26712v1 Announce Type: cross Abstract: Skin lesion segmentation is a key task in computer-aided dermatological diagnosis, where accuracy directly impacts downstream analysis and disease cla

On-board Remote-Sensing Foundation Models for Unsupervised Change Detection of Disaster Events

ResearchDGX agent

arXiv:2606.27018v1 Announce Type: cross Abstract: Remote Sensing Foundation Models (RSFMs) have emerged as a powerful alternative to supervised models for Earth Observation, allowing satellites to aut

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models

SafetyDGX agent

arXiv:2606.26694v1 Announce Type: new Abstract: Recent game world models can synthesize visually plausible, action-conditioned rollouts. However, their interaction behaviors often remain limited to ex

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't st…

IndustryDGX agent

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't stop A) global open model progress B) bad actors using them So

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

Model ReleasesDGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a …

TutorialsDGX agent

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a sentence is about' Hybrid (transformer–RNN) models are fast

Tuning Language Models by Mixture-of-Depths Ensemble

Model ReleasesDGX agent

arXiv:2410.13077v2 Announce Type: replace-cross Abstract: Transformer-based Large Language Models (LLMs) traditionally rely on final-layer loss for finetuning and final-layer representations for predi

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. …

Model ReleasesDGX agent

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. Faulty reasoning, use of inappropriate shortcuts, hallucinat

25 Jun 2026

Agentic evolution of physically constrained foundation models

Model ReleasesDGX agent

arXiv:2606.25532v1 Announce Type: cross Abstract: Artificial intelligence increasingly drives automated scientific discovery, yet contemporary generalist agents lack physical grounding, frequently hal

Conformal Orbit-Valid Trust Horizons for Equivariant World Models

ResearchDGX agent

arXiv:2606.24946v1 Announce Type: new Abstract: Learned world models are useful only over horizons on which their rollout error remains controlled. We study trust-horizon certification for latent worl

Elo-Disentangled Player-Style Embeddings for Human Chess via Rating-Conditioned Residual Move Model

Model ReleasesDGX agent

arXiv:2606.25176v1 Announce Type: new Abstract: We study representation learning for individual human chess style: a per-player embedding learned from a player's move history such that inner products

Feds deny Polestar authorization to sell cars in US from model year 2027

IndustryDGX agent

The U.S. has denied authorization for Polestar to sell 2027 model-year vehicles, effectively preventing new Polestar models from entering the U.S. market. The Connected Vehicle Rule restricts vehicles

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

Model ReleasesDGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣

Model ReleasesDGX agent

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣 llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance j

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

Model ReleasesDGX agent

arXiv:2606.25402v1 Announce Type: cross Abstract: Large software projects often depend on older versions of libraries, even as APIs continue to evolve across releases. This creates a challenge for LLM

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

SafetyDGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

Privacy-Aware Visual Language Models

Model ReleasesDGX agent

arXiv:2405.17423v4 Announce Type: replace-cross Abstract: As Visual Language Models (VLMs) become increasingly embedded in everyday applications, ensuring they can recognise and appropriately handle p

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

Model ReleasesDGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

Model ReleasesDGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

Wan-Streamer v0.1: End-to-end Real-time Interactive Foundation Models

ResearchDGX agent

arXiv:2606.25041v1 Announce Type: new Abstract: We present Wan-Streamer, a native-streaming, end-to-end interactive foundation model designed from the ground up for real-time, low-latency, full-duplex

24 Jun 2026

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

Model ReleasesDGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts (Kevin Schaul/Washington Post)

Model ReleasesDGX agent

Kevin Schaul / Washington Post: An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts — Excerp

Autonomous Video Generation with Counterfactual Controllability for Self-Evolving World Models

AgentsDGX agent

arXiv:2606.24152v1 Announce Type: new Abstract: Existing literature claims that video generation essentially is world modelling. On the one hand, the claim is productive because it pushes generative A

CALIBER: Calibrating Confidence Before and After Reasoning in Language Models

SafetyDGX agent

arXiv:2606.24281v1 Announce Type: cross Abstract: Reasoning language models are increasingly asked not only to answer difficult questions, but also to estimate their likelihood of success. Existing me

FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation

Model ReleasesDGX agent

arXiv:2507.16696v3 Announce Type: replace-cross Abstract: Industrial signal analysis is hindered by severe data heterogeneity, which we characterize as the M5 problem. Existing solutions rely on speci

Geometric Action Model for Robot Policy Learning

SafetyDGX agent

arXiv:2606.17046v2 Announce Type: replace-cross Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot actions interact in the 3D physi

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest …

ToolsDGX agent

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest open-source frontier model without changing your workflow. Y

L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Models

Model ReleasesDGX agent

arXiv:2606.24825v1 Announce Type: new Abstract: Part-of-Speech (POS) tagging is a foundational NLP task underpinning machine translation, information extraction, and syntactic parsing. Despite Marathi

Maestro Order: A Model-Agnostic Orchestration Harness

ResearchDGX agent

arXiv:2606.23983v1 Announce Type: cross Abstract: A single forward pass of a capable model is a fast, fluent, and unreliable problem-solver: it is right often enough to be useful and wrong often enoug

MambaRaw: Selective State Space Modeling for Efficient 4K Raw Image Reconstruction

Model ReleasesDGX agent

arXiv:2606.24479v1 Announce Type: new Abstract: In-camera JPEG previews are ubiquitous in raw image formats and provide an sRGB reference at negligible storage cost. Although existing metadata-based r

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

Model ReleasesDGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

On the Stability of Prompt Ranking in Large Language Model Evaluation

Model ReleasesDGX agent

arXiv:2606.24381v1 Announce Type: cross Abstract: Prompt-based interaction has become a dominant paradigm for using large language models (LLMs), where multiple candidate prompts are evaluated and the

← Previous
1…104105106107108…1009
Next →