AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
1 Jun 2026

OpenAI frontier models and Codex are now available on AWS

ApplicationsDGX agent

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new path to build with OpenAI through the AWS environments, controls, and procurement workflows they already u

Reduced-order modeling of Hamiltonian dynamics based on symplectic neural networks

ResearchDGX agent

arXiv:2508.11911v2 Announce Type: replace-cross Abstract: We introduce a novel data-driven symplectic induced-order modeling (ROM) framework for high-dimensional Hamiltonian systems that unifies laten

Representation Collapse in Sequential Post-Training of Large Language Models

SafetyDGX agent

arXiv:2605.30524v1 Announce Type: new Abstract: Large language models are now adapted through chains of post-training stages rather than through a single instruction-tuning pass. This paper studies wh

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Generation

HardwareDGX agent

arXiv:2605.30716v1 Announce Type: cross Abstract: Generating clinically useful pathology reports for pathology cases from whole-slide images (WSIs) is challenging due to gigapixel resolution, long vis

Subspace-Decomposed JEPAs: Disentangling Progression and Content in Latent World Models

AgentsDGX agent

arXiv:2605.31111v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn compact latent world models by predicting future embeddings, but no single coordinate of the late

This was right five years ago, and still is: “Large scale pretrained models are certainly likely to figure prominently in artificial intelli…

SafetyDGX agent

This was right five years ago, and still is: “Large scale pretrained models are certainly likely to figure prominently in artificial intelligence for the near future, and play an important role in com

TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments

ResearchDGX agent

arXiv:2602.02459v2 Announce Type: replace Abstract: Robots in dynamic, human-centric environments must follow language instructions while maintaining real-time reactive control. Vision-language-action

Towards Atoms of Large Language Models

ResearchDGX agent

arXiv:2509.20784v3 Announce Type: replace-cross Abstract: The fundamental representational units (FRUs) of large language models (LLMs) remain undefined, limiting further understanding of their underl

TripoSplat, an open-source image-to-3D Gaussian model from @tripoai, has Day-0 support in ComfyUI One 2D image in, a 3D Gaussian asset out, …

Local AiDGX agent

TripoSplat, an open-source image-to-3D Gaussian model from @tripoai, has Day-0 support in ComfyUI One 2D image in, a 3D Gaussian asset out, strong on creative designs, stylized props and characters. K

We'll get into @MiniMax_AI M3's model performance, the MSA architecture and what it means for long context, and how Together is optimizing i…

ToolsDGX agent

We'll get into @MiniMax_AI M3's model performance, the MSA architecture and what it means for long context, and how Together is optimizing inference and KV-cache for this new architecture. Set your re

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discoverin…

AgentsDGX agent

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discovering new capabilities every day, it's crazy for 1B active parame

What Gets Unmasked First? Trajectory Analysis of Diffusion Models for Graph-to-Text Generation

ResearchDGX agent

arXiv:2605.31564v1 Announce Type: cross Abstract: We present the first systematic study of masked diffusion language models (MDLMs) for graph-to-text generation. We analyze MDLM generation trajectorie

When English Rewrites Local Knowledge: Global Narrative Dominance in Large Language Models

Local AiDGX agent

arXiv:2605.30481v1 Announce Type: new Abstract: Large language models (LLMs) are widely used as cross-lingual knowledge interfaces. However, culturally grounded questions often reflect globally domina

World2Act: Latent Action Post-Training from World Model Dynamics

SafetyDGX agent

arXiv:2603.10422v2 Announce Type: replace Abstract: World Models (WMs) offer a promising mechanism for post-training Vision-Language-Action (VLA) policies by providing dynamics priors that improve gen

31 May 2026

/goal and other fully automated AI agents are cool, but not a great model for the future of work with people. Instead you want your AI to kn…

ApplicationsDGX agent

/goal and other fully automated AI agents are cool, but not a great model for the future of work with people. Instead you want your AI to know when to ask you GOOD questions, maybe because it is stuck

30 May 2026

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the o…

SafetyDGX agent

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the open on @huggingface, so researchers everywhere can scrutiniz

Memory chip makers are leveraging their newfound power to secure long-term agreements, a move set to reshape the industry's business model and stabilize prices (Dan Gallagher/Wall Street Journal)

IndustryDGX agent

Dan Gallagher / Wall Street Journal: Memory chip makers are leveraging their newfound power to secure long-term agreements, a move set to reshape the industry's business model and stabilize prices — M

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, an…

AgentsDGX agent

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, and multimodal workflows — and Hermes Agent users have been lo

29 May 2026

3DVLA: Enhancing Vision-Language-Action Models via 3D Spatial and Instance Understanding

ResearchDGX agent

arXiv:2605.29416v1 Announce Type: cross Abstract: Vision-Language-Action models have achieved remarkable progress in robotic manipulation, yet they suffer from a critical limitation: a lack of 3D scen

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference

Local AiDGX agent

arXiv:2605.29535v1 Announce Type: new Abstract: Vision-Language Models (VLMs) process thousands of visual tokens per image alongside comparatively few text tokens, yet existing compression methods tre

BadBlocks: Low-Cost and Stealthy Backdoor Attacks Tailored for Text-to-Image Diffusion Models

HardwareDGX agent

arXiv:2508.03221v5 Announce Type: replace-cross Abstract: Despite the remarkable progress of diffusion models in image generation, recent studies reveal their vulnerability to backdoor attacks via cov

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning

ResearchDGX agent

arXiv:2605.30231v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) often struggle with robust 3D spatial reasoning. Prevailing methods that rely on fine-tuning with 3D visual question-ans

BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inference

Local AiDGX agent

arXiv:2605.29233v1 Announce Type: cross Abstract: Diffusion language models (dLLMs) generate text by iteratively denoising multiple token positions in parallel, offering an attractive alternative to s

CA-AC-MPC: CUDA-Accelerated Actor-Critic Model Predictive Control

HardwareDGX agent

arXiv:2605.29155v1 Announce Type: cross Abstract: In the literature, actor-critic model predictive control (AC-MPC) integrates MPC with reinforcement learning to enable high-performance control of com

ComfyUI just added @OpenRouter support. Instead of being locked into a single LLM, you can now access 20+ models directly inside Comfy. More…

Local AiDGX agent

ComfyUI just added @OpenRouter support. Instead of being locked into a single LLM, you can now access 20+ models directly inside Comfy. More flexibility, less friction, same workflow. Links to the wor

Computational Modeling of Antibody-Antigen Complexes: PLM-Based and MSA-Based Approaches

ResearchDGX agent

arXiv:2605.28886v1 Announce Type: cross Abstract: Antibodies play a central role in the immune response by specifically recognizing and neutralizing antigens, and therapeutic antibodies have become ma

Convex Basins in Single-Index Model Loss Landscapes: Applications to Robust Recovery under Strong Adversarial Corruption

ResearchDGX agent

arXiv:2605.29497v1 Announce Type: new Abstract: We study the problem of robustly learning Gaussian Single Index Models (SIMs) in the presence of heavy-tailed noise and a constant fraction of adversari

Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get production-grade performance from models like @Kimi_Moon…

ApplicationsDGX agent

Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get production-grade performance from models like @Kimi_Moonshot, @Alibaba_Qwen, and @DeepSeek_ai at 20x+ lower cost tha

GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2605.29398v1 Announce Type: cross Abstract: Reinforcement learning (RL) can be used to improve the policy (denoiser) of diffusion large language models (dLLMs), while being hindered by the intra

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic …

AgentsDGX agent

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic coding. Priced at 1/m input and 2/m output, it’s extremely c

HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis

ResearchDGX agent

arXiv:2508.10566v3 Announce Type: replace Abstract: Audio-driven talking head generation faces a fundamental trade-off between personalization and generalization, limiting its practical application. I

Intent-aligned Autonomous Spacecraft Guidance via Reasoning Models

SafetyDGX agent

arXiv:2604.17176v2 Announce Type: replace-cross Abstract: Future spacecraft operations require autonomy that can interpret high-level mission intent while preserving safety. However, existing trajecto

LangSmith LLM Gateway lets you enforce spend limits and redacts PII before requests reach the model. Not after the fact.

AgentsDGX agent

LangSmith LLM Gateway is a feature that enables proactive cost and privacy controls by enforcing spending limits and redacting personally identifiable information (PII) before API requests are sent to

Masked Diffusion Modeling for Anomaly Detection

SafetyDGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

mcp-proto-okn: Natural-language access to open scientific knowledge graphs through the Model Context Protocol

AgentsDGX agent

arXiv:2605.30283v1 Announce Type: new Abstract: MCP Server Proto-OKN (mcp-proto-okn) is a Python-based Model Context Protocol server that enables AI assistants to discover, inspect, query and integrat

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

SafetyDGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

On Asymmetric Optimization of Reasoning and Perception in Vision-Language Model Post-Training

ResearchDGX agent

arXiv:2605.29496v1 Announce Type: new Abstract: Post-training has greatly improved reasoning in frontier vision-language models, yet its gains for perception remain comparatively limited, creating a b

Opus 4.8 is insane, nothing will be the same after this model 💀

SafetyDGX agent

Gary Marcus expresses strong enthusiasm about Opus 4.8, suggesting it represents a significant breakthrough in AI capabilities. The post implies the model introduces substantial improvements or novel

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Herme…

AgentsDGX agent

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Hermes Agent! ⚡️ Step 3.7 Flash is here: The new frontier is agen

Tencent bets on smaller AI models in the race with Chinese rivals, as EVP Dowson Tong says AI now contributes 20%+ of its revenue and 95%+ of new internal code (Cissy Zhou/Nikkei Asia)

IndustryDGX agent

Cissy Zhou / Nikkei Asia: Tencent bets on smaller AI models in the race with Chinese rivals, as EVP Dowson Tong says AI now contributes 20%+ of its revenue and 95%+ of new internal code — HONG KONG —

The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models

SafetyDGX agent

arXiv:2605.29123v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) uniquely support any-order generation, with confidence-based decoding currently serving as the de facto standard

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in …

HardwareDGX agent

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in under 10 seconds. This deep dive shows the systems work behi

Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge

ResearchDGX agent

arXiv:2505.16178v2 Announce Type: replace Abstract: While fine-tuning is the standard for injecting factual knowledge into large language models (LLMs), the mechanisms enabling reliable fact recall vi

UniNote: A Unified Embedding Model for Multimodal Representation and Ranking

Local AiDGX agent

arXiv:2605.29287v1 Announce Type: cross Abstract: Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines t

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

28 May 2026

Affective Music Recommendation: A Rollout-Based World Model for Offline Preference Optimization

SafetyDGX agent

arXiv:2605.28810v1 Announce Type: new Abstract: Functional music applications, from consumer focus and sleep aids to clinical interventions, share a distinctive recommendation problem: success is defi

Automatic Pruning Discovery for Large Language Models

TutorialsDGX agent

arXiv:2511.15390v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable performance on a wide range of tasks, hindering real-world deployment due to their massive siz

Chreode: A Cell World Model for One-Step Temporal Dynamics and Perturbation Prediction

ResearchDGX agent

arXiv:2605.28111v1 Announce Type: new Abstract: Predicting how a cell will change its transcriptional state under a developmental signal or a genetic perturbation is the computational core of in-silic

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models

Local AiDGX agent

arXiv:2605.28115v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face severe memory and latency bottlenecks due to high-resolution visual tokens. While current token reduction methods the

ClinicalEncoder26AM: A Multlilingual Diagnosable ColBERT Model; Evidences from the MultiClinNER Shared Task

Local AiDGX agent

arXiv:2605.28521v1 Announce Type: new Abstract: ClinicalEncoder26AM is a multilingual Diagnosable ColBERT for clinical and biomedical texts, which aligns at multiple levels its token-level semantic wi

Dell and H2O.ai target the token-cost problem with vertical AI models

ApplicationsDGX agent

As artificial intelligence adoption accelerates inside enterprises, the economics of generative AI are forcing a fundamental rethink. Runaway token costs, data sovereignty demands and a growing gap be

DODO: Discrete OCR Diffusion Models

ResearchDGX agent

arXiv:2602.16872v2 Announce Type: replace Abstract: Optical Character Recognition (OCR) is a fundamental task for digitizing information, serving as a critical bridge between visual data and textual u

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expe…

SafetyDGX agent

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expert mathematicians poured over this [transcript] and identifi

London- and SF-based Orbital Industries, which uses its Orb model to design advanced materials and then sell them directly, raised a $50M Series B led by Plural (Jeremy Kahn/Fortune)

IndustryDGX agent

Jeremy Kahn / Fortune: London- and SF-based Orbital Industries, which uses its Orb model to design advanced materials and then sell them directly, raised a 50M Series B led by Plural — Orbital Industr

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

ApplicationsDGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

AgentsDGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

Representation-Conditioned Diffusion Models for Guided Training Data Generation

ApplicationsDGX agent

arXiv:2605.27495v1 Announce Type: new Abstract: Data availability remains a critical bottleneck in many deep learning applications. Large-scale datasets are often expensive to collect, curate and anno

Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models

ResearchDGX agent

arXiv:2605.27813v1 Announce Type: cross Abstract: Text-to-image diffusion models generate images through an iterative denoising process, so internal neural layers produce trajectories of activations r

Risk-aware Selective Prompting for Hallucination Mitigation in Large Vision-Language Models

ResearchDGX agent

arXiv:2605.28123v1 Announce Type: new Abstract: Prompt-based verification is widely used to mitigate hallucinations in large vision-language models (LVLMs), yet when it helps remains poorly understood

ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains

SafetyDGX agent

arXiv:2605.28014v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) improves the reasoning performance of large language models (LLMs) by providing dense token-level supervision for on-

← Previous
1…213214215216217…1017
Next →