AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,508 results
Applications

Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get production-grade performance from models like @Kimi_Moon…

DGX agent

Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get production-grade performance from models like @Kimi_Moonshot, @Alibaba_Qwen, and @DeepSeek_ai at 20x+ lower cost tha

applicationsharrison-chase--x
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models

DGX agent

arXiv:2605.29398v1 Announce Type: cross Abstract: Reinforcement learning (RL) can be used to improve the policy (denoiser) of diffusion large language models (dLLMs), while being hindered by the intra

safetyarxiv-cs-ai
29 May 2026
Agents

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic …

DGX agent

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic coding. Priced at 1/m input and 2/m output, it’s extremely c

agentselon-musk--x
29 May 2026
Research

HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis

DGX agent

arXiv:2508.10566v3 Announce Type: replace Abstract: Audio-driven talking head generation faces a fundamental trade-off between personalization and generalization, limiting its practical application. I

researcharxiv-cs-cv
29 May 2026
Safety

Intent-aligned Autonomous Spacecraft Guidance via Reasoning Models

DGX agent

arXiv:2604.17176v2 Announce Type: replace-cross Abstract: Future spacecraft operations require autonomy that can interpret high-level mission intent while preserving safety. However, existing trajecto

safetyarxiv-cs-ai
29 May 2026
Agents

LangSmith LLM Gateway lets you enforce spend limits and redacts PII before requests reach the model. Not after the fact.

DGX agent

LangSmith LLM Gateway is a feature that enables proactive cost and privacy controls by enforcing spending limits and redacting personally identifiable information (PII) before API requests are sent to

agentsharrison-chase--x
29 May 2026
Safety

Masked Diffusion Modeling for Anomaly Detection

DGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

safetyarxiv-cs-ai
29 May 2026
Agents

mcp-proto-okn: Natural-language access to open scientific knowledge graphs through the Model Context Protocol

DGX agent

arXiv:2605.30283v1 Announce Type: new Abstract: MCP Server Proto-OKN (mcp-proto-okn) is a Python-based Model Context Protocol server that enables AI assistants to discover, inspect, query and integrat

agentsarxiv-cs-ai
29 May 2026
Safety

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

DGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

safetyarxiv-cs-lg
29 May 2026
Research

On Asymmetric Optimization of Reasoning and Perception in Vision-Language Model Post-Training

DGX agent

arXiv:2605.29496v1 Announce Type: new Abstract: Post-training has greatly improved reasoning in frontier vision-language models, yet its gains for perception remain comparatively limited, creating a b

researcharxiv-cs-cl
29 May 2026
Safety

Opus 4.8 is insane, nothing will be the same after this model 💀

DGX agent

Gary Marcus expresses strong enthusiasm about Opus 4.8, suggesting it represents a significant breakthrough in AI capabilities. The post implies the model introduces substantial improvements or novel

safetygary-marcus--x
29 May 2026
Agents

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Herme…

DGX agent

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Hermes Agent! ⚡️ Step 3.7 Flash is here: The new frontier is agen

agentsnous-research--x
29 May 2026
Industry

Tencent bets on smaller AI models in the race with Chinese rivals, as EVP Dowson Tong says AI now contributes 20%+ of its revenue and 95%+ of new internal code (Cissy Zhou/Nikkei Asia)

DGX agent

Cissy Zhou / Nikkei Asia: Tencent bets on smaller AI models in the race with Chinese rivals, as EVP Dowson Tong says AI now contributes 20%+ of its revenue and 95%+ of new internal code — HONG KONG —

industrytechmeme
29 May 2026
Safety

The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models

DGX agent

arXiv:2605.29123v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) uniquely support any-order generation, with confidence-based decoding currently serving as the de facto standard

safetyarxiv-cs-ai
29 May 2026
Hardware

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in …

DGX agent

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in under 10 seconds. This deep dive shows the systems work behi

hardwaretogether-ai--x
29 May 2026
Research

Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge

DGX agent

arXiv:2505.16178v2 Announce Type: replace Abstract: While fine-tuning is the standard for injecting factual knowledge into large language models (LLMs), the mechanisms enabling reliable fact recall vi

researcharxiv-cs-cl
29 May 2026
Local Ai

UniNote: A Unified Embedding Model for Multimodal Representation and Ranking

DGX agent

arXiv:2605.29287v1 Announce Type: cross Abstract: Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines t

local-aiarxiv-cs-cv
29 May 2026
Applications

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

DGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

applicationsarxiv-cs-ai
29 May 2026
Safety

Affective Music Recommendation: A Rollout-Based World Model for Offline Preference Optimization

DGX agent

arXiv:2605.28810v1 Announce Type: new Abstract: Functional music applications, from consumer focus and sleep aids to clinical interventions, share a distinctive recommendation problem: success is defi

safetyarxiv-cs-lg
28 May 2026
Tutorials

Automatic Pruning Discovery for Large Language Models

DGX agent

arXiv:2511.15390v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable performance on a wide range of tasks, hindering real-world deployment due to their massive siz

tutorialsarxiv-cs-cv
28 May 2026
Research

Chreode: A Cell World Model for One-Step Temporal Dynamics and Perturbation Prediction

DGX agent

arXiv:2605.28111v1 Announce Type: new Abstract: Predicting how a cell will change its transcriptional state under a developmental signal or a genetic perturbation is the computational core of in-silic

researcharxiv-cs-lg
28 May 2026
Local Ai

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models

DGX agent

arXiv:2605.28115v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face severe memory and latency bottlenecks due to high-resolution visual tokens. While current token reduction methods the

local-aiarxiv-cs-ai
28 May 2026
Local Ai

ClinicalEncoder26AM: A Multlilingual Diagnosable ColBERT Model; Evidences from the MultiClinNER Shared Task

DGX agent

arXiv:2605.28521v1 Announce Type: new Abstract: ClinicalEncoder26AM is a multilingual Diagnosable ColBERT for clinical and biomedical texts, which aligns at multiple levels its token-level semantic wi

local-aiarxiv-cs-cl
28 May 2026
Applications

Dell and H2O.ai target the token-cost problem with vertical AI models

DGX agent

As artificial intelligence adoption accelerates inside enterprises, the economics of generative AI are forcing a fundamental rethink. Runaway token costs, data sovereignty demands and a growing gap be

applicationssiliconangle
28 May 2026
Research

DODO: Discrete OCR Diffusion Models

DGX agent

arXiv:2602.16872v2 Announce Type: replace Abstract: Optical Character Recognition (OCR) is a fundamental task for digitizing information, serving as a critical bridge between visual data and textual u

researcharxiv-cs-cv
28 May 2026
Safety

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expe…

DGX agent

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expert mathematicians poured over this [transcript] and identifi

safetygary-marcus--x
28 May 2026
Industry

London- and SF-based Orbital Industries, which uses its Orb model to design advanced materials and then sell them directly, raised a $50M Series B led by Plural (Jeremy Kahn/Fortune)

DGX agent

Jeremy Kahn / Fortune: London- and SF-based Orbital Industries, which uses its Orb model to design advanced materials and then sell them directly, raised a 50M Series B led by Plural — Orbital Industr

industrytechmeme
28 May 2026
Applications

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

DGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

applicationsarxiv-cs-cl
28 May 2026
Agents

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

DGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

agentsarxiv-cs-cl
28 May 2026
Applications

Representation-Conditioned Diffusion Models for Guided Training Data Generation

DGX agent

arXiv:2605.27495v1 Announce Type: new Abstract: Data availability remains a critical bottleneck in many deep learning applications. Large-scale datasets are often expensive to collect, curate and anno

applicationsarxiv-cs-cv
28 May 2026
Research

Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models

DGX agent

arXiv:2605.27813v1 Announce Type: cross Abstract: Text-to-image diffusion models generate images through an iterative denoising process, so internal neural layers produce trajectories of activations r

researcharxiv-cs-ai
28 May 2026
Research

Risk-aware Selective Prompting for Hallucination Mitigation in Large Vision-Language Models

DGX agent

arXiv:2605.28123v1 Announce Type: new Abstract: Prompt-based verification is widely used to mitigate hallucinations in large vision-language models (LVLMs), yet when it helps remains poorly understood

researcharxiv-cs-cl
28 May 2026
Safety

ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains

DGX agent

arXiv:2605.28014v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) improves the reasoning performance of large language models (LLMs) by providing dense token-level supervision for on-

safetyarxiv-cs-cl
28 May 2026
Safety

SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Models

DGX agent

arXiv:2605.28338v1 Announce Type: new Abstract: Large language models(LLMs) increasingly match expert performance on licensing examinations, yet routine clinical use remains limited because governance

safetyarxiv-cs-ai
28 May 2026
Research

SANTS: A State-Adaptive Scheduler for World Action Models

DGX agent

arXiv:2605.27947v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by using video-based future representations to condition action generation. In pixel-space WAMs, h

researcharxiv-cs-ro
28 May 2026
Research

Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

DGX agent

arXiv:2602.04898v3 Announce Type: replace-cross Abstract: Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. E

researcharxiv-cs-ai
28 May 2026
Research

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

DGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

researcharxiv-cs-ai
28 May 2026
Hardware

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

DGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

hardwareclem-delangue--x
28 May 2026
Research

A Unified Framework for Diffusion Model Unlearning with f-Divergence

DGX agent

arXiv:2509.21167v2 Announce Type: replace-cross Abstract: Most existing methods for concept unlearning in text-to-image diffusion models minimize a mean squared error (MSE) loss between the denoiser o

researcharxiv-cs-cv
27 May 2026
Applications

Agile Online Model Selection: Resolving Adaptation Lag via Safeguarded Large Learning Rates

DGX agent

arXiv:2605.26919v1 Announce Type: new Abstract: Maintaining predictive accuracy in non-stationary environments requires online model selection to adapt autonomously to unknown distribution shifts. How

applicationsarxiv-cs-lg
27 May 2026
Research

Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis

DGX agent

arXiv:2512.14561v2 Announce Type: replace Abstract: Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compare

researcharxiv-cs-cl
27 May 2026
Research

An In-Vitro Study on Cross-Lingual Generalization in Language Models

DGX agent

arXiv:2605.26683v1 Announce Type: cross Abstract: Cross-lingual transfer in language models is difficult to study in natural corpora because lexical overlap, morphology, data imbalance, and tokenizati

researcharxiv-cs-ai
27 May 2026
Research

Clozing the Gap: Exploring Why Language Model Surprisal Outperforms Cloze Surprisal

DGX agent

arXiv:2601.09886v2 Announce Type: replace Abstract: How predictable a word is can be quantified in two ways: using human responses to the cloze task or using probabilities from language models (LMs).W

researcharxiv-cs-cl
27 May 2026
Safety

Conv-to-Bench: Evaluating Language Models Via User-Assistant Dialogues In Code Tasks

DGX agent

arXiv:2605.26440v1 Announce Type: new Abstract: The rapid advancement of Large Language Models (LLMs) has outpaced the scalability of traditional evaluation benchmarks, which remain heavily dependent

safetyarxiv-cs-cl
27 May 2026
Applications

DinoComplete: 3D Shape Completion with Distilled Semantic Priors and State Space Models

DGX agent

arXiv:2605.26949v1 Announce Type: new Abstract: 3D shape completion from partial scans remains challenging for unseen categories and noisy real-world observations, where geometry alone is often insuff

applicationsarxiv-cs-cv
27 May 2026
Research

Explainable Comparison of Feature-Based and Deep Learning Models for TROPOMI Methane Plume Screening

DGX agent

arXiv:2605.27236v1 Announce Type: new Abstract: Continuous and global detection of large methane emissions is a crucial step for global warming mitigation. Satellite observations, such as from S5P/TRO

researcharxiv-cs-lg
27 May 2026
Safety

FTibSuite: A Comprehensive Resource Suite for Tibetan Vision-Language Modeling

DGX agent

arXiv:2605.26601v1 Announce Type: new Abstract: Vision-language models have progressed rapidly, but Tibetan remains a severely underserved low-resource language due to the lack of reproducible trainin

safetyarxiv-cs-cv
27 May 2026
Safety

GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation

DGX agent

arXiv:2602.16449v2 Announce Type: replace-cross Abstract: Generative model evaluation commonly relies on high-dimensional embedding spaces to compute distances between samples. We show that dataset re

safetyarxiv-cs-ai
27 May 2026
← Previous
1…274275276277278…1303
Next →