AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
28 Jun 2026

Elon Musk sitting by the Model 3 production line in 2017 on his birthday Never give up 🙌

ApplicationsDGX agent

In 2017, Elon Musk was photographed at the Tesla Model 3 production line on his birthday, with the accompanying message 'Never give up' emphasizing persistence and determination. The image captured a

Getting regulated by a government because your model is 'too dangerous' is the best marketing (especially for enterprise sales) so everyone …

ApplicationsDGX agent

Clem Delangue suggests that government regulation of AI models perceived as 'dangerous' can paradoxically serve as effective marketing, particularly for enterprise sales. The implication is that regul

In my experience, all model routers underestimate the difficulty of non-math/coding tasks and assign them too little intelligence. This is w…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications
DGX agent

In my experience, all model routers underestimate the difficulty of non-math/coding tasks and assign them too little intelligence. This is worth addressing, as non-verifiable tasks (innovation, market

So what model is OpenAI saving the GPT-6 label for?

ApplicationsDGX agent

Ethan Mollick discusses OpenAI's naming convention and model release strategy, speculating on what capabilities or timeline the company might be reserving the GPT-6 designation for rather than applyin

the future of AI is multi-model (including a majority of open-source ones provided by @huggingface of course!)!

TutorialsDGX agent

the future of AI is multi-model (including a majority of open-source ones provided by @huggingface of course!)! How to keep AI spend flat while token usage grows exponentially: Not with friction and s

26 Jun 2026

EndoUFM: Utilizing Foundation Models for Monocular depth estimation of endoscopic images

Local AiDGX agent

arXiv:2508.17916v2 Announce Type: replace Abstract: Depth estimation is a foundational component for 3D reconstruction in minimally invasive endoscopic surgeries. However, existing monocular depth est

Generative Models on Analog Hardware with Dynamics

ResearchDGX agent

arXiv:2606.27294v1 Announce Type: cross Abstract: Analog hardware platforms such as coupled oscillators and Analog Ising Machines naturally solve differential equations at a fraction of the energy cos

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

SafetyDGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision

SafetyDGX agent

arXiv:2606.26801v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for generalizable robotic manipulation. During fine-tuning, however, action supervisio

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models

Model ReleasesDGX agent

arXiv:2606.26379v1 Announce Type: new Abstract: Visual prompt tuning has emerged as a parameter-efficient fine-tuning approach for adapting large-scale Vision Transformers (ViTs) to downstream tasks.

LithoDreamer: A Physics-Informed World Model for Multi-Stage Computational Lithography

ResearchDGX agent

arXiv:2606.26713v1 Announce Type: new Abstract: As semiconductor technology nodes scale, computational lithography is essential for ensuring yield and performance. However, lithography is a continuous

Racing a Wheeled Quadruped: Active Load Transfer Mitigation via Model Predictive Control

SafetyDGX agent

arXiv:2606.26313v1 Announce Type: new Abstract: This paper presents a hierarchical control framework using model predictive control (MPC) and reinforcement learning (RL) for active roll control to man

ReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning Models

ApplicationsDGX agent

arXiv:2606.26986v1 Announce Type: cross Abstract: Open Relation Extraction (OpenRE) requires a model to extract unseen relations between head and tail entities from unstructured text for real-world ap

Reasoning Quality Emerges Early: Data Curation for Reasoning Models

ResearchDGX agent

arXiv:2606.26797v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) on a small, high-quality set of long reasoning traces is an effective approach for eliciting strong reasoning capabilities

Reconstruction Alignment Improves Unified Multimodal Models

SafetyDGX agent

arXiv:2509.07295v4 Announce Type: replace-cross Abstract: Unified multimodal models (UMMs) unify visual understanding and generation within a single architecture. However, conventional training relies

Reducing Conversational Escalation in Large Language Model Dialogue with Nonviolent Communication Constraints

SafetyDGX agent

arXiv:2606.26106v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in emotionally charged situations involving interpersonal conflict, frustration, and distress. Whil

Risk-Aware Selective Multimodal Driver Monitoring with Driver-State World Modeling

SafetyDGX agent

arXiv:2606.26922v1 Announce Type: cross Abstract: Continuous driver monitoring in automated vehicles requires low-latency inference while avoiding unsafe decisions under uncertain driver states. Large

Tactile-WAM: Touch-Aware World Action Model with Tactile Asymmetric Attention

SafetyDGX agent

arXiv:2606.26663v1 Announce Type: new Abstract: World Action Models (WAMs) generate actions together with predicted futures, offering a powerful interface for robot decision making. In contact-rich ma

TEMPO-Diffusion: Temporally Exposed Malicious Poisoning of Diffusion Models

ResearchDGX agent

arXiv:2606.26285v1 Announce Type: cross Abstract: Noise-based backdoor attacks on diffusion models typically rely on input-time trigger injection, untargeted activation, and out-of-distribution target

The Riddle Riddle: Testing Flexible Reasoning in Large Language Models and Humans

ResearchDGX agent

arXiv:2606.27103v1 Announce Type: new Abstract: Humans flexibly adapt their reasoning strategies to the requirements of a given problem. Large language models (LLMs) have performed well on many cognit

UltraStar: Semantic-Aware Star Graph Modeling for Echocardiography Navigation

ResearchDGX agent

arXiv:2603.01461v2 Announce Type: replace Abstract: Echocardiography is critical for diagnosing cardiovascular diseases, yet the shortage of skilled sonographers hinders timely patient care, due to hi

Vulnerability of Natural Language Classifiers to Evolutionary Generated Adversarial Text

Model ReleasesDGX agent

arXiv:2606.27215v1 Announce Type: new Abstract: Deep learning models have achieved impressive performance across various fields but remain vulnerable to adversarial inputs, particularly in NLP, where

25 Jun 2026

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models

SafetyDGX agent

arXiv:2606.25380v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed across languages, but their safety behavior remains uneven across linguistic and cultural context

Confidence Sequences for Online Statistical Model Checking of Markov Decision Processes

ResearchDGX agent

arXiv:2606.25797v1 Announce Type: new Abstract: Markov decision processes (MDPs) are a classic model of decision making under uncertainty, exhibiting both non-deterministic choice as well as probabili

Cross-Modal Robustness Transfer (CMRT): Training Robust Speech Translation Models Using Adversarial Text

ApplicationsDGX agent

arXiv:2602.11933v2 Announce Type: replace Abstract: End-to-End Speech Translation (E2E-ST) has seen significant advancements, yet current models are primarily benchmarked on curated, 'clean' datasets.

Evaluating Japanese Dialect Robustness Across Speech and Text-based Large Language Models

ResearchDGX agent

arXiv:2606.25436v1 Announce Type: cross Abstract: Dialogue systems based on large language models (LLMs) have advanced significantly in recent years. However, dialectal variation remains a major chall

Evaluating performance and efficiency of the GitHub Copilot agentic harness across models and tasks

AgentsDGX agent

Explore how the GitHub Copilot agentic harness delivers strong results across multiple benchmarks and leading token efficiency, while maintaining flexibility to choose among more than 20 models. The p

Explainable Control Framework (XCF) based on Fuzzy Model-Agnostic Explanation and LLM Agent-Supported Interface

Local AiDGX agent

arXiv:2606.25941v1 Announce Type: cross Abstract: Increasing demand for precise and reliable control in complex scenarios has led to the development of increasingly sophisticated controllers, includin

ExTra: Exploratory Trajectory Optimization for Language Model Reinforcement Learning

ResearchDGX agent

arXiv:2606.24994v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) for language-model reasoning can fail at both extremes of task difficulty: easy prompts often prod

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control

SafetyDGX agent

arXiv:2502.02205v4 Announce Type: replace Abstract: The application of deep learning for partial differential equation (PDE)-constrained control is gaining increasing attention. However, existing meth

How Large Language Models Source Brand Reputation Across Languages and Markets

ResearchDGX agent

arXiv:2606.25787v1 Announce Type: cross Abstract: When a large language model (LLM) answers a question about a company, it grounds the answer in retrieved web sources, and those sources decide what th

Learning task-specific subspaces via interventional post-training of speech foundation models

TutorialsDGX agent

arXiv:2606.17967v2 Announce Type: replace Abstract: Speech foundation models, pre-trained on large corpora of unlabelled speech data, produce general-purpose representations which are useful across ta

Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure

SafetyDGX agent

arXiv:2403.11425v4 Announce Type: replace-cross Abstract: Cancer treatments are known to introduce cardiotoxicity, negatively impacting outcomes and survivorship. Identifying cancer patients at risk o

PhoneBuddy: Training Open Models for Agentic Phone Use

AgentsDGX agent

arXiv:2606.23049v2 Announce Type: replace Abstract: Phones are becoming an important execution surface for general-purpose agents, but training open models for reliable phone use remains difficult bec

RAS: Measuring LLM Safety Through Refusal Alignment

Model ReleasesDGX agent

arXiv:2606.25750v1 Announce Type: cross Abstract: Safety evaluation of large language models (LLMs) is commonly performed by querying models with unsafe or jailbreak prompts and judging whether their

Riazi-8B: An Urdu Large Language Model for Mathematical Reasoning

ResearchDGX agent

arXiv:2606.25568v1 Announce Type: new Abstract: Recent LLMs demonstrate strong mathematical reasoning capabilities, but existing gains rely heavily on English-centric training resources and benchmarks

Which tokens does a hybrid model predict better?

ToolsDGX agent

This article from Allen AI discusses how hybrid models that combine different prediction approaches perform across various token types, likely comparing their effectiveness on common tokens versus rar

24 Jun 2026

A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Levy Process Dynamics

SafetyDGX agent

arXiv:2504.01482v3 Announce Type: replace-cross Abstract: This paper develops a model-based framework for continuous-time policy evaluation (CTPE) in reinforcement learning, incorporating both Brownia

Catastrophic Compositional Generation: Why Vanilla Diffusion Models Fail to Extrapolate

ResearchDGX agent

arXiv:2606.23920v1 Announce Type: cross Abstract: The task of compositional generation involves using a conditional generative model, trained only on a subset of the possible conditions, to produce sa

Ensemble Learning for Large Language Models in Text and Code Generation: A Survey

ApplicationsDGX agent

arXiv:2503.13505v3 Announce Type: replace-cross Abstract: Generative Pretrained Transformers (GPTs) are foundational Large Language Models (LLMs) for text generation. However, individual LLMs often pr

ForensicsTok: Forensics-Guided Tokenized Modeling for Image Tampering Localization

Local AiDGX agent

arXiv:2606.24538v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) offer powerful reasoning for forensic tasks, yet existing approaches utilizing exogenous segmentation decoders

From 'Aha Moments' to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

SafetyDGX agent

arXiv:2508.04460v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) can exhibit step-by-step reasoning, reflection, and backtracking, but these behaviors are often unregulated, leading t

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A lea…

ApplicationsDGX agent

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A leaderboard winner can still lose on your codebase. Training cl

Introducing Claude for Music. You can now create songs from Claude Code, Hermes, Codex, or any agent you’re using. SOTA music model @MiniMax…

Model ReleasesDGX agent

I cannot provide an accurate summary for this entry. The URL and source attribution appear inconsistent (title credits Yohei Nakajima but URL references a different user), and the post references prod

On the Smallness of the Large Language Models Scaling Exponents

SafetyDGX agent

arXiv:2606.24504v1 Announce Type: new Abstract: We discuss reasons why the scaling exponents of current Large Language Models (LLMs) applications are indicating an unsustainable regime in terms of ene

Separating Oblivious and Adaptive Models of Variable Selection

ResearchDGX agent

arXiv:2602.16568v2 Announce Type: replace-cross Abstract: Sparse recovery is among the most well-studied problems in learning theory and high-dimensional statistics. In this work, we investigate the s

Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models

SafetyDGX agent

arXiv:2606.24165v1 Announce Type: new Abstract: Reducing visual token redundancy is critical for accelerating Multimodal Large Language Models (MLLMs) without degrading cross-modal reasoning performan

Stabilizing Physics-Informed Consistency Models via Structure-Preserving Training

ResearchDGX agent

arXiv:2602.09303v2 Announce Type: replace Abstract: We propose a physics-informed consistency modeling framework for solving partial differential equations (PDEs) via fast, few-step generative inferen

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic format…

AgentsDGX agent

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic formatting, charts (it does decent at bounding boxes). Attaching t

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

SafetyDGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

ResearchDGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

23 Jun 2026

A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models

ResearchDGX agent

arXiv:2511.15098v2 Announce Type: replace Abstract: Discrete diffusion-based multimodal large language models (dMLLMs) have emerged as a promising alternative to autoregressive MLLMs thanks to their a

An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment

ResearchDGX agent

arXiv:2509.03372v3 Announce Type: replace-cross Abstract: A recent line of research on automated speaking assessment (ASA) has benefited from self-supervised learning (SSL) representations, which capt

Bayesian Model Averaging under Predictor Redundancy via Density-Ratio Posterior Compression

TutorialsDGX agent

arXiv:2606.21080v1 Announce Type: cross Abstract: Bayesian model averaging in support-indexed regression induces a posterior distribution over active predictor supports. Under predictor redundancy, po

ByteDance unveils Seedance 2.5 in Beijing, saying the AI video model can generate 30-second clips from up to 50 reference materials, up from 12 for Seedance 2.0 (Juro Osawa/The Information)

IndustryDGX agent

Juro Osawa / The Information: ByteDance unveils Seedance 2.5 in Beijing, saying the AI video model can generate 30-second clips from up to 50 reference materials, up from 12 for Seedance 2.0 — ByteDan

Compression and Retrieval: Implicit Memory Retrieval for Video World Models

ResearchDGX agent

arXiv:2606.23105v1 Announce Type: new Abstract: Video world models hold promise for simulating interactive environments, yet maintaining consistent long-term memory across complex camera trajectories

Cultural Counterfactuals: Evaluating Cultural Biases in Large Vision-Language Models with Counterfactual Examples

ResearchDGX agent

arXiv:2603.02370v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have grown increasingly powerful in recent years, but can also exhibit harmful biases. Prior studies investigat

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

SafetyDGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

Diffusion Models Adapt to Low-Dimensional Structure Under Flexible Coefficient Choices

ResearchDGX agent

arXiv:2606.23627v1 Announce Type: cross Abstract: Diffusion models are known to exploit unknown low-dimensional structure to accelerate sampling. However, existing convergence theory under low-dimensi

Extraction and Analysis of Multimodal Concepts in Vision Language Models through Sparse Autoencoders

SafetyDGX agent

arXiv:2606.21197v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated impressive performance in tasks requiring joint understanding of images and text, such as image captioni

← Previous
1…176177178179180…1010
Next →