AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
20 Apr 2026

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

Model ReleasesDGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

Seed1.8 Model Card: Towards Generalized Real-World Agency

Model ReleasesDGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification

ResearchDGX agent

arXiv:2604.15711v1 Announce Type: cross Abstract: Pathological diagnosis is highly reliant on image analysis, where Regions of Interest (ROIs) serve as the primary basis for diagnostic evidence, while

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

Model ReleasesDGX agent

arXiv:2604.15664v1 Announce Type: new Abstract: The rise of autonomous AI agents suggests that dynamic benchmark environments with built-in feedback on scientifically grounded tasks are needed to eval

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-t…

AgentsDGX agent

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-trained and gets the best out of models by being “in-distribu

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

SafetyDGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

Unveiling Stochasticity: Universal Multi-modal Probabilistic Modeling for Traffic Forecasting

ApplicationsDGX agent

arXiv:2604.16084v1 Announce Type: cross Abstract: Traffic forecasting is a challenging spatio-temporal modeling task and a critical component of urban transportation management. Current studies mainly

When Surfaces Lie: Exploiting Wrinkle-Induced Attention Shift to Attack Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.27759v3 Announce Type: replace Abstract: Visual-Language Models (VLMs) have demonstrated exceptional cross-modal understanding across various tasks, including zero-shot classification, imag

19 Apr 2026

An obvious way to release Mythos class models with uncertain autonomous ability is to make them only available on the website, like Gemini D…

Model ReleasesDGX agent

An obvious way to release Mythos class models with uncertain autonomous ability is to make them only available on the website, like Gemini Deep Think or ChatGPT Pro. Minimal risk of being used for aut

LiteParse is the best model-free, open-source document parser for AI agents. It now gets a first-class landing page on our website 💫 Our co…

Model ReleasesDGX agent

LiteParse is the best model-free, open-source document parser for AI agents. It now gets a first-class landing page on our website 💫 Our company mission is building the world's best agentic document p

18 Apr 2026

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your ag…

Model ReleasesDGX agent

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your agent, YOU do Anthropic shut down an entire company's Claude a

17 Apr 2026

An Analysis of Regularization and Fokker-Planck Residuals in Diffusion Models for Image Generation

ResearchDGX agent

arXiv:2604.15171v1 Announce Type: new Abstract: Recent work has shown that diffusion models trained with the denoising score matching (DSM) objective often violate the Fokker--Planck (FP) equation tha

Empowerment Gain and Causal Model Construction: Children and adults are sensitive to controllability and variability in their causal interventions

AgentsDGX agent

arXiv:2512.08230v2 Announce Type: replace Abstract: Learning about the causal structure of the world is a fundamental problem for human cognition. Causal models and especially causal learning have pro

EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation

Model ReleasesDGX agent

arXiv:2604.14306v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when fac

GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning

Model ReleasesDGX agent

arXiv:2603.01410v2 Announce Type: replace Abstract: Knowledge graphs provide structured and reliable information for many real-world applications, motivating increasing interest in combining large lan

Improving Language Models with Intentional Analysis

Model ReleasesDGX agent

arXiv:2502.04689v4 Announce Type: replace Abstract: Intent, a critical cognitive notion and mental state, is ubiquitous in human communication and problem-solving. Accurately understanding the underly

It's never been easier for your agent to harness the power of Hugging Face, from account creation to getting your own finetuned model, with …

AgentsDGX agent

It's never been easier for your agent to harness the power of Hugging Face, from account creation to getting your own finetuned model, with everything in between, without any manual action on your end

Language of Thought Shapes Output Diversity in Large Language Models

SafetyDGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

Multi-Persona Thinking for Bias Mitigation in Large Language Models

SafetyDGX agent

arXiv:2601.15488v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit social biases, which can lead to harmful stereotypes and unfair outcomes. We propose extbf{Multi-Persona Thinki

Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an 'early sign' of generalization, surprising researchers (Connie Loizos/TechCrunch)

Model ReleasesDGX agent

Connie Loizos / TechCrunch: Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an “early sign” of generalization, surprising researchers — Physical Int

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization

TutorialsDGX agent

arXiv:2508.10164v2 Announce Type: replace Abstract: Recent advances in Large Reasoning Models (LRMs) have demonstrated strong performance on complex tasks through long Chain-of-Thought (CoT) reasoning

Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models

ResearchDGX agent

arXiv:2506.13139v2 Announce Type: replace-cross Abstract: Modern Machine Learning (ML) and Deep Neural Networks (DNNs) often operate on high-dimensional data and rely on overparameterized models, wher

RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

SafetyDGX agent

arXiv:2604.14951v1 Announce Type: cross Abstract: Tool learning with foundation models aims to endow AI systems with the ability to invoke external resources -- such as APIs, computational utilities,

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

SafetyDGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization

ApplicationsDGX agent

arXiv:2604.15022v1 Announce Type: cross Abstract: Cost-aware routing dynamically dispatches user queries to models of varying capability to balance performance and inference cost. However, the routing

SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits

Model ReleasesDGX agent

arXiv:2604.01473v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are powerful tools for answering user queries, yet they remain highly vulnerable to jailbreak attacks. Existing g

The Cost of Language: Centroid Erasure Exposes and Exploits Modal Competition in Multimodal Language Models

Local AiDGX agent

arXiv:2604.14363v1 Announce Type: new Abstract: Multimodal language models systematically underperform on visual perception tasks, yet the structure underlying this failure remains poorly understood.

What a week! Here’s everything we shipped: — Gemini 3.1 Flash TTS, our latest text-to-speech model, featuring native multi-speaker dialogue …

Model ReleasesDGX agent

What a week! Here’s everything we shipped: — Gemini 3.1 Flash TTS, our latest text-to-speech model, featuring native multi-speaker dialogue and improved controllability and audio tags for more natural

16 Apr 2026

A closer look at how large language models trust humans: patterns and biases

ResearchDGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

A High-Resolution Landscape Dataset for Concept-Based XAI With Application to Species Distribution Models

SafetyDGX agent

arXiv:2604.13240v1 Announce Type: new Abstract: Mapping the spatial distribution of species is essential for conservation policy and invasive species management. Species distribution models (SDMs) are

A primer on 'interpretability' and how AI researchers are figuring out how to open and understand the 'black box' that holds the formulas within most AI models (Oliver Whang/New York Times)

TutorialsDGX agent

Oliver Whang / New York Times: A primer on “interpretability” and how AI researchers are figuring out how to open and understand the “black box” that holds the formulas within most AI models — When De

Action Images: End-to-End Policy Learning via Multiview Video Generation

SafetyDGX agent

arXiv:2604.06168v2 Announce Type: replace Abstract: World action models (WAMs) have emerged as a promising direction for robot policy learning, as they can leverage powerful video backbones to model t

Anthropic releases a new Opus model amid Mythos Preview buzz

Model ReleasesDGX agent

Anthropic has released its most powerful 'generally available' model to date: Claude Opus 4.7. The company called it a step up from Opus 4.6 for advanced software engineering tasks, particularly in co

Claude remains irreducibly Claude. If you know, you know. (The fact that models have distinct personalities that are consistent across gener…

Model ReleasesDGX agent

Claude remains irreducibly Claude. If you know, you know. (The fact that models have distinct personalities that are consistent across generations is technically interesting, it also makes it very eas

Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments

Local AiDGX agent

arXiv:2508.08791v3 Announce Type: replace Abstract: Effective tool use is essential for large language models (LLMs) to interact with their environment. However, progress is limited by the lack of eff

Interpretable Stylistic Variation in Human and LLM Writing Across Genres, Models, and Decoding Strategies

TutorialsDGX agent

arXiv:2604.14111v1 Announce Type: new Abstract: Large Language Models (LLMs) are now capable of generating highly fluent, human-like text. They enable many applications, but also raise concerns such a

Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more preci…

Model ReleasesDGX agent

Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more precisely, and verifies its own outputs before reporting back. Yo

LoRA-MME: Multi-Model Ensemble of LoRA-Tuned Encoders for Code Comment Classification

Model ReleasesDGX agent

arXiv:2603.03959v4 Announce Type: replace-cross Abstract: Code comment classification is a critical task for automated software documentation and analysis. In the context of the NLBSE'26 Tool Competit

Nested Fourier-enhanced neural operator for efficient modeling of radiation transfer in fires

Local AiDGX agent

arXiv:2604.13919v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) has become an essential tool for predicting fire behavior, yet maintaining both efficiency and accuracy remains cha

Our Life Sciences model series is available as a research preview starting today for qualified customers including @Amgen, @moderna_tx, the …

Model ReleasesDGX agent

Our Life Sciences model series is available as a research preview starting today for qualified customers including @Amgen, @moderna_tx, the @AllenInstitute, and @thermofisher Scientific through ChatGP

SPG: Sandwiched Policy Gradient for Masked Diffusion Language Models

SafetyDGX agent

arXiv:2510.09541v3 Announce Type: replace Abstract: Diffusion large language models (dLLMs) are emerging as an efficient alternative to autoregressive models due to their ability to decode multiple to

The $160,000 Model X Signature Edition is officially sold out. Reservations are now closed.

IndustryDGX agent

Tesla's Model X Signature Edition, priced at $160,000, has reached full capacity and is no longer available for purchase or reservation. This announcement indicates that the high-end variant of the Mo

We’re open-sourcing HY-World 2.0, a multimodal world model that generates, reconstructs, and simulates interactive *3D worlds* from text, im…

ApplicationsDGX agent

We’re open-sourcing HY-World 2.0, a multimodal world model that generates, reconstructs, and simulates interactive *3D worlds* from text, images, and videos. Outputs can be integrated into game engine

15 Apr 2026

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Model ReleasesDGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

Adaptive Domain Models: Bayesian Evolution, Warm Rotation, and Principled Training for Geometric and Neuromorphic AI

ResearchDGX agent

arXiv:2603.18104v3 Announce Type: replace Abstract: Prevailing AI training infrastructure assumes reverse-mode automatic differentiation over IEEE-754 arithmetic. The memory overhead of training relat

Agentic Control in Variational Language Models

Local AiDGX agent

arXiv:2604.12513v1 Announce Type: new Abstract: We study whether a variational language model can support a minimal and measurable form of agentic control grounded in its own internal evidence. Our mo

ALL-FEM: Agentic Large Language models Fine-tuned for Finite Element Methods

AgentsDGX agent

arXiv:2603.21011v2 Announce Type: replace-cross Abstract: Finite element (FE) analysis guides the design and verification of nearly all manufactured objects. It is at the core of computational enginee

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic

AgentsDGX agent

arXiv:2604.12857v1 Announce Type: new Abstract: Autonomous vehicles (AVs) are now operating on public roads, which makes their testing and validation more critical than ever. Simulation offers a safe

Beyond Perception Errors: Semantic Fixation in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.12119v1 Announce Type: new Abstract: Large vision-language models (VLMs) often rely on familiar semantic priors, but existing evaluations do not cleanly separate perception failures from ru

Brain-DiT: A Universal Multi-state fMRI Foundation Model with Metadata-Conditioned Pretraining

SafetyDGX agent

arXiv:2604.12683v1 Announce Type: new Abstract: Current fMRI foundation models primarily rely on a limited range of brain states and mismatched pretraining tasks, restricting their ability to learn ge

Disposition Distillation at Small Scale: A Three-Arc Negative Result

Model ReleasesDGX agent

arXiv:2604.11867v1 Announce Type: cross Abstract: We set out to train behavioral dispositions (self-verification, uncertainty acknowledgment, feedback integration) into small language models (0.6B to

Google DeepMind introduces Gemini Robotics-ER 1.6 robotic reasoning model, says it shows significant spatial and physical reasoning improvements over ER 1.5 (Google DeepMind)

Model ReleasesDGX agent

Google DeepMind: Google DeepMind introduces Gemini Robotics-ER 1.6 robotic reasoning model, says it shows significant spatial and physical reasoning improvements over ER 1.5 — For robots to be truly h

Google rolls out Gemini 3.1 Flash TTS, a text-to-speech model with support for over 70 languages and audio tags that give developers granular speech control (Matthias Bastian/The Decoder)

Model ReleasesDGX agent

Matthias Bastian / The Decoder: Google rolls out Gemini 3.1 Flash TTS, a text-to-speech model with support for over 70 languages and audio tags that give developers granular speech control — The compa

GRACE: A Dynamic Coreset Selection Framework for Large Language Model Optimization

ResearchDGX agent

arXiv:2604.11810v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language understanding and generation. However, their immense number

Heuristic Classification of Thoughts Prompting (HCoT): Integrating Expert System Heuristics for Structured Reasoning into Large Language Models

TutorialsDGX agent

arXiv:2604.12390v1 Announce Type: new Abstract: This paper addresses two limitations of large language models (LLMs) in solving complex problems: (1) their reasoning processes exhibit Bayesian-like st

HTDC: Hesitation-Triggered Differential Calibration for Mitigating Hallucination in Large Vision-Language Models

ResearchDGX agent

arXiv:2604.12115v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong multimodal performance, but still suffer from hallucinations caused by unstable visual grounding and

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

SafetyDGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

LLM as Attention-Informed NTM and Topic Modeling as long-input Generation: Interpretability and long-Context Capability

ResearchDGX agent

arXiv:2510.03174v2 Announce Type: replace-cross Abstract: Topic modeling aims to produce interpretable topic representations and topic--document correspondences from corpora, but classical neural topi

Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

Model ReleasesDGX agent

arXiv:2604.12374v1 Announce Type: cross Abstract: We describe the pre-training, post-training, and quantization of Nemotron 3 Super, a 120 billion (active 12 billion) parameter hybrid Mamba-Attention

RPRA: Predicting an LLM-Judge for Efficient but Performant Inference

TutorialsDGX agent

arXiv:2604.12634v1 Announce Type: new Abstract: Large language models (LLMs) face a fundamental trade-off between computational efficiency (e.g., number of parameters) and output quality, especially w

← Previous
1…120121122123124…1009
Next →