AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,428 results
24 Apr 2026

SQLyzr: A Comprehensive Benchmark and Evaluation Platform for Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.21214v1 Announce Type: cross Abstract: Text-to-SQL models have significantly improved with the adoption of Large Language Models (LLMs), leading to their increasing use in real-world applic

Structured Visual Narratives Undermine Safety Alignment in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2603.21697v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) extend text-only LLMs with visual reasoning, but also introduce new safety failure modes under visual

Unlocking Multi-Spectral Data for Multi-Modal Models with Guided Inputs and Chain-of-Thought Reasoning

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.21032v1 Announce Type: new Abstract: Multi-spectral imagery is a valuable input signal for Remote Sensing applications, such as land-use and land-cover classification and environmental moni

23 Apr 2026

1. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a bi…

Model ReleasesDGX agent

1. We believe in iterative deployment; although GPT-5.5 is already a smart model, we expect rapid improvements. Iterative deployment is a big part of our safety strategy; we believe the world will be

Are there any good story writer models that I can ruj with a 5080 16gb?

Local AiDGX agent

This Reddit post from r/ollama asks about story-writing language models that can run on a 5080 GPU with 16GB of VRAM . The discussion likely covers recommended open-source or quantized models suitable

Environmental Understanding Vision-Language Model for Embodied Agent

SafetyDGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

Fairness Testing of Large Language Models in Role-Playing

Model ReleasesDGX agent

arXiv:2411.00585v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become foundational in modern language-driven software applications, profoundly influencing daily life. A cr

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

SafetyDGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models

ResearchDGX agent

arXiv:2604.19786v1 Announce Type: new Abstract: Evaluating humor in large language models (LLMs) is an open challenge because existing approaches yield isolated, incomparable metrics rather than unifi

Intersectional Fairness in Large Language Models

Model ReleasesDGX agent

arXiv:2604.20677v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive settings, raising concerns about fairness and biases, particularly across i

Language Models Learn Universal Representations of Numbers and Here's Why You Should Care

TutorialsDGX agent

arXiv:2510.26285v2 Announce Type: replace-cross Abstract: Prior work has shown that large language models (LLMs) often converge to accurate input embedding for numbers, based on sinusoidal representat

On the Quantization Robustness of Diffusion Language Models in Coding Benchmarks

ResearchDGX agent

arXiv:2604.20079v1 Announce Type: cross Abstract: Auto-regressive Large Language Models (LLMs) achieve strong performance on coding tasks, but incur high memory and inference costs. Diffusion-based la

Online Survival Analysis: A Bandit Approach under Cox PH Model

ResearchDGX agent

arXiv:2604.20296v1 Announce Type: cross Abstract: Survival analysis is a widely used statistical framework for modeling time-to-event data under censoring. Classical methods, such as the Cox proportio

Retrofitting Small Multilingual Models for Retrieval: Matching 7B Performance with 300M Parameters

ResearchDGX agent

arXiv:2510.14274v2 Announce Type: replace Abstract: Training effective multilingual embedding models presents unique challenges due to the diversity of languages and task objectives. Although small mu

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

Model ReleasesDGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for c…

Model ReleasesDGX agent

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for coding, and so I will try again with the fixes. I appreciate

X-Cache: Cross-Chunk Block Caching for Few-Step Autoregressive World Models Inference

AgentsDGX agent

arXiv:2604.20289v1 Announce Type: new Abstract: Real-time world simulation is becoming a key infrastructure for scalable evaluation and online reinforcement learning of autonomous driving systems. Rec

22 Apr 2026

Anthropic investigates unauthorized access to restricted Claude Mythos AI model

Model ReleasesDGX agent

Anthropic PBC is investigating a report that unauthorized users accessed Claude Mythos, the next-level artificial intelligence model the company says is powerful enough to enable dangerous cyberattack

Conditional Diffusion Modeling with Attention for Probabilistic Battery Capacity Prediction under Real-World Condition

ApplicationsDGX agent

arXiv:2510.17414v2 Announce Type: replace Abstract: Accurate prediction of lithium-ion battery capacity and its associated uncertainty is essential for reliable battery management but remains challeng

Do Emotions Influence Moral Judgment in Large Language Models?

SafetyDGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

Ground-Level Near Real-Time Modeling for PM2.5 Pollution Prediction

SafetyDGX agent

arXiv:2604.18973v1 Announce Type: cross Abstract: Air pollution is a worldwide public health threat that can cause or exacerbate many illnesses, including respiratory disease, cardiovascular disease,

How Out-of-Equilibrium Phase Transitions can Seed Pattern Formation in Trained Diffusion Models

Local AiDGX agent

arXiv:2603.20092v3 Announce Type: replace Abstract: Diffusion models generate structure by progressively transforming noise into data, yet the mechanisms underlying this transition remain poorly under

Lingua-SafetyBench: A Benchmark for Safety Evaluation of Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.22737v2 Announce Type: replace Abstract: The robust safety of Vision-Language Large Models (VLLMs) against joint multilingual and multimodal threats remains severely underexplored. Current

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

SafetyDGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

ml-intern by @huggingface is wild 🔥 You drop a high-level prompt (“build the best scientific reasoning model” or “crush healthcare benchmar…

Model ReleasesDGX agent

ml-intern by @huggingface is wild 🔥 You drop a high-level prompt (“build the best scientific reasoning model” or “crush healthcare benchmarks”) and this open-source agent does the entire post-training

On the Generalizability of Foundation Models for Crop Type Mapping

SafetyDGX agent

arXiv:2409.09451v5 Announce Type: replace Abstract: Foundation models pre-trained using self-supervised learning have shown powerful transfer learning capabilities on various downstream tasks, includi

One Step Forward and K Steps Back: Better Reasoning with Denoising Recursion Models

Model ReleasesDGX agent

arXiv:2604.18839v1 Announce Type: cross Abstract: Looped transformers scale computational depth without increasing parameter count by repeatedly applying a shared transformer block and can be used for

OpenAI is shutting down text-embedding-3-small?!? I strongly believe that if you shut down a closed-source embedding model that you should o…

AgentsDGX agent

OpenAI is shutting down text-embedding-3-small?!? I strongly believe that if you shut down a closed-source embedding model that you should open-source. Imaging the trillions of tokens that will no lon

OpenAI just open sourced a new 1.5B (50m active) model on HuggingFace with Apache 2.0 license! It's not a new LLM, this one is called Privac…

IndustryDGX agent

OpenAI just open sourced a new 1.5B (50m active) model on HuggingFace with Apache 2.0 license! It's not a new LLM, this one is called Privacy Filter, and it's a PII detection model (checking if text h

PREF-XAI: Preference-Based Personalized Rule Explanations of Black-Box Machine Learning Models

ApplicationsDGX agent

arXiv:2604.19684v1 Announce Type: new Abstract: Explainable artificial intelligence (XAI) has predominantly focused on generating model-centric explanations that approximate the behavior of black-box

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation

Model ReleasesDGX agent

arXiv:2604.19092v1 Announce Type: cross Abstract: Recent advances in large-scale video world models have enabled increasingly realistic future prediction, raising the prospect of leveraging imagined v

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.19638v1 Announce Type: new Abstract: Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address saf

Self-Improving Tabular Language Models via Iterative Group Alignment

Local AiDGX agent

arXiv:2604.18966v1 Announce Type: cross Abstract: While language models have been adapted for tabular data generation, two fundamental limitations remain: (1) static fine-tuning produces models that c

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

SafetyDGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citat…

Model ReleasesDGX agent

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citation quality, instruction following, and efficiency. With Qwe

21 Apr 2026

A Benchmark Study of Segmentation Models and Adaptation Strategies for Landslide Detection from Satellite Imagery

Model ReleasesDGX agent

arXiv:2604.16663v1 Announce Type: new Abstract: Landslide detection from high resolution satellite imagery is a critical task for disaster response and risk assessment, yet the relative effectiveness

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

ApplicationsDGX agent

arXiv:2604.18570v1 Announce Type: cross Abstract: Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clin

A Rapid Deployment Pipeline for Autonomous Humanoid Grasping Based on Foundation Models

AgentsDGX agent

arXiv:2604.17258v1 Announce Type: new Abstract: Deploying a humanoid robot to manipulate a new object has traditionally required one to two days of effort: data collection, manual annotation, 3D model

Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety

Model ReleasesDGX agent

arXiv:2604.18487v1 Announce Type: new Abstract: The Adversarial Humanities Benchmark (AHB) evaluates whether model safety refusals survive a shift away from familiar harmful prompt forms. Starting fro

[AINews] Moonshot Kimi K2.6: the world's leading Open Model refreshes to catch up to Opus 4.6 (ahead of DeepSeek v4?)

Model ReleasesDGX agent

Moonshot's Kimi K2.6 represents a significant update to their open-source language model, positioning it to compete with Anthropic's Claude Opus 4.6 and potentially ahead of DeepSeek v4. The refresh a

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

SafetyDGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

Appearance-free Action Recognition: Zero-shot Generalization in Humans and a Two-Pathway Model

ApplicationsDGX agent

arXiv:2604.16675v1 Announce Type: new Abstract: Action recognition is a fundamental ability for social species. Yet, its underlying computations are not well understood. Classical psychophysical studi

Applications of deep generative models to DNA reaction kinetics and to cryogenic electron microscopy

ResearchDGX agent

arXiv:2604.16851v1 Announce Type: cross Abstract: This dissertation explores how deep generative models can advance the analysis of challenging biological problems by integrating domain knowledge with

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

SafetyDGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

Auto-encoder model for faster generation of effective one-body gravitational waveform approximations

Model ReleasesDGX agent

arXiv:2511.12642v2 Announce Type: replace-cross Abstract: Upgrades to current gravitational wave detectors for the next observation run and upcoming third-generation observatories, like the Einstein t

Beyond the Failures: Rethinking Foundation Models in Pathology

ResearchDGX agent

arXiv:2510.23807v5 Announce Type: replace-cross Abstract: Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy c

CaTS-Bench: Can Language Models Describe Time Series?

Model ReleasesDGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

Clarifai says it has deleted 3M OkCupid user photos and facial-recognition models trained on them after the US FTC settled with OkCupid over privacy violations (Jody Godoy/Reuters)

ApplicationsDGX agent

Jody Godoy / Reuters: Clarifai says it has deleted 3M OkCupid user photos and facial-recognition models trained on them after the US FTC settled with OkCupid over privacy violations — Artificial intel

DifFoundMAD: Foundation Models meet Differential Morphing Attack Detection

Model ReleasesDGX agent

arXiv:2604.17961v1 Announce Type: new Abstract: In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models

Dissipative Latent Residual Physics-Informed Neural Networks for Modeling and Identification of Electromechanical Systems

TutorialsDGX agent

arXiv:2604.18277v1 Announce Type: new Abstract: Accurate dynamical modeling is essential for simulation and control of embodied systems, yet first-principles models of electromechanical systems often

Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style

ResearchDGX agent

arXiv:2603.11024v2 Announce Type: replace Abstract: VLMs have become increasingly proficient at a range of computer vision tasks, such as visual question answering and object detection. This includes

Dual-End Consistency Model

ResearchDGX agent

arXiv:2602.10764v2 Announce Type: replace Abstract: The slow iterative sampling nature remains a major bottleneck for the practical deployment of diffusion and flow-based generative models. While cons

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

Model ReleasesDGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2604.18167v1 Announce Type: new Abstract: Modern text-to-image (T2I) models amplify harmful societal biases, challenging their ethical deployment. We introduce an inference-time method that reli

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

ResearchDGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

ResearchDGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models

ApplicationsDGX agent

arXiv:2510.02370v3 Announce Type: replace Abstract: Large language models leverage both parametric knowledge acquired during pretraining and in-context knowledge provided at inference time. Crucially,

Human Cognition in Machines: A Unified Perspective of World Models

AgentsDGX agent

arXiv:2604.16592v1 Announce Type: cross Abstract: This comprehensive report distinguishes prior works by the cognitive functions they innovate. Many works claim an almost 'human-like' cognitive capabi

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's W…

Model ReleasesDGX agent

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's Waldo style image but it's where is the raccoon holding a ham

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For exampl…

Model ReleasesDGX agent

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For example, a small amount of use will show that Kimi is not as good

← Previous
1…7273747576…1008
Next →