AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
24 Apr 2026

Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control

SafetyDGX agent

arXiv:2604.20867v1 Announce Type: cross Abstract: Recent events surrounding the relationship between frontier AI suppliers and national-security customers have made a structural problem newly visible:

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

Projected Gradient Unlearning for Text-to-Image Diffusion Models: Defending Against Concept Revival Attacks

Research
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.21041v1 Announce Type: new Abstract: Machine unlearning for text-to-image diffusion models aims to selectively remove undesirable concepts from pre-trained models without costly retraining.

Schoenfeld's Anatomy of Mathematical Reasoning by Language Models

ResearchDGX agent

arXiv:2512.19995v2 Announce Type: replace-cross Abstract: Large language models increasingly expose reasoning traces, yet their underlying cognitive structure and steps remain difficult to identify an

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

StyleVAR: Controllable Image Style Transfer via Visual Autoregressive Modeling

SafetyDGX agent

arXiv:2604.21052v1 Announce Type: cross Abstract: We build on the Visual Autoregressive Modeling (VAR) framework and formulate style transfer as conditional discrete sequence modeling in a learned lat

Thinking Like a Botanist: Challenging Multimodal Language Models with Intent-Driven Chain-of-Inquiry

Model ReleasesDGX agent

arXiv:2604.20983v1 Announce Type: cross Abstract: Vision evaluations are typically done through multi-step processes. In most contemporary fields, experts analyze images using structured, evidence-bas

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21860v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This pape

Value-Conflict Diagnostics Reveal Widespread Alignment Faking in Language Models

SafetyDGX agent

arXiv:2604.20995v1 Announce Type: new Abstract: Alignment faking, where a model behaves aligned with developer policy when monitored but reverts to its own preferences when unobserved, is a concerning

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model

Model ReleasesDGX agent

arXiv:2604.21223v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated remarkable capabilities across various tasks. However, their ability to generate human-like text has ra

23 Apr 2026

Also, a ton of new Codex features coming soon! Fun little bundle w/the new model.

Model ReleasesDGX agent

Sam Altman announced upcoming new features for Codex, OpenAI's code generation model, bundled with a new model release. The post suggests these features represent a significant expansion of Codex capa

AROMA: Augmented Reasoning Over a Multimodal Architecture for Virtual Cell Genetic Perturbation Modeling

ResearchDGX agent

arXiv:2604.20263v1 Announce Type: cross Abstract: Virtual cell modeling predicts molecular state changes under genetic perturbations in silico, which is essential for biological mechanism studies. How

Assessing the Robustness of Climate Foundation Models under No-Analog Distribution Shifts

Model ReleasesDGX agent

arXiv:2603.23043v2 Announce Type: replace-cross Abstract: The accelerating pace of climate change introduces profound non-stationarities that challenge the ability of Machine Learning based climate em

Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations

ResearchDGX agent

arXiv:2509.25844v3 Announce Type: replace Abstract: When people query Vision-Language Models (VLMs) but cannot see the accompanying visual context (e.g. for blind and low-vision users), augmenting VLM

Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models

Model ReleasesDGX agent

arXiv:2508.17761v3 Announce Type: replace Abstract: In safety-critical applications data-driven models must not only be accurate but also provide reliable uncertainty estimates. This property, commonl

LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model

ResearchDGX agent

arXiv:2604.20796v1 Announce Type: new Abstract: We present LLaDA2.0-Uni, a unified discrete diffusion large language model (dLLM) that supports multimodal understanding and generation within a nativel

Meta Additive Model: Interpretable Sparse Learning With Auto Weighting

ApplicationsDGX agent

arXiv:2604.20111v1 Announce Type: cross Abstract: Sparse additive models have attracted much attention in high-dimensional data analysis due to their flexible representation and strong interpretabilit

Multi-Timescale Model Predictive Control for Slow-Fast Systems

ResearchDGX agent

arXiv:2511.14311v2 Announce Type: replace-cross Abstract: Model Predictive Control (MPC) has established itself as the primary methodology for constrained control, enabling autonomy across diverse app

OpenAI says its new GPT-5.5 model is more efficient and better at coding

Model ReleasesDGX agent

OpenAI just announced its new GPT-5.5 model, which the company calls its 'smartest and most intuitive to use model yet, and the next step toward a new way of getting work done on a computer.' OpenAI j

ParetoSlider: Diffusion Models Post-Training for Continuous Reward Control

ResearchDGX agent

arXiv:2604.20816v1 Announce Type: cross Abstract: Reinforcement Learning (RL) post-training has become the standard for aligning generative models with human preferences, yet most methods rely on a si

PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance

Model ReleasesDGX agent

arXiv:2604.20834v1 Announce Type: new Abstract: Recent advances in Vision-Language-Action (VLA) models have opened new avenues for robot manipulation, yet existing methods exhibit limited efficiency a

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

Temporally Extended Mixture-of-Experts Models

HardwareDGX agent

arXiv:2604.20156v1 Announce Type: new Abstract: Mixture-of-Experts models, now popular for scaling capacity at fixed inference speed, switch experts at nearly every token. Once a model outgrows availa

The Imperfective Paradox in Large Language Models

SafetyDGX agent

arXiv:2601.09373v2 Announce Type: replace Abstract: Do Large Language Models (LLMs) genuinely grasp the compositional semantics of events, or do they rely on surface-level probabilistic heuristics? We

Trajectory-Aware Reliability Modeling of Democratic Systems

ResearchDGX agent

arXiv:2604.20127v1 Announce Type: new Abstract: Failures in complex systems often emerge through gradual degradation and the propagation of stress across interacting components rather than through iso

Understanding Overparametrization in Survival Models through Interpolation

SafetyDGX agent

arXiv:2512.12463v3 Announce Type: replace-cross Abstract: Classical statistical learning theory predicts a U-shaped relationship between test loss and model capacity, driven by the bias-variance trade

We're the top open-weights model on Design Arena!

Model ReleasesDGX agent

We're the top open-weights model on Design Arena! BREAKING: Kimi K2.6 takes 1st overall of open weights models on Design Arena! Kimi K2.6 is in the same performance band as Claude Opus 4.7 - while est

22 Apr 2026

Curiosity-Critic: Cumulative Prediction Error Improvement as a Tractable Intrinsic Reward for World Model Training

Local AiDGX agent

arXiv:2604.18701v1 Announce Type: cross Abstract: Local prediction-error-based curiosity rewards focus on the current transition without considering the world model's cumulative prediction error acros

Efficient Autoregressive Inference for Transformer Probabilistic Models

ResearchDGX agent

arXiv:2510.09477v2 Announce Type: replace-cross Abstract: Set-based transformer models for amortized probabilistic inference and meta-learning, such as neural processes, prior-fitted networks, and tab

I've been mostly prompting the new model directly via the client.images.generate() API, I have no idea if that rewrites my prompts at all or…

Model ReleasesDGX agent

I've been mostly prompting the new model directly via the client.images.generate() API, I have no idea if that rewrites my prompts at all or if it passes them straight to the model - details here http

LatticeVision: Image to Image Networks for Modeling Non-Stationary Spatial Data

Model ReleasesDGX agent

arXiv:2505.09803v3 Announce Type: replace-cross Abstract: In many applications, we wish to fit a parametric statistical model to a small ensemble of spatially distributed random variables ('fields').

Learn more about this model release https://x.com/Alibaba_Qwen/status/2046939764428009914?s=20

AgentsDGX agent

Learn more about this model release https://x.com/Alibaba_Qwen/status/2046939764428009914?s=20 🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, a

Mask World Model: Predicting What Matters for Robust Robot Policy Learning

SafetyDGX agent

arXiv:2604.19683v1 Announce Type: new Abstract: World models derived from large-scale video generative pre-training have emerged as a promising paradigm for generalist robot policy learning. However,

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

SafetyDGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

OpenAI dropped a new model on HF today!

Model ReleasesDGX agent

OpenAI released a new model that was made available on Hugging Face, as announced by Hugging Face CEO Clem Delangue on X (formerly Twitter). The specific details about which model was released are not

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.15435v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) exhibit strong multimodal capabilities but remain vulnerable to hallucinations from intrinsic errors and

PC2Model: ISPRS benchmark on 3D point cloud to model registration

Model ReleasesDGX agent

arXiv:2604.19596v1 Announce Type: new Abstract: Point cloud registration involves aligning one point cloud with another or with a three-dimensional (3D) model, enabling the integration of multimodal d

RDP LoRA: Geometry-Driven Identification for Parameter-Efficient Adaptation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.19321v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) remains structurally uncertain despite parameter-efficient methods such as Low-Rank Adaptation (LoRA), as the

21 Apr 2026

ATLAS: Constitution-Conditioned Latent Geometry and Redistribution Across Language Models and Neural Perturbation Data

Model ReleasesDGX agent

arXiv:2604.17663v1 Announce Type: cross Abstract: Constitution-conditioned post-training can be analysed as a structured perturbation of a model's learned representational geometry. We introduce ATLAS

BengaliMoralBench: A Benchmark for Auditing Moral Reasoning in Large Language Models within Bengali Language and Culture

Model ReleasesDGX agent

arXiv:2511.03180v2 Announce Type: replace Abstract: As multilingual Large Language Models (LLMs) gain traction across South Asia, their alignment with local ethical norms, particularly for Bengali, sp

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

Model ReleasesDGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

Class-specific diffusion models improve military object detection in a low-data domain

ResearchDGX agent

arXiv:2604.18076v1 Announce Type: new Abstract: Diffusion-based image synthesis has emerged as a promising source of synthetic training data for AI-based object detection and classification. In this w

Deep Learning-Enhanced Calibration of the Heston Model: A Unified Framework

Model ReleasesDGX agent

arXiv:2510.24074v2 Announce Type: replace-cross Abstract: The Heston stochastic volatility model is a widely used tool in financial mathematics for pricing European options. However, its calibration r

Detecting LLM-Generated Spam Reviews by Integrating Language Model Embeddings and Graph Neural Network

Model ReleasesDGX agent

arXiv:2510.01801v2 Announce Type: replace Abstract: The rise of large language models (LLMs) has enabled the generation of highly persuasive spam reviews that closely mimic human writing. These review

DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models

SafetyDGX agent

arXiv:2604.16979v1 Announce Type: cross Abstract: High-quality and diverse multimodal data are essential for improving vision-language models (VLMs), yet existing datasets often contain noisy, redunda

DynaWeb: Model-Based Reinforcement Learning of Web Agents

SafetyDGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

Efficient Inference for Coupled Hidden Markov Models in Continuous Time and Discrete Space

ResearchDGX agent

arXiv:2510.12916v2 Announce Type: replace-cross Abstract: Systems of interacting continuous-time Markov chains are a powerful model class, but inference is typically intractable in high dimensional se

FM-CAC: Carbon-Aware Control for Battery-Buffered Edge AI via Time-Series Foundation Models

Local AiDGX agent

arXiv:2604.16448v1 Announce Type: cross Abstract: As edge AI deployments scale to billions of devices running always-on, real-time compound AI pipelines, they represent a massive and largely unmanaged

Forecast Sports Outcomes under Efficient Market Hypothesis: Theoretical and Experimental Analysis of Odds-Only and Generalised Linear Models

Model ReleasesDGX agent

arXiv:2604.17194v1 Announce Type: cross Abstract: Converting betting odds into accurate outcome probabilities is a fundamental challenge in order to use betting odds as a benchmark for sports forecast

HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders

Model ReleasesDGX agent

arXiv:2604.16430v1 Announce Type: new Abstract: Large Language Models (LLMs) are powerful and widely adopted, but their practical impact is limited by the well-known hallucination phenomenon. While re

Identifying Ethical Biases in Action Recognition Models

SafetyDGX agent

arXiv:2604.17971v1 Announce Type: new Abstract: Human Action Recognition (HAR) models are increasingly deployed in high-stakes environments, yet their fairness across different human appearances has n

Inertia in Moral and Value Judgments of Large Language Models

SafetyDGX agent

arXiv:2408.09049v3 Announce Type: replace Abstract: Large Language Models (LLMs) behave non-deterministically, and prompting has become a common method for steering their outputs. A popular strategy i

Kimi K2.6 autonomously overhauled exchange-core, an 8-year-old open-source financial matching engine. Over a 13-hour execution, the model it…

Model ReleasesDGX agent

Kimi K2.6 autonomously overhauled exchange-core, an 8-year-old open-source financial matching engine. Over a 13-hour execution, the model iterated through 12 optimization strategies, initiating over 1

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible th…

Model ReleasesDGX agent

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible through CLI, skills and md files) is perfect for running SOTA

Measuring Representation Robustness in Large Language Models for Geometry

Model ReleasesDGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models

ResearchDGX agent

arXiv:2511.14582v2 Announce Type: replace Abstract: Omnimodal large language models (OmniLLMs) have attracted increasing research attention of late towards unified audio-video understanding. However,

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

SafetyDGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

Parallel Test-Time Scaling for Latent Reasoning Models

Model ReleasesDGX agent

arXiv:2510.07745v4 Announce Type: replace Abstract: Parallel test-time scaling (TTS) is a pivotal approach for enhancing large language models (LLMs), typically by sampling multiple token-based chains

Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not

TutorialsDGX agent

arXiv:2604.04825v2 Announce Type: replace Abstract: Large language models achieve strong performance on many language tasks, yet it remains unclear whether they integrate world knowledge with syntacti

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

SafetyDGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

← Previous
1…5556575859…999
Next →