AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,428 results
30 Apr 2026

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution

TutorialsDGX agent

arXiv:2509.23980v2 Announce Type: replace Abstract: Diffusion models have recently shown promising results for video super-resolution (VSR). However, directly adapting generative diffusion models to V

29 Apr 2026

Adaptive Meta-Learning Stochastic Gradient Hamiltonian Monte Carlo Simulation for Bayesian Updating of Structural Dynamic Models

ResearchDGX agent

arXiv:2604.25710v1 Announce Type: cross Abstract: In the last few decades, Markov chain Monte Carlo (MCMC) methods have been widely applied to Bayesian updating of structural dynamic models in the fie

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

arXiv:2408.16322v4 Announce Type: replace Abstract: Current research in semantic bird's-eye view segmentation for autonomous driving focuses solely on optimizing neural network models using a single d

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

Model ReleasesDGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

Generative diffusion models for spatiotemporal influenza forecasting

TutorialsDGX agent

arXiv:2604.24913v1 Announce Type: new Abstract: Forecasting infectious disease incidence can provide important information to guide public health planning, yet is difficult because epidemic dynamics a

Images Amplify Misinformation Sharing in Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.13302v2 Announce Type: replace Abstract: As language and vision-language models (VLMs) become central to information access and online interaction, concerns grow about their potential to am

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

Model ReleasesDGX agent

arXiv:2507.12553v3 Announce Type: replace Abstract: Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish

Large Language Models Explore by Latent Distilling

Model ReleasesDGX agent

arXiv:2604.24927v1 Announce Type: new Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-le

LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model

ApplicationsDGX agent

arXiv:2604.25297v1 Announce Type: new Abstract: In recent years, the rapid proliferation of open-source large language models (LLMs) has spurred efforts to turn general-purpose models into domain spec

Nautile-370M: Spectral Memory Meets Attention in a Small Reasoning Model

Model ReleasesDGX agent

arXiv:2604.24809v1 Announce Type: new Abstract: We present Nautile-370M, a 371-million-parameter small language model designed for efficient reasoning under strict parameter and inference budgets. Nau

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding

Model ReleasesDGX agent

arXiv:2604.25884v1 Announce Type: cross Abstract: Quantum computing calibration depends on interpreting experimental data, and calibration plots provide the most universal human-readable representatio

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models

Model ReleasesDGX agent

arXiv:2604.25591v1 Announce Type: cross Abstract: Recent audio-aware large language models (ALLMs) have demonstrated strong capabilities across diverse audio understanding and reasoning tasks, but the

28 Apr 2026

300,000 AI builders have already added their hardware to HF to instantly see what model they can run locally. To do so, go to https://huggin…

Local AiDGX agent

300,000 AI builders have already added their hardware to HF to instantly see what model they can run locally. To do so, go to https://huggingface.co/settings/local-apps and add your hardware specs. Yo

A satellite foundation model for improved wealth monitoring

Model ReleasesDGX agent

arXiv:2604.23166v1 Announce Type: cross Abstract: Poverty statistics guide social policy, but in many low- and middle-income countries, censuses and household surveys that collect these data are costl

A Systematic Approach for Large Language Models Debugging

AgentsDGX agent

arXiv:2604.23027v1 Announce Type: new Abstract: Large language models (LLMs) have become central to modern AI workflows, powering applications from open-ended text generation to complex agent-based re

An empirical evaluation of the risks of AI model updates using clinical data: stability, arbitrariness, and fairness

SafetyDGX agent

arXiv:2604.23954v1 Announce Type: new Abstract: Artificial Intelligence and Machine Learning (AI/ML) models used in clinical settings are increasingly deployed to support clinical decision-making. How

AP-BMM: Approximating Capability-Efficiency Pareto Sets of LLMs via Asynchronous Prior-guided Bayesian Model Merging

ResearchDGX agent

arXiv:2512.09972v5 Announce Type: replace-cross Abstract: Navigating the capability--efficiency trade-off in Large Language Models (LLMs) requires approximating a high-quality Pareto set. Existing mod

Automating Categorization of Scientific Texts with In-Context Learning and Prompt-Chaining in Large Language Models

ResearchDGX agent

arXiv:2604.23430v1 Announce Type: cross Abstract: The relentless expansion of scientific literature presents significant challenges for navigation and knowledge discovery. Within Research Information

AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models

Model ReleasesDGX agent

arXiv:2604.22871v1 Announce Type: cross Abstract: Automated red-teaming methods for large language models typically optimize attack prompts within a fixed, human-designed strategy, leaving the attack

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models

SafetyDGX agent

arXiv:2502.14888v4 Announce Type: replace-cross Abstract: The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, m

CoreGuard: Safeguarding Foundational Capabilities of LLMs Against Model Stealing in Edge Deployment

ResearchDGX agent

arXiv:2410.13903v3 Announce Type: replace-cross Abstract: Proprietary large language models (LLMs) exhibit strong generalization capabilities across diverse tasks and are increasingly deployed on edge

Credal Concept Bottleneck Models for Epistemic-Aleatoric Uncertainty Decomposition

ResearchDGX agent

arXiv:2604.24170v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) predict through human-interpretable concepts, but they typically output point concept probabilities that conflate epist

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models

Model ReleasesDGX agent

arXiv:2406.10185v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) are increasingly integral to healthcare applications, including medical visual question answering and imaging r

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

SafetyDGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

DRP: Distilled Reasoning Pruning with Skill-aware Step Decomposition for Efficient Large Reasoning Models

ResearchDGX agent

arXiv:2505.13975v4 Announce Type: replace Abstract: While Large Reasoning Models (LRMs) have demonstrated success in complex reasoning tasks through long chain-of-thought (CoT) reasoning, their infere

Evaluating Language Models' Evaluations of Games

SafetyDGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

Exploring the Secondary Risks of Large Language Models

Model ReleasesDGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

GA2-CLIP: Generic Attribute Anchor for Efficient Prompt Tuningin Video-Language Models

ResearchDGX agent

arXiv:2511.22125v2 Announce Type: replace Abstract: Visual and textual soft prompt tuning can effectively improve the adaptability of Vision-Language Models (VLMs) in downstream tasks. However, fine-t

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models

Model ReleasesDGX agent

arXiv:2509.26388v3 Announce Type: replace-cross Abstract: Conversational Spoken Language Models (SLMs) are emerging as a promising paradigm for real-time speech interaction. However, their capacity of

GeoEdit: Local Frames for Fast, Training-Free On-Manifold Editing in Diffusion Models

ResearchDGX agent

arXiv:2604.24238v1 Announce Type: new Abstract: Diffusion models are a leading paradigm for data generation, but training-free editing typically re-runs the full denoising trajectory for every edit st

Geometry Preserving Loss Functions Promote Improved Adaptation of Blackbox Generative Model

ResearchDGX agent

arXiv:2604.23888v1 Announce Type: cross Abstract: Adaptation of blackbox generative models has been widely studied recently through the exploration of several methods including generator fine-tuning,

GWT: Scalable Optimizer State Compression for Large Language Model Training

ResearchDGX agent

arXiv:2501.07237v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated exceptional capabilities across diverse natural language processing benchmarks. However, the es

Integrative neurocybernetic modeling in the era of large-scale neuroscience

ResearchDGX agent

arXiv:2604.23903v1 Announce Type: cross Abstract: Large-scale neuroscience is generating rich datasets across animals, brain areas and behavioral contexts, yet our modeling efforts remains fragmented

LongFlow: Efficient KV Cache Compression for Reasoning Models

Model ReleasesDGX agent

arXiv:2603.11504v2 Announce Type: replace-cross Abstract: Recent reasoning models such as OpenAI-o1 and DeepSeek-R1 have shown strong performance on complex tasks including mathematical reasoning and

NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni is a 30B hybrid mixture-of-experts model that brings multimodal perception and reasoning into a single system, natively supporting text, image, video, and audio inputs whil

Predicting one-year clinical instability and mortality in heart failure patients using sequence modeling

Model ReleasesDGX agent

arXiv:2511.16839v3 Announce Type: replace-cross Abstract: Heart failure (HF) discharge planning depends on identifying patients at risk of deterioration or death, yet accurate prediction from routinel

Predicting Wind Loads on Container Ships in Harbor Environments through Multi-Fidelity Modeling

Model ReleasesDGX agent

arXiv:2604.22882v1 Announce Type: new Abstract: Modern container ships face higher wind loads due to increased windage areas, making accurate predictions of wind loads essential for mooring design. Ex

RegMean++: Enhancing Effectiveness and Generalization of Regression Mean for Model Merging

ResearchDGX agent

arXiv:2508.03121v3 Announce Type: replace Abstract: Regression Mean (RegMean), an approach that formulates model merging as a linear regression problem, aims to find the optimal weights for each linea

SketchVLM: Vision language models can annotate images to explain thoughts and guide users

Model ReleasesDGX agent

arXiv:2604.22875v1 Announce Type: cross Abstract: When answering questions about images, humans naturally point, label, and draw to explain their reasoning. In contrast, modern vision-language models

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

Model ReleasesDGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

The Kerimov-Alekberli Model: An Information-Geometric Framework for Real-Time System Stability

Model ReleasesDGX agent

arXiv:2604.24083v1 Announce Type: new Abstract: This study introduces the Kerimov-Alekberli model, a novel information-geometric framework that redefines AI safety by formally linking non-equilibrium

What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models

ResearchDGX agent

arXiv:2510.03075v3 Announce Type: replace-cross Abstract: Compositional generalization, the ability to generate novel combinations of known concepts, is a key ingredient for visual generative models.

Z^2-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models

SafetyDGX agent

arXiv:2604.23536v1 Announce Type: new Abstract: Diffusion models have achieved unprecedented success in text-aligned generation, largely driven by Classifier-Free Guidance (CFG). However, standard CFG

27 Apr 2026

Announcing Talkie: a new, open-weight historical LLM! We trained and finetuned a 13B model on a newly-curated dataset of only pre-1930 data.…

ResearchDGX agent

Talkie is a new 13-billion parameter open-weight language model trained exclusively on pre-1930 historical data through a newly-curated dataset. The model was trained and fine-tuned to specialize in u

Build Strands Agents with SageMaker AI models and MLflow

AgentsDGX agent

In this post, we demonstrate how to build AI agents using Strands Agents SDK with models deployed on SageMaker AI endpoints. You will learn how to deploy foundation models from SageMaker JumpStart, in

Deep Learning for Model Calibration in Simulation of Itaconic Acid Production

Model ReleasesDGX agent

arXiv:2604.22496v1 Announce Type: new Abstract: In this study, deep learning is used to estimate kinetic parameters for modeling itaconic acid production based on real batch experiments conducted at d

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model

SafetyDGX agent

arXiv:2604.22152v1 Announce Type: new Abstract: Evaluating robotics policies across thousands of environments and thousands of tasks is infeasible with existing approaches. This motivates the need for

FeatEHR-LLM: Leveraging Large Language Models for Feature Engineering in Electronic Health Records

ApplicationsDGX agent

arXiv:2604.22534v1 Announce Type: cross Abstract: Feature engineering for Electronic Health Records (EHR) is complicated by irregular observation intervals, variable measurement frequencies, and struc

Model Predictive Control of Hybrid Dynamical Systems

ResearchDGX agent

arXiv:2604.21989v1 Announce Type: cross Abstract: The problem of controlling hybrid dynamical systems using model predictive control (MPC) is formulated and sufficient conditions for asymptotic stabil

The Download: DeepSeek’s latest AI breakthrough, and the race to build world models

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Three reasons why DeepSeek’s new model matters On Friday, Chin

V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models

ResearchDGX agent

arXiv:2504.06148v3 Announce Type: replace Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in visual-text processing. However, existi

26 Apr 2026

🚨GUYS QWEN3.6 27B/35B @unslothai MLX QUANTS @Brooooook_lyn cooked 🔥 Apple Silicon users, the exact models y’all have been asking for LANDE…

Model ReleasesDGX agent

🚨GUYS QWEN3.6 27B/35B @unslothai MLX QUANTS @Brooooook_lyn cooked 🔥 Apple Silicon users, the exact models y’all have been asking for LANDED. Unsloth mixed-precision quants in native MLX format. >Fast.

Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from every frontier lab → Free mod…

AgentsDGX agent

Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from every frontier lab → Free models and discounts exclusive to the Portal → Bundled tools: w

24 Apr 2026

AFRILANGTUTOR: Advancing Language Tutoring and Culture Education in Low-Resource Languages with Large Language Models

Model ReleasesDGX agent

arXiv:2604.20996v1 Announce Type: new Abstract: How can language learning systems be developed for languages that lack sufficient training resources? This challenge is increasingly faced by developers

BackPlay: Head-Only Look-Back Self-Correction for Diffusion Language Models

ResearchDGX agent

arXiv:2601.06428v3 Announce Type: replace Abstract: Diffusion Language Models (DLMs) decode multiple tokens in parallel, but aggressive multi-token decoding amplifies cross-token dependency errors and

DeepSeek open-sources V4 large language model series

Model ReleasesDGX agent

Chinese artificial intelligence developer DeepSeek today released a new series of open-source large language models. V4, as the algorithm family is called, comprises two LLMs on launch. There’s the fl

DepthMaster: Taming Diffusion Models for Monocular Depth Estimation

SafetyDGX agent

arXiv:2501.02576v2 Announce Type: replace Abstract: Monocular depth estimation within the diffusion-denoising paradigm demonstrates impressive generalization ability but suffers from low inference spe

Fine-Grained Perspectives: Modeling Explanations with Annotator-Specific Rationales

SafetyDGX agent

arXiv:2604.21667v1 Announce Type: cross Abstract: Beyond exploring disaggregated labels for modeling perspectives, annotator rationales provide fine-grained signals of individual perspectives. In this

Generalization Properties of Score-matching Diffusion Models for Intrinsically Low-dimensional Data

ResearchDGX agent

arXiv:2603.03700v2 Announce Type: replace-cross Abstract: Despite the remarkable empirical success of score-based diffusion models, their statistical guarantees remain underdeveloped. Existing analyse

Quotient-Space Diffusion Models

SafetyDGX agent

arXiv:2604.21809v1 Announce Type: cross Abstract: Diffusion-based generative models have reformed generative AI, and have enabled new capabilities in the science domain, for example, generating 3D str

← Previous
1…7172737475…1008
Next →