AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
10 Jun 2026

What Should a Skill Remember? Quality--Cost Trade-offs in Cost-Aware Skill Rewriting for Language Model Agents

SafetyDGX agent

arXiv:2606.09421v2 Announce Type: replace Abstract: Large language model agents increasingly rely on skills: reusable procedural documents encoding workflows, tool use, implementation patterns, valida

Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music

ResearchDGX agent

arXiv:2412.11449v2 Announce Type: replace-cross Abstract: We propose WHISPER-GPT: A generative large language model (LLM) for speech and music that allows us to work with continuous audio representati

WorldKernel: A World Model is the Coupling Kernel of Admissible Possible Worlds

ResearchDGX agent

arXiv:2606.10934v1 Announce Type: new Abstract: A common assumption holds that enough observational and interventional data, given to a strong enough predictor, suffices. We report a failure mode that

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
9 Jun 2026

A Theoretical Analysis of Memory and Overfitting Phenomena in Stochastic Interpolation Models

ResearchDGX agent

arXiv:2606.08554v1 Announce Type: new Abstract: This paper provides a theoretical account of memorization in stochastic interpolation models. By leveraging closed-form expressions for the optimal velo

Active Learning with Foundation Model Priors: Efficient Learning under Class Imbalance

ApplicationsDGX agent

arXiv:2606.07630v1 Announce Type: cross Abstract: Real-world datasets across image and text domains are often characterized by skewed class distributions and noisy annotations, which jointly degrade m

AMix-1: A Pathway to Test-Time Scalable Protein Foundation Model

SafetyDGX agent

arXiv:2507.08920v4 Announce Type: replace-cross Abstract: We introduce AMix-1, a powerful protein foundation model built on Bayesian Flow Networks and empowered by a systematic training methodology, e

AUCp: Pseudo-AUC for Inference Model Selection with Unlabeled Validation Data in Abnormality Detection

ResearchDGX agent

arXiv:2606.08742v1 Announce Type: new Abstract: Abnormality detection is a crucial yet challenging task in medical image analysis. Distinguishing abnormalities from normal data by learning to reconstr

Beyond Additivity: Causal Discovery in Location-Scale Noise Models with Hidden Variables

ResearchDGX agent

arXiv:2606.08196v1 Announce Type: cross Abstract: We study causal discovery from observational data when some variables are hidden and the data-generating process follows a location-scale noise model

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly …

HardwareDGX agent

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won't notice. We

C^3ache: Accelerating World Action Models with Cross Inference Chunk Cache

TutorialsDGX agent

arXiv:2606.08962v1 Announce Type: cross Abstract: World Action Models (WAMs) generalize better than standard Vision-Language-Action (VLA) policies to novel motions and environments, because a video-mo

Decoy-Calibrated Failure Audits for Language Models

ResearchDGX agent

arXiv:2606.09046v1 Announce Type: new Abstract: Useful audits reveal not only how often a model fails, but also where its failures concentrate. An auditor may test many candidate explanations: long in

@DeepCogito needed sub-500ms time to first token at 1,000+ requests per minute for their frontier reasoning models. Together AI delivered. H…

ToolsDGX agent

@DeepCogito needed sub-500ms time to first token at 1,000+ requests per minute for their frontier reasoning models. Together AI delivered. Hear from the Deep Cogito team on what it takes to build fron

DiffSight-Former: Modeling Structural Differences and Temporal Dynamics for Glaucoma Progression Prediction

ResearchDGX agent

arXiv:2606.09140v1 Announce Type: new Abstract: Glaucoma is a leading cause of irreversible blindness worldwide, and early detection from fundus images is critical for effective disease management. Wh

Discovering Functionally Selective Brain Regions with a Deep Topographic Multimodal Model

ResearchDGX agent

arXiv:2606.09770v1 Announce Type: cross Abstract: Nearby neurons in cortex share similar response profiles, producing systematic spatial organization across sensory and cognitive systems. Recent topog

Empowering Feed-Forward Reconstruction Models with Metric Scale via Satellite Images

Local AiDGX agent

arXiv:2606.08205v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models have recently shown strong generalization across diverse scenes, yet most of them recover geometry only up to an u

Evaluate Clinical ASR Models Faster with Agent Skills and NVIDIA Nemotron Speech

Model ReleasesDGX agent

This article presents a clinical automatic speech recognition workflow for generating pronunciation-aware synthetic audio, reviewing clinical terms, and evaluating recognition quality using NVIDIA age

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/lee…

HardwareDGX agent

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/leerob/status/2052059466821198061?s=20 building on @nvidia stac

Foundation Inference Models for Ordinary Differential Equations

ResearchDGX agent

arXiv:2602.08733v2 Announce Type: replace Abstract: Ordinary differential equations (ODEs) are central to scientific modelling, but inferring their vector fields from noisy trajectories remains challe

How Much MRI Preprocessing Is Enough? A Cost-Utility Study for Brain MRI Foundation Models

ResearchDGX agent

arXiv:2606.08164v1 Announce Type: new Abstract: MRI preprocessing defines the input distribution seen by brain MRI foundation models, yet it is usually treated as routine data cleaning rather than a m

Inference-Time Conformal Reasoning with Valid Factuality Control for Large Language Models

ResearchDGX agent

arXiv:2606.08831v1 Announce Type: new Abstract: Large language models (LLMs) increasingly perform multi-step reasoning, where intermediate claims form implicit directed acyclic graphs whose node corre

Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models

SafetyDGX agent

arXiv:2602.12996v2 Announce Type: replace-cross Abstract: Knowledge augmentation has significantly enhanced the performance of Large Language Models (LLMs) in knowledge-intensive tasks. However, exist

Large Language Models Should Learn Personalized Rather Than Aggregated Human Preferences

SafetyDGX agent

arXiv:2606.07629v1 Announce Type: cross Abstract: Current approaches to aligning large language models (LLMs) aggregate diverse human preferences into a single reward signal, effectively optimizing fo

Light-WAM: Efficient World Action Models with State-Fusion Action Decoding

SafetyDGX agent

arXiv:2606.08242v1 Announce Type: new Abstract: World Action Models (WAMs) extend robot policy learning by incorporating future prediction as an additional training objective, encouraging the policy t

Locally Adaptive Conformal Inference for Operator Models

Local AiDGX agent

arXiv:2507.20975v5 Announce Type: replace-cross Abstract: Operator models are regression algorithms between Banach spaces of functions. They have become an increasingly critical tool for spatiotempora

Mitigating Diffusion Model Hallucinations with Dynamic Guidance

ResearchDGX agent

arXiv:2510.05356v2 Announce Type: replace Abstract: Hallucinations in diffusion models are samples with structural inconsistencies that can emerge due to the excessive smoothing of the learned score f

Modeling the Diachronic Evolution of Legal Norms: An LRMoo-Based, Component-Level, Event-Centric Approach to Legal Knowledge Graphs

ApplicationsDGX agent

arXiv:2506.07853v5 Announce Type: replace Abstract: Representing the temporal evolution of legal norms is a critical challenge for automated processing. While foundational frameworks exist, they lack

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)

ResearchDGX agent

arXiv:2601.21522v2 Announce Type: replace-cross Abstract: The performance of large language models (LLMs) on verifiable tasks is usually measured by pass@k, the probability of answering a question cor

Structure-Aware Modeling of Multiple-Choice Questions Improves Automatic Difficulty Estimation

ResearchDGX agent

arXiv:2606.08988v1 Announce Type: cross Abstract: Automatic Question Difficulty Estimation (AQDE) holds growing promise for educational assessment because it has the potential to yield difficulty esti

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning

ResearchDGX agent

arXiv:2606.08231v1 Announce Type: new Abstract: Test-time Scaling (TTS) has emerged as a pivotal research direction for enhancing model performance by dynamically allocating computational resources du

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

Model ReleasesDGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

SafetyDGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

this model is the opposite of mythos. Its small, cost effective, apache 2.0, and locally deployable. This is the way LLMs should go. small, …

Local AiDGX agent

this model is the opposite of mythos. Its small, cost effective, apache 2.0, and locally deployable. This is the way LLMs should go. small, open source, transparent and sovereign vs large, expensive,

Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model

TutorialsDGX agent

arXiv:2603.25184v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for post-training large language models (LLMs) in reasoning tasks. While scaling rollouts can

Vision Language Model Helps Private Information De-Identification in Vision Data

TutorialsDGX agent

arXiv:2606.09132v1 Announce Type: new Abstract: Visual Language Models (VLMs) have gained significant popularity due to their remarkable ability. While various methods exist to enhance privacy in text

8 Jun 2026

Acoustic Cue Alignment in Audio Language Models for Speech Emotion Recognition

SafetyDGX agent

arXiv:2606.07309v1 Announce Type: cross Abstract: Instruction-following audio language models (ALMs) can be augmented with explicit acoustic cues, yet it remains unclear whether such cues are used in

Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows (Hartley Charlton/MacRumors)

AgentsDGX agent

Hartley Charlton / MacRumors: Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows — Apple today

Dreaming when Necessary: Advancing World Action Models with Adaptive Multi-Modal Reasoning

TutorialsDGX agent

arXiv:2606.07089v1 Announce Type: new Abstract: World Action Models (WAMs) offer a promising approach to embodied intelligence, yet existing methods rely heavily on video prediction as action priors a

Federated Foundation Models over Vehicular Networks

ApplicationsDGX agent

arXiv:2606.06786v1 Announce Type: new Abstract: This paper presents a forward-looking vision for integrating the emerging multi-modal multi-task federated foundation models (M3T FedFMs) into vehicular

Generative Models Erode Human Temporal Learning Through Market Selection

SafetyDGX agent

arXiv:2606.06572v1 Announce Type: cross Abstract: We argue that modern generative models create structural risks for knowledge and cultural production at current, sub-AGI capability levels. We define

LuMamba: Latent Unified Mamba for Electrode Topology-Invariant and Efficient EEG Modeling

HardwareDGX agent

arXiv:2603.19100v2 Announce Type: replace Abstract: Electroencephalography (EEG) enables non-invasive monitoring of brain activity across clinical and neurotechnology applications, yet building founda

Modeling AdaGrad, RMSProp, and Adam with Integro-Differential Equations

ResearchDGX agent

arXiv:2411.09734v3 Announce Type: replace Abstract: In this paper, we propose a continuous-time formulation for the AdaGrad, RMSProp, and Adam optimization algorithms by modeling them as first-order i

One Loss to Rule Them All: Marked Time-to-Event for Structured EHR Foundation Models

ResearchDGX agent

arXiv:2602.00541v2 Announce Type: replace Abstract: Clinical events captured in Electronic Health Records (EHR) are irregularly sampled and may consist of a mixture of discrete events and numerical me

RePo: Language Models with Context Re-Positioning

ResearchDGX agent

arXiv:2512.14391v3 Announce Type: replace-cross Abstract: In-context learning is fundamental to modern Large Language Models (LLMs); however, prevailing architectures impose a rigid and fixed contextu

The Fine-Tuning Trap: Evaluating Negative Transfer and the Role of PEFT in Sub-1B Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2606.06920v1 Announce Type: cross Abstract: Deploying Small Language Models (SLMs) on edge devices requires efficient fine-tuning strategies that adapt models to new tasks without degrading thei

The Geometry of Last-Layer Model Stealing

ResearchDGX agent

arXiv:2606.06854v1 Announce Type: new Abstract: This paper uses geometry to explain how a machine learning model can be stolen using an already existing well-known method. The author has shown the exa

The next quarter of effort at the frontier of applied AI is not about model capabilities. It's about 3 key things: 1. security, guardrails a…

ApplicationsDGX agent

The next quarter of effort at the frontier of applied AI is not about model capabilities. It's about 3 key things: 1. security, guardrails around data ingress/egress, and understanding risk in your sy

Towards Efficient and Exact Forgetting Services in Pre-Trained-Model-based Continual Learning

ResearchDGX agent

arXiv:2505.12239v2 Announce Type: replace-cross Abstract: In Continual Learning (CL), using a Pre-Trained Model (PTM) as the feature extractor has become a popular practice. Accompanied by analytic cl

TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning

Local AiDGX agent

arXiv:2602.18905v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, yet their decision-making processes remain diff

We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve …

IndustryDGX agent

We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve maximum possible performance on Apple hardware Starting from

7 Jun 2026

Snowflake, Databricks and the model makers: The battle for the agentic client and AI back end

AgentsDGX agent

Agentic artificial intelligence is being misread as a set of separate battles – for example, Snowflake Inc. versus Databricks Inc., copilots versus agents, model makers versus application vendors. We

This is a pretty striking shift toward Chinese models by American AI startups since the start of the year. https://substack.com/@profgmarket…

TutorialsDGX agent

American AI startups have increasingly shifted toward adopting Chinese AI models since the beginning of the year, representing a notable change in technology sourcing preferences. This trend likely re

6 Jun 2026

Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Frequency-Domain Insights

SafetyDGX agent

arXiv:2411.18343v3 Announce Type: replace-cross Abstract: Personalized Federal learning(PFL) allows clients to cooperatively train a personalized model without disclosing their private dataset. Howeve

Do you actually use different AI models for different parts of your life, or does one subscription eventually win?

IndustryDGX agent

This Reddit discussion explores user preferences and behaviors around AI model usage, examining whether people maintain subscriptions to multiple AI services or consolidate to a single primary platfor

Explainable AI-Driven Cyber Risk Analytics and Model Reliability Assessment for Intelligent Governance of U.S. Critical Infrastructure: An XGBoost and SHAP-Based Intrusion Detection Framework

ApplicationsDGX agent

arXiv:2606.05710v1 Announce Type: cross Abstract: The increasing penetrations of the critical infrastructure sector in the United States with intelligent digital technologies have greatly increased ex

TRACE: A Temporal Conditional Estimation for Multimodal Time Series Foundation Models

TutorialsDGX agent

arXiv:2606.06285v1 Announce Type: new Abstract: Time series foundation models (TS-FMs) aim to learn generalizable temporal representations that can be adapted to a wide range of downstream tasks. In r

5 Jun 2026

A Systematic Analysis of Biases in Large Language Models

SafetyDGX agent

arXiv:2512.15792v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making. However,

Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents

SafetyDGX agent

arXiv:2606.05753v1 Announce Type: new Abstract: Latent visual reasoning (LVR) inserts supervised latent tokens between perception and answer generation in vision-language models (VLMs). The field uses

Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models

ResearchDGX agent

arXiv:2601.18383v2 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) excel at solving complex problems by explicitly generating a reasoning trace before deriving the final answer. H

EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

Local AiDGX agent

arXiv:2606.06379v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have shown increasing potential for clinical image interpretation, including lesion detection and report generatio

FontFusion: Enhancing Generative Text in Diffusion Models with Typographic Conditioning

ResearchDGX agent

arXiv:2606.06066v1 Announce Type: new Abstract: Typography generation in diffusion models faces a persistent trade-off: enabling precise font control typically degrades text legibility, while maintain

← Previous
1…178179180181182…1010
Next →