AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,860 results
Model Releases

Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models

DGX agent

arXiv:2606.08571v1 Announce Type: cross Abstract: Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to quest

model-releasesarxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Component Ablation for Efficient Hybrid Language Model Architectures: Performance, Resilience, and Compression Implications

DGX agent

arXiv:2603.22473v2 Announce Type: replace-cross Abstract: Hybrid language models combine softmax attention with linear-time sequence mechanisms such as state-space or linear-attention layers, but the

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

DisCo: World Models with Discrete Camera Motion Control

DGX agent

arXiv:2606.07967v1 Announce Type: new Abstract: Controllable video world models target interactive world exploration, where models must faithfully execute explicit action commands while preserving vis

model-releasesarxiv-cs-cv
9 Jun 2026
Tutorials

From inverse problems to neural operators: prediction, mechanism, and generalization of data-driven models

DGX agent

arXiv:2606.08956v1 Announce Type: new Abstract: Scientists have historically relied on mathematical models based on differential equations to relate system inputs -- forces, fluxes, or heat sources --

tutorialsarxiv-cs-lg
9 Jun 2026
Model Releases

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

DGX agent

arXiv:2606.08051v1 Announce Type: new Abstract: Financial transaction processing requires extracting structured merchant information from noisy, abbreviated bank transaction strings at scale. Our curr

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

IDDM: Identity-Decoupled Personalized Diffusion Models with a Tunable Privacy-Utility Trade-off

DGX agent

arXiv:2604.00903v2 Announce Type: replace Abstract: Personalized text-to-image diffusion models (e.g., DreamBooth, LoRA) enable users to synthesize high-fidelity avatars from a few reference photos fo

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

I'm really not interested in testing a new closed-weight model in the expensive tier. What's the point? I know what will happen: 1. They pre…

DGX agent

I'm really not interested in testing a new closed-weight model in the expensive tier. What's the point? I know what will happen: 1. They present the model. 2. The model beats the competition on benchm

safetygary-marcus--x
9 Jun 2026
Model Releases

Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?

DGX agent

arXiv:2606.08578v1 Announce Type: new Abstract: Recently, large time series models (LTSMs) have gained increasing attention due to their similarities to large language models, including flexible conte

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

PRISM: PRior-guided Imagination Sampling in world Models

DGX agent

arXiv:2606.07974v1 Announce Type: cross Abstract: A learned world model provides a powerful physical intuition for evaluating future states. But its effectiveness in continuous control also depends cr

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Do Local Score Models Extrapolate Across Size? A Diagnostic Theory and Benchmark

DGX agent

arXiv:2606.09705v1 Announce Type: new Abstract: Scientific generative modeling often requires size transfer, where models trained on small systems are evaluated on larger ones. While translation-invar

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Closed-Form Spectral Regularization for Multi-Task Model Merging

DGX agent

arXiv:2606.07289v1 Announce Type: cross Abstract: Model merging combines several independently fine-tuned experts into a single multi-task model without any training data, reducing the storage, servin

model-releasesarxiv-cs-cv
8 Jun 2026
Research

Drifting Models for Surrogate Flow Modeling

DGX agent

arXiv:2606.07481v1 Announce Type: new Abstract: While Computational Fluid Dynamics (CFD) provides high-fidelity flow fields for optimizing indoor environments, its computational cost limits rapid expl

researcharxiv-cs-lg
8 Jun 2026
Research

Position: A Dynamical Systems Perspective is Needed to Advance Time Series Modeling

DGX agent

arXiv:2602.16864v2 Announce Type: replace-cross Abstract: Time series (TS) modeling has come a long way from early statistical, mainly linear, approaches to the current trend in TS foundation models.

researcharxiv-cs-ai
8 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

GENEB: Why Genomic Models Are Hard to Compare

DGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

DGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

DGX agent

arXiv:2606.03827v1 Announce Type: cross Abstract: In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typ

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Echelon: Auditable Aggregate-Only Language-Model Adaptation Across Privacy Boundaries

DGX agent

arXiv:2606.02958v1 Announce Type: cross Abstract: Cross-organization language-model adaptation increasingly faces hard governance constraints: in many deployments, device-level model state-parameters,

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

From Answers to States: Verifiable Process-Level Evaluation of Chemical Reasoning in Large Language Models

DGX agent

arXiv:2606.03660v1 Announce Type: new Abstract: Large language models are increasingly used as chemistry assistants, yet most chemistry benchmarks still score only final answers. This masks a critical

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Greener Than Humans? Environmental Attitudes in Large Language Models

DGX agent

arXiv:2606.02741v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in sustainability-related decision support, reporting, and public communication, yet little systemati

model-releasesarxiv-cs-cl
3 Jun 2026
Agents

Comprehensive AI governance requires addressing non-model gains

DGX agent

arXiv:2606.00047v1 Announce Type: cross Abstract: Frontier AI governance often centres on the model-level governance paradigm, which assumes that a model's capability profile is primarily a function o

agentsarxiv-cs-ai
2 Jun 2026
Safety

Emergent Collaborative Deliberation in Multi-Model AI Systems: A BFT-Derived Protocol for Epistemic Synthesis

DGX agent

arXiv:2606.00005v1 Announce Type: new Abstract: We present the Consilium Protocol, a Byzantine Fault Tolerance-derived architecture for structured multi-model AI deliberation that treats inter-model d

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Geometry-Aware Implicit Memory for Video World Models

DGX agent

arXiv:2606.02436v1 Announce Type: new Abstract: Video world models aim to simulate controllable visual environments, but long-horizon rollouts depend on what the model remembers after observations lea

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models

DGX agent

arXiv:2606.00793v1 Announce Type: new Abstract: Recent advancements in video-based world models have demonstrated an unprecedented ability to synthesize high-fidelity visual sequences. However, a fund

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment

DGX agent

arXiv:2606.02198v1 Announce Type: new Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce differen

safetyarxiv-cs-lg
2 Jun 2026
Model Releases

Rethinking Scientific Modeling: Toward Physically Consistent and Simulation-Executable Programmatic Generation

DGX agent

arXiv:2602.07083v2 Announce Type: replace-cross Abstract: Structural modeling is a fundamental component of computational engineering science, in which even minor physical inconsistencies or specifica

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RPCASSM: Robust PCA State Space Model For Infrared Small Target Detection

DGX agent

arXiv:2606.01689v1 Announce Type: cross Abstract: The detection and segmentation of infrared small targets have important application significance in the fields of surveillance and security, maritime

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

T1: Tool-integrated Verification for Test-time Compute Scaling in Small Language Models

DGX agent

arXiv:2504.04718v2 Announce Type: replace-cross Abstract: Recent studies have demonstrated that test-time compute scaling effectively improves the performance of small language models (sLMs). However,

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?

DGX agent

arXiv:2606.01247v1 Announce Type: new Abstract: Humans can reproduce the viewpoint specified by a target image through active head and body motion, yet spatial intelligence in foundation models has la

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Diversity Matters: Revisiting Test-Time Compute in Vision-Language Models

DGX agent

arXiv:2605.30713v1 Announce Type: cross Abstract: Test-time compute (TTC) strategies have emerged as a lightweight approach to boost reasoning in large language models (LLMs). However, their applicati

researcharxiv-cs-cv
1 Jun 2026
Model Releases

MADS: Model-Aware Diverse Core Set Selection for Instruction Tuning

DGX agent

arXiv:2605.30857v1 Announce Type: new Abstract: Instruction fine-tuning is employed to enhance the instruction-following ability of large language models (LLMs). As the amount of instruction fine-tuni

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study

DGX agent

arXiv:2605.31408v1 Announce Type: cross Abstract: Skill documents provide procedural knowledge to large-language-model agents at inference time. This article studies whether the presentation granulari

model-releasesarxiv-cs-ai
1 Jun 2026
Applications

UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling

DGX agent

arXiv:2605.30898v1 Announce Type: new Abstract: In real-world deployments of large language models (LLMs), balancing inference quality and computational cost has become a central challenge. Existing a

applicationsarxiv-cs-ai
1 Jun 2026
Model Releases

Your Multimodal Speech Model Says I Have a Face for Radio

DGX agent

arXiv:2605.30472v1 Announce Type: new Abstract: As large neural models have become better at language tasks, researchers are increasingly building multi- and omnimodal models that handle more modaliti

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials

DGX agent

arXiv:2510.04704v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promising potential in scientific research, enabling tasks ranging from knowledge retrieval to propert

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation

DGX agent

arXiv:2605.28830v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in safety-critical applications, robust content moderation becomes essential. We present a c

model-releasesarxiv-cs-ai
29 May 2026
Research

Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?

DGX agent

arXiv:2510.16060v2 Announce Type: replace-cross Abstract: The recent development of foundation models for time series data has generated considerable interest in using such models across a variety of

researcharxiv-cs-ai
29 May 2026
Model Releases

CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models

DGX agent

arXiv:2605.28919v1 Announce Type: cross Abstract: Large language models have achieved strong reasoning capabilities, though often at the cost of massive parameter counts and expensive inference. In th

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

DenseSteer: Steering Small Language Models towards Dense Math Reasoning

DGX agent

arXiv:2605.29247v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong chain-of-thought (CoT) reasoning abilities, while smaller models (<= 3B parameters) significantly underp

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Label-Free Reinforcement Learning via Cross-Model Entropy

DGX agent

arXiv:2605.29009v1 Announce Type: cross Abstract: Post-training large language models with reinforcement learning is bottlenecked by the reward signal. Existing approaches require either ground-truth

model-releasesarxiv-cs-ai
29 May 2026
Research

Opt-Verifier: Unleashing the Power of LLMs for Optimization Modeling via Dual-Side Verification

DGX agent

arXiv:2605.29556v1 Announce Type: new Abstract: Building mathematical optimization models is critical in operations research (OR), while it requires substantial human expertise. Recent advancements ha

researcharxiv-cs-ai
29 May 2026
Model Releases

Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

DGX agent

arXiv:2603.05488v4 Announce Type: replace-cross Abstract: We provide evidence of performative chain-of-thought (CoT) in reasoning models, where a model becomes strongly confident in its final answer,

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

DGX agent

arXiv:2605.30161v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong performance on spatial reasoning benchmarks, yet it remains unclear whether this reflects structured 3D und

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

A Simple State Space Model Excels at Multivariate Time Series Classification

DGX agent

arXiv:2605.27406v1 Announce Type: new Abstract: Structured state space models (SSMs) have recently emerged as a promising foundation for sequence modeling, with Mamba-based architectures demonstrating

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Evaluating Local Explainability Metrics for Machine Learning Models on Tabular Data

DGX agent

arXiv:2605.27618v1 Announce Type: new Abstract: Despite the wide use of explainability techniques to attempt to understand the behavior of Artificial Intelligence (AI), the generated explanations may

model-releasesarxiv-cs-lg
28 May 2026
Tutorials

Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study

DGX agent

arXiv:2512.15791v2 Announce Type: replace-cross Abstract: In Artificial Intelligence (AI), language models have gained significant importance due to the widespread adoption of systems capable of simul

tutorialsarxiv-cs-ai
28 May 2026
Model Releases

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models

DGX agent

arXiv:2605.28347v1 Announce Type: new Abstract: Multi-Label Recognition (MLR) based on Vision-Language Models (VLMs) aims to leverage their pre-trained knowledge to better adapt complex recognition sc

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…3637383940…1248
Next →