AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies

DGX agent

arXiv:2605.29712v1 Announce Type: cross Abstract: Grounded claim factuality checking is important for large language model (LLM) applications such as retrieval-augmented generation, as it helps users

model-releasesarxiv-cs-ai
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Three-dimensional Conditional Diffusion Models for Cosmological 21 cm Lightcone Emulation

DGX agent

arXiv:2605.29016v1 Announce Type: cross Abstract: We investigate conditional diffusion modeling for three-dimensional 21 cm lightcone emulation, focusing on cubes with a sky-plane size of 64imes64 and

model-releasesarxiv-cs-lg
29 May 2026
Research

Towards a Foundation Model for the Martian Atmosphere

DGX agent

arXiv:2605.28851v1 Announce Type: cross Abstract: The martian atmosphere hosts dynamical phenomena ranging from planet-encircling dust storms to mesoscale orographic clouds and nocturnal low-level jet

researcharxiv-cs-lg
29 May 2026
Research

Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthetic Data and Relative Context Discrepancy

DGX agent

arXiv:2509.21190v4 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is a critical task, but developing models that generalize to unseen data in a zero-shot manner remains a

researcharxiv-cs-ai
29 May 2026
Local Ai

Towards Understanding the Shape of Representations in Protein Language Models

DGX agent

arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran

local-aiarxiv-cs-lg
29 May 2026
Model Releases

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

DGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

model-releasesarxiv-cs-cl
29 May 2026
Agents

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

DGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

agentsarxiv-cs-ai
28 May 2026
Model Releases

AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models

DGX agent

arXiv:2602.18481v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has led to a surge of financial benchmarks, evolving from static knowledge evaluation to

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems

DGX agent

arXiv:2605.27492v1 Announce Type: cross Abstract: LLM agents are rapidly evolving from coding assistants into autonomous software engineering systems. However, existing evaluation methodologies remain

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Cultural Binding Heads in Language Models

DGX agent

arXiv:2605.28543v1 Announce Type: new Abstract: LLMs often default to equal treatment across cultural groups, even though context warrants differentiation: this is a lack of difference awareness. Usin

model-releasesarxiv-cs-ai
28 May 2026
Safety

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes

DGX agent

arXiv:2605.28421v1 Announce Type: new Abstract: Reinforcement learning has become a central paradigm for advancing reasoning in large language models, yet most existing methods still depend on stronge

safetyarxiv-cs-ai
28 May 2026
Safety

From Pixels to Words -- Towards Native One-Vision Models at Scale

DGX agent

arXiv:2605.28820v1 Announce Type: new Abstract: Current vision-language models (VLMs) typically stitch together separate image encoders and language decoders via multi-stage alignment, a modular frame

safetyarxiv-cs-cv
28 May 2026
Model Releases

High Performance, Low Reliability: Uncertainty Benchmarking for Tabular Foundation Models

DGX agent

arXiv:2605.28554v1 Announce Type: new Abstract: Recent Tabular Foundation Models (TFMs) have demonstrated state-of-the-art predictive performance, often surpassing Gradient-Boosted Decision Trees (GBD

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Interpretability-Guided Layer Selection over Subspace Projection: SAEs as Stethoscopes, Not Scalpels, for Raw Task Vector Model Editing

DGX agent

arXiv:2605.28649v1 Announce Type: cross Abstract: LLMs increasingly require surgical model editing to enhance domain-specific capabilities without incurring the computational cost or catastrophic forg

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration

DGX agent

arXiv:2602.20497v3 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in image and video generation tasks. However, the high computational demands of Diffusion Tr

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

{Omega}-QVLA: Robust Quantization for Vision-Language-Action Models via Composite Rotation and Per-step Scaling

DGX agent

arXiv:2605.28803v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models unify perception, reasoning, and control within a single policy, yet their multi-billion-parameter backbones and dif

model-releasesarxiv-cs-cv
28 May 2026
Research

On the Fallacy of Global Token Perplexity in Spoken Language Model Evaluation

DGX agent

arXiv:2601.06329v2 Announce Type: replace-cross Abstract: Generative spoken language models pretrained on large-scale raw audio can continue a speech prompt with appropriate content while preserving a

researcharxiv-cs-ai
28 May 2026
Research

Playing with Words, Improving with Rewards: Training Language Models for Creative Association

DGX agent

arXiv:2605.27832v1 Announce Type: new Abstract: Large Language Models (LLMs) are being applied to increasingly difficult problems and use cases. To navigate their vast solution spaces effectively, LLM

researcharxiv-cs-cl
28 May 2026
Model Releases

PrunePath: Towards Highly Structured Sparse Language Models

DGX agent

arXiv:2605.28283v1 Announce Type: cross Abstract: Feed-forward networks (FFNs) dominate the parameter count and computation of modern language models, yet existing pruning methods often struggle to co

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Pruning and Distilling Mixture-of-Experts into Dense Language Models

DGX agent

arXiv:2605.28207v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) is now the dominant architecture for frontier language models, yet it requires all expert parameters to be loaded in memory,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification

DGX agent

arXiv:2512.12887v3 Announce Type: replace Abstract: 3D medical image classification is essential for modern clinical workflows. Medical foundation models (FMs) have emerged as a promising approach for

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

RGC: a radio AGN classifier based on deep learning. I. A semi-supervised multiclass model for VLA images

DGX agent

arXiv:2510.22190v2 Announce Type: replace-cross Abstract: Bent radio active galactic nuclei (RAGNs) -- wide-angle tails (WATs) and narrow-angle tails (NATs) -- trace dense environments in galaxy group

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Language Models? A Bias-Controlled Study

DGX agent

arXiv:2504.04540v2 Announce Type: replace-cross Abstract: 3D Large Language Models (LLMs) leveraging spatial information in point clouds for 3D spatial reasoning attract great attention. Despite some

model-releasesarxiv-cs-ai
28 May 2026
Safety

Trust Me, I'm an Expert: Decoding and Steering Authority Bias in Large Language Models

DGX agent

arXiv:2601.13433v3 Announce Type: replace Abstract: Prior research demonstrates that performance of language models on reasoning tasks can be influenced by suggestions, hints and endorsements. However

safetyarxiv-cs-cl
28 May 2026
Model Releases

A Deep State-Space Model Compression Method using Upper Bound on Output Error

DGX agent

arXiv:2510.14542v2 Announce Type: replace-cross Abstract: We study deep state-space models (Deep SSMs) that contain linear quadratic-output (LQO) systems as internal blocks and present a compression m

model-releasesarxiv-cs-lg
27 May 2026
Safety

AirCast-SR: A Foundation Model for Kilometer-Scale Atmospheric Super-Resolution via Latent Consistency Diffusion

DGX agent

arXiv:2605.26130v1 Announce Type: new Abstract: Operational weather prediction at kilometer scales remains computationally prohibitive for traditional numerical weather prediction (NWP) models, limiti

safetyarxiv-cs-lg
27 May 2026
Safety

Aperiodic and Low-Frequency Spectral Bias in Reconstruction based EEG Foundation Models

DGX agent

arXiv:2605.26434v1 Announce Type: cross Abstract: EEG foundation models, pre-trained on large-scale unlabelled EEG data, have emerged as a promising direction towards learning generalizable EEG repres

safetyarxiv-cs-ai
27 May 2026
Safety

Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models

DGX agent

arXiv:2605.06213v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) today rests on fixed benchmarks that apply the same set of items to any model, producing ceiling and floor e

safetyarxiv-cs-ai
27 May 2026
Model Releases

EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models

DGX agent

arXiv:2510.07231v4 Announce Type: replace-cross Abstract: Socio-economic causal effects depend heavily on their institutional and environmental contexts. The same intervention can produce different, e

model-releasesarxiv-cs-ai
27 May 2026
Safety

Erased but Exploitable: Black-box Embedding-Aware Prompting Against Unlearned Text-to-Image Diffusion Models

DGX agent

arXiv:2605.26332v1 Announce Type: cross Abstract: Machine unlearning aims to remove specific concepts from pretrained text-to-image diffusion models, yet several white- and black-box attacks have been

safetyarxiv-cs-ai
27 May 2026
Safety

Few-shot Cross-country Generalization of Tabular Machine Learning and Foundation Models for Childhood Anemia Prediction under Distribution Shift

DGX agent

arXiv:2605.26589v1 Announce Type: cross Abstract: Childhood anemia affects around 40% of children aged 6-59 months globally and arises from heterogeneous factors, limiting model generalizability. We e

safetyarxiv-cs-ai
27 May 2026
Model Releases

Hubness, Not Anisotropy, Drives Cross-Lingual Retrieval Asymmetry in Multilingual Embedding Models

DGX agent

arXiv:2605.26575v1 Announce Type: new Abstract: Multilingual embedding models are deployed under the assumption that cross-lingual retrieval is symmetric: if a query in language A retrieves its transl

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?

DGX agent

arXiv:2605.26548v1 Announce Type: cross Abstract: Large language models (LLMs) now support automated software security tasks, including vulnerability discovery and proof-of-concept (PoC) generation. E

model-releasesarxiv-cs-lg
27 May 2026
Research

Towards Controllable Image Generation through Representation-Conditioned Diffusion Models

DGX agent

arXiv:2605.27343v1 Announce Type: new Abstract: Diffusion models have emerged as powerful tools for high-quality image generation and editing, but guiding these models to produce specific outputs rema

researcharxiv-cs-cv
27 May 2026
Model Releases

ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models

DGX agent

arXiv:2605.24011v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models exhibit remarkable action generation for embodied intelligence, but their heavy compute make deployment on edge pl

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning

DGX agent

arXiv:2602.10090v3 Announce Type: replace Abstract: Recent advances in large language model (LLM) have empowered autonomous agents to perform multi-turn interactions with tools and environments. Howev

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

AstroMind: A High-Fidelity Benchmark for Spacecraft Behavior Reasoning Based on Large Language Models

DGX agent

arXiv:2605.24573v1 Announce Type: new Abstract: Understanding why a spacecraft maneuvers -- rather than simply that it did -- is an increasingly important problem for space domain awareness as Earth o

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Benchmarking Pathology Foundation Models for Spatial Domain Understanding

DGX agent

arXiv:2605.25764v1 Announce Type: cross Abstract: Pathology foundation models (PFMs) have emerged as a core approach for learning transferable representations from whole slide images (WSIs), and they

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Beyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching

DGX agent

arXiv:2605.25558v1 Announce Type: new Abstract: Optimizing the trade-off among predictive performance and computational cost is a central focus in the deployment of Large Language Models (LLMs). Curre

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models

DGX agent

arXiv:2605.25189v1 Announce Type: cross Abstract: Reward hacking arises when a model improves a proxy reward by exploiting shortcuts rather than solving the intended task. We study this failure mode t

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

DIVER-1: Scaling Intracranial EEG Foundation Models for Transferable Representations

DGX agent

arXiv:2512.19097v3 Announce Type: replace-cross Abstract: Intracranial EEG (iEEG) provides direct, millisecond-scale recordings of human neural activity, but reusable representation learning is diffic

model-releasesarxiv-cs-ai
26 May 2026
Safety

Extracting Training Data from Diffusion Language Models via Infilling

DGX agent

arXiv:2605.24173v1 Announce Type: cross Abstract: Memorization in large language models has been studied almost exclusively through prefix-conditioned extraction, a natural choice for autoregressive m

safetyarxiv-cs-ai
26 May 2026
Safety

Factored Latent Action World Models

DGX agent

arXiv:2602.16229v2 Announce Type: replace Abstract: Learning latent actions from action-free video has emerged as a powerful paradigm for scaling up controllable world model learning. Latent actions p

safetyarxiv-cs-lg
26 May 2026
Model Releases

Game-Theoretic Modeling of Heterogeneous Investor Interactions for Stock Price Forecasting

DGX agent

arXiv:2605.23953v1 Announce Type: cross Abstract: Accurate stock price forecasting has consistently remained a pivotal yet challenging FinTech task that underpins quantitative trading and investment d

model-releasesarxiv-cs-lg
26 May 2026
Safety

Generative OOD-regularized Model-based Policy Optimization

DGX agent

arXiv:2605.24405v1 Announce Type: cross Abstract: We study sequential decision-making with offline reinforcement learning (RL). Traditional offline RL policies may result in out-of-distribution (OOD)

safetyarxiv-cs-ai
26 May 2026
Model Releases

How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework

DGX agent

arXiv:2507.19219v2 Announce Type: replace Abstract: Overestimation in evaluating large language models (LLMs) has become an increasing concern. Due to the contamination of public benchmarks or imbalan

model-releasesarxiv-cs-cl
26 May 2026
Local Ai

Language Models Need Sleep

DGX agent

arXiv:2605.26099v1 Announce Type: cross Abstract: Transformer-based large language models are increasingly used for long-horizon tasks; however, their attention mechanism scales poorly with context le

local-aiarxiv-cs-ai
26 May 2026
Model Releases

M^3-Verse: A 'Spot the Difference' Challenge for Large Multimodal Models

DGX agent

arXiv:2512.18735v2 Announce Type: replace-cross Abstract: Modern Large Multimodal Models (LMMs) have demonstrated extraordinary ability in static image and single-state spatial-temporal understanding.

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…8586878889…1030
Next →