AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Transformer-based few-shot learning for modeling Electricity Consumption Profiles with minimal data across thousands of domains

DGX agent

arXiv:2408.08399v3 Announce Type: replace Abstract: Electricity Consumption Profiles (ECPs) are crucial for operating and planning power distribution systems, especially with the increasing number of

researcharxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks

DGX agent

arXiv:2605.25388v1 Announce Type: new Abstract: Nucleotide sequences constitute the fundamental genetic basis of biological systems, rendering viral genomic analysis critical for biomedical advancemen

model-releasesarxiv-cs-lg
26 May 2026
Research

Beyond Log Likelihood: Probability-Based Objectives for Supervised Fine-Tuning across the Model Capability Continuum

DGX agent

arXiv:2510.00526v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is the standard approach for post-training large language models (LLMs), yet it often shows limited generalization. We

researcharxiv-cs-cl
25 May 2026
Model Releases

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

DGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models

DGX agent

arXiv:2605.23699v1 Announce Type: new Abstract: Video prediction is increasingly viewed as a path toward generalizable world models, yet it remains unclear whether these systems learn underlying causa

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Decomposition-Based Modular Conformal Prediction for Two-Stage Modeling

DGX agent

arXiv:2510.04406v2 Announce Type: replace-cross Abstract: Conformal prediction offers finite-sample coverage guarantees under minimal assumptions. However, existing methods treat the entire modeling p

model-releasesarxiv-cs-lg
25 May 2026
Research

DiLaDiff: Distilled Latent-Augmented Diffusion for Language Modeling

DGX agent

arXiv:2605.23605v1 Announce Type: cross Abstract: Diffusion language models intrinsically fail to capture correlations between decoded tokens, which leads to a harsh trade-off between sampling quality

researcharxiv-cs-ai
25 May 2026
Model Releases

How Far Are We from Generating Missing Modalities with Foundation Models?

DGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Joint Model Parameter Scaling and Universal-Domain Data Integration for E-commerce Search Ranking

DGX agent

arXiv:2603.24226v3 Announce Type: replace-cross Abstract: Scaling studies for industrial search, advertising, and recommendation have largely emphasized enlarging model capacity or refining architectu

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

DGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Open Multimodal Datasets and Open-Source Software for Data-Driven Modeling of Multiphase Transport and Thermal Systems

DGX agent

arXiv:2605.23037v1 Announce Type: new Abstract: Data-driven modeling is becoming central to multiphase transport, electronics cooling, acoustic diagnostics, and thermal-fluid digital twins, but progre

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

DGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

model-releasesarxiv-cs-cl
25 May 2026
Research

Understanding Task Aggregation for Generalizable Ultrasound Foundation Models

DGX agent

arXiv:2603.18123v3 Announce Type: replace-cross Abstract: Foundation models promise to unify multiple clinical tasks within a single framework, but recent ultrasound studies report that unified models

researcharxiv-cs-ai
25 May 2026
Model Releases

Unextractable Protocol Models: Collaborative Training and Inference without Weight Materialization

DGX agent

arXiv:2605.23464v1 Announce Type: new Abstract: We consider a decentralized setup in which the participants collaboratively train and serve a large neural network, and where each participant only proc

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

VDE: Training-Free Accelerating Rectified Flow Model via Velocity Decomposition and Estimation

DGX agent

arXiv:2605.23381v1 Announce Type: new Abstract: Though rectified flow models have achieved remarkable performance in image, video, and 3D generation, their practical deployments are challenged by slow

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

When Symptoms Are Not Enough: Evidence-Weighting Patterns in Large Language Model Psychiatric Screening

DGX agent

arXiv:2605.23148v1 Announce Type: new Abstract: As demand for mental health care outpaces clinician-delivered assessment, scalable screening tools are increasingly needed. Large language models (LLMs)

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Billion-Scale Graph Foundation Models

DGX agent

arXiv:2602.04768v2 Announce Type: replace Abstract: Graph-structured data underpins many critical applications. While foundation models have transformed language and vision via large-scale pretraining

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

ChronoMedicalWorld: A Medical World Model for Learning Patient Trajectories from Longitudinal Care Data

DGX agent

arXiv:2605.21963v1 Announce Type: new Abstract: Long-horizon clinical simulation -- predicting how a patient's physiology evolves over years under specified interventions -- is central to chronic-dise

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

ChronoVAE-HOPE: Beyond Attention -- A Next-Generation VAE Foundation Model for Specialized Time Series Classification

DGX agent

arXiv:2605.22684v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have become a new component of the state-of-the-art in general time series forecasting. However, adapting them to

model-releasesarxiv-cs-lg
23 May 2026
Hardware

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

DGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

hardwarearxiv-cs-lg
23 May 2026
Model Releases

Tabular foundation models for robust calibration of near-infrared chemical sensing data

DGX agent

arXiv:2605.21544v1 Announce Type: new Abstract: Near-infrared spectroscopy is increasingly used as a rapid, non-destructive chemical sensing technology for the analysis of food, pharmaceutical, biolog

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark

DGX agent

arXiv:1709.03806v2 Announce Type: replace Abstract: Modern vision models have achieved strong object-recognition performance, yet it remains unclear whether their representations encode object-level s

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Linear Dynamics in the RLVR Training of Large Language Models

DGX agent

arXiv:2601.04537v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven significant performance gains in reasoning-oriented large language models (LL

model-releasesarxiv-cs-cl
22 May 2026
Research

One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluation

DGX agent

arXiv:2605.22544v1 Announce Type: new Abstract: Instruction embedding models have become common among state-of-the-art models, however are evaluated using a single prompt per task. The single-point ev

researcharxiv-cs-cl
22 May 2026
Research

Probabilistic Attribution For Large Language Models

DGX agent

arXiv:2605.21726v1 Announce Type: new Abstract: The generative nature of Large Language Models (LLMs) is reflected in the conditional probabilities they compute to sample each response token given the

researcharxiv-cs-cl
22 May 2026
Model Releases

Reflective Prompt Tuning through Language Model Function-Calling

DGX agent

arXiv:2605.21781v1 Announce Type: new Abstract: Large language models (LLMs) have become increasingly capable of following instructions and complex reasoning, making prompting a flexible interface for

model-releasesarxiv-cs-cl
22 May 2026
Research

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

DGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

researcharxiv-cs-ro
22 May 2026
Model Releases

Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception

DGX agent

arXiv:2605.21882v1 Announce Type: new Abstract: Vision-language models (VLMs) often fail under low illumination because their visual grounding is learned predominantly from RGB imagery, whereas therma

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Do LLMs Know What Luxembourgish Borrows? Probing Lexical Neology in Low-Resource Multilingual Models

DGX agent

arXiv:2605.21227v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for writing assistance in small contact languages, yet it is unclear whether they respect community n

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Do Vision--Language Models Understand 3D Scenes or Just Catalogue Objects?

DGX agent

arXiv:2605.20448v1 Announce Type: new Abstract: Vision--language models reliably name objects in a scene, but do they represent the 3D layout those objects inhabit? We introduce a 3,034-sample human-c

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Post-Hoc Understanding of Metaphor Processing in Decoder-Only Language Models via Conditional Scale Entropy

DGX agent

arXiv:2605.21391v1 Announce Type: new Abstract: Metaphor requires a language model to resolve a token whose contextual meaning diverges from its basic literal sense. Understanding how transformer mode

model-releasesarxiv-cs-cl
21 May 2026
Agents

STELLAR: Scaling 3D Perception Large Models for Autonomous Driving

DGX agent

arXiv:2605.20390v1 Announce Type: new Abstract: Model scaling has demonstrated remarkable success through large-scale training on diverse datasets. It remains an open question whether the same paradig

agentsarxiv-cs-cv
21 May 2026
Model Releases

SymbolicLight V1: Spike-Gated Dual-Path Language Modeling with High Activation Sparsity and Sub-Billion-Scale Pre-Training Evidence

DGX agent

arXiv:2605.21333v1 Announce Type: new Abstract: Natively trained spiking language models struggle to combine Transformer-like language quality, stable multi-domain pre-training, and high activation sp

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

TempGlitch: Evaluating Vision-Language Models for Temporal Glitch Detection in Gameplay Videos

DGX agent

arXiv:2605.21443v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly being explored for video game quality assurance, especially gameplay glitch detection. Most existing eval

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping

DGX agent

arXiv:2510.08482v3 Announce Type: replace-cross Abstract: Iconicity, the resemblance between linguistic form and meaning, is pervasive in signed languages, offering a natural testbed for visual ground

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

A Hybrid Modeling Framework for Crop Prediction Tasks via Dynamic Parameter Calibration and Multi-Task Learning

DGX agent

arXiv:2603.15411v2 Announce Type: replace Abstract: Accurate prediction of crop states (e.g., phenology stages and cold hardiness) is essential for timely farm management decisions such as irrigation,

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Addressing prior dependence in hierarchical Bayesian modeling for PTA data analysis II: Noise and SGWB inference through parameter decorrelation

DGX agent

arXiv:2511.01959v2 Announce Type: replace-cross Abstract: Pulsar Timing Arrays (PTA) provide a powerful framework to measure low-frequency gravitational waves, but accuracy and robustness of the resul

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Distributional Energy-Based Models for Uncertainty-Aware Structured LLM Reasoning

DGX agent

arXiv:2605.18871v1 Announce Type: cross Abstract: When Large Language Models produce structured outputs such as travel plans, code solutions, or multi-step proofs, individual reasoning steps may appea

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models

DGX agent

arXiv:2605.18824v1 Announce Type: cross Abstract: Evaluation of foundation models often rely on aggregate scores from benchmarks that lack comprehensive coverage and metadata for a fine-grained evalua

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding

DGX agent

arXiv:2605.19846v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities in general video understanding, yet they often struggle with the fine-grained

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models

DGX agent

arXiv:2605.18795v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) dominates parameter-efficient fine-tuning of large language models, yet most variants target dense architectures. Mixture-o

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

LIFT and PLACE: A Simple, Stable, and Effective Knowledge Distillation Framework for Lightweight Diffusion Models

DGX agent

arXiv:2605.19729v1 Announce Type: cross Abstract: We demonstrate that in knowledge distillation for diffusion models, the teacher network's highly complex denoising process - stemming from its substan

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models

DGX agent

arXiv:2605.20128v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into high-stakes decision-making. Inspired by the theory of inattentional blindness in human co

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

OpenCompass: A Universal Evaluation Platform for Large Language Models

DGX agent

arXiv:2605.19276v1 Announce Type: new Abstract: In recent years, the field of artificial intelligence has undergone a paradigm shift from task-specific small-scale models to general-purpose large lang

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Prompting language influences diagnostic reasoning and accuracy of large language models

DGX agent

arXiv:2605.19173v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored for clinical decision support, yet most evaluations are conducted in English, leaving their relia

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding

DGX agent

arXiv:2605.19329v1 Announce Type: cross Abstract: Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fas

model-releasesarxiv-cs-ai
20 May 2026
Safety

Stitched Value Model for Diffusion Alignment

DGX agent

arXiv:2605.19804v1 Announce Type: cross Abstract: For practical use, diffusion- or flow-based generative models must be aligned with task-specific rewards, such as prompt fidelity or aesthetic prefere

safetyarxiv-cs-ai
20 May 2026
← Previous
1…6667686970…1030
Next →