AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Fast Kernel-Space Diffusion for Remote Sensing Pansharpening

DGX agent

arXiv:2505.18991v3 Announce Type: replace Abstract: Pansharpening seeks to fuse high-resolution panchromatic (PAN) and low-resolution multispectral (LRMS) images into a single image with both fine spa

model-releasesarxiv-cs-cv
19 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Federated Martingale Posterior Samping

DGX agent

arXiv:2605.18554v1 Announce Type: new Abstract: Federated Bayesian neural networks require fixing a prior on the model parameters together with a likelihood. Eliciting meaningful priors on the weight

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints

DGX agent

arXiv:2509.06984v3 Announce Type: replace-cross Abstract: Federated Learning with LoRA fine-tuning offers an efficient and privacy-aware solution for institutions to collaboratively leverage their lar

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fidelity Probes for Specification--Code Alignment

DGX agent

arXiv:2605.17246v1 Announce Type: cross Abstract: We introduce fidelity probes: natural-language questions generated from a reference artifact with code-derived ground-truth answers, answered from a c

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

FIM-LoRA: Task-Informative Rank Allocation for LoRA via Calibration-Time Gradient-Variance Estimation

DGX agent

arXiv:2605.16800v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) assigns a uniform rank to every adapted weight matrix - a practical convenience that ignores a fundamental reality: differe

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

DGX agent

arXiv:2510.08886v3 Announce Type: replace Abstract: Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale di

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fine-grained List-wise Alignment for Generative Medication Recommendation

DGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Fine-tuning Pocket-Aware Diffusion Models via Denoising Policy Optimization

DGX agent

arXiv:2605.17693v1 Announce Type: cross Abstract: Structure-based drug design has been accelerated by pocket-aware 3D generative models, yet most methods primarily fit the training distribution and ma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Finite-Particle Rates for Regularized Stein Variational Gradient Descent

DGX agent

arXiv:2602.05172v2 Announce Type: replace-cross Abstract: We derive finite-particle rates for the regularized Stein variational gradient descent (R-SVGD) algorithm introduced by He et al. (2024) that

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FinTagging: Benchmarking LLMs for Extracting and Structuring Financial Information

DGX agent

arXiv:2505.20650v5 Announce Type: replace-cross Abstract: Accurate interpretation of numerical data in financial reports is critical for markets and regulators. Although XBRL (eXtensible Business Repo

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

DGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

DGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics

DGX agent

arXiv:2605.17373v1 Announce Type: cross Abstract: AI research agents accelerate ML research by automating hypothesis generation, experimentation, and empirical refinement. Existing agent strategies ra

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion

DGX agent

arXiv:2605.18346v1 Announce Type: cross Abstract: Recent advances in autoregressive video diffusion have enabled sequential and streaming video generation. However, long-horizon generation requires in

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Form and Function: Machine Unlearning as a Problem of Misaligned States

DGX agent

arXiv:2605.17590v1 Announce Type: new Abstract: We formulate machine unlearning for online L-BFGS as a counterfactual state-alignment problem. Given an actual event stream and a deletion-edited counte

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FormuLLA: A Large Language Model Approach to Generating Novel 3D Printable Formulations

DGX agent

arXiv:2601.02071v3 Announce Type: replace Abstract: Pharmaceutical three-dimensional (3D) printing is an advanced fabrication technology with the potential to enable truly personalised dosage forms. R

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models

DGX agent

arXiv:2508.06038v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) incur substantial computational overhead and inference latency due to the large number of vision tokens introduc

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning

DGX agent

arXiv:2605.17162v1 Announce Type: new Abstract: This paper investigates whether shallow neural network agents can master the card game Schnapsen and challenge a strong search-based baseline, RdeepBot,

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

DGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

DGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets

DGX agent

arXiv:2605.18475v1 Announce Type: cross Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. Ho

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

DGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

General Preference Reinforcement Learning

DGX agent

arXiv:2605.18721v1 Announce Type: cross Abstract: Post-training has split large language model (LLM) alignment into two largely disconnected tracks. Online reinforcement learning (RL) with verifiable

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models

DGX agent

arXiv:2605.17565v1 Announce Type: new Abstract: Recent work has fine-tuned language models on chess data and reported high benchmark scores as evidence that the resulting models can understand the rul

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Generative Artificial Intelligence for Literature Reviews

DGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

DGX agent

arXiv:2507.21035v3 Announce Type: replace Abstract: Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GenTS: A Comprehensive Benchmark Library for Generative Time Series Models

DGX agent

arXiv:2605.17804v1 Announce Type: new Abstract: Generative models have demonstrated remarkable potential in time series analysis tasks, like synthesis, forecasting, imputation, etc. However, offering

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Geometric Asymmetry in MoE Specialization: Functional Decorrelation and Representational Overlap

DGX agent

arXiv:2605.16349v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures achieve scalable capacity through sparse routing, yet the geometric structure of expert specialization remains po

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Geometric Scaling of Bayesian Inference in LLMs

DGX agent

arXiv:2512.23752v5 Announce Type: replace-cross Abstract: Recent work has shown that small transformers trained in controlled 'wind-tunnel'' settings can implement exact Bayesian inference, and that t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics

DGX agent

arXiv:2603.02531v2 Announce Type: replace-cross Abstract: Classifier-Free Guidance (CFG) improves sample quality in diffusion models, but its dual-pass inference and reliance on null-condition trainin

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology

DGX agent

arXiv:2605.18419v1 Announce Type: cross Abstract: Vision-language models (VLMs) can couple visual perception with open-ended clinical reasoning, making them attractive for computational histopathology

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GIM: Evaluating models via tasks that integrate multiple cognitive domains

DGX agent

arXiv:2605.18663v1 Announce Type: new Abstract: As LLM benchmarks saturate, the evaluation community has pursued two strategies to increase difficulty: escalating knowledge demands (GPQA, HLE) or remo

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GIST: Targeted Data Selection for Instruction Tuning via Coupled Optimization Geometry

DGX agent

arXiv:2602.18584v2 Announce Type: replace-cross Abstract: Targeted data selection has emerged as a crucial paradigm for efficient instruction tuning, aiming to identify a small yet influential subset

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GLT-PEFT: Gated Lie-Tucker Parameter-Efficient Fine-Tuning for Alzheimer's Disease Diagnosis with Hippocampal Segmentation Pretraining

DGX agent

arXiv:2605.16769v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as a promising paradigm for adapting pretrained models under limited data conditions. However, most e

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Good flavor search in SU(5): a machine learning approach

DGX agent

arXiv:2511.08154v2 Announce Type: replace-cross Abstract: We revisit the fermion mass problem of the SU(5) grand unified theory using machine learning techniques. The original SU(5) model proposed by

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics

DGX agent

arXiv:2506.14009v2 Announce Type: replace Abstract: Autonomous drones capable of interpreting and executing high-level language instructions in unstructured environments remain a long-standing goal. Y

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

Graph Embedding in the Graph Fractional Fourier Transform Domain

DGX agent

arXiv:2508.02383v2 Announce Type: replace Abstract: Spectral graph embedding plays a critical role in graph representation learning by generating low-dimensional vector representations from graph spec

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Graph Hierarchical Recurrence for Long-Range Generalization

DGX agent

arXiv:2605.18387v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) and Graph Transformers (GTs) are now a fundamental paradigm for graph learning, combining the representation-learning cap

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GVGAI-LLM: Evaluating Large Language Model Agents with Infinite Games

DGX agent

arXiv:2508.08501v3 Announce Type: replace Abstract: We introduce GVGAI-LLM, a video game benchmark for evaluating the reasoning and problem-solving capabilities of large language models (LLMs). Built

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

HalluScore: Large Language Model Hallucination Question Answering Benchmark

DGX agent

arXiv:2605.17007v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to g

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation

DGX agent

arXiv:2605.17093v1 Announce Type: cross Abstract: Distilling vision-language models into faster hybrid architectures, such as 3:1 Mamba-2/attention mixes, is now standard practice for making inference

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

High-dimensional ridge regression with random features for non-identically distributed data with a variance profile

DGX agent

arXiv:2504.03035v2 Announce Type: replace-cross Abstract: Random feature ridge regression is often analyzed in the high-dimensional regime under the homogeneous sampling model x_i=Sigma^{1/2}x_i', whe

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Hilbert-Geo: Solving Solid Geometric Problems by Neural-Symbolic Reasoning

DGX agent

arXiv:2605.16385v1 Announce Type: cross Abstract: Geometric problem solving, as a typical multimodal reasoning problem, has attracted much attention and made great progress recently, however most of w

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Homoglyph-based Adversarial Perturbation of Introductory Computer Science Theory Problems

DGX agent

arXiv:2605.16286v1 Announce Type: cross Abstract: Different AI tools such as ChatGPT, Gemini, and Claude are becoming very popular. Although they are helpful for many day-to-day tasks, they can be use

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

How Do Electrocardiogram Models Scale?

DGX agent

arXiv:2605.17276v1 Announce Type: cross Abstract: While scaling laws have established a fundamental framework for foundation models in natural language processing, their applicability to electrocardio

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

How does feature learning reshape the function space?

DGX agent

arXiv:2605.17718v1 Announce Type: cross Abstract: Feature learning is widely regarded as the key mechanism distinguishing neural networks from fixed-kernel methods, yet its impact on the induced funct

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking

DGX agent

arXiv:2605.18111v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) and Large Vision Language Models (LVLMs) have enabled general-purpose systems to demonstrate promisi

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

HPC-LLM: Practical Domain Adaptation and Retrieval-Augmented Generation for HPC Support

DGX agent

arXiv:2605.16347v1 Announce Type: new Abstract: Modern scientific research increasingly depends on High-Performance Computing (HPC) infrastructures, yet many researchers face significant operational b

model-releasesarxiv-cs-lg
19 May 2026
← Previous
1…225226227228229…361
Next →