AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Fair outputs, Biased Internals: Causal Potency and Asymmetry of Latent Bias in LLMs for High-Stakes Decisions

DGX agent

arXiv:2605.15217v1 Announce Type: new Abstract: Instruction-tuned language models exhibit behavioural fairness in high-stakes decisions while retaining biased associations in their internal representa

model-releasesarxiv-cs-ai
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Federated Imputation under Heterogeneous Feature Spaces

DGX agent

arXiv:2605.16099v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative training across decentralized clients, but most methods assume aligned feature schemas, an assumption th

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Federated Learning of Spiking Neural Networks under Heterogeneous Temporal Resolutions

DGX agent

arXiv:2605.15355v1 Announce Type: new Abstract: Spiking neural networks (SNNs) are biologically inspired energy-efficient models that use sparse binary spike-based communication between neurons, makin

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Feedback World Model Enables Precise Guidance of Diffusion Policy

DGX agent

arXiv:2605.15705v1 Announce Type: cross Abstract: World models aim to improve robotic decision making by predicting the consequences of actions. However, in practice, their predictions often become un

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

DGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

FFAvatar: Few-Shot, Feed-Forward, and Generalizable Avatar Reconstruction

DGX agent

arXiv:2605.15320v1 Announce Type: cross Abstract: Avatar reconstruction has traditionally relied on per-subject optimization that requires hours of computation or on expensive preprocessing that limit

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

DGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization

DGX agent

arXiv:2605.15980v1 Announce Type: new Abstract: Group Relative Policy Optimization has emerged as essential for aligning video diffusion models with human preferences, but faces a critical computation

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast

DGX agent

arXiv:2605.16233v1 Announce Type: new Abstract: Can LLM agents improve decision-making through self-generated memory without gradient updates? We propose FORGE (Failure-Optimized Reflective Graduation

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

ForMaT: Dataset for Visually-Grounded Multilingual PDF Translation

DGX agent

arXiv:2605.15794v1 Announce Type: new Abstract: We present ForMaT (Format-Preserving Multilingual Translation), a parallel corpus of 3,956 PDFs across 15 language pairs that preserves original layout

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

FormulaCode: Evaluating Agentic Optimization on Large Codebases

DGX agent

arXiv:2603.16011v2 Announce Type: replace-cross Abstract: Large language model (LLM) coding agents increasingly operate at the repository level, motivating benchmarks that evaluate their ability to op

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery

DGX agent

arXiv:2605.15412v1 Announce Type: cross Abstract: Modern quantitative trading increasingly relies on systematic models to extract predictive signals from large-scale financial data, where alpha factor

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

From Layers to Networks: Comparing Neural Representations via Diffusion Geometry

DGX agent

arXiv:2605.15901v1 Announce Type: new Abstract: Diffusion geometry is a manifold learning framework that uses random walks defined by Markov transition matrices to characterize the geometry of a datas

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Frontier Large Language Models Rival State-of-the-Art Planners

DGX agent

arXiv:2511.09378v2 Announce Type: replace Abstract: A series of influential studies established that large language models cannot reliably solve even simple planning tasks. We show that the latest gen

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

FRWKV+: Adaptive Periodic-Position Branch Interaction for Frequency-Space Linear Time Series Forecasting

DGX agent

arXiv:2605.15690v1 Announce Type: new Abstract: Long-term time series forecasting is essential for decision making in energy, finance, transportation, and healthcare systems. Recent lightweight foreca

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Fully Open Meditron: An Auditable Pipeline for Clinical LLMs

DGX agent

arXiv:2605.16215v1 Announce Type: new Abstract: Clinical decision support systems (CDSS) require scrutable, auditable pipelines that enable rigorous, reproducible validation. Yet current LLM-based CDS

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

DGX agent

arXiv:2509.12266v2 Announce Type: replace-cross Abstract: We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core c

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

GenShield: Unified Detection and Artifact Correction for AI-Generated Images

DGX agent

arXiv:2605.16122v1 Announce Type: cross Abstract: Diffusion-based image synthesis has made AI-generated images (AIGI) increasingly photorealistic, raising urgent concerns about authenticity in applica

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GESD: Beyond Outcome-Oriented Fairness

DGX agent

arXiv:2605.15295v1 Announce Type: cross Abstract: Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predi

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GiLT: Augmenting Transformer Language Models with Dependency Graphs

DGX agent

arXiv:2605.15562v1 Announce Type: new Abstract: Augmenting Transformers with linguistic structures effectively enhances the syntactic generalization performance of language models. Previous work in th

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis

DGX agent

arXiv:2602.20207v3 Announce Type: replace-cross Abstract: Knowledge editing in Large Language Models (LLMs) aims to update the model's prediction for a specific query to a desired target while preserv

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GQA-{mu}P: The maximal parameterization update for grouped query attention

DGX agent

arXiv:2605.15290v1 Announce Type: cross Abstract: Hyperparameter transfer across model architectures dramatically reduces the amount of compute necessary for tuning large language models (LLMs). The m

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding

DGX agent

arXiv:2605.15250v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), the attention used in DeepSeek-V2/V3, jointly compresses keys and values into a low-rank latent and matches the H10

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Graph-Regularized Sparse Autoencoders for LLM Safety Steering

DGX agent

arXiv:2512.06655v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) are increasingly used to extract activation directions for inference-time steering, but their standard sparsity obj

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero

DGX agent

arXiv:2605.15464v1 Announce Type: cross Abstract: Post-training has become a crucial step for unlocking the capabilities of large language models, with reinforcement learning (RL) emerging as a critic

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

HAI-Eval: Measuring Human-AI Synergy in Collaborative Coding

DGX agent

arXiv:2512.04111v2 Announce Type: replace-cross Abstract: LLM-powered coding agents are reshaping the development paradigm. However, existing evaluation systems, neither traditional tests for humans n

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

DGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Harnessing Unimodality in Semiparametric Contextual Pricing via Oracle Price Map Learning

DGX agent

arXiv:2605.15411v1 Announce Type: cross Abstract: We study contextual dynamic pricing in a semiparametric scalar-index valuation model where the latent value is v_t=mu_ast(mathsf c_t)+xi_t, with an un

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

DGX agent

arXiv:2605.15338v1 Announce Type: cross Abstract: Large language models are increasingly augmented with persistent memory, allowing assistants to store user-specific information across sessions for pe

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces

DGX agent

arXiv:2605.15753v1 Announce Type: cross Abstract: Functional 3D scene graphs offer a versatile and flexible representation for 3D scene understanding and robotic manipulation, defined by object nodes,

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models

DGX agent

arXiv:2510.02453v3 Announce Type: replace-cross Abstract: Frontier language models are deployed as black-box services, where model weights cannot be modified and customization is limited to prompting.

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

DGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

DGX agent

arXiv:2603.17915v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages rem

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

DGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Information-Preserving Domain Transfer with Unlabeled Data in Misspecified Simulation-Based Inference

DGX agent

arXiv:2605.05652v2 Announce Type: replace Abstract: Simulation-based inference (SBI) provides amortized Bayesian parameter inference from simulator-generated data without requiring explicit likelihood

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Interaction-Aware Influence Functions for Group Attribution

DGX agent

arXiv:2605.15675v1 Announce Type: cross Abstract: Influence functions approximate how removing a training example changes a quantity of interest, called the target function, such as a held-out loss. T

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Introducing MELI: the Mandarin-English Language Interview Corpus

DGX agent

arXiv:2603.27043v2 Announce Type: replace Abstract: We introduce the Mandarin-English Language Interview (MELI) Corpus, an open-source resource of 29.8 hours of speech from 51 Mandarin-English bilingu

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Judge Circuits

DGX agent

arXiv:2605.16023v1 Announce Type: new Abstract: LLM-as-a-judge has become the dominant paradigm for grading model outputs at scale, yet the same model assigns systematically different scores when its

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

DGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Large Language Models Could Be Rote Learners

DGX agent

arXiv:2504.08300v5 Announce Type: replace-cross Abstract: Benchmark-based evaluation, e.g., multiple-choice questions (MCQs) and open-ended questions (OEQs), is widely used for evaluating Large Langua

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

LASER: Language Model Regression for Semi-Structured Workflow Resource and Runtime Estimation

DGX agent

arXiv:2512.19701v2 Announce Type: replace-cross Abstract: Accurate prediction of resource consumption and runtime for cloud workflow jobs is critical for scheduling efficiency, yet remains challenging

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Layer Equivalence Is Not a Property of Layers Alone: How You Test Redundancy Changes What You Find

DGX agent

arXiv:2605.16234v1 Announce Type: cross Abstract: When researchers ask whether two transformer layers are 'equivalent' for compression, they often conflate distinct tests. Replacement asks whether one

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

DGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Disentangled Representations for Generalized Multi-view Clustering

DGX agent

arXiv:2605.15640v1 Announce Type: new Abstract: Multi-View Clustering (MVC) has gained significant attention for its ability to leverage complementary information across diverse views. However, existi

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Dynamic Structural Specialization for Underwater Salient Object Detection

DGX agent

arXiv:2605.15535v1 Announce Type: new Abstract: Underwater salient object detection (USOD) has attracted increasing attention for underwater visual scene understanding and vision-guided robotic applic

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning

DGX agent

arXiv:2605.15236v1 Announce Type: cross Abstract: With the coded caching, the server can use the information the users have cached to serve multiple users at a time by sending a single coded multi-cas

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

DGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

model-releasesarxiv-cs-ro
18 May 2026
← Previous
1…233234235236237…361
Next →