AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
18 May 2026

Evaluating Chinese Ambiguity Understanding in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15635v1 Announce Type: new Abstract: Linguistic ambiguity is critical to the robustness of Large Language Models (LLMs), yet existing research focuses mostly on English, with limited attent

Fair outputs, Biased Internals: Causal Potency and Asymmetry of Latent Bias in LLMs for High-Stakes Decisions

Model ReleasesDGX agent

arXiv:2605.15217v1 Announce Type: new Abstract: Instruction-tuned language models exhibit behavioural fairness in high-stakes decisions while retaining biased associations in their internal representa

Fast-tracking genetic leads to reverse cellular aging

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Google DeepMind's Co-Scientist AI tool helps researchers identify genetic changes that push cells away from senescence toward youthful states in tissues like skin, hair, and muscle. The system scans s

Federated Imputation under Heterogeneous Feature Spaces

Model ReleasesDGX agent

arXiv:2605.16099v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative training across decentralized clients, but most methods assume aligned feature schemas, an assumption th

Federated Learning of Spiking Neural Networks under Heterogeneous Temporal Resolutions

Model ReleasesDGX agent

arXiv:2605.15355v1 Announce Type: new Abstract: Spiking neural networks (SNNs) are biologically inspired energy-efficient models that use sparse binary spike-based communication between neurons, makin

Feedback World Model Enables Precise Guidance of Diffusion Policy

Model ReleasesDGX agent

arXiv:2605.15705v1 Announce Type: cross Abstract: World models aim to improve robotic decision making by predicting the consequences of actions. However, in practice, their predictions often become un

Few-Shot Large Language Models for Actionable Triage Categorization of Online Patient Inquiries

Model ReleasesDGX agent

arXiv:2605.15680v1 Announce Type: new Abstract: Online patient inquiries are often informal, incomplete, and written before professional assessment, yet they must still be routed to an appropriate lev

FFAvatar: Few-Shot, Feed-Forward, and Generalizable Avatar Reconstruction

Model ReleasesDGX agent

arXiv:2605.15320v1 Announce Type: cross Abstract: Avatar reconstruction has traditionally relied on per-subject optimization that requires hours of computation or on expensive preprocessing that limit

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization

Model ReleasesDGX agent

arXiv:2605.15980v1 Announce Type: new Abstract: Group Relative Policy Optimization has emerged as essential for aligning video diffusion models with human preferences, but faces a critical computation

FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast

Model ReleasesDGX agent

arXiv:2605.16233v1 Announce Type: new Abstract: Can LLM agents improve decision-making through self-generated memory without gradient updates? We propose FORGE (Failure-Optimized Reflective Graduation

ForMaT: Dataset for Visually-Grounded Multilingual PDF Translation

Model ReleasesDGX agent

arXiv:2605.15794v1 Announce Type: new Abstract: We present ForMaT (Format-Preserving Multilingual Translation), a parallel corpus of 3,956 PDFs across 15 language pairs that preserves original layout

FormulaCode: Evaluating Agentic Optimization on Large Codebases

Model ReleasesDGX agent

arXiv:2603.16011v2 Announce Type: replace-cross Abstract: Large language model (LLM) coding agents increasingly operate at the repository level, motivating benchmarks that evaluate their ability to op

From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery

Model ReleasesDGX agent

arXiv:2605.15412v1 Announce Type: cross Abstract: Modern quantitative trading increasingly relies on systematic models to extract predictive signals from large-scale financial data, where alpha factor

From Layers to Networks: Comparing Neural Representations via Diffusion Geometry

Model ReleasesDGX agent

arXiv:2605.15901v1 Announce Type: new Abstract: Diffusion geometry is a manifold learning framework that uses random walks defined by Markov transition matrices to characterize the geometry of a datas

Frontier Large Language Models Rival State-of-the-Art Planners

Model ReleasesDGX agent

arXiv:2511.09378v2 Announce Type: replace Abstract: A series of influential studies established that large language models cannot reliably solve even simple planning tasks. We show that the latest gen

FRWKV+: Adaptive Periodic-Position Branch Interaction for Frequency-Space Linear Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.15690v1 Announce Type: new Abstract: Long-term time series forecasting is essential for decision making in energy, finance, transportation, and healthcare systems. Recent lightweight foreca

Fully Open Meditron: An Auditable Pipeline for Clinical LLMs

Model ReleasesDGX agent

arXiv:2605.16215v1 Announce Type: new Abstract: Clinical decision support systems (CDSS) require scrutable, auditable pipelines that enable rigorous, reproducible validation. Yet current LLM-based CDS

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

Model ReleasesDGX agent

arXiv:2509.12266v2 Announce Type: replace-cross Abstract: We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core c

GenShield: Unified Detection and Artifact Correction for AI-Generated Images

Model ReleasesDGX agent

arXiv:2605.16122v1 Announce Type: cross Abstract: Diffusion-based image synthesis has made AI-generated images (AIGI) increasingly photorealistic, raising urgent concerns about authenticity in applica

GESD: Beyond Outcome-Oriented Fairness

Model ReleasesDGX agent

arXiv:2605.15295v1 Announce Type: cross Abstract: Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predi

GiLT: Augmenting Transformer Language Models with Dependency Graphs

Model ReleasesDGX agent

arXiv:2605.15562v1 Announce Type: new Abstract: Augmenting Transformers with linguistic structures effectively enhances the syntactic generalization performance of language models. Previous work in th

Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis

Model ReleasesDGX agent

arXiv:2602.20207v3 Announce Type: replace-cross Abstract: Knowledge editing in Large Language Models (LLMs) aims to update the model's prediction for a specific query to a desired target while preserv

GQA-{mu}P: The maximal parameterization update for grouped query attention

Model ReleasesDGX agent

arXiv:2605.15290v1 Announce Type: cross Abstract: Hyperparameter transfer across model architectures dramatically reduces the amount of compute necessary for tuning large language models (LLMs). The m

GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding

Model ReleasesDGX agent

arXiv:2605.15250v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), the attention used in DeepSeek-V2/V3, jointly compresses keys and values into a low-rank latent and matches the H10

Graph-Regularized Sparse Autoencoders for LLM Safety Steering

Model ReleasesDGX agent

arXiv:2512.06655v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) are increasingly used to extract activation directions for inference-time steering, but their standard sparsity obj

GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero

Model ReleasesDGX agent

arXiv:2605.15464v1 Announce Type: cross Abstract: Post-training has become a crucial step for unlocking the capabilities of large language models, with reinforcement learning (RL) emerging as a critic

HAI-Eval: Measuring Human-AI Synergy in Collaborative Coding

Model ReleasesDGX agent

arXiv:2512.04111v2 Announce Type: replace-cross Abstract: LLM-powered coding agents are reshaping the development paradigm. However, existing evaluation systems, neither traditional tests for humans n

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

Model ReleasesDGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

Harnessing Unimodality in Semiparametric Contextual Pricing via Oracle Price Map Learning

Model ReleasesDGX agent

arXiv:2605.15411v1 Announce Type: cross Abstract: We study contextual dynamic pricing in a semiparametric scalar-index valuation model where the latent value is v_t=mu_ast(mathsf c_t)+xi_t, with an un

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

Model ReleasesDGX agent

arXiv:2605.15338v1 Announce Type: cross Abstract: Large language models are increasingly augmented with persistent memory, allowing assistants to store user-specific information across sessions for pe

Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces

Model ReleasesDGX agent

arXiv:2605.15753v1 Announce Type: cross Abstract: Functional 3D scene graphs offer a versatile and flexible representation for 3D scene understanding and robotic manipulation, defined by object nodes,

How do you know your document parser is ready for production? 🤔 Existing benchmarks miss what AI agents actually need. That's the gap Parse…

Model ReleasesDGX agent

How do you know your document parser is ready for production? 🤔 Existing benchmarks miss what AI agents actually need. That's the gap ParseBench, the first doc OCR benchmark for AI agents, fills. We'l

How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models

Model ReleasesDGX agent

arXiv:2510.02453v3 Announce Type: replace-cross Abstract: Frontier language models are deployed as black-box services, where model weights cannot be modified and customization is limited to prompting.

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

Model ReleasesDGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

I love AI, it’s pure LLMs I hate. Pure LLMs *are* basically just autocomplete. Recent progress (e.g. Claude Code) doesn’t show otherwise Rat…

Model ReleasesDGX agent

I love AI, it’s pure LLMs I hate. Pure LLMs *are* basically just autocomplete. Recent progress (e.g. Claude Code) doesn’t show otherwise Rather, lot of the progress in the last two years has come from

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

Model ReleasesDGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

Model ReleasesDGX agent

arXiv:2603.17915v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages rem

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

Model ReleasesDGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

Information-Preserving Domain Transfer with Unlabeled Data in Misspecified Simulation-Based Inference

Model ReleasesDGX agent

arXiv:2605.05652v2 Announce Type: replace Abstract: Simulation-based inference (SBI) provides amortized Bayesian parameter inference from simulator-generated data without requiring explicit likelihood

Interaction-Aware Influence Functions for Group Attribution

Model ReleasesDGX agent

arXiv:2605.15675v1 Announce Type: cross Abstract: Influence functions approximate how removing a training example changes a quantity of interest, called the target function, such as a held-out loss. T

Introducing MELI: the Mandarin-English Language Interview Corpus

Model ReleasesDGX agent

arXiv:2603.27043v2 Announce Type: replace Abstract: We introduce the Mandarin-English Language Interview (MELI) Corpus, an open-source resource of 29.8 hours of speech from 51 Mandarin-English bilingu

Iranian media: Iran launches 'Hormuz Safe', a Bitcoin-backed insurance service for shipping companies transiting the Strait of Hormuz; ~1,500 ships are trapped (Golnar Motevalli/Bloomberg)

Model ReleasesDGX agent

Golnar Motevalli / Bloomberg: Iranian media: Iran launches “Hormuz Safe”, a Bitcoin-backed insurance service for shipping companies transiting the Strait of Hormuz; ~1,500 ships are trapped — Iran has

Judge Circuits

Model ReleasesDGX agent

arXiv:2605.16023v1 Announce Type: new Abstract: LLM-as-a-judge has become the dominant paradigm for grading model outputs at scale, yet the same model assigns systematically different scores when its

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

Model ReleasesDGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

Large Language Models Could Be Rote Learners

Model ReleasesDGX agent

arXiv:2504.08300v5 Announce Type: replace-cross Abstract: Benchmark-based evaluation, e.g., multiple-choice questions (MCQs) and open-ended questions (OEQs), is widely used for evaluating Large Langua

LASER: Language Model Regression for Semi-Structured Workflow Resource and Runtime Estimation

Model ReleasesDGX agent

arXiv:2512.19701v2 Announce Type: replace-cross Abstract: Accurate prediction of resource consumption and runtime for cloud workflow jobs is critical for scheduling efficiency, yet remains challenging

Layer Equivalence Is Not a Property of Layers Alone: How You Test Redundancy Changes What You Find

Model ReleasesDGX agent

arXiv:2605.16234v1 Announce Type: cross Abstract: When researchers ask whether two transformer layers are 'equivalent' for compression, they often conflate distinct tests. Replacement asks whether one

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

Model ReleasesDGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

Learning Disentangled Representations for Generalized Multi-view Clustering

Model ReleasesDGX agent

arXiv:2605.15640v1 Announce Type: new Abstract: Multi-View Clustering (MVC) has gained significant attention for its ability to leverage complementary information across diverse views. However, existi

Learning Dynamic Structural Specialization for Underwater Salient Object Detection

Model ReleasesDGX agent

arXiv:2605.15535v1 Announce Type: new Abstract: Underwater salient object detection (USOD) has attracted increasing attention for underwater visual scene understanding and vision-guided robotic applic

Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.15236v1 Announce Type: cross Abstract: With the coded caching, the server can use the information the users have cached to serve multiple users at a time by sending a single coded multi-cas

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

Model ReleasesDGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance jump with these change…

Model ReleasesDGX agent

llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance jump with these changes is massive and elevates local inference on commodity hardw

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s →…

Model ReleasesDGX agent

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s → 45 tok/s (+78%). Two flags on llama-server: --spec-type draf

LLM-EDT: Large Language Model Enhanced Cross-domain Sequential Recommendation with Dual-phase Training

Model ReleasesDGX agent

arXiv:2511.19931v2 Announce Type: replace-cross Abstract: Cross-domain Sequential Recommendation (CDSR) has been proposed to enrich user-item interactions by incorporating information from various dom

LoCO: Low-rank Compositional Rotation Fine-tuning

Model ReleasesDGX agent

arXiv:2605.15916v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as an critical technique for adapting large-scale foundation models across natural language process

Long Range Frequency Tuning for QML

Model ReleasesDGX agent

arXiv:2602.23409v2 Announce Type: replace-cross Abstract: Angle-encoded variational quantum circuits admit a truncated Fourier series representation of their output, but approximating functions with m

Looped SSMs: Depth-Recurrence and Input Reshaping for Time Series Classification

Model ReleasesDGX agent

arXiv:2605.16048v1 Announce Type: cross Abstract: State Space Models (SSMs) are inherently recurrent along the sequence dimension, yet depth-recurrence - reusing the same block repeatedly across layer

Mask-Morph Graph U-Net: A Generalisable Mesh-Based Surrogate for Crashworthiness Field Prediction under Large Geometric Variation

Model ReleasesDGX agent

arXiv:2605.15231v1 Announce Type: cross Abstract: Nonlinear finite element crash simulations are accurate but computationally expensive, limiting their use in iterative design optimisation. Machine-le

← Previous
1…242243244245246…377
Next →