AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,841 results
18 May 2026

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

Model ReleasesDGX agent

arXiv:2509.12266v2 Announce Type: replace-cross Abstract: We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core c

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

Model ReleasesDGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

15 May 2026

A Large Language Model Based Pipeline for Review of Systems Entity Recognition from Clinical Notes

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2506.11067v3 Announce Type: replace Abstract: Objective: Develop a cost-effective, large language model (LLM)-based pipeline for automatically extracting Review of Systems (ROS) entities from cl

An Interpretable Latency Model for Speculative Decoding in LLM Serving

ApplicationsDGX agent

arXiv:2605.15051v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model (LLM) inference by using a smaller draft model to propose multiple tokens that are verified b

Darwin Family: MRI-Trust-Weighted Evolutionary Merging for Training-Free Scaling of Language-Model Reasoning

Model ReleasesDGX agent

arXiv:2605.14386v1 Announce Type: cross Abstract: We present Darwin Family, a framework for training-free evolutionary merging of large language models via gradient-free weight-space recombination. We

Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models

Model ReleasesDGX agent

arXiv:2511.08565v3 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly operate in social contexts, motivating analysis of how they express and shift moral judgments. In th

NeuroMambaLLM: Dynamic Graph Learning of fMRI Functional Connectivity in Autistic Brains Using Mamba and Language Model Reasoning

Model ReleasesDGX agent

arXiv:2602.13770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong semantic reasoning across multimodal domains. However, their integration with graph-base

On the Cultural Anachronism and Temporal Reasoning in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.15071v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied to cultural heritage materials, from digital archives to educational platforms. This work ident

Tokenizer Fertility and Zero-Shot Performance of Foundation Models on Ukrainian Legal Text: A Comparative Study

Model ReleasesDGX agent

arXiv:2605.14890v1 Announce Type: new Abstract: Foundation models tokenize Ukrainian legal text with vastly different efficiency, yet no systematic comparison exists for this domain. We benchmark seve

14 May 2026

AttenA+: Rectifying Action Inequality in Robotic Foundation Models

Model ReleasesDGX agent

arXiv:2605.13548v1 Announce Type: cross Abstract: Existing robotic foundation models, while powerful, are predicated on an implicit assumption of temporal homogeneity: treating all actions as equally

Behavioral Geometric Supervision Aligns Video Foundation Models with Human Social Perception

ResearchDGX agent

arXiv:2510.01502v2 Announce Type: replace-cross Abstract: Current video foundation models, including the strongest self-supervised models such as V-JEPA2, fail to capture how humans organize social in

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

Model ReleasesDGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

No One Knows the State of the Art in Geospatial Foundation Models

Model ReleasesDGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

Sample-Efficient Optimisation over the Outputs of Generative Models

ResearchDGX agent

arXiv:2509.23800v3 Announce Type: replace-cross Abstract: Modern generative AI models, such as diffusion and flow matching models, can sample from rich data distributions. However, many applications,

13 May 2026

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

ApplicationsDGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

Large-Small Model Collaboration for Farmland Semantic Change Detection

Model ReleasesDGX agent

arXiv:2605.12282v1 Announce Type: new Abstract: Farmland Semantic Change Detection (SCD) is essential for cultivated land protection, yet existing benchmarks and models remain insufficient for fine-gr

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

Model ReleasesDGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

Strategically Deceptive Model Deployment in Performative Prediction

TutorialsDGX agent

arXiv:2506.09044v2 Announce Type: replace Abstract: Machine Learning systems are increasingly deployed in decision-making settings that shape user behavior and, in turn, the data on which future decis

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

ResearchDGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

12 May 2026

Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing

ResearchDGX agent

arXiv:2605.10794v1 Announce Type: cross Abstract: Language models are deployed in settings that require compartmentalization: system prompts should not be disclosed, chain-of-thought reasoning is hidd

Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes

Model ReleasesDGX agent

arXiv:2605.09751v1 Announce Type: new Abstract: Trainable input embedding tables are a standard component of modern language models. We ask whether they are actually necessary at the input interface.

Layer Collapse in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.06366v2 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as competitive alternatives to autoregressive (AR) language models, yet differences in their

Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies

Model ReleasesDGX agent

arXiv:2605.09665v1 Announce Type: cross Abstract: Data selection is a key component of efficient instruction tuning for large language models, as recent work has shown that data quality often matters

Machine Unlearning on Pre-trained Models by Residual Feature Alignment Using LoRA

SafetyDGX agent

arXiv:2411.08443v2 Announce Type: replace-cross Abstract: Machine unlearning is an emerging technology that removes a subset of the training data from a trained model without significantly affecting t

Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds

Model ReleasesDGX agent

arXiv:2605.09724v1 Announce Type: new Abstract: Existing accounts of grokking explain the phenomena in terms of mechanistic frameworks such as circuit efficiency or lazy-to-rich transitions. However,

Models got an order of magnitude better at following instructions in one year

Model ReleasesDGX agent

A year ago, frontier models started losing track of instructions somewhere around 200–300 simultaneous constraints. With 2026 models, that ceiling is closer to 2,000 — an order-of-magnitude jump. We r

Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift

Model ReleasesDGX agent

arXiv:2505.19519v3 Announce Type: replace Abstract: Personalizing text-to-image diffusion models involves integrating novel visual concepts from a small set of reference images while retaining the mod

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

Model ReleasesDGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

Recursive Language Models

Model ReleasesDGX agent

arXiv:2512.24601v3 Announce Type: replace Abstract: We study allowing large language models (LLMs) to process arbitrarily long prompts through the lens of inference-time scaling. We propose Recursive

TripleWin: Fixed-Point Equilibrium Pricing for Data-Model Coupled Markets

SafetyDGX agent

arXiv:2511.03368v2 Announce Type: replace Abstract: The rise of the machine learning (ML) model economy has intertwined markets for training datasets and pre-trained models. However, most pricing appr

11 May 2026

Adapting Vision-Language Models for Neutrino Event Classification in High-Energy Physics

Model ReleasesDGX agent

arXiv:2509.08461v3 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated their remarkable capacity to process and reason over structured and unstruct

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

AgentsDGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

Detecting Distillation Data from Reasoning Models

Model ReleasesDGX agent

arXiv:2510.04850v3 Announce Type: replace-cross Abstract: Reasoning distillation has emerged as a prevailing paradigm for transferring reasoning capabilities from large reasoning models to small langu

Evaluating Large Language Models in Scientific Discovery

Model ReleasesDGX agent

arXiv:2512.15567v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and

How Big Should a Wireless Foundation Model Be?

Model ReleasesDGX agent

arXiv:2605.07266v1 Announce Type: cross Abstract: Wireless foundation models are rapidly emerging as a key enabler of AI-native communication systems, yet a fundamental question remains unanswered: ho

More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models

Model ReleasesDGX agent

arXiv:2605.06672v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning and reasoning-tuned models such as DeepSeek-R1 are commonly assumed to reduce shallow heuristic biases by thinking care

Post-training makes large language models less human-like

SafetyDGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

Query-efficient model evaluation using cached responses

Model ReleasesDGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

Model ReleasesDGX agent

arXiv:2601.23143v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve remarkable performance by leveraging reinforcement learning (RL) on reasoning tasks to generate long chain-of-

7 May 2026

Emergent Hierarchical Structure in Large Language Models: An Information-Theoretic Framework for Multi-Scale Representation

Model ReleasesDGX agent

arXiv:2505.18244v3 Announce Type: replace Abstract: Why do language models from different architecture families respond so differently to the same perturbation? We argue that the answer is not scale,

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

SafetyDGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

Gyan: An Explainable Neuro-Symbolic Language Model

ResearchDGX agent

arXiv:2605.04759v1 Announce Type: new Abstract: Transformer based pre-trained large language models have become ubiquitous. There is increasing evidence to suggest that even with large scale pre-train

Low-Rank Adaptation of Geospatial Foundation Models for Wildfire Mapping Using Sentinel-2 Data

Model ReleasesDGX agent

arXiv:2605.04989v1 Announce Type: new Abstract: Wildfire burned-area mapping is essential for damage assessment, emissions modeling, and understanding fire-climate interactions across diverse ecologic

Open-Source Image Editing Models Are Zero-Shot Vision Learners

Model ReleasesDGX agent

arXiv:2605.04566v1 Announce Type: cross Abstract: Recent studies have shown that large generative models can solve vision tasks they were not explicitly trained for. However, existing evidence relies

Probing Structural Mathematical Reasoning in Language Models with Algebraic Trapdoors

Model ReleasesDGX agent

arXiv:2605.04352v1 Announce Type: new Abstract: We introduce a benchmark suite for evaluating structural mathematical reasoning in language models, built on subgroup-construction problems in SL(3, Z)

Scalable Object Detection in the Car Interior With Vision Foundation Models

Model ReleasesDGX agent

arXiv:2508.19651v2 Announce Type: replace Abstract: AI tasks in the car interior like identifying and localizing externally introduced objects is crucial for response quality of personal assistants. H

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

Model ReleasesDGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

6 May 2026

A Benchmark for Interactive World Models with a Unified Action Generation Framework

Model ReleasesDGX agent

arXiv:2605.03941v1 Announce Type: new Abstract: Achieving Artificial General Intelligence (AGI) requires agents that learn and interact adaptively, with interactive world models providing scalable env

EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving

Model ReleasesDGX agent

arXiv:2509.17677v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on mathematical reasoning under well-defined conditions. However, real-world engineering

Expanding functional protein sequence space using high entropy generative models

Model ReleasesDGX agent

arXiv:2605.03578v1 Announce Type: cross Abstract: Boltzmann Machines trained on evolutionary sequence data have emerged as a powerful paradigm for the data-driven design of artificial proteins. Howeve

Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?

Model ReleasesDGX agent

arXiv:2605.00964v1 Announce Type: cross Abstract: Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI wor

5 May 2026

Agentopic: A Generative AI Agent Workflow for Explainable Topic Modeling

Model ReleasesDGX agent

arXiv:2605.00833v1 Announce Type: new Abstract: Agentopic is a novel agent-based workflow for explainable topic modeling that leverages the reasoning capabilities of Large Language Models (LLMs). Exis

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE

Model ReleasesDGX agent

arXiv:2605.02641v1 Announce Type: new Abstract: We present Mamoda2.5, a unified AR-Diffusion framework that seamlessly integrates multimodal understanding and generation within a single architecture.

Statistically-Lossless Quantization of Large Language Models

Model ReleasesDGX agent

arXiv:2605.02404v1 Announce Type: new Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as

TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs

Model ReleasesDGX agent

arXiv:2510.15545v4 Announce Type: replace Abstract: Accelerating the inference of large language models (LLMs) has been a critical challenge in generative AI. Speculative decoding (SD) substantially i

WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild

Model ReleasesDGX agent

arXiv:2605.01018v1 Announce Type: new Abstract: Using multimodal foundation models to analyze table images is a high-value yet challenging application in consumer and enterprise scenarios. Despite its

4 May 2026

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harne…

Model ReleasesDGX agent

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harness that's truly model-agnostic, without compromising perform

Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision

ResearchDGX agent

arXiv:2605.00644v1 Announce Type: new Abstract: Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. Howeve

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

Model ReleasesDGX agent

arXiv:2605.00689v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in cross-linguistic contexts, ensuring safety in diverse regulatory and cultural environments

Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring

Model ReleasesDGX agent

arXiv:2605.00754v1 Announce Type: cross Abstract: Reward models (RMs) have become an indispensable fixture of the language model (LM) post-training playbook, enabling policy alignment and test-time sc

← Previous
1…3031323334…998
Next →