AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
17 Apr 2026

AnimationBench: Are Video Models Good at Character-Centric Animation?

Model ReleasesDGX agent

arXiv:2604.15299v1 Announce Type: new Abstract: Video generation has advanced rapidly, with recent methods producing increasingly convincing animated results. However, existing benchmarks-largely desi

Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models

ResearchDGX agent

arXiv:2601.21003v2 Announce Type: replace Abstract: Large Language Models usually put more emphasis on accuracy and therefore, will guess even when not certain about the prediction, which is especiall

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

ResearchDGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contest

ResearchDGX agent

arXiv:2510.24538v2 Announce Type: replace Abstract: Textual humor is enormously diverse and computational studies need to account for this range, including intentionally bad humor. In this paper, we c

Exploration and Exploitation Errors Are Measurable for Language Model Agents

SafetyDGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08064v2 Announce Type: replace Abstract: Existing memory benchmarks for LLM agents evaluate explicit recall of facts, yet overlook implicit memory where experience becomes automated behavio

IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation

ApplicationsDGX agent

arXiv:2604.15109v1 Announce Type: new Abstract: Despite the rapid advancement of Large Language Models (LLMs), uncertainty quantification in LLM generation is a persistent challenge. Although recent a

Keep It CALM: Toward Calibration-Free Kilometer-Level SLAM with Visual Geometry Foundation Models via an Assistant Eye

Local AiDGX agent

arXiv:2604.14795v1 Announce Type: new Abstract: Visual Geometry Foundation Models (VGFMs) demonstrate remarkable zero-shot capabilities in local reconstruction. However, deploying them for kilometer-l

Model-Free Assessment of Simulator Fidelity via Quantile Curves

SafetyDGX agent

arXiv:2512.05024v3 Announce Type: replace-cross Abstract: As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

SafetyDGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

SafetyDGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

16 Apr 2026

Alibaba's new Token Hub unit releases Happy Oyster, a new AI world model that can create 3D environments, interactive videos, films, video content, and games (Luz Ding/Bloomberg)

ApplicationsDGX agent

Luz Ding / Bloomberg: Alibaba's new Token Hub unit releases Happy Oyster, a new AI world model that can create 3D environments, interactive videos, films, video content, and games — Alibaba Group Hold

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

SafetyDGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models

AgentsDGX agent

arXiv:2604.13706v1 Announce Type: new Abstract: Professional fact-checkers rely on domain knowledge and deep contextual understanding to verify claims. Large language models (LLMs) and large reasoning

Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models

ResearchDGX agent

arXiv:2604.14129v1 Announce Type: new Abstract: While Audio-Visual Language Models (AVLMs) have achieved remarkable progress over recent years, their reliability is bottlenecked by cross-modal halluci

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints

ResearchDGX agent

arXiv:2604.13371v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly described as possessing strong reasoning capabilities, supported by high performance on mathematical, logi

Foresight Optimization for Strategic Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

Local AiDGX agent

arXiv:2604.13777v1 Announce Type: new Abstract: Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

Model ReleasesDGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

Great blogpost from @pcuenq on making a new skill + test harness to automate porting new models from Transformers to mlx-lm

Local AiDGX agent

This post discusses a blog article by @pcuenq that covers the process of creating new skills and test harnesses to automate the conversion of machine learning models from the Hugging Face Transformers

How Can We Synthesize High-Quality Pretraining Data? A Systematic Study of Prompt Design, Generator Model, and Source Data

ResearchDGX agent

arXiv:2604.13977v1 Announce Type: new Abstract: Synthetic data is a standard component in training large language models, yet systematic comparisons across design dimensions, including rephrasing stra

Introducing GPT-Rosalind, our frontier reasoning model built to support research across biology, drug discovery, and translational medicine.

IndustryDGX agent

GPT-Rosalind is a frontier reasoning model developed by OpenAI, specifically designed to support research applications in biology, drug discovery, and translational medicine. The model was introduced

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

ResearchDGX agent

arXiv:2604.13386v1 Announce Type: new Abstract: Linear probes can detect when language models produce outputs they 'know' are wrong, a capability relevant to both deception and reward hacking. However

Modeling Student Learning with 3.8 Million Program Traces

TutorialsDGX agent

arXiv:2510.05056v2 Announce Type: replace Abstract: As programmers write code, they often edit and retry multiple times, creating rich 'interaction traces' that reveal how they approach coding tasks a

Native Hybrid Attention for Efficient Sequence Modeling

ResearchDGX agent

arXiv:2510.07019v3 Announce Type: replace Abstract: Transformers excel at sequence modeling but face quadratic complexity, while linear attention offers improved efficiency but often compromises recal

Quantifying and Understanding Uncertainty in Large Reasoning Models

ResearchDGX agent

arXiv:2604.13395v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have recently demonstrated significant improvements in complex reasoning. While quantifying generation uncertainty in LR

Sign up today @ http://portal.nousresearch.com/manage-subscription Run `hermes update` and then `hermes model` to configure

ResearchDGX agent

Nous Research is promoting their subscription portal and providing instructions for users to update and configure their Hermes model through command-line tools (`hermes update` and `hermes model`). Th

15 Apr 2026

CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models

ResearchDGX agent

arXiv:2604.12767v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) suffer from substantial computational overhead due to the high redundancy in visual token sequences. Existing

Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling

ResearchDGX agent

arXiv:2604.12777v1 Announce Type: cross Abstract: The human brain constructs emotional percepts not by processing facial expressions in isolation, but through a dynamic, hierarchical integration of se

CycloneMAE: A Scalable Multi-Task Learning Model for Global Tropical Cyclone Probabilistic Forecasting

ResearchDGX agent

arXiv:2604.12180v1 Announce Type: cross Abstract: Tropical cyclones (TCs) rank among the most destructive natural hazards, yet their forecasting faces fundamental trade-offs: numerical weather predict

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

ResearchDGX agent

arXiv:2604.05546v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) enable sophisticated reasoning over images and videos, yet their inference is hindered by a systemic efficiency

Evaluating Language Models for Harmful Manipulation

Local AiDGX agent

arXiv:2603.25326v4 Announce Type: replace Abstract: Interest in the concept of AI-driven harmful manipulation is growing, yet current approaches to evaluating it are limited. This paper introduces a f

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

Model ReleasesDGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

How much more useage do you get from a $20 pro plan when using cloud models? Or OpenRouter better??

Local AiDGX agent

This Reddit thread from r/ollama discusses the value comparison between Ollama's 20/month Pro plan for cloud model usage versus using OpenRouter as an alternative. Ollama Cloud offers fixed-price subs

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

SafetyDGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

LLM-Enhanced Log Anomaly Detection: A Comprehensive Benchmark of Large Language Models for Automated System Diagnostics

Model ReleasesDGX agent

arXiv:2604.12218v1 Announce Type: new Abstract: System log anomaly detection is critical for maintaining the reliability of large-scale software systems, yet traditional methods struggle with the hete

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

Local AiDGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

Mantis: A Foundation Model for Mechanistic Disease Forecasting

ApplicationsDGX agent

arXiv:2508.12260v5 Announce Type: replace Abstract: Infectious disease forecasting in novel outbreaks or low-resource settings is hampered by the need for large disease and covariate data sets, bespok

Microsoft’s MAI-Image-2-Efficient model accelerates company’s move away from OpenAI

IndustryDGX agent

Microsoft Corp.’s push for artificial intelligence independence is gaining traction with today’s release of MAI-Image-2-Efficient, a lean and mean version of its flagship image generation model that d

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

Model ReleasesDGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints

SafetyDGX agent

arXiv:2604.12384v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) remains highly fragile during fine-tuning, where even benign adaptation can degrade pre-trained refusal

Revisiting the Reliability of Language Models in Instruction-Following

Model ReleasesDGX agent

arXiv:2512.14754v2 Announce Type: replace-cross Abstract: Advanced LLMs have achieved near-ceiling instruction-following accuracy on benchmarks such as IFEval. However, these impressive scores do not

Robotic Manipulation is Vision-to-Geometry Mapping (f(v) rightarrow G): Vision-Geometry Backbones over Language and Video Models

ApplicationsDGX agent

arXiv:2604.12908v1 Announce Type: new Abstract: At its core, robotic manipulation is a problem of vision-to-geometry mapping (f(v) rightarrow G). Physical actions are fundamentally defined by geometri

Speaker effects in language comprehension: An integrative model of language and speaker processing

ResearchDGX agent

arXiv:2412.07238v3 Announce Type: replace Abstract: The identity of a speaker influences language comprehension through modulating perception and expectation. This review explores speaker effects and

TeRA: Vector-based Random Tensor Network for High-Rank Adaptation of Large Language Models

Model ReleasesDGX agent

arXiv:2509.03234v2 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods, such as Low-Rank Adaptation (LoRA), have significantly reduced the number of trainable parameters ne

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training

SafetyDGX agent

arXiv:2509.25758v2 Announce Type: replace Abstract: The remarkable capabilities of modern large reasoning models are largely unlocked through post-training techniques such as supervised fine-tuning (S

Training code and models are live on Hugging Face. Dan Fu (Together AI's VP of Kernels) led the work. Together AI provided compute. Blog: ht…

ToolsDGX agent

Training code and models are live on Hugging Face. Dan Fu (Together AI's VP of Kernels) led the work. Together AI provided compute. Blog: https://www.together.ai/blog/parcae Paper: https://arxiv.org/a

World models are soo March 2026. What's your AllBirds strategy is what matters today. How are you competing against AllBirds? Are you buildi…

IndustryDGX agent

Cristobal Valenzuela, co-founder and CEO of Runway, posted a social media comment suggesting that 'world models' — a major AI topic of early 2026 — are already becoming passé, using the footwear brand

14 Apr 2026

A Survey of Inductive Reasoning for Large Language Models

ResearchDGX agent

arXiv:2510.10182v2 Announce Type: replace-cross Abstract: Reasoning is an important task for large language models (LLMs). Among all the reasoning paradigms, inductive reasoning is one of the fundamen

AdaQE-CG: Adaptive Query Expansion for Web-Scale Generative AI Model and Data Card Generation

Model ReleasesDGX agent

arXiv:2604.09617v1 Announce Type: new Abstract: Transparent and standardized documentation is essential for building trustworthy generative AI (GAI) systems. However, existing automated methods for ge

Asymptotic Learning Curves for Diffusion Models with Random Features Score and Manifold Data

TutorialsDGX agent

arXiv:2603.22962v2 Announce Type: replace Abstract: We study the theoretical behavior of denoising score matching--the learning task associated to diffusion models--when the data distribution is suppo

Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation

ApplicationsDGX agent

arXiv:2604.10950v1 Announce Type: new Abstract: Fully supervised Video Semantic Segmentation (VSS) relies heavily on densely annotated video data, limiting practical applicability. Alternatively, appl

Breaking the KV Cache Bottleneck: Fan Duality Model Achieves O(1) Decode Memory with Superior Associative Recall

ResearchDGX agent

arXiv:2604.07716v2 Announce Type: replace Abstract: We present FDM (Fan Duality Model), a linear sequence architecture that resolves the fundamental tension between memory efficiency and associative r

Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games

ResearchDGX agent

arXiv:2603.22479v3 Announce Type: replace-cross Abstract: Defining a constructive process to build general capabilities for language models in an automatic manner is considered an open problem in arti

Computational Implementation of a Model of Category-Theoretic Metaphor Comprehension

ResearchDGX agent

arXiv:2604.10035v1 Announce Type: cross Abstract: In this study, we developed a computational implementation for a model of metaphor comprehension based on the theory of indeterminate natural transfor

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Models

ResearchDGX agent

arXiv:2604.10681v1 Announce Type: cross Abstract: Large Language Models (LLMs), despite their impressive capabilities across domains, have been shown to be vulnerable to backdoor attacks. Prior backdo

CROP: Conservative Reward for Model-based Offline Policy Optimization

SafetyDGX agent

arXiv:2310.17245v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) aims to optimize a policy using collected data without online interactions. Model-based approaches are par

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

HardwareDGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

Delving Aleatoric Uncertainty in Medical Image Segmentation via Vision Foundation Models

ResearchDGX agent

arXiv:2604.10963v1 Announce Type: new Abstract: Medical image segmentation supports clinical workflows by precisely delineating anatomical structures and lesions. However, medical image datasets medic

← Previous
1…143144145146147…1010
Next →