AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,860 results
20 Apr 2026

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens

ResearchDGX agent

arXiv:2602.12370v2 Announce Type: replace Abstract: Recent progress in large models has led to significant advances in unified multimodal generation and understanding. However, the development of mode

Moonshot AI releases Kimi-K2.6 model with 1T parameters, attention optimizations

Model ReleasesDGX agent

Moonshot AI today released Kimi-K2.6, the latest addition to its popular Kimi series of open-source large language models. The Chinese artificial intelligence startup says that the algorithm outperfor

No Universal Courtesy: A Cross-Linguistic, Multi-Model Study of Politeness Effects on LLMs Using the PLUM Corpus

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.16275v1 Announce Type: new Abstract: This paper explores the response of Large Language Models (LLMs) to user prompts with different degrees of politeness and impoliteness. The Politeness T

Reward Modeling for Scientific Writing Evaluation

ResearchDGX agent

arXiv:2601.11374v2 Announce Type: replace Abstract: Scientific writing is an expert-domain task that demands deep domain knowledge, task-specific requirements and reasoning capabilities that leverage

Sketching the Readout of Large Language Models for Scalable Data Attribution and Valuation

ApplicationsDGX agent

arXiv:2604.16197v1 Announce Type: new Abstract: Data attribution and valuation are critical for understanding data-model synergy for Large Language Models (LLMs), yet existing gradient-based methods s

18 Apr 2026

Some Mac Mini and Mac Studio models are unavailable or facing up to 12-week wait times in the US, with analysts citing strong demand from AI agent power users (Nicole Nguyen/Wall Street Journal)

AgentsDGX agent

Nicole Nguyen / Wall Street Journal: Some Mac Mini and Mac Studio models are unavailable or facing up to 12-week wait times in the US, with analysts citing strong demand from AI agent power users — Th

17 Apr 2026

AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime

Local AiDGX agent

arXiv:2604.14661v1 Announce Type: cross Abstract: Edge AI model deployment is a multi-stage engineering process involving model conversion, operator compatibility handling, quantization calibration, r

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

Model ReleasesDGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

IMPACTX: improving model performance by appropriately constraining the training with teacher explanations

ResearchDGX agent

arXiv:2502.12222v2 Announce Type: replace Abstract: The eXplainable Artificial Intelligence (XAI) research predominantly concentrates to provide explainations about AI model decisions, especially Deep

Label-efficient underwater species classification with logistic regression on frozen foundation model embeddings

Model ReleasesDGX agent

arXiv:2604.00313v2 Announce Type: replace Abstract: Automated species classification from underwater imagery is bottlenecked by the cost of expert annotation, and supervised models trained on one data

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

SafetyDGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

To go deeper on our new Life Sciences model series, research lead @joyjiao12 and product lead Yunyun Wang joined @AndrewMayne on the OpenAI …

AgentsDGX agent

To go deeper on our new Life Sciences model series, research lead @joyjiao12 and product lead Yunyun Wang joined @AndrewMayne on the OpenAI Podcast to discuss how we’re building models for biology, dr

Towards Faster Language Model Inference Using Mixture-of-Experts Flow Matching

ResearchDGX agent

arXiv:2604.15009v1 Announce Type: cross Abstract: Flow matching retains the generation quality of diffusion models while enabling substantially faster inference, making it a compelling paradigm for ge

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just …

Model ReleasesDGX agent

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just 0.10–0.20/hour, crushing most competitors. Quiet release, loud

16 Apr 2026

Can you use Ollama models with the Codex app on Windows?

Local AiDGX agent

Yes, Ollama models can be used with the Codex app on Windows. Ollama supports all major operating systems, including Windows , and open models can be used with OpenAI's Codex CLI through Ollama — Code

Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

Model ReleasesDGX agent

arXiv:2405.19088v3 Announce Type: replace Abstract: Recent advancements in large multimodal language models have demonstrated remarkable proficiency across a wide range of tasks. Yet, these models sti

Diffusion Sequence Models for Generative In-Context Meta-Learning of Robot Dynamics

ResearchDGX agent

arXiv:2604.13366v1 Announce Type: new Abstract: Accurate modeling of robot dynamics is essential for model-based control, yet remains challenging under distributional shifts and real-time constraints.

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

ResearchDGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

Indexing Multimodal Language Models for Large-scale Image Retrieval

ResearchDGX agent

arXiv:2604.13268v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong cross-modal reasoning capabilities, yet their potential for vision-only tasks remain

Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models

ResearchDGX agent

arXiv:2601.11340v2 Announce Type: replace Abstract: Chain-of-Thought reasoning has significantly enhanced the problem-solving capabilities of Large Language Models. Unfortunately, current models gener

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

Model ReleasesDGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

15 Apr 2026

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

Model ReleasesDGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

Climate Model Tuning with Online Synchronization-Based Parameter Estimation

Model ReleasesDGX agent

arXiv:2510.06180v2 Announce Type: replace-cross Abstract: In climate science, the tuning of climate models is a computationally intensive problem due to the combination of the high-dimensionality of t

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

Model ReleasesDGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

Fast AI Model Partition for Split Learning over Edge Networks

HardwareDGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models

ResearchDGX agent

arXiv:2601.19917v2 Announce Type: replace Abstract: Strategic planning is critical for multi-step reasoning, yet compact Large Language Models (LLMs) often lack the capacity to formulate global strate

Retrievals Can Be Detrimental: Unveiling the Backdoor Vulnerability of Retrieval-Augmented Diffusion Models

ResearchDGX agent

arXiv:2501.13340v4 Announce Type: replace Abstract: Diffusion models (DMs) have recently demonstrated remarkable generation capability. However, their training generally requires huge computational re

Task Alignment: A simple and effective proxy for model merging in computer vision

SafetyDGX agent

arXiv:2604.12935v1 Announce Type: new Abstract: Efficiently merging several models fine-tuned for different tasks, but stemming from the same pretrained base model, is of great practical interest. Des

When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models

Model ReleasesDGX agent

arXiv:2604.12128v1 Announce Type: new Abstract: We investigate how self-referential inputs alter the internal matrix dynamics of large language models. Measuring 106 scalar metrics across up to 7 anal

14 Apr 2026

A Mechanistic Analysis of Looped Reasoning Language Models

TutorialsDGX agent

arXiv:2604.11791v1 Announce Type: cross Abstract: Reasoning has become a central capability in large language models. Recent research has shown that reasoning performance can be improved by looping an

Abliterated (uncensored) models

Local AiDGX agent

This r/ollama discussion covers 'abliterated' models — LLMs that have had their built-in refusal mechanisms removed through a technique called abliteration, allowing them to respond to prompts without

ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models

TutorialsDGX agent

arXiv:2511.18082v3 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have shown impressive flexibility and generalization, yet their deployment in robotic manipulation remain

Ambiguity Detection and Elimination in Automated Executable Process Modeling

Local AiDGX agent

arXiv:2604.10884v1 Announce Type: cross Abstract: Automated generation of executable Business Process Model and Notation (BPMN) models from natural-language specifications is increasingly enabled by l

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

TutorialsDGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

Everyone I know is switching over to hermes agent, in large part because it actually works with smaller open source models.

AgentsDGX agent

Nous Research's Hermes agent framework has been gaining significant adoption due to its compatibility and effectiveness with smaller open-source language models, making it accessible beyond large prop

One-click LM Studio → Ollama model linker

Local AiDGX agent

This r/ollama post discusses a tool for easily linking models between LM Studio and Ollama without duplicating disk storage. Both Ollama and LM Studio are popular local LLM tools, but they store their

Powerful Training-Free Membership Inference Against Autoregressive Language Models

Model ReleasesDGX agent

arXiv:2601.12104v2 Announce Type: replace-cross Abstract: Fine-tuned language models pose significant privacy risks, as they may memorize and expose sensitive information from their training data. Mem

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models

ResearchDGX agent

arXiv:2604.10949v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) were designed to combine the reasoning ability of large language models (LLMs) with the generation capability of visi

RAM guide: What model combinations actually fit on common Macs

TutorialsDGX agent

This r/ollama community guide provides a practical breakdown of which AI language models (and combinations of models) can realistically fit within the unified memory constraints of common Apple Silico

Towards Brain MRI Foundation Models for the Clinic: Findings from the FOMO25 Challenge

ResearchDGX agent

arXiv:2604.11679v1 Announce Type: new Abstract: Clinical deployment of automated brain MRI analysis faces a fundamental challenge: clinical data is heterogeneous and noisy, and high-quality labels are

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

13 Apr 2026

Act or Escalate? Evaluating Escalation Behavior in Automation with Language Models

SafetyDGX agent

arXiv:2604.08588v1 Announce Type: cross Abstract: Effective automation hinges on deciding when to act and when to escalate. We model this as a decision under uncertainty: an LLM forms a prediction, es

any decent model to run on 9070xt locally

Local AiDGX agent

This Reddit thread from r/ollama discusses recommendations for running local AI models on the AMD Radeon RX 9070 XT using Ollama, touching on GPU compatibility considerations given that the card uses

Balancing User Preferences by Social Networks: A Condition-Guided Social Recommendation Model for Mitigating Popularity Bias

SafetyDGX agent

arXiv:2405.16772v2 Announce Type: replace-cross Abstract: Social recommendation models weave social interactions into their design to provide uniquely personalized recommendation results for users. Ho

CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion

ResearchDGX agent

arXiv:2604.09101v1 Announce Type: cross Abstract: Organisations with limited data and computational resources increasingly outsource model training to Machine Learning as a Service (MLaaS) providers,

Constraining Sequential Model Editing with Editing Anchor Compression

Model ReleasesDGX agent

arXiv:2503.00035v2 Announce Type: replace-cross Abstract: Large language models (LLMs) struggle with hallucinations due to false or outdated knowledge. Given the high resource demands of retraining th

dnaHNet: A Scalable and Hierarchical Foundation Model for Genomic Sequence Learning

ResearchDGX agent

arXiv:2602.10603v3 Announce Type: replace Abstract: Genomic foundation models have the potential to decode DNA syntax, yet face a fundamental tradeoff in their input representation. Standard fixed-voc

FluidFlow: a flow-matching generative model for fluid dynamics surrogates on unstructured meshes

Model ReleasesDGX agent

arXiv:2604.08586v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) provides high-fidelity simulations of fluid flows but remains computationally expensive for many-query applications

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing

Model ReleasesDGX agent

arXiv:2604.08884v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have made significant strides in natural image understanding, their ability to perceive and reason over

How to find the sweet spot between cost and performance

Model ReleasesDGX agent

At Google Cloud, we often see customers asking themselves: 'How can we manage our generative AI costs effectively without sacrificing the performance and availability our applications demand?' This is

Improving Model Performance by Adapting the KGE Metric to Account for System Non-Stationarity

Model ReleasesDGX agent

arXiv:2604.03906v2 Announce Type: replace Abstract: Geoscientific systems tend to be characterized by pronounced temporal non-stationarity, arising from seasonal and climatic variability in hydrometeo

Temperature-Dependent Performance of Prompting Strategies in Extended Reasoning Large Language Models

Model ReleasesDGX agent

arXiv:2604.08563v1 Announce Type: cross Abstract: Extended reasoning models represent a transformative shift in Large Language Model (LLM) capabilities by enabling explicit test-time computation for c

The model is not the agent. The harness is. You need to read this recent study, and a blog post from @hwchase17 ... (links below). It will r…

Model ReleasesDGX agent

The model is not the agent. The harness is. You need to read this recent study, and a blog post from @hwchase17 ... (links below). It will resonate deeply. This diagram from a recent paper captures so

Try it immediately on Ollama's cloud: ollama run gemma4:31b-cloud Try it with OpenClaw: ollama launch openclaw --model gemma4:31b-cloud Try …

Model ReleasesDGX agent

Try it immediately on Ollama's cloud: ollama run gemma4:31b-cloud Try it with OpenClaw: ollama launch openclaw --model gemma4:31b-cloud Try it with Claude Code: ollama launch claude --model gemma4:31b

12 Apr 2026

I think you should build your own harness too. Build on primitives. Models come and go.

AgentsDGX agent

Harrison Chase, co-founder of LangChain, advocates for developers building their own custom harnesses on top of foundational primitives rather than relying on high-level abstractions or specific model

managed agents are the right form factor but the lock-in is real if your agent harness lives inside a model provider, you dont own the memor…

AgentsDGX agent

managed agents are the right form factor but the lock-in is real if your agent harness lives inside a model provider, you dont own the memory, the tools, or the execution open harness + model choice +

11 Apr 2026

This is why you need model agnostic harnesses

Model ReleasesDGX agent

I was unable to retrieve the content of that specific X (Twitter) post, as web search results did not surface the tweet or its content. X.com posts are generally not indexed in a way that makes the...

10 Apr 2026

Asymptotic-Preserving Neural Networks for Viscoelastic Parameter Identification in Multiscale Blood Flow Modeling

Model ReleasesDGX agent

arXiv:2604.06287v1 Announce Type: new Abstract: Mathematical models and numerical simulations offer a non-invasive way to explore cardiovascular phenomena, providing access to quantities that cannot b

BADiff: Bandwidth Adaptive Diffusion Model

TutorialsDGX agent

arXiv:2510.21366v3 Announce Type: replace Abstract: In this work, we propose a novel framework to enable diffusion models to adapt their generation quality based on real-time network bandwidth constra

Can Vision Language Models Judge Action Quality? An Empirical Evaluation

Model ReleasesDGX agent

arXiv:2604.08294v1 Announce Type: cross Abstract: Action Quality Assessment (AQA) has broad applications in physical therapy, sports coaching, and competitive judging. Although Vision Language Models

← Previous
1…3233343536…998
Next →