AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
13 Apr 2026

State Space Models are Effective Sign Language Learners: Exploiting Phonological Compositionality for Vocabulary-Scale Recognition

ResearchDGX agent

arXiv:2604.08761v1 Announce Type: new Abstract: Sign language recognition suffers from catastrophic scaling failure: models achieving high accuracy on small vocabularies collapse at realistic sizes. E

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

Model ReleasesDGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

Toward World Models for Epidemiology

SafetyDGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

TutorialsDGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

Tried Ollama Cloud, just realize only Kimi model accept images

Local AiDGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

ResearchDGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

Model ReleasesDGX agent

Boris Cherny and colleagues conducted cybersecurity evaluations of Claude Mythos Preview, finding it to be the first AI model to successfully complete an AISI (AI Safety Institute) cyber range end-to-

12 Apr 2026

How to train loras in One Trainer for Z Image using Civitai models?

TutorialsDGX agent

This r/StableDiffusion post likely discusses the community challenge of using OneTrainer — a locally-installed, GUI-based LoRA training tool — to train LoRAs specifically compatible with Z Image (Tong

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduc…

Local AiDGX agent

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduce and local LLM models might increase (save cost) and thus t

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

Model ReleasesDGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

11 Apr 2026

What's model should I run?

Local AiDGX agent

A Reddit discussion from the r/ollama community where a user seeks advice on which AI language model to run locally using Ollama. Responses likely include hardware-based recommendations (such as RAM a

10 Apr 2026

A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset

Model ReleasesDGX agent

arXiv:2604.06227v1 Announce Type: new Abstract: Accurate short-term forecasting of agricultural commodity prices is critical for food security planning and smallholder income stabilisation in developi

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

ApplicationsDGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

Model ReleasesDGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

SafetyDGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

Bootstrapping Sign Language Annotations with Sign Language Models

Model ReleasesDGX agent

arXiv:2604.07606v1 Announce Type: new Abstract: AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain

ChatGPT voice mode is a weaker model

ToolsDGX agent

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't. If

Compact Example-Based Explanations for Language Models

ResearchDGX agent

arXiv:2601.03786v2 Announce Type: replace Abstract: Training data influence estimation methods quantify the contribution of training documents to a model's output, making them a promising source of in

ConfusionPrompt: Practical Private Inference for Online Large Language Models

Local AiDGX agent

arXiv:2401.00870v5 Announce Type: replace-cross Abstract: State-of-the-art large language models (LLMs) are typically deployed as online services, requiring users to transmit detailed prompts to cloud

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning

ResearchDGX agent

arXiv:2604.08084v1 Announce Type: new Abstract: Current video captioning methods usually use an encoder-decoder structure to generate text autoregressively. However, autoregressive methods have inhere

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models

ApplicationsDGX agent

arXiv:2604.06826v1 Announce Type: cross Abstract: Environmental, Social, and Governance (ESG) considerations are increasingly integral to assessing corporate performance, reputation, and long-term sus

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexib…

AgentsDGX agent

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexibility, but to ensure we actually own the long-term memory th

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model

ToolsDGX agent

ChatGPT's voice mode, when directly queried about its knowledge cutoff date, reports April 2024, indicating it is powered by a GPT-4o era model rather than a more recent one. GPT-4o initially had ...

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads

ResearchDGX agent

arXiv:2604.06356v1 Announce Type: cross Abstract: In-Context Learning (ICL) has been extensively studied in text-only Language Models, but remains largely unexplored in the speech domain. Here, we inv

kugel-2 model (VibeVoice finetune) repo is gone. Does anyone know why?

Local AiDGX agent

The kugel-2 model is a community fine-tune of Microsoft's VibeVoice, a text-to-speech model — and its disappearance is rooted in Microsoft's own removal of the base VibeVoice repository. Microsoft...

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

SafetyDGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

LPM 1.0: Video-based Character Performance Model

Model ReleasesDGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments

ResearchDGX agent

arXiv:2604.07254v1 Announce Type: cross Abstract: Deep neural networks can predict human judgments, but this does not imply that they rely on human-like information or reveal the cues underlying those

Operator Learning for Surrogate Modeling of Wave-Induced Forces from Sea Surface Waves

ResearchDGX agent

arXiv:2604.06433v1 Announce Type: cross Abstract: Wave setup plays a significant role in transferring wave-induced energy to currents and causing an increase in water elevation. This excess momentum f

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics

ApplicationsDGX agent

arXiv:2604.08503v1 Announce Type: new Abstract: Recent advances in generative video modeling, driven by large-scale datasets and powerful architectures, have yielded remarkable visual realism. However

PlaneCycle: Training-Free 2D-to-3D Lifting of Foundation Models Without Adapters

ResearchDGX agent

arXiv:2603.04165v3 Announce Type: replace-cross Abstract: Large-scale 2D foundation models exhibit strong transferable representations, yet extending them to 3D volumetric data typically requires retr

Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models

SafetyDGX agent

arXiv:2604.07779v1 Announce Type: new Abstract: Pathology foundation models (FMs) have become central to computational histopathology, offering strong transfer performance across a wide range of diagn

Prompt reinforcing for long-term planning of large language models

Model ReleasesDGX agent

arXiv:2510.05921v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompt

Robust support vector model based on bounded asymmetric elastic net loss for binary classification

Model ReleasesDGX agent

arXiv:2603.06257v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel bounded asymmetric elastic net (L_{baen}) loss function and combine it with the support vector machine (SV

Search-R3: Unifying Reasoning and Embedding in Large Language Models

TutorialsDGX agent

arXiv:2510.07048v2 Announce Type: replace Abstract: Despite their remarkable natural language understanding capabilities, Large Language Models (LLMs) have been underutilized for retrieval tasks. We p

Tesla is down to the last few hundred Model S & X cars in inventory. Poignant end of an era.

IndustryDGX agent

Tesla is down to the last few hundred Model S & X cars in inventory. Poignant end of an era. Traded in my 2020 Model S for a brand new plaid X before they discontinue it. Car is amazing, but the FSD h

VoxCPM TTS model + LoRa training abilities right in Comfy

Local AiDGX agent

ComfyUI-VoxCPM is a custom node that integrates VoxCPM — a novel tokenizer-free Text-to-Speech system that models speech in a continuous space — directly into ComfyUI's visual workflow environmen...

9 Apr 2026

Anthropic built a model too risky to release

IndustryDGX agent

Anthropic's new frontier model, Claude Mythos, is the first model the company has publicly deemed too high-risk for general release , due to its advanced cybersecurity capabilities — including the ...

FlowInOne - A new Multimodal image model . Released on Huggingface

Model ReleasesDGX agent

FlowInOne is a vision-centric multimodal image generation framework that reformulates multimodal generation as a purely visual flow, converting all inputs into visual prompts and enabling a clean i...

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data …

IndustryDGX agent

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data from 720 people. You feed it a video, sound clip, or text, a

Understanding Amazon Bedrock model lifecycle

TutorialsDGX agent

This post shows you how to manage FM transitions in Amazon Bedrock, so you can make sure your AI applications remain operational as models evolve. We discuss the three lifecycle states, how to plan mi

8 Apr 2026

[AINews] Anthropic @ $30B ARR, Project GlassWing and Claude Mythos Preview — first model too dangerous to release since GPT-2

Model ReleasesDGX agent

Anthropic announced it has grown its ARR from $19B to $30B in just one month, coinciding with the formal unveiling of Claude Mythos Preview — described in leaked company documents as 'by far the m...

14 Aug 2026

AlayaWorld: Interactive Long-Horizon World Modeling - Full Technical Report (v1.1)

ResearchDGX agent

arXiv:2608.13492v1 Announce Type: new Abstract: This report presents an improved version of AlayaWorld. While the backbone architecture, chunk-wise autoregressive generation scheme, and training data

Apple trained its own AI model for China with help from Alibaba

IndustryDGX agent

Apple has reportedly trained a custom AI model for the China market alongside domestic tech giant Alibaba, a rare cross-border partnership that cuts across growing tensions between Beijing and Washing

Can Generalist Vision Language Models (VLMs) Rival Specialist Medical VLMs? Benchmarking and Strategic Insights

ResearchDGX agent

arXiv:2506.17337v5 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) have shown promise in automating image diagnosis and interpretation in clinical settings. However, developing sp

Diagnosing JEPA World Models with Action-Conditioned Predictive Consistency

TutorialsDGX agent

arXiv:2608.12939v1 Announce Type: new Abstract: Joint-embedding predictive architectures (JEPAs) learn world models that predict in a compact latent space rather than in pixels, reducing the pressure

Error-Aware Reverse Auction Mechanism for Large Language Model Routing

ApplicationsDGX agent

arXiv:2608.12719v1 Announce Type: cross Abstract: Routing each query to a cost-effective large language model (LLM) is critical for balancing quality and cost, yet most routers rely on a centralized t

How Good are Foundation Models in Longitudinal MRI Disease Progression Reasoning?

Model ReleasesDGX agent

arXiv:2608.13309v1 Announce Type: new Abstract: Magnetic Resonance Imaging (MRI) interpretation is fundamental to clinical decision-making, requiring radiologists to integrate multi-view anatomical pl

Measuring Task-Agnostic Training Data Influence Across Language Model Pretraining

ResearchDGX agent

arXiv:2608.13515v1 Announce Type: new Abstract: Measuring training data influence consistently across language model pretraining is challenging. It is difficult to select downstream tasks or validatio

Neural Quadratic Forms: A Unified Minimal Model for Sudden Learning and Scaling Laws

Model ReleasesDGX agent

arXiv:2608.13335v1 Announce Type: new Abstract: Neural networks trained by gradient descent on a smooth cost function can nevertheless learn in steps: the cost holds on long plateaus and then drops ab

neurosymbolic world models – *exactly* what I argued for in 2020 in the Next Decade in AI – for the win.

SafetyDGX agent

neurosymbolic world models – *exactly* what I argued for in 2020 in the Next Decade in AI – for the win. Jeremy's excellent work here is a great illustration of a very powerful type of approach: LLM-g

Shortest-Path Decomposition for Foliage-Robust 3D Tree Modeling and Above-Ground Biomass Estimation from Point Clouds

Model ReleasesDGX agent

arXiv:2506.15577v2 Announce Type: replace Abstract: Estimating above-ground biomass (AGB) from terrestrial laser scanning (TLS) via quantitative structural models (QSMs) is highly accurate under leaf-

Structured Local Differential Modeling for AI-Generated Image Detection

Local AiDGX agent

arXiv:2608.12811v1 Announce Type: new Abstract: The rapid advancement of AI-generated content has made the reliable detection of generated images an increasingly critical challenge. Existing detection

13 Aug 2026

A Simple Efficiency Incremental Learning Framework via Vision-Language Model with Nonlinear Multi-Adapters

TutorialsDGX agent

arXiv:2603.11211v3 Announce Type: replace-cross Abstract: Incremental Learning (IL) aims to learn new tasks while preserving previously acquired knowledge. Integrating the zero-shot learning capabilit

Accelerating Time Series Foundation Models with Speculative Decoding

ResearchDGX agent

arXiv:2511.18191v2 Announce Type: replace Abstract: Time series forecasting drives operational decisions under tight latency budgets, and autoregressive time series foundation models (TSFMs) increasin

Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World Models

SafetyDGX agent

arXiv:2608.12078v1 Announce Type: cross Abstract: Learning world models from offline trajectories enables agents to accomplish different tasks through planning. Object-centric (OC) representations, wh

Causal Agent based on Large Language Model

Model ReleasesDGX agent

arXiv:2408.06849v3 Announce Type: replace Abstract: The large language model (LLM) has achieved significant success across various domains. However, the inherent complexity of causal problems and caus

← Previous
1…96979899100…1009
Next →