AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,841 results
27 Apr 2026

Differentiable Filtering for Learning Hidden Markov Models

Local AiDGX agent

arXiv:2511.10571v2 Announce Type: replace Abstract: Hidden Markov Models (HMMs) are fundamental for modeling sequential data, yet learning their parameters from observations remains challenging. Class

Language Specific Knowledge: Do Models Know Better in X than in English?

Model ReleasesDGX agent

arXiv:2505.14990v3 Announce Type: replace Abstract: Often, multilingual language models are trained with the objective to map semantically similar content (in different languages) in the same latent s

When LoRA Betrays: Backdooring Text-to-Image Models by Masquerading as Benign Adapters

ResearchDGX agent

arXiv:2602.21977v4 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a leading technique for efficiently fine-tuning text-to-image diffusion models, and its widespread adoptio

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
24 Apr 2026

ChessArena: A Chess Testbed for Evaluating Strategic Reasoning Capabilities of Large Language Models

ResearchDGX agent

arXiv:2509.24239v4 Announce Type: replace-cross Abstract: Recent large language models (LLMs) have shown strong reasoning capabilities. However, a critical question remains: do these models possess ge

Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training

SafetyDGX agent

arXiv:2604.21741v1 Announce Type: new Abstract: Post-training is essential for turning pretrained generalist robot policies into reliable task-specific controllers, but existing human-in-the-loop pipe

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in …

Model ReleasesDGX agent

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in their categories while benchmarking close to the frontier mo

Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance

Model ReleasesDGX agent

arXiv:2604.21104v1 Announce Type: new Abstract: New geospatial foundation models introduce a new model architecture and pretraining dataset, often sampled using different notions of data diversity. Pe

Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors

Model ReleasesDGX agent

arXiv:2604.21893v1 Announce Type: cross Abstract: Geographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constrainin

VLA-Forget: Vision-Language-Action Unlearning for Embodied Foundation Models

SafetyDGX agent

arXiv:2604.03956v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models are emerging as embodied foundation models for robotic manipulation, but their deployment introduces a new

23 Apr 2026

LayerTracer: A Joint Task-Particle and Vulnerable-Layer Analysis framework for Arbitrary Large Language Model Architectures

Model ReleasesDGX agent

arXiv:2604.20556v1 Announce Type: cross Abstract: Currently, Large Language Models (LLMs) feature a diversified architectural landscape, including traditional Transformer, GateDeltaNet, and Mamba. How

22 Apr 2026

AlignCultura: Towards Culturally Aligned Large Language Models?

Model ReleasesDGX agent

arXiv:2604.19016v1 Announce Type: new Abstract: Cultural alignment in Large Language Models (LLMs) is essential for producing contextually aware, respectful, and trustworthy outputs. Without it, model

Automated Energy-Aware Time-Series Model Deployment on Embedded FPGAs for Resilient Combined Sewer Overflow Management

Local AiDGX agent

arXiv:2508.13905v2 Announce Type: replace Abstract: Extreme weather events, intensified by climate change, increasingly challenge aging combined sewer systems, raising the risk of untreated wastewater

Dynamic Model Routing and Cascading for Efficient LLM Inference: A Survey

ResearchDGX agent

arXiv:2603.04445v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) with diverse capabilities, costs, and domains has created a critical need for intelligent mod

Fine-Tuning Small Reasoning Models for Quantum Field Theory

Model ReleasesDGX agent

arXiv:2604.18936v1 Announce Type: cross Abstract: Despite the growing application of Large Language Models (LLMs) to theoretical physics, there is little academic exploration into how domain-specific

How Far Are Video Models from True Multimodal Reasoning?

Model ReleasesDGX agent

arXiv:2604.19193v1 Announce Type: new Abstract: Despite remarkable progress toward general-purpose video models, a critical question remains unanswered: how far are these models from achieving true mu

LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.18803v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in settings where reliable visual grounding carries operational consequences, yet their behavi

Multilingual Language Models Encode Script Over Linguistic Structure

Model ReleasesDGX agent

arXiv:2604.05090v2 Announce Type: replace Abstract: Multilingual language models (LMs) organize representations for typologically and orthographically diverse languages into a shared parameter space,

QSLM: A Performance- and Memory-aware Quantization Framework with Tiered Search Strategy for Spike-driven Language Models

ResearchDGX agent

arXiv:2601.00679v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been emerging as prominent AI models for solving many natural language tasks due to their high performance (

21 Apr 2026

Characterizing Model-Native Skills

SafetyDGX agent

arXiv:2604.17614v1 Announce Type: cross Abstract: Skills are a natural unit for describing what a language model can do and how its behavior can be changed. However, existing characterizations rely on

Cross-Family Speculative Decoding for Polish Language Models on Apple~Silicon: An Empirical Evaluation of Bielik~11B with UAG-Extended MLX-LM

Model ReleasesDGX agent

arXiv:2604.16368v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a small draft model to propose k candidate tokens for a target model to verify. While effective

Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG

Model ReleasesDGX agent

arXiv:2604.16422v1 Announce Type: new Abstract: The injection of domain-specific knowledge is crucial for adapting language models (LMs) to specialized fields such as biomedicine. While most current a

JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew

Model ReleasesDGX agent

arXiv:2604.18041v1 Announce Type: new Abstract: Despite significant advances in large language models, personalizing them for individual decision-makers remains an open problem. Here, we introduce a s

LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2603.26771v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens from a fully masked sequence. Their standard confidence-based

Reasoning Models Know What's Important, and Encode It in Their Activations

ResearchDGX agent

arXiv:2604.18307v1 Announce Type: new Abstract: Language models often solve complex tasks by generating long reasoning chains, consisting of many steps with varying importance. While some steps are cr

Scaling Recurrence-aware Foundation Models for Clinical Records via Next-Visit Prediction

Model ReleasesDGX agent

arXiv:2603.24562v2 Announce Type: replace Abstract: While large-scale pretraining has revolutionized language modeling, its potential remains underexplored in healthcare with structured electronic hea

Sharpening Lightweight Models for Generalized Polyp Segmentation: A Boundary Guided Distillation from Foundation Models

SafetyDGX agent

arXiv:2604.17865v1 Announce Type: new Abstract: Automated polyp segmentation is critical for early colorectal cancer detection and its prevention, yet remains challenging due to weak boundaries, large

SIF: Semantically In-Distribution Fingerprints for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.17041v1 Announce Type: new Abstract: The public accessibility of large vision-language models (LVLMs) raises serious concerns about unauthorized model reuse and intellectual property infrin

Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

Model ReleasesDGX agent

arXiv:2604.17159v1 Announce Type: cross Abstract: We present, to our knowledge, the most comprehensive cross-model evaluation of LLM agents on offensive cybersecurity tasks, benchmarking 10 frontier m

The Illusion of Insight in Reasoning Models

Model ReleasesDGX agent

arXiv:2601.00514v2 Announce Type: replace-cross Abstract: Do reasoning models have 'Aha!' moments? Prior work suggests that models like DeepSeek-R1-Zero undergo sudden mid-trace realizations that lead

Towards a Foundation-Model Paradigm for Aerodynamic Prediction in Three-dimensional Design

TutorialsDGX agent

arXiv:2604.18062v1 Announce Type: new Abstract: Accurate machine-learning models for aerodynamic prediction are essential for accelerating shape optimization, yet remain challenging to develop for com

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Model ReleasesDGX agent

arXiv:2505.20279v4 Announce Type: replace-cross Abstract: The rapid advancement of Large Multimodal Models (LMMs) for 2D images and videos has motivated extending these models to understand 3D scenes,

20 Apr 2026

Boston Dynamics just gave its robot dog a brain that reasons about the physical world. Google DeepMind's Gemini Robotics model is now runnin…

Model ReleasesDGX agent

Boston Dynamics just gave its robot dog a brain that reasons about the physical world. Google DeepMind's Gemini Robotics model is now running inside Spot, the four-legged robot already deployed at tho

Evaluating the Progression of Large Language Model Capabilities for Small-Molecule Drug Design

ApplicationsDGX agent

arXiv:2604.16279v1 Announce Type: new Abstract: Large Language Models (LLMs) have the potential to accelerate small molecule drug design due to their ability to reason about information from diverse s

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations

Model ReleasesDGX agent

arXiv:2603.03332v3 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) prompting has emerged as a foundational technique for eliciting reasoning from Large Language Models (LLMs), yet the ro

Information Router for Mitigating Modality Dominance in Vision-Language Models

ApplicationsDGX agent

arXiv:2604.16264v1 Announce Type: new Abstract: Vision Language models (VLMs) have demonstrated strong performance across a wide range of benchmarks, yet they often suffer from modality dominance, whe

Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding

SafetyDGX agent

arXiv:2512.04847v2 Announce Type: replace-cross Abstract: Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limi

Model page https://ollama.com/library/kimi-k2.6 More integrations https://docs.ollama.com/integrations

Local AiDGX agent

Ollama has made the Kimi K2.6 model available in its library, allowing users to run this model locally through the Ollama platform. The announcement highlights expanded integration options documented

Modeling Parkinson's Disease Progression Using Longitudinal Voice Biomarkers: A Comparative Study of Statistical and Neural Mixed-Effects Models

ResearchDGX agent

arXiv:2507.20058v3 Announce Type: replace-cross Abstract: Predicting Parkinson's Disease (PD) progression is crucial for personalized treatment, and voice biomarkers offer a promising non-invasive met

Was super interesting to chat with one of @cohere's senior PMs about its AI speech-to-text model, which it hopes to incorporate into its Nor…

Model ReleasesDGX agent

Was super interesting to chat with one of @cohere's senior PMs about its AI speech-to-text model, which it hopes to incorporate into its North platform soon. .@Cohere's Cassie Cao takes us behind the

17 Apr 2026

DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models

ResearchDGX agent

arXiv:2602.22175v2 Announce Type: replace Abstract: Understanding and reasoning over long contexts is a crucial capability for language models (LMs). Although recent models support increasingly long c

hermes @NousResearch agent with qwen3.5:35b-a3b on a 4090 is VERY good.. local models very impressive..

AgentsDGX agent

Nous Research demonstrated strong performance results using their Hermes agent with Qwen 3.5 35B model on an NVIDIA RTX 4090 GPU, highlighting competitive capabilities of locally-run models compared t

IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation

Model ReleasesDGX agent

arXiv:2603.04738v2 Announce Type: replace Abstract: Instruction-following is a foundational capability of large language models (LLMs), with its improvement hinging on scalable and accurate feedback f

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

SafetyDGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

16 Apr 2026

Beyond Static Personas: Situational Personality Steering for Large Language Models

Model ReleasesDGX agent

arXiv:2604.13846v1 Announce Type: new Abstract: Personalized Large Language Models (LLMs) facilitate more natural, human-like interactions in human-centric applications. However, existing personalizat

Most Physical AI models recognize patterns. They don’t understand the world. That’s why they fail on edge cases. BADAS 2.0 is a V-JEPA2 worl…

ApplicationsDGX agent

Most Physical AI models recognize patterns. They don’t understand the world. That’s why they fail on edge cases. BADAS 2.0 is a V-JEPA2 world model trained by @getnexar on real-world videos. We used t

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than bas…

Model ReleasesDGX agent

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than base model 🧪 Benchmarks coming soon ✅ Fits on 8GB VRAM ✍️ New mod

Peer-Predictive Self-Training for Language Model Reasoning

Model ReleasesDGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

Model ReleasesDGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are…

Model ReleasesDGX agent

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are 9x smaller than their 16-bit counterparts while outperformi

15 Apr 2026

Best realism model under 16GB VRAM

Local AiDGX agent

This r/StableDiffusion Reddit thread discusses community recommendations for photorealistic image generation models that can run within a 16GB VRAM constraint, a common hardware limit for consumer GPU

Cross-Domain Transfer with Particle Physics Foundation Models: From Jets to Neutrino Interactions

ResearchDGX agent

arXiv:2604.12364v1 Announce Type: cross Abstract: Future AI-based studies in particle physics will likely start from a foundation model to accelerate training and enhance sensitivity. As a step toward

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

Model ReleasesDGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces

Model ReleasesDGX agent

arXiv:2604.11996v1 Announce Type: cross Abstract: Should we trust Large Language Models (LLMs) with high accuracy? LLMs achieve high accuracy on reasoning benchmarks, but correctness alone does not re

Great news: the ERNIE editing model is expected to be released by the end of this month

Model ReleasesDGX agent

A Reddit post on r/StableDiffusion announces the anticipated release of an ERNIE image **editing** model from Baidu, complementing the already-available ERNIE-Image-8b generation model from Baidu, whi

LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models

ResearchDGX agent

arXiv:2604.12056v1 Announce Type: new Abstract: Block-wise diffusion language models (DLMs) generate multiple tokens in any order, offering a promising alternative to the autoregressive decoding pipel

Model S with over 800,000 kilometers on the odometer still going strong!

IndustryDGX agent

Model S with over 800,000 kilometers on the odometer still going strong! 🚗 800,361 km. Zero defects. Passed inspection again. In a Tesla Model S. You know… the “toy” that was supposed to die after a f

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

Model ReleasesDGX agent

arXiv:2604.12928v1 Announce Type: new Abstract: Speech-to-speech language models have recently emerged to enhance the naturalness of conversational AI. In particular, full-duplex models are distinguis

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏)

Model ReleasesDGX agent

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏) Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector

Model ReleasesDGX agent

arXiv:2509.07177v3 Announce Type: replace Abstract: Large language models have demonstrated impressive capabilities across various domains. However, their general-purpose nature often limits their eff

14 Apr 2026

Benchmarking Vision-Language Models under Contradictory Virtual Content Attacks in Augmented Reality

Model ReleasesDGX agent

arXiv:2604.05510v2 Announce Type: replace Abstract: Augmented reality (AR) has rapidly expanded over the past decade. As AR becomes increasingly integrated into daily life, its security and reliabilit

← Previous
1…2425262728…998
Next →