AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,859 results
Local Ai

Model page https://ollama.com/library/kimi-k2.6 More integrations https://docs.ollama.com/integrations

DGX agent

Ollama has made the Kimi K2.6 model available in its library, allowing users to run this model locally through the Ollama platform. The announcement highlights expanded integration options documented

local-aiollama--x
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Modeling Parkinson's Disease Progression Using Longitudinal Voice Biomarkers: A Comparative Study of Statistical and Neural Mixed-Effects Models

DGX agent

arXiv:2507.20058v3 Announce Type: replace-cross Abstract: Predicting Parkinson's Disease (PD) progression is crucial for personalized treatment, and voice biomarkers offer a promising non-invasive met

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Was super interesting to chat with one of @cohere's senior PMs about its AI speech-to-text model, which it hopes to incorporate into its Nor…

DGX agent

Was super interesting to chat with one of @cohere's senior PMs about its AI speech-to-text model, which it hopes to incorporate into its North platform soon. .@Cohere's Cassie Cao takes us behind the

model-releasescohere--x
20 Apr 2026
Research

DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models

DGX agent

arXiv:2602.22175v2 Announce Type: replace Abstract: Understanding and reasoning over long contexts is a crucial capability for language models (LMs). Although recent models support increasingly long c

researcharxiv-cs-cl
17 Apr 2026
Agents

hermes @NousResearch agent with qwen3.5:35b-a3b on a 4090 is VERY good.. local models very impressive..

DGX agent

Nous Research demonstrated strong performance results using their Hermes agent with Qwen 3.5 35B model on an NVIDIA RTX 4090 GPU, highlighting competitive capabilities of locally-run models compared t

agentsnous-research--x
17 Apr 2026
Model Releases

IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation

DGX agent

arXiv:2603.04738v2 Announce Type: replace Abstract: Instruction-following is a foundational capability of large language models (LLMs), with its improvement hinging on scalable and accurate feedback f

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

DGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

safetyarxiv-cs-lg
17 Apr 2026
Model Releases

Beyond Static Personas: Situational Personality Steering for Large Language Models

DGX agent

arXiv:2604.13846v1 Announce Type: new Abstract: Personalized Large Language Models (LLMs) facilitate more natural, human-like interactions in human-centric applications. However, existing personalizat

model-releasesarxiv-cs-cl
16 Apr 2026
Applications

Most Physical AI models recognize patterns. They don’t understand the world. That’s why they fail on edge cases. BADAS 2.0 is a V-JEPA2 worl…

DGX agent

Most Physical AI models recognize patterns. They don’t understand the world. That’s why they fail on edge cases. BADAS 2.0 is a V-JEPA2 world model trained by @getnexar on real-world videos. We used t

applicationsyann-lecun--x
16 Apr 2026
Model Releases

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than bas…

DGX agent

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than base model 🧪 Benchmarks coming soon ✅ Fits on 8GB VRAM ✍️ New mod

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

Peer-Predictive Self-Training for Language Model Reasoning

DGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are…

DGX agent

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are 9x smaller than their 16-bit counterparts while outperformi

model-releasesemad-mostaque--x
16 Apr 2026
Local Ai

Best realism model under 16GB VRAM

DGX agent

This r/StableDiffusion Reddit thread discusses community recommendations for photorealistic image generation models that can run within a 16GB VRAM constraint, a common hardware limit for consumer GPU

local-air-stablediffusion
15 Apr 2026
Research

Cross-Domain Transfer with Particle Physics Foundation Models: From Jets to Neutrino Interactions

DGX agent

arXiv:2604.12364v1 Announce Type: cross Abstract: Future AI-based studies in particle physics will likely start from a foundation model to accelerate training and enhance sensitivity. As a step toward

researcharxiv-cs-lg
15 Apr 2026
Model Releases

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

DGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces

DGX agent

arXiv:2604.11996v1 Announce Type: cross Abstract: Should we trust Large Language Models (LLMs) with high accuracy? LLMs achieve high accuracy on reasoning benchmarks, but correctness alone does not re

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Great news: the ERNIE editing model is expected to be released by the end of this month

DGX agent

A Reddit post on r/StableDiffusion announces the anticipated release of an ERNIE image **editing** model from Baidu, complementing the already-available ERNIE-Image-8b generation model from Baidu, whi

model-releasesr-stablediffusion
15 Apr 2026
Research

LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models

DGX agent

arXiv:2604.12056v1 Announce Type: new Abstract: Block-wise diffusion language models (DLMs) generate multiple tokens in any order, offering a promising alternative to the autoregressive decoding pipel

researcharxiv-cs-cl
15 Apr 2026
Industry

Model S with over 800,000 kilometers on the odometer still going strong!

DGX agent

Model S with over 800,000 kilometers on the odometer still going strong! 🚗 800,361 km. Zero defects. Passed inspection again. In a Tesla Model S. You know… the “toy” that was supposed to die after a f

industryelon-musk--x
15 Apr 2026
Model Releases

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

DGX agent

arXiv:2604.12928v1 Announce Type: new Abstract: Speech-to-speech language models have recently emerged to enhance the naturalness of conversational AI. In particular, full-duplex models are distinguis

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏)

DGX agent

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏) Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its

model-releasesollama--x
15 Apr 2026
Model Releases

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector

DGX agent

arXiv:2509.07177v3 Announce Type: replace Abstract: Large language models have demonstrated impressive capabilities across various domains. However, their general-purpose nature often limits their eff

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Benchmarking Vision-Language Models under Contradictory Virtual Content Attacks in Augmented Reality

DGX agent

arXiv:2604.05510v2 Announce Type: replace Abstract: Augmented reality (AR) has rapidly expanded over the past decade. As AR becomes increasingly integrated into daily life, its security and reliabilit

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

DGX agent

arXiv:2604.11287v1 Announce Type: new Abstract: Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Continuous Adversarial Flow Models

DGX agent

arXiv:2604.11521v1 Announce Type: cross Abstract: We propose continuous adversarial flow models, a type of continuous-time flow model trained with an adversarial objective. Unlike flow matching, which

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models

DGX agent

arXiv:2509.20624v5 Announce Type: replace-cross Abstract: Autoregressive language models (ARMs) deliver strong likelihoods, but are inherently serial: they generate one token per forward pass, which l

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Grid2Matrix: Revealing Digital Agnosia in Vision-Language Models

DGX agent

arXiv:2604.09687v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) excel on many multimodal reasoning benchmarks, but these evaluations often do not require an exhaustive readout of the i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks

DGX agent

arXiv:2604.10508v1 Announce Type: cross Abstract: Large language models frequently fail to produce correct code on their first attempt, yet most benchmarks evaluate them in a single-shot setting. We i

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Modular Delta Merging with Orthogonal Constraints: A Scalable Framework for Continual and Reversible Model Composition

DGX agent

arXiv:2507.20997v4 Announce Type: replace-cross Abstract: In real-world machine learning deployments, models must be continually updated, composed, and when required, selectively undone. However, exis

applicationsarxiv-cs-ai
14 Apr 2026
Research

Parallelism and Generation Order in Masked Diffusion Language Models: Limits Today, Potential Tomorrow

DGX agent

arXiv:2601.15593v2 Announce Type: replace-cross Abstract: Masked Diffusion Language Models (MDLMs) promise parallel token generation and arbitrary-order decoding, yet it remains unclear to what extent

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Self-supervised Pretraining of Cell Segmentation Models

DGX agent

arXiv:2604.10609v1 Announce Type: new Abstract: Instance segmentation enables the analysis of spatial and temporal properties of cells in microscopy images by identifying the pixels belonging to each

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Symmetry-Aware Generative Modeling through Learned Canonicalization

DGX agent

arXiv:2501.07773v3 Announce Type: replace Abstract: Generative modeling of symmetric densities has a range of applications in AI for science, from drug discovery to physics simulations. The existing g

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

DGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching

DGX agent

arXiv:2604.08574v1 Announce Type: cross Abstract: Large Genomic Foundation Models have recently achieved remarkable results and in-vivo translation capabilities. However these models quickly grow to o

researcharxiv-cs-ai
13 Apr 2026
Safety

Learning Vision-Language-Action World Models for Autonomous Driving

DGX agent

arXiv:2604.09059v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved notable progress in end-to-end autonomous driving by integrating perception, reasoning, and

safetyarxiv-cs-ai
13 Apr 2026
Safety

Post-Selection Distributional Model Evaluation

DGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Local Ai

Any models?

DGX agent

A Reddit post on r/ollama where a community member asks about model availability or recommendations for use with the Ollama local AI runtime. The discussion likely covers which open-source models (suc

local-air-ollama
12 Apr 2026
Model Releases

Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest

DGX agent

arXiv:2604.08525v1 Announce Type: cross Abstract: Today's large language models (LLMs) are trained to align with user preferences through methods such as reinforcement learning. Yet models are beginni

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Models

DGX agent

arXiv:2604.07749v1 Announce Type: new Abstract: Large language models (LLMs) can shift their answers under pressure in ways that reflect accommodation rather than reasoning. Prior work on sycophancy h

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

BREAKING: GLM 5.1 by @Zai_org overwhelmingly dominates design-centric coding tasks amongst open-weight models In the categories featured bel…

DGX agent

BREAKING: GLM 5.1 by @Zai_org overwhelmingly dominates design-centric coding tasks amongst open-weight models In the categories featured below, it is most comparable to Opus 4.6 by @AnthropicAI at ~1/

model-releaseszhipu-ai--x
10 Apr 2026
Model Releases

GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning

DGX agent

arXiv:2604.07808v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models is constrained by substantial GPU memory requirements. Low-rank adaptation methods mitigate this cha

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

MCLR: Improving Conditional Modeling via Inter-Class Likelihood-Ratio Maximization and Unifying Classifier-Free Guidance with Alignment Objectives

DGX agent

arXiv:2603.22364v2 Announce Type: replace-cross Abstract: Diffusion models have achieved state-of-the-art performance in generative modeling, but their success often relies heavily on classifier-free

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning

DGX agent

arXiv:2604.07944v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong potential for autonomous vehicle motion planning by reformulating trajectory prediction a

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

DGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning

DGX agent

arXiv:2601.04268v2 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that

model-releasesarxiv-cs-lg
10 Apr 2026
← Previous
1…3132333435…1248
Next →