AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
13 Apr 2026

Robust Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

Tango: Taming Visual Signals for Efficient Video Large Language Models

Local AiDGX agent

arXiv:2604.09547v1 Announce Type: new Abstract: Token pruning has emerged as a mainstream approach for developing efficient Video Large Language Models (Video LLMs). This work revisits and advances th

Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models

ResearchDGX agent

arXiv:2603.01400v2 Announce Type: replace Abstract: Video Large Language Models (VLLMs) demonstrate strong video understanding but suffer from inefficiency due to redundant visual tokens. Existing pru

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

We wrote a build-from-scratch Python book on the 98% of production AI systems that isn't the model call [P]

ApplicationsDGX agent

This Reddit post from r/MachineLearning announces a build-from-scratch Python book focused on the engineering infrastructure surrounding AI systems — the components beyond the model call itself, such

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

TutorialsDGX agent

arXiv:2604.08716v1 Announce Type: new Abstract: Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try

12 Apr 2026

gemma 4 going insane

Model ReleasesDGX agent

This r/ollama thread discusses users experiencing erratic and broken behavior when running Google's Gemma 4 models locally via Ollama, including issues such as strange and unrelated responses, as well

model page: https://ollama.com/library/minimax-m2.7

Local AiDGX agent

MiniMax-M1 is a large-scale open-weight reasoning model developed by MiniMax, featuring a hybrid Mixture-of-Experts (MoE) architecture with 456 billion total parameters (activating 45.9 billion per to

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the a…

AgentsDGX agent

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the alpha is deepagents is our take: built-in tools, memory, smar

11 Apr 2026

AI models are terrible at betting on soccer—especially xAI Grok

IndustryDGX agent

A report called 'KellyBench,' released by AI start-up General Reasoning, tested eight leading AI models in a virtual re-creation of the 2023–24 Premier League season, providing them with detailed hist

Does UI Preset = Base model??

Local AiDGX agent

This Reddit thread from r/StableDiffusion addresses a common point of confusion among users of Stable Diffusion UIs (such as Stable Diffusion WebUI Forge) regarding whether selecting a 'UI Preset' is

10 Apr 2026

A systematic framework for generating novel experimental hypotheses from language models

SafetyDGX agent

arXiv:2408.05086v3 Announce Type: replace Abstract: Neural language models (LMs) have been shown to capture complex linguistic patterns, yet their utility in understanding human language and more broa

ABMAMBA: Multimodal Large Language Model with Aligned Hierarchical Bidirectional Scan for Efficient Video Captioning

ResearchDGX agent

arXiv:2604.08050v1 Announce Type: new Abstract: In this study, we focus on video captioning by fully open multimodal large language models (MLLMs). The comprehension of visual sequences is challenging

AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained Models

ResearchDGX agent

arXiv:2405.16240v3 Announce Type: replace Abstract: In this paper, we introduce analytic federated learning (AFL), a new training paradigm that brings analytical (i.e., closed-form) solutions to the f

Apple's head of cloud says Open Source models will address 90% of the use case

IndustryDGX agent

I was unable to retrieve the specific Reddit thread or locate reliable sourced details about Apple's head of cloud making a statement that open source models will address 90% of use cases. The sear...

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

TutorialsDGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

Beyond the Mean: Modelling Annotation Distributions in Continuous Affect Prediction

ResearchDGX agent

arXiv:2604.07198v1 Announce Type: new Abstract: Emotion annotation is inherently subjective and cognitively demanding, producing signals that reflect diverse perceptions across annotators rather than

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

ResearchDGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

ResearchDGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

Model ReleasesDGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

ResearchDGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

Hierarchical Feature Learning for Medical Point Clouds via State Space Model

ResearchDGX agent

arXiv:2504.13015v3 Announce Type: replace Abstract: Deep learning-based point cloud modeling has been widely investigated as an indispensable component of general shape analysis. Recently, transformer

Incentive-Aware Multi-Fidelity Optimization for Generative Advertising in Large Language Models

ResearchDGX agent

arXiv:2604.06263v1 Announce Type: cross Abstract: Generative advertising in large language model (LLM) responses requires optimizing sponsorship configurations under two strict constraints: the strate

Learning is Forgetting: LLM Training As Lossy Compression

TutorialsDGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

LINE: LLM-based Iterative Neuron Explanations for Vision Models

SafetyDGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

ODE-free Neural Flow Matching for One-Step Generative Modeling

TutorialsDGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

AgentsDGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

Model ReleasesDGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training

ResearchDGX agent

arXiv:2604.06836v2 Announce Type: new Abstract: Quantization is an effective way to reduce the memory cost of large-scale model training. However, most existing methods adopt fixed-precision policies,

The quality of talks at @aiDotEngineer is insane, being able to learn about diffusion models and flow mapping from @GoogleDeepMind’s @sediel…

TutorialsDGX agent

The AI Engineer Summit (@aiDotEngineer) is a highly regarded technical conference featuring speakers from leading AI organizations including Google DeepMind, Anthropic, and OpenAI, known for its de...

Visual prompting reimagined: The power of the Activation Prompts

Model ReleasesDGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

9 Apr 2026

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments -…

AgentsDGX agent

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments - in Classical Deep Learning, we learn from each training exa

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation.…

AgentsDGX agent

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation. The most interesting finding isn't the 2-3x memory savings

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents brea…

AgentsDGX agent

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents break. The reframe: evals are the training data for your harness

8 Apr 2026

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assu…

IndustryDGX agent

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assume everyone can be compromised, and build systems that keep

GLM-5.1 gives teams a stronger model for coding, tool use, and sustained agent performance on Together AI. Learn more: http://www.together.a…

AgentsDGX agent

GLM-5.1 is Z.ai's post-training upgrade to GLM-5, now available on Together AI, delivering a 28% coding performance improvement through a refined reinforcement learning pipeline while retaining the...

> which is exactly why we believe memory should live outside of model providers open harness = open memory which everyone should want!

AgentsDGX agent

> which is exactly why we believe memory should live outside of model providers open harness = open memory which everyone should want! The new Anthropic managed agents API is basically the Letta API t

7 Apr 2026

Introducing GLM-5.1 for understanding research papers 🚀 Highlight any section of a paper to ask questions and “@” other papers for quick co…

Model ReleasesDGX agent

AlphaXiv introduced GLM-5.1 as the underlying model powering its research paper understanding features on the alphaXiv platform, enabling users to highlight any section of a paper to ask contextual...

14 Aug 2026

8/19, join us for the #ACMTechTalk, 'From Conventional LLMs to Reasoning Models to Agents,' w/AI & LLM Research Engineer @rasbt. ACM Practit…

ResearchDGX agent

8/19, join us for the #ACMTechTalk, 'From Conventional LLMs to Reasoning Models to Agents,' w/AI & LLM Research Engineer @rasbt. ACM Practitioner Board Co-Chaior @marlene_zw (@Microsoft) will moderate

A Probe Direction Is a Property of Its Prompt

Model ReleasesDGX agent

arXiv:2608.13329v1 Announce Type: new Abstract: A model that behaves differently when it senses it is being tested would undermine the evaluations we rely on, so recent work has sought to read that se

Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

ResearchDGX agent

arXiv:2608.13396v1 Announce Type: new Abstract: Shape and force sensing have long been critical bottlenecks in the development of compact capstan-driven continuum surgical robots, primarily due to the

Demand Transfer Estimation at Scale via Restricted Logit Modeling

ResearchDGX agent

arXiv:2608.12680v1 Announce Type: cross Abstract: Item demand forecasting is an integral component of store assortment optimization. Existing literature focuses on learning a suitable customer choice

DiffGRM: Diffusion-based Generative Recommendation Model

SafetyDGX agent

arXiv:2510.21805v2 Announce Type: replace-cross Abstract: Generative recommendation (GR) is an emerging paradigm that represents each item via a tokenizer as an n-digit semantic ID (SID) and predicts

From Observation to Intervention: Memory in Brains and Large Language Models

ResearchDGX agent

arXiv:2608.12377v1 Announce Type: cross Abstract: Brains and large language models (LLMs) are fundamentally different memory systems, but they can be compared through shared functional questions: wher

Intern-S2-Preview: Scientific Agentic Foundation Model

SafetyDGX agent

arXiv:2608.13505v1 Announce Type: cross Abstract: Scientific discovery increasingly requires AI systems that can reason over scientific evidence of heterogeneous modalities, interact with scientific t

JailWAM: Jailbreaking World Action Models in Robot Control

SafetyDGX agent

arXiv:2604.05498v2 Announce Type: replace Abstract: World Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation, enabling physical interaction across diverse tasks and env

Jeremy's excellent work here is a great illustration of a very powerful type of approach: LLM-guided on-the-fly synthesis of a symbolic worl…

Model ReleasesDGX agent

Jeremy's excellent work here is a great illustration of a very powerful type of approach: LLM-guided on-the-fly synthesis of a symbolic world model, i.e. making sense of the world by writing executabl

Qwen3.8-Max is live on DigitalOcean Serverless Inference. Launching side by side with DigitalOcean as our Day 0 launch partner. Big model. S…

HardwareDGX agent

Qwen3.8-Max is live on DigitalOcean Serverless Inference. Launching side by side with DigitalOcean as our Day 0 launch partner. Big model. Smooth sailing. Now on DigitalOcean.🌊🏄‍♀️ @digitalocean Now a

Request: More Transparency on Ollama Cloud Subscriptions

Local AiDGX agent

I've been an Ollama Cloud subscriber for ~6 months. I generally use the the latest GLM models available for coding as well as a personal instance of Open WebUI. I have tried to get answers directly vi

Sources: Apple trained a China-specific LLM with Alibaba's support, which would make Apple the first foreign company to offer a proprietary AI model in China (Reuters)

IndustryDGX agent

Reuters: Sources: Apple trained a China-specific LLM with Alibaba's support, which would make Apple the first foreign company to offer a proprietary AI model in China — Apple (AAPL.O) has trained a la

Structure-aware Riemannian Growth Fields for 4D Plant Modeling

ResearchDGX agent

arXiv:2608.13007v1 Announce Type: new Abstract: In this paper, we introduce a novel framework for 4D plant growth modeling that reconstructs the continuous geometric and topological evolution of plant

13 Aug 2026

A Reality Check of Language Models as Formalizers on Constraint Satisfaction Problems

ResearchDGX agent

arXiv:2505.13252v5 Announce Type: replace Abstract: Recent work shows superior performance when using large language models (LLMs) as formalizers instead of as end-to-end solvers for symbolic reasonin

Chemically Meaningful Textualization Enables Explainable Validation of Metal-Organic Frameworks by Large Language Models

ResearchDGX agent

arXiv:2608.11283v1 Announce Type: cross Abstract: Computation-ready metal-organic framework (MOF) databases are essential for high-throughput screening, yet many reported crystal structures remain che

Class Activation Mapping in Explainable Computer Vision: A Method-Centered Review of CNN, Transformer, and Foundation-Model-Era Visual Explanations

Local AiDGX agent

arXiv:2608.12299v1 Announce Type: cross Abstract: Class activation mapping (CAM) is one of the most widely used visual explanation families in explainable artificial intelligence. Its purpose is intui

CYBER SLAYER — a 1995-style action trailer made with MiniMax H3, Wan 2.2 and ComfyUI

Model ReleasesDGX agent

Hey all! This started as a way to mark my 15th anniversary working in video game cinematics. I thought it would be fun to make a completely ridiculous, fictionalized version of how I got into the indu

Fine-Tuning Generative Models for Extreme Events via CVaR-Penalized Wasserstein Gradient Flows

TutorialsDGX agent

arXiv:2608.11544v1 Announce Type: cross Abstract: We propose CVaR-penalized Generative Particle Algorithm (CVaR-GPA), a robust, tail-agnostic algorithm for fine-tuning generative models to learn heavy

Governing Agentic AI in FinTech

Model ReleasesDGX agent

arXiv:2608.11344v1 Announce Type: cross Abstract: Financial institutions are delegating consequential decisions to agentic AI systems that decompose goals, coordinate models and tools, and act with li

Grounding Large Language Models as Generalizable Policies in Network Control

SafetyDGX agent

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital

Keep the Future, Drop the Rollout: RIFT for World Action Models

ApplicationsDGX agent

arXiv:2608.11521v1 Announce Type: cross Abstract: World action models (WAMs) condition robot actions on predicted futures, but iterative video rollout increases deployment latency. We ask whether acti

← Previous
1…167168169170171…1010
Next →