AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

Intent Laundering: AI Safety Datasets Are Not What They Seem

DGX agent

arXiv:2602.16729v3 Announce Type: replace-cross Abstract: We systematically evaluate the quality of widely used adversarial safety datasets from two perspectives: in isolation and in practice. In isol

model-releasesarxiv-cs-ai
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

DGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can …

DGX agent

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit from reli

model-releasestogether-ai--x
24 Apr 2026
Model Releases

IRIS: Interpolative Renyi Iterative Self-play for Large Language Model Fine-Tuning

DGX agent

arXiv:2604.20933v1 Announce Type: cross Abstract: Self-play fine-tuning enables large language models to improve beyond supervised fine-tuning without additional human annotations by contrasting annot

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Is anyone using models to describe an image and get a prompt? Is there much difference between Qwen 3.5 9b vs Qwen 3.5 27b, vs gemma 4 27b and another model you use ?

DGX agent

I'd need to search for this specific Reddit discussion to provide an accurate summary of what was actually discussed. Let me retrieve that information. This Reddit post discusses using AI vision model

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

It's High Time: A Survey of Temporal Question Answering

DGX agent

arXiv:2505.20243v4 Announce Type: replace Abstract: Time plays a critical role in how information is generated, retrieved, and interpreted. In this survey, we provide a comprehensive overview of Tempo

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

DGX agent

arXiv:2603.28342v2 Announce Type: replace Abstract: We present Kernel-Smith, a framework for high-performance GPU kernel and operator generation that combines a stable evaluation-driven evolutionary a

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems

DGX agent

arXiv:2508.10177v3 Announce Type: replace Abstract: Recent Large Language Model (LLM)-based AutoML systems demonstrate impressive capabilities but face significant limitations such as constrained expl

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Language as a Latent Variable for Reasoning Optimization

DGX agent

arXiv:2604.21593v1 Announce Type: new Abstract: As LLMs reduce English-centric bias, a surprising trend emerges: non-English responses sometimes outperform English on reasoning tasks. We hypothesize t

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Latent Denoising Improves Visual Alignment in Large Multimodal Models

DGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Learning to Communicate: Toward End-to-End Optimization of Multi-Agent Language Systems

DGX agent

arXiv:2604.21794v1 Announce Type: new Abstract: Multi-agent systems built on large language models have shown strong performance on complex reasoning tasks, yet most work focuses on agent roles and or

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Let's go DeepSeek v4!!! 🚗🚀

DGX agent

Let's go DeepSeek v4!!! 🚗🚀 DeepSeek V4 by @deepseek_ai just dropped! SGLang is ready on Day 0 with a full stack of optimizations from architectures to low-level kernels. We also deliver a verified RL

model-releasesollama--x
24 Apr 2026
Model Releases

Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery

DGX agent

arXiv:2604.21102v1 Announce Type: cross Abstract: We present a novel framework for automatically evaluating building conditions nationwide in the United States by leveraging large language models (LLM

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in …

DGX agent

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in production. 🙌🚀 Can't wait for the @garrytan @demishassabis f

model-releasesollama--x
24 Apr 2026
Model Releases

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

DGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

LLaDA2.0-Uni Released

DGX agent

LLaDA2.0-Uni is a unified diffusion large language model (dLLM) based on Mixture-of-Experts architecture that seamlessly integrates multimodal understanding and generation. The model supports text-to-

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

llm 0.31

DGX agent

Release: llm 0.31 New GPT-5.5 OpenAI model: llm -m gpt-5.5. #1418 New option to set the text verbosity level for GPT-5+ OpenAI models: -o verbosity low. Values are low, medium, high. New option for se

model-releasessimon-willison
24 Apr 2026
Model Releases

Low-Rank Adaptation Redux for Large Models

DGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

DGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

MathDuels: Evaluating LLMs as Problem Posers and Solvers

DGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

DGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

DGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

DGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

DGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

DGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

DGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

model-releasesollama--x
24 Apr 2026
Model Releases

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in …

DGX agent

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in their categories while benchmarking close to the frontier mo

model-releasessimon-willison--x
24 Apr 2026
Model Releases

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

DGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

DGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

My first two TiKZ Sparks unicorns from DeepSeek v4. (Expert mode, from the DeepSeek site, which is supposed to be v4 Pro according to the re…

DGX agent

Ethan Mollick documents his first attempts at using DeepSeek v4's expert mode to generate TiKZ code for creating unicorn graphics, sharing results from the DeepSeek website's v4 Pro interface. The pos

model-releasesethan-mollick--x
24 Apr 2026
Model Releases

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

DGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Necessity is the mother of kv-cache optimisation

DGX agent

Necessity is the mother of kv-cache optimisation I’m still amazed that DeepSeek, Kimi, and Qwen can train very strong LLMs with far fewer and often nerfed NVIDIA GPUs, or even Huawei chips. DeepSeek V

model-releasesemad-mostaque--x
24 Apr 2026
Model Releases

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

DGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

DGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud.

DGX agent

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud. 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 Dee

model-releasesollama--x
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification

DGX agent

arXiv:2604.21602v1 Announce Type: cross Abstract: Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modes

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

DGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway

DGX agent

OpenAI's GPT-5.5 model is now available on Databricks' platform with governance capabilities provided through Unity AI Gateway, enabling enterprises to deploy and manage the model within their data in

model-releasesdatabricks
24 Apr 2026
Model Releases

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

DGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

DGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

DGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

model-releasesgoogle-ai--x
24 Apr 2026
Model Releases

pi gives you wings wherever you are

DGX agent

pi gives you wings wherever you are This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro Fo

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together!

DGX agent

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together! We built a completely free CLI agent with @badlogicgames's Pi agent, @ollama (Gemma 4), and Para

model-releasesollama--x
24 Apr 2026
Model Releases

Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance

DGX agent

arXiv:2604.21104v1 Announce Type: new Abstract: New geospatial foundation models introduce a new model architecture and pretraining dataset, often sampled using different notions of data diversity. Pe

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

PREVENT-JACK: Context Steering for Swarms of Long Heavy Articulated Vehicles

DGX agent

arXiv:2604.21337v1 Announce Type: new Abstract: In this paper, we aim to extend the traditional point-mass-like robot representation in swarm robotics and instead study a swarm of long Heavy Articulat

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence

DGX agent

arXiv:2106.01254v3 Announce Type: replace Abstract: In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in s

model-releasesarxiv-cs-lg
24 Apr 2026
← Previous
1…395396397398399…470
Next →