AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
24 Apr 2026

Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery

Model ReleasesDGX agent

arXiv:2604.21102v1 Announce Type: cross Abstract: We present a novel framework for automatically evaluating building conditions nationwide in the United States by leveraging large language models (LLM

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in …

Model ReleasesDGX agent

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in production. 🙌🚀 Can't wait for the @garrytan @demishassabis f

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

LLaDA2.0-Uni Released

Model ReleasesDGX agent

LLaDA2.0-Uni is a unified diffusion large language model (dLLM) based on Mixture-of-Experts architecture that seamlessly integrates multimodal understanding and generation. The model supports text-to-

llm 0.31

Model ReleasesDGX agent

Release: llm 0.31 New GPT-5.5 OpenAI model: llm -m gpt-5.5. #1418 New option to set the text verbosity level for GPT-5+ OpenAI models: -o verbosity low. Values are low, medium, high. New option for se

Low-Rank Adaptation Redux for Large Models

Model ReleasesDGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

Model ReleasesDGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

MathDuels: Evaluating LLMs as Problem Posers and Solvers

Model ReleasesDGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

Model ReleasesDGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

Model ReleasesDGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

Model ReleasesDGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

Model ReleasesDGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

Model ReleasesDGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in …

Model ReleasesDGX agent

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in their categories while benchmarking close to the frontier mo

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

Model ReleasesDGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

Model ReleasesDGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

My first two TiKZ Sparks unicorns from DeepSeek v4. (Expert mode, from the DeepSeek site, which is supposed to be v4 Pro according to the re…

Model ReleasesDGX agent

Ethan Mollick documents his first attempts at using DeepSeek v4's expert mode to generate TiKZ code for creating unicorn graphics, sharing results from the DeepSeek website's v4 Pro interface. The pos

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

Model ReleasesDGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

Necessity is the mother of kv-cache optimisation

Model ReleasesDGX agent

Necessity is the mother of kv-cache optimisation I’m still amazed that DeepSeek, Kimi, and Qwen can train very strong LLMs with far fewer and often nerfed NVIDIA GPUs, or even Huawei chips. DeepSeek V

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

Model ReleasesDGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

Model ReleasesDGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud.

Model ReleasesDGX agent

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud. 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 Dee

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

Model ReleasesDGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

Model ReleasesDGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification

Model ReleasesDGX agent

arXiv:2604.21602v1 Announce Type: cross Abstract: Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modes

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

Model ReleasesDGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway

Model ReleasesDGX agent

OpenAI's GPT-5.5 model is now available on Databricks' platform with governance capabilities provided through Unity AI Gateway, enabling enterprises to deploy and manage the model within their data in

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

Model ReleasesDGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

Model ReleasesDGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

Model ReleasesDGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

pi gives you wings wherever you are

Model ReleasesDGX agent

pi gives you wings wherever you are This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro Fo

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together!

Model ReleasesDGX agent

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together! We built a completely free CLI agent with @badlogicgames's Pi agent, @ollama (Gemma 4), and Para

Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance

Model ReleasesDGX agent

arXiv:2604.21104v1 Announce Type: new Abstract: New geospatial foundation models introduce a new model architecture and pretraining dataset, often sampled using different notions of data diversity. Pe

PREVENT-JACK: Context Steering for Swarms of Long Heavy Articulated Vehicles

Model ReleasesDGX agent

arXiv:2604.21337v1 Announce Type: new Abstract: In this paper, we aim to extend the traditional point-mass-like robot representation in swarm robotics and instead study a swarm of long Heavy Articulat

Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence

Model ReleasesDGX agent

arXiv:2106.01254v3 Announce Type: replace Abstract: In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in s

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

【Qwen-3.6-27B × llama.cpp】生成速度10倍の革命的スピード! Qwen-3.6-27Bで生成速度が約10倍の136.75 t/sに到達する驚異の手法が話題です!🚀 llama.cppの「ngram-mod」という投機的デコード(Speculative D…

Model ReleasesDGX agent

【Qwen-3.6-27B × llama.cpp】生成速度10倍の革命的スピード! Qwen-3.6-27Bで生成速度が約10倍の136.75 t/sに到達する驚異の手法が話題です!🚀 llama.cppの「ngram-mod」という投機的デコード(Speculative Decoding)を活用。過去の出力パターンを利用して次に来る言葉を予測し、追加のビデオメモリをほぼ消費せずに高速化を実現し

RailVQA: A Benchmark and Framework for Efficient Interpretable Visual Cognition in Automatic Train Operation

Model ReleasesDGX agent

arXiv:2603.27112v2 Announce Type: replace Abstract: As Automatic Train Operation (ATO) advances toward GoA4 and beyond, it increasingly depends on efficient, reliable cab-view visual perception and de

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

Model ReleasesDGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

RealRoute: Dynamic Query Routing System via Retrieve-then-Verify Paradigm

Model ReleasesDGX agent

arXiv:2604.20860v1 Announce Type: cross Abstract: Despite the success of Retrieval-Augmented Generation (RAG) in grounding LLMs with external knowledge, its application over heterogeneous sources (e.g

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

Model ReleasesDGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

Rectified Schrodinger Bridge Matching for Few-Step Visual Navigation

Model ReleasesDGX agent

arXiv:2604.05673v2 Announce Type: replace-cross Abstract: Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into cont

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations

Model ReleasesDGX agent

arXiv:2509.25868v3 Announce Type: replace Abstract: The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And C

Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment

Model ReleasesDGX agent

arXiv:2604.21160v1 Announce Type: new Abstract: Point-Vision-Language Models promise to empower embodied agents with executable spatial reasoning, yet they frequently succumb to geometric hallucinatio

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

Model ReleasesDGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

Remember o3 was only a year and a week ago! Also, only GPT-5.5 seemed to take the 'evolution' piece seriously and change the setting rather …

Model ReleasesDGX agent

I cannot provide an accurate summary for this entry as the text appears incomplete and lacks sufficient context. The post fragment references o3 (likely an AI model), GPT-5.5, and discusses timeline/e

Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis

Model ReleasesDGX agent

arXiv:2511.11439v2 Announce Type: replace-cross Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance ofte

Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors

Model ReleasesDGX agent

arXiv:2604.21893v1 Announce Type: cross Abstract: Geographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constrainin

RewardBench 2: Advancing Reward Model Evaluation

Model ReleasesDGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

Robust Test-time Video-Text Retrieval: Benchmarking and Adapting for Query Shifts

Model ReleasesDGX agent

arXiv:2604.20851v1 Announce Type: cross Abstract: Modern video-text retrieval (VTR) models excel on in-distribution benchmarks but are highly vulnerable to real-world query shifts, where the distribut

SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors

Model ReleasesDGX agent

arXiv:2511.18264v3 Announce Type: replace Abstract: Existing satellite video tracking methods often struggle with generalization, requiring scenario-specific training to achieve satisfactory performan

Scaling of Gaussian Kolmogorov--Arnold Networks

Model ReleasesDGX agent

arXiv:2604.21174v1 Announce Type: cross Abstract: The Gaussian scale parameter (epsilon) is central to the behavior of Gaussian Kolmogorov--Arnold Networks (KANs), yet its role in deep edge-based arch

SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation

Model ReleasesDGX agent

arXiv:2411.17061v2 Announce Type: replace Abstract: The Vision Transformer (ViT) has achieved notable success in computer vision, with its variants widely validated across various downstream tasks, in

SCM: Sleep-Consolidated Memory with Algorithmic Forgetting for Large Language Models

Model ReleasesDGX agent

arXiv:2604.20943v1 Announce Type: new Abstract: We present SCM (Sleep-Consolidated Memory), a research preview of a memory architecture for large language models that draws on neuroscientific principl

Secure LLM Fine-Tuning via Safety-Aware Probing

Model ReleasesDGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

Model ReleasesDGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

Model ReleasesDGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding

Model ReleasesDGX agent

arXiv:2510.26615v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but it must balance limited effective context, re

← Previous
1…313314315316317…373
Next →