AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
13 Apr 2026

Gemma:26b thinking issue in openWebUI

Model ReleasesDGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

Generalization and Scaling Laws for Mixture-of-Experts Transformers

Model ReleasesDGX agent

arXiv:2604.09175v1 Announce Type: cross Abstract: We develop a theory of generalization and scaling for Mixture-of-Experts (MoE) Transformers that cleanly separates active per-input capacity from rout

GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing

Model ReleasesDGX agent

arXiv:2604.08896v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) have accelerated progress in domain-oriented AI, yet their development in geoscience and rem


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GeoPAS: Geometric Probing for Algorithm Selection in Continuous Black-Box Optimisation

Model ReleasesDGX agent

arXiv:2604.09095v1 Announce Type: new Abstract: Automated algorithm selection in continuous black-box optimisation typically relies on fixed landscape descriptors computed under a limited probing budg

Got a doodle for your next project laying around? Turn it into working software using @GoogleAIStudio and Nano Banana. Watch us vibe code a …

Model ReleasesDGX agent

Got a doodle for your next project laying around? Turn it into working software using @GoogleAIStudio and Nano Banana. Watch us vibe code a weather-responsive outfit selector app from a single, hand-d

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

Model ReleasesDGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

great to see more open evals for an important problem

Model ReleasesDGX agent

great to see more open evals for an important problem We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent that wo

Grok

Model ReleasesDGX agent

Grok Grok 4.20 Reasoning just took the #1 spot on the BridgeBench reasoning benchmark. 🔥 Beating GPT-5.4, Claude Opus 4.6, Google Gemini and others. Week after week, Grok keeps climbing across benchma

Have you tried the new Claude Code renderer? What has your experience been like? If you haven't, you can enable it with: CLAUDE_CODE_NO_FLIC…

Model ReleasesDGX agent

Have you tried the new Claude Code renderer? What has your experience been like? If you haven't, you can enable it with: CLAUDE_CODE_NO_FLICKER=1 claude Today we're excited to announce NO_FLICKER mode

Hidden in Plain Sight: Visual-to-Symbolic Analytical Solution Inference from Field Visualizations

Model ReleasesDGX agent

arXiv:2604.08863v1 Announce Type: new Abstract: Recovering analytical solutions of physical fields from visual observations is a fundamental yet underexplored capability for AI-assisted scientific rea

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

Model ReleasesDGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

HiFloat4 Format for Language Model Pre-training on Ascend NPUs

Model ReleasesDGX agent

arXiv:2604.08826v1 Announce Type: cross Abstract: Large foundation models have become central to modern machine learning, with performance scaling predictably with model size and data. However, traini

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

Model ReleasesDGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing

Model ReleasesDGX agent

arXiv:2604.08884v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have made significant strides in natural image understanding, their ability to perceive and reason over

How Should Video LLMs Output Time? An Analysis of Efficient Temporal Grounding Paradigms

Model ReleasesDGX agent

arXiv:2604.08966v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have advanced Video Temporal Grounding (VTG), existing methods often couple output paradigms with differe

How to find the sweet spot between cost and performance

Model ReleasesDGX agent

At Google Cloud, we often see customers asking themselves: 'How can we manage our generative AI costs effectively without sacrificing the performance and availability our applications demand?' This is

HTNav: A Hybrid Navigation Framework with Tiered Structure for Urban Aerial Vision-and-Language Navigation

Model ReleasesDGX agent

arXiv:2604.08883v1 Announce Type: cross Abstract: Inspired by the general Vision-and-Language Navigation (VLN) task, aerial VLN has attracted widespread attention, owing to its significant practical v

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

Model ReleasesDGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

Model ReleasesDGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

Improving Model Performance by Adapting the KGE Metric to Account for System Non-Stationarity

Model ReleasesDGX agent

arXiv:2604.03906v2 Announce Type: replace Abstract: Geoscientific systems tend to be characterized by pronounced temporal non-stationarity, arising from seasonal and climatic variability in hydrometeo

Inferring Latent Temporal Sparse Coordination Graph for Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2403.19253v3 Announce Type: replace Abstract: Effective agent coordination is crucial in cooperative Multi-Agent Reinforcement Learning (MARL). While agent cooperation can be represented by grap

Inpaint workflows for z-image, qwen and flux fill onereward

Model ReleasesDGX agent

This Reddit post from r/StableDiffusion shares ComfyUI inpainting workflows for several modern AI image models, including Z-Image, Qwen Image/Edit, and Flux-series models . Flux Fill is a dedicated in

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

Model ReleasesDGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation

Model ReleasesDGX agent

arXiv:2604.08797v1 Announce Type: cross Abstract: Stories are key to transmitting values across cultures, but their interpretation varies across linguistic and cultural contexts. Thus, we introduce mu

Listener-Rewarded Thinking in VLMs for Image Preferences

Model ReleasesDGX agent

arXiv:2506.22832v3 Announce Type: replace-cross Abstract: Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generat

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

Model ReleasesDGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

Looking for people with different hardware to help benchmark local LLM behavioral reliability

Model ReleasesDGX agent

A Reddit post in the r/ollama community seeking volunteers with diverse hardware setups to participate in a collaborative effort to benchmark the **behavioral reliability** of locally-run large langua

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

Model ReleasesDGX agent

arXiv:2604.08956v1 Announce Type: new Abstract: Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite da

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

Model ReleasesDGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

LPLCv2: An Expanded Dataset for Fine-Grained License Plate Legibility Classification

Model ReleasesDGX agent

arXiv:2604.08741v1 Announce Type: new Abstract: Modern Automatic License Plate Recognition (ALPR) systems achieve outstanding performance in controlled, well-defined scenarios. However, large-scale re

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation

Model ReleasesDGX agent

arXiv:2604.09352v1 Announce Type: new Abstract: Monocular Depth Estimation (MDE) is crucial for autonomous lunar rover navigation using electro-optical cameras. However, deploying terrestrial MDE netw

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

Model ReleasesDGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

Many-Tier Instruction Hierarchy in LLM Agents

Model ReleasesDGX agent

arXiv:2604.09443v1 Announce Type: cross Abstract: Large language model agents receive instructions from many sources-system messages, user prompts, tool outputs, and more-each carrying different level

Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation

Model ReleasesDGX agent

arXiv:2604.09514v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled the large-scale generation of highly fluent and deceptive news-like content. While prior wo

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

Model ReleasesDGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs

Model ReleasesDGX agent

arXiv:2604.09124v1 Announce Type: cross Abstract: Deploying DNNs on System-on-Chips (SoC) with multiple heterogeneous acceleration engines is challenging, and the majority of deployment frameworks can

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

Medical Reasoning with Large Language Models: A Survey and MR-Bench

Model ReleasesDGX agent

arXiv:2604.08559v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved strong performance on medical exam-style tasks, motivating growing interest in their deployment in real-wor

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

Model ReleasesDGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

Model ReleasesDGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

Model ReleasesDGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Model ReleasesDGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory

Model ReleasesDGX agent

arXiv:2602.15313v2 Announce Type: replace Abstract: AI Memory, specifically how models organizes and retrieves historical messages, becomes increasingly valuable to Large Language Models (LLMs), yet e

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

Model ReleasesDGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

Model ReleasesDGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

Model ReleasesDGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

Natural Riemannian gradient for learning functional tensor networks

Model ReleasesDGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression

Model ReleasesDGX agent

arXiv:2604.08923v1 Announce Type: new Abstract: Dimensional Aspect-Based Sentiment Analysis (DimABSA) extends traditional ABSA from categorical polarity labels to continuous valence-arousal (VA) regre

Neural networks for Text-to-Speech evaluation

Model ReleasesDGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima

Model ReleasesDGX agent

arXiv:2604.09258v1 Announce Type: new Abstract: Pretraining is the cornerstone of Large Language Models (LLMs), dominating the vast majority of computational budget and data to serve as the primary en

Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs

Model ReleasesDGX agent

arXiv:2604.09021v1 Announce Type: cross Abstract: Auditory large language models (ALLMs) have demonstrated strong general capabilities in audio understanding and reasoning tasks. However, their reliab

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

Model ReleasesDGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

Ocr benchmark

Model ReleasesDGX agent

Ocr benchmark We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent that works with real-world files. ParseBench is

Ollama 0.20.6 is here with improved Gemma 4 tool calling! more improvements to come for Gemma 4!

Model ReleasesDGX agent

Ollama version 0.20.6 has been released, featuring improved tool calling support for Google's Gemma 4 model. The update focuses on enhancing the reliability and functionality of function/tool calling

Ollama / Mistral with MCP to Mempalace

Model ReleasesDGX agent

This Reddit post from r/ollama discusses integrating Ollama-served Mistral with MemPalace — a free, locally-run AI memory system — via the Model Context Protocol (MCP). MemPalace runs entirely on a us

On Semiotic-Grounded Interpretive Evaluation of Generative Art

Model ReleasesDGX agent

arXiv:2604.08641v1 Announce Type: cross Abstract: Interpretation is essential to deciphering the language of art: audiences communicate with artists by recovering meaning from visual artifacts. Howeve

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

Model ReleasesDGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

Online Intention Prediction via Control-Informed Learning

Model ReleasesDGX agent

arXiv:2604.09303v1 Announce Type: cross Abstract: This paper presents an online intention prediction framework for estimating the goal state of autonomous systems in real time, even when intention is

Open-source AI is now matching GPT-4 — and you can run it privately for free

Model ReleasesDGX agent

This Reddit post discusses how open-source AI models have advanced to rival GPT-4-level performance and can be run locally for free using Ollama — a tool that lets users download and manage large lang

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence

Model ReleasesDGX agent

arXiv:2603.11178v3 Announce Type: replace Abstract: Standard LLM distillation treats all training problems equally -- wasting compute on problems the student has already mastered or cannot yet solve.

← Previous
1…355356357358359…369
Next →