AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
Model Releases

Gemma:26b thinking issue in openWebUI

DGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

model-releasesr-ollama
13 Apr 2026
Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generalization and Scaling Laws for Mixture-of-Experts Transformers

DGX agent

arXiv:2604.09175v1 Announce Type: cross Abstract: We develop a theory of generalization and scaling for Mixture-of-Experts (MoE) Transformers that cleanly separates active per-input capacity from rout

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing

DGX agent

arXiv:2604.08896v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) have accelerated progress in domain-oriented AI, yet their development in geoscience and rem

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

GeoPAS: Geometric Probing for Algorithm Selection in Continuous Black-Box Optimisation

DGX agent

arXiv:2604.09095v1 Announce Type: new Abstract: Automated algorithm selection in continuous black-box optimisation typically relies on fixed landscape descriptors computed under a limited probing budg

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Got a doodle for your next project laying around? Turn it into working software using @GoogleAIStudio and Nano Banana. Watch us vibe code a …

DGX agent

Got a doodle for your next project laying around? Turn it into working software using @GoogleAIStudio and Nano Banana. Watch us vibe code a weather-responsive outfit selector app from a single, hand-d

model-releasesgoogle-ai--x
13 Apr 2026
Model Releases

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

DGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

great to see more open evals for an important problem

DGX agent

great to see more open evals for an important problem We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent that wo

model-releasesjerry-liu--x
13 Apr 2026
Model Releases

Grok

DGX agent

Grok Grok 4.20 Reasoning just took the #1 spot on the BridgeBench reasoning benchmark. 🔥 Beating GPT-5.4, Claude Opus 4.6, Google Gemini and others. Week after week, Grok keeps climbing across benchma

model-releaseselon-musk--x
13 Apr 2026
Model Releases

Have you tried the new Claude Code renderer? What has your experience been like? If you haven't, you can enable it with: CLAUDE_CODE_NO_FLIC…

DGX agent

Have you tried the new Claude Code renderer? What has your experience been like? If you haven't, you can enable it with: CLAUDE_CODE_NO_FLICKER=1 claude Today we're excited to announce NO_FLICKER mode

model-releasesthariq--x
13 Apr 2026
Model Releases

Hidden in Plain Sight: Visual-to-Symbolic Analytical Solution Inference from Field Visualizations

DGX agent

arXiv:2604.08863v1 Announce Type: new Abstract: Recovering analytical solutions of physical fields from visual observations is a fundamental yet underexplored capability for AI-assisted scientific rea

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

DGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

HiFloat4 Format for Language Model Pre-training on Ascend NPUs

DGX agent

arXiv:2604.08826v1 Announce Type: cross Abstract: Large foundation models have become central to modern machine learning, with performance scaling predictably with model size and data. However, traini

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

DGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing

DGX agent

arXiv:2604.08884v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have made significant strides in natural image understanding, their ability to perceive and reason over

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

How Should Video LLMs Output Time? An Analysis of Efficient Temporal Grounding Paradigms

DGX agent

arXiv:2604.08966v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have advanced Video Temporal Grounding (VTG), existing methods often couple output paradigms with differe

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

How to find the sweet spot between cost and performance

DGX agent

At Google Cloud, we often see customers asking themselves: 'How can we manage our generative AI costs effectively without sacrificing the performance and availability our applications demand?' This is

model-releasesgoogle-cloud-ai
13 Apr 2026
Model Releases

HTNav: A Hybrid Navigation Framework with Tiered Structure for Urban Aerial Vision-and-Language Navigation

DGX agent

arXiv:2604.08883v1 Announce Type: cross Abstract: Inspired by the general Vision-and-Language Navigation (VLN) task, aerial VLN has attracted widespread attention, owing to its significant practical v

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

DGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

model-releasesr-ollama
13 Apr 2026
Model Releases

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

DGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Improving Model Performance by Adapting the KGE Metric to Account for System Non-Stationarity

DGX agent

arXiv:2604.03906v2 Announce Type: replace Abstract: Geoscientific systems tend to be characterized by pronounced temporal non-stationarity, arising from seasonal and climatic variability in hydrometeo

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Inferring Latent Temporal Sparse Coordination Graph for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2403.19253v3 Announce Type: replace Abstract: Effective agent coordination is crucial in cooperative Multi-Agent Reinforcement Learning (MARL). While agent cooperation can be represented by grap

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Inpaint workflows for z-image, qwen and flux fill onereward

DGX agent

This Reddit post from r/StableDiffusion shares ComfyUI inpainting workflows for several modern AI image models, including Z-Image, Qwen Image/Edit, and Flux-series models . Flux Fill is a dedicated in

model-releasesr-stablediffusion
13 Apr 2026
Model Releases

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

DGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation

DGX agent

arXiv:2604.08797v1 Announce Type: cross Abstract: Stories are key to transmitting values across cultures, but their interpretation varies across linguistic and cultural contexts. Thus, we introduce mu

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Listener-Rewarded Thinking in VLMs for Image Preferences

DGX agent

arXiv:2506.22832v3 Announce Type: replace-cross Abstract: Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generat

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

DGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Looking for people with different hardware to help benchmark local LLM behavioral reliability

DGX agent

A Reddit post in the r/ollama community seeking volunteers with diverse hardware setups to participate in a collaborative effort to benchmark the **behavioral reliability** of locally-run large langua

model-releasesr-ollama
13 Apr 2026
Model Releases

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

DGX agent

arXiv:2604.08956v1 Announce Type: new Abstract: Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite da

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

DGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

LPLCv2: An Expanded Dataset for Fine-Grained License Plate Legibility Classification

DGX agent

arXiv:2604.08741v1 Announce Type: new Abstract: Modern Automatic License Plate Recognition (ALPR) systems achieve outstanding performance in controlled, well-defined scenarios. However, large-scale re

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation

DGX agent

arXiv:2604.09352v1 Announce Type: new Abstract: Monocular Depth Estimation (MDE) is crucial for autonomous lunar rover navigation using electro-optical cameras. However, deploying terrestrial MDE netw

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

DGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Many-Tier Instruction Hierarchy in LLM Agents

DGX agent

arXiv:2604.09443v1 Announce Type: cross Abstract: Large language model agents receive instructions from many sources-system messages, user prompts, tool outputs, and more-each carrying different level

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation

DGX agent

arXiv:2604.09514v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled the large-scale generation of highly fluent and deceptive news-like content. While prior wo

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

DGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs

DGX agent

arXiv:2604.09124v1 Announce Type: cross Abstract: Deploying DNNs on System-on-Chips (SoC) with multiple heterogeneous acceleration engines is challenging, and the majority of deployment frameworks can

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

DGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Medical Reasoning with Large Language Models: A Survey and MR-Bench

DGX agent

arXiv:2604.08559v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved strong performance on medical exam-style tasks, motivating growing interest in their deployment in real-wor

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

DGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

DGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

DGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

model-releasesharrison-chase--x
13 Apr 2026
Model Releases

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

DGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory

DGX agent

arXiv:2602.15313v2 Announce Type: replace Abstract: AI Memory, specifically how models organizes and retrieves historical messages, becomes increasingly valuable to Large Language Models (LLMs), yet e

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

DGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

DGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

DGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Natural Riemannian gradient for learning functional tensor networks

DGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression

DGX agent

arXiv:2604.08923v1 Announce Type: new Abstract: Dimensional Aspect-Based Sentiment Analysis (DimABSA) extends traditional ABSA from categorical polarity labels to continuous valence-arousal (VA) regre

model-releasesarxiv-cs-cl
13 Apr 2026
← Previous
1…444445446447448…461
Next →