AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
13 Apr 2026

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have hi…

ApplicationsDGX agent

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have high demand, which is outstripping supply (so prices will go u

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications
DGX agent

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench Reasoning • #1 on BridgeBench Speed • #1 on BridgeBench Low

How Noise Benefits AI-generated Image Detection

ResearchDGX agent

arXiv:2511.16136v2 Announce Type: replace Abstract: The rapid advancement of generative models has made real and synthetic images increasingly indistinguishable. Although extensive efforts have been d

Is an nvidia DGK Spark or similar worth it?

HardwareDGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

ResearchDGX agent

arXiv:2604.09389v1 Announce Type: cross Abstract: Training Transformer language models is expensive, as performance typically improves with increasing dataset size and computational budget. Although s

LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation

ApplicationsDGX agent

arXiv:2604.09360v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Model (LLM) providers--each exposing proprietary API formats--has created a fragmented ecosystem where appli

Looking for people with different hardware to help benchmark local LLM behavioral reliability

Model ReleasesDGX agent

A Reddit post in the r/ollama community seeking volunteers with diverse hardware setups to participate in a collaborative effort to benchmark the **behavioral reliability** of locally-run large langua

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

Model ReleasesDGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

ResearchDGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

Mechanisms of Introspective Awareness

SafetyDGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

Model ReleasesDGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

Model ReleasesDGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

SafetyDGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation

AgentsDGX agent

arXiv:2604.08782v1 Announce Type: new Abstract: Large language models (LLMs) suffer significant performance degradation when user instructions and context are distributed over multiple conversational

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

Model ReleasesDGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

Natural Riemannian gradient for learning functional tensor networks

Model ReleasesDGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

SafetyDGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution

Local AiDGX agent

arXiv:2502.06809v3 Announce Type: replace-cross Abstract: Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for m

Offline-First LLM Architecture for Adaptive Learning in Low-Connectivity Environments

Local AiDGX agent

arXiv:2603.03339v5 Announce Type: replace-cross Abstract: Artificial intelligence (AI) and large language models (LLMs) are transforming educational technology by enabling conversational tutoring, per

Ollama 0.20.6 is here with improved Gemma 4 tool calling! more improvements to come for Gemma 4!

Model ReleasesDGX agent

Ollama version 0.20.6 has been released, featuring improved tool calling support for Google's Gemma 4 model. The update focuses on enhancing the reliability and functionality of function/tool calling

p1: Better Prompt Optimization with Fewer Prompts

ResearchDGX agent

arXiv:2604.08801v1 Announce Type: cross Abstract: Prompt optimization improves language models without updating their weights by searching for a better system prompt, but its effectiveness varies wide

Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search

Model ReleasesDGX agent

arXiv:2604.08598v1 Announce Type: cross Abstract: Text-based person search faces inherent limitations due to data scarcity, driven by stringent privacy constraints and the high cost of manual annotati

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

SafetyDGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

SafetyDGX agent

arXiv:2604.09474v1 Announce Type: cross Abstract: Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise,

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

SafetyDGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

Model ReleasesDGX agent

arXiv:2604.09532v1 Announce Type: cross Abstract: Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten

Spectral Geometry of LoRA Adapters Encodes Training Objective and Predicts Harmful Compliance

Model ReleasesDGX agent

arXiv:2604.08844v1 Announce Type: new Abstract: We study whether low-rank spectral summaries of LoRA weight deltas can identify which fine-tuning objective was applied to a language model, and whether

Streaming Video Instruction Tuning

ResearchDGX agent

arXiv:2512.21334v2 Announce Type: replace Abstract: We present Streamo, a real-time streaming video LLM that serves as a general-purpose interactive assistant. Unlike existing online video models that

TaxPraBen: A Scalable Benchmark for Structured Evaluation of LLMs in Chinese Real-World Tax Practice

Model ReleasesDGX agent

arXiv:2604.08948v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel in various general domains, they exhibit notable gaps in the highly specialized, knowledge-intensive, and legal

Thinking about trying Ollama Pro — how does it compare to Claude/Codex?

Model ReleasesDGX agent

This Reddit thread from r/ollama discusses user perspectives on **Ollama Pro** as a paid/upgraded tier compared to cloud-based AI coding assistants like Anthropic's Claude and OpenAI's Codex, likely f

Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation

SafetyDGX agent

arXiv:2604.09368v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as scalable user simulators for recommender system evaluation. Yet existing simulators per

TiAb Review Plugin: A Browser-Based Tool for AI-Assisted Title and Abstract Screening

Model ReleasesDGX agent

arXiv:2604.08602v1 Announce Type: cross Abstract: Background: Server-based screening tools impose subscription costs, while open-source alternatives require coding skills. Objectives: We developed a b

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

HardwareDGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2604.08549v1 Announce Type: cross Abstract: We introduce VerifAI, an open-source expert system for biomedical question answering that integrates retrieval-augmented generation (RAG) with a novel

Visually-Guided Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning

Model ReleasesDGX agent

arXiv:2604.08639v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ m

Which Pieces Does Unigram Tokenization Really Need?

Model ReleasesDGX agent

arXiv:2512.12641v2 Announce Type: replace Abstract: The Unigram tokenization algorithm offers a probabilistic alternative to the greedy heuristics of Byte-Pair Encoding. Despite its theoretical elegan

12 Apr 2026

Hermes Agent + Ollama returns tool JSON but doesn’t actually execute anything

Local AiDGX agent

Users building agentic pipelines with Hermes models in Ollama report an issue where the model correctly generates tool call JSON in its response but the actual tool functions are never invoked or exec

Marcus Hutchins, the guy famous for stopping the WannaCry Ransomware, probably has the best take on Mythos doing vulnerability research

ResearchDGX agent

Marcus Hutchins, the cybersecurity researcher known for halting the 2017 WannaCry ransomware attack by registering a kill-switch domain, is cited as offering a notable perspective on AI systems conduc

Slay The Spire 2 - Flux.2 Klein 9b style LORAs

Local AiDGX agent

This r/StableDiffusion post covers community-created style LoRA adapters trained on the visual art style of *Slay The Spire 2*, built for use with Black Forest Labs' FLUX.2 Klein 9B model — a 9-billio

Tansan (Anime Portrait) LoRA for ZiT

Local AiDGX agent

'Tansan (Anime Portrait) LoRA for ZiT' is a community-shared Stable Diffusion LoRA model posted on r/StableDiffusion, designed to generate anime-style portrait images and optimized for use with the Zi

The mysterious science of LoRA training (sdxl)

Local AiDGX agent

This r/StableDiffusion post explores the nuanced and often empirical process of training LoRA (Low-Rank Adaptation) models on Stable Diffusion XL (SDXL), a fine-tuning technique that, rather than retr

Where to find complete illustrious/NoobAI character keywords ?

Local AiDGX agent

This r/StableDiffusion thread discusses how to locate complete lists of character prompt keywords compatible with the Illustrious and NoobAI XL Stable Diffusion models, which are trained on Danbooru a

11 Apr 2026

🔗 Codex App: https://chatgpt.com/codex/

Model ReleasesDGX agent

OpenAI's Codex App (available at chatgpt.com/codex) is a dedicated command center for agentic coding, enabling developers to manage multiple AI coding agents working in parallel across projects. T...

10 Apr 2026

3DrawAgent: Teaching LLM to Draw in 3D with Early Contrastive Experience

Model ReleasesDGX agent

arXiv:2604.08042v1 Announce Type: new Abstract: Sketching in 3D space enables expressive reasoning about shape, structure, and spatial relationships, yet generating 3D sketches through natural languag

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

SafetyDGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

ACIArena: Toward Unified Evaluation for Agent Cascading Injection

Model ReleasesDGX agent

arXiv:2604.07775v1 Announce Type: cross Abstract: Collaboration and information sharing empower Multi-Agent Systems (MAS) but also introduce a critical security risk known as Agent Cascading Injection

AdaProb: Efficient Machine Unlearning via Adaptive Probability

ResearchDGX agent

arXiv:2411.02622v3 Announce Type: replace-cross Abstract: Machine unlearning, enabling a trained model to forget specific data, is crucial for addressing erroneous data and adhering to privacy regulat

Adversarial Evasion Attacks on Computer Vision using SHAP Values

ResearchDGX agent

arXiv:2601.10587v2 Announce Type: replace Abstract: The paper introduces a white-box attack on computer vision models using SHAP values. It demonstrates how adversarial evasion attacks can compromise

arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation

Model ReleasesDGX agent

arXiv:2504.10284v5 Announce Type: replace Abstract: Literature review tables are essential for summarizing and comparing collections of scientific papers. In this paper, we study the automatic generat

ATANT: An Evaluation Framework for AI Continuity

Model ReleasesDGX agent

arXiv:2604.06710v1 Announce Type: new Abstract: We present ATANT (Automated Test for Acceptance of Narrative Truth), an open evaluation framework for measuring continuity in AI systems: the ability to

Beyond Surface Artifacts: Capturing Shared Latent Forgery Knowledge Across Modalities

Model ReleasesDGX agent

arXiv:2604.07763v1 Announce Type: new Abstract: As generative artificial intelligence evolves, deepfake attacks have escalated from single-modality manipulations to complex, multimodal threats. Existi

BiScale-GTR: Fragment-Aware Graph Transformers for Multi-Scale Molecular Representation Learning

Model ReleasesDGX agent

arXiv:2604.06336v1 Announce Type: cross Abstract: Graph Transformers have recently attracted attention for molecular property prediction by combining the inductive biases of graph neural networks (GNN

Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses

Model ReleasesDGX agent

arXiv:2604.06216v1 Announce Type: cross Abstract: As LLM-powered chatbots are increasingly deployed in mental health services, detecting hallucinations and omissions has become critical for user safet

ClawBench: Can AI Agents Complete Everyday Online Tasks?

Model ReleasesDGX agent

arXiv:2604.08523v1 Announce Type: new Abstract: AI agents may be able to automate your inbox, but can they automate other routine aspects of your life? Everyday online tasks offer a realistic yet unso

CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs

ApplicationsDGX agent

arXiv:2510.12184v2 Announce Type: replace Abstract: Recently, efficient Multimodal Large Language Models (MLLMs) have gained significant attention as a solution to their high computational complexity,

ControlNet vs LoRA

Local AiDGX agent

LoRA (Low-Rank Adaptation) and ControlNet are complementary but fundamentally different tools for controlling Stable Diffusion image generation. LoRA modifies a model's weights to teach it new sty...

CoreWeave inks multiyear cloud deal with Anthropic

Model ReleasesDGX agent

CoreWeave Inc. today announced that it has won a multiyear contract to supply Anthropic PBC with cloud infrastructure. The company’s shares closed 11% higher on the news. The data center capacity comm

CryoSplat: Gaussian Splatting for Cryo-EM Homogeneous Reconstruction

Model ReleasesDGX agent

arXiv:2508.04929v4 Announce Type: replace-cross Abstract: As a critical modality for structural biology, cryogenic electron microscopy (cryo-EM) facilitates the determination of macromolecular structu

← Previous
1…433434435436437…1060
Next →