AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,001 results
Research

SoftMCC: An MCC-Brier Calibration Bridge for Threshold-Free Model Selection under Class Imbalance

DGX agent

arXiv:2608.08984v1 Announce Type: new Abstract: Model selection for imbalanced binary classification often uses the Matthews correlation coefficient (MCC), but thresholding makes validation rankings t

researcharxiv-cs-lg
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Subjective Multi-Bias Detection with Large Language Models

DGX agent

arXiv:2608.09126v1 Announce Type: new Abstract: In this project, we delved into the pervasive challenge of bias detection within the text content. More specifically, our focus lies on the identificati

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework

DGX agent

arXiv:2601.21747v4 Announce Type: replace-cross Abstract: Sepsis remains one of the most complex and heterogeneous syndromes in intensive care. While deep learning models achieve competitive performan

researcharxiv-cs-ai
11 Aug 2026
Agents

The Belief-Desire-Intention Ontology for modelling mental reality and agency

DGX agent

arXiv:2511.17162v2 Announce Type: replace Abstract: The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, it

agentsarxiv-cs-ai
11 Aug 2026
Local Ai

The Spectral Neuron

DGX agent

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape

local-aiarxiv-cs-lg
11 Aug 2026
Model Releases

Theory-Guided Deception Detection: A RAG-Based Artificial Intelligence Exploration

DGX agent

arXiv:2608.08881v1 Announce Type: new Abstract: The current work developed seven Retrieval-Augmented Generation (RAG) models based on leading deception theories and compared how deception judgments we

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility Prediction

DGX agent

arXiv:2507.00945v2 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet e

tutorialsarxiv-cs-lg
11 Aug 2026
Research

Twin Rollouts: Noise-Coupled Counterfactual Branching in Interactive Video World Models

DGX agent

arXiv:2608.08982v1 Announce Type: new Abstract: Interactive video world models generate rollouts autoregressively under an action stream, yet they are trained and evaluated almost exclusively on factu

researcharxiv-cs-lg
11 Aug 2026
Model Releases

UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models

DGX agent

arXiv:2608.08627v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers expand recommendation capacity through conditional computation, yet a trained checkpoint still stores and routes

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

v0.32.9

DGX agent

NVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for that execution layer of always-on agents. It is designed f

model-releasesollama-releases
11 Aug 2026
Applications

Visual Distortion Detection in UGC Images Using Large Multimodal Models

DGX agent

arXiv:2608.09122v1 Announce Type: cross Abstract: The localized depiction of perceptual quality has long been a crucial, yet underexplored, challenge in image quality assessment (IQA). Existing approa

applicationsarxiv-cs-ai
11 Aug 2026
Research

WA-SpecDec: World-Aware Speculative Decoding for Vision-Language-Action Models

DGX agent

arXiv:2608.08725v1 Announce Type: new Abstract: Vision-language-action (VLA) policies generate robot controls autoregressively, making closed-loop latency dominated by repeated target-model forward pa

researcharxiv-cs-ro
11 Aug 2026
Model Releases

We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8× RTX 5090

DGX agent

We converted the model from the original safetensors and found two issues. The first one made our quantization fail several times, the second one does not fail at all, it just quietly ruins the base 1

model-releasesr-localllama
11 Aug 2026
Agents

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

DGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

b10342

DGX agent

model : Granite-Switch Architecture (#25107) granite-switch: add llama.cpp backend (POC, CPU) New 'granite-switch' architecture: a dense, all-attention Granite-4.1 model with N embedded LoRA adapters

model-releasesllama-cpp-releases
10 Aug 2026
Local Ai

Best Local LLMs - August 2026

DGX agent

Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the closed frontier, Opus level models on non-insane hardwa

local-air-localllama
10 Aug 2026
Research

EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs

DGX agent

arXiv:2608.06398v1 Announce Type: new Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by grouping bytes into dynamically sized patch

researcharxiv-cs-ai
10 Aug 2026
Local Ai

How Malachyte solves retail’s cold-start problem with managed real-time AI

DGX agent

What’s the best way to recommend products to little-known users? We’ve spent our careers trying to solve this problem for major companies like Spotify and Priceline, and it’s why Sidd founded Malachyt

local-aigoogle-cloud-ai
10 Aug 2026
Model Releases

I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8

DGX agent

There's an interactive chart and some extra data in the blog post if you're interested. There are plenty of KL-divergence benchmarks for GGUF models, but most of them compare one GGUF quant against an

model-releasesr-localllama
10 Aug 2026
Tutorials

Improving Attributed Long-form Question Answering with Intent Awareness

DGX agent

arXiv:2603.27435v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used to generate comprehensive, knowledge-intensive reports. However, while these models a

tutorialsarxiv-cs-ai
10 Aug 2026
Local Ai

M2-SMap: Memory-Efficient Semantic Mapping with Hierarchical Multi-Model Representation

DGX agent

arXiv:2608.07074v1 Announce Type: new Abstract: Dense point cloud maps, as a typically used mapping representation, are difficult to deploy on resource-constrained robots because their memory consumpt

local-aiarxiv-cs-ro
10 Aug 2026
Local Ai

MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model

DGX agent

arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy

local-aiarxiv-cs-lg
10 Aug 2026
Local Ai

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-G…

DGX agent

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF 1/ big announcement today: we will be releas

local-aisimon-willison--x
10 Aug 2026
Safety

PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model

DGX agent

arXiv:2608.06794v1 Announce Type: new Abstract: While diffusion models have made significant progress in text-to-image tasks, they still exhibit limitations when directly optimizing downstream objecti

safetyarxiv-cs-cv
10 Aug 2026
Research

Provable Training Data Identification for Large Language Models

DGX agent

arXiv:2510.09717v3 Announce Type: replace-cross Abstract: Identifying training data of large-scale models is critical for copyright litigation, privacy auditing, and ensuring fair evaluation. However,

researcharxiv-cs-ai
10 Aug 2026
Research

Recipes for Creativity: Iterative Generation and Evaluation in Large Language Models

DGX agent

arXiv:2608.07243v1 Announce Type: new Abstract: Generative models are often evaluated through singular artifacts, whereas human creativity typically emerges through iterative generation, appraisal, an

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA

DGX agent

Meta released Muse Glimmer, a 30‑billion‑parameter dense language model with a context window exceeding 120 K tokens, designed for local, long‑running agentic AI workloads. The model is optimized to r

model-releasesnvidia-developer
10 Aug 2026
Model Releases

Summary of Takeaways from the Minimax AMA

DGX agent

Summary from https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/ama_minimax_h3_team_ask_us_anything_about_our/ This summary was compiled with AI but cross-checked manually by me for accuracy. I

model-releasesr-stablediffusion
10 Aug 2026
Model Releases

Building a budget 32GB → 48GB VRAM home AI server: 2-3x RX 9060 XT 16GB vs RTX 5060 Ti 16GB, AM5 vs used EPYC?

DGX agent

I’m planning a dedicated home AI server, mainly for local LLM inference, agents/tool use, Docker services, and eventually larger MoE models with CPU offload. My plan is to start with 2x 16GB GPUs = 32

model-releasesr-localllama
8 Aug 2026
Model Releases

Showoff Saturday: Local 4x 6000 Pro (multi-year progression)

DGX agent

Not the biggest or shiniest, but it's mine From gaming machine inference on the original llama models, to a 4x RTX 6000 Pro Max Q + 4x 3090s local AI cluster. Pictures are in reverse chronological ord

model-releasesr-localllama
8 Aug 2026
Model Releases

Adapting Vision Foundation Models with Cascaded Semantics

DGX agent

arXiv:2608.05393v1 Announce Type: new Abstract: Prompt tuning, a leading parameter-efficient adaptation paradigm in NLP, has recently been extended to computer vision. Visual prompt tuning (VPT) adapt

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

DGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

safetyarxiv-cs-ai
7 Aug 2026
Tutorials

Cautious Context Steering for Language Model Personalization

DGX agent

arXiv:2608.05813v1 Announce Type: new Abstract: Personalizing language models (LMs) to individual user preferences is essential for aligning responses with diverse goals and backgrounds. Existing meth

tutorialsarxiv-cs-ai
7 Aug 2026
Agents

Disentangling 3D Modeling from Spatial Reasoning

DGX agent

arXiv:2608.05242v1 Announce Type: cross Abstract: In this work, we explore an alternative paradigm for spatial reasoning by explicitly disentangling 3D perception from reasoning, rather than jointly a

agentsarxiv-cs-cv
7 Aug 2026
Model Releases

Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents

DGX agent

arXiv:2608.06108v1 Announce Type: new Abstract: Investment competence is inherently personalized: the same market evidence can justify different actions for investors with different goals, horizons, p

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Explanations of Large Language Models Explain Language Representations in the Brain

DGX agent

arXiv:2502.14671v4 Announce Type: replace-cross Abstract: Large Language Model (LLM) representations are known to align with brain activity during language processing, but it remains unclear what driv

safetyarxiv-cs-ai
7 Aug 2026
Applications

MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation

DGX agent

arXiv:2603.25406v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models map visual observations and natural-language instructions to robot actions; however, hierarchical and autoregres

applicationsarxiv-cs-ro
7 Aug 2026
Applications

omega-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation

DGX agent

arXiv:2608.06375v1 Announce Type: new Abstract: Humanoid household tasks often require concurrent loco-manipulation, where the robot must move, adjust posture, maintain balance, and manipulate objects

applicationsarxiv-cs-ro
7 Aug 2026
Safety

Poli-Bias: Understanding and Measuring Large Language Model Biases in International Political Conflicts

DGX agent

arXiv:2608.06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, an

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Position: It's Time to Optimize LLMs for Self-Consistency

DGX agent

arXiv:2608.05188v1 Announce Type: cross Abstract: Despite ever-increasing sophistication in language model (LM) pre- and post-training pipelines, many important failures persist: models overcondition

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Relay, Don't Route: Adaptive Population Handoff for Cost-Efficient LLM-Driven Evolution

DGX agent

arXiv:2608.05651v1 Announce Type: cross Abstract: Large language model (LLM)-driven evolution has shown promise for program search and algorithm discovery, but relying on strong models throughout long

researcharxiv-cs-ai
7 Aug 2026
Applications

Revisiting Black-Box Model Ownership Verification through Information Theory

DGX agent

arXiv:2409.06130v2 Announce Type: replace-cross Abstract: Modern machine learning models require substantial computational resources and data to train, making them valuable intellectual property. Mode

applicationsarxiv-cs-ai
7 Aug 2026
Safety

Scalable estimation of VARMA models

DGX agent

arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, th

safetyarxiv-cs-lg
7 Aug 2026
Research

STAIL: Semantic Text-Anchored Incremental Learning for Medical Imaging via Large Language Models

DGX agent

arXiv:2608.05808v1 Announce Type: new Abstract: Deep learning models applied to medical image analysis suffer from severe catastrophic forgetting when continually adapting to new clinical tasks in dyn

researcharxiv-cs-cv
7 Aug 2026
Safety

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

DGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

safetyarxiv-cs-ai
7 Aug 2026
Research

A geometry-based deep equilibrium model for image restoration under multiplicative Gamma noise

DGX agent

arXiv:2608.04944v1 Announce Type: cross Abstract: We propose a deep learning framework for image restoration from images degraded by both multiplicative Gamma noise and blur. Unlike conventional deep

researcharxiv-cs-lg
6 Aug 2026
Research

A Model Merging Approach for Continual MLLM Unlearning

DGX agent

arXiv:2608.04548v1 Announce Type: cross Abstract: Multimodal large language model (MLLM) unlearning methods have been proposed to remove private, sensitive, or proprietary information from well-traine

researcharxiv-cs-ai
6 Aug 2026
Agents

AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering

DGX agent

arXiv:2411.04440v1 Announce Type: cross Abstract: Protein engineering is important for biomedical applications, but conventional approaches are often inefficient and resource-intensive. While deep lea

agentsarxiv-cs-ai
6 Aug 2026
← Previous
1…213214215216217…1271
Next →