AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
10 Aug 2026

Improving Attributed Long-form Question Answering with Intent Awareness

TutorialsDGX agent

arXiv:2603.27435v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used to generate comprehensive, knowledge-intensive reports. However, while these models a

M2-SMap: Memory-Efficient Semantic Mapping with Hierarchical Multi-Model Representation

Local AiDGX agent

arXiv:2608.07074v1 Announce Type: new Abstract: Dense point cloud maps, as a typically used mapping representation, are difficult to deploy on resource-constrained robots because their memory consumpt

MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model

Local AiDGX agent

arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-G…

Local AiDGX agent

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF 1/ big announcement today: we will be releas

PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model

SafetyDGX agent

arXiv:2608.06794v1 Announce Type: new Abstract: While diffusion models have made significant progress in text-to-image tasks, they still exhibit limitations when directly optimizing downstream objecti

Provable Training Data Identification for Large Language Models

ResearchDGX agent

arXiv:2510.09717v3 Announce Type: replace-cross Abstract: Identifying training data of large-scale models is critical for copyright litigation, privacy auditing, and ensuring fair evaluation. However,

Recipes for Creativity: Iterative Generation and Evaluation in Large Language Models

ResearchDGX agent

arXiv:2608.07243v1 Announce Type: new Abstract: Generative models are often evaluated through singular artifacts, whereas human creativity typically emerges through iterative generation, appraisal, an

Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA

Model ReleasesDGX agent

Meta released Muse Glimmer, a 30‑billion‑parameter dense language model with a context window exceeding 120 K tokens, designed for local, long‑running agentic AI workloads. The model is optimized to r

Summary of Takeaways from the Minimax AMA

Model ReleasesDGX agent

Summary from https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/ama_minimax_h3_team_ask_us_anything_about_our/ This summary was compiled with AI but cross-checked manually by me for accuracy. I

8 Aug 2026

Building a budget 32GB → 48GB VRAM home AI server: 2-3x RX 9060 XT 16GB vs RTX 5060 Ti 16GB, AM5 vs used EPYC?

Model ReleasesDGX agent

I’m planning a dedicated home AI server, mainly for local LLM inference, agents/tool use, Docker services, and eventually larger MoE models with CPU offload. My plan is to start with 2x 16GB GPUs = 32

Showoff Saturday: Local 4x 6000 Pro (multi-year progression)

Model ReleasesDGX agent

Not the biggest or shiniest, but it's mine From gaming machine inference on the original llama models, to a 4x RTX 6000 Pro Max Q + 4x 3090s local AI cluster. Pictures are in reverse chronological ord

7 Aug 2026

Adapting Vision Foundation Models with Cascaded Semantics

Model ReleasesDGX agent

arXiv:2608.05393v1 Announce Type: new Abstract: Prompt tuning, a leading parameter-efficient adaptation paradigm in NLP, has recently been extended to computer vision. Visual prompt tuning (VPT) adapt

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

SafetyDGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

Cautious Context Steering for Language Model Personalization

TutorialsDGX agent

arXiv:2608.05813v1 Announce Type: new Abstract: Personalizing language models (LMs) to individual user preferences is essential for aligning responses with diverse goals and backgrounds. Existing meth

Disentangling 3D Modeling from Spatial Reasoning

AgentsDGX agent

arXiv:2608.05242v1 Announce Type: cross Abstract: In this work, we explore an alternative paradigm for spatial reasoning by explicitly disentangling 3D perception from reasoning, rather than jointly a

Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents

Model ReleasesDGX agent

arXiv:2608.06108v1 Announce Type: new Abstract: Investment competence is inherently personalized: the same market evidence can justify different actions for investors with different goals, horizons, p

Explanations of Large Language Models Explain Language Representations in the Brain

SafetyDGX agent

arXiv:2502.14671v4 Announce Type: replace-cross Abstract: Large Language Model (LLM) representations are known to align with brain activity during language processing, but it remains unclear what driv

MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation

ApplicationsDGX agent

arXiv:2603.25406v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models map visual observations and natural-language instructions to robot actions; however, hierarchical and autoregres

omega-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation

ApplicationsDGX agent

arXiv:2608.06375v1 Announce Type: new Abstract: Humanoid household tasks often require concurrent loco-manipulation, where the robot must move, adjust posture, maintain balance, and manipulate objects

Poli-Bias: Understanding and Measuring Large Language Model Biases in International Political Conflicts

SafetyDGX agent

arXiv:2608.06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, an

Position: It's Time to Optimize LLMs for Self-Consistency

Model ReleasesDGX agent

arXiv:2608.05188v1 Announce Type: cross Abstract: Despite ever-increasing sophistication in language model (LM) pre- and post-training pipelines, many important failures persist: models overcondition

Relay, Don't Route: Adaptive Population Handoff for Cost-Efficient LLM-Driven Evolution

ResearchDGX agent

arXiv:2608.05651v1 Announce Type: cross Abstract: Large language model (LLM)-driven evolution has shown promise for program search and algorithm discovery, but relying on strong models throughout long

Revisiting Black-Box Model Ownership Verification through Information Theory

ApplicationsDGX agent

arXiv:2409.06130v2 Announce Type: replace-cross Abstract: Modern machine learning models require substantial computational resources and data to train, making them valuable intellectual property. Mode

Scalable estimation of VARMA models

SafetyDGX agent

arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, th

STAIL: Semantic Text-Anchored Incremental Learning for Medical Imaging via Large Language Models

ResearchDGX agent

arXiv:2608.05808v1 Announce Type: new Abstract: Deep learning models applied to medical image analysis suffer from severe catastrophic forgetting when continually adapting to new clinical tasks in dyn

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

SafetyDGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

6 Aug 2026

A geometry-based deep equilibrium model for image restoration under multiplicative Gamma noise

ResearchDGX agent

arXiv:2608.04944v1 Announce Type: cross Abstract: We propose a deep learning framework for image restoration from images degraded by both multiplicative Gamma noise and blur. Unlike conventional deep

A Model Merging Approach for Continual MLLM Unlearning

ResearchDGX agent

arXiv:2608.04548v1 Announce Type: cross Abstract: Multimodal large language model (MLLM) unlearning methods have been proposed to remove private, sensitive, or proprietary information from well-traine

AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering

AgentsDGX agent

arXiv:2411.04440v1 Announce Type: cross Abstract: Protein engineering is important for biomedical applications, but conventional approaches are often inefficient and resource-intensive. While deep lea

Beyond Linear Dynamics: Neural Bilinear Dynamical Models for Time Series Forecasting

ApplicationsDGX agent

arXiv:2608.04471v1 Announce Type: cross Abstract: Time series in real-world applications are often generated by nonlinear dynamical systems, making accurate forecasting challenging. Existing approache

Canonical Joint Energy-Based Model on CIFAR-10: failure modes and practical indistinguishability of Predictor-Corrector and SGLD samplers

ResearchDGX agent

arXiv:2608.05025v1 Announce Type: new Abstract: Joint Energy-Based Models (JEM) unify classification and generation within a single network and support out-of-distribution (OOD) detection. Canonical J

DataRx: Missingness-Aware Sampling for Safer Large Language Model Task-Specific Fine-Tuning

SafetyDGX agent

arXiv:2608.04322v1 Announce Type: new Abstract: Task-specific fine-tuning can improve the performance of large language models (LLMs) on downstream tasks. However, our study reveals that task-specific

DeepInvert: Semi-Supervised Embedding Inversion Against Obfuscated Language Models

ResearchDGX agent

arXiv:2608.04477v1 Announce Type: cross Abstract: Cloud-based language model services routinely process prompts containing sensitive information. Obfuscation-based defenses---including ObfusLM, Sentin

DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Models

ResearchDGX agent

arXiv:2608.04496v1 Announce Type: new Abstract: Visual inputs in vision-language models (VLMs) are often encoded into substantially longer token sequences than text, making visual tokens a major bottl

Exact Model-Free Policy Iteration for Co-safe LTL Planning

SafetyDGX agent

arXiv:2608.05047v1 Announce Type: cross Abstract: This work studies model-free reinforcement learning for co-safe linear temporal logic (sc-LTL) objectives in finite Markov decision processes, which c

Language Models Generalize to Human-like Word Order Preferences

SafetyDGX agent

arXiv:2608.05028v1 Announce Type: new Abstract: A central question in language acquisition is whether linguistic biases can emerge from general learning mechanisms operating over underdetermined input

Manipulation-Proof Oblivious Audits against Deceptive Model Providers

SafetyDGX agent

arXiv:2608.04365v1 Announce Type: new Abstract: Audits have emerged as a critical instrument for algorithmic governance, providing a mechanism for external scrutiny and governance of machine learning

Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications

SafetyDGX agent

arXiv:2509.08604v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medicine, with many studies adapting them through continued pre-traini

MOAT: Model-Agnostic Randomized Transformations for preventing Efficiency Degradation Attacks on ViTs

ResearchDGX agent

arXiv:2608.04680v1 Announce Type: cross Abstract: To adopt the Vision Transformers (ViTs) in resource-constrained environment, token pruning is widely used to reduce computational cost without impacti

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

AgentsDGX agent

arXiv:2608.05141v1 Announce Type: new Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon

Representing Visual Evidence for Item Difficulty Prediction: Visual Textualization and Image-Native Modeling

ResearchDGX agent

arXiv:2608.04554v1 Announce Type: new Abstract: Predicting item difficulty from content can provide an initial estimate for newly developed questions before sufficient student responses are available.

STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Models

SafetyDGX agent

arXiv:2608.04887v1 Announce Type: new Abstract: On-policy distillation (OPD) has become an effective approach for consolidating multiple task-specialized image generation models into a single student.

The Personalization Mirage: How LLMs Fabricate User Profiles, and Why Self-Monitoring Misleads

ResearchDGX agent

arXiv:2608.04570v1 Announce Type: new Abstract: Personalized LLMs with persistent memory are increasingly deployed, yet the faithfulness of their user models remains unexamined. We study over-inferenc

5 Aug 2026

AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities

ResearchDGX agent

arXiv:2608.03742v1 Announce Type: cross Abstract: Sound effects play a crucial role in conveying actions, events, and environmental cues across digital applications, often requiring a high degree of v

ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

ResearchDGX agent

arXiv:2608.02703v1 Announce Type: new Abstract: Weight-only quantization substantially reduces the storage of large language model (LLM) transformer blocks, but practical backends often retain the fin

Automatic Patient-Specific Microwave Ablation Planning Accelerated by a Physics-Guided Deep Learning Model

ResearchDGX agent

arXiv:2608.03086v1 Announce Type: cross Abstract: Microwave ablation (MWA) is a promising minimally invasive treatment for liver tumors, but its therapeutic outcome strongly depends on patient-specifi

Divide-and-Conquer: Towards Generalizable Amortized Bayesian Inference for the Drift Diffusion Model

TutorialsDGX agent

arXiv:2608.03566v1 Announce Type: cross Abstract: The drift diffusion model (DDM) is a cornerstone of cognitive decision-making research. Although numerous estimation methods exist, researchers contin

Enhancing VLM Reward Models Through Structure-Aware Fine-Tuning

SafetyDGX agent

arXiv:2608.03875v1 Announce Type: cross Abstract: Designing effective reward functions remains a major bottleneck in Reinforcement Learning (RL). Recent work uses large foundation Vision-Language Mode

Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing

ResearchDGX agent

arXiv:2608.02711v1 Announce Type: new Abstract: Recent advances in image generation have demonstrated the potential of unified multimodal models that integrate understanding, generation, and editing.

MissClick: Exploiting Digit-Serialized Coordinates to Attack GUI Grounding Models

ResearchDGX agent

arXiv:2608.03740v1 Announce Type: new Abstract: Recent GUI visual grounding models generate screen coordinates as sequences of digit tokens that are parsed into numerical values and mapped to executab

Modeling Matches as Language: A Generative Transformer Approach for Counterfactual Player Valuation in Football

ResearchDGX agent

arXiv:2603.15212v2 Announce Type: replace Abstract: Evaluating football player transfers is challenging because player actions depend strongly on tactical systems, teammates, and match context. Despit

Quantization Effects on Biomedical LLM Reliability

Model ReleasesDGX agent

arXiv:2608.03854v1 Announce Type: new Abstract: When decoder language models are used as classifiers, predicted class probabilities depend on implementation choices, including the prompt template, ver

Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models

SafetyDGX agent

arXiv:2608.02893v1 Announce Type: cross Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from laten

Scenema Audio Comes to ComfyUI, Runs on 8GB VRAM

Model ReleasesDGX agent

Hey everyone! Scenema Audio is now a native ComfyUI custom node. Same model that powers scenema.ai now quantized so it fits on 8GB VRAM. When we first released it a few months ago as an API and Docker

Sedentary Behavior Classification for Wearable Sensors with a CNN-BiLSTM Model

ResearchDGX agent

arXiv:2608.02946v1 Announce Type: new Abstract: Accurate detection of sedentary behavior is important for studying health risks related to prolonged sitting, but posture-based classification remains c

Simulation-free and finite-time diffusion model

ResearchDGX agent

arXiv:2608.03117v1 Announce Type: new Abstract: The performance of generative diffusion models is determined by the choice of the reference diffusion process connecting the empirical and prior distrib

Some people are surprised that APIs (aka what Anthropic, OpenAI, and others provide) are treated differently than open weights in the new AI…

SafetyDGX agent

Some people are surprised that APIs (aka what Anthropic, OpenAI, and others provide) are treated differently than open weights in the new AI model framework. I'm not surprised at all, and it's actuall

Suffix-Constrained Greedy Search Algorithms for Causal Language Models

ResearchDGX agent

arXiv:2603.01243v2 Announce Type: replace Abstract: Large language models (LLMs) are powerful tools that have found applications beyond human-machine interfaces and chatbots. Beside free-form generati

Watch a local Ollama's qwen3:8b turn one English question into a 9-node investigation graph - planned, admitted by a deterministic gate, and run live in the browser (open source, MIT)

Model ReleasesDGX agent

The video is one real run, not a mock-up: grapharc go 'why did checkout latency spike at 09:14 UTC?' --model ollama/qwen3:8b A local 8B model proposes the graph → triage fanning out into four parallel

4 Aug 2026

Active Regression for Single-Index Models with Unknown Link Functions

ResearchDGX agent

arXiv:2608.01287v1 Announce Type: cross Abstract: This paper studies active regression for single-index models under general ell_p-loss with an unknown 1-Lipschitz link function f, formulated as min_{

← Previous
1…169170171172173…1010
Next →