AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

PHOENIX: Fine-Tuned SLM-Powered Autonomous Satellite Lifetime Extension via Predictive Self-Healing and Multi-Agent AI Recovery

DGX agent

arXiv:2608.07126v1 Announce Type: cross Abstract: Most CubeSats, small and low-cost satellites roughly the size of a shoebox, do not survive as long as they were designed to: a study of 178 missions f

model-releasesarxiv-cs-ai
10 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Playing with physics is so cool in Minimax H3

DGX agent

Prompt: 'integrated_multimodal_description: [Shot 1] Live-action, ultra-realistic first-person footage at night on a rainy city street, filmed with authentic handheld smartphone qualities. The phone i

model-releasesr-stablediffusion
10 Aug 2026
Model Releases

Please Share Your Experience About Muse Glimmer

DGX agent

I have a classic test for local LLM's. I asked for 8 ball pool game with only one HTML file and Muse Glimmer spend 21k Token(I m using full context so 128k) and only created a 220 lines of HTML and sa

model-releasesr-localllama
10 Aug 2026
Model Releases

Policy-Masked Private Experts: Auditable and Reversible Capability Access Control in Sparse MoE Models

DGX agent

arXiv:2608.06690v1 Announce Type: cross Abstract: Most language-model access controls regulate behavior while leaving the same computation available to every request. We study a different systems ques

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers

DGX agent

arXiv:2608.07436v1 Announce Type: new Abstract: Under the standard split, Muon gets hidden matrices and AdamW embeddings/output head. Muon groks modular addition faster, but its solutions do not hold.

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

PURe: A Plug-and-Play Product-Unit Residual Module for Vision Networks

DGX agent

arXiv:2505.04397v3 Announce Type: replace-cross Abstract: Modern vision networks are dominated by additive local transformations, whereas explicit multiplicative local interactions remain underexplore

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Quantization Damage Is Multiplicative, Not Additive

DGX agent

arXiv:2608.06564v1 Announce Type: cross Abstract: Quantization is how large language models are actually deployed, and below four bits it is known to hurt. What nobody can say is which of the model's

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

RegionDet: A Benchmark for Region Detection Beyond Object Instances

DGX agent

arXiv:2608.06850v1 Announce Type: new Abstract: Object detection is a fundamental task in computer vision and has achieved remarkable progress on standard benchmarks by localizing discrete and well-bo

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

ResidencyRL: Reinforcement Learning in Simulated Clinical Environments

DGX agent

arXiv:2608.07418v1 Announce Type: new Abstract: In medical education, physicians convert academic knowledge into clinical expertise through residency: years of training across thousands of encounters,

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Retrieval-Constrained Policy Optimization for Attack Technique Extraction from Cyber Threat Intelligence

DGX agent

arXiv:2608.06778v1 Announce Type: cross Abstract: Mapping cyber threat intelligence (CTI) text to MITRE ATT&CK techniques is essential for structured threat analysis, yet manual annotation is costly a

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Retrofitting Linear Attention into Diffusion Language Models

DGX agent

arXiv:2608.06628v1 Announce Type: new Abstract: Diffusion language models (dLLMs) offer a promising alternative to autoregressive models by accelerating inference through parallel decoding. Recent dLL

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

RoRA: Role-Oriented Regional Allocation for Visual Token Pruning in MLLMs

DGX agent

arXiv:2608.07088v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) encode images as long visual token sequences, making prefilling and KV-cache storage expensive. Existing trai

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA

DGX agent

Meta released Muse Glimmer, a 30‑billion‑parameter dense language model with a context window exceeding 120 K tokens, designed for local, long‑running agentic AI workloads. The model is optimized to r

model-releasesnvidia-developer
10 Aug 2026
Model Releases

Running Qwen 3.5 35B A3B-Q8_0 gguf on a cheap radeon 7600 at 18 token/s

DGX agent

I also have 64 gb ddr4 ryzen 5600 Using llama.cpp Ubuntu distro Settings are as follows --n-gpu-layers 999 --n-cpu-moe 37 --no-mmap -ctk q8_0 -ctv q8_0 -fa 1 -c 9000 submitted by /u/Sweaty_Perception6

model-releasesr-localllama
10 Aug 2026
Model Releases

SABRE: Scalable and Automated Benchmarking of VLMs under Stress

DGX agent

arXiv:2608.07435v1 Announce Type: cross Abstract: Vision-language models (VLMs) are improving rapidly, but benchmark development lags behind, making weaknesses hard to identify. Building stress tests

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization

DGX agent

arXiv:2602.12187v2 Announce Type: replace-cross Abstract: Search-Augmented Generative Engines (SAGE) have emerged as a new paradigm for information access, bridging web-scale retrieval with generative

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Same physical state, different collective dynamics: state encodings select synchronization outcomes in language-model agents

DGX agent

arXiv:2608.06968v1 Announce Type: cross Abstract: Language-model agents act on state encodings of their environment, yet these are treated as interchangeable interfaces. Using pretrained language mode

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

DGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents,…

DGX agent

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents, edit videos, work with 3D/CAD, and more. From multimodal mod

model-releasesqwen--x
10 Aug 2026
Model Releases

Seeking SOTA: Time-Series Forecasting Must Adopt Taxonomy-Specific Evaluation to Dispel Illusory Gains

DGX agent

arXiv:2603.15506v2 Announce Type: replace-cross Abstract: We argue that the current practice of evaluating AI/ML time-series forecasting models, predominantly on benchmarks characterized by strong, pe

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Semantic Adapter Routing with Fine-Tuning Task Embeddings

DGX agent

arXiv:2606.19079v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) has led to model ecosystems in which a single backbone is paired with many task-specialized adapters. Given s

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Simple-OPD: Demystifying Warm-up for On-policy Distillation

DGX agent

arXiv:2608.06802v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level supervision from teacher models, but its effectiveness can depend str

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

SkySeaLand: A Wide-Format Satellite Transportation Benchmark with an Ultra-Lightweight Detection Baseline

DGX agent

arXiv:2608.07382v1 Announce Type: new Abstract: Satellite object detection is challenged by small targets and wide-format scenes that lose detail under standard square-input resizing. We introduce Sky

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

SLED: Scalable Location Encoding via Distillation

DGX agent

arXiv:2608.06612v1 Announce Type: cross Abstract: The plethora of readily available geospatial data offers exciting opportunities to learn high quality representations of the planet, but the sheer siz

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Social World Models

DGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

SparseVoxelDet: Fully Sparse Voxel Networks for Efficient Event-Based Drone Detection

DGX agent

arXiv:2603.21638v2 Announce Type: replace Abstract: Event cameras excel at detecting small, fast drones, but today's detectors give away their key advantage: they convert the sparse event stream into

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Spatiotemporal Agility: Time-Constrained Reinforcement Learning for Vision-Guided Dynamic Quadrupedal Interception

DGX agent

arXiv:2608.06907v1 Announce Type: new Abstract: Legged robots require robust agility to perceive and interact with complex and dynamic environments within a constrained time. However, most existing qu

model-releasesarxiv-cs-ro
10 Aug 2026
Model Releases

Stable Curves, Unstable Items: Item-Level Scaling Heterogeneity in Video LLMs

DGX agent

arXiv:2608.07014v1 Announce Type: new Abstract: Aggregate scaling curves suggest that Video LLMs improve smoothly or saturate as visual budgets grow. We show that this view can conceal large, opposing

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection

DGX agent

arXiv:2608.06477v1 Announce Type: cross Abstract: Computer-use agents (CUAs) face a growing threat from indirect prompt injection, where adversarial instructions are planted in the environment such as

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Stockmark-Nemotron-3-Nano-Omni-JapanDocReader: Structured Document Parsing via Capability Injection and Forgetting Control

DGX agent

arXiv:2608.06758v1 Announce Type: new Abstract: We present Stockmark-Nemotron-3-Nano-Omni-JapanDocReader, a Japanese document understanding model built from Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Stoicheia: Character-Level Masked Diffusion for Ancient Greek Textual Restoration, Parsing, and Metrical Scansion

DGX agent

arXiv:2608.07249v1 Announce Type: new Abstract: We introduce Stoicheia, a 405M-parameter character-level masked-diffusion encoder for Ancient Greek whose input factors into five aligned, independently

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Sub-Quadratic Bisimulation Metrics via Approximate Nearest Neighbors: Coverage-Augmented Guarantees and Computable Two-Sided Certificates

DGX agent

arXiv:2608.06762v1 Announce Type: new Abstract: Bisimulation metrics quantify behavioral similarity in Markov decision processes, but their Wasserstein fixed-point operator updates every state pair an

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

Summary of Takeaways from the Minimax AMA

DGX agent

Summary from https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/ama_minimax_h3_team_ask_us_anything_about_our/ This summary was compiled with AI but cross-checked manually by me for accuracy. I

model-releasesr-stablediffusion
10 Aug 2026
Model Releases

Suppress and Diversify: Refining Robust Pathways for Corruption Robustness

DGX agent

arXiv:2608.06712v1 Announce Type: new Abstract: Model robustness against natural image corruptions is essential for safety-critical applications. While existing methods primarily focus on implicit rep

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

DGX agent

arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and simulation by synthesizing realistic ins

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Symbolic Graphics Programming with Large Language Models

DGX agent

arXiv:2509.05208v2 Announce Type: replace Abstract: Large language models (LLMs) excel at program synthesis, yet their ability to produce symbolic graphics programs (SGPs) that render into precise vis

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models

DGX agent

arXiv:2608.07314v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are commonly adapted to downstream manipulation tasks via supervised fine-tuning (SFT) or online reinforcement lea

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Tensor Network Kernel Machines: A JAX Framework for Machine Learning and Nonlinear System Identification

DGX agent

arXiv:2608.07043v1 Announce Type: cross Abstract: Developing nonlinear models that are both expressive and computationally efficient remains a challenge in machine learning and nonlinear system identi

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

Test-Time Adaptation with Online Personalized Energy-Based Cache for Fine-Grained Video Expression Recognition

DGX agent

arXiv:2608.06467v1 Announce Type: new Abstract: Facial expression recognition (FER) in videos is challenging because models must identify subtle, temporally evolving affective states that vary across

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Tested Muse Glimmer locally on coding with OpenCode & agentic work

DGX agent

Ran the model with quants (Q4) by Unsloth with latest (build from master) llama.cpp server. It takes ~20GB ram running on M5 Pro with 48GB at about 17t/s. Didn't do any reasoning loops/overthinking. O

model-releasesr-localllama
10 Aug 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Model Releases

The Sparsity Whisperer

DGX agent

arXiv:2608.06630v1 Announce Type: new Abstract: Pruning reduces the inference cost of large language models, but existing criteria primarily preserve large activations or reconstruct layer outputs. We

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model desig…

DGX agent

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model designed for always-on local agent use, small enough to run on a

model-releasesollama--x
10 Aug 2026
Model Releases

There is an asymmetry in most agentic workflows that does not get talked about much: humans have many ways to talk to agents, and almost no …

DGX agent

There is an asymmetry in most agentic workflows that does not get talked about much: humans have many ways to talk to agents, and almost no standardized way for agents to talk back to humans. You can

model-releasesyohei-nakajima--x
10 Aug 2026
Model Releases

TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failure

DGX agent

arXiv:2608.06657v1 Announce Type: new Abstract: Modern cyber-physical and AI-assisted systems couple human operators, AI decision modules, and automated controllers in a single control loop, so trustw

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

TradeVerse: A Longitudinal Benchmark of Political Negotiation in International Trade

DGX agent

arXiv:2608.06549v1 Announce Type: cross Abstract: LLMs are increasingly being applied to tasks involving institutional and political texts, but existing benchmarks evaluate them on isolated documents

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Transformers are famously bad at arithmetic, so I set one's weights by hand (no training) and it multiplies with 100% accuracy [P]

DGX agent

Obviously nobody needs a transformer that's good at multiplication. I wanted to know whether a stock transformer could do exact arithmetic if I chose its weights directly. I implemented the grade-scho

model-releasesr-machinelearning
10 Aug 2026
Model Releases

Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking

DGX agent

arXiv:2608.07077v1 Announce Type: new Abstract: The Tower of Hanoi is a simple planning puzzle that in prior work has proven challenging for large reasoning models (LRMs). Current models solve the sta

model-releasesarxiv-cs-ai
10 Aug 2026
← Previous
1…2122232425…464
Next →