AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

Pin Once, Swap Light: Subspace-Aligned Centroid-Residual Training for Efficient Ultra-LoRA Serving

DGX agent

arXiv:2608.03579v1 Announce Type: cross Abstract: Modern multi-tenant Low-Rank Adapters (LoRAs) serving systems concurrently host tens to hundreds of LoRA adapters. Though powerful, this introduces a

model-releasesarxiv-cs-ai
5 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Pingala: Prosody-Aware Decoding for Sanskrit Poetry Generation

DGX agent

arXiv:2603.24413v2 Announce Type: replace Abstract: Poetry generation in Sanskrit typically requires the verse to be semantically coherent and adhere to strict prosodic rules. In Sanskrit prosody, eve

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

PLAN: Parallel Liquid-Inspired Approximation Network for Efficient Representation Learning in Flexible Job Shop Scheduling

DGX agent

arXiv:2608.03041v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) approaches for flexible job shop scheduling (FJSP) heavily rely on attention-centric architectures to achieve state-

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Predictive Enhancement Calibration for Latent Breast MRI Virtual Contrast Enhancement

DGX agent

arXiv:2608.03612v1 Announce Type: cross Abstract: Virtual contrast enhancement (VCE) synthesizes enhanced breast MR images from pre-contrast acquisitions. Modern latent generators offer strong image p

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Prime Agent - a new coding harness surpassing Codex/CC/PI

DGX agent

Prime Agent is an open-source coding and research agent for general and long-running work. A self-improving RLM harness for coding and long-running autonomous tasks. Designed to be both token-efficien

model-releasesr-localllama
5 Aug 2026
Model Releases

PRISMA: Improving the Accuracy-Latency Frontier of Diffusion-based PDE Solvers Using Physics-Informed Spectral Attention

DGX agent

arXiv:2512.01370v2 Announce Type: replace-cross Abstract: Diffusion-based solvers for partial differential equations (PDEs) are often bottle-necked by slow gradient-based test-time optimization routin

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Provably Learning Multi-Head Attention with Queries

DGX agent

arXiv:2608.03294v1 Announce Type: new Abstract: We study the problem of learning multi-head softmax attention from black-box input-output access. The learner may query arbitrary real-valued token sequ

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

PSA Update CUDA from 13.2 to 13.3 to solve DeepSeek V4 Flash 0731 Looping Problem!

DGX agent

So one of yall mentioned that cuda 13.1 or 13.2 is broken for unsloth so I looked in to it, and they were right. I had 13.2 installed, after I switched to 13.3 no more looping!!! Before the cuda updat

model-releasesr-localllama
5 Aug 2026
Model Releases

Quantifying Hallucinations in Language Language Models on Medical Textbooks

DGX agent

arXiv:2603.09986v3 Announce Type: replace-cross Abstract: Hallucinations, the tendency for large language models to provide responses with factually incorrect and unsupported claims, is a serious prob

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Quantization Effects on Biomedical LLM Reliability

DGX agent

arXiv:2608.03854v1 Announce Type: new Abstract: When decoder language models are used as classifiers, predicted class probabilities depend on implementation choices, including the prompt template, ver

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Qwen-3D: A Generalist 3D Vision-Language Model for Spatial Understanding

DGX agent

arXiv:2608.02980v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable success on images and short videos, yet scaling them to long videos remains challenging due to f

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Qwen Developers' responses from their recent Twitter/X AMA

DGX agent

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

model-releasesr-localllama
5 Aug 2026
Model Releases

Qwen-Image-3.0-Pro is live on Qwen Cloud now! Try it out👇 https://www.qwencloud.com/models/qwen-image-3.0-pro?utm_content=g_20000001188

DGX agent

Qwen-Image-3.0-Pro is live on Qwen Cloud now! Try it out👇 https://www.qwencloud.com/models/qwen-image-3.0-pro?utm_content=g_20000001188 🔔 Qwen-Image-3.0 is now live on Qwen Cloud! Ranked #1 among Chin

model-releasesqwen--x
5 Aug 2026
Model Releases

Qwen3-TTS voice cloning is now in mainline llama.cpp — the old demo finally became real support

DGX agent

People may remember the Qwen3-TTS llama.cpp demo from a few months ago. That PR said it probably wouldn’t be merged because llama.cpp was missing some of the graph and API pieces it needed. A new impl

model-releasesr-localllama
5 Aug 2026
Model Releases

Qwen3.8-Max hits #2 in Image-to-WebDev Arena! It sees, it builds~😎

DGX agent

Qwen3.8-Max hits #2 in Image-to-WebDev Arena! It sees, it builds~😎 Exciting news: Qwen3.8-Max by @Alibaba_Qwen is #2 in Image-to-WebDev Arena! With 1,631 pts, it’s trailing only Claude Opus 5 (Max) by

model-releasesqwen--x
5 Aug 2026
Model Releases

Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains

DGX agent

arXiv:2608.03219v1 Announce Type: new Abstract: Benchmark gains are often treated as evidence of greater LLM capability. Yet the same gain can reflect different changes in model behavior. A model may

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rectify Then Diffuse: Disentangling Concepts Before Denoising Trajectory Unfolds

DGX agent

arXiv:2608.03135v1 Announce Type: cross Abstract: Text-to-image diffusion models can generate individual concepts well, but they often omit or merge concepts incorrectly with multiple concepts. We tra

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

DGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Reversing Arrows in Large Language Models

DGX agent

arXiv:2608.03512v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance on text-to-knowledge graph generation and related tasks. Nevertheless, it is still unclear

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Risky Business: Measuring The Faithfulness-Safety Tension

DGX agent

arXiv:2608.03745v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output str

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Route-Align-Verify for Functional Correctness in Code Generation

DGX agent

arXiv:2608.03341v1 Announce Type: cross Abstract: Large language models (LLMs) have substantially improved code generation, yet achieving strong functional correctness remains difficult, especially fo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rubrics as Privileged Information for Open-Ended Generation

DGX agent

arXiv:2608.02948v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD), where a single model acts as both student and teacher with different contexts, has shown promise in verifiable dom

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

S^3: Improving Agent Safety through Multi-Stage Defense

DGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SAKI: Score-Aware Low-Rank Key Indexing for Long-Context KV Retrieval

DGX agent

arXiv:2608.03228v1 Announce Type: new Abstract: Existing low rank KV cache methods preserve either model weights or key variance, neither of which directly reflects the attention scores used during in

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Scaling agentic AI: How UiPath built its high-performance GPU platform on AI Hypercomputer

DGX agent

As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to

model-releasesgoogle-cloud-ai
5 Aug 2026
Model Releases

Scenema Audio Comes to ComfyUI, Runs on 8GB VRAM

DGX agent

Hey everyone! Scenema Audio is now a native ComfyUI custom node. Same model that powers scenema.ai now quantized so it fits on 8GB VRAM. When we first released it a few months ago as an API and Docker

model-releasesr-localllama
5 Aug 2026
Model Releases

SciRet: A Compute-Aware Empirical Study of Retrieval and Reranking for Scientific RAG

DGX agent

arXiv:2608.03860v1 Announce Type: cross Abstract: We introduce SciRet, a compute-aware empirical study of retrieval-augmented generation for scientific question answering over CORD-19. Rather than pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

DGX agent

arXiv:2608.03327v1 Announce Type: new Abstract: Hybrid computer-use agents can act through screenshots or call text tools. We find that having a tool available does not settle which way the effect goe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

DGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

DGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Sensitivity, Causality, and Repair Dissociate: A Layer-Wise Analysis of Perturbation Robustness and Its Scaling

DGX agent

arXiv:2608.03842v1 Announce Type: new Abstract: When a language model fails on surface-perturbed input (typos, OCR noise, homophones), 'which layer is responsible' has three natural operationalization

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SeqLLM: Augmenting LLMs with Behavioral-Sequence Modeling for High-Stakes Decisions at WeChat Pay

DGX agent

arXiv:2608.03063v1 Announce Type: new Abstract: Merchant risk control at large payment platforms screens tens of millions of merchants daily, where false positives harm legitimate merchants and false

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs

DGX agent

arXiv:2608.03573v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large langu

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces

DGX agent

arXiv:2608.03401v1 Announce Type: cross Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, b

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA

DGX agent

arXiv:2509.25459v2 Announce Type: replace Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. How

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity

DGX agent

arXiv:2608.02665v1 Announce Type: cross Abstract: A benchmark score is a measurement instrument, yet most benchmarks read each item at a single canonical surface form. We ask whether that reading is f

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The tea…

DGX agent

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The team reports 95.5% on ARC-AGI-3, above the human baseline, and

model-releasesyohei-nakajima--x
5 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Model Releases

SlimVLM: Sensitivity-aware Dynamic Structured Pruning with Adaptive Visual Token Selection for Efficient Vision-Language Models

DGX agent

arXiv:2608.03580v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have demonstrated remarkable performance in processing and understanding both text and images, their large parameter

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

SocietyBench: Forecasting Counterfactual Social-World Evolution

DGX agent

arXiv:2608.04009v1 Announce Type: new Abstract: Large language models (LLMs), and the agents built on top of them, are now benchmarked heavily on whether they can finish a task -- fix a bug, drive a b

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them

DGX agent

arXiv:2607.27703v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used in embodied agents to interpret visual inputs, reason about spatial relationships, and make task

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Sphere Retraction Normalizations

DGX agent

arXiv:2608.02668v1 Announce Type: cross Abstract: Residual connections are the de facto mechanism for training deep neural networks stably. Geodesic Normalization (GeoNorm) recasts them on a Riemannia

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatement

DGX agent

Hear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

StereoVGGT: A Training-Free Visual Geometry Transformer for Stereo Vision

DGX agent

arXiv:2603.29368v2 Announce Type: replace Abstract: Driven by the advancement of 3D devices, stereo vision tasks including stereo matching and stereo conversion have emerged as a critical research fro

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection

DGX agent

arXiv:2511.15339v3 Announce Type: replace-cross Abstract: Automotive telemetry data exhibits slow drifts and fast spikes, often within the same sequence, making reliable anomaly detection challenging.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Stuck on 'A': Diagnosing and Repairing Interface Injury in Attention-to-KDA Linearization of a 0.6B Language Model

DGX agent

arXiv:2608.02689v1 Announce Type: new Abstract: We convert 21 of 28 full-attention layers of Qwen3-0.6B-Base into KDA (Kimi Delta Attention) linear-attention layers on a single consumer-grade GPU budg

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Sulphur 3 is looking for funding

DGX agent

Hello, I'm the guy who made Sulphur 2. With the recent release of a certain video model, we are looking to mobilize and train Sulphur 3 on this new model. We are targeting $10,000 USD. This certain ne

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

SUV: Future Scene Understanding as Video Generation for End-to-End Driving

DGX agent

arXiv:2608.03084v1 Announce Type: new Abstract: End-to-end driving requires a coherent understanding of future scenes, yet existing methods model these scenes using task-specific heads and output form

model-releasesarxiv-cs-cv
5 Aug 2026
← Previous
1…3940414243…465
Next →