AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
Tutorials

Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning

DGX agent

arXiv:2607.19345v2 Announce Type: replace-cross Abstract: Large language models that generate step-by-step reasoning traces have achieved strong performance on complex tasks, and extending them to lon

tutorialsarxiv-cs-ai
3 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

KAT Coder 2.5 dev: Do yourself a favor and try it!

DGX agent

It is so good! I don't know why there aren't more people talking about it. Fewer tokens, faster and more accurate than Qwen 3.6 35b a3b. On my setup it's nearly as good as 27b, but 5x faster. And it c

model-releasesr-localllama
3 Aug 2026
Model Releases

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems

DGX agent

arXiv:2607.28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural und

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

ReLoop-UME: Recurrent Depth with Learnable Retrieval Registers for Universal Multimodal Embedding

DGX agent

arXiv:2607.28751v1 Announce Type: new Abstract: Universal multimodal embedding (UME) maps heterogeneous multimodal inputs into a shared embedding space. Existing UME models either form embeddings thro

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

The Grokked Illusion: True Equilibrium Mitigates Catastrophic Forgetting

DGX agent

arXiv:2607.29503v1 Announce Type: new Abstract: While neural networks are typically evaluated by their training and test performance, these metrics do not reveal how robust a learned representation is

model-releasesarxiv-cs-lg
3 Aug 2026
Research

Token-Level Diagnosis of Sycophancy in LLMs with Attribution-Guided Steering

DGX agent

arXiv:2607.28906v1 Announce Type: new Abstract: Sycophancy refers to the tendency for large language models (LLMs) to match user beliefs at the cost of factual correctness, thereby undermining model r

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Tokenizer-Agnostic Engram Module

DGX agent

arXiv:2607.29065v1 Announce Type: new Abstract: Deepseek's Engram, a conditional memory module, was introduced to trade-off storage versus reasoning in large language models. However, the module relie

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

DeepSeek-V4-Flash 284B on 5.3GB of memory

DGX agent

Following up on my Qwen 3.6 port, I wanted to keep adding models and ended up fixing a bunch of things along the way, so it's its own engine now: Mference. Same core idea from TurboFieldfare, MoE mode

model-releasesr-localllama
2 Aug 2026
Model Releases

How well do multiple GPUs scale for LLM inference? (Trying to understand the basics)

DGX agent

Hi everyone, I’m fairly new to the multi-GPU side of local LLMs and I’m trying to understand how inference actually scales across multiple GPUs. Suppose I have a model running on a single GPU and then

model-releasesr-localllama
2 Aug 2026
Model Releases

Real-world reality check on Qwen for autonomous coding agents

DGX agent

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

model-releasesr-localllama
2 Aug 2026
Tutorials

There's no 'one weird trick” for prompting Krea 2 art styles—just many guidelines [WF included]

DGX agent

TLDR: There is no one prompting trick that will result in Krea 2 Turbo giving you exactly the style you want and across the whole image. Instead, if you are trying to achieve styles without the use of

tutorialsr-stablediffusion
1 Aug 2026
Industry

AI-native software development requires a new engineering model

DGX agent

Artificial intelligence has quickly become a standard part of modern software development. Coding assistants, code completion tools and AI-powered integrated development environments are now widely av

industrysiliconangle
31 Jul 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Hardware

Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference

DGX agent

The article shows that dense‑attention performance in long‑context inference is governed by group size (query heads per KV head), head dimension, and sequence length, with prefill being compute‑bound

hardwarenvidia-developer
31 Jul 2026
Research

IGME: Efficient Chained Method Ensemble for Transferable Semantic Segmentation Attacks

DGX agent

arXiv:2607.27465v1 Announce Type: new Abstract: Semantic segmentation models are vulnerable to transferable adversarial perturbations, yet evaluating transfer attacks on dense prediction models can be

researcharxiv-cs-cv
31 Jul 2026
Research

Latent-Kernel Discrete Flow Maps for Few-Step Generation

DGX agent

arXiv:2607.27529v1 Announce Type: new Abstract: Discrete diffusion and flow-matching models denoise a sequence over many steps, but to keep each step cheap, they factorize the transition across positi

researcharxiv-cs-lg
31 Jul 2026
Model Releases

MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek

DGX agent

arXiv:2607.28274v1 Announce Type: new Abstract: Modern Greek is a richly inflected language, yet the language models built for it are evaluated mainly on factual knowledge, and no benchmark is dedicat

model-releasesarxiv-cs-cl
31 Jul 2026
Applications

Scalable Drift Monitoring in Medical Imaging AI

DGX agent

arXiv:2410.13174v3 Announce Type: replace-cross Abstract: The integration of artificial intelligence (AI) into medical imaging has advanced clinical diagnostics but poses challenges in managing model

applicationsarxiv-cs-cv
31 Jul 2026
Model Releases

Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups

DGX agent

arXiv:2607.27232v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping how we consume information and form our worldview. This raises concerns beyond bias in AI: do LLMs

model-releasesarxiv-cs-cl
31 Jul 2026
Hardware

Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory

DGX agent

arXiv:2607.28263v1 Announce Type: new Abstract: Transformer depth is not used uniformly: lower and middle layers build semantic representations, while upper layers increasingly specialize them for pre

hardwarearxiv-cs-cl
31 Jul 2026
Research

Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model

DGX agent

arXiv:2607.26742v1 Announce Type: cross Abstract: Zero-shot text-to-speech (TTS) clones a voice from a short audio prompt, but this reliance on reference audio is a barrier when only visual informatio

researcharxiv-cs-ai
31 Jul 2026
Model Releases

Benchmarked: MindControl for Llama.cpp

DGX agent

I recently shared the original MindControl PoC (and on github) - sampler-level guided reasoning budgets for llama.cpp, nudging the model with self-aware statements about its own thinking budget instea

model-releasesr-localllama
30 Jul 2026
Safety

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation

DGX agent

arXiv:2607.26789v1 Announce Type: new Abstract: Vision-language-action (VLA) policies commonly execute long-horizon mobile manipulation through open-loop action chunks, issuing multiple actions withou

safetyarxiv-cs-ro
30 Jul 2026
Research

ChineseBERT: Chinese Pretraining Enhanced by Glyph and Pinyin Information

DGX agent

arXiv:2106.16038v3 Announce Type: replace Abstract: Recent pretraining models in Chinese neglect two important aspects specific to the Chinese language: glyph and pinyin, which carry significant synta

researcharxiv-cs-cl
30 Jul 2026
Research

Cognitive Convergence: Deep Similarities Between Large Language Models and Human Cognition

DGX agent

arXiv:2607.26179v1 Announce Type: cross Abstract: LLMs are widely regarded as alien intelligences, systems whose cognitive operations are fundamentally unlike our own. Apparent similarities to human c

researcharxiv-cs-cl
30 Jul 2026
Model Releases

GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks

DGX agent

arXiv:2607.19083v2 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by im

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

How Kimi K3 Engineered Its Way to the Frontier [R]

DGX agent

Kimi K3 by Moonshot reached the frontier as an open-weight model. Artificial Analysis ranks it fourth of 580 models, behind only Claude Opus 5, Fable 5, and GPT-5.6 Sol. Moonshot released more than th

model-releasesr-machinelearning
30 Jul 2026
Model Releases

Nanbeige4.2-3B: I'm not impressed

DGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

model-releasesr-localllama
30 Jul 2026
Local Ai

Origins and mitigation of double descent in reduced order modeling

DGX agent

arXiv:2607.26414v1 Announce Type: cross Abstract: Latent low-dimensional structure in datasets of natural and engineered systems enables their sparse sensing, or full-state reconstruction from histori

local-aiarxiv-cs-lg
30 Jul 2026
Model Releases

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

DGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis

DGX agent

arXiv:2607.26541v1 Announce Type: cross Abstract: Audio-capable foundation models enable end-to-end spoken interaction, but they also introduce safety risks beyond transcript content. It remains uncle

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

DGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

model-releasesarxiv-cs-cl
30 Jul 2026
Research

Cinematic Compositing Using Character-Environment-Harmonized Video Generation Models

DGX agent

arXiv:2606.20233v2 Announce Type: replace Abstract: Cinematic compositing aims to integrate green-screen characters into novel environments while maintaining physical and photometric realism. Previous

researcharxiv-cs-cv
29 Jul 2026
Research

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding

DGX agent

arXiv:2601.02295v2 Announce Type: replace Abstract: Current work on robot failure detection and correction typically operates in a post hoc manner, analyzing errors and applying corrections only after

researcharxiv-cs-ro
29 Jul 2026
Model Releases

DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space

DGX agent

arXiv:2607.25675v1 Announce Type: new Abstract: Text-space optimization adapts large language models (LLMs) by editing external natural-language artifacts rather than model weights, so the optimized a

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases

DGX agent

arXiv:2607.25933v1 Announce Type: cross Abstract: Clinical diagnostic evaluation should not only assess whether models can provide correct diagnoses, but also reflect the realities of clinical practic

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales

DGX agent

arXiv:2607.25364v1 Announce Type: new Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales are neither authorization nor reliable introspection

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

Fast, accurate, and differentiable: a neural-network surrogate for NRSur7dq4 precessing binary black hole waveforms

DGX agent

arXiv:2607.24960v1 Announce Type: cross Abstract: We present a neural network surrogate model that emulates the NRSur7dq4 gravitational waveform model for precessing binary black hole mergers. The sur

model-releasesarxiv-cs-lg
29 Jul 2026
Research

Finding Optimal Cost-Bounded Plan Reductions: Refined Model

DGX agent

arXiv:2607.25484v1 Announce Type: new Abstract: In some real applications a plan may later become unfeasible due to newly imposed budget constraints, yet, at the same time, using only the original act

researcharxiv-cs-ai
29 Jul 2026
Model Releases

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

DGX agent

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as mat

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation

DGX agent

arXiv:2607.25956v1 Announce Type: new Abstract: Multi-warehouse inventory allocation is typically formulated as a mixed-integer programming (MIP) problem, yet no single formulation consistently matche

safetyarxiv-cs-ai
29 Jul 2026
Research

Long-Term PM2.5 Forecasting Using a DTW-Enhanced CNN-GRU Model

DGX agent

arXiv:2510.22863v2 Announce Type: replace-cross Abstract: Reliable long-term forecasting of PM2.5 concentrations is critical for public health early-warning systems, yet existing deep learning approac

researcharxiv-cs-ai
29 Jul 2026
Research

Measuring the State of Open Science in Transportation Using Large Language Models

DGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

researcharxiv-cs-ai
29 Jul 2026
Model Releases

RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing

DGX agent

arXiv:2607.25199v1 Announce Type: cross Abstract: Automated code generation is becoming an important tool in quantitative finance, where large language models can generate option pricing implementatio

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

DGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Simulation-based parameter estimation via a combination of embedded normalizing flows and implied empirical probabilities under moment restrictions

DGX agent

arXiv:2607.25026v1 Announce Type: cross Abstract: In this work, we present a simulation-based parameter estimation framework for a model defined by a computational simulation of a physical system. We

model-releasesarxiv-cs-lg
29 Jul 2026
Safety

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

DGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

safetyarxiv-cs-ai
29 Jul 2026
Research

Transformer Transformer: A Unified Model for Motion-Conditioned Robot Co-design

DGX agent

arXiv:2607.25798v1 Announce Type: new Abstract: An often overlooked factor of robot manipulation performance is the embodiment of the robot itself. Motivated by this problem, we study motion-condition

researcharxiv-cs-ro
29 Jul 2026
← Previous
1…315316317318319…1302
Next →