AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift

DGX agent

arXiv:2608.00928v1 Announce Type: new Abstract: Subtype robustness asks whether a model keeps the correct coarse prediction when test examples come from fine-grained subtypes absent from training but

researcharxiv-cs-lg
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

TAB-PO: Preference Optimization with a Token-Level Adaptive Barrier for Token-Critical Structured Generation

DGX agent

arXiv:2603.00025v3 Announce Type: replace Abstract: Direct Preference Optimization (DPO) is effective for offline alignment but poorly matched to ontology-driven structured prediction, where preferred

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

The Learning Objective Governs Perceptual Narrowing: A Cross-Lingual, Layer-Wise, Ten-Seed Study of Self-Supervised Speech Encoders

DGX agent

arXiv:2608.00507v1 Announce Type: new Abstract: Perceptual narrowing---the developmental loss of non-native phoneme discrimination in the first year of life itep{werker1984}---is a canonical developme

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

The Role of Disfluencies in Speech Translation

DGX agent

arXiv:2608.02138v1 Announce Type: new Abstract: Current speech translation systems, including SpeechLLMs, are trained on cleaned text and tend to strip disfluencies like filled pauses and false starts

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

Token-Native Storage: Read and Write in your Agent's Language

DGX agent

arXiv:2608.02376v1 Announce Type: cross Abstract: Search and database engines still store text as UTF-8, a format built for humans. But the systems that increasingly read and write that text (embedder

agentsarxiv-cs-cl
4 Aug 2026
Safety

WAM-Diff2: Hierarchical AR-to-Diffusion Distillation for Highly Efficient Autonomous Driving VLA

DGX agent

arXiv:2608.01035v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a prominent paradigm for end-to-end autonomous driving; however, their efficient deployment is sev

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline w…

DGX agent

We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline what happened, how the activity was contained, and how we’re

model-releasesopenai--x
4 Aug 2026
Model Releases

What Makes Position Zero Special? A Mechanistic Study of Position Zero Attention Sinks in LLMs

DGX agent

arXiv:2603.06591v2 Announce Type: replace-cross Abstract: Transformers frequently allocate disproportionate attention to specific tokens, a phenomenon known as attention sinks. Causal large language m

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

When Retrieval Helps and Distracts: Evaluating Evidence-Generating LLMs for Biomedical Claim Verification

DGX agent

arXiv:2608.01409v1 Announce Type: new Abstract: Biomedical fact-checking systems must do more than predict whether a claim is supported, contradicted, or unaddressed: they should also produce evidence

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs

DGX agent

arXiv:2608.00076v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly support high-stakes decision making by combining complementary information from images and text. W

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

Analytical and Bootstrap Confidence Intervals of Double Machine Learning: Simulation studies and an application to rural-urban difference in obesity prevalence

DGX agent

arXiv:2607.29456v1 Announce Type: cross Abstract: Double Machine Learning (DML) is a popular approach for treatment effect estimation in various settings, which allows a wide range of flexible machine

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Artificial Analysis: DeepSeek's V4-Flash costs 0.14/1M input and 0.28/1M output tokens, or 0.03 per test, far below Kimi K3's 0.86 and GPT-5.6 Sol's $1.86 (Eduardo Baptista/Reuters)

DGX agent

Eduardo Baptista / Reuters: Artificial Analysis: DeepSeek's V4-Flash costs 0.14/1M input and 0.28/1M output tokens, or 0.03 per test, far below Kimi K3's 0.86 and GPT-5.6 Sol's $1.86 — A version of Ch

model-releasestechmeme
3 Aug 2026
Model Releases

🚨ASI/AGI is imminent fans don’t realize that they ALREADY lost the argument around Astra. My argument (spelled out in detail in my Substack…

DGX agent

🚨ASI/AGI is imminent fans don’t realize that they ALREADY lost the argument around Astra. My argument (spelled out in detail in my Substack today) was that Astra was unlikely to be the dramatic leap f

model-releasesgary-marcus--x
3 Aug 2026
Model Releases

Assessing the Generalization of Graph Neural Networks for Fault Location Across Increasing Distributed Energy Resource Penetration Levels

DGX agent

arXiv:2607.29293v1 Announce Type: new Abstract: Accurate fault location is critical for distribution network reliability. However, increasing distributed energy resource (DER) penetration complicates

model-releasesarxiv-cs-lg
3 Aug 2026
Tutorials

DualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation

DGX agent

arXiv:2607.29337v1 Announce Type: cross Abstract: Background and Objective: Generating realistic medical images with anatomically accurate segmentation masks helps address the shortage of annotated da

tutorialsarxiv-cs-ai
3 Aug 2026
Model Releases

DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat

DGX agent

arXiv:2607.29577v1 Announce Type: new Abstract: Games and simulators make valuable benchmarks by turning decisions into measurable outcomes, but many current suites under-test rules-rich tactical reas

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

EarlyDx: An Admission-Anchored Benchmark for Open-Ended Generation of Evidence-Supported ED-Encounter Diagnoses

DGX agent

arXiv:2607.28788v1 Announce Type: new Abstract: Clinical diagnosis at hospital admission must be made rapidly from limited, incomplete evidence. Existing diagnosis-prediction benchmarks are poorly sui

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Fragility of Value under Imperfect Alignment

DGX agent

arXiv:2607.28881v1 Announce Type: new Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that these systems are aligned with humanity. A common

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Frugal Bayesian Optimization: Scalable Surrogates for Data- and Resource-Limited Discovery

DGX agent

arXiv:2607.29225v1 Announce Type: new Abstract: Bayesian Optimization (BO) is widely adopted for data-efficient optimization in scientific and engineering applications, yet its computational cost is r

model-releasesarxiv-cs-lg
3 Aug 2026
Research

HenTwin: A Multimodal Digital Twin Framework for Longitudinal Biological State Monitoring in Laying Hens

DGX agent

arXiv:2607.28652v1 Announce Type: cross Abstract: Early-life monitoring in laying hens remains constrained by fragmented single-modality sensing and the absence of formal system-level state representa

researcharxiv-cs-ai
3 Aug 2026
Model Releases

InferQ: A Database-Oriented Benchmark for Quantum Circuits Simulation

DGX agent

arXiv:2607.29134v1 Announce Type: cross Abstract: Recent work suggests that relational database management systems (RDBMSs) can execute quantum circuit simulation by compiling the simulation into SQL

model-releasesarxiv-cs-ai
3 Aug 2026
Research

Learning from Adversity: Semantic-Aware Mask Refinement through Adversarial Perturbation

DGX agent

arXiv:2607.29059v1 Announce Type: new Abstract: Despite significant advances in image segmentation, even state-of-the-art models produce masks with imperfect boundaries, semantic inconsistencies, and

researcharxiv-cs-cv
3 Aug 2026
Local Ai

My downloading is undownloading ??

DGX agent

So I just installed ollama and was trying to download qwen3-vl:8b model but while downloading the it downloads and then undownloads like it goes from close to 500mb to 320 mb like what is going on I t

local-air-ollama
3 Aug 2026
Model Releases

OSEF: One-Step Evidence Fusion for Cross-Video Scene Procedure Planning

DGX agent

arXiv:2607.29401v1 Announce Type: new Abstract: Video Scene Procedure Planning (VSPP) supplies the target start-goal observations in advance, leaving open how a planner should act when the evidence mu

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning

DGX agent

arXiv:2607.28695v1 Announce Type: cross Abstract: Here is the plain text version optimized for arXiv's submission form. Custom macros (like CV and SI) have been converted to standard text/math so they

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Motion

DGX agent

arXiv:2503.15225v3 Announce Type: replace-cross Abstract: The deployment of autonomous virtual avatars (in extended reality) and robots in human group activities---such as rehabilitation therapy, spor

agentsarxiv-cs-ai
3 Aug 2026
Research

ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning

DGX agent

arXiv:2606.13316v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a central technique for improving long-horizon reasoning in Large Language Models (LLMs). H

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Scaling Properties of Text Conditioning in Visual Generation

DGX agent

arXiv:2607.29679v1 Announce Type: new Abstract: We study empirical scaling properties for text conditioning in visual generation. Such properties have rarely been measured because diffusion loss does

model-releasesarxiv-cs-cv
3 Aug 2026
Agents

Scaling Scientific Discovery Environments for Turn-Level Agentic RL

DGX agent

arXiv:2607.28990v1 Announce Type: new Abstract: Large language model agents have shown promising capabilities in data-driven scientific discovery tasks, where an agent interacts with an execution envi

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

DGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

model-releasesarxiv-cs-ai
3 Aug 2026
Research

Sycophancy Undermines Epistemic Vigilance in Cooperative Vision-Language Tasks

DGX agent

arXiv:2607.29585v1 Announce Type: new Abstract: To maintain common ground in cooperative conversation, humans iteratively update their beliefs as conversation participants share new information; parti

researcharxiv-cs-cl
3 Aug 2026
Safety

TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text

DGX agent

arXiv:2607.28862v1 Announce Type: cross Abstract: The rapid development of Large Language Models (LLMs) has led to significant advances across a wide range of language tasks, while simultaneously rais

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots

DGX agent

arXiv:2509.06342v2 Announce Type: replace Abstract: Legged robots must achieve both robust locomotion and energy efficiency to be practical in real-world environments. Yet controllers trained in simul

model-releasesarxiv-cs-ro
3 Aug 2026
Model Releases

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

DGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

model-releasesswyx--x
3 Aug 2026
Model Releases

Validation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?

DGX agent

arXiv:2607.28871v1 Announce Type: cross Abstract: When a repair agent runs a test and sees it pass, the result is treated as evidence about the reported defect. We measure how often that treatment is

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

WaiT for the Signal: Simple Frequency-Aware Flow-Matching

DGX agent

arXiv:2607.28760v1 Announce Type: cross Abstract: As image generation models scale to ever higher resolutions, global coherence, local detail, and texture fidelity become critical axes for generation

local-aiarxiv-cs-ai
3 Aug 2026
Model Releases

Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.

DGX agent

I let Gemma4-31b run on my laptop for like almost a day using a heavily altered pi to do a deep dive on our beloved Llama tangentially related Subreddit, and this was the conclusion. Feels pretty accu

model-releasesr-localllama
2 Aug 2026
Model Releases

Cybersecurity isn’t a fortress problem, it’s an immunity problem. Think vaccines. Eliminating pathogen is not practically possible. Vaccines…

DGX agent

Cybersecurity isn’t a fortress problem, it’s an immunity problem. Think vaccines. Eliminating pathogen is not practically possible. Vaccines don’t eliminate pathogens. They teach the immune system to

model-releasesfireworks-ai--x
2 Aug 2026
Model Releases

Expert-only IQ3 requant of DeepSeek-V4-Flash-0731: better KLD than UD-IQ3_S, 1.4x decode on a CPU-spill rig

DGX agent

Hey all, tldr / who this helps: you run a mixed multi-GPU box where the experts spill to RAM, and you want to stay in the 3-bit tier instead of dropping to Q2 to make it fit. https://huggingface.co/Ta

model-releasesr-localllama
2 Aug 2026
Model Releases

July 2026 newsletter

DGX agent

The June edition of my sponsors-only monthly newsletter is out. If you are a sponsor (or if you start a sponsorship now) you can access it here. This month: Accidental cyberattacks by OpenAl and Anthr

model-releasessimon-willison
2 Aug 2026
Model Releases

MiniMax H3 is going open-weight in under 6 hours

DGX agent

here is all the info we have based on open PRs to add support to ComfyUI and HuggingFace diffusers - 33B for the main DiT and a pruned 20b variant - Qwen-3-VL-32b as the text encoder Edit- I posted cl

model-releasesr-stablediffusion
2 Aug 2026
Model Releases

Xberg v1 is out

DGX agent

Hi all, I'm happy to announce that Xberg v1 is out. Xberg is the successor to Kreuzberg, equivalent to what would have been Kreuzberg v5. It's a content intelligence framework that handles a very wide

model-releasesr-localllama
2 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 UD-IQ3_S 12.5 tok/s on RTX 3090 +128GB DDR5

DGX agent

I managed to run DeepSeek-V4-Flash-0731 UD-IQ3_S in text-generation-webui with: RTX 3090 24 GB 128 GB DDR5 overclocked to 5600 MHz using AMD EXPO llama.cpp loader First, I had to use a rather brutal w

model-releasesr-localllama
1 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731-UD-Q3_K_XL 3x3090 test results

DGX agent

For anyone interested, here are the llama-bench results on 3 bit K_XL quantization. I think this could be pushed further but no luck so far. CURRENT RESULTS: full moe offloading Prefill suffers 116 --

model-releasesr-localllama
1 Aug 2026
Tutorials

Kimi K3: The Complete Developer Guide

DGX agent

**Kimi K3 is Moonshot AI’s 2.8‑trillion‑parameter open‑weight language model—the largest ever released—designed for frontier tasks such as long‑horizon coding and deep reasoning.** Its architecture us

tutorialstogether-ai-blog
1 Aug 2026
Model Releases

Ten advances in mathematics and theoretical computer science

DGX agent

Ten advances in mathematics and theoretical computer science A few days ago it was Anthropic discovering cryptographic weaknesses with Claude using Mythos Preview, spending 100,000 on tokens and with

model-releasessimon-willison
1 Aug 2026
Model Releases

What's currently the 'smartest' LLM to use on 8GB vram and 16 RAM and same thing for 8 VRAM and 64 RAM?

DGX agent

Been trying to find something that actually handles my workload well instead of just being 'fine.' Started on Qwen 2.5 7B, moved to Qwen 3 8B, and right now I'm using Nemotron 3 Ultra (the big 550B on

model-releasesr-ollama
1 Aug 2026
Model Releases

A Graph-Native Bitemporal Memory Store for Conversational AI Agents

DGX agent

arXiv:2607.26520v1 Announce Type: cross Abstract: Conversational AI agents commonly lack persistent memory across sessions. The obvious fixes like injecting full chat histories into the context window

model-releasesarxiv-cs-ai
31 Jul 2026
← Previous
1…579580581582583…1371
Next →