AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,186 results
Model Releases

llama.cpp PR reports up to 169% faster quantized-KV decode at 118K context on Intel Battlemage from one SYCL kernel switch

DGX agent

A fresh llama.cpp PR (#26689) changes what looks like a tiny SYCL FlashAttention dispatch decision. With a quantized KV cache ('q4_0' / 'q8_0'), decode was being sent through the VEC kernel. On the au

model-releasesr-localllama
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LoDA: A Level of Detection Aware Method and a Multimodal Sensing Benchmark for Object Level Change Detection

DGX agent

arXiv:2608.05356v1 Announce Type: new Abstract: High-definition 3D LiDAR maps are important for autonomous driving and smart-city services, which require reliable detection of object-level changes in

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Matrix Zonotopic Attention: A Context-Adaptive Value Projection for Set Transformers

DGX agent

arXiv:2608.05472v1 Announce Type: cross Abstract: Multi-head attention combines an input-dependent softmax routing with an input-independent linear value projection, so the per-sample operator mapping

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)

DGX agent

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game

model-releasessimon-willison
7 Aug 2026
Model Releases

Multi-Representation Geometric Hierarchy Fusion: An Implicit-Submap Driven Framework for Resilient 3D Place Recognition

DGX agent

arXiv:2506.14243v4 Announce Type: replace Abstract: LiDAR-based place recognition is critical for long-term autonomous driving without GPS. Existing handcrafted feature methods face dual limitations.

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Neuro-Symbolic Closed-Loop Control of Laser Powder Bed Fusion with an In-Loop Ontology

DGX agent

arXiv:2608.05773v1 Announce Type: new Abstract: A geometry-conditioned, neuro-symbolic closed-loop architecture is proposed for laser powder bed fusion, in which a standards-aligned ontology operates

model-releasesarxiv-cs-lg
7 Aug 2026
Agents

Now we have a timeline of the OpenAI accidental attack against Hugging Face

DGX agent

OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about 'the Hugging Face Incident' (previously on this blog). The video was published yesterday. It's short and information

agentssimon-willison
7 Aug 2026
Model Releases

Outstanding cost-to-performance from DeepSeek GPT-5.6 Luna (Max) performance for a 1/4th of the cost on ARC-AGI

DGX agent

Outstanding cost-to-performance from DeepSeek GPT-5.6 Luna (Max) performance for a 1/4th of the cost on ARC-AGI DeepSeek V4 Flash from @deepseek_ai on ARC-AGI (Verified): - ARC-AGI-2: 61.4%, 0.04/task

model-releasesfrancois-chollet--x
7 Aug 2026
Model Releases

Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing

DGX agent

arXiv:2512.17426v2 Announce Type: replace-cross Abstract: We consider sparse signal reconstruction via minimization of the smoothly clipped absolute deviation (SCAD) penalty, and develop one-step repl

model-releasesarxiv-cs-lg
7 Aug 2026
Research

Perturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samples

DGX agent

arXiv:2608.05419v1 Announce Type: cross Abstract: Models trained by empirical risk minimization on data containing spurious correlations achieve high average accuracy while failing on subpopulations w

researcharxiv-cs-ai
7 Aug 2026
Safety

PolyAlign: Conditional Human-Distribution Alignment

DGX agent

arXiv:2606.13227v2 Announce Type: replace Abstract: Post-training methods such as supervised fine-tuning (SFT) and preference optimization typically align language models toward a single global assist

safetyarxiv-cs-cl
7 Aug 2026
Research

Predicting Social Media User Actions: A Hybrid Approach for Common and Rare Behavior Prediction on Bluesky

DGX agent

arXiv:2511.17241v2 Announce Type: replace Abstract: Understanding and predicting user behavior on social media platforms is crucial for content recommendation and platform design. While existing appro

researcharxiv-cs-cl
7 Aug 2026
Research

PromptForSegCXR: Prompt-Driven Multi-Organ and Multi-Disease Segmentation in Chest X-rays using a Multi-stage Fusion Mechanism

DGX agent

arXiv:2507.00673v2 Announce Type: replace-cross Abstract: Image segmentation is central to automated medical image analysis, enabling precise identification of anatomical structures and pathological r

researcharxiv-cs-cv
7 Aug 2026
Research

QEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decoding

DGX agent

arXiv:2608.05326v1 Announce Type: cross Abstract: Autoregressive large language model inference is increasingly constrained by the memory footprint of the Key-Value (KV) cache. A dominant line of work

researcharxiv-cs-cl
7 Aug 2026
Research

Quality Diversity for Reliable Data Driven Time-Use Optimization

DGX agent

arXiv:2608.05230v1 Announce Type: cross Abstract: The daily allocation of the finite 24-hour time budget is strongly associated with physical, mental, and cognitive health. While predictive models can

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Qwen 3.6 27B flags/settings in llama.cpp

DGX agent

I run the following on a 5090 and have been okay with its performance, it does most things somewhere 80-100 t/s, though that can slow down at full 262k context - more like 40 t/s at times. I use it pr

model-releasesr-localllama
7 Aug 2026
Model Releases

Rectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEG

DGX agent

arXiv:2608.05315v1 Announce Type: new Abstract: Electroencephalography (EEG) based Brain-Computer Interfaces (BCIs) often require unsupervised domain adaptation (UDA) to generalize across subjects and

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Retailers are updating their websites to rank highly in chatbot results, while making sure purchases are done on their own sites to collect customer data (Arriana McLymore/Reuters)

DGX agent

Arriana McLymore / Reuters: Retailers are updating their websites to rank highly in chatbot results, while making sure purchases are done on their own sites to collect customer data — As shoppers incr

model-releasestechmeme
7 Aug 2026
Tutorials

Reversible Unlearnable Examples: Towards the Copyright Protection in Deep Learning Era

DGX agent

arXiv:2608.06211v1 Announce Type: cross Abstract: Significant advancements in deep learning have been made possible by the utilization of large datasets, underscoring the critical importance of copyri

tutorialsarxiv-cs-cv
7 Aug 2026
Model Releases

SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation

DGX agent

arXiv:2608.05628v1 Announce Type: new Abstract: Although agent skills equip LLMs with reusable procedural knowledge, manual maintenance suffers from high costs, unscalability, and misalignment. Real-w

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

DGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

DGX agent

arXiv:2608.06270v1 Announce Type: new Abstract: The 'thinking-with-images' paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations o

safetyarxiv-cs-ai
7 Aug 2026
Applications

The Judgment-Consequence Gap: LLM Moral Reasoning in Healthcare Decisions

DGX agent

arXiv:2608.05583v1 Announce Type: cross Abstract: As large language models (LLMs) enter high-stakes domains such as healthcare, understanding their moral reasoning becomes essential. Decisions about s

applicationsarxiv-cs-ai
7 Aug 2026
Local Ai

TLNM: Externally Validated Tooth Detection, Numbering and Segmentation from Smartphone Photographs Using Mask R-CNN

DGX agent

arXiv:2608.06275v1 Announce Type: new Abstract: Oral health issues affect billions globally, but the cost and limited access to professional dental care hinder preventive oral healthcare. Research rel

local-aiarxiv-cs-cv
7 Aug 2026
Model Releases

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

DGX agent

arXiv:2608.06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging. Crit

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Tytan: Interactive Neurosymbolic Construction of Analytic Semantic Schemas from Relational Data

DGX agent

arXiv:2608.06331v1 Announce Type: cross Abstract: From natural-language query interfaces to automated report generation, data analysis tools need a description of the data: the real-world entities it

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Universal Pathologies, Conditional Consequences: A Triple-Robustness Analysis of RAG for Multi-Hop Traceability

DGX agent

arXiv:2608.05153v1 Announce Type: cross Abstract: GraphRAG underperforms vector RAG on citation precision in many reports, but where and why have remained corpus-bound. We present a triple-robustness

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text)…

DGX agent

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text) - chief assigns tasks to managers of various projects - man

model-releasesyohei-nakajima--x
7 Aug 2026
Model Releases

We analyzed DeepSeek-V4 Flash-0731 vs. GPT-5.6 Luna on software engineering tasks using DeepSWE. DeepSeek-V4 Flash-0731 delivers 80% of Luna…

DGX agent

We analyzed DeepSeek-V4 Flash-0731 vs. GPT-5.6 Luna on software engineering tasks using DeepSWE. DeepSeek-V4 Flash-0731 delivers 80% of Luna’s performance at roughly 1/6 the cost. More insights in the

model-releasestogether-ai--x
7 Aug 2026
Agents

When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters

DGX agent

arXiv:2608.05207v1 Announce Type: new Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discov

agentsarxiv-cs-lg
7 Aug 2026
Model Releases

When Self-Evolution Backfires: Pre-Commit Gating against Skill Contamination in LLM Agents

DGX agent

arXiv:2608.05810v1 Announce Type: new Abstract: Self-evolving agents accumulate capability by distilling reusable skills from their execution trajectories, but we find this process is not monotonic: p

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Where Privacy Risk Lives in English-Source Multilingual RAG: A Stage-Decomposed Audit Across Five Query Languages

DGX agent

arXiv:2608.05163v1 Announce Type: new Abstract: A common assumption holds that switching to a non-English language makes a multilingual RAG system easier to attack for personal information. We test th

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

DGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exac…

DGX agent

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exact same prompt https://x.com/simonw/status/208508951822360205

model-releasessimon-willison--x
7 Aug 2026
Model Releases

2 x 5070ti Qwen 27B full config / stats

DGX agent

Following up on yesterday's post about running everyone's faves on 2 x 16gb cards while maximizing performance and KV. Previous post data used abandoned Cu130 VLLM image. Stats here are done on cu129-

model-releasesr-localllama
6 Aug 2026
Model Releases

5.6 Sol much better in chat now and unlimited text chat for free users!

DGX agent

5.6 Sol much better in chat now and unlimited text chat for free users! We’re making better intelligence easier to access in ChatGPT for everyone: - GPT-5.6 Sol now powers both Instant and deep reason

model-releasessam-altman--x
6 Aug 2026
Model Releases

A Counterexample to Fourier Alignment in Single-Neuron Modular Addition

DGX agent

arXiv:2608.04451v1 Announce Type: cross Abstract: We give a negative solution to MAIS-O60. We first construct an example in which an initially active ReLU neuron becomes completely inactive in finite

model-releasesarxiv-cs-lg
6 Aug 2026
Agents

A quick snapshot of where Qwen3.8-Max stands today: Qwen3.8-Max now ranks #5 on the Artificial Analysis Intelligence Index, and #1 on the Ag…

DGX agent

Qwen 3.8‑Max, Alibaba’s latest large language model, was reported on August 6 2026 to rank **#5** on the Artificial Analysis Intelligence Index and **#1** on the Agentic Index. These rankings position

agentsqwen--x
6 Aug 2026
Model Releases

ACA-GS: Adaptive-Capacity Anchored Gaussian Splatting for Compact Dynamic Radiance Fields

DGX agent

arXiv:2608.04581v1 Announce Type: new Abstract: Recent advances in 4D Gaussian Splatting (4DGS) enable high-fidelity, real-time spatiotemporal rendering, but expose a fundamental trade-off between mot

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

DGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

safetyarxiv-cs-ai
6 Aug 2026
Research

Aligning Fetal Anatomy with Kinematic Tree Log-Euclidean PolyRigid Transforms

DGX agent

arXiv:2603.02371v2 Announce Type: replace Abstract: Automated analysis of articulated bodies is crucial in medical imaging. Existing surface-based models often ignore internal volumetric structures an

researcharxiv-cs-cv
6 Aug 2026
Model Releases

anyone confused about neurosymbolic AI—and its recent enormous victory—should read this. complete and total vindication for what I have been…

DGX agent

anyone confused about neurosymbolic AI—and its recent enormous victory—should read this. complete and total vindication for what I have been saying here all along. (see also my essays on Claude Code a

model-releasesgary-marcus--x
6 Aug 2026
Model Releases

ArtChart: Faithful Artistic Chart Generation with Integrated Text Rendering

DGX agent

arXiv:2607.16060v2 Announce Type: replace Abstract: Artistic charts combine data visualization with expressive marks, textures, and typography, but they are difficult for image generators: an output i

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

ATLAS: Adaptive Topological Learning with Abstract Successors for Continual Learning

DGX agent

arXiv:2608.04334v1 Announce Type: cross Abstract: Contemporary model-free reinforcement learning algorithms can achieve very high performance, but have low sample efficiency and are not robust to chan

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Automatic Statistical Test for Rationally Expressible Algorithms by Selective Inference, with Applications to Feature Selection

DGX agent

arXiv:2608.04667v1 Announce Type: cross Abstract: Selective inference (SI) provides statistically valid p-values for hypotheses selected by applying an algorithm to the data, correcting for the bias t

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

b10290

DGX agent

mtmd/ggml: add ggml_build_forward_order (#26649) ggml: add ggml_build_forward_order ggml_build_forward_expand marks the tensor and all its ancestors for compute, so using it as a pure ordering hint (k

model-releasesllama-cpp-releases
6 Aug 2026
Model Releases

b10291

DGX agent

vulkan: fix submission batching size, add debug tools for diagnosing causes of DeviceLost drivers errors (#26371) vulkan: add debug tooling to get more information about a DeviceLost error fix submiss

model-releasesllama-cpp-releases
6 Aug 2026
Model Releases

b10293

DGX agent

ci : onboard AMD ROCm CI with gfx1151 fixes (#26544) ci: prepare for amd rocm ci Signed-off-by: Aaron Teo aaron.teo1@ibm.com ci: fix editorconfig-checker Signed-off-by: Aaron Teo aaron.teo1@ibm.com ci

model-releasesllama-cpp-releases
6 Aug 2026
← Previous
1…783784785786787…1359
Next →