AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlog
88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
Model Releases

Same Question, Different Answer? Measuring and Mitigating Prompt Privilege for Equitable AI Access

DGX agent

arXiv:2608.08942v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into healthcare, education, public services, and everyday decision making. They should provide

model-releasesarxiv-cs-cl
11 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic

DGX agent

arXiv:2601.22510v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often achieve strong benchmark accuracy yet remain brittle under small distribution shifts. While recent mechanis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests

DGX agent

arXiv:2608.08640v1 Announce Type: new Abstract: Large language model agents increasingly rely on reusable skills to extend their capabilities beyond parametric knowl- edge. However, retrieving the app

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation

DGX agent

arXiv:2601.09974v2 Announce Type: replace Abstract: Personalizing Large Language Models typically relies on static retrieval or one-time adaptation, assuming user preferences remain invariant over tim

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification

DGX agent

arXiv:2608.08195v1 Announce Type: cross Abstract: Large language models (LLMs) are high-value assets that can be derived through redeployment, fine-tuning, quantization, or further alignment. Because

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

The Authority Expectancy Effect in Multi-User Conflict

DGX agent

arXiv:2608.08026v1 Announce Type: new Abstract: We investigate how social authority (SA) signals interact with severity-based prioritization in large language models, operationalizing each axis as a m

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation

DGX agent

arXiv:2511.02687v2 Announce Type: replace Abstract: The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independentl

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs

DGX agent

arXiv:2608.08506v1 Announce Type: new Abstract: Training-free low-rank compression frameworks have been gaining prominence for LLM compression given their effectiveness in reducing model parameter cou

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

v0.32.8

DGX agent

Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such

model-releasesollama-releases
11 Aug 2026
Research

When Confidence Fails: Overconfidence in LLMs under Uncertainty and Missing Clinical Information

DGX agent

arXiv:2608.09080v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance in medical question answering and clinical reasoning tasks. However, their reliability u

researcharxiv-cs-ai
11 Aug 2026
Model Releases

b10338

DGX agent

model-saver : fix expert shared/chunk FFN length key clobber (#26693) The saver called add_kv with LLM_KV_EXPERT_SHARED_FEED_FORWARD_LENGTH twice, the second time passing n_ff_chexp. gguf_set_val_u32

model-releasesllama-cpp-releases
10 Aug 2026
Model Releases

Best open-source harness like Claude Code?

DGX agent

Avid claude code user here looking to do equivalent things with local models. Just want to plug in something like Qwen and have the interface be 1:1 with claude code. Any suggestion? submitted by /u/N

model-releasesr-localllama
10 Aug 2026
Safety

CASA: Classification Augmented with Safety Attention for Robust Multimodal Alignment

DGX agent

arXiv:2604.00310v2 Announce Type: replace-cross Abstract: Multimodal large-language models (MLLMs) often experience degraded safety alignment when harmful queries exploit cross-modal interactions. Mod

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Corma launches with $60M in funding for defensive cybersecurity AI

DGX agent

Defensive cybersecurity startup Corma Labs Ltd. today announced it has raised 60 million in seed funding to build a foundation model purpose-built for security defense. Founded in 2025, Corma runs off

model-releasessiliconangle
10 Aug 2026
Model Releases

Cryptanalytic Extraction of Isolated Bias-Free GLU Feed-Forward Blocks by Antipodal Separation

DGX agent

arXiv:2608.06631v1 Announce Type: cross Abstract: Cryptanalytic extraction has been demonstrated for ReLU networks, for networks using componentwise activations such as GELU or SiLU, and for a Transfo

model-releasesarxiv-cs-ai
10 Aug 2026
Research

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality

DGX agent

arXiv:2604.06756v2 Announce Type: replace Abstract: Large language models (LLMs) has been widely adopted as a scalable surrogate for human evaluation, yet such judges remain imperfect and susceptible

researcharxiv-cs-cl
10 Aug 2026
Safety

Let's Unlearn Stereotypes Before Decision-Making: Assessing the Impact of Intrinsic Bias Mitigation on Downstream Fairness in LLMs

DGX agent

arXiv:2509.16462v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economi

safetyarxiv-cs-cl
10 Aug 2026
Research

Natural Language Processing Psychometrics

DGX agent

arXiv:2608.07316v1 Announce Type: cross Abstract: Natural Language Processing (NLP) models predicting mental health outcomes rarely specify what they measure: contextual knowledge, emotional content,

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Semantic Adapter Routing with Fine-Tuning Task Embeddings

DGX agent

arXiv:2606.19079v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) has led to model ecosystems in which a single backbone is paired with many task-specialized adapters. Given s

model-releasesarxiv-cs-ai
10 Aug 2026
Tutorials

Semi Edge Inference Idea [D]

DGX agent

Today the most important factor in AI is cost. My idea is to split ML models inference (closed ones, proprietary) across server and edge computing on clients, and I would like to hear what do you thin

tutorialsr-machinelearning
10 Aug 2026
Model Releases

SLED: Scalable Location Encoding via Distillation

DGX agent

arXiv:2608.06612v1 Announce Type: cross Abstract: The plethora of readily available geospatial data offers exciting opportunities to learn high quality representations of the planet, but the sheer siz

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Stable Curves, Unstable Items: Item-Level Scaling Heterogeneity in Video LLMs

DGX agent

arXiv:2608.07014v1 Announce Type: new Abstract: Aggregate scaling curves suggest that Video LLMs improve smoothly or saturate as visual budgets grow. We show that this view can conceal large, opposing

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Stockmark-Nemotron-3-Nano-Omni-JapanDocReader: Structured Document Parsing via Capability Injection and Forgetting Control

DGX agent

arXiv:2608.06758v1 Announce Type: new Abstract: We present Stockmark-Nemotron-3-Nano-Omni-JapanDocReader, a Japanese document understanding model built from Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Model Releases

The Sparsity Whisperer

DGX agent

arXiv:2608.06630v1 Announce Type: new Abstract: Pruning reduces the inference cost of large language models, but existing criteria primarily preserve large activations or reconstruct layer outputs. We

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys

DGX agent

arXiv:2608.06404v1 Announce Type: new Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and manage

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

We've used GPT-5.6-Cyber extensively in real-world vulnerability research, including work that uncovered previously unknown vulnerabilities …

DGX agent

OpenAI announced the release of GPT‑5.6‑Cyber as part of its Cybersecurity Initiative, “Daybreak.” The model is aimed at advanced, authorized security research and testing, helping trusted defenders d

model-releasesopenai--x
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

[2606.05682] Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillation

DGX agent

Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost constrained production environments. Quantization

model-releasesr-localllama
9 Aug 2026
Model Releases

The Gemma team will host a special event on August 20

DGX agent

Tweet by u/hackerllama Could be copium, but I would love to see Gemma 4.1 there with unified audio input for all model sizes perhaps even up to 120B, much improved tool calling (even with the latest t

model-releasesr-localllama
9 Aug 2026
Model Releases

any reasonably fast public benchmarks I should run quants of deepseek flash 0731 on?

DGX agent

I have various quants of this model and am curious how they perform. can anyone recommend which benchmark would be a good test case for quantization effects? Maybe that can be completed with about 1 m

model-releasesr-localllama
8 Aug 2026
Model Releases

Anyone else amped up over Qwen 3.8?

DGX agent

I’ve been using 3.6 27B Q4, and that quant is fast on an M5. The code has been average, but consistently “good enough.” And, after a year, I can see home LLMs being served at home much like streaming

model-releasesr-localllama
8 Aug 2026
Model Releases

Auto mode is now the default in Claude Code for Pro, Max, and Team plans

DGX agent

Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses

model-releasessimon-willison
8 Aug 2026
Model Releases

Qwen 35B-A3B MoE vs 27B dense in local coding tests: ~4× faster, much smaller quality gap than I expected

DGX agent

I compared Qwen 35B-A3B MoE against Qwen 27B dense on a series of local coding-maintenance tasks. On my R9700/llama.cpp setup, the MoE model generated about 3.9× faster (~116 vs ~30 tok/s), but the co

model-releasesr-localllama
8 Aug 2026
Model Releases

Tesla V100 Qwen3.6 27B Performance

DGX agent

Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi coding agent llama.cpp model preset: [*] spec-default =

model-releasesr-localllama
8 Aug 2026
Model Releases

Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation

DGX agent

arXiv:2503.03556v3 Announce Type: replace Abstract: Object affordance reasoning, the ability to infer object functionalities based on physical properties, is fundamental for task-oriented planning and

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

DGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Anyone running DeepSeek-V4-Flash-0731 on MI325X with vLLM? Mine is behaving completely broken

DGX agent

Is anyone here successfully running DeepSeek-V4-Flash-0731 locally with vLLM, especially on AMD MI325X? My setup: GPU: 1x AMD Instinct MI325X Model: deepseek-ai/DeepSeek-V4-Flash-0731 vLLM: 0.26.0 ROC

model-releasesr-localllama
7 Aug 2026
Model Releases

Beyond Sequence Order: Syntax-Informed Positional Embeddings for Transformers

DGX agent

arXiv:2608.06111v1 Announce Type: cross Abstract: Positional embeddings (PE) in Transformers encode token distance and order but are largely agnostic to extit{syntactic structure}. We introduce extbf{

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Continual Learning in Transition

DGX agent

arXiv:2608.06216v1 Announce Type: cross Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g.,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

GROM: Gradient-Free Rapid One-Shot Machine Unlearning

DGX agent

arXiv:2608.05783v1 Announce Type: cross Abstract: Machine unlearning has become a critical capability for safely removing specific, sensitive knowledge from large language models (LLMs). Current state

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Matching Matters: A Fair Quality-Efficiency Benchmark for Command-Line Agents

DGX agent

arXiv:2606.21140v2 Announce Type: replace-cross Abstract: Rapid advances in large language models have improved the task-solving capabilities of command-line-interface (CLI)-based agents, whose CLIs d

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

DGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Parameter-Efficient Semantic Augmentation for Enhancing Open-Vocabulary Object Detection

DGX agent

arXiv:2604.04444v2 Announce Type: replace Abstract: Open-vocabulary object detection (OVOD) enables models to detect any object category, including unseen ones. Benefiting from large-scale pre-trainin

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

DGX agent

arXiv:2608.06125v1 Announce Type: new Abstract: Latent reward models can supervise visual diffusion models without decoding intermediate states into pixel space. This makes alignment with human prefer

safetyarxiv-cs-cv
7 Aug 2026
Model Releases

SEAM: Global consistency beyond local accuracy in scientific machine learning

DGX agent

arXiv:2608.05702v1 Announce Type: new Abstract: Scientific machine learning commonly validates models at the level of a subdomain, a benchmark split, or an explanation for one prediction. Yet such loc

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

SemiAdapt-Instruct: Extensible Instruction Tuning via Latent Domain-Specialised Adapters

DGX agent

arXiv:2608.05161v1 Announce Type: new Abstract: Instruction-tuned LLMs are deployed into environments where domains evolve, yet extending a fine-tuned model's capabilities without full retraining rema

model-releasesarxiv-cs-cl
7 Aug 2026
Research

Timestep-Conditioned Transformers for Global Weather Forecasting

DGX agent

arXiv:2608.06241v1 Announce Type: new Abstract: Existing machine-learning weather forecasting models rely on predetermined and fixed autoregressive timesteps. The choice of model timestep involves a f

researcharxiv-cs-lg
7 Aug 2026
← Previous
1…380381382383384…1324
Next →