AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

DGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

model-releasesarxiv-cs-ai
7 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Align-RAG: Alignment Is All You Need for TSFM In-Context Learning

DGX agent

arXiv:2608.05571v1 Announce Type: new Abstract: Retrieval-augmented forecasting promises to adapt frozen Time Series Foundation Models (TSFMs) to new domains without fine-tuning, but recent methods ty

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Am I just hallucinating

DGX agent

Or is there any reason why I feel like model output quality seems to be better when I use higher micro-batch values (ub) in llama-cpp? I don't really have any hard numbers or anything (just running th

model-releasesr-localllama
7 Aug 2026
Model Releases

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

DGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

DGX agent

arXiv:2607.18056v2 Announce Type: replace Abstract: Frontier large language models (LLMs) are increasingly integrated into scientific workflows, yet their growing biological capabilities may outpace c

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

And here's an even better version, built by GPT-5.6 Sol Ultra running in Code Desktop https://x.com/simonw/status/2085808307865014295

DGX agent

And here's an even better version, built by GPT-5.6 Sol Ultra running in Code Desktop https://x.com/simonw/status/2085808307865014295 I had Codex Desktop and GPT-5.6 Sol Ultra take a go at building my

model-releasessimon-willison--x
7 Aug 2026
Model Releases

Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux (Marcus Mendes/9to5Mac)

DGX agent

Marcus Mendes / 9to5Mac: Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux — Users running

model-releasestechmeme
7 Aug 2026
Model Releases

Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related 'fallbacks' by ~85% in testing across product surfaces (Anthropic)

DGX agent

Anthropic: Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related “fallbacks” by ~85% in testing across product surfaces — We're making updates to Cla

model-releasestechmeme
7 Aug 2026
Model Releases

Anyone running DeepSeek-V4-Flash-0731 on MI325X with vLLM? Mine is behaving completely broken

DGX agent

Is anyone here successfully running DeepSeek-V4-Flash-0731 locally with vLLM, especially on AMD MI325X? My setup: GPU: 1x AMD Instinct MI325X Model: deepseek-ai/DeepSeek-V4-Flash-0731 vLLM: 0.26.0 ROC

model-releasesr-localllama
7 Aug 2026
Model Releases

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

DGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

DGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

b10299

DGX agent

metal : avoid threadgroup matrix array instantiation in kernel_lightning_indexer (#26646) In MSL, declaring an array of matrix types like threadgroup half4x4 causes a 'no matching constructor' compila

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10301

DGX agent

cuda: fix warnings for unused variable/function (#26688) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10303

DGX agent

sycl : fix error Error OP FLASH_ATTN_EXT on arc770 (#26441) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) i

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10305

DGX agent

sycl : Support DSv4 OPs: LIGHTNING_INDEXER,DSV4_HC_COMB,DSV4_HC_POST,DSV4_HC_PRE (#26568) support DSv4 OPs: LIGHTNING_INDEXER,DSV4_HC_COMB,DSV4_HC_POST,DSV4_HC_PREwq update ops.md fix format issue Web

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10306

DGX agent

sycl: *glu flat path (#26354) tests: add SWIGLU perf cases perf mode had no GLU coverage. Adds SWIGLU at 17408 columns, 512 and 2048 tokens, f16 and f32, with the operands both fused and split. sycl:

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10307

DGX agent

sycl: fix UE4M3 parsing (#25608) The NVFP4 quantization format stores a scaling factor for every group of 16 weights, packed into a single UE4M3 byte. The SYCL GPU code was converting these scale valu

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10308

DGX agent

Mitigate crashing issue on Windows MSYS2 UCRT64 environment (GCC 16.1.0) (#26555) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABL

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10310

DGX agent

ggml : add aarch64 HWCAP fallbacks and fix fp16 variant detection (#25554) ggml : add fallback definitions for missing aarch64 HWCAP bits ggml : require HWCAP_ASIMDHP for the aarch64 fp16 cpu variants

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10311

DGX agent

mtmd: stop feeding the text stream again during Qwen3-TTS generation (#26706) The reference implementation has two mutually exclusive prompt layouts. In non streaming mode the prefill carries the whol

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10312

DGX agent

server: (router) do not evict busy models (#26567) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10313

DGX agent

server: (router) add LRU scheduler (#26572) add lru_sched handle coalescing (req leaves waiting queue) add tests fix stream case address review comments Website: https://llama.app macOS/iOS: macOS App

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10318

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10319

DGX agent

mtmd: fix longest_edge ignoring min/max pixels (#26638) mtmd: fix longest_edge ignoring min/max pixels nits Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10321

DGX agent

metal : fix NORM/RMS_NORM for row lengths that leave a partial simdgroup (#26708) ggml_metal_op_norm sized the threadgroup with nth = std::min(nth, args.ne00_t), which can leave nth not a multiple of

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10322

DGX agent

sycl: coalesce the ssm_conv window loads (#26612) test-backend-ops perf -o SSM_CONV on an Arc Pro B70, interleaved A/B against master, 6 reps, us/run: ne_a=[515,3328,1,1] ne_b=[4,3328,1,1] n_t=512 97.

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10326

DGX agent

tts: account for the vocoder pass in the timings line (#26733) get_output runs the waveform work the pipeline defers to it, from a single trailing window to a full pass depending on the model. Measuri

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

Basically every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be signficantly higher with a bett…

DGX agent

On August 7, 2026 Ethan Mollick tweeted that “every remaining good AI benchmark score has an implied asterisk next to it which reads: * could be significantly higher with a better harness.” The commen

model-releasesethan-mollick--x
7 Aug 2026
Model Releases

Benchmarking and Enhancing LLMs for Rule-Intensive Review of National Standard Documents

DGX agent

arXiv:2608.06312v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support complex professional tasks, yet their capabilities in rule-intensive document review remain insufficie

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

DGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning

DGX agent

arXiv:2608.05250v1 Announce Type: new Abstract: Multi-task supervised fine-tuning (SFT) often casts a heterogeneous data mixture as a single optimization problem, even though different tasks may reach

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning

DGX agent

arXiv:2608.05253v1 Announce Type: new Abstract: Quantized orthogonal fine-tuning (qoft) enables parameter-efficient adaptation of low-bit language models by learning structured activation rotations be

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Beyond Sequence Order: Syntax-Informed Positional Embeddings for Transformers

DGX agent

arXiv:2608.06111v1 Announce Type: cross Abstract: Positional embeddings (PE) in Transformers encode token distance and order but are largely agnostic to extit{syntactic structure}. We introduce extbf{

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT Foundation Models

DGX agent

arXiv:2608.05960v1 Announce Type: cross Abstract: Routine CT interpretation is inherently comprehensive, capturing incidental findings across the entire scan volume. 3D CT foundation models could assi

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

C^3PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models

DGX agent

arXiv:2608.05381v1 Announce Type: new Abstract: Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modali

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Can Open-Weight LLMs Produce Kernel-Verified Coq Proofs? A Pilot Study

DGX agent

arXiv:2608.05420v1 Announce Type: cross Abstract: Large language models (LLMs) can generate text that resembles a mathematical proof, but resemblance does not establish correctness. A formal proof che

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Cancelling my subscription also it was great

DGX agent

today was the last day of my subscription on ollama cloud, to be honest it was a great price value for me and with GLM 5.2 and Deepseek V4 Pro i was able to Vibe code my custom woocomerce shop with mu

model-releasesr-ollama
7 Aug 2026
Model Releases

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

DGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Causal Episodic Memory for Feedback-Driven Agent Repair

DGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

DGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ChronoVision: Temporal Reasoning via Latent State Reconstruction

DGX agent

arXiv:2608.05631v1 Announce Type: new Abstract: Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. T

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries

DGX agent

arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific m

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

DGX agent

arXiv:2608.05886v1 Announce Type: cross Abstract: Modern LLM coding agents such as Claude Code and OpenHands share a common inefficiency: they spend much of their token budget finding the file to patc

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in thi…

DGX agent

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in this field goes into tuning the weights. We wanted to know how

model-releasesclem-delangue--x
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Continual Learning in Transition

DGX agent

arXiv:2608.06216v1 Announce Type: cross Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g.,

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…2425262728…464
Next →