AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,284 results
Model Releases

AdaMTP: An Adaptive Training Paradigm for Multi-Token Prediction

DGX agent

arXiv:2608.00434v1 Announce Type: new Abstract: Multi-Token Prediction (MTP) has emerged as an effective paradigm that augments a shared Large Language Model backbone with auxiliary heads, training th

model-releasesarxiv-cs-cl
4 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Adaptive Quantum Physics-Informed Neural Networks for Differential Equations with Applications to Fluid Dynamics

DGX agent

arXiv:2608.00850v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a versatile approach for solving nonlinear partial differential equations (PDEs), yet achieving

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Adaptive Reconstruction of Bosonic Quantum States

DGX agent

arXiv:2608.02049v1 Announce Type: cross Abstract: Bosonic quantum systems provide a hardware-efficient platform for quantum information processing but remain challenging to characterise due to their l

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification

DGX agent

arXiv:2509.24560v2 Announce Type: replace Abstract: Extended Chain-of-Thought (CoT) reasoning has significantly bolstered the capabilities of medical large language models (LLMs). However, current mod

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

AdvPlan-Bench: Adversarial Evaluation of Structured Plan-Generation Agents

DGX agent

arXiv:2608.00832v1 Announce Type: new Abstract: Structured plan-generation agents are often evaluated as if a plan has quality in isolation, yet many realistic planning tasks require asking how a cand

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AeroLLE: Constrained Pseudo-Supervision for Nighttime Aerial Image Enhancement with the AeroNight-1.5K Benchmark

DGX agent

arXiv:2608.00702v1 Announce Type: new Abstract: Nighttime aerial image enhancement is challenged by spatially nonuniform exposure, mixed illumination, and weak structural evidence, while registered no

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

DGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AgentMemBench: A Systematic Benchmark for Evaluating Long-Term Memory Management Strategies in Conversational AI Agents

DGX agent

arXiv:2608.00009v1 Announce Type: new Abstract: Long-term memory remains a critical bottleneck for conversational AI agents, whose finite context windows cannot support coherent recall across thousand

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

DGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

model-releasesnous-research--x
4 Aug 2026
Model Releases

Amortized Inference of Multi-Modal Posteriors using Likelihood-Weighted Normalizing Flows

DGX agent

arXiv:2512.04954v3 Announce Type: replace Abstract: We present a novel technique for amortized posterior estimation using Normalizing Flows trained with likelihood-weighted importance sampling. This a

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

An Accessible Solution for Deformable Image Registration Compared with Learning-Based Approaches

DGX agent

arXiv:2608.02248v1 Announce Type: cross Abstract: Deformable image registration (DIR) is a core problem in medical image analysis; but, unlike labeling decision problems such as classification and seg

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

An Embedded RISC-V Evaluation of Kolmogorov--Arnold Networks in Hard-Constrained Recurrent Physics-Informed Models

DGX agent

arXiv:2608.00737v1 Announce Type: new Abstract: Hard-constrained recurrent physics-informed networks (HRPINNs) embed known dynamics inside a recurrent numerical integrator and restrict a neural branch

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Analyzing Speech Condition Effects in Dysarthric ASR: A Layer-wise Probing Study

DGX agent

arXiv:2608.01865v1 Announce Type: new Abstract: Automatic speech recognition (ASR) performance degrades sharply on dysarthric speech, yet how disordered articulation reshapes a model's internal repres

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

AOSpec: Action and Observation Co-Speculation for Low-Latency Agent Serving

DGX agent

arXiv:2608.00881v1 Announce Type: new Abstract: Large language model agents increasingly act through stateful tools, yet model generation and environment execution remain serialized at every step. As

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Appreciate it! High performance, low cost. Try it out today. 💻

DGX agent

Appreciate it! High performance, low cost. Try it out today. 💻 Qwen 3.8 Max is available on AI Gateway. It scores mid-pack on the DeepsecBench cybersecurity leaderboard at one of the lowest costs per

model-releasesqwen--x
4 Aug 2026
Model Releases

ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification

DGX agent

arXiv:2608.01291v1 Announce Type: new Abstract: We present ArabicDialectSafety, a human-curated Arabic safety dataset of 25,071 prompts covering six Arabic varieties: Modern Standard Arabic, Syrian, E

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Asleep at the Wheel: JEPA's Limitations in Evaluating Novel Driving Data

DGX agent

arXiv:2608.01336v1 Announce Type: new Abstract: Modern autonomous-driving fleets record far more video than human reviewers can inspect. This motivates the need for an automatic clip triage mechanism,

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Assistant Placement Aria: A Benchmark for Egocentric Placement Assistance

DGX agent

arXiv:2608.00652v1 Announce Type: new Abstract: Human assistance in robotics spans around several tasks such as navigation, object manipulation, and placement, where a key challenge is selecting targe

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

AttnLink: Turning Attention into Schema Links for Text-to-SQL

DGX agent

arXiv:2608.00693v1 Announce Type: new Abstract: Schema linking is a critical component of Text-to-SQL systems, but existing approaches often trade off contextual modeling capacity, score-based control

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Auditable Release Control for Pedagogical Leakage in LLM Tutors

DGX agent

arXiv:2608.00515v1 Announce Type: cross Abstract: Large language model tutors can be correct and helpful yet disclose an answer or decisive reasoning before that disclosure is authorized. We formalize

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Augmented Inverse Hybrid Weighting: Robust Inference under Deterministic and Random Distribution Shifts

DGX agent

arXiv:2608.00701v1 Announce Type: cross Abstract: Reweighting source samples to match a target covariate distribution is a standard response to distribution shift when generalizing evidence from one p

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling

DGX agent

arXiv:2608.02602v1 Announce Type: new Abstract: Language remains an outlier in generative modeling: while images, video, and audio are increasingly modeled in continuous latent spaces, text generation

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Automatic Annotation of Ancient Greek Vowel Length

DGX agent

arXiv:2608.01935v1 Announce Type: new Abstract: Prior work in Ancient Greek NLP relies on corpora that do not disambiguate the phonemic vowel length of alpha, iota, and ypsilon, together known as the

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Automatic LV Localization and Short-Axis Plane Estimation from Arbitrary CMR Slice

DGX agent

arXiv:2608.00145v1 Announce Type: cross Abstract: Accurate estimation of left ventricular (LV) orientation is essential for cardiac magnetic resonance (CMR) imaging and downstream analysis. Existing m

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

AWS launches Kiro Crew, an autonomous agentic orchestrator for 24/7 code development

DGX agent

Amazon Web Services Inc. today launched Kiro Crew, an autonomous workspace that keeps artificial intelligence coding agents running all day and all night. The company said it developed Crew to allow d

model-releasessiliconangle
4 Aug 2026
Model Releases

b10247

DGX agent

ggml: use dynamic allocation for split graph inputs (#22789) ggml: use dynamic allocation for split graph inputs Replace fixed-size GGML_SCHED_MAX_SPLIT_INPUTS arrays with dynamically allocated buffer

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10248

DGX agent

vocab : validate default special token ids (#26506) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10249

DGX agent

server: add get_info tool (#26522) server: add get_info tool fix --rpc in docs server: harden get_info probe result handling Report the OS as unknown when the probe process fails to spawn or times out

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10250

DGX agent

tests: add model resolution test on synthetic repo listings (#26172) tests: add model resolution test on synthetic repo listings Include download.cpp and arg.cpp inside a namespace with hf_cache monke

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10251

DGX agent

model : support MTP in GLM-4.7-Flash (#24868) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10252

DGX agent

vendor : update BoringSSL to 0.20260803.0 (#26523) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10253

DGX agent

vendor : update cpp-httplib to 0.52.0 (#26485) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramewor

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10254

DGX agent

chat : add new template for DeepSeek V4 Flash 0731 (#26398) common/chat: update DeepSeek V4 templates Align the DeepSeek V4 templates with the official encoders while keeping parser behavior out of th

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10255

DGX agent

Extended SYCL oneDNN SDPA to non-FP16 KV caches (Q4_0–Q8_0 and FP32) (#25874) sycl: extend oneDNN SDPA to Q4_0-Q8_0 and F32 KV caches Extends the oneDNN SDPA path (PR #25222) to handle non-F16 KV cach

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10256

DGX agent

sycl: parallelize the non-contiguous concat kernel (#25852) sycl: parallelize the non-contiguous concat kernel Launch geometry only: the non-contiguous concat kernel launched a single-lane work-group

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10258

DGX agent

llama : move n_vocab from llama_sampler_data to penalty_sampler (#26520) This matches how it is done for logit_bias and mirostat samplers, see #25262 (comment) Website: https://llama.app macOS/iOS: ma

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10259

DGX agent

model : allow reshape of tensors during load (#26531) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCF

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10261

DGX agent

vocab : validate plamo2 byte tokens (#26511) validate plamo2 byte tokens --typo Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10262

DGX agent

vulkan backend ops: implemented GATED_LINEAR_ATTN (#25601) vulkan : add GATED_LINEAR_ATTN op docs : update Vulkan ops vulkan : remove unused GLA spec constant Updated ops.md ops.md update Website: htt

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10265

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10267

DGX agent

speculative : refactor enabled configs common_speculative_init (#26510) This commit contains a suggestion to reduce some code duplication in common_speculative_init when adding the enabled speculative

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10268

DGX agent

ci: fix pre-built binaries no longer working on macOS 15 and below (#26375) ci: fix pre-built binaries no longer working on macOS 15 and below ci: add macOS deployment target to disabled KleidiAI buil

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10269

DGX agent

models : fix dflash wo_a reshape on load (#26577) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFrame

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10270

DGX agent

mtmd: support Qwen3-TTS (note: breaking change to llama-tts binary) (#26254) convert text model main model load ok convert encoder ok speaker encoder loading ok speaker enc graph adapt vocab for backb

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10271

DGX agent

ui: CWD for agent (#26518) server : extend file_glob_search for UI pickers ui : add per-conversation working directory with picker ui : add path navigation and search scope to cwd picker Treat path-li

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10273

DGX agent

sampler : remove 'full-context windows' from history-based samplers (#26524) Resolve -1 to 1024 instead of ctx-len for samplers Because of backend-sampling we initialize samplers before the complete l

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10274

DGX agent

mtmd: correcting duplicate empty audio chunks for short inputs (#26536) correcting duplicate empty audio chunks for short inputs tests.sh code restored Website: https://llama.app macOS/iOS: macOS Appl

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

b10275

DGX agent

server: decode Windows OEM output to UTF-8 in built-in tools (#26597) a child process writes in the OEM code page, which is not UTF-8 on a western Windows install, so accented output reaches the JSON

model-releasesllama-cpp-releases
4 Aug 2026
← Previous
1…4041424344…465
Next →