AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlog
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,672 results
Model Releases

When Less Is Enough: Context Selection and Prompting Strategies for Bengali News Headline Generation

DGX agent

arXiv:2608.15879v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance in text generation tasks, yet their effectiveness on headline generation remains sensitive to

model-releasesarxiv-cs-cl
18 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

Adaptive Stopping for Multi-Turn LLM Reasoning

DGX agent

arXiv:2604.01413v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly rely on multi-turn reasoning and interaction, such as adaptive retrieval-augmented generation (RAG)

applicationsarxiv-cs-ai
17 Aug 2026
Model Releases

Attention Capture Is Not Detection: A Two-Stage Account of How Humans Miss Localized AI Image Edits

DGX agent

arXiv:2608.13865v1 Announce Type: new Abstract: As AI-generated image edits proliferate, the platforms meant to curb the resulting disinformation treat detectability as a single, undifferentiated prop

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Buy the Rumor, Sell the News: When Is News Priced In?

DGX agent

arXiv:2608.14014v1 Announce Type: new Abstract: Two old market sayings hold that news is already priced in by the time it is published, and that the rumor is bought while the news is sold. Both place

model-releasesarxiv-cs-ai
17 Aug 2026
Safety

CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

DGX agent

arXiv:2608.13925v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) accelerate language generation by predicting multiple masks in a single forward pass. However, existing dLLMs

safetyarxiv-cs-ai
17 Aug 2026
Model Releases

CVT-Bench: Probing Spatial-State Integrity through Counterfactual Viewpoint Transformations

DGX agent

arXiv:2603.21114v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) perform strongly on isolated spatial tasks, but whether their predictions remain persistent and mutually co

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis

DGX agent

arXiv:2608.13608v1 Announce Type: new Abstract: Agentic 'Continual Learning Harnesses', systems that pair an LLM with retrieval or memory to improve from feedback without retraining, have shown growin

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

LightTeaNet: A Weakly Supervised Lightweight CNN for Multi-Label Tea Leaf Disease Detection and Localization

DGX agent

arXiv:2608.14178v1 Announce Type: new Abstract: Tea is known as an important crop in many parts of South and Southeast Asia, yet the production of tea is still hampered by the multiple diseases that d

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Ling 3.0 support merged into llama.cpp

DGX agent

Support for the new ling 3.0 models has been merged into llama.cpp: https://github.com/ggml-org/llama.cpp/pull/26608#event-29549472828 Ling tiny 8b1b - https://huggingface.co/inclusionAI/Ling-3.0-tiny

model-releasesr-localllama
17 Aug 2026
Model Releases

MACS: A Hybrid Multi-Agent Framework for Reliable Conversational E-Commerce Recommendation

DGX agent

arXiv:2608.14068v1 Announce Type: cross Abstract: Conversational recommendation for e-commerce is increasingly mediated by large language models (LLMs), yet many real-world deployments operate under a

model-releasesarxiv-cs-ai
17 Aug 2026
Local Ai

Not All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM Systems

DGX agent

arXiv:2608.13571v1 Announce Type: cross Abstract: When a language model fails to answer a query on the first attempt, an agentic system retries, consuming additional tokens each time. This retry overh

local-aiarxiv-cs-ai
17 Aug 2026
Research

On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

DGX agent

arXiv:2608.13793v1 Announce Type: cross Abstract: Machine learning (ML) has become an indispensable part of modern engineering design workflows. A crucial step in training an ML model is the selection

researcharxiv-cs-lg
17 Aug 2026
Safety

Responsiveness Verification: Will Predictions Change? How Much? How Often?

DGX agent

arXiv:2507.02169v2 Announce Type: replace Abstract: Machine learning models are often used in applications where their inputs change due to routine interactions, strategic manipulation, or noise. In s

safetyarxiv-cs-lg
17 Aug 2026
Model Releases

Unpopular opinion : Qwen 3.8 27b is not an overthinker

DGX agent

Yes it uses a ton more reasoning tokens than 3.6 did But test in on the same tasks with the other chinese models, glm 5.3, deepseek v4 flash and pro, etc it's really similar, and they are needed The r

model-releasesr-localllama
17 Aug 2026
Model Releases

Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

DGX agent

arXiv:2608.14375v1 Announce Type: new Abstract: Multi-agent reasoning systems often use agreement, confidence, or automated scores to decide which messages should shape a final answer. Such filtering

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Dear Dario, 1. If Claude can cure cancer to save people like your dad, why should we 'pace the progress'? Does that mean more people with He…

DGX agent

Dear Dario, 1. If Claude can cure cancer to save people like your dad, why should we 'pace the progress'? Does that mean more people with Hepatitis C will die? 2. If Fable is so cyber-capable that it

model-releasesyann-lecun--x
16 Aug 2026
Model Releases

Qwen 3.8 2.4T at 288k tokens/s on Nvidia GB300 NVL72

DGX agent

https://developer.nvidia.com/blog/serve-qwen3-8-2-4t-a95b-a-2-4t-parameter-model-with-configurable-reasoning-on-nvidia-gb300-nvl72/ 4k tokens per second per GPU of which there are 72. 350 tokens per s

model-releasesr-localllama
16 Aug 2026
Model Releases

Qwen3.8 27B reasoning effort low/medium/xhigh comparison

DGX agent

I did a short test of the different reasoning efforts, since on default xhigh the model thinks a lot. Not very scientific, just a quick 'generate an SVG of a pelican on a bicycle' prompt with 3 differ

model-releasesr-localllama
16 Aug 2026
Model Releases

There are two separate questions regarding Dario Amodei’s post about AI and biology: motivation and reality. Motivation. Dario started in bi…

DGX agent

There are two separate questions regarding Dario Amodei’s post about AI and biology: motivation and reality. Motivation. Dario started in biology (was a PhD student of the great Bill Bialek). It might

model-releasesyann-lecun--x
16 Aug 2026
Model Releases

Qwen3.8-27B is now part of your everyday life — from smartphones to vehicles. ⚡️🚀Appreciate your work! @MediaTek

DGX agent

Qwen3.8-27B is now part of your everyday life — from smartphones to vehicles. ⚡️🚀Appreciate your work! @MediaTek Congratulations to the @Alibaba_Qwen on the launch of Qwen 3.8! MediaTek continues our

model-releasesqwen--x
15 Aug 2026
Model Releases

The new DeepSeek-V4-Pro (0813) is now fully rolled out on Ollama's cloud and included in Pro and Max subscriptions. ollama run deepseek-v4-p…

DGX agent

The new DeepSeek-V4-Pro (0813) is now fully rolled out on Ollama's cloud and included in Pro and Max subscriptions. ollama run deepseek-v4-pro:cloud Hosted in the US with Zero Data Retention (ZDR) and

model-releasesollama--x
15 Aug 2026
Model Releases

A Generative Approach for Improving Multi-Label Defect Classification in Photovoltaic Modules

DGX agent

arXiv:2608.12725v1 Announce Type: new Abstract: This paper addresses the challenge of multi-label defect classification in electroluminescence (EL) images of photovoltaic (PV) cells. Training models o

model-releasesarxiv-cs-cv
14 Aug 2026
Research

Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity

DGX agent

arXiv:2608.13430v1 Announce Type: cross Abstract: Instruction-tuned language models achieve strong performance across a range of generation tasks, but have also recently been shown to exhibit verbaliz

researcharxiv-cs-ai
14 Aug 2026
Model Releases

AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design

DGX agent

arXiv:2608.13560v1 Announce Type: cross Abstract: Transforming multimodal sources into condensed and structured media outputs can be fundamentally conceptualized as a long-horizon agentic process cent

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

CoMedBench: A Multi-Source Benchmark of Synthetic Medical Data Fidelity and Downstream Utility

DGX agent

arXiv:2608.12805v1 Announce Type: new Abstract: Access to clinical data is essential for developing reliable healthcare machine learning systems, but direct use of electronic health records is constra

model-releasesarxiv-cs-lg
14 Aug 2026
Research

Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues

DGX agent

arXiv:2608.12599v1 Announce Type: new Abstract: Multi-turn dialogues let users revoke constraints as easily as impose them, but revocation does not reliably take effect: models keep enacting withdrawn

researcharxiv-cs-ai
14 Aug 2026
Research

EEG-PRIME: Prototype-Aligned Representation Learning with Multi-Level Conditioning for EEG Decoding

DGX agent

arXiv:2608.13072v1 Announce Type: new Abstract: Electroencephalography (EEG) decoding models often generalize poorly across datasets and subjects due to domain shifts in acquisition protocols and indi

researcharxiv-cs-ai
14 Aug 2026
Local Ai

From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion

DGX agent

arXiv:2608.13043v1 Announce Type: new Abstract: Diffusion models have achieved dominant performance in visual generation but suffer from substantial inference overhead. While cache-based acceleration

local-aiarxiv-cs-ai
14 Aug 2026
Research

Geometric and Behavioral Stratification in Transformer Residual Streams

DGX agent

arXiv:2608.12447v1 Announce Type: cross Abstract: Trained transformer models develop privileged bases: coordinate axes whose statistics differ from the rest of the residual stream. But what kind of di

researcharxiv-cs-cl
14 Aug 2026
Applications

Incremental Evaluation and Training in Relational Deep Learning

DGX agent

arXiv:2608.13023v1 Announce Type: new Abstract: Relational Deep Learning (RDL) models multi-tabular databases as temporal heterogeneous graphs to enable end-to-end representation learning. However, pr

applicationsarxiv-cs-lg
14 Aug 2026
Local Ai

It's been a pleasure to partner with the @Alibaba_Qwen team.

DGX agent

On LinkedIn, the user @ollama announced that it has partnered with Alibaba's Qwen team, making the Qwen 3.8‑27B language model available on the Ollama platform. The post highlights that the model cons

local-aiollama--x
14 Aug 2026
Safety

Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence

DGX agent

arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling. Judges are typically validated by accuracy on

safetyarxiv-cs-ai
14 Aug 2026
Research

Keep, Customize, or Exit: Default Design and Token Pricing in LLM Reasoning Services

DGX agent

arXiv:2608.13315v1 Announce Type: cross Abstract: We study a large language model (LLM) service in which a provider chooses a per-token price and a default reasoning-token allocation, while a user may

researcharxiv-cs-ai
14 Aug 2026
Model Releases

Less than 2 hours to say hi👋. It's almost time! See you soon: 👀 https://huggingface.co/Qwen/Qwen3.8-27B

DGX agent

On August 14 2026 the Alibaba Qwen team tweeted a brief announcement that the Qwen 3.8‑27B model would be released on Hugging Face in less than two hours, urging followers to “say hi” and promising to

model-releasesqwen--x
14 Aug 2026
Research

Masked diffusion LLMs can use EoS tokens for hidden reasoning

DGX agent

arXiv:2603.05197v2 Announce Type: replace Abstract: Diffusion LLMs have been proposed as an alternative to autoregressive LLMs. Curiously, they are especially capable if the generation length, i.e., t

researcharxiv-cs-cl
14 Aug 2026
Tutorials

New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs

DGX agent

arXiv:2608.12361v1 Announce Type: new Abstract: Neologisms, emerging terms in meaning or form, can serve as new vehicles for toxic expression, like 'country girl' as a stigmatizing label targeting fem

tutorialsarxiv-cs-cl
14 Aug 2026
Model Releases

on DGX Spark: llama serve -hf ggml-org/Qwen3.8-27B-GGUF:Q4_K_M -hfd ggml-org/Qwen3.8-27B-GGUF:Q4_0 --spec-default --spec-type draft-…

DGX agent

Georgi Gerganov showcased running a LLaMA server on an NVIDIA DGX‑Spark, serving the Qwen 3.8‑27B model in GGUF format (`ggml-org/Qwen3.8-27B-GGUF`). The command demonstrates two quantization options—

model-releasesgeorgi-gerganov--x
14 Aug 2026
Model Releases

Predicting consumer-technology ownership without a diffusion history

DGX agent

arXiv:2608.12344v1 Announce Type: new Abstract: We test whether the perceived attributes of a consumer technology predict how widely it is owned. In a 2022 Prolific survey of US adults (n = 678), resp

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Reasoning for Social Audio-Visual Question Answering: Where Do We Stand?

DGX agent

arXiv:2608.13239v1 Announce Type: new Abstract: Training Multimodal Large Language Models for audio-visual social understanding is a crucial step toward embodied social intelligence. Chain-of-thought

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

Reliability-Aware Sexism Detection: Combining DPO with Annotator Agreement and Token-Level Confidence Scoring

DGX agent

arXiv:2608.12330v1 Announce Type: new Abstract: The detection of online sexism remains an open problem. Sexism detection is inherently subjective, yet most existing systems reduce multi-annotator labe

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

RoboSynChallenge: Mastering Real-World Dexterity via Generalizing Synthesized Manipulation Skills

DGX agent

arXiv:2608.12416v1 Announce Type: new Abstract: Achieving generalizable robotic manipulation remains a central challenge in embodied intelligence. Despite rapid advances in model architectures and lea

model-releasesarxiv-cs-ro
14 Aug 2026
Model Releases

Something has happened with post-training as shown by DeepSeek flash & GLM-5.3 updates. Same base, big improvement in perf to frontier level…

DGX agent

Something has happened with post-training as shown by DeepSeek flash & GLM-5.3 updates. Same base, big improvement in perf to frontier levels. Can't explain this by even logit distillation etc These a

model-releasesemad-mostaque--x
14 Aug 2026
Model Releases

TEMPO: Makespan-Aware Expert-Parallel Load Balancing Across Memory- and Compute-Bound Regimes

DGX agent

arXiv:2608.13057v1 Announce Type: cross Abstract: In expert-parallel (EP) MoE serving, every layer synchronizes at the slowest GPU. Dispatchers balance token counts (EPLB, LPLB, UltraEP) or activated-

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries and Referent Specificity

DGX agent

arXiv:2608.13484v1 Announce Type: cross Abstract: When asked about entities outside their knowledge boundary, LLMs routinely fabricate plausible-sounding details rather than backing off to safer, more

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Unifying Depth and Width Pruning for LLMs via Binary Knapsack Optimization

DGX agent

arXiv:2608.12953v1 Announce Type: new Abstract: Structured pruning is a promising approach for compressing large language models (LLMs), yet existing methods rely heavily on greedy heuristics that pro

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

v0.32.12

DGX agent

Qwen 3.8 - 27B model support This release adds the support of Qwen 3.8 27B. For Apple Silicon devices, Ollama has in particular optimized for maximum performance and output quality suitable for repeat

model-releasesollama-releases
14 Aug 2026
Model Releases

Web-search eating substantial quota

DGX agent

I prefer the model to have the internet as a source of truth and info and my coding work needs lot of web lookups. The middle bar section ( light blue) is the weekly quota taken by it. 120 web search

model-releasesr-ollama
14 Aug 2026
Model Releases

Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation

DGX agent

arXiv:2608.13337v1 Announce Type: new Abstract: Sparse autoencoders are meant to name the things a language model computes, and the usual way to check that a latent matters is to switch it off and see

model-releasesarxiv-cs-lg
14 Aug 2026
← Previous
1…417418419420421…1327
Next →