AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
14 Aug 2026

It's been a pleasure to partner with the @Alibaba_Qwen team.

Local AiDGX agent

On LinkedIn, the user @ollama announced that it has partnered with Alibaba's Qwen team, making the Qwen 3.8‑27B language model available on the Ollama platform. The post highlights that the model cons

Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence

SafetyDGX agent

arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling. Judges are typically validated by accuracy on

Keep, Customize, or Exit: Default Design and Token Pricing in LLM Reasoning Services

ResearchDGX agent

arXiv:2608.13315v1 Announce Type: cross Abstract: We study a large language model (LLM) service in which a provider chooses a per-token price and a default reasoning-token allocation, while a user may

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Less than 2 hours to say hi👋. It's almost time! See you soon: 👀 https://huggingface.co/Qwen/Qwen3.8-27B

Model ReleasesDGX agent

On August 14 2026 the Alibaba Qwen team tweeted a brief announcement that the Qwen 3.8‑27B model would be released on Hugging Face in less than two hours, urging followers to “say hi” and promising to

Masked diffusion LLMs can use EoS tokens for hidden reasoning

ResearchDGX agent

arXiv:2603.05197v2 Announce Type: replace Abstract: Diffusion LLMs have been proposed as an alternative to autoregressive LLMs. Curiously, they are especially capable if the generation length, i.e., t

New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs

TutorialsDGX agent

arXiv:2608.12361v1 Announce Type: new Abstract: Neologisms, emerging terms in meaning or form, can serve as new vehicles for toxic expression, like 'country girl' as a stigmatizing label targeting fem

on DGX Spark: llama serve -hf ggml-org/Qwen3.8-27B-GGUF:Q4_K_M -hfd ggml-org/Qwen3.8-27B-GGUF:Q4_0 --spec-default --spec-type draft-…

Model ReleasesDGX agent

Georgi Gerganov showcased running a LLaMA server on an NVIDIA DGX‑Spark, serving the Qwen 3.8‑27B model in GGUF format (`ggml-org/Qwen3.8-27B-GGUF`). The command demonstrates two quantization options—

Predicting consumer-technology ownership without a diffusion history

Model ReleasesDGX agent

arXiv:2608.12344v1 Announce Type: new Abstract: We test whether the perceived attributes of a consumer technology predict how widely it is owned. In a 2022 Prolific survey of US adults (n = 678), resp

Reasoning for Social Audio-Visual Question Answering: Where Do We Stand?

Model ReleasesDGX agent

arXiv:2608.13239v1 Announce Type: new Abstract: Training Multimodal Large Language Models for audio-visual social understanding is a crucial step toward embodied social intelligence. Chain-of-thought

Reliability-Aware Sexism Detection: Combining DPO with Annotator Agreement and Token-Level Confidence Scoring

Model ReleasesDGX agent

arXiv:2608.12330v1 Announce Type: new Abstract: The detection of online sexism remains an open problem. Sexism detection is inherently subjective, yet most existing systems reduce multi-annotator labe

RoboSynChallenge: Mastering Real-World Dexterity via Generalizing Synthesized Manipulation Skills

Model ReleasesDGX agent

arXiv:2608.12416v1 Announce Type: new Abstract: Achieving generalizable robotic manipulation remains a central challenge in embodied intelligence. Despite rapid advances in model architectures and lea

Something has happened with post-training as shown by DeepSeek flash & GLM-5.3 updates. Same base, big improvement in perf to frontier level…

Model ReleasesDGX agent

Something has happened with post-training as shown by DeepSeek flash & GLM-5.3 updates. Same base, big improvement in perf to frontier levels. Can't explain this by even logit distillation etc These a

TEMPO: Makespan-Aware Expert-Parallel Load Balancing Across Memory- and Compute-Bound Regimes

Model ReleasesDGX agent

arXiv:2608.13057v1 Announce Type: cross Abstract: In expert-parallel (EP) MoE serving, every layer synchronizes at the slowest GPU. Dispatchers balance token counts (EPLB, LPLB, UltraEP) or activated-

Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries and Referent Specificity

Model ReleasesDGX agent

arXiv:2608.13484v1 Announce Type: cross Abstract: When asked about entities outside their knowledge boundary, LLMs routinely fabricate plausible-sounding details rather than backing off to safer, more

Unifying Depth and Width Pruning for LLMs via Binary Knapsack Optimization

Model ReleasesDGX agent

arXiv:2608.12953v1 Announce Type: new Abstract: Structured pruning is a promising approach for compressing large language models (LLMs), yet existing methods rely heavily on greedy heuristics that pro

v0.32.12

Model ReleasesDGX agent

Qwen 3.8 - 27B model support This release adds the support of Qwen 3.8 27B. For Apple Silicon devices, Ollama has in particular optimized for maximum performance and output quality suitable for repeat

Web-search eating substantial quota

Model ReleasesDGX agent

I prefer the model to have the internet as a source of truth and info and my coding work needs lot of web lookups. The middle bar section ( light blue) is the weekly quota taken by it. 120 web search

Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation

Model ReleasesDGX agent

arXiv:2608.13337v1 Announce Type: new Abstract: Sparse autoencoders are meant to name the things a language model computes, and the usual way to check that a latent matters is to switch it off and see

13 Aug 2026

A Neighborhood Attention Transformer Network for Enhanced 3D Segmentation of the Left Anterior Descending Artery

Model ReleasesDGX agent

arXiv:2608.12274v1 Announce Type: cross Abstract: Background: Accurate segmentation of the Left Anterior Descending (LAD) artery in 3D free-breathing, non-contrast CT is critical for cardiac dose spar

Advancing MLLM-based UAV Image Understanding and Reasoning: A Benchmark and a Training-Free Multi-Agent System

Model ReleasesDGX agent

arXiv:2608.11738v1 Announce Type: cross Abstract: Multimodal Large Language Model (MLLM)-based UAV aerial image understanding and reasoning is essential for aerial intelligence yet poses distinct chal

Agent Safety Should Be a Runtime Contract

Model ReleasesDGX agent

arXiv:2608.11274v1 Announce Type: cross Abstract: The dominant paradigm treats AI safety as a property to be instilled during model training via RLHF, DPO, or Constitutional AI. We argue this is struc

APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference

ResearchDGX agent

arXiv:2608.11688v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models are attractive for edge deployment because they provide high model capacity while activating only a small subset of pa

Benchmarking Cyberattack Detection in Electric Vehicle Charging Infrastructure with Benign User Updates

Model ReleasesDGX agent

arXiv:2608.11286v1 Announce Type: cross Abstract: Cyberattack detection in electric vehicle charging infrastructure is complicated by legitimate post-activation revisions to requested energy and depar

Diffusion-Based Data-Driven Assortment Optimization

Local AiDGX agent

arXiv:2608.11419v1 Announce Type: new Abstract: Assortment optimization is a fundamental problem in revenue management, typically addressed using parametric choice models such as the multinomial logit

dots-studio/dots3-note-prev · Hugging Face

Local AiDGX agent

dots3-note preview is the first open-weight model in the dots3 family. It is a Mixture-of-Experts model with 280B total parameters, 16B activated parameters, and support for a context length of up to

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

Model ReleasesDGX agent

arXiv:2507.14267v2 Announce Type: replace Abstract: Large language model (LLM) agents can execute long-horizon scientific workflows, but their numerical outputs are difficult to trust: agents lose con

DS4 cloud (30 min) vs Qwen3.6 36B (2 min) vs Muse Glimmer 30B (3 min) on Llama.cpp (RTX 5080)

Model ReleasesDGX agent

Some people told me that the difference in richness and layout between Glimmer and Qwen wasn't clear to them. This example makes it super clear. I'm aware that comparing Glimmer 30B (a dense model) wi

GeoFlow: Efficient Driving Video Generation via Geometry-Aligned Priors

ResearchDGX agent

arXiv:2608.12203v1 Announce Type: new Abstract: Generative models like Diffusion Models and Flow Matching have demonstrated remarkable capabilities in synthesizing high-fidelity driving videos, but ar

LEMUR: Latent Entropy-aware Multimodal Unlearning via Visual-anchored Reasoning Redirection

ResearchDGX agent

arXiv:2608.11691v1 Announce Type: cross Abstract: Reinforcement-learning (RL) post-training equips multimodal large reasoning models (MLRMs) with exploratory chains of thought (CoT), substantially imp

Lifecycle-Optimal Tokenization: Vocabulary Size as a Deployment-Regime-Dependent Infrastructure Parameter

Model ReleasesDGX agent

arXiv:2608.11361v1 Announce Type: cross Abstract: Tokenizer vocabulary size is a foundational design choice in large language model (LLM) infrastructure, yet it is typically fixed at training time bas

Located but Not Releasable: Silent Gate Inversion and Bounded Linear Release

Model ReleasesDGX agent

arXiv:2608.11822v1 Announce Type: new Abstract: A growing body of work reports that language models represent task-relevant latent structure that they fail to use. Whether such structure, once located

Look What the Probes Dragged In! Real-World Chest X-ray Shortcuts in MedCLIP

ApplicationsDGX agent

arXiv:2608.12086v1 Announce Type: new Abstract: Vision-language models, such as contrastive language-image pre-training (CLIP)-based approaches, have reached state-of-the-art (SOTA) results in medical

LoRAQuant: Mixed-Precision Quantization of LoRA to Ultra-Low Bits

Model ReleasesDGX agent

arXiv:2510.26690v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a popular technique for parameter-efficient fine-tuning of large language models (LLMs). In many real-world sc

Quantifying the Relationship Between Clinical Safety and Environmental Impact in Therapeutic LLMs

SafetyDGX agent

arXiv:2608.11830v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) in mental health contexts raises questions about the relationship between clinical safety and environme

SoftWater: Class-Aware Rate Allocation for Softmax Quantization

Model ReleasesDGX agent

arXiv:2608.12026v1 Announce Type: new Abstract: Post-training quantization pipelines routinely leave the softmax output layer in high precision. Yet in small LLMs with modern vocabularies, the head ho

12 Aug 2026

Benchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxy

Model ReleasesDGX agent

arXiv:2608.10532v1 Announce Type: cross Abstract: Static load balancers cannot mitigate a backend that is degraded rather than down: round-robin and least-connections keep routing traffic to a server

ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

Model ReleasesDGX agent

arXiv:2608.10915v1 Announce Type: new Abstract: After an older adult misses a medication dose, a software agent can send another reminder and an embodied agent can bring the medication. Yet neither ex

E^3mo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise Alignment

Model ReleasesDGX agent

arXiv:2608.10796v1 Announce Type: new Abstract: Understanding both expressed and evoked emotions is critical for multimodal large language models (MLLMs) to achieve comprehensive affect-aware interact

Edge Phoneme Recognition for Children's Speech through Age-Aware Training

Model ReleasesDGX agent

arXiv:2608.10206v1 Announce Type: new Abstract: Detecting phonemes from children's speech has historically been difficult due to the scarcity of training data, and unique characteristics of children's

Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks

Model ReleasesDGX agent

arXiv:2608.10357v1 Announce Type: cross Abstract: Long-horizon tool-using agents must reason over user goals, domain policies, tool calls, simulator state, and delayed verifiable rewards. Reinforcemen

MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games

AgentsDGX agent

arXiv:2602.24188v2 Announce Type: replace Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games tha

MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment

Local AiDGX agent

arXiv:2608.11167v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) predominantly rely on image-text pairs for modality alignment pretraining, mapping global image repr

Order Matters in Retrosynthesis: Structure-aware Generation via Reaction-Center-Guided Discrete Flow Matching

Model ReleasesDGX agent

arXiv:2602.13136v2 Announce Type: replace Abstract: Template-free retrosynthesis methods treat the task as black-box sequence generation, limiting learning efficiency, while semi-template approaches r

Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference

Model ReleasesDGX agent

arXiv:2608.10288v1 Announce Type: cross Abstract: The Large Language Model from Power Law Decoder Representations (PLDR-LLM) and its attention, Power Law Graph Attention (PLGA), replace the fixed bili

Putting sign language AI into users’ hands

Model ReleasesDGX agent

Google DeepMind’s new sign‑language‑to‑text (SL2T) model powers real‑time sign-to-text dictation in Gboard and Live Transcribe on Pixel 11, currently supporting American Sign Language to English. The

Quantum Incremental Learning with Mixed State Prototypes

Model ReleasesDGX agent

arXiv:2608.10464v1 Announce Type: new Abstract: Incremental learning models are required to learn new classes sequentially without catastrophic forgetting, while operating under parameter and memory c

REDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent Systems

Model ReleasesDGX agent

arXiv:2608.10669v1 Announce Type: new Abstract: Large language model (LLM) agents combine language-based reasoning with external tools to perform complex tasks. Adversarial inputs can exploit interact

ReLTEx: Reliable LLM-based Taxonomy Expansion

Model ReleasesDGX agent

arXiv:2608.10970v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated strong capabilities in generating semantically relevant concepts and relations, maki

SafeCA: Safe Cross-Attention Localization and Regulation for Text-to-Video Jailbreak Defense

Local AiDGX agent

arXiv:2608.10933v1 Announce Type: new Abstract: Text-to-Video (T2V) generative models are vulnerable to jailbreak attacks in real-world deployment, leading them to produce harmful or inappropriate con

Static in Frames, Dynamic in Events: Rethinking Features in Event Cameras as Motion Cues

Model ReleasesDGX agent

arXiv:2608.11075v1 Announce Type: new Abstract: Event cameras capture intensity changes asynchronously with high temporal resolution, requiring novel preprocessing methods for downstream tasks. Unlike

The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs

Model ReleasesDGX agent

arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degrada

Try Grok 4.6 on tough real-world tasks!

Model ReleasesDGX agent

Try Grok 4.6 on tough real-world tasks! imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great

Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification

Model ReleasesDGX agent

arXiv:2608.10007v1 Announce Type: cross Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL

V-FiLLM: Verified Financial LLM Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar

VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection

AgentsDGX agent

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary mode

What DINO saw: ALiBi positional encoding reduces positional bias in Vision Transformers

SafetyDGX agent

arXiv:2603.16840v2 Announce Type: replace Abstract: Vision transformers (ViTs) - especially feature foundation models like DINOv2 - learn rich representations useful for many downstream tasks. However

When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning

Model ReleasesDGX agent

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of

11 Aug 2026

A Unified Issue Resolution Benchmark for Requirement Clarification, Planning, and Code Generation for Coding Agents

Model ReleasesDGX agent

arXiv:2608.09072v1 Announce Type: cross Abstract: Large language model-powered coding agents are increasingly used to modify existing code repositories, for example, by adding features or fixing bugs.

Accelerate cyber defense with OpenAI and AWS: Daybreak Red & Daybreak Blue now available to eligible customers on Amazon Bedrock

IndustryDGX agent

Daybreak Red and Daybreak Blue from OpenAI, specialized cyber defense models from OpenAI, are now available on Amazon Bedrock to eligible customers. Both models run with zero-operator access enforced

ACEvo: Adversarial Co-Evolution of Problem Distributions and Solvers for Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2506.02594v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to synthesize heuristic programs, yet most existing pipelines optimize solvers against fixed benc

← Previous
1…326327328329330…1042
Next →