AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

Zero-Shot Skeleton-Based Action Anticipation

DGX agent

arXiv:2608.14243v1 Announce Type: new Abstract: Action anticipation (AA) aims to recognize ongoing human or humanoids actions from partial observations, enabling robots to predict intentions before th

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Anyone managed to get Qwen 3.8 27B running smoothly on vLLM? Can't get rid of endless thinking

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Title pretty much says it all. I’ve deployed Qwen 3.8 27B using vLLM on an RTX 6000 Pro (tried multiple vLLM releases and launch recipes), but I can't get it into a usable state because of crazy long

model-releasesr-localllama
16 Aug 2026
Model Releases

b10453

DGX agent

model : remove some ggml_concat (#27176) Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabl

model-releasesllama-cpp-releases
16 Aug 2026
Model Releases

Best Setup for a 16 GB VRAM + 128 GB RAM System?

DGX agent

Running a 12700k + 5060 Ti 16 gb with 128 gb DDR4 ram and I'm wondering what's the ideal setup to maximize performance. I did some preliminary stuff but I have to admit, I'm still learning and kinda j

model-releasesr-localllama
16 Aug 2026
Model Releases

Huihui-ai Qwen 3.8 Ablit Available

DGX agent

At hugging face, this is the model I use for most of my analysis that most of my services have to not be refused. previous versions are quite good. It seems this might have dropped today and pulling r

model-releasesr-localllama
16 Aug 2026
Model Releases

Mindblow with Qwen 3.8 (Cline + VSC on a 5090 mobile)

DGX agent

Honestly I am surprised at the speed and performance I am getting from the model, it does not allow for complete hands off like Claude 5.0 but that is not what I want, i want to be able to iterate and

model-releasesr-ollama
16 Aug 2026
Model Releases

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈…

DGX agent

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈 It uses a harness + model set that is tuned specifically for

model-releasesjerry-liu--x
16 Aug 2026
Safety

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because …

DGX agent

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.

safetydario-amodei--x
15 Aug 2026
Model Releases

b10447

DGX agent

server: re-design yield_to_queue thread model (#27133) run common_speculative_process in worker swap worker <--> main thread design Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) ma

model-releasesllama-cpp-releases
15 Aug 2026
Model Releases

1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)

DGX agent

Processing img az99qopcg8jh1... So same as my prior post 1bit test... although this 1bit is a bit interesting you can read on unlsoth blog https://unsloth.ai/docs/models/qwen3.8 508 gigs being used I

model-releasesr-localllama
14 Aug 2026
Model Releases

A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family

DGX agent

arXiv:2608.12700v1 Announce Type: new Abstract: Systems that generate GPU kernels with language models report high correctness rates. Those rates come from a single loose test: run the kernel on a few

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

A Unified Framework for Joint Detection of Lacunes and Enlarged Perivascular Spaces

DGX agent

arXiv:2603.04243v3 Announce Type: replace Abstract: Cerebral small vessel disease (CSVD) markers, specifically enlarged perivascular spaces (EPVS) and lacunae, present a unique challenge in medical im

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

b10434

DGX agent

chat : pass reasoning_effort to template chat: add reasoning_effort to common_chat_templates_inputs Store OpenAI Chat Completions reasoning_effort and make it available to jinja templates (with model

model-releasesllama-cpp-releases
14 Aug 2026
Hardware

Balanced Adaptive Prototype Selection for Scalable TabPFN Inference on Large-Scale Tabular Data

DGX agent

arXiv:2608.12989v1 Announce Type: new Abstract: Pretrained tabular foundation models have demonstrated strong predictive capability; however, their application to large-scale datasets remains constrai

hardwarearxiv-cs-lg
14 Aug 2026
Safety

CROP: Task Relevance via Counterfactuals for Selective On-Policy Distillation

DGX agent

arXiv:2608.13387v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises a student language model on trajectories sampled from its current policy, but assigns equal credit to response t

safetyarxiv-cs-cl
14 Aug 2026
Model Releases

damn deepseek moves fast on roon tweets

DGX agent

damn deepseek moves fast on roon tweets the major ai companies should commit to a real time priced API product. ai demand varies wildly over a day/night curve and the industry is broadly extremely cap

model-releasesswyx--x
14 Aug 2026
Research

Dual-Flow Transformers: Decoupling the Primary Prefill Path from Additional Decode Computation

DGX agent

arXiv:2608.12385v1 Announce Type: new Abstract: As large language models serve more requests, cumulative inference cost is becoming increasingly important relative to one-time training cost. The two i

researcharxiv-cs-ai
14 Aug 2026
Research

Enhancing In-Hospital Mortality Prediction Using Multi-Representational Learning with LLM-Generated Expert Summaries

DGX agent

arXiv:2411.16818v2 Announce Type: replace-cross Abstract: To evaluate a multi-representational framework in which large language model (LLM)-generated expert summaries of intensive care unit (ICU) not

researcharxiv-cs-ai
14 Aug 2026
Model Releases

ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval

DGX agent

arXiv:2608.12720v1 Announce Type: cross Abstract: While Large Language Model (LLM) agents increasingly rely on long-term memory for persistent interactions, the retrieval mechanisms governing this mem

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 suppo…

DGX agent

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 support! @lightseekorg We’re proud to be the Day 0 open-source inf

model-releasesqwen--x
14 Aug 2026
Model Releases

Fixed/improved Jinja chat template for Qwen 3.8

DGX agent

'Again?' you might ask. I(*) took the original Qwen 27B 3.8 chat template and compared it to the improved template that was posted a day before. The original has issues. The improved one fixed some wh

model-releasesr-localllama
14 Aug 2026
Model Releases

FUSE: Active Functional Affordance Grounding through Adaptive Semantic-Geometric Evidence Acquisition

DGX agent

arXiv:2608.12683v1 Announce Type: cross Abstract: Embodied agents must often identify and interact with objects based on their function rather than their identity, requiring them to actively acquire o

model-releasesarxiv-cs-cv
14 Aug 2026
Applications

GENADA: efficient generative time series adversarial attack framework

DGX agent

arXiv:2608.12535v1 Announce Type: new Abstract: Deep learning models are widely used for time series analysis in domains such as healthcare, finance, energy systems, and environmental monitoring. Howe

applicationsarxiv-cs-lg
14 Aug 2026
Model Releases

HumanoidVLN: A Physics-Grounded Simulator and Benchmark for Vision-Language Navigation Across Diverse Humanoid Embodiments

DGX agent

arXiv:2608.12860v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) for humanoid robots poses challenges existing benchmarks fail to address: bipedal locomotion imposes physical constrain

model-releasesarxiv-cs-ro
14 Aug 2026
Safety

HybridSB-MoE: Dual-Domain Schrodinger Bridges with Scene-Adaptive Expert Routing for Speech Enhancement

DGX agent

arXiv:2608.12715v1 Announce Type: cross Abstract: Generative speech enhancement faces three gaps: spectral models capture harmonic structure but often disrupt phase, waveform models preserve phase but

safetyarxiv-cs-ai
14 Aug 2026
Research

LOB-ID: Evaluating Synthetic Market Data by Inception Distances

DGX agent

arXiv:2608.13082v1 Announce Type: cross Abstract: Generative models of limit orderbook (LOB) data have advanced rapidly, but their evaluation often focuses on stylised facts and selected market statis

researcharxiv-cs-ai
14 Aug 2026
Model Releases

MAG: MAnifold Guided Semi-Supervised Multi-modal In-Context Learning

DGX agent

arXiv:2608.12724v1 Announce Type: new Abstract: Few-shot in-context learning (ICL) with multi-modal large language models (MLLMs) enables task adaptation without parameter updates, but its performance

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

Mistral is now hosting GLM-5.2

DGX agent

Not directly LOCALLlama related but I thought it was interesting since Mistral and Z.ai are competitors, and more surprisingly they are pricing it (GLM-5.2) even cheaper than their current flagship mo

model-releasesr-localllama
14 Aug 2026
Agents

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

DGX agent

arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

Qwen3.8-Max is coming to Nebius Token Factory on Day 0. Great to kick things off together with Nebius as our Day 0 launch partner, bringing …

DGX agent

Qwen3.8-Max is coming to Nebius Token Factory on Day 0. Great to kick things off together with Nebius as our Day 0 launch partner, bringing dedicated inference to more users. Big model, right from the

model-releasesqwen--x
14 Aug 2026
Safety

SEMA: Simple yet Effective Learning for Multi-Turn Jailbreak Attacks

DGX agent

arXiv:2602.06854v2 Announce Type: replace Abstract: Multi-turn jailbreaks capture the real threat model for safety-aligned chatbots, where single-turn attacks are merely a special case. Yet existing a

safetyarxiv-cs-cl
14 Aug 2026
Safety

Semantic Steering for Controllable Generation: Tuning-Free Concept Erasure in Multimodal Diffusion Transformers

DGX agent

arXiv:2608.12829v1 Announce Type: new Abstract: Multimodal Diffusion Transformers (MM-DiTs) have demonstrated remarkable text-to-image generation performance, surpassing traditional U-Net-based diffus

safetyarxiv-cs-cv
14 Aug 2026
Safety

SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy

DGX agent

arXiv:2604.03181v2 Announce Type: replace-cross Abstract: Robotic manipulation requires understanding both the 3D spatial structure of the environment and its temporal evolution, yet most existing pol

safetyarxiv-cs-cv
14 Aug 2026
Safety

StateBridge: Training-free Hidden-state Alignment for Latent Communication in LLM Multi-Agent Systems

DGX agent

arXiv:2608.13317v1 Announce Type: new Abstract: Large language model based multi-agent systems usually communicate in text, i.e., using discrete tokens. However, text introduces a discrete bottleneck.

safetyarxiv-cs-ai
14 Aug 2026
Model Releases

StreamTTT: Reconciling Real-Time Perception and Long-Term Memory in Streaming VLMs

DGX agent

arXiv:2608.13416v1 Announce Type: new Abstract: Humans effortlessly perceive the present while remembering the past, yet streaming VLMs often trade off real-time perception against long-term memory. P

model-releasesarxiv-cs-cv
14 Aug 2026
Safety

Synthetic Persona Pretraining: Alignment from Token Zero

DGX agent

arXiv:2608.13482v1 Announce Type: cross Abstract: As language-model-based AI is increasingly deployed in autonomous settings, aligning its goals and values with those of humans becomes critical. Today

safetyarxiv-cs-ai
14 Aug 2026
Safety

Unmasking Conversational Bias in AI Multiagent Systems

DGX agent

arXiv:2501.14844v3 Announce Type: replace-cross Abstract: Detecting biases in the outputs produced by generative models is essential to reduce the potential risks associated with their application in

safetyarxiv-cs-ai
14 Aug 2026
Research

V-RAE: Rethinking Video Latent Spaces for Generation

DGX agent

arXiv:2608.13556v1 Announce Type: new Abstract: Latent video generation relies on autoencoders to define a compact space in which generative models operate. Although video autoencoder architectures ha

researcharxiv-cs-cv
14 Aug 2026
Research

Adaptive Online Learning with LSTM Networks for Energy Price Prediction

DGX agent

arXiv:2510.16898v2 Announce Type: replace-cross Abstract: Accurate prediction of electricity prices is crucial for stakeholders in the energy market, particularly for grid operators, energy producers,

researcharxiv-cs-ai
13 Aug 2026
Model Releases

b10415

DGX agent

spec : auto-detect mtp draft model type (#27005) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents

DGX agent

arXiv:2511.20597v2 Announce Type: replace-cross Abstract: The integration of artificial intelligence (AI) agents into web browsers introduces security challenges that go beyond traditional web applica

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval

DGX agent

arXiv:2608.11343v1 Announce Type: new Abstract: Multimodal retrieval and classification across different types of media, spanning text, images,video and audio, has traditionally relied on dual-encoder

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Causal Structure is Inducible but Functionally Decoupled: The Routing/Readout Boundary of a Typed Mechanism Library

DGX agent

arXiv:2608.11767v1 Announce Type: new Abstract: When a language model answers an interventional question, the computation it must perform depends on the type of evidence the query requires. We report

model-releasesarxiv-cs-cl
13 Aug 2026
Research

Certifying What Helps Customer-Return Timing: A Screen-and-Confirm Test for Conditioning Signals, and Why Decay Is Nearly Enough

DGX agent

arXiv:2608.11555v1 Announce Type: new Abstract: Practitioners enrich customer-return models with ever more signals (lifetime value, category, recency/frequency, calendar, geography), and the temporal-

researcharxiv-cs-lg
13 Aug 2026
Model Releases

Commonsense on Demand: Generating and Selectively Integrating Commonsense Knowledge for Natural Language Inference

DGX agent

arXiv:2507.15100v3 Announce Type: replace-cross Abstract: Natural Language Inference (NLI) determines whether a premise entails, contradicts, or is neutral with respect to a hypothesis. The task is of

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression

DGX agent

arXiv:2608.11249v1 Announce Type: cross Abstract: We study the problem of lossless text compression, motivated by the rapid growth in the collection and storage of digital textual data - including pla

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Explainability in Practice: A Survey of Explainable NLP Across Various Domains

DGX agent

arXiv:2502.00837v3 Announce Type: replace-cross Abstract: Natural Language Processing (NLP) is now embedded in critical sectors including healthcare, finance, and customer relationship management, whe

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

From Synthesis to Removal: Physics-Grounded Reflection Simulation and Diffusion-Based Video Dereflection

DGX agent

arXiv:2608.11562v1 Announce Type: cross Abstract: Videos captured through glass often contain reflections that degrade visual quality and interfere with downstream vision tasks. Although single-image

model-releasesarxiv-cs-ai
13 Aug 2026
← Previous
1…488489490491492…1359
Next →