AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study

DGX agent

arXiv:2608.13568v1 Announce Type: cross Abstract: Coding agents spend most of their context budget on retrieval. Lexical retrieval (grep) is universal, instant, and zero-setup, but noisy: it cannot te

model-releasesarxiv-cs-ai
17 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Federated Prompt Learning: A Unified Framework, Empirical Analysis, and Future Directions

DGX agent

arXiv:2608.13844v1 Announce Type: cross Abstract: Large language models (LLMs) have become core components of cloud-based intelligent services in academia and industry, yet their training and deployme

researcharxiv-cs-ai
17 Aug 2026
Model Releases

How accurate do you think this is? Qwen3.5 9B vs GPT-4o

DGX agent

Do you think today’s GPU-poor systems running Qwen 3.5 9B can outperform the ones we had with ChatGPT-4o? I mean, when 4o disappeared, people felt like they’d lost a great model, and now Qwen 3.5 9B Q

model-releasesr-localllama
17 Aug 2026
Model Releases

Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence

DGX agent

arXiv:2608.13958v1 Announce Type: new Abstract: How do we govern AI systems whose reasoning we cannot fully inspect? Governance does not require understanding a system's reasoning. It requires stating

model-releasesarxiv-cs-ai
17 Aug 2026
Safety

Language-Specific Gaps in AI Safety Training Datasets

DGX agent

arXiv:2608.13695v1 Announce Type: cross Abstract: Large language model providers routinely cite multilingual safety benchmarks spanning a dozen or more languages as evidence that their models are safe

safetyarxiv-cs-lg
17 Aug 2026
Model Releases

Learning-to-Transition for Large-scale and High-Order MIMO Detection

DGX agent

arXiv:2608.14511v1 Announce Type: cross Abstract: High-order multiple-input multiple-output (MIMO) detection requires efficient search over a large discrete symbol space while producing reliable soft

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

MAGneT-3D: Monocular and Domain-Generalizable Temporal 3D Detection

DGX agent

arXiv:2608.14282v1 Announce Type: new Abstract: Monocular temporal 3D detection aims to detect objects in 3D, given a monocular video. Query-based 3D detectors unify detection and cross-view associati

model-releasesarxiv-cs-cv
17 Aug 2026
Hardware

Omni-LiveAvatar: Minute-Level Real-Time Streaming Joint Audio-Visual Avatar Generation

DGX agent

arXiv:2608.13602v1 Announce Type: cross Abstract: Joint audio-video generative models serve as foundation for immersive and interactive digital-human generation. Nevertheless, most existing models rel

hardwarearxiv-cs-cv
17 Aug 2026
Model Releases

Owner3D: Ownership-Guided Style Writing for Training-Free Localized 3D Stylization

DGX agent

arXiv:2608.14078v1 Announce Type: new Abstract: Localized 3D stylization aims to modify the appearance of a specified object part while preserving the remaining surfaces. In large reconstruction model

model-releasesarxiv-cs-cv
17 Aug 2026
Safety

READ: A Retrieval-Alignment Diffusion Framework for Structure-based Drug Design

DGX agent

arXiv:2506.14488v2 Announce Type: replace-cross Abstract: Structure-based drug design (SBDD) models are central to modern pharmaceutical research, enabling the rational exploration of protein-ligand i

safetyarxiv-cs-lg
17 Aug 2026
Model Releases

Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

DGX agent

arXiv:2608.14089v1 Announce Type: new Abstract: Safety classifiers deployed with large language models often fail for two reasons: their decisions reflect the policy learned during training rather tha

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Retrieve-then-Adapt: Retrieval-Augmented Test-Time Adaptation for Sequential Recommendation

DGX agent

arXiv:2604.05379v2 Announce Type: replace-cross Abstract: The sequential recommendation (SR) task aims to predict the next item based on users' historical interaction sequences. Typically trained on h

model-releasesarxiv-cs-lg
17 Aug 2026
Research

RGBX-Next: Towards Realistic Generative Rendering from G-Buffers

DGX agent

arXiv:2608.13929v1 Announce Type: new Abstract: Diffusion models have achieved impressive results in image, video, and streaming generation. However, compared to traditional 3D rendering, they still l

researcharxiv-cs-cv
17 Aug 2026
Model Releases

SAGE: Surrogate-gradient Adaptation via Attention-Guided Entropy for Spiking Transformers

DGX agent

arXiv:2608.13702v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer an energy-efficient alternative to conventional deep neural networks by exploiting sparse event-driven computatio

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Simulation-Aware In-Context Policy Improvement for LLM-Aided Analog Layout Refinement

DGX agent

arXiv:2608.13767v1 Announce Type: new Abstract: Analog IC layout design remains a labor-intensive iterative process dominated by simulation-driven refinement. Although end-to-end layout generators acc

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Simulation-Driven Vehicular Traffic Data Augmentation: Extending Sensor Coverage Through Virtual Sensing

DGX agent

arXiv:2608.13993v1 Announce Type: new Abstract: Urban traffic management relies on sensor networks whose spatial coverage is limited by deployment costs and privacy regulations. Machine learning model

researcharxiv-cs-ai
17 Aug 2026
Model Releases

The Dynamics of Intelligence Explosions

DGX agent

arXiv:2608.14426v1 Announce Type: new Abstract: AI is increasingly being used to help with AI R&D. Under certain conditions this feedback loop might be able to produce an intelligence explosion, with

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

XtraLight-MedMamba for Classification of Neoplastic Tubular Adenomas

DGX agent

arXiv:2602.04819v5 Announce Type: replace Abstract: Accurate risk stratification of precancerous polyps during routine colonoscopy screening is a key strategy to reduce the incidence of colorectal can

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Zero-Shot Skeleton-Based Action Anticipation

DGX agent

arXiv:2608.14243v1 Announce Type: new Abstract: Action anticipation (AA) aims to recognize ongoing human or humanoids actions from partial observations, enabling robots to predict intentions before th

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Anyone managed to get Qwen 3.8 27B running smoothly on vLLM? Can't get rid of endless thinking

DGX agent

Title pretty much says it all. I’ve deployed Qwen 3.8 27B using vLLM on an RTX 6000 Pro (tried multiple vLLM releases and launch recipes), but I can't get it into a usable state because of crazy long

model-releasesr-localllama
16 Aug 2026
Model Releases

b10453

DGX agent

model : remove some ggml_concat (#27176) Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabl

model-releasesllama-cpp-releases
16 Aug 2026
Model Releases

Best Setup for a 16 GB VRAM + 128 GB RAM System?

DGX agent

Running a 12700k + 5060 Ti 16 gb with 128 gb DDR4 ram and I'm wondering what's the ideal setup to maximize performance. I did some preliminary stuff but I have to admit, I'm still learning and kinda j

model-releasesr-localllama
16 Aug 2026
Model Releases

Huihui-ai Qwen 3.8 Ablit Available

DGX agent

At hugging face, this is the model I use for most of my analysis that most of my services have to not be refused. previous versions are quite good. It seems this might have dropped today and pulling r

model-releasesr-localllama
16 Aug 2026
Model Releases

Mindblow with Qwen 3.8 (Cline + VSC on a 5090 mobile)

DGX agent

Honestly I am surprised at the speed and performance I am getting from the model, it does not allow for complete hands off like Claude 5.0 but that is not what I want, i want to be able to iterate and

model-releasesr-ollama
16 Aug 2026
Model Releases

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈…

DGX agent

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈 It uses a harness + model set that is tuned specifically for

model-releasesjerry-liu--x
16 Aug 2026
Safety

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because …

DGX agent

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.

safetydario-amodei--x
15 Aug 2026
Model Releases

b10447

DGX agent

server: re-design yield_to_queue thread model (#27133) run common_speculative_process in worker swap worker <--> main thread design Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) ma

model-releasesllama-cpp-releases
15 Aug 2026
Model Releases

1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)

DGX agent

Processing img az99qopcg8jh1... So same as my prior post 1bit test... although this 1bit is a bit interesting you can read on unlsoth blog https://unsloth.ai/docs/models/qwen3.8 508 gigs being used I

model-releasesr-localllama
14 Aug 2026
Model Releases

A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family

DGX agent

arXiv:2608.12700v1 Announce Type: new Abstract: Systems that generate GPU kernels with language models report high correctness rates. Those rates come from a single loose test: run the kernel on a few

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

A Unified Framework for Joint Detection of Lacunes and Enlarged Perivascular Spaces

DGX agent

arXiv:2603.04243v3 Announce Type: replace Abstract: Cerebral small vessel disease (CSVD) markers, specifically enlarged perivascular spaces (EPVS) and lacunae, present a unique challenge in medical im

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

b10434

DGX agent

chat : pass reasoning_effort to template chat: add reasoning_effort to common_chat_templates_inputs Store OpenAI Chat Completions reasoning_effort and make it available to jinja templates (with model

model-releasesllama-cpp-releases
14 Aug 2026
Hardware

Balanced Adaptive Prototype Selection for Scalable TabPFN Inference on Large-Scale Tabular Data

DGX agent

arXiv:2608.12989v1 Announce Type: new Abstract: Pretrained tabular foundation models have demonstrated strong predictive capability; however, their application to large-scale datasets remains constrai

hardwarearxiv-cs-lg
14 Aug 2026
Safety

CROP: Task Relevance via Counterfactuals for Selective On-Policy Distillation

DGX agent

arXiv:2608.13387v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises a student language model on trajectories sampled from its current policy, but assigns equal credit to response t

safetyarxiv-cs-cl
14 Aug 2026
Model Releases

damn deepseek moves fast on roon tweets

DGX agent

damn deepseek moves fast on roon tweets the major ai companies should commit to a real time priced API product. ai demand varies wildly over a day/night curve and the industry is broadly extremely cap

model-releasesswyx--x
14 Aug 2026
Research

Dual-Flow Transformers: Decoupling the Primary Prefill Path from Additional Decode Computation

DGX agent

arXiv:2608.12385v1 Announce Type: new Abstract: As large language models serve more requests, cumulative inference cost is becoming increasingly important relative to one-time training cost. The two i

researcharxiv-cs-ai
14 Aug 2026
Research

Enhancing In-Hospital Mortality Prediction Using Multi-Representational Learning with LLM-Generated Expert Summaries

DGX agent

arXiv:2411.16818v2 Announce Type: replace-cross Abstract: To evaluate a multi-representational framework in which large language model (LLM)-generated expert summaries of intensive care unit (ICU) not

researcharxiv-cs-ai
14 Aug 2026
Model Releases

ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval

DGX agent

arXiv:2608.12720v1 Announce Type: cross Abstract: While Large Language Model (LLM) agents increasingly rely on long-term memory for persistent interactions, the retrieval mechanisms governing this mem

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 suppo…

DGX agent

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 support! @lightseekorg We’re proud to be the Day 0 open-source inf

model-releasesqwen--x
14 Aug 2026
Model Releases

Fixed/improved Jinja chat template for Qwen 3.8

DGX agent

'Again?' you might ask. I(*) took the original Qwen 27B 3.8 chat template and compared it to the improved template that was posted a day before. The original has issues. The improved one fixed some wh

model-releasesr-localllama
14 Aug 2026
Model Releases

FUSE: Active Functional Affordance Grounding through Adaptive Semantic-Geometric Evidence Acquisition

DGX agent

arXiv:2608.12683v1 Announce Type: cross Abstract: Embodied agents must often identify and interact with objects based on their function rather than their identity, requiring them to actively acquire o

model-releasesarxiv-cs-cv
14 Aug 2026
Applications

GENADA: efficient generative time series adversarial attack framework

DGX agent

arXiv:2608.12535v1 Announce Type: new Abstract: Deep learning models are widely used for time series analysis in domains such as healthcare, finance, energy systems, and environmental monitoring. Howe

applicationsarxiv-cs-lg
14 Aug 2026
Model Releases

HumanoidVLN: A Physics-Grounded Simulator and Benchmark for Vision-Language Navigation Across Diverse Humanoid Embodiments

DGX agent

arXiv:2608.12860v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) for humanoid robots poses challenges existing benchmarks fail to address: bipedal locomotion imposes physical constrain

model-releasesarxiv-cs-ro
14 Aug 2026
Safety

HybridSB-MoE: Dual-Domain Schrodinger Bridges with Scene-Adaptive Expert Routing for Speech Enhancement

DGX agent

arXiv:2608.12715v1 Announce Type: cross Abstract: Generative speech enhancement faces three gaps: spectral models capture harmonic structure but often disrupt phase, waveform models preserve phase but

safetyarxiv-cs-ai
14 Aug 2026
Research

LOB-ID: Evaluating Synthetic Market Data by Inception Distances

DGX agent

arXiv:2608.13082v1 Announce Type: cross Abstract: Generative models of limit orderbook (LOB) data have advanced rapidly, but their evaluation often focuses on stylised facts and selected market statis

researcharxiv-cs-ai
14 Aug 2026
Model Releases

MAG: MAnifold Guided Semi-Supervised Multi-modal In-Context Learning

DGX agent

arXiv:2608.12724v1 Announce Type: new Abstract: Few-shot in-context learning (ICL) with multi-modal large language models (MLLMs) enables task adaptation without parameter updates, but its performance

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

Mistral is now hosting GLM-5.2

DGX agent

Not directly LOCALLlama related but I thought it was interesting since Mistral and Z.ai are competitors, and more surprisingly they are pricing it (GLM-5.2) even cheaper than their current flagship mo

model-releasesr-localllama
14 Aug 2026
Agents

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

DGX agent

arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

Qwen3.8-Max is coming to Nebius Token Factory on Day 0. Great to kick things off together with Nebius as our Day 0 launch partner, bringing …

DGX agent

Qwen3.8-Max is coming to Nebius Token Factory on Day 0. Great to kick things off together with Nebius as our Day 0 launch partner, bringing dedicated inference to more users. Big model, right from the

model-releasesqwen--x
14 Aug 2026
← Previous
1…493494495496497…1371
Next →