AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Research

Mapping General-Purpose AI Governance in Twenty AI Middle-Power Jurisdictions

DGX agent

arXiv:2608.19278v1 Announce Type: cross Abstract: The most capable general-purpose AI (GPAI) models are mostly built in two jurisdictions, the United States and China, but the risks they carry land gl

researcharxiv-cs-ai
21 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mitigating GenAI-Powered Evidence Pollution for Out-Of-Context Misinformation Detection

DGX agent

arXiv:2501.14728v2 Announce Type: replace-cross Abstract: While generative artificial intelligence (GenAI) models have achieved significant success, their misuse for generating deceptive content raise

model-releasesarxiv-cs-cl
21 Aug 2026
Model Releases

Natural Language Code Retrieval for 1C:Enterprise: An Open Benchmark and Efficient Bi-Encoder

DGX agent

arXiv:2608.19957v1 Announce Type: new Abstract: Natural language code retrieval is a rapidly evolving task in computer science. However, the 1C:Enterprise ecosystem combines Russian syntax with highly

model-releasesarxiv-cs-cl
21 Aug 2026
Model Releases

Optimal Skill Selection for LLM Agents with Provable Bicriteria Guarantees

DGX agent

arXiv:2608.19993v1 Announce Type: new Abstract: Loading reusable skill documents into a bounded context window is now the primary way large language model (LLM) agents acquire task-specific capabiliti

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

Our general-purpose coding agent just scored 100% on the ARC-AGI-3 interactive reasoning benchmark. NVIDIA AVO completed all 183 levels acro…

DGX agent

Our general-purpose coding agent just scored 100% on the ARC-AGI-3 interactive reasoning benchmark. NVIDIA AVO completed all 183 levels across all 25 public environments, figuring out what to do with

model-releasesclem-delangue--x
21 Aug 2026
Model Releases

PETA:Parameter-Efficient Test-Time Adaptation for Virtual Screening

DGX agent

arXiv:2608.19906v1 Announce Type: new Abstract: Accurately ranking active ligands for a target protein pocket from massive chemical libraries remains a central challenge in virtual screening. DrugCLIP

model-releasesarxiv-cs-lg
21 Aug 2026
Research

Phantom Gains: Auditing Self-Improvement Against a Measured Null

DGX agent

arXiv:2608.20290v1 Announce Type: new Abstract: Whether a language model has improved itself is increasingly judged not by mean accuracy but by which individual problems it gains and loses. Tracking t

researcharxiv-cs-ai
21 Aug 2026
Model Releases

Qwen 3.8 27b - PI AGENT vs OPENCODE

DGX agent

https://www.reddit.com/r/LocalLLaMA/comments/1j7r47l/i_just_made_an_animation_of_a_ball_bouncing/ This post inspired me to make that test after a year ;) That is one of my many tests I make comparing

model-releasesr-localllama
21 Aug 2026
Model Releases

Qwen 3.8 27b is strong even at Q3_xxs

DGX agent

So usually I avoid Q3 quants because I have had bad experiences with it, models were usually too degraded, so the smallest I normally do is Q4, since I only have rtx 4060 ti 16gb. But since there hasn

model-releasesr-localllama
21 Aug 2026
Model Releases

Rationally Enriched Chebyshev Trunk Bases for DeepONet Surrogates of High Peclet Entrance Transport

DGX agent

arXiv:2608.19658v1 Announce Type: new Abstract: This study demonstrates a rationally enriched Chebyshev (REC) trunk for deep operator network (DeepONet) surrogate models of singularly perturbed and hi

model-releasesarxiv-cs-lg
21 Aug 2026
Model Releases

Rethinking the Evaluation and Optimization of LLM-Based Social Simulation

DGX agent

arXiv:2608.19689v1 Announce Type: new Abstract: LLM-based social simulation is a promising complement to traditional methods such as surveys and behavioral experiments. A core question is how to evalu

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

Separating Covariate Shift from Mechanism Change with Two Discriminators: CJSD, a Conditional Discrepancy with an Exact Covariate-Concept Decomposition

DGX agent

arXiv:2608.19885v1 Announce Type: cross Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an existing expert for arriving data, spawn a new one,

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

Spiking Local Interaction and Adaptive Complementary Fusion for Spiking Transformer

DGX agent

arXiv:2608.19238v1 Announce Type: cross Abstract: Spiking Transformers model token interactions primarily through spiking self-attention (SSA). However, binary query and key representations map contin

model-releasesarxiv-cs-cv
21 Aug 2026
Model Releases

Table2Image: Lightweight Tabular Learning with Generated Proxy Representations and Reliability Diagnostics

DGX agent

arXiv:2412.06265v3 Announce Type: replace Abstract: Deep tabular models should ideally balance predictive performance, parameter efficiency, and robustness to imperfect learning signals---properties t

model-releasesarxiv-cs-lg
21 Aug 2026
Model Releases

This is very nice work from NVIDIA. Like all high-performing approaches on ARC-AGI-3, it uses deep learning-guided on-the-fly synthesis of s…

DGX agent

This is very nice work from NVIDIA. Like all high-performing approaches on ARC-AGI-3, it uses deep learning-guided on-the-fly synthesis of symbolic world models, i.e. navigating the world by generatin

model-releasesfrancois-chollet--x
21 Aug 2026
Model Releases

Towards Clinically Faithful Medical Image Captioning via Enhanced Vision-Language Alignment

DGX agent

arXiv:2608.19825v1 Announce Type: cross Abstract: Medical image captioning is a technique that accelerates early-stage diagnostic workflows and enhances the interpretability of medical diagnostic AI s

model-releasesarxiv-cs-cl
21 Aug 2026
Model Releases

YolovN-CBi: A Lightweight and Efficient Architecture for Real-Time Detection of Small UAVs

DGX agent

arXiv:2512.18046v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles, commonly known as, drones pose increasing risks in civilian and defense settings, demanding accurate and real-time drone d

model-releasesarxiv-cs-cv
21 Aug 2026
Model Releases

3 days benchmarking most llama.cpp flags on my weird 40gb vram laptop + tb4 egpu setup. Got +70% generation, +40% prefill, 60k more context, and filed a bug in llama around MTP. What I learned.

DGX agent

tldr: went from 16~ t/s to 27~ t/s generation. got my usable context up from 220k to the full 262k without sacrificing anything. prefill also increased from 376 to 573 command I ended up with, fwiw: l

model-releasesr-localllama
20 Aug 2026
Model Releases

A Few Cases Are All You Need: An Empirical Study of Annotation-Efficient LoRA Fine-Tuning of MedSAM3

DGX agent

arXiv:2608.18731v1 Announce Type: cross Abstract: Medical image segmentation is essential for clinical workflows such as treatment planning and disease assessment. While specialist tools like TotalSeg

model-releasesarxiv-cs-ai
20 Aug 2026
Safety

Accelerating Visual On-Policy Distillation with Batched Speculative Jacobi Rollouts

DGX agent

arXiv:2608.18183v1 Announce Type: new Abstract: Visual on-policy distillation (OPD) improves the training of compact visual autoregressive models by learning from trajectories generated by the current

safetyarxiv-cs-lg
20 Aug 2026
Applications

Assessing Quality of Experience in Natural Language Generation of German Text

DGX agent

arXiv:2608.18888v1 Announce Type: new Abstract: The rapid advancement of Natural Language Generation (NLG) has made the reliable evaluation of generated text increasingly critical, as these systems, s

applicationsarxiv-cs-cl
20 Aug 2026
Research

BERTilda: Explainable Topic Lifecycle Tracking with Split/Merge Detection via Similarity-and-Flow Temporal Graphs

DGX agent

arXiv:2608.18101v1 Announce Type: new Abstract: Longitudinal text streams exhibit topic birth and death, but also discrete structural reorganizations in which themes split into subtopics or merge into

researcharxiv-cs-cl
20 Aug 2026
Model Releases

CausalProfiler: Generating Synthetic Benchmarks for Rigorous and Transparent Evaluation of Causal Machine Learning

DGX agent

arXiv:2511.22842v3 Announce Type: replace-cross Abstract: Causal machine learning (Causal ML) aims to answer 'what if' questions using machine learning algorithms, making it a promising tool for high-

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents

DGX agent

arXiv:2608.18307v1 Announce Type: new Abstract: Current evaluation of computer-use agents is split between long-horizon workflow benchmarks and atomic GUI-grounding tests. This leaves an under-instrum

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence

DGX agent

arXiv:2608.18613v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) is increasingly consumed not by human analysts but by LLM agents that compose multi-step investigations at query time. T

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Event-Causal RAG: A Retrieval-Augmented Generation Framework for Long Video Reasoning in Complex Scenarios

DGX agent

arXiv:2605.06185v2 Announce Type: replace Abstract: Large vision-language models perform well on short- and medium-length video understanding but still struggle to maintain coherent event memory and r

model-releasesarxiv-cs-ai
20 Aug 2026
Safety

extsc{TestifAI}: Tomography-Based Testing for Deep Learning Systems

DGX agent

arXiv:2608.18900v1 Announce Type: new Abstract: As AI systems are increasingly deployed in safety-critical application domains (e.g., autonomous driving), associated risks increase too. Deep learning

safetyarxiv-cs-ai
20 Aug 2026
Model Releases

Fine-tuning Cactus Needle 2 can match DeepSeek v4 on the specific task

DGX agent

Hey LocalLlama, Henry from Cactus here! When we trained Needle 2, I had a strict rule to not expose the model to any data sample that remotely felt like these benchmarks. It seemed over-the-top, but b

model-releasesr-localllama
20 Aug 2026
Local Ai

Flama: a Python framework for development and deployment of production-ready APIs, machine learning, and LLM services

DGX agent

arXiv:2608.18733v1 Announce Type: cross Abstract: We present Flama, an open-source Python framework for developing and deploying production-ready web APIs, machine learning services, and large-languag

local-aiarxiv-cs-ai
20 Aug 2026
Model Releases

FrenchNews-7: Benchmarking Cross-Publisher French News Editorial Desk Classification

DGX agent

arXiv:2608.18097v1 Announce Type: new Abstract: We present FrenchNews-7, a cross-publisher France-based French-language news editorial desk classification benchmark combining a large multi-outlet corp

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

How Quantum Is the Advantage? A Fair, Calibration- and Noise-Aware Benchmark and Attribution Audit of Quantum Machine Learning for Network Intrusion Detection

DGX agent

arXiv:2608.18155v1 Announce Type: cross Abstract: Quantum machine learning (QML) for network intrusion detection (NIDS) is routinely reported to reach near-perfect accuracy, yet the most rigorous stud

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

JSL-DC: A Word-Level Japanese Sign Language Dataset with Linguist-Derived Descriptions for Distinguishing Confusable Signs

DGX agent

arXiv:2608.18412v1 Announce Type: new Abstract: Effective sign language (SL) acquisition is crucial for deaf children, yet 95% are born to hearing parents who often lack proficiency in SL. SL recognit

model-releasesarxiv-cs-cv
20 Aug 2026
Model Releases

Ling-3.0 released all 6 base checkpoints: 2 sizes × 3 stages

DGX agent

AntLing has released the full six-checkpoint matrix for the Ling-3.0 base model. tiny: pretrained, mid-trained, WSM-merged flash: pretrained, mid-trained, WSM-merged The concrete artifact is six separ

model-releasesr-localllama
20 Aug 2026
Model Releases

NanoSleep: A Parameter-Efficient Hybrid Temporal Convolutional Network for Single-Channel Sleep Stage Classification

DGX agent

arXiv:2608.18571v1 Announce Type: new Abstract: Sleep stage classification from single-channel electroencephalography (EEG) is essential for wearable and home-based sleep monitoring. However, many dee

model-releasesarxiv-cs-lg
20 Aug 2026
Model Releases

Need help choosing the right AI model/tool for a complete web app workflow

DGX agent

I currently have these models available through Ollama/cloud: And I’m using Claude, Codex, OpenCode, and Ollama as my coding/agent tools. I want to learn professional vibecoding — not just asking AI t

model-releasesr-ollama
20 Aug 2026
Model Releases

Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities

DGX agent

arXiv:2608.18090v1 Announce Type: cross Abstract: Inside a modern language model sits a single internal direction that tracks how positive or negative a sentence feels. We show how to find this valenc

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Pedagogical AI in Mental Health: A Tri-Stream Fine-Tuned LLM Framework for Automated Clinical Supervision and Risk Triage

DGX agent

arXiv:2608.18438v1 Announce Type: cross Abstract: Modern mental healthcare faces a critical shortage of senior supervisory oversight, leading to a 'supervision gap' where novice therapists manage high

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Qwen3.8-27B scored 29/30 on AIME 2026 with FP8 + xhigh reasoning — BF16 vs FP8 results

DGX agent

I benchmarked Qwen3.8-27B on MathArena/aime_2026 dataset, comparing BF16 and FP8 weights at medium and xhigh reasoning effort. Interesting findings are: quantized FP8 xhigh is better than BF 16 medium

model-releasesr-localllama
20 Aug 2026
Safety

Rethinking Privileged Information in On-Policy Self-Distillation

DGX agent

arXiv:2608.18271v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own responses using token-level supervision from the same model conditioned on privileged ref

safetyarxiv-cs-lg
20 Aug 2026
Model Releases

Safe Domain Adaptation for Physics: Overcoming Nuisances, Label Shifts, and Simulation Priors

DGX agent

arXiv:2608.18190v1 Announce Type: new Abstract: Domain adaptation is widely used to make neural networks trained on simulations applicable to experimental data. Its premise is that the two domains dif

model-releasesarxiv-cs-lg
20 Aug 2026
Model Releases

Selection, Recombination, or a Fresh Solve? A Candidate-Free Control for Single-Pass Test-Time Aggregation

DGX agent

arXiv:2608.18379v1 Announce Type: cross Abstract: When every candidate is wrong, correct-candidate selection is unavailable, yet the aggregation call can still solve the problem afresh. A correct aggr

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Stability-Aware Feature Design for Robust Watermark Detection in Machine-Generated Text

DGX agent

arXiv:2608.18102v1 Announce Type: new Abstract: The widespread adoption of large language models (LLMs) has intensified the demand for principled methods to distinguish human from machine-generated te

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

Web agent economics are set by inference volume and each step is one inference call. A simple task (extract a field, fill a form, navigate a…

DGX agent

Web agent economics are set by inference volume and each step is one inference call. A simple task (extract a field, fill a form, navigate a site) is 10 to 20 calls. A multi-site workflow runs into th

model-releasestogether-ai--x
20 Aug 2026
Model Releases

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategi…

DGX agent

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategies. It achieves state-of-the-art performance among open-sourc

model-releasesclem-delangue--x
19 Aug 2026
Model Releases

Am I doing something wrong? Qwen 3.8 27B seems useless for agentic coding

DGX agent

I have been using local models on/off for like 2 years or so but never really used them extensively because the closed ones were always much better. Once Qwen 3.8 27B was released I decided to give it

model-releasesr-localllama
19 Aug 2026
Safety

Beyond the Trace: Coupling an Interpretable Reasoning-State Readout to Native MoE Routing

DGX agent

arXiv:2608.17638v1 Announce Type: new Abstract: What a reasoning model writes is only a partial record of the process that produces it. We introduce a two-level internal readout for mixture-of-experts

safetyarxiv-cs-ai
19 Aug 2026
Model Releases

I am so tired of the PR.

DGX agent

I am so tired of the PR. How Anthropic's new results post would read without the PR: Claude orchestrated open-source protein design models, PXDesign, RFdiffusion, Genie, BoltzGen, from a 30k-token exp

model-releasesgary-marcus--x
19 Aug 2026
Model Releases

I pushed Qwen3.8-27B limits again... Dflash2 - 134 tps on a RTX 3090

DGX agent

Edit: Title says 134 tps, it's actually 138 -- keep in mind my 3090 is power limited to 250w. Three days ago I released a hyper-optimized Qwen3.8-27B inference engine for an RTX 3090 (82 tps single re

model-releasesr-localllama
19 Aug 2026
← Previous
1…431432433434435…1371
Next →