AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlog
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,672 results
Model Releases

YolovN-CBi: A Lightweight and Efficient Architecture for Real-Time Detection of Small UAVs

DGX agent

arXiv:2512.18046v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles, commonly known as, drones pose increasing risks in civilian and defense settings, demanding accurate and real-time drone d

model-releasesarxiv-cs-cv
21 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

3 days benchmarking most llama.cpp flags on my weird 40gb vram laptop + tb4 egpu setup. Got +70% generation, +40% prefill, 60k more context, and filed a bug in llama around MTP. What I learned.

DGX agent

tldr: went from 16~ t/s to 27~ t/s generation. got my usable context up from 220k to the full 262k without sacrificing anything. prefill also increased from 376 to 573 command I ended up with, fwiw: l

model-releasesr-localllama
20 Aug 2026
Model Releases

A Few Cases Are All You Need: An Empirical Study of Annotation-Efficient LoRA Fine-Tuning of MedSAM3

DGX agent

arXiv:2608.18731v1 Announce Type: cross Abstract: Medical image segmentation is essential for clinical workflows such as treatment planning and disease assessment. While specialist tools like TotalSeg

model-releasesarxiv-cs-ai
20 Aug 2026
Safety

Accelerating Visual On-Policy Distillation with Batched Speculative Jacobi Rollouts

DGX agent

arXiv:2608.18183v1 Announce Type: new Abstract: Visual on-policy distillation (OPD) improves the training of compact visual autoregressive models by learning from trajectories generated by the current

safetyarxiv-cs-lg
20 Aug 2026
Applications

Assessing Quality of Experience in Natural Language Generation of German Text

DGX agent

arXiv:2608.18888v1 Announce Type: new Abstract: The rapid advancement of Natural Language Generation (NLG) has made the reliable evaluation of generated text increasingly critical, as these systems, s

applicationsarxiv-cs-cl
20 Aug 2026
Research

BERTilda: Explainable Topic Lifecycle Tracking with Split/Merge Detection via Similarity-and-Flow Temporal Graphs

DGX agent

arXiv:2608.18101v1 Announce Type: new Abstract: Longitudinal text streams exhibit topic birth and death, but also discrete structural reorganizations in which themes split into subtopics or merge into

researcharxiv-cs-cl
20 Aug 2026
Model Releases

CausalProfiler: Generating Synthetic Benchmarks for Rigorous and Transparent Evaluation of Causal Machine Learning

DGX agent

arXiv:2511.22842v3 Announce Type: replace-cross Abstract: Causal machine learning (Causal ML) aims to answer 'what if' questions using machine learning algorithms, making it a promising tool for high-

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents

DGX agent

arXiv:2608.18307v1 Announce Type: new Abstract: Current evaluation of computer-use agents is split between long-horizon workflow benchmarks and atomic GUI-grounding tests. This leaves an under-instrum

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence

DGX agent

arXiv:2608.18613v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) is increasingly consumed not by human analysts but by LLM agents that compose multi-step investigations at query time. T

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Event-Causal RAG: A Retrieval-Augmented Generation Framework for Long Video Reasoning in Complex Scenarios

DGX agent

arXiv:2605.06185v2 Announce Type: replace Abstract: Large vision-language models perform well on short- and medium-length video understanding but still struggle to maintain coherent event memory and r

model-releasesarxiv-cs-ai
20 Aug 2026
Safety

extsc{TestifAI}: Tomography-Based Testing for Deep Learning Systems

DGX agent

arXiv:2608.18900v1 Announce Type: new Abstract: As AI systems are increasingly deployed in safety-critical application domains (e.g., autonomous driving), associated risks increase too. Deep learning

safetyarxiv-cs-ai
20 Aug 2026
Model Releases

Fine-tuning Cactus Needle 2 can match DeepSeek v4 on the specific task

DGX agent

Hey LocalLlama, Henry from Cactus here! When we trained Needle 2, I had a strict rule to not expose the model to any data sample that remotely felt like these benchmarks. It seemed over-the-top, but b

model-releasesr-localllama
20 Aug 2026
Local Ai

Flama: a Python framework for development and deployment of production-ready APIs, machine learning, and LLM services

DGX agent

arXiv:2608.18733v1 Announce Type: cross Abstract: We present Flama, an open-source Python framework for developing and deploying production-ready web APIs, machine learning services, and large-languag

local-aiarxiv-cs-ai
20 Aug 2026
Model Releases

FrenchNews-7: Benchmarking Cross-Publisher French News Editorial Desk Classification

DGX agent

arXiv:2608.18097v1 Announce Type: new Abstract: We present FrenchNews-7, a cross-publisher France-based French-language news editorial desk classification benchmark combining a large multi-outlet corp

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

How Quantum Is the Advantage? A Fair, Calibration- and Noise-Aware Benchmark and Attribution Audit of Quantum Machine Learning for Network Intrusion Detection

DGX agent

arXiv:2608.18155v1 Announce Type: cross Abstract: Quantum machine learning (QML) for network intrusion detection (NIDS) is routinely reported to reach near-perfect accuracy, yet the most rigorous stud

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

JSL-DC: A Word-Level Japanese Sign Language Dataset with Linguist-Derived Descriptions for Distinguishing Confusable Signs

DGX agent

arXiv:2608.18412v1 Announce Type: new Abstract: Effective sign language (SL) acquisition is crucial for deaf children, yet 95% are born to hearing parents who often lack proficiency in SL. SL recognit

model-releasesarxiv-cs-cv
20 Aug 2026
Model Releases

Ling-3.0 released all 6 base checkpoints: 2 sizes × 3 stages

DGX agent

AntLing has released the full six-checkpoint matrix for the Ling-3.0 base model. tiny: pretrained, mid-trained, WSM-merged flash: pretrained, mid-trained, WSM-merged The concrete artifact is six separ

model-releasesr-localllama
20 Aug 2026
Model Releases

NanoSleep: A Parameter-Efficient Hybrid Temporal Convolutional Network for Single-Channel Sleep Stage Classification

DGX agent

arXiv:2608.18571v1 Announce Type: new Abstract: Sleep stage classification from single-channel electroencephalography (EEG) is essential for wearable and home-based sleep monitoring. However, many dee

model-releasesarxiv-cs-lg
20 Aug 2026
Model Releases

Need help choosing the right AI model/tool for a complete web app workflow

DGX agent

I currently have these models available through Ollama/cloud: And I’m using Claude, Codex, OpenCode, and Ollama as my coding/agent tools. I want to learn professional vibecoding — not just asking AI t

model-releasesr-ollama
20 Aug 2026
Model Releases

Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities

DGX agent

arXiv:2608.18090v1 Announce Type: cross Abstract: Inside a modern language model sits a single internal direction that tracks how positive or negative a sentence feels. We show how to find this valenc

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Pedagogical AI in Mental Health: A Tri-Stream Fine-Tuned LLM Framework for Automated Clinical Supervision and Risk Triage

DGX agent

arXiv:2608.18438v1 Announce Type: cross Abstract: Modern mental healthcare faces a critical shortage of senior supervisory oversight, leading to a 'supervision gap' where novice therapists manage high

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Qwen3.8-27B scored 29/30 on AIME 2026 with FP8 + xhigh reasoning — BF16 vs FP8 results

DGX agent

I benchmarked Qwen3.8-27B on MathArena/aime_2026 dataset, comparing BF16 and FP8 weights at medium and xhigh reasoning effort. Interesting findings are: quantized FP8 xhigh is better than BF 16 medium

model-releasesr-localllama
20 Aug 2026
Safety

Rethinking Privileged Information in On-Policy Self-Distillation

DGX agent

arXiv:2608.18271v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own responses using token-level supervision from the same model conditioned on privileged ref

safetyarxiv-cs-lg
20 Aug 2026
Model Releases

Safe Domain Adaptation for Physics: Overcoming Nuisances, Label Shifts, and Simulation Priors

DGX agent

arXiv:2608.18190v1 Announce Type: new Abstract: Domain adaptation is widely used to make neural networks trained on simulations applicable to experimental data. Its premise is that the two domains dif

model-releasesarxiv-cs-lg
20 Aug 2026
Model Releases

Selection, Recombination, or a Fresh Solve? A Candidate-Free Control for Single-Pass Test-Time Aggregation

DGX agent

arXiv:2608.18379v1 Announce Type: cross Abstract: When every candidate is wrong, correct-candidate selection is unavailable, yet the aggregation call can still solve the problem afresh. A correct aggr

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Stability-Aware Feature Design for Robust Watermark Detection in Machine-Generated Text

DGX agent

arXiv:2608.18102v1 Announce Type: new Abstract: The widespread adoption of large language models (LLMs) has intensified the demand for principled methods to distinguish human from machine-generated te

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

Web agent economics are set by inference volume and each step is one inference call. A simple task (extract a field, fill a form, navigate a…

DGX agent

Web agent economics are set by inference volume and each step is one inference call. A simple task (extract a field, fill a form, navigate a site) is 10 to 20 calls. A multi-site workflow runs into th

model-releasestogether-ai--x
20 Aug 2026
Model Releases

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategi…

DGX agent

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategies. It achieves state-of-the-art performance among open-sourc

model-releasesclem-delangue--x
19 Aug 2026
Model Releases

Am I doing something wrong? Qwen 3.8 27B seems useless for agentic coding

DGX agent

I have been using local models on/off for like 2 years or so but never really used them extensively because the closed ones were always much better. Once Qwen 3.8 27B was released I decided to give it

model-releasesr-localllama
19 Aug 2026
Safety

Beyond the Trace: Coupling an Interpretable Reasoning-State Readout to Native MoE Routing

DGX agent

arXiv:2608.17638v1 Announce Type: new Abstract: What a reasoning model writes is only a partial record of the process that produces it. We introduce a two-level internal readout for mixture-of-experts

safetyarxiv-cs-ai
19 Aug 2026
Model Releases

I am so tired of the PR.

DGX agent

I am so tired of the PR. How Anthropic's new results post would read without the PR: Claude orchestrated open-source protein design models, PXDesign, RFdiffusion, Genie, BoltzGen, from a 30k-token exp

model-releasesgary-marcus--x
19 Aug 2026
Model Releases

I pushed Qwen3.8-27B limits again... Dflash2 - 134 tps on a RTX 3090

DGX agent

Edit: Title says 134 tps, it's actually 138 -- keep in mind my 3090 is power limited to 250w. Three days ago I released a hyper-optimized Qwen3.8-27B inference engine for an RTX 3090 (82 tps single re

model-releasesr-localllama
19 Aug 2026
Model Releases

Improving Complex Moire Removal with Generative Supervision

DGX agent

arXiv:2608.17883v1 Announce Type: new Abstract: The availability of high-quality paired data is essential for training learning-based image demoireing models. However, it remains challenging for exist

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

Key-Frame Reasoning with SAM3: Third Place Solution for the MeViS-Text Track of the 8th LSVOS Challenge

DGX agent

arXiv:2608.17279v1 Announce Type: new Abstract: This report presents a two-stage, training-free solution for the MeViS-Text track of the 8th LSVOS Challenge. The task requires a model to localize and

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

Leveraging existing sparse point annotations for benthic imagery dense segmentation

DGX agent

arXiv:2608.17561v1 Announce Type: new Abstract: The health of marine ecosystems is a critical indicator of global environmental change, yet the physical constraints of underwater observation and the i

model-releasesarxiv-cs-cv
19 Aug 2026
Safety

Likelihood Hacking in Probabilistic Program Synthesis

DGX agent

arXiv:2603.24126v2 Announce Type: replace Abstract: When language models are trained by reinforcement learning (RL) to write probabilistic programs, they can artificially inflate their marginal-likeli

safetyarxiv-cs-lg
19 Aug 2026
Model Releases

MANIGUARD: A Benchmark and Data Suite for Specification-Grounded Safety Evaluation and Improvement of Robotic Manipulation

DGX agent

arXiv:2608.17386v1 Announce Type: new Abstract: Foundation-model policies for robotic manipulation are advancing rapidly on task success, but rigorous evaluation of whether they succeed safely is stil

model-releasesarxiv-cs-ro
19 Aug 2026
Local Ai

Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals

DGX agent

arXiv:2608.17687v1 Announce Type: new Abstract: Despite their widespread use, Large Language Models (LLMs) remain limited by a fundamental problem: the generation of plausible but false content, known

local-aiarxiv-cs-ai
19 Aug 2026
Applications

Physics-Informed and Hybrid Machine Learning in Additive Manufacturing: Application to Fused Filament Fabrication

DGX agent

arXiv:2608.17246v1 Announce Type: new Abstract: This article investigates several physics-informed and hybrid machine learning strategies that incorporate physics knowledge in experimental data-driven

applicationsarxiv-cs-lg
19 Aug 2026
Safety

Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents

DGX agent

arXiv:2608.18008v1 Announce Type: cross Abstract: Combining large language models with reinforcement learning is increasingly explored, yet the theoretical status of LLM-derived reward signals is ofte

safetyarxiv-cs-ai
19 Aug 2026
Model Releases

Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations

DGX agent

arXiv:2608.16970v1 Announce Type: cross Abstract: LLM-based code generation is now embedded in mission-critical pipelines, but defenses against vulnerable output remain post-hoc -- static analyzers, f

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX

DGX agent

arXiv:2608.17379v1 Announce Type: cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to use architecture-specific PTX for GPU kernel optimizati

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

Q-Interference: Memory-Efficient Phase-Aware Quantum-Inspired Attention

DGX agent

arXiv:2608.17288v1 Announce Type: new Abstract: GPT attention measures token compatibility through dot-product similarity. This mechanism is simple, effective, and memory-efficient. But it does not ex

model-releasesarxiv-cs-cl
19 Aug 2026
Model Releases

S^3AM: A Single-Stream SAM with Reliability-Calibrated Frequency Adapter for Multi-modal Salient Object Detection

DGX agent

arXiv:2608.17475v1 Announce Type: new Abstract: Vision foundation models have recently advanced multi-modal salient object detection (MSOD) through parameter-efficient tuning and prompt learning. Howe

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

SCENARIODIFF: A Scenario-level Guidance Framework for Multimodal Time Series Forecasting--Extended Version

DGX agent

arXiv:2608.17164v1 Announce Type: new Abstract: Textual context such as news, reports, and logs can provide valuable signals for time series forecasting, especially when future dynamics are driven by

model-releasesarxiv-cs-lg
19 Aug 2026
Safety

Teach and Grow: An Agent-Centered Architecture for General Robot Learning

DGX agent

arXiv:2608.17209v1 Announce Type: cross Abstract: End-to-end vision-language-action (VLA) and world-action models offer an elegant route to general-purpose robotics, but their reliability is bounded b

safetyarxiv-cs-ai
19 Aug 2026
Research

Towards Safer RAG: Only Agents Capable of System 2 Thinking may Access Untrusted Documents

DGX agent

arXiv:2608.17153v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has significantly enhanced the performance of large language models (LLMs), yet these systems remain vulnerable to

researcharxiv-cs-cl
19 Aug 2026
Model Releases

A Large-Scale Chinese Knowledge Graph-Text Alignment Dataset for Benchmarking Knowledge-Grounded LLMs

DGX agent

arXiv:2510.06039v2 Announce Type: replace-cross Abstract: Reliable evaluation of knowledge-grounded Large Language Models (LLMs) in Chinese requires resources that explicitly align Chinese-language te

model-releasesarxiv-cs-ai
18 Aug 2026
← Previous
1…415416417418419…1327
Next →