AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
18 Aug 2026

This 1 hour workshop on context engineering will help you build agents that hold long conversations and cut your token cost by up to 90%. La…

Model ReleasesDGX agent

This 1 hour workshop on context engineering will help you build agents that hold long conversations and cut your token cost by up to 90%. Last month we gave it live at the AI Engineer World's Fair 202

TinyCast: Probabilistic Zero-Shot Forecasting with Computed Periodicity

Model ReleasesDGX agent

arXiv:2608.15767v1 Announce Type: cross Abstract: We introduce TinyCast, an attention-free zero-shot forecaster that emits a predictive distribution from 146,505 parameters, on the premise that at thi

Unified Pedestrian Path Prediction Using Inverse Reinforcement Learning

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.15929v1 Announce Type: new Abstract: Pedestrian path prediction is crucial for enhancing the safety of autonomous vehicles and advanced driver-assistance systems. Previous studies explored

VARM-Bench: Benchmarking Verifiable Structured Reasoning in Chinese Abusive Speech Moderation

Model ReleasesDGX agent

arXiv:2608.15600v1 Announce Type: new Abstract: The widespread circulation of abusive online content has increased the need for reliable moderation of Chinese social-media text. Existing Chinese bench

VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End?

Model ReleasesDGX agent

arXiv:2608.15265v1 Announce Type: new Abstract: Constructing an interactive 3D open world from a user query is important. However, existing methods are primarily evaluated on idealized, simple queries

What Does Context Compression Cost an Agent? Interaction Costs Unrevealed by Task-Completion Metrics

Model ReleasesDGX agent

arXiv:2608.16370v1 Announce Type: new Abstract: Task completion is the standard metric for evaluating context compression, yet it is incomplete: compression can increase an agent's interaction cost by

What's the best tool for offline Wikipedia RAG at the moment?

Model ReleasesDGX agent

Theoretical use case: I have a compressed version of wikipedia, and some function or tool fooRAG. Then I set my system prompt (or make a skill, or hook, etc.) to include 'You can search wikipedia with

When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification

Model ReleasesDGX agent

arXiv:2608.16147v1 Announce Type: new Abstract: Class-imbalance handling is routinely evaluated on a single benchmark dataset, and the resulting conclusions are reported as if they were properties of

😮 wow that’s huge 💰

Model ReleasesDGX agent

😮 wow that’s huge 💰 🚀Today we launched LangSmith Tuned Evaluators, starting with Perceived Error. Tuned Evaluators run on production traces to catch undesirable agent behavior and attach feedback that

YOLO26-RD: An End-to-End Road Damage Detection Network With Learnable Contrast Enhancement and Edge-Guided Downsampling

Model ReleasesDGX agent

arXiv:2608.15713v1 Announce Type: new Abstract: Automated pavement-distress detection is commonly framed as a small-object problem, motivating high-resolution P2/4 detection heads and lossless downsam

17 Aug 2026

100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s

Model ReleasesDGX agent

Qwen 27b Q3_K_M 2x rx 580 8gb (~50 each in my country, edge cases 60 per gpu) gives us 16gb vram We used it on an old already existing ddr3 motherboard with 2 gpu slots(you can buy it ror around 200$

A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation

Model ReleasesDGX agent

arXiv:2608.14329v1 Announce Type: cross Abstract: Principle-based regulation, with evaluative standards such as 'fair, clear, and not misleading' or 'deliver good outcomes', cannot be reduced to binar

Adaptive Protection for Evolutionary Feature Construction in Symbolic Regression with Application to Credit Classification

Model ReleasesDGX agent

arXiv:2608.14209v1 Announce Type: new Abstract: Evolutionary feature construction has shown strong promise in symbolic regression by automatically discovering informative transformations of input feat

Conditional Neural Optimal Transport for Predicting Cellular Phenotypes from Molecular Structure

ResearchDGX agent

arXiv:2608.14293v1 Announce Type: new Abstract: High-content microscopy enables systematic profiling of cellular responses to chemical perturbations, but the scale of the chemical space makes exhausti

Consensus-gated Multi-Agent Neural Architecture Search for Seismic Fault Segmentation

Model ReleasesDGX agent

arXiv:2608.13889v1 Announce Type: new Abstract: Neural networks for seismic fault segmentation are often borrowed from computer vision and medical imaging domains where they train under relatively muc

DeepSeek V4 Flash with Antirez Dwarfstar 4 is amazing.

Model ReleasesDGX agent

Note: I use a Mac Studio M3U with 512 GB RAM, so this is not for everyone. I have been using antirez/ds4 with DS4 Flash for a few weeks now. Top quality, I am really impressed. This combination just d

Dynamic Multi-Depot Vehicle Routing with Online Requests: Event-Driven Transformer--DRL and Rolling-Horizon Benchmarking

Model ReleasesDGX agent

arXiv:2608.13799v1 Announce Type: new Abstract: This paper presents an event-driven learning and benchmarking framework for the Dynamic Multi-Depot Vehicle Routing Problem with progressively revealed

From Fixed Grids to Moving Particles:A Transferable Latent Operator for Fluid Dynamics

ResearchDGX agent

arXiv:2608.14120v1 Announce Type: cross Abstract: Lagrangian modeling is vital to fluid dynamics, as it characterizes particle transport and complements the Eulerian description.However, Lagrangian tr

Grok is very good at agentic tasks!

Model ReleasesDGX agent

Grok is very good at agentic tasks! Agentic work is where @grok 4.6 lands hardest, taking the top spot on the Artificial Analysis Agentic Index at 59, tied with Claude Opus 5 Max. The index measures t

HERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geoscience

Model ReleasesDGX agent

arXiv:2608.14055v1 Announce Type: new Abstract: Authoritative scientific knowledge in geoscience remains largely trapped in legacy monographs and historical literature, where unstructured text and com

HI-MeshGraphNets: Efficient and Accurate Mesh-based Physics Learning with Hierarchical Multi-scale Graph Neural Networks

ResearchDGX agent

arXiv:2608.13827v1 Announce Type: new Abstract: Machine-learned physical surrogate models have become promising alternatives to mesh-based numerical solvers. Among them, graph neural networks (GNNs) a

Interactive Analysis of Global Explanations using Aggregated Class Activation Maps for Network Data

ResearchDGX agent

arXiv:2608.13575v1 Announce Type: cross Abstract: Recent machine learning (ML) advances have demonstrated that deep learning (DL) achieves impressive results in different application domains, includin

MedClaw: Heuristic Agent Harness for Long-Horizon Surgical Video Reasoning

Model ReleasesDGX agent

arXiv:2608.14015v1 Announce Type: cross Abstract: Understanding tens-of-minutes surgical videos requires long-horizon temporal reasoning, answering what happens before, after, or across stages of a pr

More Correct Mass, Worse Answers: Why Power Sampling Can Fail and How to Fix It

TutorialsDGX agent

arXiv:2608.14420v1 Announce Type: new Abstract: Power Sampling sharpens a language model's distribution over complete generation trajectories, offering a verifier-free way to improve reasoning at infe

Ordinal-Aware Calibration for Ordinal Classification

Model ReleasesDGX agent

arXiv:2410.15658v4 Announce Type: replace-cross Abstract: Deep neural networks frequently produce overconfident, miscalibrated predictions. In ordinal classification, predictions must also adhere to a

pagedMark: invisible SynthID-class watermark removal for AI images (ChatGPT, gpt-image, DALL·E, Sora, Gemini, Nano Banana), running on Metal

Model ReleasesDGX agent

pagedMark removes AI provenance from content you generated yourself. Two different things, and it is worth separating them. The first is metadata: C2PA Content Credentials, EXIF, XMP, IPTC, the genera

ProFocus: Interpreting Affective Experience in Artistic Images with Progressive Visual Focusing

ResearchDGX agent

arXiv:2608.13974v1 Announce Type: new Abstract: Interpreting the emotional responses triggered by images is central to achieving emotional intelligence. Compared with natural images, visual art is int

Qwen 3.8 27b in 24gb of VRAM

Model ReleasesDGX agent

Thought i would just add my own flags here for llama.cpp (literally pulled and rebuilt latest today). Running on a 4090 FE and 48gb of DDR4 ram on WSL2. Basically you have 3 knobs you can tune for max

Recent Advances in Deep Learning-Based Drug-Target Binding Affinity Prediction

Model ReleasesDGX agent

arXiv:2608.13797v1 Announce Type: new Abstract: Computational approaches to drug discovery involve multiple sub-problems, and among them, drug-target binding affinity prediction plays an important rol

Rethinking Auxiliary Modalities in Multimodal Zero-shot Anomaly Detection: From Semantic Fusion to Conditional Modulation

ResearchDGX agent

arXiv:2608.13973v1 Announce Type: new Abstract: Recent foundation model-based methods have endowed RGB images with strong zero-shot anomaly detection (ZSAD) through vision-language pretraining. Howeve

Self-Supervised Visual On-Policy Distillation

Model ReleasesDGX agent

arXiv:2608.14144v1 Announce Type: cross Abstract: Visual on-policy distillation relies heavily on an informative teacher-student asymmetry, through either a larger, stronger teacher or privileged supe

Spatial Message Passing in Language Space for Pathology Image Interpretation

Local AiDGX agent

arXiv:2608.14309v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can generate pathological descriptions from histological images, but gigapixel Whole Slide Images (WSIs) exceed

Structure-Guided Spatiotemporal Attention Graph Neural Network for Traffic Flow Prediction

SafetyDGX agent

arXiv:2608.14177v1 Announce Type: cross Abstract: Deep spatiotemporal models integrating graph convolutions and attention mechanisms have demonstrated excellent performance in network-level traffic fl

TimeSage-EV: A Live Benchmark for Agentic Time Series Analysis in Evolving Environments

Model ReleasesDGX agent

arXiv:2608.14270v1 Announce Type: new Abstract: Time series analysis in high-stakes domains relies on recurring data releases, where new observations can alter the evidence base and the validity of la

Tripwire: Triggering Aligned Refusal via Statistically Certified Safety Neurons

SafetyDGX agent

arXiv:2608.14392v1 Announce Type: new Abstract: Neuron- and path-level interventions offer the finest-grained route to defending large language models (LLMs) against jailbreak attacks, yet existing me

Universal Thermodynamic Interatomic Potentials for Crystalline Materials

ResearchDGX agent

arXiv:2608.14502v1 Announce Type: cross Abstract: Free energies govern solid-state phase stability, yet computational materials discovery still relies largely on ground-state energies because free ene

16 Aug 2026

5090: Windows or Linux for Qwen3.8.27b

Model ReleasesDGX agent

I've got a dedicated AI rig sitting here with a RTX 5090 and 96GB RAM and for the past few years have been using Windows 11 and primarily LM Studio, but have also used vLLM, llama.cpp and Ollama. With

Is ternary (1.58-bit) LLMs making a come back?

AgentsDGX agent

I'm just thinking, ever since microsoft announced bitnet, this sub (and myself) has been hoping for massive ternary models. In the last month alone, prismML dropped 27B ternary (though I've read commu

Qwen 3.8 27b with DSH(DeepSeek Harness) is Amazing!! Experiences so far and perfomance.

Model ReleasesDGX agent

https://preview.redd.it/wkg27e152qjh1.png?width=853&format=png&auto=webp&s=2e3f8b11ea6393041f501e95c5835f9bea0245dd So ive been trying different harnesses and coding agents with the new qwen 3.8 , and

15 Aug 2026

A short illustration of how the Claude's watermarking is supposed to work (based on my read of their released materials). In general, when w…

Model ReleasesDGX agent

A short illustration of how the Claude's watermarking is supposed to work (based on my read of their released materials). In general, when we are generating tokens, there can be multiple high-scoring

Building an AI Text Detector From Scratch

ResearchDGX agent

The article describes an end‑to‑end DIY project that builds an AI text detector from scratch, including dataset construction, model training, local deployment, and a reinforcement learning verifier (R

The 27B FP8 uncensored weights are ~31GB, so 'just run it on a 24GB card' doesn't hold up

Model ReleasesDGX agent

The uncensored FP8 build that showed up on HuggingFace this week is about 31GB in block-FP8, roughly half the BF16 weights. People keep reading '27B' and assuming it drops onto a 3060 or any single 24

14 Aug 2026

A Controlled Study of Self-Supervised Image and Video Pretraining under Limited Resources

ResearchDGX agent

arXiv:2608.13183v1 Announce Type: new Abstract: Visual foundation models are a cornerstone of image and video understanding but typically require large amounts of data and computation. The current sca

Beyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2608.13326v1 Announce Type: new Abstract: LLM benchmark scores can be precise even when the observation protocol does not identify the behavioral property they are intended to measure. In a cont

Beyond the Best Guess: Improving LLM Solution Coverage with Evolution Strategies

ResearchDGX agent

arXiv:2608.12679v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in discovery domains such as math and science. The usual approach is to present the problem to th

CAKE: Compiler-Agent Co-Design for Frontier Kernel Evolution

Model ReleasesDGX agent

arXiv:2608.12629v1 Announce Type: new Abstract: GPU kernel agents and GPU programming languages have advanced separately, leaving expert kernels difficult to reproduce. Agents usually treat the compil

Confidence-Calibrating Regularization for Robust Brain MRI Segmentation Under Domain Shift

ResearchDGX agent

arXiv:2509.23176v2 Announce Type: replace Abstract: The Segment Anything Model (SAM) exhibits strong zero-shot performance on natural images but suffers from domain shift and overconfidence when appli

CoverPrune: Coverage-Driven Token Pruning for 3D VLMs via Optimal Transport

ResearchDGX agent

arXiv:2608.13226v1 Announce Type: cross Abstract: While 3D Vision-Language Models (3D VLMs) have demonstrated remarkable spatial reasoning capabilities, they suffer from massive visual token counts th

DiG-bench: Discovery in Games

Model ReleasesDGX agent

arXiv:2608.12593v1 Announce Type: new Abstract: Discovery---formulating novel generalizations---is a central part of the scientific process. Despite its importance, there is a gap in the current AI be

Doctorina MedBench: A Dialogue-Based Benchmark and Evaluation Framework for Agent-Based Medical AI

Model ReleasesDGX agent

arXiv:2603.25821v3 Announce Type: replace-cross Abstract: We present Doctorina MedBench, an evaluation framework for agent-based medical AI based on the simulation of physician-patient interactions. U

FastThaiG2P: Lightning-fast Thai Grapheme-to-phoneme Conversion for Voice Agent Pipelines

Model ReleasesDGX agent

arXiv:2608.12814v1 Announce Type: new Abstract: FastThaiG2P provides sub-millisecond Thai grapheme-to-phoneme conversion for text-to-speech pipelines (International Phonetic Alphabet and Kokoro-TTS co

From Visual Widgets to UI Code: Efficient Tool-Grounded Generation

Model ReleasesDGX agent

arXiv:2608.12611v1 Announce Type: new Abstract: Existing screenshot-to-code systems face a trade-off between flexibility and controllability. Direct multimodal generation can hallucinate visible detai

Generalizable Operating Room Expert with Multimodal Enhancement

TutorialsDGX agent

arXiv:2508.08199v2 Announce Type: replace Abstract: Precise spatial modeling in the operating room (OR) is essential for intraoperative awareness, hazard avoidance, and surgical decision-making. Altho

If you hand-tune agent harnesses, this one is worth your time. AutoDesign puts the harness itself inside the optimization loop. A meta-harne…

Model ReleasesDGX agent

If you hand-tune agent harnesses, this one is worth your time. AutoDesign puts the harness itself inside the optimization loop. A meta-harness optimizer reads rollout feedback and directs a code agent

Interpretable Causal Discovery via Causal-Effect Constraints

Model ReleasesDGX agent

arXiv:2608.12640v1 Announce Type: cross Abstract: Causal discovery aims to uncover the underlying causal relationships given data generated from a system. The goal, however, is not merely to predict c

It's How You Ask: Gender-Associated Linguistic Bias in LLMs

SafetyDGX agent

arXiv:2608.13328v1 Announce Type: cross Abstract: Professional communication is increasingly mediated by LLMs - but do these models serve all users equally? We show that when prompts contain linguisti

Knowledge-guided Pattern Discovery via Coupled Tensor Factorizations

ResearchDGX agent

arXiv:2608.13234v1 Announce Type: new Abstract: In order to understand complex systems such as the human metabolome or human brain, different sensing technologies are used, generating complex data. Th

Learning Unified Video and Image Representation for Video Face Forgery Detection

Model ReleasesDGX agent

arXiv:2608.13064v1 Announce Type: new Abstract: Face forgery detection is crucial for preserving the security and integrity of facial data given the rapid developments in face manipulation techniques

LipCache: A Local Inference Proxy with Certified Caching for Edge Image Classification Service

Local AiDGX agent

arXiv:2608.13144v1 Announce Type: cross Abstract: As edge-side vision services continue to expand toward low-latency, high-throughput scenarios, reducing the inference cost of vision models without sa

Memorization Diagnostics for Code LLMs Should be Scale-Aware

AgentsDGX agent

arXiv:2608.12771v1 Announce Type: cross Abstract: The extent to which large language models for code rely on memorization over genuine understanding remains highly debated. While current literature fr

← Previous
1…438439440441442…1060
Next →