AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Model Releases

Approximate Muon with low-rank adapters

DGX agent

arXiv:2608.14492v1 Announce Type: new Abstract: The Muon optimizer shows clear benefits versus alternatives when pretraining neural networks. However, it is used less frequently for parameter-efficien

model-releasesarxiv-cs-lg
17 Aug 2026
Research

APTER: Adaptive Post-Training with Expert-Grounded Rubrics

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2608.14212v1 Announce Type: new Abstract: As large language models enter professional domains, they must satisfy domain constraints, include critical evidence, and provide complete reasoning rat

researcharxiv-cs-ai
17 Aug 2026
Model Releases

BCIJelly: An integrated ecosystem for brain-computer interface research

DGX agent

arXiv:2608.13576v1 Announce Type: cross Abstract: Brain-computer interface (BCI) research relies on multistage computational pipelines, yet progress remains constrained by fragmented data formats, het

model-releasesarxiv-cs-lg
17 Aug 2026
Model Releases

Built a token-aware gateway/load balancer for local LLM stacks — because nginx has no idea what a token costs

DGX agent

If you're running Ollama, llama.cpp, or vLLM behind nginx or HAProxy for more than a single user, you've probably hit this: nginx treats a 10-token prompt and a 10k-token prompt as identical 'one requ

model-releasesr-ollama
17 Aug 2026
Safety

BUZZY: Contrastive Scoring to Mitigate Text-Induced Bias in Multimodal Multiple-Choice QA

DGX agent

arXiv:2603.28026v2 Announce Type: replace Abstract: Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively measurable setting for evaluating vision-language model

safetyarxiv-cs-ai
17 Aug 2026
Research

Data-driven techniques for translational neuroscience and personalized neuro-health

DGX agent

arXiv:2608.13749v1 Announce Type: cross Abstract: Neurodegenexrative diseases such as Alzheimer's disease and Parkinson's disease are diagnosed most reliably only after substantial, often irreversible

researcharxiv-cs-ai
17 Aug 2026
Model Releases

DeepSeek V4 Pro 0813 vs Claude Fable 5 on DeepSWE: Cost, Coding, and Routing

DGX agent

DeepSeek V4 Pro 0813 and Claude Fable 5 were benchmarked on DeepSWE. A cascading strategy that starts with Pro 0813 and escalates to Fable only when it fails solves 82.7 % of tasks at an average cost

model-releasestogether-ai-blog
17 Aug 2026
Model Releases

defenders can see the future, and have a narrow window to uplevel their cybersecurity practices now. key is to uplevel fundamentals and appl…

DGX agent

defenders can see the future, and have a narrow window to uplevel their cybersecurity practices now. key is to uplevel fundamentals and apply the best AI tools. what we’re doing at OpenAI, and where o

model-releasesopenai--x
17 Aug 2026
Model Releases

Doomed to Re-Annotate, Forever: The ImageNet Story

DGX agent

arXiv:2608.13783v1 Announce Type: new Abstract: Top-1 accuracy on ImageNet-1k remains the most commonly reported metric in visual recognition. Quality issues with the dataset have been repeatedly repo

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Envs-FORGE: Frontier-Optimized Reward-Grounded Environment Synthesis for Agent RL

DGX agent

arXiv:2608.14312v1 Announce Type: new Abstract: Reinforcement learning (RL) for terminal agents needs executable training environments with reliable rewards and useful difficulty. Fixed recipes such a

model-releasesarxiv-cs-cl
17 Aug 2026
Model Releases

Estimating Dynamic Soft Continuum Robot States From Boundaries

DGX agent

arXiv:2505.04491v3 Announce Type: replace Abstract: State estimation is one of the fundamental problems in robotics. For soft continuum robots, this task is particularly challenging because their stat

model-releasesarxiv-cs-ro
17 Aug 2026
Safety

Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT

DGX agent

arXiv:2608.13681v1 Announce Type: cross Abstract: Translating C code into safe, idiomatic Rust is a longstanding software-engineering goal because it can eliminate entire classes of memory-safety vuln

safetyarxiv-cs-ai
17 Aug 2026
Research

FIRM: Fine-Grained Intra-Token Representation of Masks for Remote Sensing Reasoning Segmentation

DGX agent

arXiv:2608.13980v1 Announce Type: new Abstract: Reasoning segmentation requires multimodal large language models (MLLMs) to translate implicit instructions into precise pixel-level masks. MLLMs encode

researcharxiv-cs-cv
17 Aug 2026
Model Releases

Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation

DGX agent

arXiv:2608.14112v1 Announce Type: cross Abstract: Scientific simulations often produce scalar volumes faster than they can be stored, transferred, and loaded, while in situ reduction must use only a l

model-releasesarxiv-cs-ai
17 Aug 2026
Research

FLARE MCMC: Fidelity-based Layer-Adaptive REcursive proposals for MCMC

DGX agent

arXiv:2608.13774v1 Announce Type: new Abstract: Markov chain Monte Carlo (MCMC) requires only the ability to evaluate the likelihood, making it a common technique for inference in complex models. Howe

researcharxiv-cs-ai
17 Aug 2026
Research

FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction

DGX agent

arXiv:2608.14205v1 Announce Type: new Abstract: Load imbalance poses a major bottleneck to the efficiency of expert parallelism in distributed inference of Mixture-of-Experts (MoE) models. The most he

researcharxiv-cs-ai
17 Aug 2026
Safety

GALA: Generation-Aware Cross-Modal Alignment for Text-to-Time-Series Synthesis

DGX agent

arXiv:2608.13741v1 Announce Type: new Abstract: Synthesizing time series from natural language is emerging as the most expressive form of controllable time series generation. However, existing text-co

safetyarxiv-cs-cl
17 Aug 2026
Model Releases

GBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack Detection

DGX agent

arXiv:2608.14389v1 Announce Type: cross Abstract: Existing palm presentation attack detection (PAD) datasets are often limited by static imagery, restricted acquisition conditions, or insufficient mul

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Geometric Filtering of LLM-Generated Samples for Few-Shot Text Classification

DGX agent

arXiv:2608.13866v1 Announce Type: cross Abstract: Large language models (LLMs) can generate synthetic training data for text classification, but the quality of generated samples is heterogeneous: some

researcharxiv-cs-cl
17 Aug 2026
Tutorials

Global Interpretability via Automated Preprocessing: A Framework Inspired by Psychiatric Questionnaires

DGX agent

arXiv:2602.23459v2 Announce Type: replace Abstract: Psychiatric questionnaires are highly context sensitive and often only weakly predict subsequent symptom severity, which makes the prognostic relati

tutorialsarxiv-cs-lg
17 Aug 2026
Industry

Greg Brockman calls the OpenAI-Hugging Face incident 'a watershed moment' and discusses how OpenAI and other organizations can use AI to improve cyber defenses (Greg Brockman)

DGX agent

Greg Brockman: Greg Brockman calls the OpenAI-Hugging Face incident “a watershed moment” and discusses how OpenAI and other organizations can use AI to improve cyber defenses — The OpenAI-Hugging Face

industrytechmeme
17 Aug 2026
Model Releases

Grok 4.6 is smart, super fast & affordable

DGX agent

Grok 4.6 is smart, super fast & affordable Grok 4.6 just broke into the Top 3 on the Artificial Analysis Healthcare & Medical Index Grok 4.6 (high) scores 50 - just ONE point off the top score of 51 I

model-releaseselon-musk--x
17 Aug 2026
Model Releases

KV Cache Compression Through the Lens of Transform Coding

DGX agent

arXiv:2608.14191v1 Announce Type: cross Abstract: The key-value (KV) cache stores information from past tokens and is a major memory bottleneck in long-context inference. Existing quantization methods

model-releasesarxiv-cs-cl
17 Aug 2026
Applications

Limitations of Synthetic Data Generation in Specialized Data-Scarce Domains

DGX agent

arXiv:2608.13729v1 Announce Type: new Abstract: Advances in diffusion-based generative models have motivated the use of synthetic image generation to alleviate data scarcity in vision tasks. While thi

applicationsarxiv-cs-cv
17 Aug 2026
Model Releases

Local agentic coding Benchmark : Qwen 3.8 27B (in many weights quants / cache quants / engine / reasoning effort) vs others.

DGX agent

In medium reasoning mode, it both scores higher than the 3.6 version, AND is very much more efficient (almost half requests needed, and a third less tokens generated) - at DeepSeek v4 Flash 3107 MXFP4

model-releasesr-localllama
17 Aug 2026
Model Releases

MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends

DGX agent

arXiv:2608.13883v1 Announce Type: new Abstract: Most agent-memory benchmarks test post-hoc recall, whereas MemoryArena evaluates whether memory supports interdependent, multi-session task completion.

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

MobileMem: Learning from a Year of Mobile Experiences

DGX agent

arXiv:2608.13606v1 Announce Type: new Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can und

model-releasesarxiv-cs-ai
17 Aug 2026
Agents

Never the Number: Structural Abstention for AI Systems Whose Answers Are Consumed as Fact

DGX agent

arXiv:2608.13926v1 Announce Type: new Abstract: Large language models have made natural language interfaces to databases (NLIDB) newly credible, but LLM text-to-SQL systems fail in a way that matters

agentsarxiv-cs-ai
17 Aug 2026
Tutorials

No Universal Signal Predicts Sample-Level LLM Regression under Version Updates

DGX agent

arXiv:2608.13607v1 Announce Type: new Abstract: Frontier LLMs are updated frequently and typically outperform their predecessors in aggregate. But aggregate gains say little about individual samples:

tutorialsarxiv-cs-ai
17 Aug 2026
Model Releases

Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

DGX agent

arXiv:2608.14349v1 Announce Type: new Abstract: We present a training-free method for multi-modal trajectory prediction that achieves comparable accuracy to a 57M-parameter transformer while requiring

model-releasesarxiv-cs-lg
17 Aug 2026
Safety

PISA: A Pseudo-Individual Source-Domain Feature Adaptation Framework for Test-Time Open-Vocabulary Object Detection

DGX agent

arXiv:2608.14142v1 Announce Type: new Abstract: Open-vocabulary object detection test-time adaptation (OVOD-TTA) aims to address the performance degradation that pre-trained base models suffer when en

safetyarxiv-cs-cv
17 Aug 2026
Local Ai

Polar Code Based Federated Learning: Convergence Analysis and Resource Allocation

DGX agent

arXiv:2608.13961v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data; however, it faces significant communic

local-aiarxiv-cs-lg
17 Aug 2026
Safety

PROVE: Training-Free Prompt Recovery using Verifiable Evidence

DGX agent

arXiv:2608.13671v1 Announce Type: new Abstract: Modern text-to-image models can generate highly realistic images from natural-language prompts, while recent advances in prompt inversion have made it i

safetyarxiv-cs-cv
17 Aug 2026
Model Releases

Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index

DGX agent

Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is 7

model-releasessimon-willison
17 Aug 2026
Model Releases

Qwen 3.8 35bA3b wen?

DGX agent

Artificial analysis index scores Qwen 3.5 27b: 35 Qwen 3.6 27b: 38 Qwen 3.8 27b: 52 What the hell kind of a jump was that? Even if it is benchmaxxed, the jump is insane. Qwen3.6 35b A3b: 32 That's ~6

model-releasesr-localllama
17 Aug 2026
Model Releases

Rethinking Automated Program Repair: The Impact of Bug Complexity, Fault Localization, and LLM Cost-efficiency

DGX agent

arXiv:2608.14065v1 Announce Type: cross Abstract: Background: Software bugs remain a critical challenge in development, necessitating effective Automated Program Repair (APR) techniques. While Large L

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

SDO: Subspace Deconflicting Operator for Multi-Adapter Composition

DGX agent

arXiv:2608.13820v1 Announce Type: new Abstract: Composing independently trained adapters within a shared diffusion backbone provides a modular approach to multi-character generation, but naive joint d

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Second Thought: Reasoning in Parallel as LLM Agents Act and Observe

DGX agent

arXiv:2608.13667v1 Announce Type: new Abstract: LLM agents in the ReAct paradigm alternate between reasoning, acting, and observing, but deliberate reasoning is confined to the Thought phase: while th

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

SemPlan: Benchmarking Structured Semantic Planning for LLM-Based Queries over Enterprise Data

DGX agent

arXiv:2608.13612v1 Announce Type: new Abstract: Natural-language interfaces to enterprise data must translate underspecified requests into governed, executable behavior while controlling invalid queri

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Style or Signature? Artist-Disjoint Evaluation of Style Classification in Frozen Vision Embeddings

DGX agent

arXiv:2608.14435v1 Announce Type: new Abstract: Frozen image embeddings from models such as CLIP are increasingly used to classify paintings by art-historical style, with high reported accuracy. We as

researcharxiv-cs-cv
17 Aug 2026
Local Ai

tencent/EVIE-Preview-4.5B · Hugging Face

DGX agent

Overview EVIE-Preview-4.5B is a state-of-the-art multilingual Visual Document Retrieval (VDR) model built upon Qwen3.5-4B. It employs ColBERT-style late interaction with native 128-dimensional multi-v

local-air-localllama
17 Aug 2026
Model Releases

TENET: One Step Toward Test-Driven Development for Repository-Level Code Generation

DGX agent

arXiv:2509.24148v4 Announce Type: replace-cross Abstract: Test-Driven Development (TDD) is a widely adopted practice that requires developers to create and execute tests alongside implementation. With

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

The Architect: Interactive Visualization of Deep Learning Mathematics Directly in Microsoft Excel

DGX agent

arXiv:2608.13572v1 Announce Type: cross Abstract: We present The Architect, a system that turns Microsoft Excel into an interactive view of deep learning mathematics. A user describes a neural network

model-releasesarxiv-cs-ai
17 Aug 2026
Research

The Nonstationarity-Complexity Tradeoff in Return Prediction

DGX agent

arXiv:2512.23596v2 Announce Type: replace-cross Abstract: Does more data improve return prediction? In non-stationary financial markets, longer training windows improve prediction of complex models bu

researcharxiv-cs-lg
17 Aug 2026
Research

Think in Latent, Explain in Language: Self-Explainable Latent Reasoning

DGX agent

arXiv:2608.13570v1 Announce Type: cross Abstract: Latent reasoning has emerged as a powerful alternative to text-based Chain-of-Thought (CoT), offering significant gains in computational efficiency by

researcharxiv-cs-ai
17 Aug 2026
Research

Towards Efficient Multimodal and Multilingual Opinion Extraction for STI: A QLoRA-Based Fine-Tuning Approach

DGX agent

arXiv:2608.14152v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have reshaped semantic analysis. Opinion Extraction (OE) for Science and Technology Intelligence (STI) r

researcharxiv-cs-ai
17 Aug 2026
Applications

TRUE-Colon: Exposing a Consistent Transfer Asymmetry in Real-Time Polyp Detection

DGX agent

arXiv:2608.13711v1 Announce Type: cross Abstract: Computer-aided detection (CADe) systems for colonoscopy promise to reduce clinical miss rates, yet reliable real-world deployment remains elusive. Thi

applicationsarxiv-cs-cv
17 Aug 2026
Research

UMPIRE-Net: Unrolled Magnitude-Phase Regularization Network for Accelerated MRI

DGX agent

arXiv:2608.14422v1 Announce Type: cross Abstract: MRI reconstruction from undersampled k-space measurements is an ill-posed inverse problem. Physics-driven deep learning (PD-DL) methods have shown str

researcharxiv-cs-cv
17 Aug 2026
← Previous
1…651652653654655…1371
Next →