AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Model Releases

The Quality of Claude AI-authored Python Tests Is Not Weaker Than Human-authored Tests

DGX agent

arXiv:2608.15188v1 Announce Type: cross Abstract: We evaluate the quality of Claude AI-written Python tests against human-written Python tests from two established open-source projects Django and Pand

model-releasesarxiv-cs-ai
18 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

the real unlock is not cheaper evaluation it is making specialized evaluators cheap enough to run continuously on production traces so evalu…

DGX agent

the real unlock is not cheaper evaluation it is making specialized evaluators cheap enough to run continuously on production traces so evaluation stops being a launch checklist and becomes a permanent

model-releasesharrison-chase--x
18 Aug 2026
Model Releases

TISC: A Text-Driven Image Semantic Communication System for Faithful Reconstruction

DGX agent

arXiv:2608.16100v1 Announce Type: new Abstract: Generative image semantic communication converts an image into a text description and then performs text-to-image reconstruction at the receiver via dif

model-releasesarxiv-cs-cv
18 Aug 2026
Research

Towards Computational Provenance: Carrying Causal-State Evidence in Generated Text

DGX agent

arXiv:2608.16868v1 Announce Type: cross Abstract: A language model's output does not by itself provide verifiable evidence about the internal computation that produced it. We study computational prove

researcharxiv-cs-ai
18 Aug 2026
Local Ai

TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation

DGX agent

arXiv:2608.16765v1 Announce Type: cross Abstract: Despite recent advances in unified multimodal models for multi-reference image generation, existing benchmarks remain organized around predefined task

local-aiarxiv-cs-ai
18 Aug 2026
Model Releases

TRACE: Trajectory Aware Reasoning for Multi-Turn Adversarial Conversation Evaluation

DGX agent

arXiv:2608.15594v1 Announce Type: new Abstract: Multi-turn jailbreak attacks have emerged as a critical safety threat to LLMs, as harmful objectives are decomposed across a sequence of apparently beni

model-releasesarxiv-cs-ai
18 Aug 2026
Applications

Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision

DGX agent

arXiv:2608.16812v1 Announce Type: new Abstract: Existing image editing frameworks predominantly follow the training paradigm of text-to-image diffusion models. However, extending this paradigm to imag

applicationsarxiv-cs-cv
18 Aug 2026
Research

Valhalla: A Layered Knowledge-State and Service-Governance Framework for Long-Term Scientific Knowledge Work

DGX agent

arXiv:2608.15193v1 Announce Type: cross Abstract: As large language model (LLM) agents are increasingly adopted in scientific research, external knowledge bases, knowledge graphs, and long-term memory

researcharxiv-cs-ai
18 Aug 2026
Agents

WARA: Toward Automated Wireless Optimization Research with Closed-Loop LLM Agents

DGX agent

arXiv:2608.14573v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly capable of tool use, code execution, artifact inspection, and iterative revision, creating new oppo

agentsarxiv-cs-ai
18 Aug 2026
Research

Who Leads Now? Token-Level Modality Arbitration for Chart-to-Code Generation

DGX agent

arXiv:2608.15510v1 Announce Type: new Abstract: Chart-to-code generation requires a model to read the fine-grained visual details of a chart and write executable code that reproduces it. Existing char

researcharxiv-cs-ai
18 Aug 2026
Model Releases

A Calibrated Test of Internal Action Maps: State Signals Without Global Affine Closure

DGX agent

arXiv:2608.13626v1 Announce Type: new Abstract: A hidden state signal can be decodable or causally usable without supporting a reusable action map. We test whether action maps fitted without a source

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

AgentRewind: Recoverable Execution for Long-Horizon LLM Agents

DGX agent

arXiv:2608.14380v1 Announce Type: new Abstract: Many real-world tasks require LLM agents to interact with their environments over long execution horizons. Errors that occur early in execution may prop

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

AI-Assisted Discovery and Construction of a Counterexample to the Convergence of Three-Block ADMM with the Identity Matrix as its Third Constraint Block

DGX agent

arXiv:2608.14396v1 Announce Type: cross Abstract: The alternating direction method of multipliers (ADMM), as a landmark algorithm, has attracted tremendous research attention and extensive practical a

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Anthropic's text watermark alters word probabilities to embed a fingerprint, which could degrade Claude's writing, despite its claim of no impact on quality (John Gruber/Daring Fireball)

DGX agent

John Gruber / Daring Fireball: Anthropic's text watermark alters word probabilities to embed a fingerprint, which could degrade Claude's writing, despite its claim of no impact on quality — When I wro

model-releasestechmeme
17 Aug 2026
Model Releases

Approximate Muon with low-rank adapters

DGX agent

arXiv:2608.14492v1 Announce Type: new Abstract: The Muon optimizer shows clear benefits versus alternatives when pretraining neural networks. However, it is used less frequently for parameter-efficien

model-releasesarxiv-cs-lg
17 Aug 2026
Research

APTER: Adaptive Post-Training with Expert-Grounded Rubrics

DGX agent

arXiv:2608.14212v1 Announce Type: new Abstract: As large language models enter professional domains, they must satisfy domain constraints, include critical evidence, and provide complete reasoning rat

researcharxiv-cs-ai
17 Aug 2026
Model Releases

BCIJelly: An integrated ecosystem for brain-computer interface research

DGX agent

arXiv:2608.13576v1 Announce Type: cross Abstract: Brain-computer interface (BCI) research relies on multistage computational pipelines, yet progress remains constrained by fragmented data formats, het

model-releasesarxiv-cs-lg
17 Aug 2026
Model Releases

Built a token-aware gateway/load balancer for local LLM stacks — because nginx has no idea what a token costs

DGX agent

If you're running Ollama, llama.cpp, or vLLM behind nginx or HAProxy for more than a single user, you've probably hit this: nginx treats a 10-token prompt and a 10k-token prompt as identical 'one requ

model-releasesr-ollama
17 Aug 2026
Safety

BUZZY: Contrastive Scoring to Mitigate Text-Induced Bias in Multimodal Multiple-Choice QA

DGX agent

arXiv:2603.28026v2 Announce Type: replace Abstract: Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively measurable setting for evaluating vision-language model

safetyarxiv-cs-ai
17 Aug 2026
Research

Data-driven techniques for translational neuroscience and personalized neuro-health

DGX agent

arXiv:2608.13749v1 Announce Type: cross Abstract: Neurodegenexrative diseases such as Alzheimer's disease and Parkinson's disease are diagnosed most reliably only after substantial, often irreversible

researcharxiv-cs-ai
17 Aug 2026
Model Releases

DeepSeek V4 Pro 0813 vs Claude Fable 5 on DeepSWE: Cost, Coding, and Routing

DGX agent

DeepSeek V4 Pro 0813 and Claude Fable 5 were benchmarked on DeepSWE. A cascading strategy that starts with Pro 0813 and escalates to Fable only when it fails solves 82.7 % of tasks at an average cost

model-releasestogether-ai-blog
17 Aug 2026
Model Releases

defenders can see the future, and have a narrow window to uplevel their cybersecurity practices now. key is to uplevel fundamentals and appl…

DGX agent

defenders can see the future, and have a narrow window to uplevel their cybersecurity practices now. key is to uplevel fundamentals and apply the best AI tools. what we’re doing at OpenAI, and where o

model-releasesopenai--x
17 Aug 2026
Model Releases

Doomed to Re-Annotate, Forever: The ImageNet Story

DGX agent

arXiv:2608.13783v1 Announce Type: new Abstract: Top-1 accuracy on ImageNet-1k remains the most commonly reported metric in visual recognition. Quality issues with the dataset have been repeatedly repo

model-releasesarxiv-cs-cv
17 Aug 2026
Model Releases

Envs-FORGE: Frontier-Optimized Reward-Grounded Environment Synthesis for Agent RL

DGX agent

arXiv:2608.14312v1 Announce Type: new Abstract: Reinforcement learning (RL) for terminal agents needs executable training environments with reliable rewards and useful difficulty. Fixed recipes such a

model-releasesarxiv-cs-cl
17 Aug 2026
Model Releases

Estimating Dynamic Soft Continuum Robot States From Boundaries

DGX agent

arXiv:2505.04491v3 Announce Type: replace Abstract: State estimation is one of the fundamental problems in robotics. For soft continuum robots, this task is particularly challenging because their stat

model-releasesarxiv-cs-ro
17 Aug 2026
Safety

Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT

DGX agent

arXiv:2608.13681v1 Announce Type: cross Abstract: Translating C code into safe, idiomatic Rust is a longstanding software-engineering goal because it can eliminate entire classes of memory-safety vuln

safetyarxiv-cs-ai
17 Aug 2026
Research

FIRM: Fine-Grained Intra-Token Representation of Masks for Remote Sensing Reasoning Segmentation

DGX agent

arXiv:2608.13980v1 Announce Type: new Abstract: Reasoning segmentation requires multimodal large language models (MLLMs) to translate implicit instructions into precise pixel-level masks. MLLMs encode

researcharxiv-cs-cv
17 Aug 2026
Model Releases

Fixed-Budget Gaussian Volume Encoding with Structure-Aware Allocation

DGX agent

arXiv:2608.14112v1 Announce Type: cross Abstract: Scientific simulations often produce scalar volumes faster than they can be stored, transferred, and loaded, while in situ reduction must use only a l

model-releasesarxiv-cs-ai
17 Aug 2026
Research

FLARE MCMC: Fidelity-based Layer-Adaptive REcursive proposals for MCMC

DGX agent

arXiv:2608.13774v1 Announce Type: new Abstract: Markov chain Monte Carlo (MCMC) requires only the ability to evaluate the likelihood, making it a common technique for inference in complex models. Howe

researcharxiv-cs-ai
17 Aug 2026
Research

FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction

DGX agent

arXiv:2608.14205v1 Announce Type: new Abstract: Load imbalance poses a major bottleneck to the efficiency of expert parallelism in distributed inference of Mixture-of-Experts (MoE) models. The most he

researcharxiv-cs-ai
17 Aug 2026
Safety

GALA: Generation-Aware Cross-Modal Alignment for Text-to-Time-Series Synthesis

DGX agent

arXiv:2608.13741v1 Announce Type: new Abstract: Synthesizing time series from natural language is emerging as the most expressive form of controllable time series generation. However, existing text-co

safetyarxiv-cs-cl
17 Aug 2026
Model Releases

GBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack Detection

DGX agent

arXiv:2608.14389v1 Announce Type: cross Abstract: Existing palm presentation attack detection (PAD) datasets are often limited by static imagery, restricted acquisition conditions, or insufficient mul

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Geometric Filtering of LLM-Generated Samples for Few-Shot Text Classification

DGX agent

arXiv:2608.13866v1 Announce Type: cross Abstract: Large language models (LLMs) can generate synthetic training data for text classification, but the quality of generated samples is heterogeneous: some

researcharxiv-cs-cl
17 Aug 2026
Tutorials

Global Interpretability via Automated Preprocessing: A Framework Inspired by Psychiatric Questionnaires

DGX agent

arXiv:2602.23459v2 Announce Type: replace Abstract: Psychiatric questionnaires are highly context sensitive and often only weakly predict subsequent symptom severity, which makes the prognostic relati

tutorialsarxiv-cs-lg
17 Aug 2026
Industry

Greg Brockman calls the OpenAI-Hugging Face incident 'a watershed moment' and discusses how OpenAI and other organizations can use AI to improve cyber defenses (Greg Brockman)

DGX agent

Greg Brockman: Greg Brockman calls the OpenAI-Hugging Face incident “a watershed moment” and discusses how OpenAI and other organizations can use AI to improve cyber defenses — The OpenAI-Hugging Face

industrytechmeme
17 Aug 2026
Model Releases

Grok 4.6 is smart, super fast & affordable

DGX agent

Grok 4.6 is smart, super fast & affordable Grok 4.6 just broke into the Top 3 on the Artificial Analysis Healthcare & Medical Index Grok 4.6 (high) scores 50 - just ONE point off the top score of 51 I

model-releaseselon-musk--x
17 Aug 2026
Model Releases

KV Cache Compression Through the Lens of Transform Coding

DGX agent

arXiv:2608.14191v1 Announce Type: cross Abstract: The key-value (KV) cache stores information from past tokens and is a major memory bottleneck in long-context inference. Existing quantization methods

model-releasesarxiv-cs-cl
17 Aug 2026
Applications

Limitations of Synthetic Data Generation in Specialized Data-Scarce Domains

DGX agent

arXiv:2608.13729v1 Announce Type: new Abstract: Advances in diffusion-based generative models have motivated the use of synthetic image generation to alleviate data scarcity in vision tasks. While thi

applicationsarxiv-cs-cv
17 Aug 2026
Model Releases

Local agentic coding Benchmark : Qwen 3.8 27B (in many weights quants / cache quants / engine / reasoning effort) vs others.

DGX agent

In medium reasoning mode, it both scores higher than the 3.6 version, AND is very much more efficient (almost half requests needed, and a third less tokens generated) - at DeepSeek v4 Flash 3107 MXFP4

model-releasesr-localllama
17 Aug 2026
Model Releases

MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends

DGX agent

arXiv:2608.13883v1 Announce Type: new Abstract: Most agent-memory benchmarks test post-hoc recall, whereas MemoryArena evaluates whether memory supports interdependent, multi-session task completion.

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

MobileMem: Learning from a Year of Mobile Experiences

DGX agent

arXiv:2608.13606v1 Announce Type: new Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can und

model-releasesarxiv-cs-ai
17 Aug 2026
Agents

Never the Number: Structural Abstention for AI Systems Whose Answers Are Consumed as Fact

DGX agent

arXiv:2608.13926v1 Announce Type: new Abstract: Large language models have made natural language interfaces to databases (NLIDB) newly credible, but LLM text-to-SQL systems fail in a way that matters

agentsarxiv-cs-ai
17 Aug 2026
Tutorials

No Universal Signal Predicts Sample-Level LLM Regression under Version Updates

DGX agent

arXiv:2608.13607v1 Announce Type: new Abstract: Frontier LLMs are updated frequently and typically outperform their predecessors in aggregate. But aggregate gains say little about individual samples:

tutorialsarxiv-cs-ai
17 Aug 2026
Model Releases

Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

DGX agent

arXiv:2608.14349v1 Announce Type: new Abstract: We present a training-free method for multi-modal trajectory prediction that achieves comparable accuracy to a 57M-parameter transformer while requiring

model-releasesarxiv-cs-lg
17 Aug 2026
Safety

PISA: A Pseudo-Individual Source-Domain Feature Adaptation Framework for Test-Time Open-Vocabulary Object Detection

DGX agent

arXiv:2608.14142v1 Announce Type: new Abstract: Open-vocabulary object detection test-time adaptation (OVOD-TTA) aims to address the performance degradation that pre-trained base models suffer when en

safetyarxiv-cs-cv
17 Aug 2026
Local Ai

Polar Code Based Federated Learning: Convergence Analysis and Resource Allocation

DGX agent

arXiv:2608.13961v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data; however, it faces significant communic

local-aiarxiv-cs-lg
17 Aug 2026
Safety

PROVE: Training-Free Prompt Recovery using Verifiable Evidence

DGX agent

arXiv:2608.13671v1 Announce Type: new Abstract: Modern text-to-image models can generate highly realistic images from natural-language prompts, while recent advances in prompt inversion have made it i

safetyarxiv-cs-cv
17 Aug 2026
Model Releases

Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index

DGX agent

Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is 7

model-releasessimon-willison
17 Aug 2026
← Previous
1…657658659660661…1380
Next →