AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
Model Releases

Decoupling Code Complexity from Newcomer Participation: A Causal Study of AI Coding Agent Adoption in OSS

DGX agent

arXiv:2607.01810v1 Announce Type: cross Abstract: Open-source projects depend on a steady inflow of newcomers. A growing concern is that AI coding agents (tools such as Cursor and Claude Code that wri

model-releasesarxiv-cs-ai
3 Jul 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Dendritic In-Context Learning in a Single-Layer Spiking Neural Network

DGX agent

arXiv:2607.02283v1 Announce Type: cross Abstract: In-context learning (ICL) operates via implicit gradient descent embedded in the forward pass of modern AI architectures -- Transformers, Mamba, state

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Denser neq Better: Limits of On-Policy Self-Distillation for Continual Post-Training

DGX agent

arXiv:2607.01763v1 Announce Type: cross Abstract: Continual post-training enables foundation models to acquire new knowledge while preserving existing capabilities. Recent work suggests that on-policy

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Discrete Diffusion Language Models for Interactive Radiology Report Drafting

DGX agent

arXiv:2607.01436v1 Announce Type: new Abstract: Diffusion language models, which generate text by denoising a token canvas bidirectionally instead of emitting tokens left to right, have become competi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Distributed Attacks in Persistent-State AI Control

DGX agent

arXiv:2607.02514v1 Announce Type: new Abstract: As AI coding agents become more autonomous, they increasingly ship code iteratively, with the codebase persisting across sessions. This persistence crea

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Distributionally Robust Listwise Preference Optimization

DGX agent

arXiv:2607.01715v1 Announce Type: new Abstract: Existing robust preference optimization for language-model alignment mainly studies pairwise supervision and places robustness at the dataset, prompt, o

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Diverse Evidence, Better Forecasts: Multi-Agent Deliberation Under Information Asymmetry

DGX agent

arXiv:2607.01661v1 Announce Type: new Abstract: Multi-agent systems are increasingly used for forecasting future events, as deliberation among multiple LLMs is believed to improve reasoning and calibr

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

DL-VINS-Factory: A Modular Framework for Learned Visual Front-Ends in Visual-Inertial SLAM

DGX agent

arXiv:2607.01757v1 Announce Type: cross Abstract: Deep-learning features excel in visual matching, yet their practical value in tightly coupled visual-inertial SLAM (VI-SLAM) remains insufficiently ch

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

Do Newer Lightweight CNNs Perform Better Under Resource Constraints? A Controlled Multigenerational Study of Architecture, Initialization, Training Budget, and Efficiency

DGX agent

arXiv:2607.01984v1 Announce Type: cross Abstract: Newer lightweight convolutional neural networks are often presented as improving predictive performance and deployment efficiency, but such claims req

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

DGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

eCream-MedCorpus A Large-Scale Corpus of Clinical Notes for Italian

DGX agent

arXiv:2606.12569v2 Announce Type: replace-cross Abstract: We present eCream-MedCorpus, a new and unique large-scale dataset of clinical notes produced in Emergency Departments of Italian hospitals. Th

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

EduArt: An educational-level benchmark for evaluating art history knowledge in large language models

DGX agent

arXiv:2607.02007v1 Announce Type: new Abstract: Large language models now score near ceiling on general benchmarks, but these aggregate measures reveal little about how models behave within single dis

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots

DGX agent

arXiv:2607.02501v1 Announce Type: new Abstract: Embodied AI models now span vision-language-action (VLA) models and world-action models (WAMs), but practical deployment remains fragmented across model

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

EO-Agents: A Three-Agent LLM Pipeline for Earth Observation Hypothesis Generation

DGX agent

arXiv:2607.01584v1 Announce Type: new Abstract: Large language models have recently been explored for scientific hypothesis generation, but most prior work relies on unstructured literature and free-f

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

EPnG: Adaptive Expert Prune-and-Grow for Parameter-Efficient MoE Fine-tuning

DGX agent

arXiv:2607.01789v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale efficiently but remain costly to adapt due to redundant experts and uniform parameter allocation. Existing param

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Evidence-State Rewards for Long-Context Reasoning

DGX agent

arXiv:2607.02073v1 Announce Type: new Abstract: Long-context reasoning requires models to locate, revise, and synthesize evidence distributed across lengthy inputs. Existing long-context RL methods us

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments

DGX agent

arXiv:2607.02440v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to improve executable policies through feedback, yet existing evaluations often collapse this process into a

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

eXact-Prior Variational Autoencoder (X-VAE): Learning Data-Adaptive Gaussian Mixture Priors for Latent Distributions

DGX agent

arXiv:2607.01275v1 Announce Type: cross Abstract: Variational Autoencoders (VAEs) commonly assume a standard isotropic Gaussian prior over the latent space, an assumption that often fails to capture t

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/shea…

DGX agent

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/sheaf-admm/ Most AI models process information as one giant, mon

model-releasesdavid-ha--x
3 Jul 2026
Model Releases

Expander Sparse Autoencoders: Parameter-Efficient Dictionaries for Mechanistic Interpretability

DGX agent

arXiv:2607.01799v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) decompose internal activations of neural networks into sparse linear combinations of learned features by fitting an overcom

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Fable's judgement

DGX agent

One of the most interesting tips I got from the Fireside Chat I hosted with Cat Wu and Thariq Shihipar from the Claude Code team at AIE on Wednesday was to let Fable (and to a certain extent Opus) use

model-releasessimon-willison
3 Jul 2026
Model Releases

Fast Multi-dimensional Refusal Subspaces via RFM-AGOP

DGX agent

arXiv:2607.02396v1 Announce Type: new Abstract: Steering and monitoring activations in Large Language Models (LLMs) are increasingly used for both safety and interpretability. Early work assumed behav

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Fixed-Set Robustness in Programming by Example: Example Corruption and Semantic Partition Recovery

DGX agent

arXiv:2607.01280v1 Announce Type: new Abstract: Programming-by-example systems infer programs from a small set of input-output examples. Robust PBE work usually models wrong examples as samples from a

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Frequency Shift Physics-Informed Extreme Learning Machine for Solving High-Frequency Partial Differential Equations

DGX agent

arXiv:2607.01694v1 Announce Type: new Abstract: Solving partial differential equations (PDEs) with high-frequency solutions remains a central challenge in physics-informed machine learning due to spec

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection

DGX agent

arXiv:2512.10485v2 Announce Type: replace-cross Abstract: Vulnerability detection methods based on deep learning (DL) have shown strong performance on benchmark datasets, yet their real-world effectiv

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

From Monolingual to Multilingual: Evaluating Mamba for ASR in South African Languages

DGX agent

arXiv:2607.01502v1 Announce Type: new Abstract: Recent advances in automatic speech recognition (ASR) have explored different sequence models, including Conformer-based models and newer state space mo

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Gemini Omni Flash can manipulate objects and environments in existing videos using simple text prompts. Sometimes it takes a bit of iteratio…

DGX agent

Gemini Omni Flash can manipulate objects and environments in existing videos using simple text prompts. Sometimes it takes a bit of iteration and specific prompting, but it opens up a lot of creative

model-releasescomfyui--x
3 Jul 2026
Model Releases

Generative AI and Federated Learning for Intrusion Detection Systems: A Survey

DGX agent

arXiv:2607.01305v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDSs) are essential for monitoring network traffic and identifying malicious activities in modern cyber-physical, Interne

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Generic Expert Coverage for Pruning SparseMixture-of-Experts Language Models

DGX agent

arXiv:2607.01710v1 Announce Type: new Abstract: Sparsely activated Mixture-of-Experts (MoE) language models contain substantial structured redundancy among routed experts, but pruning them without dow

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

GLM-5.2 is now selectable in Claude Code via Hugging Face🤗 Inference Providers + hf-claude. Open models are becoming easier to plug directl…

DGX agent

GLM-5.2, an open-source model available through Hugging Face, can now be selected and used within Claude Code through Hugging Face Inference Providers and the hf-claude integration. This development d

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

Google DeepMind and A24 announce first-of-its-kind research partnership

DGX agent

Google DeepMind and A24 announced a first-of-its-kind research partnership pairing the AI research lab with the filmmaker-focused studio to help artists develop new workflows and techniques. Google is

model-releasesgoogle-deepmind
3 Jul 2026
Model Releases

Gravity-Awareness: Deep Learning Models and LLM Simulation of Human Awareness in Altered Gravity

DGX agent

arXiv:2511.05536v2 Announce Type: replace-cross Abstract: Earth s gravity fundamentally shapes human behaviour. The brain encodes this force as an internal model of gravity, enabling the prediction an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

GroundEval: A Deterministic Replacement for LLM-as-Judge in Stateful Agent Evaluation

DGX agent

arXiv:2606.22737v2 Announce Type: replace Abstract: Before letting an agent operate over real context, can you prove it used the right evidence? GroundEval turns that question into a deterministic tes

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety

DGX agent

arXiv:2607.02079v1 Announce Type: new Abstract: We present HaloGuard 1.0, an open-weights implementation of the constitutional-classifier paradigm for input safety. It achieves state-of-the-art perfor

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

DGX agent

arXiv:2606.14249v2 Announce Type: replace Abstract: AI agent performance depends critically on the runtime harness, comprising the prompts, tools, memory, and control flow that mediate how a model obs

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Has This Checkpoint Been Abliterated? A Two-Signal Audit and Its Failure Map

DGX agent

arXiv:2607.01854v1 Announce Type: cross Abstract: Can a platform tell, before deployment, whether an open-weight checkpoint has had its refusal mechanism stripped? Runtime guards cannot: they score ge

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures

DGX agent

arXiv:2607.02266v1 Announce Type: cross Abstract: Most data-mixing methods assume the corpus has already been partitioned into groups, and the choice of those groups determines what a mixer can expres

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can obj…

DGX agent

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can objectively score, so style, structure, and diversity quietly c

model-releasesdair-ai--x
3 Jul 2026
Model Releases

HNSW with Accuracy Guarantees Using Graph Spanners -- A Technical Report

DGX agent

arXiv:2607.02338v1 Announce Type: cross Abstract: Hierarchical Navigable Small World (HNSW) graphs serve as the industry standard due to their logarithmic complexity and strong empirical performance.

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

HULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read Generation

DGX agent

arXiv:2607.02381v1 Announce Type: new Abstract: This paper describes the participation of HULAT2-UC3M in the Spanish track of MER-TRANS 2026, a shared task on multilingual Easy-to-Read translation. Th

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Human Capital, Not Model Benchmarks, Predicts Hybrid Intelligence in Forecasting

DGX agent

arXiv:2607.02467v1 Announce Type: cross Abstract: Whether pairing people with AI helps or hurts is usually reported as a single average effect. Using a real-money prediction market (Polymarket) as an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

I kept asking Claude Fable to make the game 'more AAA' over and over again. The results are... interesting. In Claude's view, this meant upg…

DGX agent

I kept asking Claude Fable to make the game 'more AAA' over and over again. The results are... interesting. In Claude's view, this meant upgrading graphics, boss fights, mechanics adding custom sounds

model-releasesethan-mollick--x
3 Jul 2026
Model Releases

InduceKV: Fixed-Footprint Continual Adaptation of Multimodal LLMs via Inducing KV Memories

DGX agent

arXiv:2607.02010v1 Announce Type: new Abstract: Multimodal large language models must adapt to evolving tasks and domains, yet continual improvement under bounded deployment footprint remains difficul

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Influence of Radial Basis Activation Functions on Intelligent Controller for Robotic Manipulators

DGX agent

arXiv:2607.02167v1 Announce Type: cross Abstract: This paper presents an intelligent control framework for trajectory tracking of robotic manipulators using radial basis function (RBF) neural networks

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

IonSense-QKG: A Quantum-Readiness Metadata Framework for Lithium-Ion Battery Dataset Discovery

DGX agent

arXiv:2607.01286v1 Announce Type: new Abstract: Public lithium-ion battery datasets are increasingly used for state-of-health estimation, remaining-useful-life prediction, anomaly detection, electroch

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

IsoSci: A Benchmark of Isomorphic Cross-Domain Science Problems for Evaluating Reasoning versus Knowledge Retrieval in LLMs

DGX agent

arXiv:2607.01431v1 Announce Type: cross Abstract: We introduce ISOSCI, a benchmark of isomorphic cross-domain science problem pairs that separates reasoning ability from domain knowledge retrieval in

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

June 2026 newsletter

DGX agent

The June edition of my sponsors-only monthly newsletter is out. If you are a sponsor (or if you start a sponsorship now) you can access it here. This month: Claude Fable 5, GPT-5.6, and US export rest

model-releasessimon-willison
3 Jul 2026
Model Releases

LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning

DGX agent

arXiv:2607.02513v1 Announce Type: cross Abstract: LLMs memorize sensitive training data, including personally identifiable information (PII), creating a pressing need for reliable post hoc removal met

model-releasesarxiv-cs-ai
3 Jul 2026
← Previous
1…132133134135136…471
Next →