AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,593 results
Model Releases

ASAP: Amortized Doubly-Stochastic Attention via Sliced Dual Projection

DGX agent

arXiv:2605.12879v1 Announce Type: new Abstract: Doubly-stochastic attention has emerged as a transport-based alternative to row-softmax attention, with recent Transformer variants using it to reduce a

model-releasesarxiv-cs-lg
14 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

AttenA+: Rectifying Action Inequality in Robotic Foundation Models

DGX agent

arXiv:2605.13548v1 Announce Type: cross Abstract: Existing robotic foundation models, while powerful, are predicated on an implicit assumption of temporal homogeneity: treating all actions as equally

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Attention Once Is All You Need: Efficient Streaming Inference with Stateful Transformers

DGX agent

arXiv:2605.13784v1 Announce Type: new Abstract: Conventional transformer inference engines are request-driven, paying an O(n) prefill cost on every query. In streaming workloads, where data arrives co

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Automating AI research is the next major step in AI We let Claude Code (Opus 4.7) and Codex (GPT 5.5) run autonomously on the nanoGPT speedr…

DGX agent

Automating AI research is the next major step in AI We let Claude Code (Opus 4.7) and Codex (GPT 5.5) run autonomously on the nanoGPT speedrun optimizer track using our idle compute. ~10k runs, ~14k H

model-releasesboris-cherny--x
14 May 2026
Model Releases

Bayesian In Vivo Tracking of Synapses using Joint Poisson Deconvolution and Diffeomorphic Registration

DGX agent

arXiv:2605.13455v1 Announce Type: new Abstract: Synapses are densely packed submicron structures that dynamically reorganize during learning and memory formation. Longitudinal extit{in vivo} imaging o

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Bayesian Model Merging

DGX agent

arXiv:2605.12843v1 Announce Type: cross Abstract: Model merging aims to combine multiple task-specific expert models into a single model without joint retraining, offering a practical alternative to m

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

BEAVER: An Enterprise Benchmark for Text-to-SQL

DGX agent

arXiv:2409.02038v3 Announce Type: replace-cross Abstract: Existing text-to-SQL benchmarks have largely been constructed from public databases with well-structured schemas and simplistic question-SQL p

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Benchmarking Attribute Discrimination in Infant-Scale Vision-Language Models

DGX agent

arXiv:2512.18951v3 Announce Type: replace Abstract: Infants learn not only object categories but also fine-grained visual attributes such as color, size, and texture from limited experience. Prior inf

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Bias In, Bias Out? Finding Unbiased Subnetworks in Vanilla Models

DGX agent

arXiv:2603.05582v2 Announce Type: replace-cross Abstract: The issue of algorithmic biases in deep learning has led to the development of various debiasing techniques, many of which perform complex tra

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

big launches today!

DGX agent

big launches today! 🚀Launching: LangSmith Engine LangSmith Engine is an agent that sits on top of your traces It runs in the background and automatically identifies issues It then proactively suggests

model-releasesharrison-chase--x
14 May 2026
Model Releases

BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration

DGX agent

arXiv:2605.12520v1 Announce Type: cross Abstract: Taxonomy induction is crucial for organizing concepts into explicit and interpretable semantic hierarchies. While existing methods have achieved promi

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Building Interactive Real-Time Agents with Asynchronous I/O and Speculative Tool Calling

DGX agent

arXiv:2605.13360v1 Announce Type: new Abstract: There is a growing demand for agentic AI technologies for a range of downstream applications like customer service and personal assistants. For applicat

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

built a bird search app to show off pinecone's full text search capabilities 2,079 bird articles, full-text keyword search, Gemini Embedding…

DGX agent

built a bird search app to show off pinecone's full text search capabilities 2,079 bird articles, full-text keyword search, Gemini Embedding 2 for cross-modal visual search. the birds did not consent

model-releasespinecone--x
14 May 2026
Model Releases

Children's English Reading Story Generation via Supervised Fine-Tuning of Compact LLMs with Controllable Difficulty and Safety

DGX agent

arXiv:2605.13709v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely applied in educational practices, such as for generating children's stories. However, the generated stories ar

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation

DGX agent

arXiv:2605.12857v1 Announce Type: cross Abstract: Existing API-based agentic systems for RTL code generation are fundamentally misaligned with industrial practice: they assume a golden testbench is av

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence

DGX agent

arXiv:2605.12882v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced document understanding, yet current Doc-VQA evaluations score only the final answ

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Cloud CISO Perspectives: How Google + Wiz changes multicloud strategy for CISOs

DGX agent

Welcome to the first Cloud CISO Perspectives for May 2026. Today, Vinod D’Souza, director, Office of the CISO, shares highlights from his RSA Conference fireside chat with Anthony Belfiore, chief stra

model-releasesgoogle-cloud-ai
14 May 2026
Model Releases

Code-Centric Detection of Vulnerability-Fixing Commits: A Unified Benchmark and Empirical Study

DGX agent

arXiv:2605.13138v1 Announce Type: cross Abstract: Automated detection of vulnerability-fixing commits (VFCs) is critical for timely security patch deployment, as advisory databases lag patch releases

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

CodeClash: Benchmarking Goal-Oriented Software Engineering

DGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis

DGX agent

arXiv:2601.21577v2 Announce Type: replace Abstract: Catastrophic forgetting during knowledge injection impairs the ability of large language models to acquire new knowledge without overwriting previou

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Compact 3D Gaussian Splatting For Dense Visual SLAM

DGX agent

arXiv:2403.11247v3 Announce Type: replace Abstract: Recent work has shown that 3D Gaussian-based SLAM enables high-quality reconstruction, accurate pose estimation, and real-time rendering of scenes.

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Compact Latent Manifold Translation: A Parameter-Efficient Foundation Model for Cross-Modal and Cross-Frequency Physiological Signal Synthesis

DGX agent

arXiv:2605.13248v1 Announce Type: cross Abstract: The analysis of physiological time series, such as electrocardiograms (ECG) and photoplethysmograms (PPG), is persistently hindered by modality and fr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models

DGX agent

arXiv:2511.15743v2 Announce Type: replace Abstract: Operational forecasting of the ionosphere remains a critical space weather challenge due to sparse observations, complex coupling across geospatial

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

ConRetroBert: EMA Stabilized Dual Encoders for Template-Based Single-Step Retrosynthesis

DGX agent

arXiv:2605.12736v1 Announce Type: new Abstract: Template based single step retrosynthesis predicts reactants by selecting and applying an explicit reaction template, making each prediction traceable t

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Continual Fine-Tuning of Large Language Models via Program Memory

DGX agent

arXiv:2605.13162v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT), particularly Low-Rank Adaptation (LoRA), has become a standard approach for adapting Large Language Models (LLMs

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Controllable Quantum Memory Capacity in Quantum Reservoir Networks with Tunable partial-SWAPs

DGX agent

arXiv:2605.12713v1 Announce Type: cross Abstract: In the field of quantum reservoir computing (QRC), many different computational models and architectures have been proposed. From these models, we ide

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2

DGX agent

arXiv:2605.12968v1 Announce Type: cross Abstract: Do large language models internally encode ontological relations in a formally verifiable algebraic structure? We introduce Algebraic Ontology Project

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CoRe-Gen: Robust Spectrum-to-Structure Generation under Imperfect Fingerprint Conditions

DGX agent

arXiv:2605.12980v1 Announce Type: cross Abstract: Molecular structure elucidation from tandem mass spectra (MS/MS) remains challenging, particularly for de novo generation beyond database coverage. A

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CoT-Guard: Small Models for Strong Monitoring

DGX agent

arXiv:2605.12746v1 Announce Type: cross Abstract: Monitoring the chain-of-thought (CoT) of reasoning models is a promising approach for detecting covert misbehavior (i.e., hidden objectives) in code g

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CR-Net: Scaling Parameter-Efficient Training with Cross-Layer Low-Rank Structure

DGX agent

arXiv:2509.18993v3 Announce Type: replace Abstract: Low-rank architectures have become increasingly important for efficient large language model (LLM) pre-training, providing substantial reductions in

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

CUBic: Coordinated Unified Bimanual Perception and Control Framework

DGX agent

arXiv:2605.13452v1 Announce Type: cross Abstract: Recent advances in visuomotor policy learning have enabled robots to perform control directly from visual inputs. Yet, extending such end-to-end learn

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models

DGX agent

arXiv:2605.13276v1 Announce Type: new Abstract: The rapid evolution of Embodied AI has enabled Vision-Language-Action (VLA) models to excel in multimodal perception and task execution. However, applyi

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

datasette-ip-rate-limit 0.1a0

DGX agent

Release: datasette-ip-rate-limit 0.1a0 The datasette.io site was being hammered by poorly-behaved crawlers, so I had Codex (GPT-5.5 xhigh) build a configurable rate limiting plugin to block IPs that w

model-releasessimon-willison
14 May 2026
Model Releases

DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions

DGX agent

arXiv:2509.19538v2 Announce Type: replace-cross Abstract: Diffusion-based world models have demonstrated strong capabilities in synthesizing realistic long-horizon trajectories for offline reinforceme

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Decision Tree Learning on Product Spaces

DGX agent

arXiv:2605.12983v1 Announce Type: new Abstract: Decision tree learning has long been a central topic in theoretical computer science, driven by its practical importance. A fundamental and widely used

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Deepseek V4 Flash is now free via Nous Portal for a limited time thanks to @novita_labs!

DGX agent

Nous Research announced that Deepseek V4 Flash is temporarily available for free access through the Nous Portal, courtesy of Nous Research in collaboration with @novita_labs. This limited-time offer p

model-releasesnous-research--x
14 May 2026
Model Releases

Deepseek V4 Flash is one of the best 'worker' models out there IMHO so this is very cool! Yet ANOTHER free model available through the Nous …

DGX agent

Deepseek V4 Flash is one of the best 'worker' models out there IMHO so this is very cool! Yet ANOTHER free model available through the Nous Portal 😍😍😍 Can't believe people are still playing with toy l

model-releasesnous-research--x
14 May 2026
Model Releases

Dense vs Sparse Pretraining at Tiny Scale: Active-Parameter vs Total-Parameter Matching

DGX agent

arXiv:2605.13769v1 Announce Type: cross Abstract: We study dense and mixture-of-experts (MoE) transformers in a tiny-scale pretraining regime under a shared LLaMA-style decoder training recipe. The sp

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features

DGX agent

arXiv:2605.12874v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are now standard tools for decomposing language model activations into interpretable features, and automated interpretability

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

DistractMIA: Black-Box Membership Inference on Vision-Language Models via Semantic Distraction

DGX agent

arXiv:2605.12574v1 Announce Type: cross Abstract: Vision-language models (VLMs) are trained on large-scale image-text corpora that may contain private, copyrighted, or otherwise sensitive data, motiva

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack

DGX agent

arXiv:2605.12673v1 Announce Type: new Abstract: Agent benchmarks have become the de facto measure of frontier AI competence, guiding model selection, investment, and deployment. However, reward hackin

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

DocAtlas: Multilingual Document Understanding Across 80+ Languages

DGX agent

arXiv:2605.12623v1 Announce Type: cross Abstract: Multilingual document understanding remains limited for low-resource languages due to scarce training data and model-based annotation pipelines that p

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Domain Adaptation of Large Language Models for Polymer-Composite Additive Manufacturing Using Retrieval-Augmented Generation and Fine-Tuning

DGX agent

arXiv:2605.12516v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) often struggle to generate reliable responses in specialized engineering domains due to limited domain gr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

DRIFT: A Benchmark for Task-Free Continual Graph Learning with Continuous Distribution Shifts

DGX agent

arXiv:2605.12998v1 Announce Type: new Abstract: Continual graph learning (CGL) aims to learn from dynamically evolving graphs while mitigating catastrophic forgetting. Existing CGL approaches typicall

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

ECG-NAT: A Self-supervised Neighborhood Attention Transformer for Multi-lead Electrocardiogram Classification

DGX agent

arXiv:2605.13194v1 Announce Type: cross Abstract: Electrocardiogram (ECG) arrhythmia classification remains challenging due to signal variability, noise, limited labeled data, and the difficulty in ac

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

EcoGEO: Trajectory-Aware Evidence Ecosystems for Web-Enabled LLM Search Agents

DGX agent

arXiv:2605.12887v1 Announce Type: cross Abstract: Web-enabled LLM agents are changing how online information influences search outcomes. Existing Generative Engine Optimization (GEO) studies mainly fo

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling

DGX agent

arXiv:2605.13062v1 Announce Type: new Abstract: Recent image editing models have achieved remarkable progress in instruction following, multimodal understanding, and complex visual editing. However, e

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Effective Context in Transformers: An Analysis of Fragmentation and Tokenization

DGX agent

arXiv:2605.13485v1 Announce Type: new Abstract: Transformers predict over a representation of a sequence. The same data can be written as bytes, characters, or subword tokens, and these representation

model-releasesarxiv-cs-lg
14 May 2026
← Previous
1…314315316317318…471
Next →