AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation

DGX agent

arXiv:2605.12386v1 Announce Type: new Abstract: Robotic manipulation is typically evaluated by task success, but successful completion does not guarantee safe execution. Many safety failures are tempo

model-releasesarxiv-cs-ro
13 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation

DGX agent

arXiv:2605.12022v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong performance on standard knowledge evaluation benchmarks, yet recent work shows that their knowledge capabili

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Scaling Laws and Tradeoffs in Recurrent Networks of Expressive Neurons

DGX agent

arXiv:2605.12049v1 Announce Type: new Abstract: Cortical neurons are complex, multi-timescale processors wired into recurrent circuits, shaped by long evolutionary pressure under stringent biological

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

SCOPE: Siamese Contrastive Operon Pair Embeddings for Functional Sequence Representation and Classification

DGX agent

arXiv:2605.11022v1 Announce Type: cross Abstract: Identifying operons is a fundamental step in understanding prokaryotic gene regulation, as classifying genes into operons supports the reconstruction

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Search Your Block Floating Point Scales!

DGX agent

arXiv:2605.12464v1 Announce Type: new Abstract: Quantization has emerged as a standard technique for accelerating inference for generative models by enabling faster low-precision computations and redu

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model

DGX agent

arXiv:2605.11817v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown remarkable promise in robotics manipulation, yet their high computational cost hinders real-time deploy

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Self-Supervised Laplace Approximation for Bayesian Uncertainty Quantification

DGX agent

arXiv:2605.12208v1 Announce Type: cross Abstract: Approximate Bayesian inference typically revolves around computing the posterior parameter distribution. In practice, however, the main object of inte

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

SEMIR: Semantic Minor-Induced Representation Learning on Graphs for Visual Segmentation

DGX agent

arXiv:2605.12389v1 Announce Type: new Abstract: Segmenting small and sparse structures in large-scale images is fundamentally constrained by voxel-level, lattice-bound computation and extreme class im

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes

DGX agent

arXiv:2605.11680v1 Announce Type: new Abstract: We introduce ShapeCodeBench, a synthetic benchmark for perception-to-program reconstruction: given a rendered raster image, a model must emit an executa

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces

DGX agent

arXiv:2605.12015v1 Announce Type: cross Abstract: Reusable skills are becoming a common interface for extending large language model agents, packaging procedural guidance with access to files, tools,

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Slicing and Dicing: Configuring Optimal Mixtures of Experts

DGX agent

arXiv:2605.11689v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures have become standard in large language models, yet many of their core design choices - expert count, granularit

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval…

DGX agent

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the harness stops rationing. the

model-releasesemad-mostaque--x
13 May 2026
Model Releases

SOAR: Regression-based LiDAR Relocalization for UAVs

DGX agent

arXiv:2602.13267v3 Announce Type: replace Abstract: Regression-based LiDAR relocalization has recently emerged as a promising solution for high-precision positioning in GNSS-denied environments. Howev

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Solve the Loop: Attractor Models for Language and Reasoning

DGX agent

arXiv:2605.12466v1 Announce Type: cross Abstract: Looped Transformers offer a promising alternative to purely feed-forward computation by iteratively refining latent representations, improving languag

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Sources: Google plans to announce a new Gemini model at its I/O conference next week; the model will land roughly in the class of GPT-5.5, but short of Mythos (Alex Heath/Sources)

DGX agent

Alex Heath / Sources: Sources: Google plans to announce a new Gemini model at its I/O conference next week; the model will land roughly in the class of GPT-5.5, but short of Mythos — Sources say that

model-releasestechmeme
13 May 2026
Model Releases

Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos (Bloomberg)

DGX agent

Bloomberg: Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos — French artificial intelligence s

model-releasestechmeme
13 May 2026
Model Releases

Sparsity-Constraint Optimization via Splicing Iteration

DGX agent

arXiv:2406.12017v2 Announce Type: replace-cross Abstract: Sparsity-constrained optimization underlies many problems in signal processing, statistics, and machine learning. State-of-the-art hard-thresh

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

DGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

DGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

DGX agent

arXiv:2602.15620v4 Announce Type: replace Abstract: Reinforcement Learning (RL) has significantly improved large language model reasoning, but existing RL fine-tuning methods rely heavily on heuristic

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Starship launch next week!

DGX agent

Starship launch next week! Starship’s twelfth flight test will debut the next generation Starship and Super Heavy vehicles, powered by the next evolution of the Raptor engine and launching from a newl

model-releaseselon-musk--x
13 May 2026
Model Releases

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent …

DGX agent

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent SDK - claude -p - Claude Code GitHub Actions - Third-party a

model-releasesboris-cherny--x
13 May 2026
Model Releases

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning

DGX agent

arXiv:2605.11922v1 Announce Type: cross Abstract: Existing code reasoning methods primarily supervise final code outputs, ignoring intermediate states, often leading to reward hacking where correct an

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

DGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

DGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

DGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning

DGX agent

arXiv:2605.12179v1 Announce Type: new Abstract: Recent advancements in video-audio joint generation have achieved remarkable success in semantic correspondence. However, achieving precise temporal syn

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Targeted Neuron Modulation via Contrastive Pair Search

DGX agent

arXiv:2605.12290v1 Announce Type: new Abstract: Language models are instruction-tuned to refuse harmful requests, but the mechanisms underlying this behavior remain poorly understood. Popular steering

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

TB-AVA: Text as a Semantic Bridge for Audio-Visual Parameter Efficient Finetuning

DGX agent

arXiv:2605.11572v1 Announce Type: new Abstract: Audio-visual understanding requires effective alignment between heterogeneous modalities, yet cross-modal correspondence remains challenging when tempor

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Test-Time Compute for Dense Retrieval: Agentic Program Generation with Frozen Embedding Models

DGX agent

arXiv:2605.11374v1 Announce Type: cross Abstract: Test-time compute is widely believed to benefit only large reasoning models. We show it also helps small embedding models. Most modern embedding check

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus…

DGX agent

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus is now FREE for a limited time on Nous Portal — give it a t

model-releasesnous-research--x
13 May 2026
Model Releases

The power of LLMs on your data, more than two orders of magnitude faster and cheaper

DGX agent

Databases have introduced new AI-powered SQL functions which take natural language instructions as input and are evaluated using LLMs. They leverage the power of LLMs to answer new kinds of queries: W

model-releasesgoogle-cloud-ai
13 May 2026
Model Releases

The Price of Proportional Representation in Temporal Voting

DGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

DGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

DGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

model-releasesethan-mollick--x
13 May 2026
Model Releases

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

DGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

model-releasesjeremy-howard--x
13 May 2026
Model Releases

This jackass fought hard to remain NASA Administrator.

DGX agent

This jackass fought hard to remain NASA Administrator. SCOOP: I obtained a pitch deck in which the entity that paid for Transportation Sec’y Duffy’s new reality show outlined different partner levels

model-releasesanthropic--x
13 May 2026
Model Releases

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

DGX agent

arXiv:2605.11574v1 Announce Type: new Abstract: The literature on how large language models handle conflict between their training knowledge and a contradicting document presents a persistent empirica

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others — TikTok anno

model-releasestechmeme
13 May 2026
Model Releases

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

DGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

DGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

DGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
← Previous
1…323324325326327…471
Next →