AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
13 May 2026

Revisiting Shadow Detection from a Vision-Language Perspective

Model ReleasesDGX agent

arXiv:2605.11771v1 Announce Type: new Abstract: Shadow detection is commonly formulated as a vision-driven dense prediction problem, where models rely primarily on pixel-wise visual supervision to dis

Reviving In-domain Fine-tuning Methods for Source-Free Cross-domain Few-shot Learning

Model ReleasesDGX agent

arXiv:2605.11659v1 Announce Type: new Abstract: Cross-Domain Few-Shot Learning (CDFSL) aims to adapt large-scale pretrained models to specialized target domains with limited samples, yet the few-shot

Robust Promptable Video Object Segmentation

Model ReleasesDGX agent

arXiv:2605.12006v1 Announce Type: new Abstract: The performance of promptable video object segmentation (PVOS) models substantially degrades under input corruptions, which prevents PVOS deployment in


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems

Model ReleasesDGX agent

arXiv:2605.11800v1 Announce Type: cross Abstract: Large language models (LLMs) with mixture-of-experts (MoE) architectures achieve remarkable scalability by sparsely activating a subset of experts per

Routers Learn the Geometry of Their Experts: Geometric Coupling in Sparse Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2605.12476v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (SMoE) models enable scaling language models efficiently, but training them remains challenging, as routing can collapse ont

SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.12386v1 Announce Type: new Abstract: Robotic manipulation is typically evaluated by task success, but successful completion does not guarantee safe execution. Many safety failures are tempo

SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation

Model ReleasesDGX agent

arXiv:2605.12022v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong performance on standard knowledge evaluation benchmarks, yet recent work shows that their knowledge capabili

Scaling Laws and Tradeoffs in Recurrent Networks of Expressive Neurons

Model ReleasesDGX agent

arXiv:2605.12049v1 Announce Type: new Abstract: Cortical neurons are complex, multi-timescale processors wired into recurrent circuits, shaped by long evolutionary pressure under stringent biological

SCOPE: Siamese Contrastive Operon Pair Embeddings for Functional Sequence Representation and Classification

Model ReleasesDGX agent

arXiv:2605.11022v1 Announce Type: cross Abstract: Identifying operons is a fundamental step in understanding prokaryotic gene regulation, as classifying genes into operons supports the reconstruction

Search Your Block Floating Point Scales!

Model ReleasesDGX agent

arXiv:2605.12464v1 Announce Type: new Abstract: Quantization has emerged as a standard technique for accelerating inference for generative models by enabling faster low-precision computations and redu

See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model

Model ReleasesDGX agent

arXiv:2605.11817v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown remarkable promise in robotics manipulation, yet their high computational cost hinders real-time deploy

Self-Supervised Laplace Approximation for Bayesian Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2605.12208v1 Announce Type: cross Abstract: Approximate Bayesian inference typically revolves around computing the posterior parameter distribution. In practice, however, the main object of inte

SEMIR: Semantic Minor-Induced Representation Learning on Graphs for Visual Segmentation

Model ReleasesDGX agent

arXiv:2605.12389v1 Announce Type: new Abstract: Segmenting small and sparse structures in large-scale images is fundamentally constrained by voxel-level, lattice-bound computation and extreme class im

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes

Model ReleasesDGX agent

arXiv:2605.11680v1 Announce Type: new Abstract: We introduce ShapeCodeBench, a synthetic benchmark for perception-to-program reconstruction: given a rendered raster image, a model must emit an executa

SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces

Model ReleasesDGX agent

arXiv:2605.12015v1 Announce Type: cross Abstract: Reusable skills are becoming a common interface for extending large language model agents, packaging procedural guidance with access to files, tools,

Slicing and Dicing: Configuring Optimal Mixtures of Experts

Model ReleasesDGX agent

arXiv:2605.11689v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures have become standard in large language models, yet many of their core design choices - expert count, granularit

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval…

Model ReleasesDGX agent

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the harness stops rationing. the

SOAR: Regression-based LiDAR Relocalization for UAVs

Model ReleasesDGX agent

arXiv:2602.13267v3 Announce Type: replace Abstract: Regression-based LiDAR relocalization has recently emerged as a promising solution for high-precision positioning in GNSS-denied environments. Howev

Solve the Loop: Attractor Models for Language and Reasoning

Model ReleasesDGX agent

arXiv:2605.12466v1 Announce Type: cross Abstract: Looped Transformers offer a promising alternative to purely feed-forward computation by iteratively refining latent representations, improving languag

Sources: Google plans to announce a new Gemini model at its I/O conference next week; the model will land roughly in the class of GPT-5.5, but short of Mythos (Alex Heath/Sources)

Model ReleasesDGX agent

Alex Heath / Sources: Sources: Google plans to announce a new Gemini model at its I/O conference next week; the model will land roughly in the class of GPT-5.5, but short of Mythos — Sources say that

Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos (Bloomberg)

Model ReleasesDGX agent

Bloomberg: Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos — French artificial intelligence s

Sparsity-Constraint Optimization via Splicing Iteration

Model ReleasesDGX agent

arXiv:2406.12017v2 Announce Type: replace-cross Abstract: Sparsity-constrained optimization underlies many problems in signal processing, statistics, and machine learning. State-of-the-art hard-thresh

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

Model ReleasesDGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

Model ReleasesDGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

Model ReleasesDGX agent

arXiv:2602.15620v4 Announce Type: replace Abstract: Reinforcement Learning (RL) has significantly improved large language model reasoning, but existing RL fine-tuning methods rely heavily on heuristic

Starship launch next week!

Model ReleasesDGX agent

Starship launch next week! Starship’s twelfth flight test will debut the next generation Starship and Super Heavy vehicles, powered by the next evolution of the Raptor engine and launching from a newl

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent …

Model ReleasesDGX agent

Starting June 15, paid Claude plans can claim a dedicated monthly credit for programmatic usage. The credit covers usage of: - Claude Agent SDK - claude -p - Claude Code GitHub Actions - Third-party a

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.11922v1 Announce Type: cross Abstract: Existing code reasoning methods primarily supervise final code outputs, ignoring intermediate states, often leading to reward hacking where correct an

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

Model ReleasesDGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

Model ReleasesDGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

Model ReleasesDGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning

Model ReleasesDGX agent

arXiv:2605.12179v1 Announce Type: new Abstract: Recent advancements in video-audio joint generation have achieved remarkable success in semantic correspondence. However, achieving precise temporal syn

Targeted Neuron Modulation via Contrastive Pair Search

Model ReleasesDGX agent

arXiv:2605.12290v1 Announce Type: new Abstract: Language models are instruction-tuned to refuse harmful requests, but the mechanisms underlying this behavior remain poorly understood. Popular steering

TB-AVA: Text as a Semantic Bridge for Audio-Visual Parameter Efficient Finetuning

Model ReleasesDGX agent

arXiv:2605.11572v1 Announce Type: new Abstract: Audio-visual understanding requires effective alignment between heterogeneous modalities, yet cross-modal correspondence remains challenging when tempor

Test-Time Compute for Dense Retrieval: Agentic Program Generation with Frozen Embedding Models

Model ReleasesDGX agent

arXiv:2605.11374v1 Announce Type: cross Abstract: Test-time compute is widely believed to benefit only large reasoning models. We show it also helps small embedding models. Most modern embedding check

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus…

Model ReleasesDGX agent

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus is now FREE for a limited time on Nous Portal — give it a t

The power of LLMs on your data, more than two orders of magnitude faster and cheaper

Model ReleasesDGX agent

Databases have introduced new AI-powered SQL functions which take natural language instructions as input and are evaluated using LLMs. They leverage the power of LLMs to answer new kinds of queries: W

The Price of Proportional Representation in Temporal Voting

Model ReleasesDGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

Model ReleasesDGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

Model ReleasesDGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

Model ReleasesDGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

This jackass fought hard to remain NASA Administrator.

Model ReleasesDGX agent

This jackass fought hard to remain NASA Administrator. SCOOP: I obtained a pitch deck in which the entity that paid for Transportation Sec’y Duffy’s new reality show outlined different partner levels

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

Model ReleasesDGX agent

arXiv:2605.11574v1 Announce Type: new Abstract: The literature on how large language models handle conflict between their training knowledge and a contradicting document presents a persistent empirica

TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others (Aisha Malik/TechCrunch)

Model ReleasesDGX agent

Aisha Malik / TechCrunch: TikTok launches TikTok GO in the US for users to book hotels, attractions, and experiences directly in the app, partnering with Booking.com, Expedia, and others — TikTok anno

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

Model ReleasesDGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

Model ReleasesDGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

Model ReleasesDGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

Model ReleasesDGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

Model ReleasesDGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

Unlocking UML Class Diagram Understanding in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

Model ReleasesDGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

Model ReleasesDGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

Very Efficient Listwise Multimodal Reranking for Long Documents

Model ReleasesDGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

Model ReleasesDGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

Model ReleasesDGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability …

Model ReleasesDGX agent

We built SmithDB: the database purpose built for agent observability workloads that now powers many parts of LangSmith. Agent observability presents a challenging data problem. Agent traces can contai

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a…

Model ReleasesDGX agent

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a human would never try on another human. In one run, GPT-5.1

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

Model ReleasesDGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

What Does It Mean for a Medical AI System to Be Right?

Model ReleasesDGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

Model ReleasesDGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

← Previous
1…257258259260261…377
Next →