AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

DGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

model-releasesarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

DGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

DGX agent

arXiv:2602.15620v4 Announce Type: replace Abstract: Reinforcement Learning (RL) has significantly improved large language model reasoning, but existing RL fine-tuning methods rely heavily on heuristic

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning

DGX agent

arXiv:2605.11922v1 Announce Type: cross Abstract: Existing code reasoning methods primarily supervise final code outputs, ignoring intermediate states, often leading to reward hacking where correct an

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

DGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

DGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

DGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning

DGX agent

arXiv:2605.12179v1 Announce Type: new Abstract: Recent advancements in video-audio joint generation have achieved remarkable success in semantic correspondence. However, achieving precise temporal syn

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Targeted Neuron Modulation via Contrastive Pair Search

DGX agent

arXiv:2605.12290v1 Announce Type: new Abstract: Language models are instruction-tuned to refuse harmful requests, but the mechanisms underlying this behavior remain poorly understood. Popular steering

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

TB-AVA: Text as a Semantic Bridge for Audio-Visual Parameter Efficient Finetuning

DGX agent

arXiv:2605.11572v1 Announce Type: new Abstract: Audio-visual understanding requires effective alignment between heterogeneous modalities, yet cross-modal correspondence remains challenging when tempor

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Test-Time Compute for Dense Retrieval: Agentic Program Generation with Frozen Embedding Models

DGX agent

arXiv:2605.11374v1 Announce Type: cross Abstract: Test-time compute is widely believed to benefit only large reasoning models. We show it also helps small embedding models. Most modern embedding check

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

The Price of Proportional Representation in Temporal Voting

DGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

DGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

DGX agent

arXiv:2605.11574v1 Announce Type: new Abstract: The literature on how large language models handle conflict between their training knowledge and a contradicting document presents a persistent empirica

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

DGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

DGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

DGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

DGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

What Does It Mean for a Medical AI System to Be Right?

DGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

DGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

YFPO: A Preliminary Study of Yoked Feature Preference Optimization with Neuron-Guided Rewards for Mathematical Reasoning

DGX agent

arXiv:2605.11906v1 Announce Type: new Abstract: Preference optimization has become an important post-training paradigm for improving the reasoning abilities of large language models. Existing methods

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

DGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

DGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

DGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

DGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

DGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A new initialisation to Control Gradients in Sinusoidal Neural network

DGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

DGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Stability Benchmark of Generative Regularizers for Inverse Problems

DGX agent

arXiv:2605.10076v1 Announce Type: cross Abstract: Generative (diffusion) priors demonstrate remarkable performance in addressing inverse problems in imaging. Yet, for scientific and medical imaging, i

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

A Unified Representation of Neural Networks Architectures

DGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

DGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims

DGX agent

arXiv:2605.10575v1 Announce Type: cross Abstract: Safe fine-tuning defenses are often endorsed on the basis of a held-out gap reduction, but the same reduction can come from sampling noise, subject ar

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Action-Guided Attention for Video Action Anticipation

DGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models

DGX agent

arXiv:2605.08567v1 Announce Type: new Abstract: Action-conditioned world models (ACWMs) have shown strong promise for video prediction and decision-making. However, existing benchmarks are largely res

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

DGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

DGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

model-releasesarxiv-cs-lg
12 May 2026
← Previous
1…249250251252253…361
Next →