AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
5 May 2026

Towards Lightest Low-Light Image Enhancement Architecture for Mobile Devices

Model ReleasesDGX agent

arXiv:2507.04277v2 Announce Type: replace Abstract: Real-time low-light image enhancement on mobile and embedded devices requires models that balance visual quality and computational efficiency. Exist

Towards Visual Query Localization in the 3D World

Model ReleasesDGX agent

arXiv:2605.01498v1 Announce Type: new Abstract: Visual query localization (VQL) aims to predict the spatio-temporal response of the most recent occurrence in a sequence given a query. Currently, most

TRACED: In vivo imaging of extracellular intrinsic diffusivity, tortuosity, cell size distribution and cell density in human glioma patients

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.02615v1 Announce Type: cross Abstract: The lack of analytical models describing diffusion time dependence at intermediate time scales in complex tissue microstructure limits the accurate qu

Training-Free Time Series Classification via In-Context Reasoning with LLM Agents

Model ReleasesDGX agent

arXiv:2510.05950v2 Announce Type: replace Abstract: Time series classification (TSC) spans diverse application scenarios, yet labeled data are often scarce, making task-specific training costly and in

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

Model ReleasesDGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

Triple Spectral Fusion for Sensor-based Human Activity Recognition

Model ReleasesDGX agent

arXiv:2605.02743v1 Announce Type: cross Abstract: The field of sensor-based human activity recognition (HAR) mainly uses posture, motion and context data of Inertial Measurement Units (IMUs) to identi

Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Environments

Model ReleasesDGX agent

arXiv:2510.04142v2 Announce Type: replace Abstract: This paper identifies a critical yet underexplored challenge in reasoning alignment from multiple multi-modal large language models (MLLMs): In non-

TwistNet-2D: Learning Second-Order Channel Interactions via Spiral Twisting for Texture Recognition

Model ReleasesDGX agent

arXiv:2602.07262v3 Announce Type: replace Abstract: Second-order feature statistics are central to texture recognition, yet existing mechanisms exhibit a structural tension: bilinear pooling and Gram

Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video

Model ReleasesDGX agent

arXiv:2605.01512v1 Announce Type: new Abstract: Grounding traffic accidents in real CCTV footage is a rare-event problem where training on labeled accident video is often prohibited, yet accurate join

Understanding Emergent Misalignment via Feature Superposition Geometry

Model ReleasesDGX agent

arXiv:2605.00842v1 Announce Type: cross Abstract: Emergent misalignment, where fine-tuning on narrow, non-harmful tasks induces harmful behaviors, poses a key challenge for AI safety in LLMs. Despite

Understanding the Performance Plateau in Text-to-Video Retrieval: A Comprehensive Empirical and Linguistic Analysis

Model ReleasesDGX agent

arXiv:2605.00826v1 Announce Type: cross Abstract: Text-to-video retrieval enables users to find relevant video content using natural language queries, a task that has grown increasingly important with

Unlocking large scale AI training networks with MRC (Multipath Reliable Connection)

Model ReleasesDGX agent

MRC (Multipath Reliable Connection) is OpenAI's networking technology designed to enable efficient large-scale AI training by improving communication reliability and performance across distributed sup

Unsupervised full-field Bayesian inference of orthotropic hyperelasticity from a single biaxial test: a myocardial case study

Model ReleasesDGX agent

arXiv:2510.09498v3 Announce Type: replace-cross Abstract: Cardiac muscle tissue exhibits highly non-linear hyperelastic and orthotropic material behavior during passive deformation. Traditional consti

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation

Model ReleasesDGX agent

arXiv:2605.01517v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) animation generation is pivotal for professional design due to their structural editability and resolution independence.

Variational Matrix-Learning Fourier Networks for Parametric Multiphysics Surrogates

Model ReleasesDGX agent

arXiv:2605.02280v1 Announce Type: new Abstract: Multiphysics simulation is critical for system-technology co-optimization (STCO) in chiplet-based design, but repeated finite-element solutions of PDE-g

VeRO: An Evaluation Harness for Agents to Optimize Agents

Model ReleasesDGX agent

arXiv:2602.22480v2 Announce Type: replace-cross Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a target agent through edit-execute-eva

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.01662v1 Announce Type: new Abstract: Large vision-language models (VLMs) have advanced multimodal tasks such as video question answering (QA). However, VLMs face the challenge of selecting

VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition

Model ReleasesDGX agent

arXiv:2605.02834v1 Announce Type: new Abstract: Videos are unique in their ability to capture actions which transcend multiple frames. Accordingly, for many years action recognition was the quintessen

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

VISTA: Video Interaction Spatio-Temporal Analysis Benchmark

Model ReleasesDGX agent

arXiv:2605.01391v1 Announce Type: new Abstract: Existing benchmarks for Vision-Language Models (VLMs) primarily evaluate spatio-temporal understanding on simple single-action videos, closed attribute

Visual Implicit Autoregressive Modeling

Model ReleasesDGX agent

arXiv:2605.01220v1 Announce Type: new Abstract: Visual Autoregressive Modeling (VAR) based on next-scale prediction achieves strong generation quality, but their explicit deep stacks fix the amount of

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

Model ReleasesDGX agent

arXiv:2605.02735v1 Announce Type: new Abstract: Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evide

Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms

Model ReleasesDGX agent

arXiv:2603.14535v2 Announce Type: replace Abstract: Reinforcement learning has proven its power on various occasions. However, its performance is not always guaranteed when system dynamics change. Ins

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model

Model ReleasesDGX agent

arXiv:2605.01194v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities and generalization in embodied manipulation. However, their decision-makin

Watermarking LLM Agent Trajectories

Model ReleasesDGX agent

arXiv:2602.18700v2 Announce Type: replace-cross Abstract: LLM agents rely heavily on high-quality trajectory data to guide their problem-solving behaviors, yet producing such data requires substantial

we have very efficient models, especially for their capability level happy codexing

Model ReleasesDGX agent

we have very efficient models, especially for their capability level happy codexing yo, i'm actually worried. codex limits are genuinely insane so it's sus af .. i feel this is an intentional move for

we shipped gpt-5.5 instant today to chat; it's rolling out over the next couple days to everyone. for this model, we focused on factuality, …

Model ReleasesDGX agent

we shipped gpt-5.5 instant today to chat; it's rolling out over the next couple days to everyone. for this model, we focused on factuality, crushing hacks, and improving the baseline intelligence. 5.5

We’re also improving memory and personalization. ChatGPT can now better use context from saved memories, past chats, files, and connected Gm…

Model ReleasesDGX agent

We’re also improving memory and personalization. ChatGPT can now better use context from saved memories, past chats, files, and connected Gmail accounts to give more personalized responses. Memory sou

What Single-Prompt Accuracy Misses: A Multi-Variant Reliability Audit of Language Models

Model ReleasesDGX agent

arXiv:2605.02038v1 Announce Type: new Abstract: Single-prompt accuracy is the dominant way to benchmark language models, but it can miss reliability failures that matter. We evaluate a 15-model open-w

'What's the SemiAnalysis reader gender ratio?' Oh it's at least 90% men, so more 9's than the Claude API

Model ReleasesDGX agent

SemiAnalysis, a semiconductor industry analysis publication, has a readership that is approximately 90% male according to founder Dylan Patel. This gender skew toward male readers is notably higher th

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

Model ReleasesDGX agent

arXiv:2605.02782v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems remain brittle on dysarthric and other atypical speech. Recent audio-language models raise the possibility

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

Model ReleasesDGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.00911v1 Announce Type: new Abstract: Industrial Retrieval-Augmented Generation (RAG) systems depend on optical character recognition (OCR) to transform visual documents into text. Existing

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering

Model ReleasesDGX agent

arXiv:2601.19827v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) extends large language models (LLMs) beyond parametric knowledge, yet it is unclear when iterative retrieval-re

When Less Is More: Simplicity Beats Complexity for Physics-Constrained InSAR Phase Unwrapping

Model ReleasesDGX agent

arXiv:2605.00896v1 Announce Type: new Abstract: Operational phase unwrapping is the primary computational bottleneck in InSAR-based volcanic and seismic monitoring. We challenge the industry trend of

When RL Meets Adaptive Speculative Training: A Unified Training-Serving System

Model ReleasesDGX agent

arXiv:2602.06932v3 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating spe

WILD SAM: A Simulated-and-Real Data Augmentation for Autonomous Driving Perception under Challenging Weather

Model ReleasesDGX agent

arXiv:2605.01081v1 Announce Type: new Abstract: The performance of state-of-the-art object detectors degrades significantly under adverse weather, causing a safety-critical domain shift problem for au

WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild

Model ReleasesDGX agent

arXiv:2605.01018v1 Announce Type: new Abstract: Using multimodal foundation models to analyze table images is a high-value yet challenging application in consumer and enterprise scenarios. Despite its

WSO2 launches Agent Manager to help enterprises tame AI agent sprawl

Model ReleasesDGX agent

Open-source technology provider WS02 LLC today announced the launch of WSO2 Agent Manager, an open control plane for artificial intelligence agents that gives enterprises a unified way to identify, go

X2SAM: Any Segmentation in Images and Videos

Model ReleasesDGX agent

arXiv:2605.00891v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong image-level visual understanding and reasoning, yet their pixel-level perception acros

Zero-Shot Interpretable Image Steganalysis for Invertible Image Hiding

Model ReleasesDGX agent

arXiv:2605.01331v1 Announce Type: new Abstract: Image steganalysis, which aims at detecting secret information concealed within images, has become a critical countermeasure for assessing the security

ZNO: Stable Rational Neural Operators in the Z-Domain for Discrete-Time Dynamic

Model ReleasesDGX agent

arXiv:2605.02356v1 Announce Type: new Abstract: We introduce the Z-Domain Neural Operator (ZNO), a causal neural operator whose layers are stable low-rank multiple-input multiple-output (MIMO) rationa

4 May 2026

A challenge with AI regulation and vetting is how bad our benchmarks of AI model performance and risks are. There is no benchmark for risks …

Model ReleasesDGX agent

A challenge with AI regulation and vetting is how bad our benchmarks of AI model performance and risks are. There is no benchmark for risks and red-teaming requires experiments from dedicated speciali

A Dirac-Frenkel-Onsager principle: Instantaneous residual minimization with gauge momentum for nonlinear parametrizations of PDE solutions

Model ReleasesDGX agent

arXiv:2605.00284v1 Announce Type: new Abstract: Dirac-Frenkel instantaneous residual minimization evolves nonlinear parametrizations of PDE solutions in time, but ill-conditioning can render the param

A11y-Compressor: A Framework for Enhancing the Efficiency of GUI Agent Observations through Visual Context Reconstruction and Redundancy Reduction

Model ReleasesDGX agent

arXiv:2605.00551v1 Announce Type: new Abstract: AI agents that interact with graphical user interfaces (GUIs) require effective observation representations for reliable grounding. The accessibility tr

Adaptive Dual-Teacher Distillation with Subnetwork Rectification for Bridging Semantic Gaps in Black-Box Domain Adaptation

Model ReleasesDGX agent

arXiv:2603.22908v3 Announce Type: replace Abstract: Assuming that neither source data nor source model parameters are accessible, black-box domain adaptation (BBDA) represents a highly practical yet c

Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?

Model ReleasesDGX agent

arXiv:2603.25719v2 Announce Type: replace-cross Abstract: We present an empirical study of how far general-purpose coding agents -- without hardware-specific training -- can optimize hardware designs

AgentFloor: How Far Up the tool use Ladder Can Small Open-Weight Models Go?

Model ReleasesDGX agent

arXiv:2605.00334v1 Announce Type: cross Abstract: Production agentic systems make many model calls per user request, and most of those calls are short, structured, and routine. This raises a practical

AGoQ: Activation and Gradient Quantization for Memory-Efficient Distributed Training of LLMs

Model ReleasesDGX agent

arXiv:2605.00539v1 Announce Type: new Abstract: Quantization is a key method for reducing the GPU memory requirement of training large language models (LLMs). Yet, current approaches are ineffective f

Alethia: A Foundational Encoder for Voice Deepfakes

Model ReleasesDGX agent

arXiv:2605.00251v1 Announce Type: cross Abstract: Existing voice deepfake detection and localization models rely heavily on representations extracted from speech foundation models (SFMs). However, dow

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

Model ReleasesDGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

April 2026 newsletter

Model ReleasesDGX agent

I just sent out the April edition of my sponsors-only monthly newsletter. If you are a sponsor (or if you start a sponsorship now) you can access it here. In this month's newsletter: Opus 4.7 and GPT-

At any point in time, you can safely resume to using Anthropic's models: ollama launch claude-desktop --restore

Model ReleasesDGX agent

This post discusses Ollama's functionality for resuming work with Anthropic's Claude models, indicating that users can safely restore previous sessions or states using a command-line interface (`ollam

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction

Model ReleasesDGX agent

arXiv:2603.15949v3 Announce Type: replace Abstract: Large Language Models have demonstrated strong multilingual fluency, yet fluency alone does not guarantee socially appropriate language use. In high

Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs

Model ReleasesDGX agent

arXiv:2605.00674v1 Announce Type: new Abstract: Large language models (LLMs) are becoming increasingly capable mathematical collaborators, but static benchmarks are no longer sufficient for evaluating

Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2605.00408v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has demonstrated impressive real-time rendering performance, its efficacy remains constrained by a reliance on heuris

Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration

Model ReleasesDGX agent

arXiv:2605.00310v1 Announce Type: new Abstract: Super-resolution (SR) techniques have made major advances in reconstructing high-resolution images from low-resolution inputs. The increased resolution

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

Model ReleasesDGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs

Model ReleasesDGX agent

arXiv:2407.10853v5 Announce Type: replace Abstract: Bias and fairness risks in Large Language Models (LLMs) vary substantially across deployment contexts, yet existing approaches lack systematic guida

Can Coding Agents Reproduce Findings in Computational Materials Science?

Model ReleasesDGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

← Previous
1…290291292293294…376
Next →