AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

PHIDA: Persistence-Guided Node-to-Cluster Mapping for Online Clustering

DGX agent

arXiv:2605.08673v1 Announce Type: new Abstract: Online clustering methods that adaptively create and update nodes as data arrive often make node learning explicit, whereas the mapping from the learned

model-releasesarxiv-cs-lg
12 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Phoenix-VL 1.5 Medium Technical Report

DGX agent

arXiv:2605.10391v1 Announce Type: cross Abstract: We introduce Phoenix-VL 1.5 Medium, a 123B-parameter natively multimodal and multilingual foundation model, adapted to regional languages and the Sing

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PhyGround: Benchmarking Physical Reasoning in Generative World Models

DGX agent

arXiv:2605.10806v1 Announce Type: cross Abstract: Generative world models are increasingly used for video generation, where learned simulators are expected to capture the physical rules that govern re

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PINS: Proximal Iterations with Sparse Newton and Sinkhorn for Optimal Transport

DGX agent

arXiv:2502.03749v2 Announce Type: replace Abstract: Optimal transport (OT) is a widely used tool in machine learning, but computing high-accuracy solutions for large instances remains costly. Entropic

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes

DGX agent

arXiv:2602.00593v2 Announce Type: replace Abstract: Despite progress on general tasks, vision-language models (VLMs) still struggle with challenges that demand both fine-grained visual grounding and e

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

PlantMarkerBench: A Multi-Species Benchmark for Evidence-Grounded Plant Marker Reasoning

DGX agent

arXiv:2605.10032v1 Announce Type: new Abstract: Cell-type-specific marker genes are fundamental to plant biology, yet existing resources primarily rely on curated databases or high-throughput studies

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

PolarVSR: A Unified Framework and Benchmark for Continuous Space-Time Polarization Video Reconstruction

DGX agent

arXiv:2605.10275v1 Announce Type: new Abstract: Polarimetric imaging captures surface polarization characteristics, such as the Degree of Linear Polarization (DoLP) and the Angle of Polarization (AoP)

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark

DGX agent

arXiv:2410.14702v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) exhibit impressive problem-solving abilities in various domains, but their visual comprehension and abstra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Position: AI Security Policy Should Target Systems, Not Models

DGX agent

arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead

DGX agent

arXiv:2507.23009v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved remarkable results on a range of standardized tests originally designed to assess human cognitive a

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models

DGX agent

arXiv:2605.08800v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) may memorize sensitive cross-modal information during pretraining. However, existing MLLM unlearning benchmar

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning

DGX agent

arXiv:2602.03190v3 Announce Type: replace-cross Abstract: Reinforcement learning algorithms such as group-relative policy optimization (GRPO) have shown strong potential for improving the mathematical

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)

DGX agent

arXiv:2605.09169v1 Announce Type: cross Abstract: A Mamba state-space model trained only for next-step prediction appears to recover Granger-causal structure through a simple readout S = |W_{out} W_{i

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

DGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift

DGX agent

arXiv:2505.19519v3 Announce Type: replace Abstract: Personalizing text-to-image diffusion models involves integrating novel visual concepts from a small set of reference images while retaining the mod

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Preserving Foundational Capabilities in Flow-Matching VLAs through Conservative SFT

DGX agent

arXiv:2605.08879v1 Announce Type: new Abstract: Unconstrained fine-tuning of flow-matching Vision-Language-Action (VLA) models drives dense parameter overwrites, degrading pre-trained capabilities. We

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Pretraining large language models with MXFP4

DGX agent

arXiv:2605.09825v1 Announce Type: cross Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We a

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PRIM: Meta-Learned Bayesian Root Cause Analysis

DGX agent

arXiv:2605.08786v1 Announce Type: new Abstract: Root cause analysis (RCA) in complex systems is challenging due to error propagation across multiple variables, the need for structural causal knowledge

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs

DGX agent

arXiv:2605.10529v1 Announce Type: new Abstract: Biomedical knowledge graphs underwrite drug repurposing and clinical decision support, yet the upstream ontologies they depend on update on independent

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Priming: Hybrid State Space Models From Pre-trained Transformers

DGX agent

arXiv:2605.08301v1 Announce Type: cross Abstract: Hybrid State-Space models combine Attention with recurrent State-Space Model (SSM) layers, balancing eidetic memory from Attention with compressed fad

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Priority-Driven Control and Communication in Decentralized Multi-Agent Systems via Reinforcement Learning

DGX agent

arXiv:2605.10482v1 Announce Type: cross Abstract: Event-triggered control provides a mechanism for avoiding excessive use of constrained communication bandwidth in networked multi-agent systems. Howev

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

DGX agent

arXiv:2605.10614v1 Announce Type: new Abstract: Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Privacy Auditing Synthetic Data Release through Local Likelihood Attacks

DGX agent

arXiv:2508.21146v2 Announce Type: replace Abstract: Auditing the privacy leakage of synthetic data is an important but unresolved problem. Existing privacy auditing frameworks for synthetic data rely

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ProactBench: Beyond What The User Asked For

DGX agent

arXiv:2605.09228v1 Announce Type: cross Abstract: Most LLM benchmarks score how well a model responds to explicit requests. They leave unmeasured a different conversational ability: noticing and actin

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark

DGX agent

arXiv:2509.26574v4 Announce Type: replace Abstract: While large language models (LLMs) with reasoning capabilities are progressing rapidly on high-school math competitions and coding, can they reason

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

DGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Process Matters more than Output for Distinguishing Humans from Machines

DGX agent

arXiv:2605.06524v2 Announce Type: replace Abstract: Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online setti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Product-of-Gaussian-Mixture Diffusion Models for Joint Nonlinear MRI Reconstruction

DGX agent

arXiv:2605.10629v1 Announce Type: new Abstract: Recently, diffusion models have attracted considerable attention for magnetic resonance image reconstruction due to their high sample quality. However,

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

DGX agent

arXiv:2605.10664v1 Announce Type: cross Abstract: Activation steering controls language model behavior by adding directions to internal representations at inference time, but standard residual-stream

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prompt Estimation from Prototypes for Federated Prompt Tuning of Vision Transformers

DGX agent

arXiv:2510.25372v2 Announce Type: replace Abstract: Visual Prompt Tuning (VPT) of pre-trained Vision Transformers (ViTs) has proven highly effective as a parameter-efficient fine-tuning technique for

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

PumpSense: Real-Time Detection and Target Extraction of Crypto Pump-and-Dumps on Telegram

DGX agent

arXiv:2605.09431v1 Announce Type: new Abstract: Cryptocurrency pump-and-dump schemes coordinated via Telegram threaten market integrity. However, existing research addressing this specific threat has

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model

DGX agent

arXiv:2504.12334v2 Announce Type: replace Abstract: Large language models (LLMs) face significant challenges in specialized biomedical tasks due to the inherent complexity of medical reasoning and the

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Quantifying Concentration Phenomena of Mean-Field Transformers in the Low-Temperature Regime

DGX agent

arXiv:2605.10931v1 Announce Type: cross Abstract: Transformers with self-attention modules as their core components have become an integral architecture in modern large language and foundation models.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants

DGX agent

arXiv:2605.09808v1 Announce Type: new Abstract: User simulators are increasingly leveraged to build interactive AI assistants, yet how to measure the quality of these simulators remains an open questi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation

DGX agent

arXiv:2605.08170v1 Announce Type: new Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation proper

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quantum Circuit Simulation of Compartmental Drug Dynamics: Leveraging Variational Algorithms for Nonlinear Mixed-Effects Population Pharmacokinetics

DGX agent

arXiv:2605.09691v1 Announce Type: new Abstract: Population pharmacokinetic/pharmacodynamic (PK/PD) modeling traditionally relies on classical ordinary differential equations to simulate drug dynamics.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quasi-Linear ICA for Motor Unit Decomposition during Dynamic Contractions

DGX agent

arXiv:2406.19581v2 Announce Type: replace-cross Abstract: Decomposing surface electromyography (EMG) into the spike trains of individual motor neurons is a long-standing inverse problem and a key step

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Queryable LoRA: Instruction-Regularized Routing Over Shared Low-Rank Update Atoms

DGX agent

arXiv:2605.08423v1 Announce Type: cross Abstract: We present a data-adaptive method for parameter-efficient fine-tuning of large neural networks. Standard low-rank adaptation methods improve efficienc

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Qwen Goes Brrr: Off-the-Shelf RAG for Ukrainian Multi-Domain Document Understanding

DGX agent

arXiv:2605.10296v1 Announce Type: cross Abstract: We participated in the Fifth UNLP shared task on multi-domain document understanding, where systems must answer Ukrainian multiple-choice questions fr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Qwen-Image-2.0 Technical Report

DGX agent

arXiv:2605.10730v1 Announce Type: new Abstract: We present Qwen-Image-2.0, an omni-capable image generation foundation model that unifies high-fidelity generation and precise image editing within a si

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

R4Det: 4D Radar-Camera Fusion for High-Performance 3D Object Detection

DGX agent

arXiv:2603.11566v2 Announce Type: replace Abstract: 4D radar-camera sensing configuration has gained increasing importance in autonomous driving. However, existing 3D object detection methods that fus

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology

DGX agent

arXiv:2605.10761v1 Announce Type: new Abstract: Cancer screening is a reasoning task. A radiologist observes findings, compares them to prior scans, integrates clinical context, and reaches a diagnost

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

RareCP: Regime-Aware Retrieval for Efficient Conformal Prediction

DGX agent

arXiv:2605.08857v1 Announce Type: new Abstract: Recent advances in uncertainty quantification for time series forecasting show that conformal prediction can provide reliable prediction intervals, yet

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

RDEx-CASK: Cauchy Mutation, Archive, and Stagnation Kick for RDEx-CSOP

DGX agent

arXiv:2605.09652v1 Announce Type: cross Abstract: We extend RDEx-CSOP with 3 changes that target stagnation & late-stage variance, plus minor parameter tuning. The second scale factor in the standard

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Re^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

DGX agent

arXiv:2605.09012v1 Announce Type: new Abstract: Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation

DGX agent

arXiv:2605.10430v1 Announce Type: cross Abstract: Estimating heterogeneous treatment effects with machine learning has attracted substantial attention in both academic research and industrial practice

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

ReaMOT: A Benchmark and Framework for Reasoning-based Multi-Object Tracking

DGX agent

arXiv:2505.20381v4 Announce Type: replace Abstract: Referring Multi-Object Tracking (RMOT) aims to track targets specified by language instructions. However, existing RMOT paradigms heavily rely on ex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage

DGX agent

arXiv:2604.01527v3 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fi

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…258259260261262…361
Next →