AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

DGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

model-releasesarxiv-cs-ai
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MIRROR: A Hierarchical Benchmark for Metacognitive Calibration in Large Language Models

DGX agent

arXiv:2604.19809v1 Announce Type: new Abstract: We introduce MIRROR, a benchmark comprising eight experiments across four metacognitive levels that evaluates whether large language models can use self

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

DGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation

DGX agent

arXiv:2604.20366v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit powerful generative capabilities but frequently produce hallucinations that compromise output reliability.

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Mitigating Prompt-Induced Cognitive Biases in General-Purpose AI for Software Engineering

DGX agent

arXiv:2604.16756v2 Announce Type: replace-cross Abstract: Prompt-induced cognitive biases are changes in a general-purpose AI (GPAI) system's decisions caused solely by biased wording in the input (e.

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MixLLM: LLM Quantization with Global Mixed-precision between Output-features and Highly-efficient System Design

DGX agent

arXiv:2412.14590v2 Announce Type: replace Abstract: Quantization has become one of the most effective methodologies to compress LLMs into smaller size. However, the existing quantization solutions sti

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

MLG-Stereo: ViT Based Stereo Matching with Multi-Stage Local-Global Enhancement

DGX agent

arXiv:2604.20393v1 Announce Type: new Abstract: With the development of deep learning, ViT-based stereo matching methods have made significant progress due to their remarkable robustness and zero-shot

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Model Capability Assessment and Safeguards for Biological Weaponization

DGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MSLAU-Net: A Hybrid CNN-Transformer Network for Medical Image Segmentation

DGX agent

arXiv:2505.18823v2 Announce Type: replace Abstract: Accurate medical image segmentation allows for the precise delineation of anatomical structures and pathological regions, which is essential for tre

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Mythos and the Unverified Cage: Z3-Based Pre-Deployment Verification for Frontier-Model Sandbox Infrastructure

DGX agent

arXiv:2604.20496v1 Announce Type: cross Abstract: The April 2026 Claude Mythos sandbox escape exposed a critical weakness in frontier AI containment: the infrastructure surrounding advanced models rem

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

'Newspaper Eat' Means 'Not Tasty': A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews

DGX agent

arXiv:2601.19932v2 Announce Type: replace Abstract: Coded language is an important part of human communication. It refers to cases where users intentionally encode meaning so that the surface text dif

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

DGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

On Bayesian Softmax-Gated Mixture-of-Experts Models

DGX agent

arXiv:2604.20551v1 Announce Type: cross Abstract: Mixture-of-experts models provide a flexible framework for learning complex probabilistic input-output relationships by combining multiple expert mode

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL

DGX agent

arXiv:2506.20904v2 Announce Type: replace Abstract: We study offline reinforcement learning in average-reward MDPs, which presents increased challenges from the perspectives of distribution shift and

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Option Pricing on Noisy Intermediate-Scale Quantum Computers: A Quantum Neural Network Approach

DGX agent

arXiv:2604.19832v1 Announce Type: cross Abstract: In a global derivatives market with notional values in the hundreds of trillions of dollars, the accuracy and efficiency of pricing models are of fund

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

OVPD: A Virtual-Physical Fusion Testing Dataset of OnSite Auton-omous Driving Challenge

DGX agent

arXiv:2604.20423v1 Announce Type: new Abstract: The rapid iteration of autonomous driving algorithms has created a growing demand for high-fidelity, replayable, and diagnosable testing data. However,

model-releasesarxiv-cs-ro
23 Apr 2026
Model Releases

Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL

DGX agent

arXiv:2604.20835v1 Announce Type: new Abstract: Modern language models demonstrate impressive coding capabilities in common programming languages (PLs), such as C++ and Python, but their performance i

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Peer-Preservation in Frontier Models

DGX agent

arXiv:2604.19784v1 Announce Type: cross Abstract: Recently, it has been found that frontier AI models can resist their own shutdown, a behavior known as self-preservation. We extend this concept to th

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

PipeMFL-240K: A Large-scale Dataset and Benchmark for Object Detection in Pipeline Magnetic Flux Leakage Imaging

DGX agent

arXiv:2602.07044v2 Announce Type: replace-cross Abstract: Pipeline integrity is critical to industrial safety and environmental protection, with Magnetic Flux Leakage (MFL) detection being a primary n

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

PLR: Plackett-Luce for Reordering In-Context Learning Examples

DGX agent

arXiv:2603.21373v2 Announce Type: replace-cross Abstract: In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Amon

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance

DGX agent

arXiv:2604.20834v1 Announce Type: new Abstract: Recent advances in Vision-Language-Action (VLA) models have opened new avenues for robot manipulation, yet existing methods exhibit limited efficiency a

model-releasesarxiv-cs-ro
23 Apr 2026
Model Releases

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

DGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

DGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

QuanForge: A Mutation Testing Framework for Quantum Neural Networks

DGX agent

arXiv:2604.20706v1 Announce Type: cross Abstract: With the growing synergy between deep learning and quantum computing, Quantum Neural Networks (QNNs) have emerged as a promising paradigm by leveragin

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

QuantaAlpha: An Evolutionary Framework for LLM-Driven Alpha Mining

DGX agent

arXiv:2602.07085v2 Announce Type: replace-cross Abstract: Financial markets are noisy and non-stationary, making alpha mining highly sensitive to noise in backtesting results and sudden market regime

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

RareSpot+: A Benchmark, Model, and Active Learning Framework for Small and Rare Wildlife in Aerial Imagery

DGX agent

arXiv:2604.20000v1 Announce Type: new Abstract: Automated wildlife monitoring from aerial imagery is vital for conservation but remains limited by two persistent challenges: the difficulty of detectin

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Rashomon Sets and Model Multiplicity in Federated Learning

DGX agent

arXiv:2602.09520v2 Announce Type: replace Abstract: The Rashomon set captures the collection of models that achieve near-identical empirical performance yet may differ substantially in their decision

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability

DGX agent

arXiv:2508.07050v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

RefAerial: A Benchmark and Approach for Referring Detection in Aerial Images

DGX agent

arXiv:2604.20543v1 Announce Type: new Abstract: Referring detection refers to locate the target referred by natural languages, which has recently attracted growing research interests. However, existin

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Relative Entropy Estimation in Function Space: Theory and Applications to Trajectory Inference

DGX agent

arXiv:2604.20775v1 Announce Type: new Abstract: Trajectory Inference (TI) seeks to recover latent dynamical processes from snapshot data, where only independent samples from time-indexed marginals are

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering

DGX agent

arXiv:2604.20738v1 Announce Type: new Abstract: We introduce a benchmark dataset for question answering and translation in bilingual Latin and English settings, containing about 7,800 question-answer

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Rethinking Where to Edit: Task-Aware Localization for Instruction-Based Image Editing

DGX agent

arXiv:2604.20258v1 Announce Type: new Abstract: Instruction-based image editing (IIE) aims to modify images according to textual instructions while preserving irrelevant content. Despite recent advanc

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

retinalysis-vascx: An explainable software toolbox for the extraction of retinal vascular biomarkers

DGX agent

arXiv:2602.08580v2 Announce Type: replace-cross Abstract: Automatic extraction of retinal vascular biomarkers from color fundus images (CFI) is crucial for large-scale studies of the retinal vasculatu

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

REVNET: Rotation-Equivariant Point Cloud Completion via Vector Neuron Anchor Transformer

DGX agent

arXiv:2601.08558v2 Announce Type: replace Abstract: Incomplete point clouds captured by 3D sensors often result in the loss of both geometric and semantic information. Most existing point cloud comple

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

RExBench: Can coding agents autonomously implement AI research extensions?

DGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

RSRCC: A Remote Sensing Regional Change Comprehension Benchmark Constructed via Retrieval-Augmented Best-of-N Ranking

DGX agent

arXiv:2604.20623v1 Announce Type: cross Abstract: Traditional change detection identifies where changes occur, but does not explain what changed in natural language. Existing remote sensing change cap

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Scaling Self-Play with Self-Guidance

DGX agent

arXiv:2604.20209v1 Announce Type: new Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both i

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation

DGX agent

arXiv:2604.19907v1 Announce Type: new Abstract: Recent agentic frameworks for 3D scene synthesis have advanced realism and diversity by integrating heterogeneous generation and editing tools. These to

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

SciCoQA: Quality Assurance for Scientific Paper--Code Alignment

DGX agent

arXiv:2601.12910v3 Announce Type: replace-cross Abstract: Discrepancies between scientific papers and their code undermine reproducibility, a concern that grows as automated research agents scale scie

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images

DGX agent

arXiv:2512.08730v2 Announce Type: replace Abstract: Most existing methods for training-free open-vocabulary semantic segmentation are based on CLIP. While these approaches have made progress, they oft

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Self-Aware Vector Embeddings for Retrieval-Augmented Generation: A Neuroscience-Inspired Framework for Temporal, Confidence-Weighted, and Relational Knowledge

DGX agent

arXiv:2604.20598v1 Announce Type: cross Abstract: Modern retrieval-augmented generation (RAG) systems treat vector embeddings as static, context-free artifacts: an embedding has no notion of when it w

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

DGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Self-Describing Structured Data with Dual-Layer Guidance: A Lightweight Alternative to RAG for Precision Retrieval in Large-Scale LLM Knowledge Navigation

DGX agent

arXiv:2604.19777v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit a well-documented positional bias when processing long input contexts: information in the middle of a context win

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Semi-Supervised Flow Matching for Mosaiced and Panchromatic Fusion Imaging

DGX agent

arXiv:2604.20128v1 Announce Type: new Abstract: Fusing a low resolution (LR) mosaiced hyperspectral image (HSI) with a high resolution (HR) panchromatic (PAN) image offers a promising avenue for video

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze

DGX agent

arXiv:2604.19888v1 Announce Type: new Abstract: Driver gaze estimation is essential for understanding the driver's situational awareness of surrounding traffic. Existing gaze estimation models use dri

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

DGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

DGX agent

arXiv:2604.20087v1 Announce Type: new Abstract: Skills have become the de facto way to enable LLM agents to perform complex real-world tasks with customized instructions, workflows, and tools, but how

model-releasesarxiv-cs-cl
23 Apr 2026
← Previous
1…306307308309310…357
Next →