AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation

DGX agent

arXiv:2604.13803v1 Announce Type: new Abstract: Vision-language models are increasingly deployed in high-stakes settings, yet their susceptibility to sycophantic manipulation remains poorly understood

model-releasesarxiv-cs-cv
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

DGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs

DGX agent

arXiv:2604.13258v1 Announce Type: new Abstract: Attribution methods seek to explain language model predictions by quantifying the contribution of input tokens to generated outputs. However, most exist

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

DGX agent

arXiv:2604.14032v1 Announce Type: cross Abstract: Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its de

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

HINTBench: Horizon-agent Intrinsic Non-attack Trajectory Benchmark

DGX agent

arXiv:2604.13954v1 Announce Type: new Abstract: Existing agent-safety evaluation has focused mainly on externally induced risks. Yet agents may still enter unsafe trajectories under benign conditions.

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Hybrid Retrieval for COVID-19 Literature: Comparing Rank Fusion and Projection Fusion with Diversity Reranking

DGX agent

arXiv:2604.13728v1 Announce Type: cross Abstract: We present a hybrid retrieval system for COVID-19 scientific literature, evaluated on the TREC-COVID benchmark (171,332 papers, 50 expert queries). Th

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ID and Graph View Contrastive Learning with Multi-View Attention Fusion for Sequential Recommendation

DGX agent

arXiv:2604.14114v1 Announce Type: cross Abstract: Sequential recommendation has become increasingly prominent in both academia and industry, particularly in e-commerce. The primary goal is to extract

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

DGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

DGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Joint Representation Learning and Clustering via Gradient-Based Manifold Optimization

DGX agent

arXiv:2604.13484v1 Announce Type: cross Abstract: Clustering and dimensionality reduction have been crucial topics in machine learning and computer vision. Clustering high-dimensional data has been ch

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

KMMMU: Evaluation of Massive Multi-discipline Multimodal Understanding in Korean Language and Context

DGX agent

arXiv:2604.13058v1 Announce Type: new Abstract: We introduce KMMMU, a native Korean benchmark for evaluating multimodal understanding in Korean cultural and institutional settings. KMMMU contains 3,46

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

DGX agent

arXiv:2604.13226v1 Announce Type: new Abstract: Large Language Models (LLMs) rely heavily on Key-Value (KV) caching to minimize inference latency. However, standard KV caches are context-dependent: re

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification

DGX agent

arXiv:2604.13285v1 Announce Type: new Abstract: Clinical text classification requires choosing between specialized fine-tuned models (BERT variants) and general-purpose large language models (LLMs), y

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Language steering in latent space to mitigate unintended code-switching

DGX agent

arXiv:2510.13849v3 Announce Type: replace Abstract: Multilingual Large Language Models (LLMs) often exhibit hallucinations such as unintended code-switching, reducing reliability in downstream tasks.

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

DGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

DGX agent

arXiv:2604.13713v1 Announce Type: new Abstract: Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memo

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Leveraging LLM-GNN Integration for Open-World Question Answering over Knowledge Graphs

DGX agent

arXiv:2604.13979v1 Announce Type: new Abstract: Open-world Question Answering (OW-QA) over knowledge graphs (KGs) aims to answer questions over incomplete or evolving KGs. Traditional KGQA assumes a c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks

DGX agent

arXiv:2604.13072v1 Announce Type: new Abstract: LLM-based agents are increasingly expected to handle real-world assistant tasks, yet existing benchmarks typically evaluate them under isolated sources

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

DGX agent

arXiv:2604.14140v1 Announce Type: new Abstract: As language models are increasingly deployed for complex autonomous tasks, their ability to reason accurately over longer horizons becomes critical. An

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding

DGX agent

arXiv:2602.20913v2 Announce Type: replace Abstract: This paper addresses the critical and underexplored challenge of long video understanding with low computational budgets. We propose LongVideo-R1, a

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

LoRA-MME: Multi-Model Ensemble of LoRA-Tuned Encoders for Code Comment Classification

DGX agent

arXiv:2603.03959v4 Announce Type: replace-cross Abstract: Code comment classification is a critical task for automated software documentation and analysis. In the context of the NLBSE'26 Tool Competit

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data

DGX agent

arXiv:2604.13066v1 Announce Type: new Abstract: In-context learning has established itself as an important learning paradigm for Large Language Models (LLMs). In this paper, we demonstrate that LLMs c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MAny: Merge Anything for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2604.14016v1 Announce Type: new Abstract: Multimodal Continual Instruction Tuning (MCIT) is essential for sequential task adaptation of Multimodal Large Language Models (MLLMs) but is severely r

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

MDPs with a State Sensing Cost

DGX agent

arXiv:2505.03280v3 Announce Type: replace Abstract: In many practical sequential decision-making problems, tracking the state of the environment incurs a sensing/communication/computation cost. In the

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging

DGX agent

arXiv:2604.13756v1 Announce Type: new Abstract: The potential of Multimodal Large Language Models (MLLMs) in domain of medical imaging raise the demands of systematic and rigorous evaluation framework

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

DGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates

DGX agent

arXiv:2512.04844v2 Announce Type: replace Abstract: Expanding the linguistic diversity of instruct large language models (LLMs) is crucial for global accessibility but is often hindered by the relianc

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MM-Doc-R1: Training Agents for Long Document Visual Question Answering through Multi-turn Reinforcement Learning

DGX agent

arXiv:2604.13579v1 Announce Type: new Abstract: Conventional Retrieval-Augmented Generation (RAG) systems often struggle with complex multi-hop queries over long documents due to their single-pass ret

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals

DGX agent

arXiv:2604.13897v1 Announce Type: new Abstract: We present an open Molecular Crystal (MC) database of Machine-Learned Interatomic Potentials (MLIP) called MolCryst-MLIPs. The first release comprises f

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models

DGX agent

arXiv:2604.13287v1 Announce Type: new Abstract: Weight pruning is a common technique for compressing large neural networks. We focus on the challenging post-training one-shot setting, where a pre-trai

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Mosaic: An Extensible Framework for Composing Rule-Based and Learned Motion Planners

DGX agent

arXiv:2604.13853v1 Announce Type: new Abstract: Safe and explainable motion planning remains a central challenge in autonomous driving. While rule-based planners offer predictable and explainable beha

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models

DGX agent

arXiv:2505.07591v2 Announce Type: replace Abstract: Instruction following refers to the ability of large language models (LLMs) to generate outputs that satisfy all specified constraints. Existing res

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

DGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Olfactory pursuit: catching a moving odor source in complex flows

DGX agent

arXiv:2604.13121v1 Announce Type: new Abstract: Locating and intercepting a moving target from possibly delayed, intermittent sensory signals is a paradigmatic problem in decision-making under uncerta

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Online learning with noisy side observations

DGX agent

arXiv:2604.13740v1 Announce Type: new Abstract: We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback abo

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Optimization with SpotOptim

DGX agent

arXiv:2604.13672v1 Announce Type: new Abstract: The `spotoptim` package implements surrogate-model-based optimization of expensive black-box functions in Python. Building on two decades of Sequential

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Ordinary Least Squares is a Special Case of Transformer

DGX agent

arXiv:2604.13656v1 Announce Type: new Abstract: The statistical essence of the Transformer architecture has long remained elusive: Is it a universal approximator, or a neural network version of known

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

DGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter-efficient Quantum Multi-task Learning

DGX agent

arXiv:2604.13560v1 Announce Type: new Abstract: Multi-task learning (MTL) improves generalization and data efficiency by jointly learning related tasks through shared representations. In the widely us

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter-Free Non-Ergodic Extragradient Algorithms for Solving Monotone Variational Inequalities

DGX agent

arXiv:2604.07662v2 Announce Type: replace-cross Abstract: Monotone variational inequalities (VIs) provide a unifying framework for convex minimization, equilibrium computation, and convex-concave sadd

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning

DGX agent

arXiv:2604.14010v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) of large language models often suffers from task interference and catastrophic forgetting. Recent approaches alleviate th

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction

DGX agent

arXiv:2604.13153v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled highly photorealistic 3D reconstruction from casually captured multi-view images. However, this access

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

PBE-UNet: A light weight Progressive Boundary-Enhanced U-Net with Scale-Aware Aggregation for Ultrasound Image Segmentation

DGX agent

arXiv:2604.13791v1 Announce Type: new Abstract: Accurate lesion segmentation in ultrasound images is essential for preventive screening and clinical diagnosis, yet remains challenging due to low contr

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Peer-Predictive Self-Training for Language Model Reasoning

DGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

PersonaVLM: Long-Term Personalized Multimodal LLMs

DGX agent

arXiv:2604.13074v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) serve as daily assistants for millions. However, their ability to generate responses aligned with individual pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Predicting Time Pressure of Powered Two-Wheeler Riders for Proactive Safety Interventions

DGX agent

arXiv:2601.03173v2 Announce Type: replace Abstract: Time pressure critically influences risky maneuvers and crash proneness among powered two-wheeler riders, yet its prediction remains underexplored i

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World

DGX agent

arXiv:2604.05096v2 Announce Type: replace Abstract: Large language models (LLMs) acquire most of their knowledge during pretraining, which ties them to a fixed snapshot of the world and makes adaptati

model-releasesarxiv-cs-cl
16 Apr 2026
← Previous
1…329330331332333…357
Next →