AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

GR4CIL: Gap-compensated Routing for CLIP-based Class Incremental Learning

DGX agent

arXiv:2604.17822v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) aims to continuously acquire new categories while preserving previously learned knowledge. Recently, Contrastive Langua

model-releasesarxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

DGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Guardrails in Logit Space: Safety Token Regularization for LLM Alignment

DGX agent

arXiv:2604.17210v1 Announce Type: new Abstract: Fine-tuning well-aligned large language models (LLMs) on new domains often degrades their safety alignment, even when using benign datasets. Existing sa

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders

DGX agent

arXiv:2604.16430v1 Announce Type: new Abstract: Large Language Models (LLMs) are powerful and widely adopted, but their practical impact is limited by the well-known hallucination phenomenon. While re

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

DGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

DGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

DGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

DGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HorizonBench: Long-Horizon Personalization with Evolving Preferences

DGX agent

arXiv:2604.17283v1 Announce Type: new Abstract: User preferences evolve across months of interaction, and tracking them requires inferring when a stated preference has been changed by a subsequent lif

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

How Robustly do LLMs Understand Execution Semantics?

DGX agent

arXiv:2604.16320v1 Announce Type: cross Abstract: LLMs demonstrate remarkable reasoning capabilities, yet whether they utilize internal world models or rely on sophisticated pattern matching remains o

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

DGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HQA-VLAttack: Towards High Quality Adversarial Attack on Vision-Language Pre-Trained Models

DGX agent

arXiv:2604.16499v1 Announce Type: new Abstract: Black-box adversarial attack on vision-language pre-trained models is a practical and challenging task, as text and image perturbations need to be consi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Hybrid-Vector Retrieval for Visually Rich Documents: Combining Single-Vector Efficiency and Multi-Vector Accuracy

DGX agent

arXiv:2510.22215v2 Announce Type: replace-cross Abstract: Retrieval over visually rich documents is essential for tasks such as legal discovery, scientific search, and enterprise knowledge management.

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ICAT: Incident-Case-Grounded Adaptive Testing for Physical-Risk Prediction in Embodied World Models

DGX agent

arXiv:2604.16405v1 Announce Type: cross Abstract: Video-generative world models are increasingly used as neural simulators for embodied planning and policy learning, yet their ability to predict physi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

DGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

iDocV2: Leveraging Self-Supervision and Open-Set Detection for Improving Pattern Spotting in Historical Documents

DGX agent

arXiv:2604.16726v1 Announce Type: new Abstract: Considering the imminent massification of digital books, it has become critical to facilitate searching collections through graphical patterns. Current

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification

DGX agent

arXiv:2604.17010v1 Announce Type: new Abstract: We introduce a self-play framework for semantic equivalence in Haskell, utilizing formal verification to guide adversarial training between a generator

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

In-Context Symbolic Regression for Robustness-Improved Kolmogorov-Arnold Networks

DGX agent

arXiv:2603.15250v2 Announce Type: replace Abstract: Symbolic regression aims to replace black-box predictors with concise analytical expressions that can be inspected and validated in scientific machi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Incoherent Deformation, Not Capacity: Diagnosing and Mitigating Overfitting in Dynamic Gaussian Splatting

DGX agent

arXiv:2604.16747v1 Announce Type: new Abstract: Dynamic 3D Gaussian Splatting methods achieve strong training-view PSNR on monocular video but generalize poorly: on the D-NeRF benchmark we measure an

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

IncreFA: Breaking the Static Wall of Generative Model Attribution

DGX agent

arXiv:2604.17736v1 Announce Type: new Abstract: As AI generative models evolve at unprecedented speed, image attribution has become a moving target. New diffusion, adversarial and autoregressive gener

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation

DGX agent

arXiv:2510.09275v2 Announce Type: replace Abstract: Medical diagnostics is a high-stakes and complex domain that is critical to patient care. However, current evaluations of large language models (LLM

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG

DGX agent

arXiv:2604.16422v1 Announce Type: new Abstract: The injection of domain-specific knowledge is crucial for adapting language models (LMs) to specialized fields such as biomedicine. While most current a

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval

DGX agent

arXiv:2604.18051v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a challenging image retrieval paradigm that enables to retrieve target images based on multimodal queries consisting o

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts

DGX agent

arXiv:2509.10813v3 Announce Type: replace Abstract: The advancement of Embodied AI heavily relies on large-scale, simulatable 3D scene datasets characterized by scene diversity and realistic layouts.

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Interpolating Discrete Diffusion Models with Controllable Resampling

DGX agent

arXiv:2604.17310v1 Announce Type: new Abstract: Discrete diffusion models form a powerful class of generative models across diverse domains, including text and graphs. However, existing approaches fac

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

DGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew

DGX agent

arXiv:2604.18041v1 Announce Type: new Abstract: Despite significant advances in large language models, personalizing them for individual decision-makers remains an open problem. Here, we introduce a s

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Jupiter-N Technical Report

DGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

KIRA: Knowledge-Intensive Image Retrieval and Reasoning Architecture for Specialized Visual Domains

DGX agent

arXiv:2604.16915v1 Announce Type: new Abstract: Retrieval augmented generation (RAG) has transformed text based question answering, yet its extension to visual domains remains hindered by fundamental

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Knowing When to Quit: A Principled Framework for Dynamic Abstention in LLM Reasoning

DGX agent

arXiv:2604.18419v1 Announce Type: cross Abstract: Large language models (LLMs) using chain-of-thought reasoning often waste substantial compute by producing long, incorrect responses. Abstention can m

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact

DGX agent

arXiv:2603.00883v2 Announce Type: replace Abstract: LLMs increasingly excel on AI benchmarks, but doing so does not guarantee validity for downstream tasks. This study contrasts LLM alignment on bench

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Large Language Models Are Still Misled by Simple Bias Ensembles

DGX agent

arXiv:2505.16522v3 Announce Type: replace Abstract: With the evolution of large language models (LLMs), their robustness against individual simple biases has been enhanced. However, we observe that th

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Late Fusion Neural Operators for Extrapolation Across Parameter Space in Partial Differential Equations

DGX agent

arXiv:2604.16721v1 Announce Type: new Abstract: Developing neural operators that accurately predict the behavior of systems governed by partial differential equations (PDEs) across unseen parameter re

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Latent Preference Modeling for Cross-Session Personalized Tool Calling

DGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LayerCache: Exploiting Layer-wise Velocity Heterogeneity for Efficient Flow Matching Inference

DGX agent

arXiv:2604.16492v1 Announce Type: new Abstract: Flow Matching models achieve state-of-the-art image generation quality but incur substantial inference cost due to iterative denoising through large Tra

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

LEAF: Knowledge Distillation of Text Embedding Models with Teacher-Aligned Representations

DGX agent

arXiv:2509.12539v2 Announce Type: replace-cross Abstract: We present LEAF ('Lightweight Embedding Alignment Framework'), a knowledge distillation framework for text embedding models. A key distinguish

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Learned Nonlocal Feature Matching and Filtering for RAW Image Denoising

DGX agent

arXiv:2604.17453v1 Announce Type: cross Abstract: Being one of the oldest and most basic problems in image processing, image denoising has seen a resurgence spurred by rapid advances in deep learning.

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Learning Stable Predictors from Weak Supervision under Distribution Shift

DGX agent

arXiv:2604.05002v2 Announce Type: replace Abstract: Learning from weak, proxy, or relative supervision is common when ground-truth labels are unavailable, but robustness under distribution shift remai

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Learning to Control Summaries with Score Ranking

DGX agent

arXiv:2604.17197v1 Announce Type: new Abstract: Recent advances in summarization research focus on improving summary quality across multiple criteria, such as completeness, conciseness, and faithfulne

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Learning to Retrieve User History and Generate User Profiles for Personalized Persuasiveness Prediction

DGX agent

arXiv:2601.05654v3 Announce Type: replace Abstract: Estimating the persuasiveness of messages is critical in various applications, from recommender systems to safety assessment of LLMs. While it is im

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Leveraging Large Language Models for Sarcastic Speech Annotation in Sarcasm Detection

DGX agent

arXiv:2506.00955v2 Announce Type: replace Abstract: Sarcasm fundamentally alters meaning through tone and context, yet detecting it in speech remains a challenge due to data scarcity. In addition, exi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases

DGX agent

arXiv:2512.12643v2 Announce Type: replace Abstract: Legal relations serve as an important analytical framework for dispute resolution in civil cases. However, legal relations in Chinese civil cases re

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LiFT: Does Instruction Fine-Tuning Improve In-Context Learning for Longitudinal Modelling by Large Language Models?

DGX agent

arXiv:2604.16382v1 Announce Type: new Abstract: Longitudinal NLP tasks require reasoning over temporally ordered text to detect persistence and change in human behavior and opinions. However, in-conte

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LIFT the Veil for the Truth: Principal Weights Emerge after Rank Reduction for Reasoning-Focused Supervised Fine-Tuning

DGX agent

arXiv:2506.00772v2 Announce Type: replace-cross Abstract: Recent studies have shown that supervised fine-tuning of LLMs on a small number of high-quality datasets can yield strong reasoning capabiliti

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LiquidTAD: An Efficient Method for Temporal Action Detection via Liquid Neural Dynamics

DGX agent

arXiv:2604.18274v1 Announce Type: new Abstract: Temporal Action Detection (TAD) in untrimmed videos is currently dominated by Transformer-based architectures. While high-performing, their quadratic co

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

DGX agent

arXiv:2512.04677v5 Announce Type: replace Abstract: Audio-driven avatar interaction demands real-time, streaming, and infinite-length generation -- capabilities fundamentally at odds with the sequenti

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…315316317318319…357
Next →