AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

model-releasesarxiv-cs-cl
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Seg2Change: Adapting Open-Vocabulary Semantic Segmentation Model for Remote Sensing Change Detection

DGX agent

arXiv:2604.11231v1 Announce Type: new Abstract: Change detection is a fundamental task in remote sensing, aiming to quantify the impacts of human activities and ecological dynamics on land-cover chang

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Self-Calibrating Language Models via Test-Time Discriminative Distillation

DGX agent

arXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin

researcharxiv-cs-cl
14 Apr 2026
Research

SinkTrack: Attention Sink based Context Anchoring for Large Language Models

DGX agent

arXiv:2604.10027v1 Announce Type: new Abstract: Large language models (LLMs) suffer from hallucination and context forgetting. Prior studies suggest that attention drift is a primary cause of these pr

researcharxiv-cs-cv
14 Apr 2026
Research

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning

DGX agent

arXiv:2505.11737v4 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have demonstrated impressive capabilities, their output quality remains inconsistent across various applica

researcharxiv-cs-ai
14 Apr 2026
Research

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

DGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

researcharxiv-cs-cl
14 Apr 2026
Applications

Understanding Generalization in Role-Playing Models via Information Theory

DGX agent

arXiv:2512.17270v2 Announce Type: replace-cross Abstract: Role-playing models (RPMs) are widely used in real-world applications but underperform when deployed in the wild. This degradation can be attr

applicationsarxiv-cs-ai
14 Apr 2026
Tutorials

Using Deep Learning Models Pretrained by Self-Supervised Learning for Protein Localization

DGX agent

arXiv:2604.10970v1 Announce Type: new Abstract: Background: Task-specific microscopy datasets are often small, making it difficult to train deep learning models that learn robust features. While self-

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

DGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

DGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

YUV20K: A Complexity-Driven Benchmark and Trajectory-Aware Alignment Model for Video Camouflaged Object Detection

DGX agent

arXiv:2604.09985v1 Announce Type: new Abstract: Video Camouflaged Object Detection (VCOD) is currently constrained by the scarcity of challenging benchmarks and the limited robustness of models agains

model-releasesarxiv-cs-cv
14 Apr 2026
Research

A Predictive View on Streaming Hidden Markov Models

DGX agent

arXiv:2604.09208v1 Announce Type: cross Abstract: We develop a predictive-first optimisation framework for streaming hidden Markov models. Unlike classical approaches that prioritise full posterior re

researcharxiv-cs-lg
13 Apr 2026
Safety

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention

DGX agent

arXiv:2511.18960v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable progress in embodied tasks recently, but most methods process visual observations in

safetyarxiv-cs-cv
13 Apr 2026
Tutorials

CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation

DGX agent

arXiv:2604.09201v1 Announce Type: new Abstract: Camera-controllable video generation aims to synthesize videos with flexible and physically plausible camera movements. However, existing methods either

tutorialsarxiv-cs-cv
13 Apr 2026
Model Releases

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks

DGX agent

arXiv:2604.09535v1 Announce Type: new Abstract: Large foundation models have made significant advances in embodied intelligence, enabling synthesis and reasoning over egocentric input for household ta

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models

DGX agent

arXiv:2503.14075v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) have demonstrated remarkable capabilities in open-world multimodal understanding, yet their high computati

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

DGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym

DGX agent

arXiv:2604.09338v1 Announce Type: new Abstract: Spatial reasoning is central to navigation and robotics, yet measuring model capabilities on these tasks remains difficult. Existing benchmarks evaluate

tutorialsarxiv-cs-ai
13 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

DGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Quantisation Reshapes the Metacognitive Geometry of Language Models

DGX agent

arXiv:2604.08976v1 Announce Type: new Abstract: We report that model quantisation restructures domain-level metacognitive efficiency in LLMs rather than degrading it uniformly. Evaluating Llama-3-8B-I

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images

DGX agent

arXiv:2604.09436v1 Announce Type: new Abstract: Diffusion models trained on noisy datasets often reproduce high-frequency training artifacts, significantly degrading generation quality. To address thi

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

SessionIntentBench: A Multi-task Inter-session Intention-shift Modeling Benchmark for E-commerce Customer Behavior Understanding

DGX agent

arXiv:2507.20185v2 Announce Type: replace Abstract: Session history is a common way of recording user interacting behaviors throughout a browsing activity with multiple products. For example, if an us

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Research

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

DGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

researcharxiv-cs-ai
13 Apr 2026
Tutorials

A comparative analysis of machine learning models in SHAP analysis

DGX agent

arXiv:2604.07258v1 Announce Type: new Abstract: In this growing age of data and technology, large black-box models are becoming the norm due to their ability to handle vast amounts of data and learn i

tutorialsarxiv-cs-lg
10 Apr 2026
Model Releases

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

DGX agent

arXiv:2604.07274v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models

DGX agent

arXiv:2604.06201v1 Announce Type: cross Abstract: While most reading comprehension benchmarks for LLMs focus on factual information that can be answered by localizing specific textual evidence, many r

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

DGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

safetyarxiv-cs-ai
10 Apr 2026
Safety

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

DGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

safetyarxiv-cs-cl
10 Apr 2026
Research

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

DGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

researcharxiv-cs-ai
10 Apr 2026
Research

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

DGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

researcharxiv-cs-cl
10 Apr 2026
Safety

Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization

DGX agent

arXiv:2604.08476v1 Announce Type: new Abstract: Multimodal reasoning models (MRMs) trained with reinforcement learning with verifiable rewards (RLVR) show improved accuracy on visual reasoning benchma

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

FedSpy-LLM: Towards Scalable and Generalizable Data Reconstruction Attacks from Gradients on LLMs

DGX agent

arXiv:2604.06297v1 Announce Type: cross Abstract: Given the growing reliance on private data in training Large Language Models (LLMs), Federated Learning (FL) combined with Parameter-Efficient Fine-Tu

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

Guiding a Diffusion Model by Swapping Its Tokens

DGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathology

DGX agent

arXiv:2604.08305v1 Announce Type: cross Abstract: Immunohistochemistry (IHC) is essential for assessing specific immune biomarkers like Human Epidermal growth-factor Receptor 2 (HER2) in breast cancer

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2604.07802v1 Announce Type: new Abstract: Large-scale vision-language models (VLMs) exhibit remarkable zero-shot capabilities, yet the internal mechanisms driving their anomaly detection (AD) pe

researcharxiv-cs-cv
10 Apr 2026
Research

Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification

DGX agent

arXiv:2604.08333v1 Announce Type: new Abstract: The rise of multimodal large language models (MLLMs) has sparked an unprecedented wave of applications in the field of medical imaging analysis. However

researcharxiv-cs-cv
10 Apr 2026
Safety

LUMINA: Foundation Models for Topology Transferable ACOPF

DGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

safetyarxiv-cs-lg
10 Apr 2026
Local Ai

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models

DGX agent

arXiv:2601.04068v3 Announce Type: replace Abstract: Aligning text-to-video diffusion models with human preferences is crucial for generating high-quality videos. Existing Direct Preference Otimization

local-aiarxiv-cs-cv
10 Apr 2026
Research

Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction

DGX agent

arXiv:2604.07914v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable success across cross-modal tasks but remain hindered by hallucinations, producing textual

researcharxiv-cs-cv
10 Apr 2026
Safety

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

DGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

safetyarxiv-cs-cl
10 Apr 2026
Safety

Self-Debias: Self-correcting for Debiasing Large Language Models

DGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

DGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Human Condition as Reflected in Contemporary Large Language Models

DGX agent

arXiv:2604.06206v1 Announce Type: cross Abstract: This study seeks to uncover evidence of a latent structure in evolved human culture as it is refracted through contemporary large language models (LLM

researcharxiv-cs-ai
10 Apr 2026
Model Releases

VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents

DGX agent

arXiv:2603.15118v2 Announce Type: replace Abstract: We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

DGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning

DGX agent

arXiv:2604.08281v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

Adjustable Text-Guided Backdoor Attacks with Natural-Word Triggers on Multimodal Pretrained Models

DGX agent

arXiv:2604.05809v2 Announce Type: replace-cross Abstract: This paper presents Text-Guided Backdoor (TGB), an adjustable backdoor attack against multimodal pretrained models that uses natural-word trig

applicationsarxiv-cs-lg
14 Aug 2026
← Previous
1…120121122123124…1030
Next →