AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Conflicts Make Large Reasoning Models Vulnerable to Attacks

DGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

model-releasesarxiv-cs-ai
14 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Cultural Value Awareness in Large Vision-Language Models

DGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

safetyarxiv-cs-ai
14 Apr 2026
Research

Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models

DGX agent

arXiv:2604.11240v1 Announce Type: new Abstract: Token pruning has emerged as an effective approach to reduce the substantial computational overhead of Large Vision-Language Models (LVLMs) by discardin

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates

DGX agent

arXiv:2604.11038v1 Announce Type: new Abstract: We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive obje

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Evolutionary Token-Level Prompt Optimization for Diffusion Models

DGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks

DGX agent

arXiv:2604.11778v1 Announce Type: cross Abstract: Contemporary large language models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in specialized domains like mathematics an

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

DGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

DGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results

DGX agent

arXiv:2604.10551v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models. This challenge utili

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: AI Flash Portrait (Track 3)

DGX agent

arXiv:2604.11230v1 Announce Type: new Abstract: In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Pay Less Attention to Function Words for Free Robustness of Vision-Language Models

DGX agent

arXiv:2512.07222v3 Announce Type: replace-cross Abstract: To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs a

researcharxiv-cs-cl
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

DGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

DGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration

DGX agent

arXiv:2409.02231v5 Announce Type: replace-cross Abstract: We show that large language model (LLMs) can be transformed via supervised fine-tuning (SFT) of engineered prompts into SmileyLlama for explor

model-releasesarxiv-cs-lg
14 Apr 2026
Research

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

DGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

DGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

DGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?

DGX agent

arXiv:2510.27269v3 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, yet they still exhibit a multilingual reasoning gap, p

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

DGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models

DGX agent

arXiv:2604.04161v2 Announce Type: replace Abstract: In Vision-Language-Action (VLA) models, action chunking (i.e., executing a sequence of actions without intermediate replanning) is a key technique t

applicationsarxiv-cs-ro
13 Apr 2026
Model Releases

AssemLM: Spatial Reasoning Multimodal Large Language Models for Robotic Assembly

DGX agent

arXiv:2604.08983v1 Announce Type: new Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipulation tasks such as robotic assembly. While

model-releasesarxiv-cs-ro
13 Apr 2026
Model Releases

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

DGX agent

arXiv:2604.08867v1 Announce Type: cross Abstract: Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherentl

model-releasesarxiv-cs-ai
13 Apr 2026
Research

BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training

DGX agent

arXiv:2604.09022v1 Announce Type: new Abstract: With the rapid adoption of diffusion models, synthetic data generation has emerged as a promising approach for addressing the growing demand for large-s

researcharxiv-cs-cv
13 Apr 2026
Safety

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

DGX agent

arXiv:2604.08757v1 Announce Type: cross Abstract: Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimens

safetyarxiv-cs-ai
13 Apr 2026
Tutorials

Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?

DGX agent

arXiv:2604.08723v1 Announce Type: cross Abstract: Preference optimization methods such as DPO and KTO are widely used for aligning language models, yet little is understood about what properties of pr

tutorialsarxiv-cs-ai
13 Apr 2026
Safety

EvoLen: Evolution-Guided Tokenization for DNA Language Model

DGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

safetyarxiv-cs-lg
13 Apr 2026
Research

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

DGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

researcharxiv-cs-ai
13 Apr 2026
Research

HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

DGX agent

arXiv:2604.06165v2 Announce Type: replace Abstract: Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation s

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

DGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

DGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application

DGX agent

arXiv:2604.09421v1 Announce Type: cross Abstract: Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a si

researcharxiv-cs-cv
13 Apr 2026
Safety

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

DGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

safetyarxiv-cs-cv
13 Apr 2026
Research

Predictive Entropy Links Calibration and Paraphrase Sensitivity in Medical Vision-Language Models

DGX agent

arXiv:2604.08941v1 Announce Type: new Abstract: Medical Vision Language Models VLMs suffer from two failure modes that threaten safe deployment mis calibrated confidence and sensitivity to question re

researcharxiv-cs-lg
13 Apr 2026
Local Ai

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

DGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

local-aiarxiv-cs-ai
13 Apr 2026
Research

State Space Models are Effective Sign Language Learners: Exploiting Phonological Compositionality for Vocabulary-Scale Recognition

DGX agent

arXiv:2604.08761v1 Announce Type: new Abstract: Sign language recognition suffers from catastrophic scaling failure: models achieving high accuracy on small vocabularies collapse at realistic sizes. E

researcharxiv-cs-cv
13 Apr 2026
Model Releases

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

DGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Toward World Models for Epidemiology

DGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

safetyarxiv-cs-lg
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Research

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

DGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

researcharxiv-cs-lg
13 Apr 2026
Model Releases

A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset

DGX agent

arXiv:2604.06227v1 Announce Type: new Abstract: Accurate short-term forecasting of agricultural commodity prices is critical for food security planning and smallholder income stabilisation in developi

model-releasesarxiv-cs-lg
10 Apr 2026
Applications

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

DGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

DGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

DGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Bootstrapping Sign Language Annotations with Sign Language Models

DGX agent

arXiv:2604.07606v1 Announce Type: new Abstract: AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Compact Example-Based Explanations for Language Models

DGX agent

arXiv:2601.03786v2 Announce Type: replace Abstract: Training data influence estimation methods quantify the contribution of training documents to a model's output, making them a promising source of in

researcharxiv-cs-cl
10 Apr 2026
Local Ai

ConfusionPrompt: Practical Private Inference for Online Large Language Models

DGX agent

arXiv:2401.00870v5 Announce Type: replace-cross Abstract: State-of-the-art large language models (LLMs) are typically deployed as online services, requiring users to transmit detailed prompts to cloud

local-aiarxiv-cs-ai
10 Apr 2026
← Previous
1…9596979899…1030
Next →