AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

DGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

model-releasesarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Comparing energy consumption and accuracy in text classification inference

DGX agent

arXiv:2508.14170v2 Announce Type: replace Abstract: The increasing deployment of large language models (LLMs) in natural language processing (NLP) tasks raises concerns about energy efficiency and sus

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Concept Inconsistency in Dermoscopic Concept Bottleneck Models: A Rough-Set Analysis of the Derm7pt Dataset

DGX agent

arXiv:2604.19323v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) route predictions exclusively through a clinically grounded concept layer, binding interpretability to concept-label

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

ConvVitMamba: Efficient Multiscale Convolution, Transformer, and Mamba-Based Sequence modelling for Hyperspectral Image Classification

DGX agent

arXiv:2604.18856v1 Announce Type: new Abstract: Hyperspectral image (HSI) classification remains challenging due to high spectral dimensionality, redundancy, and limited labeled data. Although convolu

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Diff-SBSR: Learning Multimodal Feature-Enhanced Diffusion Models for Zero-Shot Sketch-Based 3D Shape Retrieval

DGX agent

arXiv:2604.19135v1 Announce Type: new Abstract: This paper presents the first exploration of text-to-image diffusion models for zero-shot sketch-based 3D shape retrieval (ZS-SBSR). Existing sketch-bas

safetyarxiv-cs-cv
22 Apr 2026
Safety

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling

DGX agent

arXiv:2604.19544v1 Announce Type: new Abstract: Multimodal reward models (MRMs) play a crucial role in aligning Multimodal Large Language Models (MLLMs) with human preferences. Training a good MRM req

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

DUALVISION: RGB-Infrared Multimodal Large Language Models for Robust Visual Reasoning

DGX agent

arXiv:2604.18829v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance on visual perception and reasoning tasks with RGB imagery, yet they remain

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

FairTree: Subgroup Fairness Auditing of Machine Learning Models with Bias-Variance Decomposition

DGX agent

arXiv:2604.19357v1 Announce Type: new Abstract: The evaluation of machine learning models typically relies mainly on performance metrics based on loss functions, which risk to overlook changes in perf

safetyarxiv-cs-lg
22 Apr 2026
Safety

Hierarchically Robust Zero-shot Vision-language Models

DGX agent

arXiv:2604.18867v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their

safetyarxiv-cs-ai
22 Apr 2026
Safety

How to Teach Large Multimodal Models New Skills

DGX agent

arXiv:2510.08564v2 Announce Type: replace Abstract: How can we teach large multimodal models (LMMs) new skills without erasing prior abilities? We study sequential fine-tuning on five target skills wh

safetyarxiv-cs-ai
22 Apr 2026
Applications

InHabit: Leveraging Image Foundation Models for Scalable 3D Human Placement

DGX agent

arXiv:2604.19673v1 Announce Type: new Abstract: Training embodied agents to understand 3D scenes as humans do requires large-scale data of people meaningfully interacting with diverse environments, ye

applicationsarxiv-cs-cv
22 Apr 2026
Hardware

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation

DGX agent

arXiv:2604.19167v1 Announce Type: cross Abstract: Deploying large language models (LLMs) in resource-constrained environments is hindered by heavy computational and memory requirements. We present LBL

hardwarearxiv-cs-ai
22 Apr 2026
Local Ai

Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control

DGX agent

arXiv:2604.19018v1 Announce Type: cross Abstract: Inference-time LLM alignment methods, particularly activation steering, offer an alternative to fine-tuning by directly modifying activations during g

local-aiarxiv-cs-ai
22 Apr 2026
Model Releases

OLLM: Options-based Large Language Models

DGX agent

arXiv:2604.19087v1 Announce Type: new Abstract: We introduce Options LLM (OLLM), a simple, general method that replaces the single next-token prediction of standard LLMs with a extit{set of learned op

model-releasesarxiv-cs-ai
22 Apr 2026
Research

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

DGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

DGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

model-releasesarxiv-cs-cl
22 Apr 2026
Applications

ParamBoost: Gradient Boosted Piecewise Cubic Polynomials

DGX agent

arXiv:2604.18864v1 Announce Type: new Abstract: Generalized Additive Models (GAMs) can be used to create non-linear glass-box (i.e. explicitly interpretable) models, where the predictive function is f

applicationsarxiv-cs-lg
22 Apr 2026
Safety

ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety

DGX agent

arXiv:2604.19083v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in cross-modal understanding and generation, yet their deployment is threate

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Q-Mask: Query-driven Causal Masks for Text Anchoring in OCR-Oriented Vision-Language Models

DGX agent

arXiv:2604.00161v2 Announce Type: replace Abstract: Optical Character Recognition (OCR) is increasingly regarded as a foundational capability for modern vision-language models (VLMs), enabling them no

model-releasesarxiv-cs-cv
22 Apr 2026
Local Ai

R^2-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction

DGX agent

arXiv:2604.18995v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to autoregressive generation by enabling parallel token prediction. Ho

local-aiarxiv-cs-ai
22 Apr 2026
Agents

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

DGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

agentsarxiv-cs-ai
22 Apr 2026
Safety

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

DGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

safetyarxiv-cs-ai
22 Apr 2026
Safety

TEMPO: Scaling Test-time Training for Large Reasoning Models

DGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

A Systematic Survey and Benchmark of Deep Learning for Molecular Property Prediction in the Foundation Model Era

DGX agent

arXiv:2604.16586v1 Announce Type: new Abstract: Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and b

model-releasesarxiv-cs-lg
21 Apr 2026
Research

A Unification of Discrete, Gaussian, and Simplicial Diffusion

DGX agent

arXiv:2512.15923v2 Announce Type: replace Abstract: To model discrete sequences such as DNA, proteins, and language using diffusion, practitioners must choose between three major methods: diffusion in

researcharxiv-cs-lg
21 Apr 2026
Applications

An LLM-Guided Query-Aware Inference System for GNN Models on Large Knowledge Graphs

DGX agent

arXiv:2603.04545v2 Announce Type: replace Abstract: Efficient inference for graph neural networks (GNNs) on large knowledge graphs (KGs) is essential for many real-world applications. GNN inference qu

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

DGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Calibrating Model-Based Evaluation Metrics for Summarization

DGX agent

arXiv:2604.17200v1 Announce Type: new Abstract: Recent advances in summary evaluation are based on model-based metrics to assess quality dimensions, such as completeness, conciseness, and faithfulness

researcharxiv-cs-cl
21 Apr 2026
Safety

Closing the Modality Reasoning Gap for Speech Large Language Models

DGX agent

arXiv:2601.05543v2 Announce Type: replace Abstract: Although Speech Large Language Models have achieved notable progress, a substantial modality reasoning gap remains: their reasoning performance on s

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

DGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Domain-Specialized Object Detection via Model-Level Mixtures of Experts

DGX agent

arXiv:2604.18256v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models provide a structured approach to combining specialized neural networks and offer greater interpretability than conventio

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

DGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models

DGX agent

arXiv:2604.16481v1 Announce Type: new Abstract: Large-scale text-to-image (T2I) diffusion models deliver remarkable visual fidelity but pose safety risks due to their capacity to reproduce undesirable

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

DGX agent

arXiv:2410.04509v3 Announce Type: replace Abstract: As the field of Multimodal Large Language Models (MLLMs) continues to evolve, their potential to revolutionize artificial intelligence is particular

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings

DGX agent

arXiv:2505.15353v3 Announce Type: replace Abstract: Log-likelihood vectors define a common space for comparing language models as probability distributions, enabling unified comparisons across heterog

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FedLLM: A Privacy-Preserving Federated Large Language Model for Explainable Traffic Flow Prediction

DGX agent

arXiv:2604.16612v1 Announce Type: new Abstract: Traffic prediction plays a central role in intelligent transportation systems (ITS) by supporting real-time decision-making, congestion management, and

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FLARE: Task-agnostic embedding model evaluation through a normalization process

DGX agent

arXiv:2604.17344v1 Announce Type: cross Abstract: When task-specific labels are not available, it becomes difficult to select an embedding model for a specific target corpus. Existing labelless measur

model-releasesarxiv-cs-cl
21 Apr 2026
Research

GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning

DGX agent

arXiv:2604.17241v1 Announce Type: new Abstract: Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However

researcharxiv-cs-ro
21 Apr 2026
Tutorials

Graph neural network for colliding particles with an application to sea ice floe modeling

DGX agent

arXiv:2602.16213v2 Announce Type: replace-cross Abstract: This paper introduces a novel approach to sea ice modeling using Graph Neural Networks (GNNs), utilizing the natural graph structure of sea ic

tutorialsarxiv-cs-cv
21 Apr 2026
Safety

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

DGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

DGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

How to Approximate Inference with Subtractive Mixture Models

DGX agent

arXiv:2604.16714v1 Announce Type: new Abstract: Classical mixture models (MMs) are widely used tractable proposals for approximate inference settings such as variational inference (VI) and importance

tutorialsarxiv-cs-lg
21 Apr 2026
Model Releases

Large Language Models Are Still Misled by Simple Bias Ensembles

DGX agent

arXiv:2505.16522v3 Announce Type: replace Abstract: With the evolution of large language models (LLMs), their robustness against individual simple biases has been enhanced. However, we observe that th

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

DGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

local-aiarxiv-cs-cl
21 Apr 2026
Research

Multilingual Training and Evaluation Resources for Vision-Language Models

DGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

researcharxiv-cs-cl
21 Apr 2026
Research

Non-Stationarity in the Embedding Space of Time Series Foundation Models

DGX agent

arXiv:2604.16428v1 Announce Type: new Abstract: Time series foundation models (TSFMs) are widely used as generic feature extractors, yet the notion of non-stationarity in their embedding spaces remain

researcharxiv-cs-lg
21 Apr 2026
Model Releases

PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking

DGX agent

arXiv:2604.17819v1 Announce Type: new Abstract: Large language models (LLMs) perform substantially below human level on existing theory-of-mind (ToM) benchmarks, even when augmented with chain-of-thou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics

DGX agent

arXiv:2604.17670v1 Announce Type: new Abstract: We introduce Prior-Fitted Functional Flows, a generative foundation model for pharmacokinetics that enables zero-shot population synthesis and individua

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…117118119120121…1030
Next →