AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models

DGX agent

arXiv:2605.08841v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit systematic bias toward visual illusions, recalling memorized facts rather than perceiving actual visual difference

model-releasesarxiv-cs-cv
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

DGX agent

arXiv:2605.08209v1 Announce Type: new Abstract: Deep learning methods are widely used under diverse resource constraints, resulting in models of varying sizes, such as the Vision Transformer (ViT) ser

researcharxiv-cs-lg
12 May 2026
Model Releases

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models

DGX agent

arXiv:2510.08592v3 Announce Type: replace-cross Abstract: Test-Time Scaling (TTS) improves LLM reasoning by exploring multiple candidate responses and then operating over this set to find the best out

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LLiMba: Sardinian on a Single GPU -- Adapting a 3B Language Model to a Vanishing Romance Language

DGX agent

arXiv:2605.09015v1 Announce Type: new Abstract: Sardinian, a Romance language with roughly one million speakers, has minimal presence in modern NLP. Commercial services do not support it, and current

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Med-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models

DGX agent

arXiv:2605.10002v1 Announce Type: new Abstract: Large vision-language models (VLMs) demonstrate strong performance in medical image understanding, but frequently generate clinically plausible yet inco

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Meow-Omni 1: A Multimodal Large Language Model for Feline Ethology

DGX agent

arXiv:2605.09152v1 Announce Type: new Abstract: Deciphering animal intent is a fundamental challenge in computational ethology, largely because of semantic aliasing, the phenomenon where identical ext

model-releasesarxiv-cs-cl
12 May 2026
Safety

On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models

DGX agent

arXiv:2605.09606v1 Announce Type: cross Abstract: Recent advances in image-to-3D models have significantly improved the fidelity and accessibility of 3D content creation. Such a powerful reconstructio

safetyarxiv-cs-cv
12 May 2026
Model Releases

Pretraining large language models with MXFP4

DGX agent

arXiv:2605.09825v1 Announce Type: cross Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We a

model-releasesarxiv-cs-ai
12 May 2026
Safety

SAID: Safety-Aware Intent Defense via Prefix Probing for Large Language Models

DGX agent

arXiv:2510.20129v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attacks, where adversarially crafted prompts induce policy-violating responses des

safetyarxiv-cs-ai
12 May 2026
Research

SLayerGen: a Crystal Generative Model for all Space and Layer Groups

DGX agent

arXiv:2605.08262v1 Announce Type: cross Abstract: Crystal generative models have shown rapid progress for accelerating the discovery of bulk, periodic materials. However, many material systems such as

researcharxiv-cs-ai
12 May 2026
Research

SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training

DGX agent

arXiv:2605.08738v1 Announce Type: cross Abstract: Structured pruning and knowledge distillation (KD) are typical techniques for compressing large language models, but it remains unclear how they shoul

researcharxiv-cs-ai
12 May 2026
Model Releases

Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition

DGX agent

arXiv:2605.10772v1 Announce Type: cross Abstract: Large language-vision models (LLVM), such as OpenAI's ChatGPT and GPT-4, have gained prominence as powerful tools for analyzing text and imagery. The

model-releasesarxiv-cs-ai
12 May 2026
Research

Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models

DGX agent

arXiv:2605.09005v1 Announce Type: cross Abstract: Vision-Language-Action models (VLAs) support generalist robotic control by enabling end-to-end decision policies directly from multi-modal inputs. As

researcharxiv-cs-ai
12 May 2026
Model Releases

Towards Compact Sign Language Translation: Frame Rate and Model Size Trade-offs

DGX agent

arXiv:2605.09554v1 Announce Type: new Abstract: Sign Language Translation (SLT) converts sign language videos into spoken-language text, bridging communication between Deaf and hearing communities. Cu

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models

DGX agent

arXiv:2605.08974v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have advanced video understanding, they remain highly prone to hallucinations in dynamic scenes. We arg

model-releasesarxiv-cs-ai
12 May 2026
Research

Uncovering Intra-expert Activation Sparsity for Efficient Mixture-of-Expert Model Execution

DGX agent

arXiv:2605.08575v1 Announce Type: cross Abstract: Mixture of Experts (MoE) architecture has become the standard for state-of-the-art large language models, owing to its computational efficiency throug

researcharxiv-cs-ai
12 May 2026
Model Releases

Understanding Asynchronous Inference Methods for Vision-Language-Action Models

DGX agent

arXiv:2605.08168v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising path to generalist robot control, but their inference latency causes observation staleness when

model-releasesarxiv-cs-ai
12 May 2026
Safety

Unsupervised Process Reward Models

DGX agent

arXiv:2605.10158v1 Announce Type: new Abstract: Process Reward Models (PRMs) are a powerful mechanism for steering large language model reasoning by providing fine-grained, step-level supervision. How

safetyarxiv-cs-lg
12 May 2026
Model Releases

Amortized Multi-Objective Optimization Across Tasks with Generative Solution Modeling

DGX agent

arXiv:2511.09598v5 Announce Type: replace Abstract: Many real-world applications require solving families of expensive multi-objective optimization problems~(EMOPs) under varying operational condition

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

BGM-IV: an AI-powered Bayesian generative modeling approach for instrumental variable analysis

DGX agent

arXiv:2605.07029v1 Announce Type: cross Abstract: Instrumental-variable (IV) regression enables causal estimation under endogeneity, but modern IV problems often involve nonlinear structural effects a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

DT-PBO: an Interpretable Tree-based Surrogate Model for Preferential Bayesian Optimization

DGX agent

arXiv:2512.14263v2 Announce Type: replace-cross Abstract: Preferential Bayesian Optimization (PBO) aims to find a decision-maker's most preferred solution in as few pairwise comparisons as possible. E

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Efficient Data Selection for Multimodal Models via Incremental Optimization Utility

DGX agent

arXiv:2605.07488v1 Announce Type: new Abstract: The scaling of Large Multimodal Models (LMMs) is constrained by the quality-quantity trade-off inherent in synthetic data. Previous approaches, such as

model-releasesarxiv-cs-ai
11 May 2026
Safety

Flow-OPD: On-Policy Distillation for Flow Matching Models

DGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

safetyarxiv-cs-ai
11 May 2026
Model Releases

Frequency-Aware Model Parameter Explorer: A new attribution method for improving explainability

DGX agent

arXiv:2510.03245v2 Announce Type: replace-cross Abstract: State-of-the-art attribution methods rely on adversarial sample generation that applies an all-pass filter across the frequency spectrum, disc

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MicroBi-ConvLSTM: An Ultra-Lightweight Efficient Model for Human Activity Recognition on Resource Constrained Devices

DGX agent

arXiv:2602.06523v2 Announce Type: replace Abstract: Human Activity Recognition (HAR) on resource constrained wearables requires models that balance accuracy against strict memory and computational bud

model-releasesarxiv-cs-cv
11 May 2026
Safety

NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models

DGX agent

arXiv:2605.07794v1 Announce Type: new Abstract: World Action Models (WAMs) are an emerging family of policies that tie robot action generation to future-observation modeling. In this work, we focus on

safetyarxiv-cs-ro
11 May 2026
Research

Normalizing Trajectory Models

DGX agent

arXiv:2605.08078v1 Announce Type: new Abstract: Diffusion-based models decompose sampling into many small Gaussian denoising steps -- an assumption that breaks down when generation is compressed to a

researcharxiv-cs-cv
11 May 2026
Model Releases

Outlier Smoothing with Closed-Form Rotations for W4A4 Large Language Model Quantization

DGX agent

arXiv:2511.22316v2 Announce Type: replace Abstract: Large Language Models (LLMs) quantization facilitates deploying LLMs in resource-limited settings, but existing methods that combine incompatible gr

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

PathPainter: Transferring the Generalization Ability of Image Generation Models to Embodied Navigation

DGX agent

arXiv:2605.07496v1 Announce Type: new Abstract: Bird's-eye-view (BEV) images have been widely demonstrated to provide valuable prior information for navigation. Given the global information provided b

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models

DGX agent

arXiv:2605.07574v1 Announce Type: new Abstract: Mainstream vision-language models (VLMs) fundamentally struggle with severe optical ambiguities, such as reflections and transparent objects, due to the

model-releasesarxiv-cs-cv
11 May 2026
Research

ProteinJEPA: Latent prediction complements protein language models

DGX agent

arXiv:2605.07554v1 Announce Type: cross Abstract: Protein language models are trained primarily with masked language modeling (MLM), which predicts amino-acid identities at masked positions. We ask wh

researcharxiv-cs-ai
11 May 2026
Safety

Reflections and New Directions for Human-Centered Large Language Models

DGX agent

arXiv:2605.06901v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping the private and professional lives of users, with numerous applications in business, education, fi

safetyarxiv-cs-cl
11 May 2026
Safety

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

DGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

safetyarxiv-cs-ai
11 May 2026
Model Releases

Structured Prototype-Guided Adaptation for EEG Foundation Models

DGX agent

arXiv:2602.17251v2 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (EFMs) have shown strong potential for transferable representation learning, yet their adaptation in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

TajPersLexon: A Tajik-Persian Lexical Resource and Hybrid Model for Cross-Script Low-Resource NLP

DGX agent

arXiv:2605.06886v1 Announce Type: new Abstract: This work introduces TajPersLexon, a curated Tajik--Persian parallel lexical resource of 40,112 word and short-phrase pairs for cross-script lexical ret

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

DGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

model-releasesarxiv-cs-lg
11 May 2026
Research

A Fast Model Counting Algorithm for Two-Variable Logic with Counting and Modulo Counting Quantifiers

DGX agent

arXiv:2605.03391v1 Announce Type: cross Abstract: Weighted first-order model counting (WFOMC) is a central task in lifted probabilistic inference: It asks for the weighted sum of all models of a first

researcharxiv-cs-ai
7 May 2026
Safety

Agent-Based Modeling of Low-Emission Fertilizer Adoption for Dairy Farm Decarbonisation using Empirical Farm Data

DGX agent

arXiv:2605.03648v1 Announce Type: new Abstract: To understand complex system dynamics in dairy farming, it is essential to use modeling tools that capture farm heterogeneity, social interactions, and

safetyarxiv-cs-ai
7 May 2026
Applications

Automatically Finding and Validating Unexpected Side-Effects of Interventions on Language Models

DGX agent

arXiv:2605.05090v1 Announce Type: new Abstract: We present an automated, contrastive evaluation pipeline for auditing the behavioral impact of interventions on large language models. Given a base mode

applicationsarxiv-cs-cl
7 May 2026
Model Releases

Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors

DGX agent

arXiv:2512.06393v5 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve high accuracy on many reasoning benchmarks but remain brittle under structural perturbations of rule-base

model-releasesarxiv-cs-cl
7 May 2026
Local Ai

Counterfactual identifiability beyond global monotonicity: non-monotone triangular structural causal models

DGX agent

arXiv:2605.04413v1 Announce Type: new Abstract: Structural causal models provide a unified semantics for interventions and counterfactuals, but most identifiability results rely on restrictive assumpt

local-aiarxiv-cs-lg
7 May 2026
Research

Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation

DGX agent

arXiv:2512.14954v2 Announce Type: replace Abstract: Computing next-token likelihood ratios between two language models (LMs) is a standard task in training paradigms such as knowledge distillation. Si

researcharxiv-cs-cl
7 May 2026
Hardware

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

DGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

hardwarearxiv-cs-lg
7 May 2026
Research

Lookahead Drifting Model

DGX agent

arXiv:2605.04060v1 Announce Type: cross Abstract: Recently, a new paradigm named drifting model has been proposed for mapping distributions, which achieves the SOTA image generation performance over I

researcharxiv-cs-cv
7 May 2026
Tutorials

Multi Language Models for On-the-Fly Syntax Highlighting

DGX agent

arXiv:2510.04166v2 Announce Type: replace-cross Abstract: Syntax highlighting is a critical feature in modern software development environments, enhancing code readability and developer productivity.

tutorialsarxiv-cs-ai
7 May 2026
Model Releases

PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentation

DGX agent

arXiv:2605.05159v1 Announce Type: new Abstract: We present our system for SemEval-2026 Task 9: Multilingual Polarization Detection, a binary classification task spanning 22 languages. Our approach fin

model-releasesarxiv-cs-cl
7 May 2026
Research

Right Model, Right Time: Real-Time Cascaded-Fidelity MPC for Bipedal Walking

DGX agent

arXiv:2605.04607v1 Announce Type: new Abstract: This paper presents a multi-phase whole-body model predictive control approach for bipedal walking, combining a detailed whole-body model in the near ho

researcharxiv-cs-ro
7 May 2026
Model Releases

RLearner-LLM: Balancing Logical Grounding and Fluency in Large Language Models via Hybrid Direct Preference Optimization

DGX agent

arXiv:2605.04539v1 Announce Type: new Abstract: Direct Preference Optimization (DPO), the efficient alternative to PPO-based RLHF, falls short on knowledge-intensive generation: standard preference si

model-releasesarxiv-cs-cl
7 May 2026
← Previous
1…113114115116117…1030
Next →