AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Continual Learning with Multilingual Foundation Model

DGX agent

arXiv:2605.13415v1 Announce Type: cross Abstract: This paper presents a multi-stage framework for detecting reclaimed slurs in multilingual social media discourse. It addresses the challenge of identi

researcharxiv-cs-ai
14 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Diffusion Model's Generalization Can Be Characterized by Inductive Biases toward a Data-Dependent Ridge Manifold

DGX agent

arXiv:2602.06021v2 Announce Type: replace-cross Abstract: We study a data-dependent notion of diffusion-model generalization: when a model does not memorize the training set, where do its generated sa

safetyarxiv-cs-lg
14 May 2026
Safety

DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense

DGX agent

arXiv:2605.13115v1 Announce Type: cross Abstract: Diffusion models depend on pseudo-random number generators (PRNGs) for latent noise sampling. We present DiffusionHijack, a supply-chain backdoor atta

safetyarxiv-cs-lg
14 May 2026
Model Releases

DistractMIA: Black-Box Membership Inference on Vision-Language Models via Semantic Distraction

DGX agent

arXiv:2605.12574v1 Announce Type: cross Abstract: Vision-language models (VLMs) are trained on large-scale image-text corpora that may contain private, copyrighted, or otherwise sensitive data, motiva

model-releasesarxiv-cs-ai
14 May 2026
Applications

It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

DGX agent

arXiv:2605.13245v1 Announce Type: new Abstract: Language models can produce convincing scientific analyses, but repeated generations on the same data do not guarantee the same result. A researcher may

applicationsarxiv-cs-ai
14 May 2026
Local Ai

KAST-BAR: Knowledge-Anchored Semantically-Dynamic Topology Brain Autoregressive Modeling for Universal Neural Interpretation

DGX agent

arXiv:2605.13133v1 Announce Type: new Abstract: While EEG foundation models have shown significant potential in universal neural decoding across tasks, their advancement remains constrained by the ina

local-aiarxiv-cs-lg
14 May 2026
Research

Latent-Augmented Discrete Diffusion Models

DGX agent

arXiv:2510.18114v3 Announce Type: replace-cross Abstract: Discrete diffusion models have emerged as a powerful class of models and a promising route to fast language generation, but practical implemen

researcharxiv-cs-ai
14 May 2026
Local Ai

LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters

DGX agent

arXiv:2605.13163v1 Announce Type: cross Abstract: Foundation models and low-rank adapters enable efficient on-device generative AI but raise risks such as intellectual property leakage and model recov

local-aiarxiv-cs-cv
14 May 2026
Agents

Mechanism Plausibility in Generative Agent-Based Modeling

DGX agent

arXiv:2605.12824v1 Announce Type: cross Abstract: Large language models (LLMs) can generate high-level diverse phenomena without explicitly programmed rules. This capability has led to their adoption

agentsarxiv-cs-ai
14 May 2026
Research

MIDST Challenge at SaTML 2025: Membership Inference over Diffusion-models-based Synthetic Tabular data

DGX agent

arXiv:2603.19185v2 Announce Type: replace Abstract: Synthetic data is often perceived as a silver-bullet solution to data anonymization and privacy-preserving data publishing. Drawn from generative mo

researcharxiv-cs-lg
14 May 2026
Research

QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling

DGX agent

arXiv:2605.13833v1 Announce Type: cross Abstract: Modeling long-range dependencies in sequential data remains a central challenge in machine learning. Transformers address this challenge through atten

researcharxiv-cs-cv
14 May 2026
Research

The Expressivity Boundary of Probabilistic Circuits: A Comparison with Large Language Models

DGX agent

arXiv:2605.12940v1 Announce Type: cross Abstract: Probabilistic Circuits (PCs) are deep generative models that support exact and efficient probabilistic inference. Yet in autoregressive language model

researcharxiv-cs-ai
14 May 2026
Research

The Score-Difference Flow for Implicit Generative Modeling

DGX agent

arXiv:2304.12906v4 Announce Type: replace Abstract: Implicit generative modeling (IGM) aims to produce samples of synthetic data matching the characteristics of a target data distribution. Recent work

researcharxiv-cs-lg
14 May 2026
Research

VIP-COP: Context Optimization for Tabular Foundation Models

DGX agent

arXiv:2605.12904v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have emerged as a powerful paradigm for in-context learning on structured data, enabling direct prediction on new tabul

researcharxiv-cs-lg
14 May 2026
Safety

What properties of reasoning supervision are associated with improved downstream model quality?

DGX agent

arXiv:2605.13290v1 Announce Type: new Abstract: Validating training data for reasoning models typically requires expensive trial-and-error fine-tuning cycles. In this work, we investigate whether the

safetyarxiv-cs-ai
14 May 2026
Model Releases

A Study on Hidden Layer Distillation for Large Language Model Pre-Training

DGX agent

arXiv:2605.11513v1 Announce Type: new Abstract: Knowledge Distillation (KD) is a critical tool for training Large Language Models (LLMs), yet the majority of research focuses on approaches that rely s

model-releasesarxiv-cs-cl
13 May 2026
Research

Active inference as a unified model of collision avoidance behavior in human drivers

DGX agent

arXiv:2506.02215v5 Announce Type: replace Abstract: Collision avoidance -- involving a rapid threat detection and quick execution of the appropriate evasive maneuver -- is a critical aspect of driving

researcharxiv-cs-ro
13 May 2026
Model Releases

ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder

DGX agent

arXiv:2605.11091v1 Announce Type: new Abstract: Automated ASD screening tools remain limited by single-architecture evaluations, axis-restricted assessment, and near-exclusive focus on adult cohorts,

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models

DGX agent

arXiv:2605.11726v1 Announce Type: new Abstract: Recently, reinforcement learning (RL) has been widely applied during post-training for diffusion large language models (dLLMs) to enhance reasoning with

model-releasesarxiv-cs-lg
13 May 2026
Safety

Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models

DGX agent

arXiv:2601.21944v2 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, thes

safetyarxiv-cs-lg
13 May 2026
Safety

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion

DGX agent

arXiv:2505.18780v3 Announce Type: replace-cross Abstract: Achieving versatile humanoid locomotion with a single policy presents a critical scalability challenge. Prevailing methods often rely on disti

safetyarxiv-cs-lg
13 May 2026
Safety

EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records

DGX agent

arXiv:2605.12335v1 Announce Type: cross Abstract: Electronic Health Records (EHR) contain rich longitudinal patient information and are widely used in predictive modeling applications. However, effect

safetyarxiv-cs-lg
13 May 2026
Research

Express Your Doubts -- Probabilistic World Modeling Should not be Based on Token logprobs

DGX agent

arXiv:2505.02072v2 Announce Type: replace Abstract: Language modeling has shifted in recent years from a distribution over strings to prediction models with textual inputs and outputs for general-purp

researcharxiv-cs-cl
13 May 2026
Safety

Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2605.12374v1 Announce Type: new Abstract: Visual latent reasoning lets a multimodal large language model (MLLM) create intermediate visual evidence as continuous tokens, avoiding external tools

safetyarxiv-cs-cv
13 May 2026
Model Releases

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

DGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

model-releasesarxiv-cs-ro
13 May 2026
Research

Molecular Design beyond Training Data with Novel Extended Objective Functionals of Generative AI Models Driven by Quantum Annealing Computer

DGX agent

arXiv:2602.15451v3 Announce Type: replace-cross Abstract: Deep generative modeling to stochastically design small molecules is an emerging technology for accelerating drug discovery and development. H

researcharxiv-cs-lg
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Model Releases

See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model

DGX agent

arXiv:2605.11817v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown remarkable promise in robotics manipulation, yet their high computational cost hinders real-time deploy

model-releasesarxiv-cs-cv
13 May 2026
Safety

STRIDE: Training-Free Diversity Guidance via PCA-Directed Feature Perturbation in Single-Step Diffusion Models

DGX agent

arXiv:2605.11494v1 Announce Type: new Abstract: Distilled one-step (T=1) or few-step (Tleq4) diffusion models enable real-time image generation but often exhibit reduced sample diversity compared to t

safetyarxiv-cs-cv
13 May 2026
Local Ai

Structural Interpretations of Protein Language Model Representations via Differentiable Graph Partitioning

DGX agent

arXiv:2605.10985v1 Announce Type: new Abstract: Protein language models such as ESM-2 learn rich residue representations that achieve strong performance on protein function prediction, but their featu

local-aiarxiv-cs-lg
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Research

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

DGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

researcharxiv-cs-cv
13 May 2026
Research

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

DGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

researcharxiv-cs-cl
13 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Attention Drift: What Autoregressive Speculative Decoding Models Learn

DGX agent

arXiv:2605.09992v1 Announce Type: cross Abstract: Speculative decoding accelerates LLM inference by drafting future tokens with a small model, but drafter models degrade sharply under template perturb

tutorialsarxiv-cs-ai
12 May 2026
Safety

Bi-CoG: Bi-Consistency-Guided Self-Training for Vision-Language Models

DGX agent

arXiv:2510.20477v2 Announce Type: replace Abstract: Exploiting unlabeled data through semi-supervised learning (SSL) or leveraging pre-trained models via fine-tuning are two prevailing paradigms for a

safetyarxiv-cs-lg
12 May 2026
Model Releases

C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving

DGX agent

arXiv:2605.10744v1 Announce Type: new Abstract: Safety-critical planning in complex environments, particularly at urban intersections, remains a fundamental challenge for autonomous driving. Existing

model-releasesarxiv-cs-cv
12 May 2026
Research

DARE: Diffusion Language Model Activation Reuse for Efficient Inference

DGX agent

arXiv:2605.08134v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to auto-regressive (AR) models, offering greater expressive capacity a

researcharxiv-cs-ai
12 May 2026
Research

Do multimodal models imagine electric sheep?

DGX agent

arXiv:2605.09693v1 Announce Type: cross Abstract: Yes. We find that large multimodal models develop mental imagery when solving spatial puzzles, and they do imagine sheep when solving sheep puzzles. W

researcharxiv-cs-ai
12 May 2026
Agents

DriveFuture: Future-Aware Latent World Models for Autonomous Driving

DGX agent

arXiv:2605.09701v1 Announce Type: new Abstract: Existing latent world models for autonomous driving have opened a promising path toward future-aware driving intelligence. However, they typically treat

agentsarxiv-cs-cv
12 May 2026
Local Ai

Dystruct: Dynamically Structured Diffusion Language Model Decoding via Bayesian Inference

DGX agent

arXiv:2605.09820v1 Announce Type: new Abstract: Diffusion language models (DLMs) have recently emerged as a promising alternative to autoregressive models, primarily due to their ability to enable par

local-aiarxiv-cs-lg
12 May 2026
Research

Edit-Based Refinement for Parallel Masked Diffusion Language Models

DGX agent

arXiv:2605.09603v1 Announce Type: new Abstract: Masked diffusion language models enable parallel token generation and offer improved decoding efficiency over autoregressive models. However, their perf

researcharxiv-cs-cl
12 May 2026
Model Releases

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models

DGX agent

arXiv:2605.09773v1 Announce Type: cross Abstract: We use sparse autoencoder (SAE) feature steering to amplify Dark Triad personality traits (Machiavellianism, narcissism, and psychopathy) in Llama-3.3

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Failing Forward: Adaptive Failure-Informed Learning for Vision-Language-Action Models

DGX agent

arXiv:2605.08434v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising paradigm for scalable robotic manipulation, yet their reliance on success-only behavioral clonin

model-releasesarxiv-cs-ro
12 May 2026
Local Ai

FedGMI: Generative Model-Driven Federated Learning for Probabilistic Mixture Inference

DGX agent

arXiv:2605.08760v1 Announce Type: new Abstract: Federated Learning (FL) facilitates collaborative model training across decentralized clients while preserving data privacy by avoiding raw data exchang

local-aiarxiv-cs-lg
12 May 2026
Model Releases

Filtering Memorization from Parameter-Space in Diffusion Models

DGX agent

arXiv:2605.10439v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely used mechanism for customizing diffusion models, enabling users to inject new visual concepts or styles t

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models

DGX agent

arXiv:2603.16253v2 Announce Type: replace-cross Abstract: Vision-language process reward models (VL-PRMs) are increasingly used to score intermediate reasoning steps and rerank candidates under test-t

model-releasesarxiv-cs-ai
12 May 2026
Research

High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models

DGX agent

arXiv:2512.21815v2 Announce Type: replace Abstract: Vision-language models (VLMs) achieve remarkable performance but remain vulnerable to adversarial attacks. Entropy, as a measure of model uncertaint

researcharxiv-cs-cv
12 May 2026
← Previous
1…112113114115116…1030
Next →