AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “concepts”

GridTimelineEvolution
2,168 results
Local Ai

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

DGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

local-aiarxiv-cs-cl
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

DGX agent

arXiv:2604.12373v1 Announce Type: new Abstract: Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether larg

local-aiarxiv-cs-cl
15 Apr 2026
Research

MAST: Mask-Guided Attention Mass Allocation for Training-Free Multi-Style Transfer

DGX agent

arXiv:2604.12281v1 Announce Type: cross Abstract: Style transfer aims to render a content image with the visual characteristics of a reference style while preserving its underlying semantic layout and

researcharxiv-cs-ai
15 Apr 2026
Research

On the Mathematical Relationship Between Layer Normalization and Dynamic Activation Functions

DGX agent

arXiv:2503.21708v4 Announce Type: replace-cross Abstract: Layer normalization (LN) is an essential component of modern neural networks. While many alternative techniques have been proposed, none of th

researcharxiv-cs-ai
15 Apr 2026
Research

RoleMAG: Learning Neighbor Roles in Multimodal Graphs

DGX agent

arXiv:2604.12271v1 Announce Type: new Abstract: Multimodal attributed graphs (MAGs) combine multimodal node attributes with structured relations. However, existing methods usually perform shared messa

researcharxiv-cs-lg
15 Apr 2026
Model Releases

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

DGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

DGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

researcharxiv-cs-cv
15 Apr 2026
Research

The Enforcement and Feasibility of Hate Speech Moderation on Twitter

DGX agent

arXiv:2604.12289v1 Announce Type: cross Abstract: Online hate speech is associated with substantial social harms, yet it remains unclear how consistently platforms enforce hate speech policies or whet

researcharxiv-cs-cl
15 Apr 2026
Safety

AI Integrity: A New Paradigm for Verifiable AI Governance

DGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Beyond Statistical Co-occurrence: Unlocking Intrinsic Semantics for Tabular Data Clustering

DGX agent

arXiv:2604.10865v1 Announce Type: new Abstract: Deep Clustering (DC) has emerged as a powerful tool for tabular data analysis in real-world domains like finance and healthcare. However, most existing

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

DGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Cross-Cultural Value Awareness in Large Vision-Language Models

DGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

safetyarxiv-cs-ai
14 Apr 2026
Safety

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

DGX agent

arXiv:2604.09665v1 Announce Type: cross Abstract: While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted

safetyarxiv-cs-ai
14 Apr 2026
Local Ai

Different types of syntactic agreement recruit the same units within large language models

DGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

local-aiarxiv-cs-cl
14 Apr 2026
Safety

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

DGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

safetyarxiv-cs-ai
14 Apr 2026
Applications

Domain-Specific Data Generation Framework for RAG Adaptation

DGX agent

arXiv:2510.11217v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) combines the language understanding and reasoning power of large language models (LLMs) with external ret

applicationsarxiv-cs-ai
14 Apr 2026
Safety

Evolutionary Token-Level Prompt Optimization for Diffusion Models

DGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

safetyarxiv-cs-ai
14 Apr 2026
Safety

FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer

DGX agent

arXiv:2604.10023v1 Announce Type: cross Abstract: With the growing availability of open-sourced adapters trained on the same diffusion backbone for diverse scenes and objects, combining these pretrain

safetyarxiv-cs-ai
14 Apr 2026
Safety

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

DGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

safetyarxiv-cs-lg
14 Apr 2026
Research

Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video

DGX agent

arXiv:2511.18322v3 Announce Type: replace-cross Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpretability or rely on prior assumpt

researcharxiv-cs-cv
14 Apr 2026
Safety

Legal2LogicICL: Improving Generalization in Transforming Legal Cases to Logical Formulas via Diverse Few-Shot Learning

DGX agent

arXiv:2604.11699v1 Announce Type: cross Abstract: This work aims to improve the generalization of logic-based legal reasoning systems by integrating recent advances in NLP with legal-domain adaptive f

safetyarxiv-cs-ai
14 Apr 2026
Research

Linear Programming for Multi-Criteria Assessment with Cardinal and Ordinal Data: A Pessimistic Virtual Gap Analysis

DGX agent

arXiv:2604.09555v1 Announce Type: new Abstract: Multi-criteria Analysis (MCA) is used to rank alternatives based on various criteria. Key MCA methods, such as Multiple Criteria Decision Making (MCDM)

researcharxiv-cs-ai
14 Apr 2026
Research

Linguistic Accommodation Between Neurodivergent Communities on Reddit:A Communication Accommodation Theory Analysis of ADHD and Autism Groups

DGX agent

arXiv:2604.10063v1 Announce Type: new Abstract: Social media research on mental health has focused predominantly on detecting and diagnosing conditions at the individual level. In this work, we shift

researcharxiv-cs-cl
14 Apr 2026
Safety

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

DGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

safetyarxiv-cs-ai
14 Apr 2026
Safety

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

DGX agent

arXiv:2603.20725v2 Announce Type: replace Abstract: Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on

safetyarxiv-cs-cv
14 Apr 2026
Safety

Prompt Injection as Role Confusion

DGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

safetyarxiv-cs-ai
14 Apr 2026
Research

RedNote-Vibe: A Dataset for Capturing Temporal Dynamics of AI-Generated Text in Lifestyle Social Media

DGX agent

arXiv:2509.22055v2 Announce Type: replace Abstract: We introduce RedNote-Vibe, a dataset spanning five years (pre-LLM to July 2025) sourced from lifestyle platform RedNote (Xiaohongshu), capturing the

researcharxiv-cs-cl
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Applications

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

DGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

applicationsarxiv-cs-ai
14 Apr 2026
Research

Vibe-driven model-based engineering

DGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

DGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Cross-Modal Knowledge Distillation from Spatial Transcriptomics to Histology

DGX agent

arXiv:2604.09076v1 Announce Type: new Abstract: Spatial transcriptomics provides a molecularly rich description of tissue organization, enabling unsupervised discovery of tissue niches -- spatially co

researcharxiv-cs-cv
13 Apr 2026
Safety

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

DGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Drift-Aware Online Dynamic Learning for Nonstationary Multivariate Time Series: Application to Sintering Quality Prediction

DGX agent

arXiv:2604.09358v1 Announce Type: new Abstract: Accurate prediction of nonstationary multivariate time series remains a critical challenge in complex industrial systems such as iron ore sintering. In

model-releasesarxiv-cs-lg
13 Apr 2026
Safety

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

DGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

safetyarxiv-cs-ai
13 Apr 2026
Research

Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations

DGX agent

arXiv:2604.02598v2 Announce Type: replace-cross Abstract: LLM-generated explanations can make technical content more accessible, but there is a ceiling on what they can support interactively. Because

researcharxiv-cs-ai
13 Apr 2026
Safety

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

DGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

safetyarxiv-cs-ai
13 Apr 2026
Safety

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

DGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

safetyarxiv-cs-ai
13 Apr 2026
Research

Learning Encodings by Maximizing State Distinguishability: Variational Quantum Error Correction

DGX agent

arXiv:2506.11552v2 Announce Type: replace-cross Abstract: Quantum error correction is crucial for protecting quantum information against decoherence. Traditional codes like the surface code require su

researcharxiv-cs-lg
13 Apr 2026
Safety

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

DGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

safetyarxiv-cs-ai
13 Apr 2026
Safety

MixFlow: Mixed Source Distributions Improve Rectified Flows

DGX agent

arXiv:2604.09181v1 Announce Type: new Abstract: Diffusion models and their variations, such as rectified flows, generate diverse and high-quality images, but they are still hindered by slow iterative

safetyarxiv-cs-cv
13 Apr 2026
Safety

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

DGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

safetyarxiv-cs-ai
13 Apr 2026
Research

Ranked Activation Shift for Post-Hoc Out-of-Distribution Detection

DGX agent

arXiv:2604.08572v1 Announce Type: cross Abstract: State-of-the-art post-hoc out-of-distribution detection methods rely on intermediate layer activation editing. However, they exhibit inconsistent perf

researcharxiv-cs-cv
13 Apr 2026
Safety

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

DGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

DGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Silhouette Loss: Differentiable Global Structure Learning for Deep Representations

DGX agent

arXiv:2604.08573v1 Announce Type: cross Abstract: Learning discriminative representations is a central goal of supervised deep learning. While cross-entropy (CE) remains the dominant objective for cla

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

DGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Verbalizing LLMs' assumptions to explain and control sycophancy

DGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

safetyarxiv-cs-ai
13 Apr 2026
← Previous
1…1920212223…46
Next →