AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “concepts”

GridTimelineEvolution
2,525 results
Safety

Cross-Cultural Value Awareness in Large Vision-Language Models

DGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

safetyarxiv-cs-ai
14 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

DGX agent

arXiv:2604.09665v1 Announce Type: cross Abstract: While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted

safetyarxiv-cs-ai
14 Apr 2026
Local Ai

Different types of syntactic agreement recruit the same units within large language models

DGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

local-aiarxiv-cs-cl
14 Apr 2026
Safety

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

DGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

safetyarxiv-cs-ai
14 Apr 2026
Applications

Domain-Specific Data Generation Framework for RAG Adaptation

DGX agent

arXiv:2510.11217v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) combines the language understanding and reasoning power of large language models (LLMs) with external ret

applicationsarxiv-cs-ai
14 Apr 2026
Safety

Evolutionary Token-Level Prompt Optimization for Diffusion Models

DGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

safetyarxiv-cs-ai
14 Apr 2026
Safety

FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer

DGX agent

arXiv:2604.10023v1 Announce Type: cross Abstract: With the growing availability of open-sourced adapters trained on the same diffusion backbone for diverse scenes and objects, combining these pretrain

safetyarxiv-cs-ai
14 Apr 2026
Safety

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

DGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

safetyarxiv-cs-lg
14 Apr 2026
Research

Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video

DGX agent

arXiv:2511.18322v3 Announce Type: replace-cross Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpretability or rely on prior assumpt

researcharxiv-cs-cv
14 Apr 2026
Safety

Legal2LogicICL: Improving Generalization in Transforming Legal Cases to Logical Formulas via Diverse Few-Shot Learning

DGX agent

arXiv:2604.11699v1 Announce Type: cross Abstract: This work aims to improve the generalization of logic-based legal reasoning systems by integrating recent advances in NLP with legal-domain adaptive f

safetyarxiv-cs-ai
14 Apr 2026
Research

Linear Programming for Multi-Criteria Assessment with Cardinal and Ordinal Data: A Pessimistic Virtual Gap Analysis

DGX agent

arXiv:2604.09555v1 Announce Type: new Abstract: Multi-criteria Analysis (MCA) is used to rank alternatives based on various criteria. Key MCA methods, such as Multiple Criteria Decision Making (MCDM)

researcharxiv-cs-ai
14 Apr 2026
Research

Linguistic Accommodation Between Neurodivergent Communities on Reddit:A Communication Accommodation Theory Analysis of ADHD and Autism Groups

DGX agent

arXiv:2604.10063v1 Announce Type: new Abstract: Social media research on mental health has focused predominantly on detecting and diagnosing conditions at the individual level. In this work, we shift

researcharxiv-cs-cl
14 Apr 2026
Safety

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

DGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

safetyarxiv-cs-ai
14 Apr 2026
Safety

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

DGX agent

arXiv:2603.20725v2 Announce Type: replace Abstract: Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on

safetyarxiv-cs-cv
14 Apr 2026
Safety

Prompt Injection as Role Confusion

DGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

safetyarxiv-cs-ai
14 Apr 2026
Research

RedNote-Vibe: A Dataset for Capturing Temporal Dynamics of AI-Generated Text in Lifestyle Social Media

DGX agent

arXiv:2509.22055v2 Announce Type: replace Abstract: We introduce RedNote-Vibe, a dataset spanning five years (pre-LLM to July 2025) sourced from lifestyle platform RedNote (Xiaohongshu), capturing the

researcharxiv-cs-cl
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Applications

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

DGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

applicationsarxiv-cs-ai
14 Apr 2026
Research

Vibe-driven model-based engineering

DGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

DGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Back further in history: single cell versus multicellular life mech battles

DGX agent

This post by Ethan Mollick likely explores AI-generated or conceptual visualizations of battles between single-celled and multicellular organisms framed as mech combat, drawing on biological history f

applicationsethan-mollick--x
13 Apr 2026
Research

Cross-Modal Knowledge Distillation from Spatial Transcriptomics to Histology

DGX agent

arXiv:2604.09076v1 Announce Type: new Abstract: Spatial transcriptomics provides a molecularly rich description of tissue organization, enabling unsupervised discovery of tissue niches -- spatially co

researcharxiv-cs-cv
13 Apr 2026
Safety

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

DGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Drift-Aware Online Dynamic Learning for Nonstationary Multivariate Time Series: Application to Sintering Quality Prediction

DGX agent

arXiv:2604.09358v1 Announce Type: new Abstract: Accurate prediction of nonstationary multivariate time series remains a critical challenge in complex industrial systems such as iron ore sintering. In

model-releasesarxiv-cs-lg
13 Apr 2026
Safety

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

DGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

safetyarxiv-cs-ai
13 Apr 2026
Research

Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations

DGX agent

arXiv:2604.02598v2 Announce Type: replace-cross Abstract: LLM-generated explanations can make technical content more accessible, but there is a ceiling on what they can support interactively. Because

researcharxiv-cs-ai
13 Apr 2026
Safety

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

DGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

safetyarxiv-cs-ai
13 Apr 2026
Safety

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

DGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

safetyarxiv-cs-ai
13 Apr 2026
Research

Learning Encodings by Maximizing State Distinguishability: Variational Quantum Error Correction

DGX agent

arXiv:2506.11552v2 Announce Type: replace-cross Abstract: Quantum error correction is crucial for protecting quantum information against decoherence. Traditional codes like the surface code require su

researcharxiv-cs-lg
13 Apr 2026
Safety

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

DGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

safetyarxiv-cs-ai
13 Apr 2026
Tutorials

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All o…

DGX agent

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All of them are variations of GRPO and all of them are incrementa

tutorialsjeremy-howard--x
13 Apr 2026
Safety

MixFlow: Mixed Source Distributions Improve Rectified Flows

DGX agent

arXiv:2604.09181v1 Announce Type: new Abstract: Diffusion models and their variations, such as rectified flows, generate diverse and high-quality images, but they are still hindered by slow iterative

safetyarxiv-cs-cv
13 Apr 2026
Safety

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

DGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

safetyarxiv-cs-ai
13 Apr 2026
Research

一つのニューラルネットに符号と記号は創発しうるか? 「Neural Computers」論文から考える @rmaruy https://rmaruy.hatenablog.com/entry/2026/04/11/223828

DGX agent

This Japanese blog post explores whether signs and symbols can emerge within a single neural network, drawing on analysis of the 'Neural Computers' paper. The discussion likely examines the intersecti

researchdavid-ha--x
13 Apr 2026
Research

Ranked Activation Shift for Post-Hoc Out-of-Distribution Detection

DGX agent

arXiv:2604.08572v1 Announce Type: cross Abstract: State-of-the-art post-hoc out-of-distribution detection methods rely on intermediate layer activation editing. However, they exhibit inconsistent perf

researcharxiv-cs-cv
13 Apr 2026
Safety

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

DGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

DGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Silhouette Loss: Differentiable Global Structure Learning for Deep Representations

DGX agent

arXiv:2604.08573v1 Announce Type: cross Abstract: Learning discriminative representations is a central goal of supervised deep learning. While cross-entropy (CE) remains the dominant objective for cla

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

DGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Verbalizing LLMs' assumptions to explain and control sycophancy

DGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

safetyarxiv-cs-ai
13 Apr 2026
Research

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

DGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

researcharxiv-cs-ai
13 Apr 2026
Tutorials

You Can't Fight in Here! This is BBS!

DGX agent

arXiv:2604.09501v1 Announce Type: new Abstract: Norm, the formal theoretical linguist, and Claudette, the computational language scientist, have a lovely time discussing whether modern language models

tutorialsarxiv-cs-cl
13 Apr 2026
Model Releases

From this perspective, Gemini is also worse than the original Bard. Sydney was the original sin of LLMs anthropormism, but also got the idea…

DGX agent

From this perspective, Gemini is also worse than the original Bard. Sydney was the original sin of LLMs anthropormism, but also got the idea that AIs can sometimes be better with personalities right.

model-releasesethan-mollick--x
12 Apr 2026
Syntheses

Wiki Lint Report — 2026-04-12

DGX agent

Automated lint: 34 errors, 0 warnings, 3 info

linthealth-checkautomated
12 Apr 2026
Research

FlashAttention (FA1–FA4) in PyTorch - educational implementations focused on algorithmic differences [P]

DGX agent

This r/MachineLearning post presents educational PyTorch implementations of FlashAttention versions 1 through 4, designed to highlight the key algorithmic differences across each iteration rather than

researchr-machinelearning
11 Apr 2026
Model Releases

ADAG: Automatically Describing Attribution Graphs

DGX agent

arXiv:2604.07615v1 Announce Type: new Abstract: In language model interpretability research, extbf{circuit tracing} aims to identify which internal features causally contributed to a particular outp

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

AI generates well-liked but templatic empathic responses

DGX agent

arXiv:2604.08479v1 Announce Type: new Abstract: Recent research shows that greater numbers of people are turning to Large Language Models (LLMs) for emotional support, and that people rate LLM respons

applicationsarxiv-cs-cl
10 Apr 2026
Model Releases

Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Models

DGX agent

arXiv:2604.07749v1 Announce Type: new Abstract: Large language models (LLMs) can shift their answers under pressure in ways that reflect accommodation rather than reasoning. Prior work on sycophancy h

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…2324252627…53
Next →