AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “concepts”

GridTimelineEvolution
2,542 results
15 Apr 2026

GeM-EA: A Generative and Meta-learning Enhanced Evolutionary Algorithm for Streaming Data-Driven Optimization

Model ReleasesDGX agent

arXiv:2604.12336v1 Announce Type: cross Abstract: Streaming Data-Driven Optimization (SDDO) problems arise in many applications where data arrive continuously and the optimization environment evolves

IMU: Influence-guided Machine Unlearning

ResearchDGX agent

arXiv:2508.01620v3 Announce Type: replace-cross Abstract: Machine Unlearning (MU) aims to selectively erase the influence of specific data points from pretrained models. However, most existing MU meth

JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence

ResearchDGX agent

arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates

TutorialsDGX agent

arXiv:2604.12397v1 Announce Type: new Abstract: Standard Large Language Model (LLM) pre-training typically treats corpora as flattened token sequences, often overlooking the real-world context that hu

Latent Planning Emerges with Scale

Model ReleasesDGX agent

arXiv:2604.12493v1 Announce Type: cross Abstract: LLMs can perform seemingly planning-intensive tasks, like writing coherent stories or functioning code, without explicitly verbalizing a plan; however

LLM as Attention-Informed NTM and Topic Modeling as long-input Generation: Interpretability and long-Context Capability

ResearchDGX agent

arXiv:2510.03174v2 Announce Type: replace-cross Abstract: Topic modeling aims to produce interpretable topic representations and topic--document correspondences from corpora, but classical neural topi

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

Local AiDGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

Local AiDGX agent

arXiv:2604.12373v1 Announce Type: new Abstract: Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether larg

MAST: Mask-Guided Attention Mass Allocation for Training-Free Multi-Style Transfer

ResearchDGX agent

arXiv:2604.12281v1 Announce Type: cross Abstract: Style transfer aims to render a content image with the visual characteristics of a reference style while preserving its underlying semantic layout and

On the Mathematical Relationship Between Layer Normalization and Dynamic Activation Functions

ResearchDGX agent

arXiv:2503.21708v4 Announce Type: replace-cross Abstract: Layer normalization (LN) is an essential component of modern neural networks. While many alternative techniques have been proposed, none of th

RoleMAG: Learning Neighbor Roles in Multimodal Graphs

ResearchDGX agent

arXiv:2604.12271v1 Announce Type: new Abstract: Multimodal attributed graphs (MAGs) combine multimodal node attributes with structured relations. However, existing methods usually perform shared messa

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

Model ReleasesDGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

ResearchDGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

The Enforcement and Feasibility of Hate Speech Moderation on Twitter

ResearchDGX agent

arXiv:2604.12289v1 Announce Type: cross Abstract: Online hate speech is associated with substantial social harms, yet it remains unclear how consistently platforms enforce hate speech policies or whet

14 Apr 2026

AI Integrity: A New Paradigm for Verifiable AI Governance

SafetyDGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

Beyond Statistical Co-occurrence: Unlocking Intrinsic Semantics for Tabular Data Clustering

Model ReleasesDGX agent

arXiv:2604.10865v1 Announce Type: new Abstract: Deep Clustering (DC) has emerged as a powerful tool for tabular data analysis in real-world domains like finance and healthcare. However, most existing

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

Model ReleasesDGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

Cross-Cultural Value Awareness in Large Vision-Language Models

SafetyDGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

SafetyDGX agent

arXiv:2604.09665v1 Announce Type: cross Abstract: While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted

Different types of syntactic agreement recruit the same units within large language models

Local AiDGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

SafetyDGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

Domain-Specific Data Generation Framework for RAG Adaptation

ApplicationsDGX agent

arXiv:2510.11217v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) combines the language understanding and reasoning power of large language models (LLMs) with external ret

Evolutionary Token-Level Prompt Optimization for Diffusion Models

SafetyDGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer

SafetyDGX agent

arXiv:2604.10023v1 Announce Type: cross Abstract: With the growing availability of open-sourced adapters trained on the same diffusion backbone for diverse scenes and objects, combining these pretrain

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

SafetyDGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video

ResearchDGX agent

arXiv:2511.18322v3 Announce Type: replace-cross Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpretability or rely on prior assumpt

Legal2LogicICL: Improving Generalization in Transforming Legal Cases to Logical Formulas via Diverse Few-Shot Learning

SafetyDGX agent

arXiv:2604.11699v1 Announce Type: cross Abstract: This work aims to improve the generalization of logic-based legal reasoning systems by integrating recent advances in NLP with legal-domain adaptive f

Linear Programming for Multi-Criteria Assessment with Cardinal and Ordinal Data: A Pessimistic Virtual Gap Analysis

ResearchDGX agent

arXiv:2604.09555v1 Announce Type: new Abstract: Multi-criteria Analysis (MCA) is used to rank alternatives based on various criteria. Key MCA methods, such as Multiple Criteria Decision Making (MCDM)

Linguistic Accommodation Between Neurodivergent Communities on Reddit:A Communication Accommodation Theory Analysis of ADHD and Autism Groups

ResearchDGX agent

arXiv:2604.10063v1 Announce Type: new Abstract: Social media research on mental health has focused predominantly on detecting and diagnosing conditions at the individual level. In this work, we shift

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

SafetyDGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

SafetyDGX agent

arXiv:2603.20725v2 Announce Type: replace Abstract: Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on

Prompt Injection as Role Confusion

SafetyDGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

RedNote-Vibe: A Dataset for Capturing Temporal Dynamics of AI-Generated Text in Lifestyle Social Media

ResearchDGX agent

arXiv:2509.22055v2 Announce Type: replace Abstract: We introduce RedNote-Vibe, a dataset spanning five years (pre-LLM to July 2025) sourced from lifestyle platform RedNote (Xiaohongshu), capturing the

Steered LLM Activations are Non-Surjective

SafetyDGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

ApplicationsDGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

Vibe-driven model-based engineering

ResearchDGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

Model ReleasesDGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

13 Apr 2026

Back further in history: single cell versus multicellular life mech battles

ApplicationsDGX agent

This post by Ethan Mollick likely explores AI-generated or conceptual visualizations of battles between single-celled and multicellular organisms framed as mech combat, drawing on biological history f

Cross-Modal Knowledge Distillation from Spatial Transcriptomics to Histology

ResearchDGX agent

arXiv:2604.09076v1 Announce Type: new Abstract: Spatial transcriptomics provides a molecularly rich description of tissue organization, enabling unsupervised discovery of tissue niches -- spatially co

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

SafetyDGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

Drift-Aware Online Dynamic Learning for Nonstationary Multivariate Time Series: Application to Sintering Quality Prediction

Model ReleasesDGX agent

arXiv:2604.09358v1 Announce Type: new Abstract: Accurate prediction of nonstationary multivariate time series remains a critical challenge in complex industrial systems such as iron ore sintering. In

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations

ResearchDGX agent

arXiv:2604.02598v2 Announce Type: replace-cross Abstract: LLM-generated explanations can make technical content more accessible, but there is a ceiling on what they can support interactively. Because

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

SafetyDGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

Learning Encodings by Maximizing State Distinguishability: Variational Quantum Error Correction

ResearchDGX agent

arXiv:2506.11552v2 Announce Type: replace-cross Abstract: Quantum error correction is crucial for protecting quantum information against decoherence. Traditional codes like the surface code require su

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

SafetyDGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All o…

TutorialsDGX agent

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All of them are variations of GRPO and all of them are incrementa

MixFlow: Mixed Source Distributions Improve Rectified Flows

SafetyDGX agent

arXiv:2604.09181v1 Announce Type: new Abstract: Diffusion models and their variations, such as rectified flows, generate diverse and high-quality images, but they are still hindered by slow iterative

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

SafetyDGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

一つのニューラルネットに符号と記号は創発しうるか? 「Neural Computers」論文から考える @rmaruy https://rmaruy.hatenablog.com/entry/2026/04/11/223828

ResearchDGX agent

This Japanese blog post explores whether signs and symbols can emerge within a single neural network, drawing on analysis of the 'Neural Computers' paper. The discussion likely examines the intersecti

Ranked Activation Shift for Post-Hoc Out-of-Distribution Detection

ResearchDGX agent

arXiv:2604.08572v1 Announce Type: cross Abstract: State-of-the-art post-hoc out-of-distribution detection methods rely on intermediate layer activation editing. However, they exhibit inconsistent perf

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

SafetyDGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

Model ReleasesDGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

Silhouette Loss: Differentiable Global Structure Learning for Deep Representations

ResearchDGX agent

arXiv:2604.08573v1 Announce Type: cross Abstract: Learning discriminative representations is a central goal of supervised deep learning. While cross-entropy (CE) remains the dominant objective for cla

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

Verbalizing LLMs' assumptions to explain and control sycophancy

SafetyDGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

ResearchDGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

You Can't Fight in Here! This is BBS!

TutorialsDGX agent

arXiv:2604.09501v1 Announce Type: new Abstract: Norm, the formal theoretical linguist, and Claudette, the computational language scientist, have a lovely time discussing whether modern language models

12 Apr 2026

From this perspective, Gemini is also worse than the original Bard. Sydney was the original sin of LLMs anthropormism, but also got the idea…

Model ReleasesDGX agent

From this perspective, Gemini is also worse than the original Bard. Sydney was the original sin of LLMs anthropormism, but also got the idea that AIs can sometimes be better with personalities right.

← Previous
1…1819202122…43
Next →