AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
13 Apr 2026

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition

SafetyDGX agent

arXiv:2604.09063v1 Announce Type: cross Abstract: Human action recognition is pivotal in computer vision, with applications ranging from surveillance to human-robot interaction. Despite the effectiven

From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales

SafetyDGX agent

arXiv:2604.08591v1 Announce Type: cross Abstract: Hallucinations in large ASR models present a critical safety risk. In this work, we propose the extit{Spectral Sensitivity Theorem}, which predicts a

From Selection to Scheduling: Federated Geometry-Aware Correction Makes Exemplar Replay Work Better under Continual Dynamic Heterogeneity

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.08617v1 Announce Type: cross Abstract: Exemplar replay has become an effective strategy for mitigating catastrophic forgetting in federated continual learning (FCL) by retaining representat

GAN-Enhanced Deep Reinforcement Learning for Semantic-Aware Resource Allocation in 6G Network Slicing

SafetyDGX agent

arXiv:2604.08576v1 Announce Type: cross Abstract: Sixth-generation (6G) wireless networks must support heterogeneous services: enhanced Mobile Broadband (eMBB) requiring 1 Tbps data rates, massive Mac

💯. “@garymarcus has won, others too” And the smart people already know it. AGI requires both symbols and neural networks. Period.

SafetyDGX agent

💯. “@garymarcus has won, others too” And the smart people already know it. AGI requires both symbols and neural networks. Period. The current regime in AI cannot possibly scale Mount AGI. Gary Marcus

Gated-SwinRMT: Unifying Swin Windowed Attention with Retentive Manhattan Decay via Input-Dependent Gating

SafetyDGX agent

arXiv:2604.06014v2 Announce Type: replace Abstract: We introduce Gated-SwinRMT, a family of hybrid vision transformers that combine the shifted-window attention of the Swin Transformer with the Manhat

Gender bias

SafetyDGX agent

This Reddit post on r/ChatGPT titled 'Gender bias' likely features a user-shared observation or experiment highlighting instances where ChatGPT exhibits gender-skewed responses, such as stereotyping o

General Purpose Technologies have downstream effects everywhere, good and bad. Those outcomes can be mitigated, or encouraged, with the righ…

SafetyDGX agent

General Purpose Technologies have downstream effects everywhere, good and bad. Those outcomes can be mitigated, or encouraged, with the right sorts of policy choices If the only options are being for

Generative Simulation for Policy Learning in Physical Human-Robot Interaction

SafetyDGX agent

arXiv:2604.08664v1 Announce Type: new Abstract: Developing autonomous physical human-robot interaction (pHRI) systems is limited by the scarcity of large-scale training data to learn robust robot beha

Geometry-Induced Long-Range Correlations in Recurrent Neural Network Quantum States

SafetyDGX agent

arXiv:2604.08661v1 Announce Type: cross Abstract: Neural Quantum States based on autoregressive recurrent neural network (RNN) wave functions enable efficient sampling without Markov-chain autocorrela

GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback

SafetyDGX agent

arXiv:2604.08553v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong performance on text-attributed graphs (TAGs) due to their superior semantic understanding ability on te

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency

SafetyDGX agent

arXiv:2604.09075v1 Announce Type: new Abstract: Large language models increasingly operate under multiple instructions from heterogeneous sources with different authority levels, including system poli

Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation

SafetyDGX agent

arXiv:2604.09231v1 Announce Type: new Abstract: Although recent advances have improved the quality of 3D texture generation, existing methods still struggle with incomplete texture coverage, cross-vie

How Similar Are Grokipedia and Wikipedia? A Multi-Dimensional Textual and Structural Comparison

SafetyDGX agent

arXiv:2510.26899v5 Announce Type: replace-cross Abstract: The launch of Grokipedia, an AI-generated encyclopedia developed by Elon Musk's xAI, was presented as a response to perceived ideological and

I am catching glimpses in my feed that there is a backlash against Mythos as 'marketing hype,' and it is a little confusing. I don't think a…

SafetyDGX agent

I am catching glimpses in my feed that there is a backlash against Mythos as 'marketing hype,' and it is a little confusing. I don't think anyone who has used the latest agentic coding tools, would th

Implicit Bias in Deep Linear Discriminant Analysis

SafetyDGX agent

arXiv:2603.02622v2 Announce Type: replace Abstract: While the Implicit Bias(or Implicit Regularization) of standard loss functions has been studied, the optimization geometry induced by discriminative

Import AI 453: Breaking AI agents; MirrorCode; and ten views on gradual disempowerment

SafetyDGX agent

Import AI issue 453 covers research and developments around vulnerabilities in AI agent systems, including methods for breaking or adversarially manipulating AI agents. The issue also features MirrorC

InstrAct: Towards Action-Centric Understanding in Instructional Videos

SafetyDGX agent

arXiv:2604.08762v1 Announce Type: cross Abstract: Understanding instructional videos requires recognizing fine-grained actions and modeling their temporal relations, which remains challenging for curr

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

SafetyDGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

Large Reasoning Models Learn Better Alignment from Flawed Thinking

SafetyDGX agent

arXiv:2510.00938v2 Announce Type: replace Abstract: Large reasoning models (LRMs) 'think' by generating structured chain-of-thought (CoT) before producing a final answer, yet they still lack the abili

Learning Vision-Language-Action World Models for Autonomous Driving

SafetyDGX agent

arXiv:2604.09059v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved notable progress in end-to-end autonomous driving by integrating perception, reasoning, and

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

SafetyDGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving

SafetyDGX agent

arXiv:2604.08719v1 Announce Type: cross Abstract: Recent years have seen remarkable progress in autonomous driving, yet generalization to long-tail and open-world scenarios remains a major bottleneck

Long-SCOPE: Fully Sparse Long-Range Cooperative 3D Perception

SafetyDGX agent

arXiv:2604.09206v1 Announce Type: new Abstract: Cooperative 3D perception via Vehicle-to-Everything communication is a promising paradigm for enhancing autonomous driving, offering extended sensing ho

MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits

SafetyDGX agent

arXiv:2604.08952v1 Announce Type: new Abstract: Document Question Answering (DQA) involves generating answers from a document based on a user's query, representing a key task in document understanding

Many Preferences, Few Policies: Towards Scalable Language Model Personalization

SafetyDGX agent

arXiv:2604.04144v2 Announce Type: replace-cross Abstract: The holy grail of LLM personalization is a single LLM for each user, perfectly aligned with that user's preferences. However, maintaining a se

MARBLE: Multi-Armed Restless Bandits in Latent Markovian Environment

SafetyDGX agent

arXiv:2511.09324v2 Announce Type: replace Abstract: Restless Multi-Armed Bandits (RMABs) are powerful models for decision-making under uncertainty, yet classical formulations typically assume fixed dy

Mechanisms of Introspective Awareness

SafetyDGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

Memo: OpenAI Chief Revenue Officer Denise Dresser says Anthropic is 'grossing up rev share with Amazon and Google' and overstating its 'run …

SafetyDGX agent

Memo: OpenAI Chief Revenue Officer Denise Dresser says Anthropic is 'grossing up rev share with Amazon and Google' and overstating its 'run rate by roughly $8B' (@haydenfield / The Verge) https://www.

MeshOn: Intersection-Free Mesh-to-Mesh Composition

SafetyDGX agent

arXiv:2604.08799v1 Announce Type: cross Abstract: We propose MeshOn, a method that finds physically and semantically realistic compositions of two input meshes. Given an accessory, a base mesh with a

MixFlow: Mixed Source Distributions Improve Rectified Flows

SafetyDGX agent

arXiv:2604.09181v1 Announce Type: new Abstract: Diffusion models and their variations, such as rectified flows, generate diverse and high-quality images, but they are still hindered by slow iterative

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

SafetyDGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

MSMO-ABSA: Multi-Scale and Multi-Objective Optimization for Cross-Lingual Aspect-Based Sentiment Analysis

SafetyDGX agent

arXiv:2502.13718v2 Announce Type: replace Abstract: Aspect-based sentiment analysis (ABSA) garnered growing research interest in multilingual contexts in the past. However, the majority of the studies

Multimodal Anomaly Detection for Human-Robot Interaction

SafetyDGX agent

arXiv:2604.09326v1 Announce Type: cross Abstract: Ensuring safety and reliability in human-robot interaction (HRI) requires the timely detection of unexpected events that could lead to system failures

Musculoskeletal Motion Imitation for Learning Personalized Exoskeleton Control Policy in Impaired Gait

SafetyDGX agent

arXiv:2604.09431v1 Announce Type: new Abstract: Designing generalizable control policies for lower-limb exoskeletons remains fundamentally constrained by exhaustive data collection or iterative optimi

MuTSE: A Human-in-the-Loop Multi-use Text Simplification Evaluator

SafetyDGX agent

arXiv:2604.08947v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly prevalent in text simplification, systematically evaluating their outputs across diverse prompting

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

SafetyDGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System

SafetyDGX agent

arXiv:2604.09069v1 Announce Type: cross Abstract: Court Judgment Prediction and Explanation (CJPE) aims to predict a judicial decision and provide a legally grounded explanation for a given case based

On Divergence Measures for Training GFlowNets

SafetyDGX agent

arXiv:2410.09355v2 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are amortized inference models designed to sample from unnormalized distributions over composable objects, with a

On the Representational Limits of Quantum-Inspired 1024-D Document Embeddings: An Experimental Evaluation Framework

SafetyDGX agent

arXiv:2604.09430v1 Announce Type: cross Abstract: Text embeddings are central to modern information retrieval and Retrieval-Augmented Generation (RAG). While dense models derived from Large Language M

On the Role of DAG topology in Energy-Aware Cloud Scheduling : A GNN-Based Deep Reinforcement Learning Approach

SafetyDGX agent

arXiv:2604.09202v1 Announce Type: cross Abstract: Cloud providers must assign heterogeneous compute resources to workflow DAGs while balancing competing objectives such as completion time, cost, and e

On the Spectral Geometry of Cross-Modal Representations: A Functional Map Diagnostic for Multimodal Alignment

SafetyDGX agent

arXiv:2604.08579v1 Announce Type: cross Abstract: We study cross-modal alignment between independently pretrained vision (DINOv2) and language (all-MiniLM-L6-v2) encoders using the functional map fram

Open secret that the big AI labs see themselves as emergent state-like entities in the vein of the distributed polities in “Diamond Age” or …

SafetyDGX agent

Open secret that the big AI labs see themselves as emergent state-like entities in the vein of the distributed polities in “Diamond Age” or “Terra Ignota”, which is pretty funny bc staff at said labs

OpenKedge: Governing Agentic Mutation with Execution-Bound Safety and Evidence Chains

SafetyDGX agent

arXiv:2604.08601v1 Announce Type: new Abstract: The rise of autonomous AI agents exposes a fundamental flaw in API-centric architectures: probabilistic systems directly execute state mutations without

PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment

SafetyDGX agent

arXiv:2604.08986v1 Announce Type: cross Abstract: Persona prompting has been widely adopted to steer large language models (LLMs) behavior and improve their instruction performance by assigning specif

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks

SafetyDGX agent

arXiv:2604.09028v1 Announce Type: cross Abstract: Unmanned aerial vehicles serving as aerial base stations can rapidly restore connectivity after disasters, yet abrupt changes in user mobility and tra

Policy-Aware Design of Large-Scale Factorial Experiments

SafetyDGX agent

arXiv:2604.08804v1 Announce Type: cross Abstract: Digital firms routinely run many online experiments on shared user populations. When product decisions are compositional, such as combinations of inte

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotia…

SafetyDGX agent

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotiations: fell apart within 21 hours of Vance’s arrival Welcome

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

SafetyDGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

Post-Selection Distributional Model Evaluation

SafetyDGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

SafetyDGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

SafetyDGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

SafetyDGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

Reducing Class Bias In Data-Balanced Datasets Through Hardness-Based Resampling

SafetyDGX agent

arXiv:2504.07031v2 Announce Type: replace Abstract: Class-bias, that is class-wise performance disparities, is typically attributed to data imbalance and addressed through frequency-based resampling.

Region-Constrained Group Relative Policy Optimization for Flow-Based Image Editing

SafetyDGX agent

arXiv:2604.09386v1 Announce Type: new Abstract: Instruction-guided image editing requires balancing target modification with non-target preservation. Recently, flow-based models have emerged as a stro

Reinforcement-aware Knowledge Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2602.22495v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training has recently driven major gains in long chain-of-thought reasoning large language models (LLMs), but

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

SafetyDGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

SafetyDGX agent

arXiv:2604.09474v1 Announce Type: cross Abstract: Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise,

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

SafetyDGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

← Previous
1…203204205206207…210
Next →