AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
28 May 2026

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

SafetyDGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

Did AI just solve math? Cal Newport podcast: https://youtu.be/fhZRWZ6J4k4?si=Gzut_5x856n5DW6o

SafetyDGX agent

Cal Newport discusses whether recent AI advances represent a breakthrough in mathematical problem-solving capabilities, examining the implications of AI systems' improved performance on complex mathem

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

SafetyDGX agent

arXiv:2512.02019v3 Announce Type: replace-cross Abstract: Diffusion models excel at sampling from complex, unnormalized distributions. In this work, we extend Maximum Entropy Reinforcement Learning (M

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Diffusion Large Language Models for Visual Speech Recognition

SafetyDGX agent

arXiv:2605.28456v1 Announce Type: new Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually

DiscoForcing: A Unified Framework for Real-Time Audio-Driven Character Control with Diffusion Forcing

SafetyDGX agent

arXiv:2605.28491v1 Announce Type: new Abstract: We study real-time audio-responsive character control as a deployment-faithful problem: strictly causal, bounded-latency streaming that must generate co

DREAM-R: Multimodal Speculative Reasoning with RL-Based Refined Drafting, Precise Verification, and Fully Parallel Execution

SafetyDGX agent

arXiv:2605.28678v1 Announce Type: new Abstract: Speculative reasoning has recently been proposed as a means to accelerate reasoning-intensive generation in large multimodal models, but its effectivene

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

SafetyDGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning

SafetyDGX agent

arXiv:2602.02150v2 Announce Type: replace-cross Abstract: Test-time reinforcement learning generates multiple candidate answers via repeated rollouts and performs online updates using pseudo-labels co

Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning

SafetyDGX agent

arXiv:2603.09882v2 Announce Type: replace-cross Abstract: Extrinsic dexterity leverages environmental contact to overcome the limitations of prehensile manipulation. However, achieving such dexterity

EntroAD: Structural Entropy-Guided Prompt Adaptation for Zero-Shot Anomaly Detection

SafetyDGX agent

arXiv:2605.28630v1 Announce Type: new Abstract: Zero-Shot Anomaly Detection (ZSAD) aims to detect anomalies in unseen domains without target-domain adaptation. Recent CLIP-based methods have shown pro

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization

SafetyDGX agent

arXiv:2605.27741v1 Announce Type: new Abstract: Audio and omni-modal large language models exhibit impressive cross-modal reasoning capabilities. However, applying standard reinforcement learning post

Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish Online News

SafetyDGX agent

arXiv:2605.28598v1 Announce Type: cross Abstract: LLM-powered social agents are increasingly used to simulate online social behavior, yet their realism remains difficult to validate. Existing work has

Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.28098v1 Announce Type: new Abstract: Multi-agent systems are increasingly deployed to support various tasks where agents interact to achieve individual and collective objectives. Although t

F Scott Fitzgerald, writing about someone who may as well have been Peter Thiel: “They were careless people, Tom and Daisy- they smashed up …

SafetyDGX agent

F Scott Fitzgerald, writing about someone who may as well have been Peter Thiel: “They were careless people, Tom and Daisy- they smashed up things and creatures and then retreated back into their mone

FABSVer: Faster Training and Better Self-Verification for LLM Mathematical Reasoning

SafetyDGX agent

arXiv:2605.28389v1 Announce Type: new Abstract: While large language models have made significant progress in mathematical reasoning, they remain unreliable at judging the correctness of their own sol

FedEHR-Gen: Federated Synthetic Time-Series EHR Generation via Latent Space Alignment and Distribution-Aware Aggregation

SafetyDGX agent

arXiv:2605.27892v1 Announce Type: new Abstract: Synthetic Electronic Health Record (EHR) generation provides a promising avenue for data augmentation and cross-hospital modeling in privacy-constrained

From Affect to Complex Behavior: Advancing Multimodal Human-Centered AI at the 10th ABAW Workshop & Competition

SafetyDGX agent

arXiv:2605.27451v1 Announce Type: new Abstract: The 10th Affective & Behavior Analysis in-the-Wild (ABAW) Workshop and Competition, held at CVPR 2026, continues to advance research on modelling, analy

From AR to Diffusion: Efficiently Adapting Large Language Models with Strictly Causal and Elastic Horizons

SafetyDGX agent

arXiv:2605.27387v1 Announce Type: cross Abstract: Diffusion models promise efficient parallel text generation but rely on bidirectional attention, creating a structural mismatch with pre-trained Autor

From Learning Resources to Competencies: LLM-Based Tagging with Evidence and Graph Constraints

SafetyDGX agent

arXiv:2605.28483v1 Announce Type: new Abstract: Linking learning resources to a structured competency framework is key to enabling competency-based search and curriculum analytics in Learning Manageme

From Pixels to Words -- Towards Native One-Vision Models at Scale

SafetyDGX agent

arXiv:2605.28820v1 Announce Type: new Abstract: Current vision-language models (VLMs) typically stitch together separate image encoders and language decoders via multi-stage alignment, a modular frame

fun thread on consciousness with @Grimezsz:

SafetyDGX agent

Gary Marcus shares a discussion thread about consciousness with user @Grimezsz on X (formerly Twitter), likely exploring philosophical, scientific, or technical perspectives on consciousness and relat

GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation

SafetyDGX agent

arXiv:2605.27491v1 Announce Type: new Abstract: We introduce GE-Sim 2.0 (Genie Envisioner World Simulator 2.0), a closed-loop video world simulator for robotic manipulation. Building on the action-con

GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization

SafetyDGX agent

arXiv:2605.27934v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves language model reasoning, but its reliance on domain-specific verifiers, sparse outcome rewards,

Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

SafetyDGX agent

arXiv:2605.27970v1 Announce Type: new Abstract: While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometr

Geometry of Relaxed Fair Regression: A Unified Framework for Aware and Unaware Settings

SafetyDGX agent

arXiv:2605.28233v1 Announce Type: cross Abstract: Fairness-accuracy trade-offs are a central concern in the deployment of fairness-aware machine learning methods. When sensitive attributes are unavail

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

SafetyDGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

SafetyDGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

Guaranteed Optimal Compositional Explanations for Neurons

SafetyDGX agent

arXiv:2511.20934v2 Announce Type: replace Abstract: Compositional explanations are a family of methods that aim to describe the spatial alignment between neurons' receptive field activations and conce

Heterogeneous Causal Discovery of Repeated Undesirable Health Outcomes

SafetyDGX agent

arXiv:2503.11477v2 Announce Type: replace Abstract: Understanding the factors that trigger or prevent undesirable health outcomes across patient subpopulations is essential for designing targeted inte

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

SafetyDGX agent

arXiv:2605.28198v1 Announce Type: new Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterog

HiRQA: Hierarchical Ranking and Quality Alignment for Opinion-Unaware Image Quality Assessment

SafetyDGX agent

arXiv:2508.15130v2 Announce Type: replace Abstract: Despite significant progress in no-reference image quality assessment (NR-IQA), dataset biases and reliance on subjective labels continue to hinder

Holy shit! They changed the rules for Elon again... They waved the profitability rule & are adding SpaceX to indices only 5 days after IPO..…

SafetyDGX agent

Holy shit! They changed the rules for Elon again... They waved the profitability rule & are adding SpaceX to indices only 5 days after IPO... normally it's 90 This forces 401k retirement & passive fun

How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks

SafetyDGX agent

arXiv:2605.27662v1 Announce Type: cross Abstract: Equivariant neural networks encode geometric symmetries by construction, yet they are often difficult to optimize and can underperform less constraine

Human-like in-group bias in instruction-tuned language model agents

SafetyDGX agent

arXiv:2605.28114v1 Announce Type: new Abstract: As autonomous AI agents are deployed in persistent, interacting networks -- coordinating tasks, routing resources, and accumulating reputational histori

I read this garbage (in a big UK newspaper) when Google can’t even count reliably and wonder why people don’t spend more time learning about…

SafetyDGX agent

I read this garbage (in a big UK newspaper) when Google can’t even count reliably and wonder why people don’t spend more time learning about AI’s actual strengths and weakness before running their mou

ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment

SafetyDGX agent

arXiv:2605.27374v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) and diffusion models (DMs) have opened new possibilities for AI-generated content. Yet, pers

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes

SafetyDGX agent

arXiv:2601.04716v3 Announce Type: replace Abstract: While Large Language Model (LLM) role-playing agents have advanced rapidly, it remains unclear which profile elements genuinely drive role-playing q

Imitating and Finetuning Model Predictive Control for Robust and Symmetric Quadrupedal Locomotion

SafetyDGX agent

arXiv:2311.02304v3 Announce Type: replace Abstract: Control of legged robots is a challenging problem that has been investigated by different approaches, such as model-based control and learning algor

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expe…

SafetyDGX agent

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expert mathematicians poured over this [transcript] and identifi

IMU Propagation as Preintegration

SafetyDGX agent

arXiv:2605.28279v1 Announce Type: new Abstract: IMU preintegration is widely used in factor-graph-based visual--inertial, lidar--inertial, and radar--inertial state estimation, yet it is often treated

Information-theoretic Multimodal Representation Learning for Electrocardiogram Signals

SafetyDGX agent

arXiv:2605.27583v1 Announce Type: new Abstract: Electrocardiograms (ECGs) are widely used non-invasive measurements of cardiac activity and play a central role in clinical diagnosis. Recent multimodal

Informing AI Policy Assessment using Large-Scale Simulation of Interventions

SafetyDGX agent

arXiv:2605.27395v1 Announce Type: cross Abstract: As the rapid proliferation of AI systems and harms spurs efforts in AI governance around the world, prioritizing among competing policy options has be

Insurance Pricing Optimization via Off-Policy Evaluation

SafetyDGX agent

arXiv:2605.28327v1 Announce Type: cross Abstract: Traditional insurance pricing relies on risk-based principles that ensure actuarial fairness and solvency but do not explicitly account for policyhold

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models

SafetyDGX agent

arXiv:2605.28609v1 Announce Type: new Abstract: Forensic vision-language models (VLMs) have recently been developed to detect image tampering and provide natural-language explanations. However, their

Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration

SafetyDGX agent

arXiv:2605.28184v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as the standard paradigm for improving reasoning capability of large language models,

just imagine what will happen to the economy and people’s retirement funds if these projections from FT are correct. brace for bailouts.

SafetyDGX agent

just imagine what will happen to the economy and people’s retirement funds if these projections from FT are correct. brace for bailouts. The AI numbers are starting to look very ugly. Even under 'best

Learning Deliberately, Acting Intuitively: Unlocking Test-Time Reasoning in Multimodal LLMs

SafetyDGX agent

arXiv:2507.06999v2 Announce Type: replace-cross Abstract: Reasoning is essential for large language models (LLMs), especially in complex tasks such as mathematical problem solving. However, multimodal

Learning High-Dimensional Parity Functions with Product Networks using Gradient Descent

SafetyDGX agent

arXiv:2605.28612v1 Announce Type: new Abstract: Parity functions are fundamental Boolean operations with critical applications across machine learning, cryptography, and error correction. Yet, learnin

Learning to Assign Prediction Tasks to Agents with Capacity Constraints

SafetyDGX agent

arXiv:2605.27999v1 Announce Type: cross Abstract: We address the problem of learning to assign prediction tasks to one agent from a set of available human or AI agents. In particular, we focus on the

Learning to Bid in Repeated Second-Price Auctions with Dynamic Values and Aggregated Feedback

SafetyDGX agent

arXiv:2605.28133v1 Announce Type: new Abstract: We study the problem of learning to bid when the bidder's value is dynamic, i.e., when the current value depends on past outcomes. Specifically, we cons

Learning with Importance Weighted Variational Inference

SafetyDGX agent

arXiv:2410.12035v2 Announce Type: replace-cross Abstract: Several variational bounds involving importance weighting ideas generalize the Evidence Lower BOund (ELBO) for marginal likelihood optimizatio

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

SafetyDGX agent

arXiv:2605.28524v1 Announce Type: new Abstract: In recent years, Large Language Models (LLMs) have shown great capability in processing graph tasks such as fraud detection. However, most existing meth

LLM Watermark Evasion via Bias Inversion

SafetyDGX agent

arXiv:2509.23019v5 Announce Type: replace-cross Abstract: Watermarking offers a promising solution for detecting LLM-generated content, yet its robustness under realistic query-free (black-box) evasio

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

SafetyDGX agent

arXiv:2605.28109v1 Announce Type: new Abstract: Recent advances in online reinforcement learning (RL) for large language models (LLMs) have demonstrated promising performance in complex reasoning task

Mag-VLA: Vision-Language-Action Model for Bimanual Magnetically Actuated Microrobot Manipulation

SafetyDGX agent

arXiv:2605.28486v1 Announce Type: new Abstract: Magnetically actuated microrobots have been used as wireless, non-contact manipulation tools at microscales, making them promising for minimally invasiv

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

SafetyDGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification

SafetyDGX agent

arXiv:2605.28604v1 Announce Type: cross Abstract: Identifying key individuals in video scenes is essential for applications such as automated video editing and intelligent surveillance. Current method

Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation

SafetyDGX agent

arXiv:2605.12515v2 Announce Type: replace Abstract: Despite their impressive capabilities, multilingual large language models (MLLMs) frequently exhibit inconsistent behaviour when the prompt's langua

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

SafetyDGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities

SafetyDGX agent

arXiv:2605.27388v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized as proxies for computational social analysis; yet, their ability to faithfully represent the 't

← Previous
1…140141142143144…242
Next →