AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
12 May 2026

Mental Health AI Safety Claims Must Preserve Temporal Evidence

SafetyDGX agent

arXiv:2605.08827v1 Announce Type: new Abstract: The safety of mental health AI is often judged at the wrong temporal scale. Current evaluations typically score isolated responses, endpoint outcomes, o

MePo: Meta Post-Refinement for Rehearsal-Free General Continual Learning

SafetyDGX agent

arXiv:2602.07940v3 Announce Type: replace Abstract: To cope with uncertain changes of the external world, intelligent systems must continually learn from complex, evolving environments and respond in

Meta offers to give rival AI chatbots free access to WhatsApp for a month while it discusses commitments with EU antitrust regulators to address their concerns (Foo Yun Chee/Reuters)

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Foo Yun Chee / Reuters: Meta offers to give rival AI chatbots free access to WhatsApp for a month while it discusses commitments with EU antitrust regulators to address their concerns — Meta Platforms

Meta-reinforcement learning with minimum attention

SafetyDGX agent

arXiv:2505.16741v4 Announce Type: replace Abstract: Minimum attention applies the least action principle to changes of control concerning state and time, first proposed by Brockett. The involved regul

Metropolis-Adjusted Diffusion Models

SafetyDGX agent

arXiv:2605.09654v1 Announce Type: cross Abstract: Sampling from score-based diffusion models incurs bias due to both time discretisation and the approximation of the score function. A common strategy

Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

SafetyDGX agent

arXiv:2605.08472v1 Announce Type: new Abstract: The effectiveness of Reinforcement Learning (RL) in Large Language Models (LLMs) depends on the nature and diversity of the data used before and during

Mismatch-Aware Adaptive Constraint Tightening for Bicycle-Model Trajectory Optimization

SafetyDGX agent

arXiv:2605.09376v1 Announce Type: new Abstract: Trajectory optimization for autonomous vehicles usually relies on the kinematic bicycle model because of its computational simplicity. However, when the

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

SafetyDGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration

SafetyDGX agent

arXiv:2605.08277v1 Announce Type: cross Abstract: Many-shot jailbreaking (MSJ) causes safety-aligned language models to answer harmful queries by preceding them with many harmful question-answer demon

MoMo: Conditioned Contrastive Representation Learning for Preference-Modulated Planning

SafetyDGX agent

arXiv:2605.08512v1 Announce Type: new Abstract: Temporally contrastive representation learning induces a latent structure capable of reducing long-horizon planning to inference in a low-dimensional li

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

SafetyDGX agent

arXiv:2605.09258v1 Announce Type: cross Abstract: Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of mo

Morphology-Aware Graph Reinforcement Learning for Tensegrity Robot Locomotion

SafetyDGX agent

arXiv:2510.26067v2 Announce Type: replace Abstract: Tensegrity robots combine rigid rods and elastic cables, offering high resilience and deployability but at the same time posing major challenges for

MTA-RL: Robust Urban Driving via Multi-modal Transformer-based 3D Affordances and Reinforcement Learning

SafetyDGX agent

arXiv:2605.10177v1 Announce Type: cross Abstract: Robust urban autonomous driving requires reliable 3D scene understanding and stable decision-making under dense interactions. However, existing end-to

Multi-layer attentive probing improves transfer of audio representations for bioacoustics

SafetyDGX agent

arXiv:2605.10494v1 Announce Type: cross Abstract: Probing heads map the representations learned from audio by a machine learning model to downstream task labels and are a key component in evaluating r

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning

SafetyDGX agent

arXiv:2605.09364v1 Announce Type: new Abstract: This paper investigates robust representation learning in offline goal-conditioned reinforcement learning (GCRL). Particularly in sparse reward scenario

Multimodal Representation Learning Conditioned on Semantic Relations

SafetyDGX agent

arXiv:2508.17497v2 Announce Type: replace-cross Abstract: Multimodal representation learning has been largely driven by contrastive models such as CLIP, which learn a shared embedding space by alignin

Muninn: Your Trajectory Diffusion Model But Faster

SafetyDGX agent

arXiv:2605.09999v1 Announce Type: new Abstract: Diffusion-based trajectory planners can synthesize rich, multimodal robot motions, but their iterative denoising makes online planning and control prohi

MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation

SafetyDGX agent

arXiv:2511.07833v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a standard recipe for post-training LLMs on reasoning tasks, with Group Relat

Mutual Information Optimal Density Control of Linear Systems and Generalized Schrodinger Bridges with Reference Refinement

SafetyDGX agent

arXiv:2605.09349v1 Announce Type: cross Abstract: We consider a mutual information (MI) regularized version of optimal density control of a discrete-time linear system. MI optimal control has been pro

MVB-Grasp: Minimum-Volume-Box Filtering of Diffusion-based Grasps for Frontal Manipulation

SafetyDGX agent

arXiv:2605.09672v1 Announce Type: new Abstract: State-of-the-art 6-DoF grasp generators excel on tabletop benchmarks with overhead cameras but struggle in frontal grasping scenarios on low-cost manipu

Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework

SafetyDGX agent

arXiv:2605.10671v1 Announce Type: new Abstract: In this work, we show that natural policy gradient, a core algorithm in reinforcement learning, admits an exact formulation as a smoothed and averaged f

Navigating EU AI Act requirements for LLM fine-tuning on Amazon SageMaker AI

SafetyDGX agent

In this post, we show you how to set up FLOPs tracking during LLM fine-tuning using the open source Fine-Tuning FLOPs Meter toolkit on Amazon SageMaker AI. You learn how to determine your compliance s

Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers

SafetyDGX agent

arXiv:2605.09176v1 Announce Type: cross Abstract: Training large language models requires optimization algorithms that are not only statistically effective, but also computationally and memory efficie

NEO: No-Optimization Test-Time Adaptation through Latent Re-Centering

SafetyDGX agent

arXiv:2510.05635v2 Announce Type: replace-cross Abstract: Test-Time Adaptation (TTA) methods are often computationally expensive, require a large amount of data for effective adaptation, or are brittl

Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs

SafetyDGX agent

arXiv:2605.10877v1 Announce Type: new Abstract: Automated question answering (QA) over electronic health records (EHRs) demands precise evidence retrieval, faithful answer generation, and explicit gro

Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning

SafetyDGX agent

arXiv:2605.05373v2 Announce Type: replace Abstract: A key capability of intelligent agents is operating under partial observability: reasoning and acting effectively despite missing or incomplete stat

Neuromorphic Reinforcement Learning for Quadruped Locomotion Control on Uneven Terrain

SafetyDGX agent

arXiv:2605.09595v1 Announce Type: cross Abstract: Reinforcement learning (RL) has enabled robust quadruped locomotion over complex terrain, but most learned controllers are trained offline with backpr

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

SafetyDGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

Not All Turns Matter: Credit Assignment for Multi-Turn Jailbreaking

SafetyDGX agent

arXiv:2605.08778v1 Announce Type: new Abstract: Deploying LLMs in multi-turn dialogues facilitates jailbreak attacks that distribute harmful intent across seemingly benign turns. Recent training-based

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor.

SafetyDGX agent

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor. hot take: unrestricted LLMs are as dangerous as weapons of mass destruction

note that I said “Even @haider1” because he is often optimistic, but he has informed me (with receipts that he shared) that he expressed som…

SafetyDGX agent

Gary Marcus references a conversation with someone named Haider, noting that despite Haider's typically optimistic outlook, he has shared documented evidence of expressing concerns about a particular

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels

SafetyDGX agent

arXiv:2602.06382v2 Announce Type: replace Abstract: Achieving robust vision-based humanoid locomotion remains challenging due to two fundamental issues: the sim-to-real gap introduces significant perc

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs

SafetyDGX agent

arXiv:2605.09433v1 Announce Type: new Abstract: Existing preference datasets for text-to-image models typically store only the final winner/loser images. This representation is insufficient for rectif

On-Policy Distillation with Best-of-N Teacher Rollout Selection

SafetyDGX agent

arXiv:2605.09725v1 Announce Type: new Abstract: On-policy distillation (OPD), which supervises a student on its own sampled trajectories, has emerged as a data-efficient post-training method for impro

On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models

SafetyDGX agent

arXiv:2605.09606v1 Announce Type: cross Abstract: Recent advances in image-to-3D models have significantly improved the fidelity and accessibility of 3D content creation. Such a powerful reconstructio

On Uniform Error Bounds for Kernel Regression under Non-Gaussian Noise

SafetyDGX agent

arXiv:2605.09757v1 Announce Type: new Abstract: Providing non-conservative uncertainty quantification for function estimates derived from noisy observations remains a fundamental challenge in statisti

On Variance Reduction in Learning Mean Flows

SafetyDGX agent

arXiv:2605.09235v1 Announce Type: cross Abstract: One-step generative modeling has emerged as a leading approach to amortize the inference cost of diffusion and flow-matching models. Among distillatio

Open Ontologies: Tool-Augmented Ontology Engineering with Stable Matching Alignment

SafetyDGX agent

arXiv:2605.09184v1 Announce Type: new Abstract: We present Open Ontologies, an open-source ontology engineering system implemented in Rust that integrates LLM-driven construction with formal OWL reaso

OpenAI’s Sam Altman’s personal investments are coming under intensifying scrutiny from Republicans following an April article in The Wall St…

SafetyDGX agent

Sam Altman's personal investments faced increased scrutiny from Republican lawmakers following reporting by The Wall Street Journal in April regarding potential conflicts of interest. The controversy

OpenClaw-RL: Train Any Agent Simply by Talking

SafetyDGX agent

arXiv:2603.10165v2 Announce Type: replace-cross Abstract: Every agent interaction generates a next-state signal, namely the user reply, tool output, terminal or GUI state change that follows each acti

Opinion: It is rare for the US public to agree on anything these days. Fear of AI is as close to a national consensus as it gets. A clear ma…

SafetyDGX agent

Opinion: It is rare for the US public to agree on anything these days. Fear of AI is as close to a national consensus as it gets. A clear majority says that AI will do more harm than good. https://ft.

Overcoming Catastrophic Forgetting in Visual Continual Learning with Reinforcement Fine-Tuning

SafetyDGX agent

arXiv:2605.09640v1 Announce Type: new Abstract: Recent studies suggest that Reinforcement Fine-Tuning (RFT) is inherently more resilient to catastrophic forgetting than Supervised Fine-Tuning (SFT). H

Path-Coupled Bellman Flows for Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.08253v1 Announce Type: cross Abstract: Distributional reinforcement learning (DRL) models the full return distribution, but existing finite-support or quantile-based methods rely on project

PATRA: Pattern-Aware Alignment and Balanced Reasoning for Time Series Question Answering

SafetyDGX agent

arXiv:2602.23161v2 Announce Type: replace Abstract: Time series reasoning demands both the perception of complex dynamics and logical depth. However, existing LLM-based approaches exhibit two limitati

Pay attention to this one if you build research or knowledge-work agents. Most research-agent systems produce uniform outputs regardless of …

SafetyDGX agent

Pay attention to this one if you build research or knowledge-work agents. Most research-agent systems produce uniform outputs regardless of who is driving them. This new work, NanoResearch, argues tha

Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs

SafetyDGX agent

arXiv:2605.09422v1 Announce Type: new Abstract: Although Large Multimodal Models (LMMs) have achieved strong performance on general video understanding, their susceptibility to textual prior shortcuts

Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework

SafetyDGX agent

arXiv:2605.10043v1 Announce Type: cross Abstract: Large Language Model (LLM) personalization aims to align model behaviors with individual user preferences. Existing methods often focus on isolated us

PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI

SafetyDGX agent

arXiv:2605.05682v2 Announce Type: replace-cross Abstract: Recent developments in AI safety research have called for red-teaming methods that effectively surface potential risks posed by generative AI

PFN-TS: Thompson Sampling for Contextual Bandits via Prior-Data Fitted Networks

SafetyDGX agent

arXiv:2605.10137v1 Announce Type: cross Abstract: Thompson sampling is a widely used strategy for contextual bandits: at each round, it samples a reward function from a Bayesian posterior and acts gre

PHAGE: Patent Heterogeneous Attention-Guided Graph Encoder for Representation Learning

SafetyDGX agent

arXiv:2605.10073v1 Announce Type: new Abstract: Patent claims form a directed dependency structure in which dependent claims inherit and refine the scope of earlier claims; however, existing patent en

PHMForge: Evaluating LLM Agents on Industrial Prognostics through MCP-Native, Algorithm-Grounded Tools

SafetyDGX agent

arXiv:2604.01532v2 Announce Type: replace Abstract: LLM agents are beginning to invoke industrial asset-management tools through the Model Context Protocol (MCP), yet whether they can act reliably on

PhysEDA: Physics-Aware Learning Framework for Efficient EDA With Manhattan Distance Decay

SafetyDGX agent

arXiv:2605.10547v1 Announce Type: new Abstract: Electronic design automation (EDA) addresses placement, routing, timing analysis, and power-integrity verification for integrated circuits. Learning met

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

SafetyDGX agent

arXiv:2605.10118v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hi

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.09638v1 Announce Type: new Abstract: Ensuring the security of reinforcement learning (RL) models is critical, particularly when they are trained by third parties and deployed in real-world

PMCTS: Particle Monte Carlo Tree Search for Principled Parallelized Inference Time Scaling

SafetyDGX agent

arXiv:2605.08982v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) is a widely used approach for policy improvement through search with increasing popularity for real world applications. D

Policy Gradient Methods for Non-Markovian Reinforcement Learning

SafetyDGX agent

arXiv:2605.10816v1 Announce Type: cross Abstract: We study policy gradient methods for reinforcement learning in non-Markovian decision processes (NMDPs), where observations and rewards depend on the

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

SafetyDGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

Position: Academic Conferences are Potentially Facing Denominator Gaming Caused by Fully Automated Scientific Agents

SafetyDGX agent

arXiv:2605.09915v1 Announce Type: cross Abstract: The implicit policy of maintaining relatively stable acceptance rates at top AI conferences, despite exponentially growing submissions, introduces a c

Positional Encoding via Token-Aware Phase Attention

SafetyDGX agent

arXiv:2509.12635v3 Announce Type: replace-cross Abstract: We prove under practical assumptions that Rotary Positional Embedding (RoPE) introduces an intrinsic distance-dependent bias in attention scor

Positional LSH: Binary Block Matrix Approximation for Attention with Linear Biases

SafetyDGX agent

arXiv:2605.09472v1 Announce Type: new Abstract: Positional encoding in transformers is commonly implemented through positional embeddings, attention masks, or bias terms, but formal connections betwee

← Previous
1…150151152153154…212
Next →