AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

DGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

safetyarxiv-cs-lg
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration

DGX agent

arXiv:2605.08277v1 Announce Type: cross Abstract: Many-shot jailbreaking (MSJ) causes safety-aligned language models to answer harmful queries by preceding them with many harmful question-answer demon

safetyarxiv-cs-ai
12 May 2026
Safety

MoMo: Conditioned Contrastive Representation Learning for Preference-Modulated Planning

DGX agent

arXiv:2605.08512v1 Announce Type: new Abstract: Temporally contrastive representation learning induces a latent structure capable of reducing long-horizon planning to inference in a low-dimensional li

safetyarxiv-cs-lg
12 May 2026
Safety

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

DGX agent

arXiv:2605.09258v1 Announce Type: cross Abstract: Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of mo

safetyarxiv-cs-ai
12 May 2026
Safety

Morphology-Aware Graph Reinforcement Learning for Tensegrity Robot Locomotion

DGX agent

arXiv:2510.26067v2 Announce Type: replace Abstract: Tensegrity robots combine rigid rods and elastic cables, offering high resilience and deployability but at the same time posing major challenges for

safetyarxiv-cs-ro
12 May 2026
Safety

MTA-RL: Robust Urban Driving via Multi-modal Transformer-based 3D Affordances and Reinforcement Learning

DGX agent

arXiv:2605.10177v1 Announce Type: cross Abstract: Robust urban autonomous driving requires reliable 3D scene understanding and stable decision-making under dense interactions. However, existing end-to

safetyarxiv-cs-ai
12 May 2026
Safety

Multi-layer attentive probing improves transfer of audio representations for bioacoustics

DGX agent

arXiv:2605.10494v1 Announce Type: cross Abstract: Probing heads map the representations learned from audio by a machine learning model to downstream task labels and are a key component in evaluating r

safetyarxiv-cs-ai
12 May 2026
Safety

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning

DGX agent

arXiv:2605.09364v1 Announce Type: new Abstract: This paper investigates robust representation learning in offline goal-conditioned reinforcement learning (GCRL). Particularly in sparse reward scenario

safetyarxiv-cs-lg
12 May 2026
Safety

Multimodal Representation Learning Conditioned on Semantic Relations

DGX agent

arXiv:2508.17497v2 Announce Type: replace-cross Abstract: Multimodal representation learning has been largely driven by contrastive models such as CLIP, which learn a shared embedding space by alignin

safetyarxiv-cs-ai
12 May 2026
Safety

Muninn: Your Trajectory Diffusion Model But Faster

DGX agent

arXiv:2605.09999v1 Announce Type: new Abstract: Diffusion-based trajectory planners can synthesize rich, multimodal robot motions, but their iterative denoising makes online planning and control prohi

safetyarxiv-cs-ro
12 May 2026
Safety

MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation

DGX agent

arXiv:2511.07833v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a standard recipe for post-training LLMs on reasoning tasks, with Group Relat

safetyarxiv-cs-ai
12 May 2026
Safety

Mutual Information Optimal Density Control of Linear Systems and Generalized Schrodinger Bridges with Reference Refinement

DGX agent

arXiv:2605.09349v1 Announce Type: cross Abstract: We consider a mutual information (MI) regularized version of optimal density control of a discrete-time linear system. MI optimal control has been pro

safetyarxiv-cs-lg
12 May 2026
Safety

MVB-Grasp: Minimum-Volume-Box Filtering of Diffusion-based Grasps for Frontal Manipulation

DGX agent

arXiv:2605.09672v1 Announce Type: new Abstract: State-of-the-art 6-DoF grasp generators excel on tabletop benchmarks with overhead cameras but struggle in frontal grasping scenarios on low-cost manipu

safetyarxiv-cs-ro
12 May 2026
Safety

Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework

DGX agent

arXiv:2605.10671v1 Announce Type: new Abstract: In this work, we show that natural policy gradient, a core algorithm in reinforcement learning, admits an exact formulation as a smoothed and averaged f

safetyarxiv-cs-lg
12 May 2026
Safety

Navigating EU AI Act requirements for LLM fine-tuning on Amazon SageMaker AI

DGX agent

In this post, we show you how to set up FLOPs tracking during LLM fine-tuning using the open source Fine-Tuning FLOPs Meter toolkit on Amazon SageMaker AI. You learn how to determine your compliance s

safetyaws-ml-blog
12 May 2026
Safety

Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers

DGX agent

arXiv:2605.09176v1 Announce Type: cross Abstract: Training large language models requires optimization algorithms that are not only statistically effective, but also computationally and memory efficie

safetyarxiv-cs-ai
12 May 2026
Safety

NEO: No-Optimization Test-Time Adaptation through Latent Re-Centering

DGX agent

arXiv:2510.05635v2 Announce Type: replace-cross Abstract: Test-Time Adaptation (TTA) methods are often computationally expensive, require a large amount of data for effective adaptation, or are brittl

safetyarxiv-cs-cv
12 May 2026
Safety

Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs

DGX agent

arXiv:2605.10877v1 Announce Type: new Abstract: Automated question answering (QA) over electronic health records (EHRs) demands precise evidence retrieval, faithful answer generation, and explicit gro

safetyarxiv-cs-cl
12 May 2026
Safety

Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning

DGX agent

arXiv:2605.05373v2 Announce Type: replace Abstract: A key capability of intelligent agents is operating under partial observability: reasoning and acting effectively despite missing or incomplete stat

safetyarxiv-cs-lg
12 May 2026
Safety

Neuromorphic Reinforcement Learning for Quadruped Locomotion Control on Uneven Terrain

DGX agent

arXiv:2605.09595v1 Announce Type: cross Abstract: Reinforcement learning (RL) has enabled robust quadruped locomotion over complex terrain, but most learned controllers are trained offline with backpr

safetyarxiv-cs-ro
12 May 2026
Safety

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

DGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

safetyarxiv-cs-ai
12 May 2026
Safety

Not All Turns Matter: Credit Assignment for Multi-Turn Jailbreaking

DGX agent

arXiv:2605.08778v1 Announce Type: new Abstract: Deploying LLMs in multi-turn dialogues facilitates jailbreak attacks that distribute harmful intent across seemingly benign turns. Recent training-based

safetyarxiv-cs-ai
12 May 2026
Safety

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor.

DGX agent

not every day @scaling01 and I agree. but he’s right. and most people won’t notice it happening, per the work of @informor. hot take: unrestricted LLMs are as dangerous as weapons of mass destruction

safetygary-marcus--x
12 May 2026
Safety

note that I said “Even @haider1” because he is often optimistic, but he has informed me (with receipts that he shared) that he expressed som…

DGX agent

Gary Marcus references a conversation with someone named Haider, noting that despite Haider's typically optimistic outlook, he has shared documented evidence of expressing concerns about a particular

safetygary-marcus--x
12 May 2026
Safety

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels

DGX agent

arXiv:2602.06382v2 Announce Type: replace Abstract: Achieving robust vision-based humanoid locomotion remains challenging due to two fundamental issues: the sim-to-real gap introduces significant perc

safetyarxiv-cs-ro
12 May 2026
Safety

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs

DGX agent

arXiv:2605.09433v1 Announce Type: new Abstract: Existing preference datasets for text-to-image models typically store only the final winner/loser images. This representation is insufficient for rectif

safetyarxiv-cs-cv
12 May 2026
Safety

On-Policy Distillation with Best-of-N Teacher Rollout Selection

DGX agent

arXiv:2605.09725v1 Announce Type: new Abstract: On-policy distillation (OPD), which supervises a student on its own sampled trajectories, has emerged as a data-efficient post-training method for impro

safetyarxiv-cs-cv
12 May 2026
Safety

On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models

DGX agent

arXiv:2605.09606v1 Announce Type: cross Abstract: Recent advances in image-to-3D models have significantly improved the fidelity and accessibility of 3D content creation. Such a powerful reconstructio

safetyarxiv-cs-cv
12 May 2026
Safety

On Uniform Error Bounds for Kernel Regression under Non-Gaussian Noise

DGX agent

arXiv:2605.09757v1 Announce Type: new Abstract: Providing non-conservative uncertainty quantification for function estimates derived from noisy observations remains a fundamental challenge in statisti

safetyarxiv-cs-lg
12 May 2026
Safety

On Variance Reduction in Learning Mean Flows

DGX agent

arXiv:2605.09235v1 Announce Type: cross Abstract: One-step generative modeling has emerged as a leading approach to amortize the inference cost of diffusion and flow-matching models. Among distillatio

safetyarxiv-cs-ai
12 May 2026
Safety

Open Ontologies: Tool-Augmented Ontology Engineering with Stable Matching Alignment

DGX agent

arXiv:2605.09184v1 Announce Type: new Abstract: We present Open Ontologies, an open-source ontology engineering system implemented in Rust that integrates LLM-driven construction with formal OWL reaso

safetyarxiv-cs-ai
12 May 2026
Safety

OpenAI’s Sam Altman’s personal investments are coming under intensifying scrutiny from Republicans following an April article in The Wall St…

DGX agent

Sam Altman's personal investments faced increased scrutiny from Republican lawmakers following reporting by The Wall Street Journal in April regarding potential conflicts of interest. The controversy

safetygary-marcus--x
12 May 2026
Safety

OpenClaw-RL: Train Any Agent Simply by Talking

DGX agent

arXiv:2603.10165v2 Announce Type: replace-cross Abstract: Every agent interaction generates a next-state signal, namely the user reply, tool output, terminal or GUI state change that follows each acti

safetyarxiv-cs-ai
12 May 2026
Safety

Opinion: It is rare for the US public to agree on anything these days. Fear of AI is as close to a national consensus as it gets. A clear ma…

DGX agent

Opinion: It is rare for the US public to agree on anything these days. Fear of AI is as close to a national consensus as it gets. A clear majority says that AI will do more harm than good. https://ft.

safetygary-marcus--x
12 May 2026
Safety

Overcoming Catastrophic Forgetting in Visual Continual Learning with Reinforcement Fine-Tuning

DGX agent

arXiv:2605.09640v1 Announce Type: new Abstract: Recent studies suggest that Reinforcement Fine-Tuning (RFT) is inherently more resilient to catastrophic forgetting than Supervised Fine-Tuning (SFT). H

safetyarxiv-cs-cv
12 May 2026
Safety

Path-Coupled Bellman Flows for Distributional Reinforcement Learning

DGX agent

arXiv:2605.08253v1 Announce Type: cross Abstract: Distributional reinforcement learning (DRL) models the full return distribution, but existing finite-support or quantile-based methods rely on project

safetyarxiv-cs-ai
12 May 2026
Safety

PATRA: Pattern-Aware Alignment and Balanced Reasoning for Time Series Question Answering

DGX agent

arXiv:2602.23161v2 Announce Type: replace Abstract: Time series reasoning demands both the perception of complex dynamics and logical depth. However, existing LLM-based approaches exhibit two limitati

safetyarxiv-cs-ai
12 May 2026
Safety

Pay attention to this one if you build research or knowledge-work agents. Most research-agent systems produce uniform outputs regardless of …

DGX agent

Pay attention to this one if you build research or knowledge-work agents. Most research-agent systems produce uniform outputs regardless of who is driving them. This new work, NanoResearch, argues tha

safetydair-ai--x
12 May 2026
Safety

Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs

DGX agent

arXiv:2605.09422v1 Announce Type: new Abstract: Although Large Multimodal Models (LMMs) have achieved strong performance on general video understanding, their susceptibility to textual prior shortcuts

safetyarxiv-cs-cl
12 May 2026
Safety

Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework

DGX agent

arXiv:2605.10043v1 Announce Type: cross Abstract: Large Language Model (LLM) personalization aims to align model behaviors with individual user preferences. Existing methods often focus on isolated us

safetyarxiv-cs-ai
12 May 2026
Safety

PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI

DGX agent

arXiv:2605.05682v2 Announce Type: replace-cross Abstract: Recent developments in AI safety research have called for red-teaming methods that effectively surface potential risks posed by generative AI

safetyarxiv-cs-ai
12 May 2026
Safety

PFN-TS: Thompson Sampling for Contextual Bandits via Prior-Data Fitted Networks

DGX agent

arXiv:2605.10137v1 Announce Type: cross Abstract: Thompson sampling is a widely used strategy for contextual bandits: at each round, it samples a reward function from a Bayesian posterior and acts gre

safetyarxiv-cs-lg
12 May 2026
Safety

PHAGE: Patent Heterogeneous Attention-Guided Graph Encoder for Representation Learning

DGX agent

arXiv:2605.10073v1 Announce Type: new Abstract: Patent claims form a directed dependency structure in which dependent claims inherit and refine the scope of earlier claims; however, existing patent en

safetyarxiv-cs-cl
12 May 2026
Safety

PHMForge: Evaluating LLM Agents on Industrial Prognostics through MCP-Native, Algorithm-Grounded Tools

DGX agent

arXiv:2604.01532v2 Announce Type: replace Abstract: LLM agents are beginning to invoke industrial asset-management tools through the Model Context Protocol (MCP), yet whether they can act reliably on

safetyarxiv-cs-ai
12 May 2026
Safety

PhysEDA: Physics-Aware Learning Framework for Efficient EDA With Manhattan Distance Decay

DGX agent

arXiv:2605.10547v1 Announce Type: new Abstract: Electronic design automation (EDA) addresses placement, routing, timing analysis, and power-integrity verification for integrated circuits. Learning met

safetyarxiv-cs-lg
12 May 2026
Safety

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

DGX agent

arXiv:2605.10118v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hi

safetyarxiv-cs-ro
12 May 2026
Safety

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning

DGX agent

arXiv:2605.09638v1 Announce Type: new Abstract: Ensuring the security of reinforcement learning (RL) models is critical, particularly when they are trained by third parties and deployed in real-world

safetyarxiv-cs-lg
12 May 2026
Safety

PMCTS: Particle Monte Carlo Tree Search for Principled Parallelized Inference Time Scaling

DGX agent

arXiv:2605.08982v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) is a widely used approach for policy improvement through search with increasing popularity for real world applications. D

safetyarxiv-cs-lg
12 May 2026
← Previous
1…190191192193194…267
Next →