AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
21 Apr 2026

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

SafetyDGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

Dual Alignment Between Language Model Layers and Human Sentence Processing

SafetyDGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

Emergency Stopping for Liquid-manipulating Robots

SafetyDGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

SafetyDGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI …

SafetyDGX agent

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI of telling people to be more coy about extinction risks. I h

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

SafetyDGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

LiDAR-based Crowd Navigation with Visible Edge Group Representation

SafetyDGX agent

arXiv:2604.16741v1 Announce Type: new Abstract: Robot navigation in crowded pedestrian environments is a well-known challenge and we explore the practical deployment of group-based representations in

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams …

SafetyDGX agent

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams actually close the evals feedback loop. Looking forward to I

Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training (Reuters)

SafetyDGX agent

Reuters: Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training — Meta (META.O) is installing new

MoCo: A One-Stop Shop for Model Collaboration Research

SafetyDGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

Multimodal Policy Internalization for Conversational Agents

SafetyDGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

SafetyDGX agent

arXiv:2604.16518v1 Announce Type: new Abstract: Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit percepti

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs

SafetyDGX agent

arXiv:2604.17543v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable success in general-domain tasks, yet their direct application to the legal domain remains challeng

SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection

SafetyDGX agent

arXiv:2604.18476v1 Announce Type: new Abstract: Camera-only 3D object detection has emerged as a cost-effective and scalable alternative to LiDAR for autonomous driving, yet existing methods primarily

Shepherding UAV Swarm with Action Prediction Based on Movement Constraints

SafetyDGX agent

arXiv:2604.17189v1 Announce Type: new Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicles (UAVs), which predicts the swarm behavior

The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning

SafetyDGX agent

arXiv:2604.17114v1 Announce Type: new Abstract: Frontier large language models generate clinically accurate outputs, but their citations are often fabricated. We term this the Provenance Gap. We teste

Towards Trustworthy Depression Estimation via Disentangled Evidential Learning

SafetyDGX agent

arXiv:2604.16579v1 Announce Type: new Abstract: Automated depression estimation is highly vulnerable to signal corruption and ambient noise in real-world deployment. Prevailing deterministic methods p

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts

SafetyDGX agent

arXiv:2604.18473v1 Announce Type: new Abstract: Extending a fully post-trained language model with new domain capabilities is fundamentally limited by monolithic training paradigms: retraining from sc

Training Language Models to Use Prolog as a Tool

SafetyDGX agent

arXiv:2512.07407v2 Announce Type: replace Abstract: Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use P

UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation

SafetyDGX agent

arXiv:2602.09130v3 Announce Type: replace Abstract: Model compression is increasingly essential for deploying large language models (LLMs), yet existing comparative studies largely focus on pruning an

20 Apr 2026

Contact-Aware Planning and Control of Continuum Robots in Highly Constrained Environments

SafetyDGX agent

arXiv:2604.15638v1 Announce Type: new Abstract: Continuum robots are well suited for navigating confined and fragile environments, such as vascular or endoluminal anatomy, where contact with surroundi

How people use Copilot for Health

SafetyDGX agent

arXiv:2604.15331v1 Announce Type: cross Abstract: We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversati

Long-Term Memory for VLA-based Agents in Open-World Task Execution

SafetyDGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

Model ReleasesDGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

On the Rejection Criterion for Proxy-based Test-time Alignment

SafetyDGX agent

arXiv:2604.16146v1 Announce Type: new Abstract: Recent works proposed test-time alignment methods that rely on a small aligned model as a proxy that guides the generation of a larger base (unaligned)

Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning

SafetyDGX agent

arXiv:2506.20020v2 Announce Type: replace Abstract: Reasoning in humans is prone to biases due to underlying motivations like identity protection, that undermine rational decision-making and judgment.

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

SafetyDGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

SafetyDGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

18 Apr 2026

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the k…

SafetyDGX agent

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the kind Anthropic loves to promote. “Accurate” would be more acc

17 Apr 2026

An unsupervised decision-support framework for multivariate biomarker analysis in athlete monitoring

SafetyDGX agent

arXiv:2604.14534v1 Announce Type: new Abstract: Purpose. Athlete monitoring is constrained by small cohorts, heterogeneous biomarker scales, limited feasibility of repeated sampling, and the lack of r

BitFlipScope: Scalable Fault Localization and Recovery for Bit-Flip Corruptions in LLMs

Local AiDGX agent

arXiv:2512.22174v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) deployed in practical and safety-critical settings are increasingly susceptible to bit-flip faults caused by hard

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

SafetyDGX agent

arXiv:2604.14892v1 Announce Type: new Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudi

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

SafetyDGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

SafetyDGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

Humanoid Factors: Design Principles for AI Humanoids in Human Worlds

SafetyDGX agent

arXiv:2602.10069v2 Announce Type: replace Abstract: Human factors research has long focused on optimizing environments, tools, and systems to account for human performance. Yet, as humanoid robots beg

IROSA: Interactive Robot Skill Adaptation using Natural Language

SafetyDGX agent

arXiv:2603.03897v3 Announce Type: replace-cross Abstract: Foundation models have demonstrated impressive capabilities across diverse domains, while imitation learning provides principled methods for r

Language Model as Planner and Formalizer under Constraints

SafetyDGX agent

arXiv:2510.05486v2 Announce Type: replace Abstract: LLMs have been widely used in planning, either as planners to generate action sequences end-to-end, or as formalizers to represent the planning doma

Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality

SafetyDGX agent

arXiv:2603.17512v4 Announce Type: replace Abstract: Large language models (LLMs) exhibit strong general intelligence, yet their multilingual performance remains highly imbalanced. Although LLMs encode

Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities

SafetyDGX agent

arXiv:2509.06921v2 Announce Type: replace-cross Abstract: Cybersecurity demands both rapid pattern recognition and deliberative reasoning, yet purely neural or purely symbolic approaches each address

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

SafetyDGX agent

arXiv:2604.15308v1 Announce Type: new Abstract: High-level autonomous driving requires motion planners capable of modeling multimodal future uncertainties while remaining robust in closed-loop interac

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

SafetyDGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

SafetyDGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

Switch: Learning Agile Skills Switching for Humanoid Robots

SafetyDGX agent

arXiv:2604.14834v1 Announce Type: new Abstract: Recent advancements in whole-body control through deep reinforcement learning have enabled humanoid robots to achieve remarkable progress in real-world

Trajectory Planning for a Multi-UAV Rigid-Payload Cascaded Transportation System Based on Enhanced Tube-RRT*

SafetyDGX agent

arXiv:2604.15074v1 Announce Type: new Abstract: This paper presents a two-stage trajectory planning framework for a multi-UAV rigid-payload cascaded transportation system, aiming to address planning c

TwinOR: Photorealistic Digital Twins of Dynamic Operating Rooms for Embodied AI Research

SafetyDGX agent

arXiv:2511.07412v2 Announce Type: replace Abstract: Developing embodied AI for intelligent surgical systems requires safe, controllable environments for continual learning and evaluation. However, saf

Vision-Based Safe Human-Robot Collaboration with Uncertainty Guarantees

SafetyDGX agent

arXiv:2604.15221v1 Announce Type: cross Abstract: We propose a framework for vision-based human pose estimation and motion prediction that gives conformal prediction guarantees for certifiably safe hu

16 Apr 2026

A Bayesian Framework for Uncertainty-Aware Explanations in Power Quality Disturbance Classification

SafetyDGX agent

arXiv:2604.13658v1 Announce Type: new Abstract: Advanced deep learning methods have shown remarkable success in power quality disturbance (PQD) classification. To enhance model transparency, explainab

Bias at the End of the Score

SafetyDGX agent

arXiv:2604.13305v1 Announce Type: new Abstract: Reward models (RMs) are inherently non-neutral value functions designed and trained to encode specific objectives, such as human preferences or text-ima

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding

SafetyDGX agent

arXiv:2603.27064v2 Announce Type: replace-cross Abstract: Understanding charts requires models to jointly reason over geometric visual patterns, structured numerical data, and natural language -- a ca

Golden Handcuffs make safer AI agents

SafetyDGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

SafetyDGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

Maybe because of this paper? https://x.com/emollick/status/1991624198855561508?s=20

SafetyDGX agent

Maybe because of this paper? https://x.com/emollick/status/1991624198855561508?s=20 Tell all the truth but tell it slant— Success in Circuit lies Too bright for our infirm Delight The Truth's superb s

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

SafetyDGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

Rethinking Uncertainty in Segmentation: From Estimation to Decision

SafetyDGX agent

arXiv:2604.13262v1 Announce Type: new Abstract: In medical image segmentation, uncertainty estimates are often reported but rarely used to guide decisions. We study the missing step: how uncertainty m

Self-adaptive Multi-Access Edge Architectures: A Robotics Case

SafetyDGX agent

arXiv:2604.13542v1 Announce Type: new Abstract: The growth of compute-intensive AI tasks highlights the need to mitigate the processing costs and improve performance and energy efficiency. This necess

15 Apr 2026

A longitudinal health agent framework

SafetyDGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

Active Imitation Learning for Thermal- and Kernel-Aware LFM Inference on 3D S-NUCA Many-Cores

SafetyDGX agent

arXiv:2604.11948v1 Announce Type: new Abstract: Large Foundation Model (LFM) inference is both memory- and compute-intensive, traditionally relying on GPUs. However, the limited availability and high

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

SafetyDGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

How Transformers Learn to Plan via Multi-Token Prediction

SafetyDGX agent

arXiv:2604.11912v1 Announce Type: cross Abstract: While next-token prediction (NTP) has been the standard objective for training language models, it often struggles to capture global structure in reas

← Previous
1…4849505152…240
Next →