AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

Documents and sources: insurers including QBE and Beazley are moving to cap cyber policy payouts for losses and regulatory fines tied to AI use and 'LLMjacking' (Lee Harris/Financial Times)

DGX agent

Lee Harris / Financial Times: Documents and sources: insurers including QBE and Beazley are moving to cap cyber policy payouts for losses and regulatory fines tied to AI use and “LLMjacking” — Beazley

safetytechmeme
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling

DGX agent

arXiv:2604.19544v1 Announce Type: new Abstract: Multimodal reward models (MRMs) play a crucial role in aligning Multimodal Large Language Models (MLLMs) with human preferences. Training a good MRM req

safetyarxiv-cs-ai
22 Apr 2026
Safety

Hierarchically Robust Zero-shot Vision-language Models

DGX agent

arXiv:2604.18867v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their

safetyarxiv-cs-ai
22 Apr 2026
Safety

Rivian and Volkswagen Group's joint venture put Devin to work on testing and ticket triage across a software platform that will power up to …

DGX agent

Rivian and Volkswagen Group's joint venture put Devin to work on testing and ticket triage across a software platform that will power up to 30M vehicles. Devin handles autonomous ticket triage in Slac

safetycognition-ai--x
22 Apr 2026
Safety

Symbolic Quantile Regression for the Interpretable Prediction of Conditional Quantiles

DGX agent

arXiv:2508.08080v2 Announce Type: replace Abstract: Symbolic Regression (SR) is a well-established framework for generating interpretable or white-box predictive models. Although SR has been successfu

safetyarxiv-cs-lg
22 Apr 2026
Safety

Task-Adaptive Admittance Control for Human-Quadrotor Cooperative Load Transportation with Dynamic Cable-Length Regulation

DGX agent

arXiv:2604.18905v1 Announce Type: new Abstract: The collaboration between humans and robots is critical in many robotic applications, especially in those requiring physical human-robot interaction (pH

safetyarxiv-cs-ro
22 Apr 2026
Safety

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

DGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

safetyarxiv-cs-ai
22 Apr 2026
Safety

User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation

DGX agent

arXiv:2501.04410v2 Announce Type: replace Abstract: User simulation is an emerging interdisciplinary topic with multiple critical applications in the era of Generative AI. It involves creating an inte

safetyarxiv-cs-ai
22 Apr 2026
Safety

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

DGX agent

arXiv:2604.17803v1 Announce Type: cross Abstract: Post-training Large Language Models requires diverse, high-quality data which is rare and costly to obtain, especially in low resource domains and for

safetyarxiv-cs-lg
21 Apr 2026
Safety

Arch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware Design

DGX agent

arXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as

safetyarxiv-cs-cl
21 Apr 2026
Safety

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

DGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

safetyarxiv-cs-cl
21 Apr 2026
Safety

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

DGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

safetyarxiv-cs-cl
21 Apr 2026
Safety

Deep learning based Non-Rigid Volume-to-Surface Registration for Brain Shift compensation Using Point Cloud

DGX agent

arXiv:2604.17389v1 Announce Type: new Abstract: Soft-tissue deformation remains a major limitation in image-guided neurosurgery, where intra-operative anatomy can deviate substantially from pre-operat

safetyarxiv-cs-cv
21 Apr 2026
Safety

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

DGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

safetyarxiv-cs-lg
21 Apr 2026
Safety

Dual Alignment Between Language Model Layers and Human Sentence Processing

DGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

safetyarxiv-cs-cl
21 Apr 2026
Safety

Emergency Stopping for Liquid-manipulating Robots

DGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

safetyarxiv-cs-ro
21 Apr 2026
Safety

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

DGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

safetyarxiv-cs-lg
21 Apr 2026
Safety

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI …

DGX agent

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI of telling people to be more coy about extinction risks. I h

safetyconnor-leahy--x
21 Apr 2026
Safety

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

DGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

safetyarxiv-cs-cv
21 Apr 2026
Safety

LiDAR-based Crowd Navigation with Visible Edge Group Representation

DGX agent

arXiv:2604.16741v1 Announce Type: new Abstract: Robot navigation in crowded pedestrian environments is a well-known challenge and we explore the practical deployment of group-based representations in

safetyarxiv-cs-ro
21 Apr 2026
Safety

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams …

DGX agent

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams actually close the evals feedback loop. Looking forward to I

safetyharrison-chase--x
21 Apr 2026
Safety

Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training (Reuters)

DGX agent

Reuters: Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training — Meta (META.O) is installing new

safetytechmeme
21 Apr 2026
Safety

MoCo: A One-Stop Shop for Model Collaboration Research

DGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

safetyarxiv-cs-cl
21 Apr 2026
Safety

Multimodal Policy Internalization for Conversational Agents

DGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

safetyarxiv-cs-cl
21 Apr 2026
Safety

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

DGX agent

arXiv:2604.16518v1 Announce Type: new Abstract: Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit percepti

safetyarxiv-cs-ro
21 Apr 2026
Safety

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs

DGX agent

arXiv:2604.17543v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable success in general-domain tasks, yet their direct application to the legal domain remains challeng

safetyarxiv-cs-cl
21 Apr 2026
Safety

SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection

DGX agent

arXiv:2604.18476v1 Announce Type: new Abstract: Camera-only 3D object detection has emerged as a cost-effective and scalable alternative to LiDAR for autonomous driving, yet existing methods primarily

safetyarxiv-cs-cv
21 Apr 2026
Safety

Shepherding UAV Swarm with Action Prediction Based on Movement Constraints

DGX agent

arXiv:2604.17189v1 Announce Type: new Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicles (UAVs), which predicts the swarm behavior

safetyarxiv-cs-ro
21 Apr 2026
Safety

The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning

DGX agent

arXiv:2604.17114v1 Announce Type: new Abstract: Frontier large language models generate clinically accurate outputs, but their citations are often fabricated. We term this the Provenance Gap. We teste

safetyarxiv-cs-cl
21 Apr 2026
Safety

Towards Trustworthy Depression Estimation via Disentangled Evidential Learning

DGX agent

arXiv:2604.16579v1 Announce Type: new Abstract: Automated depression estimation is highly vulnerable to signal corruption and ambient noise in real-world deployment. Prevailing deterministic methods p

safetyarxiv-cs-lg
21 Apr 2026
Safety

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts

DGX agent

arXiv:2604.18473v1 Announce Type: new Abstract: Extending a fully post-trained language model with new domain capabilities is fundamentally limited by monolithic training paradigms: retraining from sc

safetyarxiv-cs-lg
21 Apr 2026
Safety

Training Language Models to Use Prolog as a Tool

DGX agent

arXiv:2512.07407v2 Announce Type: replace Abstract: Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use P

safetyarxiv-cs-cl
21 Apr 2026
Safety

UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation

DGX agent

arXiv:2602.09130v3 Announce Type: replace Abstract: Model compression is increasingly essential for deploying large language models (LLMs), yet existing comparative studies largely focus on pruning an

safetyarxiv-cs-lg
21 Apr 2026
Safety

Contact-Aware Planning and Control of Continuum Robots in Highly Constrained Environments

DGX agent

arXiv:2604.15638v1 Announce Type: new Abstract: Continuum robots are well suited for navigating confined and fragile environments, such as vascular or endoluminal anatomy, where contact with surroundi

safetyarxiv-cs-ro
20 Apr 2026
Safety

How people use Copilot for Health

DGX agent

arXiv:2604.15331v1 Announce Type: cross Abstract: We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversati

safetyarxiv-cs-ai
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Model Releases

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

DGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

On the Rejection Criterion for Proxy-based Test-time Alignment

DGX agent

arXiv:2604.16146v1 Announce Type: new Abstract: Recent works proposed test-time alignment methods that rely on a small aligned model as a proxy that guides the generation of a larger base (unaligned)

safetyarxiv-cs-cl
20 Apr 2026
Safety

Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning

DGX agent

arXiv:2506.20020v2 Announce Type: replace Abstract: Reasoning in humans is prone to biases due to underlying motivations like identity protection, that undermine rational decision-making and judgment.

safetyarxiv-cs-ai
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Safety

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

DGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

safetyarxiv-cs-lg
20 Apr 2026
Safety

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the k…

DGX agent

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the kind Anthropic loves to promote. “Accurate” would be more acc

safetygary-marcus--x
18 Apr 2026
Safety

An unsupervised decision-support framework for multivariate biomarker analysis in athlete monitoring

DGX agent

arXiv:2604.14534v1 Announce Type: new Abstract: Purpose. Athlete monitoring is constrained by small cohorts, heterogeneous biomarker scales, limited feasibility of repeated sampling, and the lack of r

safetyarxiv-cs-lg
17 Apr 2026
Local Ai

BitFlipScope: Scalable Fault Localization and Recovery for Bit-Flip Corruptions in LLMs

DGX agent

arXiv:2512.22174v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) deployed in practical and safety-critical settings are increasingly susceptible to bit-flip faults caused by hard

local-aiarxiv-cs-ai
17 Apr 2026
Safety

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

DGX agent

arXiv:2604.14892v1 Announce Type: new Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudi

safetyarxiv-cs-lg
17 Apr 2026
Safety

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

DGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

safetyarxiv-cs-cl
17 Apr 2026
Safety

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

DGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

safetyarxiv-cs-lg
17 Apr 2026
Safety

Humanoid Factors: Design Principles for AI Humanoids in Human Worlds

DGX agent

arXiv:2602.10069v2 Announce Type: replace Abstract: Human factors research has long focused on optimizing environments, tools, and systems to account for human performance. Yet, as humanoid robots beg

safetyarxiv-cs-ro
17 Apr 2026
← Previous
1…6061626364…300
Next →