AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

UGAF-ITS: A Standards Harmonization Framework and Validation Tool for Multi-Framework AI Governance in Distributed Intelligent Transportation Systems

DGX agent

arXiv:2604.22789v1 Announce Type: cross Abstract: Organizations deploying AI-enabled Intelligent Transportation Systems face fragmented governance: ISO/IEC 42001 demands a certifiable management syste

safetyarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems

DGX agent

arXiv:2604.23362v1 Announce Type: cross Abstract: Adversarial attacks play a pivotal role in testing and improving the reliability of deep learning (DL) systems. Existing literature has demonstrated t

safetyarxiv-cs-lg
28 Apr 2026
Safety

Verifying Quantized GNNs With Readout Is Decidable But Highly Intractable

DGX agent

arXiv:2510.08045v2 Announce Type: replace-cross Abstract: We introduce a logical language for reasoning about quantized aggregate-combine graph neural networks with global readout (ACR-GNNs). We provi

safetyarxiv-cs-ai
28 Apr 2026
Safety

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning

DGX agent

arXiv:2604.22873v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) can learn effective policies from fixed datasets, but deployment objectives may change after training, and in many

safetyarxiv-cs-ai
28 Apr 2026
Safety

A Co-Evolutionary Theory of Human-AI Coexistence: Mutualism, Governance, and Dynamics in Complex Societies

DGX agent

arXiv:2604.22227v1 Announce Type: cross Abstract: Classical robot ethics is often framed around obedience, most famously through Asimov's laws. This framing is too narrow for contemporary AI systems,

safetyarxiv-cs-ai
27 Apr 2026
Safety

How Large Language Models Balance Internal Knowledge with User and Document Assertions

DGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

safetyarxiv-cs-cl
27 Apr 2026
Safety

Is it me or X starting to look like a vibe coded mess? Polls are broken. Accounts are getting hacked. My DMs are full of phishing scams. Bas…

DGX agent

Gary Marcus discusses technical and security issues affecting the X platform, including malfunctioning polls, compromised accounts, and increased phishing scams in direct messages. The post appears to

safetygary-marcus--x
27 Apr 2026
Safety

Learning-augmented robotic automation for real-world manufacturing

DGX agent

arXiv:2604.22235v1 Announce Type: cross Abstract: Industrial robots are widely used in manufacturing, yet most manipulation still depends on fixed waypoint scripts that are brittle to environmental ch

safetyarxiv-cs-ai
27 Apr 2026
Safety

On the Properties of Feature Attribution for Supervised Contrastive Learning

DGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

safetyarxiv-cs-ai
27 Apr 2026
Safety

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

DGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

safetyarxiv-cs-ai
27 Apr 2026
Safety

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

DGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

safetyarxiv-cs-ai
27 Apr 2026
Safety

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have f…

DGX agent

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have financial gain from OpenAI. Ex-board member of OpenAI calls S

safetyelon-musk--x
27 Apr 2026
Safety

The only AGI that Sam Altman is after is Adjusted Gross Income.

DGX agent

Gary Marcus made a critical commentary on Sam Altman's priorities, using a pun on 'AGI' (Artificial General Intelligence) to suggest that Altman's actual focus is on 'Adjusted Gross Income' rather tha

safetygary-marcus--x
27 Apr 2026
Safety

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We s…

DGX agent

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We started the Torchbearer Community for exactly this reason. Re

safetyconnor-leahy--x
27 Apr 2026
Safety

Coders and software engineers ONLY:

DGX agent

Gary Marcus, a prominent AI researcher and critic, posted a message on X (formerly Twitter) addressing software engineers and coders, likely discussing technical aspects of AI development, programming

safetygary-marcus--x
26 Apr 2026
Safety

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good f…

DGX agent

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good fake crisis, which is why they're perfect for Bernie to try t

safetyyann-lecun--x
26 Apr 2026
Safety

What if I told you there was a technology where 1.5 million people would die every year and injure 50 million would you sign up for that tec…

DGX agent

What if I told you there was a technology where 1.5 million people would die every year and injure 50 million would you sign up for that tech? Hell no, right? But the answer is actually 'hell yes' bec

safetyyann-lecun--x
26 Apr 2026
Safety

A Survey of Legged Robotics in Non-Inertial Environments: Past, Present, and Future

DGX agent

arXiv:2604.20990v1 Announce Type: new Abstract: Legged robots have demonstrated remarkable agility on rigid, stationary ground, but their locomotion reliability remains limited in non-inertial environ

safetyarxiv-cs-ro
24 Apr 2026
Safety

Automated Annotation of Shearographic Measurements Enabling Weakly Supervised Defect Detection

DGX agent

arXiv:2512.06171v2 Announce Type: replace Abstract: Shearography is an interferometric technique sensitive to surface displacement gradients, providing high sensitivity for detecting subsurface defect

safetyarxiv-cs-cv
24 Apr 2026
Safety

CARE: Counselor-Aligned Response Engine for Online Mental-Health Support

DGX agent

arXiv:2604.21352v1 Announce Type: new Abstract: Mental health challenges are increasing worldwide, straining emotional support services and leading to counselor overload. This can result in delayed re

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI

DGX agent

arXiv:2604.20972v1 Announce Type: new Abstract: Content moderation systems are typically evaluated by measuring agreement with human labels. In rule-governed environments this assumption fails: multip

safetyarxiv-cs-ai
24 Apr 2026
Safety

Fairness Evaluation and Inference Level Mitigation in LLMs

DGX agent

arXiv:2510.18914v4 Announce Type: replace-cross Abstract: Large language models often display undesirable behaviors embedded in their internal representations, undermining fairness, inconsistency drif

safetyarxiv-cs-ai
24 Apr 2026
Safety

In 1982, Blade Runner was speculation. In 2026, it's the conversation. Memory, personhood, alignment, what we owe the minds we build. Watchi…

DGX agent

In 1982, Blade Runner was speculation. In 2026, it's the conversation. Memory, personhood, alignment, what we owe the minds we build. Watching it Thursday 28 May, Kensington Central Library. Come thin

safetyemad-mostaque--x
24 Apr 2026
Safety

Inferring High-Level Events from Timestamped Data: Complexity and Medical Applications

DGX agent

arXiv:2604.21793v1 Announce Type: new Abstract: In this paper, we develop a novel logic-based approach to detecting high-level temporally extended events from timestamped data and background knowledge

safetyarxiv-cs-ai
24 Apr 2026
Safety

Language-Conditioned Safe Trajectory Generation for Spacecraft Rendezvous

DGX agent

arXiv:2512.09111v3 Announce Type: replace-cross Abstract: Reliable real-time trajectory generation is essential for future autonomous spacecraft. While recent progress in nonconvex guidance and contro

safetyarxiv-cs-ai
24 Apr 2026
Safety

Probabilistic Verification of Neural Networks via Efficient Probabilistic Hull Generation

DGX agent

arXiv:2604.21556v1 Announce Type: new Abstract: The problem of probabilistic verification of a neural network investigates the probability of satisfying the safe constraints in the output space when t

safetyarxiv-cs-ai
24 Apr 2026
Safety

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

DGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

safetyarxiv-cs-ai
24 Apr 2026
Safety

TraceScope: Interactive URL Triage via Decoupled Checklist Adjudication

DGX agent

arXiv:2604.21840v1 Announce Type: cross Abstract: Modern phishing campaigns increasingly evade snapshot-based URL classifiers using interaction gates (e.g., checkbox/slider challenges), delayed conten

safetyarxiv-cs-ai
24 Apr 2026
Safety

Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

DGX agent

arXiv:2512.08216v3 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessm

safetyarxiv-cs-cv
24 Apr 2026
Safety

Unbiased Prevalence Estimation with Multicalibrated LLMs

DGX agent

arXiv:2604.21549v1 Announce Type: new Abstract: Estimating the prevalence of a category in a population using imperfect measurement devices (diagnostic tests, classifiers, or large language models) is

safetyarxiv-cs-ai
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

Environmental Understanding Vision-Language Model for Embodied Agent

DGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

safetyarxiv-cs-ai
23 Apr 2026
Safety

Explainable AML Triage with LLMs: Evidence Retrieval and Counterfactual Checks

DGX agent

arXiv:2604.19755v1 Announce Type: new Abstract: Anti-money laundering (AML) transaction monitoring generates large volumes of alerts that must be rapidly triaged by investigators under strict audit an

safetyarxiv-cs-ai
23 Apr 2026
Safety

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

DGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

safetyarxiv-cs-ai
23 Apr 2026
Safety

From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP

DGX agent

arXiv:2510.12817v3 Announce Type: replace-cross Abstract: Human Label Variation (HLV) refers to legitimate disagreement in annotation that reflects the diversity of human perspectives rather than mere

safetyarxiv-cs-ai
23 Apr 2026
Safety

FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory

DGX agent

arXiv:2604.20300v1 Announce Type: new Abstract: For LLM agents, memory management critically impacts efficiency, quality, and security. While much research focuses on retention, selective forgetting--

safetyarxiv-cs-ai
23 Apr 2026
Safety

ICLR 2026: 12 papers on making AI systems reliable, efficient, and secure

DGX agent

A 7B agent that beats GPT-4o. Lossless weight compression that speeds up inference by 177%. An arena where 23 teams battled across 103,000 adversarial rounds. This year at ICLR, Lambda is presenting t

safetylambda-labs
23 Apr 2026
Safety

Large language models perceive cities through a culturally uneven baseline

DGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

safetyarxiv-cs-cl
23 Apr 2026
Safety

MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

DGX agent

arXiv:2604.20441v1 Announce Type: new Abstract: Background: Agent skills are increasingly deployed as modular, reusable capability units in AI agent systems. Medical research agent skills require safe

safetyarxiv-cs-ai
23 Apr 2026
Safety

NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answering

DGX agent

arXiv:2602.15353v2 Announce Type: replace-cross Abstract: Large pretrained language models and neural reasoning systems have advanced many natural language tasks, yet they remain challenged by knowled

safetyarxiv-cs-ai
23 Apr 2026
Safety

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

DGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains

DGX agent

arXiv:2508.15832v2 Announce Type: replace-cross Abstract: Web agents have shown great promise in performing many tasks on ecommerce website. To assess their capabilities, several benchmarks have been

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

DGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

safetyarxiv-cs-ro
22 Apr 2026
Safety

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

DGX agent

arXiv:2506.22174v2 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway

safetyarxiv-cs-lg
22 Apr 2026
Safety

CAHAL: Clinically Applicable resolution enHAncement for Low-resolution MRI scans

DGX agent

arXiv:2604.18781v1 Announce Type: new Abstract: Large-scale automated morphometric analysis of brain MRI is limited by the thick-slice, anisotropic acquisitions prevalent in routine clinical practice.

safetyarxiv-cs-cv
22 Apr 2026
Safety

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

DGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

safetyarxiv-cs-ai
22 Apr 2026
Safety

Developing a Robotic Surgery Training System for Wide Accessibility and Research

DGX agent

arXiv:2505.20562v2 Announce Type: replace Abstract: Robotic surgery represents a major breakthrough in medical interventions, which has revolutionized surgical procedures. However, the high cost and l

safetyarxiv-cs-ro
22 Apr 2026
← Previous
1…5960616263…300
Next →