AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning

DGX agent

arXiv:2605.24216v1 Announce Type: cross Abstract: Monitoring autonomous large language model (LLM) agents for covert malicious behavior is challenging due to delayed, context-dependent, and long-horiz

safetyarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Agents and AI responsibility; nice clip from @thsottiaux and @siliconvalleymm

DGX agent

Gary Marcus shares a video clip discussing the intersection of AI agents and questions of responsibility, featuring contributors Thierry Souttiaux and Silicon Valley commentators. The post highlights

safetygary-marcus--x
26 May 2026
Safety

AI-Assisted Systematization for Evaluating GenAI Systems

DGX agent

arXiv:2605.26001v1 Announce Type: cross Abstract: Evaluating generative AI (GenAI) systems is challenging because many targets of evaluation are broad, contested concepts, such as 'reasoning,' 'fairne

safetyarxiv-cs-ai
26 May 2026
Safety

AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains

DGX agent

arXiv:2605.23946v1 Announce Type: cross Abstract: Climate volatility, regional production concentration, labor constraints, cyber risk, and dependence on long-distance fresh-produce supply chains expo

safetyarxiv-cs-ai
26 May 2026
Safety

Analogies between Transformer Layers and Power Method

DGX agent

arXiv:2605.25619v1 Announce Type: new Abstract: In the paper we show that there is an analogy between the operations occurring in a layer of a transformer (projections and layer normalizations, disreg

safetyarxiv-cs-lg
26 May 2026
Safety

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

DGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth

safetyarxiv-cs-ai
26 May 2026
Safety

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to t…

DGX agent

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to take a big chunk of your retirement, and unless you call your

safetygary-marcus--x
26 May 2026
Safety

Anthropic: fantasy versus reality.

DGX agent

Gary Marcus critiques Anthropic's claims about their AI capabilities, likely contrasting inflated marketing promises with the actual technical limitations and performance of their systems. The post pr

safetygary-marcus--x
26 May 2026
Safety

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond

DGX agent

arXiv:2605.26113v1 Announce Type: new Abstract: Generating high-fidelity and controllable synthetic data is critical for advancing end-to-end autonomous driving, particularly for addressing the long t

safetyarxiv-cs-ro
26 May 2026
Safety

Approximating Safety Feedback Without a Safety Oracle via Model Predictive Control

DGX agent

arXiv:2510.20955v2 Announce Type: replace Abstract: Safe decision-making algorithms for control of mobile robots often require the existence of feedback to verify the safety of proposed actions. This

safetyarxiv-cs-lg
26 May 2026
Safety

Auditing medical multi-agent AI reveals risks of false consensus

DGX agent

arXiv:2510.10185v2 Announce Type: replace-cross Abstract: Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through sp

safetyarxiv-cs-ai
26 May 2026
Safety

Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

DGX agent

arXiv:2605.23938v1 Announce Type: new Abstract: Large language models (LLMs) increasingly fuse heterogeneous inputs in ubiquitous systems. Yet, how LLMs implicitly allocate authority when sensor measu

safetyarxiv-cs-ai
26 May 2026
Safety

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

DGX agent

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

safetyarxiv-cs-lg
26 May 2026
Safety

AvAtar: Learning to Align via Active Optimal Transport

DGX agent

arXiv:2605.24395v1 Announce Type: new Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration.

safetyarxiv-cs-lg
26 May 2026
Safety

Balancing Fairness, Privacy, and Accuracy: A Multitask Adversarial Framework for Centralized Data-Driven Systems

DGX agent

arXiv:2605.24458v1 Announce Type: cross Abstract: The integration of fairness and privacy in centralized data-driven applications is critical, especially as these systems increasingly influence sector

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

DGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

DGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

DGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

safetyarxiv-cs-cl
26 May 2026
Safety

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

DGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

safetyarxiv-cs-lg
26 May 2026
Safety

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

DGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

safetyarxiv-cs-ai
26 May 2026
Safety

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

DGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

safetyarxiv-cs-cl
26 May 2026
Safety

Capability and Robustness Cannot Both Be Free: An Information-Theoretic Bound for Vision-Language-Action Models

DGX agent

arXiv:2605.25889v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed on real robots, where each predicted action is executed and each failure carries a safet

safetyarxiv-cs-lg
26 May 2026
Safety

Causal methods for LLM development and evaluation

DGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

safetyarxiv-cs-lg
26 May 2026
Safety

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

DGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

safetyarxiv-cs-ai
26 May 2026
Safety

Characterizing Linear Alignment Across Language Models

DGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

safetyarxiv-cs-ai
26 May 2026
Safety

Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems

DGX agent

arXiv:2605.25290v1 Announce Type: cross Abstract: Online experiments in ads, recommendation, and member-experience systems are often planned before the dominant interference mechanism is known. A trea

safetyarxiv-cs-lg
26 May 2026
Safety

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

DGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

safetyarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Safety

ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

DGX agent

arXiv:2605.25553v1 Announce Type: cross Abstract: Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with th

safetyarxiv-cs-ro
26 May 2026
Safety

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

DGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

safetyarxiv-cs-ai
26 May 2026
Safety

Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2602.08499v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an effective paradigm for improving the reasoning capabilities of large language mode

safetyarxiv-cs-ai
26 May 2026
Safety

Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents

DGX agent

arXiv:2605.22634v2 Announce Type: replace-cross Abstract: Skills have become a practical packaging mechanism for agent instructions, workflows, scripts, and reference materials. In enterprise settings

safetyarxiv-cs-ai
26 May 2026
Safety

corruption

DGX agent

corruption Elon Musk has used Trump’s war in Iran to quintuple the amount he’s charging the Pentagon for SpaceX services. https://www.thedailybeast.com/elon-musks-spacex-uses-trumps-war-to-squeeze-mor

safetygary-marcus--x
26 May 2026
Safety

Counterfactually Safe Reinforcement Learning

DGX agent

arXiv:2605.25114v1 Announce Type: cross Abstract: Reinforcement learning algorithms are generally designed to maximize the expected return across a population. However, a policy that is optimal on ave

safetyarxiv-cs-lg
26 May 2026
Safety

Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning

DGX agent

arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this pa

safetyarxiv-cs-ai
26 May 2026
Safety

CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data

DGX agent

arXiv:2601.10494v2 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Managemen

safetyarxiv-cs-lg
26 May 2026
Safety

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

DGX agent

arXiv:2605.25511v1 Announce Type: new Abstract: Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning ca

safetyarxiv-cs-cl
26 May 2026
Safety

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

DGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

safetyarxiv-cs-ai
26 May 2026
Safety

CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum

DGX agent

arXiv:2605.24663v1 Announce Type: cross Abstract: This paper presents CyBOKClaw, an interpretable human-in-the-loop retrieval framework for mapping cybersecurity keywords or phrases (KWoPs) to the Cyb

safetyarxiv-cs-ai
26 May 2026
Safety

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

DGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

safetyarxiv-cs-ai
26 May 2026
Safety

Dale meets Langevin: A Multiplicative Denoising Diffusion Model

DGX agent

arXiv:2510.02730v2 Announce Type: replace Abstract: Exponentiated gradient descent (EGD), a biologically motivated optimisation algorithm that respects Dale's law, produces log-normally distributed sy

safetyarxiv-cs-lg
26 May 2026
Safety

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

DGX agent

arXiv:2409.20473v3 Announce Type: replace Abstract: Tactile sensing is critical for learning-based dexterous manipulation, yet principled guidelines for sensor placement remain largely absent. While d

safetyarxiv-cs-ro
26 May 2026
Safety

DBPnet: Damper Characteristics-Based Bayesian Physics-Informed Neural Network for Wheel Load Estimation

DGX agent

arXiv:2605.24860v1 Announce Type: cross Abstract: Advanced driver assistance systems (ADAS) play an important role in modern automotive intelligence, significantly enhancing vehicle safety and stabili

safetyarxiv-cs-ai
26 May 2026
Safety

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care

DGX agent

arXiv:2510.08350v3 Announce Type: replace-cross Abstract: Objective: Enteral nutrition (EN) delivery in the ICU remains suboptimal due to limited personalization and uncertainty regarding appropriate

safetyarxiv-cs-ai
26 May 2026
Safety

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

DGX agent

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradien

safetyarxiv-cs-lg
26 May 2026
Safety

DeGRe: Dense-supervised Generative Reranking for Recommendation

DGX agent

arXiv:2605.25749v1 Announce Type: cross Abstract: In multi-stage recommender systems, reranking optimizes overall utility by capturing intra-list contextual dependencies, yet its central challenge lie

safetyarxiv-cs-ai
26 May 2026
Safety

Delta Energy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization

DGX agent

arXiv:2510.11296v3 Announce Type: replace-cross Abstract: Recent approaches for vision-language models (VLMs) have shown remarkable success in achieving fast downstream adaptation. When applied to rea

safetyarxiv-cs-lg
26 May 2026
Safety

Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL

DGX agent

arXiv:2605.24001v1 Announce Type: cross Abstract: Recent advances in one-step text-to-image generation have enabled real-time synthesis with remarkable efficiency and quality. Previous reinforcement l

safetyarxiv-cs-ai
26 May 2026
← Previous
1…144145146147148…267
Next →