AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
Safety

A comparative study of accuracy and rollout stability of temporal surrogate models

DGX agent

arXiv:2605.24868v1 Announce Type: new Abstract: Temporal surrogate models are effective for predicting chaotic dynamical systems where computational cost can be prohibitive. Several deep neural networ

safetyarxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

A Decentralized LiDAR-SLAM System with Certifiably Optimal Pose Graph Optimization

DGX agent

arXiv:2605.25051v1 Announce Type: new Abstract: Decentralized multi-robot LiDAR-SLAM is essential for collaborative missions but faces significant challenges in maintaining global consistency. Existin

safetyarxiv-cs-ro
26 May 2026
Safety

A governance horizon for ethical-use constraints in open-weight AI models

DGX agent

arXiv:2605.24383v1 Announce Type: new Abstract: Ethical constraints on open-weight AI models are both a reflection of societal concerns and a foundation for AI governance policy. They are expected to

safetyarxiv-cs-ai
26 May 2026
Safety

A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring

DGX agent

arXiv:2605.26026v1 Announce Type: cross Abstract: Light sheet fluorescence microscopy (LSM) enables high-resolution, three-dimensional (3D) imaging of biological specimens, providing rich volumetric d

safetyarxiv-cs-ai
26 May 2026
Safety

A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning

DGX agent

arXiv:2605.25540v1 Announce Type: cross Abstract: Alzheimer's disease (AD) is a progressive neurodegenerative disorder and the leading cause of dementia, affecting memory, reasoning, communication, an

safetyarxiv-cs-lg
26 May 2026
Safety

A post about Pope Leo XIV's encyclical on AI. Why the Pope is right, but perhaps not right enough. Artificial intelligence is reshaping the …

DGX agent

A post about Pope Leo XIV's encyclical on AI. Why the Pope is right, but perhaps not right enough. Artificial intelligence is reshaping the world in front of our eyes: how we communicate, how we acces

safetyyann-lecun--x
26 May 2026
Safety

a shout out to the paper: https://arxiv.org/html/2605.25376v1 'KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable …

DGX agent

KYA is a framework-agnostic trust layer designed for autonomous systems that provides verifiable guarantees, addressing the need for trustworthy and transparent operation of AI agents across different

safetyyohei-nakajima--x
26 May 2026
Safety

A Sober Look at Agentic Misalignment in Automated Workflows

DGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

safetyarxiv-cs-ai
26 May 2026
Safety

A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions

DGX agent

arXiv:2605.25536v1 Announce Type: cross Abstract: Context. Large language models (LLMs) are increasingly applied to code-generating tasks (CGTs) in software engineering. While reported results are pro

safetyarxiv-cs-ai
26 May 2026
Safety

A Unified Python Framework for Direct PPO-based Control of AHUs with Economizer Logic and CO2-Constrained Ventilation

DGX agent

arXiv:2605.24406v1 Announce Type: new Abstract: Optimizing HVAC (Heating, Ventilation and Air Conditioning) can enhance a building's energy efficiency while providing comfort levels for its occupants.

safetyarxiv-cs-lg
26 May 2026
Safety

Active Learning for Stochastic Contextual Linear Bandits

DGX agent

arXiv:2605.24803v1 Announce Type: new Abstract: A key goal in stochastic contextual linear bandits is to efficiently learn a near-optimal policy. Prior algorithms for this problem learn a policy by st

safetyarxiv-cs-lg
26 May 2026
Safety

Adaptive Human-AI Coordination via Hierarchical Action Disentanglement

DGX agent

arXiv:2605.24343v1 Announce Type: new Abstract: Human-AI collaboration requires agents that can adapt to diverse partner behaviors and skill levels while remaining robust to unseen partners. Existing

safetyarxiv-cs-ai
26 May 2026
Safety

Adaptive Preference Optimization with Uncertainty-aware Utility Anchor

DGX agent

arXiv:2509.10515v1 Announce Type: cross Abstract: Offline preference optimization methods are efficient for large language models (LLMs) alignment. Direct Preference optimization (DPO)-like learning,

safetyarxiv-cs-cl
26 May 2026
Safety

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models

DGX agent

arXiv:2605.26013v1 Announce Type: cross Abstract: We introduce AdvantageFlow, a forward-process reinforcement learning algorithm for rectified flow models. Unlike Flow-GRPO, which optimizes the revers

safetyarxiv-cs-ai
26 May 2026
Safety

Adversarial Error Correction for Visual Autoregressive Generation

DGX agent

arXiv:2605.24843v1 Announce Type: cross Abstract: Visual Autoregressive (VAR) models have emerged as a powerful paradigm for image synthesis by performing hierarchical next-scale prediction. However,

safetyarxiv-cs-ai
26 May 2026
Model Releases

AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue

DGX agent

arXiv:2605.23974v1 Announce Type: new Abstract: Current language models create two safety challenges: risk must be detected early enough to avoid exposing harmful continuation, and the harmfulness its

model-releasesarxiv-cs-cl
26 May 2026
Safety

Agent-Centric Social Trajectory Prediction: A Free Energy Principle Perspective

DGX agent

arXiv:2605.25748v1 Announce Type: new Abstract: Trajectory prediction methods have demonstrated remarkable capabilities in capturing complex motion patterns. However, existing methods rely on global s

safetyarxiv-cs-ai
26 May 2026
Safety

Agent-Facing Information Design in LLM Tool Registries

DGX agent

arXiv:2605.23916v1 Announce Type: cross Abstract: LLM tool registries function as unregulated advertising platforms: providers write free-text descriptions that agents use for selection, yet no measur

safetyarxiv-cs-ai
26 May 2026
Safety

Agent Learning via Early Experience

DGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

safetyarxiv-cs-ai
26 May 2026
Safety

Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning

DGX agent

arXiv:2605.24216v1 Announce Type: cross Abstract: Monitoring autonomous large language model (LLM) agents for covert malicious behavior is challenging due to delayed, context-dependent, and long-horiz

safetyarxiv-cs-ai
26 May 2026
Safety

Agents and AI responsibility; nice clip from @thsottiaux and @siliconvalleymm

DGX agent

Gary Marcus shares a video clip discussing the intersection of AI agents and questions of responsibility, featuring contributors Thierry Souttiaux and Silicon Valley commentators. The post highlights

safetygary-marcus--x
26 May 2026
Safety

AI-Assisted Systematization for Evaluating GenAI Systems

DGX agent

arXiv:2605.26001v1 Announce Type: cross Abstract: Evaluating generative AI (GenAI) systems is challenging because many targets of evaluation are broad, contested concepts, such as 'reasoning,' 'fairne

safetyarxiv-cs-ai
26 May 2026
Safety

AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains

DGX agent

arXiv:2605.23946v1 Announce Type: cross Abstract: Climate volatility, regional production concentration, labor constraints, cyber risk, and dependence on long-distance fresh-produce supply chains expo

safetyarxiv-cs-ai
26 May 2026
Safety

Analogies between Transformer Layers and Power Method

DGX agent

arXiv:2605.25619v1 Announce Type: new Abstract: In the paper we show that there is an analogy between the operations occurring in a layer of a transformer (projections and layer normalizations, disreg

safetyarxiv-cs-lg
26 May 2026
Safety

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

DGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth

safetyarxiv-cs-ai
26 May 2026
Safety

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to t…

DGX agent

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to take a big chunk of your retirement, and unless you call your

safetygary-marcus--x
26 May 2026
Safety

Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

DGX agent

arXiv:2605.23938v1 Announce Type: new Abstract: Large language models (LLMs) increasingly fuse heterogeneous inputs in ubiquitous systems. Yet, how LLMs implicitly allocate authority when sensor measu

safetyarxiv-cs-ai
26 May 2026
Safety

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

DGX agent

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

safetyarxiv-cs-lg
26 May 2026
Safety

AvAtar: Learning to Align via Active Optimal Transport

DGX agent

arXiv:2605.24395v1 Announce Type: new Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration.

safetyarxiv-cs-lg
26 May 2026
Safety

Balancing Fairness, Privacy, and Accuracy: A Multitask Adversarial Framework for Centralized Data-Driven Systems

DGX agent

arXiv:2605.24458v1 Announce Type: cross Abstract: The integration of fairness and privacy in centralized data-driven applications is critical, especially as these systems increasingly influence sector

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

DGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

DGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

DGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

safetyarxiv-cs-cl
26 May 2026
Safety

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

DGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

safetyarxiv-cs-lg
26 May 2026
Safety

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

DGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

safetyarxiv-cs-ai
26 May 2026
Safety

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

DGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

safetyarxiv-cs-cl
26 May 2026
Safety

Causal methods for LLM development and evaluation

DGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

safetyarxiv-cs-lg
26 May 2026
Safety

Characterizing Linear Alignment Across Language Models

DGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

safetyarxiv-cs-ai
26 May 2026
Safety

Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems

DGX agent

arXiv:2605.25290v1 Announce Type: cross Abstract: Online experiments in ads, recommendation, and member-experience systems are often planned before the dominant interference mechanism is known. A trea

safetyarxiv-cs-lg
26 May 2026
Safety

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

DGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

safetyarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Safety

ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

DGX agent

arXiv:2605.25553v1 Announce Type: cross Abstract: Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with th

safetyarxiv-cs-ro
26 May 2026
Safety

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

DGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

safetyarxiv-cs-ai
26 May 2026
Safety

Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2602.08499v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an effective paradigm for improving the reasoning capabilities of large language mode

safetyarxiv-cs-ai
26 May 2026
Safety

corruption

DGX agent

corruption Elon Musk has used Trump’s war in Iran to quintuple the amount he’s charging the Pentagon for SpaceX services. https://www.thedailybeast.com/elon-musks-spacex-uses-trumps-war-to-squeeze-mor

safetygary-marcus--x
26 May 2026
Safety

Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning

DGX agent

arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this pa

safetyarxiv-cs-ai
26 May 2026
Safety

CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data

DGX agent

arXiv:2601.10494v2 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Managemen

safetyarxiv-cs-lg
26 May 2026
Safety

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

DGX agent

arXiv:2605.25511v1 Announce Type: new Abstract: Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning ca

safetyarxiv-cs-cl
26 May 2026
← Previous
1…181182183184185…302
Next →