AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
Safety

lol. OpenAI as the WeWork of AI. Literally called it, in those exact words, @CNBC w @carlquintanilla, 2024. Now even SoftBank is worried.

DGX agent

lol. OpenAI as the WeWork of AI. Literally called it, in those exact words, @CNBC w @carlquintanilla, 2024. Now even SoftBank is worried. SoftBank's own executives think Sam Altman is scamming their C

safetygary-marcus--x
26 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Looking forward to speaking at BloombergTech next week with @shiringhaffary to discuss AI safety and our research progress at @LawZero_!

DGX agent

Looking forward to speaking at BloombergTech next week with @shiringhaffary to discuss AI safety and our research progress at @LawZero_! Is AI development progressing too quickly? @business' @shiringh

safetyyoshua-bengio--x
26 May 2026
Safety

Machine Psychometrics: A Mathematical Psychology of Artificial Intelligence

DGX agent

arXiv:2605.23952v1 Announce Type: new Abstract: Artificial agents now generate behavior rich enough to invite trust, surprise, and concern, yet our evaluation tools still privilege capability scores o

safetyarxiv-cs-ai
26 May 2026
Safety

MAGIC: Multimodal Alignment & Grounding-aware Instruction Coreset for Vision-Language Models

DGX agent

arXiv:2605.26004v1 Announce Type: cross Abstract: Instruction tuning of large vision-language models (LVLMs) increasingly depends on massive multimodal corpora, yet these datasets contain samples with

safetyarxiv-cs-cl
26 May 2026
Safety

MAPLE: Multi-State Aggregated Policy Evaluation for AlphaZero in Imperfect-Information Games

DGX agent

arXiv:2605.24139v1 Announce Type: new Abstract: Imperfect-information games (IIGs) are challenging, as players must make decisions without fully observing the true game state. While AlphaZero has achi

safetyarxiv-cs-ai
26 May 2026
Safety

MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling

DGX agent

arXiv:2602.17658v2 Announce Type: replace-cross Abstract: Reward modeling is central to alignment pipelines such as RLHF, RLAIF, and PPO-based policy optimization, yet its reliability is constrained b

safetyarxiv-cs-ai
26 May 2026
Safety

MATO: Multi-objective Personalized Alignment with Test-time Optimization for Large Language Models

DGX agent

arXiv:2605.25342v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse and multifaceted user preferences is a fundamental challenge in personalized AI systems. Existing mul

safetyarxiv-cs-cl
26 May 2026
Safety

Measuring the Depth of LLM Unlearning via Activation Patching

DGX agent

arXiv:2605.24614v1 Announce Type: cross Abstract: Large language model (LLM) unlearning has emerged as a crucial post-hoc mechanism for privacy protection and AI safety, yet auditing whether target kn

safetyarxiv-cs-ai
26 May 2026
Safety

MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents

DGX agent

arXiv:2602.02474v2 Announce Type: replace-cross Abstract: Most Large Language Model (LLM) agent memory systems rely on a small set of static, hand-designed operations for extracting memory. These fixe

safetyarxiv-cs-ai
26 May 2026
Safety

Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI

DGX agent

arXiv:2605.23981v1 Announce Type: cross Abstract: Generative AI research increasingly confronts a shared problem: systems must sustain yet govern their own generative activity when uncertainty is high

safetyarxiv-cs-ai
26 May 2026
Safety

Micro-Swarm Locomotion Optimization in Dynamic Flow using Multi-Objective Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.25025v1 Announce Type: new Abstract: Coordinating micro-robotic swarms in physiologically realistic, time-dependent fluid environments remains an unsolved challenge for biomedical and envir

safetyarxiv-cs-ro
26 May 2026
Safety

Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation

DGX agent

arXiv:2512.16189v3 Announce Type: replace Abstract: In healthcare, it is essential for any LLM-generated output to be reliable and accurate, particularly in cases involving decision-making and patient

safetyarxiv-cs-cl
26 May 2026
Safety

MMUEChange: A Generalized LLM Agent Framework for Intelligent Multi-Modal Urban Environment Change Analysis

DGX agent

arXiv:2601.05483v2 Announce Type: replace Abstract: Understanding urban environment change is essential for sustainable development. However, current approaches, particularly remote sensing change det

safetyarxiv-cs-ai
26 May 2026
Safety

Motion-Compensated Weight Compression

DGX agent

arXiv:2605.24754v1 Announce Type: cross Abstract: Neural network weights are increasingly a bottleneck for deployment, yet most compression pipelines treat layers independently and overlook cross-laye

safetyarxiv-cs-ai
26 May 2026
Safety

MuGen: Multi-Skill Generative Locomotion Controller for Humanoid Robots

DGX agent

arXiv:2605.24592v1 Announce Type: new Abstract: This paper presents MuGen, a data-driven framework for learning and deploying multi-skill locomotion on humanoid robots. MuGen enables a robot to perfor

safetyarxiv-cs-ro
26 May 2026
Safety

Multi-Agent Coordination Adaptation via Structure-Guided Orchestration

DGX agent

arXiv:2605.25746v1 Announce Type: cross Abstract: As large language model (LLM)-based multi-agent systems scale to handle increasingly complex tasks, balancing structural stability and dynamic adaptab

safetyarxiv-cs-ai
26 May 2026
Safety

Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?

DGX agent

arXiv:2605.25929v1 Announce Type: cross Abstract: The effectiveness of multi-agent LLM deliberation depends not only on the agents' individual predictions, but also on how they communicate and collabo

safetyarxiv-cs-lg
26 May 2026
Safety

Multi-Alignment Contrastive Learning for Enzyme--Reaction Retrieval

DGX agent

arXiv:2512.08508v2 Announce Type: replace-cross Abstract: Identifying enzymes that catalyze target biochemical reactions is a key step in computational enzyme discovery and biocatalyst design. Recent

safetyarxiv-cs-lg
26 May 2026
Safety

Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval

DGX agent

arXiv:2209.11572v3 Announce Type: replace-cross Abstract: As an increasingly popular task in multimedia information retrieval, video moment retrieval (VMR) aims to localize the target moment from an u

safetyarxiv-cs-ai
26 May 2026
Safety

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning

DGX agent

arXiv:2605.25210v1 Announce Type: cross Abstract: Diffusion models are increasingly used as powerful conditional generators, yet real deployments often involve multiple target distributions arising fr

safetyarxiv-cs-ai
26 May 2026
Safety

Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation

DGX agent

arXiv:2605.25220v1 Announce Type: cross Abstract: High-fidelity 3D Gaussian head avatar generation is critical for applications such as AR/VR, telepresence, and digital humans. Existing methods depend

safetyarxiv-cs-ro
26 May 2026
Safety

Multicalibration Boosting: Theory, Convergence, and Transferability

DGX agent

arXiv:2605.24364v1 Announce Type: cross Abstract: Multicalibration extends classical calibration by requiring predictions to be unbiased over a rich collection of functions, encompassing both predicti

safetyarxiv-cs-lg
26 May 2026
Safety

Multimodal Alignment and Preference Optimization for Zero-Shot Conditional RNA Generation

DGX agent

arXiv:2605.23961v1 Announce Type: cross Abstract: The design of RNA molecules that interact with specific proteins is a critical challenge in experimental and computational biology. Despite recent pro

safetyarxiv-cs-ai
26 May 2026
Safety

Multimodal Functional Maximum Correlation for Emotion Recognition

DGX agent

arXiv:2512.23076v2 Announce Type: replace-cross Abstract: Emotional states manifest as coordinated yet heterogeneous physiological responses across central and autonomic systems, posing a fundamental

safetyarxiv-cs-ai
26 May 2026
Safety

MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection

DGX agent

arXiv:2505.23803v2 Announce Type: replace-cross Abstract: Phishing email detection faces significant challenges due to evolving adversarial tactics and heterogeneous attack patterns. Traditional appro

safetyarxiv-cs-ai
26 May 2026
Safety

Music Transcription with (Almost) No Supervision

DGX agent

arXiv:2605.24193v1 Announce Type: cross Abstract: Competitive music transcription models require large amounts of paired audio-score data, which is scarce due to collection costs, alignment difficulty

safetyarxiv-cs-lg
26 May 2026
Safety

NeuralTouch: Neural Descriptors for Precise Sim-to-Real Tactile Robot Control

DGX agent

arXiv:2510.20390v2 Announce Type: replace Abstract: Grasping accuracy is a critical prerequisite for precise object manipulation, often requiring careful alignment between the robot hand and object. N

safetyarxiv-cs-ro
26 May 2026
Safety

Neuro-Inspired Inverse Learning for Planning and Control

DGX agent

arXiv:2605.24152v1 Announce Type: new Abstract: We present a neuro-inspired framework for embodied planning and control. Building on three principles that enable fast and highly effective goal-directe

safetyarxiv-cs-ai
26 May 2026
Safety

Not All Transitions Matter: Evidence from PPO

DGX agent

arXiv:2605.24071v1 Announce Type: cross Abstract: Training a reinforcement learning agent on-policy means collecting fresh experience at every update, and that experience comes with a hidden problem.

safetyarxiv-cs-ai
26 May 2026
Safety

Not only where, But when: Temporal Scheduling for RLVR

DGX agent

arXiv:2605.25381v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a core technique for post-training of Large Language Models (LLMs). While policy optimi

safetyarxiv-cs-lg
26 May 2026
Safety

OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation

DGX agent

arXiv:2605.25829v1 Announce Type: cross Abstract: Recent vision-language-action (VLA) models and world action models (WAMs) advance robotic manipulation by enriching intermediate representations with

safetyarxiv-cs-ai
26 May 2026
Safety

oh. my. god. could this word cloud diagram be … conscious?

DGX agent

Gary Marcus, a prominent AI researcher and critic, questions whether a word cloud diagram could possess consciousness, likely engaging in ironic commentary on overclaimed AI capabilities or consciousn

safetygary-marcus--x
26 May 2026
Safety

OmniSapiens: A Foundation Model for Social Behavior Processing via Heterogeneity-Aware Relative Policy Optimization

DGX agent

arXiv:2602.10635v2 Announce Type: replace Abstract: Socially intelligent AI systems must entail reasoning across diverse human behavioral tasks, and generalization to new contexts. However, AI has yet

safetyarxiv-cs-ai
26 May 2026
Safety

On Reliability of Efficient Membership Inference Vulnerability Evaluation

DGX agent

arXiv:2605.25819v1 Announce Type: new Abstract: Membership inference attacks (MIAs) are popular methods for empirically assessing the leakage of sensitive information in the training data through mode

safetyarxiv-cs-lg
26 May 2026
Safety

On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits

DGX agent

arXiv:2605.25789v1 Announce Type: cross Abstract: We study a stochastic multi-armed bandit problem where an agent is granted a free exploration budget before regret accumulates, a setting not captured

safetyarxiv-cs-ai
26 May 2026
Safety

On the Stability and Realizability of Recurrent Polynomial Surrogate Ternary Logic Gate Networks

DGX agent

arXiv:2605.24649v1 Announce Type: cross Abstract: Recurrent Neural Networks (RNNs) can learn to predict Signal Temporal Logic (STL) verdicts online from partial trajectories, but deploying them as run

safetyarxiv-cs-ai
26 May 2026
Safety

One-Step Bellman Alignment Enables Provably Efficient Transfer in Online RL

DGX agent

arXiv:2601.21924v2 Announce Type: replace Abstract: We study online transfer reinforcement learning (RL) in episodic Markov decision processes, where experience from related source tasks is available

safetyarxiv-cs-lg
26 May 2026
Safety

openai truly tanked a golden reputation, and squandered its lead, yet its board stands by its CEO, and looks for a band aid.

DGX agent

openai truly tanked a golden reputation, and squandered its lead, yet its board stands by its CEO, and looks for a band aid. OpenAI has a PR challenge, and although it's spoken to several candidates f

safetygary-marcus--x
26 May 2026
Safety

Operationalizing Reconstructive Authority: Runtime Construction, Dependency Resolution, and Execution Gating in Autonomous Agent Systems

DGX agent

arXiv:2605.23935v1 Announce Type: new Abstract: Autonomous agent systems fail not only due to incorrect decisions, but due to executing decisions whose authority no longer holds at runtime. Prior work

safetyarxiv-cs-ai
26 May 2026
Safety

Optimizing Token Choice for Code Watermarking: An RL Approach

DGX agent

arXiv:2508.11925v3 Announce Type: replace-cross Abstract: Protecting intellectual property on LLM-generated code necessitates effective watermarking systems that can operate within code's highly struc

safetyarxiv-cs-cl
26 May 2026
Safety

PageLLM: A Multi-Grained Reward Framework for Whole-Page Optimization with Large Language Models

DGX agent

arXiv:2506.09084v2 Announce Type: replace-cross Abstract: Whole-page optimization (WPO) decides how search and recommendation results are surfaced to users, and large language models (LLMs) open a new

safetyarxiv-cs-ai
26 May 2026
Safety

ParkingWorld: End-to-End Autonomous Parking Reinforcement Learning from Corrective Experience in 3DGS Simulation

DGX agent

arXiv:2605.25029v1 Announce Type: new Abstract: Autonomous parking demands precise low-speed maneuvering within narrow, cluttered, and highly constrained environments, where vehicles must navigate tig

safetyarxiv-cs-ro
26 May 2026
Safety

“pass me the crack pipe” @edels0n on SpaceX’s ludicrous valuation:

DGX agent

Gary Marcus critiques SpaceX's valuation as unreasonably inflated, using hyperbolic language to suggest the company's market assessment lacks rational justification. The post likely discusses concerns

safetygary-marcus--x
26 May 2026
Safety

PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs

DGX agent

arXiv:2601.20539v3 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled automated heuristic design (AHD) for combinatorial optimization problems (COPs), but existing frameworks'

safetyarxiv-cs-ai
26 May 2026
Safety

Peak-Then-Collapse and the Four Interface Channels of Knowledge-Graph Tool Use

DGX agent

arXiv:2605.26037v1 Announce Type: new Abstract: We test the standard RLVR tool-use recipe -- GRPO on Qwen2.5-7B-Instruct -- on a deliberately minimal knowledge-graph tool API: four Freebase navigation

safetyarxiv-cs-cl
26 May 2026
Safety

PID-Guided Partial Alignment for Multimodal Decentralized Federated Learning

DGX agent

arXiv:2601.10012v2 Announce Type: replace Abstract: Multimodal decentralized federated learning (DFL) must support collaboration among agents that hold different modality subsets and often different m

safetyarxiv-cs-lg
26 May 2026
Safety

PILOT: Policy-Informed Learned Optimization for Adaptive Deep Network Training

DGX agent

arXiv:2605.24570v1 Announce Type: cross Abstract: Despite the central role of optimization in deep learning, most optimizers rely on update structures whose functional form is fixed before training be

safetyarxiv-cs-ai
26 May 2026
Safety

PolyGnosis 2.0: Enhancing LLM Reasoning via Agentic Harness Engineering for Polymarket and OSINT Insight Extraction

DGX agent

arXiv:2605.25958v1 Announce Type: new Abstract: This paper introduces PolyGnosis 2.0, a pioneering multi-agent architecture designed to extract predictive intelligence by synthesizing Polymarket anoma

safetyarxiv-cs-cl
26 May 2026
← Previous
1…147148149150151…267
Next →