AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

OpenAI’s market share is slipping.

DGX agent

OpenAI’s market share is slipping. ChatGPT is dying. OpenAI lost the crown as their market share fell below 40 percent for the first time ever. It stands at just 38.7 percent now after four straight m

safetygary-marcus--x
20 Apr 2026
Safety

Order: India's CCI sets May hearing on penalties after Apple failed to submit info on alleged app market abuse; Apple said it fears it could be fined up to $38B (Aditya Kalra/Reuters)

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Aditya Kalra / Reuters: Order: India's CCI sets May hearing on penalties after Apple failed to submit info on alleged app market abuse; Apple said it fears it could be fined up to $38B — Apple has not

safetytechmeme
20 Apr 2026
Safety

PAWN: Piece Value Analysis with Neural Networks

DGX agent

arXiv:2604.15585v1 Announce Type: cross Abstract: Predicting the relative value of any given chess piece in a position remains an open challenge, as a piece's contribution depends on its spatial relat

safetyarxiv-cs-ai
20 Apr 2026
Safety

Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning

DGX agent

arXiv:2506.20020v2 Announce Type: replace Abstract: Reasoning in humans is prone to biases due to underlying motivations like identity protection, that undermine rational decision-making and judgment.

safetyarxiv-cs-ai
20 Apr 2026
Safety

PLAF: Pixel-wise Language-Aligned Feature Extraction for Efficient 3D Scene Understanding

DGX agent

arXiv:2604.15770v1 Announce Type: new Abstract: Accurate open-vocabulary 3D scene understanding requires semantic representations that are both language-aligned and spatially precise at the pixel leve

safetyarxiv-cs-cv
20 Apr 2026
Safety

Preregistered Belief Revision Contracts

DGX agent

arXiv:2604.15558v1 Announce Type: new Abstract: Deliberative multi-agent systems allow agents to exchange messages and revise beliefs over time. While this interaction is meant to improve performance,

safetyarxiv-cs-ai
20 Apr 2026
Safety

Prototype-Grounded Concept Models for Verifiable Concept Alignment

DGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

safetyarxiv-cs-ai
20 Apr 2026
Safety

Puppets or partners? Governing cyborg propaganda in the digital public square

DGX agent

arXiv:2602.13088v2 Announce Type: replace-cross Abstract: The distinction between genuine grassroots activism and automated influence operations is collapsing. While contemporary policy debates priori

safetyarxiv-cs-ai
20 Apr 2026
Safety

Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability

DGX agent

arXiv:2604.16106v1 Announce Type: cross Abstract: The Project of AI is a world-building endeavor, wherein those who fund and develop AI systems both operate through and seek to sustain networks of pow

safetyarxiv-cs-ai
20 Apr 2026
Safety

Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning

DGX agent

arXiv:2510.10959v3 Announce Type: replace-cross Abstract: Reasoning ability has become a defining capability of Large Language Models (LLMs), with Reinforcement Learning with Verifiable Rewards (RLVR)

safetyarxiv-cs-ai
20 Apr 2026
Safety

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

DGX agent

arXiv:2604.15577v1 Announce Type: cross Abstract: Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vect

safetyarxiv-cs-ai
20 Apr 2026
Safety

Robust Multispectral Semantic Segmentation under Missing or Full Modalities via Structured Latent Projection

DGX agent

arXiv:2604.15856v1 Announce Type: cross Abstract: Multimodal remote sensing data provide complementary information for semantic segmentation, but in real-world deployments, some modalities may be unav

safetyarxiv-cs-ai
20 Apr 2026
Safety

Robust Synchronisation for Federated Learning in The Face of Correlated Device Failure

DGX agent

arXiv:2604.16090v1 Announce Type: cross Abstract: Probabilistic Synchronous Parallel (PSP) is a technique in distributed learning systems to reduce synchronization bottlenecks by sampling a subset of

safetyarxiv-cs-ai
20 Apr 2026
Safety

Safe and Energy-Aware Multi-Robot Density Control via PDE-Constrained Optimization for Long-Duration Autonomy

DGX agent

arXiv:2604.15524v1 Announce Type: cross Abstract: This paper presents a novel density control framework for multi-robot systems with spatial safety and energy sustainability guarantees. Stochastic rob

safetyarxiv-cs-ro
20 Apr 2026
Safety

Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility

DGX agent

arXiv:2604.16033v1 Announce Type: cross Abstract: Buildings account for approximately 40% of global energy consumption, and with the growing share of intermittent renewable energy sources, enabling de

safetyarxiv-cs-ai
20 Apr 2026
Safety

Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model

DGX agent

arXiv:2604.16111v1 Announce Type: new Abstract: We study the sample complexity of learning an epsilon-optimal policy in the Stochastic Shortest Path (SSP) problem. We first derive sample complexity bo

safetyarxiv-cs-lg
20 Apr 2026
Safety

Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agents

DGX agent

arXiv:2504.13541v5 Announce Type: replace-cross Abstract: Training resource-constrained autonomous agents on multiple tasks simultaneously is crucial for adapting to diverse real-world environments. R

safetyarxiv-cs-ai
20 Apr 2026
Safety

Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration

DGX agent

arXiv:2503.05578v4 Announce Type: replace Abstract: Pose estimation-guided unseen object 6-DoF robotic manipulation is a key task in robotics. However, the scalability of current pose estimation metho

safetyarxiv-cs-cv
20 Apr 2026
Safety

Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting

DGX agent

arXiv:2604.15794v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degra

safetyarxiv-cs-ai
20 Apr 2026
Safety

Seriously. If you don’t understand the below, you really shouldn’t speculate about AI. It’s absolutely fundamental.

DGX agent

Seriously. If you don’t understand the below, you really shouldn’t speculate about AI. It’s absolutely fundamental. literally my basic model since 1998. crazy that some people still haven’t figured th

safetygary-marcus--x
20 Apr 2026
Safety

SIMMER: Cross-Modal Food Image--Recipe Retrieval via MLLM-Based Embedding

DGX agent

arXiv:2604.15628v1 Announce Type: cross Abstract: Cross-modal retrieval between food images and recipe texts is an important task with applications in nutritional management, dietary logging, and cook

safetyarxiv-cs-cl
20 Apr 2026
Safety

Skill-RAG: Failure-State-Aware Retrieval Augmentation via Hidden-State Probing and Skill Routing

DGX agent

arXiv:2604.15771v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a foundational paradigm for grounding large language models in external knowledge. While adaptive re

safetyarxiv-cs-cl
20 Apr 2026
Safety

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

DGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

safetyarxiv-cs-ai
20 Apr 2026
Safety

Symbolic Guardrails for Domain-Specific Agents: Stronger Safety and Security Guarantees Without Sacrificing Utility

DGX agent

arXiv:2604.15579v1 Announce Type: cross Abstract: AI agents that interact with their environments through tools enable powerful applications, but in high-stakes business settings, unintended actions c

safetyarxiv-cs-ai
20 Apr 2026
Safety

'Taking Stock at FAccT': Using Participatory Design to Co-Create a Vision for the Fairness, Accountability and Transparency Community

DGX agent

arXiv:2604.16224v1 Announce Type: cross Abstract: As a relatively new forum, ACM FAccT has become a key space for activists and scholars to critically examine emerging AI and ML technologies. It bring

safetyarxiv-cs-ai
20 Apr 2026
Safety

Targeted Exploration via Unified Entropy Control for Reinforcement Learning

DGX agent

arXiv:2604.14646v2 Announce Type: replace Abstract: Recent advances in reinforcement learning (RL) have improved the reasoning capabilities of large language models (LLMs) and vision-language models (

safetyarxiv-cs-ai
20 Apr 2026
Safety

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

DGX agent

arXiv:2604.15383v1 Announce Type: cross Abstract: Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a temporal smoothing bias: transient

safetyarxiv-cs-ai
20 Apr 2026
Safety

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to…

DGX agent

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to the problem of how one could even control something vastly

safetyconnor-leahy--x
20 Apr 2026
Safety

The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback

DGX agent

arXiv:2604.16087v1 Announce Type: new Abstract: We study the problem of learning in zero-sum matrix games with repeated play and bandit feedback. Specifically, we focus on developing uncoupled algorit

safetyarxiv-cs-lg
20 Apr 2026
Safety

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning

DGX agent

arXiv:2603.01283v2 Announce Type: replace Abstract: Deployed RL agents operate in closed-loop systems where reliable performance depends on maintaining coherent coupling between observations, actions,

safetyarxiv-cs-ai
20 Apr 2026
Safety

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

safetyarxiv-cs-ai
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Safety

Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning

DGX agent

arXiv:2604.15705v1 Announce Type: new Abstract: Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with comple

safetyarxiv-cs-lg
20 Apr 2026
Safety

Trajectory Planning for Safe Dual Control with Active Exploration

DGX agent

arXiv:2604.15507v1 Announce Type: new Abstract: Planning safe trajectories under model uncertainty is a fundamental challenge. Robust planning ensures safety by considering worst-case realizations, ye

safetyarxiv-cs-ro
20 Apr 2026
Safety

Truncated Kernel Stochastic Gradient Descent with General Losses and Spherical Radial Basis Functions

DGX agent

arXiv:2510.04237v5 Announce Type: replace Abstract: In this paper, we propose a novel kernel stochastic gradient descent (SGD) algorithm for large-scale supervised learning with general losses. Compar

safetyarxiv-cs-lg
20 Apr 2026
Safety

Unsupervised domain adaptation for radioisotope identification in gamma spectroscopy

DGX agent

arXiv:2603.05719v2 Announce Type: replace Abstract: Training machine learning models for radioisotope identification using gamma spectroscopy remains an elusive challenge for many practical applicatio

safetyarxiv-cs-lg
20 Apr 2026
Safety

UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval

DGX agent

arXiv:2604.15827v1 Announce Type: cross Abstract: Conventional information retrieval is concerned with identifying the relevance of texts for a given query. Yet, the conventional definition of relevan

safetyarxiv-cs-cl
20 Apr 2026
Safety

VADF: Vision-Adaptive Diffusion Policy Framework for Efficient Robotic Manipulation

DGX agent

arXiv:2604.15938v1 Announce Type: new Abstract: Diffusion policies are becoming mainstream in robotic manipulation but suffer from hard negative class imbalance due to uniform sampling and lack of sam

safetyarxiv-cs-ro
20 Apr 2026
Safety

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

DGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

safetyarxiv-cs-lg
20 Apr 2026
Safety

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models

DGX agent

arXiv:2510.09689v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been augmented with web search to overcome the limitations of the static knowledge boundary by accessing up-

safetyarxiv-cs-ai
20 Apr 2026
Safety

Whose Facts Win? LLM Source Preferences under Knowledge Conflicts

DGX agent

arXiv:2601.03746v3 Announce Type: replace Abstract: As large language models (LLMs) are more frequently used in retrieval-augmented generation pipelines, it is increasingly relevant to study their beh

safetyarxiv-cs-cl
20 Apr 2026
Safety

Why Colors Make Clustering Harder:Global Integrality Gaps, the Price of Fairness, and Color-Coupled Algorithms in Chromatic Correlation Clustering

DGX agent

arXiv:2604.15738v1 Announce Type: new Abstract: Chromatic Correlation Clustering (CCC) extends Correlation Clustering by assigning semantic colors to edges and requiring each cluster to receive a sing

safetyarxiv-cs-lg
20 Apr 2026
Safety

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

DGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

safetyarxiv-cs-cl
20 Apr 2026
Safety

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

DGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

safetyarxiv-cs-lg
20 Apr 2026
Safety

good to know that a trillion dollars’ investment scaling has thoroughly solved the challenges of common sense.

DGX agent

Gary Marcus critiques the assumption that massive financial investment in AI scaling alone can solve fundamental challenges related to common sense reasoning in artificial intelligence systems. The po

safetygary-marcus--x
19 Apr 2026
Safety

Tesla expands its robotaxi service to Dallas and Houston after launching in Austin last year and starting to offer rides without safety drivers in January 2026 (Anthony Ha/TechCrunch)

DGX agent

Anthony Ha / TechCrunch: Tesla expands its robotaxi service to Dallas and Houston after launching in Austin last year and starting to offer rides without safety drivers in January 2026 — Tesla is expa

safetytechmeme
19 Apr 2026
Safety

Dario is not as different from Sam as people seem to think.

DGX agent

Gary Marcus argues that Dario Amodei (Anthropic CEO) and Sam Altman (OpenAI CEO) share more similarities in their approaches to AI development and safety than commonly perceived, despite their public

safetygary-marcus--x
18 Apr 2026
Safety

easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P]

DGX agent

easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization

safetyr-machinelearning
18 Apr 2026
← Previous
1…240241242243244…265
Next →