AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,498 results
Safety

Wow. This is worth watching, might have huge impact. @SenWarren makes some valid points, and the SEC owes her public answers

DGX agent

Wow. This is worth watching, might have huge impact. @SenWarren makes some valid points, and the SEC owes her public answers Sen. Warren calls on SEC to delay SpaceX IPO @CNBC https://www.cnbc.com/202

safetygary-marcus--x
10 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

YUBI: Yielding Universal Bidigital Interface for Bimanual Dexterous Manipulation at Scale

DGX agent

arXiv:2606.10244v1 Announce Type: cross Abstract: We introduce Yielding Universal Bidigital Interface (YUBI), a finger-aligned gripper designed to enable intuitive, ergonomic, and scalable data collec

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales

DGX agent

arXiv:2606.09470v1 Announce Type: cross Abstract: Automated L2 speech assessment can assign proficiency labels, but often lacks interpretability. We propose a rubric-guided SpeechLLM for multi-aspect,

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Geometric Unification of Concept Learning with Concept Cones

DGX agent

arXiv:2512.07355v2 Announce Type: replace Abstract: Two traditions of interpretability have evolved side by side but seldom spoken to each other: Concept Bottleneck Models (CBMs), which prescribe what

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Joint Finite-Sample Certificate for Adaptive Selective Conformal Risk Control

DGX agent

arXiv:2606.08517v1 Announce Type: new Abstract: Selective predictors answer on confident inputs and abstain elsewhere; deploying one safely needs a single finite-sample certificate that simultaneously

safetyarxiv-cs-lg
9 Jun 2026
Safety

A Mixed Diet Makes DINO An Omnivorous Vision Encoder

DGX agent

arXiv:2602.24181v2 Announce Type: replace-cross Abstract: Pre-trained vision encoders like DINOv2 have demonstrated exceptional performance on unimodal tasks. However, we observe that their features a

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Unifying Lens on Reward Uncertainty in RLHF

DGX agent

arXiv:2606.09073v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) is bottlenecked by reward hacking, where the policy exploits errors in a proxy reward model (RM) and

safetyarxiv-cs-ai
9 Jun 2026
Safety

Ablation-Reversible Heads Don't Transfer: A Stress Test for Mechanistic Role Claims in Transformers

DGX agent

arXiv:2606.08292v1 Announce Type: new Abstract: In mechanistic interpretability, attention heads are commonly elevated to role claims (e.g., 'this head represents addition') when they are necessary fo

safetyarxiv-cs-ai
9 Jun 2026
Safety

absolutely correct

DGX agent

absolutely correct @GaryMarcus I've recently learned not to believe anything that anyone says anymore when it comes to AI companies. They're all competing to win against each other while exploiting, s

safetygary-marcus--x
9 Jun 2026
Safety

ActProbe: Action-Space Probe for Early Failure Detection of Generative Robot Policies

DGX agent

arXiv:2606.08508v1 Announce Type: cross Abstract: Generative robot policies fail unpredictably at deployment: they hesitate at critical moments, drift off-task, or commit to unrecoverable actions. Exi

safetyarxiv-cs-ai
9 Jun 2026
Safety

Adaptive Loss Balancing for Noise-Robust GRPO in Generative Recommendation

DGX agent

arXiv:2606.08480v1 Announce Type: cross Abstract: Reinforcement learning (RL) presents a promising avenue for enhancing generative recommendation beyond supervised imitation, leveraging reward signals

safetyarxiv-cs-ai
9 Jun 2026
Safety

Agent Economics: An Entropy-Controlled Pluralistic Alignment Framework for Preventing Artificial Hivemind in Autonomous Agents

DGX agent

arXiv:2606.09039v1 Announce Type: new Abstract: This study proposes the Behavioral Protocol Framework (BPF), an entropy-controlled pluralistic alignment framework designed to address two critical chal

safetyarxiv-cs-ai
9 Jun 2026
Safety

AgriGov: A Structured Multilingual Dataset Curation for Indian Government Schemes for Farmers

DGX agent

arXiv:2606.08272v1 Announce Type: cross Abstract: AgriGov is a curated, trilingual (English-Hindi-Marathi) dataset designed to address the scarcity of domain-grounded multilingual resources for agricu

safetyarxiv-cs-ai
9 Jun 2026
Safety

AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing

DGX agent

arXiv:2606.09811v1 Announce Type: cross Abstract: World-action models have emerged as a promising paradigm for robot manipulation, jointly modeling visual scene dynamics and actions to inject physical

safetyarxiv-cs-ai
9 Jun 2026
Safety

AI Code Sandboxes: A Comparative Security Study. Part 1 of 2 -- Engine-Level Properties (Attack Surface, Leakage, Stackability, CVE History, Patch Cadence, Fuzzing)

DGX agent

arXiv:2606.08433v1 Announce Type: cross Abstract: This paper reads six engine-level measurements together -- 1.1 host attack surface, 1.2 information leakage, 1.3 defense-in-depth stackability, 1.4 pu

safetyarxiv-cs-ai
9 Jun 2026
Safety

AI-Integrated Learning Management System for Middle School: A Longitudinal Study of Learning Outcomes Through High School and Beyond

DGX agent

arXiv:2606.07544v1 Announce Type: cross Abstract: Middle school is a key window for building core academic skills and the learning routines students carry into later grades, yet many students still fa

safetyarxiv-cs-ai
9 Jun 2026
Safety

Aligned but Not Partner-Specific: Distinguishing How Multimodal LLM Agents Succeed in Reference Games Without Human-Like Conventions

DGX agent

arXiv:2606.08081v1 Announce Type: cross Abstract: Repeated reference games test whether interlocutors replace their initially long descriptions with shorter, partner-specific conventions grounded in s

safetyarxiv-cs-ai
9 Jun 2026
Safety

AMix-1: A Pathway to Test-Time Scalable Protein Foundation Model

DGX agent

arXiv:2507.08920v4 Announce Type: replace-cross Abstract: We introduce AMix-1, a powerful protein foundation model built on Bayesian Flow Networks and empowered by a systematic training methodology, e

safetyarxiv-cs-ai
9 Jun 2026
Safety

An Agency-Transferring Model-Free Policy Enhancement Technique

DGX agent

arXiv:2606.09825v1 Announce Type: cross Abstract: Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substan

safetyarxiv-cs-ai
9 Jun 2026
Safety

Anchor-Conditioned Compositional Control for Landscape Image Generation

DGX agent

arXiv:2606.07638v1 Announce Type: cross Abstract: Image generative models, though widely used as creative tools, offer limited support for the kind of compositional control that photographers and visu

safetyarxiv-cs-ai
9 Jun 2026
Safety

Anthropic didn’t just add guardrails to make Mythos safer; they added guardrails to protect their own IP. *Their own IP*. They are still as …

DGX agent

Anthropic didn’t just add guardrails to make Mythos safer; they added guardrails to protect their own IP. *Their own IP*. They are still as happy as fuck to build their AI on other people’s IP. Intere

safetygary-marcus--x
9 Jun 2026
Safety

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman…

DGX agent

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman) This is the scary phase of AI — a model deemed so powerful

safetygary-marcus--x
9 Jun 2026
Safety

Autonomous Aerial Manipulation via Contextual Contrastive Meta Reinforcement Learning

DGX agent

arXiv:2606.08533v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) are increasingly being deployed in logistics, service robotics, and other real-world applications, creating a growing de

safetyarxiv-cs-lg
9 Jun 2026
Safety

Autonomous FPV Flight with Translational Optical Flow and Uncertainty Mask

DGX agent

arXiv:2606.09088v1 Announce Type: new Abstract: Autonomous FPV quadrotor flight in complex environments using a monocular RGB camera as the sole exteroceptive sensor remains a fundamental challenge. R

safetyarxiv-cs-ro
9 Jun 2026
Safety

Autonomous Obstacle Removal for Excavators through Policy Learning with Particle Simulation

DGX agent

arXiv:2606.09183v1 Announce Type: new Abstract: Autonomous obstacle removal from the ground is an important earthwork task, but this is difficult to automate because an excavator must adapt its excava

safetyarxiv-cs-ro
9 Jun 2026
Safety

Back to the Familiar Future: Failure Recovery for VLA Policies via Pre-Imagined Milestone Selection

DGX agent

arXiv:2606.09258v1 Announce Type: new Abstract: Vision-language-action (VLA) policies can deviate from nominal trajectories during manipulation, even when tasks remain physically feasible. Recovering

safetyarxiv-cs-ro
9 Jun 2026
Safety

Bandits for Efficient Experimentation: Adapting to Control Group, Preferences, and Context Drifts

DGX agent

arXiv:2606.09802v1 Announce Type: cross Abstract: We consider a variant of the linear contextual stochastic multi-armed bandits, where the learner must provide recommendations to a group of users, eac

safetyarxiv-cs-ai
9 Jun 2026
Safety

BareWave: Waveform-Native Flow-Matching Text-to-Speech

DGX agent

arXiv:2606.09048v1 Announce Type: cross Abstract: Removing intermediate representations and separately trained decoding stages has become an important direction in generative modeling. In text-to-spee

safetyarxiv-cs-ai
9 Jun 2026
Safety

Beware of GeeksBearing Gifts: Building True EU Frontier AI Sovereignty

DGX agent

arXiv:2606.07536v1 Announce Type: cross Abstract: Frontier artificial intelligence is reshaping all aspects of society, from economic output or military capability to democratic institutions. The EU i

safetyarxiv-cs-ai
9 Jun 2026
Safety

Beyond Homophily: Towards Generalized Graph Reconstruction Attack and Defense

DGX agent

arXiv:2606.08067v1 Announce Type: new Abstract: Graph neural networks (GNNs) are widely deployed on relational data, yet they can leak sensitive or proprietary information about the training graph adj

safetyarxiv-cs-lg
9 Jun 2026
Safety

Beyond Neural Collapse: Task-Intrinsic Geometry Governs Neural Representations in Modular Arithmetic

DGX agent

arXiv:2606.08985v1 Announce Type: new Abstract: While neural collapse (NC) predicts that a K-class-balanced classifier should organize terminal representations as a (K-1)-dimensional simplex equiangul

safetyarxiv-cs-lg
9 Jun 2026
Safety

Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions

DGX agent

arXiv:2606.09076v1 Announce Type: new Abstract: Reward models are central to text-to-image post-training, but visual preference is subjective and better represented as a distribution over rubric score

safetyarxiv-cs-cv
9 Jun 2026
Safety

Boundary Variance Inflation Causes Acquisition Bias in Gaussian Processes

DGX agent

arXiv:2606.07561v1 Announce Type: new Abstract: Gaussian processes with stationary kernels on bounded domains exhibit inflated posterior variance near the boundary. Despite being a long-recognized art

safetyarxiv-cs-lg
9 Jun 2026
Safety

Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families

DGX agent

arXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp

safetyarxiv-cs-lg
9 Jun 2026
Safety

Bridging Expert Knowledge and Automated Feature Engineering via Self-Evolution

DGX agent

arXiv:2606.08800v1 Announce Type: new Abstract: In high-stakes settings such as brand compliance, clinical care, and content moderation, machine learning cannot be deployed as opaque oracles: practiti

safetyarxiv-cs-ai
9 Jun 2026
Safety

Bridging Traditional Explainability Methods and Multimodal Multilingual Models: An XAI-Based Analysis

DGX agent

arXiv:2606.07533v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) effectively integrate text and audio to interpret context in complex interactive dialogues. However, the inte

safetyarxiv-cs-ai
9 Jun 2026
Safety

Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs

DGX agent

arXiv:2606.09371v1 Announce Type: new Abstract: Tool learning enables LLMs to invoke external tools to accomplish tasks. Prior studies have demonstrated the effectiveness of a hierarchical structure:

safetyarxiv-cs-ai
9 Jun 2026
Safety

Causal Semantic Alignment for LLM-based Time Series Forecasting

DGX agent

arXiv:2606.08262v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new possibilities for time series forecasting by enabling alignment between temporal pattern

safetyarxiv-cs-lg
9 Jun 2026
Safety

Causal Transfer in Medical Image Analysis

DGX agent

arXiv:2603.24388v2 Announce Type: replace Abstract: Medical imaging models frequently fail when deployed across hospitals, scanners, populations, or imaging protocols due to domain shift, limiting the

safetyarxiv-cs-cv
9 Jun 2026
Safety

CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs

DGX agent

arXiv:2606.08420v1 Announce Type: new Abstract: Vision-language models (VLMs) pretrained on large-scale image-text pairs demonstrate strong image-level understanding, but are primarily optimized for g

safetyarxiv-cs-cv
9 Jun 2026
Safety

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation

DGX agent

arXiv:2606.09639v1 Announce Type: new Abstract: The fidelity and structural diversity of training datasets fundamentally determine the capabilities of video generation models. While commercial systems

safetyarxiv-cs-cv
9 Jun 2026
Safety

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

DGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

safetyarxiv-cs-lg
9 Jun 2026
Safety

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning

DGX agent

arXiv:2509.25004v2 Announce Type: replace Abstract: Online reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning abilities of large languag

safetyarxiv-cs-ai
9 Jun 2026
Safety

Cold feet about coating the entire surface of the earth in data centers?

DGX agent

Gary Marcus likely discusses concerns about the environmental and practical implications of exponentially expanding data center infrastructure across the globe, questioning whether covering Earth's su

safetygary-marcus--x
9 Jun 2026
Safety

Comparative evaluation of training strategies using partially labelled datasets for segmentation of white matter hyperintensities and stroke lesions in FLAIR MRI

DGX agent

arXiv:2601.20503v2 Announce Type: replace-cross Abstract: White matter hyperintensities (WMH) and ischaemic stroke lesions (ISL) are key imaging biomarkers of cerebral small vessel disease (SVD) detec

safetyarxiv-cs-ai
9 Jun 2026
Safety

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

DGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

safetyarxiv-cs-lg
9 Jun 2026
Safety

Constrained Paraphrase Consistency for LLM Hallucination Detection

DGX agent

arXiv:2606.08158v1 Announce Type: cross Abstract: Large language models (LLMs) can generate factually inconsistent claims, motivating accurate and scalable hallucination detectors. Prior work largely

safetyarxiv-cs-ai
9 Jun 2026
Safety

Constrained user-item allocation for e-commerce marketing campaigns

DGX agent

arXiv:2606.09623v1 Announce Type: new Abstract: When running marketing campaigns, retailers must decide which products to promote and which users to target. These decisions are inherently coupled: eff

safetyarxiv-cs-lg
9 Jun 2026
← Previous
1…143144145146147…303
Next →