AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

Altman’s AI safety proposal: bail me out as we badly missed our revenue runway, or i will not be a multi-billionaire

DGX agent

Altman’s AI safety proposal: bail me out as we badly missed our revenue runway, or i will not be a multi-billionaire Altman’s AI safety proposal: let us win, or everybody loses https://ft.trib.al/UrDI

safetygary-marcus--x
2 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

ASPIRE: Agentic /Skills Discovery for Robotics

DGX agent

arXiv:2607.00272v1 Announce Type: cross Abstract: Traditional robot programming is challenging: it requires orchestrating multimodal perception, managing physical contact dynamics, and handling divers

safetyarxiv-cs-ai
2 Jul 2026
Safety

Attribute-Prompted Kernel Hashing for Unsupervised Data-Efficient Cross-Modal Retrieval

DGX agent

arXiv:2607.00379v1 Announce Type: cross Abstract: Unsupervised cross-modal hashing enables efficient retrieval of semantically related instances across different modalities without requiring manual se

safetyarxiv-cs-cv
2 Jul 2026
Safety

AutoSpeed: Annotation-Free Stage-Adaptive Motion Speed Learning for Robot Manipulation

DGX agent

arXiv:2607.01051v1 Announce Type: new Abstract: Different stages of manipulation tasks exhibit varying levels of difficulty, suggesting stage-dependent motion speeds and temporal prediction horizons.

safetyarxiv-cs-ro
2 Jul 2026
Safety

Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces

DGX agent

arXiv:2607.00481v1 Announce Type: cross Abstract: Jailbreak attacks remain a critical threat to the safe deployment of large language models (LLMs). While prior work has primarily studied attacks and

safetyarxiv-cs-ai
2 Jul 2026
Safety

Bounded Morality: Defining the Space of Moral Computation

DGX agent

arXiv:2607.00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as stati

safetyarxiv-cs-ai
2 Jul 2026
Safety

BrainFIBRE: A Foundation Model via Information Decomposition for Brain Microstructure

DGX agent

arXiv:2607.00573v1 Announce Type: new Abstract: Diffusion MRI probes brain microstructure with particular sensitivity to early cerebrovascular and neurodegenerative changes. Neurite Orientation Disper

safetyarxiv-cs-cv
2 Jul 2026
Safety

Caption Bottleneck Models

DGX agent

arXiv:2607.00578v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide interpretability by routing predictions through a layer of human-understandable concepts. However, defining an

safetyarxiv-cs-cv
2 Jul 2026
Safety

ClinRAG-GRAPH: Clinical-prior Retrieval-Augmented Graph Model with Domain Adversarial Learning for Breast pCR Prediction

DGX agent

arXiv:2607.00798v1 Announce Type: new Abstract: Neoadjuvant chemotherapy (NAC) response prediction is clinically important for treatment stratification in breast cancer. However, robust pre-treatment

safetyarxiv-cs-cv
2 Jul 2026
Safety

congrats @yudapearl!

DGX agent

congrats @yudapearl! Judea Pearl Named AI Pioneer by Boston Global Forum in Honor of America’s 250th Anniversary https://samueli.ucla.edu/judea-pearl-named-ai-pioneer-by-boston-global-forum-in-honor-o

safetygary-marcus--x
2 Jul 2026
Safety

'consensus' can't just built by big tech! let's all be careful of regulatory capture and extreme concentration of power!

DGX agent

'consensus' can't just built by big tech! let's all be careful of regulatory capture and extreme concentration of power! NEW: Anthropic announces it is drafting a consensus framework with Amazon, Micr

safetygary-marcus--x
2 Jul 2026
Safety

Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction

DGX agent

arXiv:2607.00001v1 Announce Type: new Abstract: Most approaches to AI alignment treat human preferences as fixed targets to be inferred and optimized. This assumption conflicts with extensive empirica

safetyarxiv-cs-ai
2 Jul 2026
Safety

Continuous Speculative Decoding for Autoregressive Image Generation

DGX agent

arXiv:2411.11925v3 Announce Type: replace Abstract: Continuous visual autoregressive (AR) models have demonstrated promising performance in image generation, but their inherently sequential nature res

safetyarxiv-cs-cv
2 Jul 2026
Safety

Dataset Biases and Shortcut Learning in Motion-Based AI-Generated Video Detection

DGX agent

arXiv:2607.00948v1 Announce Type: new Abstract: The visual quality of AI-generated videos has improved drastically in recent years, making it increasingly difficult for humans to distinguish between r

safetyarxiv-cs-cv
2 Jul 2026
Safety

Decentralized Geometric Control for Cable-Suspended Payload Transport with Adaptive Mass Estimation

DGX agent

arXiv:2607.00024v1 Announce Type: new Abstract: Cooperative aerial transport requires controllers that respect nonlinear manifold geometry, operate without centralized coordination, and respect operat

safetyarxiv-cs-ro
2 Jul 2026
Safety

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

DGX agent

arXiv:2607.01170v1 Announce Type: cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they ar

safetyarxiv-cs-ai
2 Jul 2026
Safety

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation

DGX agent

arXiv:2607.01208v1 Announce Type: cross Abstract: Language models deployed in high-stakes roles can potentially favor certain entities, brands, or viewpoints, steering user decisions at scale. Such pr

safetyarxiv-cs-ai
2 Jul 2026
Safety

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

DGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

safetyarxiv-cs-ro
2 Jul 2026
Safety

Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts

DGX agent

arXiv:2607.00666v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models often fail to perform the same learned tasks under environmental shifts, such as changes in camera pose and shifts

safetyarxiv-cs-cv
2 Jul 2026
Safety

Dual-Informed Vertical Expansion for Multi-Objective Node Selection in Anytime Conflict-Based Search

DGX agent

arXiv:2607.00156v1 Announce Type: new Abstract: Conflict-Based Search (CBS) is a leading exact algorithm for Multi-Agent Path Finding (MAPF), but its high-level node-selection rule is usually treated

safetyarxiv-cs-ro
2 Jul 2026
Safety

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

DGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

safetyarxiv-cs-cv
2 Jul 2026
Safety

ELMP: Efficient Learning for Motion Planning via Analytical Policy Gradients

DGX agent

arXiv:2607.00215v1 Announce Type: new Abstract: Neural Motion Planners (NMPs) enable fast reactive motion generation, but adapting them to new environments typically requires recollecting large expert

safetyarxiv-cs-ro
2 Jul 2026
Safety

Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured (Wall Street Journal)

DGX agent

Wall Street Journal: Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured — Undersecretary E

safetytechmeme
2 Jul 2026
Safety

Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms

DGX agent

arXiv:2407.15283v2 Announce Type: replace-cross Abstract: Industry is moving toward autonomous, network-connected machines that detect and adapt to changing conditions, including hardware faults. Conv

safetyarxiv-cs-ai
2 Jul 2026
Safety

EPO: Boosting 3D Foundation Models with Edge-based Pose Optimization

DGX agent

arXiv:2607.00579v1 Announce Type: new Abstract: We introduce extbf{Edge-based Pose Optimization (EPO)}, a trackless geometric optimization framework specifically designed to boost the Structure-from-M

safetyarxiv-cs-cv
2 Jul 2026
Safety

EquiSteer: Cross-Attention Steering Towards a Fairer Text-Guided Image Generation

DGX agent

arXiv:2607.01147v1 Announce Type: new Abstract: Text-to-image diffusion models power everyday creative tasks, but they still reproduce the demographic biases in their training data. On common prompts

safetyarxiv-cs-cv
2 Jul 2026
Safety

Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles

DGX agent

arXiv:2511.06160v2 Announce Type: replace Abstract: While recent safety guardrails effectively suppress overtly biased outputs, subtler forms of social bias emerge during complex logical reasoning tas

safetyarxiv-cs-ai
2 Jul 2026
Safety

Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows

DGX agent

arXiv:2607.00828v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate queries, invoke tools, and construct analytical workflows. Although recent advances hav

safetyarxiv-cs-ai
2 Jul 2026
Safety

FAR: Failure-Aware Retry for Test-Time Recovery and Continual Policy Improvement

DGX agent

arXiv:2607.01111v1 Announce Type: cross Abstract: Robot policies inevitably encounter failures when deployed in real environments. Naive retries often repeat the same mistakes, while many existing rec

safetyarxiv-cs-ai
2 Jul 2026
Safety

FastBridge: Closing the Model-Based Realization Gap in Safety Filters on 3D Gaussian Splatting for Fast Quadrotor Flight

DGX agent

arXiv:2607.01200v1 Announce Type: new Abstract: Fast quadrotor flight requires safe obstacle avoidance under tight onboard compute limits. While 3D Gaussian Splatting (3DGS) provides a continuous, geo

safetyarxiv-cs-ro
2 Jul 2026
Safety

ForAug: Mitigating Biases in Image Classification via Controlled Image Compositions

DGX agent

arXiv:2503.09399v4 Announce Type: replace-cross Abstract: Large-scale image classification datasets exhibit strong compositional biases: objects tend to be centered, appear at characteristic scales, a

safetyarxiv-cs-ai
2 Jul 2026
Safety

FrameONE: Hierarchical Motion Modeling for Universal Multi-View Echocardiographic Keyframe Detection

DGX agent

arXiv:2607.00748v1 Announce Type: new Abstract: Accurate detection of end-systole (ES) and end-diastole (ED) frames is fundamental to echocardiographic assessment. Existing methods are typically devel

safetyarxiv-cs-cv
2 Jul 2026
Safety

From Holistic Evaluation to Structured Criteria: Rubrics Across the Evolving LLM Landscape

DGX agent

arXiv:2606.08625v2 Announce Type: replace Abstract: As Large Language Models (LLMs) advance toward open-ended autonomous agents, the mechanisms used to evaluate and guide their behavior must evolve ac

safetyarxiv-cs-cl
2 Jul 2026
Safety

From Pixels to Temporal Correlations: Learning Informative Representations for Reinforcement Learning Pre-training

DGX agent

arXiv:2607.00811v1 Announce Type: new Abstract: Unsupervised pre-training on large-scale datasets has demonstrated significant potential for improving the sample efficiency and performance of Reinforc

safetyarxiv-cs-lg
2 Jul 2026
Safety

From Prediction Uncertainty to Conformalized Distance Fields for Safe Motion Planning

DGX agent

arXiv:2607.00776v1 Announce Type: new Abstract: Safe motion planning in dynamic environments requires reasoning about the uncertainty in predicted obstacle motion without sacrificing real-time perform

safetyarxiv-cs-ro
2 Jul 2026
Safety

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning

DGX agent

arXiv:2603.10263v2 Announce Type: replace-cross Abstract: We introduce Distribution Contractive Reinforcement Learning (DICE-RL), a framework that uses reinforcement learning (RL) as a 'distribution c

safetyarxiv-cs-lg
2 Jul 2026
Safety

From Silos to Systems: Process-Oriented Hazard Analysis for AI Systems

DGX agent

arXiv:2410.22526v2 Announce Type: replace Abstract: To effectively address potential harms from Artificial Intelligence (AI) systems, it is essential to identify and mitigate system-level hazards. Cur

safetyarxiv-cs-ai
2 Jul 2026
Safety

From World Models to World Action Models: A Concise Tutorial for Robotics

DGX agent

arXiv:2607.00836v1 Announce Type: cross Abstract: World models are increasingly used in embodied intelligence and generative simulation, yet their scope remains ambiguous across communities. This tuto

safetyarxiv-cs-ai
2 Jul 2026
Safety

Gauging, Measuring, and Controlling Critic Complexity in Actor-Critic Reinforcement Learning

DGX agent

arXiv:2607.00452v1 Announce Type: cross Abstract: Actor-critic methods depend on learned critics, but critic quality is often evaluated only indirectly through return, temporal-difference error, or va

safetyarxiv-cs-ai
2 Jul 2026
Safety

GaussianFusion: Unified 3D Gaussian Representation for Multi-Modal Fusion Perception

DGX agent

arXiv:2607.00746v1 Announce Type: cross Abstract: The bird's-eye view (BEV) representation enables multi-sensor features to be fused within a unified space, serving as the primary approach for achievi

safetyarxiv-cs-ai
2 Jul 2026
Safety

Google’s Continued Disruption of Malicious Residential Proxy Networks

DGX agent

Background Today, in coordination with the FBI, Lumen, and others, Google took action against the NetNut residential proxy network, also known as Popa. This action builds on our disruption of the IPID

safetygoogle-cloud-ai
2 Jul 2026
Safety

Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination

DGX agent

arXiv:2607.00924v1 Announce Type: new Abstract: Accelerating materials discovery requires AI systems that can generate scientifically valid hypotheses through multi-step, domain-grounded reasoning. St

safetyarxiv-cs-ai
2 Jul 2026
Safety

GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity

DGX agent

arXiv:2607.00152v1 Announce Type: cross Abstract: Three of the most popular methods for training language models to reason look like three different tricks. They are not. All three adjust a single num

safetyarxiv-cs-ai
2 Jul 2026
Safety

HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment

DGX agent

arXiv:2607.00572v1 Announce Type: new Abstract: Understanding how aligned LLMs internally represent safety is critical for diagnosing alignment vulnerabilities, as it explains why jailbreaks succeed a

safetyarxiv-cs-ai
2 Jul 2026
Safety

How Environment and Urbanization Shape Bird Diversity in Sri Lanka

DGX agent

arXiv:2607.00582v1 Announce Type: cross Abstract: This study presents a comprehensive analysis of bird diversity across Sri Lanka by integrating spatial, temporal, and environmental data. Bird observa

safetyarxiv-cs-lg
2 Jul 2026
Safety

HyFL-CLIP: Hyperbolic Fine-Tuning of CLIP for Robust Long-Context Understanding

DGX agent

arXiv:2607.00428v1 Announce Type: new Abstract: CLIP (Contrastive Language-Image Pre-training) has become a de facto paradigm for image-text alignment, but it struggles with long-context descriptions

safetyarxiv-cs-cv
2 Jul 2026
Safety

Identifying Latent Concepts and Structures for Generalized Category Discovery

DGX agent

arXiv:2607.00620v1 Announce Type: cross Abstract: Generalized Category Discovery (GCD) aims to recognize known classes while autonomously discovering novel ones in open-world settings. However, curren

safetyarxiv-cs-ai
2 Jul 2026
Safety

in case you needed more evidence that tokenmaxxing was a stupid idea

DGX agent

in case you needed more evidence that tokenmaxxing was a stupid idea Meta burns 2.65B a year on AI tokens. at 300K for a Meta engineer, that's enough to pay ~9,000 engineers for a full year. now ask y

safetygary-marcus--x
2 Jul 2026
← Previous
1…6061626364…265
Next →