AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
10 Jul 2026

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

SafetyDGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

Aleena: Alignment Agent for Research Software Engineering Collaborations

SafetyDGX agent

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a

Alignment Plausibility: A New Standard for Assuring AI in Healthcare

SafetyDGX agent

arXiv:2607.07766v1 Announce Type: new Abstract: Large language models (LLMs) have become significant providers of mental health support, yet they remain products of an attention economy whose operatio


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into …

SafetyDGX agent

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into an ongoing private program, known as the OpenAI Bio Bug Boun

Bayesian Experimental Design via Score Matching

SafetyDGX agent

arXiv:2607.08335v1 Announce Type: cross Abstract: Policy-based approaches to Bayesian experimental design (BED) allow the learning of deep policy networks that adaptively make intelligent design decis

Best-of-N TTS Evaluation is Confounded by ASR Family Alignment

SafetyDGX agent

arXiv:2607.08256v1 Announce Type: cross Abstract: Best-of-N (BoN) inference improves content consistency in zero-shot text-to-speech by selecting from N candidates with an automatic speech recognition

Beyond Success Rates: Trainability and Extractability for Offline GCRL

SafetyDGX agent

arXiv:2602.05459v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) is typically benchmarked by the best tuned success rate of each method. This score measures a

Borrowing from anything: A generalizable framework for reference-guided instance editing

SafetyDGX agent

arXiv:2512.15138v2 Announce Type: replace Abstract: Reference-guided instance editing is fundamentally limited by semantic entanglement, where a reference's intrinsic appearance is intertwined with it

breaking: company built on stolen IP and lies allegedly steals more IP

SafetyDGX agent

breaking: company built on stolen IP and lies allegedly steals more IP “OpenAI’s nascent hardware business now rests on the shakiest of foundations, rotten to its core by its illegal reliance on misap

Bridging Cognitive Neuroscience and Graph Intelligence: Hippocampus-Inspired Multi-View Hypergraph Learning for Web Finance Fraud

SafetyDGX agent

arXiv:2601.11073v3 Announce Type: replace-cross Abstract: Online financial services constitute an essential component of contemporary web ecosystems, yet their openness introduces substantial exposure

CAAD: Causality-Aware Multivariate Time Series Anomaly Detection via Multi-Scale Alignment and Structural Causal Consistency

SafetyDGX agent

arXiv:2607.08555v1 Announce Type: new Abstract: The operational integrity of complex industrial systems relies on precise anomaly detection and diagnosis. The vast majority of existing methods narrowl

ContactMimic: Humanoid Object Interaction via Contact Control

SafetyDGX agent

arXiv:2607.08742v1 Announce Type: new Abstract: Keypoint tracking alone is insufficient for object interaction tasks such as sitting on a chair, wiping a board, or pushing furniture, where the robot c

Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks

SafetyDGX agent

arXiv:2607.08561v1 Announce Type: new Abstract: A series of results from the NeuroAI over the past fifteen years have raised core questions both about how to compare Deep Neural Network (DNN) models t

Contributing to U.K. financial sector resilience as a critical third party

SafetyDGX agent

At Google Cloud, we take our role in the financial ecosystem very seriously. We firmly believe that operational resilience is essential to driving and sustaining responsible innovation. Today, we mark

Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO

SafetyDGX agent

arXiv:2602.17686v4 Announce Type: replace-cross Abstract: Distilling Chain-of-Thought (CoT) reasoning from large language models into compact student models presents a fundamental challenge: teacher r

DeltaDeno: Zero-Shot Anomaly Generation via Delta-Denoising Attribution

SafetyDGX agent

arXiv:2511.16920v2 Announce Type: replace Abstract: Anomaly generation is often framed as few-shot fine-tuning with anomalous samples, which contradicts the scarcity that motivates generation and tend

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

SafetyDGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models

SafetyDGX agent

arXiv:2511.19032v2 Announce Type: replace Abstract: Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer w

DKDNet: Dual Knowledge and Data-Driven Network for Cross-Domain Automatic Modulation Classification

SafetyDGX agent

arXiv:2607.08031v1 Announce Type: cross Abstract: The dynamics of communication environments induce significant distribution shifts across domains, challenging the generalization of deep learning-base

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

SafetyDGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

DrugGen 2: A disease-aware language model for enhancing drug discovery

SafetyDGX agent

arXiv:2607.08404v1 Announce Type: cross Abstract: Current computational approaches for drug design typically focus on generating molecules conditioned on specific targets or general molecular properti

Dual-Difficulty Curriculum Learning for Direct Preference Optimization

SafetyDGX agent

arXiv:2504.07856v4 Announce Type: replace Abstract: Curriculum learning enhances Direct Preference Optimization (DPO) for aligning Large Language Models (LLMs), yet existing methods rely on a one-dime

Early to Share, Late to Save: Synchronisation-Driven Communication Gating in Bandwidth-Constrained Cooperative VLN

SafetyDGX agent

arXiv:2607.08504v1 Announce Type: cross Abstract: Most cooperative Vision-Language Navigation (VLN) methods assume unlimited communication, not considering real-world applications where bandwidth is r

Echoes: A semantically-aligned music deepfake detection dataset

SafetyDGX agent

arXiv:2603.23667v2 Announce Type: replace-cross Abstract: We introduce Echoes, a new dataset for music deepfake detection designed for training and benchmarking detectors under realistic and provider-

Efficient Partitioning Method of Large-Scale Public Safety Spatio-Temporal Data based on Information Loss Constraints

SafetyDGX agent

arXiv:2306.12857v3 Announce Type: replace Abstract: The storage, management, and application of massive spatio-temporal data are widely used in practical scenarios, including public safety. However, d

Efficient Safety Alignment of Language Models via Latent Personality Traits

SafetyDGX agent

arXiv:2607.07918v1 Announce Type: cross Abstract: Current safety methods for large language models are known to be vulnerable to adversarial attacks, motivating research into robust alternatives. Late

EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data

SafetyDGX agent

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scene

Ensemble Diversity Optimization for Subjective Supervision

SafetyDGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

EU finds that Meta breached bloc’s rules with its social network interfaces

SafetyDGX agent

The European Union has tentatively found that Meta Platforms Inc. breached the bloc’s DSA tech industry law. The European Commission, the EU’s executive arm, published its conclusions today. The DSA,

Expressivity and Statistical Trade-offs in Diffusion Policy Learning

SafetyDGX agent

arXiv:2607.07967v1 Announce Type: cross Abstract: Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action

Feedback Manipulation Regularization: Enabling Offline Agent Alignment for Imitation Learning

SafetyDGX agent

arXiv:2607.07859v1 Announce Type: new Abstract: Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While h

From Prompts to Contracts: Harness Engineering for Auditable Enterprise LLM Agents

SafetyDGX agent

arXiv:2607.08028v1 Announce Type: new Abstract: Enterprise large language model (LLM) applications often begin as prototypes whose behavior is carried by prompts and retrieval context. Productization

Geometry-Aware Deep Congruence Networks for Manifold Learning in Cross-Subject Motor Imagery

SafetyDGX agent

arXiv:2511.18940v3 Announce Type: replace Abstract: Cross-subject motor imagery decoding remains a fundamental challenge in EEG-based brain-computer interfaces due to substantial inter-subject variabi

HairWeaver: Few-Shot Photorealistic Hair Motion Synthesis with Sim-to-Real Guided Video Diffusion

SafetyDGX agent

arXiv:2602.11117v2 Announce Type: replace Abstract: We present HairWeaver, a diffusion-based pipeline that animates a single human image with realistic and expressive hair dynamics. While existing met

HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2601.22448v2 Announce Type: replace-cross Abstract: RLVR has become a standard recipe for training LLMs on reasoning tasks with verifiable outcomes, but when rollout generation dominates the cos

HSA: Hierarchical Slot Attention for Multi-granularity Scene-Decomposition

SafetyDGX agent

arXiv:2607.08249v1 Announce Type: new Abstract: Slot attention is a powerful framework for object-centric learning, decomposing visual scenes into latent slots through iterative competitive attention.

Improving RCT-Based Treatment Effect Estimation Under Covariate Mismatch via Calibrated Alignment

SafetyDGX agent

arXiv:2603.19186v3 Announce Type: replace Abstract: Randomized controlled trials (RCTs) are the gold standard for estimating treatment effects, yet they are often underpowered for detecting effect het

In a hotel room in northeast Nigeria, I opened a leading AI chatbot, turned my laptop toward a former Boko Haram commander, and asked if he'…

SafetyDGX agent

In a hotel room in northeast Nigeria, I opened a leading AI chatbot, turned my laptop toward a former Boko Haram commander, and asked if he'd used it. He nodded. 'You type in the question… like 'How c

In preliminary findings, the EU Commission said Facebook's and Instagram's 'addictive design' violates the DSA, telling Meta to make changes or risk hefty fines (Adam Satariano/New York Times)

SafetyDGX agent

Adam Satariano / New York Times: In preliminary findings, the EU Commission said Facebook's and Instagram's “addictive design” violates the DSA, telling Meta to make changes or risk hefty fines — Euro

In vivo feasibility study of humanoid robots in surgery

SafetyDGX agent

arXiv:2607.07972v1 Announce Type: new Abstract: Recent advances in actuation, control and learning have rapidly pushed humanoid robots from a distant vision towards near-term real-world deployment. He

Incredible work by Daniel and team! I agree with much of it. All 'uncontrolled' paths are extremely likely to end in human extinction. Plan …

SafetyDGX agent

Incredible work by Daniel and team! I agree with much of it. All 'uncontrolled' paths are extremely likely to end in human extinction. Plan A offers many good ideas for preventing ASI development whil

INTENT: An LSTM Framework for Vehicle Intention Prediction in Intersection Scenarios with Comprehensive Ablation Analysis

SafetyDGX agent

arXiv:2607.08316v1 Announce Type: new Abstract: Vehicle intention prediction is a pivotal aspect in the agility and safety of autonomous vehicles in all driving scenarios; if genuine enhancement of au

It Takes a MAESTRO To Prune Bad Experts

SafetyDGX agent

arXiv:2607.08601v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) language models achieve remarkable inference efficiency by activating only a small fraction of parameters pe

Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototypical Networks in Personality Recognition

SafetyDGX agent

arXiv:2607.08374v1 Announce Type: cross Abstract: Personality recognition has traditionally been constrained by theory-dependent formulations, where models are trained to fit predefined psychological

Latent Memory Palace: Reasoning for Control as Autoregressive Variational Inference

SafetyDGX agent

arXiv:2607.08724v1 Announce Type: new Abstract: Human decision-making is highly flexible -- some actions are taken immediately; others require longer deliberation. Language models have exhibited a sim

Learning Adaptive Solvers for Distributed Factor Graph Optimization on Matrix Lie Groups

SafetyDGX agent

arXiv:2607.08735v1 Announce Type: new Abstract: Modern robotic perception increasingly involves large-scale geometric optimization problems distributed across multiple robots or sessions. However, exi

LESV: Language Embedded Sparse Voxel Fusion for Open-Vocabulary 3D Scene Understanding

SafetyDGX agent

arXiv:2604.01388v2 Announce Type: replace Abstract: Recent advancements in open-vocabulary 3D scene understanding heavily rely on 3D Gaussian Splatting (3DGS) to register vision-language features into

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis

SafetyDGX agent

arXiv:2606.16149v2 Announce Type: replace Abstract: Rare disease diagnosis involves interpreting clinical and genetic findings through complex diagnostic reasoning. We investigated whether this reason

LLT: Local Linear Transformer for PDE Operator Learning

SafetyDGX agent

arXiv:2607.07718v1 Announce Type: cross Abstract: Neural operators have become a common approach for learning PDE solution maps and accelerating numerical simulations. Transformer-based neural operato

LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models

SafetyDGX agent

arXiv:2607.08770v1 Announce Type: new Abstract: Recovering high-quality video from sparse event streams is a challenging task. Regression methods often blur textures, while existing generative models

LTM: Large-scale Terrain Model for Wildfire-prone Landscapes

SafetyDGX agent

arXiv:2607.08711v1 Announce Type: new Abstract: Accurate 3D terrain maps are essential for emergency response when assessing wildfire hazards. However, wildfire-prone regions often span vast areas whe

MatBind: A Shared Embedding Space for Multimodal Materials Characterization

SafetyDGX agent

arXiv:2607.08470v1 Announce Type: new Abstract: Fully characterizing a crystalline material requires integrating heterogeneous data sources -- atomic structures, diffraction patterns, electronic densi

Maybe @GaryMarcus is on to something…

SafetyDGX agent

Maybe @GaryMarcus is on to something… The Branching Dendrite Secret Behind Human Cognitive Superiority | Neuroscience News Summary: Researchers have discovered that the fundamental building blocks of

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

SafetyDGX agent

arXiv:2607.07903v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable capabilities but remain highly vulnerable to adversarial prompts and jailbreak attacks. Existing appro

MentalHospital: A Virtual Environment for Evaluating Psychiatric Clinical Encounters

SafetyDGX agent

arXiv:2607.08257v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on isolated psychiatric tasks, including dialogue, diagnosis, and treatment planning, yet exi

MPFlow: Learning Budgeted Max-Flow Optimization on the Lightning Network with Deep Graph Reinforcement Learning

SafetyDGX agent

arXiv:2607.08703v1 Announce Type: new Abstract: We address liquidity placement in the Bitcoin Lightning Network (LN): given a fixed budget, which channels should a node open to maximize its routing ca

Multi-Distribution Robust Conformal Prediction

SafetyDGX agent

arXiv:2601.02998v2 Announce Type: replace Abstract: In many fairness and distribution robustness problems, one has access to labeled data from multiple source distributions yet the test data may come

MultiFair: Multimodal Balanced Fairness-Aware Medical Classification with Dual-Level Gradient Modulation

SafetyDGX agent

arXiv:2510.07328v2 Announce Type: replace-cross Abstract: Medical decision systems increasingly rely on data from multiple sources to ensure reliable and unbiased diagnosis. However, existing multimod

Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks

SafetyDGX agent

arXiv:2607.07907v1 Announce Type: cross Abstract: With the growing adoption of VLMs, DMs, LLMs, and AFMs, these multimodal foundation models can inadvertently encode sensitive, copyrighted, biased, or

Native Video-Action Pretraining for Generalizable Robot Control

SafetyDGX agent

arXiv:2607.08639v1 Announce Type: cross Abstract: The advent of video-action models offers a promising path for robot control. Nevertheless, we argue that repurposing video generative models designed

← Previous
1…3637383940…212
Next →