AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

General Frameworks for Conditional Two-Sample Testing

DGX agent

arXiv:2410.16636v2 Announce Type: replace-cross Abstract: We study the problem of conditional two-sample testing, which aims to determine whether two populations have the same distribution after accou

safetyarxiv-cs-lg
5 May 2026
Safety

Generalized Distributional Alignment Games for Unbiased Answer-Level Fine-Tuning

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.02435v1 Announce Type: new Abstract: The Distributional Alignment Game framework provides a powerful variational perspective on Answer-Level Fine-Tuning (ALFT). However, standard algorithms

safetyarxiv-cs-lg
5 May 2026
Safety

Geometric and Spectral Alignment for Deep Neural Network I

DGX agent

arXiv:2605.02108v1 Announce Type: new Abstract: Deep residual architectures are modeled as products of near-identity Jacobians. This paper proves deterministic quotient-geometric estimates for singula

safetyarxiv-cs-lg
5 May 2026
Safety

Geometric and Spectral Alignment for Deep Neural Network II

DGX agent

arXiv:2605.02111v1 Announce Type: new Abstract: This paper develops the angular and static-channel component of Geometric and Spectral Alignment for residual Jacobian chains. Starting from Cartan-coor

safetyarxiv-cs-lg
5 May 2026
Safety

GETA-3DGS: Automatic Joint Structured Pruning and Quantization for 3D Gaussian Splatting

DGX agent

arXiv:2605.02086v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) is a state-of-the-art representation for real-time photorealistic novel-view synthesis, yet a single high-fidelity scene ty

safetyarxiv-cs-lg
5 May 2026
Safety

Good in Bad (GiB): Sifting Through End-user Demonstrations for Learning a Better Policy

DGX agent

arXiv:2605.01529v1 Announce Type: new Abstract: Imitation learning offers a promising framework for enabling robots to acquire diverse skills from human users. However, most imitation learning algorit

safetyarxiv-cs-ro
5 May 2026
Safety

Gradient-Gated DPO: Stabilizing Preference Optimization in Language Models

DGX agent

arXiv:2605.02626v1 Announce Type: new Abstract: Preference optimization has become a central paradigm for aligning large language models with human feedback. Direct Preference Optimization (DPO) simpl

safetyarxiv-cs-lg
5 May 2026
Safety

Green Energy Management for Sustainable Data Centers Using Deep Reinforcement Learning

DGX agent

arXiv:2507.21153v2 Announce Type: replace Abstract: The exponential growth of digital services has positioned data centers among the most energy-intensive infrastructures in the modern economy, raisin

safetyarxiv-cs-lg
5 May 2026
Safety

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies

DGX agent

arXiv:2603.12243v3 Announce Type: replace Abstract: Mastering dexterous manipulation with multi-fingered hands has been a grand challenge in robotics for decades. Despite its potential, the difficulty

safetyarxiv-cs-ro
5 May 2026
Safety

Hazard-Aware Traffic Scene Graph Generation

DGX agent

arXiv:2603.03584v2 Announce Type: replace Abstract: Maintaining situational awareness in complex driving scenarios is challenging. It requires continuously prioritizing attention among extensive scene

safetyarxiv-cs-cv
5 May 2026
Safety

HeteroRAG: A Heterogeneous Retrieval-Augmented Generation Framework for Medical Vision Language Tasks

DGX agent

arXiv:2508.12778v2 Announce Type: replace Abstract: Medical large vision-language Models (Med-LVLMs) have shown promise in clinical applications but suffer from factual inaccuracies and unreliable out

safetyarxiv-cs-cl
5 May 2026
Safety

High entropy leads to symmetry equivariant policies in Dec-POMDPs

DGX agent

arXiv:2511.22581v3 Announce Type: replace Abstract: We prove that in any Dec-POMDP, sufficiently high entropy regularization ensures that the policy gradient flow with tabular softmax parametrization

safetyarxiv-cs-lg
5 May 2026
Safety

How Can One Choose the Best CAM-Based Explainability Method for a CNN Model?

DGX agent

arXiv:2605.02007v1 Announce Type: cross Abstract: In recent years, several advances have been observed in Deep Learning with surprising results. Models in this area have been increasingly used in nume

safetyarxiv-cs-cv
5 May 2026
Safety

HumanSplatHMR: Closing the Loop Between Human Mesh Recovery and Gaussian Splatting Avatar

DGX agent

arXiv:2605.02784v1 Announce Type: new Abstract: Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motio

safetyarxiv-cs-cv
5 May 2026
Safety

Hybrid Quantum Reinforcement Learning with QAOA for Improved Vehicle Routing Optimization

DGX agent

arXiv:2605.01574v1 Announce Type: new Abstract: Vehicle Routing Problem (VRP) is one of the most complex NP-hard combinatorial optimization problem in transportation and logistics that requires a dyna

safetyarxiv-cs-lg
5 May 2026
Safety

Hydra-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control

DGX agent

arXiv:2605.01581v1 Announce Type: new Abstract: Diffusion-based visuomotor policies perform well in robotic manipulation, yet current methods still inherit image-generation-style decoders and multi-st

safetyarxiv-cs-ro
5 May 2026
Safety

Hyp2Former: Hierarchy-Aware Hyperbolic Embeddings for Open-Set Panoptic Segmentation

DGX agent

arXiv:2605.02580v1 Announce Type: new Abstract: Recognizing unknown objects is crucial for safety-critical applications such as autonomous driving and robotics. Open-Set Panoptic Segmentation (OPS) ai

safetyarxiv-cs-cv
5 May 2026
Safety

I am old enough to remember when people used to believe every word of Sam’s bullshit (and to call me a “hater” for doubting him).

DGX agent

Gary Marcus reflects on changing public perception of Sam Altman, noting that skepticism toward Altman's claims was once dismissed as hatred but has become more mainstream. The post suggests a shift i

safetygary-marcus--x
5 May 2026
Safety

I predict this will be the most embarrassing cover in the history of this magazine. This is like putting Enron executives on your cover, or …

DGX agent

I predict this will be the most embarrassing cover in the history of this magazine. This is like putting Enron executives on your cover, or doing a gauzy special report on Bernie Madoff. Sam Altman is

safetygary-marcus--x
5 May 2026
Safety

Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction

DGX agent

arXiv:2510.25426v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) is positioning language at the core of human-computer interaction (HCI). We argue that advanci

safetyarxiv-cs-cl
5 May 2026
Safety

Important new development. AI company employees have an enormous amount of power — far more than they realize. Absent legislation, AI co wor…

DGX agent

Important new development. AI company employees have an enormous amount of power — far more than they realize. Absent legislation, AI co worker power is one of the key levers to shaping what the indus

safetygary-marcus--x
5 May 2026
Safety

Improving Model Safety by Targeted Error Correction

DGX agent

arXiv:2605.02544v1 Announce Type: cross Abstract: The widespread adoption of machine learning in critical applications demands techniques to mitigate high-consequence errors. Our method utilizes a dua

safetyarxiv-cs-cv
5 May 2026
Safety

In vividly explaining how he deceived Musk about his commitment to the nonprofit, without a trace of remorse, OpenAI’s Brockman has done Mus…

DGX agent

In vividly explaining how he deceived Musk about his commitment to the nonprofit, without a trace of remorse, OpenAI’s Brockman has done Musk’s counsel quite a service. Brockman’s also been pretty gre

safetygary-marcus--x
5 May 2026
Safety

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression

DGX agent

arXiv:2605.01402v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) struggle with numerical regression under long-tailed target distributions. Token-level supervised fine-tuning (

safetyarxiv-cs-cl
5 May 2026
Safety

Investigating Anthropometric Fidelity in SAM 3D Body

DGX agent

arXiv:2601.06035v2 Announce Type: replace-cross Abstract: The release of SAM 3D Body is a recent development in human mesh recovery, demonstrating improved performance in producing clean, topologicall

safetyarxiv-cs-cv
5 May 2026
Safety

IPS: In-Prompt Process Supervision for Short Video Content Moderation

DGX agent

arXiv:2412.15251v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are effective at capturing the semantics of short video content; however, they often fail to attend to the

safetyarxiv-cs-cl
5 May 2026
Safety

🆕 @katiemiller has started following @GaryMarcus

DGX agent

Katie Miller began following Gary Marcus on X (formerly Twitter). Gary Marcus is a cognitive scientist and AI researcher known for his public commentary on artificial intelligence and technology polic

safetygary-marcus--x
5 May 2026
Safety

Knowledge-Based Design Requirements for Generative Social Robots in Higher Education

DGX agent

arXiv:2602.12873v4 Announce Type: replace-cross Abstract: Generative social robots (GSRs) powered by large language models enable adaptive, conversational tutoring but also introduce risks such as mis

safetyarxiv-cs-ai
5 May 2026
Safety

Lateral String Stability for Vehicle Platoons: Formulation, Definition, and Analysis

DGX agent

arXiv:2605.01731v1 Announce Type: new Abstract: Platooning of connected and automated vehicles provides significant benefits in terms of energy efficiency, traffic throughput, and, most critically, sa

safetyarxiv-cs-ro
5 May 2026
Safety

Learning to Act Through Contact: A Unified View of Multi-Task Robot Learning

DGX agent

arXiv:2510.03599v2 Announce Type: replace Abstract: We present a unified framework for multi-task locomotion and manipulation policy learning grounded in a contact-explicit representation. Instead of

safetyarxiv-cs-ro
5 May 2026
Safety

Less is More: Geometric Unlearning for LLMs with Minimal Data Disclosure

DGX agent

arXiv:2605.01735v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in real-world systems, they must support post-hoc removal of specific content to meet privacy

safetyarxiv-cs-cl
5 May 2026
Safety

Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy

DGX agent

arXiv:2509.21173v5 Announce Type: replace Abstract: Vision-Language Models (VLMs) such as CLIP have revolutionized zero-shot classification and safety-critical tasks, including Out-of-Distribution (OO

safetyarxiv-cs-cv
5 May 2026
Safety

Linking spatial biology and clinical histology via Haiku

DGX agent

arXiv:2605.00925v1 Announce Type: cross Abstract: Integrating molecular, morphological, and clinical data is essential for basic and translational biomedical research, yet systematic frameworks for jo

safetyarxiv-cs-cv
5 May 2026
Safety

LLM-Augmented Semantic Steering of Text Embedding Projection Spaces

DGX agent

arXiv:2605.01957v1 Announce Type: cross Abstract: Low-dimensional projections of text embeddings support visual analysis of document collections, but their spatial organization may not reflect the rel

safetyarxiv-cs-cl
5 May 2026
Safety

LLM-Based Agentic Negotiation for 6G: Addressing Uncertainty Neglect and Tail-Event Risk

DGX agent

arXiv:2511.19175v2 Announce Type: replace-cross Abstract: A critical barrier to the trustworthiness of sixth-generation (6G) agentic autonomous networks is the uncertainty neglect bias; a cognitive te

safetyarxiv-cs-ai
5 May 2026
Safety

LLM-VA: Resolving the Jailbreak-Overrefusal Trade-off via Vector Alignment

DGX agent

arXiv:2601.19487v2 Announce Type: replace Abstract: Safety-aligned LLMs suffer from two failure modes: jailbreak (answering harmful inputs) and over-refusal (declining benign queries). Existing vector

safetyarxiv-cs-lg
5 May 2026
Safety

Logit-Gap Steering: A Forward-Pass Diagnostic for Alignment Robustness

DGX agent

arXiv:2506.24056v2 Announce Type: replace-cross Abstract: RLHF-style alignment trains language models to refuse unsafe requests, but how much operational margin does this refusal rest on? We introduce

safetyarxiv-cs-cl
5 May 2026
Safety

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance

DGX agent

arXiv:2410.18717v2 Announce Type: replace Abstract: Recent advancements in artificial intelligence hold ample potential for monitoring applications using surveillance cameras. However, concerns about

safetyarxiv-cs-cv
5 May 2026
Safety

LVLM-Aided Alignment of Task-Specific Vision Models

DGX agent

arXiv:2512.21985v2 Announce Type: replace Abstract: In high-stakes domains, small task-specific vision models are crucial due to their low computational requirements and the availability of numerous m

safetyarxiv-cs-cv
5 May 2026
Safety

Machine Learning Enhanced Laser Spectroscopy for Multi-Species Gas Detection in Complex and Harsh Environments

DGX agent

arXiv:2605.01306v1 Announce Type: cross Abstract: Laser absorption spectroscopy (LAS) is a well-established technique for non-intrusive measurement of gas species in combustion and atmospheric environ

safetyarxiv-cs-lg
5 May 2026
Safety

MAD-OPD: Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate

DGX agent

arXiv:2605.01347v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own trajectories under token-level teacher supervision, but existing methods are capped by a single

safetyarxiv-cs-cl
5 May 2026
Safety

Major new class action lawsuit accuses Meta of copyright infringement around AI training. It says they trained on pirated books, and that AI…

DGX agent

Major new class action lawsuit accuses Meta of copyright infringement around AI training. It says they trained on pirated books, and that AI books flooding the market demonstrates market harm. These l

safetygary-marcus--x
5 May 2026
Safety

Manifold-Constrained Adversarial Training for Long-Tailed Robustness via Geometric Alignment

DGX agent

arXiv:2605.02183v1 Announce Type: new Abstract: Adversarial training is effective on balanced datasets, but its robustness degrades under longtailed class distributions, where tail classes suffer high

safetyarxiv-cs-lg
5 May 2026
Safety

Mean Testing under Truncation beyond Gaussian

DGX agent

arXiv:2605.01335v1 Announce Type: cross Abstract: We characterize the fundamental limits of high-dimensional mean testing under arbitrary truncation, where samples are drawn from the conditional distr

safetyarxiv-cs-lg
5 May 2026
Safety

Meta says it will expand Instagram teen account safeguards to 27 EU countries, and plans to roll them out on Facebook in the US, ahead of the UK and EU in June (Foo Yun Chee/Reuters)

DGX agent

Foo Yun Chee / Reuters: Meta says it will expand Instagram teen account safeguards to 27 EU countries, and plans to roll them out on Facebook in the US, ahead of the UK and EU in June — Meta Platforms

safetytechmeme
5 May 2026
Safety

Minimizing Collateral Damage in Activation Steering

DGX agent

arXiv:2605.01167v1 Announce Type: new Abstract: Activation steering is a method for controlling Large Language Model (LLM) behavior by intervening in its internal representations to increase the align

safetyarxiv-cs-lg
5 May 2026
Safety

MIRA: A Score for Conditional Distribution Accuracy and Model Comparison

DGX agent

arXiv:2605.02014v1 Announce Type: cross Abstract: We introduce Mira, a sample-based score for assessing the accuracy of a candidate conditional distribution using only joint samples from the true data

safetyarxiv-cs-lg
5 May 2026
Safety

Mitigating Misalignment Contagion by Steering with Implicit Traits

DGX agent

arXiv:2605.02751v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used in high-stakes, multi-agent settings, where following instructions and maintaining value alignment are cri

safetyarxiv-cs-cl
5 May 2026
← Previous
1…206207208209210…265
Next →