AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy

DGX agent

arXiv:2605.04295v1 Announce Type: new Abstract: LLMs' overconfidence, particularly when hallucinating, poses a significant challenge for the deployment of the models in safety-critical settings and ma

safetyarxiv-cs-lg
7 May 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Look Once, Beam Twice: Camera-Primed Real-Time Double-Directional mmWave Beam Management for Vehicular Connectivity

DGX agent

arXiv:2605.05071v1 Announce Type: cross Abstract: Millimeter-wave (mmWave) frequencies promise multi-gigabit connectivity for vehicle-to-everything (V2X) networks, but face challenges in terms of seve

safetyarxiv-cs-cv
7 May 2026
Safety

Many, many OpenAI employees quit over safety concerns, including @DKokotajlo, William Saunders, @sjgadler, etc as well @Janleike. The founde…

DGX agent

Many, many OpenAI employees quit over safety concerns, including @DKokotajlo, William Saunders, @sjgadler, etc as well @Janleike. The founders of Anthropic such as @DarioAmodei and @jackclarkSF may ha

safetygary-marcus--x
7 May 2026
Safety

Marcus (rightly) Mocks X influencer accounts 😁

DGX agent

Gary Marcus criticizes the credibility and practices of X (formerly Twitter) influencer accounts, likely highlighting misleading claims, engagement manipulation, or questionable expertise common among

safetygary-marcus--x
7 May 2026
Safety

Mechanical Conscience: A Mathematical Framework for Dependability of Machine Intelligenc

DGX agent

arXiv:2605.03847v1 Announce Type: new Abstract: Distributed collaborative intelligence (DCI), encompassing edge-to-edge architectures, federated learning, transfer learning, and swarm systems, creates

safetyarxiv-cs-ai
7 May 2026
Safety

MedFabric and EtHER: A Data-Centric Framework for Word-Level Fabrication Generation and Detection in Medical LLMs

DGX agent

arXiv:2605.04180v1 Announce Type: new Abstract: Large Language Models exhibit strong reasoning and semantic understanding capabilities but often hallucinate in domains that require expert knowledge, a

safetyarxiv-cs-cl
7 May 2026
Safety

MenuNet: A Strategy-Proof Mechanism for Matching Markets

DGX agent

arXiv:2605.03216v1 Announce Type: cross Abstract: Strategy-proofness is a fundamental desideratum in mechanism design, ensuring truthful reporting and robust participation. Stability is another centra

safetyarxiv-cs-ai
7 May 2026
Safety

Meta challenges Ofcom in UK High Court over the Online Safety Act, which calculates levies based on global, not UK, revenue, in a case scheduled for October (Sam Tobin/Reuters)

DGX agent

Sam Tobin / Reuters: Meta challenges Ofcom in UK High Court over the Online Safety Act, which calculates levies based on global, not UK, revenue, in a case scheduled for October — Facebook and Instagr

safetytechmeme
7 May 2026
Safety

Misaligned by Reward: Socially Undesirable Preferences in LLMs

DGX agent

arXiv:2605.05003v1 Announce Type: new Abstract: Reward models are a key component of large language model alignment, serving as proxies for human preferences during training. However, existing evaluat

safetyarxiv-cs-cl
7 May 2026
Safety

Multi-Level Bidirectional Biomimetic Learning for EEG-Based Visual Decoding

DGX agent

arXiv:2605.04680v1 Announce Type: new Abstract: EEG-based visual neural decoding aims to align neural responses with visual stimuli for tasks such as image retrieval. However, limited paired data and

safetyarxiv-cs-cv
7 May 2026
Safety

Multi-Scale Wavelet Transformers for Operator Learning of Dynamical Systems

DGX agent

arXiv:2602.01486v2 Announce Type: replace Abstract: Recent years have seen a surge in data-driven surrogates for dynamical systems that can be orders of magnitude faster than numerical solvers. Howeve

safetyarxiv-cs-lg
7 May 2026
Safety

Multivariate Time Series Data Imputation via Distributionally Robust Regularization

DGX agent

arXiv:2602.00844v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is often compromised by mismatch between the observed and true data distributions, a bias induced by the c

safetyarxiv-cs-lg
7 May 2026
Safety

NEAT: Neighborhood-Guided, Efficient, Autoregressive Set Transformer for 3D Molecular Generation

DGX agent

arXiv:2512.05844v3 Announce Type: replace Abstract: Transformer-based autoregressive models offer an efficient alternative to diffusion- and flow-matching-based approaches for generating 3D molecules.

safetyarxiv-cs-lg
7 May 2026
Safety

New Bigtable in-memory tier for sub-millisecond read latency

DGX agent

In the high-stakes world of digital infrastructure, speed isn't just a metric — it’s currency. At Google Cloud Next ‘26 we announced the Bigtable in-memory tier, a breakthrough for our fully managed c

safetygoogle-cloud-ai
7 May 2026
Safety

On-line Learning in Tree MDPs by Treating Policies as Bandit Arms

DGX agent

arXiv:2605.04979v1 Announce Type: cross Abstract: A Tree Markov Decision Problem (T-MDP) is a finite-horizon MDP with a starting state s_{1}, in which every state is reachable from s_{1} through exact

safetyarxiv-cs-lg
7 May 2026
Safety

On the Hardness of Junking LLMs

DGX agent

arXiv:2605.05116v1 Announce Type: new Abstract: Large language models (LLMs) are known to be vulnerable to jailbreak attacks, which typically rely on carefully designed prompts containing explicit sem

safetyarxiv-cs-lg
7 May 2026
Safety

One of the things that made the Mythos release hard to interpret is that Anthropic held back details on most vulns they found, to give defen…

DGX agent

One of the things that made the Mythos release hard to interpret is that Anthropic held back details on most vulns they found, to give defenders time to patch. 1 month later, info from orgs with acces

safetygary-marcus--x
7 May 2026
Safety

One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving

DGX agent

arXiv:2605.04450v1 Announce Type: cross Abstract: Generative Recommender (GR) inference places embedding hot caches (EMB) and KV caches in direct competition for limited GPU HBM: allocating more memor

safetyarxiv-cs-lg
7 May 2026
Safety

OracleProto: A Reproducible Framework for Benchmarking LLM Native Forecasting via Knowledge Cutoff and Temporal Masking

DGX agent

arXiv:2605.03762v1 Announce Type: new Abstract: Large language models are moving from static text generators toward real-world decision-support systems, where forecasting is a composite capability tha

safetyarxiv-cs-ai
7 May 2026
Safety

Order Matters: Improving Domain Adaptation by Reordering Data

DGX agent

arXiv:2605.05084v1 Announce Type: new Abstract: Domain shift remains a key challenge in deploying machine learning models to the real world. Unsupervised domain adaptation (UDA) aims to address this b

safetyarxiv-cs-lg
7 May 2026
Safety

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage

DGX agent

arXiv:2506.07548v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) has reached competitive performance on cooperative tasks against scripted adversaries, yet most meth

safetyarxiv-cs-ro
7 May 2026
Safety

Overcoming reward signal challenges: Verifiable rewards-based reinforcement learning with GRPO on SageMaker AI

DGX agent

In this post, you will learn how to implement reinforcement learning with verifiable rewards (RLVR) to introduce verification and transparency into reward signals to improve training performance. This

safetyaws-ml-blog
7 May 2026
Safety

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal

DGX agent

arXiv:2603.22844v4 Announce Type: replace Abstract: Surgical smoke severely degrades intraoperative video quality, obscuring anatomical structures and limiting surgical perception. Existing learning-b

safetyarxiv-cs-ai
7 May 2026
Safety

POMA-3D: The Point Map Way to 3D Scene Understanding

DGX agent

arXiv:2511.16567v3 Announce Type: replace Abstract: In this paper, we introduce POMA-3D, the first self-supervised 3D representation model learned from point maps. Point maps encode explicit 3D coordi

safetyarxiv-cs-cv
7 May 2026
Safety

Practical validation of synthetic pre-crash scenarios

DGX agent

arXiv:2605.04564v1 Announce Type: new Abstract: The representativeness of synthetic pre-crash scenarios is crucial for assessing the safety impact of Driving Automation Systems through virtual simulat

safetyarxiv-cs-ro
7 May 2026
Safety

Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs

DGX agent

arXiv:2605.04215v1 Announce Type: new Abstract: Diffusion-based Large Language Models (D-LLMs) represent a promising frontier in generative AI, offering fully parallel token generation that can lead t

safetyarxiv-cs-lg
7 May 2026
Safety

Preference-Based Self-Distillation: Beyond KL Matching via Reward Regularization

DGX agent

arXiv:2605.05040v1 Announce Type: new Abstract: On-policy distillation is an efficient alternative to reinforcement learning, offering dense token-level training signals. However, its reliance on a st

safetyarxiv-cs-lg
7 May 2026
Safety

ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection

DGX agent

arXiv:2601.09195v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is a fundamental post-training strategy to align Large Language Models (LLMs) with human intent. However, traditional S

safetyarxiv-cs-cl
7 May 2026
Safety

Provable imitation learning for control of instability in partially-observed Vlasov--Poisson equations

DGX agent

arXiv:2605.05081v1 Announce Type: new Abstract: We consider the stabilization of Vlasov--Poisson plasma dynamics, a central control problem in nuclear fusion. Our focus is the gap between what an idea

safetyarxiv-cs-lg
7 May 2026
Safety

Purdah and Patriarchy: Evaluating and Mitigating South Asian Biases in Open-Ended Multilingual LLM Generations

DGX agent

arXiv:2505.18466v2 Announce Type: replace Abstract: Evaluations of Large Language Models (LLMs) often overlook intersectional and culturally specific biases, particularly in underrepresented multiling

safetyarxiv-cs-cl
7 May 2026
Safety

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents

DGX agent

arXiv:2604.03976v2 Announce Type: replace Abstract: Prior work on trustworthy AI emphasizes model-internal properties such as bias mitigation, adversarial robustness, and interpretability. As AI syste

safetyarxiv-cs-ai
7 May 2026
Safety

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments

DGX agent

arXiv:2508.04204v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have demonstrated impressive performance in reasoning-intensive tasks, but they remain vulnerable to harmful content g

safetyarxiv-cs-cl
7 May 2026
Safety

Reinforcement Learning for Compositional Generalization with Outcome-Level Optimization

DGX agent

arXiv:2605.04920v1 Announce Type: cross Abstract: Compositional generalization refers to correctly interpret novel combinations of known primitives, which remains a major challenge. Existing approache

safetyarxiv-cs-cl
7 May 2026
Safety

Road Risk Monitor: A Deployable U.S. Road Incident Forecasting System with Live Weather and Road-Level Tiles

DGX agent

arXiv:2605.04242v1 Announce Type: new Abstract: Nationwide road-incident forecasting is a systems problem before it is a modeling problem. A usable service must connect historical incident archives, h

safetyarxiv-cs-lg
7 May 2026
Safety

Robust Agent Compensation (RAC): Teaching AI Agents to Compensate

DGX agent

arXiv:2605.03409v1 Announce Type: new Abstract: We present Robust Agent Compensation (RAC), a log-based recovery paradigm (providing a safety net) implemented through an architectural extension that c

safetyarxiv-cs-ai
7 May 2026
Safety

Rollout Pass-Rate Control: Steering Binary-Reward RL Toward Its Most Informative Regime

DGX agent

arXiv:2605.05112v1 Announce Type: new Abstract: SWE-bench-style agentic reinforcement learning relies on expensive stateful trajectories, yet substantial compute is wasted on sampled rollout groups wi

safetyarxiv-cs-lg
7 May 2026
Safety

S1-MMAlign: A Large-Scale, Multi-Disciplinary Dataset for Scientific Figure-Text Understanding

DGX agent

arXiv:2601.00264v2 Announce Type: replace Abstract: Multimodal learning has revolutionized general domain tasks, yet its application in scientific discovery is hindered by the profound semantic gap be

safetyarxiv-cs-cv
7 May 2026
Safety

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

DGX agent

arXiv:2601.08623v2 Announce Type: replace Abstract: Image generation models (IGMs), while capable of producing impressive and creative content, often memorize a wide range of undesirable concepts from

safetyarxiv-cs-cv
7 May 2026
Safety

Safety by Invariance, Liveness through Refinement: Heterogeneous Contract Framework for Co-Design of Layered Control

DGX agent

arXiv:2605.04222v1 Announce Type: cross Abstract: Real-world control systems must achieve long-horizon objectives (liveness) while respecting continuous-time safety constraints, a combination that mot

safetyarxiv-cs-ro
7 May 2026
Safety

Safety Must Precede the Deployment of Open-Ended AI

DGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

safetyarxiv-cs-ai
7 May 2026
Safety

Scalable inference of spatial regions and temporal signatures from time series

DGX agent

arXiv:2605.05008v1 Announce Type: cross Abstract: Regionalization aims to partition a spatial domain into contiguous regions that share similar characteristics, enabling more effective spatial analysi

safetyarxiv-cs-lg
7 May 2026
Safety

Scalable Multi Agent Diffusion Policies for Coverage Control

DGX agent

arXiv:2509.17244v2 Announce Type: replace Abstract: We propose MADP, a novel diffusion-model-based approach for collaboration in decentralized robot swarms. MADP leverages diffusion models to generate

safetyarxiv-cs-ro
7 May 2026
Safety

Scalable Policy Maximization Under Network Interference

DGX agent

arXiv:2505.18118v2 Announce Type: replace-cross Abstract: Many interventions, such as vaccines in clinical trials or coupons in online marketplaces, must be assigned sequentially without full knowledg

safetyarxiv-cs-lg
7 May 2026
Safety

Sequential Strategic Classification with Multi-Stage Selective Classifiers

DGX agent

arXiv:2605.04202v1 Announce Type: new Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outcomes

safetyarxiv-cs-lg
7 May 2026
Safety

Software Engineering for Self-Adaptive Robotics: A Research Agenda

DGX agent

arXiv:2505.19629v3 Announce Type: replace-cross Abstract: Self-adaptive robotic systems operate autonomously in dynamic and uncertain environments, requiring robust real-time monitoring and adaptive b

safetyarxiv-cs-ro
7 May 2026
Safety

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization

DGX agent

arXiv:2605.04700v1 Announce Type: cross Abstract: Jailbreak attacks on audio language models (ALMs) optimize audio perturbations to elicit unsafe generations, and they typically update the entire wave

safetyarxiv-cs-cl
7 May 2026
Safety

Structural Equivalence and Learning Dynamics in Delayed MARL

DGX agent

arXiv:2605.04345v1 Announce Type: new Abstract: We formally establish the equivalence between Observation Delay (OD) and Action Delay (AD) in cooperative partially observable multi-agent systems using

safetyarxiv-cs-lg
7 May 2026
Safety

Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition

DGX agent

arXiv:2605.04617v1 Announce Type: new Abstract: Wearable human activity recognition (WHAR) models often suffer from performance degradation under real-world cross-user distribution shifts. Test-time a

safetyarxiv-cs-cv
7 May 2026
← Previous
1…200201202203204…265
Next →