AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

MESA: Improving MoE Safety Alignment via Decentralized Expertise

DGX agent

arXiv:2606.00651v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures scale Large Language Models (LLMs) efficiently, enabling greater capacity with reduced computational cost by dy

safetyarxiv-cs-ai
2 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Meta-Black-Box Optimization with Ensemble Surrogate Modeling for Robustness-Accuracy Trade-off within SAEA

DGX agent

arXiv:2606.00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems. However, their reliance on rig

safetyarxiv-cs-lg
2 Jun 2026
Safety

Meta expands Teen Accounts safety features to limit harmful content on Instagram, Facebook, and Messenger, including on nutrition, weight lifting, and anxiety (Eli Tan/New York Times)

DGX agent

Eli Tan / New York Times: Meta expands Teen Accounts safety features to limit harmful content on Instagram, Facebook, and Messenger, including on nutrition, weight lifting, and anxiety — The changes,

safetytechmeme
2 Jun 2026
Safety

Microsoft releases ASSERT, an open-source framework that lets developers generate and run AI behavior tests using natural-language descriptions (Ram Iyer/TechCrunch)

DGX agent

Ram Iyer / TechCrunch: Microsoft releases ASSERT, an open-source framework that lets developers generate and run AI behavior tests using natural-language descriptions — AI researchers and labs have ad

safetytechmeme
2 Jun 2026
Safety

MidSteer: Optimal Affine Framework for Steering Generative Models

DGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

safetyarxiv-cs-ai
2 Jun 2026
Safety

Minimax-Optimal Policy Regret in Partially Observable Markov Games

DGX agent

arXiv:2606.02363v1 Announce Type: new Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov g

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mitigating Bias in Locally Constrained Decoding via Tractable Proposals

DGX agent

arXiv:2606.01926v1 Announce Type: new Abstract: Generations from large language models often fail to conform to desired constraints such as JSON schema. Existing locally constrained decoding (LCD) app

safetyarxiv-cs-cl
2 Jun 2026
Safety

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

DGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

safetyarxiv-cs-ai
2 Jun 2026
Safety

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

DGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

safetyarxiv-cs-ai
2 Jun 2026
Safety

Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment

DGX agent

arXiv:2606.02198v1 Announce Type: new Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce differen

safetyarxiv-cs-lg
2 Jun 2026
Safety

MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts

DGX agent

arXiv:2606.00844v1 Announce Type: cross Abstract: Bounding-box regression is a fundamental component of object detection, playing a critical role in precise object localization. Existing Intersection-

safetyarxiv-cs-ai
2 Jun 2026
Safety

Morningstar: Get real, SpaceX just isn’t worth a trillion dollars, let alone two.

DGX agent

Gary Marcus argues that SpaceX's valuation is significantly inflated, contending that the company is not worth the trillion-dollar valuations that have been suggested. The critique appears to challeng

safetygary-marcus--x
2 Jun 2026
Safety

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

DGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

safetyarxiv-cs-ai
2 Jun 2026
Safety

Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection

DGX agent

arXiv:2606.02352v1 Announce Type: new Abstract: Robust self-supervised learning of multi-modal video representations is critical for real-world applications such as driver distraction detection, where

safetyarxiv-cs-cv
2 Jun 2026
Safety

Multi-Objective Reference-Aligned Machine Unlearning

DGX agent

arXiv:2606.00399v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training samples while preserving the model's utility. Existing single-objective approaches,

safetyarxiv-cs-lg
2 Jun 2026
Safety

Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic

DGX agent

arXiv:2601.18783v2 Announce Type: replace-cross Abstract: Balancing safety, efficiency, and operational costs in highway driving poses a challenging decision-making problem for heavy-duty vehicles. A

safetyarxiv-cs-ai
2 Jun 2026
Safety

MURMUR: An Efficient Inference System for Long-Form ASR

DGX agent

arXiv:2606.01483v1 Announce Type: cross Abstract: Long-form automatic speech recognition (ASR) requires both high accuracy and low latency, but existing systems force a trade-off between the two. Chun

safetyarxiv-cs-ai
2 Jun 2026
Safety

MViewRouter: Internalizing Geometric Equivariance via Multi-view Alternating Attention for Combinatorial Routing

DGX agent

arXiv:2606.01084v1 Announce Type: cross Abstract: Combinatorial routing problems such as the Traveling Salesman Problem (TSP) and the Capacitated Vehicle Routing Problem (CVRP) are fundamental NP-hard

safetyarxiv-cs-ai
2 Jun 2026
Safety

MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding

DGX agent

arXiv:2606.00174v1 Announce Type: cross Abstract: Electromyography (EMG) directly reflects muscle activation and is a key sensing modality for gesture recognition, prosthetic control, and wearable int

safetyarxiv-cs-ai
2 Jun 2026
Safety

NDPP-Grasp: Non-Differentiable Physical Plausibility Constraint-Guided Task-Oriented Dexterous Grasp Generation

DGX agent

arXiv:2606.02432v1 Announce Type: new Abstract: Task-oriented dexterous grasp generation aims to produce dexterous grasp poses that are both physically plausible and functionally suitable for specifie

safetyarxiv-cs-ro
2 Jun 2026
Safety

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

DGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

safetyarxiv-cs-ai
2 Jun 2026
Safety

NEW: Ahead of the SpaceX IPO, we tracked hundreds of promises that Musk has made over the years (FSD, Mars etc). His success rate is…not goo…

DGX agent

NEW: Ahead of the SpaceX IPO, we tracked hundreds of promises that Musk has made over the years (FSD, Mars etc). His success rate is…not good. And it’s getting worse. Some actual data in our months lo

safetygary-marcus--x
2 Jun 2026
Safety

Non-Uniform Noise-to-Signal Ratio in the REINFORCE Policy-Gradient Estimator

DGX agent

arXiv:2602.01460v3 Announce Type: replace-cross Abstract: Policy-gradient methods are widely used in reinforcement learning, yet training often becomes unstable or slows down as learning progresses. W

safetyarxiv-cs-lg
2 Jun 2026
Safety

NormEval: A Unified Multi-Metric Framework for Evaluating Semantic Fidelity in Text Normalization

DGX agent

arXiv:2511.20409v2 Announce Type: replace Abstract: Text normalization methods such as stemming and lemmatization are fundamental components of NLP pipelines. As new normalization tools are developed

safetyarxiv-cs-cl
2 Jun 2026
Safety

Not convinced that this kind of nationalization by fiat is at all the right way to go (and for that matter don’t expect the current breed of…

DGX agent

Not convinced that this kind of nationalization by fiat is at all the right way to go (and for that matter don’t expect the current breed of technology to generate trillions), but I am glad that Sande

safetygary-marcus--x
2 Jun 2026
Safety

ObjEmbed: Towards Universal Multimodal Object Embeddings

DGX agent

arXiv:2602.01753v3 Announce Type: replace Abstract: Aligning objects with corresponding textual descriptions is a fundamental challenge and a realistic requirement in vision-language understanding. Wh

safetyarxiv-cs-cv
2 Jun 2026
Safety

Off-Policy Learning in Large Action Spaces: Optimization Matters More Than Estimation

DGX agent

arXiv:2509.03456v2 Announce Type: replace-cross Abstract: Off-policy evaluation (OPE) and off-policy learning (OPL) are foundational for decision-making in offline contextual bandits. Recent advances

safetyarxiv-cs-lg
2 Jun 2026
Safety

On Effectiveness and Efficiency of Agentic Tool-calling and RL Training

DGX agent

arXiv:2606.00135v1 Announce Type: cross Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This pa

safetyarxiv-cs-ai
2 Jun 2026
Safety

On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance

DGX agent

arXiv:2606.00467v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for zero-shot annotation and LLM-as-a-judge tasks, yet their reliability hinges on how model-intern

safetyarxiv-cs-ai
2 Jun 2026
Safety

One Bias After Another: Mechanistic Reward Shaping and Persistent Biases in Language Reward Models

DGX agent

arXiv:2603.03291v2 Announce Type: replace-cross Abstract: Reward Models (RMs) are crucial for online alignment of language models (LMs) with human preferences. However, RM-based preference-tuning is v

safetyarxiv-cs-ai
2 Jun 2026
Safety

OPD+: Rethinking the Advantage Design for On-Policy Distillation

DGX agent

arXiv:2606.01039v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a widely used technique to transfer capabilities from capable teacher language models to the base student models, and

safetyarxiv-cs-ai
2 Jun 2026
Safety

OpenAI says it has not donated to any super PACs and does not have an employee-funded PAC, and that Greg Brockman's support for Leading the Future is personal (OpenAI)

DGX agent

OpenAI: OpenAI says it has not donated to any super PACs and does not have an employee-funded PAC, and that Greg Brockman's support for Leading the Future is personal — AI is going to be one of the mo

safetytechmeme
2 Jun 2026
Safety

Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers

DGX agent

arXiv:2602.05395v2 Announce Type: replace-cross Abstract: A simple strategy for improving LLM accuracy, especially in math and reasoning problems, is to sample multiple responses and submit the answer

safetyarxiv-cs-ai
2 Jun 2026
Safety

Optimizing Diversity and Quality through Base-Aligned Model Collaboration

DGX agent

arXiv:2511.05650v2 Announce Type: replace-cross Abstract: Alignment has greatly improved large language models (LLMs)' output quality at the cost of diversity, yielding highly similar outputs across g

safetyarxiv-cs-ai
2 Jun 2026
Safety

OSCAR: Obstacle Survival Curves for Adaptive Robot Navigation

DGX agent

arXiv:2606.00990v1 Announce Type: new Abstract: A mobile robot following a graph of known routes can make costly navigation errors when a temporary obstacle blocks a critical edge: waiting too long be

safetyarxiv-cs-ro
2 Jun 2026
Safety

PACE: Phase-Aware Chunk Execution for Robot Policies with Action Chunking

DGX agent

arXiv:2606.00537v1 Announce Type: new Abstract: Recent vision-language-action and diffusion-based robot policies often use action chunking, where each policy query predicts a sequence of future action

safetyarxiv-cs-ro
2 Jun 2026
Safety

Paradoxical noise preference in RNNs

DGX agent

arXiv:2601.04539v2 Announce Type: replace-cross Abstract: In recurrent neural networks (RNNs) used to model biological neural networks, noise is typically introduced during training to emulate biologi

safetyarxiv-cs-ai
2 Jun 2026
Safety

Partial Fairness Awareness: Belief-Guided Strategic Mechanism for Strategic Agents

DGX agent

arXiv:2606.00826v1 Announce Type: new Abstract: Strategic machine learning investigates scenarios where agents manipulate their features to receive favorable decisions from predictive models. To addre

safetyarxiv-cs-lg
2 Jun 2026
Safety

Pave-GRPO: Beyond Instantaneous Guidance through Principled Average Velocity Decomposition

DGX agent

arXiv:2606.01636v1 Announce Type: new Abstract: Post-training via Group Relative Policy Optimization (GRPO) has emerged as a powerful paradigm for aligning flow-based generative models with human pref

safetyarxiv-cs-cv
2 Jun 2026
Safety

Perspective on Bias in Biomedical AI: Preventing Downstream Healthcare Disparities

DGX agent

arXiv:2604.14514v2 Announce Type: replace Abstract: Healthcare disparities persist across socioeconomic boundaries, often attributed to unequal access to screening, diagnostics, and therapeutics. Howe

safetyarxiv-cs-ai
2 Jun 2026
Safety

Perturbation Effects on Accuracy and Fairness among Similar Individuals

DGX agent

arXiv:2404.01356v3 Announce Type: replace-cross Abstract: Deep neural networks are vulnerable to adversarial perturbations that can simultaneously degrade prediction robustness and individual fairness

safetyarxiv-cs-ai
2 Jun 2026
Safety

PHASOR: Phase-Anchored Universal Action Representations for Humanoid Embodiments

DGX agent

arXiv:2606.01851v1 Announce Type: new Abstract: Learning a good action embedding space is fundamental to scalable robot policy learning, yet existing methods treat action latents as task-specific inte

safetyarxiv-cs-ro
2 Jun 2026
Safety

PhyScene3D: Physically Consistent Interactive 3D Tabletop Scene Generation

DGX agent

arXiv:2606.01649v1 Announce Type: new Abstract: Generating physically consistent 3D tabletop scenes is a fundamental yet underexplored problem for interactive and generalist robotic learning. The chal

safetyarxiv-cs-cv
2 Jun 2026
Safety

Physically-Constrained Mamba-SDE for Remaining Useful Life Prediction under Irregular Observations

DGX agent

arXiv:2606.01894v1 Announce Type: new Abstract: Accurate Remaining Useful Life prediction is critical for industrial predictive maintenance. However, real-world deployment is challenging due to the ir

safetyarxiv-cs-ai
2 Jun 2026
Safety

Policy and World Modeling Co-Training for Language Agents

DGX agent

arXiv:2606.02388v1 Announce Type: cross Abstract: Reinforcement learning (RL) improves large language model (LLM) agents by teaching them which actions lead to high rewards, but provides little superv

safetyarxiv-cs-ai
2 Jun 2026
Safety

Policy-based Foveated Imaging and Perception

DGX agent

arXiv:2606.02565v1 Announce Type: new Abstract: Ultra-high-resolution image sensors offer the potential to capture fine spatial details critical for many visual perception tasks, but acquiring and pro

safetyarxiv-cs-cv
2 Jun 2026
Safety

Pool-Select-Refine: Allocation-Aware Generative Dataset Distillation with Soft-Label-Guided Latent Refinement

DGX agent

arXiv:2606.01920v1 Announce Type: new Abstract: Diffusion-based dataset distillation has recently emerged as a promising paradigm for condensing large-scale datasets into compact synthetic sets. By le

safetyarxiv-cs-cv
2 Jun 2026
Safety

Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants

DGX agent

arXiv:2508.08337v3 Announce Type: replace-cross Abstract: Algorithmic fairness research has largely framed unfairness as discrimination along sensitive attributes. However, this approach limits visibi

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…123124125126127…267
Next →