AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

How Much Progress Has There Been in NVIDIA Datacenter GPUs?

DGX agent

arXiv:2601.20115v3 Announce Type: replace-cross Abstract: As the role of modern Graphics Processing Units (GPUs) becomes increasingly essential for several computing tasks, analyzing their past and cu

safetyarxiv-cs-ai
2 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space

DGX agent

arXiv:2603.01302v2 Announce Type: replace Abstract: Reinforcement learning in discrete-continuous hybrid action spaces presents fundamental challenges for robotic manipulation, where high-level task d

safetyarxiv-cs-ro
2 Jun 2026
Safety

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual…

DGX agent

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual return – IF it works out. Who’s in? Don’t worry about the r

safetygary-marcus--x
2 Jun 2026
Safety

Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry

DGX agent

arXiv:2606.01098v1 Announce Type: cross Abstract: Generative action policies based on diffusion or flow matching excel in behavior cloning, yet their iterative sampling is prohibitive for high-frequen

safetyarxiv-cs-ai
2 Jun 2026
Safety

Improving Visual Representation Alignment Generation with GRPO

DGX agent

arXiv:2606.00583v1 Announce Type: cross Abstract: Recent diffusion transformers have demonstrated strong image synthesis capabilities but remain inefficient to train due to weak alignment between gene

safetyarxiv-cs-ai
2 Jun 2026
Safety

Infeasible optimization problems and the hierarchical augmented Lagrangian method in imitation learning

DGX agent

arXiv:2606.00730v1 Announce Type: new Abstract: Imitation learning (IL) is an effective approach to train complex robotics policies. Recent works have introduced hard constraints into imitation-learni

safetyarxiv-cs-ro
2 Jun 2026
Safety

InFerActive: Interactive Tree-Based Exploration of LLM Sampling for Safety Evaluation

DGX agent

arXiv:2512.10234v2 Announce Type: replace-cross Abstract: Even LLMs that appear safe during evaluation can still produce harmful responses in deployment. Because stochastic sampling yields different r

safetyarxiv-cs-ai
2 Jun 2026
Safety

Interaction-Limited Safe Continuous-Time RL for Dynamical Medical Treatment

DGX agent

arXiv:2606.01051v1 Announce Type: new Abstract: Dynamic medical treatment requires deciding treatment intensity and intervention timing, while patient states evolve continuously and adverse events may

safetyarxiv-cs-lg
2 Jun 2026
Safety

Internalize the Temperature: On-Policy Self-Distillation as Policy Reheater for Reinforcement Learning

DGX agent

arXiv:2606.00755v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards improves the reasoning ability of large language models, but often suffers from entropy collapse, in whic

safetyarxiv-cs-cl
2 Jun 2026
Safety

Interpretability in Deep Time Series Models Demands Semantic Alignment

DGX agent

arXiv:2602.02239v2 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, exi

safetyarxiv-cs-lg
2 Jun 2026
Safety

Interpretable Modeling of Driver Attention Shifts with a Vision--Language Model

DGX agent

arXiv:2508.05852v2 Announce Type: replace Abstract: Driver gaze is commonly modeled as a spatial heatmap, but heatmaps alone are difficult for humans to interpret because they do not explain which roa

safetyarxiv-cs-cv
2 Jun 2026
Safety

Interpretable Multimodal Gesture Recognition for Drone and Mobile Robot Teleoperation via Log-Likelihood Ratio Fusion

DGX agent

arXiv:2602.23694v3 Announce Type: replace-cross Abstract: Human operators are still frequently exposed to hazardous environments such as disaster zones and industrial facilities, where intuitive and r

safetyarxiv-cs-ai
2 Jun 2026
Safety

Interpretable Policy Distillation for Power Grid Topology Control

DGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

safetyarxiv-cs-ai
2 Jun 2026
Safety

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

DGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

safetyarxiv-cs-lg
2 Jun 2026
Safety

IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation

DGX agent

arXiv:2601.00212v2 Announce Type: replace Abstract: Segmentation of vestibular schwannoma and cochlea from T2 MRI is clinically important yet annotation-intensive. Domain adaptation (DA) has been wide

safetyarxiv-cs-cv
2 Jun 2026
Safety

Inverse Depth Scaling From Most Layers Being Similar

DGX agent

arXiv:2602.05970v2 Announce Type: replace-cross Abstract: Neural scaling laws relate loss to model size in large language models (LLMs), yet depth and width may contribute to performance differently,

safetyarxiv-cs-ai
2 Jun 2026
Safety

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

DGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

safetyarxiv-cs-ai
2 Jun 2026
Safety

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

DGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

safetyarxiv-cs-lg
2 Jun 2026
Safety

Jailbreaking Multimodal Large Language Models using Multi-Clip Video

DGX agent

arXiv:2606.02111v1 Announce Type: cross Abstract: As multimodal large language models (MLLMs) have advanced to process video inputs, concerns have emerged about their potential for malicious misuse. P

safetyarxiv-cs-ai
2 Jun 2026
Safety

Joint Agent Memory and Exploration Learning via Novelty Signals

DGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Jointly Optimizing Debiased CTR and Uplift for Coupons Marketing: A Unified Causal Framework

DGX agent

arXiv:2602.12972v2 Announce Type: replace-cross Abstract: In online advertising, marketing interventions such as coupons introduce significant confounding bias into Click-Through Rate (CTR) prediction

safetyarxiv-cs-lg
2 Jun 2026
Safety

KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation

DGX agent

arXiv:2606.01282v1 Announce Type: new Abstract: Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cu

safetyarxiv-cs-cv
2 Jun 2026
Safety

KISS: Keeping it Simple and Slotted when Learning to Communicate over Wireless

DGX agent

arXiv:2606.00266v1 Announce Type: cross Abstract: A long-standing challenge in distributed wireless systems is ensuring efficient and fair random channel access. Existing solutions often address speci

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies

DGX agent

arXiv:2606.01151v1 Announce Type: new Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distri

safetyarxiv-cs-lg
2 Jun 2026
Safety

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

safetyarxiv-cs-ai
2 Jun 2026
Safety

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

DGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

safetyarxiv-cs-ai
2 Jun 2026
Safety

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

DGX agent

arXiv:2602.08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network. Once the denoiser is fixed, the samp

safetyarxiv-cs-lg
2 Jun 2026
Safety

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

DGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

safetyarxiv-cs-ai
2 Jun 2026
Safety

LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World

DGX agent

arXiv:2606.01458v1 Announce Type: new Abstract: Training vision-language-action (VLA) policies for humanoid loco-manipulation is constrained by the high cost and complexity of collecting human teleope

safetyarxiv-cs-ro
2 Jun 2026
Safety

Leyline: KV Cache Directives for Agentic Inference

DGX agent

arXiv:2606.01065v1 Announce Type: cross Abstract: Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only evict

safetyarxiv-cs-ai
2 Jun 2026
Safety

LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving

DGX agent

arXiv:2606.00372v1 Announce Type: new Abstract: Reliable object detection is critical for automated driving, yet even state-of-the-art detectors inevitably make errors that can compromise safety. Intr

safetyarxiv-cs-cv
2 Jun 2026
Safety

LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification

DGX agent

arXiv:2606.00647v1 Announce Type: cross Abstract: Detecting psychological defense mechanisms in conversational text remains a challenging clinical NLP problem. For the PsyDefDetect 2026 shared task (n

safetyarxiv-cs-ai
2 Jun 2026
Safety

LLM as a Meta-Judge: Synthetic Data for NLP Evaluation Metric Validation

DGX agent

arXiv:2603.09403v2 Announce Type: replace Abstract: Validating evaluation metrics for NLG typically relies on expensive and time-consuming human annotations, which predominantly exist only for English

safetyarxiv-cs-cl
2 Jun 2026
Safety

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

DGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

safetyarxiv-cs-ro
2 Jun 2026
Safety

Longitudinal Multimodal Sensing of Physical Activity and Well-Being in Older Adults

DGX agent

arXiv:2606.00345v1 Announce Type: new Abstract: Wearable and mobile sensing technologies enable continuous monitoring of human behavior and health in real-world settings. However, predictive modeling

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

DGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

safetyarxiv-cs-ai
2 Jun 2026
Safety

Looped Transformers with Layer Normalization Provably Learn the Power Method

DGX agent

arXiv:2606.00605v1 Announce Type: new Abstract: Transformers have achieved remarkable success across a wide range of applications, and a growing body of work suggests that part of their strength comes

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lost in Delusion: Examining LLM Safety Under User Delusions and Distress

DGX agent

arXiv:2606.00975v1 Announce Type: new Abstract: LLM chatbots increasingly serve as a first source of support for people in psychological distress, including those whose distress is entangled with delu

safetyarxiv-cs-cl
2 Jun 2026
Safety

Low-Pass Flow Matching

DGX agent

arXiv:2606.02177v1 Announce Type: new Abstract: Flow Matching typically relies on white noise sources, a choice often misaligned with the power spectra of natural data, which tend to decay with freque

safetyarxiv-cs-lg
2 Jun 2026
Safety

Make Mechanistic Interpretability Auditable: A Call to Develop Guidelines via Continuous Collaborative Reviewing

DGX agent

arXiv:2606.00033v1 Announce Type: cross Abstract: While mechanistic interpretability (MI) has produced important insights into neural network internals, the field has yet to establish a standardized s

safetyarxiv-cs-ai
2 Jun 2026
Safety

Markerless Augmented Reality Registration for Surgical Guidance: A Multi-Anatomy Clinical Accuracy Study

DGX agent

arXiv:2511.02086v2 Announce Type: replace Abstract: Purpose: In this paper, we develop and clinically evaluate a depth-only, markerless augmented reality (AR) registration pipeline on a head-mounted d

safetyarxiv-cs-cv
2 Jun 2026
Safety

Market-Based Replanning for Safety-Critical UAV Swarms in Search and Rescue Missions

DGX agent

arXiv:2606.01970v1 Announce Type: new Abstract: Reliable autonomous UAV swarms in Search and Rescue (SAR) missions require fault-tolerant coordination capable of sustaining operations despite agent de

safetyarxiv-cs-ro
2 Jun 2026
Safety

MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems

DGX agent

arXiv:2601.14230v2 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) are emerging as promising socio-collaborative companions for emotional and cognitive support. However, existing syst

safetyarxiv-cs-ai
2 Jun 2026
Safety

Massive Spikes in LLMs are Bias Vectors: Mechanistic Uncovering and Spike-Free Quantization

DGX agent

arXiv:2606.02288v1 Announce Type: new Abstract: Massive activation spikes in Large Language Models (LLMs) severely degrade quantization by stretching dynamic ranges. While prior hypotheses characteriz

safetyarxiv-cs-lg
2 Jun 2026
Safety

Maybe @ylecun can be automated after all, @SchmidhuberAI?

DGX agent

Gary Marcus poses a question to Yann LeCun and Jürgen Schmidhuber about whether automation of AI systems (possibly referring to AI development or reasoning processes) might be feasible, suggesting a d

safetygary-marcus--x
2 Jun 2026
Safety

Measurement Geometry and Design for Trustworthy Generative Inverse Problems

DGX agent

arXiv:2606.02309v1 Announce Type: cross Abstract: Generative models are increasingly used as priors for inverse problems, but their ability to produce realistic images creates a basic trust problem: a

safetyarxiv-cs-cv
2 Jun 2026
Safety

Measuring the Symmetry--Data Exchange Rate

DGX agent

arXiv:2606.01090v1 Announce Type: cross Abstract: Equivariance theory predicts that an architectural symmetry prior reduces sample complexity by a factor of |G|; this is widely cited but rarely measur

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning

DGX agent

arXiv:2606.01914v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) remain unreliable on spatial multiple-choice questions, and their failures are often attributed to poorly atten

safetyarxiv-cs-cl
2 Jun 2026
← Previous
1…122123124125126…267
Next →