AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

How Much Progress Has There Been in NVIDIA Datacenter GPUs?

DGX agent

arXiv:2601.20115v3 Announce Type: replace-cross Abstract: As the role of modern Graphics Processing Units (GPUs) becomes increasingly essential for several computing tasks, analyzing their past and cu

safetyarxiv-cs-ai
2 Jun 2026
Safety

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2603.01302v2 Announce Type: replace Abstract: Reinforcement learning in discrete-continuous hybrid action spaces presents fundamental challenges for robotic manipulation, where high-level task d

safetyarxiv-cs-ro
2 Jun 2026
Safety

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual…

DGX agent

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual return – IF it works out. Who’s in? Don’t worry about the r

safetygary-marcus--x
2 Jun 2026
Safety

Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry

DGX agent

arXiv:2606.01098v1 Announce Type: cross Abstract: Generative action policies based on diffusion or flow matching excel in behavior cloning, yet their iterative sampling is prohibitive for high-frequen

safetyarxiv-cs-ai
2 Jun 2026
Safety

Improving Visual Representation Alignment Generation with GRPO

DGX agent

arXiv:2606.00583v1 Announce Type: cross Abstract: Recent diffusion transformers have demonstrated strong image synthesis capabilities but remain inefficient to train due to weak alignment between gene

safetyarxiv-cs-ai
2 Jun 2026
Safety

Internalize the Temperature: On-Policy Self-Distillation as Policy Reheater for Reinforcement Learning

DGX agent

arXiv:2606.00755v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards improves the reasoning ability of large language models, but often suffers from entropy collapse, in whic

safetyarxiv-cs-cl
2 Jun 2026
Safety

Interpretability in Deep Time Series Models Demands Semantic Alignment

DGX agent

arXiv:2602.02239v2 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, exi

safetyarxiv-cs-lg
2 Jun 2026
Safety

Interpretable Modeling of Driver Attention Shifts with a Vision--Language Model

DGX agent

arXiv:2508.05852v2 Announce Type: replace Abstract: Driver gaze is commonly modeled as a spatial heatmap, but heatmaps alone are difficult for humans to interpret because they do not explain which roa

safetyarxiv-cs-cv
2 Jun 2026
Safety

Interpretable Policy Distillation for Power Grid Topology Control

DGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

safetyarxiv-cs-ai
2 Jun 2026
Safety

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

DGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

safetyarxiv-cs-lg
2 Jun 2026
Safety

IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation

DGX agent

arXiv:2601.00212v2 Announce Type: replace Abstract: Segmentation of vestibular schwannoma and cochlea from T2 MRI is clinically important yet annotation-intensive. Domain adaptation (DA) has been wide

safetyarxiv-cs-cv
2 Jun 2026
Safety

Inverse Depth Scaling From Most Layers Being Similar

DGX agent

arXiv:2602.05970v2 Announce Type: replace-cross Abstract: Neural scaling laws relate loss to model size in large language models (LLMs), yet depth and width may contribute to performance differently,

safetyarxiv-cs-ai
2 Jun 2026
Safety

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

DGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

safetyarxiv-cs-ai
2 Jun 2026
Safety

Joint Agent Memory and Exploration Learning via Novelty Signals

DGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Jointly Optimizing Debiased CTR and Uplift for Coupons Marketing: A Unified Causal Framework

DGX agent

arXiv:2602.12972v2 Announce Type: replace-cross Abstract: In online advertising, marketing interventions such as coupons introduce significant confounding bias into Click-Through Rate (CTR) prediction

safetyarxiv-cs-lg
2 Jun 2026
Safety

KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation

DGX agent

arXiv:2606.01282v1 Announce Type: new Abstract: Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cu

safetyarxiv-cs-cv
2 Jun 2026
Safety

KISS: Keeping it Simple and Slotted when Learning to Communicate over Wireless

DGX agent

arXiv:2606.00266v1 Announce Type: cross Abstract: A long-standing challenge in distributed wireless systems is ensuring efficient and fair random channel access. Existing solutions often address speci

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies

DGX agent

arXiv:2606.01151v1 Announce Type: new Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distri

safetyarxiv-cs-lg
2 Jun 2026
Safety

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

safetyarxiv-cs-ai
2 Jun 2026
Safety

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

DGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

safetyarxiv-cs-ai
2 Jun 2026
Safety

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

DGX agent

arXiv:2602.08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network. Once the denoiser is fixed, the samp

safetyarxiv-cs-lg
2 Jun 2026
Safety

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

DGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

safetyarxiv-cs-ai
2 Jun 2026
Safety

LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World

DGX agent

arXiv:2606.01458v1 Announce Type: new Abstract: Training vision-language-action (VLA) policies for humanoid loco-manipulation is constrained by the high cost and complexity of collecting human teleope

safetyarxiv-cs-ro
2 Jun 2026
Safety

Leyline: KV Cache Directives for Agentic Inference

DGX agent

arXiv:2606.01065v1 Announce Type: cross Abstract: Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only evict

safetyarxiv-cs-ai
2 Jun 2026
Safety

LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification

DGX agent

arXiv:2606.00647v1 Announce Type: cross Abstract: Detecting psychological defense mechanisms in conversational text remains a challenging clinical NLP problem. For the PsyDefDetect 2026 shared task (n

safetyarxiv-cs-ai
2 Jun 2026
Safety

LLM as a Meta-Judge: Synthetic Data for NLP Evaluation Metric Validation

DGX agent

arXiv:2603.09403v2 Announce Type: replace Abstract: Validating evaluation metrics for NLG typically relies on expensive and time-consuming human annotations, which predominantly exist only for English

safetyarxiv-cs-cl
2 Jun 2026
Safety

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

DGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

safetyarxiv-cs-ro
2 Jun 2026
Safety

Longitudinal Multimodal Sensing of Physical Activity and Well-Being in Older Adults

DGX agent

arXiv:2606.00345v1 Announce Type: new Abstract: Wearable and mobile sensing technologies enable continuous monitoring of human behavior and health in real-world settings. However, predictive modeling

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

DGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

safetyarxiv-cs-ai
2 Jun 2026
Safety

Looped Transformers with Layer Normalization Provably Learn the Power Method

DGX agent

arXiv:2606.00605v1 Announce Type: new Abstract: Transformers have achieved remarkable success across a wide range of applications, and a growing body of work suggests that part of their strength comes

safetyarxiv-cs-lg
2 Jun 2026
Safety

Low-Pass Flow Matching

DGX agent

arXiv:2606.02177v1 Announce Type: new Abstract: Flow Matching typically relies on white noise sources, a choice often misaligned with the power spectra of natural data, which tend to decay with freque

safetyarxiv-cs-lg
2 Jun 2026
Safety

Markerless Augmented Reality Registration for Surgical Guidance: A Multi-Anatomy Clinical Accuracy Study

DGX agent

arXiv:2511.02086v2 Announce Type: replace Abstract: Purpose: In this paper, we develop and clinically evaluate a depth-only, markerless augmented reality (AR) registration pipeline on a head-mounted d

safetyarxiv-cs-cv
2 Jun 2026
Safety

MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems

DGX agent

arXiv:2601.14230v2 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) are emerging as promising socio-collaborative companions for emotional and cognitive support. However, existing syst

safetyarxiv-cs-ai
2 Jun 2026
Safety

Massive Spikes in LLMs are Bias Vectors: Mechanistic Uncovering and Spike-Free Quantization

DGX agent

arXiv:2606.02288v1 Announce Type: new Abstract: Massive activation spikes in Large Language Models (LLMs) severely degrade quantization by stretching dynamic ranges. While prior hypotheses characteriz

safetyarxiv-cs-lg
2 Jun 2026
Safety

Maybe @ylecun can be automated after all, @SchmidhuberAI?

DGX agent

Gary Marcus poses a question to Yann LeCun and Jürgen Schmidhuber about whether automation of AI systems (possibly referring to AI development or reasoning processes) might be feasible, suggesting a d

safetygary-marcus--x
2 Jun 2026
Safety

Measurement Geometry and Design for Trustworthy Generative Inverse Problems

DGX agent

arXiv:2606.02309v1 Announce Type: cross Abstract: Generative models are increasingly used as priors for inverse problems, but their ability to produce realistic images creates a basic trust problem: a

safetyarxiv-cs-cv
2 Jun 2026
Safety

Measuring the Symmetry--Data Exchange Rate

DGX agent

arXiv:2606.01090v1 Announce Type: cross Abstract: Equivariance theory predicts that an architectural symmetry prior reduces sample complexity by a factor of |G|; this is widely cited but rarely measur

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning

DGX agent

arXiv:2606.01914v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) remain unreliable on spatial multiple-choice questions, and their failures are often attributed to poorly atten

safetyarxiv-cs-cl
2 Jun 2026
Safety

Meta-Black-Box Optimization with Ensemble Surrogate Modeling for Robustness-Accuracy Trade-off within SAEA

DGX agent

arXiv:2606.00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems. However, their reliance on rig

safetyarxiv-cs-lg
2 Jun 2026
Safety

Minimax-Optimal Policy Regret in Partially Observable Markov Games

DGX agent

arXiv:2606.02363v1 Announce Type: new Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov g

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mitigating Bias in Locally Constrained Decoding via Tractable Proposals

DGX agent

arXiv:2606.01926v1 Announce Type: new Abstract: Generations from large language models often fail to conform to desired constraints such as JSON schema. Existing locally constrained decoding (LCD) app

safetyarxiv-cs-cl
2 Jun 2026
Safety

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

DGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

safetyarxiv-cs-ai
2 Jun 2026
Safety

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

DGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

safetyarxiv-cs-ai
2 Jun 2026
Safety

Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment

DGX agent

arXiv:2606.02198v1 Announce Type: new Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce differen

safetyarxiv-cs-lg
2 Jun 2026
Safety

MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts

DGX agent

arXiv:2606.00844v1 Announce Type: cross Abstract: Bounding-box regression is a fundamental component of object detection, playing a critical role in precise object localization. Existing Intersection-

safetyarxiv-cs-ai
2 Jun 2026
Safety

Morningstar: Get real, SpaceX just isn’t worth a trillion dollars, let alone two.

DGX agent

Gary Marcus argues that SpaceX's valuation is significantly inflated, contending that the company is not worth the trillion-dollar valuations that have been suggested. The critique appears to challeng

safetygary-marcus--x
2 Jun 2026
Safety

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

DGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

safetyarxiv-cs-ai
2 Jun 2026
Safety

Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection

DGX agent

arXiv:2606.02352v1 Announce Type: new Abstract: Robust self-supervised learning of multi-modal video representations is critical for real-world applications such as driver distraction detection, where

safetyarxiv-cs-cv
2 Jun 2026
← Previous
1…163164165166167…302
Next →