AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Interpretable Policy Distillation for Power Grid Topology Control

DGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

safetyarxiv-cs-ai
2 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

DGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

safetyarxiv-cs-lg
2 Jun 2026
Safety

IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation

DGX agent

arXiv:2601.00212v2 Announce Type: replace Abstract: Segmentation of vestibular schwannoma and cochlea from T2 MRI is clinically important yet annotation-intensive. Domain adaptation (DA) has been wide

safetyarxiv-cs-cv
2 Jun 2026
Safety

Inverse Depth Scaling From Most Layers Being Similar

DGX agent

arXiv:2602.05970v2 Announce Type: replace-cross Abstract: Neural scaling laws relate loss to model size in large language models (LLMs), yet depth and width may contribute to performance differently,

safetyarxiv-cs-ai
2 Jun 2026
Safety

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

DGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

safetyarxiv-cs-ai
2 Jun 2026
Safety

Joint Agent Memory and Exploration Learning via Novelty Signals

DGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Jointly Optimizing Debiased CTR and Uplift for Coupons Marketing: A Unified Causal Framework

DGX agent

arXiv:2602.12972v2 Announce Type: replace-cross Abstract: In online advertising, marketing interventions such as coupons introduce significant confounding bias into Click-Through Rate (CTR) prediction

safetyarxiv-cs-lg
2 Jun 2026
Safety

KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation

DGX agent

arXiv:2606.01282v1 Announce Type: new Abstract: Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cu

safetyarxiv-cs-cv
2 Jun 2026
Safety

KISS: Keeping it Simple and Slotted when Learning to Communicate over Wireless

DGX agent

arXiv:2606.00266v1 Announce Type: cross Abstract: A long-standing challenge in distributed wireless systems is ensuring efficient and fair random channel access. Existing solutions often address speci

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies

DGX agent

arXiv:2606.01151v1 Announce Type: new Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distri

safetyarxiv-cs-lg
2 Jun 2026
Safety

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

safetyarxiv-cs-ai
2 Jun 2026
Safety

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

DGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

safetyarxiv-cs-ai
2 Jun 2026
Safety

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

DGX agent

arXiv:2602.08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network. Once the denoiser is fixed, the samp

safetyarxiv-cs-lg
2 Jun 2026
Safety

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

DGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

safetyarxiv-cs-ai
2 Jun 2026
Safety

LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World

DGX agent

arXiv:2606.01458v1 Announce Type: new Abstract: Training vision-language-action (VLA) policies for humanoid loco-manipulation is constrained by the high cost and complexity of collecting human teleope

safetyarxiv-cs-ro
2 Jun 2026
Safety

Leyline: KV Cache Directives for Agentic Inference

DGX agent

arXiv:2606.01065v1 Announce Type: cross Abstract: Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only evict

safetyarxiv-cs-ai
2 Jun 2026
Safety

LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification

DGX agent

arXiv:2606.00647v1 Announce Type: cross Abstract: Detecting psychological defense mechanisms in conversational text remains a challenging clinical NLP problem. For the PsyDefDetect 2026 shared task (n

safetyarxiv-cs-ai
2 Jun 2026
Safety

LLM as a Meta-Judge: Synthetic Data for NLP Evaluation Metric Validation

DGX agent

arXiv:2603.09403v2 Announce Type: replace Abstract: Validating evaluation metrics for NLG typically relies on expensive and time-consuming human annotations, which predominantly exist only for English

safetyarxiv-cs-cl
2 Jun 2026
Safety

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

DGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

safetyarxiv-cs-ro
2 Jun 2026
Safety

Longitudinal Multimodal Sensing of Physical Activity and Well-Being in Older Adults

DGX agent

arXiv:2606.00345v1 Announce Type: new Abstract: Wearable and mobile sensing technologies enable continuous monitoring of human behavior and health in real-world settings. However, predictive modeling

safetyarxiv-cs-lg
2 Jun 2026
Safety

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

DGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

safetyarxiv-cs-ai
2 Jun 2026
Safety

Looped Transformers with Layer Normalization Provably Learn the Power Method

DGX agent

arXiv:2606.00605v1 Announce Type: new Abstract: Transformers have achieved remarkable success across a wide range of applications, and a growing body of work suggests that part of their strength comes

safetyarxiv-cs-lg
2 Jun 2026
Safety

Low-Pass Flow Matching

DGX agent

arXiv:2606.02177v1 Announce Type: new Abstract: Flow Matching typically relies on white noise sources, a choice often misaligned with the power spectra of natural data, which tend to decay with freque

safetyarxiv-cs-lg
2 Jun 2026
Safety

Markerless Augmented Reality Registration for Surgical Guidance: A Multi-Anatomy Clinical Accuracy Study

DGX agent

arXiv:2511.02086v2 Announce Type: replace Abstract: Purpose: In this paper, we develop and clinically evaluate a depth-only, markerless augmented reality (AR) registration pipeline on a head-mounted d

safetyarxiv-cs-cv
2 Jun 2026
Safety

MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems

DGX agent

arXiv:2601.14230v2 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) are emerging as promising socio-collaborative companions for emotional and cognitive support. However, existing syst

safetyarxiv-cs-ai
2 Jun 2026
Safety

Massive Spikes in LLMs are Bias Vectors: Mechanistic Uncovering and Spike-Free Quantization

DGX agent

arXiv:2606.02288v1 Announce Type: new Abstract: Massive activation spikes in Large Language Models (LLMs) severely degrade quantization by stretching dynamic ranges. While prior hypotheses characteriz

safetyarxiv-cs-lg
2 Jun 2026
Safety

Measurement Geometry and Design for Trustworthy Generative Inverse Problems

DGX agent

arXiv:2606.02309v1 Announce Type: cross Abstract: Generative models are increasingly used as priors for inverse problems, but their ability to produce realistic images creates a basic trust problem: a

safetyarxiv-cs-cv
2 Jun 2026
Safety

Measuring the Symmetry--Data Exchange Rate

DGX agent

arXiv:2606.01090v1 Announce Type: cross Abstract: Equivariance theory predicts that an architectural symmetry prior reduces sample complexity by a factor of |G|; this is widely cited but rarely measur

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning

DGX agent

arXiv:2606.01914v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) remain unreliable on spatial multiple-choice questions, and their failures are often attributed to poorly atten

safetyarxiv-cs-cl
2 Jun 2026
Safety

Meta-Black-Box Optimization with Ensemble Surrogate Modeling for Robustness-Accuracy Trade-off within SAEA

DGX agent

arXiv:2606.00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems. However, their reliance on rig

safetyarxiv-cs-lg
2 Jun 2026
Safety

Minimax-Optimal Policy Regret in Partially Observable Markov Games

DGX agent

arXiv:2606.02363v1 Announce Type: new Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov g

safetyarxiv-cs-lg
2 Jun 2026
Safety

Mitigating Bias in Locally Constrained Decoding via Tractable Proposals

DGX agent

arXiv:2606.01926v1 Announce Type: new Abstract: Generations from large language models often fail to conform to desired constraints such as JSON schema. Existing locally constrained decoding (LCD) app

safetyarxiv-cs-cl
2 Jun 2026
Safety

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

DGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

safetyarxiv-cs-ai
2 Jun 2026
Safety

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

DGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

safetyarxiv-cs-ai
2 Jun 2026
Safety

Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment

DGX agent

arXiv:2606.02198v1 Announce Type: new Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce differen

safetyarxiv-cs-lg
2 Jun 2026
Safety

MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts

DGX agent

arXiv:2606.00844v1 Announce Type: cross Abstract: Bounding-box regression is a fundamental component of object detection, playing a critical role in precise object localization. Existing Intersection-

safetyarxiv-cs-ai
2 Jun 2026
Safety

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

DGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

safetyarxiv-cs-ai
2 Jun 2026
Safety

Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection

DGX agent

arXiv:2606.02352v1 Announce Type: new Abstract: Robust self-supervised learning of multi-modal video representations is critical for real-world applications such as driver distraction detection, where

safetyarxiv-cs-cv
2 Jun 2026
Safety

Multi-Objective Reference-Aligned Machine Unlearning

DGX agent

arXiv:2606.00399v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training samples while preserving the model's utility. Existing single-objective approaches,

safetyarxiv-cs-lg
2 Jun 2026
Safety

MURMUR: An Efficient Inference System for Long-Form ASR

DGX agent

arXiv:2606.01483v1 Announce Type: cross Abstract: Long-form automatic speech recognition (ASR) requires both high accuracy and low latency, but existing systems force a trade-off between the two. Chun

safetyarxiv-cs-ai
2 Jun 2026
Safety

MViewRouter: Internalizing Geometric Equivariance via Multi-view Alternating Attention for Combinatorial Routing

DGX agent

arXiv:2606.01084v1 Announce Type: cross Abstract: Combinatorial routing problems such as the Traveling Salesman Problem (TSP) and the Capacitated Vehicle Routing Problem (CVRP) are fundamental NP-hard

safetyarxiv-cs-ai
2 Jun 2026
Safety

MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding

DGX agent

arXiv:2606.00174v1 Announce Type: cross Abstract: Electromyography (EMG) directly reflects muscle activation and is a key sensing modality for gesture recognition, prosthetic control, and wearable int

safetyarxiv-cs-ai
2 Jun 2026
Safety

NDPP-Grasp: Non-Differentiable Physical Plausibility Constraint-Guided Task-Oriented Dexterous Grasp Generation

DGX agent

arXiv:2606.02432v1 Announce Type: new Abstract: Task-oriented dexterous grasp generation aims to produce dexterous grasp poses that are both physically plausible and functionally suitable for specifie

safetyarxiv-cs-ro
2 Jun 2026
Safety

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

DGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

safetyarxiv-cs-ai
2 Jun 2026
Safety

Non-Uniform Noise-to-Signal Ratio in the REINFORCE Policy-Gradient Estimator

DGX agent

arXiv:2602.01460v3 Announce Type: replace-cross Abstract: Policy-gradient methods are widely used in reinforcement learning, yet training often becomes unstable or slows down as learning progresses. W

safetyarxiv-cs-lg
2 Jun 2026
Safety

ObjEmbed: Towards Universal Multimodal Object Embeddings

DGX agent

arXiv:2602.01753v3 Announce Type: replace Abstract: Aligning objects with corresponding textual descriptions is a fundamental challenge and a realistic requirement in vision-language understanding. Wh

safetyarxiv-cs-cv
2 Jun 2026
Safety

Off-Policy Learning in Large Action Spaces: Optimization Matters More Than Estimation

DGX agent

arXiv:2509.03456v2 Announce Type: replace-cross Abstract: Off-policy evaluation (OPE) and off-policy learning (OPL) are foundational for decision-making in offline contextual bandits. Recent advances

safetyarxiv-cs-lg
2 Jun 2026
Safety

On Effectiveness and Efficiency of Agentic Tool-calling and RL Training

DGX agent

arXiv:2606.00135v1 Announce Type: cross Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This pa

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…143144145146147…260
Next →