AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Update-Free On-Policy Steering via Verifiers

DGX agent

arXiv:2603.10282v2 Announce Type: replace Abstract: In recent years, Behavior Cloning (BC) has become one of the most prevalent methods for learning manipulation from human demonstrations. Despite the

safetyarxiv-cs-ro
2 Jun 2026
Safety

Update Opacity: Epistemic Accessibility and Governance Under AI System Change

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.00037v1 Announce Type: cross Abstract: Machine learning models embedded in deployed AI systems are routinely updated to maintain correct functioning over time. Yet such updates can generate

safetyarxiv-cs-ai
2 Jun 2026
Safety

V-LynX: Token Interface Alignment for Video+X LLMs

DGX agent

arXiv:2606.00508v1 Announce Type: cross Abstract: This study introduces an intriguing phenomenon in Video LLMs: rather than merely translating frames into textual embeddings, Video LLMs establish a co

safetyarxiv-cs-ai
2 Jun 2026
Safety

Value-Free Policy Optimization via Reward Partitioning

DGX agent

arXiv:2506.13702v4 Announce Type: replace-cross Abstract: Single-trajectory preference optimization methods learn from datasets of ((prompt, response, reward)) tuples, offering a practical alternative

safetyarxiv-cs-ai
2 Jun 2026
Safety

VERA: Variational Inference Framework for Jailbreaking Large Language Models

DGX agent

arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabi

safetyarxiv-cs-cl
2 Jun 2026
Safety

Visualizing definitional divergence in high-dimensional data by manifold alignment: Application to 3D right ventricular strain computations

DGX agent

arXiv:2501.12178v2 Announce Type: replace Abstract: Medical imaging studies often rely on a single sample per subject, assuming it is representative of their physiological traits. However, variations

safetyarxiv-cs-cv
2 Jun 2026
Safety

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models

DGX agent

arXiv:2601.03309v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models, which integrate pretrained large Vision-Language Models (VLM) into their policy backbone, are gaining sig

safetyarxiv-cs-ai
2 Jun 2026
Safety

Wavelet-Fusion Diffusion Model for Multimodal Brain MRI Synthesis with Modality and Metadata Conditioning

DGX agent

arXiv:2606.00689v1 Announce Type: new Abstract: Multimodal MRI provides complementary information for neuroimaging analysis, where different imaging modalities capture distinct anatomical, tissue, and

safetyarxiv-cs-cv
2 Jun 2026
Safety

Weak Critics Make Strong Learners: On-Policy Critique Distillation for Scalable Oversight

DGX agent

arXiv:2606.00424v1 Announce Type: new Abstract: As large language models become stronger, weak supervisors may fail to provide reliable labels, preferences, or final judgments for complex outputs, lim

safetyarxiv-cs-ai
2 Jun 2026
Safety

When Does Predictive Inverse Dynamics Outperform Behavior Cloning?

DGX agent

arXiv:2601.21718v2 Announce Type: replace-cross Abstract: Behavior cloning (BC) is a practical offline imitation learning method, but it often fails when expert demonstrations are limited. Recent work

safetyarxiv-cs-ai
2 Jun 2026
Safety

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models

DGX agent

arXiv:2606.01671v1 Announce Type: new Abstract: In the contemporary epoch of multilingual education, learning idioms provides a fascinating gateway towards creativity, cultural values, historical cont

safetyarxiv-cs-cl
2 Jun 2026
Safety

Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

DGX agent

arXiv:2511.05613v2 Announce Type: replace-cross Abstract: Foundation models are increasingly central to high-stakes AI systems, and governance frameworks now depend on evaluations to assess their risk

safetyarxiv-cs-ai
2 Jun 2026
Safety

World Models for Robotic Manipulation: A Survey

DGX agent

arXiv:2606.00113v1 Announce Type: new Abstract: Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world mode

safetyarxiv-cs-ro
2 Jun 2026
Safety

World-Task Factorization for Robot Learning

DGX agent

arXiv:2606.02027v1 Announce Type: cross Abstract: Robot learning must produce policies that generalize to new combinations of constraints, teammates, and environments. To achieve this, we must structu

safetyarxiv-cs-lg
2 Jun 2026
Safety

A hitchhiker's guide to Poisson gradient estimation

DGX agent

arXiv:2602.03896v2 Announce Type: replace-cross Abstract: Poisson-distributed latent variable models are widely used in computational neuroscience, but differentiating through discrete stochastic samp

safetyarxiv-cs-lg
1 Jun 2026
Safety

A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models

DGX agent

arXiv:2605.30843v1 Announce Type: new Abstract: In the forward reinforcement-learning problem, the reward is fixed and known; the learner is asked to find a good policy or value function. Here we turn

safetyarxiv-cs-lg
1 Jun 2026
Safety

A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI

DGX agent

arXiv:2605.31021v1 Announce Type: new Abstract: Current alignment paradigms for generative artificial intelligence rely predominantly on monolithic benchmarking frameworks that reduce the plurality of

safetyarxiv-cs-ai
1 Jun 2026
Safety

A Unified Framework for Gradient Aggregation in Multi-Objective Optimization

DGX agent

arXiv:2605.30452v1 Announce Type: cross Abstract: Many machine learning problems involve multiple inherent trade-offs that are best addressed by gradient-based multi-objective optimization (MOO) algor

safetyarxiv-cs-ai
1 Jun 2026
Safety

Active Timepoint Selection for Learning Measure-Valued Trajectories

DGX agent

arXiv:2605.30625v1 Announce Type: cross Abstract: Inferring continuous probability paths from sparse snapshots is a fundamental challenge in domains like single-cell biology, where high-fidelity data

safetyarxiv-cs-ai
1 Jun 2026
Safety

AI Loss of Control Incident Management: Response & Resilience

DGX agent

arXiv:2605.30406v1 Announce Type: cross Abstract: Recent research demonstrating AI systems exhibiting deception and shutdown resistance suggests that AI loss of control (LOC) is an urgent policy conce

safetyarxiv-cs-ai
1 Jun 2026
Safety

Annealed Softmax Greedy in Many-Armed Bayesian Bandits

DGX agent

arXiv:2605.31034v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and group-based policy optimization methods such as GRPO update a stochastic policy by sampling

safetyarxiv-cs-ai
1 Jun 2026
Safety

Annotations Are Not All You Need: A Cross-modal Knowledge Transfer Network for Unsupervised Temporal Sentence Grounding

DGX agent

arXiv:2605.30742v1 Announce Type: new Abstract: This paper addresses the task of temporal sentence grounding (TSG). Although many respectable works have made decent achievements in this important topi

safetyarxiv-cs-cv
1 Jun 2026
Safety

Are Full Rollouts Necessary for On-Policy Distillation?

DGX agent

arXiv:2605.31490v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense teacher feedback along rollouts generated by the student and has emerged as a promising post-training paradi

safetyarxiv-cs-cl
1 Jun 2026
Safety

BIAS-ID: A Framework for Analyzing Transformation Biases in AI-Generated Image Detectors

DGX agent

arXiv:2605.31153v1 Announce Type: new Abstract: Given the surge of harmful AI-generated imagery online, reliably distinguishing authentic images from generated ones has become an urgent research topic

safetyarxiv-cs-cv
1 Jun 2026
Safety

Biases in the Blind Spot: Detecting What LLMs Fail to Mention

DGX agent

arXiv:2602.10117v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often provide chain-of-thought (CoT) reasoning traces that appear plausible, but may hide internal biases. We cal

safetyarxiv-cs-ai
1 Jun 2026
Safety

BiSegMamba: Efficient Bidirectional Tri-Oriented Mamba for 3D Medical Image Segmentation

DGX agent

arXiv:2605.30972v1 Announce Type: new Abstract: Accurate 3D medical image segmentation requires both long-range volumetric context and fine boundary preservation. CNN-based methods have limited global

safetyarxiv-cs-cv
1 Jun 2026
Safety

Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models

DGX agent

arXiv:2510.11683v3 Announce Type: replace-cross Abstract: A key challenge in applying reinforcement learning (RL) to diffusion large language models (dLLMs) is the intractability of their likelihood f

safetyarxiv-cs-ai
1 Jun 2026
Safety

Breaking Information Cocoons: A Hyperbolic Framework for Balancing Exploration and Exploitation in Recommender Systems

DGX agent

arXiv:2411.13865v4 Announce Type: replace-cross Abstract: Modern recommender systems often create information cocoons, restricting users' exposure to diverse content. The central challenge is to balan

safetyarxiv-cs-ai
1 Jun 2026
Safety

Building Generalization Into Behavior Generation Via Adaptive Compositions of Regularities

DGX agent

arXiv:2605.31110v1 Announce Type: new Abstract: Generalization in robotics requires prior knowledge about how the world is structured, yet this structure changes from one situation to the next. This p

safetyarxiv-cs-ro
1 Jun 2026
Safety

Calibrated Uncertainty for Trustworthy Clinical Gait Analysis Using Probabilistic Multiview Markerless Motion Capture

DGX agent

arXiv:2601.22412v2 Announce Type: replace Abstract: Video-based human movement analysis holds potential for movement assessment in clinical practice and research. However, the clinical implementation

safetyarxiv-cs-cv
1 Jun 2026
Safety

Can Aerial VLA Models Cooperate? Evaluating Closed-Loop Air-Ground Coordination with CARLA-Air

DGX agent

arXiv:2605.31066v1 Announce Type: new Abstract: Recent aerial vision-language-action (VLA) models show promising single-UAV capabilities, such as tracking moving objects and navigating to language-spe

safetyarxiv-cs-ro
1 Jun 2026
Safety

Causal Evaluation of Membership Inference Attacks

DGX agent

arXiv:2602.02819v3 Announce Type: replace Abstract: Membership Inference Attacks (MIAs) aim to distinguish training points (members) from unseen data (non-members), and are widely used to quantify mem

safetyarxiv-cs-lg
1 Jun 2026
Safety

CellBRIDGE: Learning Cellular Trajectories via Interaction-Aware Alignment

DGX agent

arXiv:2605.30635v1 Announce Type: new Abstract: Inferring dynamics from population snapshots is a fundamental challenge in machine learning and biology. In scRNA-sequencing (scRNA-seq), destructive me

safetyarxiv-cs-lg
1 Jun 2026
Safety

COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Language Models

DGX agent

arXiv:2605.30641v1 Announce Type: cross Abstract: Large language models (LLMs) can reveal and amplify societal biases during chain-of-thought (CoT) generation. We present COFT (Chain of Fair Thought),

safetyarxiv-cs-ai
1 Jun 2026
Safety

ConSensus: Multi-Agent Collaboration for Multimodal Sensing

DGX agent

arXiv:2601.06453v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly grounded in sensor data to perceive and reason about human physiology and the physical world. However,

safetyarxiv-cs-ai
1 Jun 2026
Safety

Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion

DGX agent

arXiv:2605.31388v1 Announce Type: new Abstract: Multi-Objective Reinforcement Learning (MORL) extends standard RL by optimizing policies with respect to multiple, often conflicting, objectives. While

safetyarxiv-cs-lg
1 Jun 2026
Safety

Contextual Scalarisation Thompson Sampling for multi-objective decisions in public media

DGX agent

arXiv:2605.31291v1 Announce Type: cross Abstract: Recommender systems may operate under multiple, competing objectives. For example, audience reach, cultural values, public service mandate, and operat

safetyarxiv-cs-lg
1 Jun 2026
Safety

Cost-aware Stopping for Bayesian Optimization

DGX agent

arXiv:2507.12453v5 Announce Type: replace Abstract: In automated machine learning, scientific discovery, and other applications of Bayesian optimization, deciding when to stop evaluating expensive bla

safetyarxiv-cs-lg
1 Jun 2026
Safety

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation

DGX agent

arXiv:2501.01926v3 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have shown remarkable capabilities in visual-language understanding. Despite their success, LVLMs still s

safetyarxiv-cs-ai
1 Jun 2026
Safety

Detect in Any Scene: An Agentic Framework for Object Detection with Experience-Aware Reasoning

DGX agent

arXiv:2605.31174v1 Announce Type: new Abstract: Object detection in real-world scenarios remains challenging due to diverse image degradations and heterogeneous object distributions, which significant

safetyarxiv-cs-cv
1 Jun 2026
Safety

Diagnosing Failure Modes of Shared-State Collaboration in Resource-Constrained Visual Agents

DGX agent

arXiv:2605.31354v1 Announce Type: new Abstract: Modular visual reasoning systems increasingly rely on shared working memory for multi-step collaboration, yet the failure dynamics of intermediate state

safetyarxiv-cs-ai
1 Jun 2026
Safety

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

DGX agent

arXiv:2602.00521v2 Announce Type: replace Abstract: While LLM-as-a-Judge is widely used in automated evaluation, existing validation practices primarily operate at the level of observed outputs, offer

safetyarxiv-cs-ai
1 Jun 2026
Safety

Differentially Private Preference Data Synthesis for Large Language Model Alignment

DGX agent

arXiv:2605.30808v1 Announce Type: cross Abstract: Preference alignment is a crucial post-training step for large language models (LLMs) to ensure their outputs align with human values. However, post-t

safetyarxiv-cs-ai
1 Jun 2026
Safety

DISCO: Mitigating Bias in Deep Learning with Conditional Distance Correlation

DGX agent

arXiv:2506.11653v3 Announce Type: replace-cross Abstract: Dataset bias often leads deep learning models to exploit spurious correlations instead of task-relevant signals. We introduce the Standard Ant

safetyarxiv-cs-ai
1 Jun 2026
Safety

Distilling LLM Feedback for Lean Theorem Proving

DGX agent

arXiv:2605.30861v1 Announce Type: new Abstract: Post-training for reasoning models typically combines supervised fine-tuning with reinforcement learning from verifiable rewards, most commonly with GRP

safetyarxiv-cs-ai
1 Jun 2026
Safety

DiTTo: Scalable Order-aware All-in-One Image Restoration Agent

DGX agent

arXiv:2605.30915v1 Announce Type: new Abstract: Real-world images rarely suffer from a single degradation, and the order in which degradations are removed substantially affects the final restoration q

safetyarxiv-cs-cv
1 Jun 2026
Safety

DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation with SpeechLLMs

DGX agent

arXiv:2605.31432v1 Announce Type: cross Abstract: Simultaneous speech-to-text translation (SimulST) generates translations while speech is still unfolding, requiring a streaming policy that decides wh

safetyarxiv-cs-ai
1 Jun 2026
Safety

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization

DGX agent

arXiv:2605.31455v1 Announce Type: cross Abstract: Large language models are increasingly deployed in multi-turn interactive settings where users or environments can iteratively provide lightweight fee

safetyarxiv-cs-cl
1 Jun 2026
← Previous
1…146147148149150…260
Next →