AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success

DGX agent

arXiv:2601.18175v2 Announce Type: replace Abstract: A widely used technique for improving policies is success conditioning, in which one collects trajectories, identifies those that achieve a desired

safetyarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Test-time reward-guided alignment of language models by importance sampling on pre-logit space

DGX agent

arXiv:2510.26219v3 Announce Type: replace-cross Abstract: Test-time alignment of large language models (LLMs) attracts attention because fine-tuning of LLMs requires high computational costs. In this

safetyarxiv-cs-ai
4 Jun 2026
Safety

The Accountability Horizon: An Impossibility Theorem for Governing Human-Agent Collectives

DGX agent

arXiv:2604.07778v2 Announce Type: replace Abstract: Existing accountability frameworks for AI systems, legal, ethical, and regulatory, rest on a shared assumption: for any consequential outcome, at le

safetyarxiv-cs-ai
4 Jun 2026
Safety

The Digital Apprentice: A Framework for Human-Directed Agentic AI Development

DGX agent

arXiv:2606.04321v1 Announce Type: new Abstract: Agentic AI deployments face a recurring design tension: heavy human oversight limits scale, while broad autonomy outruns accountability. Neither posture

safetyarxiv-cs-ai
4 Jun 2026
Safety

The Invisible Lottery: How Subtle Cues Steer Algorithm Choice in LLM Code Generation

DGX agent

arXiv:2606.04057v1 Announce Type: cross Abstract: Large language models (LLMs) now generate substantial production code, often for tasks with multiple valid algorithmic solutions. Incidental prompt cu

safetyarxiv-cs-ai
4 Jun 2026
Safety

The Loss Is Not Enough: Sampling Conditions and Inductive Bias in Contrastive Representation Learning

DGX agent

arXiv:2606.04280v1 Announce Type: cross Abstract: Contrastive learning has become a leading paradigm for self-supervised representation learning, yet the conditions under which it recovers meaningful

safetyarxiv-cs-ai
4 Jun 2026
Safety

The Right Measure for Physics-Constrained Generation: A Co-Area Correction for Posterior-Consistent PDE Inverse Problems

DGX agent

arXiv:2606.04804v1 Announce Type: new Abstract: Generative models -- diffusion and flow matching -- are increasingly used to solve partial differential equation (PDE) inverse problems, enforcing the g

safetyarxiv-cs-lg
4 Jun 2026
Safety

Think Fast and Far: Long-Horizon Online POMDP Planning via Rapid State Sampling

DGX agent

arXiv:2606.04355v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are a general and principled framework for motion planning under uncertainty. Despite tremendous

safetyarxiv-cs-ro
4 Jun 2026
Safety

Toward Multi-Domain and Long-Tailed Quantization via Feature Alignment and Scaling

DGX agent

arXiv:2606.04920v1 Announce Type: cross Abstract: Quantizing deep neural networks is essential for efficient inference on resource-constrained devices. However, most existing methods are designed for

safetyarxiv-cs-cv
4 Jun 2026
Safety

Towards Pretraining Text Encoders for TabPFN

DGX agent

arXiv:2606.04876v1 Announce Type: new Abstract: Tabular foundation models, such as TabPFN, achieve strong performance on tabular datasets with numerical and categorical data, but do not natively handl

safetyarxiv-cs-lg
4 Jun 2026
Safety

Trace-Mediated Peak Bias: Bridging Temporal Credit Assignment and Cognitive Heuristics in Deep Reinforcement Learning

DGX agent

arXiv:2606.04735v1 Announce Type: cross Abstract: Temporal credit assignment is central to both biological and artificial intelligence, yet its interaction with non-linear function approximation is po

safetyarxiv-cs-ai
4 Jun 2026
Safety

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

DGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

safetyarxiv-cs-cv
4 Jun 2026
Safety

TransTac: Visuo-Tactile Modality Transition via Ultraviolet-Encoded Transparent Elastomers

DGX agent

arXiv:2606.04477v1 Announce Type: new Abstract: Vision-based tactile sensors (VBTS) recover high-resolution contact geometry but typically rely on opaque elastomer layers that prevent visual transpare

safetyarxiv-cs-ro
4 Jun 2026
Safety

U-Net-Accelerated Quality-Diversity Optimization for Climate-Adaptive Urban Layouts

DGX agent

arXiv:2606.04658v1 Announce Type: cross Abstract: Optimizing urban layouts for climate adaptation requires balancing building density with cold-air ventilation. Because physics-based climate simulatio

safetyarxiv-cs-lg
4 Jun 2026
Safety

UniFair: A unified fair clustering approach based on separation and compactness

DGX agent

arXiv:2606.04777v1 Announce Type: new Abstract: Clustering is increasingly used to support high-impact decisions, yet standard objectives such as k-means can produce clusterings that treat demographic

safetyarxiv-cs-lg
4 Jun 2026
Safety

Unlocking Proactivity in Task-Oriented Dialogue

DGX agent

arXiv:2605.22240v2 Announce Type: replace Abstract: Proactive task-oriented dialogue (TOD), such as outbound sales, demands a persuasive agent that actively probes the user's concerns and steers the c

safetyarxiv-cs-ai
4 Jun 2026
Safety

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

DGX agent

arXiv:2602.19101v2 Announce Type: replace-cross Abstract: Value alignment of Large Language Models (LLMs) requires us to empirically measure these models' actual, acquired representation of value. Amo

safetyarxiv-cs-ai
4 Jun 2026
Safety

VentAgent: When LLMs Learn to Breathe -- Multi-Objective Arbitration for ARDS Ventilation

DGX agent

arXiv:2606.04632v1 Announce Type: cross Abstract: Mechanical ventilation for Acute Respiratory Distress Syndrome (ARDS) requires balancing competing physiological goals, including oxygenation, lung pr

safetyarxiv-cs-cl
4 Jun 2026
Safety

VT-3DAD: Cross-Category 3D Anomaly Detection via Visual-Text Normal Space Alignment

DGX agent

arXiv:2606.04369v1 Announce Type: new Abstract: Few-shot cross-category 3D anomaly detection aims to determine whether an unknown point cloud belongs to a target normal category using only a few norma

safetyarxiv-cs-cv
4 Jun 2026
Safety

WAM-Nav: Asymmetric Latent World-Action Modeling for Unified Visual Navigation

DGX agent

arXiv:2606.04907v1 Announce Type: new Abstract: Visual navigation requires generating smooth and collision-free trajectories under complex geometric and physical constraints. Existing reactive policie

safetyarxiv-cs-ro
4 Jun 2026
Safety

What Type of Inference is Active Inference?

DGX agent

arXiv:2606.04935v1 Announce Type: new Abstract: Active inference casts decision-making as inference, with the Expected Free Energy (EFE) unifying goal-directed and information-seeking behavior. Recent

safetyarxiv-cs-ai
4 Jun 2026
Safety

When Both Layers Learn: Training Dynamics of Representing Linear Models via ReLU Networks

DGX agent

arXiv:2606.04476v1 Announce Type: new Abstract: In this paper, we study the gradient descent dynamics for jointly training both layers of a one-hidden-layer ReLU network to fit a linear target functio

safetyarxiv-cs-lg
4 Jun 2026
Safety

X4Val: Learning Neural Surrogates for Variance-Reduced Policy Evaluation

DGX agent

arXiv:2606.05159v1 Announce Type: new Abstract: Rigorous evaluation of learning-based robotic systems is an essential prerequisite for deployment. However, real-world test data is expensive to gather;

safetyarxiv-cs-ro
4 Jun 2026
Safety

ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data

DGX agent

arXiv:2603.09170v2 Announce Type: replace-cross Abstract: Achieving versatile and natural whole-body humanoid interaction control remains challenging due to the high cost of whole-body teleoperation d

safetyarxiv-cs-ai
4 Jun 2026
Safety

A Cartesian-3j Framework for Machine Learning Interatomic Potentials

DGX agent

arXiv:2512.16882v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have brought substantial gains in the extrapolation capability in computational chemistry. How

safetyarxiv-cs-lg
3 Jun 2026
Safety

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

DGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

safetyarxiv-cs-ai
3 Jun 2026
Safety

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

DGX agent

arXiv:2606.03948v1 Announce Type: new Abstract: We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy Alig

safetyarxiv-cs-cl
3 Jun 2026
Safety

Adaptive Causal Alignment for High-Confidence Adversarial Training

DGX agent

arXiv:2606.03925v1 Announce Type: new Abstract: Inverse adversarial training leverages high-confidence predictions to stabilize robust learning, yet we uncover a critical paradox: high confidence ofte

safetyarxiv-cs-cv
3 Jun 2026
Safety

AirDreamer: Generalist Drone Navigation with World Models

DGX agent

arXiv:2606.03252v1 Announce Type: cross Abstract: Navigating a drone in unseen and cluttered environments requires reliable generalization to unseen scene layouts and understanding of environmental st

safetyarxiv-cs-ai
3 Jun 2026
Safety

Aletheia: What Makes RLVR For Code Verifiers Tick?

DGX agent

arXiv:2601.12186v3 Announce Type: replace-cross Abstract: Multi-domain thinking verifiers trained via Reinforcement Learning with Verifiable Rewards (RLVR) are a cornerstone of modern post-training. H

safetyarxiv-cs-ai
3 Jun 2026
Safety

Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model Enhancement

DGX agent

arXiv:2412.01282v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) bring powerful understanding and reasoning capabilities to multimodal tasks. Meanwhile, the great need for capab

safetyarxiv-cs-ai
3 Jun 2026
Safety

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

DGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

safetyarxiv-cs-ai
3 Jun 2026
Safety

Alignment-Aware Decoding

DGX agent

arXiv:2509.26169v2 Announce Type: replace Abstract: Alignment of large language models remains a central challenge in natural language processing. Preference optimization has emerged as a popular and

safetyarxiv-cs-lg
3 Jun 2026
Safety

Are we really tilting? The mechanics of reward guidance in flow and diffusion models

DGX agent

arXiv:2606.02884v1 Announce Type: cross Abstract: Reward guidance algorithms steer a learned generative process toward the reward-tilted measure at inference time. While empirically powerful, these me

safetyarxiv-cs-ai
3 Jun 2026
Safety

ASAP: Exploiting the Satisficing Generalization Edge in Neural Combinatorial Optimization

DGX agent

arXiv:2501.17377v4 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) has emerged as a promising approach for solving Combinatorial Optimization (CO) problems, such as the 3D Bin

safetyarxiv-cs-ai
3 Jun 2026
Safety

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

DGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Attention Calibration for Position-Fair Dense Information Retrieval

DGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

safetyarxiv-cs-ai
3 Jun 2026
Safety

Bayesian Tensor Decomposition with Diffusion Model Prior

DGX agent

arXiv:2606.03212v1 Announce Type: new Abstract: Low-rank tensor decomposition (TD) is usually effective on clean, fully observed data, but it often degrades under severe missingness or noise. Low-rank

safetyarxiv-cs-lg
3 Jun 2026
Safety

Best of Both Worlds: Multimodal Reasoning and Generation via Unified Discrete Flow Matching

DGX agent

arXiv:2602.12221v2 Announce Type: replace Abstract: We propose UniDFlow, a unified discrete flow-matching framework for multimodal understanding, generation, and editing. It decouples understanding an

safetyarxiv-cs-cv
3 Jun 2026
Safety

Bionic Human-Motion Style Transfer for Physically Executable Whole-Body Control of Humanoid Robots

DGX agent

arXiv:2606.03536v1 Announce Type: new Abstract: Expressive whole-body motion is important for humanoid robots operating in human environments, where robots are expected to move stably while presenting

safetyarxiv-cs-ro
3 Jun 2026
Safety

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

DGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

safetyarxiv-cs-ai
3 Jun 2026
Safety

Breaking the Self-Confirming Loop: Diagnosing and Mitigating Systemic Reward Bias in Self-Rewarding RL

DGX agent

arXiv:2510.08977v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) efficiently scales the reasoning ability of large language models (LLMs) but is bottlene

safetyarxiv-cs-cl
3 Jun 2026
Safety

Brief Announcement: Generative Markov Model for Distributed Computing Systems

DGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

safetyarxiv-cs-ai
3 Jun 2026
Safety

Building Better Activation Oracles

DGX agent

arXiv:2606.02609v1 Announce Type: cross Abstract: Activation Oracles (AOs) are promising methods for interpreting residual stream activations. However, current AOs face important issues, such as hallu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Coherence Maximization Improves Pluralistic Alignment

DGX agent

arXiv:2606.03110v1 Announce Type: new Abstract: Aligning AI systems with diverse human values requires value specifications grounded in concrete examples, but generating such examples without extensiv

safetyarxiv-cs-cl
3 Jun 2026
Safety

Collab-REC: An LLM-based Agentic Framework for Balancing Recommendations in Tourism

DGX agent

arXiv:2508.15030v5 Announce Type: replace Abstract: We propose COLLAB-REC, a multi-agent framework designed to counteract popularity bias and improve diversity in tourism recommendations. In our setup

safetyarxiv-cs-ai
3 Jun 2026
Safety

Consistency Training Can Entrench Misalignment

DGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

safetyarxiv-cs-ai
3 Jun 2026
Safety

ConTrack: Constrained Hand Motion Tracking with Adaptive Trade-off Control

DGX agent

arXiv:2606.03177v1 Announce Type: new Abstract: Human demonstrations provide strong priors for robot manipulation, yet it is non-trivial to transfer them to execute on real robots due to the kinematic

safetyarxiv-cs-ro
3 Jun 2026
← Previous
1…137138139140141…260
Next →