AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence

DGX agent

arXiv:2605.16291v1 Announce Type: cross Abstract: With the growing adoption of AI systems, reasoning about how society can exert control over AI becomes an increasingly urgent problem. Existing work o

safetyarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

AIM: Adversarial Information Masking for Faithfulness Evaluation of Saliency Maps

DGX agent

arXiv:2605.16905v1 Announce Type: cross Abstract: Post-hoc saliency methods are widely used to interpret deep neural networks, but their faithfulness is difficult to evaluate reliably. Existing evalua

safetyarxiv-cs-cv
19 May 2026
Safety

Algorithmic Cultivation: How Social Media Feeds Shape User Language

DGX agent

arXiv:2605.17010v1 Announce Type: cross Abstract: Algorithmic feeds have become primary environments for encountering information online, yet while they shape what people see, less is known about how

safetyarxiv-cs-ai
19 May 2026
Safety

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey

DGX agent

arXiv:2505.17352v2 Announce Type: replace Abstract: Diffusion models have become a central paradigm for image and multimodal generation, yet their deployment raises persistent questions about alignmen

safetyarxiv-cs-cv
19 May 2026
Safety

Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework

DGX agent

arXiv:2605.16516v1 Announce Type: cross Abstract: Long-term interaction with LLM-based systems may produce alignment drift: a gradual process in which system outputs become less constrained by the use

safetyarxiv-cs-ai
19 May 2026
Safety

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering

DGX agent

arXiv:2605.17352v1 Announce Type: new Abstract: Despite substantial advances in large language models (LLMs), generating factually consistent responses for knowledge-intensive question answering remai

safetyarxiv-cs-cl
19 May 2026
Safety

AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment

DGX agent

arXiv:2605.18529v1 Announce Type: new Abstract: The alignment of Large Language Models (LLMs) for complex reasoning heavily relies on Reinforcement Learning with Verifiable Rewards (RLVR). However, st

safetyarxiv-cs-ai
19 May 2026
Safety

An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration

DGX agent

arXiv:2605.18648v1 Announce Type: cross Abstract: Central to human-aligned AI is understanding the benefits of human-elicited labels over synthetic alternatives. While human soft-labels improve calibr

safetyarxiv-cs-ai
19 May 2026
Safety

An Efficient Streaming Video Understanding Framework with Agentic Control

DGX agent

arXiv:2605.17921v1 Announce Type: new Abstract: Streaming video requires handling dynamic information density under strict latency budgets. Yet, existing methods typically employ static strategies, su

safetyarxiv-cs-cv
19 May 2026
Safety

An Empirical Study of Privacy Leakage Chains via Prompt Injection in Black-Box Chatbot Environments

DGX agent

arXiv:2605.18133v1 Announce Type: cross Abstract: LLM-based chatbot agents increasingly process user requests by combining natural-language reasoning with external tools such as web browsing. These ca

safetyarxiv-cs-ai
19 May 2026
Safety

AnchorDiff: Topology-Aware Masked Diffusion with Confidence-based Rewriting for Radiology Report Generation

DGX agent

arXiv:2605.17071v1 Announce Type: new Abstract: Radiology report generation (RRG) aims to automatically produce clinically accurate textual reports from medical images. Existing methods predominantly

safetyarxiv-cs-ai
19 May 2026
Safety

Anytime and Difficulty-Adaptive PAC-Bayes for Constrained Density-Ratio Network with Continual Learning Guarantees

DGX agent

arXiv:2605.17212v1 Announce Type: new Abstract: A unified framework for learning under covariate shift is presented, in which a constrained density-ratio network approximates the Radon-Nikodym derivat

safetyarxiv-cs-lg
19 May 2026
Safety

Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild

DGX agent

arXiv:2603.04727v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have demonstrated impressive general competence in video understanding, yet their reliability for rea

safetyarxiv-cs-ai
19 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Safety

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making

DGX agent

arXiv:2605.17228v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes domains such as clinical decision support and medical documentation. However, the

safetyarxiv-cs-cl
19 May 2026
Safety

As a Soviet historian who has spent years writing about the extreme, repressive control Soviet Communism exercised over its unfortunate citi…

DGX agent

As a Soviet historian who has spent years writing about the extreme, repressive control Soviet Communism exercised over its unfortunate citizens, I find it really hard to bring a similar accusation ag

safetyelon-musk--x
19 May 2026
Safety

Assessing Localization Technologies for Pedestrian Collision Avoidance

DGX agent

arXiv:2605.18295v1 Announce Type: new Abstract: Robust pedestrian safety is crucial to the next-generation of intelligent transportation systems. Such systems rely on active pedestrian localization an

safetyarxiv-cs-ro
19 May 2026
Safety

Assured autonomy: How operations research powers and orchestrates generative AI systems

DGX agent

arXiv:2512.23978v2 Announce Type: replace Abstract: Generative artificial intelligence (GenAI) is shifting from conversational assistants toward agentic systems -- autonomous decision-making systems t

safetyarxiv-cs-lg
19 May 2026
Safety

Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?

DGX agent

arXiv:2605.16354v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as automated evaluators of AI systems, including in high-stakes applications. In this role, LLMs ar

safetyarxiv-cs-ai
19 May 2026
Safety

AURORA: Contextual Orthogonalization for Geometric Representation Learning in Healthcare Foundation Models

DGX agent

arXiv:2605.17765v1 Announce Type: new Abstract: Recent healthcare foundation models have achieved strong predictive performance through large scale self supervised learning, yet their latent represent

safetyarxiv-cs-lg
19 May 2026
Safety

Automatic Generation of High-Performance RL Environments

DGX agent

arXiv:2603.12145v2 Announce Type: replace-cross Abstract: Translating complex reinforcement learning (RL) environments into high-performance implementations has traditionally required months of specia

safetyarxiv-cs-ai
19 May 2026
Safety

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment

DGX agent

arXiv:2605.17602v1 Announce Type: new Abstract: Aligning Text-to-Image (T2I) generation models with human preferences increasingly relies on image reward models that score or rank generated images acc

safetyarxiv-cs-ai
19 May 2026
Safety

Avoiding Structural Failure Modes in Tabular Fair SSL: Online Primal-Dual Allocation under Confidence Gating

DGX agent

arXiv:2605.16446v1 Announce Type: cross Abstract: Semi-supervised learning (SSL) enables prediction with limited labels, but high-stakes tabular applications (medical, credit, recidivism) require stat

safetyarxiv-cs-ai
19 May 2026
Safety

Benchmarking transferability of SSL pretraining to same and different modality segmentation tasks

DGX agent

arXiv:2605.18491v1 Announce Type: new Abstract: Methods: Nine SSL methods spanning four pretext-task families were pretrained from scratch using the same 10{,}412 3D CT scans (1.89~M 2D axial slices)

safetyarxiv-cs-cv
19 May 2026
Safety

Beyond Compliance: How AI Could Help Creative Writers by Refusing Them

DGX agent

arXiv:2605.16272v1 Announce Type: cross Abstract: Mainstream creativity support design prioritizes compliant AI for seamless writing interactions, but concerns over inappropriate AI reliance highlight

safetyarxiv-cs-ai
19 May 2026
Safety

Beyond Policy Optimization: A Data Curation Flywheel for Sparse-Reward Long-Horizon Planning

DGX agent

arXiv:2508.03018v2 Announce Type: replace Abstract: Large Language Reasoning Models have demonstrated remarkable success on static tasks, yet their application to multi-round agentic planning in inter

safetyarxiv-cs-ai
19 May 2026
Safety

Beyond RLHF: A Unified Theoretical Framework of Alignment

DGX agent

arXiv:2506.01523v2 Announce Type: replace Abstract: Alignment via reinforcement learning from human feedback (RLHF) has become the dominant paradigm for controlling the quality of outputs from large l

safetyarxiv-cs-lg
19 May 2026
Safety

Beyond Safety Filtering: Control Barrier Function-Informed Reinforcement Learning for Connected and Automated Vehicles

DGX agent

arXiv:2605.16894v1 Announce Type: new Abstract: Reinforcement Learning (RL) uses rewards to guide learning, yet reward design is typically hand-crafted using heuristics that can be difficult to tune.

safetyarxiv-cs-ro
19 May 2026
Safety

Beyond Scaling: Agents Are Heading to the Edge

DGX agent

arXiv:2605.18535v1 Announce Type: new Abstract: The bottleneck of useful agentic intelligence has shifted from compressing world knowledge into a single model to executing a coordinated system. This p

safetyarxiv-cs-lg
19 May 2026
Safety

Beyond the Final Actor: Modeling the Dual Roles of Creator and Editor for Fine-Grained LLM-Generated Text Detection

DGX agent

arXiv:2604.04932v3 Announce Type: replace Abstract: The misuse of large language models (LLMs) requires precise detection of synthetic text. Existing works mainly follow binary or ternary classificati

safetyarxiv-cs-cl
19 May 2026
Safety

Beyond Transcripts: Iterative Peer-Editing with Audio Unlocks High-Quality Human Summaries of Conversational Speech

DGX agent

arXiv:2605.17652v1 Announce Type: new Abstract: There are not enough established benchmarks for the task fo speech summarization. Creating new benchmarks demands human annotation, as LLMs could embed

safetyarxiv-cs-cl
19 May 2026
Safety

BIDO: A Biometric Identity Online Authentication Framework

DGX agent

arXiv:2605.16908v1 Announce Type: cross Abstract: Security systems demand continuous, cryptograph- ically robust identity verification without requiring subjects to carry physical tokens, smart cards,

safetyarxiv-cs-cv
19 May 2026
Safety

Body-Grounded Perspective Formation and Conative Attunement in Artificial Agents

DGX agent

arXiv:2605.16728v1 Announce Type: new Abstract: This paper proposes a minimal architecture for body-grounded perspective formation in artificial agents. Extending prior work, the model introduces an i

safetyarxiv-cs-ai
19 May 2026
Safety

Bridging the Intention-Expression Gap: Aligning Multi-Dimensional Preferences via Hierarchical Relevance Feedback in Text-to-Image Diffusion

DGX agent

arXiv:2603.14936v3 Announce Type: replace Abstract: Users often possess a clear visual intent but struggle to articulate it precisely in language. This intention-expression gap makes aligning generate

safetyarxiv-cs-cv
19 May 2026
Safety

Building Reliable Arithmetic Multipliers Under NBTI Aging and Process Variations

DGX agent

arXiv:2605.18444v1 Announce Type: cross Abstract: Hardware aging poses a significant challenge for integrated circuits (ICs), leading to performance degradation and eventual failure. In this work, we

safetyarxiv-cs-ai
19 May 2026
Safety

Canonical Regularisation of Wide Feature-Learning Neural Networks

DGX agent

arXiv:2605.18180v1 Announce Type: cross Abstract: Wide neural networks in the feature-learning regime drive modern deep learning, and yet they remain far less studied than their kernel-regime counterp

safetyarxiv-cs-lg
19 May 2026
Safety

CatalyticMLLM: A Graph-Text Multimodal Large Language Model for Catalytic Materials

DGX agent

arXiv:2605.17254v1 Announce Type: new Abstract: Property prediction and inverse structural design of catalytic materials are typically modeled as two independent tasks: the former predicts target prop

safetyarxiv-cs-ai
19 May 2026
Safety

ChartDesign: Towards LLM Designer of Data Visualization

DGX agent

arXiv:2605.16274v1 Announce Type: cross Abstract: Charts are the dominant medium for visualizing data, discovering patterns and trends, and communicating data driven insights, yet designing them still

safetyarxiv-cs-ai
19 May 2026
Safety

ChemVA: Advancing Large Language Models on Chemical Reaction Diagrams Understanding

DGX agent

arXiv:2605.17214v1 Announce Type: new Abstract: While Large Language Models (LLMs) have revolutionized scientific text processing, they exhibit a significant capability gap when interpreting chemical

safetyarxiv-cs-ai
19 May 2026
Safety

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks

DGX agent

arXiv:2605.17458v1 Announce Type: new Abstract: Text classification models are typically trained via supervised fine-tuning (SFT). However, SFT essentially performs behavior cloning from instance-wise

safetyarxiv-cs-lg
19 May 2026
Safety

Code as Agent Harness

DGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

safetyarxiv-cs-ai
19 May 2026
Safety

CodeBind: Decoupled Representation Learning for Multimodal Alignment with Unified Compositional Codebook

DGX agent

arXiv:2605.18257v1 Announce Type: cross Abstract: Multimodal representation alignment is pivotal for large language models and robotics. Traditional methods are often hindered by cross-modal informati

safetyarxiv-cs-ai
19 May 2026
Safety

Collaborative Learning for Semi-Supervised LiDAR Semantic Segmentation

DGX agent

arXiv:2605.17135v1 Announce Type: new Abstract: Annotating large-scale LiDAR point clouds for 3D semantic segmentation is costly and time-consuming, which motivates the use of semi-supervised learning

safetyarxiv-cs-cv
19 May 2026
Safety

COLSON: Controllable Learning-Based Social Navigation via Diffusion-Based Reinforcement Learning

DGX agent

arXiv:2503.13934v2 Announce Type: replace-cross Abstract: Mobile robot navigation in dynamic environments with pedestrian traffic is a key challenge in the development of autonomous mobile service rob

safetyarxiv-cs-ai
19 May 2026
Safety

Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression

DGX agent

arXiv:2605.17304v1 Announce Type: cross Abstract: LLM context is not just tokens; it is a set of commitments. Long-running conversations accumulate goals, constraints, decisions, preferences, tool res

safetyarxiv-cs-cl
19 May 2026
Safety

Confidence-Gated Robot Autonomy: When Does Uncertainty Actually Help?

DGX agent

arXiv:2605.18045v1 Announce Type: cross Abstract: Robotic systems often use predictive uncertainty to decide whether to act autonomously or defer to a fallback policy. In threshold-gated autonomy, unc

safetyarxiv-cs-ai
19 May 2026
Safety

Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics

DGX agent

arXiv:2605.16300v1 Announce Type: cross Abstract: Robotic systems are moving from isolated platforms to interconnected multi-agent ecosystems that operate in human environments. This shift raises a go

safetyarxiv-cs-ai
19 May 2026
Safety

Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States

DGX agent

arXiv:2605.17144v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models leverage powerful perceptual priors from web-scale Vision-Language Model (VLM) pre-training, yet they remain surpr

safetyarxiv-cs-ai
19 May 2026
← Previous
1…164165166167168…267
Next →