AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas

DGX agent

arXiv:2505.19212v2 Announce Type: replace Abstract: Recent advances in LLMs have enabled their use in complex agentic roles, involving decision-making with humans or other agents, making ethical align

safetyarxiv-cs-cl
27 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity

DGX agent

arXiv:2607.22218v1 Announce Type: new Abstract: Despite the growing use of large language models (LLMs) as creativity evaluators, evidence of their alignment with human evaluations remains mixed, rais

safetyarxiv-cs-cl
27 Jul 2026
Safety

A Framework for Reputation Aware Uninorm-driven Consensus Algorithms for Blockchain Networks

DGX agent

arXiv:2607.20700v1 Announce Type: cross Abstract: The operation of blockchain is governed by consensus algorithms (CA). Several consensus mechanisms require significant computational power, while othe

safetyarxiv-cs-ai
24 Jul 2026
Safety

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space

DGX agent

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to

safetyarxiv-cs-lg
24 Jul 2026
Safety

A Unified Moral-Value Dataset for Instruction Tuning

DGX agent

arXiv:2607.21279v1 Announce Type: new Abstract: Large language models (LLMs) have developed rapidly and become valuable tools in everyday life. However, how to align LLMs to a particular set of human

safetyarxiv-cs-cl
24 Jul 2026
Safety

Adaptive Depth Sparse Framework: Similarity-Driven Resource Allocation for Pre-Trained LLMs

DGX agent

arXiv:2607.21291v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong generation and reasoning performance, but the Transformer architecture incurs high inference cost. Existing

safetyarxiv-cs-cl
24 Jul 2026
Safety

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation

DGX agent

arXiv:2607.18042v2 Announce Type: replace-cross Abstract: End-to-end vision-language navigation (VLN) with causal vision-language models maps instructions and egocentric observations directly to actio

safetyarxiv-cs-ai
24 Jul 2026
Safety

Approximate Quantum State Preparation Through Proximal Policy Optimization

DGX agent

arXiv:2607.21121v1 Announce Type: cross Abstract: In this work, a quantum architecture search framework for approximate quantum state preparation (QSP) is proposed. QSP is a challenging task, since th

safetyarxiv-cs-lg
24 Jul 2026
Safety

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

DGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

safetyarxiv-cs-ai
24 Jul 2026
Safety

ASTRA-Net: Anatomy-Specific Transfer and Representation Alignment for Drug-Induced Sleep Endoscopy Segmentation

DGX agent

arXiv:2607.21370v1 Announce Type: new Abstract: Quantitative drug-induced sleep endoscopy (DISE) requires reliable airway boundaries at specific anatomical levels. Pixel-level DISE annotations are sca

safetyarxiv-cs-cv
24 Jul 2026
Safety

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

DGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

safetyarxiv-cs-ai
24 Jul 2026
Safety

Axolotl3D: a Unified Framework for Faithful 3D Shape Completion

DGX agent

arXiv:2607.20660v1 Announce Type: new Abstract: Recent 3D generative models produce high-quality geometry from a single image using large-scale priors and diffusion architectures. However, they assume

safetyarxiv-cs-cv
24 Jul 2026
Safety

Belief Propagation in LLM World Models: Measuring Strategic Information Bias with Prediction Markets

DGX agent

arXiv:2607.20441v1 Announce Type: new Abstract: Every information ecosystem produces beliefs that shape strategic decisions. Both human analysts and AI systems inherit the blind spots of their informa

safetyarxiv-cs-cl
24 Jul 2026
Safety

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

DGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

safetyarxiv-cs-ai
24 Jul 2026
Safety

Climate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobility

DGX agent

arXiv:2607.21444v1 Announce Type: cross Abstract: Reliable electric vehicle (EV) charging infrastructure is a cornerstone of sustainable, low-carbon cities, yet urban climate stress such as extreme he

safetyarxiv-cs-lg
24 Jul 2026
Safety

Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception

DGX agent

arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service of a contextually or experimentally induced goal.

safetyarxiv-cs-ai
24 Jul 2026
Safety

CSPF: A Constrained Shared-Private Fusion Method for Non-Verifiable Preference Evaluation

DGX agent

arXiv:2607.20862v1 Announce Type: new Abstract: At present, reliable evaluation of non-verifiable tasks remains challenging. Existing approaches often fail to adequately capture the diverse evaluative

safetyarxiv-cs-cl
24 Jul 2026
Safety

DINOde: Continuous Vision-Text Alignment for Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2607.21371v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) leverages textual semantics to segment objects beyond predefined categories. While the self-supervised mo

safetyarxiv-cs-ai
24 Jul 2026
Safety

Distribution-Alignment Bridge for Uncertainty-Aware Text-to-Video Retrieval

DGX agent

arXiv:2607.20984v1 Announce Type: new Abstract: This paper proposes the Distribution-Alignment Bridge (DAB), a framework that reconceptualizes text-to-video retrieval as a distribution alignment task

safetyarxiv-cs-cv
24 Jul 2026
Safety

DTIF: Robust Loop Closure Detection via Delaunay Triangle Topology in Complex Forests

DGX agent

arXiv:2607.21138v1 Announce Type: new Abstract: Accurate forest inventory and large-scale mapping are essential for ecosystem monitoring and sustainable forest management. Multiple low-cost edge platf

safetyarxiv-cs-cv
24 Jul 2026
Safety

DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers

DGX agent

arXiv:2508.21797v2 Announce Type: replace-cross Abstract: Industry 4.0's highly networked Machine Tool Controllers (MTCs) are prime targets for replay attacks that use outdated sensor data to manipula

safetyarxiv-cs-ai
24 Jul 2026
Safety

EmoAgent-R1: Towards Multimodal Emotion Understanding with Reinforcement Learning-based Dynamic Agent Specialization

DGX agent

arXiv:2607.21013v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance in multimodal emotion recognition (MER) tasks and lifted MER to a new leve

safetyarxiv-cs-ai
24 Jul 2026
Safety

EmoSpace: Immersive Affective Image Generation Guided by Fine-Grained Emotion Prototypes

DGX agent

arXiv:2602.11658v2 Announce Type: replace Abstract: Immersive affective content generation aims to create visually compelling VR imagery with controllable emotional nuance, yet existing methods typica

safetyarxiv-cs-cv
24 Jul 2026
Safety

Enhancing Explainable Cardiac Diagnosis with Guide-Grounded Multimodal LLMs

DGX agent

arXiv:2607.20814v1 Announce Type: new Abstract: The electrocardiogram (ECG) is a cornerstone of cardiac as- sessment, yet clinical deployment of deep learning models remains con- strained by limited i

safetyarxiv-cs-ai
24 Jul 2026
Safety

Environment-Aware Channel Inference via Cross-Modal Flow: From Multimodal Sensing to Wireless Channels

DGX agent

arXiv:2512.04966v2 Announce Type: replace-cross Abstract: Accurate channel state information (CSI) underpins reliable and efficient wireless communication. However, acquiring CSI via pilot estimation

safetyarxiv-cs-lg
24 Jul 2026
Safety

Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective

DGX agent

arXiv:2604.25077v2 Announce Type: replace Abstract: Weak-to-strong alignment offers a promising route to scalable supervision, but it can fail when a strong model becomes confidently wrong on examples

safetyarxiv-cs-ai
24 Jul 2026
Safety

EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL

DGX agent

arXiv:2607.20489v1 Announce Type: new Abstract: Text-to-SQL has advanced rapidly with large language models, but complex database queries still require reasoning beyond one-shot generation, including

safetyarxiv-cs-ai
24 Jul 2026
Safety

Expert Behavior Prior Reinforcement Learning

DGX agent

arXiv:2607.21302v1 Announce Type: new Abstract: Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by l

safetyarxiv-cs-ai
24 Jul 2026
Safety

Explainability Framework for Policy-Aware Autonomous Agents

DGX agent

arXiv:2607.21209v1 Announce Type: cross Abstract: In the field of Artificial Intelligence, an agent is a system which is able to autonomously make decisions in order to reach a desired goal. As these

safetyarxiv-cs-ai
24 Jul 2026
Safety

Explainable graph attention network for stress recognition (StressGAT) via differential action units

DGX agent

arXiv:2607.20819v1 Announce Type: new Abstract: Stress is a dynamic process characterized by significant individual variability in facial expression. Traditional architectures, such as Recurrent Neura

safetyarxiv-cs-cv
24 Jul 2026
Safety

FELT: Generating Tactile Signals from Vision for Visuo-Tactile Manipulation

DGX agent

arXiv:2607.20683v1 Announce Type: new Abstract: The sense of touch is central to manipulation, especially when vision is occluded or ambiguous. Although combining vision and touch improves manipulatio

safetyarxiv-cs-ro
24 Jul 2026
Safety

FORGE-plus: Force-Budgeted Recovery for Contact-Rich Assembly with a Frozen LLM Supervisor

DGX agent

arXiv:2607.21227v1 Announce Type: new Abstract: Force-conditioned reinforcement learning (RL) enables tight-clearance assembly under a commanded force ceiling, but practical deployment requires determ

safetyarxiv-cs-ro
24 Jul 2026
Safety

From Agent Failures to Text Policies: What Works and What Breaks

DGX agent

arXiv:2607.20668v1 Announce Type: cross Abstract: TextGrad improves language-model systems by revising text from feedback. Its core thesis is that natural-language feedback can act as a gradient for o

safetyarxiv-cs-ai
24 Jul 2026
Safety

From Static Bibliometrics to Dynamic Knowledge Graphs: An LLM-Powered Framework for Modernizing Science, Technology, and Innovation (STI) Analytics

DGX agent

arXiv:2607.21327v1 Announce Type: cross Abstract: Bibliometric indicators - citation counts, h-indexes, co-authorship networks - have long anchored science, technology, and innovation (STI) analytics,

safetyarxiv-cs-ai
24 Jul 2026
Safety

Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis

DGX agent

arXiv:2411.00631v2 Announce Type: replace-cross Abstract: This scoping review examines the relationship between Generative AI (GenAI) and agency in education, analyzing the literature available throug

safetyarxiv-cs-ai
24 Jul 2026
Safety

Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances

DGX agent

arXiv:2511.03354v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) is transforming bioinformatics by advancing genomics, proteomics, transcriptomics, structural biolo

safetyarxiv-cs-ai
24 Jul 2026
Safety

GLAM-SLAM: Real-time Gaussian Large-scale Mapping via Flow Densification and Spatial Decomposition

DGX agent

arXiv:2607.21416v1 Announce Type: cross Abstract: Existing Gaussian-splatting-based monocular Simultaneous Localization and Mapping (SLAM) systems are either tailored to short sequences, are not real-

safetyarxiv-cs-cv
24 Jul 2026
Safety

GroupVideo: Multi-Identity Customized Text-to-Video Generation

DGX agent

arXiv:2607.21027v1 Announce Type: new Abstract: Current identity customized video generation methodologies are predominantly limited to single-identity scenarios, as the lack of explicit identity sepa

safetyarxiv-cs-cv
24 Jul 2026
Safety

GuidedAttention: Interpretable and Correctable Visual Attention for OOD-Robust Robot Manipulation via Imitation Learning

DGX agent

arXiv:2607.21049v1 Announce Type: new Abstract: End-to-end visuomotor policies provide little opportunity for humans to understand or correct the policy's visual attention. We propose GuidedAttention,

safetyarxiv-cs-ro
24 Jul 2026
Safety

Inference-Time Scaling of Diffusion Models via Progressive Seed Pruning

DGX agent

arXiv:2607.21591v1 Announce Type: new Abstract: Diffusion and flow-matching models dominate conditional image generation, yet inference-time scaling for these models is far less developed than for aut

safetyarxiv-cs-cv
24 Jul 2026
Safety

KeySI: An Interaction Framework for Tuning Text Embeddings Based on Human Feedback

DGX agent

arXiv:2607.20556v1 Announce Type: new Abstract: In large-scale text analysis tasks, pre-trained language models are often used to embed text corpora for downstream analysis. However, such models may s

safetyarxiv-cs-ai
24 Jul 2026
Safety

MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs

DGX agent

arXiv:2508.05502v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) perform strongly in high-resource languages, yet often produce fluent but culturally 'thin' descripti

safetyarxiv-cs-ai
24 Jul 2026
Safety

Multivariate Planar Curves: A Statistical Framework for Shape Analysis in Images

DGX agent

arXiv:2508.11780v3 Announce Type: replace-cross Abstract: Recent developments in computer vision have made segmented images widely available across many domains, such as medicine, where segmented radi

safetyarxiv-cs-cv
24 Jul 2026
Safety

One Round Is All You Need: Analytic Federated Learning for Task-Heterogeneous Multi-Label Medical Image Classification

DGX agent

arXiv:2607.20641v1 Announce Type: new Abstract: Federated learning (FL) enables multiple clinical institutions to collaboratively train a shared disease classifier without centralizing patient data. I

safetyarxiv-cs-lg
24 Jul 2026
Safety

Operational Identity: A Finite Audit of Declared and Implemented Rules of Sameness

DGX agent

arXiv:2607.20729v1 Announce Type: cross Abstract: A record system declares when two records refer to the same entity, occurrence, scope, or rule. Its disclosed implementation mechanisms induce a corre

safetyarxiv-cs-ai
24 Jul 2026
Safety

OPOD: On-Policy Omni Distillation

DGX agent

arXiv:2607.20918v1 Announce Type: new Abstract: Omni-modal models can handle text, images, and audio in one system, but improving all of these abilities together remains difficult. Training a single m

safetyarxiv-cs-ai
24 Jul 2026
Safety

PATS: Policy-Aware Training Scaffolding for Agentic Reinforcement Learning

DGX agent

arXiv:2607.21419v1 Announce Type: new Abstract: In long-horizon LLM agent reinforcement learning, weak policies often repeat similar failures, producing uninformative rollout trajectories and limiting

safetyarxiv-cs-ai
24 Jul 2026
Safety

Perspective Latents as an Architectural Condition for Causal Emergence in Active Inference Agents

DGX agent

arXiv:2607.20708v1 Announce Type: new Abstract: A recent line of work measures causal emergence in reinforcement learning agents through Integrated Information Decomposition, reporting that Phi_r grow

safetyarxiv-cs-lg
24 Jul 2026
← Previous
1…8586878889…260
Next →