AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Approximate Equivariance via Projection-based Regularisation

DGX agent

arXiv:2601.05028v2 Announce Type: replace Abstract: Equivariance is a powerful inductive bias in neural networks, improving generalisation and physical consistency. Recently, however, non-equivariant

safetyarxiv-cs-lg
27 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

DGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

safetyarxiv-cs-ai
27 May 2026
Safety

Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice

DGX agent

arXiv:2605.26559v1 Announce Type: cross Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks requi

safetyarxiv-cs-ai
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Safety

BASIS: Batchwise Advantage Estimation from Single-Rollout Information Sharing for LLM Reasoning

DGX agent

arXiv:2605.27293v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become a standard recipe for improving the reasoning abilities of large language models. Existing alg

safetyarxiv-cs-lg
27 May 2026
Safety

Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility

DGX agent

arXiv:2603.03585v2 Announce Type: replace-cross Abstract: Misinformation is a growing societal threat, and susceptibility to misinformative claims varies across demographic groups due to differences i

safetyarxiv-cs-ai
27 May 2026
Safety

Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation

DGX agent

arXiv:2601.03525v3 Announce Type: replace-cross Abstract: Effective reward design is a central challenge in Reinforcement Learning (RL) for code generation. Mainstream test-suite-level outcome rewards

safetyarxiv-cs-ai
27 May 2026
Safety

Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models

DGX agent

arXiv:2605.26491v1 Announce Type: cross Abstract: Preference optimization has emerged as an efficient alternative to online reinforcement learning from human feedback (RLHF) for aligning text-to-image

safetyarxiv-cs-cv
27 May 2026
Safety

Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap Between Theory and Practice

DGX agent

arXiv:2605.27131v1 Announce Type: cross Abstract: Enterprise data platforms face an enduring tension between domain self-service and holistic governance. The data mesh paradigm proposed decentralized

safetyarxiv-cs-ai
27 May 2026
Safety

Beyond Trajectory-Level Attribution: Graph-Based Credit Assignment for Agentic Reinforcement Learning

DGX agent

arXiv:2605.26684v1 Announce Type: cross Abstract: Group-based reinforcement learning (RL) methods have achieved remarkable success in improving the performance of large language models (LLMs) and have

safetyarxiv-cs-ai
27 May 2026
Safety

Bilevel Optimization over Saddle Points of Zero-Sum Markov Games

DGX agent

arXiv:2605.26654v1 Announce Type: cross Abstract: Reinforcement learning (RL) often has a hierarchical structure, where an upper-level (UL) learner selects model parameters and a lower-level (LL) deci

safetyarxiv-cs-ai
27 May 2026
Safety

BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization

DGX agent

arXiv:2605.26182v1 Announce Type: new Abstract: Generating physically buildable brick structures from 3D shapes requires more than geometric reconstruction: the output must also satisfy discrete part

safetyarxiv-cs-ai
27 May 2026
Safety

CFG-OEC: Classifier Free Guidance with Orthogonal Error Correction

DGX agent

arXiv:2511.14075v2 Announce Type: replace-cross Abstract: Classifier free guidance is a standard method for conditional sampling in diffusion models, but its sampling rule is not aligned with the obje

safetyarxiv-cs-ai
27 May 2026
Safety

Completion vs Optimality: Policy Gradient in Long-Horizon Cumulative-Damage Problems

DGX agent

arXiv:2605.26657v1 Announce Type: new Abstract: Long-horizon decision problems with cumulative damage couple locally attractive actions to globally adverse outcomes. We identify two orthogonal failure

safetyarxiv-cs-ai
27 May 2026
Safety

Constrained Bayesian Experimental Design via Online Planning

DGX agent

arXiv:2605.26990v1 Announce Type: cross Abstract: Bayesian experimental design (BED) is a principled framework for data-efficient design of sequential experiments. However, existing BED methods are un

safetyarxiv-cs-lg
27 May 2026
Safety

Conv-to-Bench: Evaluating Language Models Via User-Assistant Dialogues In Code Tasks

DGX agent

arXiv:2605.26440v1 Announce Type: new Abstract: The rapid advancement of Large Language Models (LLMs) has outpaced the scalability of traditional evaluation benchmarks, which remain heavily dependent

safetyarxiv-cs-cl
27 May 2026
Safety

Cost of Structural Learning Under Censored Feedback: A Threshold-Bandit Approach

DGX agent

arXiv:2605.27076v1 Announce Type: cross Abstract: In many multi-agent applications, tasks yield rewards only when executed by a coalition meeting an unknown size threshold; otherwise, feedback is full

safetyarxiv-cs-lg
27 May 2026
Safety

Counteraction-Aware Multi-Teacher On-Policy Distillation for General Capability Recovery with Domain Preservation

DGX agent

arXiv:2605.27115v1 Announce Type: new Abstract: Domain specialization can improve LLM behavior in vertical domains, but often weakens the general capabilities inherited from the original model. Recent

safetyarxiv-cs-ai
27 May 2026
Safety

Counterfactual Credit Policy Optimization for Multi-Agent Collaboration

DGX agent

arXiv:2603.21563v2 Announce Type: replace Abstract: Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such s

safetyarxiv-cs-ai
27 May 2026
Safety

Credit-assigned Policy Gradient for Early Stage Retrieval in Two-stage Ranking

DGX agent

arXiv:2605.26385v1 Announce Type: cross Abstract: Large-scale search, recommendation, and retrieval-augmented generation (RAG) systems typically employ a two-stage architecture: an early-stage ranker

safetyarxiv-cs-ai
27 May 2026
Safety

CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations

DGX agent

arXiv:2605.26293v1 Announce Type: cross Abstract: Prior work establishes that controlled contrastiveness between self-generated responses from large language models, set via reward scores, improves do

safetyarxiv-cs-ai
27 May 2026
Safety

Cross-Receiver Generalization for RF Fingerprint Identification via Feature Disentanglement and Adversarial Training

DGX agent

arXiv:2510.09405v2 Announce Type: replace Abstract: Radio frequency fingerprint identification (RFFI) is a key technique for wireless network security, leveraging intrinsic hardware imperfections to e

safetyarxiv-cs-lg
27 May 2026
Safety

Dimensional Distribution Emotion State: Leveraging Valence and Arousal as a Common Embedding Space for Visual Emotion Analysis

DGX agent

arXiv:2605.26262v1 Announce Type: new Abstract: Museums are important sites for the dissemination of culture and art. They are institutions rooted in history and tradition; their exhibitions are often

safetyarxiv-cs-cv
27 May 2026
Safety

DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation

DGX agent

arXiv:2605.26236v1 Announce Type: new Abstract: Co-speech gesture generation requires both semantic expressivity and biomechanically plausible rhythmic motion. Existing holistic gesture models mix lex

safetyarxiv-cs-cv
27 May 2026
Safety

DV-SFT: Direct Vision Supervision for Fine-Grained Visual Understanding

DGX agent

arXiv:2605.26656v1 Announce Type: new Abstract: Multimodal large language models are typically trained end-to-end to predict ground-truth answers, yet supervision signals are applied exclusively to te

safetyarxiv-cs-cv
27 May 2026
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Safety

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient

DGX agent

arXiv:2605.26478v1 Announce Type: cross Abstract: We present the stochastic decoupled policy gradient (SDPG), a lightweight visual reinforcement learning (RL) method that trains diverse visuomotor con

safetyarxiv-cs-ai
27 May 2026
Safety

Elias in the Lighthouse, Again? Diagnosing Low Diversity in LLM Stories

DGX agent

arXiv:2605.26492v1 Announce Type: cross Abstract: LLM-generated stories are a popular use case, but they show very low variability. We sample 20,000 total stories from four current models using five p

safetyarxiv-cs-ai
27 May 2026
Safety

EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation

DGX agent

arXiv:2605.26785v1 Announce Type: cross Abstract: Post-trained LLMs are often optimized to align responses with human preferences, making them safe, polite, and conversationally appropriate. In advers

safetyarxiv-cs-ai
27 May 2026
Safety

Enabling Extensible Embodied Capabilities with Tools

DGX agent

arXiv:2605.26637v1 Announce Type: new Abstract: Most existing embodied intelligence methods formulate perception, reasoning, planning, and control within a unified parameterized policy. Yet these capa

safetyarxiv-cs-ro
27 May 2026
Safety

Ethical Fairness without Demographics in Human-Centered AI

DGX agent

arXiv:2603.13373v3 Announce Type: replace-cross Abstract: In ubiquitous and mobile health systems, computational models infer human states from wearable, behavioral, and physiological sensing data. In

safetyarxiv-cs-ai
27 May 2026
Safety

Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights

DGX agent

arXiv:2501.06708v5 Announce Type: replace-cross Abstract: Large-scale web-crawled datasets contain noise, bias, and irrelevant information, necessitating data selection techniques. Existing methods de

safetyarxiv-cs-ai
27 May 2026
Safety

FalAR: A Large-scale Speaker-Annotated European Portuguese Speech Corpus of Parliamentary Sessions

DGX agent

arXiv:2605.27062v1 Announce Type: new Abstract: State-of-the-art performance for Automatic Speech Recognition (ASR) largely depends on the availability of large-scale labeled corpora. This creates a d

safetyarxiv-cs-cl
27 May 2026
Safety

Few-shot Cross-country Generalization of Tabular Machine Learning and Foundation Models for Childhood Anemia Prediction under Distribution Shift

DGX agent

arXiv:2605.26589v1 Announce Type: cross Abstract: Childhood anemia affects around 40% of children aged 6-59 months globally and arises from heterogeneous factors, limiting model generalizability. We e

safetyarxiv-cs-ai
27 May 2026
Safety

Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints

DGX agent

arXiv:2603.17685v3 Announce Type: replace Abstract: Balancing policy expressiveness with the exploration-exploitation trade-off is a core challenge in online Reinforcement Learning (RL). While Stochas

safetyarxiv-cs-lg
27 May 2026
Safety

FM-fMRI: Event Conditioned Flow Matching for Rest-to-Task fMRI Time-Series Synthesis

DGX agent

arXiv:2605.26423v1 Announce Type: new Abstract: Task-based fMRI provides a direct readout of task-evoked neural dynamics, but it is expensive and difficult to acquire at scale, motivating rest-to-task

safetyarxiv-cs-lg
27 May 2026
Safety

Foundations of a Time-Consistent Counterfactual Actuarial Runtime for Autonomous AI Agents

DGX agent

arXiv:2605.26508v1 Announce Type: cross Abstract: We propose a foundational runtime actuarial layer for autonomous AI agents in which every side-effect-bearing action carries a time-consistent, counte

safetyarxiv-cs-ai
27 May 2026
Safety

From Norms to Indicators (N2I-RAG): An Agentic Retrieval-Augmented Generation Framework for Legal Indicator Computation

DGX agent

arXiv:2605.26926v1 Announce Type: new Abstract: Computing legal indicators from normative texts is a key task in legal monitoring and policy evaluation, but presents significant challenges due to the

safetyarxiv-cs-ai
27 May 2026
Safety

From Static Context to Calibrated Interactive RL: Mitigating Distribution Shift in Multi-turn Dialogue with Aligned Simulator

DGX agent

arXiv:2605.26403v1 Announce Type: new Abstract: A long-standing goal of the research community is to develop highly interactive LLM-based dialogue agents. Recent research focuses on optimizing policie

safetyarxiv-cs-ai
27 May 2026
Safety

FTibSuite: A Comprehensive Resource Suite for Tibetan Vision-Language Modeling

DGX agent

arXiv:2605.26601v1 Announce Type: new Abstract: Vision-language models have progressed rapidly, but Tibetan remains a severely underserved low-resource language due to the lack of reproducible trainin

safetyarxiv-cs-cv
27 May 2026
Safety

Generalist Graph Anomaly Detection via Prototype-Based Distillation

DGX agent

arXiv:2605.26857v1 Announce Type: new Abstract: Driven by the pressing demand for graph anomaly detection (GAD) in high-stakes domains, the generalist GAD paradigm, which trains a single detector tran

safetyarxiv-cs-lg
27 May 2026
Safety

GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation

DGX agent

arXiv:2602.16449v2 Announce Type: replace-cross Abstract: Generative model evaluation commonly relies on high-dimensional embedding spaces to compute distances between samples. We show that dataset re

safetyarxiv-cs-ai
27 May 2026
Safety

Grounding Text Embeddings in Stakeholder Associations

DGX agent

arXiv:2605.27168v1 Announce Type: cross Abstract: Text embeddings are widely used to analyse large corpora of complex texts. However, it is unclear whether the embeddings capture the same semantic dis

safetyarxiv-cs-ai
27 May 2026
Safety

Heterogeneous AAV Logistics Task Allocation: A Reinforcement Learning Enhanced Overlapping Coalition Formation Game Approach

DGX agent

arXiv:2605.26471v1 Announce Type: new Abstract: In dynamic urban logistics, the stochastic emergence of time-sensitive tasks poses a significant optimality challenge for heterogeneous AAVs logistics t

safetyarxiv-cs-ro
27 May 2026
Safety

Hi-SAM: A Hierarchical Structure-Aware Multi-modal Framework for Large-Scale Recommendation

DGX agent

arXiv:2602.11799v2 Announce Type: replace Abstract: Multi-modal recommendation has gained traction as items possess rich attributes like text and images. Semantic ID-based approaches effectively discr

safetyarxiv-cs-ai
27 May 2026
Safety

How Reliable are LLMs for Reasoning on the Re-ranking task?

DGX agent

arXiv:2508.18444v2 Announce Type: replace-cross Abstract: With the improving semantic understanding capability of Large Language Models (LLMs), they exhibit a greater awareness and alignment with huma

safetyarxiv-cs-ai
27 May 2026
Safety

HyperSim: A Holistic Sim-To-Real Framework For Robust Robotic Manipulation

DGX agent

arXiv:2605.26638v1 Announce Type: new Abstract: Scaling data volume and diversity is critical for generalizing embodied intelligence. While synthetic data generation offers a scalable alternative to e

safetyarxiv-cs-ro
27 May 2026
Safety

Image Thresholding: Understanding Bias of Evaluation Metrics towards Specific Evaluation Functions

DGX agent

arXiv:2605.27132v1 Announce Type: new Abstract: Multilevel image thresholding is widely used for segmentation in applications ranging from medical imaging to remote sensing. Classical objective functi

safetyarxiv-cs-cv
27 May 2026
← Previous
1…156157158159160…260
Next →