AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

VLA-Forget: Vision-Language-Action Unlearning for Embodied Foundation Models

DGX agent

arXiv:2604.03956v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models are emerging as embodied foundation models for robotic manipulation, but their deployment introduces a new

safetyarxiv-cs-ai
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

DGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

safetyarxiv-cs-cl
24 Apr 2026
Safety

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

DGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

safetyarxiv-cs-ai
24 Apr 2026
Safety

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

DGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

safetyarxiv-cs-ai
24 Apr 2026
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Safety

A Survey of Scaling in Large Language Model Reasoning

DGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

safetyarxiv-cs-ai
23 Apr 2026
Safety

A Synchronized Audio-Visual Multi-View Capture System

DGX agent

arXiv:2603.23089v2 Announce Type: replace Abstract: Multi-view capture systems have been an important tool in research for recording human motion under controlling conditions. Most existing systems ar

safetyarxiv-cs-cv
23 Apr 2026
Safety

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking

DGX agent

arXiv:2604.20305v1 Announce Type: new Abstract: Realizing active visual tracking with a single unified model across diverse robots is challenging, as the physical constraints and motion dynamics vary

safetyarxiv-cs-ro
23 Apr 2026
Safety

AI models of unstable flow exhibit hallucination

DGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

safetyarxiv-cs-ai
23 Apr 2026
Safety

All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG

DGX agent

arXiv:2604.20199v1 Announce Type: new Abstract: Multilingual Retrieval-Augmented Generation (mRAG) leverages cross-lingual evidence to ground Large Language Models (LLMs) in global knowledge. However,

safetyarxiv-cs-cl
23 Apr 2026
Safety

Auto-ART: Structured Literature Synthesis and Automated Adversarial Robustness Testing

DGX agent

arXiv:2604.20704v1 Announce Type: cross Abstract: Adversarial robustness evaluation underpins every claim of trustworthy ML deployment, yet the field suffers from fragmented protocols and undetected g

safetyarxiv-cs-lg
23 Apr 2026
Safety

Avoiding Overthinking and Underthinking: Curriculum-Aware Budget Scheduling for LLMs

DGX agent

arXiv:2604.19780v1 Announce Type: new Abstract: Scaling test-time compute via extended reasoning has become a key paradigm for improving the capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-cl
23 Apr 2026
Safety

Best Policy Learning from Trajectory Preference Feedback

DGX agent

arXiv:2501.18873v4 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) has emerged as a powerful approach for aligning generative models, but its reliance on learned rew

safetyarxiv-cs-lg
23 Apr 2026
Safety

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

DGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

safetyarxiv-cs-ai
23 Apr 2026
Safety

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

DGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

safetyarxiv-cs-cl
23 Apr 2026
Safety

CAST: Achieving Stable LLM-based Text Analysis for Data Analytics

DGX agent

arXiv:2602.15861v2 Announce Type: replace-cross Abstract: Text analysis of tabular data relies on two core operations: summarization for corpus-level theme extraction and tagging for row-level labelin

safetyarxiv-cs-ai
23 Apr 2026
Safety

ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration

DGX agent

arXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene

safetyarxiv-cs-ai
23 Apr 2026
Safety

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation

DGX agent

arXiv:2603.25383v3 Announce Type: replace Abstract: CLIP aligns image and text embeddings via contrastive learning and demonstrates strong zero-shot generalization. Its large-scale architecture requir

safetyarxiv-cs-cv
23 Apr 2026
Safety

CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment

DGX agent

arXiv:2510.18471v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) excel at code generation by learning from vast code corpora, a fundamental semantic gap remains between the

safetyarxiv-cs-ai
23 Apr 2026
Safety

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

DGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

safetyarxiv-cs-ai
23 Apr 2026
Safety

CoRe: Joint Optimization with Contrastive Learning for Medical Image Registration

DGX agent

arXiv:2603.23694v2 Announce Type: replace Abstract: Medical image registration is a fundamental task in medical image analysis, enabling the alignment of images from different modalities or time point

safetyarxiv-cs-cv
23 Apr 2026
Safety

Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation

DGX agent

arXiv:2604.20763v1 Announce Type: cross Abstract: Retrieval quality is the primary bottleneck for accuracy and robustness in retrieval-augmented generation (RAG). Current evaluation relies on heuristi

safetyarxiv-cs-ai
23 Apr 2026
Safety

CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction

DGX agent

arXiv:2505.04897v2 Announce Type: replace-cross Abstract: Interactive imitation learning makes an agent's control policy robust by stepwise supervisions from an expert. The recent algorithms mostly em

safetyarxiv-cs-lg
23 Apr 2026
Safety

Decentralized Machine Learning with Centralized Performance Guarantees via Gibbs Algorithms

DGX agent

arXiv:2604.20492v1 Announce Type: cross Abstract: In this paper, it is shown, for the first time, that centralized performance is achievable in decentralized learning without sharing the local dataset

safetyarxiv-cs-lg
23 Apr 2026
Safety

Distributional Inverse Reinforcement Learning

DGX agent

arXiv:2510.03013v3 Announce Type: replace Abstract: We propose a distributional framework for offline Inverse Reinforcement Learning (IRL) that jointly models uncertainty over reward functions and ful

safetyarxiv-cs-lg
23 Apr 2026
Safety

DRIV-EX: Counterfactual Explanations for Driving LLMs

DGX agent

arXiv:2603.00696v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as reasoning engines in autonomous driving, yet their decision-making remains opaque. We propose

safetyarxiv-cs-cl
23 Apr 2026
Safety

Efficient Reinforcement Learning using Linear Koopman Dynamics for Nonlinear Robotic Systems

DGX agent

arXiv:2604.19980v1 Announce Type: new Abstract: This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed appro

safetyarxiv-cs-ro
23 Apr 2026
Safety

Efficiently Closing Loops in LiDAR-Based SLAM Using Point Cloud Density Maps

DGX agent

arXiv:2501.07399v2 Announce Type: replace Abstract: Consistent maps are key for most autonomous mobile robots, and they often use SLAM approaches to build such maps. Loop closures via place recognitio

safetyarxiv-cs-ro
23 Apr 2026
Safety

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

DGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

safetyarxiv-cs-ai
23 Apr 2026
Safety

Epistemic Constitutionalism Or: how to avoid coherence bias

DGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

safetyarxiv-cs-ai
23 Apr 2026
Safety

Epistemology gives a Future to Complementarity in Human-AI Interactions

DGX agent

arXiv:2601.09871v2 Announce Type: replace Abstract: Human-AI complementarity is the claim that a human supported by an AI system can outperform either alone in a decision-making process. Since its int

safetyarxiv-cs-ai
23 Apr 2026
Safety

ETac: A Lightweight and Efficient Tactile Simulation Framework for Learning Dexterous Manipulation

DGX agent

arXiv:2604.20295v1 Announce Type: new Abstract: Tactile sensors are increasingly integrated into dexterous robotic manipulators to enhance contact perception. However, learning manipulation policies t

safetyarxiv-cs-ro
23 Apr 2026
Safety

Evaluating Black-Box Vulnerabilities with Wasserstein-Constrained Data Perturbations

DGX agent

arXiv:2603.15867v2 Announce Type: replace Abstract: The growing use of Machine Learning (ML) tools comes with critical challenges, such as limited model explainability. We propose a global explainabil

safetyarxiv-cs-lg
23 Apr 2026
Safety

Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias

DGX agent

arXiv:2604.19763v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) systems have growing applications in sensitive domains such as mental health and education, where biased predictions

safetyarxiv-cs-ai
23 Apr 2026
Safety

Fairness-Aware Multi-Group Target Detection in Online Discussion

DGX agent

arXiv:2407.11933v4 Announce Type: replace Abstract: Target-group detection is the task of detecting which group(s) a piece of content is ``directed at or about''. Applications include targeted marketi

safetyarxiv-cs-lg
23 Apr 2026
Safety

Faster Fixed-Point Methods for Multichain MDPs

DGX agent

arXiv:2506.20910v2 Announce Type: replace-cross Abstract: We study value-iteration (VI) algorithms for solving general (a.k.a. multichain) Markov decision processes (MDPs) under the average-reward cri

safetyarxiv-cs-lg
23 Apr 2026
Safety

FingerEye: Continuous and Unified Vision-Tactile Sensing for Dexterous Manipulation

DGX agent

arXiv:2604.20689v1 Announce Type: new Abstract: Dexterous robotic manipulation requires comprehensive perception across all phases of interaction: pre-contact, contact initiation, and post-contact. Su

safetyarxiv-cs-ro
23 Apr 2026
Safety

FluSplat: Sparse-View 3D Editing without Test-Time Optimization

DGX agent

arXiv:2604.20038v1 Announce Type: new Abstract: Recent advances in text-guided image editing and 3D Gaussian Splatting (3DGS) have enabled high-quality 3D scene manipulation. However, existing pipelin

safetyarxiv-cs-cv
23 Apr 2026
Safety

Frictionless Love: Associations Between AI Companion Roles and Behavioral Addiction

DGX agent

arXiv:2604.20011v1 Announce Type: cross Abstract: AI companion chatbots increasingly shape how people seek social and emotional connection, sometimes substituting for relationships with romantic partn

safetyarxiv-cs-ai
23 Apr 2026
Safety

FurnSet: Exploiting Repeats for 3D Scene Reconstruction

DGX agent

arXiv:2604.20093v1 Announce Type: new Abstract: Single-view 3D scene reconstruction involves inferring both object geometry and spatial layout. Existing methods typically reconstruct objects independe

safetyarxiv-cs-cv
23 Apr 2026
Safety

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

DGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

safetyarxiv-cs-lg
23 Apr 2026
Safety

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning

DGX agent

arXiv:2604.20659v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced the reasoning capabilities of Large Language Models (LLMs) by leveraging direct out

safetyarxiv-cs-ai
23 Apr 2026
Safety

Hidden Reliability Risks in Large Language Models: Systematic Identification of Precision-Induced Output Disagreements

DGX agent

arXiv:2604.19790v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed under diverse numerical precision configurations, including standard floating-point formats (e.g.

safetyarxiv-cs-ai
23 Apr 2026
Safety

Hybrid Latent Reasoning with Decoupled Policy Optimization

DGX agent

arXiv:2604.20328v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning significantly elevates the complex problem-solving capabilities of multimodal large language models (MLLMs). However, a

safetyarxiv-cs-cv
23 Apr 2026
Safety

Hybrid Policy Distillation for LLMs

DGX agent

arXiv:2604.20244v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a powerful paradigm for compressing large language models (LLMs), whose effectiveness depends on intertwined choices of

safetyarxiv-cs-ai
23 Apr 2026
Safety

i-WiViG: Interpretable Window Vision GNN

DGX agent

arXiv:2503.08321v2 Announce Type: replace Abstract: Vision graph neural networks have emerged as a popular approach for modeling the global and spatial context for image recognition. However, a signif

safetyarxiv-cs-cv
23 Apr 2026
Safety

Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning

DGX agent

arXiv:2505.07527v5 Announce Type: replace Abstract: The advantage function is a central concept in RL that helps reduce variance in policy gradient estimates. For language modeling, Group Relative Pol

safetyarxiv-cs-lg
23 Apr 2026
Safety

Language-Coupled Reinforcement Learning for Multilingual Retrieval-Augmented Generation

DGX agent

arXiv:2601.14896v2 Announce Type: replace Abstract: Multilingual retrieval-augmented generation (MRAG) requires models to effectively acquire and integrate beneficial external knowledge from multiling

safetyarxiv-cs-cl
23 Apr 2026
← Previous
1…213214215216217…257
Next →