AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models

DGX agent

arXiv:2604.25903v1 Announce Type: cross Abstract: The accelerating adoption of Large Language Models (LLMs) in software engineering (SE) has brought with it a silent crisis: unsustainable computationa

safetyarxiv-cs-lg
29 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CHUCKLE -- When Humans Teach AI To Learn Emotions The Easy Way

DGX agent

arXiv:2510.09382v2 Announce Type: replace Abstract: Curriculum learning (CL) structures training from simple to complex samples, facilitating progressive learning. However, existing CL approaches for

safetyarxiv-cs-lg
29 Apr 2026
Safety

Compute Aligned Training: Optimizing for Test Time Inference

DGX agent

arXiv:2604.24957v1 Announce Type: new Abstract: Scaling test-time compute has emerged as a powerful mechanism for enhancing Large Language Model (LLM) performance. However, standard post-training para

safetyarxiv-cs-lg
29 Apr 2026
Safety

Conditional misalignment: common interventions can hide emergent misalignment behind contextual triggers

DGX agent

arXiv:2604.25891v1 Announce Type: new Abstract: Finetuning a language model can lead to emergent misalignment (EM) [Betley et al., 2025b]. Models trained on a narrow distribution of misaligned behavio

safetyarxiv-cs-lg
29 Apr 2026
Safety

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

DGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

safetyarxiv-cs-cl
29 Apr 2026
Safety

CroSearch-R1: Better Leveraging Cross-lingual Knowledge for Retrieval-Augmented Generation

DGX agent

arXiv:2604.25182v1 Announce Type: new Abstract: A multilingual collection may contain useful knowledge in other languages to supplement and correct the facts in the original language for Retrieval-Aug

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Cross-Lingual Jailbreak Detection via Semantic Codebooks

DGX agent

arXiv:2604.25716v1 Announce Type: new Abstract: Safety mechanisms for large language models (LLMs) remain predominantly English-centric, creating systematic vulnerabilities in multilingual deployment.

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

DEGround: An Effective Baseline for Ego-centric 3D Visual Grounding with a Homogeneous Framework

DGX agent

arXiv:2506.05199v3 Announce Type: replace Abstract: A core task in embodied intelligence is ego-centric 3D visual grounding. Existing methods typically adopt two-stage, heterogeneous pipelines that pa

safetyarxiv-cs-cv
29 Apr 2026
Safety

DGLight: DQN-Guided GRPO Fine-Tuning of Large Language Models for Traffic Signal Control

DGX agent

arXiv:2604.25259v1 Announce Type: new Abstract: Traffic signal control (TSC) plays a central role in reducing congestion and maintaining urban mobility. This dissertation introduces DGLight, a critic-

safetyarxiv-cs-lg
29 Apr 2026
Safety

DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors

DGX agent

arXiv:2604.25050v1 Announce Type: new Abstract: Unlike chatbots, physical AI must act while the world keeps evolving. Therefore, the inter-chunk pause of synchronous executors are fatal for dynamic ta

safetyarxiv-cs-ro
29 Apr 2026
Safety

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

DGX agent

arXiv:2604.25136v1 Announce Type: new Abstract: We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to

safetyarxiv-cs-cl
29 Apr 2026
Safety

How Fast Should a Model Commit to Supervision? Training Reasoning Models on the Tsallis Loss Continuum

DGX agent

arXiv:2604.25907v1 Announce Type: new Abstract: Adapting reasoning models to new tasks during post-training with only output-level supervision stalls under reinforcement learning from verifiable rewar

safetyarxiv-cs-lg
29 Apr 2026
Safety

How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning

DGX agent

arXiv:2603.01070v2 Announce Type: replace Abstract: Solving complex geometric problems inherently requires interleaved reasoning: a tight alternation between constructing diagrams and performing logic

safetyarxiv-cs-cl
29 Apr 2026
Safety

I-INR: Iterative Implicit Neural Representations

DGX agent

arXiv:2504.17364v4 Announce Type: replace Abstract: Implicit Neural Representations (INRs) have revolutionized signal processing and computer vision by modeling signals as continuous, differentiable f

safetyarxiv-cs-cv
29 Apr 2026
Safety

Interactive Episodic Memory with User Feedback

DGX agent

arXiv:2604.24893v1 Announce Type: new Abstract: In episodic memory with natural language queries (EM-NLQ), a user may ask a question (e.g., 'Where did I place the mug?') that requires searching a long

safetyarxiv-cs-cv
29 Apr 2026
Safety

Learning-Based Dynamics Modeling and Robust Control for Tendon-Driven Continuum Robots

DGX agent

arXiv:2604.25691v1 Announce Type: new Abstract: Tendon-Driven Continuum Robots (TDCRs) pose significant modeling and control challenges due to complex nonlinearities, such as frictional hysteresis and

safetyarxiv-cs-ro
29 Apr 2026
Safety

Learning from Noisy Preferences: A Semi-Supervised Learning Approach to Direct Preference Optimization

DGX agent

arXiv:2604.24952v1 Announce Type: new Abstract: Human visual preferences are inherently multi-dimensional, encompassing aesthetics, detail fidelity, and semantic alignment. However, existing datasets

safetyarxiv-cs-cv
29 Apr 2026
Safety

Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System

DGX agent

arXiv:2604.24921v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into ex

safetyarxiv-cs-cl
29 Apr 2026
Safety

MAIC-UI: Making Interactive Courseware with Generative UI

DGX agent

arXiv:2604.25806v1 Announce Type: new Abstract: Creating interactive STEM courseware traditionally requires HTML/CSS/JavaScript expertise, leaving barriers for educators. While generative AI can produ

safetyarxiv-cs-cl
29 Apr 2026
Safety

MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts

DGX agent

arXiv:2411.14721v2 Announce Type: replace Abstract: Molecule discovery is a pivotal research field, impacting everything from medicine to materials. Recently, Large Language Models (LLMs) have been wi

safetyarxiv-cs-cl
29 Apr 2026
Safety

Navigating Global AI Regulation: A Multi-Jurisdictional Retrieval-Augmented Generation System

DGX agent

arXiv:2604.25448v1 Announce Type: new Abstract: Navigating AI regulation across jurisdictions is increasingly difficult for policymakers, legal professionals, and researchers. To address this, we pres

safetyarxiv-cs-cl
29 Apr 2026
Safety

NimbleReg: A light-weight deep-learning framework for diffeomorphic image registration

DGX agent

arXiv:2503.07768v2 Announce Type: replace Abstract: This paper presents NimbleReg, a light-weight deep-learning (DL) framework for diffeomorphic image registration leveraging surface representation of

safetyarxiv-cs-cv
29 Apr 2026
Safety

One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement

DGX agent

arXiv:2604.25444v1 Announce Type: new Abstract: Large Language Models (LLMs) often fail to utilize their latent reasoning capabilities due to a distributional mismatch between ambiguous human inquirie

safetyarxiv-cs-cl
29 Apr 2026
Safety

Policy Improvement Reinforcement Learning

DGX agent

arXiv:2604.00860v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a central post-training paradigm for improving the reasoning capabilities of large

safetyarxiv-cs-lg
29 Apr 2026
Safety

Progressing beyond Art Masterpieces or Touristic Cliches: how to assess your LLMs for cultural alignment?

DGX agent

arXiv:2604.25654v1 Announce Type: new Abstract: Although the cultural (mis)alignment of Large Language Models (LLMs) has attracted increasing attention -- often framed in terms of cultural bias -- unt

safetyarxiv-cs-cl
29 Apr 2026
Safety

Reference-Augmented Learning for Precise Tracking Policy of Tendon-Driven Continuum Robots

DGX agent

arXiv:2604.25698v1 Announce Type: new Abstract: Tendon-Driven Continuum Robots (TDCRs) pose significant control challenges due to their highly nonlinear, path-dependent dynamics and non-Markovian char

safetyarxiv-cs-ro
29 Apr 2026
Safety

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

DGX agent

arXiv:2604.25636v1 Announce Type: new Abstract: Unified multimodal models (UMMs) integrate visual understanding and generation within a single framework. For text-to-image (T2I) tasks, this unified ca

safetyarxiv-cs-cv
29 Apr 2026
Safety

ResetEdit: Precise Text-guided Editing of Generated Image via Resettable Starting Latent

DGX agent

arXiv:2604.25128v1 Announce Type: new Abstract: Recent advances in diffusion models have enabled high-quality image generation, leading to increasing demand for post-generation editing that modifies l

safetyarxiv-cs-cv
29 Apr 2026
Safety

ReSim: Reliable World Simulation for Autonomous Driving

DGX agent

arXiv:2506.09981v2 Announce Type: replace Abstract: How can we reliably simulate future driving scenarios under a wide range of ego driving behaviors? Recent driving world models, developed exclusivel

safetyarxiv-cs-cv
29 Apr 2026
Safety

Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective

DGX agent

arXiv:2510.10150v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) serves as a cornerstone technique for enhancing the reasoning capabilities of Large Language M

safetyarxiv-cs-lg
29 Apr 2026
Safety

Spark Policy Toolkit: Semantic Contracts and Scalable Execution for Policy Learning in Spark

DGX agent

arXiv:2604.25061v1 Announce Type: cross Abstract: Custom policy-learning pipelines in Spark fail for two coupled systems reasons: rowwise Python execution makes inference impractical, and driver-side

safetyarxiv-cs-lg
29 Apr 2026
Safety

Spectral bandits

DGX agent

arXiv:2604.25272v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this work, we study a bandit problem where the payoffs

safetyarxiv-cs-lg
29 Apr 2026
Safety

Subliminal Steering: Stronger Encoding of Hidden Signals

DGX agent

arXiv:2604.25783v1 Announce Type: new Abstract: Subliminal learning describes a student language model inheriting a behavioral bias by fine-tuning on seemingly innocuous data generated by a biased tea

safetyarxiv-cs-cl
29 Apr 2026
Safety

The Rashomon Effect for Visualizing High-Dimensional Data

DGX agent

arXiv:2604.00485v2 Announce Type: replace Abstract: Dimension reduction (DR) is inherently non-unique: multiple embeddings can preserve the structure of high-dimensional data equally well while differ

safetyarxiv-cs-lg
29 Apr 2026
Safety

The Russian Legislative Corpus

DGX agent

arXiv:2406.04855v3 Announce Type: replace Abstract: We present a comprehensive corpus of Russian primary and secondary legislation adopted between 1991 and 2025, comprising 304,382 texts (194,425,905

safetyarxiv-cs-cl
29 Apr 2026
Safety

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

DGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

safetyarxiv-cs-cl
29 Apr 2026
Safety

Three Models of RLHF Annotation: Extension, Evidence, and Authority

DGX agent

arXiv:2604.25895v1 Announce Type: cross Abstract: Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape

safetyarxiv-cs-cl
29 Apr 2026
Safety

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

DGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

safetyarxiv-cs-cl
29 Apr 2026
Safety

Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research

DGX agent

arXiv:2604.25776v1 Announce Type: new Abstract: Critical analyses of emotion recognition technology have raised ethical concerns around task validity and potential downstream impacts, urging researche

safetyarxiv-cs-cl
29 Apr 2026
Safety

Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation

DGX agent

arXiv:2511.21517v2 Announce Type: replace Abstract: Unlike text, speech conveys information about the speaker, such as gender, through acoustic cues like pitch. This gives rise to modality-specific bi

safetyarxiv-cs-cl
29 Apr 2026
Safety

When Errors Can Be Beneficial: A Categorization of Imperfect Rewards for Policy Gradient

DGX agent

arXiv:2604.25872v1 Announce Type: new Abstract: Training language models via reinforcement learning often relies on imperfect proxy rewards, since ground truth rewards that precisely define the intend

safetyarxiv-cs-lg
29 Apr 2026
Model Releases

A BERTology View of LLM Orchestrations: Token- and Layer-Selective Probes for Efficient Single-Pass Classification

DGX agent

arXiv:2601.13288v2 Announce Type: replace Abstract: Production LLM systems often rely on separate models for safety and other classification-heavy steps, increasing latency, VRAM footprint, and operat

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs

DGX agent

arXiv:2603.07475v2 Announce Type: replace Abstract: Autoregressive (AR) language models build representations incrementally via left-to-right prediction, while diffusion language models (dLLMs) are tr

safetyarxiv-cs-cl
28 Apr 2026
Safety

A Differentiable Framework for Global Circulation Model Precipitation Bias Correction

DGX agent

arXiv:2604.23045v1 Announce Type: new Abstract: Systematic biases in Global Circulation Model (GCM) outputs limit their direct applicability in regional planning, necessitating bias correction. Correc

safetyarxiv-cs-lg
28 Apr 2026
Safety

A Multi-Dimensional Audit of Politically Aligned Large Language Models

DGX agent

arXiv:2604.24429v1 Announce Type: new Abstract: As the application of Large Language Models (LLMs) spreads across various industries, there are increasing concerns about the potential for their misuse

safetyarxiv-cs-cl
28 Apr 2026
Safety

A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning

DGX agent

arXiv:2604.24532v1 Announce Type: new Abstract: Many sequential decision-making tasks involve optimizing multiple conflicting objectives, requiring policies that adapt to different user preferences. I

safetyarxiv-cs-lg
28 Apr 2026
Safety

A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning

DGX agent

arXiv:2604.23386v1 Announce Type: cross Abstract: Federated Learning (FL) typically assumes unconditional collaboration, a premise that overlooks the complexities of real-world, multi-stakeholder envi

safetyarxiv-cs-ai
28 Apr 2026
Safety

AdaRubric: Task-Adaptive Rubrics for LLM Agent Evaluation

DGX agent

arXiv:2603.21362v2 Announce Type: replace Abstract: LLM-as-Judge evaluation fails agent tasks because a fixed rubric cannot capture what matters for this task: code debugging demands Correctness and E

safetyarxiv-cs-ai
28 Apr 2026
← Previous
1…207208209210211…257
Next →