AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
Safety

Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective

DGX agent

arXiv:2601.03154v2 Announce Type: replace Abstract: Reasoning-tuned LLMs utilizing long Chain-of-Thought (CoT) excel at single-answer tasks, yet their ability to model Human Label Variation--which req

safetyarxiv-cs-cl
21 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models

DGX agent

arXiv:2511.15669v2 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision-Language-Action (VLA) models, or does it merely add overhead? Existing CoT-VLA system

safetyarxiv-cs-lg
21 Apr 2026
Safety

Demystifying the unreasonable effectiveness of online alignment methods

DGX agent

arXiv:2604.17207v1 Announce Type: cross Abstract: Iterative alignment methods based on purely greedy updates are remarkably effective in practice, yet existing theoretical guarantees of (O(log T)) KL-

safetyarxiv-cs-cl
21 Apr 2026
Safety

DFedReweighting: A Unified Framework for Objective-Oriented Reweighting in Decentralized Federated Learning

DGX agent

arXiv:2512.12022v2 Announce Type: replace Abstract: Decentralized federated learning (DFL) has emerged as a promising paradigm that enables multiple clients to collaboratively train machine learning m

safetyarxiv-cs-lg
21 Apr 2026
Safety

Diagnosing LLM-based Rerankers in Cold-Start Recommender Systems: Coverage, Exposure and Practical Mitigations

DGX agent

arXiv:2604.16318v1 Announce Type: cross Abstract: Large language models (LLMs) and cross-encoder rerankers have gained attention for improving recommender systems, particularly in cold-start scenarios

safetyarxiv-cs-cl
21 Apr 2026
Safety

DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs

DGX agent

arXiv:2601.03559v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning improves multi-step mathematical problem solving in large language models but remains vulnerable to exposure bias a

safetyarxiv-cs-cl
21 Apr 2026
Safety

Differential Privacy in Two-Layer Networks: How DP-SGD Harms Fairness and Robustness

DGX agent

arXiv:2603.04881v2 Announce Type: replace Abstract: Differentially private learning is essential for training models on sensitive data, but empirical studies consistently show that it can degrade perf

safetyarxiv-cs-lg
21 Apr 2026
Safety

Distributional Off-Policy Evaluation with Deep Quantile Process Regression

DGX agent

arXiv:2604.18143v1 Announce Type: cross Abstract: This paper investigates the off-policy evaluation (OPE) problem from a distributional perspective. Rather than focusing solely on the expectation of t

safetyarxiv-cs-lg
21 Apr 2026
Safety

Diverse Dictionary Learning

DGX agent

arXiv:2604.17568v1 Announce Type: new Abstract: Given only observational data X = g(Z), where both the latent variables Z and the generating process g are unknown, recovering Z is ill-posed without ad

safetyarxiv-cs-lg
21 Apr 2026
Safety

DMax: Aggressive Parallel Decoding for dLLMs

DGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

safetyarxiv-cs-lg
21 Apr 2026
Safety

Do LLMs Use Cultural Knowledge Without Being Told? A Multilingual Evaluation of Implicit Pragmatic Adaptation

DGX agent

arXiv:2604.17718v1 Announce Type: new Abstract: Many benchmarks show that large language models can answer direct questions about culture. We study a different question: do they also change how they s

safetyarxiv-cs-cl
21 Apr 2026
Safety

Does 'Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?

DGX agent

arXiv:2604.18161v1 Announce Type: new Abstract: In policy gradient reinforcement learning, access to a differentiable model enables 1st-order gradient estimation that accelerates learning compared to

safetyarxiv-cs-lg
21 Apr 2026
Safety

Does Welsh media need a review? Detecting bias in Nation.Cymru's political reporting

DGX agent

arXiv:2604.17628v1 Announce Type: new Abstract: Wales' political landscape has been marked by growing accusations of bias in Welsh media. This paper takes the first computational step toward testing t

safetyarxiv-cs-cl
21 Apr 2026
Safety

DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models

DGX agent

arXiv:2604.16979v1 Announce Type: cross Abstract: High-quality and diverse multimodal data are essential for improving vision-language models (VLMs), yet existing datasets often contain noisy, redunda

safetyarxiv-cs-cl
21 Apr 2026
Safety

DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty

DGX agent

arXiv:2506.12622v2 Announce Type: replace Abstract: Deep reinforcement learning (RL) has achieved remarkable success, yet its deployment in real-world scenarios is often limited by vulnerability to en

safetyarxiv-cs-lg
21 Apr 2026
Safety

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

DGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

safetyarxiv-cs-cv
21 Apr 2026
Safety

Dynamic Emotion and Personality Profiling for Multimodal Deception Detection

DGX agent

arXiv:2604.17037v1 Announce Type: new Abstract: Deception detection is of great significance for ensuring information security and conducting public opinion analysis, with personality factors and emot

safetyarxiv-cs-cl
21 Apr 2026
Safety

Dynamic Visual-semantic Alignment for Zero-shot Learning with Ambiguous Labels

DGX agent

arXiv:2604.17710v1 Announce Type: new Abstract: Zero-shot learning (ZSL) aims to recognize unseen classes without visual instances. However, existing methods usually assume clean labels, overlooking r

safetyarxiv-cs-cv
21 Apr 2026
Safety

DynaWeb: Model-Based Reinforcement Learning of Web Agents

DGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

safetyarxiv-cs-cl
21 Apr 2026
Safety

Efficient Federated RLHF via Zeroth-Order Policy Optimization

DGX agent

arXiv:2604.17747v1 Announce Type: new Abstract: This paper considers reinforcement learning from human feedback in a federated learning setting with resource-constrained agents, such as edge devices.

safetyarxiv-cs-lg
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Safety

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

DGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi

safetyarxiv-cs-lg
21 Apr 2026
Safety

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

DGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

safetyarxiv-cs-cv
21 Apr 2026
Safety

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

DGX agent

arXiv:2601.10306v2 Announce Type: replace-cross Abstract: While Reinforcement Learning (RL) has advanced LLM reasoning, applying it to long-context scenarios is hindered by sparsity of outcome rewards

safetyarxiv-cs-cl
21 Apr 2026
Safety

Evolutionary Negative Module Pruning for Better LoRA Merging

DGX agent

arXiv:2604.17753v1 Announce Type: cross Abstract: Merging multiple Low-Rank Adaptation (LoRA) experts into a single backbone is a promising approach for efficient multi-task deployment. While existing

safetyarxiv-cs-cl
21 Apr 2026
Safety

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

DGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

safetyarxiv-cs-cl
21 Apr 2026
Safety

FairLogue: Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using the All of Us Research Program

DGX agent

arXiv:2604.16450v1 Announce Type: cross Abstract: Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess de

safetyarxiv-cs-lg
21 Apr 2026
Safety

Fairness Constraints in High-Dimensional Generalized Linear Models

DGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

safetyarxiv-cs-lg
21 Apr 2026
Safety

FairNVT: Improving Fairness via Noise Injection in Vision Transformers

DGX agent

arXiv:2604.16780v1 Announce Type: new Abstract: This paper presents FairNVT, a lightweight debiasing framework for pretrained transformer-based encoders that improves both representation and predictio

safetyarxiv-cs-cv
21 Apr 2026
Safety

'Faithful to What?' On the Limits of Fidelity-Based Explanations

DGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

safetyarxiv-cs-lg
21 Apr 2026
Safety

Fisher Decorator: Refining Flow Policy via A Local Transport Map

DGX agent

arXiv:2604.17919v1 Announce Type: new Abstract: Recent advances in flow-based offline reinforcement learning (RL) have achieved strong performance by parameterizing policies via flow matching. However

safetyarxiv-cs-lg
21 Apr 2026
Safety

Foundation Model for Cardiac Time Series via Masked Latent Attention

DGX agent

arXiv:2603.26475v2 Announce Type: replace Abstract: Electrocardiograms (ECGs) are among the most widely available clinical signals and play a central role in cardiovascular diagnosis. While recent fou

safetyarxiv-cs-lg
21 Apr 2026
Safety

Function Words as Statistical Cues for Language Learning

DGX agent

arXiv:2601.21191v2 Announce Type: replace Abstract: What statistical properties might support learning abstract grammatical knowledge from linear input? We address this question by examining the stati

safetyarxiv-cs-cl
21 Apr 2026
Safety

Generative Semantic Communication via Alternating Dual-Domain Posterior Sampling

DGX agent

arXiv:2604.16796v1 Announce Type: new Abstract: Generative semantic communication (SemCom) harnesses pretrained generative priors to improve the perceptual quality of wireless image transmission. Exis

safetyarxiv-cs-cv
21 Apr 2026
Safety

Geometric Stability: The Missing Axis of Representations

DGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

safetyarxiv-cs-cl
21 Apr 2026
Safety

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

DGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

safetyarxiv-cs-cl
21 Apr 2026
Safety

GS-STVSR: Ultra-Efficient Continuous Spatio-Temporal Video Super-Resolution via 2D Gaussian Splatting

DGX agent

arXiv:2604.18047v1 Announce Type: new Abstract: Continuous Spatio-Temporal Video Super-Resolution (C-STVSR) aims to simultaneously enhance the spatial resolution and frame rate of videos by arbitrary

safetyarxiv-cs-cv
21 Apr 2026
Safety

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment

DGX agent

arXiv:2604.17928v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has proven effective for training reasoning-oriented large language models, but existing methods la

safetyarxiv-cs-lg
21 Apr 2026
Safety

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

DGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

safetyarxiv-cs-cv
21 Apr 2026
Safety

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

DGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

safetyarxiv-cs-cl
21 Apr 2026
Safety

Human-Centered Supervision for Sentiment Analysis in Telugu: A Systematic Inquiry Beyond Accuracy

DGX agent

arXiv:2508.01486v3 Announce Type: replace Abstract: Sentiment analysis for low-resource languages remains challenging in an era where interpretability, human alignment, and fairness are increasingly n

safetyarxiv-cs-cl
21 Apr 2026
Safety

Hybrid Multi-Dimensional MRI Prostate Cancer Detection via Hadamard Network-Based Bias Correction and Residual Networks

DGX agent

arXiv:2604.17107v1 Announce Type: new Abstract: Magnetic Resonance Imaging (MRI) is vital for prostate cancer (PCa) diagnosis. While advanced techniques such as Hybrid Multi-dimensional MRI (HM-MRI) h

safetyarxiv-cs-cv
21 Apr 2026
Safety

Hyperbolic Enhanced Representation Learning for Incomplete Multi-view Clustering

DGX agent

arXiv:2604.16959v1 Announce Type: cross Abstract: Incomplete Multi-View Clustering (IMVC) faces the challenge of learning discriminative representations from fragmentary observations while maintaining

safetyarxiv-cs-cv
21 Apr 2026
Safety

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to res…

DGX agent

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to restore fairness in the upcoming elections.' Modern American po

safetyelon-musk--x
21 Apr 2026
Safety

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

DGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

safetyarxiv-cs-cl
21 Apr 2026
Safety

Identifying Ethical Biases in Action Recognition Models

DGX agent

arXiv:2604.17971v1 Announce Type: new Abstract: Human Action Recognition (HAR) models are increasingly deployed in high-stakes environments, yet their fairness across different human appearances has n

safetyarxiv-cs-cv
21 Apr 2026
Safety

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

DGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

safetyarxiv-cs-lg
21 Apr 2026
Safety

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

DGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

safetyarxiv-cs-cv
21 Apr 2026
← Previous
1…253254255256257…299
Next →