AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models

DGX agent

arXiv:2604.16979v1 Announce Type: cross Abstract: High-quality and diverse multimodal data are essential for improving vision-language models (VLMs), yet existing datasets often contain noisy, redunda

safetyarxiv-cs-cl
21 Apr 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

DGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

safetyarxiv-cs-lg
21 Apr 2026
Safety

DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty

DGX agent

arXiv:2506.12622v2 Announce Type: replace Abstract: Deep reinforcement learning (RL) has achieved remarkable success, yet its deployment in real-world scenarios is often limited by vulnerability to en

safetyarxiv-cs-lg
21 Apr 2026
Safety

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

DGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

safetyarxiv-cs-cv
21 Apr 2026
Safety

Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving

DGX agent

arXiv:2512.16055v2 Announce Type: replace Abstract: Safety-critical corner cases, difficult to collect in the real world, are crucial for evaluating end-to-end autonomous driving. Adversarial interact

safetyarxiv-cs-cv
21 Apr 2026
Safety

Driving risk emerges from the required two-dimensional joint evasive acceleration

DGX agent

arXiv:2604.17841v1 Announce Type: new Abstract: Most autonomous driving safety benchmarks use time-to-collision (TTC) to assess risk and guide safe behaviour. However, TTC-based methods treat risk as

safetyarxiv-cs-ro
21 Apr 2026
Safety

Dual Alignment Between Language Model Layers and Human Sentence Processing

DGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

safetyarxiv-cs-cl
21 Apr 2026
Safety

Dynamic Emotion and Personality Profiling for Multimodal Deception Detection

DGX agent

arXiv:2604.17037v1 Announce Type: new Abstract: Deception detection is of great significance for ensuring information security and conducting public opinion analysis, with personality factors and emot

safetyarxiv-cs-cl
21 Apr 2026
Safety

Dynamic Visual-semantic Alignment for Zero-shot Learning with Ambiguous Labels

DGX agent

arXiv:2604.17710v1 Announce Type: new Abstract: Zero-shot learning (ZSL) aims to recognize unseen classes without visual instances. However, existing methods usually assume clean labels, overlooking r

safetyarxiv-cs-cv
21 Apr 2026
Safety

DynaWeb: Model-Based Reinforcement Learning of Web Agents

DGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

safetyarxiv-cs-cl
21 Apr 2026
Safety

Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing

DGX agent

arXiv:2604.17838v1 Announce Type: new Abstract: Generative modeling within constrained sets is essential for scientific and engineering applications involving physical, geometric, or safety requiremen

safetyarxiv-cs-lg
21 Apr 2026
Safety

Efficient Federated RLHF via Zeroth-Order Policy Optimization

DGX agent

arXiv:2604.17747v1 Announce Type: new Abstract: This paper considers reinforcement learning from human feedback in a federated learning setting with resource-constrained agents, such as edge devices.

safetyarxiv-cs-lg
21 Apr 2026
Safety

Emergency Stopping for Liquid-manipulating Robots

DGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

safetyarxiv-cs-ro
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Safety

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

DGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi

safetyarxiv-cs-lg
21 Apr 2026
Safety

Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models

DGX agent

arXiv:2604.16481v1 Announce Type: new Abstract: Large-scale text-to-image (T2I) diffusion models deliver remarkable visual fidelity but pose safety risks due to their capacity to reproduce undesirable

safetyarxiv-cs-cv
21 Apr 2026
Safety

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

DGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

safetyarxiv-cs-cv
21 Apr 2026
Safety

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

DGX agent

arXiv:2601.10306v2 Announce Type: replace-cross Abstract: While Reinforcement Learning (RL) has advanced LLM reasoning, applying it to long-context scenarios is hindered by sparsity of outcome rewards

safetyarxiv-cs-cl
21 Apr 2026
Safety

Evolutionary Negative Module Pruning for Better LoRA Merging

DGX agent

arXiv:2604.17753v1 Announce Type: cross Abstract: Merging multiple Low-Rank Adaptation (LoRA) experts into a single backbone is a promising approach for efficient multi-task deployment. While existing

safetyarxiv-cs-cl
21 Apr 2026
Safety

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

DGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

safetyarxiv-cs-cl
21 Apr 2026
Safety

FairLogue: Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using the All of Us Research Program

DGX agent

arXiv:2604.16450v1 Announce Type: cross Abstract: Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess de

safetyarxiv-cs-lg
21 Apr 2026
Safety

Fairness Constraints in High-Dimensional Generalized Linear Models

DGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

safetyarxiv-cs-lg
21 Apr 2026
Safety

FairNVT: Improving Fairness via Noise Injection in Vision Transformers

DGX agent

arXiv:2604.16780v1 Announce Type: new Abstract: This paper presents FairNVT, a lightweight debiasing framework for pretrained transformer-based encoders that improves both representation and predictio

safetyarxiv-cs-cv
21 Apr 2026
Safety

'Faithful to What?' On the Limits of Fidelity-Based Explanations

DGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

safetyarxiv-cs-lg
21 Apr 2026
Safety

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

DGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

safetyarxiv-cs-cl
21 Apr 2026
Safety

Fisher Decorator: Refining Flow Policy via A Local Transport Map

DGX agent

arXiv:2604.17919v1 Announce Type: new Abstract: Recent advances in flow-based offline reinforcement learning (RL) have achieved strong performance by parameterizing policies via flow matching. However

safetyarxiv-cs-lg
21 Apr 2026
Safety

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

DGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

safetyarxiv-cs-lg
21 Apr 2026
Safety

Foundation Model for Cardiac Time Series via Masked Latent Attention

DGX agent

arXiv:2603.26475v2 Announce Type: replace Abstract: Electrocardiograms (ECGs) are among the most widely available clinical signals and play a central role in cardiovascular diagnosis. While recent fou

safetyarxiv-cs-lg
21 Apr 2026
Safety

Function Words as Statistical Cues for Language Learning

DGX agent

arXiv:2601.21191v2 Announce Type: replace Abstract: What statistical properties might support learning abstract grammatical knowledge from linear input? We address this question by examining the stati

safetyarxiv-cs-cl
21 Apr 2026
Safety

Generative Semantic Communication via Alternating Dual-Domain Posterior Sampling

DGX agent

arXiv:2604.16796v1 Announce Type: new Abstract: Generative semantic communication (SemCom) harnesses pretrained generative priors to improve the perceptual quality of wireless image transmission. Exis

safetyarxiv-cs-cv
21 Apr 2026
Safety

Geometric Stability: The Missing Axis of Representations

DGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

safetyarxiv-cs-cl
21 Apr 2026
Safety

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

DGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

safetyarxiv-cs-cl
21 Apr 2026
Safety

GS-STVSR: Ultra-Efficient Continuous Spatio-Temporal Video Super-Resolution via 2D Gaussian Splatting

DGX agent

arXiv:2604.18047v1 Announce Type: new Abstract: Continuous Spatio-Temporal Video Super-Resolution (C-STVSR) aims to simultaneously enhance the spatial resolution and frame rate of videos by arbitrary

safetyarxiv-cs-cv
21 Apr 2026
Safety

HAVEN: Hierarchical Adversary-aware Visibility-Enabled Navigation with Cover Utilization using Deep Transformer Q-Networks

DGX agent

arXiv:2512.00592v2 Announce Type: replace Abstract: Autonomous navigation in partially observable environments requires agents to reason beyond immediate sensor input, exploit occlusion, and ensure sa

safetyarxiv-cs-ro
21 Apr 2026
Safety

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment

DGX agent

arXiv:2604.17928v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has proven effective for training reasoning-oriented large language models, but existing methods la

safetyarxiv-cs-lg
21 Apr 2026
Safety

Heterogeneous Self-Play for Realistic Highway Traffic Simulation

DGX agent

arXiv:2604.16406v1 Announce Type: cross Abstract: Realistic highway simulation is critical for scalable safety evaluation of autonomous vehicles, particularly for interactions that are too rare to stu

safetyarxiv-cs-lg
21 Apr 2026
Safety

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

DGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

safetyarxiv-cs-cv
21 Apr 2026
Safety

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

DGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

safetyarxiv-cs-cl
21 Apr 2026
Safety

Human-Centered Supervision for Sentiment Analysis in Telugu: A Systematic Inquiry Beyond Accuracy

DGX agent

arXiv:2508.01486v3 Announce Type: replace Abstract: Sentiment analysis for low-resource languages remains challenging in an era where interpretability, human alignment, and fairness are increasingly n

safetyarxiv-cs-cl
21 Apr 2026
Safety

Hybrid Multi-Dimensional MRI Prostate Cancer Detection via Hadamard Network-Based Bias Correction and Residual Networks

DGX agent

arXiv:2604.17107v1 Announce Type: new Abstract: Magnetic Resonance Imaging (MRI) is vital for prostate cancer (PCa) diagnosis. While advanced techniques such as Hybrid Multi-dimensional MRI (HM-MRI) h

safetyarxiv-cs-cv
21 Apr 2026
Safety

Hybrid Spectro-Temporal Fusion Framework for Structural Health Monitoring

DGX agent

arXiv:2604.16589v1 Announce Type: new Abstract: Structural health monitoring plays a critical role in ensuring structural safety by analyzing vibration responses from engineering systems. This paper p

safetyarxiv-cs-lg
21 Apr 2026
Safety

Hyperbolic Enhanced Representation Learning for Incomplete Multi-view Clustering

DGX agent

arXiv:2604.16959v1 Announce Type: cross Abstract: Incomplete Multi-View Clustering (IMVC) faces the challenge of learning discriminative representations from fragmentary observations while maintaining

safetyarxiv-cs-cv
21 Apr 2026
Safety

I agree with Connor that most folks concerned about AI have been far too coy about extinction risk, and that ControlAI is one of few excepti…

DGX agent

I agree with Connor that most folks concerned about AI have been far too coy about extinction risk, and that ControlAI is one of few exceptions. The world needs more earnest communication efforts. We

safetyconnor-leahy--x
21 Apr 2026
Safety

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to res…

DGX agent

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to restore fairness in the upcoming elections.' Modern American po

safetyelon-musk--x
21 Apr 2026
Safety

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

DGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

safetyarxiv-cs-cl
21 Apr 2026
Safety

Identifying Ethical Biases in Action Recognition Models

DGX agent

arXiv:2604.17971v1 Announce Type: new Abstract: Human Action Recognition (HAR) models are increasingly deployed in high-stakes environments, yet their fairness across different human appearances has n

safetyarxiv-cs-cv
21 Apr 2026
Safety

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI …

DGX agent

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI of telling people to be more coy about extinction risks. I h

safetyconnor-leahy--x
21 Apr 2026
Safety

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

DGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

safetyarxiv-cs-lg
21 Apr 2026
← Previous
1…235236237238239…265
Next →