AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
14 Apr 2026

Exploring the impact of fairness-aware criteria in AutoML

SafetyDGX agent

arXiv:2604.10224v1 Announce Type: cross Abstract: Machine Learning (ML) systems are increasingly used to support decision-making processes that affect individuals. However, these systems often rely on

ExpressMM: Expressive Mobile Manipulation Behaviors in Human-Robot Interactions

SafetyDGX agent

arXiv:2604.05320v2 Announce Type: replace Abstract: Mobile manipulators are increasingly deployed in human-centered environments to perform tasks. While completing such tasks, they should also be able

F2F-AP: Flow-to-Future Asynchronous Policy for Real-time Dynamic Manipulation

SafetyDGX agent

arXiv:2604.02408v2 Announce Type: replace Abstract: Asynchronous inference has emerged as a prevalent paradigm in robotic manipulation, achieving significant progress in ensuring trajectory smoothness


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count

SafetyDGX agent

arXiv:2604.09689v1 Announce Type: cross Abstract: Machine learning progress has historically prioritized model-centric innovations, yet achievable performance is frequently capped by the intrinsic com

FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

SafetyDGX agent

arXiv:2604.10693v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has improved LLM reasoning, but models often generate explanations that appear coherent while containing unfaithful int

Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning

SafetyDGX agent

arXiv:2604.11704v1 Announce Type: cross Abstract: Deep Neural Networks are highly susceptible to shortcut learning, frequently memorizing low-dimensional spurious correlations instead of underlying ca

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

SafetyDGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection

SafetyDGX agent

arXiv:2602.10042v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability

Fatigue-PINN: Physics-Informed Fatigue-Driven Motion Modulation and Synthesis

SafetyDGX agent

arXiv:2502.19056v2 Announce Type: replace-cross Abstract: Fatigue modeling is essential for motion synthesis tasks to model human motions under fatigued conditions and biomechanical engineering applic

Federated Single-Agent Robotics: Multi-Robot Coordination Without Intra-Robot Multi-Agent Fragmentation

SafetyDGX agent

arXiv:2604.11028v1 Announce Type: cross Abstract: As embodied robots move toward fleet-scale operation, multi-robot coordination is becoming a central systems challenge. Existing approaches often trea

FGML-DG: Feynman-Inspired Cognitive Science Paradigm for Cross-Domain Medical Image Segmentation

SafetyDGX agent

arXiv:2604.10524v1 Announce Type: new Abstract: In medical image segmentation across multiple modalities (e.g., MRI, CT, etc.) and heterogeneous data sources (e.g., different hospitals and devices), D

FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling

SafetyDGX agent

arXiv:2604.11083v1 Announce Type: cross Abstract: Text-to-motion generation is driven by learning motion representations for semantic alignment with language. Existing methods rely on either continuou

FRAMER: Frequency-Aligned Self-Distillation with Adaptive Modulation Leveraging Diffusion Priors for Real-World Image Super-Resolution

SafetyDGX agent

arXiv:2512.01390v3 Announce Type: replace Abstract: Real-image super-resolution (Real-ISR) seeks to recover HR images from LR inputs with mixed, unknown degradations. While diffusion models surpass GA

FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer

SafetyDGX agent

arXiv:2604.10023v1 Announce Type: cross Abstract: With the growing availability of open-sourced adapters trained on the same diffusion backbone for diverse scenes and objects, combining these pretrain

From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning

SafetyDGX agent

arXiv:2604.11137v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into clinical decision support is critically obstructed by their opaque and often unreliable reasoning.

From Recency Bias to Stable Convergence Block Kaczmarz Methods for Online Preference Learning in Matchmaking Applications

SafetyDGX agent

arXiv:2604.09964v1 Announce Type: new Abstract: We present a family of Kaczmarz-based preference learning algorithms for real-time personalized matchmaking in reciprocal recommender systems. Post-step

GenProve: Learning to Generate Text with Fine-Grained Provenance

SafetyDGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

SafetyDGX agent

arXiv:2604.10591v1 Announce Type: cross Abstract: Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, ye

GLEaN: A Text-to-image Bias Detection Approach for Public Comprehension

SafetyDGX agent

arXiv:2604.09923v1 Announce Type: new Abstract: Text-to-image (T2I) models, and their encoded biases, increasingly shape the visual media the public encounters. While researchers have produced a rich

Google designates 'back button hijacking' as malicious, saying sites interfering with a browser's back button function may be demoted in Search starting in June (Abner Li/9to5Google)

SafetyDGX agent

Abner Li / 9to5Google: Google designates “back button hijacking” as malicious, saying sites interfering with a browser's back button function may be demoted in Search starting in June — Websites that

GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models

SafetyDGX agent

arXiv:2604.10385v1 Announce Type: new Abstract: Generating complex multi-actor scenario videos remains difficult even for state-of-the-art neural generators, while evaluating them is hard due to the l

HDR Video Generation via Latent Alignment with Logarithmic Encoding

SafetyDGX agent

arXiv:2604.11788v1 Announce Type: new Abstract: High dynamic range (HDR) imagery offers a rich and faithful representation of scene radiance, but remains challenging for generative models due to its m

Heartbreaking news, like losing a close friend. I learned so much at Hampshire College. For a tiny college it has had a disproportionate—and…

SafetyDGX agent

Heartbreaking news, like losing a close friend. I learned so much at Hampshire College. For a tiny college it has had a disproportionate—and exceptionally positive—effect on the world. The world is a

Here's my first drive on Tesla FSD V14.3.1. It feels polished vs 14.3 and it my opinion is ready for wide release. • With 14.3.1, you can no…

SafetyDGX agent

Here's my first drive on Tesla FSD V14.3.1. It feels polished vs 14.3 and it my opinion is ready for wide release. • With 14.3.1, you can now tap on the new 'P' parking icon and bring up your differen

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at thi…

SafetyDGX agent

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at this scalable oversight problem! Progress would let AARs work o

Hubble: An LLM-Driven Agentic Framework for Safe and Automated Alpha Factor Discovery

SafetyDGX agent

arXiv:2604.09601v1 Announce Type: new Abstract: Discovering predictive alpha factors in quantitative finance remains a formidable challenge due to the vast combinatorial search space and inherently lo

HyperGraphPro: Progress-Aware Reinforcement Learning for Structure-Guided Hypergraph RAG

SafetyDGX agent

arXiv:2601.17755v2 Announce Type: replace Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has emerged as a promising paradigm that organizes external knowledge into structured graphs of enti

🚨 In 2 weeks, a final decision on amendments to the EU AI Act and the GDPR will be made. What is at stake is nothing other than the future …

SafetyDGX agent

🚨 In 2 weeks, a final decision on amendments to the EU AI Act and the GDPR will be made. What is at stake is nothing other than the future of Europe. Many don't know, but the stream of events leading

Influencing Humans to Conform to Preference Models for RLHF

SafetyDGX agent

arXiv:2501.06416v3 Announce Type: replace-cross Abstract: Designing a reinforcement learning from human feedback (RLHF) algorithm to approximate a human's unobservable reward function requires assumin

Interactive Learning for LLM Reasoning

SafetyDGX agent

arXiv:2509.26306v4 Announce Type: replace Abstract: Existing multi-agent learning approaches have developed interactive training environments to explicitly promote collaboration among multiple Large L

Is there a workflow to relight videos with perfect pixel-level alignment?

SafetyDGX agent

This r/StableDiffusion thread discusses community-driven approaches to relighting videos using Stable Diffusion-based tools, with a focus on the challenge of maintaining pixel-perfect alignment betwee

Isomorphic Functionalities between Ant Colony and Ensemble Learning: Part III -- Gradient Descent, Neural Plasticity, and the Emergence of Deep Intelligence

SafetyDGX agent

arXiv:2604.09677v1 Announce Type: cross Abstract: In Parts I and II of this series, we established isomorphisms between ant colony decision-making and two major families of ensemble learning: random f

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion

SafetyDGX agent

arXiv:2604.10326v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks -- inputs designed to bypass safety mechanisms and elicit harmful responses -- despite ad

Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers

SafetyDGX agent

arXiv:2604.11246v1 Announce Type: new Abstract: Evaluating the quality of model responses remains challenging in generative tasks with long-form answers, as the expected answers usually contain multip

Large Language Model as An Operator: An Experience-Driven Solution for Distribution Network Voltage Control

SafetyDGX agent

arXiv:2507.14800v2 Announce Type: replace-cross Abstract: With the advanced reasoning, contextual understanding, and information synthesis capabilities of large language models (LLMs), a novel paradig

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

SafetyDGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

Latent Structure of Affective Representations in Large Language Models

SafetyDGX agent

arXiv:2604.07382v2 Announce Type: replace-cross Abstract: The geometric structure of latent representations in large language models (LLMs) is an active area of research, driven in part by its implica

LayerNorm Induces Recency Bias in Transformer Decoders

SafetyDGX agent

arXiv:2509.21042v3 Announce Type: replace Abstract: Causal self-attention provides positional information to Transformer decoders. Prior work has shown that stacks of causal self-attention layers alon

Layerwise Dynamics for In-Context Classification in Transformers

SafetyDGX agent

arXiv:2604.11613v1 Announce Type: cross Abstract: Transformers can perform in-context classification from a few labeled examples, yet the inference-time algorithm remains opaque. We study multi-class

LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving

SafetyDGX agent

arXiv:2512.20563v2 Announce Type: replace-cross Abstract: Simulators can generate virtually unlimited driving data, yet imitation learning policies in simulation still struggle to achieve robust close

Learning Aligned Stability in Neural ODEs Reconciling Accuracy with Robustness

SafetyDGX agent

arXiv:2509.21879v2 Announce Type: replace Abstract: Despite Neural Ordinary Differential Equations (Neural ODEs) exhibiting intrinsic robustness, existing methods often impose Lyapunov stability for f

Learning from Emptiness: De-biasing Listwise Rerankers with Content-Agnostic Probability Calibration

SafetyDGX agent

arXiv:2604.10150v1 Announce Type: new Abstract: Generative listwise reranking leverages global context for superior retrieval but is plagued by intrinsic position bias, where models exhibit structural

Learning to Focus: CSI-Free Hierarchical MARL for Reconfigurable Reflectors

SafetyDGX agent

arXiv:2604.05165v2 Announce Type: replace Abstract: Reconfigurable Intelligent Surfaces (RIS) has a potential to engineer smart radio environments for next-generation millimeter-wave (mmWave) networks

Learning to Test: Physics-Informed Representation for Dynamical Instability Detection

SafetyDGX agent

arXiv:2604.10967v1 Announce Type: new Abstract: Many safety-critical scientific and engineering systems evolve according to differential-algebraic equations (DAEs), where dynamical behavior is constra

Learning to Unscramble: Simplifying Symbolic Expressions via Self-Supervised Oracle Trajectories

SafetyDGX agent

arXiv:2603.11164v2 Announce Type: replace-cross Abstract: We present a new self-supervised machine learning approach for symbolic simplification of complex mathematical expressions. Training data is g

Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards

SafetyDGX agent

arXiv:2510.14884v3 Announce Type: replace-cross Abstract: In high-stakes AI applications, even a single action can cause irreparable damage. However, nearly all of sequential decision-making theory as

Legal2LogicICL: Improving Generalization in Transforming Legal Cases to Logical Formulas via Diverse Few-Shot Learning

SafetyDGX agent

arXiv:2604.11699v1 Announce Type: cross Abstract: This work aims to improve the generalization of logic-based legal reasoning systems by integrating recent advances in NLP with legal-domain adaptive f

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

SafetyDGX agent

arXiv:2604.10677v1 Announce Type: cross Abstract: Scaling up robot learning is hindered by the scarcity of robotic demonstrations, whereas human videos offer a vast, untapped source of interaction dat

Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Adoption in Cybersecurity Operations on Reddit

SafetyDGX agent

arXiv:2604.09998v1 Announce Type: cross Abstract: Large language models (LLMs) have recently emerged as promising tools for augmenting Security Operations Center (SOC) workflows, with vendors increasi

LLM-as-Judge on a Budget

SafetyDGX agent

arXiv:2602.15481v2 Announce Type: replace Abstract: LLM-as-a-judge has emerged as a cornerstone technique for evaluating large language models by leveraging LLM reasoning to score prompt-response pair

LLM-based Realistic Safety-Critical Driving Video Generation

SafetyDGX agent

arXiv:2507.01264v2 Announce Type: replace-cross Abstract: Designing diverse and safety-critical driving scenarios is essential for evaluating autonomous driving systems. In this paper, we propose a no

LLM Nepotism in Organizational Governance

SafetyDGX agent

arXiv:2604.09620v1 Announce Type: cross Abstract: Large language models are increasingly used to support organizational decisions from hiring to governance, raising fairness concerns in AI-assisted ev

LPNSR: Optimal Noise-Guided Diffusion Image Super-Resolution Via Learnable Noise Prediction

SafetyDGX agent

arXiv:2603.21045v4 Announce Type: replace-cross Abstract: Diffusion-based image super-resolution (SR) aims to reconstruct high-resolution (HR) images from low-resolution (LR) observations, yet faces a

MADQRL: Distributed Quantum Reinforcement Learning Framework for Multi-Agent Environments

SafetyDGX agent

arXiv:2604.11131v1 Announce Type: new Abstract: Reinforcement learning (RL) is one of the most practical ways to learn from real-life use-cases. Motivated from the cognitive methods used by humans mak

MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization

SafetyDGX agent

arXiv:2601.07208v2 Announce Type: replace-cross Abstract: Group-Relative Policy Optimization (GRPO) has emerged as an efficient paradigm for aligning Large Language Models (LLMs), yet its efficacy is

Maine's legislature passed a bill blocking new data centers that exceed 20 MW capacity until November 2027, making it the first state to enact such a measure (Alyssa Lukpat/Wall Street Journal)

SafetyDGX agent

Alyssa Lukpat / Wall Street Journal: Maine's legislature passed a bill blocking new data centers that exceed 20 MW capacity until November 2027, making it the first state to enact such a measure — The

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

SafetyDGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

SafetyDGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

Maximum Entropy Relaxation of Multi-Way Cardinality Constraints for Synthetic Population Generation

SafetyDGX agent

arXiv:2603.22558v2 Announce Type: replace Abstract: Generating synthetic populations from aggregate statistics is a core component of microsimulation, agent-based modeling, policy analysis, and privac

MDP Planning as Policy Inference

SafetyDGX agent

arXiv:2602.17375v2 Announce Type: replace Abstract: We cast episodic Markov decision process (MDP) planning as Bayesian inference over policies. A policy is treated as the latent variable and is assig

← Previous
1…199200201202203…210
Next →