AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
14 Apr 2026

Heartbreaking news, like losing a close friend. I learned so much at Hampshire College. For a tiny college it has had a disproportionate—and…

SafetyDGX agent

Heartbreaking news, like losing a close friend. I learned so much at Hampshire College. For a tiny college it has had a disproportionate—and exceptionally positive—effect on the world. The world is a

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at thi…

SafetyDGX agent

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at this scalable oversight problem! Progress would let AARs work o

HyperGraphPro: Progress-Aware Reinforcement Learning for Structure-Guided Hypergraph RAG

Safety
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.17755v2 Announce Type: replace Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has emerged as a promising paradigm that organizes external knowledge into structured graphs of enti

🚨 In 2 weeks, a final decision on amendments to the EU AI Act and the GDPR will be made. What is at stake is nothing other than the future …

SafetyDGX agent

🚨 In 2 weeks, a final decision on amendments to the EU AI Act and the GDPR will be made. What is at stake is nothing other than the future of Europe. Many don't know, but the stream of events leading

Influencing Humans to Conform to Preference Models for RLHF

SafetyDGX agent

arXiv:2501.06416v3 Announce Type: replace-cross Abstract: Designing a reinforcement learning from human feedback (RLHF) algorithm to approximate a human's unobservable reward function requires assumin

Interactive Learning for LLM Reasoning

SafetyDGX agent

arXiv:2509.26306v4 Announce Type: replace Abstract: Existing multi-agent learning approaches have developed interactive training environments to explicitly promote collaboration among multiple Large L

Is there a workflow to relight videos with perfect pixel-level alignment?

SafetyDGX agent

This r/StableDiffusion thread discusses community-driven approaches to relighting videos using Stable Diffusion-based tools, with a focus on the challenge of maintaining pixel-perfect alignment betwee

Isomorphic Functionalities between Ant Colony and Ensemble Learning: Part III -- Gradient Descent, Neural Plasticity, and the Emergence of Deep Intelligence

SafetyDGX agent

arXiv:2604.09677v1 Announce Type: cross Abstract: In Parts I and II of this series, we established isomorphisms between ant colony decision-making and two major families of ensemble learning: random f

Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers

SafetyDGX agent

arXiv:2604.11246v1 Announce Type: new Abstract: Evaluating the quality of model responses remains challenging in generative tasks with long-form answers, as the expected answers usually contain multip

Large Language Model as An Operator: An Experience-Driven Solution for Distribution Network Voltage Control

SafetyDGX agent

arXiv:2507.14800v2 Announce Type: replace-cross Abstract: With the advanced reasoning, contextual understanding, and information synthesis capabilities of large language models (LLMs), a novel paradig

LayerNorm Induces Recency Bias in Transformer Decoders

SafetyDGX agent

arXiv:2509.21042v3 Announce Type: replace Abstract: Causal self-attention provides positional information to Transformer decoders. Prior work has shown that stacks of causal self-attention layers alon

Layerwise Dynamics for In-Context Classification in Transformers

SafetyDGX agent

arXiv:2604.11613v1 Announce Type: cross Abstract: Transformers can perform in-context classification from a few labeled examples, yet the inference-time algorithm remains opaque. We study multi-class

LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving

SafetyDGX agent

arXiv:2512.20563v2 Announce Type: replace-cross Abstract: Simulators can generate virtually unlimited driving data, yet imitation learning policies in simulation still struggle to achieve robust close

Learning Aligned Stability in Neural ODEs Reconciling Accuracy with Robustness

SafetyDGX agent

arXiv:2509.21879v2 Announce Type: replace Abstract: Despite Neural Ordinary Differential Equations (Neural ODEs) exhibiting intrinsic robustness, existing methods often impose Lyapunov stability for f

Learning from Emptiness: De-biasing Listwise Rerankers with Content-Agnostic Probability Calibration

SafetyDGX agent

arXiv:2604.10150v1 Announce Type: new Abstract: Generative listwise reranking leverages global context for superior retrieval but is plagued by intrinsic position bias, where models exhibit structural

Learning to Unscramble: Simplifying Symbolic Expressions via Self-Supervised Oracle Trajectories

SafetyDGX agent

arXiv:2603.11164v2 Announce Type: replace-cross Abstract: We present a new self-supervised machine learning approach for symbolic simplification of complex mathematical expressions. Training data is g

Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards

SafetyDGX agent

arXiv:2510.14884v3 Announce Type: replace-cross Abstract: In high-stakes AI applications, even a single action can cause irreparable damage. However, nearly all of sequential decision-making theory as

LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment

SafetyDGX agent

arXiv:2604.10677v1 Announce Type: cross Abstract: Scaling up robot learning is hindered by the scarcity of robotic demonstrations, whereas human videos offer a vast, untapped source of interaction dat

LLM Nepotism in Organizational Governance

SafetyDGX agent

arXiv:2604.09620v1 Announce Type: cross Abstract: Large language models are increasingly used to support organizational decisions from hiring to governance, raising fairness concerns in AI-assisted ev

LPNSR: Optimal Noise-Guided Diffusion Image Super-Resolution Via Learnable Noise Prediction

SafetyDGX agent

arXiv:2603.21045v4 Announce Type: replace-cross Abstract: Diffusion-based image super-resolution (SR) aims to reconstruct high-resolution (HR) images from low-resolution (LR) observations, yet faces a

MADQRL: Distributed Quantum Reinforcement Learning Framework for Multi-Agent Environments

SafetyDGX agent

arXiv:2604.11131v1 Announce Type: new Abstract: Reinforcement learning (RL) is one of the most practical ways to learn from real-life use-cases. Motivated from the cognitive methods used by humans mak

MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization

SafetyDGX agent

arXiv:2601.07208v2 Announce Type: replace-cross Abstract: Group-Relative Policy Optimization (GRPO) has emerged as an efficient paradigm for aligning Large Language Models (LLMs), yet its efficacy is

Maine's legislature passed a bill blocking new data centers that exceed 20 MW capacity until November 2027, making it the first state to enact such a measure (Alyssa Lukpat/Wall Street Journal)

SafetyDGX agent

Alyssa Lukpat / Wall Street Journal: Maine's legislature passed a bill blocking new data centers that exceed 20 MW capacity until November 2027, making it the first state to enact such a measure — The

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

SafetyDGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

SafetyDGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

Maximum Entropy Relaxation of Multi-Way Cardinality Constraints for Synthetic Population Generation

SafetyDGX agent

arXiv:2603.22558v2 Announce Type: replace Abstract: Generating synthetic populations from aggregate statistics is a core component of microsimulation, agent-based modeling, policy analysis, and privac

MDP Planning as Policy Inference

SafetyDGX agent

arXiv:2602.17375v2 Announce Type: replace Abstract: We cast episodic Markov decision process (MDP) planning as Bayesian inference over policies. A policy is treated as the latent variable and is assig

MGA: Memory-Driven GUI Agent for Observation-Centric Interaction

SafetyDGX agent

arXiv:2510.24168v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced GUI agents, yet long-horizon automation remains constrained by two critical bot

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

SafetyDGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization

SafetyDGX agent

arXiv:2506.13331v3 Announce Type: replace Abstract: Human cognitive behavior arises from the interaction of specialized brain networks dedicated to distinct functions, such as language, logic, and soc

MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion

SafetyDGX agent

arXiv:2604.09587v1 Announce Type: new Abstract: Mobile agents can autonomously complete user-assigned tasks through GUI interactions. However, existing mainstream evaluation benchmarks, such as Androi

MonoEM-GS: Monocular Expectation-Maximization Gaussian Splatting SLAM

SafetyDGX agent

arXiv:2604.10593v1 Announce Type: new Abstract: Feed-forward geometric foundation models can infer dense point clouds and camera motion directly from RGB streams, providing priors for monocular SLAM.

Multi-Frequency Local Plasticity for Visual Representation Learning

SafetyDGX agent

arXiv:2604.09734v1 Announce Type: cross Abstract: We study how far structured architectural bias can compensate for the absence of end-to-end gradient-based representation learning in visual recogniti

Multi-Granularity Reasoning for Image Quality Assessment via Attribute-Aware Reinforcement Learning to Rank

SafetyDGX agent

arXiv:2604.09704v1 Announce Type: new Abstract: Recent advances in reasoning-induced image quality assessment (IQA) have demonstrated the power of reinforcement learning to rank (RL2R) for training vi

Multi-Modal Learning meets Genetic Programming: Analyzing Alignment in Latent Space Optimization

SafetyDGX agent

arXiv:2604.08324v2 Announce Type: replace-cross Abstract: Symbolic regression (SR) aims to discover mathematical expressions from data, a task traditionally tackled using Genetic Programming (GP) thro

Multi-ORFT: Stable Online Reinforcement Fine-Tuning for Multi-Agent Diffusion Planning in Cooperative Driving

Model ReleasesDGX agent

arXiv:2604.11734v1 Announce Type: cross Abstract: Closed-loop cooperative driving requires planners that generate realistic multimodal multi-agent trajectories while improving safety and traffic effic

Multimodal Dataset Normalization and Perceptual Validation for Music-Taste Correspondences

SafetyDGX agent

arXiv:2604.10632v1 Announce Type: cross Abstract: Collecting large, aligned cross-modal datasets for music-flavor research is difficult because perceptual experiments are costly and small by design. W

Naka-GS: A Bionics-inspired Dual-Branch Naka Correction and Progressive Point Pruning for Low-Light 3DGS

SafetyDGX agent

arXiv:2604.11142v1 Announce Type: new Abstract: Low-light conditions severely hinder 3D restoration and reconstruction by degrading image visibility, introducing color distortions, and contaminating g

NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data

SafetyDGX agent

arXiv:2604.10401v1 Announce Type: new Abstract: Inferring nationality from personal names is a critical capability for equity and bias monitoring, personalization, and a valuable tool in biomedical an

Normative Common Ground Replication (NormCoRe): Replication-by-Translation for Studying Norms in Multi-Agent AI

SafetyDGX agent

arXiv:2603.11974v2 Announce Type: replace Abstract: In the late 2010s, the fashion trend NormCore framed sameness as a signal of belonging, illustrating how norms emerge through collective coordinatio

NOSE: Neural Olfactory-Semantic Embedding with Tri-Modal Orthogonal Contrastive Learning

SafetyDGX agent

arXiv:2604.10452v1 Announce Type: new Abstract: Olfaction lies at the intersection of chemical structure, neural encoding, and linguistic perception, yet existing representation methods fail to fully

Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2504.13818v4 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as the leading approach for enhancing reasoning capabilities in large langua

OmniUMI: Towards Physically Grounded Robot Learning via Human-Aligned Multimodal Interaction

SafetyDGX agent

arXiv:2604.10647v1 Announce Type: new Abstract: UMI-style interfaces enable scalable robot learning, but existing systems remain largely visuomotor, relying primarily on RGB observations and trajector

On the Effectiveness of Textual Prompting with Lightweight Fine-Tuning for SAM3 Remote Sensing Segmentation

SafetyDGX agent

arXiv:2512.15564v2 Announce Type: replace Abstract: Remote sensing (RS) image segmentation is constrained by the limited availability of annotated data and a gap between overhead imagery and natural i

Orthogonal machine learning for conditional odds and risk ratios

SafetyDGX agent

arXiv:2604.10412v1 Announce Type: cross Abstract: Conditional effects are commonly used measures for understanding how treatment effects vary across different groups, and are often used to target trea

PACO: Proxy-Task Alignment and Online Calibration for On-the-Fly Category Discovery

SafetyDGX agent

arXiv:2604.11484v1 Announce Type: new Abstract: On-the-Fly Category Discovery (OCD) requires a model, trained on an offline support set, to recognize known classes while discovering new ones from an o

Particle Diffusion Matching: Random Walk Correspondence Search for the Alignment of Standard and Ultra-Widefield Fundus Images

SafetyDGX agent

arXiv:2604.10085v1 Announce Type: new Abstract: We propose a robust alignment technique for Standard Fundus Images (SFIs) and Ultra-Widefield Fundus Images (UWFIs), which are challenging to align due

PAT: Privacy-Preserving Adversarial Transfer for Accurate, Robust and Privacy-Preserving EEG Decoding

SafetyDGX agent

arXiv:2412.11390v3 Announce Type: replace-cross Abstract: An electroencephalogram (EEG)-based brain-computer interface (BCI) enables direct communication between the brain and external devices. Howeve

Perceptual Inductive Bias Is What You Need Before Contrastive Learning

SafetyDGX agent

arXiv:2506.01201v2 Announce Type: replace Abstract: David Marr's seminal theory of human perception stipulates that visual processing is a multi-stage process, prioritizing the derivation of boundary

PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency

SafetyDGX agent

arXiv:2603.25620v2 Announce Type: replace Abstract: Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet the

Policy Split: Incentivizing Dual-Mode Exploration in LLM Reinforcement with Dual-Mode Entropy Regularization

SafetyDGX agent

arXiv:2604.11510v1 Announce Type: cross Abstract: To encourage diverse exploration in reinforcement learning (RL) for large language models (LLMs) without compromising accuracy, we propose Policy Spli

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2509.21882v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a practical, scalable way to improve large language models on math, code, and other s

Predictions can be weapons of power. They only work if we believe them. -@carissaveliz @TEDTalks 2026

SafetyDGX agent

Carissa Véliz delivered a TED Talk in 2026 arguing that predictions function as instruments of power, shaping behavior and outcomes through the act of belief itself. Her thesis suggests that predictiv

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

SafetyDGX agent

arXiv:2603.20725v2 Announce Type: replace Abstract: Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation

SafetyDGX agent

arXiv:2604.10030v1 Announce Type: new Abstract: Video diffusion models have achieved remarkable progress in generating high-quality videos. However, these models struggle to represent the temporal suc

ProUIE: A Macro-to-Micro Progressive Learning Method for LLM-based Universal Information Extraction

SafetyDGX agent

arXiv:2604.10633v1 Announce Type: new Abstract: LLM-based universal information extraction (UIE) methods often rely on additional information beyond the original training data, which increases trainin

Proximal Supervised Fine-Tuning

SafetyDGX agent

arXiv:2508.17784v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) of foundation models often leads to poor generalization, where prior capabilities deteriorate after tuning on new

QFS-Composer: Query-focused summarization pipeline for less resourced languages

SafetyDGX agent

arXiv:2604.10687v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in text summarization, yet their effectiveness drops significantly across languages with res

Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid

SafetyDGX agent

arXiv:2511.04776v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (GenAI) represents a rapidly expanding digital infrastructure whose energy demand and associated CO2 emissi

RAG-KT: Cross-platform Explainable Knowledge Tracing with Multi-view Fusion Retrieval Generation

SafetyDGX agent

arXiv:2604.10960v1 Announce Type: new Abstract: Knowledge Tracing (KT) infers a student's knowledge state from past interactions to predict future performance. Conventional Deep Learning (DL)-based KT

← Previous
1…212213214215216…240
Next →