AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
23 Jul 2026

Drift-Aware RL-based Wavelet Denoising for Network-Traffic Anomaly Detection

SafetyDGX agent

arXiv:2607.20011v1 Announce Type: cross Abstract: Traffic-utilisation measurements for network monitoring are corrupted by additive noise and statistical drift: time-dependent change in the signal's m

Dual Adversarial Fine-tuning for Enhancing Robustness of Large Vision Language Model

SafetyDGX agent

arXiv:2607.18958v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs), represented by LLaVA and GPT-4V, have demonstrated remarkable capabilities, their visual inputs remain vulne

Evaluating and Mitigating Gender Bias in Pre-trained Embeddings for ML-based Recruitment

SafetyDGX agent

arXiv:2607.20073v1 Announce Type: new Abstract: AI-based recruitment systems that rely on machine learning models trained on historical CV data, risk perpetuating and amplifying social biases. A key c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evolving Cache Schedules for Fast Diffusion Policy Inference

SafetyDGX agent

arXiv:2607.20293v1 Announce Type: new Abstract: Diffusion policies achieve strong visuomotor control by iteratively denoising action chunks, but repeated denoising makes real-time deployment computati

Extreme-RGMT: Continual Learning of Highly Dynamic Skills for Robust Generalist Humanoid Control

SafetyDGX agent

arXiv:2607.20110v1 Announce Type: new Abstract: Humans can progressively acquire highly dynamic motor skills while preserving reliable everyday motor abilities. In contrast, existing humanoid controll

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation

SafetyDGX agent

arXiv:2607.19395v1 Announce Type: cross Abstract: Small language models are attractive backbones for interactive agents, but direct distillation from strong teacher trajectories often turns rich multi

FSDBN: Foreground-Aware EEG-Visual Alignment via Dynamic Brain Networks

SafetyDGX agent

arXiv:2607.18344v2 Announce Type: replace-cross Abstract: EEG-based visual decoding provides a non-invasive pathway for interpreting visual semantics. However, existing methods often overlook the perc

Geometry-Guided Generative Representation for Functional Brain Graphs

SafetyDGX agent

arXiv:2511.04539v2 Announce Type: replace-cross Abstract: In network neuroscience, functional brain systems are often characterized using separate yet related graph-theoretic or spectral descriptors,

Global Building Area Estimation Products: How Accurate Are They?

SafetyDGX agent

arXiv:2607.19766v1 Announce Type: new Abstract: Geo-spatial rasters of building footprint area are useful for a variety of tasks, such as monitoring urbanization, improving energy efficiency, and trac

Great X: A Unified Multi-Modal Simulator Bridging the Sim2Real Gap for 6G

SafetyDGX agent

arXiv:2507.08716v4 Announce Type: replace Abstract: Large-scale, precisely synchronized multi-modal datasets are critical for data-driven sixth-generation (6G) wireless research, yet real-world collec

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement

SafetyDGX agent

arXiv:2607.18217v2 Announce Type: replace Abstract: Human-object centric video personalization (HOCVP) is a core task within subject-driven video generation. However, existing methods suffer from two

How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes for RLHF

SafetyDGX agent

arXiv:2607.19712v1 Announce Type: new Abstract: In RLHF pipelines, reward scoring blocks policy updates. Slow scoring bottlenecks the entire loop, since no update runs until every rollout gets a score

HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions

SafetyDGX agent

arXiv:2607.19398v1 Announce Type: new Abstract: Multi-entity compositional questions pose significant challenges to existing retrieval-augmented language models. Conventional methods fall into a dilem

In-Run Data Shapley for Adam Optimizer

SafetyDGX agent

arXiv:2602.00329v4 Announce Type: replace-cross Abstract: Reliable data attribution is essential for mitigating bias and reducing computational waste in modern machine learning, with the Shapley value

Isaac Sim-to-Real: Reinforcement Learning based Locomotion for Quadrupeds

SafetyDGX agent

arXiv:2607.18135v1 Announce Type: cross Abstract: Learning-based approaches to locomotion have risen in popularity in recent years, showing the capability for complex legged locomotion and whole-body

'It's really hard to overstate how crazy it is what happened here.' ControlAI's US Director Connor Leahy (@NPCollapse) speaks with @LindseyR…

SafetyDGX agent

'It's really hard to overstate how crazy it is what happened here.' ControlAI's US Director Connor Leahy (@NPCollapse) speaks with @LindseyReiser on CBS News, after OpenAI's own AI autonomously escape

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models

SafetyDGX agent

arXiv:2607.17572v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. Whil

LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs

SafetyDGX agent

arXiv:2602.00462v5 Announce Type: replace-cross Abstract: Transforming a large language model (LLM) into a vision-language model (VLM) can be achieved by mapping the visual tokens from a vision encode

Lean-SAM2: Target-Anchored Memory and Encoder Acceleration for SAM2

SafetyDGX agent

arXiv:2607.19811v1 Announce Type: new Abstract: The Segment Anything Model 2 (SAM2) has advanced temporal promptable segmentation, yet its deployment remains hindered by heavy memory cross-attention o

Learning Encoding-Decoding Direction Pairs to Unveil Concepts of Influence in Deep Vision Networks

SafetyDGX agent

arXiv:2509.23926v4 Announce Type: replace Abstract: Empirical evidence shows that deep vision networks often represent concepts as directions in latent space with concept information written along dir

Local Causal Structure Learning in the Presence of Latent Variables and Selection Bias

SafetyDGX agent

arXiv:2607.19866v1 Announce Type: new Abstract: Discovering the direct causes and effects of a target variable from observational data is a fundamental problem in causal discovery, with broad applicat

Masked Visual Actions for Unified World Modeling

SafetyDGX agent

arXiv:2607.19343v1 Announce Type: new Abstract: Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world

MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators

SafetyDGX agent

arXiv:2607.15273v2 Announce Type: replace Abstract: MeanFlow generators achieve fast few-step sampling by predicting average velocities over time intervals, making them attractive for efficient genera

Meta-Learning Preferences for Multilingual LLM Alignment

SafetyDGX agent

arXiv:2607.13315v2 Announce Type: replace Abstract: Unequal availability of human preference data across languages poses a significant challenge for aligning large language models in multilingual sett

Milo, a Fully Autonomous Indoor/Outdoor Robotic Guide Dog

SafetyDGX agent

arXiv:2607.19530v1 Announce Type: new Abstract: Many Blind and Low-Vision (BLV) people rely on guide dogs for moment-to-moment navigation, such as staying on path and avoiding obstacles and pedestrian

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval

SafetyDGX agent

arXiv:2607.19027v1 Announce Type: new Abstract: Zero-shot video moment retrieval aims to overcome the limitations of traditional approaches that require large-scale datasets annotated with text and it

Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment

SafetyDGX agent

arXiv:2607.19371v1 Announce Type: new Abstract: Large language model (LLM)-based Socratic tutors increasingly guide students through multi-turn questioning, but they can suffer from scaffolding collap

ModPack: An Extensible Teleoperation Interface for Bimanual Mobile Manipulation

SafetyDGX agent

arXiv:2607.19479v1 Announce Type: cross Abstract: Existing teleoperation systems are often tailored to specific robot hardware and task domains, limiting their scalability and adaptability. We present

MOF-Sleuth: Tool-Grounded Reward Alignment for Explainable Fine-Grained MOF CIF Auditing

SafetyDGX agent

arXiv:2607.19935v1 Announce Type: new Abstract: Large metal-organic framework (MOF) databases support simulation, screening, and machine learning through crystallographic information files (CIFs). Sub

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation

SafetyDGX agent

arXiv:2607.19886v1 Announce Type: new Abstract: Thermal-to-visible face translation presents fundamental challenges including geometric discontinuities, semantic attribute mismatches, and identity deg

Nonlinear Bias-Compensated Adaptive Filter and Its Application for Time-Series Prediction

SafetyDGX agent

arXiv:2607.19902v1 Announce Type: new Abstract: Most existing nonlinear adaptive filtering algorithms only account for output noise, neglecting the fact that input noise is also prevalent in practice.

Norm or Direction? Decoding Vision Mambas for High-Resolution Vision

SafetyDGX agent

arXiv:2607.18625v1 Announce Type: new Abstract: Vision Mamba models replace quadratic self-attention with linear complexity selective state space models (SSMs), emerging as efficient visual backbones.

One4Many-StablePacker: An Efficient Deep Reinforcement Learning Framework for the 3D Bin Packing Problem

SafetyDGX agent

arXiv:2510.10057v2 Announce Type: replace Abstract: The three-dimensional bin packing problem (3D-BPP) is widely applied in logistics and warehousing. Existing learning-based approaches often neglect

Online Variance Reduction for Domain Adaptation on Streaming Data

SafetyDGX agent

arXiv:2607.20374v1 Announce Type: new Abstract: This paper studies the problem of stochastic variance reduction (SVR) for the maximum mean discrepancy (MMD) and correlation alignment (CORAL) loss func

Pathologist Attention-Aligned Report Generation for Prostate Histopathology

SafetyDGX agent

arXiv:2607.19624v1 Announce Type: new Abstract: The allocation of visual attention by pathologists during cancer diagnosis is a highly selective process that critically shapes the information extracte

PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration

SafetyDGX agent

arXiv:2607.16602v2 Announce Type: replace Abstract: Action-conditioned world models are a key component of embodied AI, serving as scalable policy evaluators that reduce reliance on expensive real-wor

PGTT: Phase-Guided Terrain Traversal for Perceptive Legged Locomotion

SafetyDGX agent

arXiv:2510.18348v2 Announce Type: replace-cross Abstract: State-of-the-art perceptive Reinforcement Learning controllers for legged robots typically either (i) impose oscillator-or IK-based gait prior

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

SafetyDGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

Privileged Lesion-Context Relational Distillation for Mask-Free Skin Lesion Classification

SafetyDGX agent

arXiv:2607.18773v1 Announce Type: new Abstract: Accurate skin lesion classification can benefit from lesion segmentation masks, but requiring masks or an auxiliary segmentation model during inference

Prompt Programming for Cultural Bias and Alignment of Large Language Models

SafetyDGX agent

arXiv:2603.16827v2 Announce Type: replace Abstract: Culture shapes reasoning, values, prioritization, and strategic decision-making, yet large language models (LLMs) often exhibit cultural biases that

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference

SafetyDGX agent

arXiv:2607.20327v1 Announce Type: new Abstract: Large language models (LLMs) provide strong reasoning capabilities but are expensive to serve at scale, whereas small language models (SLMs) are cheaper

Really insightful Substack post from @GaryMarcus ref the recent OpenAI -HuggingFace incident. It provided much need context around initial m…

SafetyDGX agent

Really insightful Substack post from @GaryMarcus ref the recent OpenAI -HuggingFace incident. It provided much need context around initial media reporting. Two points especially stood out to me: #Tech

REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning

SafetyDGX agent

arXiv:2607.19450v1 Announce Type: cross Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool

Rewarding Better Thinking for LLM Preference Alignment

SafetyDGX agent

arXiv:2607.19824v1 Announce Type: new Abstract: LLM preference alignment aims to optimize models toward human preferences across diverse user instructions. Reinforcement learning has become a major po

Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses

SafetyDGX agent

arXiv:2607.19387v1 Announce Type: cross Abstract: Surrogate modeling for high-dimensional nonlinear dynamical systems that exhibit chaos requires mechanisms that preserve not only pointwise accuracy b

SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding

SafetyDGX agent

arXiv:2503.06437v3 Announce Type: replace-cross Abstract: We present SEED (Semantic Evaluation for Visual Brain Decoding), a novel metric for evaluating the semantic decoding performance of visual bra

SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments

SafetyDGX agent

arXiv:2607.20207v1 Announce Type: new Abstract: Practical robotic grasping in complex scenes requires both 3D spatial reasoning and alignment with task-specific requirements. Vision-language models (V

Self-Explaining Reinforcement Learning for Mobile Network Resource Allocation

SafetyDGX agent

arXiv:2509.14925v2 Announce Type: replace Abstract: Deep reinforcement learning (DRL) methods, though powerful, often lack transparency, which limits their adoption in critical domains. We apply Self-

SemICP: Semantic Non-Rigid Point Cloud Registration with Elastic Energy Regularization

SafetyDGX agent

arXiv:2503.00972v4 Announce Type: replace Abstract: Purpose: Accurate point cloud registration is essential in computer-aided interventions (CAI) to align multi-modal medical images for intraoperative

Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics

SafetyDGX agent

arXiv:2607.19389v1 Announce Type: cross Abstract: As AI-driven Decision Makers (ADMs) influence our socioeconomic reality, their roles in both enhancing efficiency and amplifying the social biases hav

SLPO: Scaling Latent Reasoning via a Surrogate Policy

SafetyDGX agent

arXiv:2607.19691v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become the predominant recipe for eliciting test-time scaling in explicit Chain-of-Thought reasoner

SOPD-SocialNav: Selective On-Policy Distillation for Vision-Language Social Navigation

SafetyDGX agent

arXiv:2607.19850v1 Announce Type: new Abstract: Vision-language models have shown strong potential for social robot navigation by leveraging rich semantic understanding of complex environments and hum

Spatiotemporal Facial Action Unit Detection using Twin Cycle Autoencoders for Driver Monitoring

SafetyDGX agent

arXiv:2607.16760v2 Announce Type: replace-cross Abstract: Driver monitoring systems (DMS) increasingly rely on facial cues to infer drowsiness, distraction, and cognitive load in real time. Facial Act

Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning

SafetyDGX agent

arXiv:2607.18722v2 Announce Type: replace-cross Abstract: Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byp

Statevector-Referenced Geometry Survival of a Four-Qubit ZZ Quantum Kernel on IBM Quantum Hardware: A Fixed-Subset Diagnostic Across Three Execution Configurations

SafetyDGX agent

arXiv:2607.20377v1 Announce Type: cross Abstract: Quantum-kernel methods encode a dataset's geometry in a Gram matrix, so learning claims on hardware kernels assume the intended geometry survives exec

STEREOFLOW: Progressive Stereo Matching with StereoDiT and Transition Flow Matching

SafetyDGX agent

arXiv:2607.19986v1 Announce Type: new Abstract: Stereo matching is a fundamental task in 3D reconstruction. Despite remarkable advances, the prevailing paradigms formulate stereo matching as a determi

Stochastic Primal-Dual Decoding for Multiobjective Generative Recommender Systems

SafetyDGX agent

arXiv:2607.19357v1 Announce Type: new Abstract: Recent advances in recommender systems (RS) have shown substantial performance gains through generative modelling. In practice, recommendation often inv

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation

SafetyDGX agent

arXiv:2607.20174v1 Announce Type: cross Abstract: Existing human--object interaction (HOI) video generation methods are largely limited to offline short-video generation with complex driving condition

Stress Testing Concept Erasure with Large Language Model Agents

SafetyDGX agent

arXiv:2607.17890v2 Announce Type: replace Abstract: Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. Howeve

← Previous
1…7778798081…242
Next →