AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
1 May 2026

When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry

SafetyDGX agent

arXiv:2604.27656v1 Announce Type: cross Abstract: To preserve previously learned representations, continual learning systems must strike a balance between plasticity, the ability to acquire new knowle

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't G…

SafetyDGX agent

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't Go in the Office [hard rock] ♫ Nine-Nine-Six [hip-hop/rap] ♫ Fog F

30 Apr 2026

A Multimodal Pre-trained Network for Integrated EEG-Video Seizure Detection


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

arXiv:2604.26379v1 Announce Type: new Abstract: Reliable seizure detection in mouse models is essential for preclinical epilepsy research, yet manual review of synchronized video-EEG recordings is lab

A Scaled Three-Vehicle Platooning Platform

SafetyDGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework

SafetyDGX agent

arXiv:2604.25933v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly generate and process clinical text, scalable evaluation has become critical. LLM-as-a-Judge (LaaJ), which

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

A Survey of Safe Reinforcement Learning and Constrained MDPs: A Technical Survey on Single-Agent and Multi-Agent Safety

SafetyDGX agent

arXiv:2505.17342v2 Announce Type: replace Abstract: Safe Reinforcement Learning (SafeRL) is the subfield of reinforcement learning that explicitly deals with safety constraints during the learning and

A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

SafetyDGX agent

arXiv:2505.10924v4 Announce Type: replace-cross Abstract: Recently, AI-driven interactions with computing devices have advanced from basic prototype tools to sophisticated, LLM-based systems that emul

Accelerating RL Post-Training Rollouts via System-Integrated Speculative Decoding

SafetyDGX agent

arXiv:2604.26779v1 Announce Type: cross Abstract: RL post-training of frontier language models is increasingly bottlenecked by autoregressive rollout generation, making rollout acceleration a central

Adversarial Robustness of NTK Neural Networks

SafetyDGX agent

arXiv:2604.25965v1 Announce Type: cross Abstract: Deep learning models are widely deployed in safety-critical domains, but remain vulnerable to adversarial attacks. In this paper, we study the adversa

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alig…

SafetyDGX agent

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alignment in 2026: yeah we can not make it stop talking about go

ATLAS: An Annotation Tool for Long-horizon Robotic Action Segmentation

SafetyDGX agent

arXiv:2604.26637v1 Announce Type: cross Abstract: Annotating long-horizon robotic demonstrations with precise temporal action boundaries is crucial for training and evaluating action segmentation and

Atomic-Probe Governance for Skill Updates in Compositional Robot Policies

SafetyDGX agent

arXiv:2604.26689v1 Announce Type: cross Abstract: Skill libraries in deployed robotic systems are continually updated through fine-tuning, fresh demonstrations, or domain adaptation, yet existing type

Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints

SafetyDGX agent

arXiv:2604.26453v1 Announce Type: new Abstract: Audio-visual deepfakes have reached a level of realism that makes perceptual detection unreliable, threatening media integrity and biometric security. W

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control

SafetyDGX agent

arXiv:2604.26577v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly considered for deployment as the control component of robotic health attendants, yet their safety in this

Beyond Shortcuts: Mitigating Visual Illusions in Frozen VLMs via Qualitative Reasoning

SafetyDGX agent

arXiv:2604.26250v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved state-of-the-art performance in general visual tasks, their perceptual robustness remains remarkably b

Big Tech’s $700 billion spending on AI this year is called the ‘greatest capital misallocation in history’ https://trib.al/qUQIbrJ

SafetyDGX agent

Big Tech companies are projected to spend approximately $700 billion on AI infrastructure and development in the current year, a figure that AI researcher and entrepreneur Gary Marcus has criticized a

Classification of Public Opinion on the Free Nutritional Meal Program on YouTube Media Using the LSTM Method

SafetyDGX agent

arXiv:2604.26312v1 Announce Type: new Abstract: Public opinion towards the Free Nutritious Meal Program (MBG) on YouTube social media reflects diverse community responses. This study applies the Long

Co-Learning Port-Hamiltonian Systems and Optimal Energy-Shaping Control

SafetyDGX agent

arXiv:2604.26172v1 Announce Type: cross Abstract: We develop a physics-informed learning framework for energy-shaping control of port-Hamiltonian (pH) systems from trajectory data. The proposed approa

Consist-Retinex: One-Step Noise-Emphasized Consistency Training Accelerates High-Quality Retinex Enhancement

SafetyDGX agent

arXiv:2512.08982v2 Announce Type: replace-cross Abstract: Retinex-based low-light image enhancement benefits from separating reflectance and illumination, yet recent generative approaches often rely o

Correcting Performance Estimation Bias in Imbalanced Classification with Minority Subconcepts

SafetyDGX agent

arXiv:2604.26024v1 Announce Type: cross Abstract: Class-level evaluation can conceal substantial performance disparities across subconcepts within the same class, causing models that perform well on a

Crime Hotspot Prediction Using Deep Graph Convolutional Networks

SafetyDGX agent

arXiv:2506.13116v2 Announce Type: replace-cross Abstract: Crime hotspot prediction is critical for ensuring urban safety and effective law enforcement, it remains challenging due to complex spatial de

Cuando sucede un problema en un puesto automatizado por la IA... ¿Quién es el responsable?

SafetyDGX agent

This post discusses the liability and accountability questions that arise when problems occur in AI-automated workplaces, exploring who bears responsibility—whether the AI developer, the employer, the

Culturally Aware GenAI Risks for Youth: Perspectives from Youth, Parents, and Teachers in a Non-Western Context

SafetyDGX agent

arXiv:2604.26494v1 Announce Type: cross Abstract: Generative AI tools are widely used by youth and have introduced new privacy and safety challenges. While prior research has explored youth's safety i

Data-Centric Foundation Models in Computational Healthcare: A Survey

SafetyDGX agent

arXiv:2401.02458v3 Announce Type: replace-cross Abstract: The advent of foundation models (FMs) as an emerging suite of AI techniques has struck a wave of opportunities in computational healthcare. Th

DC-Ada: Reward-Only Decentralized Sensor Adaptation for Heterogeneous Multi-Robot Teams

SafetyDGX agent

arXiv:2604.03905v2 Announce Type: replace-cross Abstract: Heterogeneity is a defining feature of deployed multi-robot teams: platforms often differ in sensing modalities, ranges, fields of view, and f

Dear @elonmusk, If you still genuinely care about AI safety, you can’t let the Trump administration leave the AI industry almost entirely un…

SafetyDGX agent

Dear @elonmusk, If you still genuinely care about AI safety, you can’t let the Trump administration leave the AI industry almost entirely unregulated. You just can’t. - Gary The judge just instructed

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

SafetyDGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training

SafetyDGX agent

arXiv:2604.26256v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time --

Efficient and Interpretable Transformer for Counterfactual Fairness

SafetyDGX agent

arXiv:2604.26188v1 Announce Type: new Abstract: The growing reliance of machine learning models in high-stakes, highly regulated domains such as finance and insurance has created a growing tension bet

Evaluating Strategic Reasoning in Forecasting Agents

SafetyDGX agent

arXiv:2604.26106v1 Announce Type: new Abstract: Forecasting benchmarks produce accuracy leaderboards but little insight into why some forecasters are more accurate than others. We introduce Bench to t

Evaluating the Alignment Between GeoAI Explanations and Domain Knowledge in Satellite-Based Flood Mapping

SafetyDGX agent

arXiv:2604.26051v1 Announce Type: cross Abstract: The increasing number of satellites has improved the temporal resolution of Earth observation, making satellite-based flood mapping a promising approa

EvoSelect: Data-Efficient LLM Evolution for Targeted Task Adaptation

SafetyDGX agent

arXiv:2604.26170v1 Announce Type: new Abstract: Adapting large language models (LLMs) to a targeted task efficiently and effectively remains a fundamental challenge. Such adaptation often requires ite

FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility

SafetyDGX agent

arXiv:2510.26841v2 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training without data sharing, yet participants face a fundamental challenge, e.g., simult

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding

SafetyDGX agent

arXiv:2504.09925v3 Announce Type: replace Abstract: We introduce FLARE, a family of vision language models (VLMs) with a fully vision-language alignment and integration paradigm. Unlike existing appro

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how y…

SafetyDGX agent

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how you regulate open-source models that can be served by a range

From Prompt Risk to Response Risk: Paired Analysis of Safety Behavior of Large Language Model

SafetyDGX agent

arXiv:2604.26052v1 Announce Type: new Abstract: Safety evaluations of large language models (LLMs) typically report binary outcomes such as attack success rate, refusal rate, or harmful/not-harmful re

Fundamental Physics, Existential Risks and Human Futures

SafetyDGX agent

arXiv:2604.26530v1 Announce Type: cross Abstract: Over the past 25 years, I have been involved in some intriguing developments in the foundations of physics, exploring the quantum reality problem, the

Generative Bid Shading in Real-Time Bidding Advertising

SafetyDGX agent

arXiv:2508.06550v3 Announce Type: replace-cross Abstract: Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainst

Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning

SafetyDGX agent

arXiv:2601.13942v2 Announce Type: replace-cross Abstract: Large Multimodal Models (LMMs) have achieved remarkable success in visual understanding, yet they struggle with knowledge-intensive queries in

GNC-Pose: Geometry-Aware GNC-PnP for Accurate 6D Pose Estimation

SafetyDGX agent

arXiv:2512.06565v2 Announce Type: replace Abstract: We present GNC-Pose, a fully learning-free monocular 6D object pose estimation pipeline for textured objects that combines rendering-based initializ

Heterogeneous Adaptive Policy Optimization: Tailoring Optimization to Every Token's Nature

SafetyDGX agent

arXiv:2509.16591v2 Announce Type: replace Abstract: Using entropy as a measure of heterogeneity to guide optimization has emerged as a crucial research direction in Reinforcement Learning for LLMs. Ho

Hierarchical Multi-Persona Induction from User Behavioral Logs: Learning Evidence-Grounded and Truthful Personas

SafetyDGX agent

arXiv:2604.26120v1 Announce Type: new Abstract: Behavioral logs provide rich signals for user modeling, but are noisy and interleaved across diverse intents. Recent work uses LLMs to generate interpre

HiPAN: Hierarchical Posture-Adaptive Navigation for Quadruped Robots in Unstructured 3D Environments

SafetyDGX agent

arXiv:2604.26504v1 Announce Type: new Abstract: Navigating quadruped robots in unstructured 3D environments poses significant challenges, requiring goal-directed motion, effective exploration to escap

Improving Bayesian Optimization for Portfolio Management with an Adaptive Scheduling

SafetyDGX agent

arXiv:2504.13529v4 Announce Type: replace Abstract: Existing black-box portfolio management systems are prevalent in the financial industry due to commercial and safety constraints, though their perfo

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or…

SafetyDGX agent

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or to the library or maker space? I fear most will make the ea

Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification

SafetyDGX agent

arXiv:2601.15808v2 Announce Type: replace Abstract: Recent advances in Deep Research Agents (DRAs) are transforming automated knowledge discovery and problem-solving. While the majority of existing ef

John Oliver @LastWeekTonight slays a pile of greedy tech CEOs, gives props to @GaryMarcus 😻 The tides are changing, and other folks like @w…

SafetyDGX agent

John Oliver @LastWeekTonight slays a pile of greedy tech CEOs, gives props to @GaryMarcus 😻 The tides are changing, and other folks like @wendyweeww are speaking out in support. Even though I’ve alrea

Learning Vision-Based Omnidirectional Navigation: A Teacher-Student Approach Using Monocular Depth Estimation

SafetyDGX agent

arXiv:2603.01999v2 Announce Type: replace-cross Abstract: Reliable obstacle avoidance in industrial settings demands 3D scene understanding, but widely used 2D LiDAR sensors perceive only a single hor

Lifting Embodied World Models for Planning and Control

SafetyDGX agent

arXiv:2604.26182v1 Announce Type: cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are hi

Lights Out: A Nighttime UAV Localization Framework Using Thermal Imagery and Semantic 3D Maps

SafetyDGX agent

arXiv:2604.26201v1 Announce Type: new Abstract: Reliable backup localization for unmanned aerial vehicles (UAVs) operating in GNSS-denied nighttime conditions remains an open challenge due to the seve

Mapping the maturation of TCM as an adjuvant to radiotherapy

SafetyDGX agent

arXiv:2601.11923v2 Announce Type: replace Abstract: The integration of complementary medicine into oncology represents a paradigm shift that has seen to increasing adoption of Traditional Chinese Medi

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution

SafetyDGX agent

arXiv:2604.26283v1 Announce Type: cross Abstract: High-precision medical diagnosis relies not only on static imaging features but also on the implicit diagnostic memory experts instantly invoke during

Meta has dropped 10% today.

SafetyDGX agent

Meta has dropped 10% today. BREAKING: Meta stock, META, extends losses to over -10% on the day and is now on track for its biggest daily decline since October 2025. Meta has erased -170 billion in mar

Mini-Batch Class Composition Bias in Link Prediction

SafetyDGX agent

arXiv:2604.25978v1 Announce Type: cross Abstract: Prior work on node classification has shown that Graph Neural Networks (GNNs) can learn representations that transfer across graphs, when underlying g

MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text

SafetyDGX agent

arXiv:2506.02494v2 Announce Type: replace-cross Abstract: Evaluation is important for multimodal generation tasks, while traditional multimodal evaluation metrics suffer from several limitations. With

My argument since day one:

SafetyDGX agent

My argument since day one: @GaryMarcus Going all-in on LLMs, as a technology meant to be an approach to creating general purpose ML tech whose capabilities could be described as 'AGI,' *could* perhaps

Near-Optimal Cryptographic Hardness of Learning With Homogeneous Halfspaces Under Gaussian Marginals

SafetyDGX agent

arXiv:2604.26446v1 Announce Type: new Abstract: We study three problems that involve identifying homogeneous halfspaces under Gaussian distributions: agnostic learning, one-sided reliable learning, an

One underdiscussed part of the Google-DOD deal is how it blindsided the company’s own employees. One told @cogcelia that senior management h…

SafetyDGX agent

One underdiscussed part of the Google-DOD deal is how it blindsided the company’s own employees. One told @cogcelia that senior management had repeatedly insisted Google wouldn’t cave to the Pentagon’

One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety

SafetyDGX agent

arXiv:2604.25921v1 Announce Type: new Abstract: Large Language Models (LLMs) are trained to refuse harmful requests, yet they remain vulnerable to jailbreak attacks that exploit weaknesses in conversa

← Previous
1…171172173174175…212
Next →