AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

EVERYONE who heard about @SeismicOrg (Seismic Foundation)'s report that AI 'salience' is low NEEDS to hear about this! New research from dat…

DGX agent

EVERYONE who heard about @SeismicOrg (Seismic Foundation)'s report that AI 'salience' is low NEEDS to hear about this! New research from data science wiz @davidshor shows AI is now ahead of ABORTION a

safetyconnor-leahy--x
18 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the k…

DGX agent

Is a 92% “honest”* AI really good enough? Or a disaster waiting to happen? —- *”honest” is itself a misleading anthropomorphization of the kind Anthropic loves to promote. “Accurate” would be more acc

safetygary-marcus--x
18 Apr 2026
Safety

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants …

DGX agent

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants to be the company that didn't invest in AI when everyone els

safetygary-marcus--x
18 Apr 2026
Safety

A Mechanistic Account of Attention Sinks in GPT-2: One Circuit, Broader Implications for Mitigation

DGX agent

arXiv:2604.14722v1 Announce Type: new Abstract: Transformers commonly exhibit an attention sink: disproportionately high attention to the first position. We study this behavior in GPT-2-style models w

safetyarxiv-cs-lg
17 Apr 2026
Safety

Abstract Sim2Real through Approximate Information States

DGX agent

arXiv:2604.15289v1 Announce Type: new Abstract: In recent years, reinforcement learning (RL) has shown remarkable success in robotics when a fast and accurate simulator is available for a given task.

safetyarxiv-cs-ro
17 Apr 2026
Safety

AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation

DGX agent

arXiv:2510.01433v2 Announce Type: replace Abstract: Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background fe

safetyarxiv-cs-ro
17 Apr 2026
Safety

An unsupervised decision-support framework for multivariate biomarker analysis in athlete monitoring

DGX agent

arXiv:2604.14534v1 Announce Type: new Abstract: Purpose. Athlete monitoring is constrained by small cohorts, heterogeneous biomarker scales, limited feasibility of repeated sampling, and the lack of r

safetyarxiv-cs-lg
17 Apr 2026
Safety

Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs

DGX agent

arXiv:2509.05367v4 Announce Type: replace-cross Abstract: Large Language Model safety alignment predominantly operates on a binary assumption that requests are either safe or unsafe. This classificati

safetyarxiv-cs-ai
17 Apr 2026
Safety

Beyond Importance Sampling: Rejection-Gated Policy Optimization

DGX agent

arXiv:2604.14895v1 Announce Type: new Abstract: We propose a new perspective on policy optimization: rather than reweighting all samples by their importance ratios, an optimizer should select which sa

safetyarxiv-cs-lg
17 Apr 2026
Safety

Bias in Surface Electromyography Features across a Demographically Diverse Cohort

DGX agent

arXiv:2604.14460v1 Announce Type: cross Abstract: Neuromotor decoding from upper-limb electromyography (sEMG) can enhance human-machine interfaces and offer a more natural means of controlling prosthe

safetyarxiv-cs-lg
17 Apr 2026
Safety

Bird-SR: Bidirectional Reward-Guided Diffusion for Real-World Image Super-Resolution

DGX agent

arXiv:2602.07069v2 Announce Type: replace Abstract: Powered by multimodal text-to-image priors, diffusion-based super-resolution excels at synthesizing intricate details; however, models trained on sy

safetyarxiv-cs-cv
17 Apr 2026
Safety

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

DGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

safetyarxiv-cs-cl
17 Apr 2026
Safety

BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation

DGX agent

arXiv:2510.20151v2 Announce Type: replace Abstract: Structured texts refer to texts containing structured elements beyond plain texts, such as code snippets and placeholders. Such structured texts inc

safetyarxiv-cs-cl
17 Apr 2026
Safety

Building Trust in the Skies: A Knowledge-Grounded LLM-based Framework for Aviation Safety

DGX agent

arXiv:2604.13101v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into aviation safety decision-making represents a significant technological advancement, yet their sta

safetyarxiv-cs-ai
17 Apr 2026
Safety

Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades

DGX agent

arXiv:2604.14251v1 Announce Type: new Abstract: Monitoring LLM safety at scale requires balancing cost and accuracy: a cheap latent-space probe can screen every input, but hard cases should be escalat

safetyarxiv-cs-lg
17 Apr 2026
Safety

Calibration-Gated LLM Pseudo-Observations for Online Contextual Bandits

DGX agent

arXiv:2604.14961v1 Announce Type: new Abstract: Contextual bandit algorithms suffer from high regret during cold-start, when the learner has insufficient data to distinguish good arms from bad. We pro

safetyarxiv-cs-lg
17 Apr 2026
Safety

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

DGX agent

arXiv:2604.14892v1 Announce Type: new Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudi

safetyarxiv-cs-lg
17 Apr 2026
Safety

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs

DGX agent

arXiv:2604.14520v1 Announce Type: new Abstract: Omni-modal Large Language Models (Omni-MLLMs) promise a unified integration of diverse sensory streams. However, recent evaluations reveal a critical pe

safetyarxiv-cs-cv
17 Apr 2026
Safety

ClimateCause: Complex and Implicit Causal Structures in Climate Reports

DGX agent

arXiv:2604.14856v1 Announce Type: new Abstract: Understanding climate change requires reasoning over complex causal networks. Yet, existing causal discovery datasets predominantly capture explicit, di

safetyarxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Safety

Conformal Policy Control

DGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

safetyarxiv-cs-lg
17 Apr 2026
Safety

Context Over Content: Exposing Evaluation Faking in Automated Judges

DGX agent

arXiv:2604.15224v1 Announce Type: cross Abstract: The extit{LLM-as-a-judge} paradigm has become the operational backbone of automated AI evaluation pipelines, yet rests on an unverified assumption: th

safetyarxiv-cs-cl
17 Apr 2026
Safety

Continuous-time reinforcement learning: ellipticity enables model-free value function approximation

DGX agent

arXiv:2602.06930v2 Announce Type: replace Abstract: We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. W

safetyarxiv-cs-lg
17 Apr 2026
Safety

Controllable Video Object Insertion via Multiview Priors

DGX agent

arXiv:2604.14556v1 Announce Type: new Abstract: Video object insertion is a critical task for dynamically inserting new objects into existing environments. Previous video generation methods focus prim

safetyarxiv-cs-cv
17 Apr 2026
Safety

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

DGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

safetyarxiv-cs-cl
17 Apr 2026
Safety

Crowdsourcing of Real-world Image Annotation via Visual Properties

DGX agent

arXiv:2604.14449v1 Announce Type: new Abstract: Recent advances in data-centric artificial intelligence highlight inherent limitations in object recognition datasets. One of the primary issues stems f

safetyarxiv-cs-cv
17 Apr 2026
Safety

CURA: Clinical Uncertainty Risk Alignment for Language Model-Based Risk Prediction

DGX agent

arXiv:2604.14651v1 Announce Type: new Abstract: Clinical language models (LMs) are increasingly applied to support clinical risk prediction from free-text notes, yet their uncertainty estimates often

safetyarxiv-cs-cl
17 Apr 2026
Safety

Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value

DGX agent

arXiv:2506.13763v2 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in generative modeling. Despite more stable training, the loss of diffusion models is not in

safetyarxiv-cs-cv
17 Apr 2026
Safety

Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach

DGX agent

arXiv:2411.00361v4 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) enables agents to solve complex, long-horizon tasks by decomposing them into manageable sub-tasks. However

safetyarxiv-cs-lg
17 Apr 2026
Safety

Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error

DGX agent

arXiv:2510.26109v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly boosted the reasoning capability of language models (LMs). However, existing

safetyarxiv-cs-lg
17 Apr 2026
Safety

DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation

DGX agent

arXiv:2604.15023v1 Announce Type: new Abstract: Mobile manipulation is a fundamental capability that enables robots to interact in expansive environments such as homes and factories. Most existing app

safetyarxiv-cs-ro
17 Apr 2026
Safety

Emergent Neural Automaton Policies: Learning Symbolic Structure from Visuomotor Trajectories

DGX agent

arXiv:2603.25903v2 Announce Type: replace Abstract: Scaling robot learning to long-horizon tasks remains a formidable challenge. While end-to-end policies often lack the structural priors needed for e

safetyarxiv-cs-ro
17 Apr 2026
Safety

[Emerging Ideas] Artificial Tripartite Intelligence: A Bio-Inspired, Sensor-First Architecture for Physical AI

DGX agent

arXiv:2604.13959v1 Announce Type: new Abstract: As AI moves from data centers to robots and wearables, scaling ever-larger models becomes insufficient. Physical AI operates under tight latency, energy

safetyarxiv-cs-ai
17 Apr 2026
Safety

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

DGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

safetyarxiv-cs-lg
17 Apr 2026
Safety

Everyone’s quitting but everything’s going great. AGI will be here next week. Scout’s honor!

DGX agent

Everyone’s quitting but everything’s going great. AGI will be here next week. Scout’s honor! This story has now been updated with more details. Three leaders departed from OpenAI today: - Kevin Weil,

safetygary-marcus--x
17 Apr 2026
Safety

Exploration and Exploitation Errors Are Measurable for Language Model Agents

DGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

safetyarxiv-cs-ai
17 Apr 2026
Safety

Filling in the Mechanisms: How do LMs Learn Filler-Gap Dependencies under Developmental Constraints?

DGX agent

arXiv:2604.14459v1 Announce Type: new Abstract: For humans, filler-gap dependencies require a shared representation across different syntactic constructions. Although causal analyses suggest this may

safetyarxiv-cs-cl
17 Apr 2026
Safety

Flow with the Force Field: Learning 3D Compliant Flow Matching Policies from Force and Demonstration-Guided Simulation Data

DGX agent

arXiv:2510.02738v3 Announce Type: replace-cross Abstract: While visuomotor policy has made advancements in recent years, contact-rich tasks still remain a challenge. Robotic manipulation tasks that re

safetyarxiv-cs-lg
17 Apr 2026
Safety

Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems

DGX agent

arXiv:2510.14133v2 Announce Type: replace Abstract: Agentic AI systems, which leverage multiple autonomous agents and large language models (LLMs), are increasingly used to address complex, multi-step

safetyarxiv-cs-ai
17 Apr 2026
Safety

From Boundaries to Semantics: Prompt-Guided Multi-Task Learning for Petrographic Thin-section Segmentation

DGX agent

arXiv:2604.14805v1 Announce Type: new Abstract: Grain-edge segmentation (GES) and lithology semantic segmentation (LSS) are two pivotal tasks for quantifying rock fabric and composition. However, thes

safetyarxiv-cs-cv
17 Apr 2026
Safety

From Plausible to Causal: Counterfactual Semantics for Policy Evaluation in Simulated Online Communities

DGX agent

arXiv:2604.03920v2 Announce Type: replace Abstract: LLM-based social simulations can generate believable community interactions, enabling ``policy wind tunnels'' where governance interventions are tes

safetyarxiv-cs-cl
17 Apr 2026
Safety

From Risk to Rescue: An Agentic Survival Analysis Framework for Liquidation Prevention

DGX agent

arXiv:2604.14583v1 Announce Type: new Abstract: Decentralized Finance (DeFi) lending protocols like Aave v3 rely on over-collateralization to secure loans, yet users frequently face liquidation due to

safetyarxiv-cs-lg
17 Apr 2026
Safety

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

DGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

safetyarxiv-cs-lg
17 Apr 2026
Safety

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

DGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

safetyarxiv-cs-lg
17 Apr 2026
Safety

GSNR: Graph Smooth Null-Space Representation for Inverse Problems

DGX agent

arXiv:2602.20328v2 Announce Type: replace Abstract: Inverse problems in imaging are ill-posed, leading to infinitely many solutions consistent with the measurements due to the non-trivial null-space o

safetyarxiv-cs-cv
17 Apr 2026
Safety

Heat and Matern Kernels on Matchings

DGX agent

arXiv:2604.14331v1 Announce Type: new Abstract: Applying kernel methods to matchings is challenging due to their discrete, non-Euclidean nature. In this paper, we develop a principled framework for co

safetyarxiv-cs-lg
17 Apr 2026
Safety

Hierarchical Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text

DGX agent

arXiv:2604.14166v1 Announce Type: new Abstract: Mapping Cyber Threat Intelligence (CTI) text to MITRE ATT&CK technique IDs is a critical task for understanding adversary behaviors and automating threa

safetyarxiv-cs-cl
17 Apr 2026
Safety

Humanoid Factors: Design Principles for AI Humanoids in Human Worlds

DGX agent

arXiv:2602.10069v2 Announce Type: replace Abstract: Human factors research has long focused on optimizing environments, tools, and systems to account for human performance. Yet, as humanoid robots beg

safetyarxiv-cs-ro
17 Apr 2026
← Previous
1…241242243244245…265
Next →