AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
25 Jun 2026

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

SafetyDGX agent

arXiv:2606.25651v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare settings, accurate error detection and correction in generated or existing text

Memory Retrieval in Visuomotor Policies for Long-Horizon Robot Control

SafetyDGX agent

arXiv:2606.25136v1 Announce Type: new Abstract: General-purpose robots operating in partially observable environments, such as homes, require memory to support autonomy. They must recall diverse infor

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

SafetyDGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

SafetyDGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

More Soviet-style deceptions that I never expected to see in the United States 😢

SafetyDGX agent

Gary Marcus expresses concern about deceptive practices in the United States that he compares to Soviet-style tactics, suggesting he has observed propaganda, misinformation, or manipulative government

Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure

SafetyDGX agent

arXiv:2403.11425v4 Announce Type: replace-cross Abstract: Cancer treatments are known to introduce cardiotoxicity, negatively impacting outcomes and survivorship. Identifying cancer patients at risk o

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

SafetyDGX agent

arXiv:2606.26080v1 Announce Type: new Abstract: Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings remains prohibitively difficult: long-h

Neural Machine Translation for Low-Resource Tangkhul--English

SafetyDGX agent

arXiv:2606.25365v1 Announce Type: new Abstract: We present a study on low-resource machine translation for the Tangkhul-English (nmf-en) language pair. Tangkhul is a severely under-resourced Tibeto-Bu

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a qu…

SafetyDGX agent

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a quarter in which Elon gave them a big one time subsidy. that’s

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

SafetyDGX agent

arXiv:2606.25389v1 Announce Type: new Abstract: Extracting skills from multi-agent offline dataset improves learning efficiency via sharing task-invariant coordination skills among tasks. In settings

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity

SafetyDGX agent

arXiv:2606.26091v1 Announce Type: new Abstract: On-policy self-distillation achieves strong pass@1 accuracy by using a single model as both teacher and student, with the teacher conditioned on a corre

One Body, Two Minds: Variable Autonomy Approach for a Co-embodied Robotic Hand

SafetyDGX agent

arXiv:2606.25575v1 Announce Type: new Abstract: Assistive robotic systems face a fundamental trade-off: fully autonomous systems lack user agency, while fully user-controlled systems demand continuous

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

SafetyDGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

Paid Voices vs. Public Feeds: Interpretable Cross-Platform Theme-Based Analysis of Climate Discourse

SafetyDGX agent

arXiv:2601.13317v2 Announce Type: replace Abstract: Climate discourse online shapes public understanding of climate change and informs political and policy debate, yet it unfolds across structurally d

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca …

SafetyDGX agent

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca and @DavidSacks did everything in their power to keep the Wh

PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data

SafetyDGX agent

arXiv:2507.20068v3 Announce Type: replace Abstract: Off-policy evaluation (OPE) methods estimate the value of a new reinforcement learning (RL) policy prior to deployment. Recent advances have shown t

Phoneme-Level Mispronunciation Screening in Polish-Speaking Children with an Explainable Assistant

SafetyDGX agent

arXiv:2606.25181v1 Announce Type: cross Abstract: Early identification of speech sound errors in children is often limited by access to specialists, motivating lightweight screening tools that can ope

Point Cloud Diffusion with Global and Local Reconstruction for Instance-Level 3D Anomaly Detection

SafetyDGX agent

arXiv:2606.25740v1 Announce Type: new Abstract: 3D anomaly detection in point clouds is critical for high-precision industrial manufacturing. Reconstruction-based methods have laid a strong foundation

Position Spaces and Graphs

SafetyDGX agent

arXiv:2606.25719v1 Announce Type: new Abstract: In this paper, we introduce position graphs, a graph-based reasoning framework based on the formalization of position spaces. This framework utilizes tw

Power-Budgeted Underwater Vehicle Control via Constrained Reinforcement Learning

SafetyDGX agent

arXiv:2606.25680v1 Announce Type: new Abstract: Underwater vehicles operate from a fixed onboard energy budget that propulsion rapidly depletes, so a controller that completes its task while drawing l

Reflective VLA: In-Context Action Consequences Make VLAs Generalize

SafetyDGX agent

arXiv:2606.25215v1 Announce Type: new Abstract: Most vision-language-action (VLA) models are reactive: they predict the next action from the current instruction and observation, implicitly assuming th

Reliability-Asymmetric Spacecraft Autonomy: Co-Designing a Capable Learned GNC Stack with a Verified, Adaptation-Aware Runtime Shield

SafetyDGX agent

arXiv:2606.25366v1 Announce Type: new Abstract: Deep-space missions need onboard autonomy that is both capable and certifiable. Rule-based autonomy is certifiable but brittle, while learned autonomy i

Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See

SafetyDGX agent

arXiv:2606.25127v1 Announce Type: new Abstract: We investigate how reward design shapes the internal attention patterns of reinforcement learning agents trained for autonomous driving. Using three Per

RGB: RL Guided Whole-Body MPPI for Humanoid Control

SafetyDGX agent

arXiv:2606.25123v1 Announce Type: new Abstract: Humanoid robots require whole-body controllers that are both robust and precise in contact-rich environments. While deep reinforcement learning (RL) ach

RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2601.23075v2 Announce Type: replace Abstract: On-policy Reinforcement Learning (RL) remains a dominant paradigm for continuous control, yet standard implementations rely on Gaussian actors and r

ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.25800v1 Announce Type: new Abstract: Effective online adaptation of vision-language-action (VLA) models remains challenging, as sparse rewards provide weak supervision for high-dimensional

SAC^2-Net: Semantic Anchoring and Complementary-Consensus Fusion for Multimodal Micro-Expression Recognition

SafetyDGX agent

arXiv:2606.25542v1 Announce Type: new Abstract: Micro-expression recognition (MER) is challenging due to subtle facial movements, limited data, and the ambiguous relationship between Action Units (AUs

Safe Learning Control with Optimality and Stability Guarantees

SafetyDGX agent

arXiv:2501.15373v2 Announce Type: replace-cross Abstract: Merely pursuing performance may adversely affect safety, while a conservative policy for safe exploration will degrade the performance. How to

SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided Navigation

SafetyDGX agent

arXiv:2606.25497v1 Announce Type: new Abstract: Object-Goal Navigation (ObjNav) requires embodied agents to autonomously locate specified targets using only egocentric visual observations. Existing mo

ScalingAR: Scaling Confidence for Autoregressive Image Generation

SafetyDGX agent

arXiv:2509.26376v3 Announce Type: replace Abstract: Test-time strategies have shown remarkable success in improving large language models, but their application to next-token prediction (NTP) autoregr

Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents

SafetyDGX agent

arXiv:2606.25852v1 Announce Type: new Abstract: Group-based reinforcement learning effectively post-trains LLM agents for long-horizon, sparse-reward tasks by deriving step-level credit from trajector

Solving Markov Decision Processes with Future Information via MPC

SafetyDGX agent

arXiv:2606.24991v1 Announce Type: cross Abstract: Model Predictive Control (MPC) is widely used in industrial and robotic systems for enforcing constraints and embedding domain knowledge through finit

Specifying AI-SDLC Processes: A Protocol Language for Human-Agent Boundaries

SafetyDGX agent

arXiv:2606.20615v2 Announce Type: replace Abstract: AI agents now participate as first-class team members across the software development lifecycle, yet no specification language exists for expressing

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion

SafetyDGX agent

arXiv:2606.25097v1 Announce Type: new Abstract: Speculative decoding accelerates inference by letting a draft model propose tokens for a target model to verify, raising a concrete safety question: at

StairMaster: Learning to Conquer Risky Hollow Stairs for Agile Quadrupedal Robots

SafetyDGX agent

arXiv:2606.25765v1 Announce Type: new Abstract: Climbing hollow stairs remains a challenging problem for quadruped robots due to the high risk of leg trapping, severe depth sparsity, and high-frequenc

Statistically Valid Hyperparameter Selection: From Tuning to Guarantees

SafetyDGX agent

arXiv:2606.25601v1 Announce Type: cross Abstract: Hyperparameter selection is a critical step in the deployment of modern artificial intelligence systems, given the need to tune degrees of freedom suc

STOCKSTAY Another Day: The Latest Addition to Turla’s Intelligence Gathering Apparatus

SafetyDGX agent

Written by: Jordan Jones Introduction Google Threat Intelligence Group (GTIG) has conducted an in-depth analysis of a .NET backdoor, tracked as STOCKSTAY, that has been continually developed and deplo

Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources

SafetyDGX agent

arXiv:2606.24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challe

SycoEval-EM: Sycophancy Evaluation of Large Language Models in Simulated Clinical Encounters for Emergency Care

SafetyDGX agent

arXiv:2601.16529v3 Announce Type: replace Abstract: Large language models (LLMs) deployed in clinical decision support may acquiesce to patient requests for care that conflicts with evidence-based gui

Taxonomy-aware deep learning for hierarchical marine species classification in underwater imagery

SafetyDGX agent

arXiv:2606.25989v1 Announce Type: new Abstract: Automated classification of marine species from underwater imagery is essential for scalable ocean biodiversity monitoring and conservation policy. Exis

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

SafetyDGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

SafetyDGX agent

arXiv:2606.24937v1 Announce Type: cross Abstract: The Hitchhiker's Guide to Agentic AI is a comprehensive practitioner's reference for building autonomous AI systems. The book covers the full stack fr

The saddest thing about AI is how many people are using it to become worse, lesser, and disempower themselves.

SafetyDGX agent

Connor Leahy expresses concern that AI tools are being used by many people in ways that diminish their capabilities and agency rather than enhance them. The post suggests that rather than leveraging A

The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar

SafetyDGX agent

arXiv:2606.26015v1 Announce Type: new Abstract: Text detoxification, the automated detection and mitigation of abusive and harmful content, is essential for ensuring the safety of online communities a

The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapable AI Systems

SafetyDGX agent

arXiv:2606.26057v1 Announce Type: cross Abstract: AI agents are granted access to tools, APIs, and other infrastructure, making them active principals in those systems. The dominant approach places co

TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control

SafetyDGX agent

arXiv:2601.14945v2 Announce Type: replace Abstract: Large-scale Vision-Language-Action (VLA) models offer semantic generalization but suffer from high inference latency, limiting them to low-frequency

Towards a Bathroom-Centered Human-Building Digital Twin Framework for Indoor Safety Analysis

SafetyDGX agent

arXiv:2606.23292v2 Announce Type: replace-cross Abstract: Bathroom use is a critical safety challenge for older adults because wet surfaces, constrained layouts, limited support, and frequent posture

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models

SafetyDGX agent

arXiv:2606.24962v1 Announce Type: new Abstract: Recent progress in large-scale sequence modeling has shown that a single model can learn useful representations across highly diverse data distributions

Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization

SafetyDGX agent

arXiv:2505.23866v2 Announce Type: replace Abstract: Deep neural networks have been increasingly used in safety-critical applications such as medical diagnosis and autonomous driving. However, many stu

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR

SafetyDGX agent

arXiv:2606.25178v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has been extended from single-domain training to multi-domain reasoning suites spanning mathematic

TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction

SafetyDGX agent

arXiv:2601.22615v3 Announce Type: replace Abstract: Streaming recurrent models enable efficient 3D reconstruction by maintaining persistent state representations. However, they suffer from catastrophi

Uncertainty-aware reinforcement learning for chemical language models

SafetyDGX agent

arXiv:2606.24990v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a powerful paradigm for de novo molecular design, enabling Chemical Language Models (CLMs) to navigate and explor

update with some context and additional thoughts: https://open.substack.com/pub/garymarcus/p/the-generative-ai-fizzle?utm_source=app-post-st…

SafetyDGX agent

Gary Marcus discusses concerns about generative AI's limitations and potential overhyping of the technology, arguing that initial enthusiasm may not translate into sustained practical breakthroughs as

VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction

SafetyDGX agent

arXiv:2509.19297v3 Announce Type: replace Abstract: Feed-forward 3D Gaussian Splatting (3DGS) has emerged as a highly effective solution for novel view synthesis. Existing methods predominantly rely o

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria…

SafetyDGX agent

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria. You can’t be pro-growth, pro-innovation, and opaque at all

What Does It Mean to Break a Distillation Defense?

SafetyDGX agent

arXiv:2606.25059v1 Announce Type: cross Abstract: Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its ou

When Do Conservation Laws Survive Learned Representations? Certified Horizons for Latent World Models

SafetyDGX agent

arXiv:2606.24945v1 Announce Type: new Abstract: We ask a representation-learning question about physical world models: when does a conservation law remain certifiable after a model learns a latent rep

When Does Synthetic Data Augmentation Improve Score-Based Imbalanced Classification?

SafetyDGX agent

arXiv:2606.26053v1 Announce Type: cross Abstract: Synthetic data augmentation is widely used to mitigate class imbalance, but its theoretical effects on score-based classification remain poorly unders

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

SafetyDGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

wouldn’t it be funny if greed undid OpenAI?

SafetyDGX agent

wouldn’t it be funny if greed undid OpenAI? 🚨BREAKING: OPENAI IPO DELAYED Altman told everyone OpenAI is worth a TRILLION dollars His own advisors warned him that retail investors aren’t buying it… ga

← Previous
1…6465666768…214
Next →