AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion

DGX agent

arXiv:2606.25097v1 Announce Type: new Abstract: Speculative decoding accelerates inference by letting a draft model propose tokens for a target model to verify, raising a concrete safety question: at

safetyarxiv-cs-lg
25 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

StairMaster: Learning to Conquer Risky Hollow Stairs for Agile Quadrupedal Robots

DGX agent

arXiv:2606.25765v1 Announce Type: new Abstract: Climbing hollow stairs remains a challenging problem for quadruped robots due to the high risk of leg trapping, severe depth sparsity, and high-frequenc

safetyarxiv-cs-ro
25 Jun 2026
Safety

Statistically Valid Hyperparameter Selection: From Tuning to Guarantees

DGX agent

arXiv:2606.25601v1 Announce Type: cross Abstract: Hyperparameter selection is a critical step in the deployment of modern artificial intelligence systems, given the need to tune degrees of freedom suc

safetyarxiv-cs-lg
25 Jun 2026
Safety

STOCKSTAY Another Day: The Latest Addition to Turla’s Intelligence Gathering Apparatus

DGX agent

Written by: Jordan Jones Introduction Google Threat Intelligence Group (GTIG) has conducted an in-depth analysis of a .NET backdoor, tracked as STOCKSTAY, that has been continually developed and deplo

safetygoogle-cloud-ai
25 Jun 2026
Safety

Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources

DGX agent

arXiv:2606.24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challe

safetyarxiv-cs-lg
25 Jun 2026
Safety

SycoEval-EM: Sycophancy Evaluation of Large Language Models in Simulated Clinical Encounters for Emergency Care

DGX agent

arXiv:2601.16529v3 Announce Type: replace Abstract: Large language models (LLMs) deployed in clinical decision support may acquiesce to patient requests for care that conflicts with evidence-based gui

safetyarxiv-cs-ai
25 Jun 2026
Safety

Taxonomy-aware deep learning for hierarchical marine species classification in underwater imagery

DGX agent

arXiv:2606.25989v1 Announce Type: new Abstract: Automated classification of marine species from underwater imagery is essential for scalable ocean biodiversity monitoring and conservation policy. Exis

safetyarxiv-cs-cv
25 Jun 2026
Safety

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

DGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

safetyarxiv-cs-ai
25 Jun 2026
Safety

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

DGX agent

arXiv:2606.24937v1 Announce Type: cross Abstract: The Hitchhiker's Guide to Agentic AI is a comprehensive practitioner's reference for building autonomous AI systems. The book covers the full stack fr

safetyarxiv-cs-cl
25 Jun 2026
Safety

The saddest thing about AI is how many people are using it to become worse, lesser, and disempower themselves.

DGX agent

Connor Leahy expresses concern that AI tools are being used by many people in ways that diminish their capabilities and agency rather than enhance them. The post suggests that rather than leveraging A

safetyconnor-leahy--x
25 Jun 2026
Safety

The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar

DGX agent

arXiv:2606.26015v1 Announce Type: new Abstract: Text detoxification, the automated detection and mitigation of abusive and harmful content, is essential for ensuring the safety of online communities a

safetyarxiv-cs-cl
25 Jun 2026
Safety

The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapable AI Systems

DGX agent

arXiv:2606.26057v1 Announce Type: cross Abstract: AI agents are granted access to tools, APIs, and other infrastructure, making them active principals in those systems. The dominant approach places co

safetyarxiv-cs-lg
25 Jun 2026
Safety

TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control

DGX agent

arXiv:2601.14945v2 Announce Type: replace Abstract: Large-scale Vision-Language-Action (VLA) models offer semantic generalization but suffer from high inference latency, limiting them to low-frequency

safetyarxiv-cs-ro
25 Jun 2026
Safety

Towards a Bathroom-Centered Human-Building Digital Twin Framework for Indoor Safety Analysis

DGX agent

arXiv:2606.23292v2 Announce Type: replace-cross Abstract: Bathroom use is a critical safety challenge for older adults because wet surfaces, constrained layouts, limited support, and frequent posture

safetyarxiv-cs-ai
25 Jun 2026
Safety

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models

DGX agent

arXiv:2606.24962v1 Announce Type: new Abstract: Recent progress in large-scale sequence modeling has shown that a single model can learn useful representations across highly diverse data distributions

safetyarxiv-cs-lg
25 Jun 2026
Safety

Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization

DGX agent

arXiv:2505.23866v2 Announce Type: replace Abstract: Deep neural networks have been increasingly used in safety-critical applications such as medical diagnosis and autonomous driving. However, many stu

safetyarxiv-cs-lg
25 Jun 2026
Safety

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR

DGX agent

arXiv:2606.25178v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has been extended from single-domain training to multi-domain reasoning suites spanning mathematic

safetyarxiv-cs-ai
25 Jun 2026
Safety

TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction

DGX agent

arXiv:2601.22615v3 Announce Type: replace Abstract: Streaming recurrent models enable efficient 3D reconstruction by maintaining persistent state representations. However, they suffer from catastrophi

safetyarxiv-cs-cv
25 Jun 2026
Safety

Uncertainty-aware reinforcement learning for chemical language models

DGX agent

arXiv:2606.24990v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a powerful paradigm for de novo molecular design, enabling Chemical Language Models (CLMs) to navigate and explor

safetyarxiv-cs-lg
25 Jun 2026
Safety

update with some context and additional thoughts: https://open.substack.com/pub/garymarcus/p/the-generative-ai-fizzle?utm_source=app-post-st…

DGX agent

Gary Marcus discusses concerns about generative AI's limitations and potential overhyping of the technology, arguing that initial enthusiasm may not translate into sustained practical breakthroughs as

safetygary-marcus--x
25 Jun 2026
Safety

VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction

DGX agent

arXiv:2509.19297v3 Announce Type: replace Abstract: Feed-forward 3D Gaussian Splatting (3DGS) has emerged as a highly effective solution for novel view synthesis. Existing methods predominantly rely o

safetyarxiv-cs-cv
25 Jun 2026
Safety

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria…

DGX agent

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria. You can’t be pro-growth, pro-innovation, and opaque at all

safetygary-marcus--x
25 Jun 2026
Safety

What Does It Mean to Break a Distillation Defense?

DGX agent

arXiv:2606.25059v1 Announce Type: cross Abstract: Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its ou

safetyarxiv-cs-ai
25 Jun 2026
Safety

When Do Conservation Laws Survive Learned Representations? Certified Horizons for Latent World Models

DGX agent

arXiv:2606.24945v1 Announce Type: new Abstract: We ask a representation-learning question about physical world models: when does a conservation law remain certifiable after a model learns a latent rep

safetyarxiv-cs-lg
25 Jun 2026
Safety

When Does Synthetic Data Augmentation Improve Score-Based Imbalanced Classification?

DGX agent

arXiv:2606.26053v1 Announce Type: cross Abstract: Synthetic data augmentation is widely used to mitigate class imbalance, but its theoretical effects on score-based classification remain poorly unders

safetyarxiv-cs-lg
25 Jun 2026
Safety

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

DGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

safetyarxiv-cs-cl
25 Jun 2026
Safety

wouldn’t it be funny if greed undid OpenAI?

DGX agent

wouldn’t it be funny if greed undid OpenAI? 🚨BREAKING: OPENAI IPO DELAYED Altman told everyone OpenAI is worth a TRILLION dollars His own advisors warned him that retail investors aren’t buying it… ga

safetygary-marcus--x
25 Jun 2026
Safety

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

DGX agent

arXiv:2606.25034v1 Announce Type: new Abstract: General-purpose models often struggle to reliably identify and understand real-world multimodal risks, largely due to the inherent multimodal adversaria

safetyarxiv-cs-cv
25 Jun 2026
Safety

A Comparative Study of Bayesian Contextual Bandits for Real-Time Warehouse Sorter Optimization

DGX agent

arXiv:2606.23977v1 Announce Type: new Abstract: Efficient sorter diversion control of automated material handling systems (MHS) is critical for optimizing operational efficiency in large-scale warehou

safetyarxiv-cs-lg
24 Jun 2026
Safety

A Geometry-Informed Computer Vision Method for Detecting and Examining Overtaking Vehicles From A Bicycle

DGX agent

arXiv:2606.23699v1 Announce Type: new Abstract: Instrumented bicycle studies have produced direct field evidence on vehicle passing behavior, but extracting overtaking events from continuous rear-faci

safetyarxiv-cs-cv
24 Jun 2026
Safety

A global log for medical AI

DGX agent

arXiv:2510.04033v2 Announce Type: replace Abstract: Modern computer systems rely on syslog, a universal protocol that records critical events across heterogeneous infrastructure. Medicine's rapidly gr

safetyarxiv-cs-ai
24 Jun 2026
Safety

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be …

DGX agent

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be fine tuned and start treating it like the existential threat

safetyelon-musk--x
24 Jun 2026
Safety

A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Levy Process Dynamics

DGX agent

arXiv:2504.01482v3 Announce Type: replace-cross Abstract: This paper develops a model-based framework for continuous-time policy evaluation (CTPE) in reinforcement learning, incorporating both Brownia

safetyarxiv-cs-lg
24 Jun 2026
Safety

Abstractions of Queries in Ontology-Based Data Access

DGX agent

arXiv:2606.24618v1 Announce Type: new Abstract: In ontology-based data access (OBDA), multiple data sources are integrated via mappings to an ontology. We consider an OBDA setting based on existential

safetyarxiv-cs-ai
24 Jun 2026
Safety

Accelerated Stochastic Min-Max Optimization Based on Bias-corrected Momentum

DGX agent

arXiv:2406.13041v3 Announce Type: replace Abstract: Lower-bound analyses for nonconvex strongly-concave minimax optimization problems have shown that stochastic first-order algorithms require at least

safetyarxiv-cs-lg
24 Jun 2026
Safety

Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems

DGX agent

arXiv:2606.24416v1 Announce Type: new Abstract: Network operators' changing policies, service requirements, and stringent real-time constraints render existing methods designed with fixed objectives a

safetyarxiv-cs-ai
24 Jun 2026
Safety

Aligning Audio Captions with Human Preferences

DGX agent

arXiv:2509.14659v3 Announce Type: replace-cross Abstract: Current audio captioning relies on supervised learning with paired audio-caption data, which is costly to curate and may not reflect human pre

safetyarxiv-cs-lg
24 Jun 2026
Safety

An Introduction to Causal Reinforcement Learning

DGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

safetyarxiv-cs-ai
24 Jun 2026
Safety

An LLM-based Two-Stage Transformer Framework for Cross-Domain Bearing Fault Diagnosis with Limited Data

DGX agent

arXiv:2606.24459v1 Announce Type: cross Abstract: Bearing fault diagnosis faces critical challenges when dataset heterogeneity, operating condition variations, and limited labeled data occur simultane

safetyarxiv-cs-cl
24 Jun 2026
Safety

Are LLM Evaluators Really Narcissists? Sanity Checking Self-Preference Evaluations

DGX agent

arXiv:2601.22548v4 Announce Type: replace-cross Abstract: Recent research has shown that large language models (LLMs) favor their own outputs when acting as judges, undermining the integrity of automa

safetyarxiv-cs-ai
24 Jun 2026
Safety

Are Safety Guarantees in Neural Networks Safe? How to Compute Trustworthy Robustness Certifications

DGX agent

arXiv:2606.23858v1 Announce Type: cross Abstract: A primary challenge in AI safety is the existence of adversarial examples -- slightly distorted inputs that cause a neural network (NN) to misclassify

safetyarxiv-cs-ai
24 Jun 2026
Safety

ARIA: Adaptive Region-Based Importance Allocation for Conditional Diffusion Distillation

DGX agent

arXiv:2606.23898v1 Announce Type: cross Abstract: Distilling conditional diffusion models aims to transfer the behavior of a large teacher to a smaller student while preserving alignment across condit

safetyarxiv-cs-ai
24 Jun 2026
Safety

AsyncOPD: How Stale Can On-Policy Distillation Be?

DGX agent

arXiv:2606.24143v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts guided by teacher feedback and is becoming increasingly important for large language m

safetyarxiv-cs-lg
24 Jun 2026
Safety

Attention in Motion: Secure Platooning via Transformer-based Misbehavior Detection

DGX agent

arXiv:2512.15503v3 Announce Type: replace-cross Abstract: Vehicular platooning promises transformative improvements in transportation efficiency and safety through the coordination of multi-vehicle fo

safetyarxiv-cs-ai
24 Jun 2026
Safety

Audio-visual Contrastive Alignment for Diffusion-based Visual-conditioned Speech Enhancement

DGX agent

arXiv:2606.23712v1 Announce Type: cross Abstract: Audio-visual speech enhancement (AVSE) exploits visual cues such as lip movements to recover speech in noisy environments. Recent work introduced diff

safetyarxiv-cs-ai
24 Jun 2026
Safety

AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming

DGX agent

arXiv:2606.24245v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly automate complex tasks by integrating language models with external tools and environments. However, th

safetyarxiv-cs-ai
24 Jun 2026
Safety

Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.24064v1 Announce Type: new Abstract: Distilling reasoning capabilities from strong to weak language models typically involves imitating specific solution trajectories, effectively transferr

safetyarxiv-cs-ai
24 Jun 2026
Safety

Beyond U-Net: A Latent-Representation-Aligned Skip-Free Backbone for Flow-Matching Speech Enhancement

DGX agent

arXiv:2606.24745v1 Announce Type: cross Abstract: Generative models, particularly diffusion and score-based approaches, have recently achieved strong performance in speech enhancement, but their itera

safetyarxiv-cs-ai
24 Jun 2026
← Previous
1…8182838485…267
Next →