AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
Safety

The SpaceX IPO is a shitshow. When you peel back the layers, you realize how thoroughly corrupt it is.

DGX agent

The SpaceX IPO is a shitshow. When you peel back the layers, you realize how thoroughly corrupt it is. SpaceX’s unconventional corporate arrangements appear to benefit Elon Musk at the expense of othe

safetygary-marcus--x
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

TopoAlign: Topology-Aware Visual Representation Alignment

DGX agent

arXiv:2605.25541v1 Announce Type: cross Abstract: Neural networks encode inputs as high-dimensional vectors, known as representations, that capture how models process data by encoding task-relevant st

safetyarxiv-cs-ai
26 May 2026
Safety

TorchLean: Formalizing Neural Networks in Lean

DGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

safetyarxiv-cs-lg
26 May 2026
Safety

Toward Reliable Design of LLM-Enabled Agentic Workflows: Optimizing Latency-Reliability-Cost Tradeoffs

DGX agent

arXiv:2605.23929v1 Announce Type: new Abstract: Modern AI systems increasingly rely on workflows composed of multiple interacting agents, some powered by large language models (LLMs) and others by con

safetyarxiv-cs-ai
26 May 2026
Safety

Towards Cognitively-Faithful Decision-Making Models to Improve AI Alignment

DGX agent

arXiv:2509.04445v2 Announce Type: replace Abstract: Recent AI trends seek to align AI models to learned human-centric objectives, such as personal preferences, utility, or societal values. Using stand

safetyarxiv-cs-lg
26 May 2026
Safety

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

DGX agent

arXiv:2509.12672v2 Announce Type: replace Abstract: The volume of machine-generated content online has grown dramatically due to the widespread use of Large Language Models (LLMs), leading to new chal

safetyarxiv-cs-cl
26 May 2026
Safety

Towards Low-Gravity Planetary Exploration using Reinforcement Learning for Walking, Jumping, and In-flight Attitude Control

DGX agent

arXiv:2605.24643v1 Announce Type: new Abstract: This paper presents reinforcement learning (RL) policies for dynamic quadrupedal locomotion in planetary exploration scenarios. Building on a taskoptimi

safetyarxiv-cs-ro
26 May 2026
Safety

Towards the Connection between Activation Sparsity and Flat Minima

DGX agent

arXiv:2605.25612v1 Announce Type: cross Abstract: The observation that activation sparsity emerges in MLP blocks of standardly trained Transformers offers an opportunity to drastically reduce computat

safetyarxiv-cs-ai
26 May 2026
Safety

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

DGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

safetyarxiv-cs-ai
26 May 2026
Safety

Towards Understanding Adam Convergence on Highly Degenerate Polynomials

DGX agent

arXiv:2603.09581v2 Announce Type: replace Abstract: Adam is a widely used optimization algorithm in deep learning, yet the specific class of objective functions where it exhibits inherent advantages r

safetyarxiv-cs-lg
26 May 2026
Safety

Trait-Aware Policy Optimization for Autoregressive Multi-Trait Essay Scoring

DGX agent

arXiv:2605.25731v1 Announce Type: new Abstract: Multi-trait essay scoring aims to provide fine-grained evaluation of writing quality across multiple dimensions. However, how to effectively post-train

safetyarxiv-cs-cl
26 May 2026
Safety

Trust-Aware Joint Feature-Prediction Discrepancy for Robust Domain Adaptation

DGX agent

arXiv:2605.25119v1 Announce Type: cross Abstract: Domain adaptation aims to mitigate performance degradation caused by distribution shifts between a labeled source domain and an unlabeled or sparsely

safetyarxiv-cs-ai
26 May 2026
Safety

Uncertainty-DTW for Sequences and Visual Tokens

DGX agent

arXiv:2605.25110v1 Announce Type: cross Abstract: Aligning structured data is a fundamental problem in computer vision and machine learning, underlying tasks such as time series analysis, human action

safetyarxiv-cs-ai
26 May 2026
Safety

Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets

DGX agent

arXiv:2605.24862v1 Announce Type: new Abstract: Cross-domain offline reinforcement learning (RL) aims to learn a policy in the target domain with a limited target domain dataset and a source domain da

safetyarxiv-cs-lg
26 May 2026
Safety

UniRank: End-to-End Domain-Specific Reranking of Hybrid Text-Image Candidates

DGX agent

arXiv:2603.29897v2 Announce Type: replace-cross Abstract: Reranking is a critical component in many information retrieval pipelines. Despite remarkable progress in text-only settings, multimodal reran

safetyarxiv-cs-ai
26 May 2026
Safety

Universal Boosts, Specific Suppressors: Sparse Autoencoder Steering of Medical Vision-Language Models

DGX agent

arXiv:2605.24977v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) often hallucinate findings when generating chest X-ray reports: they fabricate findings that are not present in

safetyarxiv-cs-cl
26 May 2026
Safety

University of California STEM professors want standardized tests back due to severe math deficiencies among students: “We now observe prepar…

DGX agent

University of California STEM professors want standardized tests back due to severe math deficiencies among students: “We now observe preparation gaps so severe that instructors must reteach middle sc

safetygary-marcus--x
26 May 2026
Safety

“Unserious, empty, hallucinatory, and borderline dishonest” - the prospectus for a company that the S&P 500 is about to jam down your throat…

DGX agent

Gary Marcus criticizes a company's prospectus as containing unserious, empty, and potentially dishonest claims before its inclusion in the S&P 500 index. The post appears to highlight concerns about i

safetygary-marcus--x
26 May 2026
Safety

VEN-VL: A Visual Ensemble MoE Framework for Effective and Efficient Multi-Modal Understanding

DGX agent

arXiv:2605.25952v1 Announce Type: cross Abstract: Despite the remarkable progress achieved by recent efficient methods in accelerating multimodal understanding, they still suffer from noticeable perfo

safetyarxiv-cs-ai
26 May 2026
Safety

Vision-Guided Outdoor Flight and Obstacle Evasion via Reinforcement Learning

DGX agent

arXiv:2605.24449v1 Announce Type: cross Abstract: Although quadcopters boast impressive traversal capabilities enabled by their omnidirectional maneuverability, the need for continuous pilot control i

safetyarxiv-cs-lg
26 May 2026
Safety

What Gets Cited: Competitive GEO in AI Answer Engines

DGX agent

arXiv:2605.25517v1 Announce Type: new Abstract: AI answer engines generate answers from retrieved pages but cite only a few sources. This makes visibility depend not just on ranking, but on being cite

safetyarxiv-cs-ai
26 May 2026
Safety

When Does Multi-Agent RL Improve LLM Workflows? Workflow, Scale, and Policy-Sharing Tradeoffs

DGX agent

arXiv:2605.24202v1 Announce Type: new Abstract: Multi-agent LLM workflows route inference through specialized roles to lift end-task accuracy, but jointly training those roles with reinforcement learn

safetyarxiv-cs-ai
26 May 2026
Safety

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2605.25864v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable advancements in reasoning capabilities empowered by Reinforcement Learning with Verifiable Rewar

safetyarxiv-cs-cl
26 May 2026
Safety

when you think about it SpaceX has cumulatively lost a lot less than OpenAI so maybe it’s a bargain? or maybe we just shouldn’t value compan…

DGX agent

when you think about it SpaceX has cumulatively lost a lot less than OpenAI so maybe it’s a bargain? or maybe we just shouldn’t value companies at a trillion dollars until they actually show evidence

safetygary-marcus--x
26 May 2026
Safety

Workflow cleanup tools for ComfyUI: Visual Fold, group folding, and node alignment

DGX agent

Visual Fold is a tool for simple visual organization of ComfyUI workflows that does not turn selected nodes into a subgraph or change workflow logic. Group folding and node alignment features enable c

safetyr-stablediffusion
26 May 2026
Safety

World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy

DGX agent

arXiv:2602.06508v2 Announce Type: replace Abstract: Reinforcement learning (RL) can refine Vision-Language-Action (VLA) policies beyond behavior cloning, but real-world RL remains expensive due to ext

safetyarxiv-cs-ro
26 May 2026
Safety

X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models

DGX agent

arXiv:2605.25044v1 Announce Type: new Abstract: Learning universal policies from cross-embodied data remains a fundamental challenge in robotics. Although Vision-Language-Action (VLA) models are pre-t

safetyarxiv-cs-ro
26 May 2026
Safety

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation

DGX agent

arXiv:2510.06672v3 Announce Type: replace Abstract: Reinforcement learning algorithms such as GRPO have driven recent advances in large language model (LLM) reasoning. While scaling the number of roll

safetyarxiv-cs-lg
26 May 2026
Safety

100% agree. and this is important. and it will affect you, personally.

DGX agent

100% agree. and this is important. and it will affect you, personally. The SpaceX IPO is the most brazen retail fleecing in modern market history. NASDAQ has REWRITTEN the index rules specifically for

safetygary-marcus--x
25 May 2026
Safety

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding

DGX agent

arXiv:2605.05997v2 Announce Type: replace Abstract: Dynamic spatial reasoning from monocular video is essential for bridging visual intelligence and the physical world, yet remains challenging for vis

safetyarxiv-cs-cv
25 May 2026
Safety

A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies (New York Times)

DGX agent

New York Times: A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies — The government's A.I. Se

safetytechmeme
25 May 2026
Safety

A Novel Approach for the Counting of Wood Logs Using cGANs and Image Processing Techniques

DGX agent

arXiv:2605.23775v1 Announce Type: new Abstract: This study tackles the challenge of precise wood log counting, where applications of the proposed methodology can span from automated approaches for mat

safetyarxiv-cs-cv
25 May 2026
Safety

ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation

DGX agent

arXiv:2602.05472v2 Announce Type: replace Abstract: The quest for expert-level reasoning in Large Language Models (LLMs) has been hampered by a persistent extit{reward bottleneck}: traditional reinfor

safetyarxiv-cs-ai
25 May 2026
Safety

ARMS: Automatic Reward Shaping for Sparse-Reward Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.23562v1 Announce Type: cross Abstract: Sparse rewards are a major bottleneck in multi-agent reinforcement learning (MARL), where simultaneous learning induces non-stationarity and makes rew

safetyarxiv-cs-ai
25 May 2026
Safety

As one of the first people to warn about a possible AI backlash—years ago—let me tell you this: it’s going to get much, much worse. It break…

DGX agent

As one of the first people to warn about a possible AI backlash—years ago—let me tell you this: it’s going to get much, much worse. It breaks my heart that AI—something I spent my whole life thinking

safetygary-marcus--x
25 May 2026
Safety

Assessing Predictive Models for Fairness Based on Movement Patterns

DGX agent

arXiv:2605.23234v1 Announce Type: new Abstract: Assessing the spatial fairness of predictive models involves establishing whether they are statistically penalizing (favoring) individuals associated wi

safetyarxiv-cs-lg
25 May 2026
Safety

B-GRTO: Bootstrapped Group Relative Tool Optimization for Referring Segmentation

DGX agent

arXiv:2605.23500v1 Announce Type: new Abstract: Segmentation is a fundamental task in computer vision, underpinning pixel-level scene understanding and serving as a cornerstone for applications rangin

safetyarxiv-cs-cv
25 May 2026
Safety

BarrierSteer: LLM Safety via Learning Barrier Steering

DGX agent

arXiv:2602.20102v2 Announce Type: replace-cross Abstract: Despite the strong performance of large language models (LLMs) across diverse tasks, their susceptibility to adversarial attacks and unsafe co

safetyarxiv-cs-ai
25 May 2026
Safety

Beyond Binary Edits Robust Multimodal Knowledge Editing with Adversarial Subspace Alignment

DGX agent

arXiv:2605.23780v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) need efficient mechanisms to update knowledge without degrading existing capabilities. While intrinsic multimod

safetyarxiv-cs-ai
25 May 2026
Safety

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling

DGX agent

arXiv:2602.11146v2 Announce Type: replace-cross Abstract: Preference optimization for diffusion and flow-matching models relies on reward functions that are both discriminatively robust and computatio

safetyarxiv-cs-ai
25 May 2026
Safety

Bridging AI and Clinical Reasoning: Abductive Explanations for Alignment on Critical Symptoms

DGX agent

arXiv:2602.13985v2 Announce Type: replace Abstract: Artificial intelligence (AI) has demonstrated strong potential in clinical diagnostics, often achieving accuracy comparable to or exceeding that of

safetyarxiv-cs-ai
25 May 2026
Safety

CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training

DGX agent

arXiv:2603.06610v2 Announce Type: replace Abstract: Large language model (LLM) post-training enhances latent skills, unlocks value alignment, improves performance, and enables domain adaptation. Unfor

safetyarxiv-cs-lg
25 May 2026
Safety

CarlaNCAP: A Framework for Quantifying the Safety of Vulnerable Road Users in Infrastructure-Assisted Collective Perception Using EuroNCAP Scenarios

DGX agent

arXiv:2512.11551v2 Announce Type: replace Abstract: The growing number of road users has significantly increased the risk of accidents in recent years. Vulnerable Road Users (VRUs) are particularly at

safetyarxiv-cs-ro
25 May 2026
Safety

CBANet: A Compact Attention-Based CNN-BiLSTM Network for Aggressive Driving Event Detection

DGX agent

arXiv:2605.23471v1 Announce Type: cross Abstract: Aggressive driving is a major cause of traffic accidents and poses a serious threat to road safety. Although deep learning methods have shown promisin

safetyarxiv-cs-ai
25 May 2026
Safety

ChainFlow-VLA: Causal Flow Planning with Vision-Language Models

DGX agent

arXiv:2605.23270v1 Announce Type: cross Abstract: Current end-to-end autonomous driving systems are fundamentally limited by a mismatch between temporal causal reasoning and global trajectory consiste

safetyarxiv-cs-ai
25 May 2026
Safety

Class-Dependent Hybrid Data Augmentation for Multiclass Migraine Classification under Severe Class Imbalance

DGX agent

arXiv:2605.23453v1 Announce Type: new Abstract: We conducted a reproducibility-oriented re-evaluation of prior migraine classification studies, correcting for data leakage and metric bias. We then int

safetyarxiv-cs-lg
25 May 2026
Safety

Classical State Preparation for Variational Quantum Algorithms via Reinforcement Learning

DGX agent

arXiv:2605.23138v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) potentially offer a pathway to practical quantum advantage, but their optimization is heavily hindered by barren

safetyarxiv-cs-ai
25 May 2026
Safety

ClimateChat-300K: A Multi-Modal Facebook Dataset for Understanding Diverse Perspectives in Climate Communication

DGX agent

arXiv:2605.23326v1 Announce Type: new Abstract: We present ClimateChat-300K, a large-scale dataset of 299,329 public Facebook posts about climate change collected between May 2020 and May 2024 through

safetyarxiv-cs-cl
25 May 2026
← Previous
1…149150151152153…267
Next →