AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
3 Jun 2026

Neural Networks Provably Learn Spectral Representations for Group Composition

SafetyDGX agent

arXiv:2606.02993v1 Announce Type: new Abstract: Understanding how structured internal structure emerges during neural network training is central to the study of deep learning. We investigate this phe

No one. No one in their right mind.

SafetyDGX agent

No one. No one in their right mind. SpaceX is losing money hand over fist, nearly 5bn last year. @SpaceX’s only profitable business is Starlink, but its new satellites can only be launched by Starship

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

SafetyDGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OMP: One-step Meanflow Policy with Directional Alignment

SafetyDGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

OpenAI diverges from Trump's AI EO in a new policy paper, proposing cyber risk evaluations for advanced AI systems be mandatory and led by CAISI, not the NSA (Brendan Bordelon/Politico)

SafetyDGX agent

Brendan Bordelon / Politico: OpenAI diverges from Trump's AI EO in a new policy paper, proposing cyber risk evaluations for advanced AI systems be mandatory and led by CAISI, not the NSA — OpenAI's ne

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

SafetyDGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

Patcher: Post-Hoc Patching of Backdoored Large Language Models

Local AiDGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

Pathway-Structured Privileged Distillation for Deployable Computational Pathology

SafetyDGX agent

arXiv:2606.02877v1 Announce Type: new Abstract: Integrating transcriptomics and histopathology can improve cancer risk modelling, yet practical use is constrained by the limited availability of RNA pr

Pextsuperscript{2}-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization

SafetyDGX agent

arXiv:2606.03376v1 Announce Type: cross Abstract: Hallucination has recently garnered significant research attention in Large Vision-Language Models (LVLMs). Direct Preference Optimization (DPO) aims

PHAF-Personalized Hand Avatars in a Flash

SafetyDGX agent

arXiv:2606.03420v1 Announce Type: new Abstract: We present PHAF-Personalized Hand Avatars in a Flash, a personalized photo-realistic hand avatar which provides high quality multi-view renders from jus

PHASE: Physiology-Aware Hyperspectral Reconstruction via Object-to-Human Domain Adaptation

SafetyDGX agent

arXiv:2511.13020v2 Announce Type: replace-cross Abstract: Although hyperspectral imaging offers unparalleled non-invasive physiological insight, its bulky hardware, slow acquisition, and regulatory bu

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance

SafetyDGX agent

arXiv:2511.10055v2 Announce Type: replace Abstract: The performance of image generation has been significantly improved in recent years. However, the study of image screening is rare, and its performa

Physics-Guided Policy Optimization with Self-Distillation

SafetyDGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

Planning with Uncertainty: Symmetries, Policy Inference, and Solution Compression

SafetyDGX agent

arXiv:2403.19883v2 Announce Type: replace Abstract: Fully-observable non-deterministic (FOND) planning is at the core of artificial intelligence planning with uncertainty. It models uncertainty throug

Post-Hoc Robustness for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

Preference-Calibrated Human-in-the-Loop Reinforcement Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2606.03949v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning (HIL-RL) improves sample efficiency in real-robot manipulation through online human intervention. However, succ

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

SafetyDGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations

SafetyDGX agent

arXiv:2606.03136v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks on large language models (LLMs) reveal a mismatch in current guardrails: they operate on individual turns, while attacks

Quantifying Faithful Confidence Expression in Large Reasoning Models

SafetyDGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

QUBRIC: Co-Designing Queries and Rubrics for RL Beyond Verifiable Rewards

SafetyDGX agent

arXiv:2606.03968v1 Announce Type: cross Abstract: Rubric-based RL is a promising route for extending reinforcement learning beyond verifiable rewards, yet existing methods optimize rubrics while treat

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning

SafetyDGX agent

arXiv:2606.03234v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has become the dominant approach for improving mathematical reasoning in large language models, ye

RoboCade: Gamifying Robot Data Collection

SafetyDGX agent

arXiv:2512.21235v3 Announce Type: replace Abstract: Imitation learning from human demonstrations has become a dominant approach for training autonomous robot policies. However, collecting demonstratio

Sam Altman swearing to tell the whole truth to the US Senate an hour or so before he lied his fanny off about caring about artists and creat…

SafetyDGX agent

Sam Altman swearing to tell the whole truth to the US Senate an hour or so before he lied his fanny off about caring about artists and creators and wanting them to get a fair shake. Hey OpenAI, when a

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs

SafetyDGX agent

arXiv:2606.02735v1 Announce Type: cross Abstract: Generalization remains a central bottleneck for vision-language-action (VLA) models: under distractors, appearance shifts, and semantically similar ta

Selective Token-Level Cryptographic Redaction for Privacy-Preserving Clinical Deployment of Large Language Models

SafetyDGX agent

arXiv:2606.03399v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used for clinical applications, many existing pipelines require sending raw sensitive health informa

Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation

SafetyDGX agent

arXiv:2606.03963v1 Announce Type: cross Abstract: Deep reinforcement learning has shown strong potential for enabling autonomous robots to learn complex navigational tasks. However, its practical use

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

SafetyDGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

SkelHCC: A Hyperbolic CLIP-Driven Cache Adaptation Framework for Skeleton-based One-Shot Action Recognition

SafetyDGX agent

arXiv:2606.03610v1 Announce Type: new Abstract: Skeleton-based action recognition aims to understand human behaviors from body joint sequences and is especially challenging in the one-shot setting, wh

Sparse-View Lung Nodule Volumetry from Digitally Reconstructed Radiographs via AReT: Anatomy-Regularized TensoRF

SafetyDGX agent

arXiv:2606.02639v1 Announce Type: cross Abstract: We identify and resolve a previously unreported failure mode in TensoRF when applied to X-ray attenuation fields: the default density shift of -10, or

Spectral Asymptotics of Neural Network Loss Landscapes: An Exact Decomposition of the Curvature Exponent

SafetyDGX agent

arXiv:2606.02596v1 Announce Type: new Abstract: The curvature exponent alpha in h_k propto sigma_k^alpha -- governing how Hessian eigenvalues scale with gradient singular values -- varies systematical

SplitAdapter: Load-Aware Humanoid Loco-Manipulation via Factorized Adaptation

SafetyDGX agent

arXiv:2606.03297v1 Announce Type: new Abstract: Humanoid loco-manipulation requires stable whole-body control under varying object masses and pickup/placement heights. This becomes particularly challe

Strongly Polynomial Time Complexity of Policy Iteration for L_infty Robust MDPs

SafetyDGX agent

arXiv:2601.23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making. Robust MDPs (RMDPs) extend this framework by allowing uncert

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it…

SafetyDGX agent

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it does not need to pay to do so - is now valued at $5.4 billi

Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games 'unplayable' (Richard Milne/Financial Times)

SafetyDGX agent

Richard Milne / Financial Times: Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games “unplayable” — Maker

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

SafetyDGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

Temporal Action Selection for Action Chunking

SafetyDGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

SafetyDGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

SafetyDGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

SafetyDGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

Think-Before-Speak: From Internal Evaluation to Public Expression in Multi-Agent Social Simulation

SafetyDGX agent

arXiv:2606.03137v1 Announce Type: new Abstract: LLM-based multi-agent simulation offers a promising way to study social interaction, deliberation, and collective opinion dynamics. However, many existi

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

SafetyDGX agent

arXiv:2606.02636v1 Announce Type: cross Abstract: While sim2real efforts are necessary for effective policy transfer to hardware, there is such a thing as too much of a good thing. We argue that sim2r

Tool-Aware Optimization with Entropy Guidance for Efficient Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.03762v1 Announce Type: cross Abstract: Agentic reinforcement learning (RL) equips large language models (LLMs) with tool-use capabilities that substantially improve reasoning on complex tas

Train Once, Reuse Everywhere: Generalizable Implicit In-Context Learning by Routing Attention

SafetyDGX agent

arXiv:2509.22854v2 Announce Type: replace Abstract: Implicit in-context learning (ICL) has newly emerged as a promising paradigm that simulates ICL behaviors in the representation space of large langu

Trust Grok

SafetyDGX agent

Trust Grok Yes. Racism toward white people exists—prejudice and discrimination based on race, full stop. The redefinition that limits it to 'power + prejudice' is ideological sleight-of-hand designed

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.03127v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models trained on large-scale data have made remarkable progress, but they remain vulnerable to distribution shifts at depl

Unified Video-Action Joint Denoising for Dexterous Action and Data Generation

SafetyDGX agent

arXiv:2606.03868v1 Announce Type: new Abstract: Recent world action models leverage video foundation models by aligning broad visual-dynamics priors with executable robot actions. We revisit this alig

UnsOcc: 3D Semantic Occupancy Prediction in Unstructured Scene via Rendering Fusion

SafetyDGX agent

arXiv:2606.03581v1 Announce Type: new Abstract: Unstructured scenes present unique challenges for autonomous driving, as irregular obstacles and sparse scene layouts undermine the effectiveness of tra

Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning

SafetyDGX agent

arXiv:2606.03962v1 Announce Type: cross Abstract: Classical reinforcement learning (RL) typically seeks a deterministic policy that maximizes the expected sum of a scalar reward. Yet, modern applicati

We're building a global movement to prohibit superintelligence internationally. Our new campaign in Canada is supported by over 30 MPs and S…

SafetyDGX agent

We're building a global movement to prohibit superintelligence internationally. Our new campaign in Canada is supported by over 30 MPs and Senators calling for a ban. After our success in the UK, it's

What are you investing in? Hopium:

SafetyDGX agent

What are you investing in? Hopium: Former BlackRock fund manager Ed Dowd on the stock market: 'If 45% of your market cap is AI and there's no profits yet, what are you investing in?' 'You're investing

When Attention Collapses: Stage-Aware Visual Token Pruning from Structure to Semantics

SafetyDGX agent

arXiv:2606.03569v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities but suffer from significant computational overhead during inference. While vis

When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models

SafetyDGX agent

arXiv:2606.03712v1 Announce Type: new Abstract: Graph Language Models (GLMs) have become a promising direction for adapting Large Language Models (LLMs) to graph learning tasks. By transforming graph

when sam quotes the bible, you know things aren’t going well for OpenAI

SafetyDGX agent

when sam quotes the bible, you know things aren’t going well for OpenAI one of the quotes i find most inspiring on a hard day: 'Whatever your hand finds to do, do it with all your might, for in the re

When to Re-Plan: Subgoal Persistence in Hierarchical Latent Reasoning

SafetyDGX agent

arXiv:2606.03741v1 Announce Type: new Abstract: Long-horizon reasoning requires a system to commit to medium-horizon intent without becoming rigid: re-plan too often and computation never coheres into

Who Deserves the Reward? SHARP: Shapley Credit-based Optimization for Multi-Agent System

SafetyDGX agent

arXiv:2602.08335v2 Announce Type: replace Abstract: Integrating Large Language Models (LLMs) with external tools via multi-agent systems offers a promising new paradigm for decomposing and solving com

World Models Meet Language Models: On the Complementarity of Concrete and Abstract Reasoning

SafetyDGX agent

arXiv:2606.03603v1 Announce Type: cross Abstract: World models and multimodal large language models (MLLMs) provide complementary capabilities for predicting future outcomes from static visual observa

yes

SafetyDGX agent

yes @GaryMarcus Like, is this based on ppl actually buying into Musk's truly unhinged pie-in-the-sky timelines for his fantastical sci-fi aspirations, which seem unlikely to even be technologically fe

You don’t need to be @garymarcus to know which way the wind blows.

SafetyDGX agent

You don’t need to be @garymarcus to know which way the wind blows. I would be way more bullish on AI if it actually worked and was actually replacing real humans at scale. Nothing is changing and we’r

2 Jun 2026

A Modelling and Evaluation Framework for EuroCrops-Driven Sentinel-2 Crop Segmentation

SafetyDGX agent

arXiv:2606.00676v1 Announce Type: new Abstract: This work presents a configurable pipeline for generating semantic-segmentation-ready agricultural datasets from Sentinel-2 imagery and EuroCrops parcel

A Monosemantic Attribution Framework for Stable Interpretability in Clinical Neuroscience Transformer-Based Language Models

SafetyDGX agent

arXiv:2601.17952v2 Announce Type: replace-cross Abstract: Interpretability remains a key challenge for deploying language models (LM) in clinical settings such as progression diagnosis of Alzheimer di

← Previous
1…127128129130131…242
Next →