AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

MajinBook: An open catalogue of digitally mediated world literature

DGX agent

arXiv:2511.11412v5 Announce Type: replace Abstract: This data paper introduces MajinBook, an open catalogue designed to facilitate the use of shadow libraries-such as Library Genesis and Z-Library-for

safetyarxiv-cs-cl
13 May 2026
Safety

Maximin Robust Bayesian Experimental Design

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2603.14094v2 Announce Type: replace-cross Abstract: We address the brittleness of Bayesian experimental design under model misspecification by formulating the problem as a max--min game between

safetyarxiv-cs-lg
13 May 2026
Safety

Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

DGX agent

arXiv:2605.12070v1 Announce Type: new Abstract: Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization

safetyarxiv-cs-lg
13 May 2026
Safety

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics

DGX agent

arXiv:2605.12119v1 Announce Type: new Abstract: Generative novel view synthesis faces a fundamental dilemma: geometric priors provide spatial alignment but become sparse and inaccurate under view chan

safetyarxiv-cs-cv
13 May 2026
Safety

Model-based Bootstrap of Controlled Markov Chains

DGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

safetyarxiv-cs-lg
13 May 2026
Safety

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible.…

DGX agent

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible. Even less plausible is the claim that there will be an AI j

safetygary-marcus--x
13 May 2026
Safety

More Than Meets the Eye: A Semantics-Aware Traffic Augmentation Framework for Generalizable Website Fingerprinting

DGX agent

arXiv:2605.11402v1 Announce Type: new Abstract: Deep learning-based website fingerprinting has emerged as an effective technique for inferring the websites users visit. Although existing methods achie

safetyarxiv-cs-lg
13 May 2026
Safety

Morphologically Equivariant Flow Matching for Bimanual Mobile Manipulation

DGX agent

arXiv:2605.12228v1 Announce Type: new Abstract: Mobile manipulation requires coordinated control of high-dimensional, bimanual robots. Imitation learning methods have been broadly used to solve these

safetyarxiv-cs-ro
13 May 2026
Safety

Multimodal Abstractive Summarization of Instructional Videos with Vision-Language Models

DGX agent

arXiv:2605.11959v1 Announce Type: cross Abstract: Multimodal video summarization requires visual features that align semantically with language generation. Traditional approaches rely on CNN features

safetyarxiv-cs-cl
13 May 2026
Safety

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs …

DGX agent

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs we get from AI. By scraping the open web, LLMs are unwitting

safetygary-marcus--x
13 May 2026
Safety

Newton's Lantern: A Reinforcement Learning Framework for Finetuning AC Power Flow Warm Start Models

DGX agent

arXiv:2605.11102v1 Announce Type: new Abstract: Neural warm starts can sharply reduce the number of Newton-Raphson iterations required to solve the AC power flow problem, but existing supervised appro

safetyarxiv-cs-lg
13 May 2026
Safety

@Nima292 LLMs are not AGI but will lead to some job losses; true AGI would likely lead to many more.

DGX agent

Gary Marcus argues that current large language models (LLMs) do not constitute artificial general intelligence (AGI), though they will cause some job displacement. He suggests that true AGI, if achiev

safetygary-marcus--x
13 May 2026
Safety

no remorse, just further evasion. so slick; so dangerous.

DGX agent

no remorse, just further evasion. so slick; so dangerous. 🚨 SEVEN OPENAI INSIDERS HAVE ACCUSED SAM ALTMAN OF LYING Today on cross, Musk's lawyer walked Altman through all of them: >Ilya Sutskever (co-

safetygary-marcus--x
13 May 2026
Safety

Off-Policy Learning with Limited Supply

DGX agent

arXiv:2603.18702v3 Announce Type: replace Abstract: We study off-policy learning (OPL) in contextual bandits, which plays a key role in a wide range of real-world applications such as recommendation s

safetyarxiv-cs-lg
13 May 2026
Safety

Offline Constrained Reinforcement Learning under Partial Data Coverage

DGX agent

arXiv:2505.17506v2 Announce Type: replace-cross Abstract: We study offline constrained reinforcement learning with general function approximation in discounted constrained Markov decision processes. P

safetyarxiv-cs-lg
13 May 2026
Safety

Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation

DGX agent

arXiv:2605.11479v1 Announce Type: new Abstract: Policy evaluation is a fundamental component of the development and deployment pipeline for robotic policies. In modern manipulation systems, this probl

safetyarxiv-cs-ro
13 May 2026
Safety

OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning

DGX agent

arXiv:2605.12400v1 Announce Type: new Abstract: We study {on-policy self-distillation} (OPSD), where a language model improves its reasoning ability by distilling privileged teacher distributions alon

safetyarxiv-cs-lg
13 May 2026
Safety

OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation

DGX agent

arXiv:2605.12480v1 Announce Type: new Abstract: Recent advances in joint audio-video generation have been remarkable, yet real-world applications demand strong per-modality fidelity, cross-modal align

safetyarxiv-cs-cv
13 May 2026
Safety

On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

DGX agent

arXiv:2605.12206v1 Announce Type: new Abstract: In reinforcement learning (RL), agents acting in partially observable Markov decision processes (POMDPs) must rely on memory, typically encoded in a rec

safetyarxiv-cs-lg
13 May 2026
Safety

Optimal Policy Learning under Budget and Coverage Constraints

DGX agent

arXiv:2605.12235v1 Announce Type: cross Abstract: We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and

safetyarxiv-cs-lg
13 May 2026
Safety

Optimizing 4D Wires for Sparse 3D Abstraction

DGX agent

arXiv:2605.11977v1 Announce Type: new Abstract: We present a unified framework for 3D geometric abstraction using a single continuous 4D wire, parameterized as a B-spline with spatial coordinates and

safetyarxiv-cs-cv
13 May 2026
Safety

ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models

DGX agent

arXiv:2605.12446v1 Announce Type: cross Abstract: Large language models (LLMs) often produce answers with high certainty even when they are incorrect, making reliable confidence estimation essential f

safetyarxiv-cs-cl
13 May 2026
Safety

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has…

DGX agent

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has been doubling every few months, and this rate has become fa

safetygary-marcus--x
13 May 2026
Safety

OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness

DGX agent

arXiv:2605.11525v1 Announce Type: new Abstract: Missing values are routinely treated as defects to be eliminated through deletion or imputation prior to machine learning. In many applied domains, howe

safetyarxiv-cs-lg
13 May 2026
Safety

Oversmoothing as Representation Degeneracy in Neural Sheaf Diffusion

DGX agent

arXiv:2605.11178v1 Announce Type: new Abstract: Neural Sheaf Diffusion (NSD) generalizes diffusion-based Graph Neural Networks by replacing scalar graph Laplacians with sheaf Laplacians whose learned

safetyarxiv-cs-lg
13 May 2026
Safety

Physics-Informed Graph Neural Networks for Frequency-Aware Optical Aberration Correction

DGX agent

arXiv:2512.05683v2 Announce Type: replace Abstract: Optical aberrations significantly degrade image quality in microscopy, particularly when imaging deeper into samples. These aberrations arise from d

safetyarxiv-cs-cv
13 May 2026
Safety

PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian Splatting

DGX agent

arXiv:2605.11520v1 Announce Type: new Abstract: Unsupervised point cloud segmentation is critical for embodied artificial intelligence and autonomous driving, as it mitigates the prohibitive cost of d

safetyarxiv-cs-cv
13 May 2026
Safety

Position: Universal Aesthetic Alignment Narrows Artistic Expression

DGX agent

arXiv:2512.11883v3 Announce Type: replace-cross Abstract: Over-aligning image generation models to a generalized aesthetic preference conflicts with user intent, particularly when 'anti-aesthetic' out

safetyarxiv-cs-cv
13 May 2026
Safety

Post-ADC Inference: Valid Inference After Active Data Collection

DGX agent

arXiv:2605.11511v1 Announce Type: cross Abstract: The validity of statistical inference depends critically on how data are collected. When data gathered through active data collection (ADC) are reused

safetyarxiv-cs-lg
13 May 2026
Safety

Predictive Maps of Multi-Agent Reasoning: A Successor-Representation Spectrum for LLM Communication Topologies

DGX agent

arXiv:2605.11453v1 Announce Type: cross Abstract: Practitioners deploying multi-agent large language model (LLM) systems must currently choose between communication topologies such as chain, star, mes

safetyarxiv-cs-lg
13 May 2026
Safety

Pretraining Exposure Explains Popularity Judgments in Large Language Models

DGX agent

arXiv:2605.12382v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic preferences for well-known entities, a phenomenon often attributed to popularity bias. However, the exte

safetyarxiv-cs-cl
13 May 2026
Safety

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses

DGX agent

arXiv:2605.11535v1 Announce Type: new Abstract: Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are eithe

safetyarxiv-cs-lg
13 May 2026
Safety

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling

DGX agent

arXiv:2605.11299v1 Announce Type: cross Abstract: Code generation is typically trained in the primal space of programs: a model produces a candidate solution and receives sparse execution feedback, of

safetyarxiv-cs-cl
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Safety

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks

DGX agent

arXiv:2509.06701v2 Announce Type: replace Abstract: We develop a theory of intelligent agency grounded in probabilistic modeling for neural models. Agents are represented as outcome distributions with

safetyarxiv-cs-lg
13 May 2026
Safety

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s mo…

DGX agent

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s models? never forget that humans are remarkably intelligent (t

safetygary-marcus--x
13 May 2026
Safety

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

DGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

safetyarxiv-cs-cl
13 May 2026
Safety

Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning

DGX agent

arXiv:2605.11289v1 Announce Type: new Abstract: Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct dis

safetyarxiv-cs-lg
13 May 2026
Safety

Rainbow Deep Q-Learning with Kinematics-Aware Design for Cooperative Delta and 3-RRS Parallel Robot Insertion

DGX agent

arXiv:2605.11697v1 Announce Type: new Abstract: This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipula

safetyarxiv-cs-ro
13 May 2026
Safety

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking

DGX agent

arXiv:2605.11151v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning (RL) improves sample efficiency by leveraging pre-collected datasets prior to online interaction. A key chall

safetyarxiv-cs-ro
13 May 2026
Safety

Real-Scale Island Area and Coastline Estimation using Only its Place Name or Coordinates

DGX agent

arXiv:2605.11267v1 Announce Type: new Abstract: Accurate measurement of island area and coastline length is crucial for coastal zone monitoring and oceanographic analysis. However, traditional measure

safetyarxiv-cs-cv
13 May 2026
Safety

REFNet++: Multi-Task Efficient Fusion of Camera and Radar Sensor Data in Bird's-Eye Polar View

DGX agent

arXiv:2605.11824v1 Announce Type: new Abstract: A realistic view of the vehicle's surroundings is generally offered by camera sensors, which is crucial for environmental perception. Affordable radar s

safetyarxiv-cs-cv
13 May 2026
Safety

Rethink the Role of Neural Decoders in Quantum Error Correction

DGX agent

arXiv:2605.12046v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for enabling quantum advantages, with decoding as a central algorithmic primitive. Owing to its importance

safetyarxiv-cs-lg
13 May 2026
Safety

Rethinking external validation for the target population: Capturing patient-level similarity with a generative model

DGX agent

arXiv:2605.11284v1 Announce Type: cross Abstract: Background: External validation is essential for assessing the transportability of predictive models. However, its interpretation is often confounded

safetyarxiv-cs-lg
13 May 2026
Safety

RIO: Flexible Real-Time Robot I/O for Cross-Embodiment Robot Learning

DGX agent

arXiv:2605.11564v1 Announce Type: new Abstract: Despite recent efforts to collect multi-task, multi-embodiment datasets, to design recipes for training Vision-Language-Action models (VLAs), and to sho

safetyarxiv-cs-ro
13 May 2026
Safety

Robust Multi-Agent Path Finding under Observation Attacks: A Principled Adversarial-Plus-Smoothing Training Recipe

DGX agent

arXiv:2605.11469v1 Announce Type: new Abstract: Decentralized multi-agent path finding (MAPF) routes a team of agents on a shared grid, each acting from its own local view. The standard solution train

safetyarxiv-cs-lg
13 May 2026
Safety

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

DGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

safetyarxiv-cs-ro
13 May 2026
Safety

Sequential Off-Policy Learning with Logarithmic Smoothing

DGX agent

arXiv:2506.10664v2 Announce Type: replace-cross Abstract: Off-policy learning enables training policies from logged interaction data. Most prior work considers the batch setting, where a policy is lea

safetyarxiv-cs-lg
13 May 2026
← Previous
1…214215216217218…302
Next →