AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
13 May 2026

MajinBook: An open catalogue of digitally mediated world literature

SafetyDGX agent

arXiv:2511.11412v5 Announce Type: replace Abstract: This data paper introduces MajinBook, an open catalogue designed to facilitate the use of shadow libraries-such as Library Genesis and Z-Library-for

Maximin Robust Bayesian Experimental Design

SafetyDGX agent

arXiv:2603.14094v2 Announce Type: replace-cross Abstract: We address the brittleness of Bayesian experimental design under model misspecification by formulating the problem as a max--min game between

Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

SafetyDGX agent

arXiv:2605.12070v1 Announce Type: new Abstract: Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics

SafetyDGX agent

arXiv:2605.12119v1 Announce Type: new Abstract: Generative novel view synthesis faces a fundamental dilemma: geometric priors provide spatial alignment but become sparse and inaccurate under view chan

Model-based Bootstrap of Controlled Markov Chains

SafetyDGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible.…

SafetyDGX agent

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible. Even less plausible is the claim that there will be an AI j

More Than Meets the Eye: A Semantics-Aware Traffic Augmentation Framework for Generalizable Website Fingerprinting

SafetyDGX agent

arXiv:2605.11402v1 Announce Type: new Abstract: Deep learning-based website fingerprinting has emerged as an effective technique for inferring the websites users visit. Although existing methods achie

Morphologically Equivariant Flow Matching for Bimanual Mobile Manipulation

SafetyDGX agent

arXiv:2605.12228v1 Announce Type: new Abstract: Mobile manipulation requires coordinated control of high-dimensional, bimanual robots. Imitation learning methods have been broadly used to solve these

Multimodal Abstractive Summarization of Instructional Videos with Vision-Language Models

SafetyDGX agent

arXiv:2605.11959v1 Announce Type: cross Abstract: Multimodal video summarization requires visual features that align semantically with language generation. Traditional approaches rely on CNN features

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs …

SafetyDGX agent

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs we get from AI. By scraping the open web, LLMs are unwitting

Newton's Lantern: A Reinforcement Learning Framework for Finetuning AC Power Flow Warm Start Models

SafetyDGX agent

arXiv:2605.11102v1 Announce Type: new Abstract: Neural warm starts can sharply reduce the number of Newton-Raphson iterations required to solve the AC power flow problem, but existing supervised appro

@Nima292 LLMs are not AGI but will lead to some job losses; true AGI would likely lead to many more.

SafetyDGX agent

Gary Marcus argues that current large language models (LLMs) do not constitute artificial general intelligence (AGI), though they will cause some job displacement. He suggests that true AGI, if achiev

no remorse, just further evasion. so slick; so dangerous.

SafetyDGX agent

no remorse, just further evasion. so slick; so dangerous. 🚨 SEVEN OPENAI INSIDERS HAVE ACCUSED SAM ALTMAN OF LYING Today on cross, Musk's lawyer walked Altman through all of them: >Ilya Sutskever (co-

Off-Policy Learning with Limited Supply

SafetyDGX agent

arXiv:2603.18702v3 Announce Type: replace Abstract: We study off-policy learning (OPL) in contextual bandits, which plays a key role in a wide range of real-world applications such as recommendation s

Offline Constrained Reinforcement Learning under Partial Data Coverage

SafetyDGX agent

arXiv:2505.17506v2 Announce Type: replace-cross Abstract: We study offline constrained reinforcement learning with general function approximation in discounted constrained Markov decision processes. P

Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation

SafetyDGX agent

arXiv:2605.11479v1 Announce Type: new Abstract: Policy evaluation is a fundamental component of the development and deployment pipeline for robotic policies. In modern manipulation systems, this probl

OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning

SafetyDGX agent

arXiv:2605.12400v1 Announce Type: new Abstract: We study {on-policy self-distillation} (OPSD), where a language model improves its reasoning ability by distilling privileged teacher distributions alon

OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation

SafetyDGX agent

arXiv:2605.12480v1 Announce Type: new Abstract: Recent advances in joint audio-video generation have been remarkable, yet real-world applications demand strong per-modality fidelity, cross-modal align

On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

SafetyDGX agent

arXiv:2605.12206v1 Announce Type: new Abstract: In reinforcement learning (RL), agents acting in partially observable Markov decision processes (POMDPs) must rely on memory, typically encoded in a rec

Optimal Policy Learning under Budget and Coverage Constraints

SafetyDGX agent

arXiv:2605.12235v1 Announce Type: cross Abstract: We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and

Optimizing 4D Wires for Sparse 3D Abstraction

SafetyDGX agent

arXiv:2605.11977v1 Announce Type: new Abstract: We present a unified framework for 3D geometric abstraction using a single continuous 4D wire, parameterized as a B-spline with spatial coordinates and

ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models

SafetyDGX agent

arXiv:2605.12446v1 Announce Type: cross Abstract: Large language models (LLMs) often produce answers with high certainty even when they are incorrect, making reliable confidence estimation essential f

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has…

SafetyDGX agent

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has been doubling every few months, and this rate has become fa

OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness

SafetyDGX agent

arXiv:2605.11525v1 Announce Type: new Abstract: Missing values are routinely treated as defects to be eliminated through deletion or imputation prior to machine learning. In many applied domains, howe

Oversmoothing as Representation Degeneracy in Neural Sheaf Diffusion

SafetyDGX agent

arXiv:2605.11178v1 Announce Type: new Abstract: Neural Sheaf Diffusion (NSD) generalizes diffusion-based Graph Neural Networks by replacing scalar graph Laplacians with sheaf Laplacians whose learned

Physics-Informed Graph Neural Networks for Frequency-Aware Optical Aberration Correction

SafetyDGX agent

arXiv:2512.05683v2 Announce Type: replace Abstract: Optical aberrations significantly degrade image quality in microscopy, particularly when imaging deeper into samples. These aberrations arise from d

PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian Splatting

SafetyDGX agent

arXiv:2605.11520v1 Announce Type: new Abstract: Unsupervised point cloud segmentation is critical for embodied artificial intelligence and autonomous driving, as it mitigates the prohibitive cost of d

Position: Universal Aesthetic Alignment Narrows Artistic Expression

SafetyDGX agent

arXiv:2512.11883v3 Announce Type: replace-cross Abstract: Over-aligning image generation models to a generalized aesthetic preference conflicts with user intent, particularly when 'anti-aesthetic' out

Post-ADC Inference: Valid Inference After Active Data Collection

SafetyDGX agent

arXiv:2605.11511v1 Announce Type: cross Abstract: The validity of statistical inference depends critically on how data are collected. When data gathered through active data collection (ADC) are reused

Predictive Maps of Multi-Agent Reasoning: A Successor-Representation Spectrum for LLM Communication Topologies

SafetyDGX agent

arXiv:2605.11453v1 Announce Type: cross Abstract: Practitioners deploying multi-agent large language model (LLM) systems must currently choose between communication topologies such as chain, star, mes

Pretraining Exposure Explains Popularity Judgments in Large Language Models

SafetyDGX agent

arXiv:2605.12382v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic preferences for well-known entities, a phenomenon often attributed to popularity bias. However, the exte

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses

SafetyDGX agent

arXiv:2605.11535v1 Announce Type: new Abstract: Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are eithe

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling

SafetyDGX agent

arXiv:2605.11299v1 Announce Type: cross Abstract: Code generation is typically trained in the primal space of programs: a model produces a candidate solution and receives sparse execution feedback, of

PriorZero: Bridging Language Priors and World Models for Decision Making

SafetyDGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks

SafetyDGX agent

arXiv:2509.06701v2 Announce Type: replace Abstract: We develop a theory of intelligent agency grounded in probabilistic modeling for neural models. Agents are represented as outcome distributions with

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s mo…

SafetyDGX agent

probably correct, from @polynoamial: “with today’s AI models, intelligence is a function of inference compute.” but what about tomorrow’s models? never forget that humans are remarkably intelligent (t

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

SafetyDGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.11289v1 Announce Type: new Abstract: Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct dis

Rainbow Deep Q-Learning with Kinematics-Aware Design for Cooperative Delta and 3-RRS Parallel Robot Insertion

SafetyDGX agent

arXiv:2605.11697v1 Announce Type: new Abstract: This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipula

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking

SafetyDGX agent

arXiv:2605.11151v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning (RL) improves sample efficiency by leveraging pre-collected datasets prior to online interaction. A key chall

Real-Scale Island Area and Coastline Estimation using Only its Place Name or Coordinates

SafetyDGX agent

arXiv:2605.11267v1 Announce Type: new Abstract: Accurate measurement of island area and coastline length is crucial for coastal zone monitoring and oceanographic analysis. However, traditional measure

REFNet++: Multi-Task Efficient Fusion of Camera and Radar Sensor Data in Bird's-Eye Polar View

SafetyDGX agent

arXiv:2605.11824v1 Announce Type: new Abstract: A realistic view of the vehicle's surroundings is generally offered by camera sensors, which is crucial for environmental perception. Affordable radar s

Rethink the Role of Neural Decoders in Quantum Error Correction

SafetyDGX agent

arXiv:2605.12046v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for enabling quantum advantages, with decoding as a central algorithmic primitive. Owing to its importance

Rethinking external validation for the target population: Capturing patient-level similarity with a generative model

SafetyDGX agent

arXiv:2605.11284v1 Announce Type: cross Abstract: Background: External validation is essential for assessing the transportability of predictive models. However, its interpretation is often confounded

RIO: Flexible Real-Time Robot I/O for Cross-Embodiment Robot Learning

SafetyDGX agent

arXiv:2605.11564v1 Announce Type: new Abstract: Despite recent efforts to collect multi-task, multi-embodiment datasets, to design recipes for training Vision-Language-Action models (VLAs), and to sho

Robust Multi-Agent Path Finding under Observation Attacks: A Principled Adversarial-Plus-Smoothing Training Recipe

SafetyDGX agent

arXiv:2605.11469v1 Announce Type: new Abstract: Decentralized multi-agent path finding (MAPF) routes a team of agents on a shared grid, each acting from its own local view. The standard solution train

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

SafetyDGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

Sequential Off-Policy Learning with Logarithmic Smoothing

SafetyDGX agent

arXiv:2506.10664v2 Announce Type: replace-cross Abstract: Off-policy learning enables training policies from logged interaction data. Most prior work considers the batch setting, where a policy is lea

SEVO: Semantic-Enhanced Virtual Observation for Robust VLA Manipulation via Active Illumination and Data-Centric Collection

SafetyDGX agent

arXiv:2605.11114v1 Announce Type: new Abstract: Vision-Language-Action (VLA) and imitation-learning policies trained via community toolchains on low-cost hardware frequently fail when deployed outside

SI-Diff: A Framework for Learning Search and High-Precision Insertion with a Force-Domain Diffusion Policy

SafetyDGX agent

arXiv:2605.12247v1 Announce Type: new Abstract: Contact-rich assembly is fundamental in robotics but poses significant challenges due to uncertainties in relative poses, such as misalignments and smal

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

SafetyDGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation

SafetyDGX agent

arXiv:2603.15759v2 Announce Type: replace-cross Abstract: Robot learning requires adaptation methods that improve reliably from limited, mixed-quality interaction data. This is especially challenging

Simulation-Ready Cluttered Scene Estimation via Physics-aware Joint Shape and Pose Optimization

SafetyDGX agent

arXiv:2602.20150v2 Announce Type: replace-cross Abstract: Estimating simulation-ready scenes from real-world observations is crucial for downstream planning and policy learning tasks. Regretfully, exi

SkillGraph: Skill-Augmented Reinforcement Learning for Agents via Evolving Skill Graphs

SafetyDGX agent

arXiv:2605.12039v1 Announce Type: new Abstract: Skill libraries enable large language model agents to reuse experience from past interactions, but most existing libraries store skills as isolated entr

Space Syntax-guided Post-training for Residential Floor Plan Generation

SafetyDGX agent

arXiv:2602.22507v2 Announce Type: replace-cross Abstract: Residential floor plan generation requires not only geometric fidelity but also spatial configurational logic: shared living spaces should be

Sparse Offline Reinforcement Learning with Corruption Robustness

SafetyDGX agent

arXiv:2512.24768v3 Announce Type: replace-cross Abstract: We investigate robustness to strong data corruption in offline sparse reinforcement learning (RL). In our setting, an adversary may arbitraril

Sparsity and Out-of-Distribution Generalization

SafetyDGX agent

arXiv:2603.07388v2 Announce Type: replace Abstract: Explaining out-of-distribution generalization has been a central problem in epistemology since Goodman's 'grue' puzzle in 1946. Today it's a central

Spectral-Adaptive Modulation Networks for Visual Perception

SafetyDGX agent

arXiv:2503.23947v2 Announce Type: replace Abstract: Recent studies have shown that 2D convolution and self-attention exhibit distinct spectral behaviors, and optimizing their spectral properties can e

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training

SafetyDGX agent

arXiv:2605.11134v1 Announce Type: new Abstract: Preference learning methods such as Direct Preference Optimization (DPO) are known to induce reliance on spurious correlations, leading to sycophancy an

STRIDE: Training-Free Diversity Guidance via PCA-Directed Feature Perturbation in Single-Step Diffusion Models

SafetyDGX agent

arXiv:2605.11494v1 Announce Type: new Abstract: Distilled one-step (T=1) or few-step (Tleq4) diffusion models enable real-time image generation but often exhibit reduced sample diversity compared to t

← Previous
1…171172173174175…242
Next →