AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
13 Apr 2026

MARBLE: Multi-Armed Restless Bandits in Latent Markovian Environment

SafetyDGX agent

arXiv:2511.09324v2 Announce Type: replace Abstract: Restless Multi-Armed Bandits (RMABs) are powerful models for decision-making under uncertainty, yet classical formulations typically assume fixed dy

Mechanisms of Introspective Awareness

SafetyDGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

Memo: OpenAI Chief Revenue Officer Denise Dresser says Anthropic is 'grossing up rev share with Amazon and Google' and overstating its 'run …

SafetyDGX agent

Memo: OpenAI Chief Revenue Officer Denise Dresser says Anthropic is 'grossing up rev share with Amazon and Google' and overstating its 'run rate by roughly $8B' (@haydenfield / The Verge) https://www.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MeshOn: Intersection-Free Mesh-to-Mesh Composition

SafetyDGX agent

arXiv:2604.08799v1 Announce Type: cross Abstract: We propose MeshOn, a method that finds physically and semantically realistic compositions of two input meshes. Given an accessory, a base mesh with a

MixFlow: Mixed Source Distributions Improve Rectified Flows

SafetyDGX agent

arXiv:2604.09181v1 Announce Type: new Abstract: Diffusion models and their variations, such as rectified flows, generate diverse and high-quality images, but they are still hindered by slow iterative

MSMO-ABSA: Multi-Scale and Multi-Objective Optimization for Cross-Lingual Aspect-Based Sentiment Analysis

SafetyDGX agent

arXiv:2502.13718v2 Announce Type: replace Abstract: Aspect-based sentiment analysis (ABSA) garnered growing research interest in multilingual contexts in the past. However, the majority of the studies

Musculoskeletal Motion Imitation for Learning Personalized Exoskeleton Control Policy in Impaired Gait

SafetyDGX agent

arXiv:2604.09431v1 Announce Type: new Abstract: Designing generalizable control policies for lower-limb exoskeletons remains fundamentally constrained by exhaustive data collection or iterative optimi

MuTSE: A Human-in-the-Loop Multi-use Text Simplification Evaluator

SafetyDGX agent

arXiv:2604.08947v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly prevalent in text simplification, systematically evaluating their outputs across diverse prompting

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

SafetyDGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System

SafetyDGX agent

arXiv:2604.09069v1 Announce Type: cross Abstract: Court Judgment Prediction and Explanation (CJPE) aims to predict a judicial decision and provide a legally grounded explanation for a given case based

On Divergence Measures for Training GFlowNets

SafetyDGX agent

arXiv:2410.09355v2 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are amortized inference models designed to sample from unnormalized distributions over composable objects, with a

On the Representational Limits of Quantum-Inspired 1024-D Document Embeddings: An Experimental Evaluation Framework

SafetyDGX agent

arXiv:2604.09430v1 Announce Type: cross Abstract: Text embeddings are central to modern information retrieval and Retrieval-Augmented Generation (RAG). While dense models derived from Large Language M

On the Role of DAG topology in Energy-Aware Cloud Scheduling : A GNN-Based Deep Reinforcement Learning Approach

SafetyDGX agent

arXiv:2604.09202v1 Announce Type: cross Abstract: Cloud providers must assign heterogeneous compute resources to workflow DAGs while balancing competing objectives such as completion time, cost, and e

Open secret that the big AI labs see themselves as emergent state-like entities in the vein of the distributed polities in “Diamond Age” or …

SafetyDGX agent

Open secret that the big AI labs see themselves as emergent state-like entities in the vein of the distributed polities in “Diamond Age” or “Terra Ignota”, which is pretty funny bc staff at said labs

PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment

SafetyDGX agent

arXiv:2604.08986v1 Announce Type: cross Abstract: Persona prompting has been widely adopted to steer large language models (LLMs) behavior and improve their instruction performance by assigning specif

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks

SafetyDGX agent

arXiv:2604.09028v1 Announce Type: cross Abstract: Unmanned aerial vehicles serving as aerial base stations can rapidly restore connectivity after disasters, yet abrupt changes in user mobility and tra

Policy-Aware Design of Large-Scale Factorial Experiments

SafetyDGX agent

arXiv:2604.08804v1 Announce Type: cross Abstract: Digital firms routinely run many online experiments on shared user populations. When product decisions are compositional, such as combinations of inte

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotia…

SafetyDGX agent

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotiations: fell apart within 21 hours of Vance’s arrival Welcome

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

SafetyDGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

Post-Selection Distributional Model Evaluation

SafetyDGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

SafetyDGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

SafetyDGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

Reducing Class Bias In Data-Balanced Datasets Through Hardness-Based Resampling

SafetyDGX agent

arXiv:2504.07031v2 Announce Type: replace Abstract: Class-bias, that is class-wise performance disparities, is typically attributed to data imbalance and addressed through frequency-based resampling.

Region-Constrained Group Relative Policy Optimization for Flow-Based Image Editing

SafetyDGX agent

arXiv:2604.09386v1 Announce Type: new Abstract: Instruction-guided image editing requires balancing target modification with non-target preservation. Recently, flow-based models have emerged as a stro

Reinforcement-aware Knowledge Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2602.22495v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training has recently driven major gains in long chain-of-thought reasoning large language models (LLMs), but

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

SafetyDGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

SafetyDGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images

SafetyDGX agent

arXiv:2604.09436v1 Announce Type: new Abstract: Diffusion models trained on noisy datasets often reproduce high-frequency training artifacts, significantly degrading generation quality. To address thi

Score-Driven Rating System for Sports

SafetyDGX agent

arXiv:2604.09143v1 Announce Type: new Abstract: This paper introduces a score-driven rating system, a generalization of the classical Elo rating system that employs the score, i.e. the gradient of the

SHIFT: Steering Hidden Intermediates in Flow Transformers

SafetyDGX agent

arXiv:2604.09213v1 Announce Type: new Abstract: Diffusion models have become leading approaches for high-fidelity image generation. Recent DiT-based diffusion models, in particular, achieve strong pro

Sim-to-Real Transfer for Muscle-Actuated Robots via Generalized Actuator Networks

SafetyDGX agent

arXiv:2604.09487v1 Announce Type: cross Abstract: Tendon drives paired with soft muscle actuation enable faster and safer robots while potentially accelerating skill acquisition. Still, these systems

SPEAR: An Engineering Case Study of Multi-Agent Coordination for Smart Contract Auditing

SafetyDGX agent

arXiv:2602.04418v3 Announce Type: replace-cross Abstract: We present SPEAR, a multi-agent coordination framework for smart contract auditing that applies established MAS patterns in a realistic securi

SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasks

SafetyDGX agent

arXiv:2604.08865v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is central to aligning Large Language Models (LLMs) in reasoning tasks with verifiable rewards. However, standard tok

SSPO: Subsentence-level Policy Optimization

SafetyDGX agent

arXiv:2511.04256v2 Announce Type: replace Abstract: As a key component of large language model (LLM) post-training, Reinforcement Learning from Verifiable Rewards (RLVR) has substantially improved rea

StaRPO: Stability-Augmented Reinforcement Policy Optimization

SafetyDGX agent

arXiv:2604.08905v1 Announce Type: new Abstract: Reinforcement learning (RL) is effective in enhancing the accuracy of large language models in complex reasoning tasks. Existing RL policy optimization

STCast: Adaptive Boundary Alignment for Global and Regional Weather Forecasting

SafetyDGX agent

arXiv:2509.25210v3 Announce Type: replace-cross Abstract: To gain finer regional forecasts, many works have explored the regional integration from the global atmosphere, e.g., by solving boundary equa

StructRL: Recovering Dynamic Programming Structure from Learning Dynamics in Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2604.08620v1 Announce Type: cross Abstract: Reinforcement learning is typically treated as a uniform, data-driven optimization process, where updates are guided by rewards and temporal-differenc

SubQuad: Near-Quadratic-Free Structure Inference with Distribution-Balanced Objectives in Adaptive Receptor framework

SafetyDGX agent

arXiv:2602.17330v3 Announce Type: replace-cross Abstract: Comparative analysis of adaptive immune repertoires at population scale is hampered by two practical bottlenecks: the near-quadratic cost of p

Summary: AI Governance to Avoid Extinction

SafetyDGX agent

With AI capabilities rapidly increasing, humans appear close to developing AI systems that are better than human experts across all domains. This raises a series of questions about how the world will—

The causal relation between off-street parking and electric vehicle adoption in Scotland

SafetyDGX agent

arXiv:2604.09271v1 Announce Type: new Abstract: The transition to electric mobility hinges on maximising aggregate adoption while also facilitating equitable access. This study examines whether the 'c

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

SafetyDGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

The trend of treating all of AI as One Big Thing that always includes data centers & job changes & education changes & power & accelerating …

SafetyDGX agent

The trend of treating all of AI as One Big Thing that always includes data centers & job changes & education changes & power & accelerating science & misinformation & national security & corporate con

The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs

SafetyDGX agent

arXiv:2601.01580v2 Announce Type: replace-cross Abstract: Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT c

'There is no advantage from being first. If you get first to a superintelligence you don't control, the superintelligence wins. Not the USA.…

SafetyDGX agent

'There is no advantage from being first. If you get first to a superintelligence you don't control, the superintelligence wins. Not the USA. Not China. Not the UK.' @andreamiotti of @ControlAI & TBC o

Think Less, Know More: State-Aware Reasoning Compression with Knowledge Guidance for Efficient Reasoning

SafetyDGX agent

arXiv:2604.09150v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on complex tasks by leveraging long Chain-of-Thought (CoT), but often suffer from overthinking,

Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation

SafetyDGX agent

arXiv:2604.09368v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as scalable user simulators for recommender system evaluation. Yet existing simulators per

TME-PSR: Time-aware, Multi-interest, and Explanation Personalization for Sequential Recommendation

SafetyDGX agent

arXiv:2604.09439v1 Announce Type: cross Abstract: In this paper, we propose a sequential recommendation model that integrates Time-aware personalization, Multi-interest personalization, and Explanatio

Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence

SafetyDGX agent

arXiv:2604.09057v1 Announce Type: new Abstract: Audio-video (AV) generation has recently made strong progress in perceptual quality and multimodal coherence, yet generating content with plausible moti

Toward World Models for Epidemiology

SafetyDGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

SafetyDGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

SafetyDGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

SafetyDGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

SafetyDGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

SafetyDGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

SafetyDGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

Violence is not the answer. But maybe boycotts are?

SafetyDGX agent

Violence is not the answer. But maybe boycotts are? 🚨 NOW: The FBI is RAIDING the home of a 20-year-old man who threw a molotov cocktail at the home of OpenAI CEO Sam Altman Over a DOZEN federal agent

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

SafetyDGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

Visually-Guided Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

Model ReleasesDGX agent

Boris Cherny and colleagues conducted cybersecurity evaluations of Claude Mythos Preview, finding it to be the first AI model to successfully complete an AISI (AI Safety Institute) cyber range end-to-

When & How to Write for Personalized Demand-aware Query Rewriting in Video Search

SafetyDGX agent

arXiv:2602.17667v2 Announce Type: replace-cross Abstract: In video search systems, user historical behaviors provide rich context for identifying search intent and resolving ambiguity. However, tradit

← Previous
1…215216217218219…240
Next →