AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
Safety

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotia…

DGX agent

Pope Francis: Died within hours after meeting JD Vance Viktor Orban: Lost within a few days after meeting JD Vance Middle east peace negotiations: fell apart within 21 hours of Vance’s arrival Welcome

safetygary-marcus--x
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

DGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

safetyarxiv-cs-cv
13 Apr 2026
Safety

Post-Selection Distributional Model Evaluation

DGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

safetyarxiv-cs-lg
13 Apr 2026
Safety

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

DGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

safetyarxiv-cs-lg
13 Apr 2026
Safety

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

DGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

safetyarxiv-cs-ai
13 Apr 2026
Safety

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

DGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

safetyarxiv-cs-ai
13 Apr 2026
Safety

Reducing Class Bias In Data-Balanced Datasets Through Hardness-Based Resampling

DGX agent

arXiv:2504.07031v2 Announce Type: replace Abstract: Class-bias, that is class-wise performance disparities, is typically attributed to data imbalance and addressed through frequency-based resampling.

safetyarxiv-cs-lg
13 Apr 2026
Safety

Region-Constrained Group Relative Policy Optimization for Flow-Based Image Editing

DGX agent

arXiv:2604.09386v1 Announce Type: new Abstract: Instruction-guided image editing requires balancing target modification with non-target preservation. Recently, flow-based models have emerged as a stro

safetyarxiv-cs-cv
13 Apr 2026
Safety

Reinforcement-aware Knowledge Distillation for LLM Reasoning

DGX agent

arXiv:2602.22495v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) post-training has recently driven major gains in long chain-of-thought reasoning large language models (LLMs), but

safetyarxiv-cs-ai
13 Apr 2026
Safety

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

DGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

safetyarxiv-cs-ai
13 Apr 2026
Safety

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

DGX agent

arXiv:2604.09474v1 Announce Type: cross Abstract: Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise,

safetyarxiv-cs-ai
13 Apr 2026
Safety

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

DGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

safetyarxiv-cs-cv
13 Apr 2026
Safety

Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence

DGX agent

arXiv:2604.09104v1 Announce Type: cross Abstract: Scheming, the covert pursuit of misaligned goals by AI systems, represents a potentially catastrophic risk, yet scheming research suffers from signifi

safetyarxiv-cs-ai
13 Apr 2026
Safety

SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images

DGX agent

arXiv:2604.09436v1 Announce Type: new Abstract: Diffusion models trained on noisy datasets often reproduce high-frequency training artifacts, significantly degrading generation quality. To address thi

safetyarxiv-cs-cv
13 Apr 2026
Safety

Score-Driven Rating System for Sports

DGX agent

arXiv:2604.09143v1 Announce Type: new Abstract: This paper introduces a score-driven rating system, a generalization of the classical Elo rating system that employs the score, i.e. the gradient of the

safetyarxiv-cs-lg
13 Apr 2026
Safety

Semantic Intent Fragmentation: A Single-Shot Compositional Attack on Multi-Agent AI Pipelines

DGX agent

arXiv:2604.08608v1 Announce Type: cross Abstract: We introduce Semantic Intent Fragmentation (SIF), an attack class against LLM orchestration systems where a single, legitimately phrased request cause

safetyarxiv-cs-ai
13 Apr 2026
Safety

SHIFT: Steering Hidden Intermediates in Flow Transformers

DGX agent

arXiv:2604.09213v1 Announce Type: new Abstract: Diffusion models have become leading approaches for high-fidelity image generation. Recent DiT-based diffusion models, in particular, achieve strong pro

safetyarxiv-cs-cv
13 Apr 2026
Safety

Sim-to-Real Transfer for Muscle-Actuated Robots via Generalized Actuator Networks

DGX agent

arXiv:2604.09487v1 Announce Type: cross Abstract: Tendon drives paired with soft muscle actuation enable faster and safer robots while potentially accelerating skill acquisition. Still, these systems

safetyarxiv-cs-lg
13 Apr 2026
Safety

SPEAR: An Engineering Case Study of Multi-Agent Coordination for Smart Contract Auditing

DGX agent

arXiv:2602.04418v3 Announce Type: replace-cross Abstract: We present SPEAR, a multi-agent coordination framework for smart contract auditing that applies established MAS patterns in a realistic securi

safetyarxiv-cs-ai
13 Apr 2026
Safety

SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasks

DGX agent

arXiv:2604.08865v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is central to aligning Large Language Models (LLMs) in reasoning tasks with verifiable rewards. However, standard tok

safetyarxiv-cs-ai
13 Apr 2026
Safety

SSPO: Subsentence-level Policy Optimization

DGX agent

arXiv:2511.04256v2 Announce Type: replace Abstract: As a key component of large language model (LLM) post-training, Reinforcement Learning from Verifiable Rewards (RLVR) has substantially improved rea

safetyarxiv-cs-cl
13 Apr 2026
Safety

StaRPO: Stability-Augmented Reinforcement Policy Optimization

DGX agent

arXiv:2604.08905v1 Announce Type: new Abstract: Reinforcement learning (RL) is effective in enhancing the accuracy of large language models in complex reasoning tasks. Existing RL policy optimization

safetyarxiv-cs-ai
13 Apr 2026
Safety

STCast: Adaptive Boundary Alignment for Global and Regional Weather Forecasting

DGX agent

arXiv:2509.25210v3 Announce Type: replace-cross Abstract: To gain finer regional forecasts, many works have explored the regional integration from the global atmosphere, e.g., by solving boundary equa

safetyarxiv-cs-ai
13 Apr 2026
Safety

StructRL: Recovering Dynamic Programming Structure from Learning Dynamics in Distributional Reinforcement Learning

DGX agent

arXiv:2604.08620v1 Announce Type: cross Abstract: Reinforcement learning is typically treated as a uniform, data-driven optimization process, where updates are guided by rewards and temporal-differenc

safetyarxiv-cs-ai
13 Apr 2026
Safety

SubQuad: Near-Quadratic-Free Structure Inference with Distribution-Balanced Objectives in Adaptive Receptor framework

DGX agent

arXiv:2602.17330v3 Announce Type: replace-cross Abstract: Comparative analysis of adaptive immune repertoires at population scale is hampered by two practical bottlenecks: the near-quadratic cost of p

safetyarxiv-cs-ai
13 Apr 2026
Safety

Summary: AI Governance to Avoid Extinction

DGX agent

With AI capabilities rapidly increasing, humans appear close to developing AI systems that are better than human experts across all domains. This raises a series of questions about how the world will—

safetymiri
13 Apr 2026
Safety

The causal relation between off-street parking and electric vehicle adoption in Scotland

DGX agent

arXiv:2604.09271v1 Announce Type: new Abstract: The transition to electric mobility hinges on maximising aggregate adoption while also facilitating equitable access. This study examines whether the 'c

safetyarxiv-cs-lg
13 Apr 2026
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Safety

The trend of treating all of AI as One Big Thing that always includes data centers & job changes & education changes & power & accelerating …

DGX agent

The trend of treating all of AI as One Big Thing that always includes data centers & job changes & education changes & power & accelerating science & misinformation & national security & corporate con

safetyethan-mollick--x
13 Apr 2026
Safety

The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs

DGX agent

arXiv:2601.01580v2 Announce Type: replace-cross Abstract: Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT c

safetyarxiv-cs-ai
13 Apr 2026
Safety

'There is no advantage from being first. If you get first to a superintelligence you don't control, the superintelligence wins. Not the USA.…

DGX agent

'There is no advantage from being first. If you get first to a superintelligence you don't control, the superintelligence wins. Not the USA. Not China. Not the UK.' @andreamiotti of @ControlAI & TBC o

safetyconnor-leahy--x
13 Apr 2026
Safety

Think Less, Know More: State-Aware Reasoning Compression with Knowledge Guidance for Efficient Reasoning

DGX agent

arXiv:2604.09150v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on complex tasks by leveraging long Chain-of-Thought (CoT), but often suffer from overthinking,

safetyarxiv-cs-cl
13 Apr 2026
Safety

Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation

DGX agent

arXiv:2604.09368v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as scalable user simulators for recommender system evaluation. Yet existing simulators per

safetyarxiv-cs-cv
13 Apr 2026
Safety

TME-PSR: Time-aware, Multi-interest, and Explanation Personalization for Sequential Recommendation

DGX agent

arXiv:2604.09439v1 Announce Type: cross Abstract: In this paper, we propose a sequential recommendation model that integrates Time-aware personalization, Multi-interest personalization, and Explanatio

safetyarxiv-cs-ai
13 Apr 2026
Safety

Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence

DGX agent

arXiv:2604.09057v1 Announce Type: new Abstract: Audio-video (AV) generation has recently made strong progress in perceptual quality and multimodal coherence, yet generating content with plausible moti

safetyarxiv-cs-cv
13 Apr 2026
Safety

Toward World Models for Epidemiology

DGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

safetyarxiv-cs-lg
13 Apr 2026
Safety

Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models

DGX agent

arXiv:2604.08815v1 Announce Type: new Abstract: Medical vision-language models (VLMs) show strong performance on radiology tasks but often produce fluent yet weakly grounded conclusions due to over-re

safetyarxiv-cs-cv
13 Apr 2026
Safety

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

DGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

safetyarxiv-cs-lg
13 Apr 2026
Safety

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

DGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

safetyarxiv-cs-ai
13 Apr 2026
Safety

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

DGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

safetyarxiv-cs-lg
13 Apr 2026
Safety

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

DGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

safetyarxiv-cs-ai
13 Apr 2026
Safety

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

DGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

safetyarxiv-cs-cv
13 Apr 2026
Safety

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

DGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

safetyarxiv-cs-cv
13 Apr 2026
Safety

Verbalizing LLMs' assumptions to explain and control sycophancy

DGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

safetyarxiv-cs-ai
13 Apr 2026
Safety

Violence is not the answer. But maybe boycotts are?

DGX agent

Violence is not the answer. But maybe boycotts are? 🚨 NOW: The FBI is RAIDING the home of a 20-year-old man who threw a molotov cocktail at the home of OpenAI CEO Sam Altman Over a DOZEN federal agent

safetygary-marcus--x
13 Apr 2026
Safety

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

DGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

safetyarxiv-cs-ai
13 Apr 2026
Safety

Visually-Guided Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

safetyarxiv-cs-ai
13 Apr 2026
Safety

When & How to Write for Personalized Demand-aware Query Rewriting in Video Search

DGX agent

arXiv:2602.17667v2 Announce Type: replace-cross Abstract: In video search systems, user historical behaviors provide rich context for identifying search intent and resolving ambiguity. However, tradit

safetyarxiv-cs-cv
13 Apr 2026
← Previous
1…255256257258259…263
Next →