AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects

DGX agent

arXiv:2605.19587v1 Announce Type: new Abstract: Indoor scene synthesis underpins embodied AI, robotic manipulation, and simulation-based policy evaluation, where a useful scene must specify not only w

safetyarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting

DGX agent

arXiv:2605.19554v1 Announce Type: new Abstract: Instilling creativity in text-to-image (T2I) generation presents a significant challenge, as it requires synthesized images to exhibit not only visual n

safetyarxiv-cs-cv
20 May 2026
Safety

Set-Valued Policy Learning

DGX agent

arXiv:2605.19830v1 Announce Type: new Abstract: Conventional treatment policies map patient covariates to a single recommended intervention in order to maximize expected clinical outcomes. Although a

safetyarxiv-cs-lg
20 May 2026
Safety

SimGym: A Framework for A/B Test Simulation in E-Commerce with Traffic-Grounded VLM Agents

DGX agent

arXiv:2605.19219v1 Announce Type: new Abstract: A/B testing remains the gold standard for evaluating modifications to e-commerce storefronts, yet it diverts traffic, requires weeks to reach statistica

safetyarxiv-cs-ai
20 May 2026
Safety

Smooth Partial Lotteries for Stable Randomized Selection

DGX agent

arXiv:2605.20069v1 Announce Type: new Abstract: Competitive selection processes, from scientific funding to admissions and hiring, use evaluations to score candidates, and eventually choose a subset o

safetyarxiv-cs-lg
20 May 2026
Safety

Spatially Prompted Visual Trajectory Prediction for Egocentric Manipulation

DGX agent

arXiv:2605.20085v1 Announce Type: new Abstract: Robotic manipulation is often specified through language instructions or task identifiers, yet cluttered environments with similar objects are better ha

safetyarxiv-cs-cv
20 May 2026
Safety

Stitched Value Model for Diffusion Alignment

DGX agent

arXiv:2605.19804v1 Announce Type: cross Abstract: For practical use, diffusion- or flow-based generative models must be aligned with task-specific rewards, such as prompt fidelity or aesthetic prefere

safetyarxiv-cs-ai
20 May 2026
Safety

Structural Energy Guidance for View-Consistent Text-to-3D Generation

DGX agent

arXiv:2605.19876v1 Announce Type: new Abstract: Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work iden

safetyarxiv-cs-cv
20 May 2026
Safety

StruMPL: Multi-task Dense Regression under Disjoint Partial Supervision and MNAR Labels

DGX agent

arXiv:2605.19931v1 Announce Type: cross Abstract: Estimating forest aboveground biomass (AGB) from Earth observation combines two structurally incompatible label sources: spaceborne lidar provides can

safetyarxiv-cs-ai
20 May 2026
Safety

Surviving the Unseen: Predictive Defense for Novel Multi-Turn Multimodal Attacks

DGX agent

arXiv:2605.18988v1 Announce Type: cross Abstract: The expansion of Multimodal Large Language Models (MLLMs) and their integration into autonomous agentic workflows has introduced a non-stationary atta

safetyarxiv-cs-ai
20 May 2026
Safety

Swimming with Whales: Analysis of Power Imbalances in Stake-Weighted Governance

DGX agent

arXiv:2605.19264v1 Announce Type: new Abstract: Voting methods weighted by stakes are the fundamental governance paradigm in Proof-of-Stake (PoS) blockchains. Such a paradigm is known to be prone to p

safetyarxiv-cs-ai
20 May 2026
Safety

Symmetry in the Wild: The Role of Equivariance in Neural Fluid Surrogates

DGX agent

arXiv:2605.18816v1 Announce Type: cross Abstract: Neural surrogates enable orders-of-magnitude acceleration of computational fluid dynamics (CFD) simulations, with the potential to transform engineeri

safetyarxiv-cs-ai
20 May 2026
Safety

TEA-Time: Transporting Effects Across Time

DGX agent

arXiv:2603.07018v2 Announce Type: replace-cross Abstract: Treatment effects estimated from a randomized controlled trial are local not only to the study population but also to the time at which the tr

safetyarxiv-cs-lg
20 May 2026
Safety

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

DGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

safetyarxiv-cs-lg
20 May 2026
Safety

Text-to-SPARQL Generation with Reinforcement Learning: A GRPO-based Approach on DBLP

DGX agent

arXiv:2605.20066v1 Announce Type: new Abstract: Knowledge graph question answering seeks to translate natural language questions into executable queries over knowledge graphs, but existing approaches

safetyarxiv-cs-cl
20 May 2026
Safety

The Accessibility Capability Boundary: Operational Limits and Expansion Potential of AI-Generated Browser-Native Accessibility Systems

DGX agent

arXiv:2605.19638v1 Announce Type: cross Abstract: As large language models (LLMs) demonstrate increasing competence in synthesizing functional user interfaces, a fundamental question emerges in access

safetyarxiv-cs-ai
20 May 2026
Safety

The consensus of analysts is that AI capital investments will rise by 20% a year for five years while revenues are expected to grow 15% resu…

DGX agent

The consensus of analysts is that AI capital investments will rise by 20% a year for five years while revenues are expected to grow 15% resulting in negative returns. 'The AI boom will become a story

safetygary-marcus--x
20 May 2026
Safety

the crazy part is that people are “clowning” me without knowing anything about the training or whether anything else other than scaled chang…

DGX agent

the crazy part is that people are “clowning” me without knowing anything about the training or whether anything else other than scaled changed or how the model does on anything else. (or what it costs

safetygary-marcus--x
20 May 2026
Safety

The Uncomfortable Truth About AI “Reasoning” | World Science Festival https://youtu.be/iFYF_e1GSGI?si=pFNjvhp48TMKp-PE via @YouTube @GaryMar…

DGX agent

This video from the World Science Festival, shared by cognitive scientist Gary Marcus, examines the gap between AI systems' claimed reasoning capabilities and their actual mechanisms, likely arguing t

safetygary-marcus--x
20 May 2026
Safety

think you know @garymarcus because read a few of his tweets? try watching this, to see the real deal, and why, for example, the US Senate in…

DGX agent

think you know @garymarcus because read a few of his tweets? try watching this, to see the real deal, and why, for example, the US Senate invited him to testify. Superb conversation between @bgreene a

safetygary-marcus--x
20 May 2026
Safety

ThoughtTrace: Understanding User Thoughts in Real-World LLM Interactions

DGX agent

arXiv:2605.20087v1 Announce Type: cross Abstract: Conversational AI has now reached billions of users, yet existing datasets capture only what people say, not what they think. We introduce ThoughtTrac

safetyarxiv-cs-ai
20 May 2026
Safety

Toward an AI-Powered Computational Testbed for Workforce Policy

DGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

safetyarxiv-cs-ai
20 May 2026
Safety

Towards Distillation Guarantees under Algorithmic Alignment for Combinatorial Optimization

DGX agent

arXiv:2605.20074v1 Announce Type: new Abstract: Distillation transfers knowledge from a large model trained on broad data to a smaller, more efficient model suitable for deployment. In structured pred

safetyarxiv-cs-lg
20 May 2026
Safety

Trustworthy Agent Network: Trust in Agent Networks Must Be Baked In, Not Bolted On

DGX agent

arXiv:2605.19035v1 Announce Type: new Abstract: The rapid advancement of Large Language Models has given rise to autonomous LLM-based agents capable of complex reasoning and execution. As these agents

safetyarxiv-cs-ai
20 May 2026
Safety

TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents

DGX agent

arXiv:2602.11767v3 Announce Type: replace Abstract: Advances in large language models (LLMs) are driving a shift toward using reinforcement learning (RL) to train agents from iterative, multi-turn int

safetyarxiv-cs-ai
20 May 2026
Safety

two trillion dollars to build “pathologically dishonest” AI

DGX agent

two trillion dollars to build “pathologically dishonest” AI I think @RyanPGreenblatt's recent post summarizes the vibe of this behavior well. Not the sort of thing that would be acceptable for a human

safetygary-marcus--x
20 May 2026
Safety

Universal Skeleton Understanding via Differentiable Rendering and MLLMs

DGX agent

arXiv:2603.18003v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) exhibit strong visual-language reasoning, yet cannot process structured, non-visual data such as human skel

safetyarxiv-cs-cv
20 May 2026
Safety

When Critics Disagree: Adaptive Reward Poisoning Attacks in RIS-Aided Wireless Control System

DGX agent

arXiv:2605.20037v1 Announce Type: cross Abstract: Reward-poisoning attacks present a significant risk to learning-based wireless control systems. Given this, we propose a Disagreement-Guided Reward Po

safetyarxiv-cs-ai
20 May 2026
Safety

When Preference Labels Fall Short: Aligning Diffusion Models from Real Data

DGX agent

arXiv:2605.19839v1 Announce Type: new Abstract: Preference alignment aims to guide generative models by learning from comparisons between preferred and non-preferred samples. In practice, most existin

safetyarxiv-cs-cv
20 May 2026
Safety

When Tabular Foundation Models Meet Strategic Tabular Data: A Prior Alignment Approach

DGX agent

arXiv:2605.19662v1 Announce Type: new Abstract: Tabular foundation models based on pretrained prior-data fitted networks~(PFNs) have shown strong generalization on diverse tabular tasks, but they are

safetyarxiv-cs-ai
20 May 2026
Safety

When to Stop Reusing: Dynamic Gradient Gating for Sample-Efficient RLVR

DGX agent

arXiv:2605.19425v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for advanced reasoning in Large Language Models (LLMs), but rol

safetyarxiv-cs-ai
20 May 2026
Safety

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making

DGX agent

arXiv:2602.07008v2 Announce Type: replace Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typical

safetyarxiv-cs-cv
20 May 2026
Safety

Worst-Group Equalized Odds Regularization for Multi-Attribute Fair Medical Image Classification

DGX agent

arXiv:2605.19214v1 Announce Type: cross Abstract: Diagnostic performance in medical AI varies systematically across demographic groups, yet subgroup AUC can mask clinically important disparities. At a

safetyarxiv-cs-cv
20 May 2026
Safety

A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insights

DGX agent

arXiv:2605.16913v1 Announce Type: cross Abstract: Neural networks trained with gradient-based methods exhibit a strong simplicity bias: they learn simpler statistical features of their data before mov

safetyarxiv-cs-lg
19 May 2026
Safety

A Simplex Witness Certificate for Constant Collapse in Variational Autoencoders

DGX agent

arXiv:2605.18224v1 Announce Type: cross Abstract: This note studies exact constant collapse in variational autoencoders, where the encoder mean becomes independent of the input. The goal is to make th

safetyarxiv-cs-ai
19 May 2026
Safety

A study done on 361,645 job applications in almost 30 countries over the last 40 years discovered the hiring bias in society is actually aga…

DGX agent

A large-scale meta-analysis examining over 361,000 job applications across nearly 30 countries spanning 40 years found that hiring discrimination based on protected characteristics (such as race, gend

safetyelon-musk--x
19 May 2026
Safety

A Visual Reinforcement Learning-Based Separate Primitive Policy for Peg-in-Hole Tasks

DGX agent

arXiv:2504.14820v2 Announce Type: replace Abstract: For peg-in-hole tasks, humans rely on binocular visual perception to locate the peg above the hole surface and then proceed with insertion. This pap

safetyarxiv-cs-ro
19 May 2026
Safety

Ablating Safety: Mechanisms for Removing Alignment in Language Models for Security Applications

DGX agent

arXiv:2605.17413v1 Announce Type: cross Abstract: Safety-aligned language models often refuse cybersecurity requests whose wording resembles misuse, even when the task is authorized and defensive. Thi

safetyarxiv-cs-ai
19 May 2026
Safety

Actionable World Representation

DGX agent

arXiv:2605.18743v1 Announce Type: new Abstract: Inspired by the emergent behaviors in large language models that generalized human intelligence, the research community is pursuing similar emergent cap

safetyarxiv-cs-ai
19 May 2026
Safety

Activation Steering with a Feedback Controller

DGX agent

arXiv:2510.04309v3 Announce Type: replace Abstract: Controlling the behaviors of large language models (LLM) is fundamental to their safety alignment and reliable deployment. However, existing steerin

safetyarxiv-cs-lg
19 May 2026
Safety

Adaptive Control in Autonomous Driving via Real-Time Recurrent RL

DGX agent

arXiv:2602.02236v4 Announce Type: replace-cross Abstract: We study online fine-tuning of pretrained control policies for autonomous driving using Real-Time Recurrent Reinforcement Learning (RTRRL), a

safetyarxiv-cs-lg
19 May 2026
Safety

Adaptive Experimentation for Censored Survival Outcomes

DGX agent

arXiv:2605.18459v1 Announce Type: new Abstract: Adaptive experimentation enables efficient estimation of causal effects, but existing methods are not designed for survival data with censoring, where e

safetyarxiv-cs-lg
19 May 2026
Safety

Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification

DGX agent

arXiv:2605.17609v1 Announce Type: new Abstract: Many inference-time language-model pipelines combine a cheap reward signal with an expensive verifier, such as exact answer checking in mathematical rea

safetyarxiv-cs-lg
19 May 2026
Safety

Adversarial Fragility and Language Vulnerability in Clinical AI: A Systematic Audit of Diagnostic Collapse Under Imperceptible Perturbations and Cross-Lingual Drift in Low-Resource Healthcare Settings

DGX agent

arXiv:2605.16993v1 Announce Type: cross Abstract: Current clinical artificial intelligence (AI) systems are evaluated almost exclusively on clean, standardised, English-language inputs, conditions tha

safetyarxiv-cs-ai
19 May 2026
Safety

AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment

DGX agent

arXiv:2605.17517v1 Announce Type: new Abstract: Recent advances in Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation. However, the visual represe

safetyarxiv-cs-ro
19 May 2026
Safety

Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces

DGX agent

arXiv:2605.17698v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) as autonomous economic agents introduces systemic risks that extend beyond individual capability failures

safetyarxiv-cs-lg
19 May 2026
Safety

AI Agents May Always Fall for Prompt Injections

DGX agent

arXiv:2605.17634v1 Announce Type: cross Abstract: Prompt injection is the most critical vulnerability in deployed AI agents. Despite recent progress, we show that the prevailing defense paradigm (data

safetyarxiv-cs-cl
19 May 2026
Safety

AI Alignment Breaks at the Edge

DGX agent

arXiv:2602.20042v2 Announce Type: replace Abstract: General Alignment has improved average-case helpfulness and safety, but current alignment practice still rewards confident, single-turn responses. T

safetyarxiv-cs-cl
19 May 2026
← Previous
1…163164165166167…267
Next →