AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Intrinsic Vicarious Conditioning for Deep Reinforcement Learning

DGX agent

arXiv:2605.12224v1 Announce Type: new Abstract: Advancements in reinforcement learning have produced a variety of complex and useful intrinsic driving forces; crucially, these drivers operate under a

safetyarxiv-cs-lg
13 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation

DGX agent

arXiv:2510.17062v2 Announce Type: replace Abstract: While reasoning-based large language models excel at complex tasks through an internal, structured thinking process, a concerning phenomenon has eme

safetyarxiv-cs-cl
13 May 2026
Safety

Invisible failures in human-AI interactions

DGX agent

arXiv:2603.15423v2 Announce Type: replace Abstract: AI systems fail silently far more often than they fail visibly. In an analysis of 100K human-AI interactions from the WildChat dataset, we find that

safetyarxiv-cs-cl
13 May 2026
Safety

JACoP: Joint Alignment for Compliant Multi-Agent Prediction

DGX agent

arXiv:2605.11385v1 Announce Type: new Abstract: Stochastic Human Trajectory Prediction (HTP) using generative modeling has emerged as a significant area of research. Although state-of-the-art models e

safetyarxiv-cs-cv
13 May 2026
Safety

LA-Sign: Looped Transformers with Geometry-aware Alignment for Skeleton-based Sign Language Recognition

DGX agent

arXiv:2603.29057v2 Announce Type: replace Abstract: Skeleton-based isolated sign language recognition (ISLR) demands fine-grained understanding of articulated motion across multiple spatial scales, fr

safetyarxiv-cs-cv
13 May 2026
Safety

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer?

DGX agent

arXiv:2605.11301v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have heterogeneous strengths across OCR, chart understanding, spatial reasoning, visual question answering, c

safetyarxiv-cs-cl
13 May 2026
Safety

LEAP: Unlocking dLLM Parallelism via Lookahead Early-Convergence Token Detection

DGX agent

arXiv:2605.10980v1 Announce Type: new Abstract: Diffusion Language Models (dLLMs) have garnered significant attention for their potential in highly parallel processing. The parallel capabilities of ex

safetyarxiv-cs-lg
13 May 2026
Safety

Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training

DGX agent

arXiv:2605.11931v1 Announce Type: new Abstract: Post-training with explicit reasoning traces is common to improve the reasoning capabilities of Multimodal Large Language Models (MLLMs). However, acqui

safetyarxiv-cs-cv
13 May 2026
Safety

Learning Adapter Rank via Symmetry Breaking

DGX agent

arXiv:2506.22809v4 Announce Type: replace-cross Abstract: Low-rank adaptation is effective partly because downstream updates lie in a low-dimensional subspace, but the latent rank coordinates of LoRA

safetyarxiv-cs-cl
13 May 2026
Safety

Learning Agentic Policy from Action Guidance

DGX agent

arXiv:2605.12004v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) for Large Language Models (LLMs) critically depends on the exploration capability of the base policy, as training si

safetyarxiv-cs-cl
13 May 2026
Safety

Learning Minimally Rigid Graphs with High Realization Counts

DGX agent

arXiv:2605.12427v1 Announce Type: new Abstract: For minimally rigid graphs, the same edge-length data can admit multiple realizations (up to translations and rotations). Finding graphs with exceptiona

safetyarxiv-cs-lg
13 May 2026
Safety

Leveraging Multimodal Large Language Models for All-in-One Image Restoration via a Mixture of Frequency Experts

DGX agent

arXiv:2605.11444v1 Announce Type: new Abstract: All-in-one image restoration seeks to recover clean images from inputs affected by diverse and unknown degradations using a unified framework. Recent me

safetyarxiv-cs-cv
13 May 2026
Safety

Leveraging RAG for Training-Free Alignment of LLMs

DGX agent

arXiv:2605.11217v1 Announce Type: new Abstract: Large language model (LLM) alignment algorithms typically consist of post-training over preference pairs. While such algorithms are widely used to enabl

safetyarxiv-cs-lg
13 May 2026
Safety

Logit-Attention Divergence: Mitigating Position Bias in Multi-Image Retrieval via Attention-Guided Calibration

DGX agent

arXiv:2605.11591v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance in multi-image cross-modal retrieval, yet suffer from severe position bias, where

safetyarxiv-cs-cv
13 May 2026
Safety

Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making

DGX agent

arXiv:2602.07668v2 Announce Type: replace Abstract: The looking-in-looking-out (LILO) framework has enabled intelligent vehicle applications that understand both the outside scene and the driver state

safetyarxiv-cs-cv
13 May 2026
Safety

LucidFlux: Caption-Free Photo-Realistic Image Restoration via a Large-Scale Diffusion Transformer

DGX agent

arXiv:2509.22414v4 Announce Type: replace Abstract: Image restoration (IR) aims to recover images degraded by unknown mixtures while preserving semanticsconditions under which discriminative restorers

safetyarxiv-cs-cv
13 May 2026
Safety

LucidNFT: LR-Anchored Multi-Reward Preference Optimization for Flow-Based Real-World Super-Resolution

DGX agent

arXiv:2603.05947v3 Announce Type: replace Abstract: Generative real-world image super-resolution (Real-ISR) can synthesize visually convincing details from severely degraded low-resolution (LR) inputs

safetyarxiv-cs-cv
13 May 2026
Safety

MajinBook: An open catalogue of digitally mediated world literature

DGX agent

arXiv:2511.11412v5 Announce Type: replace Abstract: This data paper introduces MajinBook, an open catalogue designed to facilitate the use of shadow libraries-such as Library Genesis and Z-Library-for

safetyarxiv-cs-cl
13 May 2026
Safety

Maximin Robust Bayesian Experimental Design

DGX agent

arXiv:2603.14094v2 Announce Type: replace-cross Abstract: We address the brittleness of Bayesian experimental design under model misspecification by formulating the problem as a max--min game between

safetyarxiv-cs-lg
13 May 2026
Safety

Metaphor Is Not All Attention Needs

DGX agent

arXiv:2605.12128v1 Announce Type: new Abstract: Large language models are increasingly deployed in safety-critical applications, where their ability to resist harmful instructions is essential. Althou

safetyarxiv-cs-cl
13 May 2026
Safety

Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

DGX agent

arXiv:2605.12070v1 Announce Type: new Abstract: Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization

safetyarxiv-cs-lg
13 May 2026
Safety

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics

DGX agent

arXiv:2605.12119v1 Announce Type: new Abstract: Generative novel view synthesis faces a fundamental dilemma: geometric priors provide spatial alignment but become sparse and inaccurate under view chan

safetyarxiv-cs-cv
13 May 2026
Safety

Model-based Bootstrap of Controlled Markov Chains

DGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

safetyarxiv-cs-lg
13 May 2026
Safety

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible.…

DGX agent

More accurate statement IMHO would be: there won’t immediately be an AI jobpocalyspe. Saying there never will be one hardly seems plausible. Even less plausible is the claim that there will be an AI j

safetygary-marcus--x
13 May 2026
Safety

More Than Meets the Eye: A Semantics-Aware Traffic Augmentation Framework for Generalizable Website Fingerprinting

DGX agent

arXiv:2605.11402v1 Announce Type: new Abstract: Deep learning-based website fingerprinting has emerged as an effective technique for inferring the websites users visit. Although existing methods achie

safetyarxiv-cs-lg
13 May 2026
Safety

Morphologically Equivariant Flow Matching for Bimanual Mobile Manipulation

DGX agent

arXiv:2605.12228v1 Announce Type: new Abstract: Mobile manipulation requires coordinated control of high-dimensional, bimanual robots. Imitation learning methods have been broadly used to solve these

safetyarxiv-cs-ro
13 May 2026
Safety

Multimodal Abstractive Summarization of Instructional Videos with Vision-Language Models

DGX agent

arXiv:2605.11959v1 Announce Type: cross Abstract: Multimodal video summarization requires visual features that align semantically with language generation. Traditional approaches rely on CNN features

safetyarxiv-cs-cl
13 May 2026
Safety

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs …

DGX agent

New paper in Nature. The more a government controls its domestic media, the more it dominates AI training data, the more pro-regime outputs we get from AI. By scraping the open web, LLMs are unwitting

safetygary-marcus--x
13 May 2026
Safety

Newton's Lantern: A Reinforcement Learning Framework for Finetuning AC Power Flow Warm Start Models

DGX agent

arXiv:2605.11102v1 Announce Type: new Abstract: Neural warm starts can sharply reduce the number of Newton-Raphson iterations required to solve the AC power flow problem, but existing supervised appro

safetyarxiv-cs-lg
13 May 2026
Safety

@Nima292 LLMs are not AGI but will lead to some job losses; true AGI would likely lead to many more.

DGX agent

Gary Marcus argues that current large language models (LLMs) do not constitute artificial general intelligence (AGI), though they will cause some job displacement. He suggests that true AGI, if achiev

safetygary-marcus--x
13 May 2026
Safety

no remorse, just further evasion. so slick; so dangerous.

DGX agent

no remorse, just further evasion. so slick; so dangerous. 🚨 SEVEN OPENAI INSIDERS HAVE ACCUSED SAM ALTMAN OF LYING Today on cross, Musk's lawyer walked Altman through all of them: >Ilya Sutskever (co-

safetygary-marcus--x
13 May 2026
Safety

Off-Policy Learning with Limited Supply

DGX agent

arXiv:2603.18702v3 Announce Type: replace Abstract: We study off-policy learning (OPL) in contextual bandits, which plays a key role in a wide range of real-world applications such as recommendation s

safetyarxiv-cs-lg
13 May 2026
Safety

Offline Constrained Reinforcement Learning under Partial Data Coverage

DGX agent

arXiv:2505.17506v2 Announce Type: replace-cross Abstract: We study offline constrained reinforcement learning with general function approximation in discounted constrained Markov decision processes. P

safetyarxiv-cs-lg
13 May 2026
Safety

Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation

DGX agent

arXiv:2605.11479v1 Announce Type: new Abstract: Policy evaluation is a fundamental component of the development and deployment pipeline for robotic policies. In modern manipulation systems, this probl

safetyarxiv-cs-ro
13 May 2026
Safety

OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning

DGX agent

arXiv:2605.12400v1 Announce Type: new Abstract: We study {on-policy self-distillation} (OPSD), where a language model improves its reasoning ability by distilling privileged teacher distributions alon

safetyarxiv-cs-lg
13 May 2026
Safety

OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation

DGX agent

arXiv:2605.12480v1 Announce Type: new Abstract: Recent advances in joint audio-video generation have been remarkable, yet real-world applications demand strong per-modality fidelity, cross-modal align

safetyarxiv-cs-cv
13 May 2026
Safety

On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

DGX agent

arXiv:2605.12206v1 Announce Type: new Abstract: In reinforcement learning (RL), agents acting in partially observable Markov decision processes (POMDPs) must rely on memory, typically encoded in a rec

safetyarxiv-cs-lg
13 May 2026
Safety

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue

DGX agent

arXiv:2605.05630v2 Announce Type: replace Abstract: Hidden malicious intent in multi-turn dialogue poses a growing threat to deployed large language models (LLMs). Rather than exposing a harmful objec

safetyarxiv-cs-cl
13 May 2026
Safety

OpenAI endorses the Kids Online Safety Act and Illinois SB 315, an AI safety bill to create requirements around transparency, incident reporting, and more (OpenAI Global Affairs)

DGX agent

OpenAI Global Affairs: OpenAI endorses the Kids Online Safety Act and Illinois SB 315, an AI safety bill to create requirements around transparency, incident reporting, and more — Welcome (back) to Th

safetytechmeme
13 May 2026
Safety

Optimal Policy Learning under Budget and Coverage Constraints

DGX agent

arXiv:2605.12235v1 Announce Type: cross Abstract: We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and

safetyarxiv-cs-lg
13 May 2026
Safety

Optimizing 4D Wires for Sparse 3D Abstraction

DGX agent

arXiv:2605.11977v1 Announce Type: new Abstract: We present a unified framework for 3D geometric abstraction using a single continuous 4D wire, parameterized as a B-spline with spatial coordinates and

safetyarxiv-cs-cv
13 May 2026
Safety

ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models

DGX agent

arXiv:2605.12446v1 Announce Type: cross Abstract: Large language models (LLMs) often produce answers with high certainty even when they are incorrect, making reliable confidence estimation essential f

safetyarxiv-cs-cl
13 May 2026
Safety

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has…

DGX agent

Our evaluations show that frontier AI's cyber capabilities are advancing quickly. The length of cyber tasks frontier models can complete has been doubling every few months, and this rate has become fa

safetygary-marcus--x
13 May 2026
Safety

OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness

DGX agent

arXiv:2605.11525v1 Announce Type: new Abstract: Missing values are routinely treated as defects to be eliminated through deletion or imputation prior to machine learning. In many applied domains, howe

safetyarxiv-cs-lg
13 May 2026
Safety

Oversmoothing as Representation Degeneracy in Neural Sheaf Diffusion

DGX agent

arXiv:2605.11178v1 Announce Type: new Abstract: Neural Sheaf Diffusion (NSD) generalizes diffusion-based Graph Neural Networks by replacing scalar graph Laplacians with sheaf Laplacians whose learned

safetyarxiv-cs-lg
13 May 2026
Safety

Persona-Conditioned Adversarial Prompting: Multi-Identity Red-Teaming for Adversarial Discovery and Mitigation

DGX agent

arXiv:2605.11730v1 Announce Type: new Abstract: Automated red-teaming for LLMs often discovers narrow attack slices, missing diverse real-world threats, and yielding insufficient data for safety fine-

safetyarxiv-cs-lg
13 May 2026
Safety

Physics-Informed Graph Neural Networks for Frequency-Aware Optical Aberration Correction

DGX agent

arXiv:2512.05683v2 Announce Type: replace Abstract: Optical aberrations significantly degrade image quality in microscopy, particularly when imaging deeper into samples. These aberrations arise from d

safetyarxiv-cs-cv
13 May 2026
Safety

PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian Splatting

DGX agent

arXiv:2605.11520v1 Announce Type: new Abstract: Unsupervised point cloud segmentation is critical for embodied artificial intelligence and autonomous driving, as it mitigates the prohibitive cost of d

safetyarxiv-cs-cv
13 May 2026
← Previous
1…183184185186187…267
Next →