AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth takin…

DGX agent

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth taking the time to understand the issues, if you want to understa

safetygary-marcus--x
2 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the…

DGX agent

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the issues, if you want to understand a lot of what is driving

safetygary-marcus--x
2 Jun 2026
Safety

“The public has swung 49 points against data centers in just nine months, underscoring the heightened political salience of the facilities a…

DGX agent

Public opinion has shifted dramatically against data centers over a nine-month period, with a 49-point swing in unfavorable sentiment, reflecting growing political concern about these facilities. This

safetygary-marcus--x
2 Jun 2026
Safety

The role of class encoding in neural collapse

DGX agent

arXiv:2606.00344v1 Announce Type: new Abstract: Neural collapse is a structural property of the last-hidden-layer activations in neural network classification models, when trained beyond a zero classi

safetyarxiv-cs-lg
2 Jun 2026
Safety

The Social Cost of Intelligence: Emergence, Propagation, and Amplification of Stereotypical Bias in Multi-Agent Systems

DGX agent

arXiv:2510.10943v2 Announce Type: replace-cross Abstract: Bias in large language models (LLMs) remains a persistent challenge, often leading to stereotyping and unfair treatment across social groups.

safetyarxiv-cs-cl
2 Jun 2026
Safety

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to!

DGX agent

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to! Anthropic faces AI spending backlash before IPO https://www.axios.com/2026/06/02/anthropic-

safetygary-marcus--x
2 Jun 2026
Safety

Threading Optimization for Vision-Language-Action Model Inference in Low-Cost Smart Agricultural Manipulation

DGX agent

arXiv:2606.00966v1 Announce Type: new Abstract: Vision-Language Action (VLA) models continue to face challenges such as slow inference speed and difficulty performing fine-grained motion adjustments,

safetyarxiv-cs-ro
2 Jun 2026
Safety

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most i…

DGX agent

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most important policy to implement, at the US Senate in May 2023, an

safetygary-marcus--x
2 Jun 2026
Safety

ToolFG: Towards Well-Grounded Fine-Grained Image Classification

DGX agent

arXiv:2606.02518v1 Announce Type: new Abstract: Fine-grained image classification (FGIC) has broad applications and has attracted significant research attention. In this paper, we explore a novel para

safetyarxiv-cs-cv
2 Jun 2026
Safety

ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation

DGX agent

arXiv:2602.07883v3 Announce Type: replace Abstract: LLM-powered agentic systems excel at complex long-horizon tasks, but remain constrained by static configurations fixed before execution. Such rigidi

safetyarxiv-cs-ai
2 Jun 2026
Safety

Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities

DGX agent

arXiv:2606.00596v1 Announce Type: new Abstract: Large language models have rapidly evolved in multilingual competence and reasoning capacity, enabling their integration into Social Sciences and Humani

safetyarxiv-cs-cl
2 Jun 2026
Safety

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO

DGX agent

arXiv:2606.02313v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising end-to-end paradigm for unmanned aerial vehicles (UAVs) to accomplish complex tasks specified by f

safetyarxiv-cs-ro
2 Jun 2026
Safety

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

DGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

safetyarxiv-cs-ai
2 Jun 2026
Safety

Training-free image inversion for one-step diffusion models

DGX agent

arXiv:2606.01380v1 Announce Type: new Abstract: In this work, we introduce a novel training-free inversion (TFinv) framework for one-step diffusion models,addressing key challenges in real image inver

safetyarxiv-cs-cv
2 Jun 2026
Safety

Training-Free Imitation Learning with Closed-Form Diffusion Policies

DGX agent

arXiv:2606.01238v1 Announce Type: cross Abstract: While diffusion-based policies have impressive performance and expressivity, their long offline training slows down the data collection and policy dep

safetyarxiv-cs-lg
2 Jun 2026
Safety

Trajectory Data Suffices for Statistically Efficient Policy Evaluation in Fixed-Horizon Offline RL with Linear q^pi-Realizability and Concentrability

DGX agent

arXiv:2510.03494v2 Announce Type: replace Abstract: We study finite-horizon offline reinforcement learning (RL) with function approximation for both policy evaluation and policy optimization. Prior wo

safetyarxiv-cs-lg
2 Jun 2026
Safety

TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment

DGX agent

arXiv:2606.01755v1 Announce Type: new Abstract: Personalized large language models adapt responses to users' preferences and social attributes, but can introduce substantial universal truth inconsiste

safetyarxiv-cs-ai
2 Jun 2026
Safety

TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos

DGX agent

arXiv:2606.02350v1 Announce Type: new Abstract: Reconstructing humans and their surrounding environments in a globally consistent 4D space is essential for comprehensive perception. However, prior wor

safetyarxiv-cs-cv
2 Jun 2026
Safety

Trust Region On-Policy Distillation

DGX agent

arXiv:2606.01249v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) is a fundamental technique for efficient post-training of large language models (LLMs), with broad applications in agent

safetyarxiv-cs-cl
2 Jun 2026
Model Releases

TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages

DGX agent

arXiv:2606.01322v1 Announce Type: cross Abstract: Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, c

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval

DGX agent

arXiv:2606.01615v1 Announce Type: new Abstract: Video-language models are pivotal for tasks such as moment retrieval and highlight detection, yet they often struggle to capture the dynamic, non-linear

safetyarxiv-cs-cv
2 Jun 2026
Safety

UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment

DGX agent

arXiv:2606.00170v1 Announce Type: cross Abstract: In recent years, emotion recognition based on physiological signals such as electroencephalogram (EEG) has gained considerable attention, as internal

safetyarxiv-cs-ai
2 Jun 2026
Safety

Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances

DGX agent

arXiv:2606.02153v1 Announce Type: new Abstract: Methods using inertial measurement units (IMUs) provide a wearable alternative to camera-based motion capture. To mitigate drift from inertial signals,

safetyarxiv-cs-cv
2 Jun 2026
Safety

Until today I never once in my life thought about what the plural of backlash was.

DGX agent

Gary Marcus observes the unusual linguistic phenomenon of rarely contemplating the plural form of the word 'backlash,' highlighting how certain words are infrequently used in plural despite being gram

safetygary-marcus--x
2 Jun 2026
Safety

Update-Free On-Policy Steering via Verifiers

DGX agent

arXiv:2603.10282v2 Announce Type: replace Abstract: In recent years, Behavior Cloning (BC) has become one of the most prevalent methods for learning manipulation from human demonstrations. Despite the

safetyarxiv-cs-ro
2 Jun 2026
Safety

Update Opacity: Epistemic Accessibility and Governance Under AI System Change

DGX agent

arXiv:2606.00037v1 Announce Type: cross Abstract: Machine learning models embedded in deployed AI systems are routinely updated to maintain correct functioning over time. Yet such updates can generate

safetyarxiv-cs-ai
2 Jun 2026
Safety

V-LynX: Token Interface Alignment for Video+X LLMs

DGX agent

arXiv:2606.00508v1 Announce Type: cross Abstract: This study introduces an intriguing phenomenon in Video LLMs: rather than merely translating frames into textual embeddings, Video LLMs establish a co

safetyarxiv-cs-ai
2 Jun 2026
Safety

Value-Free Policy Optimization via Reward Partitioning

DGX agent

arXiv:2506.13702v4 Announce Type: replace-cross Abstract: Single-trajectory preference optimization methods learn from datasets of ((prompt, response, reward)) tuples, offering a practical alternative

safetyarxiv-cs-ai
2 Jun 2026
Safety

VERA: Variational Inference Framework for Jailbreaking Large Language Models

DGX agent

arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabi

safetyarxiv-cs-cl
2 Jun 2026
Safety

Visualizing definitional divergence in high-dimensional data by manifold alignment: Application to 3D right ventricular strain computations

DGX agent

arXiv:2501.12178v2 Announce Type: replace Abstract: Medical imaging studies often rely on a single sample per subject, assuming it is representative of their physiological traits. However, variations

safetyarxiv-cs-cv
2 Jun 2026
Safety

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models

DGX agent

arXiv:2601.03309v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models, which integrate pretrained large Vision-Language Models (VLM) into their policy backbone, are gaining sig

safetyarxiv-cs-ai
2 Jun 2026
Safety

Wavelet-Fusion Diffusion Model for Multimodal Brain MRI Synthesis with Modality and Metadata Conditioning

DGX agent

arXiv:2606.00689v1 Announce Type: new Abstract: Multimodal MRI provides complementary information for neuroimaging analysis, where different imaging modalities capture distinct anatomical, tissue, and

safetyarxiv-cs-cv
2 Jun 2026
Safety

Weak Critics Make Strong Learners: On-Policy Critique Distillation for Scalable Oversight

DGX agent

arXiv:2606.00424v1 Announce Type: new Abstract: As large language models become stronger, weak supervisors may fail to provide reliable labels, preferences, or final judgments for complex outputs, lim

safetyarxiv-cs-ai
2 Jun 2026
Safety

When Does Predictive Inverse Dynamics Outperform Behavior Cloning?

DGX agent

arXiv:2601.21718v2 Announce Type: replace-cross Abstract: Behavior cloning (BC) is a practical offline imitation learning method, but it often fails when expert demonstrations are limited. Recent work

safetyarxiv-cs-ai
2 Jun 2026
Safety

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models

DGX agent

arXiv:2606.01671v1 Announce Type: new Abstract: In the contemporary epoch of multilingual education, learning idioms provides a fascinating gateway towards creativity, cultural values, historical cont

safetyarxiv-cs-cl
2 Jun 2026
Safety

Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

DGX agent

arXiv:2511.05613v2 Announce Type: replace-cross Abstract: Foundation models are increasingly central to high-stakes AI systems, and governance frameworks now depend on evaluations to assess their risk

safetyarxiv-cs-ai
2 Jun 2026
Safety

Why things will eventually fall apart: 1. Everybody, even Google, seems to be treating AI as if it were some kind of winner take all competi…

DGX agent

Why things will eventually fall apart: 1. Everybody, even Google, seems to be treating AI as if it were some kind of winner take all competition like web search was, in which Google taking over 95% 2.

safetygary-marcus--x
2 Jun 2026
Safety

World Models for Robotic Manipulation: A Survey

DGX agent

arXiv:2606.00113v1 Announce Type: new Abstract: Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world mode

safetyarxiv-cs-ro
2 Jun 2026
Safety

World-Task Factorization for Robot Learning

DGX agent

arXiv:2606.02027v1 Announce Type: cross Abstract: Robot learning must produce policies that generalize to new combinations of constraints, teammates, and environments. To achieve this, we must structu

safetyarxiv-cs-lg
2 Jun 2026
Safety

A hitchhiker's guide to Poisson gradient estimation

DGX agent

arXiv:2602.03896v2 Announce Type: replace-cross Abstract: Poisson-distributed latent variable models are widely used in computational neuroscience, but differentiating through discrete stochastic samp

safetyarxiv-cs-lg
1 Jun 2026
Safety

A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models

DGX agent

arXiv:2605.30843v1 Announce Type: new Abstract: In the forward reinforcement-learning problem, the reward is fixed and known; the learner is asked to find a good policy or value function. Here we turn

safetyarxiv-cs-lg
1 Jun 2026
Safety

A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI

DGX agent

arXiv:2605.31021v1 Announce Type: new Abstract: Current alignment paradigms for generative artificial intelligence rely predominantly on monolithic benchmarking frameworks that reduce the plurality of

safetyarxiv-cs-ai
1 Jun 2026
Safety

A Unified Framework for Gradient Aggregation in Multi-Objective Optimization

DGX agent

arXiv:2605.30452v1 Announce Type: cross Abstract: Many machine learning problems involve multiple inherent trade-offs that are best addressed by gradient-based multi-objective optimization (MOO) algor

safetyarxiv-cs-ai
1 Jun 2026
Safety

Active Timepoint Selection for Learning Measure-Valued Trajectories

DGX agent

arXiv:2605.30625v1 Announce Type: cross Abstract: Inferring continuous probability paths from sparse snapshots is a fundamental challenge in domains like single-cell biology, where high-fidelity data

safetyarxiv-cs-ai
1 Jun 2026
Safety

AI Loss of Control Incident Management: Response & Resilience

DGX agent

arXiv:2605.30406v1 Announce Type: cross Abstract: Recent research demonstrating AI systems exhibiting deception and shutdown resistance suggests that AI loss of control (LOC) is an urgent policy conce

safetyarxiv-cs-ai
1 Jun 2026
Safety

Annealed Softmax Greedy in Many-Armed Bayesian Bandits

DGX agent

arXiv:2605.31034v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and group-based policy optimization methods such as GRPO update a stochastic policy by sampling

safetyarxiv-cs-ai
1 Jun 2026
Safety

Annotations Are Not All You Need: A Cross-modal Knowledge Transfer Network for Unsupervised Temporal Sentence Grounding

DGX agent

arXiv:2605.30742v1 Announce Type: new Abstract: This paper addresses the task of temporal sentence grounding (TSG). Although many respectable works have made decent achievements in this important topi

safetyarxiv-cs-cv
1 Jun 2026
Safety

Are Full Rollouts Necessary for On-Policy Distillation?

DGX agent

arXiv:2605.31490v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense teacher feedback along rollouts generated by the student and has emerged as a promising post-training paradi

safetyarxiv-cs-cl
1 Jun 2026
← Previous
1…166167168169170…302
Next →