AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
2 Jun 2026

Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation

SafetyDGX agent

arXiv:2606.02441v1 Announce Type: new Abstract: Identity-preserving video generation (IPVG) aims to synthesize high-fidelity videos that follow text prompts while faithfully preserving a reference ide

Spatiotemporal Multi-Task Graph Transformer for Trip-Level Transit Prediction

SafetyDGX agent

arXiv:2606.00572v1 Announce Type: new Abstract: Passenger count data from public transit systems reveals urban mobility patterns and is essential for planning, operation, and optimisation. However, no

SpeedAug: Policy Acceleration via Tempo-Enriched Policy and RL Fine-Tuning

SafetyDGX agent

arXiv:2512.00062v2 Announce Type: replace-cross Abstract: Robotic policy learning for complex real-world manipulation tasks has seen rapid recent progress, enabled in large part by the ability to coll


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SS-ZKR: Spatial-Semantic Zero-Knowledge Routing for Privacy-Preserving Multi-Agent Collaboration

SafetyDGX agent

arXiv:2606.00962v1 Announce Type: cross Abstract: Foundational agent interoperability standards, notably the Agent-to-Agent (A2A) protocol and the Model Context Protocol (MCP), have advanced multi-age

Stabilizing Policy Optimization via Logits Convexity

SafetyDGX agent

arXiv:2603.00963v2 Announce Type: replace-cross Abstract: While reinforcement learning (RL) has been central to the recent success of large language models (LLMs), RL optimization is notoriously unsta

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that th…

SafetyDGX agent

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that the environment can maintain reliably, it probably doesn't bel

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement

SafetyDGX agent

arXiv:2606.00267v1 Announce Type: cross Abstract: Video world models (WMs) have shown promise for policy evaluation and improvement by imagining realistic future observations conditioned on ego-robot

SurrogateSHAP: Training-Free Contributor Attribution for Text-to-Image (T2I) Models

SafetyDGX agent

arXiv:2601.22276v2 Announce Type: replace-cross Abstract: As Text-to-Image (T2I) diffusion models are increasingly used in real-world creative workflows, a principled framework for valuing contributor

SWARD: Stochastic Window-Attention-Based Relational Distillation for Cross-Architectural Semantic Segmentation

SafetyDGX agent

arXiv:2606.00999v1 Announce Type: new Abstract: Large-scale vision foundation models have driven substantial gains on dense prediction tasks such as semantic segmentation, but their size makes deploym

Sympatheia: Emotionally Adaptive Voice Assistant with Continuous Affect Conditioning

SafetyDGX agent

arXiv:2606.00851v1 Announce Type: cross Abstract: Empathetic spoken dialogue systems must infer a user's emotional state to respond appropriately, yet everyday speech often carries weak, neutral, or a

T-POP: Test-Time Personalization with Online Preference Feedback

SafetyDGX agent

arXiv:2509.24696v2 Announce Type: replace-cross Abstract: Personalizing large language models (LLMs) to individual user preferences is a critical step beyond generating generically helpful responses.

Task-Induced Representational Invariances Depend on Learning Objective in Deep RL

SafetyDGX agent

arXiv:2606.01868v1 Announce Type: new Abstract: Reinforcement Learning (RL) has long served as a model for goal-directed animal behavior in neuroscience. Modern deep RL has shown remarkable success ac

Tether-Aware Dynamic Collision Avoidance for USV-HROV Systems

SafetyDGX agent

arXiv:2606.01112v1 Announce Type: new Abstract: Heterogeneous marine robotic systems composed of an unmanned surface vehicle (USV) and a hybrid remotely operated vehicle (HROV) have shown great potent

Thanks everyone for reading! Longer version some additional points on crowd psychology and links to a pair of great new podcasts with the le…

SafetyDGX agent

Thanks everyone for reading! Longer version some additional points on crowd psychology and links to a pair of great new podcasts with the legendary investors @realsteveeisman and @gnoble79 and converg

The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer

SafetyDGX agent

arXiv:2602.02557v2 Announce Type: replace-cross Abstract: Recent advances in end-to-end trained omni-models have substantially improved audio capabilities by strengthening text-audio modality alignmen

The Harsh Truth: Segment-Level Analysis of Harsh Driving Events in Milan Using Large-Scale Telematics, Street Networks, and Google Street View

SafetyDGX agent

arXiv:2606.00261v1 Announce Type: new Abstract: Police-reported crash statistics remain the standard input for urban road-safety assessment, but their incompleteness and reporting lag limit their usef

The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs

SafetyDGX agent

arXiv:2606.00674v1 Announce Type: cross Abstract: Large Language Models (LLMs) aligned via outcome-based Reinforcement Learning (RL) frequently exhibit a critical failure mode: they achieve high perfo

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth takin…

SafetyDGX agent

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth taking the time to understand the issues, if you want to understa

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the…

SafetyDGX agent

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the issues, if you want to understand a lot of what is driving

“The public has swung 49 points against data centers in just nine months, underscoring the heightened political salience of the facilities a…

SafetyDGX agent

Public opinion has shifted dramatically against data centers over a nine-month period, with a 49-point swing in unfavorable sentiment, reflecting growing political concern about these facilities. This

The role of class encoding in neural collapse

SafetyDGX agent

arXiv:2606.00344v1 Announce Type: new Abstract: Neural collapse is a structural property of the last-hidden-layer activations in neural network classification models, when trained beyond a zero classi

The Social Cost of Intelligence: Emergence, Propagation, and Amplification of Stereotypical Bias in Multi-Agent Systems

SafetyDGX agent

arXiv:2510.10943v2 Announce Type: replace-cross Abstract: Bias in large language models (LLMs) remains a persistent challenge, often leading to stereotyping and unfair treatment across social groups.

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to!

SafetyDGX agent

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to! Anthropic faces AI spending backlash before IPO https://www.axios.com/2026/06/02/anthropic-

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

SafetyDGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

Threading Optimization for Vision-Language-Action Model Inference in Low-Cost Smart Agricultural Manipulation

SafetyDGX agent

arXiv:2606.00966v1 Announce Type: new Abstract: Vision-Language Action (VLA) models continue to face challenges such as slow inference speed and difficulty performing fine-grained motion adjustments,

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most i…

SafetyDGX agent

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most important policy to implement, at the US Senate in May 2023, an

ToolFG: Towards Well-Grounded Fine-Grained Image Classification

SafetyDGX agent

arXiv:2606.02518v1 Announce Type: new Abstract: Fine-grained image classification (FGIC) has broad applications and has attracted significant research attention. In this paper, we explore a novel para

ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation

SafetyDGX agent

arXiv:2602.07883v3 Announce Type: replace Abstract: LLM-powered agentic systems excel at complex long-horizon tasks, but remain constrained by static configurations fixed before execution. Such rigidi

Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities

SafetyDGX agent

arXiv:2606.00596v1 Announce Type: new Abstract: Large language models have rapidly evolved in multilingual competence and reasoning capacity, enabling their integration into Social Sciences and Humani

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO

SafetyDGX agent

arXiv:2606.02313v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising end-to-end paradigm for unmanned aerial vehicles (UAVs) to accomplish complex tasks specified by f

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

SafetyDGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

TRACE: Trajectory Risk-Aware Compression for Long-Horizon Agent Safety

SafetyDGX agent

arXiv:2606.00611v1 Announce Type: new Abstract: Long-horizon LLM agents produce safety evidence across long trajectories, where sparse, delayed, and compositional risk signals often escape local moder

Train, Test, Re-evaluate: Schedule-Sensitive Evaluation of Generative Data for Hand Detection

SafetyDGX agent

arXiv:2606.01896v1 Announce Type: cross Abstract: Generated (or synthetic) image data is increasingly used to augment or replace real training datasets when target imagery is scarce, expensive, or bia

Training-free image inversion for one-step diffusion models

SafetyDGX agent

arXiv:2606.01380v1 Announce Type: new Abstract: In this work, we introduce a novel training-free inversion (TFinv) framework for one-step diffusion models,addressing key challenges in real image inver

Training-Free Imitation Learning with Closed-Form Diffusion Policies

SafetyDGX agent

arXiv:2606.01238v1 Announce Type: cross Abstract: While diffusion-based policies have impressive performance and expressivity, their long offline training slows down the data collection and policy dep

Trajectory Data Suffices for Statistically Efficient Policy Evaluation in Fixed-Horizon Offline RL with Linear q^pi-Realizability and Concentrability

SafetyDGX agent

arXiv:2510.03494v2 Announce Type: replace Abstract: We study finite-horizon offline reinforcement learning (RL) with function approximation for both policy evaluation and policy optimization. Prior wo

TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment

SafetyDGX agent

arXiv:2606.01755v1 Announce Type: new Abstract: Personalized large language models adapt responses to users' preferences and social attributes, but can introduce substantial universal truth inconsiste

TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos

SafetyDGX agent

arXiv:2606.02350v1 Announce Type: new Abstract: Reconstructing humans and their surrounding environments in a globally consistent 4D space is essential for comprehensive perception. However, prior wor

Trust Region On-Policy Distillation

SafetyDGX agent

arXiv:2606.01249v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) is a fundamental technique for efficient post-training of large language models (LLMs), with broad applications in agent

Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval

SafetyDGX agent

arXiv:2606.01615v1 Announce Type: new Abstract: Video-language models are pivotal for tasks such as moment retrieval and highlight detection, yet they often struggle to capture the dynamic, non-linear

UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment

SafetyDGX agent

arXiv:2606.00170v1 Announce Type: cross Abstract: In recent years, emotion recognition based on physiological signals such as electroencephalogram (EEG) has gained considerable attention, as internal

Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances

SafetyDGX agent

arXiv:2606.02153v1 Announce Type: new Abstract: Methods using inertial measurement units (IMUs) provide a wearable alternative to camera-based motion capture. To mitigate drift from inertial signals,

Until today I never once in my life thought about what the plural of backlash was.

SafetyDGX agent

Gary Marcus observes the unusual linguistic phenomenon of rarely contemplating the plural form of the word 'backlash,' highlighting how certain words are infrequently used in plural despite being gram

Update-Free On-Policy Steering via Verifiers

SafetyDGX agent

arXiv:2603.10282v2 Announce Type: replace Abstract: In recent years, Behavior Cloning (BC) has become one of the most prevalent methods for learning manipulation from human demonstrations. Despite the

Update Opacity: Epistemic Accessibility and Governance Under AI System Change

SafetyDGX agent

arXiv:2606.00037v1 Announce Type: cross Abstract: Machine learning models embedded in deployed AI systems are routinely updated to maintain correct functioning over time. Yet such updates can generate

V-LynX: Token Interface Alignment for Video+X LLMs

SafetyDGX agent

arXiv:2606.00508v1 Announce Type: cross Abstract: This study introduces an intriguing phenomenon in Video LLMs: rather than merely translating frames into textual embeddings, Video LLMs establish a co

Value-Free Policy Optimization via Reward Partitioning

SafetyDGX agent

arXiv:2506.13702v4 Announce Type: replace-cross Abstract: Single-trajectory preference optimization methods learn from datasets of ((prompt, response, reward)) tuples, offering a practical alternative

VERA: Variational Inference Framework for Jailbreaking Large Language Models

SafetyDGX agent

arXiv:2506.22666v3 Announce Type: replace-cross Abstract: The rise of API-only access to state-of-the-art LLMs highlights the need for effective black-box jailbreak methods to identify model vulnerabi

Visual Persuasion: What Influences Decisions of Vision-Language Models?

SafetyDGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

Visualizing definitional divergence in high-dimensional data by manifold alignment: Application to 3D right ventricular strain computations

SafetyDGX agent

arXiv:2501.12178v2 Announce Type: replace Abstract: Medical imaging studies often rely on a single sample per subject, assuming it is representative of their physiological traits. However, variations

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models

SafetyDGX agent

arXiv:2601.03309v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models, which integrate pretrained large Vision-Language Models (VLM) into their policy backbone, are gaining sig

Wavelet-Fusion Diffusion Model for Multimodal Brain MRI Synthesis with Modality and Metadata Conditioning

SafetyDGX agent

arXiv:2606.00689v1 Announce Type: new Abstract: Multimodal MRI provides complementary information for neuroimaging analysis, where different imaging modalities capture distinct anatomical, tissue, and

Weak Critics Make Strong Learners: On-Policy Critique Distillation for Scalable Oversight

SafetyDGX agent

arXiv:2606.00424v1 Announce Type: new Abstract: As large language models become stronger, weak supervisors may fail to provide reliable labels, preferences, or final judgments for complex outputs, lim

What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs

SafetyDGX agent

arXiv:2606.01624v1 Announce Type: new Abstract: Driving vision-language models (VLMs) must accurately understand scenes across diverse conditions defined by Operational Design Domains (ODDs), yet veri

When Does Predictive Inverse Dynamics Outperform Behavior Cloning?

SafetyDGX agent

arXiv:2601.21718v2 Announce Type: replace-cross Abstract: Behavior cloning (BC) is a practical offline imitation learning method, but it often fails when expert demonstrations are limited. Recent work

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models

SafetyDGX agent

arXiv:2606.01671v1 Announce Type: new Abstract: In the contemporary epoch of multilingual education, learning idioms provides a fascinating gateway towards creativity, cultural values, historical cont

Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

SafetyDGX agent

arXiv:2511.05613v2 Announce Type: replace-cross Abstract: Foundation models are increasingly central to high-stakes AI systems, and governance frameworks now depend on evaluations to assess their risk

Why things will eventually fall apart: 1. Everybody, even Google, seems to be treating AI as if it were some kind of winner take all competi…

SafetyDGX agent

Why things will eventually fall apart: 1. Everybody, even Google, seems to be treating AI as if it were some kind of winner take all competition like web search was, in which Google taking over 95% 2.

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications

SafetyDGX agent

arXiv:2606.00133v1 Announce Type: new Abstract: World models, internal simulators that learn the structure and dynamics of an environment, have emerged as a central paradigm in the pursuit of artifici

World Models for Robotic Manipulation: A Survey

SafetyDGX agent

arXiv:2606.00113v1 Announce Type: new Abstract: Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world mode

← Previous
1…100101102103104…214
Next →