AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
6 Jun 2026

Towards World Models in Biomedical Research

SafetyDGX agent

arXiv:2606.05925v1 Announce Type: new Abstract: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbatio

UniVoice: A Unified Model for Speech and Singing Voice Generation

SafetyDGX agent

arXiv:2606.05852v1 Announce Type: cross Abstract: Text-to-speech (TTS) and singing voice synthesis (SVS) both aim to generate human vocal audio from symbolic inputs, but they impose different requirem

Unsupervised Pattern Analysis in Japanese Veterinary Toxicology: A Regulatory-Compliant Framework for Cross-Species Risk Assessment

SafetyDGX agent

arXiv:2606.06207v1 Announce Type: new Abstract: Veterinary pharmacovigilance systems are essential for monitoring adverse drug events (ADEs), yet existing approaches often fail to capture region-speci


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

We live at a delicate, tragic moment in history, and greed and desperation is probably about to make it much worse.

SafetyDGX agent

Gary Marcus expresses concern about contemporary global instability, suggesting that human greed and desperation pose significant risks to an already precarious historical moment. The post implies tha

Where Should Knowledge Enter? A Layered Framework for Knowledge Infusion in Multimodal Iterative Generative Mo

SafetyDGX agent

arXiv:2606.06356v1 Announce Type: new Abstract: Multimodal generative models produce fluent outputs but remain unreliable when generation must respect structured, domain-specific, or safety-critical k

Whether SpaceX/ Xai is making money on the deals with Google and Anthropic or losing money, they are waving the towel on winning the frontie…

SafetyDGX agent

Whether SpaceX/ Xai is making money on the deals with Google and Anthropic or losing money, they are waving the towel on winning the frontier model race— by arming their competitors rather than themse

White House AI advisor Sriram Krishnan says he will leave his role at the end of June; sources: Krishnan plans to start a pro-Trump AI policy institution (Leo Schwartz/The Information)

SafetyDGX agent

Leo Schwartz / The Information: White House AI advisor Sriram Krishnan says he will leave his role at the end of June; sources: Krishnan plans to start a pro-Trump AI policy institution — Sriram Krish

wild thought: will US invest in Anthropic, the alleged supply chain risk? weird if they do. but if they don’t (but do invest in OpenAI) the …

SafetyDGX agent

wild thought: will US invest in Anthropic, the alleged supply chain risk? weird if they do. but if they don’t (but do invest in OpenAI) the US government may immensely and immediately increase Anthrop

Willing but Unable: Separating Refusal from Capability in Code LLMs via Abliteration

SafetyDGX agent

arXiv:2606.05396v1 Announce Type: cross Abstract: Producing a labeled vulnerable code at scale is a recurring obstacle for learning-based vulnerability detection: mined corpora carry substantial label

Your GFlowNet Secretly Learns an Optimal Transport Plan

SafetyDGX agent

arXiv:2606.06272v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are a framework for sampling structured objects via stochastic trajectories in a directed graph. In this work, we

Zero knowledge verification for frontier AI training is possible

SafetyDGX agent

arXiv:2606.05433v1 Announce Type: new Abstract: Frontier AI governance frameworks increasingly use cumulative training compute as the primary criterion for designating high-impact models, but enforcem

5 Jun 2026

A Komi-Yazva--Russian Parallel Corpus and Evaluation Protocol for Zero- and Few-Shot LLM Translation

SafetyDGX agent

arXiv:2606.06420v1 Announce Type: new Abstract: We present the first Komi-Yazva--Russian parallel corpus together with an explicit evaluation protocol for studying LLM translation in an endangered, ex

A Systematic Analysis of Biases in Large Language Models

SafetyDGX agent

arXiv:2512.15792v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making. However,

Absofuckinglutely called it. Nationalized stakes are just a bailout by a different name. Here we are seventeen months later, and the fleece …

SafetyDGX agent

Absofuckinglutely called it. Nationalized stakes are just a bailout by a different name. Here we are seventeen months later, and the fleece the taxpayer game is on. The countdown until we are told tha

ACE-SQL: Adaptive Co-Optimization via Empirical Credit Assignment for Text-to-SQL

SafetyDGX agent

arXiv:2606.05906v1 Announce Type: new Abstract: Text-to-SQL maps natural language questions to executable SQL queries. Modern databases often contain large and complex schemas, making schema linking a

Adversarial Attacks Already Tell the Answer: Directional Bias-Guided Test-time Defense for Vision-Language Models

SafetyDGX agent

arXiv:2606.06186v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have shown strong zero-shot generalization but remain highly vulnerable to adversarial perturbations, posin

Alignment Risks from Capability-Seeking RL Training

SafetyDGX agent

arXiv:2602.12124v2 Announce Type: replace-cross Abstract: While most AI alignment research focuses on preventing models from generating explicitly harmful content, a more subtle risk arises from capab

“America won’t win the AI race if we beat China but end up with a CCP-style social credit system in the U.S. — and that is the danger as the…

SafetyDGX agent

“America won’t win the AI race if we beat China but end up with a CCP-style social credit system in the U.S. — and that is the danger as the government becomes more deeply involved in AI development a

Analysis of the Neglect-Zero Effect in Large Language Models

SafetyDGX agent

arXiv:2606.05864v1 Announce Type: new Abstract: We investigate the extent to which the language processing of LLMs resembles human cognitive processes, focusing on a human cognitive bias called the ex

Attitude-Aided Linear Calibration of Triaxial Accelerometers

SafetyDGX agent

arXiv:2606.06308v1 Announce Type: new Abstract: Triaxial MEMS accelerometers are widely used for inertial sensing, navigation, and sensor fusion, but existing calibration methods often rely on costly

Auditing Demonstration Curation Metrics: Action-Only Scorers Fail on the Structural Defects That Degrade Imitation Policies

SafetyDGX agent

arXiv:2606.05588v1 Announce Type: new Abstract: Imitation-learning policies inherit the quality of the demonstrations they are trained on, and a growing set of curation metrics promise to score and fi

Beyond Alignment: Value Diversity as a Collective Property in Multicultural Agent Systems

SafetyDGX agent

arXiv:2606.05985v1 Announce Type: new Abstract: Multicultural multi-agent systems are increasingly deployed in globally diverse settings, where different agents are grounded in different cultural back

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

SafetyDGX agent

arXiv:2602.12628v4 Announce Type: replace Abstract: Simulation offers a scalable and low-cost way to enrich vision-language-action (VLA) training, reducing reliance on expensive real-robot demonstrati

Beyond tokens: a unified framework for latent communication in LLM-based multi-agent systems

SafetyDGX agent

arXiv:2606.05711v1 Announce Type: new Abstract: Multi-agent systems built on large language models (LLMs) have become a prevailing paradigm for tackling complex reasoning, planning, and tool-use tasks

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning

SafetyDGX agent

arXiv:2606.05523v1 Announce Type: new Abstract: Despite advances in safety alignment, prompt-rewriting attacks such as persona modulation, fictional framing and persuasion-based reformulation, can byp

'Chi nas dal soch el sent de legn' -- Auditing Text Corpora for Lombard

SafetyDGX agent

arXiv:2606.06349v1 Announce Type: new Abstract: Several of the world's languages are still under-resourced in terms of Natural Language Processing (NLP) tools. This is mostly due to the lack of high-q

Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents

SafetyDGX agent

arXiv:2606.05753v1 Announce Type: new Abstract: Latent visual reasoning (LVR) inserts supervised latent tokens between perception and answer generation in vision-language models (VLMs). The field uses

DexFuture: Hierarchical Future-State Visuomotor Targeting for Bimanual Dexterous Tool Use

SafetyDGX agent

arXiv:2606.05699v1 Announce Type: new Abstract: Bimanual dexterous tool use remains challenging for robots due to high-dimensional hand configurations and complex hand-tool-object dynamics and contact

Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning

SafetyDGX agent

arXiv:2606.05645v1 Announce Type: new Abstract: Autonomous driving requires reasoning about how ego actions shape the evolution of the surrounding world. However, most end-to-end methods rely on direc

Disentangled Fine-Grained Prototype Learning for Incomplete Image-Tabular Classification

SafetyDGX agent

arXiv:2606.05455v1 Announce Type: new Abstract: The missing-modality problem poses a significant challenge in image-tabular multimodal learning across a wide range of multimedia applications, includin

Do Models Share Safety Representations? Cross-Model Steering for Safe Visual Generation

SafetyDGX agent

arXiv:2606.05290v1 Announce Type: new Abstract: Recent progress in generative modeling has made safety control a central challenge, yet existing approaches remain largely model-specific, requiring ret

Drishti AI-Event Guardian: An Intelligent Real-Time Crowd Monitoring and Emergency Response System for Mass Gathering Events

SafetyDGX agent

arXiv:2606.05185v1 Announce Type: cross Abstract: Mass gathering events are associated with critical safety incidents caused by insufficient crowd monitoring and inadequate emergency response coordina

EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration

SafetyDGX agent

arXiv:2602.10106v2 Announce Type: replace Abstract: Human demonstrations offer rich environmental diversity and scale naturally, making them an appealing alternative to robot teleoperation. While this

EMBER: Efficient Memory via Budgeted Evidence Retention for Long-Horizon Agents

SafetyDGX agent

arXiv:2606.05894v1 Announce Type: new Abstract: Long-horizon agents can archive large histories, but future answers still incur retrieval, rereading, and context costs. When retained memory misses ans

Ensuring Interaction Safety in Multitask Exoskeleton Control: A Simulation-Trained Variable Impedance Framework

SafetyDGX agent

arXiv:2606.06370v1 Announce Type: new Abstract: Wearable exoskeletons can augment human phys ical capabilities during complex activities. However, ensuring adaptation across diverse tasks while guaran

Epistemic Injustice in Language Models: An Audit of Pretraining Filters and Guardrails

SafetyDGX agent

arXiv:2606.05936v1 Announce Type: new Abstract: Modern language models rely on pretraining filters to remove undesirable content from training corpora and inference-time guardrails to suppress undesir

EVE: A Generator-Verifier System for Generative Policies

SafetyDGX agent

arXiv:2512.21430v2 Announce Type: replace Abstract: Visuomotor policies based on generative such as diffusion and flow-matching have shown strong performance for robotics applications but degrade unde

Finally, a commencement speaker who calls bullshit. Great oped by @mollyjongfast on “billionaire brain”, and why young people have a right t…

SafetyDGX agent

Finally, a commencement speaker who calls bullshit. Great oped by @mollyjongfast on “billionaire brain”, and why young people have a right to boo what the AI industry has become. I wrote about the com

Flow-based Policy Adaptation without Policy Updates

SafetyDGX agent

arXiv:2606.06461v1 Announce Type: new Abstract: Leveraging prior knowledge from pretrained policies, foundation models, or human operators offers an efficient alternative to learning robot skills from

FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization

SafetyDGX agent

arXiv:2606.05468v1 Announce Type: new Abstract: Post-training Vision-Language-Action (VLA) models into policies that can be reliably deployed on real robots remains a major bottleneck. SFT and DAgger

Forgive or forget: Understanding the context of hate in audio retrieval systems

SafetyDGX agent

arXiv:2606.05857v1 Announce Type: new Abstract: Handling toxic retrieval in text-to-audio systems is challenging due to contextual dependencies. Existing strategies (e.g., rephrasing, summarization) r

Funny how so many people read Anthropic is calling for a pause when they did NOT actually call for a pause. Read what they said, carefully. …

SafetyDGX agent

Funny how so many people read Anthropic is calling for a pause when they did NOT actually call for a pause. Read what they said, carefully. They want it both ways. They *don’t* actually want a pause -

Geometry-Aware Dataset Condensation for Diffusion Model Training

SafetyDGX agent

arXiv:2606.05883v1 Announce Type: new Abstract: Dataset condensation aims to construct compact datasets from real data via synthesis or selection. However, existing approaches are ill-suited for diffu

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech

SafetyDGX agent

arXiv:2606.05889v1 Announce Type: cross Abstract: We propose GLASS, a framework for composable acoustic style control in zero-shot autoregressive text-to-speech (TTS) that learns controls from post-ge

Grounded but Misleading: Evaluating Semantic Alignment in AI-Generated Security Explanations

SafetyDGX agent

arXiv:2602.05056v2 Announce Type: replace-cross Abstract: Online scams increasingly leverage fluent and context-aware social engineering strategies, creating growing demand for AI systems that explain

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

SafetyDGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

SafetyDGX agent

arXiv:2602.16705v3 Announce Type: replace-cross Abstract: Visual loco-manipulation of arbitrary in-the-wild objects requires accurate end-effector (EE) control and a generalizable understanding of the

Human Adults and LLMs as Scientists: Who Benefits from Active Exploration?

SafetyDGX agent

arXiv:2606.06464v1 Announce Type: new Abstract: A long-standing finding in the causal learning literature is that adults struggle to identify conjunctive causal rules, where an effect requires the sim

IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization

SafetyDGX agent

arXiv:2407.10486v3 Announce Type: replace-cross Abstract: Query-focused summarization (QFS) aims to produce summaries that answer particular questions of interest, enabling greater user control and pe

Inverse Design of Realizable Metasurface based Absorbers using Improved Conditioning and Diversity Enhanced Progressively Growing GANs

SafetyDGX agent

arXiv:2606.05849v1 Announce Type: cross Abstract: Metasurfaces enable precise manipulation of electromagnetic waves for applications such as beam steering, sensing, and stealth technology. However, in

Inverse Manipulation through Symbolic Planning and Residual Operator Learning

SafetyDGX agent

arXiv:2606.05248v1 Announce Type: new Abstract: Inverting a robotic task requires more than reversing symbolic state transitions or rewinding motor trajectories. In robot manipulation tasks, symbolic

Is Diversity All You Need for Scalable Robotic Manipulation?

SafetyDGX agent

arXiv:2507.06219v2 Announce Type: replace Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles o

It was a pleasure to discuss all of these topics on @NxThompson's podcast, The Most Interesting Thing in AI. Thanks for having me on! https:…

SafetyDGX agent

Yoshua Bengio appeared as a guest on 'The Most Interesting Thing in AI' podcast hosted by N.X. Thompson, where he discussed various topics related to artificial intelligence. The appearance was shared

⚠️ Keep your eye on the ball, and don’t panic over Anthropic’s new blog. Here’s why: Anthropic is trying to strike terror into everyone’s he…

SafetyDGX agent

⚠️ Keep your eye on the ball, and don’t panic over Anthropic’s new blog. Here’s why: Anthropic is trying to strike terror into everyone’s hearts – “full recursive self-improvement also might increase

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation

SafetyDGX agent

arXiv:2606.06049v1 Announce Type: new Abstract: Intra-vehicular robots in spacecraft help reduce astronaut workload and improve mission efficiency. Recent research focuses on using deep learning metho

LadderMan: Learning Humanoid Perceptive Ladder Climbing

SafetyDGX agent

arXiv:2606.05873v1 Announce Type: cross Abstract: Humanoid robots hold great promise for operating in human-centered environments, yet ladder climbing remains one of the most challenging tasks due to

Large Language Models are Perplexed by some Political Parties

SafetyDGX agent

arXiv:2606.05937v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used, including in political applications, but their political fairness has been little studied. We assess

Latent Reasoning with Normalizing Flows

SafetyDGX agent

arXiv:2606.06447v1 Announce Type: new Abstract: Large language models often improve reasoning by generating explicit chain-of-thought (CoT), demonstrating the importance of intermediate computation. H

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

SafetyDGX agent

arXiv:2606.05952v1 Announce Type: new Abstract: In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model s

Learning Visual Spatial Planning from Symbolic State via Modality-Gap-Aware Self-Distillation

SafetyDGX agent

arXiv:2606.06076v1 Announce Type: cross Abstract: While vision-language models excel at general multimodal understanding, they still struggle with visual spatial planning. We attribute this to a perce

← Previous
1…8788899091…214
Next →