AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
5 Aug 2026

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR

SafetyDGX agent

arXiv:2608.03610v1 Announce Type: new Abstract: Modern LLM-based ASR systems have established multilingual capability as a standard feature, leveraging large-scale multilingual corpora and LLMs' cross

Learning Attribute-aware Representations for Few-shot Scene Text Segmentation

SafetyDGX agent

arXiv:2504.11164v2 Announce Type: replace Abstract: Supervised scene text segmentation has achieved notable progress in recent years. However, its development is largely constrained by the scarcity of

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

SafetyDGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning Context-Aware Motion Priors for Humanoid Control

SafetyDGX agent

arXiv:2608.03234v1 Announce Type: new Abstract: Motion priors provide powerful guidance for learning naturalistic humanoid behaviors. However, existing methods typically learn a general, task-agnostic

Learning Molecular Representations from Cellular Phenotypes with Structure Preservation

SafetyDGX agent

arXiv:2608.02688v1 Announce Type: cross Abstract: Phenotypic drug discovery enables the discovery of functional relationships between molecular structures and cellular responses. However, existing mul

Learning Music Style for Piano Arrangement Through Cross-Modal Bootstrapping

SafetyDGX agent

arXiv:2608.03050v1 Announce Type: cross Abstract: What is music style? Though often described using text labels such as 'swing,' 'classical,' or 'emotional,' the real style remains implicit and hidden

Less Traffic, Better Outcomes: Competition-Aware Request Dispatch in Real-Time Ad Exchanges

SafetyDGX agent

arXiv:2608.03705v1 Announce Type: new Abstract: Real-time bidding (RTB) ad exchanges typically forward nearly all incoming requests to demand-side platforms (DSPs), even though only a small fraction r

Light-Loco-Parkour: Versatile Perceptive Whole-Body Locomotion via Multi-Skill Distillation

SafetyDGX agent

arXiv:2608.02653v1 Announce Type: new Abstract: Existing humanoid whole-body control systems still fall short of the way humans move through cluttered terrain: they either track expressive whole-body

Lightweight 3D Object Detection via Mamba-Based Knowledge Distillation

SafetyDGX agent

arXiv:2608.03490v1 Announce Type: cross Abstract: 3D object detection using light detection and ranging (LiDAR) sensors requires a balance between accuracy and computational efficiency for onboard per

LLM-Derived Priors for Thompson Sampling in Cold-Start Comment Recommendation

SafetyDGX agent

arXiv:2608.03382v1 Announce Type: cross Abstract: Multi-armed bandit algorithms, especially Thompson sampling, are widely used in online recommendation. Despite their ability to adapt from online feed

MambaTS: Improved Selective State Space Models for Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2405.16440v2 Announce Type: replace-cross Abstract: In recent years, Transformers have become the de-facto architecture for long-term time series forecasting (LTSF), yet they face challenges ass

MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion

SafetyDGX agent

arXiv:2508.15169v4 Announce Type: replace Abstract: Mesh models have become increasingly accessible for numerous cities; however, the lack of realistic textures restricts their application in virtual

MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory

SafetyDGX agent

arXiv:2608.02843v1 Announce Type: cross Abstract: Persistent agent memory must adapt as later outcomes change earlier evidence, yet mutable retrieval weights create an attribution problem: reviewers m

OPTD: On-Policy Transition Distillation with Consistency-Guided Adaptive Compression for Few-Step Diffusion Language Models

SafetyDGX agent

arXiv:2608.02942v1 Announce Type: new Abstract: Diffusion language models (dLLMs) can predict many tokens in parallel, but accurate generation still requires many iterative denoising steps. Few-step d

Optimal Liability Design for Medical AI

SafetyDGX agent

arXiv:2608.03114v1 Announce Type: cross Abstract: Artificial intelligence (AI) is increasingly integrated into medical decision-making, yet its liability implications remain complex, particularly when

PACE: Physics Augmentation for Coordinated End-to-end Reinforcement Learning toward Versatile Humanoid Table Tennis

SafetyDGX agent

arXiv:2509.21690v5 Announce Type: replace Abstract: Humanoid table tennis (TT) demands rapid perception, proactive whole-body motion, and agile footwork under strict timing--capabilities that remain d

Path-conditioned Reinforcement Learning-based Local Planning for Long-Range Navigation

SafetyDGX agent

arXiv:2603.13888v2 Announce Type: replace Abstract: Long-range navigation is commonly addressed through hierarchical pipelines in which a global planner generates a path, decomposed into waypoints, an

Perceptual Anchoring: Prototype-Guided Text Calibration for Training-free Open-Vocabulary Semantic Segmentation

SafetyDGX agent

arXiv:2608.03991v1 Announce Type: new Abstract: Training-free open-vocabulary semantic segmentation (OVSS) partitions an image into semantically distinct regions based on arbitrary text descriptions,

Permission Denied: Policy-Graded Evaluation of Coding Agents in Hardened Environments

SafetyDGX agent

arXiv:2608.02670v1 Announce Type: cross Abstract: Coding agents increasingly run inside organizations whose security controls (scoped credentials, restricted egress, read-only filesystems, non-root ex

PFM-HR: Pose Flow Matching for Humanoid Robots

SafetyDGX agent

arXiv:2608.03227v1 Announce Type: new Abstract: Motion priors improve reinforcement learning for physics-based humanoid tracking, but temporal priors require ordered motion clips, while pose priors pr

PhyAI: Real-Time Physical AI at the Edge, Scalable Rollouts in the Cloud

SafetyDGX agent

arXiv:2608.03682v1 Announce Type: new Abstract: Physical AI policies require inference throughout their lifecycle, including model evaluation, cloud reinforcement learning rollout, edge GPU serving, a

Policy Fragmentation or Institutional Alignment? Institutional Governance of AI in Universities and Business Schools

SafetyDGX agent

arXiv:2608.03584v1 Announce Type: new Abstract: Artificial intelligence (AI) is rapidly transforming high-skilled domains, requiring higher education institutions (HEI) to balance the teaching of foun

Predicting Multilingual Classification and Translation Performance of LLMs with Cross-Lingual Alignment nicode{x2013} Is English Enough?

SafetyDGX agent

arXiv:2608.03446v1 Announce Type: new Abstract: Multilingual large language models (LLMs) have been shown to perform better on non-English classification tasks when the representations of the given la

Prescribed-Basis Coefficient-to-Coefficient Neural Operator for Partial Differential Equations

SafetyDGX agent

arXiv:2510.10350v3 Announce Type: replace-cross Abstract: Operator learning provides a data-driven approach to approximating solution operators of partial differential equations, but its effectiveness

Privacy-Preserving AI Verification via Minimal Information Disclosure

SafetyDGX agent

arXiv:2608.02774v1 Announce Type: cross Abstract: AI verification crosses a trust boundary: a verifier must learn enough to establish an authorized claim, yet the same evidence can reveal sensitive de

Quo Vadis, World Modeling?

SafetyDGX agent

arXiv:2608.02713v1 Announce Type: cross Abstract: Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow,

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning

SafetyDGX agent

arXiv:2608.03972v1 Announce Type: new Abstract: On-policy training has emerged as a powerful post-training paradigm for improving the reasoning capabilities of large language models, and is often enha

Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning

SafetyDGX agent

arXiv:2608.02831v1 Announce Type: cross Abstract: Audio reasoning is essential for machine understanding of the acoustic world. Reinforcement learning with verifiable rewards can elicit such reasoning

Relational Priors as Convergence Pressure in LLM-Based Multi-Agent Systems

SafetyDGX agent

arXiv:2608.03239v1 Announce Type: new Abstract: Large language model-based multi-agent systems (LLM-MAS) are designed through roles, debate protocols, and aggregation rules. These choices create impli

Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations

SafetyDGX agent

arXiv:2608.03611v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) integrates text, audio, and vision to infer human affect, yet real-world multimodal observations are often incomplet

RIDGE: Re-Noising with Internal Dynamic Guidance for Image Editing

SafetyDGX agent

arXiv:2608.03059v1 Announce Type: new Abstract: Inversion-free flow-based image editing avoids latent inversion, but still requires a target-side state at every editing step. The widely used equal-dis

Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models

SafetyDGX agent

arXiv:2608.02893v1 Announce Type: cross Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from laten

Scalable Frequency- and Length-Aware Subdocument Deduplication for Large Language Model Pretraining

SafetyDGX agent

arXiv:2608.03089v1 Announce Type: new Abstract: Large-scale pretraining corpora contain substantial duplicate content. Although document-level deduplication is widely used, removing subdocument-level

Self-Organising Digital Circuits

SafetyDGX agent

arXiv:2608.02606v1 Announce Type: new Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware redundancy and error-correcting codes. Biological sys

Self-Supervised Representation-Guided Generative Dataset Distillation

SafetyDGX agent

arXiv:2608.03218v1 Announce Type: cross Abstract: Dataset distillation compresses a large training set into a compact synthetic set while retaining its downstream utility. Most existing methods target

SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation

SafetyDGX agent

arXiv:2608.03092v1 Announce Type: cross Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Opti

so do i have this right? the market is up on optimism about Microsoft’s revenue but MSFT’S biggest customer by far is burning billions a mon…

SafetyDGX agent

so do i have this right? the market is up on optimism about Microsoft’s revenue but MSFT’S biggest customer by far is burning billions a month with no obvious way to meet all of its future obligations

Socially Grounded Agentic AI: Coordinating Plural Perspectives through Social Theory

SafetyDGX agent

arXiv:2608.03910v1 Announce Type: new Abstract: As AI systems are deployed across increasingly diverse social contexts, alignment can no longer be framed as the optimization of a single, unified set o

Solver-Aware Decompositions for Programming-by-Example: When Dividing Requires Knowing how to Conquer

SafetyDGX agent

arXiv:2608.03461v1 Announce Type: new Abstract: Decomposition-based Programming-by-example (PBE) scales performance by splitting tasks into subtasks that a learned synthesizer solves: a decomposer pre

SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling

SafetyDGX agent

arXiv:2608.02951v1 Announce Type: cross Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference

SafetyDGX agent

arXiv:2608.03335v1 Announce Type: new Abstract: Video diffusion transformers (vDiTs) generate high quality but pay quadratic self-attention cost, making inference prohibitive at video-token scales. Th

StreamDAM: Presence-Aware Memory for Real-Time Streaming Video Object Segmentation

SafetyDGX agent

arXiv:2608.03912v1 Announce Type: new Abstract: Quality-tier video object segmentation (VOS) trackers such as DAM4SAM top accuracy leaderboards, but they are measured offline, one frame at a time with

string2string Studio: An Interactive, In-Browser Platform for String-to-String Algorithms

SafetyDGX agent

arXiv:2608.03984v1 Announce Type: new Abstract: We present string2string Studio, an interactive in-browser platform for string-to-string analysis across natural language processing, computational biol

T2VAttack: Adversarial Attack on Text-to-Video Diffusion Models

SafetyDGX agent

arXiv:2512.23953v2 Announce Type: replace Abstract: The rapid evolution of Text-to-Video (T2V) diffusion models has driven remarkable advancements in generating high-quality, temporally coherent video

Test Time Adaptation Methods for Point Cloud Registration in Laparoscopic Surgery

SafetyDGX agent

arXiv:2608.02883v1 Announce Type: new Abstract: 3D point cloud registration in laparoscopic surgery estimates the transformation between an intraoperative organ reconstructed from video and its preope

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

SafetyDGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But tha…

SafetyDGX agent

🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But that moment has come. TBD whether OpenAI and Anthropic can survi

The Evolutionary Origin of Values: implications for AI alignment, sentience and existential risk

SafetyDGX agent

arXiv:2608.03361v1 Announce Type: cross Abstract: AI systems based on Large Language Models (LLMs) have prompted fears that they may harbor hidden goals, seek to dominate or eliminate humanity, or eve

The Signal Horizon: Local Blindness and the Contraction of Pauli-Weight Spectra in Noisy Quantum Encodings

SafetyDGX agent

arXiv:2602.14735v2 Announce Type: replace-cross Abstract: The performance of quantum classifiers is typically analyzed through global state distinguishability or the trainability of variational models

Tired Actor: Fatigue-Informed Character Control

SafetyDGX agent

arXiv:2608.03528v1 Announce Type: new Abstract: Replicating human behavior with physics simulation has been a long-expected goal in character animation. Existing efforts have achieved impressive perfo

Track4Action: Distilling World-Centric 3D Tracker into Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.03727v1 Announce Type: new Abstract: Action labels tell a vision-language-action (VLA) policy which robot commands to imitate, but not how those commands change the 3D world. The aligned de

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2608.04007v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) enables LLMs to solve complex tasks through iterative tool interactions. However, existing reinforcement learning meth

Unified Visuomotor Targets: Supervising VLAs Beyond Physical Actions

SafetyDGX agent

arXiv:2608.03563v1 Announce Type: new Abstract: VLA models are trained to predict robot actions from visual and language observations. This is a natural choice, but it creates a mismatch: VLMs encode

ValueFormer: A Causal Transformer Value Function with Stage-Aware Labels for Semi-Autonomous Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.02958v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies trained by behavior cloning fail silently: from the action stream alone, a collapsing rollout looks much like on

Welcome to the thunderdome of commoditization

SafetyDGX agent

Welcome to the thunderdome of commoditization 🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But that moment has co

What is the Right Embedding Space for Contrastive Learning in REC?

SafetyDGX agent

arXiv:2505.22850v2 Announce Type: replace Abstract: Referring Expression Counting (REC) requires distinguishing visually similar objects described by fine-grained text cues. Existing methods tackle th

When Oracle Conditioning Misleads Deployment: Conditioning-Availability Bias in Echocardiographic Segmentation

SafetyDGX agent

arXiv:2608.03342v1 Announce Type: cross Abstract: Conditional segmentation models may be trained and evaluated with auxiliary signals cleaner than those available at deployment. We study this protocol

When Policies Change Probabilities: Modular Decision-Making for LLM Code Review

SafetyDGX agent

arXiv:2608.02677v1 Announce Type: cross Abstract: LLM code reviewers often estimate patch risk and make approval decisions in one prompt. A probability should depend on evidence; costs should determin

When Search Teaches Style: Causal Internalization of Tactical Priors in AlphaZero

SafetyDGX agent

arXiv:2504.14636v3 Announce Type: replace-cross Abstract: AlphaZero is normally evaluated as one agent: a policy-value network fused with Monte Carlo tree search. That fusion hides a causal question.

When Teachers Mislead: Spurious-Signal-Aware On-Policy Distillation

SafetyDGX agent

arXiv:2608.03632v1 Announce Type: new Abstract: On-Policy distillation (OPD) transfers teacher capabilities by supervising student-sampled trajectories with dense token-level teacher signals. Recent s

← Previous
1…6364656667…242
Next →