AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
13 Aug 2026

CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation

SafetyDGX agent

arXiv:2608.11590v1 Announce Type: cross Abstract: Human voice generation has made rapid progress in speech generation, singing voice generation, voice cloning, and voice editing. However, most existin

Defending against Model Extraction for GNNs with Model Reprogramming

SafetyDGX agent

arXiv:2608.11495v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications in Machine-Learning-as-a-Service (MLaaS). Still, their black-box deploym

Diagnosis Before Recovery: Turning Agent Failures into Selective Self-Correction

SafetyDGX agent

arXiv:2608.11772v1 Announce Type: new Abstract: Self-correction is particularly useful when a failure constrains the next repair. Coding agents benefit from this property because compilers, tests, and

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks

SafetyDGX agent

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater v

Disentangling the Expressivity of RoPE

SafetyDGX agent

arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information wit

Dual Anchors, Do It Better: Hierarchical Group Merging for Zero-Shot Anomaly Detection

SafetyDGX agent

arXiv:2608.11933v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to identify anomalies in unseen domains, a setting that is particularly critical for industrial and medical appl

Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation

SafetyDGX agent

arXiv:2608.11335v1 Announce Type: cross Abstract: Clinical text can narrow down what to segment, but recent text-guided designs emphasize spatial alignment while overlooking frequency content that gov

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

SafetyDGX agent

arXiv:2608.11207v1 Announce Type: new Abstract: When two LLM agents with structurally opposed objectives interact across multiple turns, the absence of a shared goal function produces not competition

Enhancing Visual Domain Robustness in Behaviour Cloning via Saliency-Guided Augmentation

SafetyDGX agent

arXiv:2608.11870v1 Announce Type: new Abstract: In vision-based behavior cloning (BC), conventional image augmentations such as Random Crop and Color Jitter often fall short under substantial visual d

Epiplexity Guided Data Selection and Generation for Out-of-Distribution Generalization

SafetyDGX agent

arXiv:2608.11746v1 Announce Type: cross Abstract: Modern systems are increasingly expected to transfer across tasks not specified during training. What data facilitates generalization in these new, un

FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting

SafetyDGX agent

arXiv:2608.11623v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily

Foresight Without Seeing: Latent Futures for World Action Models

SafetyDGX agent

arXiv:2608.11605v1 Announce Type: new Abstract: World Action Models (WAMs) couple future visual prediction with robot action generation, enabling policies to model how the physical world evolves durin

FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees

SafetyDGX agent

arXiv:2608.12140v1 Announce Type: cross Abstract: Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing design

From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation

SafetyDGX agent

arXiv:2608.11493v1 Announce Type: new Abstract: Traditional offline recommendation evaluation relies heavily on complex, manually maintained feature pipelines that are difficult to scale. While Large

GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs

SafetyDGX agent

arXiv:2608.11674v1 Announce Type: cross Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities,

Generative Learning for Quantum Measurement Design

SafetyDGX agent

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting

Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment

SafetyDGX agent

arXiv:2608.11537v1 Announce Type: cross Abstract: Generative semantic segmentation exposes structured predictions as images, but direct color decoding is susceptible to color drift and boundary mixing

Grounding Large Language Models as Generalizable Policies in Network Control

SafetyDGX agent

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital

Group Alignment-Induced Sycophancy: A Two-Sided Evaluation of Steerable Pluralistic Alignment

SafetyDGX agent

arXiv:2608.11528v1 Announce Type: new Abstract: Group alignment adapts a language model to a demographic group to produce responses that reflect the group's opinions, values, and preferences. Sycophan

HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion

SafetyDGX agent

arXiv:2608.11913v1 Announce Type: new Abstract: Video-to-audio (V2A) generation faces significant challenges in achieving precise temporal synchronization and high perceptual quality due to the comple

HCGRec: Hint-Conditioned Generative Recommendation with Semantic IDs

SafetyDGX agent

arXiv:2608.11980v1 Announce Type: cross Abstract: Semantic-ID generative recommenders represent each item as a short sequence of discrete semantic tokens and predict the next item by autoregressively

Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing

SafetyDGX agent

arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge

I am a scientist. I don’t know for absolute sure that my conjectures about Generative AI will prove correct. But I do know for sure that peo…

SafetyDGX agent

Gary Marcus posted on 13 Aug 2026 that, as a scientist, he is uncertain whether his conjectures about generative AI will be proven correct but remains confident that critics of the technology are pres

Instruction Alignment for Binary Code Representation Learning

SafetyDGX agent

arXiv:2608.11766v1 Announce Type: cross Abstract: Binary code representation learning is a fundamental problem in software security and reverse engineering. Existing methods mainly learn function-leve

Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces

SafetyDGX agent

arXiv:2608.11354v1 Announce Type: new Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet interactions often reflect exploration or comparison rat

IoT-Enabled Autonomous Maritime Navigation in Smart Ports: A Curriculum-Guided Shared Policy Learning Framework

SafetyDGX agent

arXiv:2608.11597v1 Announce Type: cross Abstract: As smart port infrastructures increasingly rely on autonomous maritime devices enabled by the Internet of Things (IoT), ensuring reliable onboard navi

Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models

SafetyDGX agent

arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unk

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

SafetyDGX agent

arXiv:2608.11753v1 Announce Type: new Abstract: Financial text is produced and interpreted within a market environment, yet financial text classifiers almost always receive text alone. We study whethe

Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation

SafetyDGX agent

arXiv:2603.13891v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for automated text annotation in tasks ranging from academic research to content moderation

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

SafetyDGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic Segmentation

SafetyDGX agent

arXiv:2608.11681v1 Announce Type: cross Abstract: This work addresses the challenge of open-vocabulary instance segmentation (OVIS) and open-set panoptic segmentation (OSPS), which aim to recognize bo

Learning from Online User Feedback for Shopping Agents

SafetyDGX agent

arXiv:2608.11604v1 Announce Type: new Abstract: Large language model-based shopping agents are increasingly deployed in real-world e-commerce platforms, generating massive amounts of user interaction

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL

SafetyDGX agent

arXiv:2608.12063v1 Announce Type: cross Abstract: Integrating locomotion and manipulation is essential for robot autonomy, but scaling standard Reinforcement Learning (RL) to complex tasks is severely

Let it Cook: Learning to Wait in Sequential Decision Making

SafetyDGX agent

arXiv:2608.11511v1 Announce Type: cross Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep. However, such active participation may not alwa

LiDAR-based 3D Change Detection at City Scale

SafetyDGX agent

arXiv:2510.21112v3 Announce Type: replace-cross Abstract: High-definition 3D city maps enable city planning and change detection, which is essential for municipal compliance, map maintenance, and asse

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

SafetyDGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

Machine Learning-Based Cyber Defense for Cloud Infrastructure: An Adaptive Deep Q-Network Architecture for Intelligent Intrusion Detection and Automated Threat Mitigation

SafetyDGX agent

arXiv:2608.12190v1 Announce Type: cross Abstract: With the increasing complexity of cyber assaults in cloud environments, adaptable security solutions are needed that can support real-time detection a

NAE: Normalizing AutoEncoder

SafetyDGX agent

arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional (d=D) and bottleneck (d<D

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

SafetyDGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

PAIR: Pairwise-Aware Inclusion Reweighting for Adaptive Rollout Allocation in RLVR

SafetyDGX agent

arXiv:2608.11368v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) spends most of its compute generating groups of long reasoning trajectories. Recent allocators red

Policy-as-logic for robust reasoning over rules

SafetyDGX agent

arXiv:2608.11905v1 Announce Type: new Abstract: In many practical applications of generative AI systems, from tax rules to airline baggage allowance, responses to natural language queries must respect

Policy-Induced Hand Priors in Humanoid Dual-Arm Manipulation: Diagnosing and Mitigating Initial-Pose Dependence

SafetyDGX agent

arXiv:2608.11769v1 Announce Type: new Abstract: Vision-language-action (VLA) policies are expected to operate robustly across variations in the robot's initial configuration, yet aggregate task succes

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

SafetyDGX agent

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context oldsymbol{x}, the model must predict the

Prompt-Driven Exploration

SafetyDGX agent

arXiv:2607.08837v2 Announce Type: replace-cross Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject

Proportional Committee Elections with Positive and Negative Votes

SafetyDGX agent

arXiv:2503.01985v2 Announce Type: replace-cross Abstract: In the classic committee election setting each voter approves a subset of candidates and the goal is to select k winners based on these prefer

RA-ClipScore: Making Generative Model Evaluation More Interpretable

SafetyDGX agent

arXiv:2608.12088v1 Announce Type: new Abstract: Generative models can produce images nearly indistinguishable from real data, yet rigorous and interpretable evaluation remains challenging. Conventiona

Redistribution-based Cost Inference Improves Sparse Safe Offline RL

SafetyDGX agent

arXiv:2608.12306v1 Announce Type: cross Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback:

REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation

SafetyDGX agent

arXiv:2608.11698v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods

Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints

SafetyDGX agent

arXiv:2608.11383v1 Announce Type: new Abstract: We study new algorithms for Contextual Bandits with Knapsack. In these problems, there are finitely many types of customers, products, and resources. Ea

Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning

SafetyDGX agent

arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalist

ScaleVid: Geometry-Aware Video Object Scaling with Mesh-Free Inference

SafetyDGX agent

arXiv:2608.12232v1 Announce Type: new Abstract: Geometry-aware video object scaling aims to anisotropically resize the object along object-centric axes while preserving geometric plausibility, tempora

Small Data Explainer -- The impact of small data methods in everyday life

SafetyDGX agent

arXiv:2507.11773v2 Announce Type: replace-cross Abstract: The emergence of breakthrough artificial intelligence (AI) techniques has led to a renewed focus on how small data settings, i.e., settings wi

Stolen LLM Reasoning: How come OpenAI, Anthrophic, Google have the same vulnerabilities?

SafetyDGX agent

If you haven't checked the paper: https://arxiv.org/abs/2608.09867 TLDR: the authors show that you can swap out the 'encrypted' reasoning of the biggest model, like Opus, Sol, and put them into weaker

Through Van Gogh's Eyes: Global Style Transfer with Diffusion Mod

SafetyDGX agent

arXiv:2608.11546v1 Announce Type: new Abstract: Artistic image synthesis aims to recreate the expressive visual identity of a target artist, yet existing methods often fail to capture an artist's glob

ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

SafetyDGX agent

arXiv:2608.11878v1 Announce Type: cross Abstract: Large language model (LLM) agents integrated with external tools are vulnerable to indirect prompt injections embedded in environmental states. Howeve

Toward Meaningful Transparency for AI Chatbots: Disclosing Persuasive Intent Reduces Persuasion

SafetyDGX agent

arXiv:2608.11794v1 Announce Type: cross Abstract: The growing role of AI-generated content and AI-enabled systems in public communication has led regulators to demand clear disclosure of content prove

Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem

SafetyDGX agent

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper ta

Towards Human Motion World Models via Executable Behaviour Representations

SafetyDGX agent

arXiv:2604.18064v2 Announce Type: replace Abstract: Human motion world models should capture motion's intentionality by being executable: adaptable to different actions and capable of assessing motion

Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling

SafetyDGX agent

arXiv:2608.11829v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the stu

UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos

SafetyDGX agent

arXiv:2608.11752v1 Announce Type: new Abstract: Talking-video character replacement requires coordinated transfer of appearance and voice while preserving the source motion, scene, linguistic content,

← Previous
1…5152535455…240
Next →