AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation

DGX agent

arXiv:2608.11590v1 Announce Type: cross Abstract: Human voice generation has made rapid progress in speech generation, singing voice generation, voice cloning, and voice editing. However, most existin

safetyarxiv-cs-lg
13 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Defending against Model Extraction for GNNs with Model Reprogramming

DGX agent

arXiv:2608.11495v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications in Machine-Learning-as-a-Service (MLaaS). Still, their black-box deploym

safetyarxiv-cs-lg
13 Aug 2026
Safety

Diagnosis Before Recovery: Turning Agent Failures into Selective Self-Correction

DGX agent

arXiv:2608.11772v1 Announce Type: new Abstract: Self-correction is particularly useful when a failure constrains the next repair. Coding agents benefit from this property because compilers, tests, and

safetyarxiv-cs-cl
13 Aug 2026
Safety

Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks

DGX agent

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater v

safetyarxiv-cs-lg
13 Aug 2026
Safety

Disentangling the Expressivity of RoPE

DGX agent

arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information wit

safetyarxiv-cs-lg
13 Aug 2026
Safety

Dual Anchors, Do It Better: Hierarchical Group Merging for Zero-Shot Anomaly Detection

DGX agent

arXiv:2608.11933v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to identify anomalies in unseen domains, a setting that is particularly critical for industrial and medical appl

safetyarxiv-cs-cv
13 Aug 2026
Safety

Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation

DGX agent

arXiv:2608.11335v1 Announce Type: cross Abstract: Clinical text can narrow down what to segment, but recent text-guided designs emphasize spatial alignment while overlooking frequency content that gov

safetyarxiv-cs-ai
13 Aug 2026
Safety

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

DGX agent

arXiv:2608.11207v1 Announce Type: new Abstract: When two LLM agents with structurally opposed objectives interact across multiple turns, the absence of a shared goal function produces not competition

safetyarxiv-cs-ai
13 Aug 2026
Safety

Enhancing Visual Domain Robustness in Behaviour Cloning via Saliency-Guided Augmentation

DGX agent

arXiv:2608.11870v1 Announce Type: new Abstract: In vision-based behavior cloning (BC), conventional image augmentations such as Random Crop and Color Jitter often fall short under substantial visual d

safetyarxiv-cs-ro
13 Aug 2026
Safety

Epiplexity Guided Data Selection and Generation for Out-of-Distribution Generalization

DGX agent

arXiv:2608.11746v1 Announce Type: cross Abstract: Modern systems are increasingly expected to transfer across tasks not specified during training. What data facilitates generalization in these new, un

safetyarxiv-cs-cl
13 Aug 2026
Safety

FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting

DGX agent

arXiv:2608.11623v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily

safetyarxiv-cs-ai
13 Aug 2026
Safety

Foresight Without Seeing: Latent Futures for World Action Models

DGX agent

arXiv:2608.11605v1 Announce Type: new Abstract: World Action Models (WAMs) couple future visual prediction with robot action generation, enabling policies to model how the physical world evolves durin

safetyarxiv-cs-ai
13 Aug 2026
Safety

FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees

DGX agent

arXiv:2608.12140v1 Announce Type: cross Abstract: Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing design

safetyarxiv-cs-lg
13 Aug 2026
Safety

From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation

DGX agent

arXiv:2608.11493v1 Announce Type: new Abstract: Traditional offline recommendation evaluation relies heavily on complex, manually maintained feature pipelines that are difficult to scale. While Large

safetyarxiv-cs-ai
13 Aug 2026
Safety

GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs

DGX agent

arXiv:2608.11674v1 Announce Type: cross Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities,

safetyarxiv-cs-ai
13 Aug 2026
Safety

Generative Learning for Quantum Measurement Design

DGX agent

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting

safetyarxiv-cs-lg
13 Aug 2026
Safety

Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment

DGX agent

arXiv:2608.11537v1 Announce Type: cross Abstract: Generative semantic segmentation exposes structured predictions as images, but direct color decoding is susceptible to color drift and boundary mixing

safetyarxiv-cs-ai
13 Aug 2026
Safety

Grounding Large Language Models as Generalizable Policies in Network Control

DGX agent

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital

safetyarxiv-cs-lg
13 Aug 2026
Safety

Group Alignment-Induced Sycophancy: A Two-Sided Evaluation of Steerable Pluralistic Alignment

DGX agent

arXiv:2608.11528v1 Announce Type: new Abstract: Group alignment adapts a language model to a demographic group to produce responses that reflect the group's opinions, values, and preferences. Sycophan

safetyarxiv-cs-cl
13 Aug 2026
Safety

HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion

DGX agent

arXiv:2608.11913v1 Announce Type: new Abstract: Video-to-audio (V2A) generation faces significant challenges in achieving precise temporal synchronization and high perceptual quality due to the comple

safetyarxiv-cs-cv
13 Aug 2026
Safety

HCGRec: Hint-Conditioned Generative Recommendation with Semantic IDs

DGX agent

arXiv:2608.11980v1 Announce Type: cross Abstract: Semantic-ID generative recommenders represent each item as a short sequence of discrete semantic tokens and predict the next item by autoregressively

safetyarxiv-cs-ai
13 Aug 2026
Safety

Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing

DGX agent

arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge

safetyarxiv-cs-ai
13 Aug 2026
Safety

I am a scientist. I don’t know for absolute sure that my conjectures about Generative AI will prove correct. But I do know for sure that peo…

DGX agent

Gary Marcus posted on 13 Aug 2026 that, as a scientist, he is uncertain whether his conjectures about generative AI will be proven correct but remains confident that critics of the technology are pres

safetygary-marcus--x
13 Aug 2026
Safety

Instruction Alignment for Binary Code Representation Learning

DGX agent

arXiv:2608.11766v1 Announce Type: cross Abstract: Binary code representation learning is a fundamental problem in software security and reverse engineering. Existing methods mainly learn function-leve

safetyarxiv-cs-ai
13 Aug 2026
Safety

Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces

DGX agent

arXiv:2608.11354v1 Announce Type: new Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet interactions often reflect exploration or comparison rat

safetyarxiv-cs-ai
13 Aug 2026
Safety

IoT-Enabled Autonomous Maritime Navigation in Smart Ports: A Curriculum-Guided Shared Policy Learning Framework

DGX agent

arXiv:2608.11597v1 Announce Type: cross Abstract: As smart port infrastructures increasingly rely on autonomous maritime devices enabled by the Internet of Things (IoT), ensuring reliable onboard navi

safetyarxiv-cs-lg
13 Aug 2026
Safety

Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models

DGX agent

arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unk

safetyarxiv-cs-cl
13 Aug 2026
Safety

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

DGX agent

arXiv:2608.11753v1 Announce Type: new Abstract: Financial text is produced and interpreted within a market environment, yet financial text classifiers almost always receive text alone. We study whethe

safetyarxiv-cs-cl
13 Aug 2026
Safety

Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation

DGX agent

arXiv:2603.13891v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for automated text annotation in tasks ranging from academic research to content moderation

safetyarxiv-cs-ai
13 Aug 2026
Safety

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

DGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

safetyarxiv-cs-cl
13 Aug 2026
Safety

Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic Segmentation

DGX agent

arXiv:2608.11681v1 Announce Type: cross Abstract: This work addresses the challenge of open-vocabulary instance segmentation (OVIS) and open-set panoptic segmentation (OSPS), which aim to recognize bo

safetyarxiv-cs-ai
13 Aug 2026
Safety

Learning from Online User Feedback for Shopping Agents

DGX agent

arXiv:2608.11604v1 Announce Type: new Abstract: Large language model-based shopping agents are increasingly deployed in real-world e-commerce platforms, generating massive amounts of user interaction

safetyarxiv-cs-ai
13 Aug 2026
Safety

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL

DGX agent

arXiv:2608.12063v1 Announce Type: cross Abstract: Integrating locomotion and manipulation is essential for robot autonomy, but scaling standard Reinforcement Learning (RL) to complex tasks is severely

safetyarxiv-cs-ai
13 Aug 2026
Safety

Let it Cook: Learning to Wait in Sequential Decision Making

DGX agent

arXiv:2608.11511v1 Announce Type: cross Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep. However, such active participation may not alwa

safetyarxiv-cs-ai
13 Aug 2026
Safety

LiDAR-based 3D Change Detection at City Scale

DGX agent

arXiv:2510.21112v3 Announce Type: replace-cross Abstract: High-definition 3D city maps enable city planning and change detection, which is essential for municipal compliance, map maintenance, and asse

safetyarxiv-cs-ai
13 Aug 2026
Safety

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

DGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

safetyarxiv-cs-ai
13 Aug 2026
Safety

Machine Learning-Based Cyber Defense for Cloud Infrastructure: An Adaptive Deep Q-Network Architecture for Intelligent Intrusion Detection and Automated Threat Mitigation

DGX agent

arXiv:2608.12190v1 Announce Type: cross Abstract: With the increasing complexity of cyber assaults in cloud environments, adaptable security solutions are needed that can support real-time detection a

safetyarxiv-cs-ai
13 Aug 2026
Safety

NAE: Normalizing AutoEncoder

DGX agent

arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional (d=D) and bottleneck (d<D

safetyarxiv-cs-lg
13 Aug 2026
Safety

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

DGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

safetyarxiv-cs-ai
13 Aug 2026
Safety

PAIR: Pairwise-Aware Inclusion Reweighting for Adaptive Rollout Allocation in RLVR

DGX agent

arXiv:2608.11368v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) spends most of its compute generating groups of long reasoning trajectories. Recent allocators red

safetyarxiv-cs-lg
13 Aug 2026
Safety

Policy-as-logic for robust reasoning over rules

DGX agent

arXiv:2608.11905v1 Announce Type: new Abstract: In many practical applications of generative AI systems, from tax rules to airline baggage allowance, responses to natural language queries must respect

safetyarxiv-cs-ai
13 Aug 2026
Safety

Policy-Induced Hand Priors in Humanoid Dual-Arm Manipulation: Diagnosing and Mitigating Initial-Pose Dependence

DGX agent

arXiv:2608.11769v1 Announce Type: new Abstract: Vision-language-action (VLA) policies are expected to operate robustly across variations in the robot's initial configuration, yet aggregate task succes

safetyarxiv-cs-ro
13 Aug 2026
Safety

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

DGX agent

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context oldsymbol{x}, the model must predict the

safetyarxiv-cs-ai
13 Aug 2026
Safety

Prompt-Driven Exploration

DGX agent

arXiv:2607.08837v2 Announce Type: replace-cross Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject

safetyarxiv-cs-ai
13 Aug 2026
Safety

Proportional Committee Elections with Positive and Negative Votes

DGX agent

arXiv:2503.01985v2 Announce Type: replace-cross Abstract: In the classic committee election setting each voter approves a subset of candidates and the goal is to select k winners based on these prefer

safetyarxiv-cs-ai
13 Aug 2026
Safety

RA-ClipScore: Making Generative Model Evaluation More Interpretable

DGX agent

arXiv:2608.12088v1 Announce Type: new Abstract: Generative models can produce images nearly indistinguishable from real data, yet rigorous and interpretable evaluation remains challenging. Conventiona

safetyarxiv-cs-cv
13 Aug 2026
Safety

Redistribution-based Cost Inference Improves Sparse Safe Offline RL

DGX agent

arXiv:2608.12306v1 Announce Type: cross Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback:

safetyarxiv-cs-ai
13 Aug 2026
Safety

REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation

DGX agent

arXiv:2608.11698v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods

safetyarxiv-cs-ai
13 Aug 2026
← Previous
1…6465666768…300
Next →