AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Generative Learning for Quantum Measurement Design

DGX agent

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting

safetyarxiv-cs-lg
13 Aug 2026
Safety

Generative Semantic Segmentation via an Observable Semantic-Image Interface and Hierarchical Generator Evidence Alignment

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.11537v1 Announce Type: cross Abstract: Generative semantic segmentation exposes structured predictions as images, but direct color decoding is susceptible to color drift and boundary mixing

safetyarxiv-cs-ai
13 Aug 2026
Safety

Grounding Large Language Models as Generalizable Policies in Network Control

DGX agent

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital

safetyarxiv-cs-lg
13 Aug 2026
Safety

Group Alignment-Induced Sycophancy: A Two-Sided Evaluation of Steerable Pluralistic Alignment

DGX agent

arXiv:2608.11528v1 Announce Type: new Abstract: Group alignment adapts a language model to a demographic group to produce responses that reflect the group's opinions, values, and preferences. Sycophan

safetyarxiv-cs-cl
13 Aug 2026
Safety

HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion

DGX agent

arXiv:2608.11913v1 Announce Type: new Abstract: Video-to-audio (V2A) generation faces significant challenges in achieving precise temporal synchronization and high perceptual quality due to the comple

safetyarxiv-cs-cv
13 Aug 2026
Safety

HCGRec: Hint-Conditioned Generative Recommendation with Semantic IDs

DGX agent

arXiv:2608.11980v1 Announce Type: cross Abstract: Semantic-ID generative recommenders represent each item as a short sequence of discrete semantic tokens and predict the next item by autoregressively

safetyarxiv-cs-ai
13 Aug 2026
Safety

Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing

DGX agent

arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge

safetyarxiv-cs-ai
13 Aug 2026
Safety

Instruction Alignment for Binary Code Representation Learning

DGX agent

arXiv:2608.11766v1 Announce Type: cross Abstract: Binary code representation learning is a fundamental problem in software security and reverse engineering. Existing methods mainly learn function-leve

safetyarxiv-cs-ai
13 Aug 2026
Safety

Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces

DGX agent

arXiv:2608.11354v1 Announce Type: new Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet interactions often reflect exploration or comparison rat

safetyarxiv-cs-ai
13 Aug 2026
Safety

IoT-Enabled Autonomous Maritime Navigation in Smart Ports: A Curriculum-Guided Shared Policy Learning Framework

DGX agent

arXiv:2608.11597v1 Announce Type: cross Abstract: As smart port infrastructures increasingly rely on autonomous maritime devices enabled by the Internet of Things (IoT), ensuring reliable onboard navi

safetyarxiv-cs-lg
13 Aug 2026
Safety

Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models

DGX agent

arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unk

safetyarxiv-cs-cl
13 Aug 2026
Safety

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

DGX agent

arXiv:2608.11753v1 Announce Type: new Abstract: Financial text is produced and interpreted within a market environment, yet financial text classifiers almost always receive text alone. We study whethe

safetyarxiv-cs-cl
13 Aug 2026
Safety

Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation

DGX agent

arXiv:2603.13891v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for automated text annotation in tasks ranging from academic research to content moderation

safetyarxiv-cs-ai
13 Aug 2026
Safety

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

DGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

safetyarxiv-cs-cl
13 Aug 2026
Safety

Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic Segmentation

DGX agent

arXiv:2608.11681v1 Announce Type: cross Abstract: This work addresses the challenge of open-vocabulary instance segmentation (OVIS) and open-set panoptic segmentation (OSPS), which aim to recognize bo

safetyarxiv-cs-ai
13 Aug 2026
Safety

Learning from Online User Feedback for Shopping Agents

DGX agent

arXiv:2608.11604v1 Announce Type: new Abstract: Large language model-based shopping agents are increasingly deployed in real-world e-commerce platforms, generating massive amounts of user interaction

safetyarxiv-cs-ai
13 Aug 2026
Safety

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL

DGX agent

arXiv:2608.12063v1 Announce Type: cross Abstract: Integrating locomotion and manipulation is essential for robot autonomy, but scaling standard Reinforcement Learning (RL) to complex tasks is severely

safetyarxiv-cs-ai
13 Aug 2026
Safety

Let it Cook: Learning to Wait in Sequential Decision Making

DGX agent

arXiv:2608.11511v1 Announce Type: cross Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep. However, such active participation may not alwa

safetyarxiv-cs-ai
13 Aug 2026
Safety

LiDAR-based 3D Change Detection at City Scale

DGX agent

arXiv:2510.21112v3 Announce Type: replace-cross Abstract: High-definition 3D city maps enable city planning and change detection, which is essential for municipal compliance, map maintenance, and asse

safetyarxiv-cs-ai
13 Aug 2026
Safety

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

DGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

safetyarxiv-cs-ai
13 Aug 2026
Safety

Machine Learning-Based Cyber Defense for Cloud Infrastructure: An Adaptive Deep Q-Network Architecture for Intelligent Intrusion Detection and Automated Threat Mitigation

DGX agent

arXiv:2608.12190v1 Announce Type: cross Abstract: With the increasing complexity of cyber assaults in cloud environments, adaptable security solutions are needed that can support real-time detection a

safetyarxiv-cs-ai
13 Aug 2026
Safety

NAE: Normalizing AutoEncoder

DGX agent

arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional (d=D) and bottleneck (d<D

safetyarxiv-cs-lg
13 Aug 2026
Safety

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

DGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

safetyarxiv-cs-ai
13 Aug 2026
Safety

PAIR: Pairwise-Aware Inclusion Reweighting for Adaptive Rollout Allocation in RLVR

DGX agent

arXiv:2608.11368v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) spends most of its compute generating groups of long reasoning trajectories. Recent allocators red

safetyarxiv-cs-lg
13 Aug 2026
Safety

Policy-as-logic for robust reasoning over rules

DGX agent

arXiv:2608.11905v1 Announce Type: new Abstract: In many practical applications of generative AI systems, from tax rules to airline baggage allowance, responses to natural language queries must respect

safetyarxiv-cs-ai
13 Aug 2026
Safety

Policy-Induced Hand Priors in Humanoid Dual-Arm Manipulation: Diagnosing and Mitigating Initial-Pose Dependence

DGX agent

arXiv:2608.11769v1 Announce Type: new Abstract: Vision-language-action (VLA) policies are expected to operate robustly across variations in the robot's initial configuration, yet aggregate task succes

safetyarxiv-cs-ro
13 Aug 2026
Safety

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

DGX agent

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context oldsymbol{x}, the model must predict the

safetyarxiv-cs-ai
13 Aug 2026
Safety

Prompt-Driven Exploration

DGX agent

arXiv:2607.08837v2 Announce Type: replace-cross Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject

safetyarxiv-cs-ai
13 Aug 2026
Safety

Proportional Committee Elections with Positive and Negative Votes

DGX agent

arXiv:2503.01985v2 Announce Type: replace-cross Abstract: In the classic committee election setting each voter approves a subset of candidates and the goal is to select k winners based on these prefer

safetyarxiv-cs-ai
13 Aug 2026
Safety

RA-ClipScore: Making Generative Model Evaluation More Interpretable

DGX agent

arXiv:2608.12088v1 Announce Type: new Abstract: Generative models can produce images nearly indistinguishable from real data, yet rigorous and interpretable evaluation remains challenging. Conventiona

safetyarxiv-cs-cv
13 Aug 2026
Safety

Redistribution-based Cost Inference Improves Sparse Safe Offline RL

DGX agent

arXiv:2608.12306v1 Announce Type: cross Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback:

safetyarxiv-cs-ai
13 Aug 2026
Safety

REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation

DGX agent

arXiv:2608.11698v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level supervision from a teacher. Reward-extrapolation methods

safetyarxiv-cs-ai
13 Aug 2026
Safety

Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints

DGX agent

arXiv:2608.11383v1 Announce Type: new Abstract: We study new algorithms for Contextual Bandits with Knapsack. In these problems, there are finitely many types of customers, products, and resources. Ea

safetyarxiv-cs-lg
13 Aug 2026
Safety

Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning

DGX agent

arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalist

safetyarxiv-cs-cl
13 Aug 2026
Safety

ScaleVid: Geometry-Aware Video Object Scaling with Mesh-Free Inference

DGX agent

arXiv:2608.12232v1 Announce Type: new Abstract: Geometry-aware video object scaling aims to anisotropically resize the object along object-centric axes while preserving geometric plausibility, tempora

safetyarxiv-cs-cv
13 Aug 2026
Safety

Small Data Explainer -- The impact of small data methods in everyday life

DGX agent

arXiv:2507.11773v2 Announce Type: replace-cross Abstract: The emergence of breakthrough artificial intelligence (AI) techniques has led to a renewed focus on how small data settings, i.e., settings wi

safetyarxiv-cs-ai
13 Aug 2026
Safety

Through Van Gogh's Eyes: Global Style Transfer with Diffusion Mod

DGX agent

arXiv:2608.11546v1 Announce Type: new Abstract: Artistic image synthesis aims to recreate the expressive visual identity of a target artist, yet existing methods often fail to capture an artist's glob

safetyarxiv-cs-cv
13 Aug 2026
Safety

ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

DGX agent

arXiv:2608.11878v1 Announce Type: cross Abstract: Large language model (LLM) agents integrated with external tools are vulnerable to indirect prompt injections embedded in environmental states. Howeve

safetyarxiv-cs-cl
13 Aug 2026
Safety

Toward Meaningful Transparency for AI Chatbots: Disclosing Persuasive Intent Reduces Persuasion

DGX agent

arXiv:2608.11794v1 Announce Type: cross Abstract: The growing role of AI-generated content and AI-enabled systems in public communication has led regulators to demand clear disclosure of content prove

safetyarxiv-cs-ai
13 Aug 2026
Safety

Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem

DGX agent

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper ta

safetyarxiv-cs-lg
13 Aug 2026
Safety

Towards Human Motion World Models via Executable Behaviour Representations

DGX agent

arXiv:2604.18064v2 Announce Type: replace Abstract: Human motion world models should capture motion's intentionality by being executable: adaptable to different actions and capable of assessing motion

safetyarxiv-cs-ai
13 Aug 2026
Safety

Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling

DGX agent

arXiv:2608.11829v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the stu

safetyarxiv-cs-cl
13 Aug 2026
Safety

UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos

DGX agent

arXiv:2608.11752v1 Announce Type: new Abstract: Talking-video character replacement requires coordinated transfer of appearance and voice while preserving the source motion, scene, linguistic content,

safetyarxiv-cs-cv
13 Aug 2026
Safety

Variable Selection in the Context of AI Fairness

DGX agent

arXiv:2608.11251v1 Announce Type: cross Abstract: Fairness in AI systems has become more important with recent regulatory demands, such as the EU AI Act. Traditional approaches often do not take into

safetyarxiv-cs-ai
13 Aug 2026
Safety

Who Would You Vote For? Auditing Political Alignment in LLMs: An Italian Case-Study

DGX agent

arXiv:2608.11649v1 Announce Type: new Abstract: As users increasingly turn to Large Language Models (LLMs) for information and advice on political matters, particularly during election periods, the po

safetyarxiv-cs-cl
13 Aug 2026
Safety

Why AI Detection Fails for Academic Integrity

DGX agent

arXiv:2608.11256v1 Announce Type: new Abstract: Institutions use commercial AI detectors for academic integrity, yet detectors cannot distinguish AI editing from full LLM drafts and may treat both as

safetyarxiv-cs-lg
13 Aug 2026
Safety

A Joint-Distribution Route to Fair Representations with Continuous Sensitive Attributes

DGX agent

arXiv:2608.10470v1 Announce Type: new Abstract: Fair representation learning with a continuous sensitive attribute S requires a representation Z that is statistically independent of S. Existing criter

safetyarxiv-cs-lg
12 Aug 2026
Safety

Actions Speak Louder than Words: Measuring Cross-Lingual Policy Retention in Tool-Using Agents

DGX agent

arXiv:2608.11110v1 Announce Type: new Abstract: When a tool-using agent is given the same task in a different language, does it still take the same steps? Multilingual evaluation rarely asks: it compa

safetyarxiv-cs-cl
12 Aug 2026
← Previous
1…5758596061…257
Next →