AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

HSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs

DGX agent

arXiv:2607.27379v1 Announce Type: new Abstract: High-quality, diverse data are vital for large language models (LLMs) but remain scarce and costly. Data synthesis is a viable alternative and succeeds

safetyarxiv-cs-cl
31 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning

DGX agent

arXiv:2607.27564v1 Announce Type: new Abstract: Multi-image medical VQA is not merely a prompt-length problem; it is a fundamental challenge of agentic decision-making. Medical vision-language agents

safetyarxiv-cs-cv
31 Jul 2026
Safety

Integrating Contextual Embeddings into Evaluation of Expressive MIDI Piano Performances

DGX agent

arXiv:2607.27909v1 Announce Type: cross Abstract: Objective evaluation of expressive MIDI piano performances typically relies on attribute statistics such as timing, velocity, and duration of individu

safetyarxiv-cs-lg
31 Jul 2026
Safety

It's Not Just More Demos: Counterfactual Action Sensitivity Coverage for Data-Efficient Robust Robot Imitation

DGX agent

arXiv:2607.27261v1 Announce Type: new Abstract: Visuomotor imitation learning has demonstrated success for manipulation tasks. However, the trained policies remain brittle to visual `nuisances', with

safetyarxiv-cs-ro
31 Jul 2026
Safety

Learning Social Robot Navigation By Sensing Human Legs

DGX agent

arXiv:2607.27922v1 Announce Type: new Abstract: Robots navigating among pedestrians typically sense their surroundings with a 2D LiDAR mounted close to the ground. At that height, the sensor mostly se

safetyarxiv-cs-ro
31 Jul 2026
Safety

MedXplore: Towards Reliable and Unbiased Generalized Category Discovery in Medical Imaging

DGX agent

arXiv:2607.27620v1 Announce Type: new Abstract: Deep learning has shown strong potential in medical image analysis, but most existing methods rely on large-scale annotations and a closed-world assumpt

safetyarxiv-cs-cv
31 Jul 2026
Safety

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

DGX agent

arXiv:2607.28565v1 Announce Type: new Abstract: Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typical

safetyarxiv-cs-cv
31 Jul 2026
Safety

MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation

DGX agent

arXiv:2607.26698v1 Announce Type: cross Abstract: Cover song generation (CSG) should preserve the melodic and linguistic content of a reference song while recreating the remaining musical components.

safetyarxiv-cs-ai
31 Jul 2026
Safety

MUGEN: A Unified Framework for Efficient Motion Understanding and Generation

DGX agent

arXiv:2607.27581v1 Announce Type: new Abstract: Grounding human motion in language, and language in motion, is a central step toward physical AI systems that can understand, generate, and communicate

safetyarxiv-cs-lg
31 Jul 2026
Safety

Multi-channel Uplift Policy Learning

DGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

safetyarxiv-cs-lg
31 Jul 2026
Safety

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

DGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

safetyarxiv-cs-cv
31 Jul 2026
Safety

On-Policy and Off-Policy Learning for Large Action Spaces

DGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

safetyarxiv-cs-lg
31 Jul 2026
Safety

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

DGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

safetyarxiv-cs-lg
31 Jul 2026
Safety

OPLD: On-Policy Latent Distillation for Multimodal Reasoning

DGX agent

arXiv:2607.28154v1 Announce Type: new Abstract: Interleaved multimodal Chain-of-Thought (CoT) improves visual reasoning by incorporating auxiliary visual evidence into intermediate reasoning. However,

safetyarxiv-cs-cv
31 Jul 2026
Safety

Optimizing Regret

DGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

safetyarxiv-cs-lg
31 Jul 2026
Safety

Policy Gradient Steering: Interventions from Behavioral Objectives

DGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

safetyarxiv-cs-lg
31 Jul 2026
Safety

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation

DGX agent

arXiv:2506.21076v4 Announce Type: replace Abstract: Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domain

safetyarxiv-cs-cv
31 Jul 2026
Safety

Procedural Fairness in Multi-Agent Bandits

DGX agent

arXiv:2601.10600v2 Announce Type: replace-cross Abstract: In the context of multi-agent multi-armed bandits (MA-MAB), fairness is often reduced to outcomes: maximizing welfare, reducing inequality, or

safetyarxiv-cs-lg
31 Jul 2026
Safety

QQWorld: Quantile-Quantile Matching for World Model Regularization

DGX agent

arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically

safetyarxiv-cs-cv
31 Jul 2026
Safety

Recognition and Label-Free Adaptation Across Recording Sessions in Surface-EMG Gesture Decoding

DGX agent

arXiv:2607.27568v1 Announce Type: new Abstract: Recognition accuracy obtained during a recording session does not persist when a user puts on the electrodes again after the electrodes had previously b

safetyarxiv-cs-lg
31 Jul 2026
Safety

ReDiPPO: Reference-Guided Value Calibration and Discrepancy-Aware Token Reweighting for Mathematical Reasoning

DGX agent

arXiv:2607.27631v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for enhancing the mathematical reasoning capabilities of large language models. Among exis

safetyarxiv-cs-cl
31 Jul 2026
Safety

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

DGX agent

arXiv:2603.13707v3 Announce Type: replace-cross Abstract: Humanoid loco-manipulation requires coordinated task-space motion planning with stable loco-manipulation command tracking under complex robot-

safetyarxiv-cs-lg
31 Jul 2026
Safety

Regularizing modality contribution drift in multimodal continual learning

DGX agent

arXiv:2607.27260v1 Announce Type: new Abstract: Multimodal continual learning (MMCL) aims to learn emerging knowledge from multimodal data while preserving knowledge. To mitigate forgetting, current M

safetyarxiv-cs-lg
31 Jul 2026
Safety

Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review

DGX agent

arXiv:2607.27209v1 Announce Type: cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to

safetyarxiv-cs-lg
31 Jul 2026
Safety

ROAD: Reciprocal-Objective Alignment of Discriminative Semantics for 3D Shape Generation

DGX agent

arXiv:2607.28581v1 Announce Type: new Abstract: High-fidelity 3D generation predominantly relies on scaling model capacity and data, which incurs prohibitive computational costs. This paradigm typical

safetyarxiv-cs-cv
31 Jul 2026
Safety

Robust Estimation of Sparse Numerical Vectors under Local Differential Privacy

DGX agent

arXiv:2607.27815v1 Announce Type: cross Abstract: Local differential privacy (LDP) protocols are vulnerable to poisoning attacks. Existing research have proposed efficient defense strategies for singl

safetyarxiv-cs-lg
31 Jul 2026
Safety

SCOPE: Supply-Chain Operations through Coupled Policies for End-to-End Coordination

DGX agent

arXiv:2607.28488v1 Announce Type: cross Abstract: Can supply-chain AI move beyond isolated decision modules toward unified operational planning? A complete replenishment plan specifies which products

safetyarxiv-cs-lg
31 Jul 2026
Safety

ServerlessT2I: Efficient Text-to-Image Workflow Serving on a Serverless Platform

DGX agent

arXiv:2607.26566v1 Announce Type: cross Abstract: Text-to-image (T2I) workflows are increasingly deployed on serverless platforms because users often compose customized workflows and invoke them inter

safetyarxiv-cs-ai
31 Jul 2026
Safety

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

DGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

safetyarxiv-cs-ai
31 Jul 2026
Safety

SkillSight: Calibrating Generic Content Bias for Skill Retrieval

DGX agent

arXiv:2607.18785v2 Announce Type: replace Abstract: As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability

safetyarxiv-cs-ai
31 Jul 2026
Safety

Static In, Dynamic Out: Counterfactual Action Augmentation for Moving Object Manipulation

DGX agent

arXiv:2607.27890v1 Announce Type: new Abstract: Visuomotor policies have advanced on manipulation tasks where the target object stays static during execution, but real deployments break this assumptio

safetyarxiv-cs-ro
31 Jul 2026
Safety

Strategies for Milestone-driven Start-ups in Multi-activity Settings

DGX agent

arXiv:2607.27563v1 Announce Type: new Abstract: New venture start-ups need to ``survive'' through multiple stages of reaching milestone targets. We investigate the strategies for start-ups in a milest

safetyarxiv-cs-lg
31 Jul 2026
Safety

SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute

DGX agent

arXiv:2607.28457v1 Announce Type: cross Abstract: Scaling test-time computation can improve language-model reasoning, but uniform budgets waste computation on easy inputs, while verifier-guided refine

safetyarxiv-cs-cl
31 Jul 2026
Safety

TAPO: Transition-Aware Policy Optimization for LLM Agents

DGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

safetyarxiv-cs-lg
31 Jul 2026
Safety

Temporal Concentration from Rollout Errors: Implicit Preference Optimization for Text-to-Video Diffusion

DGX agent

arXiv:2607.28058v1 Announce Type: new Abstract: Recent advances in preference alignment for diffusion-based video generation, particularly via Direct Preference Optimization (DPO), have significantly

safetyarxiv-cs-cv
31 Jul 2026
Safety

The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models

DGX agent

arXiv:2602.08159v2 Announce Type: replace-cross Abstract: When a language model asserts that 'the capital of Australia is Sydney,' does it know this is wrong? Models assert misconceptions with the sam

safetyarxiv-cs-cl
31 Jul 2026
Safety

The Easy Trap: Why LLMs Underestimate Misconception-Driven Difficulty

DGX agent

arXiv:2607.26067v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for estimating item difficulty in educational assessment. However, it remains unclear whether such

safetyarxiv-cs-ai
31 Jul 2026
Safety

The Kinetics of Training: A Driven-Nucleation Rate Law for Emergence, Plasticity Loss, and Circuit Control in Language Models

DGX agent

arXiv:2607.27281v1 Announce Type: new Abstract: A capability appears in a language model when the last parts of its circuit align in one stochastic attempt, and getting all but one right is worth noth

safetyarxiv-cs-lg
31 Jul 2026
Safety

UniCross: Unified Cross-Skill Dexterous Manipulation Synthesis

DGX agent

arXiv:2607.28198v1 Announce Type: cross Abstract: Many dexterous manipulation tasks require the object to remain securely held throughout the interaction. From the perspective of hand-object relationa

safetyarxiv-cs-cv
31 Jul 2026
Safety

Unifying Adversarially Robust Model Experts in Vision-Language Models

DGX agent

arXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.

safetyarxiv-cs-cv
31 Jul 2026
Safety

VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation

DGX agent

arXiv:2607.28590v1 Announce Type: cross Abstract: Multimodal on-policy distillation (OPD) transfers fine-grained visual knowledge by supervising student-generated trajectories with a privileged-view t

safetyarxiv-cs-cl
31 Jul 2026
Safety

Variance-Aware Baselines and Adaptive Learning Rates for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2511.23310v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the de

safetyarxiv-cs-lg
31 Jul 2026
Safety

When Does Explicit View Routing Work? A Controlled Study of Multi-View Graph-Text Alignment

DGX agent

arXiv:2607.27530v1 Announce Type: new Abstract: Graph-text retrieval typically maps a graph and its description to a single embedding, even when a query concerns only one semantic aspect, such as a cl

safetyarxiv-cs-lg
31 Jul 2026
Safety

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models

DGX agent

arXiv:2607.27599v1 Announce Type: cross Abstract: Building generalizable agents for diverse applications remains a fundamental challenge. While imitation learning-based policies succeed in specific tr

safetyarxiv-cs-ro
31 Jul 2026
Safety

A Persona-based Rate Action Index

DGX agent

arXiv:2607.26545v1 Announce Type: cross Abstract: We propose an index for predicting the U.S. Federal Open Market Committee (FOMC) decision to hike/hold/cut the current federal funds target rate based

safetyarxiv-cs-lg
30 Jul 2026
Safety

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

DGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

safetyarxiv-cs-lg
30 Jul 2026
Safety

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

DGX agent

arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by

safetyarxiv-cs-cl
30 Jul 2026
Safety

Anatomy Contextualized Adaption of CT Foundation Models

DGX agent

arXiv:2607.27154v1 Announce Type: new Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume repres

safetyarxiv-cs-cv
30 Jul 2026
← Previous
1…7879808182…260
Next →