AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
Safety

Multi-channel Uplift Policy Learning

DGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

safetyarxiv-cs-lg
31 Jul 2026
Safety

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

safetyarxiv-cs-cv
31 Jul 2026
Safety

On-Policy and Off-Policy Learning for Large Action Spaces

DGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

safetyarxiv-cs-lg
31 Jul 2026
Safety

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

DGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

safetyarxiv-cs-lg
31 Jul 2026
Safety

OPLD: On-Policy Latent Distillation for Multimodal Reasoning

DGX agent

arXiv:2607.28154v1 Announce Type: new Abstract: Interleaved multimodal Chain-of-Thought (CoT) improves visual reasoning by incorporating auxiliary visual evidence into intermediate reasoning. However,

safetyarxiv-cs-cv
31 Jul 2026
Safety

Optimizing Regret

DGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

safetyarxiv-cs-lg
31 Jul 2026
Safety

Policy Gradient Steering: Interventions from Behavioral Objectives

DGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

safetyarxiv-cs-lg
31 Jul 2026
Safety

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation

DGX agent

arXiv:2506.21076v4 Announce Type: replace Abstract: Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domain

safetyarxiv-cs-cv
31 Jul 2026
Safety

Procedural Fairness in Multi-Agent Bandits

DGX agent

arXiv:2601.10600v2 Announce Type: replace-cross Abstract: In the context of multi-agent multi-armed bandits (MA-MAB), fairness is often reduced to outcomes: maximizing welfare, reducing inequality, or

safetyarxiv-cs-lg
31 Jul 2026
Safety

QQWorld: Quantile-Quantile Matching for World Model Regularization

DGX agent

arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically

safetyarxiv-cs-cv
31 Jul 2026
Safety

Recognition and Label-Free Adaptation Across Recording Sessions in Surface-EMG Gesture Decoding

DGX agent

arXiv:2607.27568v1 Announce Type: new Abstract: Recognition accuracy obtained during a recording session does not persist when a user puts on the electrodes again after the electrodes had previously b

safetyarxiv-cs-lg
31 Jul 2026
Safety

ReDiPPO: Reference-Guided Value Calibration and Discrepancy-Aware Token Reweighting for Mathematical Reasoning

DGX agent

arXiv:2607.27631v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for enhancing the mathematical reasoning capabilities of large language models. Among exis

safetyarxiv-cs-cl
31 Jul 2026
Safety

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

DGX agent

arXiv:2603.13707v3 Announce Type: replace-cross Abstract: Humanoid loco-manipulation requires coordinated task-space motion planning with stable loco-manipulation command tracking under complex robot-

safetyarxiv-cs-lg
31 Jul 2026
Safety

Regularizing modality contribution drift in multimodal continual learning

DGX agent

arXiv:2607.27260v1 Announce Type: new Abstract: Multimodal continual learning (MMCL) aims to learn emerging knowledge from multimodal data while preserving knowledge. To mitigate forgetting, current M

safetyarxiv-cs-lg
31 Jul 2026
Safety

Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review

DGX agent

arXiv:2607.27209v1 Announce Type: cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to

safetyarxiv-cs-lg
31 Jul 2026
Safety

ROAD: Reciprocal-Objective Alignment of Discriminative Semantics for 3D Shape Generation

DGX agent

arXiv:2607.28581v1 Announce Type: new Abstract: High-fidelity 3D generation predominantly relies on scaling model capacity and data, which incurs prohibitive computational costs. This paradigm typical

safetyarxiv-cs-cv
31 Jul 2026
Safety

Robust Estimation of Sparse Numerical Vectors under Local Differential Privacy

DGX agent

arXiv:2607.27815v1 Announce Type: cross Abstract: Local differential privacy (LDP) protocols are vulnerable to poisoning attacks. Existing research have proposed efficient defense strategies for singl

safetyarxiv-cs-lg
31 Jul 2026
Safety

SCOPE: Supply-Chain Operations through Coupled Policies for End-to-End Coordination

DGX agent

arXiv:2607.28488v1 Announce Type: cross Abstract: Can supply-chain AI move beyond isolated decision modules toward unified operational planning? A complete replenishment plan specifies which products

safetyarxiv-cs-lg
31 Jul 2026
Safety

ServerlessT2I: Efficient Text-to-Image Workflow Serving on a Serverless Platform

DGX agent

arXiv:2607.26566v1 Announce Type: cross Abstract: Text-to-image (T2I) workflows are increasingly deployed on serverless platforms because users often compose customized workflows and invoke them inter

safetyarxiv-cs-ai
31 Jul 2026
Safety

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

DGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

safetyarxiv-cs-ai
31 Jul 2026
Safety

SkillSight: Calibrating Generic Content Bias for Skill Retrieval

DGX agent

arXiv:2607.18785v2 Announce Type: replace Abstract: As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability

safetyarxiv-cs-ai
31 Jul 2026
Safety

Static In, Dynamic Out: Counterfactual Action Augmentation for Moving Object Manipulation

DGX agent

arXiv:2607.27890v1 Announce Type: new Abstract: Visuomotor policies have advanced on manipulation tasks where the target object stays static during execution, but real deployments break this assumptio

safetyarxiv-cs-ro
31 Jul 2026
Safety

Strategies for Milestone-driven Start-ups in Multi-activity Settings

DGX agent

arXiv:2607.27563v1 Announce Type: new Abstract: New venture start-ups need to ``survive'' through multiple stages of reaching milestone targets. We investigate the strategies for start-ups in a milest

safetyarxiv-cs-lg
31 Jul 2026
Safety

SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute

DGX agent

arXiv:2607.28457v1 Announce Type: cross Abstract: Scaling test-time computation can improve language-model reasoning, but uniform budgets waste computation on easy inputs, while verifier-guided refine

safetyarxiv-cs-cl
31 Jul 2026
Safety

TAPO: Transition-Aware Policy Optimization for LLM Agents

DGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

safetyarxiv-cs-lg
31 Jul 2026
Safety

Temporal Concentration from Rollout Errors: Implicit Preference Optimization for Text-to-Video Diffusion

DGX agent

arXiv:2607.28058v1 Announce Type: new Abstract: Recent advances in preference alignment for diffusion-based video generation, particularly via Direct Preference Optimization (DPO), have significantly

safetyarxiv-cs-cv
31 Jul 2026
Safety

The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models

DGX agent

arXiv:2602.08159v2 Announce Type: replace-cross Abstract: When a language model asserts that 'the capital of Australia is Sydney,' does it know this is wrong? Models assert misconceptions with the sam

safetyarxiv-cs-cl
31 Jul 2026
Safety

The Easy Trap: Why LLMs Underestimate Misconception-Driven Difficulty

DGX agent

arXiv:2607.26067v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for estimating item difficulty in educational assessment. However, it remains unclear whether such

safetyarxiv-cs-ai
31 Jul 2026
Safety

The Kinetics of Training: A Driven-Nucleation Rate Law for Emergence, Plasticity Loss, and Circuit Control in Language Models

DGX agent

arXiv:2607.27281v1 Announce Type: new Abstract: A capability appears in a language model when the last parts of its circuit align in one stochastic attempt, and getting all but one right is worth noth

safetyarxiv-cs-lg
31 Jul 2026
Safety

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alte…

DGX agent

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alternatives that can, or we are screwed. Anybody remember this

safetygary-marcus--x
31 Jul 2026
Safety

UniCross: Unified Cross-Skill Dexterous Manipulation Synthesis

DGX agent

arXiv:2607.28198v1 Announce Type: cross Abstract: Many dexterous manipulation tasks require the object to remain securely held throughout the interaction. From the perspective of hand-object relationa

safetyarxiv-cs-cv
31 Jul 2026
Safety

Unifying Adversarially Robust Model Experts in Vision-Language Models

DGX agent

arXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.

safetyarxiv-cs-cv
31 Jul 2026
Safety

VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation

DGX agent

arXiv:2607.28590v1 Announce Type: cross Abstract: Multimodal on-policy distillation (OPD) transfers fine-grained visual knowledge by supervising student-generated trajectories with a privileged-view t

safetyarxiv-cs-cl
31 Jul 2026
Safety

Variance-Aware Baselines and Adaptive Learning Rates for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2511.23310v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the de

safetyarxiv-cs-lg
31 Jul 2026
Safety

When Does Explicit View Routing Work? A Controlled Study of Multi-View Graph-Text Alignment

DGX agent

arXiv:2607.27530v1 Announce Type: new Abstract: Graph-text retrieval typically maps a graph and its description to a single embedding, even when a query concerns only one semantic aspect, such as a cl

safetyarxiv-cs-lg
31 Jul 2026
Safety

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models

DGX agent

arXiv:2607.27599v1 Announce Type: cross Abstract: Building generalizable agents for diverse applications remains a fundamental challenge. While imitation learning-based policies succeed in specific tr

safetyarxiv-cs-ro
31 Jul 2026
Safety

A Persona-based Rate Action Index

DGX agent

arXiv:2607.26545v1 Announce Type: cross Abstract: We propose an index for predicting the U.S. Federal Open Market Committee (FOMC) decision to hike/hold/cut the current federal funds target rate based

safetyarxiv-cs-lg
30 Jul 2026
Safety

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

DGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

safetyarxiv-cs-lg
30 Jul 2026
Safety

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

DGX agent

arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by

safetyarxiv-cs-cl
30 Jul 2026
Safety

AlloyDB adds group authentication to secure enterprise scale and AI agents

DGX agent

Database security traditionally relies on a fragile balance between the granular control developers need and the administrative overhead of managing thousands of individual database passwords. Between

safetygoogle-cloud-ai
30 Jul 2026
Safety

Anatomy Contextualized Adaption of CT Foundation Models

DGX agent

arXiv:2607.27154v1 Announce Type: new Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume repres

safetyarxiv-cs-cv
30 Jul 2026
Safety

Anchoring and Steering Diffusion: Enhancing the Faithfulness of Text-to-Image Generation at Inference Time

DGX agent

arXiv:2607.26647v1 Announce Type: new Abstract: While text-to-image diffusion models achieve impressive visual quality, they frequently struggle to maintain precise alignment with complex compositiona

safetyarxiv-cs-cv
30 Jul 2026
Safety

CASIAL: Geometric Distortion Robust Image Watermarking

DGX agent

arXiv:2607.26729v1 Announce Type: new Abstract: Deep learning-based watermarking has shown strong robustness against non-geometric distortions, yet its performance under geometric transformations rema

safetyarxiv-cs-cv
30 Jul 2026
Safety

CG-World: A Large-Scale World-State Dataset and Protocol for World Models

DGX agent

arXiv:2607.26452v1 Announce Type: cross Abstract: World models must learn the joint dynamics of states, actions, events, and observations, yet existing video, robotics, and simulation datasets usually

safetyarxiv-cs-cv
30 Jul 2026
Safety

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation

DGX agent

arXiv:2607.26789v1 Announce Type: new Abstract: Vision-language-action (VLA) policies commonly execute long-horizon mobile manipulation through open-loop action chunks, issuing multiple actions withou

safetyarxiv-cs-ro
30 Jul 2026
Model Releases

Choosing Where and How to Moderate: End-to-End Trade-offs in Filter Placement and Response Rewriting

DGX agent

arXiv:2607.26200v1 Announce Type: new Abstract: Content-moderation classifiers are usually evaluated in isolation, but deployment requires choosing where to intervene and what follows a flag. We evalu

model-releasesarxiv-cs-cl
30 Jul 2026
Safety

CineWeaver: Training-Free Reference-Controllable Multi-Shot Long Video Generation for Cinematic Storytelling

DGX agent

arXiv:2607.26529v1 Announce Type: new Abstract: Cinematic video generation is challenging for text-to-video diffusion models due to concurrent requirements on multi-shot generation, fine-grained contr

safetyarxiv-cs-cv
30 Jul 2026
Safety

Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning

DGX agent

arXiv:2607.26509v1 Announce Type: new Abstract: Deep off-policy reinforcement learning algorithms for continuous control typically rely on neural value function approximation to guide policy improveme

safetyarxiv-cs-lg
30 Jul 2026
← Previous
1…8788899091…302
Next →