AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
31 Jul 2026

FiRE: Enhancing MLLMs with Fine-Grained Context Learning for Complex Image Retrieval

SafetyDGX agent

arXiv:2607.27959v1 Announce Type: new Abstract: Due to their strong generalizable multimodal processing and reasoning capabilities, Multimodal Large Language Models (MLLMs) have demonstrated significa

Flux-OPD: On-Policy Distillation with Evolving Contexts

SafetyDGX agent

arXiv:2607.28022v1 Announce Type: new Abstract: Large language model training in open-ended domains lacks verifiable rewards, making task preferences difficult to formalize as effective supervision. C

Foundation-Model Earth Representations Enable Regional-Scale Forest Aboveground Biomass Monitoring Across the Northeastern United States

SafetyDGX agent

arXiv:2607.27217v1 Announce Type: cross Abstract: Forest aboveground biomass (AGB) is a critical indicator of ecosystem productivity and terrestrial carbon storage, yet regional carbon monitoring rema

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Graph Is the Verifier: Agentic Reinforcement Learning for Interprocedural Vulnerability Detection

SafetyDGX agent

arXiv:2607.26656v1 Announce Type: cross Abstract: Real-world vulnerabilities often span multiple functions, yet most learning-based detectors classify each function in isolation: on a sample of real C

Group-Reflective Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.28076v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is effective for training large language model agents. However, terminal rewards provide only co

GuidedRAG: Semantic Steering of Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2607.26071v1 Announce Type: cross Abstract: In this work, we propose GuidedRAG, a novel extension to traditional Retrieval-Augmented Generation (RAG) that introduces a dedicated selection stage

HALO: Heterogeneous Admission through Localized Obligations for Safe Agentic Execution

SafetyDGX agent

arXiv:2607.27636v1 Announce Type: cross Abstract: Recent agentic AI systems may return a heterogeneous response containing notices, requests, handoffs, and actions. Conditions can change before extern

Harness-G: A Graph-Structured Harness for Search Agents

SafetyDGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

Harnessing the Potential of Optimizing Data Mixtures via Bayesian Domain Reweighting

SafetyDGX agent

arXiv:2607.27928v1 Announce Type: new Abstract: The performance of Large Language Models (LLMs) is fundamentally influenced by the distributional composition of multi-domain pre-training data. While m

Heterogeneous Ranking in Industrial-Scale Recommender Systems: A Case Study

SafetyDGX agent

arXiv:2607.27577v1 Announce Type: cross Abstract: Heterogeneous recommendation feeds present complex challenges that extend beyond those found in highly homogeneous environments (e.g., music-only or v

HSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs

SafetyDGX agent

arXiv:2607.27379v1 Announce Type: new Abstract: High-quality, diverse data are vital for large language models (LLMs) but remain scarce and costly. Data synthesis is a viable alternative and succeeds

Huawei opensouced openPangu-2.0-Pro, 505B-A18B

SafetyDGX agent

openPangu-2.0-Pro is an MoE model trained on Ascend. The model has 505B total parameters and 18B activated parameters. Its context length is 512k. The total pretraining data contains 34T tokens. Durin

Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning

SafetyDGX agent

arXiv:2607.27564v1 Announce Type: new Abstract: Multi-image medical VQA is not merely a prompt-length problem; it is a fundamental challenge of agentic decision-making. Medical vision-language agents

Integrating Contextual Embeddings into Evaluation of Expressive MIDI Piano Performances

SafetyDGX agent

arXiv:2607.27909v1 Announce Type: cross Abstract: Objective evaluation of expressive MIDI piano performances typically relies on attribute statistics such as timing, velocity, and duration of individu

It's Not Just More Demos: Counterfactual Action Sensitivity Coverage for Data-Efficient Robust Robot Imitation

SafetyDGX agent

arXiv:2607.27261v1 Announce Type: new Abstract: Visuomotor imitation learning has demonstrated success for manipulation tasks. However, the trained policies remain brittle to visual `nuisances', with

Learning Social Robot Navigation By Sensing Human Legs

SafetyDGX agent

arXiv:2607.27922v1 Announce Type: new Abstract: Robots navigating among pedestrians typically sense their surroundings with a 2D LiDAR mounted close to the ground. At that height, the sensor mostly se

Massive Update to my Krea 2 Multi-Lora Bounding Box workflow, now bounding boxes control placement with better accuracy. Also introduced Edit features like Scene and Outfit transfer, put multiple character loras in a scene or outfit of your choosing! Token drift also fixed by facial detailer stage

SafetyDGX agent

Krea 2 has been my favorite base model for character work, but the moment you put two character LoRAs in the same generation they smear into one blended face. Attention bias, prompt engineering, and C

MedXplore: Towards Reliable and Unbiased Generalized Category Discovery in Medical Imaging

SafetyDGX agent

arXiv:2607.27620v1 Announce Type: new Abstract: Deep learning has shown strong potential in medical image analysis, but most existing methods rely on large-scale annotations and a closed-world assumpt

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

SafetyDGX agent

arXiv:2607.28565v1 Announce Type: new Abstract: Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typical

MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation

SafetyDGX agent

arXiv:2607.26698v1 Announce Type: cross Abstract: Cover song generation (CSG) should preserve the melodic and linguistic content of a reference song while recreating the remaining musical components.

MUGEN: A Unified Framework for Efficient Motion Understanding and Generation

SafetyDGX agent

arXiv:2607.27581v1 Announce Type: new Abstract: Grounding human motion in language, and language in motion, is a central step toward physical AI systems that can understand, generate, and communicate

Multi-channel Uplift Policy Learning

SafetyDGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

SafetyDGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

On-Policy and Off-Policy Learning for Large Action Spaces

SafetyDGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

SafetyDGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

OPLD: On-Policy Latent Distillation for Multimodal Reasoning

SafetyDGX agent

arXiv:2607.28154v1 Announce Type: new Abstract: Interleaved multimodal Chain-of-Thought (CoT) improves visual reasoning by incorporating auxiliary visual evidence into intermediate reasoning. However,

Optimizing Regret

SafetyDGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

Policy Gradient Steering: Interventions from Behavioral Objectives

SafetyDGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation

SafetyDGX agent

arXiv:2506.21076v4 Announce Type: replace Abstract: Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domain

Procedural Fairness in Multi-Agent Bandits

SafetyDGX agent

arXiv:2601.10600v2 Announce Type: replace-cross Abstract: In the context of multi-agent multi-armed bandits (MA-MAB), fairness is often reduced to outcomes: maximizing welfare, reducing inequality, or

QQWorld: Quantile-Quantile Matching for World Model Regularization

SafetyDGX agent

arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically

Recognition and Label-Free Adaptation Across Recording Sessions in Surface-EMG Gesture Decoding

SafetyDGX agent

arXiv:2607.27568v1 Announce Type: new Abstract: Recognition accuracy obtained during a recording session does not persist when a user puts on the electrodes again after the electrodes had previously b

ReDiPPO: Reference-Guided Value Calibration and Discrepancy-Aware Token Reweighting for Mathematical Reasoning

SafetyDGX agent

arXiv:2607.27631v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for enhancing the mathematical reasoning capabilities of large language models. Among exis

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

SafetyDGX agent

arXiv:2603.13707v3 Announce Type: replace-cross Abstract: Humanoid loco-manipulation requires coordinated task-space motion planning with stable loco-manipulation command tracking under complex robot-

Regularizing modality contribution drift in multimodal continual learning

SafetyDGX agent

arXiv:2607.27260v1 Announce Type: new Abstract: Multimodal continual learning (MMCL) aims to learn emerging knowledge from multimodal data while preserving knowledge. To mitigate forgetting, current M

Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review

SafetyDGX agent

arXiv:2607.27209v1 Announce Type: cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to

ROAD: Reciprocal-Objective Alignment of Discriminative Semantics for 3D Shape Generation

SafetyDGX agent

arXiv:2607.28581v1 Announce Type: new Abstract: High-fidelity 3D generation predominantly relies on scaling model capacity and data, which incurs prohibitive computational costs. This paradigm typical

Robust Estimation of Sparse Numerical Vectors under Local Differential Privacy

SafetyDGX agent

arXiv:2607.27815v1 Announce Type: cross Abstract: Local differential privacy (LDP) protocols are vulnerable to poisoning attacks. Existing research have proposed efficient defense strategies for singl

SCOPE: Supply-Chain Operations through Coupled Policies for End-to-End Coordination

SafetyDGX agent

arXiv:2607.28488v1 Announce Type: cross Abstract: Can supply-chain AI move beyond isolated decision modules toward unified operational planning? A complete replenishment plan specifies which products

ServerlessT2I: Efficient Text-to-Image Workflow Serving on a Serverless Platform

SafetyDGX agent

arXiv:2607.26566v1 Announce Type: cross Abstract: Text-to-image (T2I) workflows are increasingly deployed on serverless platforms because users often compose customized workflows and invoke them inter

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

SafetyDGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

SkillSight: Calibrating Generic Content Bias for Skill Retrieval

SafetyDGX agent

arXiv:2607.18785v2 Announce Type: replace Abstract: As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability

Static In, Dynamic Out: Counterfactual Action Augmentation for Moving Object Manipulation

SafetyDGX agent

arXiv:2607.27890v1 Announce Type: new Abstract: Visuomotor policies have advanced on manipulation tasks where the target object stays static during execution, but real deployments break this assumptio

Strategies for Milestone-driven Start-ups in Multi-activity Settings

SafetyDGX agent

arXiv:2607.27563v1 Announce Type: new Abstract: New venture start-ups need to ``survive'' through multiple stages of reaching milestone targets. We investigate the strategies for start-ups in a milest

SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute

SafetyDGX agent

arXiv:2607.28457v1 Announce Type: cross Abstract: Scaling test-time computation can improve language-model reasoning, but uniform budgets waste computation on easy inputs, while verifier-guided refine

TAPO: Transition-Aware Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

Temporal Concentration from Rollout Errors: Implicit Preference Optimization for Text-to-Video Diffusion

SafetyDGX agent

arXiv:2607.28058v1 Announce Type: new Abstract: Recent advances in preference alignment for diffusion-based video generation, particularly via Direct Preference Optimization (DPO), have significantly

The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models

SafetyDGX agent

arXiv:2602.08159v2 Announce Type: replace-cross Abstract: When a language model asserts that 'the capital of Australia is Sydney,' does it know this is wrong? Models assert misconceptions with the sam

The Easy Trap: Why LLMs Underestimate Misconception-Driven Difficulty

SafetyDGX agent

arXiv:2607.26067v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for estimating item difficulty in educational assessment. However, it remains unclear whether such

The Kinetics of Training: A Driven-Nucleation Rate Law for Emergence, Plasticity Loss, and Circuit Control in Language Models

SafetyDGX agent

arXiv:2607.27281v1 Announce Type: new Abstract: A capability appears in a language model when the last parts of its circuit align in one stochastic attempt, and getting all but one right is worth noth

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alte…

SafetyDGX agent

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alternatives that can, or we are screwed. Anybody remember this

UniCross: Unified Cross-Skill Dexterous Manipulation Synthesis

SafetyDGX agent

arXiv:2607.28198v1 Announce Type: cross Abstract: Many dexterous manipulation tasks require the object to remain securely held throughout the interaction. From the perspective of hand-object relationa

Unifying Adversarially Robust Model Experts in Vision-Language Models

SafetyDGX agent

arXiv:2607.27897v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment.

VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation

SafetyDGX agent

arXiv:2607.28590v1 Announce Type: cross Abstract: Multimodal on-policy distillation (OPD) transfers fine-grained visual knowledge by supervising student-generated trajectories with a privileged-view t

Variance-Aware Baselines and Adaptive Learning Rates for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2511.23310v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the de

When Does Explicit View Routing Work? A Controlled Study of Multi-View Graph-Text Alignment

SafetyDGX agent

arXiv:2607.27530v1 Announce Type: new Abstract: Graph-text retrieval typically maps a graph and its description to a single embedding, even when a query concerns only one semantic aspect, such as a cl

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models

SafetyDGX agent

arXiv:2607.27599v1 Announce Type: cross Abstract: Building generalizable agents for diverse applications remains a fundamental challenge. While imitation learning-based policies succeed in specific tr

30 Jul 2026

A Persona-based Rate Action Index

SafetyDGX agent

arXiv:2607.26545v1 Announce Type: cross Abstract: We propose an index for predicting the U.S. Federal Open Market Committee (FOMC) decision to hike/hold/cut the current federal funds target rate based

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

SafetyDGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

SafetyDGX agent

arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by

← Previous
1…6970717273…242
Next →