AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
2 Jun 2026

From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

SafetyDGX agent

arXiv:2606.00083v1 Announce Type: cross Abstract: Reinforcement learning relies on accurate reward functions, which are often hand-crafted or even unavailable in real-world applications, such as robot

From Graph Retrieval to Schema Realization: Counterfactual Validation for Text-to-SPARQL over Heterogeneous Knowledge Graphs

SafetyDGX agent

arXiv:2508.01815v2 Announce Type: replace-cross Abstract: Text-to-SPARQL maps natural-language questions to executable SPARQL queries over RDF knowledge graphs. While standard evaluations often fix th

From Noise to Control: Parameterized Diffusion Policies

SafetyDGX agent

arXiv:2606.00336v1 Announce Type: new Abstract: We propose Parameterized Diffusion Policy (PDP), a framework for learning diffusion policies conditioned on low-dimensional, continuous parameters embed


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From 'Weak' Signals to Strong Models: Preference Delta Aggregation with LoRA Merging

SafetyDGX agent

arXiv:2606.00357v1 Announce Type: new Abstract: Training strong large language models (LLMs) requires high-quality supervision, which is often scarce. Recent work shows that paired preference data fro

FROST-STA: Frozen Dense Features for the Ego4D Short-Term Object Interaction Anticipation

SafetyDGX agent

arXiv:2606.00694v1 Announce Type: new Abstract: Short-term anticipation in egocentric video requires more than recognizing the current scene: a system must infer which object the camera wearer will co

GFlowGR: Fine-tuning Generative Recommendation Frameworks with Generative Flow Networks

SafetyDGX agent

arXiv:2506.16114v3 Announce Type: replace-cross Abstract: Generative recommendations (GR), which usually include item tokenizers and generative Large Language Models (LLMs), have demonstrated remarkab

Global-Local Attention Decomposition for Terrain Encoding in Humanoid Perceptive Locomotion

SafetyDGX agent

arXiv:2606.00637v1 Announce Type: new Abstract: Although reinforcement learning has significantly advanced humanoid locomotion, perceptive policies still struggle on sparse-foothold terrain and constr

GovAI-Pipe: A Layered AI Governance Pipeline for Citizen-Facing AI in Turkey's e-Government Gateway

SafetyDGX agent

arXiv:2606.01417v1 Announce Type: new Abstract: Turkey's e-Government Gateway (e-Devlet) serves over 68 million registered users with more than 9,200 government services, and is increasingly integrati

Grounding or Guessing? Visual Signals for Detecting Hallucinations in Sign Language Translation

SafetyDGX agent

arXiv:2510.18439v3 Announce Type: replace Abstract: Hallucination, where models generate fluent text unsupported by visual evidence, remains a major flaw in vision-language models and is particularly

Hard Labels In! Rethinking the Role of Hard Labels in Mitigating Local Semantic Drift

SafetyDGX agent

arXiv:2512.15647v3 Announce Type: replace Abstract: Soft labels from teacher models are a de facto practice for knowledge transfer and large-scale dataset distillation (e.g., SRe2L, LPLD). However, wh

Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses

SafetyDGX agent

arXiv:2606.02373v1 Announce Type: new Abstract: Search agents are often trained as policies over growing transcripts: the model must decide how to search while also remembering what it has seen, which

HarnessForge: Joint Harness and Policy Evolution for Adaptive Agent Systems

SafetyDGX agent

arXiv:2606.01779v1 Announce Type: new Abstract: LLM agents are increasingly expected to operate across heterogeneous task regimes that require distinct execution paradigms. This challenges fixed agent

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

SafetyDGX agent

arXiv:2606.01117v1 Announce Type: cross Abstract: Extreme multi-label classification (XMC) involves learning models over large output spaces with millions of labels, making the output layer a memory-c

Hierarchical Object Representation for Spatial Robot Perception: Points, Meshes, and Superquadrics

SafetyDGX agent

arXiv:2606.01545v1 Announce Type: new Abstract: Hierarchical 3D Scene Graphs (3DSG) have emerged as an actionable and scalable representation for long-term autonomy incorporating metric, semantic, and

Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation

SafetyDGX agent

arXiv:2606.01565v1 Announce Type: cross Abstract: Vision-Language Navigation in Continuous Environments (VLN-CE) poses a formidable challenge for autonomous agents, requiring seamless integration of n

HiTokSR: A Coarse-to-Fine Tokenizer with Hierarchical Codebooks for High-Fidelity Real-World Image Super-Resolution

SafetyDGX agent

arXiv:2606.01157v1 Announce Type: new Abstract: Vector-quantized (VQ) generative models have shown promising results in real-world image super-resolution (Real-ISR). However, existing methods typicall

HMPO: Hybrid Median-length Policy Optimization for Chain-of-Thought Compression

SafetyDGX agent

arXiv:2606.01934v1 Announce Type: cross Abstract: Large language models achieve remarkable performance via extended chain-of-thought (CoT) reasoning, yet this lengthy process incurs substantial infere

HOIST: Humanoid Optimization with Imitation and Sample-efficient Tuning for Manipulating Suspended Loads

SafetyDGX agent

arXiv:2606.00252v1 Announce Type: cross Abstract: Manipulating suspended payloads with humanoid robots is challenging because the robot can only influence an underactuated, oscillatory load through wh

HOLA: Holistic Multi-Modal Alignment for Open-Set 3D Recognition

SafetyDGX agent

arXiv:2606.01334v1 Announce Type: new Abstract: Open-set 3D recognition requires models that generalize to rare or unseen categories. Recent approaches address this by distilling language-vision knowl

Hot-Start Chinese Language Modeling:Visual Glyphs Accelerate Sample-Efficient Learning

SafetyDGX agent

arXiv:2601.09566v4 Announce Type: replace-cross Abstract: In this work, we study whether rendering Chinese characters as visual glyph images, rather than discrete token IDs as mainstream LLMs do, prov

How Hard Can It Be? Hardness-Aware Multi-Objective Unlearning

SafetyDGX agent

arXiv:2606.02119v1 Announce Type: cross Abstract: Machine unlearning aims to remove the influence of specific forget training data due to privacy, copyright or bias concerns while maintaining the mode

How Much Progress Has There Been in NVIDIA Datacenter GPUs?

SafetyDGX agent

arXiv:2601.20115v3 Announce Type: replace-cross Abstract: As the role of modern Graphics Processing Units (GPUs) becomes increasingly essential for several computing tasks, analyzing their past and cu

Hybrid TD3: Overestimation Bias Analysis and Stable Policy Optimization for Hybrid Action Space

SafetyDGX agent

arXiv:2603.01302v2 Announce Type: replace Abstract: Reinforcement learning in discrete-continuous hybrid action spaces presents fundamental challenges for robotic manipulation, where high-level task d

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual…

SafetyDGX agent

I have a great idea. I am going to spend a trillion dollars so I can make $10 billion a year in profit, if all goes well. That’s a 1% annual return – IF it works out. Who’s in? Don’t worry about the r

Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry

SafetyDGX agent

arXiv:2606.01098v1 Announce Type: cross Abstract: Generative action policies based on diffusion or flow matching excel in behavior cloning, yet their iterative sampling is prohibitive for high-frequen

Improving Visual Representation Alignment Generation with GRPO

SafetyDGX agent

arXiv:2606.00583v1 Announce Type: cross Abstract: Recent diffusion transformers have demonstrated strong image synthesis capabilities but remain inefficient to train due to weak alignment between gene

Infeasible optimization problems and the hierarchical augmented Lagrangian method in imitation learning

SafetyDGX agent

arXiv:2606.00730v1 Announce Type: new Abstract: Imitation learning (IL) is an effective approach to train complex robotics policies. Recent works have introduced hard constraints into imitation-learni

InFerActive: Interactive Tree-Based Exploration of LLM Sampling for Safety Evaluation

SafetyDGX agent

arXiv:2512.10234v2 Announce Type: replace-cross Abstract: Even LLMs that appear safe during evaluation can still produce harmful responses in deployment. Because stochastic sampling yields different r

Interaction-Limited Safe Continuous-Time RL for Dynamical Medical Treatment

SafetyDGX agent

arXiv:2606.01051v1 Announce Type: new Abstract: Dynamic medical treatment requires deciding treatment intensity and intervention timing, while patient states evolve continuously and adverse events may

Internalize the Temperature: On-Policy Self-Distillation as Policy Reheater for Reinforcement Learning

SafetyDGX agent

arXiv:2606.00755v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards improves the reasoning ability of large language models, but often suffers from entropy collapse, in whic

Interpretability in Deep Time Series Models Demands Semantic Alignment

SafetyDGX agent

arXiv:2602.02239v2 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, exi

Interpretable Modeling of Driver Attention Shifts with a Vision--Language Model

SafetyDGX agent

arXiv:2508.05852v2 Announce Type: replace Abstract: Driver gaze is commonly modeled as a spatial heatmap, but heatmaps alone are difficult for humans to interpret because they do not explain which roa

Interpretable Multimodal Gesture Recognition for Drone and Mobile Robot Teleoperation via Log-Likelihood Ratio Fusion

SafetyDGX agent

arXiv:2602.23694v3 Announce Type: replace-cross Abstract: Human operators are still frequently exposed to hazardous environments such as disaster zones and industrial facilities, where intuitive and r

Interpretable Policy Distillation for Power Grid Topology Control

SafetyDGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

SafetyDGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation

SafetyDGX agent

arXiv:2601.00212v2 Announce Type: replace Abstract: Segmentation of vestibular schwannoma and cochlea from T2 MRI is clinically important yet annotation-intensive. Domain adaptation (DA) has been wide

Inverse Depth Scaling From Most Layers Being Similar

SafetyDGX agent

arXiv:2602.05970v2 Announce Type: replace-cross Abstract: Neural scaling laws relate loss to model size in large language models (LLMs), yet depth and width may contribute to performance differently,

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

SafetyDGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

SafetyDGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

Jailbreaking Multimodal Large Language Models using Multi-Clip Video

SafetyDGX agent

arXiv:2606.02111v1 Announce Type: cross Abstract: As multimodal large language models (MLLMs) have advanced to process video inputs, concerns have emerged about their potential for malicious misuse. P

Joint Agent Memory and Exploration Learning via Novelty Signals

SafetyDGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

Jointly Optimizing Debiased CTR and Uplift for Coupons Marketing: A Unified Causal Framework

SafetyDGX agent

arXiv:2602.12972v2 Announce Type: replace-cross Abstract: In online advertising, marketing interventions such as coupons introduce significant confounding bias into Click-Through Rate (CTR) prediction

KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation

SafetyDGX agent

arXiv:2606.01282v1 Announce Type: new Abstract: Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cu

KISS: Keeping it Simple and Slotted when Learning to Communicate over Wireless

SafetyDGX agent

arXiv:2606.00266v1 Announce Type: cross Abstract: A long-standing challenge in distributed wireless systems is ensuring efficient and fair random channel access. Existing solutions often address speci

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies

SafetyDGX agent

arXiv:2606.01151v1 Announce Type: new Abstract: Behavior cloning with high-capacity generative policies achieves strong imitation performance, but is often limited by demonstration coverage and distri

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

SafetyDGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

SafetyDGX agent

arXiv:2602.08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network. Once the denoiser is fixed, the samp

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World

SafetyDGX agent

arXiv:2606.01458v1 Announce Type: new Abstract: Training vision-language-action (VLA) policies for humanoid loco-manipulation is constrained by the high cost and complexity of collecting human teleope

Leyline: KV Cache Directives for Agentic Inference

SafetyDGX agent

arXiv:2606.01065v1 Announce Type: cross Abstract: Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only evict

LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving

SafetyDGX agent

arXiv:2606.00372v1 Announce Type: new Abstract: Reliable object detection is critical for automated driving, yet even state-of-the-art detectors inevitably make errors that can compromise safety. Intr

LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psychological Defense Mechanism Classification

SafetyDGX agent

arXiv:2606.00647v1 Announce Type: cross Abstract: Detecting psychological defense mechanisms in conversational text remains a challenging clinical NLP problem. For the PsyDefDetect 2026 shared task (n

LLM as a Meta-Judge: Synthetic Data for NLP Evaluation Metric Validation

SafetyDGX agent

arXiv:2603.09403v2 Announce Type: replace Abstract: Validating evaluation metrics for NLG typically relies on expensive and time-consuming human annotations, which predominantly exist only for English

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

SafetyDGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

Longitudinal Multimodal Sensing of Physical Activity and Well-Being in Older Adults

SafetyDGX agent

arXiv:2606.00345v1 Announce Type: new Abstract: Wearable and mobile sensing technologies enable continuous monitoring of human behavior and health in real-world settings. However, predictive modeling

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

SafetyDGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

Looped Transformers with Layer Normalization Provably Learn the Power Method

SafetyDGX agent

arXiv:2606.00605v1 Announce Type: new Abstract: Transformers have achieved remarkable success across a wide range of applications, and a growing body of work suggests that part of their strength comes

Lost in Delusion: Examining LLM Safety Under User Delusions and Distress

SafetyDGX agent

arXiv:2606.00975v1 Announce Type: new Abstract: LLM chatbots increasingly serve as a first source of support for people in psychological distress, including those whose distress is entangled with delu

Low-Pass Flow Matching

SafetyDGX agent

arXiv:2606.02177v1 Announce Type: new Abstract: Flow Matching typically relies on white noise sources, a choice often misaligned with the power spectra of natural data, which tend to decay with freque

← Previous
1…979899100101…214
Next →