AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos?

DGX agent

arXiv:2606.06627v1 Announce Type: cross Abstract: Human video datasets used for cotraining robot manipulation policies largely consist of curated demonstrations where motions are orchestrated to resem

safetyarxiv-cs-ai
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Where to Touch, How to Contact: Hierarchical RL-MPC Framework for Geometry-Aware Long-Horizon Dexterous Manipulation

DGX agent

arXiv:2601.10930v3 Announce Type: replace Abstract: A key challenge in contact-rich dexterous manipulation is the need to jointly reason over global geometry and nonsmooth contact dynamics. End-to-end

safetyarxiv-cs-ro
8 Jun 2026
Safety

A Pre-Registered Causal Partition of Self-Consistency Elicitation and Reward Design in RLVR

DGX agent

arXiv:2606.05932v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) improves reasoning even when the reward signal is spurious -- assigning credit to the group-plural

safetyarxiv-cs-ai
6 Jun 2026
Safety

AdaMEM: Test-Time Adaptive Memory for Language Agents

DGX agent

arXiv:2606.05684v1 Announce Type: new Abstract: A central challenge for language agents is utilizing past experience to adapt to dynamic test-time conditions. While recent work demonstrates the promis

safetyarxiv-cs-ai
6 Jun 2026
Safety

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models

DGX agent

arXiv:2606.06154v1 Announce Type: new Abstract: Federated fine-tuning of foundation models using Low-Rank Adaptation (LoRA) offers a communication efficient solution for distributed learning. However,

safetyarxiv-cs-ai
6 Jun 2026
Safety

An Infectious Disease Spread Simulation Based on Large Language Model Decision Making

DGX agent

arXiv:2606.06360v1 Announce Type: new Abstract: Modelling individual decision-making during infectious disease outbreaks is crucial for understanding behavioural dynamics and informing effective publi

safetyarxiv-cs-ai
6 Jun 2026
Safety

Assessing the Geographic Diversity of AI's Platial Representations in Image Generation

DGX agent

arXiv:2606.05188v1 Announce Type: cross Abstract: (Gen)AI diversity is not merely an ethical issue. From the perspective of geographic information science (GIScience), it could be interpreted as a fun

safetyarxiv-cs-ai
6 Jun 2026
Safety

Beyond Rewards in Reinforcement Learning for Cyber Defence

DGX agent

arXiv:2602.04809v3 Announce Type: replace-cross Abstract: Recent years have seen an explosion of interest in autonomous cyber defence agents trained to defend computer networks using deep reinforcemen

safetyarxiv-cs-ai
6 Jun 2026
Safety

Bridging Domain Expertise and Generalization for Performance Estimation

DGX agent

arXiv:2606.06335v1 Announce Type: cross Abstract: Performance estimation under distribution shift aims to predict how a model behaves on an unlabeled test set whose distribution differs from the train

safetyarxiv-cs-ai
6 Jun 2026
Safety

Class-Specific Branch Attention for Mitigating Gradient Interference under Class Imbalance

DGX agent

arXiv:2606.05740v1 Announce Type: new Abstract: Deep neural networks trained under severe class imbalance often exhibit degraded performance, typically attributed to statistical bias. In this work, we

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

CogManip: Benchmarking Manipulative Behavior in Multi-Turn Interactions with Large Language Model

DGX agent

arXiv:2606.06099v1 Announce Type: new Abstract: Whether Large Language Models (LLMs) exhibit covert psychological manipulation in complex human-AI interactions has garnered increasing safety concerns.

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Frequency-Domain Insights

DGX agent

arXiv:2411.18343v3 Announce Type: replace-cross Abstract: Personalized Federal learning(PFL) allows clients to cooperatively train a personalized model without disclosing their private dataset. Howeve

safetyarxiv-cs-ai
6 Jun 2026
Safety

Differentiable Efficient Operator Search

DGX agent

arXiv:2606.05232v1 Announce Type: cross Abstract: Efficient multimodal foundation models often rely on manually designed token-reduction operators, such as pruning, merging, pooling, and adaptive rewe

safetyarxiv-cs-ai
6 Jun 2026
Safety

Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss

DGX agent

arXiv:2606.06418v1 Announce Type: cross Abstract: Many modern applications of deep learning involve training a neural network via a one-step prediction loss (e.g., L^2 regression, cross-entropy), but

safetyarxiv-cs-ai
6 Jun 2026
Safety

Escaping the Verifier: Learning to Reason via Demonstrations

DGX agent

arXiv:2511.21667v4 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) to reason often relies on Reinforcement Learning (RL) with task-specific verifiers. However, many real-w

safetyarxiv-cs-ai
6 Jun 2026
Safety

FIDES: Faithful Inference via Deep Evidence Signals for Retrieval-Memory Conflict in RAG

DGX agent

arXiv:2606.05644v1 Announce Type: new Abstract: When retrieved evidence contradicts parametric memory, language models frequently ignore context and default to memorized priors -- a failure that under

safetyarxiv-cs-ai
6 Jun 2026
Safety

GIPO: Gaussian Importance Sampling Policy Optimization

DGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

safetyarxiv-cs-ai
6 Jun 2026
Safety

HypRAG: Hyperbolic Dense Retrieval for Retrieval Augmented Generation

DGX agent

arXiv:2602.07739v2 Announce Type: replace-cross Abstract: Embedding geometry plays a fundamental role in retrieval quality, yet dense retrievers for retrieval-augmented generation (RAG) remain largely

safetyarxiv-cs-ai
6 Jun 2026
Safety

LatentWave: JEPA Pretraining for Wireless Foundation Models

DGX agent

arXiv:2606.06373v1 Announce Type: cross Abstract: Wireless foundation models have emerged as a promising alternative to building separate models for each wireless task. However, existing approaches re

safetyarxiv-cs-ai
6 Jun 2026
Safety

Learning to replenish: A hybrid deep reinforcement learning for dynamic inventory management in the pharmaceutical supply chains

DGX agent

arXiv:2606.06201v1 Announce Type: new Abstract: Pharmaceutical supply chains (PSCs) struggle with inventory management (IM) due to unpredictable demand patterns and variable lead times associated with

safetyarxiv-cs-ai
6 Jun 2026
Safety

Mutation Without Variation: Convergence Dynamics in LLM-Driven Program Evolution

DGX agent

arXiv:2606.05408v1 Announce Type: new Abstract: When an LLM repeatedly mutates a program, does it explore new forms or circle back to the same ones? We study this question by analyzing LLM-driven muta

safetyarxiv-cs-ai
6 Jun 2026
Safety

Policy-Conditioned Counterfactual Credit for Verifiable Reinforcement Learning of Long-Horizon Language Agents

DGX agent

arXiv:2606.05263v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards improves reasoning and tool use, yet long-horizon language agents still learn unsupported evidence chai

safetyarxiv-cs-ai
6 Jun 2026
Safety

Regret Minimization with Adaptive Opponents in Repeated Games

DGX agent

arXiv:2606.06486v1 Announce Type: cross Abstract: In this paper, we study regret minimization in repeated games with adaptive opponents who can respond based on histories of play. The standard metric

safetyarxiv-cs-ai
6 Jun 2026
Safety

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

DGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

safetyarxiv-cs-ai
6 Jun 2026
Safety

RREDCoT: Segment-Level Reward Redistribution for Reasoning Models

DGX agent

arXiv:2606.06475v1 Announce Type: cross Abstract: Recent advancements in reasoning language models have been driven by Reinforcement Learning (RL) fine-tuning. Most often, these rely on the Group Rela

safetyarxiv-cs-ai
6 Jun 2026
Safety

SAGE: Scalable AI Governance & Evaluation

DGX agent

arXiv:2602.07840v3 Announce Type: replace-cross Abstract: Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained hu

safetyarxiv-cs-ai
6 Jun 2026
Safety

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

DGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

safetyarxiv-cs-ai
6 Jun 2026
Safety

Soft Sequence Policy Optimization

DGX agent

arXiv:2602.19327v3 Announce Type: replace-cross Abstract: A significant portion of recent research on Large Language Model (LLM) alignment focuses on developing new policy optimization methods based o

safetyarxiv-cs-ai
6 Jun 2026
Safety

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

DGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

safetyarxiv-cs-ai
6 Jun 2026
Safety

UniVoice: A Unified Model for Speech and Singing Voice Generation

DGX agent

arXiv:2606.05852v1 Announce Type: cross Abstract: Text-to-speech (TTS) and singing voice synthesis (SVS) both aim to generate human vocal audio from symbolic inputs, but they impose different requirem

safetyarxiv-cs-ai
6 Jun 2026
Safety

Your GFlowNet Secretly Learns an Optimal Transport Plan

DGX agent

arXiv:2606.06272v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are a framework for sampling structured objects via stochastic trajectories in a directed graph. In this work, we

safetyarxiv-cs-ai
6 Jun 2026
Safety

Zero knowledge verification for frontier AI training is possible

DGX agent

arXiv:2606.05433v1 Announce Type: new Abstract: Frontier AI governance frameworks increasingly use cumulative training compute as the primary criterion for designating high-impact models, but enforcem

safetyarxiv-cs-ai
6 Jun 2026
Safety

A Komi-Yazva--Russian Parallel Corpus and Evaluation Protocol for Zero- and Few-Shot LLM Translation

DGX agent

arXiv:2606.06420v1 Announce Type: new Abstract: We present the first Komi-Yazva--Russian parallel corpus together with an explicit evaluation protocol for studying LLM translation in an endangered, ex

safetyarxiv-cs-cl
5 Jun 2026
Safety

A Systematic Analysis of Biases in Large Language Models

DGX agent

arXiv:2512.15792v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making. However,

safetyarxiv-cs-cl
5 Jun 2026
Safety

ACE-SQL: Adaptive Co-Optimization via Empirical Credit Assignment for Text-to-SQL

DGX agent

arXiv:2606.05906v1 Announce Type: new Abstract: Text-to-SQL maps natural language questions to executable SQL queries. Modern databases often contain large and complex schemas, making schema linking a

safetyarxiv-cs-cl
5 Jun 2026
Safety

Adversarial Attacks Already Tell the Answer: Directional Bias-Guided Test-time Defense for Vision-Language Models

DGX agent

arXiv:2606.06186v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have shown strong zero-shot generalization but remain highly vulnerable to adversarial perturbations, posin

safetyarxiv-cs-cv
5 Jun 2026
Safety

Analysis of the Neglect-Zero Effect in Large Language Models

DGX agent

arXiv:2606.05864v1 Announce Type: new Abstract: We investigate the extent to which the language processing of LLMs resembles human cognitive processes, focusing on a human cognitive bias called the ex

safetyarxiv-cs-cl
5 Jun 2026
Safety

Attitude-Aided Linear Calibration of Triaxial Accelerometers

DGX agent

arXiv:2606.06308v1 Announce Type: new Abstract: Triaxial MEMS accelerometers are widely used for inertial sensing, navigation, and sensor fusion, but existing calibration methods often rely on costly

safetyarxiv-cs-ro
5 Jun 2026
Safety

Auditing Demonstration Curation Metrics: Action-Only Scorers Fail on the Structural Defects That Degrade Imitation Policies

DGX agent

arXiv:2606.05588v1 Announce Type: new Abstract: Imitation-learning policies inherit the quality of the demonstrations they are trained on, and a growing set of curation metrics promise to score and fi

safetyarxiv-cs-ro
5 Jun 2026
Safety

Beyond Alignment: Value Diversity as a Collective Property in Multicultural Agent Systems

DGX agent

arXiv:2606.05985v1 Announce Type: new Abstract: Multicultural multi-agent systems are increasingly deployed in globally diverse settings, where different agents are grounded in different cultural back

safetyarxiv-cs-cl
5 Jun 2026
Safety

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

DGX agent

arXiv:2602.12628v4 Announce Type: replace Abstract: Simulation offers a scalable and low-cost way to enrich vision-language-action (VLA) training, reducing reliance on expensive real-robot demonstrati

safetyarxiv-cs-ro
5 Jun 2026
Safety

Beyond tokens: a unified framework for latent communication in LLM-based multi-agent systems

DGX agent

arXiv:2606.05711v1 Announce Type: new Abstract: Multi-agent systems built on large language models (LLMs) have become a prevailing paradigm for tackling complex reasoning, planning, and tool-use tasks

safetyarxiv-cs-cl
5 Jun 2026
Safety

'Chi nas dal soch el sent de legn' -- Auditing Text Corpora for Lombard

DGX agent

arXiv:2606.06349v1 Announce Type: new Abstract: Several of the world's languages are still under-resourced in terms of Natural Language Processing (NLP) tools. This is mostly due to the lack of high-q

safetyarxiv-cs-cl
5 Jun 2026
Safety

Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents

DGX agent

arXiv:2606.05753v1 Announce Type: new Abstract: Latent visual reasoning (LVR) inserts supervised latent tokens between perception and answer generation in vision-language models (VLMs). The field uses

safetyarxiv-cs-cv
5 Jun 2026
Safety

DexFuture: Hierarchical Future-State Visuomotor Targeting for Bimanual Dexterous Tool Use

DGX agent

arXiv:2606.05699v1 Announce Type: new Abstract: Bimanual dexterous tool use remains challenging for robots due to high-dimensional hand configurations and complex hand-tool-object dynamics and contact

safetyarxiv-cs-ro
5 Jun 2026
Safety

Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning

DGX agent

arXiv:2606.05645v1 Announce Type: new Abstract: Autonomous driving requires reasoning about how ego actions shape the evolution of the surrounding world. However, most end-to-end methods rely on direc

safetyarxiv-cs-ro
5 Jun 2026
Safety

Disentangled Fine-Grained Prototype Learning for Incomplete Image-Tabular Classification

DGX agent

arXiv:2606.05455v1 Announce Type: new Abstract: The missing-modality problem poses a significant challenge in image-tabular multimodal learning across a wide range of multimedia applications, includin

safetyarxiv-cs-cv
5 Jun 2026
Safety

EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration

DGX agent

arXiv:2602.10106v2 Announce Type: replace Abstract: Human demonstrations offer rich environmental diversity and scale naturally, making them an appealing alternative to robot teleoperation. While this

safetyarxiv-cs-ro
5 Jun 2026
← Previous
1…133134135136137…260
Next →