AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

DV-SFT: Direct Vision Supervision for Fine-Grained Visual Understanding

DGX agent

arXiv:2605.26656v1 Announce Type: new Abstract: Multimodal large language models are typically trained end-to-end to predict ground-truth answers, yet supervision signals are applied exclusively to te

safetyarxiv-cs-cv
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Safety

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient

DGX agent

arXiv:2605.26478v1 Announce Type: cross Abstract: We present the stochastic decoupled policy gradient (SDPG), a lightweight visual reinforcement learning (RL) method that trains diverse visuomotor con

safetyarxiv-cs-ai
27 May 2026
Safety

Elias in the Lighthouse, Again? Diagnosing Low Diversity in LLM Stories

DGX agent

arXiv:2605.26492v1 Announce Type: cross Abstract: LLM-generated stories are a popular use case, but they show very low variability. We sample 20,000 total stories from four current models using five p

safetyarxiv-cs-ai
27 May 2026
Safety

Elon Musk on why humanity must become multiplanetary: “I think it's important for the long-term preservation and ultimately the expansion an…

DGX agent

Elon Musk on why humanity must become multiplanetary: “I think it's important for the long-term preservation and ultimately the expansion and extension of the scope and scale of consciousness... that

safetyelon-musk--x
27 May 2026
Safety

EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation

DGX agent

arXiv:2605.26785v1 Announce Type: cross Abstract: Post-trained LLMs are often optimized to align responses with human preferences, making them safe, polite, and conversationally appropriate. In advers

safetyarxiv-cs-ai
27 May 2026
Safety

Enabling Extensible Embodied Capabilities with Tools

DGX agent

arXiv:2605.26637v1 Announce Type: new Abstract: Most existing embodied intelligence methods formulate perception, reasoning, planning, and control within a unified parameterized policy. Yet these capa

safetyarxiv-cs-ro
27 May 2026
Safety

Erased but Exploitable: Black-box Embedding-Aware Prompting Against Unlearned Text-to-Image Diffusion Models

DGX agent

arXiv:2605.26332v1 Announce Type: cross Abstract: Machine unlearning aims to remove specific concepts from pretrained text-to-image diffusion models, yet several white- and black-box attacks have been

safetyarxiv-cs-ai
27 May 2026
Safety

🔬ESMFold2: The Bitter Lesson is Coming for Proteins - Alex Rives, BioHub

DGX agent

ESMFold2 represents an advancement in protein structure prediction leveraging scaling laws and transformer-based language models, building on principles that favor compute and data scale over hand-cra

safetylatent-space
27 May 2026
Safety

Ethical Fairness without Demographics in Human-Centered AI

DGX agent

arXiv:2603.13373v3 Announce Type: replace-cross Abstract: In ubiquitous and mobile health systems, computational models infer human states from wearable, behavioral, and physiological sensing data. In

safetyarxiv-cs-ai
27 May 2026
Safety

Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights

DGX agent

arXiv:2501.06708v5 Announce Type: replace-cross Abstract: Large-scale web-crawled datasets contain noise, bias, and irrelevant information, necessitating data selection techniques. Existing methods de

safetyarxiv-cs-ai
27 May 2026
Safety

FalAR: A Large-scale Speaker-Annotated European Portuguese Speech Corpus of Parliamentary Sessions

DGX agent

arXiv:2605.27062v1 Announce Type: new Abstract: State-of-the-art performance for Automatic Speech Recognition (ASR) largely depends on the availability of large-scale labeled corpora. This creates a d

safetyarxiv-cs-cl
27 May 2026
Safety

Few-shot Cross-country Generalization of Tabular Machine Learning and Foundation Models for Childhood Anemia Prediction under Distribution Shift

DGX agent

arXiv:2605.26589v1 Announce Type: cross Abstract: Childhood anemia affects around 40% of children aged 6-59 months globally and arises from heterogeneous factors, limiting model generalizability. We e

safetyarxiv-cs-ai
27 May 2026
Safety

FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

DGX agent

arXiv:2605.27333v1 Announce Type: new Abstract: Finance LLM agents must simultaneously block prompt-induced unauthorized actions and approve legitimate multi-step business workflows. However, boundary

safetyarxiv-cs-cl
27 May 2026
Safety

Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints

DGX agent

arXiv:2603.17685v3 Announce Type: replace Abstract: Balancing policy expressiveness with the exploration-exploitation trade-off is a core challenge in online Reinforcement Learning (RL). While Stochas

safetyarxiv-cs-lg
27 May 2026
Safety

FM-fMRI: Event Conditioned Flow Matching for Rest-to-Task fMRI Time-Series Synthesis

DGX agent

arXiv:2605.26423v1 Announce Type: new Abstract: Task-based fMRI provides a direct readout of task-evoked neural dynamics, but it is expensive and difficult to acquire at scale, motivating rest-to-task

safetyarxiv-cs-lg
27 May 2026
Safety

Foundations of a Time-Consistent Counterfactual Actuarial Runtime for Autonomous AI Agents

DGX agent

arXiv:2605.26508v1 Announce Type: cross Abstract: We propose a foundational runtime actuarial layer for autonomous AI agents in which every side-effect-bearing action carries a time-consistent, counte

safetyarxiv-cs-ai
27 May 2026
Safety

From Norms to Indicators (N2I-RAG): An Agentic Retrieval-Augmented Generation Framework for Legal Indicator Computation

DGX agent

arXiv:2605.26926v1 Announce Type: new Abstract: Computing legal indicators from normative texts is a key task in legal monitoring and policy evaluation, but presents significant challenges due to the

safetyarxiv-cs-ai
27 May 2026
Safety

From Static Context to Calibrated Interactive RL: Mitigating Distribution Shift in Multi-turn Dialogue with Aligned Simulator

DGX agent

arXiv:2605.26403v1 Announce Type: new Abstract: A long-standing goal of the research community is to develop highly interactive LLM-based dialogue agents. Recent research focuses on optimizing policie

safetyarxiv-cs-ai
27 May 2026
Safety

FTibSuite: A Comprehensive Resource Suite for Tibetan Vision-Language Modeling

DGX agent

arXiv:2605.26601v1 Announce Type: new Abstract: Vision-language models have progressed rapidly, but Tibetan remains a severely underserved low-resource language due to the lack of reproducible trainin

safetyarxiv-cs-cv
27 May 2026
Safety

Furina: Fragmented Uncertainty-Driven Refusal Instability Attack

DGX agent

arXiv:2605.26158v1 Announce Type: cross Abstract: Safety alignment in large language models (LLMs) and multimodal large language models (MLLMs) is commonly assumed to operate as a near-binary threshol

safetyarxiv-cs-ai
27 May 2026
Safety

Gary's no way as bad as some people make him out to be.

DGX agent

Gary's no way as bad as some people make him out to be. some data i shared yesterday on anthropic revenue possibly slowing down aren’t as a compelling as i thought; i have deleted my posts and await b

safetygary-marcus--x
27 May 2026
Safety

Generalist Graph Anomaly Detection via Prototype-Based Distillation

DGX agent

arXiv:2605.26857v1 Announce Type: new Abstract: Driven by the pressing demand for graph anomaly detection (GAD) in high-stakes domains, the generalist GAD paradigm, which trains a single detector tran

safetyarxiv-cs-lg
27 May 2026
Safety

GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation

DGX agent

arXiv:2602.16449v2 Announce Type: replace-cross Abstract: Generative model evaluation commonly relies on high-dimensional embedding spaces to compute distances between samples. We show that dataset re

safetyarxiv-cs-ai
27 May 2026
Safety

Grounding Text Embeddings in Stakeholder Associations

DGX agent

arXiv:2605.27168v1 Announce Type: cross Abstract: Text embeddings are widely used to analyse large corpora of complex texts. However, it is unclear whether the embeddings capture the same semantic dis

safetyarxiv-cs-ai
27 May 2026
Safety

Heterogeneous AAV Logistics Task Allocation: A Reinforcement Learning Enhanced Overlapping Coalition Formation Game Approach

DGX agent

arXiv:2605.26471v1 Announce Type: new Abstract: In dynamic urban logistics, the stochastic emergence of time-sensitive tasks poses a significant optimality challenge for heterogeneous AAVs logistics t

safetyarxiv-cs-ro
27 May 2026
Safety

Hi-SAM: A Hierarchical Structure-Aware Multi-modal Framework for Large-Scale Recommendation

DGX agent

arXiv:2602.11799v2 Announce Type: replace Abstract: Multi-modal recommendation has gained traction as items possess rich attributes like text and images. Semantic ID-based approaches effectively discr

safetyarxiv-cs-ai
27 May 2026
Safety

How Reliable are LLMs for Reasoning on the Re-ranking task?

DGX agent

arXiv:2508.18444v2 Announce Type: replace-cross Abstract: With the improving semantic understanding capability of Large Language Models (LLMs), they exhibit a greater awareness and alignment with huma

safetyarxiv-cs-ai
27 May 2026
Safety

HyperSim: A Holistic Sim-To-Real Framework For Robust Robotic Manipulation

DGX agent

arXiv:2605.26638v1 Announce Type: new Abstract: Scaling data volume and diversity is critical for generalizing embodied intelligence. While synthetic data generation offers a scalable alternative to e

safetyarxiv-cs-ro
27 May 2026
Safety

i so wish i could fast forward a few years to see how all this turned out.

DGX agent

Gary Marcus expresses curiosity about the long-term outcomes of current developments, likely related to artificial intelligence or technology trends given his expertise in AI and cognitive science. Th

safetygary-marcus--x
27 May 2026
Safety

Illinois Legislature passes SB 315, a bill requiring annual independent third-party safety audits of leading AI companies; the bill heads to the governor's desk (Jared Perlo/NBC News)

DGX agent

Jared Perlo / NBC News: Illinois Legislature passes SB 315, a bill requiring annual independent third-party safety audits of leading AI companies; the bill heads to the governor's desk — The measure,

safetytechmeme
27 May 2026
Safety

Image Thresholding: Understanding Bias of Evaluation Metrics towards Specific Evaluation Functions

DGX agent

arXiv:2605.27132v1 Announce Type: new Abstract: Multilevel image thresholding is widely used for segmentation in applications ranging from medical imaging to remote sensing. Classical objective functi

safetyarxiv-cs-cv
27 May 2026
Safety

Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinforcement Learning Approaches and Architectures

DGX agent

arXiv:2502.06963v3 Announce Type: replace-cross Abstract: The increasing complexity of Intelligent Transportation Systems (ITS) has led to significant interest in computational offloading to external

safetyarxiv-cs-ai
27 May 2026
Safety

Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification

DGX agent

arXiv:2510.00902v2 Announce Type: replace Abstract: Transfer learning is crucial for medical imaging, yet the selection of source datasets often relies on researchers' intuition rather than systematic

safetyarxiv-cs-cv
27 May 2026
Safety

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

DGX agent

arXiv:2605.27288v1 Announce Type: cross Abstract: Large language models (LLMs) are known to abandon their initial stance to conform to user pushback. While prior research largely attributes this behav

safetyarxiv-cs-ai
27 May 2026
Safety

Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models

DGX agent

arXiv:2605.26409v1 Announce Type: cross Abstract: Evaluating and mitigating a generative system's susceptibility to jailbreak attacks is critical to its safe deployment. Given the number of deployable

safetyarxiv-cs-ai
27 May 2026
Safety

KARMA: Karma-Aligned Reward Model Adaptation

DGX agent

arXiv:2605.26738v1 Announce Type: new Abstract: Human communication depends on implicit social signals where effectiveness is shaped by tone, context, and conversational norms rather than semantic con

safetyarxiv-cs-cl
27 May 2026
Safety

KZ-SafetyPrompts: A Kazakh Safety Evaluation Prompt Dataset for Large Language Models

DGX agent

arXiv:2605.26947v1 Announce Type: new Abstract: Kazakh is underrepresented in resources for evaluating the safety behavior of large language models. We present KZ-SafetyPrompts, a Kazakh prompt datase

safetyarxiv-cs-cl
27 May 2026
Safety

LAD-VF: LLM-Automatic Differentiation Enables Fine-Tuning-Free Robot Planning from Formal Methods Feedback

DGX agent

arXiv:2509.18384v2 Announce Type: replace Abstract: Large language models (LLMs) can translate natural language instructions into executable action plans for robotics, autonomous driving, and other do

safetyarxiv-cs-ro
27 May 2026
Safety

LearnedCache: An eBPF-Integrated Perceptron-Based Eviction Policy for the Linux Page Cache

DGX agent

arXiv:2605.26168v1 Announce Type: cross Abstract: Linux is the foundation of the digital age, accounting for the majority of the cloud and mobile OS markets. Any device that runs Linux uses the Linux

safetyarxiv-cs-lg
27 May 2026
Safety

Learning Dynamic Graph Representations through Timespan View Contrasts

DGX agent

arXiv:2605.27063v1 Announce Type: new Abstract: The rich information underlying graphs has inspired further investigation of unsupervised graph representation. Existing studies mainly depend on node f

safetyarxiv-cs-lg
27 May 2026
Safety

Learning to Balance Motor Thermal Safety and Quadrupedal Locomotion Performance with Residual Policy

DGX agent

arXiv:2605.27046v1 Announce Type: new Abstract: Motor thermal management is often overlooked in the context of electrically-actuated robots, particularly legged robots, but motor overheating is a key

safetyarxiv-cs-ro
27 May 2026
Safety

Learning to Orchestrate Agents under Uncertainty

DGX agent

arXiv:2605.27073v1 Announce Type: new Abstract: Adaptive orchestration of heterogeneous agents requires making sequential delegation decisions under uncertain and evolving agent behaviour, e.g., coord

safetyarxiv-cs-lg
27 May 2026
Safety

Learning to Reason Efficiently with Discounted Reinforcement Learning

DGX agent

arXiv:2510.23486v2 Announce Type: replace Abstract: Large reasoning models (LRMs) often consume excessive tokens, inflating computational cost and latency. More broadly, in goal reaching sequential de

safetyarxiv-cs-lg
27 May 2026
Safety

Less is More: Early Stopping Rollout for On-Policy Distillation

DGX agent

arXiv:2605.27028v1 Announce Type: cross Abstract: On-policy distillation has recently emerged as a promising alternative to standard sequence-level imitation, training a student by scoring its own rol

safetyarxiv-cs-ai
27 May 2026
Safety

Linear and Neural Dueling Bandits with Delayed Feedback

DGX agent

arXiv:2605.26554v1 Announce Type: cross Abstract: Contextual dueling bandits form a cornerstone of preference-based decision-making, with critical applications in recommender systems and large languag

safetyarxiv-cs-ai
27 May 2026
Safety

Look Further: Socially-Compliant Navigation System in Residential Buildings

DGX agent

arXiv:2605.26710v1 Announce Type: new Abstract: The distance at which a mobile robot reacts to a person strongly impacts various qualities of the human-robot interaction. In this paper, we focus on th

safetyarxiv-cs-ro
27 May 2026
Safety

MAIGO: Mitigating Lost-in-Conversation with History-Cleaned On-Policy Self-Distillation

DGX agent

arXiv:2605.27186v1 Announce Type: new Abstract: Large language models often solve tasks from a fully specified prompt but degrade when the same requirements unfold over multiple turns, known as the lo

safetyarxiv-cs-cl
27 May 2026
← Previous
1…141142143144145…267
Next →