AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

When AI Says It Feels

DGX agent

arXiv:2606.05734v1 Announce Type: cross Abstract: Large language models (LLMs) are generally constrained from expressing feelings through human-preference alignment in post-training processes. This po

safetyarxiv-cs-cl
5 Jun 2026
Safety

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

DGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
safetyarxiv-cs-cl
5 Jun 2026
Safety

3PoinTr: 3D Point Tracks for Learning Manipulation from Unconstrained Human Videos

DGX agent

arXiv:2603.08485v2 Announce Type: replace Abstract: Learning manipulation policies from human videos could greatly reduce the need for expensive robot demonstrations, but existing approaches typically

safetyarxiv-cs-ro
4 Jun 2026
Safety

A Goal-Set Characterization of Task Composition in the Boolean Task Algebra

DGX agent

arXiv:2606.04053v1 Announce Type: cross Abstract: The Boolean Task Algebra (BTA) provides a principled framework for zero-shot task composition in reinforcement learning by equipping goal-reaching tas

safetyarxiv-cs-ai
4 Jun 2026
Safety

A Unified Framework for Locality in Scalable MARL

DGX agent

arXiv:2602.16966v2 Announce Type: replace-cross Abstract: Scalable methods for networked multi-agent reinforcement learning let each agent plan using only a small neighborhood of the agent graph. This

safetyarxiv-cs-ai
4 Jun 2026
Safety

A Unified Geometric Space for Topological Alignment Between Transformer-Based Models and Human Brain Networks

DGX agent

arXiv:2510.24342v2 Announce Type: replace Abstract: Prior brain-AI alignment studies are typically constrained by specific inputs and tasks, limiting their ability to capture organizational properties

safetyarxiv-cs-ai
4 Jun 2026
Safety

Achieving Rotation-Invariant Convolution via Non-Learnable Orientation Alignment Operators

DGX agent

arXiv:2404.11309v2 Announce Type: replace Abstract: Achieving rotational invariance in deep neural networks without data augmentation is a research hotspot. Intrinsic invariance enables features to ca

safetyarxiv-cs-cv
4 Jun 2026
Safety

Adaptive Calibration for Fair and Performant Facial Recognition

DGX agent

arXiv:2606.04469v1 Announce Type: cross Abstract: We introduce Adaptive Calibration (AC), a novel calibration strategy for facial recognition that maps cosine similarity between normalized embeddings

safetyarxiv-cs-ai
4 Jun 2026
Safety

Adaptive Information Control for Search-Augmented LLM Reasoning

DGX agent

arXiv:2602.01672v2 Announce Type: replace Abstract: Search-augmented reasoning agents interleave multi-step reasoning with external retrieval, but uncontrolled retrieval can introduce redundant eviden

safetyarxiv-cs-cl
4 Jun 2026
Safety

Anycast Performance in Context

DGX agent

arXiv:2606.04298v1 Announce Type: cross Abstract: IP anycast lets a service advertise one address from many physical sites, leaving BGP to map each client to a site. It is central to the DNS root serv

safetyarxiv-cs-ai
4 Jun 2026
Safety

Be Fair! Can Machine Learning Engineering Agents Adhere to Fairness Constraints?

DGX agent

arXiv:2606.04971v1 Announce Type: new Abstract: Machine learning engineering (MLE) agents promise to automate end-to-end ML pipeline development from raw data and natural language instructions, potent

safetyarxiv-cs-lg
4 Jun 2026
Safety

Beyond Static Priors: Dynamic Neural Guidance for Large-Scale Ant Colony Optimization

DGX agent

arXiv:2606.04039v1 Announce Type: cross Abstract: Neural-guided Ant Colony Optimization (ACO) suffers from a fundamental training-inference misalignment: policies are typically trained to generate sta

safetyarxiv-cs-ai
4 Jun 2026
Safety

Beyond Symmetric Alignment: Spectral Diagnostics of Modality Imbalance in Vision-Language Models in the Medical Domain

DGX agent

arXiv:2606.04613v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle when applied to medical image-text data, yet the tools available to diagnose this failure remain limited. Existin

safetyarxiv-cs-cv
4 Jun 2026
Safety

BiasGRPO: Stabilizing Bias Mitigation in High-Variance Reward Landscapes via Group-Relative Policy Optimization

DGX agent

arXiv:2606.04807v1 Announce Type: new Abstract: Mitigating social bias in Large Language Models (LLMs) presents a distinct alignment challenge: unlike verifiable tasks, bias lacks a single ground trut

safetyarxiv-cs-ai
4 Jun 2026
Safety

Blessing from Human-AI Interaction: Super Reinforcement Learning in Confounded Environments

DGX agent

arXiv:2209.15448v3 Announce Type: replace Abstract: As AI becomes more prevalent throughout society, effective methods of integrating humans and AI systems that leverage their respective strengths and

safetyarxiv-cs-lg
4 Jun 2026
Safety

Causal Multi-fidelity Surrogate Forward and Inverse Models for ICF Implosions

DGX agent

arXiv:2509.05510v3 Announce Type: replace-cross Abstract: Continued progress in inertial confinement fusion (ICF) requires solving inverse problems relating experimental observations to simulation inp

safetyarxiv-cs-lg
4 Jun 2026
Safety

Channel-Oriented Design for EEG-to-Music Reconstruction

DGX agent

arXiv:2606.04040v1 Announce Type: cross Abstract: Brain-computer interfaces aim to decode naturalistic stimuli from neural signals, yet most progress to date has focused on vision and language. In thi

safetyarxiv-cs-ai
4 Jun 2026
Safety

Confidence Before Answering: A Paradigm Shift for Efficient LLM Uncertainty Estimation

DGX agent

arXiv:2603.05881v2 Announce Type: replace Abstract: Reliable deployment of large language models (LLMs) requires accurate uncertainty estimation. Existing methods are predominantly answer-first, produ

safetyarxiv-cs-cl
4 Jun 2026
Safety

CoRe-MoE: Contrastive Reweighted Mixture of Experts for Multi-Terrain Humanoid Locomotion with Gait Adaptation

DGX agent

arXiv:2606.04718v1 Announce Type: cross Abstract: Humans primarily rely on walking and running to traverse complex terrains, without resorting to unnecessarily complex motion patterns. Similarly, huma

safetyarxiv-cs-ai
4 Jun 2026
Safety

Covert Influence Between Language Models

DGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

safetyarxiv-cs-cl
4 Jun 2026
Safety

Crafting Your Evolving Dreams: Concept-Incremental Versatile Customization

DGX agent

arXiv:2606.04797v1 Announce Type: new Abstract: Custom diffusion models (CDMs) have garnered significant interest owing to their remarkable capacity for generating personalized concepts. However, the

safetyarxiv-cs-cv
4 Jun 2026
Safety

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

DGX agent

arXiv:2601.22396v2 Announce Type: replace-cross Abstract: Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accura

safetyarxiv-cs-ai
4 Jun 2026
Safety

Customizing the Inductive Biases of Softmax Attention using Structured Matrices

DGX agent

arXiv:2509.07963v2 Announce Type: replace Abstract: The core component of attention is the scoring function, which transforms the inputs into low-dimensional queries and keys and takes the dot product

safetyarxiv-cs-lg
4 Jun 2026
Safety

DiffAero: A GPU-Accelerated Differentiable Simulation Framework for Efficient Quadrotor Policy Learning

DGX agent

arXiv:2509.10247v1 Announce Type: cross Abstract: This letter introduces DiffAero, a lightweight, GPU-accelerated, and fully differentiable simulation framework designed for efficient quadrotor contro

safetyarxiv-cs-ai
4 Jun 2026
Safety

DPM++: Dynamic Masked Metric Learning for Occluded Person Re-identification

DGX agent

arXiv:2605.06637v2 Announce Type: replace Abstract: Although person re-identification has made impressive progress, occlusion caused by obstacles remains an unsettled issue in real applications. The d

safetyarxiv-cs-cv
4 Jun 2026
Safety

DuDi: Dual-Signal Distillation with Cross-Lingual Verbalizer

DGX agent

arXiv:2606.04694v1 Announce Type: new Abstract: Small language models (SLMs) are efficient and scalable, but their multilingual capabilities degrade severely at sub-billion scales, especially for Sout

safetyarxiv-cs-cl
4 Jun 2026
Safety

DVGT: Driving Visual Geometry Transformer

DGX agent

arXiv:2512.16919v2 Announce Type: replace-cross Abstract: Perceiving and reconstructing 3D scene geometry from visual inputs is crucial for autonomous driving. However, there still lacks a driving-tar

safetyarxiv-cs-ai
4 Jun 2026
Safety

Dynamic Multi-Pair Trading Strategy in Cryptocurrency Markets with Deep Reinforcement Learning

DGX agent

arXiv:2606.04574v1 Announce Type: new Abstract: This study aims to determine whether the application of Deep Reinforcement Learning (DRL) as a specialized execution overlay can enhance pair trading in

safetyarxiv-cs-lg
4 Jun 2026
Safety

Dynamic Policy Learning for Legged Robot with Simplified Model Pretraining and Model-Homotopy-Inspired Transfer

DGX agent

arXiv:2512.24698v2 Announce Type: replace Abstract: Generating dynamic motions for legged robots remains a challenging problem. While reinforcement learning has achieved notable success in various leg

safetyarxiv-cs-ro
4 Jun 2026
Safety

Edge of Stability Selectively Shapes Learning Across the Data Distribution

DGX agent

arXiv:2606.04212v1 Announce Type: new Abstract: Existing analyses of the edge of stability (EoS) treat it as a global property of optimization. We show that it is also selective: the stability constra

safetyarxiv-cs-lg
4 Jun 2026
Safety

Efficient Adversarial Attacks on High-dimensional Offline Bandits

DGX agent

arXiv:2602.01658v2 Announce Type: replace-cross Abstract: Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large

safetyarxiv-cs-ai
4 Jun 2026
Safety

Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling

DGX agent

arXiv:2606.05021v1 Announce Type: new Abstract: We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algor

safetyarxiv-cs-lg
4 Jun 2026
Safety

Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods

DGX agent

arXiv:2511.03304v2 Announce Type: replace-cross Abstract: With the on-going integration of machine learning systems into the everyday social life of millions the notion of fairness becomes an ever inc

safetyarxiv-cs-ai
4 Jun 2026
Safety

FLAGG: Flexible Autoregressive Graph Generation

DGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

safetyarxiv-cs-lg
4 Jun 2026
Safety

Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learning in a Game Environment

DGX agent

arXiv:2606.04750v1 Announce Type: new Abstract: Instilling virtuous behavior in artificial intelligence has seen increasing interest. One of the techniques proposed is known as affinity-based reinforc

safetyarxiv-cs-ai
4 Jun 2026
Safety

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

DGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

safetyarxiv-cs-cl
4 Jun 2026
Safety

Generalizable Multi-Task Learning for Wireless Networks Using Prompt Decision Transformers

DGX agent

arXiv:2606.04328v1 Announce Type: cross Abstract: Future wireless networks demand rapid adaptation to highly heterogeneous environments and dynamic task configurations, necessitating a shift from conv

safetyarxiv-cs-ai
4 Jun 2026
Safety

Generalization of World Models under Environmental Variability for Vision-based Quadrotor Navigation

DGX agent

arXiv:2606.05015v1 Announce Type: new Abstract: World models, learned generative models that predict how an environment evolves, have become a promising tool for sample-efficient robot learning. Yet h

safetyarxiv-cs-ro
4 Jun 2026
Safety

Geometry-Aware Distillation for Prompt Tuning Biomedical Vision-Language Models

DGX agent

arXiv:2606.04922v1 Announce Type: cross Abstract: Current prompt-based and adapter-based tuning of vision-language models (VLMs) is attractive for medical imaging, where clinical data sensitivity favo

safetyarxiv-cs-ai
4 Jun 2026
Safety

Geospatial Foundation Models to Enable Progress on Sustainable Development Goals

DGX agent

arXiv:2505.24528v3 Announce Type: replace Abstract: Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and c

safetyarxiv-cs-cv
4 Jun 2026
Safety

Global Sketch-Based Watermarking for Diffusion Language Models

DGX agent

arXiv:2606.04486v1 Announce Type: cross Abstract: Watermarking methods for language models have been studied extensively in the autoregressive setting, where tokens are generated sequentially. These w

safetyarxiv-cs-cl
4 Jun 2026
Safety

Good Reasoning Makes Good Demonstrations: Implicit Reasoning Quality Supervision via In-Context Reinforcement Learning

DGX agent

arXiv:2603.09803v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves reasoning in large language models but treats all correct solutions equally, potentia

safetyarxiv-cs-lg
4 Jun 2026
Safety

GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2606.04889v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (e.g. GRPO) is now a common way to improve mathematical reasoning in Large Language Models (LLMs). Howeve

safetyarxiv-cs-cl
4 Jun 2026
Safety

HapTile: A Haptic-Informed Vision-Tactile-Language-Action Dataset for Contact-Rich Imitation Learning

DGX agent

arXiv:2606.04825v1 Announce Type: new Abstract: Despite the importance of tactile sensing for reliable manipulation, most existing Vision-Language-Action (VLA) datasets remain vision-only, and those t

safetyarxiv-cs-ro
4 Jun 2026
Safety

Hybrid Adversarial Defence for Natural Language Understanding Tasks

DGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

safetyarxiv-cs-cl
4 Jun 2026
Safety

Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

DGX agent

arXiv:2606.05030v1 Announce Type: new Abstract: Autoregressive chain-of-thought (CoT) reasoning in large language models (LLMs) is fundamentally forward-directed: each step conditions only on prior to

safetyarxiv-cs-cl
4 Jun 2026
Safety

In-Context Graphical Inference

DGX agent

arXiv:2606.05042v1 Announce Type: cross Abstract: Marginal inference in discrete graphical models forces a choice between exactness and scalability: exact algorithms are intractable for high-treewidth

safetyarxiv-cs-cl
4 Jun 2026
Safety

Instant-Fold: In-Context Imitation Learning for Deformable Object Manipulation

DGX agent

arXiv:2606.04269v1 Announce Type: cross Abstract: Deformable object manipulation (DOM) is challenging due to high-dimensional, partially observable states that evolve through long-horizon, topology-ch

safetyarxiv-cs-ai
4 Jun 2026
← Previous
1…135136137138139…260
Next →