AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
20 Jul 2026

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example.…

SafetyDGX agent

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example. S&P cut Oracle's credit rating to one notch above junk, mean

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss.…

SafetyDGX agent

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss. We’re sharing what we learned from studying a long-running

*None* of the four of us believe that LLMs on their own are sufficient for AGI.

Safety

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

*None* of the four of us believe that LLMs on their own are sufficient for AGI. Key people in AI worth following • François Chollet —> @fchollet • Yann LeCun —-> @ylecun • Gary Marcus —-> @GaryMarcus

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute alloc…

SafetyDGX agent

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute allocation on inference, releasing os will enable continued full

16 Jul 2026

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

SafetyDGX agent

arXiv:2606.16944v2 Announce Type: replace Abstract: Theory of mind (ToM), the capacity to ascribe mental states to others and use those ascriptions for prediction and inference, is widely assumed to b

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

SafetyDGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

A Generalised Exponentiated Gradient Approach to Enhance Fairness in Binary and Multi-class Classification Tasks

SafetyDGX agent

arXiv:2603.21393v2 Announce Type: replace Abstract: The widespread use of AI and ML models in sensitive areas raises significant concerns about fairness. While the research community has introduced va

A Hybrid Sampling-Based Trajectory Planner with Game-Theoretic Guidance for Autonomous Racing

SafetyDGX agent

arXiv:2607.13354v1 Announce Type: new Abstract: Autonomous racing demands planning algorithms that balance vehicle dynamics at the limits of handling with strategic decision-making in competitive mult

A VAE-Driven Multi-Task Satellite-Aided Semantic Communication Framework for 6G-Enabled Connected Autonomous Vehicles

SafetyDGX agent

arXiv:2607.13494v1 Announce Type: new Abstract: The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle networ

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

SafetyDGX agent

arXiv:2607.13160v1 Announce Type: cross Abstract: Active beyond-diagonal reconfigurable intelligent surfaces (BD-RISs) enables hybrid transmitting and reflecting mode to achieve effective signal ampli

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

SafetyDGX agent

arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accum

Adversarial Prompting Framework for AI Safety Assessment

SafetyDGX agent

arXiv:2607.13453v1 Announce Type: cross Abstract: Artificial Intelligence (AI), especially Generative AI (GenAI), adoption has increased in industries significantly in recent years. However, the use o

Agile perceptive multi-skill locomotion for quadrupedal robots in the wild

SafetyDGX agent

arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multipl

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

SafetyDGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

An Empirical Study on Stage-Information Interfaces for VLA Fine-Tuning

SafetyDGX agent

arXiv:2607.13605v1 Announce Type: new Abstract: One high-level instruction in long-horizon manipulation can cover several action stages. We use segmented action annotations as an intermediate represen

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization

SafetyDGX agent

arXiv:2607.13805v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining learns a shared representation space through large-scale contrastive learning. However, existing methods that enf

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

SafetyDGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

BARS: Benign-Anchored Ranking and Selection for False Alarm Reduction in Network Intrusion Detection

SafetyDGX agent

arXiv:2607.13203v1 Announce Type: cross Abstract: False alarms remain a major barrier to deploying network intrusion detection systems (NIDS). In high-volume environments, even a sub-1% false positive

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

SafetyDGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

C-Norm: Cell-Distribution Normalization Enables Precision Recognition of Medical-Cell Image

SafetyDGX agent

arXiv:2607.13116v1 Announce Type: new Abstract: ThinPrep Cytologic Test (TCT) enables early cervical cancer screening, but manual reading is time-consuming and yields inconsistent diagnostic results a

Cost-Optimal Foundation Model Deployment Portfolio for Transportation Management

SafetyDGX agent

arXiv:2607.13239v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs) and vision-language models (VLMs), are increasingly used for transportation management center

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention

SafetyDGX agent

arXiv:2607.13731v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders di

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

SafetyDGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

SafetyDGX agent

arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systemat

Designing Safety-Constrained LLM Systems for Public Health Information Access

SafetyDGX agent

arXiv:2607.13038v1 Announce Type: cross Abstract: We present the design and implementation of a safety constrained large language model (LLM) system for public health information access, focusing on m

Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for Disaster Governance

SafetyDGX agent

arXiv:2607.13260v1 Announce Type: cross Abstract: Policy documents shape governance outcomes, but their reasoning is often implicit. Participatory commitments and managerial control routinely coexist

Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing

SafetyDGX agent

arXiv:2607.13103v1 Announce Type: cross Abstract: Knowledge tracing (KT) aims to predict students' future performance by modeling their evolving knowledge states from historical interactions. Existing

Distributionally Robust and Safe Imitation Learning

SafetyDGX agent

arXiv:2607.13436v1 Announce Type: new Abstract: Imitation learning (IL) has achieved remarkable success in complex decision-making tasks. However, its performance is highly sensitive to distribution s

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education

SafetyDGX agent

arXiv:2607.14046v1 Announce Type: new Abstract: This paper presents Earthquaker-AI, a hybrid educational framework building upon a previously implemented educational robotics project by integrating a

Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-Loop Table Recognition

SafetyDGX agent

arXiv:2607.13347v1 Announce Type: cross Abstract: LLM-as-a-judge is widely used to provide feedback and selection signals in closedloop regeneration, but this use remains insufficiently validated. We

Explaining Reinforcement Learning Agents via Inductive Logic Programming

SafetyDGX agent

arXiv:2607.13655v1 Announce Type: new Abstract: Explainable Reinforcement Learning (XRL) seeks to make Reinforcement Learning (RL) policies more transparent and interpretable, a key requirement in saf

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

SafetyDGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

Fine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT Understanding

SafetyDGX agent

arXiv:2607.13892v1 Announce Type: new Abstract: Computed tomography (CT) vision-language pretraining from paired volumes and radiology reports is a scalable yet challenging task. Existing methods comm

FreeLit: Paired-Free Indoor Relighting via Physics-Guided Diffusion

SafetyDGX agent

arXiv:2607.13656v1 Announce Type: new Abstract: Image-based indoor scene relighting remains challenging due to the complex interplay between cluttered geometry and local illumination, requiring precis

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment

SafetyDGX agent

arXiv:2607.13429v1 Announce Type: cross Abstract: Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-langua

GNN-DIP: Neural Corridor Selection for Decomposition-Based Motion Planning

SafetyDGX agent

arXiv:2603.12361v2 Announce Type: replace Abstract: Motion planning through narrow passages remains a core challenge: sampling-based planners rarely place samples inside these narrow but critical regi

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs

SafetyDGX agent

arXiv:2607.13712v1 Announce Type: cross Abstract: Despite the rapid progress of Multimodal Large Language Models (MLLMs), they still suffer from untruthfulness issues, such as visual hallucinations, c

HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation

SafetyDGX agent

arXiv:2607.09776v2 Announce Type: replace-cross Abstract: When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively ad

Human4K: A Large-Scale 4K Multi-View Mocap Dataset for Whole-Body 3D Human Reconstruction

SafetyDGX agent

arXiv:2607.13646v1 Announce Type: cross Abstract: Recent advances in 3D human reconstruction have improved overall performance, yet current models still fail in the most challenging real-world scenari

Introducing Human-Centeredness in AI-Assisted Lexicography

SafetyDGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

Inverse-LLaVA: Rethinking Multimodal Alignment via Text-to-Vision Mapping

SafetyDGX agent

arXiv:2508.12466v2 Announce Type: replace-cross Abstract: Traditional multimodal learning approaches rely on alignment pre-training to bridge vision and language modalities, typically by projecting vi

Joint On-and-Off Policy Learning for Vision-and-Language Navigation

SafetyDGX agent

arXiv:2607.13461v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) necessitates an embodied agent to navigate in the physical world by adhering to natural language instructions. Rece

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

SafetyDGX agent

arXiv:2607.13501v1 Announce Type: new Abstract: Reinforcement learning for multi-turn search reasoning typically relies on terminal outcome rewards, which cannot distinguish useful, redundant, and har

Layered Risk Mapping for Autonomous Patient Transport in Expeditionary Medical Facilities

SafetyDGX agent

arXiv:2607.13497v1 Announce Type: new Abstract: In expeditionary medical facilities, routine patient transport imposes a compounding burden of personal protective equipment consumption, staff diversio

Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents

SafetyDGX agent

arXiv:2606.18223v2 Announce Type: replace-cross Abstract: With sophisticated cyber-attacks becoming increasingly prevalent, modern networks require intelligent autonomous cyber-defense agents trained

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

SafetyDGX agent

arXiv:2607.13172v1 Announce Type: new Abstract: We address the problem of safely training an agent policy and deploying a good and safe policy, in settings where the environment dynamics are unknown a

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

SafetyDGX agent

arXiv:2604.00830v3 Announce Type: replace-cross Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at

Marker-free deformable registration and fusion for augmented reality-guided positive margin localization during tumor resection surgery

SafetyDGX agent

arXiv:2607.13343v1 Announce Type: new Abstract: Positive margins in head and neck oncologic surgery require mapping specimen-side pathology findings to the patient resection bed. This is challenging b

MASPRM: Multi-Agent System Process Reward Model

SafetyDGX agent

arXiv:2510.24803v3 Announce Type: replace-cross Abstract: Inference-time search over multi-agent systems (MAS) wastes compute when it cannot identify which agent's intermediate message advanced progre

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

SafetyDGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

Mind the Gap: Action Rebinding Attacks against Android GUI Agents

SafetyDGX agent

arXiv:2601.12349v3 Announce Type: replace-cross Abstract: Large multimodal model powered GUI agents are emerging as high-privilege operators on mobile platforms, entrusted to perceive screen content a

Music-to-Dance Generation via Atomic Movements

SafetyDGX agent

arXiv:2607.13978v1 Announce Type: cross Abstract: Music-driven dance generation aims to produce human motion that is both rhythmically synchronized and semantically consistent with music. While recent

Non-Expansive Two-Time-Scale Stochastic Approximation: A Fixed-Schedule One-Quarter Barrier and Bias-Corrected Acceleration

SafetyDGX agent

arXiv:2607.13414v1 Announce Type: cross Abstract: Non-expansive two-time-scale stochastic approximation is governed by a slow stochastic Krasnoselskii--Mann fixed-point iteration rather than by contra

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache

SafetyDGX agent

arXiv:2505.18231v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference is typically memory-intensive, especially when processing large batch sizes and long sequences, due to th

On the Sublinear Regret of Continuous K-Max Bandits

SafetyDGX agent

arXiv:2502.13467v2 Announce Type: replace Abstract: The K-Max combinatorial multi-armed bandit problem arises in applications such as recommendation and distributed decision making, where the reward i

Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows

SafetyDGX agent

arXiv:2607.13078v1 Announce Type: cross Abstract: LLMs are now proposed for fraud detection, scam investigation, content moderation, and other trust-and-safety workflows. Much of the public literature

Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation

SafetyDGX agent

arXiv:2607.13686v1 Announce Type: new Abstract: We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a conte

OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations

SafetyDGX agent

arXiv:2607.13420v1 Announce Type: new Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-P

PC-Diffuser: Path-Consistent Capsule CBF Safety Filtering for Diffusion-Based Trajectory Planner

SafetyDGX agent

arXiv:2603.10330v2 Announce Type: replace-cross Abstract: Autonomous driving in complex traffic requires planners that generalize beyond hand-crafted rules, motivating data-driven approaches that lear

PhysClaw-0: A Symbiotic Agentic System for Robot Autonomy via Language Corrections

SafetyDGX agent

arXiv:2607.14047v1 Announce Type: new Abstract: Autonomous data collection governs the volume and quality of real-world trajectories for manipulation policy learning. Existing pipelines reduce human e

← Previous
1…3334353637…212
Next →