AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

some news: it turns out 34,000 Instagram accounts got hit by a Meta AI exploit the other day, per internal company docs hackers also used a …

DGX agent

some news: it turns out 34,000 Instagram accounts got hit by a Meta AI exploit the other day, per internal company docs hackers also used a senior 'Space Force' official's account to post anti-Iran wa

safetygary-marcus--x
9 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning

DGX agent

arXiv:2606.08992v1 Announce Type: cross Abstract: Vision-and-Language Navigation in continuous environments requires agents to understand the spatial structure of previously unseen environments in ord

safetyarxiv-cs-ai
9 Jun 2026
Safety

Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models

DGX agent

arXiv:2606.08446v1 Announce Type: cross Abstract: Despite being powerful, reinforcement learning with verifiable rewards (RLVR) induces extremely long COT, making it computationally expensive. Since R

safetyarxiv-cs-ai
9 Jun 2026
Safety

Speaker-Invariant Representation Learning for Spoofing Detection via Gradient Reversal and A Variational Information Bottleneck

DGX agent

arXiv:2606.08678v1 Announce Type: cross Abstract: Sophisticated generative speech technology can undermined the reliability of voice biometrics. While spoofing detection systems excel when assessed un

safetyarxiv-cs-lg
9 Jun 2026
Safety

SPIN: Decentralized Swarm Control via Tensorized Policy Coordination

DGX agent

arXiv:2606.07557v1 Announce Type: new Abstract: Decentralized multi-agent swarm coordination on resource-constrained edge platforms remains fundamentally bottlenecked by the exponential scaling of joi

safetyarxiv-cs-lg
9 Jun 2026
Safety

Stage-1 Controls the Entropy Regime, Not the Outcome

DGX agent

arXiv:2606.09059v1 Announce Type: cross Abstract: Two-stage post-training -- a Stage-1 warm-start (supervised fine-tuning, SFT, or on-policy distillation, OPD) followed by Stage-2 reinforcement learni

safetyarxiv-cs-ai
9 Jun 2026
Safety

Stain-Aware Wavelet Regularization for Instant Adversarial Purification in Histopathology

DGX agent

arXiv:2606.08745v1 Announce Type: new Abstract: Deep learning has become prevalent in computational pathology pipelines that support tasks such as cancer screening and digital pathology analysis. Howe

safetyarxiv-cs-cv
9 Jun 2026
Safety

STARIXNet: Multivariate and Multi-attribute Deep Learning Approach to Real-Time Resource Allocation in Cloud Platforms

DGX agent

arXiv:2606.07565v1 Announce Type: new Abstract: Intelligent scaling of microservices in cloud platforms is crucial for mitigating escalating compute costs while avoiding service disruptions. Current s

safetyarxiv-cs-lg
9 Jun 2026
Safety

State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space

DGX agent

arXiv:2601.04266v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are widely deployed in safety-critical embodied AI applications such as robotics. However, their complex m

safetyarxiv-cs-lg
9 Jun 2026
Safety

Steganography Without Modification: Hidden Communication via LLM Seeds

DGX agent

arXiv:2606.09135v1 Announce Type: cross Abstract: We demonstrate that widely deployed Large Language Model (LLM) inference stacks harbor a steganographic channel that requires no modification to model

safetyarxiv-cs-ai
9 Jun 2026
Safety

STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for Long-Tailed Species Distribution Modeling

DGX agent

arXiv:2606.08484v1 Announce Type: cross Abstract: Joint Species Distribution Modeling (JSDM) is a key enabler for biodiversity monitoring and conservation planning. However, accurate JSDM faces two co

safetyarxiv-cs-ai
9 Jun 2026
Safety

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning

DGX agent

arXiv:2606.08735v1 Announce Type: new Abstract: Quality-diversity reinforcement learning (QD-RL) aims to construct policy repertoires that contain both high-performing and behaviorally diverse policie

safetyarxiv-cs-ai
9 Jun 2026
Safety

Summarization is Not Dead Yet

DGX agent

arXiv:2606.08000v1 Announce Type: cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries rival or even surpass human-written references, raising

safetyarxiv-cs-ai
9 Jun 2026
Safety

Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models

DGX agent

arXiv:2606.08451v1 Announce Type: cross Abstract: Safety-aligned large language models often exhibit sycophancy, which is the tendency to affirm users' opinions regardless of factual accuracy. Althoug

safetyarxiv-cs-ai
9 Jun 2026
Safety

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

DGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

safetyarxiv-cs-ai
9 Jun 2026
Safety

SynthICL: Scalable In-context Imitation Learning with Synthetic Data

DGX agent

arXiv:2606.08154v1 Announce Type: new Abstract: In-context imitation learning (ICIL) enables robots to learn new tasks from a small number of demonstrations by conditioning a pre-trained policy on tas

safetyarxiv-cs-ro
9 Jun 2026
Safety

Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks

DGX agent

arXiv:2606.08738v1 Announce Type: cross Abstract: Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication,

safetyarxiv-cs-ro
9 Jun 2026
Safety

Targeting World Models to Compromise Robot Learning Pipelines

DGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

safetyarxiv-cs-ai
9 Jun 2026
Safety

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

DGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

safetyarxiv-cs-ai
9 Jun 2026
Safety

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

DGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Confidence Trap: Calibration Attacks for Graph Neural Networks

DGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

DGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

DGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

DGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

DGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Spectral Dynamics and Noise Geometry of Muon

DGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

safetyarxiv-cs-lg
9 Jun 2026
Safety

Think Before You Act: Intention-Guided Reasoning for LLM-Based Location Prediction

DGX agent

arXiv:2606.08122v1 Announce Type: new Abstract: Predicting a user's next Point-of-Interest (POI) based on their historical check-in records is a fundamental task in location-based services. While rece

safetyarxiv-cs-ai
9 Jun 2026
Safety

This is actually a pretty measured bet compared to spending of the U.S companies. OpenAI alone is about twice as much. If things falls apart…

DGX agent

This is actually a pretty measured bet compared to spending of the U.S companies. OpenAI alone is about twice as much. If things falls apart, the US will be hit harder. 🚨BREAKING: CHINA IS AGI-PILLED

safetygary-marcus--x
9 Jun 2026
Safety

This quote from OpenAI is telling. Translation: We have a rapidly closing window to get this IPO out the door before the bubble bursts, our …

DGX agent

This quote from OpenAI is telling. Translation: We have a rapidly closing window to get this IPO out the door before the bubble bursts, our CFO doesn't want to go through with it because our books are

safetygary-marcus--x
9 Jun 2026
Safety

TinyJudge: Unverifiable Constraint Alignment via Lightweight Specialist Ensembles

DGX agent

arXiv:2606.07520v1 Announce Type: cross Abstract: Instruction Following (IF) is a core capability of LLMs, requiring strict adherence to diverse constraints, ranging from verifiable ones (e.g., output

safetyarxiv-cs-lg
9 Jun 2026
Safety

TORL-VLA: Tactile Guided Online Reinforcement Learning for Contact-Rich Manipulation

DGX agent

arXiv:2606.09337v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a powerful framework for robotic manipulation, and recent studies have introduced tactile or force feedb

safetyarxiv-cs-ro
9 Jun 2026
Safety

Toward autocorrection of chemical process flowsheets using large language models

DGX agent

arXiv:2312.02873v2 Announce Type: replace-cross Abstract: The process engineering domain widely uses Process Flow Diagrams (PFDs) and Process and Instrumentation Diagrams (P&IDs) to represent process

safetyarxiv-cs-ai
9 Jun 2026
Safety

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

DGX agent

arXiv:2606.08566v1 Announce Type: new Abstract: Emotional Video Captioning (EVC) is a challenging task that aims to generate factually accurate and emotionally rich descriptions for videos. Existing E

safetyarxiv-cs-cv
9 Jun 2026
Safety

Towards End to End Motion Planning and Execution for Autonomous Underwater Vehicles Using Reinforcement Learning

DGX agent

arXiv:2606.08513v1 Announce Type: cross Abstract: Autonomous Underwater Vehicles (AUVs) traditionally rely on complex, heavily engineered pipelines for perception, path planning, and motion control. T

safetyarxiv-cs-lg
9 Jun 2026
Safety

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings

DGX agent

arXiv:2511.05017v2 Announce Type: replace Abstract: Hallucinations in Large Vision-Language Models (LVLMs) remain a persistent challenge, often stemming from inadequate integration of visual informati

safetyarxiv-cs-cv
9 Jun 2026
Safety

TRACER: Token ReAssignment for Concept ERasure in Generative Recommendation

DGX agent

arXiv:2606.07688v1 Announce Type: cross Abstract: Generative recommendation formulates next-item prediction as autoregressive generation over semantic ID (SID) sequences derived from users' historical

safetyarxiv-cs-ai
9 Jun 2026
Safety

Training-Inference Kernel Contracts: Bounding Divergence in Post-Training and Deployment

DGX agent

arXiv:2606.07581v1 Announce Type: cross Abstract: A modern post-training pipeline often writes one symbol for its policy, pi_theta, while evaluating it through two different programs: a training kerne

safetyarxiv-cs-ai
9 Jun 2026
Safety

Trait-space Monitoring for Emergent Misalignment During Supervised Finetuning

DGX agent

arXiv:2606.07631v1 Announce Type: cross Abstract: Emergent misalignment (EM) occurs when narrow finetuning causes a model to behave dangerously outside the finetuning task. Standard training signals c

safetyarxiv-cs-ai
9 Jun 2026
Safety

Transforming Police-Car Swerving for Mitigating Isolated Stop-and-Go Traffic Waves: A Practice-Oriented Jam-Absorption Driving Strategy

DGX agent

arXiv:2602.10234v3 Announce Type: replace-cross Abstract: Stop-and-go traffic waves, a major form of freeway congestion, impose severe and persistent adverse impacts, including reduced traffic efficie

safetyarxiv-cs-ai
9 Jun 2026
Safety

TRUST-SCF: Transformer-based Risk Understanding and Scoring for Transactional Supply Chain Finance

DGX agent

arXiv:2606.08140v1 Announce Type: new Abstract: Supply Chain Finance (SCF) and LendTech platforms need credit scoring systems that respond to evolving transaction behavior, repayment delays, and activ

safetyarxiv-cs-lg
9 Jun 2026
Safety

Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data

DGX agent

arXiv:2606.08520v1 Announce Type: new Abstract: Vision-language models (VLMs) are powerful general-purpose reasoners, yet converting them into robot control policies (VLAs) is surprisingly difficult.

safetyarxiv-cs-ro
9 Jun 2026
Safety

Uncertainty-Aware Hierarchical Re-Localization in OpenStreetMap via Semantic Alignment

DGX agent

arXiv:2603.01613v2 Announce Type: replace Abstract: Monocular re-localization enables robots to estimate camera poses from visual observations. However, many existing methods rely on dense maps or lar

safetyarxiv-cs-cv
9 Jun 2026
Safety

Unifying Object-Centric World Models and Diffusion Policy: A Hierarchical Framework for Multi-Stage Robotic Tasks

DGX agent

arXiv:2606.08775v1 Announce Type: cross Abstract: Visual world models have shown great potential in learning complex system dynamics. Recent advancements leverage these models as transition functions

safetyarxiv-cs-ai
9 Jun 2026
Safety

VAIC: Vision-Guided Humanoid Agile Object Interaction Control via Decoupled Commands

DGX agent

arXiv:2606.09286v1 Announce Type: new Abstract: Humanoid robots hold immense potential for real-world assistance, yet agile interaction with objects in unstructured environments demands tightly couple

safetyarxiv-cs-ro
9 Jun 2026
Safety

Vessel Traffic Flow Prediction on Sparse Data via Spatio-Temporal Graph Neural Networks with a Learnable Tweedie Head

DGX agent

arXiv:2606.07694v1 Announce Type: new Abstract: Accurate vessel traffic flow prediction is crucial for smart port operations and navigational safety. However, maritime traffic flow data are often high

safetyarxiv-cs-lg
9 Jun 2026
Safety

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

DGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

safetyarxiv-cs-ai
9 Jun 2026
Safety

Video Understanding by Design: How Datasets Shape Video Models

DGX agent

arXiv:2509.09151v2 Announce Type: replace-cross Abstract: Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While exi

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…103104105106107…267
Next →