AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,494 results
Safety

sGPO: Trading Inference FLOPs for Training Efficiency in RLVR

DGX agent

arXiv:2606.08854v1 Announce Type: cross Abstract: Standard Reinforcement Learning with Verifiable Rewards (RLVR) training allocates a fixed rollout budget to every query, without regard for what each

safetyarxiv-cs-ai
9 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SMI: Efficient Self-Supervised Learning via Mutual-Information-Inspired Dependency Optimization

DGX agent

arXiv:2606.08332v1 Announce Type: new Abstract: Self-supervised learning (SSL) has achieved remarkable representation learning performance, but many existing methods rely on large batch sizes, memory

safetyarxiv-cs-cv
9 Jun 2026
Safety

'So There's a Catch-22 Here': How Early Adopters Who Build Multi-Agent LLM Systems Conceptualize Transparency

DGX agent

arXiv:2606.08323v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems are rapidly emerging, yet transparency, a cornerstone of responsible AI, remains under-defined in these

safetyarxiv-cs-ai
9 Jun 2026
Safety

some news: it turns out 34,000 Instagram accounts got hit by a Meta AI exploit the other day, per internal company docs hackers also used a …

DGX agent

some news: it turns out 34,000 Instagram accounts got hit by a Meta AI exploit the other day, per internal company docs hackers also used a senior 'Space Force' official's account to post anti-Iran wa

safetygary-marcus--x
9 Jun 2026
Safety

SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning

DGX agent

arXiv:2606.08992v1 Announce Type: cross Abstract: Vision-and-Language Navigation in continuous environments requires agents to understand the spatial structure of previously unseen environments in ord

safetyarxiv-cs-ai
9 Jun 2026
Safety

Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models

DGX agent

arXiv:2606.08446v1 Announce Type: cross Abstract: Despite being powerful, reinforcement learning with verifiable rewards (RLVR) induces extremely long COT, making it computationally expensive. Since R

safetyarxiv-cs-ai
9 Jun 2026
Safety

Speaker-Invariant Representation Learning for Spoofing Detection via Gradient Reversal and A Variational Information Bottleneck

DGX agent

arXiv:2606.08678v1 Announce Type: cross Abstract: Sophisticated generative speech technology can undermined the reliability of voice biometrics. While spoofing detection systems excel when assessed un

safetyarxiv-cs-lg
9 Jun 2026
Safety

SPIN: Decentralized Swarm Control via Tensorized Policy Coordination

DGX agent

arXiv:2606.07557v1 Announce Type: new Abstract: Decentralized multi-agent swarm coordination on resource-constrained edge platforms remains fundamentally bottlenecked by the exponential scaling of joi

safetyarxiv-cs-lg
9 Jun 2026
Safety

Stage-1 Controls the Entropy Regime, Not the Outcome

DGX agent

arXiv:2606.09059v1 Announce Type: cross Abstract: Two-stage post-training -- a Stage-1 warm-start (supervised fine-tuning, SFT, or on-policy distillation, OPD) followed by Stage-2 reinforcement learni

safetyarxiv-cs-ai
9 Jun 2026
Safety

STARIXNet: Multivariate and Multi-attribute Deep Learning Approach to Real-Time Resource Allocation in Cloud Platforms

DGX agent

arXiv:2606.07565v1 Announce Type: new Abstract: Intelligent scaling of microservices in cloud platforms is crucial for mitigating escalating compute costs while avoiding service disruptions. Current s

safetyarxiv-cs-lg
9 Jun 2026
Safety

Steganography Without Modification: Hidden Communication via LLM Seeds

DGX agent

arXiv:2606.09135v1 Announce Type: cross Abstract: We demonstrate that widely deployed Large Language Model (LLM) inference stacks harbor a steganographic channel that requires no modification to model

safetyarxiv-cs-ai
9 Jun 2026
Safety

STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for Long-Tailed Species Distribution Modeling

DGX agent

arXiv:2606.08484v1 Announce Type: cross Abstract: Joint Species Distribution Modeling (JSDM) is a key enabler for biodiversity monitoring and conservation planning. However, accurate JSDM faces two co

safetyarxiv-cs-ai
9 Jun 2026
Safety

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning

DGX agent

arXiv:2606.08735v1 Announce Type: new Abstract: Quality-diversity reinforcement learning (QD-RL) aims to construct policy repertoires that contain both high-performing and behaviorally diverse policie

safetyarxiv-cs-ai
9 Jun 2026
Safety

Summarization is Not Dead Yet

DGX agent

arXiv:2606.08000v1 Announce Type: cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries rival or even surpass human-written references, raising

safetyarxiv-cs-ai
9 Jun 2026
Safety

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

DGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

safetyarxiv-cs-ai
9 Jun 2026
Safety

SynthICL: Scalable In-context Imitation Learning with Synthetic Data

DGX agent

arXiv:2606.08154v1 Announce Type: new Abstract: In-context imitation learning (ICIL) enables robots to learn new tasks from a small number of demonstrations by conditioning a pre-trained policy on tas

safetyarxiv-cs-ro
9 Jun 2026
Safety

Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks

DGX agent

arXiv:2606.08738v1 Announce Type: cross Abstract: Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication,

safetyarxiv-cs-ro
9 Jun 2026
Safety

Targeting World Models to Compromise Robot Learning Pipelines

DGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

safetyarxiv-cs-ai
9 Jun 2026
Safety

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

DGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

safetyarxiv-cs-ai
9 Jun 2026
Safety

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

DGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

DGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

DGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

safetyarxiv-cs-ai
9 Jun 2026
Safety

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

DGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Spectral Dynamics and Noise Geometry of Muon

DGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

safetyarxiv-cs-lg
9 Jun 2026
Safety

Think Before You Act: Intention-Guided Reasoning for LLM-Based Location Prediction

DGX agent

arXiv:2606.08122v1 Announce Type: new Abstract: Predicting a user's next Point-of-Interest (POI) based on their historical check-in records is a fundamental task in location-based services. While rece

safetyarxiv-cs-ai
9 Jun 2026
Safety

This is actually a pretty measured bet compared to spending of the U.S companies. OpenAI alone is about twice as much. If things falls apart…

DGX agent

This is actually a pretty measured bet compared to spending of the U.S companies. OpenAI alone is about twice as much. If things falls apart, the US will be hit harder. 🚨BREAKING: CHINA IS AGI-PILLED

safetygary-marcus--x
9 Jun 2026
Safety

This quote from OpenAI is telling. Translation: We have a rapidly closing window to get this IPO out the door before the bubble bursts, our …

DGX agent

This quote from OpenAI is telling. Translation: We have a rapidly closing window to get this IPO out the door before the bubble bursts, our CFO doesn't want to go through with it because our books are

safetygary-marcus--x
9 Jun 2026
Safety

TinyJudge: Unverifiable Constraint Alignment via Lightweight Specialist Ensembles

DGX agent

arXiv:2606.07520v1 Announce Type: cross Abstract: Instruction Following (IF) is a core capability of LLMs, requiring strict adherence to diverse constraints, ranging from verifiable ones (e.g., output

safetyarxiv-cs-lg
9 Jun 2026
Safety

TORL-VLA: Tactile Guided Online Reinforcement Learning for Contact-Rich Manipulation

DGX agent

arXiv:2606.09337v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a powerful framework for robotic manipulation, and recent studies have introduced tactile or force feedb

safetyarxiv-cs-ro
9 Jun 2026
Safety

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

DGX agent

arXiv:2606.08566v1 Announce Type: new Abstract: Emotional Video Captioning (EVC) is a challenging task that aims to generate factually accurate and emotionally rich descriptions for videos. Existing E

safetyarxiv-cs-cv
9 Jun 2026
Safety

Towards End to End Motion Planning and Execution for Autonomous Underwater Vehicles Using Reinforcement Learning

DGX agent

arXiv:2606.08513v1 Announce Type: cross Abstract: Autonomous Underwater Vehicles (AUVs) traditionally rely on complex, heavily engineered pipelines for perception, path planning, and motion control. T

safetyarxiv-cs-lg
9 Jun 2026
Safety

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings

DGX agent

arXiv:2511.05017v2 Announce Type: replace Abstract: Hallucinations in Large Vision-Language Models (LVLMs) remain a persistent challenge, often stemming from inadequate integration of visual informati

safetyarxiv-cs-cv
9 Jun 2026
Safety

Training-Inference Kernel Contracts: Bounding Divergence in Post-Training and Deployment

DGX agent

arXiv:2606.07581v1 Announce Type: cross Abstract: A modern post-training pipeline often writes one symbol for its policy, pi_theta, while evaluating it through two different programs: a training kerne

safetyarxiv-cs-ai
9 Jun 2026
Safety

Trait-space Monitoring for Emergent Misalignment During Supervised Finetuning

DGX agent

arXiv:2606.07631v1 Announce Type: cross Abstract: Emergent misalignment (EM) occurs when narrow finetuning causes a model to behave dangerously outside the finetuning task. Standard training signals c

safetyarxiv-cs-ai
9 Jun 2026
Safety

TRUST-SCF: Transformer-based Risk Understanding and Scoring for Transactional Supply Chain Finance

DGX agent

arXiv:2606.08140v1 Announce Type: new Abstract: Supply Chain Finance (SCF) and LendTech platforms need credit scoring systems that respond to evolving transaction behavior, repayment delays, and activ

safetyarxiv-cs-lg
9 Jun 2026
Safety

Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data

DGX agent

arXiv:2606.08520v1 Announce Type: new Abstract: Vision-language models (VLMs) are powerful general-purpose reasoners, yet converting them into robot control policies (VLAs) is surprisingly difficult.

safetyarxiv-cs-ro
9 Jun 2026
Safety

Uncertainty-Aware Hierarchical Re-Localization in OpenStreetMap via Semantic Alignment

DGX agent

arXiv:2603.01613v2 Announce Type: replace Abstract: Monocular re-localization enables robots to estimate camera poses from visual observations. However, many existing methods rely on dense maps or lar

safetyarxiv-cs-cv
9 Jun 2026
Safety

Unifying Object-Centric World Models and Diffusion Policy: A Hierarchical Framework for Multi-Stage Robotic Tasks

DGX agent

arXiv:2606.08775v1 Announce Type: cross Abstract: Visual world models have shown great potential in learning complex system dynamics. Recent advancements leverage these models as transition functions

safetyarxiv-cs-ai
9 Jun 2026
Safety

VAIC: Vision-Guided Humanoid Agile Object Interaction Control via Decoupled Commands

DGX agent

arXiv:2606.09286v1 Announce Type: new Abstract: Humanoid robots hold immense potential for real-world assistance, yet agile interaction with objects in unstructured environments demands tightly couple

safetyarxiv-cs-ro
9 Jun 2026
Safety

Video Understanding by Design: How Datasets Shape Video Models

DGX agent

arXiv:2509.09151v2 Announce Type: replace-cross Abstract: Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While exi

safetyarxiv-cs-ai
9 Jun 2026
Safety

Vision-Language Asymmetry in Bistable Image Captioning

DGX agent

arXiv:2606.08031v1 Announce Type: new Abstract: Wittgenstein's duck-rabbit poses a question for vision-language models: when a model captions an ambiguous image, where in the model is the commitment t

safetyarxiv-cs-cv
9 Jun 2026
Safety

Visual Para-Thinker++: A Single-Policy Multi-Agent Framework for Visual Reasoning

DGX agent

arXiv:2606.09290v1 Announce Type: new Abstract: Visual reasoning requires integrating evidence distributed across regions, attributes, and relations, making single-chain reasoning prone to early perce

safetyarxiv-cs-cv
9 Jun 2026
Safety

wait til you see what happens with SpaceX.

DGX agent

Gary Marcus, a prominent AI researcher and critic, posted a teaser statement on X about upcoming SpaceX developments, likely referring to significant announcements or milestones in the company's space

safetygary-marcus--x
9 Jun 2026
Safety

WaveDiT: Distribution-Aware Wavelet Flow Matching for Efficient 3D Brain MRI Synthesis

DGX agent

arXiv:2606.08670v1 Announce Type: new Abstract: Large and demographically balanced datasets are essential for reliable neuroimaging biomarkers. Full-resolution 3D brain MRI synthesis can support data

safetyarxiv-cs-cv
9 Jun 2026
Safety

Web Agents Should Use Typed Actions Instead of Click-Based Browsing

DGX agent

arXiv:2602.17245v2 Announce Type: replace Abstract: This position paper argues that building a reliable agentic Web requires shifting from low-level interaction primitives to typed actions supported b

safetyarxiv-cs-ai
9 Jun 2026
Safety

what constitutes reasoning in AI is a critical debate. i hope that @dwarkesh_sp will respond.

DGX agent

what constitutes reasoning in AI is a critical debate. i hope that @dwarkesh_sp will respond. what does this even mean, @dwarkesh_sp, “the real deal”? is it even a falsifiable conjecture? what’s the e

safetygary-marcus--x
9 Jun 2026
Safety

what does this even mean, @dwarkesh_sp, “the real deal”? is it even a falsifiable conjecture? what’s the evidence? and if you can agree that…

DGX agent

what does this even mean, @dwarkesh_sp, “the real deal”? is it even a falsifiable conjecture? what’s the evidence? and if you can agree that “If you ask an LLM a question it can't answer, sometimes it

safetygary-marcus--x
9 Jun 2026
← Previous
1…147148149150151…302
Next →