AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

Incentivizing Temporal-Awareness in Egocentric Video Understanding Models

DGX agent

Multimodal large language models (MLLMs) have recently shown strong performance in visual understanding, yet they often lack temporal awareness, particularly in egocentric settings where reasoning dep

safetyapple-ml-research
9 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

InductWave: Inductive Multi-Hop Logical Query Answering on Knowledge Graphs

DGX agent

arXiv:2607.07422v1 Announce Type: new Abstract: Logical Multi-Hop Query Answering over Knowledge Graphs (KGs) can be formulated as querying, with an implicit completeness assumption. Current works mai

safetyarxiv-cs-ai
9 Jul 2026
Safety

Initiation Safety: A Missing Dimension in Generalist-Robot Safety

DGX agent

arXiv:2607.07420v1 Announce Type: new Abstract: Safety for generalist robots is usually discussed in terms of motion or dialogue. We argue a third question is missing: should the robot take its first

safetyarxiv-cs-ro
9 Jul 2026
Safety

Latent Policy Steering through One-Step Flow Policies

DGX agent

arXiv:2603.05296v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) allows robots to learn from offline datasets without risky exploration. Yet, offline RL's performance ofte

safetyarxiv-cs-lg
9 Jul 2026
Safety

Learning social norms enhances compatibility in dynamic human-AI coordination

DGX agent

arXiv:2607.07021v1 Announce Type: new Abstract: Humans continuously coordinate with others in dynamic interactions, often through implicit, hard-to-quantify social norms that act as shared tacit expec

safetyarxiv-cs-ai
9 Jul 2026
Safety

LHM-Humanoid: Long-Horizon Human Motion Control for Continuous Object Transport in Cluttered Scenes

DGX agent

arXiv:2508.16943v3 Announce Type: replace-cross Abstract: Physics-based human motion control can make a simulated character walk, sit, and manipulate objects with high physical realism. Almost always,

safetyarxiv-cs-ai
9 Jul 2026
Safety

Lipschitz-Regularized Critics Lead to Policy Robustness Against Transition Dynamics Uncertainty

DGX agent

arXiv:2404.13879v5 Announce Type: replace Abstract: Uncertainties in transition dynamics pose a critical challenge in reinforcement learning (RL), often resulting in performance degradation of trained

safetyarxiv-cs-lg
9 Jul 2026
Safety

LipSSD: Lipschitz-Constrained Single-Shot Detection for Adversarially Robust Object Detection

DGX agent

arXiv:2607.06592v1 Announce Type: cross Abstract: Object detectors have many applications in safety-critical systems, but they are known to be sensitive to worst-case perturbations such as adversarial

safetyarxiv-cs-ai
9 Jul 2026
Safety

LLM-Guided Task-Semantic Field Factorization for Industrial Process Forecasting

DGX agent

arXiv:2607.06623v1 Announce Type: cross Abstract: Process industries rely on time-series forecasting and soft sensing to estimate quality variables that are hard to measure online. Labeled data are sc

safetyarxiv-cs-ai
9 Jul 2026
Safety

LLM-powered reasoning in agent-based modeling

DGX agent

arXiv:2607.06757v1 Announce Type: new Abstract: Agent-based modeling (ABM) has the capability to model millions of individuals and their interactions, which is useful for policy making. However, ABMs

safetyarxiv-cs-ai
9 Jul 2026
Safety

LLMs Silently Correct African American English: Auditing and Mitigating Dialect Bias via Activation Steering

DGX agent

arXiv:2607.06845v1 Announce Type: new Abstract: African American English (AAE), a rule-governed dialect spoken by over 30 million people, is routinely misinterpreted and 'corrected' by large language

safetyarxiv-cs-cl
9 Jul 2026
Safety

Manual, Joystick, or Haptic Control? An In Vitro Comparison of Navigation Strategies for Robotic Interventional Neuroradiology Procedures

DGX agent

arXiv:2607.07253v1 Announce Type: new Abstract: Objective: To evaluate robotic controller interfaces for interventional neuroradiology procedures in-vitro incorporating a force-sensing platform to ass

safetyarxiv-cs-ro
9 Jul 2026
Safety

Mathematical methods of reinforcement learning

DGX agent

arXiv:2607.06935v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly grounded in tools from probability, optimization, and operator theory. This survey organizes the mathemati

safetyarxiv-cs-lg
9 Jul 2026
Safety

Max Out GRPO Signal: Adaptive Trace Prefix Control for Hard Reasoning Problems

DGX agent

arXiv:2607.07674v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) stalls on a model's hardest problems: when no rollout in a group succeeds, the group-relative advantages van

safetyarxiv-cs-cl
9 Jul 2026
Safety

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning

DGX agent

arXiv:2607.07316v1 Announce Type: new Abstract: This article offers a comprehensive overview of mechanistic interpretability, an emerging field that seeks to reverse-engineer the internal algorithms o

safetyarxiv-cs-lg
9 Jul 2026
Safety

Microsoft President Brad Smith says the US now has AI 'regulation without transparent or complete rules', and adds that 'without rules, businesses can't plan' (Beatrice Nolan/Fortune)

DGX agent

Beatrice Nolan / Fortune: Microsoft President Brad Smith says the US now has AI “regulation without transparent or complete rules”, and adds that “without rules, businesses can't plan” — Microsoft Pre

safetytechmeme
9 Jul 2026
Safety

Multi-Agent AI Control: Distributed Attacks Hamper Per-Instance Monitors

DGX agent

arXiv:2607.07368v1 Announce Type: cross Abstract: AI control is a family of techniques to prevent an AI with malicious goals from subverting its operator's intent. AI Control usually studies a single

safetyarxiv-cs-ai
9 Jul 2026
Safety

NativeMEM: Native Memory Compression for Long-Horizon Robotic Manipulation

DGX agent

arXiv:2607.06678v1 Announce Type: new Abstract: How can pretrained Vision-Language-Action (VLA) models retain long-horizon visual histories with high-frequency updates without sacrificing efficiency?

safetyarxiv-cs-ro
9 Jul 2026
Safety

NonTextual Target Attack

DGX agent

arXiv:2510.02999v5 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks on Large Language Models (LLMs) typically optimize adversarial suffixes to align the LLM output with

safetyarxiv-cs-ai
9 Jul 2026
Safety

Online Data Selection Is Implicit Alignment

DGX agent

arXiv:2607.07023v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is often treated as a capability-adaptation step, while alignment is attributed to later preference optimization or reinfor

safetyarxiv-cs-lg
9 Jul 2026
Safety

Open-Ended Scenario Reasoning for Specialist Model Adaptation

DGX agent

arXiv:2607.06625v1 Announce Type: cross Abstract: Process industries have accumulated validated specialist models, yet sensor drift, feedstock variation, and regime switching cause these models to deg

safetyarxiv-cs-ai
9 Jul 2026
Safety

ORAN-DEFEND: Subspace Detection and Sanitization of Backdoor DRL xApps in Open RAN

DGX agent

arXiv:2607.06647v1 Announce Type: cross Abstract: Open Radio Access Networks (O-RAN) increasingly delegate near-real-time control to deep reinforcement learning (DRL) xApps obtained from third-party v

safetyarxiv-cs-lg
9 Jul 2026
Safety

ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies

DGX agent

arXiv:2607.07235v1 Announce Type: cross Abstract: Explainability remains a key issue in reinforcement learning (RL). Distilling an interpretable policy from an agent trained in a complex environment i

safetyarxiv-cs-ai
9 Jul 2026
Safety

PB-OEL: A Performance-Bounded Online Ensemble Learning Framework With Mixed Feedback for Real-Time Safety Assessment

DGX agent

arXiv:2503.15581v2 Announce Type: replace Abstract: Real-time safety assessment is critical for ensuring the reliable operation of complex dynamic systems. However, obtaining full safety labels in rea

safetyarxiv-cs-lg
9 Jul 2026
Safety

Pelican-VLA 0.5: Attending Before Acting Benefits Generalization

DGX agent

arXiv:2607.06655v1 Announce Type: cross Abstract: In this report, we present Pelican-VLA 0.5, a unified VLA model that integrates vision-language understanding, future-frame generation, and action pre

safetyarxiv-cs-lg
9 Jul 2026
Safety

Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production

DGX agent

arXiv:2607.07052v1 Announce Type: cross Abstract: AI agents deployed for IT operations are typically permanent cost centers because every execution requires full LLM inference, even for previously sol

safetyarxiv-cs-ai
9 Jul 2026
Safety

Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees

DGX agent

arXiv:2405.16668v2 Announce Type: replace Abstract: Adversarial Imitation Learning (AIL) faces challenges with sample inefficiency because of its reliance on sufficient on-policy data to evaluate the

safetyarxiv-cs-lg
9 Jul 2026
Safety

R^3: Advertisement Compliance Rectification via Group-Relative Experience Extractor and Curriculum Reinforcement

DGX agent

arXiv:2607.07318v1 Announce Type: new Abstract: Rigorous content moderation is crucial for online advertising but leads to millions of daily rejections. This scale renders manual rectification infeasi

safetyarxiv-cs-cl
9 Jul 2026
Safety

Rapidly Learning Soft Robot Control via Implicit Time-Stepping

DGX agent

arXiv:2511.06667v2 Announce Type: replace-cross Abstract: With the explosive growth of rigid-body simulators, policy learning in simulation has become the de facto standard for most rigid morphologies

safetyarxiv-cs-ai
9 Jul 2026
Safety

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops

DGX agent

arXiv:2607.07663v1 Announce Type: new Abstract: AI systems increasingly participate in their own improvement: revising their outputs, adapting their own harnesses during deployment, training on data t

safetyarxiv-cs-ai
9 Jul 2026
Safety

Residual-Conservative Model Predictive Path Integral Control

DGX agent

arXiv:2607.06950v1 Announce Type: cross Abstract: Sampling-based model predictive control methods handle nonlinear dynamics and complex cost landscapes through Monte Carlo rollouts, yet typically empl

safetyarxiv-cs-ro
9 Jul 2026
Safety

RoboSnap: One-Shot Real-to-Sim Scene Generation for Generalizable Robot Learning and Evaluation

DGX agent

arXiv:2607.06699v1 Announce Type: new Abstract: Recovering real-world scenes as interactive simulation environments can enable generalizable robot learning and reproducible policy evaluation. However,

safetyarxiv-cs-ro
9 Jul 2026
Safety

Safe Reinforcement Learning using Ideas from Model Predictive Control

DGX agent

arXiv:2607.07252v1 Announce Type: new Abstract: Reinforcement learning (RL) enables the synthesis of control policies directly from data, making it highly appealing for complex cyber-physical systems

safetyarxiv-cs-lg
9 Jul 2026
Safety

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

DGX agent

arXiv:2607.07675v1 Announce Type: new Abstract: Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For e

safetyarxiv-cs-cv
9 Jul 2026
Safety

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

DGX agent

arXiv:2607.07693v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful paradigm for aligning generative models with human preferences. However, a

safetyarxiv-cs-ai
9 Jul 2026
Safety

SHTA: Semantic Hard Token Correction and Center Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2607.07019v1 Announce Type: new Abstract: Recent advances in semi-supervised medical image segmentation have achieved remarkable performance through prediction consistency, pseudo-label supervis

safetyarxiv-cs-cv
9 Jul 2026
Safety

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2607.07508v1 Announce Type: cross Abstract: Reinforcement learning (RL) is becoming increasingly important for post-training large language models (LLMs). Previous RL pipelines for LLMs were mos

safetyarxiv-cs-ai
9 Jul 2026
Safety

SOMtime the World Ain't Fair: Violating Fairness Using Self-Organizing Maps

DGX agent

arXiv:2602.18201v2 Announce Type: replace Abstract: Unsupervised representations are widely assumed to be neutral with respect to sensitive attributes when those attributes are withheld from training.

safetyarxiv-cs-ai
9 Jul 2026
Safety

Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction

DGX agent

arXiv:2601.17216v3 Announce Type: replace-cross Abstract: Intelligent Transportation Systems (ITS) demand real-time collision prediction to ensure road safety and reduce accident severity. Conventiona

safetyarxiv-cs-ai
9 Jul 2026
Safety

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs

DGX agent

arXiv:2511.16107v3 Announce Type: replace-cross Abstract: Visual in-context learning (VICL) solves visual tasks by conditioning on a few input-output demonstrations without any model training. Recent

safetyarxiv-cs-ai
9 Jul 2026
Safety

TACoS: Weakly Supervised Learning of Two-Dimensional Materials from Scribble Annotations to Precise Segmentation

DGX agent

arXiv:2607.07169v1 Announce Type: new Abstract: The precise pixel-level localization of 2D material flakes is crucial for high-throughput screening. However, traditional fully supervised methods rely

safetyarxiv-cs-cv
9 Jul 2026
Safety

The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents

DGX agent

arXiv:2607.07436v1 Announce Type: new Abstract: A self-evolving agent retires its bad skills by watching them fail, so what happens when the judge cannot see the failures? Skill retirement is the stru

safetyarxiv-cs-ai
9 Jul 2026
Safety

The New York Times and several news outlets are accusing OpenAI of hiding its ability to search its AI training data and output logs for mor…

DGX agent

The New York Times and several news outlets are accusing OpenAI of hiding its ability to search its AI training data and output logs for more than two years during an ongoing copyright lawsuit: • The

safetygary-marcus--x
9 Jul 2026
Safety

The Power of Backdoor Absorption in Community Training

DGX agent

arXiv:2607.06643v1 Announce Type: cross Abstract: Backdoor attacks severely threaten large-scale AI models. When model owners delegate training to external compute providers within a decentralized tra

safetyarxiv-cs-lg
9 Jul 2026
Safety

Thinking Seeds: Leveraging Historical Diversity for Position-Aware RL in LLMs

DGX agent

arXiv:2601.21476v2 Announce Type: replace Abstract: On-policy reinforcement learning (RL) for language model post-training suffers from a fundamental tension: as training progresses, policy entropy co

safetyarxiv-cs-cl
9 Jul 2026
Safety

This is a very important observation. AI is not adopting the language and shaping it to fit its known ideas. It has no known ideas. It is re…

DGX agent

This is a very important observation. AI is not adopting the language and shaping it to fit its known ideas. It has no known ideas. It is repeating patterns found in the language. super interesting -

safetygary-marcus--x
9 Jul 2026
Safety

TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration

DGX agent

arXiv:2603.27742v2 Announce Type: replace Abstract: Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing framework

safetyarxiv-cs-cv
9 Jul 2026
Safety

Towards a Pseudo-Labeling Workflow for Celltype-Classification from Explanted Brain Slice Recordings

DGX agent

arXiv:2607.06569v1 Announce Type: cross Abstract: This paper proposes an unsupervised workflow to pseudo-label extracellular spikes from human brain slice MEA recordings into two putative cell types:

safetyarxiv-cs-lg
9 Jul 2026
← Previous
1…4849505152…265
Next →