AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Probing the Origins of Reasoning Performance: Representational Quality for Mathematical Problem-Solving in RL vs. SFT Fine-Tuned Models

DGX agent

arXiv:2607.26119v1 Announce Type: cross Abstract: Large reasoning models trained via reinforcement learning (RL) have been increasingly shown to outperform their supervised fine-tuned (SFT) counterpar

safetyarxiv-cs-cl
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Quoting Bruce Schneier

DGX agent

The writing assignments I give my students are gym tasks, not work tasks. I ask them to write policy memos not because the world needs more policy memos. I assign them because the very act of writing,

safetysimon-willison
30 Jul 2026
Safety

R-SLPR: Region-based Small-to-Large Point-cloud Registration with Contrastive Learning

DGX agent

arXiv:2607.26583v1 Announce Type: new Abstract: Point-cloud (PC) registration is fundamental to three-dimensional (3D) perception in robotic systems. However, classic registration algorithms falter wh

safetyarxiv-cs-cv
30 Jul 2026
Safety

Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbreaks

DGX agent

arXiv:2607.26574v1 Announce Type: cross Abstract: Safety classifiers ('guards') are the dominant black-box defense for vision-language models, yet they judge an input's surface form, not its meaning:

safetyarxiv-cs-lg
30 Jul 2026
Safety

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D

DGX agent

arXiv:2607.19321v2 Announce Type: replace-cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be

safetyarxiv-cs-lg
30 Jul 2026
Safety

Rethinking Clinical Relevance in Chest X-ray Machine Learning: How Evaluation References Define Performance

DGX agent

arXiv:2607.26333v1 Announce Type: cross Abstract: Chest X-ray (CXR) machine learning relies heavily on automated evaluation using reference standards that aim to approximate clinical judgment. However

safetyarxiv-cs-cv
30 Jul 2026
Safety

Retrospective Orthogonal Design: Response-Surface Reconstruction from Observational Data

DGX agent

arXiv:2607.26219v1 Announce Type: cross Abstract: Regression estimates from observational data can depend on specification under multicollinearity, while sequential sums of squares (SS) depend on term

safetyarxiv-cs-lg
30 Jul 2026
Safety

RL^2-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models

DGX agent

arXiv:2607.26991v1 Announce Type: new Abstract: Despite the impressive visuomotor capabilities enabled by Vision-Language-Action (VLA) models, their performance often degrades on challenging and out-o

safetyarxiv-cs-ro
30 Jul 2026
Safety

RLMM-Flow: A Flow-based Mobile Manipulation Framework with Latent-Space Reinforcement Learning

DGX agent

arXiv:2607.26460v1 Announce Type: new Abstract: Mobile manipulation requires generating whole-body action chunks that jointly satisfy goal reaching, collision avoidance, base kinematic constraints, ma

safetyarxiv-cs-ro
30 Jul 2026
Safety

Robostreet Flow: A Lightweight, Ultra-Low-Drag Electric Tractor and Four-Truck Hybrid Convoy Architecture for Minimum-Cost Point-to-Point Freight

DGX agent

arXiv:2607.26250v1 Announce Type: new Abstract: Line-haul trucking costs are dominated by three comparably sized components: energy, driver labor, and equipment. Most efficiency technologies address o

safetyarxiv-cs-cl
30 Jul 2026
Safety

SCALPEL: Semantic Cross-modal Alignment via LLM-Powered Encoder Learning for Medical Vision-Language Representation

DGX agent

arXiv:2607.26885v1 Announce Type: new Abstract: Vision-language pre-training (VLP) serves as a cornerstone for medical multimodal representation learning. However, existing medical VLP frameworks are

safetyarxiv-cs-cv
30 Jul 2026
Safety

SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence

DGX agent

arXiv:2607.27066v1 Announce Type: new Abstract: Scientific figure assessment in peer review differs fundamentally from general image quality evaluation: a figure must be visually legible, faithfully s

safetyarxiv-cs-cv
30 Jul 2026
Safety

Searching for Robust Augmentations to Improve Out-of-Domain Generalization in Dermoscopic Skin Cancer Classification

DGX agent

arXiv:2607.26765v1 Announce Type: new Abstract: Background/Objectives: Dermoscopic skin lesion classifiers often lose accuracy under domain shift across imaging devices, illumination, and capture arti

safetyarxiv-cs-cv
30 Jul 2026
Safety

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

DGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

safetyarxiv-cs-lg
30 Jul 2026
Safety

Shape-Based Inductive Bias for Glioma Grading from Tumor Contours

DGX agent

arXiv:2607.26090v1 Announce Type: cross Abstract: Glioma grading from tumor contours is often treated as a pixel problem even when the signal of interest is shape. We align closed contours with a func

safetyarxiv-cs-cv
30 Jul 2026
Safety

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from …

DGX agent

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from investors after suffering heavy losses during the recent rou

safetygary-marcus--x
30 Jul 2026
Safety

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

DGX agent

arXiv:2607.26784v1 Announce Type: new Abstract: Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. Yet standard agentic reinforcement learnin

safetyarxiv-cs-lg
30 Jul 2026
Safety

SMSP: A Plug-and-Play Strategy of Multi-Scale Perception for MLLMs to Perceive Visual Illusions

DGX agent

arXiv:2603.23118v2 Announce Type: replace Abstract: Recent works have shown that multimodal large language models (MLLMs) are highly vulnerable to hidden-pattern visual illusions, where the hidden con

safetyarxiv-cs-cv
30 Jul 2026
Safety

Stable and Budget-Feasible Coalition Formation for Clustered Federated Learning: A Hedonic Potential-Game Approach

DGX agent

arXiv:2607.26788v1 Announce Type: cross Abstract: Clustered federated learning benefits from organizing heterogeneous participants into coalitions that train coalition-specific models, but such cluste

safetyarxiv-cs-lg
30 Jul 2026
Safety

Steering Instruction Hierarchies at Inference Time

DGX agent

arXiv:2607.26228v1 Announce Type: new Abstract: Instruction hierarchies are a core safety assumption of language model deployment: higher priority inputs, such as system prompts, should override confl

safetyarxiv-cs-cl
30 Jul 2026
Safety

SymmGrid: Super-Scaling On-Robot Learning with Parallelized Symmetries and Egocentric-Exocentric Visual Perception

DGX agent

arXiv:2607.26985v1 Announce Type: cross Abstract: Deep reinforcement policy learning directly in physical robots (on-robot learning) remains bottlenecked by slow wall-clock training times. We present

safetyarxiv-cs-lg
30 Jul 2026
Safety

The Advantage of Fine-Grained Training

DGX agent

arXiv:2509.05130v2 Announce Type: replace Abstract: In classification problems, models are trained to predict a class label based on the input data features. However, class labels are organized hierar

safetyarxiv-cs-lg
30 Jul 2026
Safety

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that lat…

DGX agent

// The agent is its own best speculator // Agents spend a large share of wall-clock time waiting on tool results. Speculation hides that latency by predicting and pre-executing the next call, but exte

safetydair-ai--x
30 Jul 2026
Safety

The Confounder Trap: Treatment-Encoding Representations in Causal Inference with Text

DGX agent

arXiv:2607.26309v1 Announce Type: cross Abstract: Estimating causal effects of linguistic properties from observational text is difficult because the same document can contain both the treatment of in

safetyarxiv-cs-cl
30 Jul 2026
Safety

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, man…

DGX agent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, many forefront members of the AI-Safety community, in their fer

safetyyann-lecun--x
30 Jul 2026
Safety

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models

DGX agent

arXiv:2607.26845v1 Announce Type: new Abstract: Inference-time thinking improves the performance of large language models, but aggregate outcomes do not reveal whether models use available evidence mo

safetyarxiv-cs-lg
30 Jul 2026
Safety

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs

DGX agent

arXiv:2607.27122v1 Announce Type: new Abstract: Gastrointestinal (GI) endoscopic image analysis has shifted from single-label classification toward visual question answering (VQA), where a model must

safetyarxiv-cs-cv
30 Jul 2026
Safety

TraceCLIP: Recovering Local Semantics from Patch-to-CLS Contributions

DGX agent

arXiv:2607.26107v1 Announce Type: new Abstract: Dense vision-language understanding, including object localization, region recognition, and open-vocabulary semantic segmentation, requires associating

safetyarxiv-cs-cv
30 Jul 2026
Safety

Veritas++: Value-aware On-Policy Distillation for Perception-Enhanced AIGI Detection

DGX agent

arXiv:2607.27113v1 Announce Type: new Abstract: The growing capability of image generation models has made synthetic images a routine presence in open media, making robust and generalizable AI-Generat

safetyarxiv-cs-cv
30 Jul 2026
Safety

Visibility-Aware Cooperative Tracking with Decentralized LiDAR-Based Aerial Swarms

DGX agent

arXiv:2512.01280v2 Announce Type: replace Abstract: Autonomous aerial tracking with drones offers vast potential for surveillance, cinematography, and industrial inspection applications. While single-

safetyarxiv-cs-ro
30 Jul 2026
Safety

Weak-to-Strong On-Policy Distillation

DGX agent

arXiv:2607.26246v1 Announce Type: new Abstract: On-policy distillation (OPD), which aligns a student with the teacher's token-level distribution on the student's own rollouts, is an effective paradigm

safetyarxiv-cs-lg
30 Jul 2026
Safety

WikiLoop: Jointly Learning to Build and Navigate Agent-Native Wikis with Downstream Feedback

DGX agent

arXiv:2607.26604v1 Announce Type: new Abstract: Knowledge-base construction and querying are typically optimized in isolation: retrieval-augmented agents operate over a fixed, externally maintained in

safetyarxiv-cs-cl
30 Jul 2026
Safety

A context-adaptive policy framework for robust and reactive robotic manipulation via uncertainty-aware imitation learning

DGX agent

arXiv:2410.24035v2 Announce Type: replace-cross Abstract: Generating robust and reactive manipulation strategies that can adapt to changing context information is a challenging task in robotics. Over

safetyarxiv-cs-ai
29 Jul 2026
Safety

A Functional Approach to Curve Alignment and Shape Analysis

DGX agent

arXiv:2503.05632v2 Announce Type: cross Abstract: In many image analysis problems, the contours of objects carry important statistical information about shape. Such contours are typically affected by

safetyarxiv-cs-cv
29 Jul 2026
Safety

A Unified Algorithmic Framework for Hybrid Reinforcement Learning in Tabular MDPs with Shifted Transition Dynamics

DGX agent

arXiv:2607.25207v1 Announce Type: new Abstract: This paper investigates a hybrid reinforcement learning setting in tabular Markov Decision Processes (MDPs), where an agent aims to learn an optimal pol

safetyarxiv-cs-lg
29 Jul 2026
Safety

A2D2: Audi Autonomous Driving Dataset

DGX agent

arXiv:2004.06320v2 Announce Type: replace Abstract: Research in machine learning, mobile robotics, and autonomous driving is accelerated by the availability of high quality annotated data. To this end

safetyarxiv-cs-cv
29 Jul 2026
Safety

AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning

DGX agent

arXiv:2607.24833v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards is a powerful paradigm for eliciting reasoning in large language models, yet it suffers from severe rewar

safetyarxiv-cs-ai
29 Jul 2026
Safety

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

DGX agent

arXiv:2607.25489v1 Announce Type: new Abstract: Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and

safetyarxiv-cs-cv
29 Jul 2026
Safety

AlphaCrafter: Harnessing Multi-Agent Workflows for Cross-Sectional Quantitative Trading

DGX agent

arXiv:2605.05580v2 Announce Type: replace Abstract: Quantitative trading agents have demonstrated substantial promise in automating factor discovery, signal aggregation, and portfolio execution. Howev

safetyarxiv-cs-ai
29 Jul 2026
Safety

Architectural Backdoors in Vision-Language Model Supply Chains via Representation Steering

DGX agent

arXiv:2607.25479v1 Announce Type: cross Abstract: Vision--Language Models (VLMs) are increasingly deployed through a model supply chain in which pretrained checkpoints, architecture definitions, text

safetyarxiv-cs-ai
29 Jul 2026
Safety

Behavior-Driven Explainability

DGX agent

arXiv:2607.24881v1 Announce Type: new Abstract: As system complexity has vastly increased, it has become significantly more challenging for a single person or a team to fully understand all aspects of

safetyarxiv-cs-lg
29 Jul 2026
Safety

Beyond Predictive Accuracy: A Reliability-Aware Audit of Molecular Representations for Human Olfaction

DGX agent

arXiv:2607.24848v1 Announce Type: cross Abstract: Pretrained molecular encoders are commonly evaluated through downstream prediction, but predictive accuracy alone does not establish that a learned re

safetyarxiv-cs-ai
29 Jul 2026
Safety

Beyond Self-Knowledge: Propagating Uncertainty Across Reasoning and Retrieval in LLMs

DGX agent

arXiv:2607.25600v1 Announce Type: cross Abstract: Retrieval-augmented generation improves knowledge-intensive question answering, but indiscriminate retrieval can introduce irrelevant evidence and unn

safetyarxiv-cs-ai
29 Jul 2026
Safety

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates

DGX agent

arXiv:2602.10430v2 Announce Type: replace-cross Abstract: Off-policy policy optimization reuses historical behavior, including negative-advantage samples that suppress known failures. We show that rep

safetyarxiv-cs-ai
29 Jul 2026
Safety

Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning

DGX agent

arXiv:2607.24996v1 Announce Type: cross Abstract: Neural networks are hindered by accumulating dormant neurons and loss of expressivity throughout training, particularly in non-stationary data setting

safetyarxiv-cs-ai
29 Jul 2026
Safety

CAST: Game Solvers as Turn-Level Teachers for LLM Agents

DGX agent

arXiv:2607.25308v1 Announce Type: cross Abstract: Training large language models (LLMs) to act in long-horizon games is a promising step toward generalist decision-making, yet reinforcement learning w

safetyarxiv-cs-ai
29 Jul 2026
Safety

CIFNet: An Analytic Neural Learning Framework for Efficient and Calibrated Class-Incremental Learning

DGX agent

arXiv:2509.11285v2 Announce Type: replace-cross Abstract: Class-Incremental Learning (CIL) in deep neural networks is conventionally framed as an iterative gradient-based optimization problem, incurri

safetyarxiv-cs-ai
29 Jul 2026
Safety

Construction-Driven Injection: Linguistically-Grounded Edit-Based Code-Mixing Fingerprints for Large Language Models

DGX agent

arXiv:2607.25633v1 Announce Type: cross Abstract: Large language models (LLMs) are costly intellectual assets that remain exposed to unauthorized redistribution and commercial misuse. Injected fingerp

safetyarxiv-cs-ai
29 Jul 2026
← Previous
1…2829303132…265
Next →