AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
6 Aug 2026

Agent Skills for Automated Reasoning policies in Amazon Bedrock

SafetyDGX agent

Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custo

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

SafetyDGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (…

SafetyDGX agent

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (i.e. a harness that calls neural models) or whether the oute

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Arnold: A multi-task, multi-embodiment muscle transformer policy

SafetyDGX agent

arXiv:2508.18066v2 Announce Type: replace-cross Abstract: Controlling high-dimensional and nonlinear musculoskeletal models of the human body is a foundational scientific challenge. Recent machine lea

ATLAS: Adaptive Topological Learning with Abstract Successors for Continual Learning

SafetyDGX agent

arXiv:2608.04334v1 Announce Type: cross Abstract: Contemporary model-free reinforcement learning algorithms can achieve very high performance, but have low sample efficiency and are not robust to chan

Attention Fusion for Bridge Deck Delamination Detection

SafetyDGX agent

arXiv:2512.20113v4 Announce Type: replace Abstract: Subsurface delaminations in reinforced concrete bridge decks escape conventional visual inspection, and the two principal sensing techniques used to

Beyond the Dirac Delta: Mitigating Diversity Collapse in Reinforcement Fine-Tuning for Versatile Image Generation

SafetyDGX agent

arXiv:2601.12401v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as a powerful paradigm for fine-tuning large-scale generative models, such as diffusion and flow model

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

SafetyDGX agent

arXiv:2608.05042v1 Announce Type: new Abstract: Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as a promising paradigm for 3D robot m

Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.04663v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often adds social terms to individual rewards, yet the scale of those terms is usually chosen by hand. We

Calibrating Transformer Attention via Task-Space Sensitivity Feedback

SafetyDGX agent

arXiv:2512.20661v2 Announce Type: replace Abstract: Transformer-based pre-trained language models (PLMs) excel in text classification but suffer from attention dilution and attention sink effects, for

CARGO-VL: Counterfactual Arbitration with Risk-Constrained Group Optimization for Vision-Language Models

SafetyDGX agent

arXiv:2608.04509v1 Announce Type: new Abstract: Vision-language systems combine images with retrieved text, but these sources can disagree or jointly fail to support an answer. Reliable models must id

CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models

SafetyDGX agent

arXiv:2608.04302v1 Announce Type: new Abstract: Benchmarking video-language models has largely focused on short clips and single-sentence metrics, leaving open whether current systems can generate acc

C’mon, it’s not game over for Google Seven reasons why not, excerpted from my newsletter:

SafetyDGX agent

C’mon, it’s not game over for Google Seven reasons why not, excerpted from my newsletter: Jeff Dean and Demis Hassabis are the two most important AI executives at Google. Jeff is leaving and Demis is

CofactVLA: Deconfounding Vision-Language-Action Models via Counterfactual Intervention

SafetyDGX agent

arXiv:2608.04396v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have driven significant progress in robotic manipulation, yet they fundamentally struggle with the vision-override p

Compass: Continuously Aligning Social Media Feeds via In-Situ Reflections

SafetyDGX agent

arXiv:2608.04274v1 Announce Type: cross Abstract: Social media recommendation feeds often optimize for users' immediate impulses rather than preferences they would hold after deeper reflection. Some s

Contrastive Diffusion Alignment: Learning Structured Latents for Controllable Generation

SafetyDGX agent

arXiv:2510.14190v3 Announce Type: replace Abstract: Diffusion models excel at generation, but their latent spaces are high dimensional and not explicitly organized for interpretation or control. We in

Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore

SafetyDGX agent

Learn about new capabilities in Amazon Bedrock AgentCore: temporal policies powered by Dogwood, a new open source policy language for AI agents, and rate limiting on the gateway. These features give y

COSMO: Consensus-Driven Shift Modulation for Source-Free Domain Adaptation

SafetyDGX agent

arXiv:2608.04604v1 Announce Type: new Abstract: Source-free domain adaptation (SFDA) adapts a source-trained model to an unlabeled target domain without source data, a practical setting under privacy

critical history and context that a lot of people have conveniently forgotten

SafetyDGX agent

critical history and context that a lot of people have conveniently forgotten For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with onl

DAC-Pose: Dual-Agent Collaborative Framework for Pose-Guided Human Generation

SafetyDGX agent

arXiv:2608.04622v1 Announce Type: new Abstract: AI agents have emerged as a powerful new paradigm in generative image synthesis, enabling systems to perform complex semantic reasoning rather than pass

Data-Aware and Scalable Sensitivity Analysis for Decision Tree Ensembles

SafetyDGX agent

arXiv:2602.07453v2 Announce Type: replace Abstract: Decision tree ensembles are widely used in critical domains, making robustness and sensitivity analysis essential to their trustworthiness. We study

Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints

SafetyDGX agent

arXiv:2608.04669v1 Announce Type: new Abstract: Many social services assign scarce resources, such as housing assistance or hospital interventions, to people who arrive one at a time: each arrival mus

DXC partners with Primary on zero-trust security for enterprise AI

SafetyDGX agent

DXC Technology Co. today announced a partnership with security startup Primary that makes the information technology services company the exclusive managed services partner for Primary’s zero-trust pl

Enabling Urgency-aware Robot Swarm Intralogistics using Smart IoT Tags

SafetyDGX agent

arXiv:2608.04721v1 Announce Type: new Abstract: Warehouse items differ in how urgently they must be moved: perishable goods, pharmaceutical shipments, and just-in-time production materials must be del

EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment

SafetyDGX agent

arXiv:2608.04472v1 Announce Type: cross Abstract: The development of foundation models (FMs) is crucial for advancing endoscopic image analysis. However, existing endoscopy FMs mainly rely on self-sup

Exact Model-Free Policy Iteration for Co-safe LTL Planning

SafetyDGX agent

arXiv:2608.05047v1 Announce Type: cross Abstract: This work studies model-free reinforcement learning for co-safe linear temporal logic (sc-LTL) objectives in finite Markov decision processes, which c

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

SafetyDGX agent

arXiv:2608.04771v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) excel on complex tasks through long chain-of-thought (CoT) reasoning, but their lengthy intermediate steps cause severe ov

Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation

SafetyDGX agent

arXiv:2602.19161v2 Announce Type: replace Abstract: Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers

FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory

SafetyDGX agent

arXiv:2608.04530v1 Announce Type: new Abstract: GUI agents must remember both useful experience from earlier tasks and unfinished progress in the current interaction. Latent memory offers a compact so

From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

SafetyDGX agent

arXiv:2601.03808v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved notable performance in code synthesis; however, data-aware augmentation remains a limiting factor, handle

Gary Marcus won! @GaryMarcus

SafetyDGX agent

Gary Marcus won! @GaryMarcus I would have assumed it was fairly obvious, but in case it's not: a million-line codebase (also known as a 'harness'), running at inference time, orchestrating thousands o

Generative Optimization for Incentivized Advertising with Global Level Constraints

SafetyDGX agent

arXiv:2608.04421v1 Announce Type: cross Abstract: Incentivized advertising allocates monetary or virtual rewards to drive user engagement, where a key challenge is optimizing continuous incentive magn

GeoReward: Mitigating Contextual Variable Overestimation in Vision-Language Models for Cross-Market Preference Prediction

SafetyDGX agent

arXiv:2608.04504v1 Announce Type: cross Abstract: Vision-language models excel in many multimodal tasks but remain prone to a subtle yet impactful failure mode: they tend to overestimate dominant visu

GFlowNet Training by Policy Gradients

SafetyDGX agent

arXiv:2408.05885v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) have been shown effective to generate combinatorial objects with desired properties. We here propose a new GFlo

Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming

SafetyDGX agent

arXiv:2608.04018v1 Announce Type: cross Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to per

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning

Model ReleasesDGX agent

arXiv:2608.05045v1 Announce Type: cross Abstract: Released aligned large language models remain vulnerable to malicious downstream finetuning. Existing defenses are largely designed for the fine-tunin

HALT: Verification-Aware Stopping for Retrieval-Augmented Search Agents

SafetyDGX agent

arXiv:2608.02009v2 Announce Type: replace Abstract: Retrieval-augmented search agents answer multi-hop questions by repeatedly issuing search queries and accumulating evidence. This creates a stopping

HCRide: Harmonizing Passenger Fairness and Driver Preference for Human-Centered Ride-Hailing

SafetyDGX agent

arXiv:2508.04811v2 Announce Type: replace Abstract: Order dispatch systems play a vital role in ride-hailing services, which directly influence operator revenue, driver profit, and passenger experienc

Interpreting GFlowNets for Drug Discovery: What probes can and cannot show

SafetyDGX agent

arXiv:2511.19264v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) construct molecules through sequential decisions, but their internal policies remain opaque, limiting ado

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does t…

SafetyDGX agent

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does the observer observe itself and know that it changed for the

Joint UAV Flight and Opportunistic Routing under Reinforcement Learning for Delay-Tolerant Networks

SafetyDGX agent

arXiv:2608.04590v1 Announce Type: new Abstract: The growing deployment of delay-tolerant networks (DTNs) has made store-carry-forward (SCF) communication indispensable under sparse connectivity. Howev

Language Models Generalize to Human-like Word Order Preferences

SafetyDGX agent

arXiv:2608.05028v1 Announce Type: new Abstract: A central question in language acquisition is whether linguistic biases can emerge from general learning mechanisms operating over underdetermined input

Learning to Resolve Neutron Resonances with Fully Convolutional Neural Networks

SafetyDGX agent

arXiv:2608.04027v1 Announce Type: new Abstract: This work investigates the feasibility of augmenting traditional R-Matrix codes with a robust machine learning framework for automatically detecting neu

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control

SafetyDGX agent

arXiv:2608.05084v1 Announce Type: new Abstract: Diffusion policies are a powerful policy class for continuous control, but their iterative denoising process creates a substantial computational bottlen

Lesion Detection in CT with Frozen Self-Distilled Features: SALT, a Spatially Adaptive Label-Guided Temperature

SafetyDGX agent

arXiv:2608.05100v1 Announce Type: new Abstract: Self-supervised pretraining objectives are spatially uniform: the teacher temperature and the per-patch loss weight are identical everywhere in the imag

LLMs Struggle to Measure What Distinguishes Students of Different Proficiency Levels: A Study of Item Discrimination in Reading Comprehension Assessment

SafetyDGX agent

arXiv:2606.18709v2 Announce Type: replace Abstract: Existing work on LLM-based educational assessment has focused largely on item difficulty, but difficulty alone does not indicate whether an item mea

Manipulation-Proof Oblivious Audits against Deceptive Model Providers

SafetyDGX agent

arXiv:2608.04365v1 Announce Type: new Abstract: Audits have emerged as a critical instrument for algorithmic governance, providing a mechanism for external scrutiny and governance of machine learning

MGSB: Manifold Gated Signature Branch Pressure-Domain Baseline Architecture for Two-Phase Pipeline Flows Under Distributional Shift

SafetyDGX agent

arXiv:2608.04805v1 Announce Type: new Abstract: Leak detection models for multiphase pipelines often degrade when deployed under flow regimes that differ from training. Existing evaluations typically

Multi-Objective Ranking for Live-Streaming: Balancing Fresh and Delayed Signals with Segment-Aware Targeting

SafetyDGX agent

arXiv:2608.04455v1 Announce Type: cross Abstract: One of the most challenging problems entertainment live-streaming services face in recommendation systems is that user behaviors are sparse and delaye

Multicalibration Yields Better Matchings

SafetyDGX agent

arXiv:2511.11413v2 Announce Type: replace Abstract: Consider the problem of finding the best matching in a weighted graph where we only have access to predictions of the actual stochastic weights, bas

Multimodal Alignment Through Joint Kernel Entropic Gromov--Wasserstein Optimal Transport

SafetyDGX agent

arXiv:2608.04234v1 Announce Type: cross Abstract: We study the problem of aligning data from multiple modalities into a shared representation space, focusing on settings where strong pretrained unimod

Non-Stationary Inventory Control with Lead Times

SafetyDGX agent

arXiv:2602.05799v2 Announce Type: replace-cross Abstract: We study non-stationary single-item, periodic-review inventory control problems in which the demand distribution is unknown and may change ove

Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles

SafetyDGX agent

arXiv:2608.04483v1 Announce Type: new Abstract: Vision-language models (VLMs) process an image as a sequence of visual tokens, which creates a substantial computational bottleneck during inference. Re

Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation

SafetyDGX agent

arXiv:2608.04408v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises student-visited trajectories, yet divergence-based rules cannot determine whether an erroneous prefix remains

ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance

SafetyDGX agent

arXiv:2608.04524v1 Announce Type: new Abstract: Synthetic generation of Cognitive Behavioral Therapy (CBT) sessions is challenged by two competing demands: adhering to strict therapeutic structure whi

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school…

SafetyDGX agent

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school me on the definition 🤦‍♂️ coding harness and tools calls ar

OPD-V: Visual On-Policy Self-Distillation with Modality Balance

SafetyDGX agent

arXiv:2608.05131v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has become a standard post-training approach for improving visual reasoning in multimodal large language models (ML

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

SafetyDGX agent

arXiv:2608.05080v1 Announce Type: cross Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods

OutLangSplat: 3D Language Gaussian Splatting for UAV Outdoor Scenes

SafetyDGX agent

arXiv:2608.04560v1 Announce Type: new Abstract: 3D Language Gaussian Splatting embeds open-vocabulary language features into 3D Gaussian Splatting, providing an efficient explicit representation for t

Overcoming Statistical Bias in Action-Controllable World Models

SafetyDGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

← Previous
1…5960616263…240
Next →