AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
1 Jul 2026

3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance

SafetyDGX agent

arXiv:2606.31329v1 Announce Type: cross Abstract: Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve generalization in robot manipulation.

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most random…

SafetyDGX agent

A basic sigmoid extrapolation suggests that this measurement will be close to its performance ceiling in around a year---meaning most randomly sampled remote work projects would be highly automatable

A Theory of How Pretraining Shapes Inductive Bias in Fine-Tuning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.20062v2 Announce Type: replace Abstract: Pretraining and fine-tuning are central stages in modern machine learning systems. In practice, feature learning plays an important role across both

AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation

SafetyDGX agent

arXiv:2606.31204v1 Announce Type: new Abstract: Synthetic data generation has emerged as a powerful tool for improving data scalability in computer vision. Recent diffusion-based pipelines have demons

ADAPT: Attention Dynamics Alignment with Preference Tuning for Faithful MLLMs

SafetyDGX agent

arXiv:2606.31054v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are critically hampered by hallucination, generating content inconsistent with the provided image. In this pa

Adapting Generalist Robot Policies with Semantic Reinforcement Learning

SafetyDGX agent

arXiv:2606.31958v1 Announce Type: new Abstract: Generalist robot policies learn a diverse repertoire of behaviors from large-scale pretraining. In principle, this makes them excellent priors for downs

AeroVerse-SatAgent: UAV-Satellite Collaborative Spatial Reasoning Inspired by the Dual Visual Pathway Theory of Cognitive Neuroscience

SafetyDGX agent

arXiv:2606.31467v1 Announce Type: new Abstract: With the rapid advancement of aerospace embodied intelligence, enabling Unmanned Aerial Vehicles (UAVs) to autonomously understand and reason about comp

AETDICE: Unified Framework and Offline Optimization for Nonlinear Multi-Objective RL

SafetyDGX agent

arXiv:2606.31178v1 Announce Type: cross Abstract: Optimizing nonlinear preferences in multi-objective reinforcement learning (MORL) is essential for capturing complex trade-offs like risk aversion or

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model

SafetyDGX agent

arXiv:2606.19100v2 Announce Type: replace Abstract: Large Vision and Language Models (LVLMs) have advanced rapidly, yet European Portuguese (pt-PT) remains systematically underserved by existing open-

Analysis of Atomic Charge State and Atomic Number for VAMOS++ Magnetic Spectrometer using Deep Neural Networks and Fractionally Labelled Events

SafetyDGX agent

arXiv:2507.07109v2 Announce Type: cross Abstract: The VAMOS++ magnetic spectrometer is a multi-parametric system that integrates ion optical magnetic elements with a multi-detector stack. The magnetic

Anchoring on Reality: Breaking the Pseudo-Target Ceiling in Makeup Transfer

SafetyDGX agent

arXiv:2606.31089v1 Announce Type: new Abstract: Makeup transfer applies a reference cosmetic style to a source face while preserving its identity and geometry. However, this task is severely hindered

And as we say in our blog, we're continuing to refine these safeguards to better distinguish genuine misuse from legitimate requests and red…

ToolsDGX agent

This post discusses ongoing efforts to improve AI safety safeguards, specifically refining systems to accurately differentiate between genuine misuse and legitimate user requests while addressing edge

AR-CoPO: Align Autoregressive Video Generation with Contrastive Policy Optimization

SafetyDGX agent

arXiv:2603.17461v2 Announce Type: replace Abstract: Streaming autoregressive (AR) video generators combined with few-step distillation achieve low-latency, high-quality synthesis, yet remain difficult

Ask the World Before Acting: Budgeted Environment Probing for World-Model Calibration

SafetyDGX agent

arXiv:2606.31422v1 Announce Type: new Abstract: Long-horizon language agents do not only choose actions; they carry a private model of the world from one decision to the next. When that model drifts,

Autonomous UAV Navigation for Individual Wildlife Re-Identification

SafetyDGX agent

arXiv:2606.31772v1 Announce Type: new Abstract: Reliable individual re-identification (re-ID) of wildlife is essential for population monitoring, behavioral tracking, and conservation policy evaluatio

AVTok: 1D Unified Tokenization for Holistic Audio-Video Generation

SafetyDGX agent

arXiv:2606.30811v1 Announce Type: new Abstract: Audio-video generation has recently gained unprecedented research attention, aiming to synthesize high-quality sounding video content with fine-grained

Behavior Cloning is Not All You Need: The Optimality of On-Policy Distillation for Noisy Expert Feedback

SafetyDGX agent

arXiv:2606.30923v1 Announce Type: cross Abstract: Imitation Learning is a natural framework for learning in sequential decision-making systems and has emerged as the dominant paradigm through which we

Beyond the Expressivity-Trainability Paradox: A Dynamical Lie Algebra Perspective on Navigating Barren Plateaus in Quantum Machine Learning

SafetyDGX agent

arXiv:2606.31536v1 Announce Type: new Abstract: As Quantum Machine Learning (QML) transitions toward practical implementation, the field faces a critical architectural bottleneck that challenges the f

BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding

SafetyDGX agent

arXiv:2606.31315v1 Announce Type: new Abstract: Speculative decoding accelerates inference by using a lightweight draft model to generate candidate tokens in parallel, and are then verified by the tar

BP-TTA: Balanced and Prototype-Guided Test-Time Adaptation in Dynamic Scenarios

SafetyDGX agent

arXiv:2606.31420v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) enables models trained on a source domain to adapt online to unlabeled test data under distribution shifts. While recent TTA

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

SafetyDGX agent

arXiv:2606.31825v1 Announce Type: cross Abstract: Recent multimodal large language models have shown great promise in clinical image reasoning, but existing post-training pipelines remain predominantl

Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration

SafetyDGX agent

arXiv:2601.19506v4 Announce Type: replace Abstract: Blind face restoration remains a persistent challenge due to the inherent ill-posedness of reconstructing holistic structures from severely constrai

Can Tabular In-Context Learners Generalize to Biomolecular Property Prediction?

SafetyDGX agent

arXiv:2606.31126v1 Announce Type: new Abstract: Predicting biomolecular properties from limited labeled data is a central bottleneck in protein engineering and small-molecule design. As strong pretrai

CasaMaestro: Multi-View Panoramas for House-Scale 3D Reconstruction

SafetyDGX agent

arXiv:2606.31086v1 Announce Type: new Abstract: The rise of home-deployed embodied AI systems is driving a growing need for fast, metric 3D reconstruction of residential spaces to support navigation,

ChronoFlow-Policy: Unifying Past-Current-Future Interaction Flow in Visuomotor Policy Learning

SafetyDGX agent

arXiv:2606.31493v1 Announce Type: new Abstract: Visual signals play a crucial role in policy learning by enabling models to capture object motion and interaction dynamics. Just as humans reason about

CLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical Reasoning

SafetyDGX agent

arXiv:2606.31608v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong results on many medical benchmarks, but their clinical reasoning remains difficult to evaluate reliably. A c

CLOUDADV: Decision-Aligned Instance Sizing with Zero-Shot Foundation Models under Drift

SafetyDGX agent

arXiv:2606.31470v1 Announce Type: new Abstract: Cloud virtual machines are often overprovisioned, creating avoidable cost and operational inefficiency. We present CLOUDADV, an interactive engineer-fac

Corruption Robust Offline Reinforcement Learning with Human Feedback

SafetyDGX agent

arXiv:2402.06734v2 Announce Type: replace-cross Abstract: We study data corruption robustness for reinforcement learning with human feedback (RLHF) in an offline setting. Given an offline dataset of p

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers

SafetyDGX agent

arXiv:2606.32020v1 Announce Type: new Abstract: Modern one-step diffusion models achieve impressive quality through distribution-based timestep distillation. Yet, they rely on a critical assumption: T

DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation

SafetyDGX agent

arXiv:2606.31537v1 Announce Type: new Abstract: Text-rich image generation is one of the most challenging settings in image generation, since models must simultaneously produce visually realistic imag

Deep Reinforcement Learning for Spacecraft Attitude Control During Atmospheric Re-Entry

SafetyDGX agent

arXiv:2606.31291v1 Announce Type: new Abstract: Deep reinforcement learning has the potential to solve attitude control problems more adaptively, precisely, and robustly by handling nonlinear dynamics

Designing Privacy-Preserving Visual Perception for Robot Navigation Based on User Privacy Preferences

SafetyDGX agent

arXiv:2604.06382v2 Announce Type: replace Abstract: Visual navigation is a fundamental capability of mobile service robots, yet the onboard cameras required for such navigation can capture privacy-sen

DISCOVER: A Solver for Distributional Counterfactual Explanations

SafetyDGX agent

arXiv:2603.16436v2 Announce Type: replace Abstract: Counterfactual explanations (CE) explain model decisions by identifying input modifications that lead to different predictions. Most existing method

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

SafetyDGX agent

arXiv:2606.31650v1 Announce Type: cross Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Existing cont

Efficient Sim-to-Real Transfer of World-Action Models from Synthetic Priors

SafetyDGX agent

arXiv:2606.31101v1 Announce Type: new Abstract: Bridging the sim-to-real gap is a core challenge in deploying learned manipulation policies. Sim-to-real learning is attractive because it can replace e

EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning

SafetyDGX agent

arXiv:2511.18242v3 Announce Type: replace Abstract: Egocentric video understanding requires procedural reasoning under partial observability and continuously shifting viewpoints. Current multimodal la

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies

SafetyDGX agent

arXiv:2606.31132v1 Announce Type: new Abstract: Generative control policies (GCPs), such as diffusion policies and flow-based vision-language-action models, enable test-time scaling in robot control.

End-to-End Efficient RL for Linear Bellman Complete MDPs with Deterministic Transitions

SafetyDGX agent

arXiv:2603.23461v2 Announce Type: replace Abstract: We study reinforcement learning (RL) with linear function approximation in Markov Decision Processes (MDPs) satisfying linear Bellman completeness -

ERA: Entropy-Guided Visual Token Pruning with Rectified Attention for Efficient MLLMs

SafetyDGX agent

arXiv:2606.31982v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) incur prohibitive inference costs due to long visual token sequences. Training-free visual token reduction prov

Evil Spectra: How Optimisers can Amplify or Suppress Emergent Misalignment

SafetyDGX agent

arXiv:2606.31591v1 Announce Type: cross Abstract: Emergent misalignment (EM) is a recently discovered phenomenon in LLMs where fine-tuning on a narrow misaligned task, such as writing insecure code, l

Evo-PI: Aligning Medical Reasoning via Evolving Principle-Guided Supervision

SafetyDGX agent

arXiv:2606.31800v1 Announce Type: new Abstract: Despite recent progress, the reasoning capabilities of large multimodal language models (MLLMs) remain fundamentally constrained by static supervision,

Explainable Artificial Intelligence For The Detection and Characterisation of Stage B Heart Failure

SafetyDGX agent

arXiv:2606.30665v1 Announce Type: cross Abstract: Stage B heart failure is characterized by asymptomatic structural or functional cardiac abnormalities. Identifying individuals at this stage is clinic

ExPLoRe: Expert Patch-Level Loss Routing for Multi-Objective Masked Image Modeling

SafetyDGX agent

arXiv:2606.31201v1 Announce Type: new Abstract: Multi-objective masked image modeling (MIM) combines complementary learning signals (token distillation, CLS alignment, and pixel reconstruction) but ex

FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion

SafetyDGX agent

arXiv:2606.31691v1 Announce Type: new Abstract: Scalable reinforcement learning has popularized high-throughput sampling architectures, which significantly compresses the training time for off-policy

Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction

SafetyDGX agent

arXiv:2603.09930v2 Announce Type: replace Abstract: Text-motion retrieval aims to learn a semantically aligned latent space between natural language descriptions and 3D human motion skeleton sequences

Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models

SafetyDGX agent

arXiv:2603.12893v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a standard technique for post-training diffusion-based image synthesis models, as it enables learning f

From Failure to Alignment: A Requirements Engineering Framework for Machine Learning Systems

SafetyDGX agent

arXiv:2606.31589v1 Announce Type: cross Abstract: Organisations designing, developing, and deploying machine learning systems (MLS) need to be able to check that these systems are trustworthy, and com

From Propositional to Perceptual Asymmetry: Extending Frictive Policy Optimization to Asymmetric Partial Information Dialogue

SafetyDGX agent

arXiv:2606.30973v1 Announce Type: new Abstract: Frictive Policy Optimization (FPO; Pustejovsky et al., 2025) treats friction in collaborative dialogue -- misalignment, misunderstanding, repair -- as a

G2P: Gaussian-to-Point Attribute Alignment for Boundary-Aware 3D Segmentation

SafetyDGX agent

arXiv:2601.03510v3 Announce Type: replace Abstract: Point cloud segmentation is critical for 3D scene understanding. However, sparse and irregular point distributions provide limited appearance eviden

GEAR: Guided End-to-End AutoRegression for Image Synthesis

SafetyDGX agent

arXiv:2606.32039v1 Announce Type: new Abstract: Visual generative models are typically trained in two stages. A tokenizer is first trained for reconstruction and then frozen, after which a generator i

GR2 Technical Report

SafetyDGX agent

arXiv:2606.31984v1 Announce Type: cross Abstract: Industrial recommendation systems serve billions of users through a multi-stage funnel -- retrieval, early-stage ranking, and re-ranking -- where the

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

SafetyDGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

HistoriQA-ThirdRepublic: Multi-Hop Question Answering Corpus for Historical Research, Parliamentary Debates from the French Third Republic (1870-1940)

SafetyDGX agent

arXiv:2606.31325v1 Announce Type: new Abstract: We present HistoriQA-ThirdRepublic: a French-language dataset of multi-hop historical questions derived from parliamentary debates and newspapers of the

How Human Feedback Shapes AI-generated Community Notes

SafetyDGX agent

arXiv:2606.30905v1 Announce Type: cross Abstract: Community Notes, a bridging-based crowd-sourced fact-checking system, has emerged as a new mechanism for moderating misleading information on social m

Incentive Aware AI Regulations: A Credal Characterisation

SafetyDGX agent

arXiv:2603.05175v2 Announce Type: replace Abstract: The rapid proliferation of AI applications has intensified debate on effective regulation of these black-box services. Effective regulation must bal

InfiniVerse: Occupancy Guided Unbounded Scene Generation for Autonomous Driving

SafetyDGX agent

arXiv:2606.31109v1 Announce Type: new Abstract: Generating realistic, controllable, and temporally coherent urban environments is a critical yet unresolved challenge in the autonomous driving communit

IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing

SafetyDGX agent

arXiv:2606.13368v2 Announce Type: replace Abstract: Computer-Aided Design is pivotal in modern manufacturing, yet existing automated methods predominantly rely on open-loop, one-shot generation, creat

Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents

SafetyDGX agent

arXiv:2606.12634v2 Announce Type: replace-cross Abstract: Long-horizon tool-use reinforcement learning learns from outcome verification, but trajectory-level advantages are broadcast over reasoning, A

LaMP: Learning Vision-Language-Action Policy with 3D Scene Flow as Latent Motion Prior

SafetyDGX agent

arXiv:2603.25399v2 Announce Type: replace Abstract: We introduce extbf{LaMP}, a dual-expert Vision-Language-Action framework that embeds dense 3D scene flow as a latent motion prior for robotic manipu

Language-Assisted Super-Resolution from Real-World Low-Resolution Patches

SafetyDGX agent

arXiv:2606.31363v1 Announce Type: new Abstract: Single image super-resolution aims to reconstruct high-resolution (HR) images from low-resolution (LR) inputs. Training SR models typically requires pai

← Previous
1…9293949596…242
Next →