AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

Change My View? The Dynamics of Persuasion and Polarization in Online Discourse

DGX agent

arXiv:2605.08383v1 Announce Type: new Abstract: Philosophical accounts of persuasion often assume that shared evidence and rational argumentation should lead to a convergence of views between peers, y

safetyarxiv-cs-cl
12 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

CoLA-Flow Policy: Temporally Coherent Imitation Learning via Continuous Latent Action Flow Matching for Robotic Manipulation

DGX agent

arXiv:2601.23087v3 Announce Type: replace Abstract: Learning long-horizon robotic manipulation requires jointly achieving expressive behavior modeling, real-time inference, and stable execution, which

safetyarxiv-cs-ro
12 May 2026
Safety

Composing Policy Gradients and Prompt Optimization for Language Model Programs

DGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

safetyarxiv-cs-cl
12 May 2026
Safety

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision

DGX agent

arXiv:2509.14234v3 Announce Type: replace Abstract: Where do learning signals come from when there is no ground truth in post-training? We show that inference compute itself can serve as supervision.

safetyarxiv-cs-lg
12 May 2026
Safety

Compute Where it Counts: Self Optimizing Language Models

DGX agent

arXiv:2605.10875v1 Announce Type: cross Abstract: Efficient LLM inference research has largely focused on reducing the cost of each decoding step (e.g., using quantization, pruning, or sparse attentio

safetyarxiv-cs-cl
12 May 2026
Safety

Continuity Laws for Sequential Models

DGX agent

arXiv:2605.08539v1 Announce Type: cross Abstract: Inductive biases influence the behavior and performance of sequential models. In this work, we study an underexplored inductive bias in sequential mod

safetyarxiv-cs-ai
12 May 2026
Safety

Control-Augmented Autoregressive Diffusion for Data Assimilation

DGX agent

arXiv:2510.06637v3 Announce Type: replace-cross Abstract: Despite advances in test-time scaling and diffusion finetuning, guidance for Auto-Regressive Diffusion Models (ARDMs) remains underexplored. W

safetyarxiv-cs-ai
12 May 2026
Safety

Core-Halo Decomposition: Decentralizing Large-Scale Fixed-Point Problems

DGX agent

arXiv:2605.08681v1 Announce Type: cross Abstract: We study solving large-scale fixed-point equation (x^star=ar F(x^star)) with decomposition. Standard strict decomposition assigns each agent a disjoin

safetyarxiv-cs-ai
12 May 2026
Safety

Cornerstones or Stumbling Blocks? Deciphering the Rock Tokens in On-Policy Distillation

DGX agent

arXiv:2605.09253v1 Announce Type: cross Abstract: While recent work in Reinforcement Learning with Verifiable Rewards (RLVR) has shown that a small subset of critical tokens disproportionately drives

safetyarxiv-cs-ai
12 May 2026
Safety

Crosslingual On-Policy Self-Distillation for Multilingual Reasoning

DGX agent

arXiv:2605.09548v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in mathematical reasoning, but this ability is not equally accessible across languages. E

safetyarxiv-cs-cl
12 May 2026
Safety

Crowding Out The Noise: Algorithmic Collective Action Under Differential Privacy

DGX agent

arXiv:2505.05707v2 Announce Type: replace Abstract: The integration of AI into daily life has generated considerable attention and excitement, while also raising concerns about automating algorithmic

safetyarxiv-cs-lg
12 May 2026
Safety

DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition

DGX agent

arXiv:2605.09604v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar provides privacy-preserving sensing and is valuable for human action recognition (HAR). Existing mmWave point cloud datas

safetyarxiv-cs-cv
12 May 2026
Safety

DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models

DGX agent

arXiv:2605.08902v1 Announce Type: cross Abstract: In recent years, pre-trained visual-linguistic models have demonstrated tremendous potential, becoming a crucial foundational framework for numerous d

safetyarxiv-cs-ai
12 May 2026
Safety

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation

DGX agent

arXiv:2605.09188v1 Announce Type: cross Abstract: Reinforcement learning improves the reasoning ability of large language models but remains costly and sample-inefficient, as many rollouts provide wea

safetyarxiv-cs-ai
12 May 2026
Safety

Data-driven transport modelling without overfit

DGX agent

arXiv:2605.08801v1 Announce Type: new Abstract: Macroscopic transport modelling aims to predict traffic flows after proposed public policy interventions, such as a new road or railway section or a tem

safetyarxiv-cs-lg
12 May 2026
Safety

Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents

DGX agent

arXiv:2605.08717v1 Announce Type: cross Abstract: Software engineering agents are increasingly deployed in evaluable engineering environments, yet post-failure recovery remains costly, manual, and ad

safetyarxiv-cs-ai
12 May 2026
Safety

Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2605.08389v1 Announce Type: cross Abstract: Zero-shot composed image retrieval (ZS-CIR) retrieves a target image from a reference image and a text modification without human-annotated CIR triple

safetyarxiv-cs-ai
12 May 2026
Safety

Dependency-Aware Discrete Diffusion for Scene Graph Generation

DGX agent

arXiv:2605.09065v1 Announce Type: new Abstract: Scene graphs (SGs) represent objects and their relationships as structured graphs, enabling applications in image generation, robotics, and 3D understan

safetyarxiv-cs-cv
12 May 2026
Safety

DexWrist: A Robotic Wrist for Constrained and Dynamic Manipulation

DGX agent

arXiv:2507.01008v3 Announce Type: replace Abstract: Development of dexterous manipulation hardware has primarily focused on hands and grippers. However, these end-effectors are often paired with bulky

safetyarxiv-cs-ro
12 May 2026
Safety

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

DGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

safetyarxiv-cs-lg
12 May 2026
Safety

DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization

DGX agent

arXiv:2605.10863v1 Announce Type: new Abstract: Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consist

safetyarxiv-cs-cl
12 May 2026
Safety

Disentangled Representation Learning via Flow Matching

DGX agent

arXiv:2602.05214v2 Announce Type: replace Abstract: Disentangled representation learning aims to capture the underlying explanatory factors of observed data, enabling a principled understanding of the

safetyarxiv-cs-lg
12 May 2026
Safety

Do Linear Probes Generalize Better in Persona Coordinates?

DGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

safetyarxiv-cs-ai
12 May 2026
Safety

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs

DGX agent

arXiv:2605.10281v1 Announce Type: cross Abstract: Generating realistic drum audio directly from symbolic representations is a challenging task at the intersection of music perception and machine learn

safetyarxiv-cs-ai
12 May 2026
Safety

DuetFair: Coupling Inter- and Intra-Subgroup Robustness for Fair Medical Image Segmentation

DGX agent

arXiv:2605.10521v1 Announce Type: cross Abstract: Medical image segmentation models can perform unevenly across subgroups. Most existing fairness methods focus on improving average subgroup performanc

safetyarxiv-cs-ai
12 May 2026
Safety

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning

DGX agent

arXiv:2605.10923v1 Announce Type: cross Abstract: Large language model agents increasingly rely on external skills to solve complex tasks, where skills act as modular units that extend their capabilit

safetyarxiv-cs-cl
12 May 2026
Safety

E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability

DGX agent

arXiv:2605.10261v1 Announce Type: new Abstract: TCAV (Testing with Concept Activation Vectors) is an interpretability method that assesses the alignment between the internal representations of a train

safetyarxiv-cs-ai
12 May 2026
Safety

Effective Explanations Support Planning Under Uncertainty

DGX agent

arXiv:2605.08406v1 Announce Type: cross Abstract: Explaining how to get from A to B can be challenging. It requires mentally simulating what the listener will do based on what they are told. To captur

safetyarxiv-cs-ai
12 May 2026
Safety

EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents

DGX agent

arXiv:2605.10366v1 Announce Type: new Abstract: Graph reasoning agents operating from natural-language inputs must solve a coupled problem: they must reconstruct a structured graph instance from text,

safetyarxiv-cs-ai
12 May 2026
Safety

ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation

DGX agent

arXiv:2605.08799v1 Announce Type: new Abstract: Diffusion policies have demonstrated exceptional performance in embodied AI. However, their iterative denoising process results in high latency, and exi

safetyarxiv-cs-ro
12 May 2026
Safety

Emergence of Physical Intelligence via Controllable Information Production

DGX agent

arXiv:2601.22449v2 Announce Type: replace Abstract: Intrinsic Motivation (IM) aims to train agents without external rewards, enabling useful behavior to emerge from the agent's interaction with its en

safetyarxiv-cs-ai
12 May 2026
Safety

Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning

DGX agent

arXiv:2605.09395v1 Announce Type: new Abstract: In this paper, we propose the first VLnderline{extbf{M}} nderline{extbf{a}}gentic nderline{extbf{r}}easoning framework for few-nderline{extbf{s}}hot mul

safetyarxiv-cs-ai
12 May 2026
Safety

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

DGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

safetyarxiv-cs-lg
12 May 2026
Safety

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

DGX agent

arXiv:2601.21484v2 Announce Type: replace Abstract: Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complic

safetyarxiv-cs-lg
12 May 2026
Safety

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

DGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

safetyarxiv-cs-ai
12 May 2026
Safety

EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent

DGX agent

arXiv:2605.09777v1 Announce Type: cross Abstract: Gradient-based preference optimization methods for large language model (LLM) alignment suffer from preference collapse, converging to narrow behavior

safetyarxiv-cs-ai
12 May 2026
Safety

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

DGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

safetyarxiv-cs-ai
12 May 2026
Safety

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

DGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

safetyarxiv-cs-ai
12 May 2026
Safety

Explanation-Aware Learning for Enhanced Interpretability in Biomedical Imaging

DGX agent

arXiv:2605.10054v1 Announce Type: new Abstract: Deep neural networks for medical image diagnosis often achieve high predictive accuracy while relying on spurious or clinically irrelevant visual cues,

safetyarxiv-cs-cv
12 May 2026
Safety

Explicit Stair Geometry Conditioning for Robust Humanoid Locomotion

DGX agent

arXiv:2605.09944v1 Announce Type: new Abstract: Robust humanoid stair climbing remains challenging due to geometric discontinuities, sensitivity to step height variations, and perception uncertainty i

safetyarxiv-cs-ro
12 May 2026
Safety

Exploration-Driven Optimization for Test-Time Large Language Model Reasoning

DGX agent

arXiv:2605.09853v1 Announce Type: new Abstract: Post-training techniques combined with inference-time scaling significantly enhance the reasoning and alignment capabilities of large language models (L

safetyarxiv-cs-lg
12 May 2026
Safety

Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality

DGX agent

arXiv:2605.10206v1 Announce Type: cross Abstract: Distributional causal inference requires estimating not only average treatment effects but also interventional outcome distributions, including quanti

safetyarxiv-cs-lg
12 May 2026
Safety

FairHealth: An Open-Source Python Library for Trustworthy Healthcare AI in Low-Resource Settings

DGX agent

arXiv:2605.08198v1 Announce Type: cross Abstract: We present FairHealth, an open-source Python library that provides a unified, modular framework for trustworthy machine learning in healthcare applica

safetyarxiv-cs-ai
12 May 2026
Safety

Fairness of Explanations in Artificial Intelligence (AI): A Unifying Framework, Axioms, and Future Direction toward Responsible AI

DGX agent

arXiv:2605.09852v1 Announce Type: new Abstract: Machine learning algorithms are being used in high-stakes decisions, including those in criminal justice, healthcare, credit, and employment. The resear

safetyarxiv-cs-ai
12 May 2026
Safety

Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems

DGX agent

arXiv:2605.10604v1 Announce Type: cross Abstract: Designing fair algorithmic decision systems requires balancing model performance with fairness toward affected individuals: More fairness might requir

safetyarxiv-cs-ai
12 May 2026
Safety

Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability

DGX agent

arXiv:2605.09214v1 Announce Type: cross Abstract: Kullback-Leibler (KL) regularization is ubiquitous in reinforcement learning algorithms in the form of reverse or forward KL. Recent studies have demo

safetyarxiv-cs-ai
12 May 2026
Safety

Few-Click-Driven Interactive 3D Segmentation with Semantic Embedding

DGX agent

arXiv:2605.08925v1 Announce Type: new Abstract: Interactive segmentation allows efficient label generation by leveraging user-provided clicks to progressively refine predictions, which is critical whe

safetyarxiv-cs-cv
12 May 2026
Safety

Flag Varieties: A Geometric Framework for Deep Network Alignment

DGX agent

arXiv:2605.09861v1 Announce Type: cross Abstract: Alignment, the tendency of adjacent weight matrices in deep networks to develop compatible subspace orientations, underlies gradient flow, Neural Coll

safetyarxiv-cs-ai
12 May 2026
← Previous
1…217218219220221…302
Next →