AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

Where Pretraining writes and Alignment reads: the asymmetry of Transformer weight space

DGX agent

arXiv:2605.16600v1 Announce Type: cross Abstract: Cross-entropy pretraining and preference alignment update the same transformer weights, but leave geometrically distinct traces. We characterise this

safetyarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

DGX agent

arXiv:2605.18150v1 Announce Type: new Abstract: Diffusion models (DMs) are widely used for text-to-image generation, but their strong generative capabilities also raise concerns about unsafe or undesi

safetyarxiv-cs-ai
19 May 2026
Safety

White-Box Sensitivity Auditing with Steering Vectors

DGX agent

arXiv:2601.16398v2 Announce Type: replace-cross Abstract: Algorithmic audits are essential tools for examining systems for properties required by regulators or desired by operators. Current audits of

safetyarxiv-cs-cl
19 May 2026
Safety

World Model-Enabled Causal Digital Twins for Semantic Communications in Physical AI Systems

DGX agent

arXiv:2605.16547v1 Announce Type: new Abstract: Semantic communication has emerged as a promising paradigm for enabling goal-oriented networking. However, most existing semantic communication solution

safetyarxiv-cs-lg
19 May 2026
Safety

Zero-Shot Textual Explanations via Translating Decision-Critical Features

DGX agent

arXiv:2512.07245v2 Announce Type: replace Abstract: Textual explanations make image classifier decisions transparent by describing the prediction rationale in natural language. Large vision-language m

safetyarxiv-cs-cv
19 May 2026
Safety

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse

DGX agent

arXiv:2509.23183v3 Announce Type: replace Abstract: Test-time entropy minimization helps adapt a model to novel environments and incentivize its reasoning capability, unleashing the model's potential

safetyarxiv-cs-lg
19 May 2026
Safety

A Differentiable Measure of Algebraic Complexity: Provably Exact Discovery of Group Structures

DGX agent

arXiv:2511.23152v3 Announce Type: replace Abstract: Discovering discrete algebraic rules from data is a fundamental challenge in machine learning. We formalize this problem through Cayley-table comple

safetyarxiv-cs-lg
18 May 2026
Safety

A Generative AI Framework for Intelligent Utility Billing CO 2 Analytics and Sustainable Resource Optimisation

DGX agent

arXiv:2605.16250v1 Announce Type: cross Abstract: Distribution utilities are now expected to deliver bills that customers can actually read attach a defensible carbon number to every kWh sold and sche

safetyarxiv-cs-ai
18 May 2026
Safety

A Split-Client Approach to Second-Order Optimization

DGX agent

arXiv:2510.15714v3 Announce Type: replace-cross Abstract: Second-order optimization methods offer superior convergence rates but are often bottlenecked by the wall-clock cost of Hessian computation an

safetyarxiv-cs-lg
18 May 2026
Safety

a very procedural end. we will never know what the world might be like had OpenAI been forced to fully follow its original mission.

DGX agent

a very procedural end. we will never know what the world might be like had OpenAI been forced to fully follow its original mission. Breaking News: A jury rejected Elon Musk’s lawsuit accusing OpenAI o

safetygary-marcus--x
18 May 2026
Safety

Accelerated Gradient Descent for Faster Convergence with Minimal Overhead

DGX agent

arXiv:2605.16017v1 Announce Type: new Abstract: In this paper, we present CT-AGD (Curvature-Tuned Accelerated Gradient Descent), an optimization method for non-convex optimization problems in deep lea

safetyarxiv-cs-lg
18 May 2026
Safety

ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment

DGX agent

arXiv:2505.19241v2 Announce Type: replace-cross Abstract: The recent success in using human preferences to align large language models (LLMs) has significantly improved their performance in various do

safetyarxiv-cs-ai
18 May 2026
Safety

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making

DGX agent

arXiv:2605.16054v1 Announce Type: cross Abstract: Recent work has framed decision-making as a sequence modeling problem using generative models such as diffusion models. Although promising, these appr

safetyarxiv-cs-ai
18 May 2026
Safety

Adaptive Outer-Loop Control of Quadrotors via Reinforcement Learning

DGX agent

arXiv:2605.16015v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) for quadrotor flight control typically relies on Domain Randomization (DR) for sim-to-real transfer, resulting in ov

safetyarxiv-cs-lg
18 May 2026
Safety

AI-Mediated Communication Can Steer Collective Opinion

DGX agent

arXiv:2605.16245v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is increasingly integrated into the online platforms where humans exchange opinions; large language models (LL

safetyarxiv-cs-ai
18 May 2026
Safety

Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time

DGX agent

arXiv:2605.15220v1 Announce Type: cross Abstract: Data mixing decides how to combine different sources or types of data and is a consequential problem throughout language model training. In pretrainin

safetyarxiv-cs-ai
18 May 2026
Safety

“Americans are now more comfortable living near a nuclear power plant than an AI data center” -@RachelBitecofer The AI oligarchs took a winn…

DGX agent

“Americans are now more comfortable living near a nuclear power plant than an AI data center” -@RachelBitecofer The AI oligarchs took a winning hand, and with a mixture cigarette-industry level greed

safetygary-marcus--x
18 May 2026
Safety

An Algebraic Exposition of the Theory of Dyadic Morality

DGX agent

arXiv:2605.16153v1 Announce Type: new Abstract: This paper provides an algebraic exposition of the theory of dyadic morality (TDM), a psychological model of moral judgment grounded in a simple two-nod

safetyarxiv-cs-ai
18 May 2026
Safety

An Introduction to Deep Reinforcement and Imitation Learning

DGX agent

arXiv:2512.08052v3 Announce Type: replace-cross Abstract: Embodied agents, such as robots and virtual characters, must continuously select actions to execute tasks effectively, solving complex sequent

safetyarxiv-cs-lg
18 May 2026
Safety

AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs

DGX agent

arXiv:2605.15565v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly used to improve the reasoning, coding, and tool-use capabilities of large language models, but agentic RL

safetyarxiv-cs-ai
18 May 2026
Safety

Best-of-Both-Worlds for Heavy-Tailed Markov Decision Processes

DGX agent

arXiv:2602.01295v3 Announce Type: replace Abstract: We investigate episodic Markov Decision Processes with heavy-tailed losses (HTMDPs). Existing approaches for HTMDPs are conservative in stochastic e

safetyarxiv-cs-lg
18 May 2026
Safety

Beyond Objective-Based Improvement: Stationarity-Aware Expected Improvement for Bayesian Optimization

DGX agent

arXiv:2601.21357v2 Announce Type: replace Abstract: Bayesian Optimization (BO) is a principled framework for optimizing expensive black-box functions, with Expected Improvement (EI) among its most wid

safetyarxiv-cs-lg
18 May 2026
Safety

Beyond Performance Disparities: A Three-Level Audit of Representational Harm in CelebA

DGX agent

arXiv:2605.15312v1 Announce Type: cross Abstract: Large-scale facial datasets like CelebA are widely used in computer vision, yet the cultural biases embedded in their labels remain underexplored. Fai

safetyarxiv-cs-cv
18 May 2026
Safety

Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling

DGX agent

arXiv:2507.01679v3 Announce Type: replace-cross Abstract: Existing LLMs-post-training techniques are broadly categorized into supervised fine-tuning (SFT) and reinforcement fine-tuning (RFT). Each par

safetyarxiv-cs-ai
18 May 2026
Safety

Can we all agree that Dario played the “Ooh! AI scary!” card one time too many?

DGX agent

Can we all agree that Dario played the “Ooh! AI scary!” card one time too many? “Americans are now more comfortable living near a nuclear power plant than an AI data center” -@RachelBitecofer The AI o

safetygary-marcus--x
18 May 2026
Safety

Controllable Molecular Generative Foundation Models

DGX agent

arXiv:2605.15354v1 Announce Type: new Abstract: Despite the success of foundation models in language and vision, molecular graph generation still lacks a unified framework for heterogeneous design tas

safetyarxiv-cs-lg
18 May 2026
Safety

DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation

DGX agent

arXiv:2605.16113v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved unprecedented success due to their exceptional generative capabilities. However, because they depend on kno

safetyarxiv-cs-ai
18 May 2026
Safety

Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation

DGX agent

arXiv:2605.15942v1 Announce Type: cross Abstract: Open-vocabulary segmentation models often struggle to generalize to unseen combinations of object categories and attributes, because fine-grained desc

safetyarxiv-cs-ai
18 May 2026
Safety

Deep Double Q-learning

DGX agent

arXiv:2507.00275v2 Announce Type: replace-cross Abstract: Double Q-learning is a classical control algorithm that mitigates the maximization bias of Q-learning. To do so, it explicitly trains two inde

safetyarxiv-cs-ai
18 May 2026
Safety

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

DGX agent

arXiv:2605.15532v1 Announce Type: cross Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically

safetyarxiv-cs-ai
18 May 2026
Safety

Designing Datacenter Power Delivery Hierarchies for the AI Era

DGX agent

arXiv:2605.16255v1 Announce Type: cross Abstract: Demand for AI accelerators is rapidly increasing rack power density, with projections approaching 1MW per deployment by 2027. This poses a major chall

safetyarxiv-cs-ai
18 May 2026
Safety

Detecting Heel Strike and toe off Events Using Kinematic Methods and LSTM Models

DGX agent

arXiv:2503.00794v2 Announce Type: replace Abstract: Accurate gait event detection is crucial for gait analysis, rehabilitation, and assistive technology, particularly in exoskeleton control, where pre

safetyarxiv-cs-ro
18 May 2026
Safety

Differentially Private Motif-Preserving Multi-modal Hashing

DGX agent

arXiv:2605.15460v1 Announce Type: cross Abstract: Cross-modal hashing enables efficient retrieval by encoding images and text into compact binary codes. State-of-the-art methods rely on semantic simil

safetyarxiv-cs-ai
18 May 2026
Safety

Diffusion Policy for Coordinated Control of a Nonholonomic Mobile Base and Dual Arms in Door Opening and Passing

DGX agent

arXiv:2605.15352v1 Announce Type: new Abstract: Opening heavy, self closing doors, especially those that require pulling remains a long standing challenge in robotics. Humans naturally employ both arm

safetyarxiv-cs-ro
18 May 2026
Safety

DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments

DGX agent

arXiv:2605.15519v1 Announce Type: cross Abstract: Visual active search (VAS) has been introduced as a modeling framework that leverages visual cues to direct aerial (e.g., UAV-based) exploration and p

safetyarxiv-cs-ai
18 May 2026
Safety

Discretizing Group-Convolutional Neural Networks for 3D Geometry in Feature Space

DGX agent

arXiv:2605.15368v1 Announce Type: new Abstract: Group-convolutional neural networks (GCNNs) are among the most important methods for introducing symmetry as an inductive bias in deep learning: In each

safetyarxiv-cs-cv
18 May 2026
Safety

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?

DGX agent

arXiv:2605.15855v1 Announce Type: new Abstract: Despite strong image-generation performance, diffusion models' reconstruction objectives limit alignment with human preferences. RL enables such alignme

safetyarxiv-cs-cv
18 May 2026
Safety

DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

DGX agent

arXiv:2511.19399v3 Announce Type: replace-cross Abstract: Deep research agents perform multi-step research to produce long-form, well-attributed answers. However, most open deep research agents are tr

safetyarxiv-cs-ai
18 May 2026
Safety

Drawback of Enforcing Equivariance and its Compensation via the Lens of Expressive Power

DGX agent

arXiv:2512.09673v3 Announce Type: replace-cross Abstract: Equivariant neural networks encode the intrinsic symmetry of data as an inductive bias, which has achieved impressive performance in wide doma

safetyarxiv-cs-ai
18 May 2026
Safety

DualKV: Shared-Prompt Flash Attention for Efficient RL Training with Large Rollouts and Long Contexts

DGX agent

arXiv:2605.15422v1 Announce Type: new Abstract: Modern RL post-training methods such as GRPO and DAPO train on N response sequences of R tokens sampled from a shared prompt of P tokens, but standard F

safetyarxiv-cs-lg
18 May 2026
Safety

DualReg: Dual-Space Filtering and Reinforcement for Rigid Registration

DGX agent

arXiv:2508.17034v2 Announce Type: cross Abstract: Noisy, partially overlapping data and the need for real-time processing pose major challenges for rigid registration. Considering that feature-based m

safetyarxiv-cs-cv
18 May 2026
Safety

Dynamic Plasma Shape Control with Arbitrary Sensor Subsets

DGX agent

arXiv:2605.15935v1 Announce Type: new Abstract: Plasma shape control in tokamaks requires a real-time controller that tracks dynamically changing shape targets while tolerating diagnostic failures. Cl

safetyarxiv-cs-ro
18 May 2026
Safety

Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling

DGX agent

arXiv:2509.23352v3 Announce Type: replace-cross Abstract: The integration of Reinforcement Learning (RL) into flow matching models for text-to-image (T2I) generation has driven substantial advances in

safetyarxiv-cs-ai
18 May 2026
Safety

Efficiently Solving Mixed-Hierarchy Games with Quasi-Policy Approximations

DGX agent

arXiv:2602.01568v2 Announce Type: replace-cross Abstract: Multi-robot coordination often exhibits hierarchical structure, with some robots' decisions depending on the planned behaviors of others. Whil

safetyarxiv-cs-ro
18 May 2026
Safety

EgoExo-WM: Unlocking Exo Video for Ego World Models

DGX agent

arXiv:2605.15477v1 Announce Type: new Abstract: Egocentric world models present a promising direction for enabling agents to predict and plan, but their performance is constrained by the limited avail

safetyarxiv-cs-cv
18 May 2026
Safety

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices

DGX agent

arXiv:2605.15684v1 Announce Type: new Abstract: The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffus

safetyarxiv-cs-cv
18 May 2026
Safety

Embedding-perturbed Exploration Preference Optimization for Flow Models

DGX agent

arXiv:2605.15803v1 Announce Type: new Abstract: Recent advancements have established Reinforcement Learning (RL) as a pivotal paradigm for aligning generative models with human intent. However, group-

safetyarxiv-cs-cv
18 May 2026
Safety

Embracing Biased Transition Matrices for Complementary-Label Learning with Many Classes

DGX agent

arXiv:2605.15586v1 Announce Type: cross Abstract: Complementary-label learning (CLL) is a weakly supervised paradigm where instances are labeled with classes they do not belong to. Despite a decade of

safetyarxiv-cs-ai
18 May 2026
← Previous
1…203204205206207…302
Next →