AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
25 Jun 2026

The saddest thing about AI is how many people are using it to become worse, lesser, and disempower themselves.

SafetyDGX agent

Connor Leahy expresses concern that AI tools are being used by many people in ways that diminish their capabilities and agency rather than enhance them. The post suggests that rather than leveraging A

TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control

SafetyDGX agent

arXiv:2601.14945v2 Announce Type: replace Abstract: Large-scale Vision-Language-Action (VLA) models offer semantic generalization but suffer from high inference latency, limiting them to low-frequency

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models

SafetyDGX agent

arXiv:2606.24962v1 Announce Type: new Abstract: Recent progress in large-scale sequence modeling has shown that a single model can learn useful representations across highly diverse data distributions

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR

SafetyDGX agent

arXiv:2606.25178v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has been extended from single-domain training to multi-domain reasoning suites spanning mathematic

TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction

SafetyDGX agent

arXiv:2601.22615v3 Announce Type: replace Abstract: Streaming recurrent models enable efficient 3D reconstruction by maintaining persistent state representations. However, they suffer from catastrophi

Uncertainty-aware reinforcement learning for chemical language models

SafetyDGX agent

arXiv:2606.24990v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a powerful paradigm for de novo molecular design, enabling Chemical Language Models (CLMs) to navigate and explor

update with some context and additional thoughts: https://open.substack.com/pub/garymarcus/p/the-generative-ai-fizzle?utm_source=app-post-st…

SafetyDGX agent

Gary Marcus discusses concerns about generative AI's limitations and potential overhyping of the technology, arguing that initial enthusiasm may not translate into sustained practical breakthroughs as

VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction

SafetyDGX agent

arXiv:2509.19297v3 Announce Type: replace Abstract: Feed-forward 3D Gaussian Splatting (3DGS) has emerged as a highly effective solution for novel view synthesis. Existing methods predominantly rely o

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria…

SafetyDGX agent

Weird to see this just as the administration is delaying a model for the second time in weeks, apparently without clarity about its criteria. You can’t be pro-growth, pro-innovation, and opaque at all

What Does It Mean to Break a Distillation Defense?

SafetyDGX agent

arXiv:2606.25059v1 Announce Type: cross Abstract: Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its ou

When Do Conservation Laws Survive Learned Representations? Certified Horizons for Latent World Models

SafetyDGX agent

arXiv:2606.24945v1 Announce Type: new Abstract: We ask a representation-learning question about physical world models: when does a conservation law remain certifiable after a model learns a latent rep

When Does Synthetic Data Augmentation Improve Score-Based Imbalanced Classification?

SafetyDGX agent

arXiv:2606.26053v1 Announce Type: cross Abstract: Synthetic data augmentation is widely used to mitigate class imbalance, but its theoretical effects on score-based classification remain poorly unders

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

SafetyDGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

wouldn’t it be funny if greed undid OpenAI?

SafetyDGX agent

wouldn’t it be funny if greed undid OpenAI? 🚨BREAKING: OPENAI IPO DELAYED Altman told everyone OpenAI is worth a TRILLION dollars His own advisors warned him that retail investors aren’t buying it… ga

24 Jun 2026

A Comparative Study of Bayesian Contextual Bandits for Real-Time Warehouse Sorter Optimization

SafetyDGX agent

arXiv:2606.23977v1 Announce Type: new Abstract: Efficient sorter diversion control of automated material handling systems (MHS) is critical for optimizing operational efficiency in large-scale warehou

A global log for medical AI

SafetyDGX agent

arXiv:2510.04033v2 Announce Type: replace Abstract: Modern computer systems rely on syslog, a universal protocol that records critical events across heterogeneous infrastructure. Medicine's rapidly gr

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be …

SafetyDGX agent

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be fine tuned and start treating it like the existential threat

A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Levy Process Dynamics

SafetyDGX agent

arXiv:2504.01482v3 Announce Type: replace-cross Abstract: This paper develops a model-based framework for continuous-time policy evaluation (CTPE) in reinforcement learning, incorporating both Brownia

Abstractions of Queries in Ontology-Based Data Access

SafetyDGX agent

arXiv:2606.24618v1 Announce Type: new Abstract: In ontology-based data access (OBDA), multiple data sources are integrated via mappings to an ontology. We consider an OBDA setting based on existential

Accelerated Stochastic Min-Max Optimization Based on Bias-corrected Momentum

SafetyDGX agent

arXiv:2406.13041v3 Announce Type: replace Abstract: Lower-bound analyses for nonconvex strongly-concave minimax optimization problems have shown that stochastic first-order algorithms require at least

Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems

SafetyDGX agent

arXiv:2606.24416v1 Announce Type: new Abstract: Network operators' changing policies, service requirements, and stringent real-time constraints render existing methods designed with fixed objectives a

Aligning Audio Captions with Human Preferences

SafetyDGX agent

arXiv:2509.14659v3 Announce Type: replace-cross Abstract: Current audio captioning relies on supervised learning with paired audio-caption data, which is costly to curate and may not reflect human pre

An Introduction to Causal Reinforcement Learning

SafetyDGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

An LLM-based Two-Stage Transformer Framework for Cross-Domain Bearing Fault Diagnosis with Limited Data

SafetyDGX agent

arXiv:2606.24459v1 Announce Type: cross Abstract: Bearing fault diagnosis faces critical challenges when dataset heterogeneity, operating condition variations, and limited labeled data occur simultane

Are LLM Evaluators Really Narcissists? Sanity Checking Self-Preference Evaluations

SafetyDGX agent

arXiv:2601.22548v4 Announce Type: replace-cross Abstract: Recent research has shown that large language models (LLMs) favor their own outputs when acting as judges, undermining the integrity of automa

ARIA: Adaptive Region-Based Importance Allocation for Conditional Diffusion Distillation

SafetyDGX agent

arXiv:2606.23898v1 Announce Type: cross Abstract: Distilling conditional diffusion models aims to transfer the behavior of a large teacher to a smaller student while preserving alignment across condit

AsyncOPD: How Stale Can On-Policy Distillation Be?

SafetyDGX agent

arXiv:2606.24143v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts guided by teacher feedback and is becoming increasingly important for large language m

Audio-visual Contrastive Alignment for Diffusion-based Visual-conditioned Speech Enhancement

SafetyDGX agent

arXiv:2606.23712v1 Announce Type: cross Abstract: Audio-visual speech enhancement (AVSE) exploits visual cues such as lip movements to recover speech in noisy environments. Recent work introduced diff

Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.24064v1 Announce Type: new Abstract: Distilling reasoning capabilities from strong to weak language models typically involves imitating specific solution trajectories, effectively transferr

Beyond U-Net: A Latent-Representation-Aligned Skip-Free Backbone for Flow-Matching Speech Enhancement

SafetyDGX agent

arXiv:2606.24745v1 Announce Type: cross Abstract: Generative models, particularly diffusion and score-based approaches, have recently achieved strong performance in speech enhancement, but their itera

bit-equivalent on-policy rl for glm-5.2 has been achieved internally developing…

SafetyDGX agent

I cannot provide an accurate summary of this entry as the title appears incomplete and the URL/source information seems corrupted or mismatched (the handle doesn't match the URL). To create a reliable

Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation

SafetyDGX agent

arXiv:2606.24464v1 Announce Type: new Abstract: Text-driven Referring Video Object Segmentation (RVOS) aims to locate and segment target objects in videos given natural language. However, existing mod

Breaking Shortcut Learning for Cross-Trial EEG-Guided Target Speech Extraction via Two-Stage Training

SafetyDGX agent

arXiv:2606.24164v1 Announce Type: cross Abstract: Recent end-to-end models for EEG-guided target speech extraction report impressive results, underscoring potential for neuro-steered hearing technolog

Breaking the Filter Bubble: A Semantic Pareto-DQN Framework for Multi-Objective Recommendation

SafetyDGX agent

arXiv:2606.24042v1 Announce Type: new Abstract: Recommender systems often induce filter bubbles and semantic homogenization by monolithically optimizing for immediate user engagement. Standard single-

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

SafetyDGX agent

arXiv:2509.03647v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly serve as automated evaluators, yet they suffer from 'self-preference bias': a tendency to favor thei

Bridging the Manifold Gap: Riemannian Residual Line Search for One-Step Image Editing

SafetyDGX agent

arXiv:2606.24844v1 Announce Type: new Abstract: One-step diffusion editors are fast because they avoid inversion and iterative optimization, but a single transport update must be aggressive enough to

CALIBER: Calibrating Confidence Before and After Reasoning in Language Models

SafetyDGX agent

arXiv:2606.24281v1 Announce Type: cross Abstract: Reasoning language models are increasingly asked not only to answer difficult questions, but also to estimate their likelihood of success. Existing me

COMPANY BUILT ON IP THEFT RAGES AGAINST THEFT OF ITS OWN IP

SafetyDGX agent

COMPANY BUILT ON IP THEFT RAGES AGAINST THEFT OF ITS OWN IP ANTHROPIC ACCUSED ALIBABA OF UNAUTHORIZED ACCESS TO ITS AI MODELS, OUTLINING ITS ALLEGATIONS IN LETTERS SENT TO U.S. SENATORS AND THE WHITE

Congresswoman denies staff used AI to write defense funding amendment

SafetyDGX agent

Rep. Anna Paulina Luna (R-FL) says her staff used AI for 'spellcheck' in an amendment summary for a major defense bill, but denies it was used for the bill text itself and says 'NO Legislation is ever

Cryptographic certificates of validity for trustworthy AI

SafetyDGX agent

arXiv:2606.23768v1 Announce Type: cross Abstract: We propose cryptographic certificates of validity for agentic AI systems. The core idea is to formally specify a correctness or policy condition as a

Decentralized SGD with Controlled Disagreement Finds Flatter Minima

SafetyDGX agent

arXiv:2602.02899v2 Announce Type: replace Abstract: Decentralized training is often regarded as inferior to centralized training because the consensus errors between workers are thought to undermine c

DREG: A Layer-Wise Jacobian Regularization as a General-Purpose Penalty

SafetyDGX agent

arXiv:2606.23942v1 Announce Type: new Abstract: We present a large-scale empirical study isolating the contributions of the Derivative Regularization penalty (DREG). Across a fully-crossed factorial s

Dynamic Symmetric Point Tracking: Tackling Non-ideal Reference in Analog In-memory Training

SafetyDGX agent

arXiv:2602.21321v2 Announce Type: replace Abstract: Analog in-memory computing (AIMC) performs computation directly within resistive crossbar arrays, offering an energy-efficient platform to scale lar

Enabling Robust Cloth Manipulation via Inference-Time Simulator-in-the-Loop Refinement

SafetyDGX agent

arXiv:2606.24552v1 Announce Type: new Abstract: Simulator-in-the-loop optimization offers a promising inference-time mechanism for robot manipulation. It uses a physical simulator as a backend rollout

Enforcing Human-like Kinematics in Dexterous Piano Playing via Adversarial Posture Regularization

SafetyDGX agent

arXiv:2606.23848v1 Announce Type: new Abstract: Reinforcement learning can train bimanual dexterous hands to play piano in physics simulation with high note accuracy, but for high-DoF dexterous hands,

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

SafetyDGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations

SafetyDGX agent

arXiv:2606.24716v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are increasingly used to extract interpretable concepts from vision and vision language models, yet existing evaluation met

EvidenceLens: A Claim-Evidence Matrix for Auditing Financial Question Answering

SafetyDGX agent

arXiv:2606.23724v1 Announce Type: cross Abstract: Large language models are increasingly used to answer questions over annual reports, earnings decks, and analyst notes, yet their outputs remain diffi

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL

SafetyDGX agent

arXiv:2606.23693v1 Announce Type: new Abstract: Text-to-SQL enables users to query databases using natural language by generating executable SQL queries. Recent methods have increasingly adopted Large

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection

SafetyDGX agent

arXiv:2508.18684v2 Announce Type: replace-cross Abstract: Signature-based Intrusion Detection Systems (IDS) detect malicious activity by matching network or host events against predefined rules. Secur

From 'Aha Moments' to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

SafetyDGX agent

arXiv:2508.04460v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) can exhibit step-by-step reasoning, reflection, and backtracking, but these behaviors are often unregulated, leading t

From Local Corrections to Generalized Skills: Improving Neuro-Symbolic Policies with MEMO

SafetyDGX agent

arXiv:2603.04560v2 Announce Type: replace Abstract: Recent works use a neuro-symbolic framework for general manipulation policies. The advantage of this framework is that -- by applying off-the-shelf

G^3VLA: Geometric inductive bias for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.24472v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have made rapid progress in generalist robot manipulation by harnessing semantic knowledge from pretrained vision-

GENA3D: Generative Amodal 3D Modeling by Bridging 2D Priors and 3D Coherence

SafetyDGX agent

arXiv:2511.21945v3 Announce Type: replace Abstract: Generating complete 3D objects under partial occlusions (i.e., amodal scenarios) is a practically important yet challenging problem, as large portio

Generative Manifold Distillation: Aligning Restoration Trajectories with Natural Image Prior

SafetyDGX agent

arXiv:2512.11121v2 Announce Type: replace Abstract: Pre-trained image restoration models often fail on out-of-distribution (OOD) real-world degradations. Adapting to these domains is challenging as re

Geometric Action Model for Robot Policy Learning

SafetyDGX agent

arXiv:2606.17046v2 Announce Type: replace-cross Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot actions interact in the 3D physi

Governed Shared Memory for Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2606.24535v1 Announce Type: new Abstract: Multi-agent LLM environments require robust mechanisms for shared knowledge management. This paper formalizes the fleet-memory problem and identifies fo

Hierarchical Spatial and Channel Aggregation for Cross-domain Few-shot Segmentation

SafetyDGX agent

arXiv:2606.24296v1 Announce Type: new Abstract: Cross-domain Few-shot Segmentation (CD-FSS) aims to learn generalizable segmentation capability from abundant annotated samples in the source domain, en

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making

SafetyDGX agent

arXiv:2508.16420v3 Announce Type: replace Abstract: Target-conditioned sequence models provide a simple interface for controllable offline decision making, but the requested target return can be an un

JEDEL: Zero-Shot DNA-Encoded Library Design for Early-Stage Drug Discovery

SafetyDGX agent

arXiv:2606.23745v1 Announce Type: cross Abstract: We present JEDEL, a framework for generating synthesis-ready DNA-encoded libraries (DELs) directly from three-dimensional pharmacophore representation

← Previous
1…103104105106107…242
Next →