AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
25 May 2026

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

SafetyDGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

Diffusion and Flow Matching Models for Tabular Data: A Survey

SafetyDGX agent

arXiv:2502.17119v2 Announce Type: replace-cross Abstract: Deep generative models have made rapid progress in image, text, audio, and video generation, and are increasingly being applied to structured

Direct Dynamic Retargeting for Humanoid Imitation Learning from Videos

SafetyDGX agent

arXiv:2605.23762v1 Announce Type: new Abstract: Imitation Learning from monocular video demonstrations provides a scalable approach for teaching complex skills to humanoid robots. However, translating

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models

SafetyDGX agent

arXiv:2509.06858v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly used to simulate human opinion dynamics, yet the effect of genuine interaction is often obscured by sys

Dreaming Smoothly and Sample Efficiently with Gradient Penalized Latent Dynamics

SafetyDGX agent

arXiv:2605.23089v1 Announce Type: cross Abstract: Model-based reinforcement learning improves sample efficiency by learning a world model. However, existing latent world models such as DreamerV3 do no

Droneulator: A Portable UAV Simulator for Agricultural Workflows with RotorPy and Godot 4

SafetyDGX agent

arXiv:2605.23386v1 Announce Type: new Abstract: Agricultural UAV research requires simulators that integrate realistic 3D scenes, high-fidelity vehicle dynamics, and robotics middleware, while remaini

Dynamic Weight-based Temporal Aggregation for Low-light Video Enhancement Under Extreme Noise

SafetyDGX agent

arXiv:2510.09450v2 Announce Type: replace Abstract: Low-light video enhancement (LLVE) is challenging due to noise, low contrast, and color degradation. While learning-based methods enable fast infere

EDGE-OPD: Internalizing Privileged Context with Evidence Guided On-Policy Distillation

SafetyDGX agent

arXiv:2605.23493v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has gained wide attraction as an LLM post-training paradigm due to its effectiveness in improving capabilities without intr

Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems

SafetyDGX agent

arXiv:2605.22883v1 Announce Type: new Abstract: Current AI energy benchmarks measure consumption at the granularity of a single model invocation or training run. For classical single-turn workloads th

Entropy-Aware On-Policy Distillation of Language Models

SafetyDGX agent

arXiv:2603.07079v2 Announce Type: replace-cross Abstract: On-policy distillation is a promising approach for transferring knowledge between language models, where a student learns from dense token-lev

EquiSumm : A Gender Bias-Aware Framework for Inclusive Tweet Summarization

SafetyDGX agent

arXiv:2605.23412v1 Announce Type: new Abstract: While social media platforms, such as Twitter, provide a medium for large-scale opinion sharing during news events, it is manually impossible for indivi

FIRMA: FIbonacci Ring Model Aggregation for Privacy-preserving Federated Learning

SafetyDGX agent

arXiv:2605.22898v1 Announce Type: new Abstract: Federated learning protocols face a structural trilemma: canonical server-based aggregation~ite{mcmahan2017} creates a single point of failure and gradi

for the years it was the opposite bashing AI gets more views than it used to BECAUSE PEOPLE FIGURED OUT HOW GREEDY MOST OF THE TECH OLIGARCH…

SafetyDGX agent

for the years it was the opposite bashing AI gets more views than it used to BECAUSE PEOPLE FIGURED OUT HOW GREEDY MOST OF THE TECH OLIGARCHS ARE. even the Pope felt it was time to speak out. Bashing

Foundation Protocol: A Coordination Layer for Agentic Society

SafetyDGX agent

arXiv:2605.23218v1 Announce Type: new Abstract: Autonomous agents are moving from tools into a layer of social infrastructure: they browse, purchase, deploy software, manage systems, and increasingly

Four Simple Proprioceptive Estimators for Legged Robots

SafetyDGX agent

arXiv:2605.23100v1 Announce Type: new Abstract: Legged robots carry an IMU, but the inertial solution drifts because consumer-grade IMUs are noisy. However, the feet create intermittent contacts with

From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.23382v1 Announce Type: new Abstract: Agentic reinforcement learning (Agentic RL) has achieved strong progress in tasks with clear success signals. However, many real-world agent application

From the Patrick Boyle Channel on YouTube discussing the upcoming SpaceX IPO: Based on its filing, SpaceX has lost ~ $37 Billion over 24 yea…

SafetyDGX agent

From the Patrick Boyle Channel on YouTube discussing the upcoming SpaceX IPO: Based on its filing, SpaceX has lost ~ $37 Billion over 24 years. The worst performance of any company which has gone publ

Geo-Align: Video Generation Alignment via Metric Geometry Reward

SafetyDGX agent

arXiv:2605.23903v1 Announce Type: new Abstract: Camera-controlled video generation has achieved remarkable progress in recent years. However, existing video-to-video re-rendering methods primarily rel

Goal-Conditioned Agents that Learn Everything All at Once

SafetyDGX agent

arXiv:2605.23551v1 Announce Type: cross Abstract: A goal-conditioned reinforcement learning agent exploring an environment will see a wealth of information throughout a trajectory, most of which is di

🚨 Google DeepMind CEO Sir Demis Hassabis: “Today’s systems, are nowhere near [AGI]. Doesn’t matter how many Erdős problems you solve… I thi…

SafetyDGX agent

🚨 Google DeepMind CEO Sir Demis Hassabis: “Today’s systems, are nowhere near [AGI]. Doesn’t matter how many Erdős problems you solve… I think it’s far, far from what a true invention or someone like a

Graph Alignment Topology as an Inductive Bias for Grounding Detection

SafetyDGX agent

arXiv:2605.22963v1 Announce Type: cross Abstract: Large Language Models (LLMs) are optimized to produce distributionally plausible continuations rather than to explicitly verify whether generated prop

Great that so many of you have enjoyed my deconstructions of Elon, Zuckerberg, and that OpenAI dude, but this is much more important:

SafetyDGX agent

Great that so many of you have enjoyed my deconstructions of Elon, Zuckerberg, and that OpenAI dude, but this is much more important: ⚠️⚠️⚠️i don’t think most people understand the implications of the

How about if you change your social media name to Kekius Maximus?

SafetyDGX agent

Gary Marcus humorously suggests changing one's social media name to 'Kekius Maximus,' likely commenting on internet culture, meme terminology, or the absurdity of online personas. The post appears to

If enough other companies report the same, the bubble pops. 🫧

SafetyDGX agent

Gary Marcus comments on how widespread reporting of similar issues or problems across multiple companies can lead to a collapse of confidence or a 'bubble pop' in a particular sector or narrative. The

If you live in the UK, and care about the risk from superintelligence, please read below!

SafetyDGX agent

If you live in the UK, and care about the risk from superintelligence, please read below! URGENT: The UK Parliament just selected 20 MPs to introduce a bill of their choosing. This is an unprecedented

Information Access of the Oppressed: Freirean Design for Emancipatory Information Access

SafetyDGX agent

arXiv:2601.09600v3 Announce Type: replace-cross Abstract: Online information access (IA) platforms are targets of authoritarian capture. We explore the question of how to safeguard our platforms and e

Instrumentation for Imitation Learning: Enhancing Training Datasets for Clothes Hanger Insertion

SafetyDGX agent

arXiv:2605.23847v1 Announce Type: new Abstract: Large behaviour models have transformed the field of robotic manipulation, but prohibitive data requirements have thus far prevented a revolution simila

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents

SafetyDGX agent

arXiv:2604.05157v2 Announce Type: replace Abstract: Computer-Use Agents (CUAs) leverage large language models to execute GUI operations on desktop environments, yet they generate actions without evalu

interesting to listen to this years later. what held up? what changed?

SafetyDGX agent

interesting to listen to this years later. what held up? what changed? I had a podcast where I was fortunate to have amazing guests. This was just as GPT 3 came out. Despite LLMs exceeding so many exp

it’s great! you can pay for the infrastructure that will eventually take your jobs! and if it fails and the bubble bursts? you can bail the …

SafetyDGX agent

it’s great! you can pay for the infrastructure that will eventually take your jobs! and if it fails and the bubble bursts? you can bail the hyperscalers out, and watch your pension fund die. The CEO o

I've yet to see, or hear of a company that is winning against its competition, because said company is spending more on AI tools, or using i…

SafetyDGX agent

I've yet to see, or hear of a company that is winning against its competition, because said company is spending more on AI tools, or using it better than the competition. Ways I see companies win: - b

Learning Kernel-Based MDPs from Episodic Preferential Feedback

SafetyDGX agent

arXiv:2605.23650v1 Announce Type: cross Abstract: Human feedback often arrives as preferences rather than calibrated numeric rewards, motivating reinforcement learning from preferential feedback, also

“Like nuclear energy, AI must be at the service of all and of the common good. Decisions about technology must never be separated from consc…

SafetyDGX agent

“Like nuclear energy, AI must be at the service of all and of the common good. Decisions about technology must never be separated from conscience and responsibility.” I agree with this sentiment by @P

Lipschitz Optimization for Formal Verification of Homographies

Model ReleasesDGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

Millimeter-wave Imaging for Anthropometric Body Measurement

SafetyDGX agent

arXiv:2605.23064v1 Announce Type: new Abstract: Body shape and circumferences are clinically informative biomarkers for risk stratification, including measures such as waist to hip ratio, limb and tru

Multimodal Distribution Matching for Vision-Language Dataset Distillation

SafetyDGX agent

arXiv:2605.23482v1 Announce Type: cross Abstract: Dataset distillation compresses large training sets into compact synthetic datasets while preserving downstream performance. As modern systems increas

Musk hates Altman because Altman deceived him. Altman hates Musk because Musk is an egomaniac. LeCun (who is an equally big egomaniac) hates…

SafetyDGX agent

Musk hates Altman because Altman deceived him. Altman hates Musk because Musk is an egomaniac. LeCun (who is an equally big egomaniac) hates Musk because Musk is a jerk. Musk hates LeCun because LeCun

Naturalistic measure of social norms alignment

SafetyDGX agent

arXiv:2605.23420v1 Announce Type: new Abstract: Social norms reflect shared expectations on acceptable behavior. Measuring social norms alignment remains challenging, with existing approaches typicall

NeuralBoneReg: An Instance-Specific Label-Free Point Cloud-Based Method for Multi-Modal Bone Surface Registration

SafetyDGX agent

arXiv:2511.14286v2 Announce Type: replace Abstract: In computer- and robot-assisted orthopedic surgery (CAOS), patient-specific surgical plans derived from preoperative imaging define target locations

Next-Latent Prediction Transformers Learn Compact World Models

SafetyDGX agent

arXiv:2511.05963v2 Announce Type: replace Abstract: Transformers replace recurrence with a memory that grows with sequence length and self-attention that enables ad-hoc lookups over past tokens. Conse

Not Too Generative, Not Too Discriminative: The Human Alignment Sweet Spot

SafetyDGX agent

arXiv:2605.23819v1 Announce Type: cross Abstract: A central question in computational vision is whether human-like visual representations are better explained by discriminative or generative learning.

on this, @DavidSacks and I agree.

SafetyDGX agent

Gary Marcus and David Sacks found common ground on an unspecified topic, as indicated by Marcus's post on X (formerly Twitter). Without access to the full thread context, the specific point of agreeme

Online Learning with Multiple Fairness Regularizers via Graph-Structured Feedback

SafetyDGX agent

arXiv:2508.14311v2 Announce Type: replace-cross Abstract: There is an increasing need to enforce multiple, often competing, measures of fairness within automated decision systems. The appropriate weig

Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models

SafetyDGX agent

arXiv:2605.23145v1 Announce Type: cross Abstract: Individual fairness, the notion that 'similar individuals should be treated similarly,' provides a strong and flexible fairness guarantee for algorith

people in glass houses should not throw stones, Zuck

SafetyDGX agent

people in glass houses should not throw stones, Zuck Mark Zuckerberg says Apple's lack of innovation since the iPhone will lead to its decline 'They haven't really invented anything great in a while.

pi_0-EqM: Equilibrium Matching for Closed-Loop Vision-Language-Action Control

SafetyDGX agent

arXiv:2605.23128v1 Announce Type: new Abstract: Currently, Vision-Language-Action (VLA) models have become the most adopted paradigm for robotic manipulation for its great potential for task generaliz

PilotWiMAE: Pilot-Native Representation Learning for Wireless Channels

SafetyDGX agent

arXiv:2605.22856v1 Announce Type: cross Abstract: Channel foundation models assume access to fully observed channels, an assumption that fails in deployment. We introduce PilotWiMAE, a self-supervised

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

SafetyDGX agent

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

Point Tracking Improves World Action Models

SafetyDGX agent

arXiv:2605.23856v1 Announce Type: new Abstract: Robot policy learning benefits from world-action models that capture environment dynamics, but pixel-level prediction entangles dynamics with nuisance f

Pope Leo warns AI revolution driven by ‘idolatry of profit’ https://ft.trib.al/UGP9H8b

SafetyDGX agent

Pope Leo has warned that the artificial intelligence revolution is being driven primarily by profit motives rather than ethical considerations, characterizing this as a form of 'idolatry of profit.' T

Precise: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models

SafetyDGX agent

arXiv:2605.23522v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become an effective way to improve prompt alignment and perceptual quality in diffusion and flow-matching generators.

Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback

SafetyDGX agent

arXiv:2605.23182v1 Announce Type: new Abstract: Pure exploration in episodic Reinforcement Learning has primarily focused on Best Policy Identification (BPI), which seeks to identify a (near)-optimal

RAG4Outcome: A Retrieval-Augmented Multimodal Framework for Prognostic Prediction in Chronic Osteomyelitis

SafetyDGX agent

arXiv:2605.22833v1 Announce Type: cross Abstract: Chronic osteomyelitis presents substantial prognostic challenges due to its high recurrence risk and complex postoperative recovery trajectories. Trad

ReCoVer: Resilient LLM Pre-Training System via Fault-Tolerant Collective and Versatile Workload

SafetyDGX agent

arXiv:2605.11215v2 Announce Type: replace-cross Abstract: Pre-training large language models on massive GPU clusters has made hardware faults routine rather than rare, driving the need for resilient t

Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control

SafetyDGX agent

arXiv:2605.23415v1 Announce Type: cross Abstract: Reinforcement learning has long struggled with poor sample efficiency. One promising approach to mitigate this problem is leveraging group-invariant M

Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers

SafetyDGX agent

arXiv:2510.00915v4 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) replaces costly human labeling with automated verifiers. To reduce verifier hacking, man

Robotic Strawberry Harvesting with Robust Vision and Deep Reinforcement Learning based Sim-to-Real Control

SafetyDGX agent

arXiv:2605.23863v1 Announce Type: new Abstract: This study presents a closed-loop robotic strawberry harvesting system that combines a robust vision module, simulation-trained deep reinforcement learn

Sample-wise Targeted Adversarial Attacks on Test-time Adaptation

SafetyDGX agent

arXiv:2605.23411v1 Announce Type: cross Abstract: Test-time adaptation (TTA) effectively counters distribution shifts but exposes models to adversarial manipulation via the unlabeled test stream. Exis

Score-Based One-step MeanFlow Policy Optimization

SafetyDGX agent

arXiv:2605.23365v1 Announce Type: cross Abstract: Diffusion and flow matching have emerged as expressive policy classes in reinforcement learning, but their reliance on multi-step denoising imposes su

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

SafetyDGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

← Previous
1…149150151152153…242
Next →