AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
13 Apr 2026

Wireless Communication Enhanced Value Decomposition for Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.08728v1 Announce Type: new Abstract: Cooperation in multi-agent reinforcement learning (MARL) benefits from inter-agent communication, yet most approaches assume idealized channels and exis

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector

SafetyDGX agent

arXiv:2603.15757v2 Announce Type: replace-cross Abstract: What happens when a pretrained generative robot policy is provided a constant initial noise as input, rather than repeatedly sampling it from

12 Apr 2026

I just love the language of this study...it speaks of the shifting 'community language'...and that is so true...have you noticed the new 'co…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

I just love the language of this study...it speaks of the shifting 'community language'...and that is so true...have you noticed the new 'community language' is 'harness', it was 'contextual prompting

Sources: the US' AI chip export push risks being undermined by licensing bottlenecks, staff attrition, and unclear policy at the Bureau of Industry and Security (Maggie Eastland/Bloomberg)

SafetyDGX agent

Maggie Eastland / Bloomberg: Sources: the US' AI chip export push risks being undermined by licensing bottlenecks, staff attrition, and unclear policy at the Bureau of Industry and Security — Presiden

Wow, time for a social media detox. Just today - a call for violence against me - insults (that’s every day) - flagrant lies about my creden…

SafetyDGX agent

Wow, time for a social media detox. Just today - a call for violence against me - insults (that’s every day) - flagrant lies about my credentials - claims that I advocated violence when i repeatedly c

11 Apr 2026

After @aiDotEngineer, which was full of useful criticism, I remembered that the most confident takes on AI often came from the least exposur…

SafetyDGX agent

After @aiDotEngineer, which was full of useful criticism, I remembered that the most confident takes on AI often came from the least exposure. Rejection is easy, trial and error is expensive. I wrote

Angel investor calls for violence against AI skeptic.* *AI skeptic himself repeatedly decried violence, in favor of boycotts, despite multip…

SafetyDGX agent

An angel investor reportedly called for violence against AI skeptic Gary Marcus, who has consistently and publicly advocated against violent responses, instead promoting boycotts as a form of protest.

At this point how can anybody take seriously @sama’s claim that “Working towards prosperity for everyone, empowering all people, and advanci…

SafetyDGX agent

At this point how can anybody take seriously @sama’s claim that “Working towards prosperity for everyone, empowering all people, and advancing science and technology are moral obligations for me”, whe

Don’t fuck him. Don’t bomb him. Boycott him.

SafetyDGX agent

Gary Marcus, an AI researcher and cognitive scientist, posted a tweet advocating for a boycott as a nonviolent, non-confrontational form of protest or opposition against an unnamed individual. The pos

@GaryMarcus @sama the gap between altmans public statements and openais actual behavior has been widening steadily for 2 years now. at some …

SafetyDGX agent

@GaryMarcus @sama the gap between altmans public statements and openais actual behavior has been widening steadily for 2 years now. at some point 'we want to benefit humanity' and 'we want zero liabil

👇 “I still maintain that LLMs are really dumb and of limited use. it’s my opinion as a practitioner and professional that the hype being pe…

SafetyDGX agent

👇 “I still maintain that LLMs are really dumb and of limited use. it’s my opinion as a practitioner and professional that the hype being peddled by the AI corporations and self promoters (from X grift

memory isn't a retrieval widget. it's write policy. the harness decides what survives compaction, what gets promoted, and what becomes reusa…

SafetyDGX agent

memory isn't a retrieval widget. it's write policy. the harness decides what survives compaction, what gets promoted, and what becomes reusable state. https://x.com/hwchase17/status/204297850056760973

postscript:

SafetyDGX agent

postscript: Violence is not the answer. Boycott is the answer. This one’s easy. There is ample evidence that Altman is a dishonest person with inordinate power that we should not trust. But the way fo

@sterlingcrispin @sama Violence was unjustified, and i don’t support it. But many many people may die at Altman’s hands and it is important …

SafetyDGX agent

@sterlingcrispin @sama Violence was unjustified, and i don’t support it. But many many people may die at Altman’s hands and it is important to speak out when he lies about his moral stance. The headli

Violence is not the answer. Boycott is the answer. This one’s easy. There is ample evidence that Altman is a dishonest person with inordinat…

SafetyDGX agent

Violence is not the answer. Boycott is the answer. This one’s easy. There is ample evidence that Altman is a dishonest person with inordinate power that we should not trust. But the way forward to is

What happened to Sam Altman and his family is really awful. It is hard to reconcile his call to “de-escalate the rhetoric and tactics” with …

SafetyDGX agent

What happened to Sam Altman and his family is really awful. It is hard to reconcile his call to “de-escalate the rhetoric and tactics” with his implication that a piece of critical journalism (Ronan F

X search is completely unreliable. An undergrad could make a better system for classifying trends. And yet they are supposed to be part of a…

SafetyDGX agent

Gary Marcus, a cognitive scientist and AI critic, posted on X criticizing the platform's search functionality, arguing it is poorly designed for classifying and surfacing trends. He contrasts its inad

You know when you’re making the right calls when the haters start to get louder. We may not always see to eye on everything, I tend to encou…

SafetyDGX agent

You know when you’re making the right calls when the haters start to get louder. We may not always see to eye on everything, I tend to encourage development in areas that you don’t but I still highly

10 Apr 2026

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

SafetyDGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

A cry for the kind of extreme unfettered capitalism that could get us all killed and make all of @ESYudkowsky’s worst nightmares come true.

SafetyDGX agent

A cry for the kind of extreme unfettered capitalism that could get us all killed and make all of @ESYudkowsky’s worst nightmares come true. AI needs open markets and open access not whatever is happen

A First Guess is Rarely the Final Answer: Learning to Search in the Travelling Salesperson Problem

SafetyDGX agent

arXiv:2604.06940v1 Announce Type: cross Abstract: Most neural solvers for the Traveling Salesperson Problem (TSP) are trained to output a single solution, even though practitioners rarely stop there:

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring

SafetyDGX agent

arXiv:2604.07395v1 Announce Type: cross Abstract: Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an act

A systematic framework for generating novel experimental hypotheses from language models

SafetyDGX agent

arXiv:2408.05086v3 Announce Type: replace Abstract: Neural language models (LMs) have been shown to capture complex linguistic patterns, yet their utility in understanding human language and more broa

Active Reward Machine Inference From Raw State Trajectories

SafetyDGX agent

arXiv:2604.07480v1 Announce Type: new Abstract: Reward machines are automaton-like structures that capture the memory required to accomplish a multi-stage task. When combined with reinforcement learni

ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration

SafetyDGX agent

arXiv:2604.08534v1 Announce Type: new Abstract: Large-scale real-world robot data collection is a prerequisite for bringing robots into everyday deployment. However, existing pipelines often rely on s

Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2512.10510v2 Announce Type: replace-cross Abstract: Offline-to-Online Reinforcement Learning (O2O RL) faces a critical dilemma in balancing the use of a fixed offline dataset with newly collecte

AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power

SafetyDGX agent

arXiv:2604.07007v1 Announce Type: cross Abstract: Autonomous AI agents are beginning to operate across organizational boundaries on the open internet -- discovering, transacting with, and delegating t

AI-Driven Research for Databases

SafetyDGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

Alternatives to the Laplacian for Scalable Spectral Clustering with Group Fairness Constraints

SafetyDGX agent

arXiv:2510.20220v3 Announce Type: replace Abstract: Recent research has focused on mitigating algorithmic bias in clustering by incorporating fairness constraints into algorithmic design. Notions such

An Agentic Evaluation Architecture for Historical Bias Detection in Educational Textbooks

SafetyDGX agent

arXiv:2604.07883v1 Announce Type: cross Abstract: History textbooks often contain implicit biases, nationalist framing, and selective omissions that are difficult to audit at scale. We propose an agen

Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions

SafetyDGX agent

arXiv:2604.07277v1 Announce Type: cross Abstract: Online reinforcement learning (RL) serves as an effective method for enhancing the capabilities of Android agents. However, guiding agents to learn th

Are Face Embeddings Compatible Across Deep Neural Network Models?

SafetyDGX agent

arXiv:2604.07282v1 Announce Type: cross Abstract: Automated face recognition has made rapid strides over the past decade due to the unprecedented rise of deep neural network (DNN) models that can be t

Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries

SafetyDGX agent

arXiv:2604.06416v1 Announce Type: cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We ev

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

SafetyDGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

Beyond Loss Values: Robust Dynamic Pruning via Loss Trajectory Alignment

SafetyDGX agent

arXiv:2604.07306v1 Announce Type: cross Abstract: Existing dynamic data pruning methods often fail under noisy-label settings, as they typically rely on per-sample loss as the ranking criterion. This

Beyond Pessimism: Offline Learning in KL-regularized Games

SafetyDGX agent

arXiv:2604.06738v1 Announce Type: cross Abstract: We study offline learning in KL-regularized two-player zero-sum games, where policies are optimized under a KL constraint to a fixed reference policy.

Beyond Surface Judgments: Human-Grounded Risk Evaluation of LLM-Generated Disinformation

SafetyDGX agent

arXiv:2604.06820v1 Announce Type: new Abstract: Large language models (LLMs) can generate persuasive narratives at scale, raising concerns about their potential use in disinformation campaigns. Assess

Bias Redistribution in Visual Machine Unlearning: Does Forgetting One Group Harm Another?

SafetyDGX agent

arXiv:2604.08111v1 Announce Type: cross Abstract: Machine unlearning enables models to selectively forget training data, driven by privacy regulations such as GDPR and CCPA. However, its fairness impl

Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules

Model ReleasesDGX agent

arXiv:2604.06233v1 Announce Type: new Abstract: Safety-trained language models routinely refuse requests for help circumventing rules. But not all rules deserve compliance. When users ask for help eva

Brain3D: EEG-to-3D Decoding of Visual Representations via Multimodal Reasoning

SafetyDGX agent

arXiv:2604.08068v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) has recently achieved promising results, primarily focusing on reconstructing two-dimensio

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

SafetyDGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

Candles are more regulated than AI. And candle manufacturers aren’t AFAIK lobbying for absolute freedom from liability if they fuck up. Open…

SafetyDGX agent

Candles are more regulated than AI. And candle manufacturers aren’t AFAIK lobbying for absolute freedom from liability if they fuck up. OpenAI is truly appalling. @GaryMarcus @ESYudkowsky Seriously fu

CNN-based Surface Temperature Forecasts with Ensemble Numerical Weather Prediction

SafetyDGX agent

arXiv:2507.18937v3 Announce Type: replace-cross Abstract: Due to limited computational resources, medium-range temperature forecasts typically rely on low-resolution numerical weather prediction (NWP)

Contextualising (Im)plausible Events Triggers Figurative Language

SafetyDGX agent

arXiv:2604.07885v1 Announce Type: new Abstract: This work explores the connection between (non-)literalness and plausibility at the example of subject-verb-object events in English. We design a system

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

SafetyDGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

SafetyDGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

Discrete Flow Matching Policy Optimization

SafetyDGX agent

arXiv:2604.06491v1 Announce Type: cross Abstract: We introduce Discrete flow Matching policy Optimization (DoMinO), a unified framework for Reinforcement Learning (RL) fine-tuning Discrete Flow Matchi

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

SafetyDGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation

SafetyDGX agent

arXiv:2602.13669v4 Announce Type: replace Abstract: Recent multi-modal video generation models have achieved high visual quality, but their prohibitive latency and limited temporal stability hinder re

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

SafetyDGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

Epic corruption, endless propaganda, security force crackdowns and false flag operations, election interference by a hostile foreign dictato…

SafetyDGX agent

Epic corruption, endless propaganda, security force crackdowns and false flag operations, election interference by a hostile foreign dictatorship… That’s Orbán’s Hungary today and, unless America wake

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

SafetyDGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization

SafetyDGX agent

arXiv:2604.08476v1 Announce Type: new Abstract: Multimodal reasoning models (MRMs) trained with reinforcement learning with verifiable rewards (RLVR) show improved accuracy on visual reasoning benchma

Force-Aware Residual DAgger via Trajectory Editing for Precision Insertion with Impedance Control

SafetyDGX agent

arXiv:2603.04038v2 Announce Type: replace Abstract: Imitation learning (IL) has shown strong potential for contact-rich precision insertion tasks. However, its practical deployment is often hindered b

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

SafetyDGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

SafetyDGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

SafetyDGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly avai…

SafetyDGX agent

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly available - and yet I bet he still could crush any LLM that didn

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

SafetyDGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

SafetyDGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

← Previous
1…216217218219220…240
Next →