AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

The Sound of Absence: Audio-Language Embedding Models Struggle with Negation

DGX agent

arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-o

safetyarxiv-cs-ai
15 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the th…

DGX agent

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the the 13th root of input intelligence [1,2]. That means the inte

safetygary-marcus--x
15 Jul 2026
Safety

Thompson Sampling Is 2-Competitive for Mistakes

DGX agent

arXiv:2607.12389v1 Announce Type: cross Abstract: We consider Bayesian bandit models and prove that Thompson sampling makes at most twice the expected number of mistakes (selections of a suboptimal ar

safetyarxiv-cs-lg
15 Jul 2026
Safety

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

DGX agent

arXiv:2511.18089v2 Announce Type: replace Abstract: Multimodal survival analysis aims to improve cancer prognosis using heterogeneous biomedical data, such as histopathology images and genomic profile

safetyarxiv-cs-cv
15 Jul 2026
Safety

TRAIL: A Platform for Configurable Human--AI Teaming Experiments

DGX agent

arXiv:2607.12180v1 Announce Type: cross Abstract: An AI teammate's design properties (personality, communication style, when it speaks) can shape a team's trust, coordination, and decisions. Studying

safetyarxiv-cs-ai
15 Jul 2026
Safety

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

DGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

safetyarxiv-cs-ro
15 Jul 2026
Safety

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

DGX agent

arXiv:2603.04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems. In X-ray imaging, acquisition

safetyarxiv-cs-ai
15 Jul 2026
Safety

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

DGX agent

arXiv:2607.12588v1 Announce Type: new Abstract: According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, li

safetyarxiv-cs-ai
15 Jul 2026
Safety

Vision-Based Dribbling for Humanoid Soccer via Privileged Representation Learning

DGX agent

arXiv:2607.12702v1 Announce Type: new Abstract: Recent advances in humanoid robotics have highlighted the importance of deployable loco-manipulation skills. Dribbling a soccer ball while evading activ

safetyarxiv-cs-ro
15 Jul 2026
Safety

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation

DGX agent

arXiv:2607.12356v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a powerful end-to-end paradigm for robotic manipulation by mapping language instructions and 2D visu

safetyarxiv-cs-ro
15 Jul 2026
Safety

Watermark Forensics for Generative Models: An Information-Theoretic Perspective

DGX agent

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user w

safetyarxiv-cs-lg
15 Jul 2026
Safety

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

DGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

safetyarxiv-cs-cl
15 Jul 2026
Safety

What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking

DGX agent

arXiv:2607.12735v1 Announce Type: new Abstract: Companion work showed the grokking delay is causally the time to form task-structured representations, injectable via a contrastive prior. Here we chara

safetyarxiv-cs-lg
15 Jul 2026
Safety

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

DGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

safetyarxiv-cs-lg
15 Jul 2026
Safety

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

DGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

safetygary-marcus--x
14 Jul 2026
Safety

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

DGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

safetyyoshua-bengio--x
14 Jul 2026
Safety

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

DGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

safetyaws-ml-blog
14 Jul 2026
Safety

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years t…

DGX agent

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years to reflect on Kenneth Stanley's ideas as originally communica

safetydavid-ha--x
11 Jul 2026
Safety

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But tha…

DGX agent

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But that’s exactly what loop engineering is: adding symbols to deep

safetygary-marcus--x
11 Jul 2026
Safety

@theo What @GaryMarcus has been saying ... the engineering around LLMs matters even more than the LLMs themselves today

DGX agent

Gary Marcus argues that the engineering and infrastructure surrounding large language models are more critical to their practical success than the models themselves. This reflects his broader perspect

safetygary-marcus--x
11 Jul 2026
Safety

A First-Principles Theory of Slow Thinking and Active Perception

DGX agent

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perce

safetyarxiv-cs-ai
10 Jul 2026
Safety

A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation (Noam Scheiber/New York Times)

DGX agent

Noam Scheiber / New York Times: A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation — A fed

safetytechmeme
10 Jul 2026
Safety

ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning

DGX agent

arXiv:2607.08443v1 Announce Type: cross Abstract: Dynamic traffic variations in Open Radio Access Networks (O-RAN) lead to drift, which degrades the performance of Artificial Intelligence/Machine Lear

safetyarxiv-cs-ai
10 Jul 2026
Safety

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

DGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

safetyconnor-leahy--x
10 Jul 2026
Safety

Aleena: Alignment Agent for Research Software Engineering Collaborations

DGX agent

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a

safetyarxiv-cs-ai
10 Jul 2026
Safety

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into …

DGX agent

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into an ongoing private program, known as the OpenAI Bio Bug Boun

safetyopenai--x
10 Jul 2026
Model Releases

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

DGX agent

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as sc

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

Bayesian Experimental Design via Score Matching

DGX agent

arXiv:2607.08335v1 Announce Type: cross Abstract: Policy-based approaches to Bayesian experimental design (BED) allow the learning of deep policy networks that adaptively make intelligent design decis

safetyarxiv-cs-lg
10 Jul 2026
Safety

Best-of-N TTS Evaluation is Confounded by ASR Family Alignment

DGX agent

arXiv:2607.08256v1 Announce Type: cross Abstract: Best-of-N (BoN) inference improves content consistency in zero-shot text-to-speech by selecting from N candidates with an automatic speech recognition

safetyarxiv-cs-ai
10 Jul 2026
Safety

Beyond Success Rates: Trainability and Extractability for Offline GCRL

DGX agent

arXiv:2602.05459v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) is typically benchmarked by the best tuned success rate of each method. This score measures a

safetyarxiv-cs-lg
10 Jul 2026
Safety

Borrowing from anything: A generalizable framework for reference-guided instance editing

DGX agent

arXiv:2512.15138v2 Announce Type: replace Abstract: Reference-guided instance editing is fundamentally limited by semantic entanglement, where a reference's intrinsic appearance is intertwined with it

safetyarxiv-cs-cv
10 Jul 2026
Safety

breaking: company built on stolen IP and lies allegedly steals more IP

DGX agent

breaking: company built on stolen IP and lies allegedly steals more IP “OpenAI’s nascent hardware business now rests on the shakiest of foundations, rotten to its core by its illegal reliance on misap

safetygary-marcus--x
10 Jul 2026
Safety

Bridging Cognitive Neuroscience and Graph Intelligence: Hippocampus-Inspired Multi-View Hypergraph Learning for Web Finance Fraud

DGX agent

arXiv:2601.11073v3 Announce Type: replace-cross Abstract: Online financial services constitute an essential component of contemporary web ecosystems, yet their openness introduces substantial exposure

safetyarxiv-cs-ai
10 Jul 2026
Safety

CAAD: Causality-Aware Multivariate Time Series Anomaly Detection via Multi-Scale Alignment and Structural Causal Consistency

DGX agent

arXiv:2607.08555v1 Announce Type: new Abstract: The operational integrity of complex industrial systems relies on precise anomaly detection and diagnosis. The vast majority of existing methods narrowl

safetyarxiv-cs-lg
10 Jul 2026
Safety

ContactMimic: Humanoid Object Interaction via Contact Control

DGX agent

arXiv:2607.08742v1 Announce Type: new Abstract: Keypoint tracking alone is insufficient for object interaction tasks such as sitting on a chair, wiping a board, or pushing furniture, where the robot c

safetyarxiv-cs-ro
10 Jul 2026
Safety

Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks

DGX agent

arXiv:2607.08561v1 Announce Type: new Abstract: A series of results from the NeuroAI over the past fifteen years have raised core questions both about how to compare Deep Neural Network (DNN) models t

safetyarxiv-cs-lg
10 Jul 2026
Safety

Contributing to U.K. financial sector resilience as a critical third party

DGX agent

At Google Cloud, we take our role in the financial ecosystem very seriously. We firmly believe that operational resilience is essential to driving and sustaining responsible innovation. Today, we mark

safetygoogle-cloud-ai
10 Jul 2026
Safety

Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO

DGX agent

arXiv:2602.17686v4 Announce Type: replace-cross Abstract: Distilling Chain-of-Thought (CoT) reasoning from large language models into compact student models presents a fundamental challenge: teacher r

safetyarxiv-cs-ai
10 Jul 2026
Safety

DeltaDeno: Zero-Shot Anomaly Generation via Delta-Denoising Attribution

DGX agent

arXiv:2511.16920v2 Announce Type: replace Abstract: Anomaly generation is often framed as few-shot fine-tuning with anomalous samples, which contradicts the scarcity that motivates generation and tend

safetyarxiv-cs-cv
10 Jul 2026
Safety

Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models

DGX agent

arXiv:2511.19032v2 Announce Type: replace Abstract: Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer w

safetyarxiv-cs-cv
10 Jul 2026
Safety

DKDNet: Dual Knowledge and Data-Driven Network for Cross-Domain Automatic Modulation Classification

DGX agent

arXiv:2607.08031v1 Announce Type: cross Abstract: The dynamics of communication environments induce significant distribution shifts across domains, challenging the generalization of deep learning-base

safetyarxiv-cs-ai
10 Jul 2026
Safety

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

DGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

safetyarxiv-cs-cl
10 Jul 2026
Safety

DrugGen 2: A disease-aware language model for enhancing drug discovery

DGX agent

arXiv:2607.08404v1 Announce Type: cross Abstract: Current computational approaches for drug design typically focus on generating molecules conditioned on specific targets or general molecular properti

safetyarxiv-cs-ai
10 Jul 2026
Safety

Dual-Difficulty Curriculum Learning for Direct Preference Optimization

DGX agent

arXiv:2504.07856v4 Announce Type: replace Abstract: Curriculum learning enhances Direct Preference Optimization (DPO) for aligning Large Language Models (LLMs), yet existing methods rely on a one-dime

safetyarxiv-cs-ai
10 Jul 2026
Safety

Early to Share, Late to Save: Synchronisation-Driven Communication Gating in Bandwidth-Constrained Cooperative VLN

DGX agent

arXiv:2607.08504v1 Announce Type: cross Abstract: Most cooperative Vision-Language Navigation (VLN) methods assume unlimited communication, not considering real-world applications where bandwidth is r

safetyarxiv-cs-ro
10 Jul 2026
Safety

Echoes: A semantically-aligned music deepfake detection dataset

DGX agent

arXiv:2603.23667v2 Announce Type: replace-cross Abstract: We introduce Echoes, a new dataset for music deepfake detection designed for training and benchmarking detectors under realistic and provider-

safetyarxiv-cs-ai
10 Jul 2026
Safety

EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data

DGX agent

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scene

safetyarxiv-cs-ai
10 Jul 2026
Safety

Ensemble Diversity Optimization for Subjective Supervision

DGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

safetyarxiv-cs-cl
10 Jul 2026
← Previous
1…101102103104105…302
Next →