AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents

DGX agent

arXiv:2607.12790v1 Announce Type: new Abstract: Self-evolving agent systems improve by creating, revising, and retiring their own skills, but every such loop rests on a hidden assumption: a reliable e

safetyarxiv-cs-ai
15 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

DGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

safetygary-marcus--x
14 Jul 2026
Safety

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

DGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

safetyyoshua-bengio--x
14 Jul 2026
Safety

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

DGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

safetyaws-ml-blog
14 Jul 2026
Safety

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years t…

DGX agent

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years to reflect on Kenneth Stanley's ideas as originally communica

safetydavid-ha--x
11 Jul 2026
Safety

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But tha…

DGX agent

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But that’s exactly what loop engineering is: adding symbols to deep

safetygary-marcus--x
11 Jul 2026
Safety

The Anglo-Scottish Enlightenment – the real antidote to Rousseau and Voltaire The French Enlightenment and the Anglo-Scottish Enlightenment …

DGX agent

The Anglo-Scottish Enlightenment – the real antidote to Rousseau and Voltaire The French Enlightenment and the Anglo-Scottish Enlightenment happened simultaneously, in the same century, reading the sa

safetyelon-musk--x
11 Jul 2026
Safety

@theo What @GaryMarcus has been saying ... the engineering around LLMs matters even more than the LLMs themselves today

DGX agent

Gary Marcus argues that the engineering and infrastructure surrounding large language models are more critical to their practical success than the models themselves. This reflects his broader perspect

safetygary-marcus--x
11 Jul 2026
Safety

A Collaborative Reasoning Framework for Anomaly Diagnostics in Underwater Robotics

DGX agent

arXiv:2511.03075v2 Announce Type: replace Abstract: The safe deployment of autonomous systems in safety-critical settings requires a paradigm that combines human expertise with AI-driven analysis, esp

safetyarxiv-cs-ro
10 Jul 2026
Safety

A First-Principles Theory of Slow Thinking and Active Perception

DGX agent

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perce

safetyarxiv-cs-ai
10 Jul 2026
Safety

A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation (Noam Scheiber/New York Times)

DGX agent

Noam Scheiber / New York Times: A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation — A fed

safetytechmeme
10 Jul 2026
Safety

ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning

DGX agent

arXiv:2607.08443v1 Announce Type: cross Abstract: Dynamic traffic variations in Open Radio Access Networks (O-RAN) lead to drift, which degrades the performance of Artificial Intelligence/Machine Lear

safetyarxiv-cs-ai
10 Jul 2026
Safety

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

DGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

safetyconnor-leahy--x
10 Jul 2026
Safety

Aleena: Alignment Agent for Research Software Engineering Collaborations

DGX agent

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a

safetyarxiv-cs-ai
10 Jul 2026
Safety

Alignment Plausibility: A New Standard for Assuring AI in Healthcare

DGX agent

arXiv:2607.07766v1 Announce Type: new Abstract: Large language models (LLMs) have become significant providers of mental health support, yet they remain products of an attention economy whose operatio

safetyarxiv-cs-ai
10 Jul 2026
Safety

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into …

DGX agent

As part of our ongoing efforts to strengthen our safeguards for advanced AI capabilities in biology, we’re evolving our Bio Bug Bounty into an ongoing private program, known as the OpenAI Bio Bug Boun

safetyopenai--x
10 Jul 2026
Safety

Bayesian Experimental Design via Score Matching

DGX agent

arXiv:2607.08335v1 Announce Type: cross Abstract: Policy-based approaches to Bayesian experimental design (BED) allow the learning of deep policy networks that adaptively make intelligent design decis

safetyarxiv-cs-lg
10 Jul 2026
Safety

Best-of-N TTS Evaluation is Confounded by ASR Family Alignment

DGX agent

arXiv:2607.08256v1 Announce Type: cross Abstract: Best-of-N (BoN) inference improves content consistency in zero-shot text-to-speech by selecting from N candidates with an automatic speech recognition

safetyarxiv-cs-ai
10 Jul 2026
Safety

Beyond Success Rates: Trainability and Extractability for Offline GCRL

DGX agent

arXiv:2602.05459v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) is typically benchmarked by the best tuned success rate of each method. This score measures a

safetyarxiv-cs-lg
10 Jul 2026
Safety

Borrowing from anything: A generalizable framework for reference-guided instance editing

DGX agent

arXiv:2512.15138v2 Announce Type: replace Abstract: Reference-guided instance editing is fundamentally limited by semantic entanglement, where a reference's intrinsic appearance is intertwined with it

safetyarxiv-cs-cv
10 Jul 2026
Safety

breaking: company built on stolen IP and lies allegedly steals more IP

DGX agent

breaking: company built on stolen IP and lies allegedly steals more IP “OpenAI’s nascent hardware business now rests on the shakiest of foundations, rotten to its core by its illegal reliance on misap

safetygary-marcus--x
10 Jul 2026
Safety

Bridging Cognitive Neuroscience and Graph Intelligence: Hippocampus-Inspired Multi-View Hypergraph Learning for Web Finance Fraud

DGX agent

arXiv:2601.11073v3 Announce Type: replace-cross Abstract: Online financial services constitute an essential component of contemporary web ecosystems, yet their openness introduces substantial exposure

safetyarxiv-cs-ai
10 Jul 2026
Safety

CAAD: Causality-Aware Multivariate Time Series Anomaly Detection via Multi-Scale Alignment and Structural Causal Consistency

DGX agent

arXiv:2607.08555v1 Announce Type: new Abstract: The operational integrity of complex industrial systems relies on precise anomaly detection and diagnosis. The vast majority of existing methods narrowl

safetyarxiv-cs-lg
10 Jul 2026
Safety

ContactMimic: Humanoid Object Interaction via Contact Control

DGX agent

arXiv:2607.08742v1 Announce Type: new Abstract: Keypoint tracking alone is insufficient for object interaction tasks such as sitting on a chair, wiping a board, or pushing furniture, where the robot c

safetyarxiv-cs-ro
10 Jul 2026
Safety

Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks

DGX agent

arXiv:2607.08561v1 Announce Type: new Abstract: A series of results from the NeuroAI over the past fifteen years have raised core questions both about how to compare Deep Neural Network (DNN) models t

safetyarxiv-cs-lg
10 Jul 2026
Safety

Contributing to U.K. financial sector resilience as a critical third party

DGX agent

At Google Cloud, we take our role in the financial ecosystem very seriously. We firmly believe that operational resilience is essential to driving and sustaining responsible innovation. Today, we mark

safetygoogle-cloud-ai
10 Jul 2026
Safety

Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO

DGX agent

arXiv:2602.17686v4 Announce Type: replace-cross Abstract: Distilling Chain-of-Thought (CoT) reasoning from large language models into compact student models presents a fundamental challenge: teacher r

safetyarxiv-cs-ai
10 Jul 2026
Safety

DeltaDeno: Zero-Shot Anomaly Generation via Delta-Denoising Attribution

DGX agent

arXiv:2511.16920v2 Announce Type: replace Abstract: Anomaly generation is often framed as few-shot fine-tuning with anomalous samples, which contradicts the scarcity that motivates generation and tend

safetyarxiv-cs-cv
10 Jul 2026
Safety

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

DGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

safetyarxiv-cs-cl
10 Jul 2026
Safety

Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models

DGX agent

arXiv:2511.19032v2 Announce Type: replace Abstract: Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer w

safetyarxiv-cs-cv
10 Jul 2026
Safety

DKDNet: Dual Knowledge and Data-Driven Network for Cross-Domain Automatic Modulation Classification

DGX agent

arXiv:2607.08031v1 Announce Type: cross Abstract: The dynamics of communication environments induce significant distribution shifts across domains, challenging the generalization of deep learning-base

safetyarxiv-cs-ai
10 Jul 2026
Safety

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

DGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

safetyarxiv-cs-cl
10 Jul 2026
Safety

DrugGen 2: A disease-aware language model for enhancing drug discovery

DGX agent

arXiv:2607.08404v1 Announce Type: cross Abstract: Current computational approaches for drug design typically focus on generating molecules conditioned on specific targets or general molecular properti

safetyarxiv-cs-ai
10 Jul 2026
Safety

Dual-Difficulty Curriculum Learning for Direct Preference Optimization

DGX agent

arXiv:2504.07856v4 Announce Type: replace Abstract: Curriculum learning enhances Direct Preference Optimization (DPO) for aligning Large Language Models (LLMs), yet existing methods rely on a one-dime

safetyarxiv-cs-ai
10 Jul 2026
Safety

Early to Share, Late to Save: Synchronisation-Driven Communication Gating in Bandwidth-Constrained Cooperative VLN

DGX agent

arXiv:2607.08504v1 Announce Type: cross Abstract: Most cooperative Vision-Language Navigation (VLN) methods assume unlimited communication, not considering real-world applications where bandwidth is r

safetyarxiv-cs-ro
10 Jul 2026
Safety

Echoes: A semantically-aligned music deepfake detection dataset

DGX agent

arXiv:2603.23667v2 Announce Type: replace-cross Abstract: We introduce Echoes, a new dataset for music deepfake detection designed for training and benchmarking detectors under realistic and provider-

safetyarxiv-cs-ai
10 Jul 2026
Safety

Efficient Partitioning Method of Large-Scale Public Safety Spatio-Temporal Data based on Information Loss Constraints

DGX agent

arXiv:2306.12857v3 Announce Type: replace Abstract: The storage, management, and application of massive spatio-temporal data are widely used in practical scenarios, including public safety. However, d

safetyarxiv-cs-lg
10 Jul 2026
Safety

Efficient Safety Alignment of Language Models via Latent Personality Traits

DGX agent

arXiv:2607.07918v1 Announce Type: cross Abstract: Current safety methods for large language models are known to be vulnerable to adversarial attacks, motivating research into robust alternatives. Late

safetyarxiv-cs-ai
10 Jul 2026
Safety

EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data

DGX agent

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scene

safetyarxiv-cs-ai
10 Jul 2026
Safety

Ensemble Diversity Optimization for Subjective Supervision

DGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

safetyarxiv-cs-cl
10 Jul 2026
Safety

EU finds that Meta breached bloc’s rules with its social network interfaces

DGX agent

The European Union has tentatively found that Meta Platforms Inc. breached the bloc’s DSA tech industry law. The European Commission, the EU’s executive arm, published its conclusions today. The DSA,

safetysiliconangle
10 Jul 2026
Safety

Expressivity and Statistical Trade-offs in Diffusion Policy Learning

DGX agent

arXiv:2607.07967v1 Announce Type: cross Abstract: Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action

safetyarxiv-cs-lg
10 Jul 2026
Safety

Feedback Manipulation Regularization: Enabling Offline Agent Alignment for Imitation Learning

DGX agent

arXiv:2607.07859v1 Announce Type: new Abstract: Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While h

safetyarxiv-cs-ai
10 Jul 2026
Safety

From Prompts to Contracts: Harness Engineering for Auditable Enterprise LLM Agents

DGX agent

arXiv:2607.08028v1 Announce Type: new Abstract: Enterprise large language model (LLM) applications often begin as prototypes whose behavior is carried by prompts and retrieval context. Productization

safetyarxiv-cs-ai
10 Jul 2026
Safety

Geometry-Aware Deep Congruence Networks for Manifold Learning in Cross-Subject Motor Imagery

DGX agent

arXiv:2511.18940v3 Announce Type: replace Abstract: Cross-subject motor imagery decoding remains a fundamental challenge in EEG-based brain-computer interfaces due to substantial inter-subject variabi

safetyarxiv-cs-lg
10 Jul 2026
Safety

HairWeaver: Few-Shot Photorealistic Hair Motion Synthesis with Sim-to-Real Guided Video Diffusion

DGX agent

arXiv:2602.11117v2 Announce Type: replace Abstract: We present HairWeaver, a diffusion-based pipeline that animates a single human image with realistic and expressive hair dynamics. While existing met

safetyarxiv-cs-cv
10 Jul 2026
Safety

HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

DGX agent

arXiv:2601.22448v2 Announce Type: replace-cross Abstract: RLVR has become a standard recipe for training LLMs on reasoning tasks with verifiable outcomes, but when rollout generation dominates the cos

safetyarxiv-cs-cl
10 Jul 2026
Safety

HSA: Hierarchical Slot Attention for Multi-granularity Scene-Decomposition

DGX agent

arXiv:2607.08249v1 Announce Type: new Abstract: Slot attention is a powerful framework for object-centric learning, decomposing visual scenes into latent slots through iterative competitive attention.

safetyarxiv-cs-cv
10 Jul 2026
← Previous
1…4546474849…265
Next →