AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
15 Jul 2026

ExToken: Structured Exploration for Efficient Vision-Language-Action Reinforcement Fine-tuning

SafetyDGX agent

arXiv:2607.12931v1 Announce Type: new Abstract: Reinforcement Learning (RL) has demonstrated significant potential for improving Vision-Language-Action (VLA) models on complex manipulation tasks. Howe

FlowWAM: Optical Flow as a Unified Action Representation for World Action Models

SafetyDGX agent

arXiv:2607.13017v1 Announce Type: cross Abstract: World Action Models (WAMs) are able to leverage pretrained video generators for both world modeling and action prediction. However, directly leveragin

From Geometric Recovery to Causal Validation: A Reproducible Audit of Sparse Autoencoder Features, from Superposition Geometry to Causal Inertness

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.12166v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are the standard for decomposing superposed neural representations into interpretable features, and evaluation relies predomi

Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

SafetyDGX agent

arXiv:2607.12463v1 Announce Type: new Abstract: Coding agents must integrate external tool returns into ongoing reasoning - a capability that standard left-to-right pretraining on code exposes only in

GaitSpan: Growing Humanoid Locomotion from Walking to Running

SafetyDGX agent

arXiv:2607.12114v1 Announce Type: cross Abstract: A humanoid that can walk should not relearn locomotion from scratch to jog or run. Yet current approaches often obtain gait diversity by prescribing g

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and O…

Model ReleasesDGX agent

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and OpenEvidence in a Stanford-Harvard clinical AI safety study.

Gradient Flow Dynamics and Implicit Bias of Diagonal Linear Networks under Infinitesimal Initialization

SafetyDGX agent

arXiv:2607.12332v1 Announce Type: new Abstract: We study the gradient flow dynamics of diagonal linear networks for regression tasks under infinitesimal initialization. Extending Theorem 1 from Pesme

Gradient-free learning of a closed-loop wall controller for turbulent drag reduction

SafetyDGX agent

arXiv:2607.12626v1 Announce Type: cross Abstract: Closed-loop wall control learnt by multi-agent reinforcement learning can lower skin-friction drag in turbulent channels, but these gradient-based pol

Growing a Tail: Increasing Output Diversity in Large Language Models

SafetyDGX agent

arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

SafetyDGX agent

arXiv:2607.12752v1 Announce Type: cross Abstract: While recent advances in 3D generation have enabled impressive visual synthesis, existing methods often rely on 2D diffusion supervision without expli

How to Analyze and Govern Gemini Enterprise App Usage at Scale with BigQuery

Model ReleasesDGX agent

Deploying the Gemini Enterprise app across an organization marks a transformative leap forward in workforce productivity, providing employees with an amazing, high-performance suite of agentic AI tool

HSEmotion Team at the 11th ABAW Challenge: Multi-Task Learning and Ambivalence/Hesitancy Video Recognition

SafetyDGX agent

arXiv:2607.12774v1 Announce Type: cross Abstract: This article presents our results for the 11th Affective Behavior Analysis in-the-Wild (ABAW) competition. For multi-task learning with simultaneous p

IDC: Why the right networking approach is foundational to agentic AI

SafetyDGX agent

Editor’s note: Today we hear from IDC on the results of its 2026 AI in Networking Special Report Survey exploring the enterprises' concerns about networking infrastructure to support the rise of agent

I'm Sorry, but I Can't Help with Braille: Revealing Accessibility Failures in State-of-the-Art LLMs

SafetyDGX agent

arXiv:2607.11893v1 Announce Type: cross Abstract: Large Language Models (LLMs) perform strongly on many language tasks, but their capability in structurally constrained, accessibility-critical modalit

Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts

SafetyDGX agent

arXiv:2607.11656v2 Announce Type: replace-cross Abstract: Accurate diagnostic classification and disease-severity prediction for Alzheimer's disease are hampered by the incompleteness and heterogeneit

JADE: Expert-Grounded Dynamic Evaluation for Open-Ended Professional Tasks

SafetyDGX agent

arXiv:2602.06486v4 Announce Type: replace Abstract: Evaluating agentic AI on open-ended professional tasks faces a fundamental dilemma between rigor and flexibility. Static rubrics provide rigorous, r

Learning When to Trust in Contextual Social Bandits

SafetyDGX agent

arXiv:2603.13356v2 Announce Type: replace Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget. We

Lost in Visual Translation: A VLM-Assisted Perceptual-Semantic Coherence Framework for EEG-to-Image Reconstruction

SafetyDGX agent

arXiv:2607.12364v1 Announce Type: cross Abstract: EEG-to-image evaluation should distinguish visual fidelity from recoverable meaning. Yet EEG-derived reconstructions are blurry, distorted, and low-de

M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention

SafetyDGX agent

arXiv:2601.14776v3 Announce Type: replace Abstract: Recent advances in multi-modal detection have significantly improved detection accuracy in challenging environments (e.g., low light, overexposure).

MAMMOTH: A Multi-Modal End-to-End Policy for Off-Road Mobility Robust to Missing Modality

SafetyDGX agent

arXiv:2607.12965v1 Announce Type: new Abstract: Reliable autonomous navigation in unstructured off-road environments remains a critical unsolved challenge due to extreme terrain diversity, drastic ill

MESH: Scaling Up Retrieval with Heterogeneous Content Unification

SafetyDGX agent

arXiv:2607.12392v1 Announce Type: cross Abstract: Optimizing large-scale retrieval hinges on the ability to efficiently surface candidates across diverse content tiers. However, to capture segments su

Mistake gating leads to energy and memory efficient continual learning

SafetyDGX agent

arXiv:2604.14336v2 Announce Type: replace Abstract: Synaptic plasticity is metabolically expensive, yet animals continuously update their internal models without exhausting energy reserves. However, w

MUSA-PINN: Multi-scale Weak-form Physics-Informed Neural Networks for Fluid Flow in Complex Geometries

SafetyDGX agent

arXiv:2603.08465v3 Announce Type: replace Abstract: While Physics-Informed Neural Networks (PINNs) offer a mesh-free approach to solving fluid-flow PDEs, standard point-wise residual minimization suff

“one 80% price cut ends OpenAI and Anthropic” possibly overstated but directionally correct.

SafetyDGX agent

“one 80% price cut ends OpenAI and Anthropic” possibly overstated but directionally correct. Everyone is debating the price of tokens. The bigger question is why they keep getting cheaper. As AI syste

Operationalising Multi-Dimensional Evaluation for Conversational Agents: A Scalable, Governed Pipeline with Selective Re-evaluation and Model Benchmarking

SafetyDGX agent

arXiv:2607.12085v1 Announce Type: new Abstract: Evaluating retail conversational agents requires methods beyond lexical-overlap metrics to assess intent alignment, factuality, helpfulness, clarity, to

Policy-Conditioned Constrained Decoding for Column-Level Access Control in Text-to-SQL

SafetyDGX agent

arXiv:2607.12341v1 Announce Type: new Abstract: Text-to-SQL is increasingly deployed across trust boundaries between data providers and users. Such deployment must balance three competing requirements

PoseAlign: Sculpting Pose-Consistent Meshes via Text-Guided Deformation

SafetyDGX agent

arXiv:2607.10560v2 Announce Type: replace-cross Abstract: Mesh deformation, the process of altering the vertex positions of a 3D mesh while preserving its topological structure, is a cornerstone of co

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs

Model ReleasesDGX agent

arXiv:2607.12985v1 Announce Type: new Abstract: Aligned language models routinely misreport under non-evidential incentive pressure: they agree with a confident user or overstate certainty even when t

Robot Drummer: Learning Rhythmic Skills for Humanoid Drumming

SafetyDGX agent

arXiv:2507.11498v3 Announce Type: replace Abstract: Humanoid robots have seen remarkable advances in dexterity, balance, and locomotion, yet their role in expressive domains such as music performance

Sat2RealCity: Geometry-Aware and Appearance-Controllable 3D Urban Generation from Satellite Imagery

SafetyDGX agent

arXiv:2511.11470v3 Announce Type: replace Abstract: 3D urban generation from satellite imagery is an important task for scalable digital twins and real-world simulation environments. Existing approach

Scalable Optimal Transport Algorithm for Network Alignment

SafetyDGX agent

arXiv:2607.11952v1 Announce Type: new Abstract: Network alignment identifies node correspondences across different networks and is a fundamental primitive in many data science applications, including

Structure-Semantic Co-optimized Latent Diffusion Model for Fast Visual Anagram Synthesis

SafetyDGX agent

arXiv:2606.16241v3 Announce Type: replace Abstract: Visual anagram is an intriguing form of art creation wherein a single image presents different conceptual interpretations under transformations such

TADPO: Reinforcement Learning Goes Off-road

SafetyDGX agent

arXiv:2603.05995v2 Announce Type: replace-cross Abstract: Off-road autonomous driving poses significant challenges such as navigating unmapped, variable terrain with uncertain and diverse dynamics. Ad

Text-Aided Multi-Modal Panoptic Symbol Spotting for CAD Floor Plan Drawings

SafetyDGX agent

arXiv:2607.12678v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) floor plan drawings contain both graphical primitives and textual annotations, which provide complementary geometric and s

The creation of a regulatory body to enforce and implement standards, as suggested here by @demishassabis or by @AnthropicAI last month, wou…

SafetyDGX agent

The creation of a regulatory body to enforce and implement standards, as suggested here by @demishassabis or by @AnthropicAI last month, would be a huge step in the right direction to make AI developm

The Sound of Absence: Audio-Language Embedding Models Struggle with Negation

SafetyDGX agent

arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-o

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the th…

SafetyDGX agent

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the the 13th root of input intelligence [1,2]. That means the inte

Thompson Sampling Is 2-Competitive for Mistakes

SafetyDGX agent

arXiv:2607.12389v1 Announce Type: cross Abstract: We consider Bayesian bandit models and prove that Thompson sampling makes at most twice the expected number of mistakes (selections of a suboptimal ar

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

SafetyDGX agent

arXiv:2511.18089v2 Announce Type: replace Abstract: Multimodal survival analysis aims to improve cancer prognosis using heterogeneous biomedical data, such as histopathology images and genomic profile

TRAIL: A Platform for Configurable Human--AI Teaming Experiments

SafetyDGX agent

arXiv:2607.12180v1 Announce Type: cross Abstract: An AI teammate's design properties (personality, communication style, when it speaks) can shape a team's trust, coordination, and decisions. Studying

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

SafetyDGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

SafetyDGX agent

arXiv:2603.04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems. In X-ray imaging, acquisition

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

SafetyDGX agent

arXiv:2607.12588v1 Announce Type: new Abstract: According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, li

Vision-Based Dribbling for Humanoid Soccer via Privileged Representation Learning

SafetyDGX agent

arXiv:2607.12702v1 Announce Type: new Abstract: Recent advances in humanoid robotics have highlighted the importance of deployable loco-manipulation skills. Dribbling a soccer ball while evading activ

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation

SafetyDGX agent

arXiv:2607.12356v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a powerful end-to-end paradigm for robotic manipulation by mapping language instructions and 2D visu

Watermark Forensics for Generative Models: An Information-Theoretic Perspective

SafetyDGX agent

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user w

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

SafetyDGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking

SafetyDGX agent

arXiv:2607.12735v1 Announce Type: new Abstract: Companion work showed the grokking delay is causally the time to form task-structured representations, injectable via a contrastive prior. Here we chara

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

SafetyDGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

14 Jul 2026

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

SafetyDGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

SafetyDGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

SafetyDGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

11 Jul 2026

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years t…

SafetyDGX agent

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years to reflect on Kenneth Stanley's ideas as originally communica

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But tha…

SafetyDGX agent

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But that’s exactly what loop engineering is: adding symbols to deep

@theo What @GaryMarcus has been saying ... the engineering around LLMs matters even more than the LLMs themselves today

SafetyDGX agent

Gary Marcus argues that the engineering and infrastructure surrounding large language models are more critical to their practical success than the models themselves. This reflects his broader perspect

10 Jul 2026

A First-Principles Theory of Slow Thinking and Active Perception

SafetyDGX agent

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perce

A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation (Noam Scheiber/New York Times)

SafetyDGX agent

Noam Scheiber / New York Times: A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation — A fed

ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning

SafetyDGX agent

arXiv:2607.08443v1 Announce Type: cross Abstract: Dynamic traffic variations in Open Radio Access Networks (O-RAN) lead to drift, which degrades the performance of Artificial Intelligence/Machine Lear

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it…

SafetyDGX agent

Ahead of a dinner with a US senator, AI researcher Nate Soares (@So8res) was told: 'Don't give them any of the crazy crap. You know, play it cool.' His friends opened with the concern that someone cou

Aleena: Alignment Agent for Research Software Engineering Collaborations

SafetyDGX agent

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a

← Previous
1…8081828384…242
Next →