AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
15 Jul 2026

Growing a Tail: Increasing Output Diversity in Large Language Models

SafetyDGX agent

arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

SafetyDGX agent

arXiv:2607.12752v1 Announce Type: cross Abstract: While recent advances in 3D generation have enabled impressive visual synthesis, existing methods often rely on 2D diffusion supervision without expli

HSEmotion Team at the 11th ABAW Challenge: Multi-Task Learning and Ambivalence/Hesitancy Video Recognition

SafetyDGX agent

arXiv:2607.12774v1 Announce Type: cross Abstract: This article presents our results for the 11th Affective Behavior Analysis in-the-Wild (ABAW) competition. For multi-task learning with simultaneous p


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

IDC: Why the right networking approach is foundational to agentic AI

SafetyDGX agent

Editor’s note: Today we hear from IDC on the results of its 2026 AI in Networking Special Report Survey exploring the enterprises' concerns about networking infrastructure to support the rise of agent

I'm Sorry, but I Can't Help with Braille: Revealing Accessibility Failures in State-of-the-Art LLMs

SafetyDGX agent

arXiv:2607.11893v1 Announce Type: cross Abstract: Large Language Models (LLMs) perform strongly on many language tasks, but their capability in structurally constrained, accessibility-critical modalit

Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts

SafetyDGX agent

arXiv:2607.11656v2 Announce Type: replace-cross Abstract: Accurate diagnostic classification and disease-severity prediction for Alzheimer's disease are hampered by the incompleteness and heterogeneit

Internet of Agentic Things: Networked AI Agents for Closed-Loop IoT Orchestration

SafetyDGX agent

arXiv:2607.12662v1 Announce Type: new Abstract: The paper introduces the Internet of Agentic Things (IoAT), an architectural framework that integrates agentic AI, IoT, cyber-physical systems, Physical

Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions

SafetyDGX agent

arXiv:2607.12406v1 Announce Type: new Abstract: The capability of LLM agents to function as the ``brain'' of a system fundamentally expands the scope of analysis beyond a standalone model. Consequentl

JADE: Expert-Grounded Dynamic Evaluation for Open-Ended Professional Tasks

SafetyDGX agent

arXiv:2602.06486v4 Announce Type: replace Abstract: Evaluating agentic AI on open-ended professional tasks faces a fundamental dilemma between rigor and flexibility. Static rubrics provide rigorous, r

Learning When to Trust in Contextual Social Bandits

SafetyDGX agent

arXiv:2603.13356v2 Announce Type: replace Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget. We

Lost in Visual Translation: A VLM-Assisted Perceptual-Semantic Coherence Framework for EEG-to-Image Reconstruction

SafetyDGX agent

arXiv:2607.12364v1 Announce Type: cross Abstract: EEG-to-image evaluation should distinguish visual fidelity from recoverable meaning. Yet EEG-derived reconstructions are blurry, distorted, and low-de

M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention

SafetyDGX agent

arXiv:2601.14776v3 Announce Type: replace Abstract: Recent advances in multi-modal detection have significantly improved detection accuracy in challenging environments (e.g., low light, overexposure).

MAMMOTH: A Multi-Modal End-to-End Policy for Off-Road Mobility Robust to Missing Modality

SafetyDGX agent

arXiv:2607.12965v1 Announce Type: new Abstract: Reliable autonomous navigation in unstructured off-road environments remains a critical unsolved challenge due to extreme terrain diversity, drastic ill

MESH: Scaling Up Retrieval with Heterogeneous Content Unification

SafetyDGX agent

arXiv:2607.12392v1 Announce Type: cross Abstract: Optimizing large-scale retrieval hinges on the ability to efficiently surface candidates across diverse content tiers. However, to capture segments su

Mistake gating leads to energy and memory efficient continual learning

SafetyDGX agent

arXiv:2604.14336v2 Announce Type: replace Abstract: Synaptic plasticity is metabolically expensive, yet animals continuously update their internal models without exhausting energy reserves. However, w

Model-Based Diffusion Optimal Control for Multi-Robot Motion Planning

SafetyDGX agent

arXiv:2607.12423v1 Announce Type: new Abstract: Multi-Robot Motion Planning in continuous environments, where robots must generate dynamically feasible, collision-free trajectories, is challenging due

MUSA-PINN: Multi-scale Weak-form Physics-Informed Neural Networks for Fluid Flow in Complex Geometries

SafetyDGX agent

arXiv:2603.08465v3 Announce Type: replace Abstract: While Physics-Informed Neural Networks (PINNs) offer a mesh-free approach to solving fluid-flow PDEs, standard point-wise residual minimization suff

Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering

SafetyDGX agent

arXiv:2603.28583v2 Announce Type: replace-cross Abstract: Despite the success of Vision-Language Models (VLMs), misleading charts remain a significant challenge due to their deceptive visual structure

“one 80% price cut ends OpenAI and Anthropic” possibly overstated but directionally correct.

SafetyDGX agent

“one 80% price cut ends OpenAI and Anthropic” possibly overstated but directionally correct. Everyone is debating the price of tokens. The bigger question is why they keep getting cheaper. As AI syste

Operationalising Multi-Dimensional Evaluation for Conversational Agents: A Scalable, Governed Pipeline with Selective Re-evaluation and Model Benchmarking

SafetyDGX agent

arXiv:2607.12085v1 Announce Type: new Abstract: Evaluating retail conversational agents requires methods beyond lexical-overlap metrics to assess intent alignment, factuality, helpfulness, clarity, to

Policy-Conditioned Constrained Decoding for Column-Level Access Control in Text-to-SQL

SafetyDGX agent

arXiv:2607.12341v1 Announce Type: new Abstract: Text-to-SQL is increasingly deployed across trust boundaries between data providers and users. Such deployment must balance three competing requirements

PoseAlign: Sculpting Pose-Consistent Meshes via Text-Guided Deformation

SafetyDGX agent

arXiv:2607.10560v2 Announce Type: replace-cross Abstract: Mesh deformation, the process of altering the vertex positions of a 3D mesh while preserving its topological structure, is a cornerstone of co

Predictive Modeling of High-Altitude Clear Air Turbulence in the United States: A Machine Learning Approach

SafetyDGX agent

arXiv:2607.11899v1 Announce Type: cross Abstract: High-altitude Clear Air Turbulence (CAT) poses significant risks to aviation safety due to its unpredictability and challenges in detection. This stud

Removable Defects: The Economics and Limits of Deliberate Deficiency

SafetyDGX agent

arXiv:2607.11983v1 Announce Type: cross Abstract: A specialist tolerates blind spots that a generalist does not. Usually this is treated as a cost to be minimized. We treat it as a design variable: a

Robot Drummer: Learning Rhythmic Skills for Humanoid Drumming

SafetyDGX agent

arXiv:2507.11498v3 Announce Type: replace Abstract: Humanoid robots have seen remarkable advances in dexterity, balance, and locomotion, yet their role in expressive domains such as music performance

Sat2RealCity: Geometry-Aware and Appearance-Controllable 3D Urban Generation from Satellite Imagery

SafetyDGX agent

arXiv:2511.11470v3 Announce Type: replace Abstract: 3D urban generation from satellite imagery is an important task for scalable digital twins and real-world simulation environments. Existing approach

Scalable Optimal Transport Algorithm for Network Alignment

SafetyDGX agent

arXiv:2607.11952v1 Announce Type: new Abstract: Network alignment identifies node correspondences across different networks and is a fundamental primitive in many data science applications, including

StratMamba: Strategic and Reactive Stream Partitioning for Path-Efficient LiDAR-Based Obstacle Avoidance

SafetyDGX agent

arXiv:2607.12370v1 Announce Type: new Abstract: This paper proposes StratMamba, a dual-stream Mamba-based temporal modeling architecture, to more efficiently capture long-horizon temporal dependencies

Streamlining stereo differentiable rendering for marker-free real-time tracking of surgical robots

SafetyDGX agent

arXiv:2607.12604v1 Announce Type: new Abstract: Purpose: Marker-based tracking of surgical robots is occlusion-prone in cluttered operating rooms. We evaluate stereo differentiable rendering for marke

Structure-Semantic Co-optimized Latent Diffusion Model for Fast Visual Anagram Synthesis

SafetyDGX agent

arXiv:2606.16241v3 Announce Type: replace Abstract: Visual anagram is an intriguing form of art creation wherein a single image presents different conceptual interpretations under transformations such

TADPO: Reinforcement Learning Goes Off-road

SafetyDGX agent

arXiv:2603.05995v2 Announce Type: replace-cross Abstract: Off-road autonomous driving poses significant challenges such as navigating unmapped, variable terrain with uncertain and diverse dynamics. Ad

Text-Aided Multi-Modal Panoptic Symbol Spotting for CAD Floor Plan Drawings

SafetyDGX agent

arXiv:2607.12678v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) floor plan drawings contain both graphical primitives and textual annotations, which provide complementary geometric and s

The creation of a regulatory body to enforce and implement standards, as suggested here by @demishassabis or by @AnthropicAI last month, wou…

SafetyDGX agent

The creation of a regulatory body to enforce and implement standards, as suggested here by @demishassabis or by @AnthropicAI last month, would be a huge step in the right direction to make AI developm

The Sound of Absence: Audio-Language Embedding Models Struggle with Negation

SafetyDGX agent

arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-o

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the th…

SafetyDGX agent

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the the 13th root of input intelligence [1,2]. That means the inte

Thompson Sampling Is 2-Competitive for Mistakes

SafetyDGX agent

arXiv:2607.12389v1 Announce Type: cross Abstract: We consider Bayesian bandit models and prove that Thompson sampling makes at most twice the expected number of mistakes (selections of a suboptimal ar

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

SafetyDGX agent

arXiv:2511.18089v2 Announce Type: replace Abstract: Multimodal survival analysis aims to improve cancer prognosis using heterogeneous biomedical data, such as histopathology images and genomic profile

Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap

SafetyDGX agent

arXiv:2607.12113v1 Announce Type: cross Abstract: One year ago, the AISLE roadmap argued that autonomous laboratories operated as isolated islands and proposed a grassroots network organized around fi

TRAIL: A Platform for Configurable Human--AI Teaming Experiments

SafetyDGX agent

arXiv:2607.12180v1 Announce Type: cross Abstract: An AI teammate's design properties (personality, communication style, when it speaks) can shape a team's trust, coordination, and decisions. Studying

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

SafetyDGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

SafetyDGX agent

arXiv:2603.04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems. In X-ray imaging, acquisition

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

SafetyDGX agent

arXiv:2607.12588v1 Announce Type: new Abstract: According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, li

Vision-Based Dribbling for Humanoid Soccer via Privileged Representation Learning

SafetyDGX agent

arXiv:2607.12702v1 Announce Type: new Abstract: Recent advances in humanoid robotics have highlighted the importance of deployable loco-manipulation skills. Dribbling a soccer ball while evading activ

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation

SafetyDGX agent

arXiv:2607.12356v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a powerful end-to-end paradigm for robotic manipulation by mapping language instructions and 2D visu

Watermark Forensics for Generative Models: An Information-Theoretic Perspective

SafetyDGX agent

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user w

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

SafetyDGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking

SafetyDGX agent

arXiv:2607.12735v1 Announce Type: new Abstract: Companion work showed the grokking delay is causally the time to form task-structured representations, injectable via a contrastive prior. Here we chara

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

SafetyDGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents

SafetyDGX agent

arXiv:2607.12790v1 Announce Type: new Abstract: Self-evolving agent systems improve by creating, revising, and retiring their own skills, but every such loop rests on a hidden assumption: a reliable e

14 Jul 2026

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

SafetyDGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

SafetyDGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS

SafetyDGX agent

In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to build a Health Insurance Portability and A

11 Jul 2026

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years t…

SafetyDGX agent

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years to reflect on Kenneth Stanley's ideas as originally communica

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But tha…

SafetyDGX agent

For almost two decades people like @YLeCun and @geoffreyhinton dumped on me for saying we need symbols in addition to deep learning. But that’s exactly what loop engineering is: adding symbols to deep

The Anglo-Scottish Enlightenment – the real antidote to Rousseau and Voltaire The French Enlightenment and the Anglo-Scottish Enlightenment …

SafetyDGX agent

The Anglo-Scottish Enlightenment – the real antidote to Rousseau and Voltaire The French Enlightenment and the Anglo-Scottish Enlightenment happened simultaneously, in the same century, reading the sa

@theo What @GaryMarcus has been saying ... the engineering around LLMs matters even more than the LLMs themselves today

SafetyDGX agent

Gary Marcus argues that the engineering and infrastructure surrounding large language models are more critical to their practical success than the models themselves. This reflects his broader perspect

10 Jul 2026

A Collaborative Reasoning Framework for Anomaly Diagnostics in Underwater Robotics

SafetyDGX agent

arXiv:2511.03075v2 Announce Type: replace Abstract: The safe deployment of autonomous systems in safety-critical settings requires a paradigm that combines human expertise with AI-driven analysis, esp

A First-Principles Theory of Slow Thinking and Active Perception

SafetyDGX agent

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perce

A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation (Noam Scheiber/New York Times)

SafetyDGX agent

Noam Scheiber / New York Times: A US NLRB judge rules that Atlassian had illegally fired an employee in 2023 for pushing back against manager layoffs, and orders reinstatement and compensation — A fed

ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning

SafetyDGX agent

arXiv:2607.08443v1 Announce Type: cross Abstract: Dynamic traffic variations in Open Radio Access Networks (O-RAN) lead to drift, which degrades the performance of Artificial Intelligence/Machine Lear

← Previous
1…3536373839…212
Next →