AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
10 Aug 2026

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

SafetyDGX agent

arXiv:2608.07068v1 Announce Type: new Abstract: Long-horizon agents accumulate growing contexts during interaction, impairing performance and stability. Compact memory mitigates this problem by compre

MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents

SafetyDGX agent

arXiv:2608.06745v1 Announce Type: new Abstract: Long-horizon agents rely on memory to reuse experiences, yet existing memory systems often assume that evidence can be directly consumed through a fixed

Mind the Gap: A Dual Knowledge Graph Framework for Unified Multi-task User Intent Inference

SafetyDGX agent

arXiv:2608.06752v1 Announce Type: new Abstract: This paper proposes DKG-MTI, a dual knowledge graph framework for unified multi-task user intent inference from online travel reviews. Existing approach

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Minimal Ingredients for Reward Assignment from Expert Demonstrations

SafetyDGX agent

arXiv:2506.06793v2 Announce Type: replace-cross Abstract: Reward assignment from scarce demonstrations is a key challenge in both offline and online imitation learning. A common and intuitive strategy

MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resolution Structural Anchoring

SafetyDGX agent

arXiv:2608.06713v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) accelerate drug discovery, but standard pipelines assume query molecules already exist as graph entities, leaving unre

NTDH: Complex Reasoning for Comprehensive Affective Analysis

SafetyDGX agent

arXiv:2608.06425v1 Announce Type: new Abstract: Comprehensive affective analysis is challenging for two reasons: it spans heterogeneous prediction tasks with continuous, ordinal, and multi-label outpu

PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model

SafetyDGX agent

arXiv:2608.06794v1 Announce Type: new Abstract: While diffusion models have made significant progress in text-to-image tasks, they still exhibit limitations when directly optimizing downstream objecti

People Are Not Just Their Countries. Disentangling Social Determinants of LLM Value Alignment Across Europe

SafetyDGX agent

arXiv:2608.07367v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly used as a primary source of information and advice, understanding their alignment to humans in terms of

Progress-Certified Reversible Simplex Supervision of Goal-Reaching Reinforcement Learning

SafetyDGX agent

arXiv:2601.19499v2 Announce Type: replace Abstract: Task completion is difficult to certify when state aggregation, model mismatch, and disturbances invalidate nominal RL transitions. We present a rev

Progressive Alignment of Recommender Foundation Model through Multi-Phase Post-Training

SafetyDGX agent

arXiv:2608.06792v1 Announce Type: cross Abstract: Foundation model(FM) for recommendation has shown strong ability to model long-horizon sequential user behavior. In practice, a single pretrained foun

R2S-EGO: Dual-Proxy Refinement for Sparse-Capture Real-to-Sim

SafetyDGX agent

arXiv:2608.06827v1 Announce Type: cross Abstract: Real-to-sim (R2S) depends on scene representations that render observations along robot ego trajectories, yet dense multi-view capture limits per-envi

Reading Copom's Tone: A Weighted LLM Framework for Hawkish-Dovish Sentiment, Forward Guidance, and Uncertainty

SafetyDGX agent

arXiv:2608.07251v1 Announce Type: cross Abstract: This paper documents an applied natural-language-processing framework for measuring the tone of Brazilian Monetary Policy Committee (Copom) statements

Really important point and nicely explained! Two further points worth noting: (1) if an AI system relies on a harness (which as Gary notes i…

SafetyDGX agent

Really important point and nicely explained! Two further points worth noting: (1) if an AI system relies on a harness (which as Gary notes in other writing is increasingly improtant in top models), th

Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation

SafetyDGX agent

arXiv:2608.07176v1 Announce Type: cross Abstract: Developing foundation generative models for endoscopy is limited by the gap between natural and clinical images and the computational cost of training

RibAssist 3D: Biplanar Rib-Fracture Detection, Addressing, and Selective 3D Localization from CT-Derived Projections

SafetyDGX agent

arXiv:2608.06914v1 Announce Type: new Abstract: Rib fractures are common, clinically significant, and time-consuming to localize on computed tomography (CT). We ask whether fractures detected in two o

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions

SafetyDGX agent

arXiv:2608.06545v1 Announce Type: new Abstract: Distributionally robust Markov decision processes provide a principled framework for sequential decision making under model uncertainty. We study how ma

Self-Distillation Enables Continual Learning

SafetyDGX agent

arXiv:2601.19897v2 Announce Type: replace Abstract: Continual learning, enabling models to acquire new skills and knowledge without degrading existing capabilities, remains a fundamental challenge for

Shape Your Feed: An LLM-based Agentic System for Conversational Recommendation

SafetyDGX agent

arXiv:2608.06632v1 Announce Type: new Abstract: Industrial recommendation systems predominantly adopt a passive ranking paradigm that infers user preferences from implicit behavioral signals (e.g., cl

SNI-GNN: SmartNIC-Assisted Full-Graph GNN Training with In-Network Embedding Prediction

SafetyDGX agent

arXiv:2608.06441v1 Announce Type: new Abstract: Full-graph GNN training delivers high accuracy but scales poorly on multi-server clusters due to heavy, irregular inter-node embedding exchanges. We pre

Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration (Thomas Barrabi/New York Post)

SafetyDGX agent

Thomas Barrabi / New York Post: Sources: officials say OpenAI risks its White House relationship by hiring Dean Ball, who has criticized Trump's AI strategy after leaving the administration — Top Open

TA-RAG: Tone Awareness as a Design Imperative for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2608.06672v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a robust architecture for grounding large language models (LLMs) in trusted knowledge. However, standard

The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2608.06663v1 Announce Type: new Abstract: Frontier language models solve reasoning problems in a single forward pass that would have been research contributions years ago, yet fail at multi-hour

The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows

SafetyDGX agent

arXiv:2608.06714v1 Announce Type: new Abstract: Recent systems for optimizing prompts, programs, and ML workflows typically rely on explicit outer-loop controllers such as evolutionary search, bandits

TMTE: Effective Multimodal Graph Learning with Task-aware Modality and Topology Co-evolution

SafetyDGX agent

arXiv:2603.27723v2 Announce Type: replace Abstract: Multimodal-attributed graphs (MAGs) are a fundamental data structure for multimodal graph learning (MGL), enabling both graph-centric and modality-c

Toward surface-based registration of a virtual preoperative cutting guide onto the mandible for reconstruction surgery

SafetyDGX agent

arXiv:2608.06599v1 Announce Type: new Abstract: Mandibular reconstruction restores facial continuity and oral function after segmental resection. Patient-specific cutting guides transfer a computed to

Unmasking Removal-Budget Confounding: A Matched Operating-Point Evaluation Framework for Adaptive Data Cleaning

SafetyDGX agent

arXiv:2608.06511v1 Announce Type: new Abstract: Adaptive data-cleaning methods replace manual filtering thresholds with data-driven partitions. However, changing the partition granularity, the number

Wasserstein Policy Gradient for Entropy-Regularized Linear-Quadratic Control

SafetyDGX agent

arXiv:2608.07433v1 Announce Type: cross Abstract: Wasserstein policy gradient (WPG) updates state-conditional action laws by transport in the action space. We study entropy-regularized discounted line

WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN

SafetyDGX agent

arXiv:2608.07267v1 Announce Type: new Abstract: Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into vision-language-action (VLA) policies t

9 Aug 2026

New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizo…

SafetyDGX agent

New research from Meta. Agent harnesses are still mostly authored by hand. This makes it hard to tune robust agent harnesses for long-horizon tasks. In this new work, agents learn harness policies off

8 Aug 2026

and we agree with @GaryMarcus. Training and running frontier class AI on CPUs, at a fractional cost, saving the planet, while building human…

SafetyDGX agent

and we agree with @GaryMarcus. Training and running frontier class AI on CPUs, at a fractional cost, saving the planet, while building human aligned AI is the 2nd innings of AI race. CPUs and the rise

CPUs and the rise of neurosymbolic AI by @garymarcus (with implications for what Artificial Intelligence implies about natural intelligence …

SafetyDGX agent

CPUs and the rise of neurosymbolic AI by @garymarcus (with implications for what Artificial Intelligence implies about natural intelligence -- Gary's & my interest ever since we worked together 37 yea

7 Aug 2026

A Bridge from Audio to Video: Phoneme-Viseme Alignment Allows Every Face to Speak Multiple Languages

SafetyDGX agent

arXiv:2510.06612v2 Announce Type: replace Abstract: Speech-driven talking face synthesis (TFS) focuses on generating lifelike facial animations from speech input. Current TFS models perform well in En

A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation

SafetyDGX agent

arXiv:2608.05244v1 Announce Type: cross Abstract: We developed a unified covariate-adjusted causal inference framework for estimating the desirability of outcome ranking (DOOR) probability for benefit

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.05987v1 Announce Type: new Abstract: Reinforcement learning (RL) with verifiable rewards constructs trajectory-level advantage estimates, yet it often fails to credit the few pivotal decisi

All-Quadrant Bounded Clipping GRPO: Closing the Unbounded Blind Spot for Stable and Generalizable Training

SafetyDGX agent

arXiv:2601.03895v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as a popular algorithm for reinforcement learning with large language models (LLMs). How

Alternating Levenberg-Marquardt Training of Physics-Informed Neural Networks with Fourier-Enhanced Features

SafetyDGX agent

arXiv:2608.05892v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) often fail to accurately resolve partial differential equations (PDEs) with high-frequency or multi-scale solut

An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals

SafetyDGX agent

arXiv:2608.05255v1 Announce Type: cross Abstract: Retail investors lack access to the kind of personalized, tax-aware portfolio management that institutional clients take for granted -- existing robo-

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

SafetyDGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

ARGUS: Aligning Robot Scene Geometry Under Shifting Views with Large 3D Vision Models

SafetyDGX agent

arXiv:2608.05579v1 Announce Type: new Abstract: Large-scale visuomotor policies have demonstrated impressive performance across a wide range of robot manipulation tasks. However, despite this success,

As You Wish: Mission Planning with Formal Verification using LLMs in Precision Agriculture

SafetyDGX agent

arXiv:2606.18519v2 Announce Type: replace-cross Abstract: Though robotic systems are now being commercialized and deployed in various industries, many of these systems are highly specialized and often

Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with Negative Feedback

SafetyDGX agent

arXiv:2509.03206v2 Announce Type: replace-cross Abstract: Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that

Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction

SafetyDGX agent

arXiv:2608.06206v1 Announce Type: cross Abstract: Conformal prediction endows arbitrary black-box predictors with finite-sample, distribution-free marginal coverage, yet marginal validity can hide sev

Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for Political News Evaluation

SafetyDGX agent

arXiv:2608.05155v1 Announce Type: cross Abstract: Traditional sentiment analysis (SA) models, while effective for polarity classification, provide limited insight into the rhetorical, ideological, and

Bias Analysis of L2 Speaking Assessment Systems Using Concept Activation Vectors

SafetyDGX agent

arXiv:2608.06300v1 Announce Type: new Abstract: Automatic speaking assessment systems are increasingly deployed in high-stakes settings to mark second language (L2) learners' speaking tests, making it

Challenges in Evaluating Explanation Methods for Static and Evolving Data

SafetyDGX agent

arXiv:2608.06351v1 Announce Type: new Abstract: This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through

CircuitSteer: Geometrically Aligned Multi-Layer Steering via Sparse Autoencoder Circuits

SafetyDGX agent

arXiv:2608.05732v1 Announce Type: new Abstract: Controlling the behavior of large language models (LLMs) remains a critical challenge for AI alignment. Existing steering methods, such as Contrastive A

Contextual Information Policy Optimization for Search Agents

SafetyDGX agent

arXiv:2608.06128v1 Announce Type: new Abstract: Search agents extend large language models beyond static parametric memory by enabling them to acquire and use ex ternal evidence during multi-step reas

CoordRefer: Coordinate-Aware 3D Visual Grounding from Multiview Images

SafetyDGX agent

arXiv:2608.05569v1 Announce Type: new Abstract: Multiview image-based 3D visual grounding predicts a coordinate frame to define a coordinate system and then regresses a 3D bounding box for localizatio

CourseGraph: Finding overlaps and differences in Computer Science courses across universities

SafetyDGX agent

arXiv:2608.05910v1 Announce Type: new Abstract: Student mobility programs such as Erasmus+ enable students to take courses at other universities, broadening their academic and cultural horizons. Howev

DARAD: Dual Adapters and Ranking-Aware Distillation for Continual Remote Sensing Image-Text Retrieval

SafetyDGX agent

arXiv:2608.06059v1 Announce Type: new Abstract: With the rapid growth of Earth observation technologies, remote sensing archives are rapidly expanding, making remote sensing image-text retrieval (RS-I

Decolonizing Linguistic Policies in Automated Speech Recognition: A Framework for Cross-Culturally Competent Speech AI

SafetyDGX agent

arXiv:2608.06141v1 Announce Type: new Abstract: This paper focuses on automatic speech recognition (ASR) and ASR-mediated voice interfaces that shape access to public services, healthcare, and educati

Different Perturbations, Different Mechanisms: Understanding Continued Pre-training for Zero-Shot Dialect Robustness

SafetyDGX agent

arXiv:2608.05510v1 Announce Type: new Abstract: Dialectal variation remains a major challenge for multilingual language models. Perturbation-based continued pre-training (CPT) has emerged as a promisi

DistMedVL: Distributional Vision-Language Alignment for Uncertainty-Aware Medical Image Segmentation

SafetyDGX agent

arXiv:2608.05683v1 Announce Type: cross Abstract: Cross-modal alignment of visual and textual representations is fundamental to multimodal medical image understanding, yet remains hindered by uncertai

DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation

SafetyDGX agent

arXiv:2608.06374v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a powerful paradigm for robot manipulation, but training a single generalist policy for heterogeneous ro

EmoWorld: A Decoupled Affective Field for Controllable Emotional Video Generation

SafetyDGX agent

arXiv:2608.06231v1 Announce Type: new Abstract: Emotion shapes how viewers interpret a scene, yet existing video generators entangle global atmosphere, affect-bearing semantic cues, and temporal progr

EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.06197v1 Announce Type: new Abstract: Training large language model agents for long-horizon tool use typically relies on interactions with real or synthesized executable environments, whose

Epistemic Trustworthiness in Generative AI: A Normative Framework for Warranted Reliance in High-Stakes Workflows

SafetyDGX agent

arXiv:2608.05602v1 Announce Type: new Abstract: Generative AI systems are increasingly deployed in high-stakes professional contexts, where their outputs shape what users believe, how they reason, and

Estimating time spent on work tasks

SafetyDGX agent

arXiv:2608.05172v1 Announce Type: cross Abstract: The task-based framework in economics models occupations as bundles of tasks. It is the standard lens for understanding how technology affects work: a

Evidence-Driven Dynamic Visual Selector for Efficient Long Video Understanding

SafetyDGX agent

arXiv:2608.05780v1 Announce Type: new Abstract: Recent advancements in MLLM-based long-form video understanding have mitigated inference-time computational cost and limited context lengths by selectin

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

← Previous
1…5758596061…240
Next →