AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
29 Jul 2026

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding

SafetyDGX agent

arXiv:2607.24794v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate superior generalization in fundamental video tasks, restricted context windows limit their lo

RoboHarness: Memory-Driven Orchestration of Heterogeneous Robot Policies for Long-Horizon Planning

SafetyDGX agent

arXiv:2607.18060v2 Announce Type: replace Abstract: Long-horizon robotic tasks require diverse capabilities that no single policy can reliably provide. Heterogeneous policies offer complementary stren

S2A2: Audio-Visual Imitation Learning for Manipulation Tasks Using Acoustic Spatial Information

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.26047v1 Announce Type: new Abstract: Acoustic information provides rich cues about object location, material properties, and changes caused by contact or motion. This paper introduces a new

SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task

SafetyDGX agent

arXiv:2607.24850v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning h

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller

SafetyDGX agent

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MAS

Spectral Truncation in Synthetic Control

SafetyDGX agent

arXiv:2607.25074v1 Announce Type: cross Abstract: Synthetic control (SC) matches a treated unit's pre-treatment trajectory to a weighted combination of donor units. We study Spectral SC, which instead

Structure-aware Relative Policy Optimization for Ranking

SafetyDGX agent

arXiv:2607.25268v1 Announce Type: cross Abstract: Ranking is a fundamental component of modern information access systems. Reinforcement learning (RL) provides a flexible framework for directly optimi

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

SafetyDGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem

SafetyDGX agent

arXiv:2604.14990v2 Announce Type: replace Abstract: The prospect of Artificial General Intelligence (AGI) is increasingly driving institutional decisions, and alignment of AGI is a hard problem. The c

The problem: agent workflows alternate between GPU-heavy reasoning and GPU-idle waiting on tools. Run hundreds concurrently and their KV cac…

SafetyDGX agent

The problem: agent workflows alternate between GPU-heavy reasoning and GPU-idle waiting on tools. Run hundreds concurrently and their KV caches fight for memory. Engines evict on a dumb LRU policy, ev

The Scaling Properties of Implicit Deductive Reasoning in Transformers

SafetyDGX agent

arXiv:2605.04330v2 Announce Type: replace Abstract: We investigate the scaling properties of implicit deductive reasoning over Horn clauses in depth-bounded Transformers. By systematically decorrelati

Time-Frequency Consistency Learning for Robust Speech Deepfake Detection

SafetyDGX agent

arXiv:2607.17761v2 Announce Type: replace-cross Abstract: Recently, speech deepfake detection (SDD) has achieved significant progress. However, its robustness evaluation remains largely confined to co

Tri-Manual Visuomotor Imitation Learning of Robot Policies

SafetyDGX agent

arXiv:2607.25731v1 Announce Type: new Abstract: Bimanual teleoperation provides an effective way to collect robot demonstrations, but it assumes that the operator and robot have matching numbers of si

Tripody: An Overconstrained 3-SPR-like Parallel Robot for High-Reach Construction Tasks

SafetyDGX agent

arXiv:2607.25781v1 Announce Type: new Abstract: Many ceiling construction tasks still rely on heavy serial manipulators that are difficult to deploy in cluttered interiors, motivating lightweight, fie

Unlocking Spatial Grounding in Large Audio-Visual Retrieval models

SafetyDGX agent

arXiv:2607.24786v1 Announce Type: cross Abstract: Weak supervision sets a practical regime for audio-visual sound source localization as dense spatial annotations are costly to obtain at scale. The ta

Variance-Reduced Conditional Gradient Methods under Markovian Sampling for Nonconvex Composite Optimization

SafetyDGX agent

arXiv:2607.25785v1 Announce Type: cross Abstract: We study stochastic composite nonconvex optimization over a compact convex set when gradient samples arrive along a single trajectory of a fixed ergod

Vector-Valued Distributional Reinforcement Learning Policy Evaluation: A Hilbert Space Embedding Approach

SafetyDGX agent

arXiv:2601.18952v2 Announce Type: replace Abstract: We propose an (offline) multi-dimensional distributional reinforcement learning framework (KE-DRL) that leverages Hilbert space mappings to estimate

We ran a large-scale distillation attack on the Kimi K3 technical report by reading it in parallel at the Hugging Face Journal Club :) https…

SafetyDGX agent

We ran a large-scale distillation attack on the Kimi K3 technical report by reading it in parallel at the Hugging Face Journal Club :) https://youtu.be/MW8-kqd2SD8?si=jSKDogcUWbJ8N2k7 Our main takeawa

When Do Agent Loops Mistake Stagnation for Progress? Self-Evaluation Bias and Externally Grounded Verification in Long-Running Autonomous LLM Agent Loops

SafetyDGX agent

arXiv:2607.25152v1 Announce Type: new Abstract: Long-running autonomous agents plan, act, and judge their own completion without human intervention. When an agent grades its own work, self-evaluation

When Does Legacy Data Start to Help? Emergent Transfer in Cross-Configuration Robot Learning

SafetyDGX agent

arXiv:2607.25593v1 Announce Type: new Abstract: Robotic hardware evolves over time, but demonstration data is often tied to a specific sensor and actuator configuration. This raises a practical and un

When Thinking Before Retrieval Hurts: TraceBound Diagnostics for Adaptive Knowledge-Graph Retrieval

SafetyDGX agent

arXiv:2607.24800v1 Announce Type: cross Abstract: Adaptive retrieval promises to make knowledge-graph question answering more robust by letting a controller search, inspect neighborhoods, revise actio

Where Steering Signals Come From: Activation Source Selection in Activation Steering

SafetyDGX agent

arXiv:2607.25270v1 Announce Type: cross Abstract: Activation steering controls language models by adding vectors or features to hidden states at inference time, but the upstream source of these steeri

28 Jul 2026

A Coulomb Particle Model for Learning Kernel Attention in Transformers

SafetyDGX agent

arXiv:2607.23869v1 Announce Type: cross Abstract: Randomized features provide a scalable approximation to kernel machines, but their performance depends strongly on the choice of feature distribution.

A Cyclic Adaptation-Generalization Framework with Uncertainty-Guided Self-Paced Learning for Long-Term Brain-Machine Interfaces

SafetyDGX agent

arXiv:2607.24031v1 Announce Type: new Abstract: Brain-Machine Interfaces (BMIs), which link the brain to external devices, hold great potential in rehabilitation, human performance augmentation, and h

A Multi-stage Constrained Optimization Framework for Data-driven Problems

SafetyDGX agent

arXiv:2607.23480v1 Announce Type: new Abstract: Variational autoencoders (VAEs) transform high-dimensional, often noisy data into a compact latent representation, making downstream optimization more t

ACRL: Adaptive Control of Training-Inference Discrepancy for Stable Reinforcement Learning

SafetyDGX agent

arXiv:2607.24062v1 Announce Type: cross Abstract: Reinforcement Learning (RL) training for Large Language Models (LLMs) often suffers from instability due to the discrepancy between training and infer

Action-Sufficient Goal Representations

SafetyDGX agent

arXiv:2601.22496v2 Announce Type: replace-cross Abstract: In offline goal-conditioned reinforcement learning (GCRL), hierarchical approaches decompose long-horizon tasks into high-level subgoal predic

AGI smarter than the smartest humans, my ass. @Kasparov63 (peak rating 2851) probably could’ve crushed the best commercial large language mo…

SafetyDGX agent

AGI smarter than the smartest humans, my ass. @Kasparov63 (peak rating 2851) probably could’ve crushed the best commercial large language models in chess when he was 7 years old. graph courtesy @chess

Answering Path Queries under Linear and Guarded Existential Rules

SafetyDGX agent

arXiv:2607.22636v1 Announce Type: new Abstract: Ontology-mediated query answering is concerned with the problem of answering queries over knowledge bases consisting of a database instance and an ontol

Are Flat Minima an Illusion?

SafetyDGX agent

arXiv:2605.05209v2 Announce Type: replace-cross Abstract: Flat minima are an account of why deep networks generalise. However flatness is a matter of form (parameters), while generalisation is of func

Asymmetric Hierarchical Anchoring for Robust Audio-Visual Cross-Modal Generalization

SafetyDGX agent

arXiv:2602.03570v2 Announce Type: replace Abstract: Audio-visual joint representation learning under Cross-Modal Generalization (CMG) aims to transfer knowledge from a labeled source modality to an un

Attribution and Uncertainty Behavior of Learned Residual Gyro Correction for Gyro-Stellar Estimation

SafetyDGX agent

arXiv:2607.24608v1 Announce Type: new Abstract: This work investigates uncertainty decomposition and explainability in a deep learning-based framework for gyroscope bias correction. A 1-D Convolutiona

Auditing Alignment Controllability in LLMs via Political Axes

Model ReleasesDGX agent

arXiv:2607.23519v1 Announce Type: cross Abstract: Political audits of large language models (LLMs) usually reduce each to one point on a political compass. But that resting point barely matters in dep

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

SafetyDGX agent

arXiv:2607.23175v1 Announce Type: cross Abstract: Reducing toxicity is often framed as a global alignment problem, yet perceptions of harmful language are subjective and context-dependent. We present

Beyond Squared Error: Exploring Loss Design for Enhanced Training of Generative Flow Networks

SafetyDGX agent

arXiv:2410.02596v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) are a novel class of generative models designed to sample from unnormalized distributions and have found

Breaking the Synthetic-Real Domain Shortcut for Training-Free Generative Replay-based Class Incremental Learning

SafetyDGX agent

arXiv:2607.22994v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to continuously acquire new knowledge while avoiding catastrophic forgetting. While exemplar replay is

Bridging Reinforcement Learning and Optimal Control via Feasible Action Mapping

Model ReleasesDGX agent

arXiv:2607.23930v1 Announce Type: cross Abstract: Operating constrained dynamical systems requires controllers to efficiently solve complex tasks while enforcing recursive feasibility and safety const

Concurrent Prehensile and Nonprehensile Manipulation: A Practical Approach to Multi-Stage Dexterous Tasks

SafetyDGX agent

arXiv:2603.11655v3 Announce Type: replace Abstract: Dexterous hands enable concurrent prehensile and nonprehensile manipulation, such as holding one object while interacting with another, a capability

ConFusion: Continuous Fusion Space Learning for Fine-Grained Controllable Infrared and Visible Image Fusion

SafetyDGX agent

arXiv:2607.23600v1 Announce Type: new Abstract: Controllable infrared-visible image fusion aims to integrate complementary thermal and structural information with flexible region-aware modulation, pro

Continual-RL for Generalization in Autonomous Racing on the RoboRacer Platform

SafetyDGX agent

arXiv:2607.24320v1 Announce Type: new Abstract: A key challenge in modern robotics is to adapt to changing environments, a challenge that is exacerbated when simulations cannot encompass every possibl

CoopReflect: Towards Natural Language Communication for Cooperative Autonomous Driving via Multi-Agent Learning

SafetyDGX agent

arXiv:2505.18334v2 Announce Type: replace-cross Abstract: Past work has demonstrated that autonomous vehicles can drive more safely if they communicate with each other. However, this communication is

CRAFT: Learn the Schema, Execute the Plan

SafetyDGX agent

arXiv:2607.22642v1 Announce Type: new Abstract: Enterprise coding agents translate natural-language analytical requests into executable code over proprietary APIs, schemas, and metric definitions. Yet

D3O: Dynamic Distribution Distillation for Ordinal Regression

SafetyDGX agent

arXiv:2607.23575v1 Announce Type: cross Abstract: Ordinal regression is widely used in scenarios where labels are discrete yet inherently ordered. In practice, however, ordinal labels are often obtain

Data Pyramid for Embodied Manipulation

SafetyDGX agent

arXiv:2607.24744v1 Announce Type: cross Abstract: Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit no such shortcut, since they require d

Dementia Etiology Diagnosis via Collaborative Meta Knowledge Enhancement

SafetyDGX agent

arXiv:2607.22770v1 Announce Type: cross Abstract: Although artificial intelligence (AI) has shown promising performance in several medical tasks, accurate dementia etiology diagnosis with AI remains c

Dependency-Guided Code Generation: Structured Matrix Decomposition and Consistency-Guided Refinement

SafetyDGX agent

arXiv:2607.16692v2 Announce Type: replace-cross Abstract: The increasing complexity of modern software systems has made automated code generation a fundamental task in software engineering. However, e

Designing Service Systems from Textual Evidence

SafetyDGX agent

arXiv:2603.10400v2 Announce Type: replace-cross Abstract: Designing service systems requires selecting among alternative configurations -- choosing the best chatbot variant, the optimal routing policy

DeVA: Decoupled Video-Action Model with physical guidance for robot policy learning

SafetyDGX agent

arXiv:2607.24159v1 Announce Type: cross Abstract: Generalizable robot manipulation requires policies that can anticipate how visual scenes evolve while executing language instructions. While recent Vi

DICA: Dual-Indicator Guided Contrastive Alignment in Multimodal Large Language Models

SafetyDGX agent

arXiv:2607.23944v1 Announce Type: new Abstract: Human visual reasoning typically follows a coarse-to-fine attention process, starting from global scene understanding and gradually focusing on question

Disentangling Acoustic Cues in Alzheimer's Pathology and Perception: The Roles of Language and Gender

SafetyDGX agent

arXiv:2607.23977v1 Announce Type: cross Abstract: Acoustic biomarkers show promise for detecting Alzheimer's Disease (AD), yet whether the cues driving diagnostic AI align with those salient to human

Disentangling Semantic Attention from Structural Bias in the Attention Manifold

SafetyDGX agent

arXiv:2607.24017v1 Announce Type: cross Abstract: The empirical success of attention mechanism in Multimodal Large Language Models (MLLMs) often obscures its inherent, subtle flaws. Specifically, MLLM

Do LLM Debates Repeat Arguments Differently Across Languages?

SafetyDGX agent

arXiv:2607.23442v1 Announce Type: new Abstract: LLM debate is usually evaluated by final answers, but transcripts also reveal whether later turns develop new argumentative content or return to earlier

Does Faithfulness-Guided Alignment Hurt Accuracy? Unlocking Accurate and Faithful Post-Retrieval Reasoning

SafetyDGX agent

arXiv:2602.01348v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) can achieve strong answer accuracy on multi-hop questions, but outcome-level rewards often leave reasoning trac

EchoBridge: Long-Tail-Aware ECG-Echocardiography Text Alignment for Echocardiography-Derived Cardiac Findings

SafetyDGX agent

arXiv:2607.24553v1 Announce Type: cross Abstract: Standardized echocardiography conclusions provide meaningful supervision for learning ECG representations of echocardiography-derived cardiac findings

Efficiency Matters in Autonomous Research

SafetyDGX agent

arXiv:2607.24647v1 Announce Type: new Abstract: AI-driven autonomous research (AR) systems are becoming increasingly effective across a broad range of tasks. Their performance, however, is still evalu

Eviction as Estimation: A Fixed-Lag Smoothing View of Test-Time Memory, and When Measuring Beats Accumulating

SafetyDGX agent

arXiv:2607.24667v1 Announce Type: new Abstract: A language model with a bounded working memory must repeatedly decide which stored items to keep. Every deployed method decides the moment an item arriv

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

SafetyDGX agent

arXiv:2607.22569v1 Announce Type: new Abstract: Coding agents are increasingly integrated into system operations, where their tool use can directly modify project artifacts, execution environments, an

Face Age Verification Vulnerabilities Under Simple Appearance Manipulations

SafetyDGX agent

arXiv:2607.24194v1 Announce Type: new Abstract: Online platforms increasingly rely on automated age estimation systems to enforce minimum-age policies. Focusing on vision-based models designed for thi

Finite-Time Analysis of the Natural Policy Gradient in Finite-Horizon Markov Decision Processes

SafetyDGX agent

arXiv:2607.22982v1 Announce Type: new Abstract: Natural Policy Gradient (NPG) is a well-established Reinforcement Learning algorithm that underlies widely used methods such as Trust Region Policy Opti

FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models

SafetyDGX agent

arXiv:2607.24522v1 Announce Type: cross Abstract: While on-policy distillation (OPD) effectively addresses sparse rewards and exposure bias in large language model post-training, its extension to flow

← Previous
1…7273747576…242
Next →