AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Context Assembly as the Controlled Variable: A Control-Theoretic View of Harness Policies for Frozen LLM Agents

DGX agent

arXiv:2607.25408v1 Announce Type: new Abstract: A growing body of 2026 work applies control theory to LLM agents: Lyapunov-certified stability for tool-mediated controllers (Prinos et al., 'Stable Age

safetyarxiv-cs-ai
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Contextual Deconvolution for Variance-Stable Demand Sensing: Kernel-Modulated Operators in Promotional Retail

DGX agent

arXiv:2607.25664v1 Announce Type: new Abstract: Machine learning demand forecasts optimize statistical accuracy yet leave excess operational volatility that inflates safety stock and amplifies the Bul

safetyarxiv-cs-lg
29 Jul 2026
Safety

ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS Design

DGX agent

arXiv:2607.25283v1 Announce Type: new Abstract: This paper presents ContractHIL-HLS, a contract-aligned multi-agent workflow for practical high-level synthesis (HLS) engineering. The workflow makes th

safetyarxiv-cs-ai
29 Jul 2026
Safety

Contrastive Weak-to-strong Generalization

DGX agent

arXiv:2510.07884v3 Announce Type: replace-cross Abstract: Weak-to-strong generalization provides a promising paradigm for scaling large language models (LLMs) by training stronger models on samples fr

safetyarxiv-cs-ai
29 Jul 2026
Safety

CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization

DGX agent

arXiv:2607.25659v1 Announce Type: new Abstract: Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit criteria. Yet in GRPO-style pipelines,

safetyarxiv-cs-ai
29 Jul 2026
Safety

Data-Dependent Regret and Polyak Corrections for Constrained Online Convex Optimization

DGX agent

arXiv:2607.25480v1 Announce Type: new Abstract: Constrained online convex optimization requires minimizing regret against adversarial convex costs while satisfying a convex constraint at every round,

safetyarxiv-cs-lg
29 Jul 2026
Safety

Dataset Poisoning Attacks on Behavioral Cloning Policies

DGX agent

arXiv:2511.20992v2 Announce Type: replace Abstract: Behavior Cloning (BC) is a popular framework for training sequential decision policies from expert demonstrations via supervised learning. As these

safetyarxiv-cs-lg
29 Jul 2026
Safety

DDSNet: Dual-domain Symmetry-aware Network for PCSEL Property Prediction

DGX agent

arXiv:2607.24785v1 Announce Type: cross Abstract: Efficient exploration of the photonic crystal (PhC) lattice design space is essential for developing photonic crystal surface-emitting lasers. While c

safetyarxiv-cs-ai
29 Jul 2026
Safety

Decentralized Scalable Exploration via Emergent Adaptive Levy Walks on Minimal-Sensing Platforms

DGX agent

arXiv:2607.25195v1 Announce Type: new Abstract: Efficient autonomous exploration with palm-sized nano-UAVs remains challenging due to severe limitations in sensing, computation, and flight endurance.

safetyarxiv-cs-ro
29 Jul 2026
Safety

DensFiLM: Density-Conditioned Video Saliency for Crowd Scenes

DGX agent

arXiv:2607.25465v1 Announce Type: new Abstract: Video saliency models typically apply a single fixation strategy across crowd scenes, despite systematic changes in attention with crowd density. Sparse

safetyarxiv-cs-cv
29 Jul 2026
Safety

Do Models Fake Alignment Without Clear Consequences?

DGX agent

arXiv:2607.24758v1 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical d

safetyarxiv-cs-ai
29 Jul 2026
Safety

Does Runtime Topology Context Improve LLM-Generated Kubernetes Security Patches?

DGX agent

arXiv:2607.25995v1 Announce Type: cross Abstract: Kubernetes is central to the cloud-native ecosystem, orchestrating containerised workloads. Recent work suggests that large language models (LLMs) can

safetyarxiv-cs-ai
29 Jul 2026
Safety

DSCD-Nav: Dual-Stance Cooperative Debate for Object Navigation

DGX agent

arXiv:2601.21409v3 Announce Type: replace Abstract: Adaptive navigation in unfamiliar indoor environments is crucial for household service robots. Despite advances in zero-shot perception and reasonin

safetyarxiv-cs-ro
29 Jul 2026
Safety

Early Detection of Distributed Backdoors in Multi-Agent LLM Systems: A Characterization Study

DGX agent

arXiv:2607.24893v1 Announce Type: cross Abstract: Multi-agent LLM systems can be attacked by a payload that no single agent ever holds in full: a poisoned tool hides encrypted fragments in its observa

safetyarxiv-cs-ai
29 Jul 2026
Safety

Egocentric Station Holding of Robotic Fish in Unknown Turbulent Background Flow

DGX agent

arXiv:2607.24860v1 Announce Type: new Abstract: Approaching a target position and holding station in flowing water is a fundamental and critical capability for robotic fish operating in natural aquati

safetyarxiv-cs-ro
29 Jul 2026
Safety

Evaluation of Adversarial Robustness in Arabic Language Models

DGX agent

arXiv:2607.25814v1 Announce Type: new Abstract: The emergence of the recent outstanding capabilities of Arabic Language Models has opened doors for exposing their vulnerabilities. One of the major sec

safetyarxiv-cs-cl
29 Jul 2026
Safety

Evaluation of forced alignment of code-mixed speech: the case of Hindi-English

DGX agent

arXiv:2607.25581v1 Announce Type: new Abstract: Code-mixed speech poses unique challenges to forced alignment: expanded inventories, orthographic errors, and speaker variation. We evaluate forced alig

safetyarxiv-cs-cl
29 Jul 2026
Safety

Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales

DGX agent

arXiv:2607.25364v1 Announce Type: new Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales are neither authorization nor reliable introspection

safetyarxiv-cs-ai
29 Jul 2026
Safety

Face De-Identification: A Domain-Centric Survey from Capture to Processing

DGX agent

arXiv:2607.25926v1 Announce Type: cross Abstract: Face de-identification (De-ID) aims to remove or conceal personally identifiable facial features in images or videos to prevent identity recognition w

safetyarxiv-cs-ai
29 Jul 2026
Safety

Faces of Fairness: Examining Bias in Facial Expression Recognition Datasets and Models

DGX agent

arXiv:2502.11049v3 Announce Type: replace Abstract: Automated Facial Expression Recognition (FER), involves two critical aspects: data and model design. Both significantly influence bias and fairness

safetyarxiv-cs-cv
29 Jul 2026
Safety

Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume Screening

DGX agent

arXiv:2507.11548v3 Announce Type: replace-cross Abstract: The use of publicly available generative AI systems for resume evaluation is often justified by the assumption that these tools reduce bias re

safetyarxiv-cs-ai
29 Jul 2026
Safety

Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment

DGX agent

arXiv:2607.26034v1 Announce Type: new Abstract: Technological races create tension between speed and safety: actors may gain by moving faster than competitors, even when risky development is harmful.

safetyarxiv-cs-ai
29 Jul 2026
Safety

Fine-Grained Food Image Understanding via Target-Aware Data Alignment

DGX agent

arXiv:2607.25794v1 Announce Type: new Abstract: Fine-grained food visual--semantic understanding requires models to capture subtle distinctions across ingredients, cooking methods, doneness, color, te

safetyarxiv-cs-cv
29 Jul 2026
Safety

From Dyad to Triad: Eliciting XAI Requirements in Stroke Rehabilitation

DGX agent

arXiv:2607.25423v1 Announce Type: cross Abstract: Eliciting explainable AI (XAI) requirements from stroke survivors presents a methodological challenge with direct implications for the design of trust

safetyarxiv-cs-ai
29 Jul 2026
Safety

@GaryMarcus OpenAI and Anthropic desperately need to raise prices. We seem to be faced with one of three choices for them: 1. Bailout 2. Ban…

DGX agent

Gary Marcus has argued that both OpenAI and Anthropic must increase their pricing, citing the rapid decline in token costs and competition from open‑source models. He presents a binary choice: either

safetygary-marcus--x
29 Jul 2026
Safety

Group Equivariant Diffusion for Anomaly Detection in Computational Cytology

DGX agent

arXiv:2607.25503v1 Announce Type: new Abstract: Computational cytology on whole-slide images is challenging because malignant cells are rare, heterogeneous, and annotated slides are scarce. Anomaly de

safetyarxiv-cs-cv
29 Jul 2026
Safety

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM

DGX agent

arXiv:2508.05775v3 Announce Type: replace Abstract: Large Language Models (LLMs) have revolutionized content creation across digital platforms, offering unprecedented capabilities in natural language

safetyarxiv-cs-cl
29 Jul 2026
Safety

Handy dandy AI crisis flowchart from @klonick

DGX agent

Gary Marcus shared a “handy dandy AI crisis flowchart” created by Kate Klonick (@Klonick) on Saturday, July 28. The tweet also notes that he had planned to write an article about Hugging Face and Open

safetygary-marcus--x
29 Jul 2026
Safety

Hybrid Analysis for Secure MCP Tool Use in LLM Agents

DGX agent

arXiv:2607.25297v1 Announce Type: cross Abstract: The rapid development of large language model (LLM) agents has enabled their broad adoption across diverse real-world tasks. To standardize interactio

safetyarxiv-cs-ai
29 Jul 2026
Safety

Improving Rare Medication Recommendation with Counterfactual Data Augmentation and Large Language Models

DGX agent

arXiv:2607.24829v1 Announce Type: cross Abstract: AI-based medication recommendation systems have attracted substantial attention due to their potential to enhance patient safety and therapeutic outco

safetyarxiv-cs-lg
29 Jul 2026
Safety

Interpretable GOHR Agents via Sparse Autoencoders

DGX agent

arXiv:2607.25132v1 Announce Type: new Abstract: A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help ex

safetyarxiv-cs-lg
29 Jul 2026
Safety

Inverse RL Helps Align AI by Imitating Humans

DGX agent

arXiv:2607.24900v1 Announce Type: new Abstract: Language model alignment aims to make model behavior reliably reflect desirable properties such as helpfulness, safety, and instruction following. Curre

safetyarxiv-cs-lg
29 Jul 2026
Safety

IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment

DGX agent

arXiv:2607.25579v1 Announce Type: cross Abstract: Entity alignment (EA) identifies entities across knowledge graphs (KGs) that refer to the same real-world object. Conventional EA methods mainly explo

safetyarxiv-cs-ai
29 Jul 2026
Safety

Keypoint-Guided Optimal Transport: Models, Algorithms, and Applications

DGX agent

arXiv:2303.13102v2 Announce Type: replace Abstract: Existing Optimal Transport (OT) methods mainly derive the optimal transport plan/matching under the criterion of transport cost/distance minimizatio

safetyarxiv-cs-cv
29 Jul 2026
Safety

Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation

DGX agent

arXiv:2607.25956v1 Announce Type: new Abstract: Multi-warehouse inventory allocation is typically formulated as a mixed-integer programming (MIP) problem, yet no single formulation consistently matche

safetyarxiv-cs-ai
29 Jul 2026
Safety

Learning from Local Walks on Dynamic Graphs with Bandit Feedback

DGX agent

arXiv:2607.10571v2 Announce Type: replace-cross Abstract: We study stochastic multi-armed bandits on dynamic graphs, where arms correspond to the vertices of a network with time-varying edges. In this

safetyarxiv-cs-ai
29 Jul 2026
Safety

Learning from the Unseen: Offline Reinforcement Learning with Hidden Actions

DGX agent

arXiv:2607.25241v1 Announce Type: cross Abstract: Standard offline reinforcement learning (RL) algorithms typically assume that the actions in the dataset are observed without error. However, in many

safetyarxiv-cs-lg
29 Jul 2026
Safety

LGFNet: A CTC-Guided Local-Global Fusion Framework for Single-Channel Sleep Staging

DGX agent

arXiv:2607.25197v1 Announce Type: new Abstract: Sleep staging remains challenging due to long-range temporal dependencies, ambiguous stage transitions-particularly in N1-and substantial distribution s

safetyarxiv-cs-cv
29 Jul 2026
Safety

LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Models

DGX agent

arXiv:2607.24892v1 Announce Type: cross Abstract: Text-conditioned time-series forecasting predicts a series from both its numerical history and natural-language context, allowing forecasts to account

safetyarxiv-cs-ai
29 Jul 2026
Safety

LLM Scheming Inversely Scales with Pretraining Language Coverage

DGX agent

arXiv:2607.24769v1 Announce Type: new Abstract: With the growing capabilities of frontier models, AI alignment becomes increasingly critical in high-risk deployment settings. While recent work has emp

safetyarxiv-cs-ai
29 Jul 2026
Safety

Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation

DGX agent

arXiv:2602.10508v2 Announce Type: replace Abstract: Modern segmentation models achieve strong predictive performance but remain largely opaque, limiting our ability to diagnose failures, understand da

safetyarxiv-cs-cv
29 Jul 2026
Safety

Medical world models in healthcare: foundations, applications, and challenges for trustworthy clinical translation

DGX agent

arXiv:2607.25242v1 Announce Type: new Abstract: Medical world models offer a framework for extending medical artificial intelligence beyond static prediction by representing evolving patient states an

safetyarxiv-cs-cv
29 Jul 2026
Safety

Motion-Acceleration Calibration and Compensation in IMUs without External Equipment for Attitude Estimation Filters

DGX agent

arXiv:2607.25784v1 Announce Type: new Abstract: Attitude estimation based on inertial sensing requires measurements of local angular velocities and local gravity via gyroscopes and accelerometers. How

safetyarxiv-cs-ro
29 Jul 2026
Safety

Motion Generation With Environmental Constraints

DGX agent

arXiv:2607.25053v1 Announce Type: new Abstract: Robot motion planning faces challenges in high-dimensional spaces and uncertain environments, often constrained by the need for collision-free motions.

safetyarxiv-cs-ro
29 Jul 2026
Safety

Multi-Sensor Alignment for Weather Simulations

DGX agent

arXiv:2607.25612v1 Announce Type: new Abstract: Perception tasks for autonomous vehicles need to work satisfactorily in adverse weather conditions. Due to lack of real-world weather datasets, weather

safetyarxiv-cs-ai
29 Jul 2026
Safety

NEXT: Reasoning-Driven Video Recommendation via a Vision-Language Model

DGX agent

arXiv:2607.24789v1 Announce Type: cross Abstract: We present NEXT (Next-interest EXploration Transformer), a reasoning-driven video recommendation framework that reasons over the video a user has just

safetyarxiv-cs-cv
29 Jul 2026
Safety

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning

DGX agent

arXiv:2607.25369v1 Announce Type: new Abstract: Agentic systems have rapidly advanced in their ability to interact with real-world environments, leverage external tools, and provide services for users

safetyarxiv-cs-ai
29 Jul 2026
Safety

On the Use of Synthetic Data for Threshold Calibration in Face Recognition: Performance and Security Implications for Border Control Systems

DGX agent

arXiv:2607.25990v1 Announce Type: new Abstract: The recently deployed Entry/Exit System (EES) introduces large-scale biometric verification into European border control, requiring face recognition sys

safetyarxiv-cs-cv
29 Jul 2026
← Previous
1…2930313233…265
Next →