AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling

DGX agent

arXiv:2608.08553v1 Announce Type: new Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from

safetyarxiv-cs-cv
11 Aug 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-Branch Policy Optimization for Multimodal Large Language Models

DGX agent

arXiv:2608.07581v1 Announce Type: cross Abstract: Group-based reinforcement learning methods for multimodal large language models typically rely on trajectory-level credit assignment that applies a si

safetyarxiv-cs-ai
11 Aug 2026
Safety

MultiShadow: Multi-Object Shadow Generation for Image Compositing via Diffusion Model

DGX agent

arXiv:2603.02743v4 Announce Type: replace Abstract: Realistic shadow generation is crucial for achieving seamless image compositing, yet existing methods primarily focus on single-object insertion and

safetyarxiv-cs-cv
11 Aug 2026
Safety

Neural Message Passing on Structural Interaction Graphs for Fully-Inductive Graph Neural Networks

DGX agent

arXiv:2608.08567v1 Announce Type: new Abstract: A central obstacle in building graph foundation models is the input heterogeneity in terms of feature space dimensionality, semantics, and structure. Su

safetyarxiv-cs-lg
11 Aug 2026
Safety

North Africa's Missing Framework: NLP-Driven Mental Healthcare in Algeria and Implications for Low-resource Settings

DGX agent

arXiv:2608.08607v1 Announce Type: new Abstract: Mental health disorders are a leading cause of disability worldwide, yet Natural Language Processing (NLP) research for mental healthcare has remained c

safetyarxiv-cs-cl
11 Aug 2026
Safety

OD-Gear: Online Decomposition and Group Sampling for Expert-Guided Adversarial Routing in Scalable Capacitated Vehicle Routing

DGX agent

arXiv:2602.00488v3 Announce Type: replace Abstract: Solving large-scale capacitated vehicle routing problems (CVRP) is hindered by the high complexity of classical heuristics and the limited generaliz

safetyarxiv-cs-lg
11 Aug 2026
Safety

On the use of foundation models in cognitive science

DGX agent

arXiv:2608.07812v1 Announce Type: new Abstract: A host of recent studies have evaluated the cognitive and developmental alignment of Foundation Models (FMs). These investigations include evaluations o

safetyarxiv-cs-cl
11 Aug 2026
Safety

OnEvoMemory: Evolving Memory through Online Robot Rollouts for Pretrained Robot Policies

DGX agent

arXiv:2608.08749v1 Announce Type: new Abstract: Long-horizon robot manipulation requires policies to track completed subtasks and critical interaction events. However, existing memory mechanisms heavi

safetyarxiv-cs-ro
11 Aug 2026
Safety

PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distillation

DGX agent

arXiv:2608.08726v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) uses a privileged teacher to supervise a reasoning model on prefixes sampled from its own rollouts. Yet each rollou

safetyarxiv-cs-ai
11 Aug 2026
Safety

Penalizing Length: Uncovering Systematic Bias in Quality Estimation Metrics

DGX agent

arXiv:2510.22028v4 Announce Type: replace Abstract: Quality Estimation (QE) metrics are vital in machine translation for reference-free evaluation and increasingly serve as selection criteria in data

safetyarxiv-cs-cl
11 Aug 2026
Safety

Perception Before Supervision: Self-Contained Visual Distillation from Counterfactual Blind Spots

DGX agent

arXiv:2608.09931v1 Announce Type: new Abstract: Self-improvement for multimodal large language models (MLLMs) is typically driven by reward-based methods that provide only coarse scalar feedback. Dist

safetyarxiv-cs-cv
11 Aug 2026
Safety

PhysAttNet: Enhancing Predictive Performance in Industrial and Astrophysical Time Series via Physics-Informed Attention

DGX agent

arXiv:2608.07681v1 Announce Type: new Abstract: Accurate and robust time series forecasting is essential in many applications involving physical processes, such as manufacturing monitoring and astroph

safetyarxiv-cs-lg
11 Aug 2026
Safety

Physics-Informed Policy Iteration for High-Dimensional Hamilton--Jacobi--Bellman Equations: Interior Error Bounds without Boundary Data

DGX agent

arXiv:2508.01718v2 Announce Type: replace Abstract: We develop a physics-informed policy-iteration method for stationary second-order Hamilton--Jacobi--Bellman equations arising in continuous-time sto

safetyarxiv-cs-lg
11 Aug 2026
Safety

PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering

DGX agent

arXiv:2608.07509v1 Announce Type: cross Abstract: LLMs are increasingly used for conversational tutoring, but effective tutoring requires more than correct answers. Tutors must choose when to scaffold

safetyarxiv-cs-ai
11 Aug 2026
Safety

Position Bias in Ordinal Classification: A Systematic Evaluation

DGX agent

arXiv:2608.08869v1 Announce Type: new Abstract: Large language models are increasingly used for ordinal classification, yet semantically equivalent changes to prompt organization can alter their predi

safetyarxiv-cs-cl
11 Aug 2026
Safety

Privacy-Preserving Data Drift Detection and Recovery for Large-Scale LLM Applications via Proxy Representations

DGX agent

arXiv:2608.08245v1 Announce Type: cross Abstract: LLM applications deployed at scale face a fundamental challenge: privacy constraints prevent direct inspection of user interactions, making it difficu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Private Anytime Selective-Risk Certification for Federated Retrieval-Augmented Generation: Guarantees and Empirical Limits

DGX agent

arXiv:2608.07913v1 Announce Type: cross Abstract: Selective-risk certificates promise that accepted outputs meet a declared error target. We develop Fed-SRC, a score-agnostic certificate for federated

safetyarxiv-cs-ai
11 Aug 2026
Safety

Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation

DGX agent

arXiv:2608.09263v1 Announce Type: new Abstract: Outcome verifiers score completed reasoning traces but do not assign credit to intermediate tokens. Privileged self-distillation attempts to fill this g

safetyarxiv-cs-ai
11 Aug 2026
Safety

Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation

DGX agent

arXiv:2608.09228v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) is commonly interpreted as the transfer of privileged information: a teacher observes the verified solution to the

safetyarxiv-cs-ai
11 Aug 2026
Safety

Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update

DGX agent

arXiv:2607.11505v2 Announce Type: replace-cross Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behav

safetyarxiv-cs-ai
11 Aug 2026
Safety

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

DGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

safetyarxiv-cs-ai
11 Aug 2026
Safety

Real-Time Nonlinear MPC via Sequential Quadratic Programming with Structure-Exploiting ADMM and Interior-Point Methods for Underactuated Double-Pendulum Swing-Up

DGX agent

arXiv:2608.09272v1 Announce Type: cross Abstract: The 4th 'AI Olympics with RealAIGym' competition, to be held at IJCAI-ECAI 2026 in Bremen, challenges participants to develop a global control policy

safetyarxiv-cs-ro
11 Aug 2026
Safety

Reconfigurable Structural Robotic Assembly: Interlocking 3D Aggregations with Self-Aligning Compound Nested Lattice Modules

DGX agent

arXiv:2608.07576v1 Announce Type: new Abstract: Robotic construction systems often treat the material system and the robot as separate design problems, locating intelligence primarily in hardware, sen

safetyarxiv-cs-ro
11 Aug 2026
Safety

Regret of exploratory policy improvement and q-learning

DGX agent

arXiv:2411.01302v2 Announce Type: replace Abstract: We study the convergence of q-learning and related algorithms introduced by Jia and Zhou (J. Mach. Learn. Res., 24 (2023), 161) for controlled diffu

safetyarxiv-cs-lg
11 Aug 2026
Safety

REIN: Bridging the Gap between Reasoning and Reliability via Reflection and Abstention Alignment

DGX agent

arXiv:2608.07931v1 Announce Type: new Abstract: Large reasoning models (LRMs) are prone to hallucination, which undermines their reliability and poses challenges for safe deployment. Hallucinations in

safetyarxiv-cs-ai
11 Aug 2026
Safety

Retrieval-Augmented Generation-Based Color Restoration for Low-Light Image Enhancement

DGX agent

arXiv:2608.08211v1 Announce Type: cross Abstract: Recent low-light image enhancement (LLIE) methods have driven brightness and structural fidelity close to that of normally-exposed images, yet their o

safetyarxiv-cs-cv
11 Aug 2026
Safety

RL-Native Distillation: Exploiting Scored Trajectories for Few-Step Image Generation

DGX agent

arXiv:2608.09226v1 Announce Type: cross Abstract: Efficient text-to-image generation requires both reinforcement-learning (RL)-based reward alignment and few-step distillation, yet these procedures ar

safetyarxiv-cs-ai
11 Aug 2026
Safety

RobustDefect-LLM: Explainable and Robustness-Aware Industrial Surface Defect Classification with Decision Support and AI-Assisted Reporting

DGX agent

arXiv:2608.08589v1 Announce Type: new Abstract: This paper presents RobustDefect-LLM, an industrial surface-defect inspection framework integrating deep-learning classification, operator-facing visual

safetyarxiv-cs-cv
11 Aug 2026
Safety

RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

DGX agent

arXiv:2608.09853v1 Announce Type: cross Abstract: General-purpose reward models are increasingly the bottleneck for scaling robot learning, yet the recipe for learning value-related capabilities from

safetyarxiv-cs-cv
11 Aug 2026
Safety

Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance

DGX agent

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO)

safetyarxiv-cs-lg
11 Aug 2026
Safety

SC-Diff: Semantically Calibrated Diffusion for Visible-to-Infrared Image Translation

DGX agent

arXiv:2608.08555v1 Announce Type: new Abstract: Visible-to-infrared image translation provides a practical way to expand infrared training data using abundant visible images. Diffusion models are prom

safetyarxiv-cs-cv
11 Aug 2026
Safety

Scalable extensions to given-data Sobol' index estimators

DGX agent

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computa

safetyarxiv-cs-lg
11 Aug 2026
Safety

SCOUT: Self-Checking and Recovery-Aware Tool-Thought Agents for Ultra-Long Egocentric Video Reasoning

DGX agent

arXiv:2608.07959v1 Announce Type: new Abstract: Ultra-long egocentric video understanding requires reasoning over temporally sparse evidence distributed across hours or days, challenging current multi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

DGX agent

arXiv:2608.07531v1 Announce Type: cross Abstract: Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing ext

safetyarxiv-cs-ai
11 Aug 2026
Safety

Self Supervised Learning from Automatically Generated Demonstrations for Visual Robotic Manipulation

DGX agent

arXiv:2608.07553v1 Announce Type: new Abstract: Robotic manipulation often requires object specific programming, manual data annotation, or calibrated perception pipelines, which limits rapid deployme

safetyarxiv-cs-ro
11 Aug 2026
Safety

SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer

DGX agent

arXiv:2602.17632v3 Announce Type: replace-cross Abstract: Modern offline Reinforcement Learning (RL) methods find performant actor-critics, however, fine-tuning these actor-critics online with value-b

safetyarxiv-cs-ai
11 Aug 2026
Safety

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

DGX agent

arXiv:2608.09138v1 Announce Type: cross Abstract: While learned robotic policies hold promise for advancing generalizable manipulation, their practical deployment is often hindered by suboptimal execu

safetyarxiv-cs-ai
11 Aug 2026
Safety

SR-OPSD: Self-Referenced On-Policy Self-Distillation

DGX agent

arXiv:2608.09745v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, provi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Safety

StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoning

DGX agent

arXiv:2608.08326v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective approach for improving multimodal reasoning. However, most existing me

safetyarxiv-cs-ai
11 Aug 2026
Safety

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

DGX agent

arXiv:2608.00143v2 Announce Type: replace-cross Abstract: Automated attack chain generation is critical for modern cybersecurity, yet manual construction fails to scale as adversary behaviors expand.

safetyarxiv-cs-ai
11 Aug 2026
Safety

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

DGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

safetyarxiv-cs-cv
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Sample Complexity of Policy Learning with Mu-Resets

DGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

safetyarxiv-cs-lg
11 Aug 2026
Safety

The Scaling Paradox in Human-AI Collaboration

DGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

DGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

DGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

safetyarxiv-cs-cl
11 Aug 2026
Safety

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

DGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

safetyarxiv-cs-cl
11 Aug 2026
← Previous
1…6465666768…260
Next →