AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
Safety

Learning an Interior Layout Policy in a Domain Specific Language Action Space

DGX agent

arXiv:2608.07547v1 Announce Type: cross Abstract: Indoor scene layout generation is a challenging task in interior design. Existing methods often oversimplify the task by reducing room conditions to c

safetyarxiv-cs-ai
11 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Learning Deep Modality-Shared Self-Expressiveness for Image Clustering with Textual Information

DGX agent

arXiv:2608.08418v1 Announce Type: new Abstract: Leveraging textual information for image clustering has emerged as a promising direction, largely owing to the powerful representations learned by Visio

safetyarxiv-cs-cv
11 Aug 2026
Safety

Learning from Consensus and Disagreement: Unsupervised On-Policy Self-Distillation with Minority-Trajectory Contrast

DGX agent

arXiv:2608.08764v1 Announce Type: cross Abstract: On-policy self-distillation improves language-model reasoning by querying a teacher on states actually visited by the student. Recent methods create a

safetyarxiv-cs-ai
11 Aug 2026
Safety

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

DGX agent

arXiv:2608.09507v1 Announce Type: cross Abstract: Natural language user preferences provide an interpretable interface for LLM personalization. However, universal preference summaries often contain in

safetyarxiv-cs-ai
11 Aug 2026
Safety

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates

DGX agent

arXiv:2608.00326v2 Announce Type: replace Abstract: Tool calling allows large language models (LLMs) to invoke external computation during problem solving, a useful capability in various fields includ

safetyarxiv-cs-ai
11 Aug 2026
Safety

Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump Control

DGX agent

arXiv:2608.09453v1 Announce Type: cross Abstract: On--off cycling is the main cause of compressor wear in residential heat pumps, yet reinforcement learning (RL) controllers for buildings typically op

safetyarxiv-cs-ai
11 Aug 2026
Safety

Learning When to See and When to Feel: Adaptive Vision-Torque Fusion for Contact-Aware Manipulation

DGX agent

arXiv:2604.01414v2 Announce Type: replace Abstract: Vision-based policies have achieved a good performance in robotic manipulation due to the accessibility and richness of visual observations. However

safetyarxiv-cs-ro
11 Aug 2026
Safety

Legal Responsibilities Using Autonomous Agents For Artificial Intelligence

DGX agent

arXiv:2608.08022v1 Announce Type: new Abstract: Recent incidents involving Artificial Intelligence (AI) agents, which were reported escaping their containment `unintentionally' to gain unauthorized ac

safetyarxiv-cs-ai
11 Aug 2026
Safety

LF{}^{2}AR: Accounting for Layerwise Dynamics to Improve Multimodal Adaptation of Language Models

DGX agent

arXiv:2503.06211v3 Announce Type: replace-cross Abstract: Text-pretrained language models (LMs) encode rich world knowledge, but adapting them to process and generate perceptual modalities such as aud

safetyarxiv-cs-ai
11 Aug 2026
Safety

Linguistically-Aligned and Visually-Grounded Preference Optimization for Clinically-Augmented Medical Report Generation

DGX agent

arXiv:2608.08494v1 Announce Type: new Abstract: Despite significant advances in Medical Report Generation (MRG), the reliability remains constrained by the prevalence of factual errors. While Direct P

safetyarxiv-cs-cv
11 Aug 2026
Safety

LLMs Remember First, Forget Last: Dual-Process Interference in Large Language Models

DGX agent

arXiv:2603.00270v3 Announce Type: replace-cross Abstract: Large language models can process millions of tokens, yet how they handle conflicting information within context remains poorly understood. Fr

safetyarxiv-cs-ai
11 Aug 2026
Safety

LUCID: Latent-Skill Unified Control via Imagined Dynamics for Long-Horizon Humanoid Loco-Manipulation

DGX agent

arXiv:2608.07746v1 Announce Type: new Abstract: Long-horizon humanoid loco-manipulation requires composing versatile whole-body skills and reliable high-level decision making. Existing methods often c

safetyarxiv-cs-lg
11 Aug 2026
Safety

MAGIC-SSCIL: Manifold Anchoring and Geometric Incremental Calibration for Semi-Supervised Class Incremental Learning

DGX agent

arXiv:2608.07586v1 Announce Type: new Abstract: Semi-supervised Class Incremental Learning (SSCIL) is a severe challenge for neural networks, and it is hardest in the exemplar-free setting where no pa

safetyarxiv-cs-cv
11 Aug 2026
Safety

MARA: Flow-Matching-Guided Multi-Agent Resource Allocation for Computational Resource Efficient Learning

DGX agent

arXiv:2608.09130v1 Announce Type: cross Abstract: Allocating limited computation among concurrent learning tasks is difficult when each task must reach a target loss before a deadline but its required

safetyarxiv-cs-ai
11 Aug 2026
Safety

Matching Supervision to the Student's Learning Capacity: A Unified Framework for On-Policy Self-Distillation

DGX agent

arXiv:2608.08176v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) improves the reasoning abilities of LLMs by internalizing privileged context into model parameters through self-disti

safetyarxiv-cs-ai
11 Aug 2026
Safety

Metanormative Theory for RL-Based Moral Agents

DGX agent

arXiv:2608.08220v1 Announce Type: new Abstract: The overlapping disciplines of machine ethics and value alignment are concerned with designing artificial agents that are aligned with human values and

safetyarxiv-cs-ai
11 Aug 2026
Safety

MGMCL: Multi-Granularity Manifold Contrastive Learning With Neural ODEs for Cross-Subject EEG Emotion Recognition

DGX agent

arXiv:2608.08440v1 Announce Type: new Abstract: Cross-subject electroencephalogram (EEG)-based emotion recognition remains challenging due to substantial inter-individual variability and discrete form

safetyarxiv-cs-lg
11 Aug 2026
Safety

Mismatch Matters: On-Policy Distillation Beyond Token Agreement

DGX agent

arXiv:2608.09836v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as a core component of modern LLM post-training pipelines, yet we reveal a failure mode: degenerate agreement,

safetyarxiv-cs-ai
11 Aug 2026
Safety

Mitigating Gender Bias in English to Romanian Machine Translation

DGX agent

arXiv:2608.08606v1 Announce Type: cross Abstract: Machine translation (MT) systems often fail to correctly translate gender, especially when converting from a gender-neutral language like English to a

safetyarxiv-cs-ai
11 Aug 2026
Safety

ML-Based Hierarchical Prediction for Practical Energy Scheduling in Dynamic NTN-WPT Systems

DGX agent

arXiv:2608.08804v1 Announce Type: cross Abstract: With advancements in long-distance wireless power transfer (WPT) and space-based energy technologies, integrating WPT into non-terrestrial networks (N

safetyarxiv-cs-lg
11 Aug 2026
Safety

Model the Edit, Not the Image: Visual Autoregressive Editing from a Source-Centric Perspective

DGX agent

arXiv:2608.09057v1 Announce Type: new Abstract: Next-scale visual autoregressive models (VARs) have emerged as a powerful generative paradigm, producing high-quality images through efficient coarse-to

safetyarxiv-cs-cv
11 Aug 2026
Safety

Motif 3: Technical Report

DGX agent

arXiv:2608.09119v1 Announce Type: new Abstract: We introduce Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. Each spar

safetyarxiv-cs-ai
11 Aug 2026
Safety

MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling

DGX agent

arXiv:2608.08553v1 Announce Type: new Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from

safetyarxiv-cs-cv
11 Aug 2026
Safety

Multi-Branch Policy Optimization for Multimodal Large Language Models

DGX agent

arXiv:2608.07581v1 Announce Type: cross Abstract: Group-based reinforcement learning methods for multimodal large language models typically rely on trajectory-level credit assignment that applies a si

safetyarxiv-cs-ai
11 Aug 2026
Safety

MultiShadow: Multi-Object Shadow Generation for Image Compositing via Diffusion Model

DGX agent

arXiv:2603.02743v4 Announce Type: replace Abstract: Realistic shadow generation is crucial for achieving seamless image compositing, yet existing methods primarily focus on single-object insertion and

safetyarxiv-cs-cv
11 Aug 2026
Safety

Neural Message Passing on Structural Interaction Graphs for Fully-Inductive Graph Neural Networks

DGX agent

arXiv:2608.08567v1 Announce Type: new Abstract: A central obstacle in building graph foundation models is the input heterogeneity in terms of feature space dimensionality, semantics, and structure. Su

safetyarxiv-cs-lg
11 Aug 2026
Safety

North Africa's Missing Framework: NLP-Driven Mental Healthcare in Algeria and Implications for Low-resource Settings

DGX agent

arXiv:2608.08607v1 Announce Type: new Abstract: Mental health disorders are a leading cause of disability worldwide, yet Natural Language Processing (NLP) research for mental healthcare has remained c

safetyarxiv-cs-cl
11 Aug 2026
Safety

OD-Gear: Online Decomposition and Group Sampling for Expert-Guided Adversarial Routing in Scalable Capacitated Vehicle Routing

DGX agent

arXiv:2602.00488v3 Announce Type: replace Abstract: Solving large-scale capacitated vehicle routing problems (CVRP) is hindered by the high complexity of classical heuristics and the limited generaliz

safetyarxiv-cs-lg
11 Aug 2026
Safety

On the use of foundation models in cognitive science

DGX agent

arXiv:2608.07812v1 Announce Type: new Abstract: A host of recent studies have evaluated the cognitive and developmental alignment of Foundation Models (FMs). These investigations include evaluations o

safetyarxiv-cs-cl
11 Aug 2026
Safety

OnEvoMemory: Evolving Memory through Online Robot Rollouts for Pretrained Robot Policies

DGX agent

arXiv:2608.08749v1 Announce Type: new Abstract: Long-horizon robot manipulation requires policies to track completed subtasks and critical interaction events. However, existing memory mechanisms heavi

safetyarxiv-cs-ro
11 Aug 2026
Safety

OWN YOUR INTELLIGENCE Last year, building on open-weight models was primarily a cost rationalization exercise. Slightly worse performance fo…

DGX agent

OWN YOUR INTELLIGENCE Last year, building on open-weight models was primarily a cost rationalization exercise. Slightly worse performance for a much cheaper price. Now, it is increasingly an existenti

safetysonya-huang--x
11 Aug 2026
Safety

PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distillation

DGX agent

arXiv:2608.08726v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) uses a privileged teacher to supervise a reasoning model on prefixes sampled from its own rollouts. Yet each rollou

safetyarxiv-cs-ai
11 Aug 2026
Safety

Penalizing Length: Uncovering Systematic Bias in Quality Estimation Metrics

DGX agent

arXiv:2510.22028v4 Announce Type: replace Abstract: Quality Estimation (QE) metrics are vital in machine translation for reference-free evaluation and increasingly serve as selection criteria in data

safetyarxiv-cs-cl
11 Aug 2026
Safety

Perception Before Supervision: Self-Contained Visual Distillation from Counterfactual Blind Spots

DGX agent

arXiv:2608.09931v1 Announce Type: new Abstract: Self-improvement for multimodal large language models (MLLMs) is typically driven by reward-based methods that provide only coarse scalar feedback. Dist

safetyarxiv-cs-cv
11 Aug 2026
Safety

PhysAttNet: Enhancing Predictive Performance in Industrial and Astrophysical Time Series via Physics-Informed Attention

DGX agent

arXiv:2608.07681v1 Announce Type: new Abstract: Accurate and robust time series forecasting is essential in many applications involving physical processes, such as manufacturing monitoring and astroph

safetyarxiv-cs-lg
11 Aug 2026
Safety

Physics-Informed Policy Iteration for High-Dimensional Hamilton--Jacobi--Bellman Equations: Interior Error Bounds without Boundary Data

DGX agent

arXiv:2508.01718v2 Announce Type: replace Abstract: We develop a physics-informed policy-iteration method for stationary second-order Hamilton--Jacobi--Bellman equations arising in continuous-time sto

safetyarxiv-cs-lg
11 Aug 2026
Safety

PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering

DGX agent

arXiv:2608.07509v1 Announce Type: cross Abstract: LLMs are increasingly used for conversational tutoring, but effective tutoring requires more than correct answers. Tutors must choose when to scaffold

safetyarxiv-cs-ai
11 Aug 2026
Safety

Planning/RL for a stochastic single-player merge puzzle: afterstates, previewed chance events, and long-horizon throughput [D]

DGX agent

I am working on an AI for a small single-player merge puzzle and would appreciate pointers to related algorithms, papers, or existing implementations. It resembles 2048 in its action -> afterstate ->

safetyr-machinelearning
11 Aug 2026
Safety

Position Bias in Ordinal Classification: A Systematic Evaluation

DGX agent

arXiv:2608.08869v1 Announce Type: new Abstract: Large language models are increasingly used for ordinal classification, yet semantically equivalent changes to prompt organization can alter their predi

safetyarxiv-cs-cl
11 Aug 2026
Safety

Privacy-Preserving Data Drift Detection and Recovery for Large-Scale LLM Applications via Proxy Representations

DGX agent

arXiv:2608.08245v1 Announce Type: cross Abstract: LLM applications deployed at scale face a fundamental challenge: privacy constraints prevent direct inspection of user interactions, making it difficu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Private Anytime Selective-Risk Certification for Federated Retrieval-Augmented Generation: Guarantees and Empirical Limits

DGX agent

arXiv:2608.07913v1 Announce Type: cross Abstract: Selective-risk certificates promise that accepted outputs meet a declared error target. We develop Fed-SRC, a score-agnostic certificate for federated

safetyarxiv-cs-ai
11 Aug 2026
Safety

Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation

DGX agent

arXiv:2608.09263v1 Announce Type: new Abstract: Outcome verifiers score completed reasoning traces but do not assign credit to intermediate tokens. Privileged self-distillation attempts to fill this g

safetyarxiv-cs-ai
11 Aug 2026
Safety

Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation

DGX agent

arXiv:2608.09228v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) is commonly interpreted as the transfer of privileged information: a teacher observes the verified solution to the

safetyarxiv-cs-ai
11 Aug 2026
Safety

Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update

DGX agent

arXiv:2607.11505v2 Announce Type: replace-cross Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behav

safetyarxiv-cs-ai
11 Aug 2026
Safety

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

DGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

safetyarxiv-cs-ai
11 Aug 2026
Safety

Real-Time Nonlinear MPC via Sequential Quadratic Programming with Structure-Exploiting ADMM and Interior-Point Methods for Underactuated Double-Pendulum Swing-Up

DGX agent

arXiv:2608.09272v1 Announce Type: cross Abstract: The 4th 'AI Olympics with RealAIGym' competition, to be held at IJCAI-ECAI 2026 in Bremen, challenges participants to develop a global control policy

safetyarxiv-cs-ro
11 Aug 2026
Safety

Reconfigurable Structural Robotic Assembly: Interlocking 3D Aggregations with Self-Aligning Compound Nested Lattice Modules

DGX agent

arXiv:2608.07576v1 Announce Type: new Abstract: Robotic construction systems often treat the material system and the robot as separate design problems, locating intelligence primarily in hardware, sen

safetyarxiv-cs-ro
11 Aug 2026
Safety

Regret of exploratory policy improvement and q-learning

DGX agent

arXiv:2411.01302v2 Announce Type: replace Abstract: We study the convergence of q-learning and related algorithms introduced by Jia and Zhou (J. Mach. Learn. Res., 24 (2023), 161) for controlled diffu

safetyarxiv-cs-lg
11 Aug 2026
← Previous
1…7172737475…302
Next →