AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Inverse Critical Experiment Design via Gradient Optimization and a Multigroup Attention-Based Neural Network Architecture

DGX agent

arXiv:2606.04033v1 Announce Type: new Abstract: The validation of advanced nuclear reactor designs and fuel concepts requires critical experiments with high neutronic similarity to the target technolo

safetyarxiv-cs-lg
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

KODA: Contrastive Representation Comparison and Alignment for Vision-Language Foundation Models

DGX agent

arXiv:2606.04180v1 Announce Type: new Abstract: Vision-language foundation models such as CLIP and SigLIP provide widely used representations for multimodal learning systems. While these models are ty

safetyarxiv-cs-lg
4 Jun 2026
Safety

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

DGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

safetyarxiv-cs-cl
4 Jun 2026
Safety

LaVIDE: Language-Prompted Satellite Change Detection via Map-Image Alignment

DGX agent

arXiv:2411.19758v2 Announce Type: replace-cross Abstract: Remote sensing change detection based on a map reference and an up-to-date image boosts timely observation of the Earth's surface when earlier

safetyarxiv-cs-ai
4 Jun 2026
Safety

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents

DGX agent

arXiv:2606.04815v1 Announce Type: cross Abstract: Lifelong learning is essential for Large Language Model (LLM) agents operating in dynamic, interactive environments. However, existing lifelong learni

safetyarxiv-cs-ai
4 Jun 2026
Safety

M3imic: Learning a Versatile Whole-Body Controller for Multimodal Motion Mimicking

DGX agent

arXiv:2606.04829v1 Announce Type: new Abstract: Building a general-purpose whole-body controller is essential for enabling diverse motion capabilities in humanoid robots across a wide range of downstr

safetyarxiv-cs-ro
4 Jun 2026
Safety

MATCH: Multi-faceted Adaptive Topo-Consistency for Semi-Supervised Histopathology Segmentation

DGX agent

arXiv:2510.01532v2 Announce Type: replace Abstract: In semi-supervised segmentation, capturing meaningful semantic structures from unlabeled data is essential. This is particularly challenging in hist

safetyarxiv-cs-cv
4 Jun 2026
Safety

(Mis)generalization of Helpful-only Fine-tuning

DGX agent

arXiv:2606.04413v1 Announce Type: new Abstract: Helpful-only models, that is, models that are trained to always follow user intent, are valuable for dangerous capability evaluations and other areas of

safetyarxiv-cs-lg
4 Jun 2026
Safety

MM-BizRAG: Rethinking Multimodal Retrieval-Augmented Generation for General Purpose Enterprise Q&A

DGX agent

arXiv:2606.04231v1 Announce Type: cross Abstract: Recent advances in multimodal retrieval-augmented generation (MM-RAG) have shifted toward minimal parsing, relying on page-level images for producing

safetyarxiv-cs-ai
4 Jun 2026
Safety

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

DGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

safetyarxiv-cs-ai
4 Jun 2026
Safety

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

DGX agent

arXiv:2606.04847v1 Announce Type: cross Abstract: Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggl

safetyarxiv-cs-cl
4 Jun 2026
Safety

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

DGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

safetyarxiv-cs-ai
4 Jun 2026
Safety

OA-CutMix: Correcting the Label Bias of CutMix

DGX agent

arXiv:2606.04820v1 Announce Type: cross Abstract: CutMix has become the de facto standard mixing augmentation, yet its label assignment rests on a flawed assumption: The area of the pasted patch faith

safetyarxiv-cs-ai
4 Jun 2026
Safety

On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

DGX agent

arXiv:2603.28762v2 Announce Type: replace-cross Abstract: Modern Text-to-Image (T2I) diffusion models have achieved remarkable semantic alignment, yet they often suffer from a significant lack of vari

safetyarxiv-cs-ai
4 Jun 2026
Safety

Optimal Transport Flow Matching by Design

DGX agent

arXiv:2606.04092v1 Announce Type: new Abstract: Flow matching models learn to transport samples from a simple prior distribution to a complex data distribution. When prior-data pairs are coupled via o

safetyarxiv-cs-cv
4 Jun 2026
Safety

Optimal Transport under Group Fairness Constraints

DGX agent

arXiv:2601.07144v3 Announce Type: replace-cross Abstract: Ensuring fairness in matching algorithms is a key challenge in allocating scarce resources and positions. Focusing on Optimal Transport (OT),

safetyarxiv-cs-lg
4 Jun 2026
Safety

OSCAR: Omni-Embodiment Skeleton-Conditioned World Action Model for Robotics

DGX agent

arXiv:2606.04463v1 Announce Type: new Abstract: We present OSCAR, a precise action-conditioned video world model that generalizes across different robot embodiments and enables robot policy evaluation

safetyarxiv-cs-ro
4 Jun 2026
Safety

Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data

DGX agent

arXiv:2601.15158v4 Announce Type: replace-cross Abstract: Transformers trained via Reinforcement Learning (RL) with outcome-based supervision can spontaneously develop the ability to generate intermed

safetyarxiv-cs-ai
4 Jun 2026
Safety

Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning

DGX agent

arXiv:2601.07408v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has emerged as a promising critic-free reinforcement learning paradigm for reasoning tasks. However, stand

safetyarxiv-cs-cl
4 Jun 2026
Safety

Path-conditioned training: a principled way to rescale ReLU neural networks

DGX agent

arXiv:2602.19799v2 Announce Type: replace-cross Abstract: Despite recent algorithmic advances, we still lack principled ways to leverage the well-documented rescaling symmetries in ReLU neural network

safetyarxiv-cs-lg
4 Jun 2026
Safety

PersonaTree: Structured Lifecycle Memory for Person Understanding in LLM Agents

DGX agent

arXiv:2606.04780v1 Announce Type: new Abstract: Persistent LLM agents require memory representations that make the formation of person understanding explicit across long term interaction. Existing age

safetyarxiv-cs-cl
4 Jun 2026
Safety

Physics-Informed Machine Learning for Short-Term Flood Prediction

DGX agent

arXiv:2606.04143v1 Announce Type: cross Abstract: Accurate flood forecasting is essential for mitigating disaster risks and protecting communities. However, purely data-driven machine learning models

safetyarxiv-cs-ai
4 Jun 2026
Safety

Plug-and-Play Diffusion Meets ADMM: Dual-Variable Coupling for Robust Medical Image Reconstruction

DGX agent

arXiv:2602.23214v2 Announce Type: replace Abstract: Plug-and-Play diffusion prior (PnPDP) frameworks have emerged as a powerful paradigm for solving imaging inverse problems by treating pretrained gen

safetyarxiv-cs-cv
4 Jun 2026
Safety

POLARIS: Guiding Small Models to Write Long Stories

DGX agent

arXiv:2606.04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quali

safetyarxiv-cs-ai
4 Jun 2026
Safety

Policy Gradient for Continuous-Time Robust Markov Decision Processes

DGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

safetyarxiv-cs-lg
4 Jun 2026
Safety

Potential-Guided Flow Matching for Vision-Language-Action Policy Improvement

DGX agent

arXiv:2606.04968v1 Announce Type: new Abstract: Large vision-language-action (VLA) policies are increasingly trained as conditional generative models over action chunks. Yet deployment produces mixed-

safetyarxiv-cs-ro
4 Jun 2026
Safety

Probing Outcome-Level Resemblance and Mechanism-Level Alignment in LLM Risk Decisions: Evidence from the St. Petersburg Game

DGX agent

arXiv:2606.04978v1 Announce Type: new Abstract: LLMs can appear cautious in risk decision-making tasks, yet cautious-looking outputs do not necessarily indicate alignment with human decision-making me

safetyarxiv-cs-cl
4 Jun 2026
Safety

Reducing the Filtering Effect in Public School Admissions: A Bias-aware Analysis for Targeted Interventions

DGX agent

arXiv:2004.10846v5 Announce Type: replace-cross Abstract: Problem definition: Traditionally, New York City's top 8 public schools have selected candidates solely based on their scores in the Specializ

safetyarxiv-cs-lg
4 Jun 2026
Safety

REGAIN: REconciliation GAIN-driven Auxiliary Direction Learning

DGX agent

arXiv:2606.04380v1 Announce Type: cross Abstract: Forecast reconciliation usually starts from a fixed measurement system and asks how forecasts should be projected onto a coherent space. We ask a diff

safetyarxiv-cs-lg
4 Jun 2026
Safety

Reinforcement Learning from Rich Feedback with Distributional DAgger

DGX agent

arXiv:2606.05152v1 Announce Type: cross Abstract: Reasoning models have advanced rapidly, but the dominant reinforcement learning from verifiable rewards (RLVR) recipe remains surprisingly narrow: sam

safetyarxiv-cs-ai
4 Jun 2026
Safety

RePercENT: Scaling Disentangled Representation Learning Beyond Two Modalities

DGX agent

arXiv:2606.05109v1 Announce Type: new Abstract: To leverage the full potential of multimodal data, we need representations that go beyond the state-of-the-art alignment and fusion approaches and explo

safetyarxiv-cs-lg
4 Jun 2026
Safety

Representation Matters in Randomized Smoothing for Audio Classification

DGX agent

arXiv:2606.04210v1 Announce Type: cross Abstract: Randomized smoothing (RS) certifies robustness in the vector space where Gaussian noise is added. In audio classification, this space is often not uni

safetyarxiv-cs-lg
4 Jun 2026
Safety

Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.04923v1 Announce Type: cross Abstract: Rubric-based reinforcement learning (RL) uses an LLM-as-a-Judge (LaaJ) to score model outputs according to rubrics as rewards. However, policy models

safetyarxiv-cs-ai
4 Jun 2026
Safety

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

DGX agent

arXiv:2606.04703v1 Announce Type: new Abstract: Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward c

safetyarxiv-cs-cl
4 Jun 2026
Safety

Rethinking Sales Lead Scoring with LLM-based Hierarchical Preference Ranking

DGX agent

arXiv:2606.04387v1 Announce Type: cross Abstract: Sales lead conversion in high-stakes domains (e.g., automotive, real estate) differs fundamentally from e-commerce recommendation due to prolonged dec

safetyarxiv-cs-ai
4 Jun 2026
Safety

Reusing Trajectories in Policy Gradients Enables Fast Convergence

DGX agent

arXiv:2506.06178v3 Announce Type: replace Abstract: Policy gradient (PG) methods are a class of effective reinforcement learning algorithms, particularly when dealing with continuous control problems.

safetyarxiv-cs-lg
4 Jun 2026
Safety

RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training

DGX agent

arXiv:2606.04272v1 Announce Type: new Abstract: The standard LLM training pipeline applies reinforcement learning (RL) only after pre-training and supervised fine-tuning (SFT). We question this status

safetyarxiv-cs-lg
4 Jun 2026
Safety

Scaling Self-Evolving Agents via Parametric Memory

DGX agent

arXiv:2606.04536v1 Announce Type: new Abstract: Existing memory-augmented LLM agents store past experience exclusively in prompt space, as textual summaries or retrieved passages, while keeping model

safetyarxiv-cs-ai
4 Jun 2026
Safety

Segment, Embed, and Align: A Universal Recipe for Aligning Subtitles to Signing

DGX agent

arXiv:2512.08094v2 Announce Type: replace Abstract: The goal of this work is to develop a universal approach for aligning subtitles (i.e., spoken language text with corresponding timestamps) to contin

safetyarxiv-cs-cl
4 Jun 2026
Safety

Selective Coupling of Decoupled Informative Regions: Masked Attention Alignment for Data-Free Quantization of Vision Transformers

DGX agent

arXiv:2606.04373v1 Announce Type: cross Abstract: Data-Free Quantization (DFQ) addresses data security concerns by synthesizing samples, without accessing real data. It has garnered increasing attenti

safetyarxiv-cs-ai
4 Jun 2026
Safety

Self-Distilled Policy Gradient

DGX agent

arXiv:2606.04036v1 Announce Type: new Abstract: On-policy self-distillation, where a language model conditions on privileged context to supervise its own generations, is a promising source of dense su

safetyarxiv-cs-lg
4 Jun 2026
Safety

Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model

DGX agent

arXiv:2512.21917v3 Announce Type: replace-cross Abstract: Policy alignment to preference data typically assumes a known link function between observed preferences and latent rewards (e.g., Bradley-Ter

safetyarxiv-cs-ai
4 Jun 2026
Safety

Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents

DGX agent

arXiv:2510.13704v2 Announce Type: replace-cross Abstract: Recent works have proposed accelerating the wall-clock training time of actor-critic methods via the use of large-scale environment paralleliz

safetyarxiv-cs-ai
4 Jun 2026
Safety

Smart Transportation Without Neurons -- Fair Metro Network Expansion with Tabular Reinforcement Learning

DGX agent

arXiv:2606.04167v1 Announce Type: cross Abstract: We tackle the Metro Network Expansion Problem (MNEP), a subset of the Transport Network Design Problem (TNDP), which focuses on expanding metro system

safetyarxiv-cs-ai
4 Jun 2026
Safety

SocialCoach: Personalized Social Skill Learning with RL-based Agentic Tutoring and Practice

DGX agent

arXiv:2606.04155v1 Announce Type: cross Abstract: Social skills such as negotiation and leadership are crucial for personal and professional success in today's interconnected world. However, scalable

safetyarxiv-cs-cl
4 Jun 2026
Safety

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

DGX agent

arXiv:2505.11166v3 Announce Type: replace-cross Abstract: Despite advances in pretraining with extended context sizes, large language models (LLMs) still face challenges in effectively utilizing real-

safetyarxiv-cs-ai
4 Jun 2026
Safety

SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space

DGX agent

arXiv:2511.20102v3 Announce Type: replace Abstract: Sparse attention reduces the quadratic complexity of full self-attention but faces two challenges: (1) an attention gap, where applying sparse atten

safetyarxiv-cs-cl
4 Jun 2026
Safety

Stumbling Into AI Emotional Dependence: How Routine AI Interactions Reshape Human Connection

DGX agent

arXiv:2606.04150v1 Announce Type: new Abstract: Public discourse and emerging policy typically assume that AI emotional support is a deliberate act: a lonely user consciously seeking comfort from a de

safetyarxiv-cs-ai
4 Jun 2026
← Previous
1…136137138139140…260
Next →