AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

How Off-Policy Can GRPO Be? Mu-GRPO for Efficient LLM Reinforcement Learning

DGX agent

arXiv:2605.17570v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has been a key driver of recent progress in reinforcement learning with verifiable rewards (RLVR) for large

safetyarxiv-cs-cl
19 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How to Instruct Your Robot: Dense Language Annotations Power Robot Policy Learning

DGX agent

arXiv:2605.17077v1 Announce Type: cross Abstract: Scaling robot policy learning is bottlenecked by the cost of collecting demonstrations, while language annotations for existing demonstrations are com

safetyarxiv-cs-ai
19 May 2026
Safety

How Wrong Can Your Counterfactual Be? Quantifying Confounding Bias for Continuous Treatments without a Control Group

DGX agent

arXiv:2603.07438v2 Announce Type: replace Abstract: Stress testing poses a causal question: how would portfolio credit losses change if the macroeconomy followed an adverse counterfactual path? Yet st

safetyarxiv-cs-ai
19 May 2026
Safety

Identifying Latent Actions and Dynamics from Offline Data via Demonstrator Diversity

DGX agent

arXiv:2603.17577v2 Announce Type: replace-cross Abstract: Can latent actions and environment dynamics be recovered from offline trajectories when actions are never observed? We study this question in

safetyarxiv-cs-ai
19 May 2026
Safety

'I'm Not Mad, Just Focused'': Understanding Human Emotions in Human-Robot Collaboration

DGX agent

arXiv:2605.16816v1 Announce Type: new Abstract: Human-robot collaboration (HRC) can benefit from robots' abilities to interpret human emotional states. However, current emotion recognition (ER) models

safetyarxiv-cs-ro
19 May 2026
Safety

Implicit Hierarchical GRPO: Decoupling Tool Invocation from Execution for Tool-Integrated Mathematical Reasoning

DGX agent

arXiv:2605.18500v1 Announce Type: new Abstract: Large language models (LLMs) have increasingly leveraged tool invocation to enhance their reasoning capabilities. However, existing approaches typically

safetyarxiv-cs-cl
19 May 2026
Safety

Improved Baselines with Representation Autoencoders

DGX agent

arXiv:2605.18324v1 Announce Type: cross Abstract: Representation Autoencoders (RAE) replace traditional VAE with pretrained vision encoders. In this paper, we systematically investigate several design

safetyarxiv-cs-ai
19 May 2026
Safety

Improving MLLM Training Efficiency via Stage-Aware Sparsity

DGX agent

arXiv:2509.18150v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated outstanding performance across a variety of domains. However, training MLLMs is oft

safetyarxiv-cs-ai
19 May 2026
Safety

Individual utilities of life satisfaction reveal inequality aversion unrelated to political alignment

DGX agent

arXiv:2509.07793v4 Announce Type: replace-cross Abstract: How should well-being be prioritised in society, and what trade-offs are people willing to make between fairness and personal well-being? We i

safetyarxiv-cs-ai
19 May 2026
Safety

InFeR: Informed Failure Resilience in Learned Visual Navigation Control

DGX agent

arXiv:2510.24680v2 Announce Type: replace Abstract: While imitation learning (IL) has enabled successful visual navigation in many common environments, IL policies are prone to unpredictable failures

safetyarxiv-cs-ro
19 May 2026
Safety

Interpretable epistemic uncertainty decomposition in sequential generative models via polynomial chaos surrogates

DGX agent

arXiv:2510.21523v2 Announce Type: replace Abstract: Sequential generative models conditioned on uncertain rewards are central to AI-driven scientific discovery, yet the epistemic uncertainty they inhe

safetyarxiv-cs-lg
19 May 2026
Safety

it’s strange how many podcasters shy away from airing both sides of an argument that may totally shape our lives. https://x.com/benjamin_hor…

DGX agent

Gary Marcus critiques podcasters for avoiding balanced discussion of contentious issues that significantly impact society, suggesting there is an reluctance to present opposing viewpoints on important

safetygary-marcus--x
19 May 2026
Safety

LACE: Latent Visual Representation for Cross-Embodiment Learning

DGX agent

arXiv:2605.16743v1 Announce Type: new Abstract: Cross-embodiment learning from human demonstrations is hindered by the visual gap between human and robot embodiments. While self-supervised learning (S

safetyarxiv-cs-ro
19 May 2026
Safety

Lance: Unified Multimodal Modeling by Multi-Task Synergy

DGX agent

arXiv:2605.18678v1 Announce Type: cross Abstract: We present Lance, a lightweight native unified model supporting multimodal understanding, generation, and editing for both images and videos. Rather t

safetyarxiv-cs-ai
19 May 2026
Safety

Latent Action Control for Reasoning-Guided Unified Image Generation

DGX agent

arXiv:2605.16961v1 Announce Type: cross Abstract: Unified multimodal models can encode visual understanding and image generation within a shared backbone, yet understanding does not automatically tran

safetyarxiv-cs-ai
19 May 2026
Safety

LatentUMM: Dual Latent Alignment for Unified Multimodal Models

DGX agent

arXiv:2605.17766v1 Announce Type: new Abstract: Unified multimodal models (UMMs) achieve strong performance in both understanding and generation by learning a shared latent space, yet they often exhib

safetyarxiv-cs-cv
19 May 2026
Safety

Learning Fill-in Reduction Ordering via Graph Policy Optimization for Sparse Matrices

DGX agent

arXiv:2605.17362v1 Announce Type: new Abstract: Matrix reordering in large sparse solvers seeks a permutation that minimizes factorization fill-in to reduce memory and computation. Because the minimum

safetyarxiv-cs-lg
19 May 2026
Safety

Learning Multi-Timescale Abstractions for Hierarchical Combinatorial Planning

DGX agent

arXiv:2605.17058v1 Announce Type: new Abstract: The combination of exponentially large action spaces, stochastic dynamics, and long-horizon decision-making under limited resources makes Sequential Sto

safetyarxiv-cs-lg
19 May 2026
Safety

Learning Native Continuation for Action Chunking Flow Policies

DGX agent

arXiv:2602.12978v2 Announce Type: replace-cross Abstract: Action chunking enables Vision Language Action (VLA) models to run in real time, but naive chunked execution often exhibits discontinuities at

safetyarxiv-cs-ai
19 May 2026
Safety

Learning Relative Representations for Fine-Grained Multimodal Alignment with Limited Data

DGX agent

arXiv:2605.16834v1 Announce Type: cross Abstract: Multimodal pre-training demonstrates strong generalization performance, but this paradigm is often impractical in domains where paired data are scarce

safetyarxiv-cs-ai
19 May 2026
Safety

Learning to Reason without External Rewards

DGX agent

arXiv:2505.19590v5 Announce Type: replace-cross Abstract: Training large language models (LLMs) for complex reasoning via Reinforcement Learning with Verifiable Rewards (RLVR) is effective but limited

safetyarxiv-cs-cl
19 May 2026
Safety

Learning Transferable Topology Priors for Multi-Agent LLM Collaboration Across Domains

DGX agent

arXiv:2605.17359v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential for complex reasoning by coordinating specialized agents through struct

safetyarxiv-cs-cl
19 May 2026
Safety

Learning under Distributional Drift: Prequential Reproducibility as an Intrinsic Statistical Resource

DGX agent

arXiv:2512.13506v4 Announce Type: replace Abstract: Statistical learning under distributional drift remains poorly characterized, especially in closed-loop settings where learning alters the data-gene

safetyarxiv-cs-lg
19 May 2026
Safety

LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters

DGX agent

arXiv:2605.12012v2 Announce Type: replace Abstract: Public-sector legal departments in the Netherlands face acute staff shortages, increased case volumes, and increased pressure to meet regulatory com

safetyarxiv-cs-ai
19 May 2026
Safety

Linguistic Uncertainty and Reply Engagement on X: A Cross-Domain Replication of the Uncertainty-Reply Asymmetry

DGX agent

arXiv:2605.16289v1 Announce Type: cross Abstract: Linguistic uncertainty is common in social media, but its relationship with engagement remains unclear across languages and topics. Using 2,258 Englis

safetyarxiv-cs-cl
19 May 2026
Safety

LISTEN to Your Preferences: An LLM Framework for Multi-Objective Selection

DGX agent

arXiv:2510.25799v2 Announce Type: replace Abstract: Human experts often struggle to select the best option from a large set of items with multiple competing objectives, a process bottlenecked by the d

safetyarxiv-cs-cl
19 May 2026
Safety

Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory

DGX agent

arXiv:2605.17478v1 Announce Type: new Abstract: Visual Geometry Grounded Transformers (VGGT) have set new benchmarks in high-fidelity 3D scene reconstruction. However, as the sequence length increases

safetyarxiv-cs-cv
19 May 2026
Safety

MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization

DGX agent

arXiv:2605.17997v1 Announce Type: cross Abstract: Recently, residual reconstruction-based model quantization methods have achieved promising performance in low-bit post-training quantization (PTQ) by

safetyarxiv-cs-ai
19 May 2026
Safety

Masking Causality and Conditional Dependence

DGX agent

arXiv:2603.06984v2 Announce Type: replace-cross Abstract: Many regulatory and analytic problems require that a prohibited variable influence a decision only through a designated allowable channel -- a

safetyarxiv-cs-ai
19 May 2026
Safety

Measuring Changes in Instructor Class Design and Student Learning After the Release of Large Language Models (LLMs)

DGX agent

arXiv:2605.16284v1 Announce Type: cross Abstract: Student use of Generative AI (GenAI) products in completing their classwork, with or without their professors' knowledge and/or approval, has resulted

safetyarxiv-cs-ai
19 May 2026
Safety

Medical Context Distorts Decisions in Clinical Vision Language Models

DGX agent

arXiv:2605.17436v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly proposed for clinical decision support, yet their reliability in real-world scenarios that require inte

safetyarxiv-cs-cl
19 May 2026
Safety

Minor First, Major Last: A Depth-Induced Implicit Bias of Sharpness-Aware Minimization

DGX agent

arXiv:2603.08290v2 Announce Type: replace-cross Abstract: We study the implicit bias of Sharpness-Aware Minimization (SAM) when training L-layer linear diagonal networks on linearly separable binary c

safetyarxiv-cs-ai
19 May 2026
Safety

Mitigating Conversational Inertia in Multi-Turn Agents

DGX agent

arXiv:2602.03664v3 Announce Type: replace Abstract: Large language models excel as few-shot learners when provided with appropriate demonstrations, yet this strength becomes problematic in multiturn a

safetyarxiv-cs-ai
19 May 2026
Safety

MoASE++: Mixture of Activation Sparsity Experts with Domain-Adaptive On-policy Distillation for Continual Test Time Adaptation

DGX agent

arXiv:2605.17743v1 Announce Type: new Abstract: Continual test-time adaptation adapts a source-pretrained model to non-stationary, unlabeled target streams while retaining past competence, yet texture

safetyarxiv-cs-cv
19 May 2026
Safety

MSIQ: Moment-based Scale-Invariant Quality Measure for Single Image Super-Resolution

DGX agent

arXiv:2605.17588v1 Announce Type: new Abstract: Assessing the quality of single image super-resolution (SISR) results remains an open methodological problem. Common full-reference metrics (PSNR, SSIM,

safetyarxiv-cs-cv
19 May 2026
Safety

Multi-Order Matching Network for Alignment-Free Depth Super-Resolution

DGX agent

arXiv:2511.16361v3 Announce Type: replace Abstract: Recent guided depth super-resolution methods are premised on the assumption of strict spatial alignment between depth and RGB, achieving high-qualit

safetyarxiv-cs-cv
19 May 2026
Safety

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages

DGX agent

arXiv:2605.17152v1 Announce Type: new Abstract: Multimodal LLMs are evolving from vision-language to tri-modality that see, hear, and read, yet pipelines and benchmarks remain English-centric and comp

safetyarxiv-cs-cl
19 May 2026
Model Releases

Multilingual jailbreaking of LLMs using low-resource languages

DGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

model-releasesarxiv-cs-ai
19 May 2026
Safety

Natural-Language Agent Harnesses

DGX agent

arXiv:2603.25723v2 Announce Type: replace-cross Abstract: Agent performance is strongly shaped by the surrounding harness: the external execution system around a model that organizes a task run. Yet t

safetyarxiv-cs-ai
19 May 2026
Safety

Neural equilibria for long-term prediction of nonlinear conservation laws

DGX agent

arXiv:2501.06933v3 Announce Type: replace Abstract: Nonlinear conservation laws govern a broad class of important physical systems in science and industry and are central to scientific machine learnin

safetyarxiv-cs-lg
19 May 2026
Safety

Neural Visual Decoding via Cognitive guided Adaptive Blurring and Information Constrained Alignment

DGX agent

arXiv:2605.16418v1 Announce Type: cross Abstract: EEG-based visual decoding aims to establish a mapping between neural signals and visual semantics. However, it remains constrained by the dual challen

safetyarxiv-cs-ai
19 May 2026
Safety

NEWTON: Agentic Planning for Physically Grounded Video Generation

DGX agent

arXiv:2605.18396v1 Announce Type: new Abstract: Video generation models produce visually compelling results but systematically violate physical commonsense -- on VideoPhy-2, the best model achieves on

safetyarxiv-cs-cv
19 May 2026
Safety

Old Habits Die Hard: How Conversational History Geometrically Traps LLMs

DGX agent

arXiv:2603.03308v2 Announce Type: replace-cross Abstract: How does the conversational past of large language models (LLMs) influence their future performance? Recent work suggests that LLMs are affect

safetyarxiv-cs-ai
19 May 2026
Safety

oldsymbol{f}-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control

DGX agent

arXiv:2605.17862v1 Announce Type: cross Abstract: Scaling on-policy distillation (OPD) for large language models (LLMs) confronts a fundamental tension: asynchronous execution is necessary for system

safetyarxiv-cs-ai
19 May 2026
Safety

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression

DGX agent

arXiv:2601.21531v2 Announce Type: replace-cross Abstract: Visual token compression is widely used to accelerate large vision-language models (LVLMs) by pruning or merging visual tokens, yet its advers

safetyarxiv-cs-ai
19 May 2026
Safety

OPERA: A Reinforcement Learning--Enhanced Orchestrated Planner-Executor Architecture for Reasoning-Oriented Multi-Hop Retrieval

DGX agent

arXiv:2508.16438v4 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) and dense retrievers have driven significant progress in retrieval-augmented generation (RAG).

safetyarxiv-cs-ai
19 May 2026
Safety

Optimal Control of Multiclass Fluid Queueing Networks: A Machine Learning Approach

DGX agent

arXiv:2307.12405v2 Announce Type: replace Abstract: We propose a machine learning approach to the optimal control of multiclass fluid queueing networks (MFQNETs) that provides explicit and insightful

safetyarxiv-cs-lg
19 May 2026
Safety

OrbiSim: World Models as Differentiable Physics Engines for Embodied Intelligence

DGX agent

arXiv:2605.16395v1 Announce Type: cross Abstract: We present OrbiSim, a novel robotic simulation paradigm that redefines world models as a fully differentiable physics engine for embodied intelligence

safetyarxiv-cs-lg
19 May 2026
← Previous
1…200201202203204…302
Next →