AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
15 Apr 2026

Continuous Knowledge Metabolism: Generating Scientific Hypotheses from Evolving Literature

SafetyDGX agent

arXiv:2604.12243v1 Announce Type: cross Abstract: Scientific hypothesis generation requires tracking how knowledge evolves, not just what is currently known. We introduce Continuous Knowledge Metaboli

CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing

SafetyDGX agent

arXiv:2604.12292v1 Announce Type: cross Abstract: Movie dubbing aims to synthesize speech that preserves the vocal identity of a reference audio while synchronizing with the lip movements in a target

Cross-Cultural Simulation of Citizen Emotional Responses to Bureaucratic Red Tape Using LLM Agents

SafetyDGX agent

arXiv:2604.12545v1 Announce Type: new Abstract: Improving policymaking is a central concern in public administration. Prior human subject studies reveal substantial cross-cultural differences in citiz

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Modal Knowledge Distillation for PET-Free Amyloid-Beta Detection from MRI

SafetyDGX agent

arXiv:2604.12574v1 Announce Type: new Abstract: Detecting amyloid-eta (Aeta) positivity is crucial for early diagnosis of Alzheimer's disease but typically requires PET imaging, which is costly, invas

Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training

SafetyDGX agent

arXiv:2604.12967v1 Announce Type: new Abstract: Reinforcement Learning (RL) has shown strong potential for optimizing search agents in complex information retrieval tasks. However, existing approaches

DBGL: Decay-aware Bipartite Graph Learning for Irregular Medical Time Series Classification

SafetyDGX agent

arXiv:2604.11842v1 Announce Type: cross Abstract: Irregular Medical Time Series play a critical role in the clinical domain to better understand the patient's condition. However, inherent irregularity

Designing Reliable LLM-Assisted Rubric Scoring for Constructed Responses: Evidence from Physics Exams

SafetyDGX agent

arXiv:2604.12227v1 Announce Type: new Abstract: Student responses in STEM assessments are often handwritten and combine symbolic expressions, calculations, and diagrams, creating substantial variation

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

SafetyDGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models

SafetyDGX agent

arXiv:2511.00710v4 Announce Type: replace Abstract: Recent studies posit that Reinforcement Learning with Verifiable Rewards (RLVR) primarily amplifies behaviors inherent to the pre-training distribut

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

SafetyDGX agent

arXiv:2509.19695v3 Announce Type: replace-cross Abstract: Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient

Dynamic Multi-Robot Task Allocation under Uncertainty and Communication Constraints: A Game-Theoretic Approach

SafetyDGX agent

arXiv:2604.11954v1 Announce Type: cross Abstract: We study dynamic multi-robot task allocation under uncertain task completion, time-window constraints, and incomplete information. Tasks arrive online

E2E-Fly: An Integrated Training-to-Deployment System for End-to-End Quadrotor Autonomy

SafetyDGX agent

arXiv:2604.12916v1 Announce Type: new Abstract: Training and transferring learning-based policies for quadrotors from simulation to reality remains challenging due to inefficient visual rendering, phy

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

SafetyDGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

Euler-inspired Decoupling Neural Operator for Efficient Pansharpening

SafetyDGX agent

arXiv:2604.12463v1 Announce Type: cross Abstract: Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spe

Evaluating the Limitations of Protein Sequence Representations for Parkinson's Disease Classification

SafetyDGX agent

arXiv:2604.11852v1 Announce Type: cross Abstract: The identification of reliable molecular biomarkers for Parkinson's disease remains challenging due to its multifactorial nature. Although protein seq

Evolution-Inspired Sample Competition for Deep Neural Network Optimization

SafetyDGX agent

arXiv:2604.12568v1 Announce Type: new Abstract: Conventional deep network training generally optimizes all samples under a largely uniform learning paradigm, without explicitly modeling the heterogene

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

SafetyDGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

From Myopic Selection to Long-Horizon Awareness: Sequential LLM Routing for Multi-Turn Dialogue

SafetyDGX agent

arXiv:2604.12385v1 Announce Type: new Abstract: Multi-turn dialogue is the predominant form of interaction with large language models (LLMs). While LLM routing is effective in single-turn settings, ex

Generative AI in nutshell: 'If the technology is as formidable as the [tech CEO's] claim, then they could be leading us toward existential d…

SafetyDGX agent

Generative AI in nutshell: 'If the technology is as formidable as the [tech CEO's] claim, then they could be leading us toward existential disaster; if the technology proves less transformative, and t

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

SafetyDGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles

SafetyDGX agent

arXiv:2506.01256v4 Announce Type: replace-cross Abstract: Forced alignment is a common tool to align audio with orthographic and phonetic transcriptions. Most forced alignment tools provide only point

Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs

SafetyDGX agent

arXiv:2206.00939v3 Announce Type: replace-cross Abstract: The training of neural networks by gradient descent methods is a cornerstone of the deep learning revolution. Yet, despite some recent progres

Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs

SafetyDGX agent

arXiv:2604.05643v2 Announce Type: replace Abstract: Extending CoT through RL has been widely used to enhance the reasoning capabilities of LLMs. However, due to the sparsity of reward signals, it can

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions

SafetyDGX agent

arXiv:2604.12929v1 Announce Type: new Abstract: We present Grasp in Gaussians (GraG), a fast and robust method for reconstructing dynamic 3D hand-object interactions from a single monocular video. Unl

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since …

SafetyDGX agent

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since February People love Grok because it's the only AI they can

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

SafetyDGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

Hail to the Thief: Exploring Attacks and Defenses in Decentralised GRPO

SafetyDGX agent

arXiv:2511.09780v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has demonstrated wide adoption in the post-training of Large Language Models (LLMs). In GRPO, prompts are

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

SafetyDGX agent

arXiv:2509.19742v4 Announce Type: replace-cross Abstract: Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without co

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

SafetyDGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance

SafetyDGX agent

arXiv:2511.21356v2 Announce Type: replace-cross Abstract: Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate f…

SafetyDGX agent

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate for an Illinois bill that would shield them from liability if

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- …

SafetyDGX agent

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- into a clear list of beliefs. Here they are in full. 1. It’s

Incentivizing High-Quality Human Annotations with Golden Questions

SafetyDGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

Information-Geometric Decomposition of Generalization Error in Unsupervised Learning

SafetyDGX agent

arXiv:2604.12340v1 Announce Type: cross Abstract: We decompose the Kullback--Leibler generalization error (GE) -- the expected KL divergence from the data distribution to the trained model -- of unsup

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

SafetyDGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning

SafetyDGX agent

arXiv:2604.12303v1 Announce Type: new Abstract: Batch active learning (BAL) is a crucial technique for reducing labeling costs and improving data efficiency in training large-scale deep learning model

Learning step-level dynamic soaring in shear flow

SafetyDGX agent

arXiv:2604.12413v1 Announce Type: cross Abstract: Dynamic soaring enables sustained flight by extracting energy from wind shear, yet it is commonly understood as a cycle-level maneuver that assumes st

Learning Versatile Humanoid Manipulation with Touch Dreaming

SafetyDGX agent

arXiv:2604.13015v1 Announce Type: new Abstract: Humanoid robots promise general-purpose assistance, yet real-world humanoid loco-manipulation remains challenging because it requires whole-body stabili

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

SafetyDGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion

SafetyDGX agent

arXiv:2604.12286v1 Announce Type: new Abstract: Live Photo captures both a high-quality key photo and a short video clip to preserve the precious dynamics around the captured moment. While users may c

Man and machine: artificial intelligence and judicial decision making

SafetyDGX agent

arXiv:2603.19042v4 Announce Type: replace Abstract: The integration of artificial intelligence (AI) technologies into judicial decision-making, particularly in pretrial, sentencing, and parole context

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

SafetyDGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation

SafetyDGX agent

arXiv:2604.12277v1 Announce Type: new Abstract: Pretrained language models often rely on superficial features that appear predictive during training yet fail to generalize at test time, a phenomenon k

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization

SafetyDGX agent

arXiv:2604.12237v1 Announce Type: cross Abstract: In drug discovery, molecular optimization aims to iteratively refine a lead compound to improve molecular properties while preserving structural simil

Mutual Information Surprise: Rethinking Unexpectedness in Autonomous Systems

SafetyDGX agent

arXiv:2508.17403v3 Announce Type: replace Abstract: A community of researchers appears to think that a machine can be surprised and have introduced various surprise measures, principally the Shannon S

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning

SafetyDGX agent

arXiv:2601.06794v2 Announce Type: replace Abstract: Critique-guided reinforcement learning (RL) has emerged as a powerful paradigm for training LLM agents by augmenting sparse outcome rewards with nat

Not All Turns Are Equally Hard: Adaptive Thinking Budgets For Efficient Multi-Turn Reasoning

SafetyDGX agent

arXiv:2604.05164v2 Announce Type: replace-cross Abstract: As LLM reasoning performance plateau, improving inference-time compute efficiency is crucial to mitigate overthinking and long thinking traces

Offline-Online Reinforcement Learning for Linear Mixture MDPs

SafetyDGX agent

arXiv:2604.11994v1 Announce Type: new Abstract: We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data ar

PAINT: Partner-Agnostic Intent-Aware Cooperative Transport with Legged Robots

SafetyDGX agent

arXiv:2604.12852v1 Announce Type: new Abstract: Collaborative transport requires robots to infer partner intent through physical interaction while maintaining stable loco-manipulation. This becomes pa

Perception-Aware Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation

SafetyDGX agent

arXiv:2604.12113v1 Announce Type: cross Abstract: Visual Foundation Models (VFMs) such as the Segment Anything Model (SAM) have significantly advanced broad use of image segmentation. However, SAM and

Prediction from a year ago about AI backlash that looks to be on track:

SafetyDGX agent

Prediction from a year ago about AI backlash that looks to be on track: Public backlash against AI will so be strong by 2028 that anti-AI sentiment will probably be a major factor in the 2028 US Presi

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation

SafetyDGX agent

arXiv:2511.17097v2 Announce Type: replace Abstract: Vision-Language Navigation requires agents to act coherently over long horizons by understanding not only local visual context but also how far they

PubSwap: Public-Data Off-Policy Coordination for Federated RLVR

SafetyDGX agent

arXiv:2604.12160v1 Announce Type: new Abstract: Reasoning post-training with reinforcement learning from verifiable rewards (RLVR) is typically studied in centralized settings, yet many realistic appl

Redefining Quality Criteria and Distance-Aware Score Modeling for Image Editing Assessment

SafetyDGX agent

arXiv:2604.12175v1 Announce Type: new Abstract: Recent advances in image editing have heightened the need for reliable Image Editing Quality Assessment (IEQA). Unlike traditional methods, IEQA require

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

SafetyDGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

SafetyDGX agent

arXiv:2604.13016v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a core technique in the post-training of large language models, yet its training dynamics remain poorly unders

Retrieval as a Decision: Training-Free Adaptive Gating for Efficient RAG

SafetyDGX agent

arXiv:2511.09803v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves factuality but retrieving for every query often hurts quality while inflating tokens and latency. We p

SAM3-I: Segment Anything with Instructions

SafetyDGX agent

arXiv:2512.04585v3 Announce Type: replace Abstract: Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instanc

← Previous
1…209210211212213…240
Next →