AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

DGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

safetyarxiv-cs-cv
21 Apr 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Inertia in Moral and Value Judgments of Large Language Models

DGX agent

arXiv:2408.09049v3 Announce Type: replace Abstract: Large Language Models (LLMs) behave non-deterministically, and prompting has become a common method for steering their outputs. A popular strategy i

safetyarxiv-cs-cl
21 Apr 2026
Safety

Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias

DGX agent

arXiv:2601.16934v2 Announce Type: replace Abstract: To be discoverable in an embedding-based search process, each part of a document should be reflected in its embedding representation. To quantify an

safetyarxiv-cs-cl
21 Apr 2026
Safety

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

DGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

safetyarxiv-cs-cv
21 Apr 2026
Safety

Instinct vs. Reflection: Unifying Token and Verbalized Confidence in Multimodal Large Models

DGX agent

arXiv:2604.17274v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in various perception and reasoning tasks. Despite this success, ens

safetyarxiv-cs-cv
21 Apr 2026
Safety

Integrated Wheel Sensor Communication using ESP32 -- A Contribution towards a Digital Twin of the Road System

DGX agent

arXiv:2509.04061v2 Announce Type: replace Abstract: While current onboard state estimation methods are adequate for most driving and safety-related applications, they do not provide insights into the

safetyarxiv-cs-ro
21 Apr 2026
Safety

Inter-Agent Relative Representations for Multi-Agent Option Discovery

DGX agent

arXiv:2512.24827v3 Announce Type: replace Abstract: Temporally extended actions improve the ability to explore and plan in single-agent settings. In multi-agent settings, the exponential growth of the

safetyarxiv-cs-lg
21 Apr 2026
Safety

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

DGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

safetyarxiv-cs-cl
21 Apr 2026
Safety

J-PARSE: Jacobian-based Projection Algorithm for Resolving Singularities Effectively in Inverse Kinematic Control of Serial Manipulators

DGX agent

arXiv:2505.00306v5 Announce Type: replace Abstract: J-PARSE is an algorithm for smooth first-order inverse kinematic control of a serial manipulator near kinematic singularities. The commanded end-eff

safetyarxiv-cs-ro
21 Apr 2026
Safety

Jailbreaking Large Language Models with Morality Attacks

DGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

safetyarxiv-cs-cl
21 Apr 2026
Safety

Last Wednesday, our founder and scientific advisor, @Yoshua_Bengio, was officially appointed an Officer of the Order of the British Empire (…

DGX agent

Last Wednesday, our founder and scientific advisor, @Yoshua_Bengio, was officially appointed an Officer of the Order of the British Empire (OBE). This prestigious distinction recognizes his contributi

safetyyoshua-bengio--x
21 Apr 2026
Safety

LatentMimic: Terrain-Adaptive Locomotion via Latent Space Imitation

DGX agent

arXiv:2604.16440v1 Announce Type: new Abstract: Developing natural and diverse locomotion controllers for quadruped robots that can adapt to complex terrains while preserving motion style remains a si

safetyarxiv-cs-ro
21 Apr 2026
Safety

Learning-Based Sparsification of Dynamic Graphs in Robotic Exploration Algorithms

DGX agent

arXiv:2604.16509v1 Announce Type: cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapi

safetyarxiv-cs-lg
21 Apr 2026
Safety

Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification

DGX agent

arXiv:2601.21244v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced LLM reasoning, but remains constrained by inefficient exploration under lim

safetyarxiv-cs-cl
21 Apr 2026
Safety

LiDAR-based Crowd Navigation with Visible Edge Group Representation

DGX agent

arXiv:2604.16741v1 Announce Type: new Abstract: Robot navigation in crowded pedestrian environments is a well-known challenge and we explore the practical deployment of group-based representations in

safetyarxiv-cs-ro
21 Apr 2026
Safety

LLM-Extracted Covariates for Clinical Causal Inference: Rethinking Integration Strategies

DGX agent

arXiv:2604.16763v1 Announce Type: new Abstract: Causal inference from electronic health records (EHR) is fundamentally limited by unmeasured confounding: critical clinical states such as frailty, goal

safetyarxiv-cs-lg
21 Apr 2026
Safety

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams …

DGX agent

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams actually close the evals feedback loop. Looking forward to I

safetyharrison-chase--x
21 Apr 2026
Safety

Mammo-FM: Breast-specific foundational model for Integrated Mammographic Diagnosis, Prognosis, and Reporting

DGX agent

arXiv:2512.00198v2 Announce Type: replace Abstract: Breast cancer is one of the leading causes of death among women worldwide. We introduce Mammo-FM, the first foundation model specifically for mammog

safetyarxiv-cs-cv
21 Apr 2026
Safety

Mark Zuckerberg and Meta Platforms $META just sent a memo to employees saying Meta Platforms is installing a new tracking software on the co…

DGX agent

Mark Zuckerberg and Meta Platforms $META just sent a memo to employees saying Meta Platforms is installing a new tracking software on the computers of all employees in the United States 🇺🇸 so it can t

safetygary-marcus--x
21 Apr 2026
Safety

MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning

DGX agent

arXiv:2602.17550v3 Announce Type: replace Abstract: Existing Reinforcement Learning with Verifiable Rewards (RLVR) algorithms, such as GRPO, rely on rigid, uniform, and symmetric trust region mechanis

safetyarxiv-cs-lg
21 Apr 2026
Safety

MASSIVE: 🇺🇸 The BBC just validated everything we've been saying. A clear pattern of trades right before major Trump announcements. Iran wa…

DGX agent

MASSIVE: 🇺🇸 The BBC just validated everything we've been saying. A clear pattern of trades right before major Trump announcements. Iran war. Tariff reversals. Policy shifts. We tracked a whale for wee

safetyyann-lecun--x
21 Apr 2026
Safety

Mechanisms of Multimodal Synchronization: Insights from Decoder-Based Video-Text-to-Speech Synthesis

DGX agent

arXiv:2411.17690v3 Announce Type: replace-cross Abstract: Unified decoder-only transformers have shown promise for multimodal generation, yet the mechanisms by which they synchronize modalities with h

safetyarxiv-cs-cv
21 Apr 2026
Safety

MESA: A Training-Free Multi-Exemplar Deep Framework for Restoring Ancient Inscription Textures

DGX agent

arXiv:2604.17390v1 Announce Type: new Abstract: Ancient inscriptions frequently suffer missing or corrupted regions from fragmentation, erosion, or other damage, hindering reading, and analysis. We re

safetyarxiv-cs-cv
21 Apr 2026
Safety

Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training (Reuters)

DGX agent

Reuters: Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training — Meta (META.O) is installing new

safetytechmeme
21 Apr 2026
Safety

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

DGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

safetyarxiv-cs-cl
21 Apr 2026
Safety

Mix and Match: Context Pairing for Scalable Topic-Controlled Educational Summarisation

DGX agent

arXiv:2604.18087v1 Announce Type: new Abstract: Topic-controlled summarisation enables users to generate summaries focused on specific aspects of source documents. This paper investigates a data augme

safetyarxiv-cs-cl
21 Apr 2026
Safety

MoCo: A One-Stop Shop for Model Collaboration Research

DGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

safetyarxiv-cs-cl
21 Apr 2026
Safety

Modeling User Exploration Saturation: When Recommender Systems Should Stop Pushing Novelty

DGX agent

arXiv:2604.16419v1 Announce Type: cross Abstract: Fairness-aware recommender systems often mitigate bias by increasing exposure to under-represented or long-tail content, commonly through mechanisms t

safetyarxiv-cs-lg
21 Apr 2026
Safety

Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition

DGX agent

arXiv:2604.17062v1 Announce Type: new Abstract: Zero-shot action recognition is challenging due to the semantic gap between seen and unseen classes. We present a novel framework that enhances CLIP wit

safetyarxiv-cs-cv
21 Apr 2026
Safety

Multimodal Policy Internalization for Conversational Agents

DGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

safetyarxiv-cs-cl
21 Apr 2026
Safety

Navigating Distribution Shifts in Medical Image Analysis: A Survey

DGX agent

arXiv:2411.05824v3 Announce Type: replace-cross Abstract: Medical Image Analysis (MedIA) has become indispensable in modern healthcare, enhancing clinical diagnostics and personalized treatment. Despi

safetyarxiv-cs-cv
21 Apr 2026
Safety

Navigating the Conceptual Multiverse

DGX agent

arXiv:2604.17815v1 Announce Type: cross Abstract: When language models answer open-ended problems, they implicitly make hidden decisions that shape their outputs, leaving users with uncontextualized a

safetyarxiv-cs-cl
21 Apr 2026
Safety

Negative Advantage Is a Double-Edged Sword: Calibrating Advantage in GRPO for Deep Search

DGX agent

arXiv:2604.18235v1 Announce Type: new Abstract: Deep search agents can autonomously initiate multi-turn interactions with search engines, thereby exhibiting strong question-answering capabilities. Suc

safetyarxiv-cs-cl
21 Apr 2026
Safety

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL

DGX agent

arXiv:2604.17706v1 Announce Type: new Abstract: Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception,

safetyarxiv-cs-ro
21 Apr 2026
Safety

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

DGX agent

arXiv:2604.16518v1 Announce Type: new Abstract: Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit percepti

safetyarxiv-cs-ro
21 Apr 2026
Safety

On Safety Risks in Experience-Driven Self-Evolving Agents

DGX agent

arXiv:2604.16968v1 Announce Type: new Abstract: Experience-driven self-evolution has emerged as a promising paradigm for improving the autonomy of large language model agents, yet its reliance on self

safetyarxiv-cs-cl
21 Apr 2026
Safety

On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks

DGX agent

arXiv:2510.03923v2 Announce Type: replace Abstract: Continuous-depth graph neural networks, also known as Graph Neural Differential Equations (GNDEs), combine the structural inductive bias of Graph Ne

safetyarxiv-cs-lg
21 Apr 2026
Safety

On the Importance of Tactile Sensing for Imitation Learning: A Case Study on Robotic Match Lighting

DGX agent

arXiv:2504.13618v4 Announce Type: replace Abstract: The field of robotic manipulation has advanced significantly in recent years. At the sensing level, several novel tactile sensors have been develope

safetyarxiv-cs-ro
21 Apr 2026
Safety

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization

DGX agent

arXiv:2509.23542v2 Announce Type: replace Abstract: The LLM-as-a-judge paradigm is widely used in both evaluating free-text model responses and reward modeling for model alignment and fine-tuning. Rec

safetyarxiv-cs-cl
21 Apr 2026
Safety

One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

DGX agent

arXiv:2601.18731v2 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individu

safetyarxiv-cs-cl
21 Apr 2026
Safety

Online Conformal Prediction with Adversarial Semi-bandit Feedback via Regret Minimization

DGX agent

arXiv:2604.17984v1 Announce Type: new Abstract: Uncertainty quantification is crucial in safety-critical systems, where decisions must be made under uncertainty. In particular, we consider the problem

safetyarxiv-cs-lg
21 Apr 2026
Safety

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

safetyarxiv-cs-cv
21 Apr 2026
Safety

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

DGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

safetyarxiv-cs-cl
21 Apr 2026
Safety

OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection

DGX agent

arXiv:2511.21064v2 Announce Type: replace-cross Abstract: Open-Vocabulary Object Detection (OVOD) aims to enable detectors to generalize across categories by leveraging semantic information. Although

safetyarxiv-cs-cv
21 Apr 2026
Safety

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling

DGX agent

arXiv:2510.24235v3 Announce Type: replace Abstract: Reward models (RMs) are central to reinforcement learning from human feedback (RLHF), providing the critical supervision signals that align large la

safetyarxiv-cs-lg
21 Apr 2026
Safety

Peerispect: Claim Verification in Scientific Peer Reviews

DGX agent

arXiv:2604.17667v1 Announce Type: new Abstract: Peer review is central to scientific publishing, yet reviewers frequently include claims that are subjective, rhetorical, or misaligned with the submitt

safetyarxiv-cs-cl
21 Apr 2026
Safety

PEPR: Privileged Event-based Predictive Regularization for Domain Generalization

DGX agent

arXiv:2602.04583v2 Announce Type: replace Abstract: Deep neural networks for visual perception are highly susceptible to domain shift, which poses a critical challenge for real-world deployment under

safetyarxiv-cs-cv
21 Apr 2026
Safety

Plasticity Loss in Deep Reinforcement Learning: A Survey

DGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

safetyarxiv-cs-lg
21 Apr 2026
← Previous
1…236237238239240…265
Next →