AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
10 Apr 2026

Governed Capability Evolution for Embodied Agents: Safe Upgrade, Compatibility Checking, and Runtime Rollback for Embodied Capability Modules

SafetyDGX agent

arXiv:2604.08059v1 Announce Type: new Abstract: Embodied agents are increasingly expected to improve over time by updating their executable capabilities rather than rewriting the agent itself. Prior w

Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations

SafetyDGX agent

arXiv:2604.07517v1 Announce Type: new Abstract: Building generalist robots capable of performing functional grasping in everyday, open-world environments remains a significant challenge due to the vas

Guiding a Diffusion Model by Swapping Its Tokens

SafetyDGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles

SafetyDGX agent

arXiv:2604.07650v1 Announce Type: cross Abstract: The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretra

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

SafetyDGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

How to Evaluate Speech Translation with Source-Aware Neural MT Metrics

SafetyDGX agent

arXiv:2511.03295v3 Announce Type: replace-cross Abstract: Automatic evaluation of ST systems is typically performed by comparing translation hypotheses with one or more reference translations. While e

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

SafetyDGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

I agree totally @Gary. I’ve been saying this since the ChatGPT moment in Nov 22. Thank you for saying it out loud @demishassabis. Too much t…

SafetyDGX agent

I agree totally @Gary. I’ve been saying this since the ChatGPT moment in Nov 22. Thank you for saying it out loud @demishassabis. Too much time and energy being spent on mitigating the unintended cons

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance …

SafetyDGX agent

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance towards general intelligence. Same as it ever was. @GaryMarc

Improving Semantic Uncertainty Quantification in Language Model Question-Answering via Token-Level Temperature Scaling

SafetyDGX agent

arXiv:2604.07172v1 Announce Type: new Abstract: Calibration is central to reliable semantic uncertainty quantification, yet prior work has largely focused on discrimination, neglecting calibration. As

Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation

SafetyDGX agent

arXiv:2604.07945v1 Announce Type: new Abstract: As the demand for mobile robots continues to increase, social navigation has emerged as a critical task, driving active research into deep reinforcement

Indeed, if the LLM crew would just stick to this narrative, I would have a *lot* less to say 🤷‍♂️

SafetyDGX agent

Indeed, if the LLM crew would just stick to this narrative, I would have a *lot* less to say 🤷‍♂️ @GaryMarcus All we want is truth. Instead of overhyping LLMs, they should keep narative: -LLMs are use

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

SafetyDGX agent

arXiv:2604.08192v1 Announce Type: cross Abstract: Reliable generalization metrics are fundamental to the evaluation of machine learning models. Especially in high-stakes applications where labeled tar

Just launched at @aiDotEngineer : our official AGI Pills! prescribe one (1) if your colleague is saying we are hitting a wall and/or trying …

SafetyDGX agent

Just launched at @aiDotEngineer : our official AGI Pills! prescribe one (1) if your colleague is saying we are hitting a wall and/or trying to add inductive bias instead of Trusting The Model Media bt

Karma Mechanisms for Decentralised, Cooperative Multi Agent Path Finding

SafetyDGX agent

arXiv:2604.07970v1 Announce Type: cross Abstract: Multi-Agent Path Finding (MAPF) is a fundamental coordination problem in large-scale robotic and cyber-physical systems, where multiple agents must co

KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.06691v1 Announce Type: new Abstract: Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. Whil

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents

SafetyDGX agent

arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

SafetyDGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

SafetyDGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

SafetyDGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

MCLR: Improving Conditional Modeling via Inter-Class Likelihood-Ratio Maximization and Unifying Classifier-Free Guidance with Alignment Objectives

SafetyDGX agent

arXiv:2603.22364v2 Announce Type: replace-cross Abstract: Diffusion models have achieved state-of-the-art performance in generative modeling, but their success often relies heavily on classifier-free

MDP modeling for multi-stage stochastic programs

SafetyDGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

Mixture Proportion Estimation and Weakly-supervised Kernel Test for Conditional Independence

SafetyDGX agent

arXiv:2604.07191v1 Announce Type: cross Abstract: Mixture proportion estimation (MPE) aims to estimate class priors from unlabeled data. This task is a critical component in weakly supervised learning

MO-RiskVAE: A Multi-Omics Variational Autoencoder for Survival Risk Modeling in Multiple MyelomaMO-RiskVAE

SafetyDGX agent

arXiv:2604.06267v1 Announce Type: cross Abstract: Multimodal variational autoencoders (VAEs) have emerged as a powerful framework for survival risk modeling in multiple myeloma by integrating heteroge

MonoUNet: A Robust Tiny Neural Network for Automated Knee Cartilage Segmentation on Point-of-Care Ultrasound Devices

SafetyDGX agent

arXiv:2604.07780v1 Announce Type: cross Abstract: Objective: To develop a robust and compact deep learning model for automated knee cartilage segmentation on point-of-care ultrasound (POCUS) devices.

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

SafetyDGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models

SafetyDGX agent

arXiv:2604.07991v1 Announce Type: new Abstract: Recent advances in world models have demonstrated strong capabilities in simulating physical reality, making them an increasingly important foundation f

MSCT: Differential Cross-Modal Attention for Deepfake Detection

SafetyDGX agent

arXiv:2604.07741v1 Announce Type: new Abstract: Audio-visual deepfake detection typically employs a complementary multi-modal model to check the forgery traces in the video. These methods primarily ex

Multi-agent Reach-avoid MDP via Potential Games and Low-rank Policy Structure

SafetyDGX agent

arXiv:2410.17690v2 Announce Type: replace-cross Abstract: We optimize finite horizon multi-agent reach-avoid Markov decision process (MDP) via local feedback policies. The global feedback polic

Multi-Faceted Self-Consistent Preference Alignment for Query Rewriting in Conversational Search

SafetyDGX agent

arXiv:2604.06771v1 Announce Type: cross Abstract: Conversational Query Rewriting (CQR) aims to rewrite ambiguous queries to achieve more efficient conversational search. Early studies have predominant

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

SafetyDGX agent

arXiv:2604.07148v1 Announce Type: new Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) ad

Neural Computers

SafetyDGX agent

arXiv:2604.06425v1 Announce Type: cross Abstract: We propose a new frontier: Neural Computers (NCs) -- an emerging machine form that unifies computation, memory, and I/O in a learned runtime state. Un

On the Global Photometric Alignment for Low-Level Vision

SafetyDGX agent

arXiv:2604.08172v1 Announce Type: new Abstract: Supervised low-level vision models rely on pixel-wise losses against paired references, yet paired training sets exhibit per-pair photometric inconsiste

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

SafetyDGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

Oracle is down more than 50% since two men (jointly) took over Safra Catz’s CEO job. Sexism is up 500%? 5000%?

SafetyDGX agent

Oracle is down more than 50% since two men (jointly) took over Safra Catz’s CEO job. Sexism is up 500%? 5000%? Nick Fuentes says women can only do three things “Women can be three things: they can be

OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance

SafetyDGX agent

arXiv:2604.08461v1 Announce Type: new Abstract: Open-Vocabulary Segmentation (OVS) aims to segment image regions beyond predefined category sets by leveraging semantic descriptions. While CLIP based a

OxEnsemble: Fair Ensembles for Low-Data Classification

SafetyDGX agent

arXiv:2512.09665v2 Announce Type: replace Abstract: We address the problem of fair classification in settings where data is scarce and unbalanced across demographic groups. Such low-data regimes are c

Part^{2}GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting

SafetyDGX agent

arXiv:2506.17212v2 Announce Type: replace Abstract: Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods.

Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon

SafetyDGX agent

arXiv:2503.02129v2 Announce Type: replace-cross Abstract: Path regularization has shown to be a very effective regularization to train neural networks, leading to a better generalization property than

PEER: Unified Process-Outcome Reinforcement Learning for Structured Empathetic Reasoning

SafetyDGX agent

arXiv:2508.09521v2 Announce Type: replace Abstract: Emotional support conversations require more than fluent responses. Supporters need to understand the seeker's situation and emotions, adopt an appr

People in Washington get played, yet again We really should worry about cybersecurity - a lot – but Mythos is not the model these guys think…

SafetyDGX agent

People in Washington get played, yet again We really should worry about cybersecurity - a lot – but Mythos is not the model these guys think it is. (See my newsletter today for three reasons why it is

Personalizing Text-to-Image Generation to Individual Taste

SafetyDGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models

SafetyDGX agent

arXiv:2604.07779v1 Announce Type: new Abstract: Pathology foundation models (FMs) have become central to computational histopathology, offering strong transfer performance across a wide range of diagn

PolySLGen: Online Multimodal Speaking-Listening Reaction Generation in Polyadic Interaction

SafetyDGX agent

arXiv:2604.08125v1 Announce Type: new Abstract: Human-like multimodal reaction generation is essential for natural group interactions between humans and embodied AI. However, existing approaches are l

PriPG-RL: Privileged Planner-Guided Reinforcement Learning for Partially Observable Systems with Anytime-Feasible MPC

SafetyDGX agent

arXiv:2604.08036v1 Announce Type: cross Abstract: This paper addresses the problem of training a reinforcement learning (RL) policy under partial observability by exploiting a privileged, anytime-feas

Probabilistic Language Tries: A Unified Framework for Compression, Decision Policies, and Execution Reuse

SafetyDGX agent

arXiv:2604.06228v1 Announce Type: cross Abstract: We introduce probabilistic language tries (PLTs), a unified representation that makes explicit the prefix structure implicitly defined by any generati

Quality-preserving Model for Electronics Production Quality Tests Reduction

SafetyDGX agent

arXiv:2604.06451v1 Announce Type: new Abstract: Manufacturing test flows in high-volume electronics production are typically fixed during product development and executed unchanged on every unit, even

Qualixar OS: A Universal Operating System for AI Agent Orchestration

SafetyDGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization

SafetyDGX agent

arXiv:2604.07165v1 Announce Type: new Abstract: Reinforcement learning for Large Language Model agents is often hindered by sparse rewards in multi-step reasoning tasks. Existing approaches like Group

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

SafetyDGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

SafetyDGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

Reflection-Based Task Adaptation for Self-Improving VLA

SafetyDGX agent

arXiv:2510.12710v3 Announce Type: replace Abstract: Pre-trained Vision-Language-Action (VLA) models represent a major leap towards general-purpose robots, yet efficiently adapting them to novel, speci

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

SafetyDGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

Region-R1: Reinforcing Query-Side Region Cropping for Multi-Modal Re-Ranking

SafetyDGX agent

arXiv:2604.05268v2 Announce Type: replace-cross Abstract: Multi-modal retrieval-augmented generation (MM-RAG) relies heavily on re-rankers to surface the most relevant evidence for image-question quer

Remember when ChatGPT helped someone plan a mass shooting, and a dozen OpenAI employees notified management, and they did nothing and 8 peop…

SafetyDGX agent

Remember when ChatGPT helped someone plan a mass shooting, and a dozen OpenAI employees notified management, and they did nothing and 8 people died? They don't want to be liable for their incompetence

Reset-Free Reinforcement Learning for Real-World Agile Driving: An Empirical Study

SafetyDGX agent

arXiv:2604.07672v1 Announce Type: new Abstract: This paper presents an empirical study of reset-free reinforcement learning (RL) for real-world agile driving, in which a physical 1/10-scale vehicle le

Revisiting Fairness Impossibility with Endogenous Behavior

SafetyDGX agent

arXiv:2604.06378v1 Announce Type: cross Abstract: In many real-world settings, institutions can and do adjust the consequences attached to algorithmic classification decisions, such as the size of fin

RewardFlow: Generate Images by Optimizing What You Reward

SafetyDGX agent

arXiv:2604.08536v1 Announce Type: new Abstract: We introduce RewardFlow, an inversion-free framework that steers pretrained diffusion and flow-matching models at inference time through multi-reward La

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

SafetyDGX agent

arXiv:2604.07774v1 Announce Type: cross Abstract: This paper focuses on embodied task planning, where an agent acquires visual observations from the environment and executes atomic actions to accompli

← Previous
1…217218219220221…240
Next →