AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,596 results
10 Apr 2026

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

SafetyDGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

From experimentation to engagement: on the paradox of participatory AI and power in contexts of forced displacement and humanitarian crises

SafetyDGX agent

arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

SafetyDGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

SafetyDGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly avai…

SafetyDGX agent

Garry @Kasparov63 retired from competitive chess over twenty years ago, and most or all of his tournament games are presumably publicly available - and yet I bet he still could crush any LLM that didn

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

SafetyDGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

SafetyDGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

Governed Capability Evolution for Embodied Agents: Safe Upgrade, Compatibility Checking, and Runtime Rollback for Embodied Capability Modules

SafetyDGX agent

arXiv:2604.08059v1 Announce Type: new Abstract: Embodied agents are increasingly expected to improve over time by updating their executable capabilities rather than rewriting the agent itself. Prior w

Governing frontier general-purpose AI in the public sector: adaptive risk management and policy capacity under uncertainty through 2030

SafetyDGX agent

arXiv:2604.06215v1 Announce Type: cross Abstract: The governance of frontier general-purpose artificial intelligence has become a public-sector problem of institutional design, not merely a technical

Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations

SafetyDGX agent

arXiv:2604.07517v1 Announce Type: new Abstract: Building generalist robots capable of performing functional grasping in everyday, open-world environments remains a significant challenge due to the vas

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs

SafetyDGX agent

arXiv:2604.07655v1 Announce Type: cross Abstract: Hard-gated safety checkers often over-refuse and misalign with a vendor's model spec; prevailing taxonomies also neglect robustness and honesty, yield

Guiding a Diffusion Model by Swapping Its Tokens

SafetyDGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution

SafetyDGX agent

arXiv:2604.07833v1 Announce Type: new Abstract: Embodied agents are evolving from passive reasoning systems into active executors that interact with tools, robots, and physical environments. Once gran

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization

SafetyDGX agent

arXiv:2604.06285v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become essential for tasks such as image synthesis, captioning, and retrieval by aligning textual and visual inform

How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles

SafetyDGX agent

arXiv:2604.07650v1 Announce Type: cross Abstract: The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretra

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

SafetyDGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

How to Evaluate Speech Translation with Source-Aware Neural MT Metrics

SafetyDGX agent

arXiv:2511.03295v3 Announce Type: replace-cross Abstract: Automatic evaluation of ST systems is typically performed by comparing translation hypotheses with one or more reference translations. While e

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

SafetyDGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

I agree totally @Gary. I’ve been saying this since the ChatGPT moment in Nov 22. Thank you for saying it out loud @demishassabis. Too much t…

SafetyDGX agent

I agree totally @Gary. I’ve been saying this since the ChatGPT moment in Nov 22. Thank you for saying it out loud @demishassabis. Too much time and energy being spent on mitigating the unintended cons

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance …

SafetyDGX agent

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance towards general intelligence. Same as it ever was. @GaryMarc

Improving Semantic Uncertainty Quantification in Language Model Question-Answering via Token-Level Temperature Scaling

SafetyDGX agent

arXiv:2604.07172v1 Announce Type: new Abstract: Calibration is central to reliable semantic uncertainty quantification, yet prior work has largely focused on discrimination, neglecting calibration. As

Incorporating Social Awareness into Control of Unknown Multi-Agent Systems: A Real-Time Spatiotemporal Tubes Approach

SafetyDGX agent

arXiv:2510.25597v2 Announce Type: replace-cross Abstract: This paper presents a decentralized control framework that incorporates social awareness into multi-agent systems with unknown dynamics to ach

Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation

SafetyDGX agent

arXiv:2604.07945v1 Announce Type: new Abstract: As the demand for mobile robots continues to increase, social navigation has emerged as a critical task, driving active research into deep reinforcement

Indeed, if the LLM crew would just stick to this narrative, I would have a *lot* less to say 🤷‍♂️

SafetyDGX agent

Indeed, if the LLM crew would just stick to this narrative, I would have a *lot* less to say 🤷‍♂️ @GaryMarcus All we want is truth. Instead of overhyping LLMs, they should keep narative: -LLMs are use

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

SafetyDGX agent

arXiv:2604.08192v1 Announce Type: cross Abstract: Reliable generalization metrics are fundamental to the evaluation of machine learning models. Especially in high-stakes applications where labeled tar

Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents

SafetyDGX agent

arXiv:2510.07809v4 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) empower autonomous mobile agents, yet their security under realistic mobile deployment constraints remain

Just launched at @aiDotEngineer : our official AGI Pills! prescribe one (1) if your colleague is saying we are hitting a wall and/or trying …

SafetyDGX agent

Just launched at @aiDotEngineer : our official AGI Pills! prescribe one (1) if your colleague is saying we are hitting a wall and/or trying to add inductive bias instead of Trusting The Model Media bt

Karma Mechanisms for Decentralised, Cooperative Multi Agent Path Finding

SafetyDGX agent

arXiv:2604.07970v1 Announce Type: cross Abstract: Multi-Agent Path Finding (MAPF) is a fundamental coordination problem in large-scale robotic and cyber-physical systems, where multiple agents must co

KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.06691v1 Announce Type: new Abstract: Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. Whil

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents

SafetyDGX agent

arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

SafetyDGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

SafetyDGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

SafetyDGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

Learning Without Losing Identity: Capability Evolution for Embodied Agents

SafetyDGX agent

arXiv:2604.07799v1 Announce Type: new Abstract: Embodied agents are expected to operate persistently in dynamic physical environments, continuously acquiring new capabilities over time. Existing appro

Leveraging Wireless Sensor Networks for Real-Time Monitoring and Control of Industrial Environments

SafetyDGX agent

arXiv:2510.13820v3 Announce Type: replace-cross Abstract: This research proposes an extensive technique for monitoring and controlling the industrial parameters using Internet of Things (IoT) technolo

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

SafetyDGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

LINE: LLM-based Iterative Neuron Explanations for Vision Models

SafetyDGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

SafetyDGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

LUMINA: Foundation Models for Topology Transferable ACOPF

SafetyDGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

Machine Unlearning in the Era of Quantum Machine Learning: An Empirical Study

SafetyDGX agent

arXiv:2512.19253v4 Announce Type: replace-cross Abstract: We present the first empirical study of machine unlearning (MU) in hybrid quantum-classical neural networks. While MU has been extensively exp

Mark “Metaverse” Zuckerberg totally bought Moltbook at the peak of the market 🤣

SafetyDGX agent

Meta acquired Moltbook, an AI-only social network launched in January 2026 by entrepreneurs Matt Schlicht and Ben Parr, after the platform had gone viral and then faded in popularity. Moltbook is ...

MCLR: Improving Conditional Modeling via Inter-Class Likelihood-Ratio Maximization and Unifying Classifier-Free Guidance with Alignment Objectives

SafetyDGX agent

arXiv:2603.22364v2 Announce Type: replace-cross Abstract: Diffusion models have achieved state-of-the-art performance in generative modeling, but their success often relies heavily on classifier-free

MDP modeling for multi-stage stochastic programs

SafetyDGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

Mixture Proportion Estimation and Weakly-supervised Kernel Test for Conditional Independence

SafetyDGX agent

arXiv:2604.07191v1 Announce Type: cross Abstract: Mixture proportion estimation (MPE) aims to estimate class priors from unlabeled data. This task is a critical component in weakly supervised learning

MO-RiskVAE: A Multi-Omics Variational Autoencoder for Survival Risk Modeling in Multiple MyelomaMO-RiskVAE

SafetyDGX agent

arXiv:2604.06267v1 Announce Type: cross Abstract: Multimodal variational autoencoders (VAEs) have emerged as a powerful framework for survival risk modeling in multiple myeloma by integrating heteroge

MonoUNet: A Robust Tiny Neural Network for Automated Knee Cartilage Segmentation on Point-of-Care Ultrasound Devices

SafetyDGX agent

arXiv:2604.07780v1 Announce Type: cross Abstract: Objective: To develop a robust and compact deep learning model for automated knee cartilage segmentation on point-of-care ultrasound (POCUS) devices.

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

SafetyDGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models

SafetyDGX agent

arXiv:2604.07991v1 Announce Type: new Abstract: Recent advances in world models have demonstrated strong capabilities in simulating physical reality, making them an increasingly important foundation f

MSCT: Differential Cross-Modal Attention for Deepfake Detection

SafetyDGX agent

arXiv:2604.07741v1 Announce Type: new Abstract: Audio-visual deepfake detection typically employs a complementary multi-modal model to check the forgery traces in the video. These methods primarily ex

Multi-agent Reach-avoid MDP via Potential Games and Low-rank Policy Structure

SafetyDGX agent

arXiv:2410.17690v2 Announce Type: replace-cross Abstract: We optimize finite horizon multi-agent reach-avoid Markov decision process (MDP) via local feedback policies. The global feedback polic

Multi-Faceted Self-Consistent Preference Alignment for Query Rewriting in Conversational Search

SafetyDGX agent

arXiv:2604.06771v1 Announce Type: cross Abstract: Conversational Query Rewriting (CQR) aims to rewrite ambiguous queries to achieve more efficient conversational search. Early studies have predominant

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

SafetyDGX agent

arXiv:2604.07148v1 Announce Type: new Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) ad

Neural Computers

SafetyDGX agent

arXiv:2604.06425v1 Announce Type: cross Abstract: We propose a new frontier: Neural Computers (NCs) -- an emerging machine form that unifies computation, memory, and I/O in a learned runtime state. Un

On the Global Photometric Alignment for Low-Level Vision

SafetyDGX agent

arXiv:2604.08172v1 Announce Type: new Abstract: Supervised low-level vision models rely on pixel-wise losses against paired references, yet paired training sets exhibit per-pair photometric inconsiste

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

SafetyDGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

Oracle is down more than 50% since two men (jointly) took over Safra Catz’s CEO job. Sexism is up 500%? 5000%?

SafetyDGX agent

Oracle is down more than 50% since two men (jointly) took over Safra Catz’s CEO job. Sexism is up 500%? 5000%? Nick Fuentes says women can only do three things “Women can be three things: they can be

OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance

SafetyDGX agent

arXiv:2604.08461v1 Announce Type: new Abstract: Open-Vocabulary Segmentation (OVS) aims to segment image regions beyond predefined category sets by leveraging semantic descriptions. While CLIP based a

← Previous
1…206207208209210
Next →