AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
1 May 2026

Mapping how LLMs debate societal issues when shadowing human personality traits, sociodemographics and social media behavior

SafetyDGX agent

arXiv:2604.27624v1 Announce Type: cross Abstract: Large Language Models (LLMs) can strongly shape social discourse, yet datasets investigating how LLM outputs vary across controlled social and context

“Marcus’ specific point about coding is structurally important: a model that produces code which compiles and passes the tests it was given …

SafetyDGX agent

“Marcus’ specific point about coding is structurally important: a model that produces code which compiles and passes the tests it was given is not the same as a model that produces correct, secure, ma

Meta is basically Black Mirror incarnate.

SafetyDGX agent

Meta is basically Black Mirror incarnate. This is a confusingly written piece, but the upshot is that Meta's smart glasses record even when you don't want them to, and that Meta's data analysis teams

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

SafetyDGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

MIFair: A Mutual-Information Framework for Intersectionality and Multiclass Fairness

SafetyDGX agent

arXiv:2604.28030v1 Announce Type: cross Abstract: Fairness in machine learning remains challenging due to its ethical complexity, the absence of a universal definition, and the need for context-specif

Mind the Gap: Structure-Aware Consistency in Preference Learning

SafetyDGX agent

arXiv:2604.27733v1 Announce Type: new Abstract: Preference learning has become the foundation of aligning Large Language Models (LLMs) with human intent. Popular methods, such as Direct Preference Opt

Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO

SafetyDGX agent

arXiv:2603.21016v2 Announce Type: replace-cross Abstract: Large language models (LLMs) used for multiple-choice and pairwise evaluation tasks often exhibit selection bias due to non-semantic factors l

MotuBrain: An Advanced World Action Model for Robot Control

SafetyDGX agent

arXiv:2604.27792v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong semantic generalization but often lack fine-grained modeling of world dynamics. Recent work explores

MSR:Hybrid Field Modeling for CT-MRI Rigid-Deformable Registration of the Cervical Spine with an Annotated Dataset

SafetyDGX agent

arXiv:2604.27654v1 Announce Type: new Abstract: Accurate CT-MRI registration of the cervical spine is essential for preoperative planning because this region is anatomically complex,highly variable,an

One of the things I hate the most about this site is the consistent lack of nuance. That’s why everything is an argument, and progress here …

SafetyDGX agent

Gary Marcus critiques social media platforms for lacking nuance in discourse, which he identifies as a root cause of persistent arguments and stalled progress on the site. The post reflects concerns a

Online semi-supervised perception: Real-time learning without explicit feedback

SafetyDGX agent

arXiv:2604.27562v1 Announce Type: new Abstract: This paper proposes an algorithm for real-time learning without explicit feedback. The algorithm combines the ideas of semi-supervised learning on graph

OpAgent: Operator Agent for Web Navigation

SafetyDGX agent

arXiv:2602.13559v2 Announce Type: replace Abstract: To fulfill user instructions, autonomous web agents must contend with the inherent complexity and volatile nature of real-world websites. Convention

Performance-Driven QUBO for Recommender Systems on Quantum Annealers

SafetyDGX agent

arXiv:2410.15272v3 Announce Type: replace-cross Abstract: Quantum annealers offer a promising hardware platform for solving combinatorial optimization problems, especially those formulated as Quadrati

Political Bias Audits of LLMs Capture Sycophancy to the Inferred Auditor

SafetyDGX agent

arXiv:2604.27633v1 Announce Type: new Abstract: Large language models (LLMs) are commonly evaluated for political bias based on their responses to fixed questionnaires, which typically place frontier

Preserving Temporal Dynamics in Time Series Generation

SafetyDGX agent

arXiv:2604.27182v1 Announce Type: cross Abstract: Time-series data augmentation plays a crucial role in regression-oriented forecasting tasks, where limited data restricts the performance of deep lear

Residual Gaussian Splatting for Ultra Sparse-View CBCT Reconstruction

SafetyDGX agent

arXiv:2604.27552v1 Announce Type: new Abstract: While 3D Gaussian splatting (3DGS) offers explicit and efficient scene representations for cone-beam computed tomography reconstruction, conventional ph

RHyVE: Competence-Aware Verification and Phase-Aware Deployment for LLM-Generated Reward Hypotheses

SafetyDGX agent

arXiv:2604.28056v1 Announce Type: new Abstract: Large language models (LLMs) make reward design in reinforcement learning substantially more scalable, but generated rewards are not automatically relia

Robot Learning from Human Videos: A Survey

SafetyDGX agent

arXiv:2604.27621v1 Announce Type: cross Abstract: A critical bottleneck hindering further advancement in embodied AI and robotics is the challenge of scaling robot data. To address this, the field of

Sam Altman is a master at insincerity, Author @GaryMarcus claims. 'He is a master of projecting insincerity, a master at telling the room wh…

SafetyDGX agent

Sam Altman is a master at insincerity, Author @GaryMarcus claims. 'He is a master of projecting insincerity, a master at telling the room what it wants to hear and not always truthful'. “Altman, sitti

Sample-efficient evidence estimation of score based priors for model selection

SafetyDGX agent

arXiv:2602.20549v2 Announce Type: replace-cross Abstract: The choice of prior is central to solving ill-posed imaging inverse problems, making it essential to select one consistent with the measuremen

SCOOP: A pro-AI dark money group backed by a powerful super PAC funded by execs tied to Palantir and OpenAI, has been secretly paying influe…

SafetyDGX agent

SCOOP: A pro-AI dark money group backed by a powerful super PAC funded by execs tied to Palantir and OpenAI, has been secretly paying influencers to push pro-AI, anti-China propaganda on TikTok and IG

Stable Behavior, Limited Variation: Persona Validity in LLM Agents for Urban Sentiment Perception

SafetyDGX agent

arXiv:2604.28048v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used as proxies for human perception in urban analysis, yet it remains unclear whether persona prompting p

Stating that women do not have penises is conservative. Stating that the scientific method is superior to ancestral tribal dances for seekin…

SafetyDGX agent

Stating that women do not have penises is conservative. Stating that the scientific method is superior to ancestral tribal dances for seeking truth is conservative. Supporting a rational immigration p

Supercharging Agenda Setting Research: The ParlaCAP Dataset of 28 European Parliaments and a Scalable Multilingual LLM-Based Classification

SafetyDGX agent

arXiv:2602.16516v2 Announce Type: replace Abstract: This paper introduces ParlaCAP, a large-scale dataset for analyzing parliamentary agenda setting across Europe, and proposes a cost-effective method

Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance

SafetyDGX agent

arXiv:2601.08418v2 Announce Type: replace-cross Abstract: Tax code prediction is a crucial yet underexplored task in automating invoicing and compliance management for large-scale e-commerce platforms

Test-Time Distillation for Continual Model Adaptation

SafetyDGX agent

arXiv:2506.02671v3 Announce Type: replace Abstract: Deep neural networks often suffer performance degradation upon deployment due to distribution shifts. Continual Test-Time Adaptation (CTTA) aims to

The Likelihood Ratio Wall: Structural Limits on Accurate Risk Assessment for Rare Violence

SafetyDGX agent

arXiv:2604.27282v1 Announce Type: cross Abstract: Pretrial risk assessment tools are used on over one million U.S. defendants each year, yet their use for predicting rare violent re-offense faces a ba

The Two Boundaries: Why Behavioral AI Governance Fails Structurally

SafetyDGX agent

arXiv:2604.27292v1 Announce Type: new Abstract: Every system that performs effects has two boundaries: what it can do (expressiveness) and what governance covers (governance). In nearly all deployed A

Tokenmaxxing is stupid. Change my mind?

SafetyDGX agent

Gary Marcus critiques the AI industry's focus on scaling model parameters and training data (tokenmaxxing) as an inefficient approach to advancing AI capabilities. He argues that simply increasing tok

TouchGuide: Inference-Time Steering of Visuomotor Policies via Touch Guidance

SafetyDGX agent

arXiv:2601.20239v4 Announce Type: replace Abstract: Fine-grained and contact-rich manipulation remain challenging for robots, largely due to the underutilization of tactile feedback. To address this,

Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues

SafetyDGX agent

arXiv:2506.05412v3 Announce Type: replace-cross Abstract: Where someone looks is a nonverbal communication cue that children and adults readily use. How well can Vision-Language Models (VLMs) infer ga

When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry

SafetyDGX agent

arXiv:2604.27656v1 Announce Type: cross Abstract: To preserve previously learned representations, continual learning systems must strike a balance between plasticity, the ability to acquire new knowle

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't G…

SafetyDGX agent

who's ready to dance? 😆🕺🎶 uploaded 11 new favorite @suno songs to @spotify [search: 'captain yohei'] 💿 Captain Yohei Vol 1. 💿 ♫ Don't Go in the Office [hard rock] ♫ Nine-Nine-Six [hip-hop/rap] ♫ Fog F

30 Apr 2026

A Multimodal Pre-trained Network for Integrated EEG-Video Seizure Detection

SafetyDGX agent

arXiv:2604.26379v1 Announce Type: new Abstract: Reliable seizure detection in mouse models is essential for preclinical epilepsy research, yet manual review of synchronized video-EEG recordings is lab

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

Accelerating RL Post-Training Rollouts via System-Integrated Speculative Decoding

SafetyDGX agent

arXiv:2604.26779v1 Announce Type: cross Abstract: RL post-training of frontier language models is increasingly bottlenecked by autoregressive rollout generation, making rollout acceleration a central

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alig…

SafetyDGX agent

alignment in 2016: obviously any real AI will be made inside a faraday cage magnetically suspended in a 10×10×10 cube of telekill alloy alignment in 2026: yeah we can not make it stop talking about go

ATLAS: An Annotation Tool for Long-horizon Robotic Action Segmentation

SafetyDGX agent

arXiv:2604.26637v1 Announce Type: cross Abstract: Annotating long-horizon robotic demonstrations with precise temporal action boundaries is crucial for training and evaluating action segmentation and

Atomic-Probe Governance for Skill Updates in Compositional Robot Policies

SafetyDGX agent

arXiv:2604.26689v1 Announce Type: cross Abstract: Skill libraries in deployed robotic systems are continually updated through fine-tuning, fresh demonstrations, or domain adaptation, yet existing type

Attribution-Guided Multimodal Deepfake Detection via Cross-Modal Forensic Fingerprints

SafetyDGX agent

arXiv:2604.26453v1 Announce Type: new Abstract: Audio-visual deepfakes have reached a level of realism that makes perceptual detection unreliable, threatening media integrity and biometric security. W

Beyond Shortcuts: Mitigating Visual Illusions in Frozen VLMs via Qualitative Reasoning

SafetyDGX agent

arXiv:2604.26250v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved state-of-the-art performance in general visual tasks, their perceptual robustness remains remarkably b

Big Tech’s $700 billion spending on AI this year is called the ‘greatest capital misallocation in history’ https://trib.al/qUQIbrJ

SafetyDGX agent

Big Tech companies are projected to spend approximately $700 billion on AI infrastructure and development in the current year, a figure that AI researcher and entrepreneur Gary Marcus has criticized a

Classification of Public Opinion on the Free Nutritional Meal Program on YouTube Media Using the LSTM Method

SafetyDGX agent

arXiv:2604.26312v1 Announce Type: new Abstract: Public opinion towards the Free Nutritious Meal Program (MBG) on YouTube social media reflects diverse community responses. This study applies the Long

Co-Learning Port-Hamiltonian Systems and Optimal Energy-Shaping Control

SafetyDGX agent

arXiv:2604.26172v1 Announce Type: cross Abstract: We develop a physics-informed learning framework for energy-shaping control of port-Hamiltonian (pH) systems from trajectory data. The proposed approa

Consist-Retinex: One-Step Noise-Emphasized Consistency Training Accelerates High-Quality Retinex Enhancement

SafetyDGX agent

arXiv:2512.08982v2 Announce Type: replace-cross Abstract: Retinex-based low-light image enhancement benefits from separating reflectance and illumination, yet recent generative approaches often rely o

Correcting Performance Estimation Bias in Imbalanced Classification with Minority Subconcepts

SafetyDGX agent

arXiv:2604.26024v1 Announce Type: cross Abstract: Class-level evaluation can conceal substantial performance disparities across subconcepts within the same class, causing models that perform well on a

Data-Centric Foundation Models in Computational Healthcare: A Survey

SafetyDGX agent

arXiv:2401.02458v3 Announce Type: replace-cross Abstract: The advent of foundation models (FMs) as an emerging suite of AI techniques has struck a wave of opportunities in computational healthcare. Th

DC-Ada: Reward-Only Decentralized Sensor Adaptation for Heterogeneous Multi-Robot Teams

SafetyDGX agent

arXiv:2604.03905v2 Announce Type: replace-cross Abstract: Heterogeneity is a defining feature of deployed multi-robot teams: platforms often differ in sensing modalities, ranges, fields of view, and f

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

SafetyDGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training

SafetyDGX agent

arXiv:2604.26256v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time --

Efficient and Interpretable Transformer for Counterfactual Fairness

SafetyDGX agent

arXiv:2604.26188v1 Announce Type: new Abstract: The growing reliance of machine learning models in high-stakes, highly regulated domains such as finance and insurance has created a growing tension bet

Evaluating Strategic Reasoning in Forecasting Agents

SafetyDGX agent

arXiv:2604.26106v1 Announce Type: new Abstract: Forecasting benchmarks produce accuracy leaderboards but little insight into why some forecasters are more accurate than others. We introduce Bench to t

Evaluating the Alignment Between GeoAI Explanations and Domain Knowledge in Satellite-Based Flood Mapping

SafetyDGX agent

arXiv:2604.26051v1 Announce Type: cross Abstract: The increasing number of satellites has improved the temporal resolution of Earth observation, making satellite-based flood mapping a promising approa

EvoSelect: Data-Efficient LLM Evolution for Targeted Task Adaptation

SafetyDGX agent

arXiv:2604.26170v1 Announce Type: new Abstract: Adapting large language models (LLMs) to a targeted task efficiently and effectively remains a fundamental challenge. Such adaptation often requires ite

FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility

SafetyDGX agent

arXiv:2510.26841v2 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training without data sharing, yet participants face a fundamental challenge, e.g., simult

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding

SafetyDGX agent

arXiv:2504.09925v3 Announce Type: replace Abstract: We introduce FLARE, a family of vision language models (VLMs) with a fully vision-language alignment and integration paradigm. Unlike existing appro

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how y…

SafetyDGX agent

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how you regulate open-source models that can be served by a range

Fundamental Physics, Existential Risks and Human Futures

SafetyDGX agent

arXiv:2604.26530v1 Announce Type: cross Abstract: Over the past 25 years, I have been involved in some intriguing developments in the foundations of physics, exploring the quantum reality problem, the

Generative Bid Shading in Real-Time Bidding Advertising

SafetyDGX agent

arXiv:2508.06550v3 Announce Type: replace-cross Abstract: Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainst

Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning

SafetyDGX agent

arXiv:2601.13942v2 Announce Type: replace-cross Abstract: Large Multimodal Models (LMMs) have achieved remarkable success in visual understanding, yet they struggle with knowledge-intensive queries in

← Previous
1…190191192193194…240
Next →