AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
29 Apr 2026

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

SafetyDGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

True or False: OpenAI will eventually become a massively profitable company, more than earning out all the money that went into it.

SafetyDGX agent

Gary Marcus poses a question about whether OpenAI will achieve sufficient profitability to justify the substantial capital investments made into the company. This reflects broader industry speculation

Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.25776v1 Announce Type: new Abstract: Critical analyses of emotion recognition technology have raised ethical concerns around task validity and potential downstream impacts, urging researche

Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation

SafetyDGX agent

arXiv:2511.21517v2 Announce Type: replace Abstract: Unlike text, speech conveys information about the speaker, such as gender, through acoustic cues like pitch. This gives rise to modality-specific bi

We at @ControlAI are sometimes asked what we think of banning datacenter construction. At CAI, we focus on one issue: The risk of extinction…

SafetyDGX agent

We at @ControlAI are sometimes asked what we think of banning datacenter construction. At CAI, we focus on one issue: The risk of extinction from superintelligent AI. The only way to prevent this is t

“we might as well stop training radiologists” Geoff Hinton, 2016, vs the actual data, via Torsten Slok at Apollo

SafetyDGX agent

Gary Marcus shares Torsten Slok's analysis comparing Geoffrey Hinton's 2016 prediction that radiologist training should cease due to AI capabilities against actual empirical data on AI performance in

When Errors Can Be Beneficial: A Categorization of Imperfect Rewards for Policy Gradient

SafetyDGX agent

arXiv:2604.25872v1 Announce Type: new Abstract: Training language models via reinforcement learning often relies on imperfect proxy rewards, since ground truth rewards that precisely define the intend

Zuckerberg bets big on biology? About as much he put into 5 employees in his AI startup for a year 🙄 He is *far* more interested AI for sel…

SafetyDGX agent

Gary Marcus critiques Mark Zuckerberg's claimed focus on biology research, suggesting the financial commitment is minimal compared to his AI investments and stating that Zuckerberg's actual priorities

28 Apr 2026

A BERTology View of LLM Orchestrations: Token- and Layer-Selective Probes for Efficient Single-Pass Classification

Model ReleasesDGX agent

arXiv:2601.13288v2 Announce Type: replace Abstract: Production LLM systems often rely on separate models for safety and other classification-heavy steps, increasing latency, VRAM footprint, and operat

A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs

SafetyDGX agent

arXiv:2603.07475v2 Announce Type: replace Abstract: Autoregressive (AR) language models build representations incrementally via left-to-right prediction, while diffusion language models (dLLMs) are tr

A Differentiable Framework for Global Circulation Model Precipitation Bias Correction

SafetyDGX agent

arXiv:2604.23045v1 Announce Type: new Abstract: Systematic biases in Global Circulation Model (GCM) outputs limit their direct applicability in regional planning, necessitating bias correction. Correc

A Multi-Dimensional Audit of Politically Aligned Large Language Models

SafetyDGX agent

arXiv:2604.24429v1 Announce Type: new Abstract: As the application of Large Language Models (LLMs) spreads across various industries, there are increasing concerns about the potential for their misuse

A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning

SafetyDGX agent

arXiv:2604.24532v1 Announce Type: new Abstract: Many sequential decision-making tasks involve optimizing multiple conflicting objectives, requiring policies that adapt to different user preferences. I

A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning

SafetyDGX agent

arXiv:2604.23386v1 Announce Type: cross Abstract: Federated Learning (FL) typically assumes unconditional collaboration, a premise that overlooks the complexities of real-world, multi-stakeholder envi

AdaRubric: Task-Adaptive Rubrics for LLM Agent Evaluation

SafetyDGX agent

arXiv:2603.21362v2 Announce Type: replace Abstract: LLM-as-Judge evaluation fails agent tasks because a fixed rubric cannot capture what matters for this task: code debugging demands Correctness and E

Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation

SafetyDGX agent

arXiv:2602.14914v2 Announce Type: replace Abstract: Off-policy evaluation (OPE) is essential for assessing ranking and recommendation systems without costly online interventions. Self-Normalised Inver

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

SafetyDGX agent

arXiv:2510.15479v2 Announce Type: replace Abstract: We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatmen

Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI

SafetyDGX agent

arXiv:2604.22765v1 Announce Type: cross Abstract: The increasing use of artificial intelligence (AI) by public authorities introduces both opportunities for innovation and significant challenges for t

Aligning with Your Own Voice: Self-Corrected Preference Learning for Hallucination Mitigation in LVLMs

SafetyDGX agent

arXiv:2604.24395v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) frequently suffer from hallucinations. Existing preference learning-based approaches largely rely on proprietary mo

ANCHOR: LLM-driven Subject Conditioning for Text-to-Image Synthesis

SafetyDGX agent

arXiv:2404.10141v2 Announce Type: replace-cross Abstract: Text-to-image (T2I) models have achieved remarkable progress in high-quality image synthesis, yet most benchmarks rely on simple, self-contain

AnemiaVision: Non-Invasive Anemia Detection via Smartphone Imagery Using EfficientNet-B3 with TrivialAugmentWide, Mixup Augmentation, and Persistent Patient History Management

SafetyDGX agent

arXiv:2604.22964v1 Announce Type: new Abstract: Anemia affects over one billion people globally and remains severely under-diagnosed in low-resource regions where laboratory blood tests are inaccessib

Animalbooth: multimodal feature enhancement for animal subject personalization

SafetyDGX agent

arXiv:2509.16702v2 Announce Type: replace Abstract: Personalized animal image generation is challenging due to rich appearance cues and large morphological variability. Existing approaches often exhib

As an avid cyclist, I was amused to see ChatGPT's “powerful new image engine' draw a bicycle with the 'brake' label pointing to empty space …

SafetyDGX agent

As an avid cyclist, I was amused to see ChatGPT's “powerful new image engine' draw a bicycle with the 'brake' label pointing to empty space where brakes are sometimes found on other bicycles. The poin

Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence

SafetyDGX agent

arXiv:2601.18840v3 Announce Type: replace Abstract: Markov decision problems are most commonly solved via dynamic programming. Another approach is Bellman residual minimization, which directly minimiz

Beyond Match Maximization and Fairness: Retention-Optimized Two-Sided Matching

SafetyDGX agent

arXiv:2602.15752v2 Announce Type: replace Abstract: On two-sided matching platforms such as online dating and recruiting, recommendation algorithms often aim to maximize the total number of matches. H

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training

SafetyDGX agent

arXiv:2604.23121v1 Announce Type: cross Abstract: Have you ever post-trained a generalist vision-language-action (VLA) policy on a small demonstration dataset, only to find that it stops responding to

Bridging Reasoning and Action: Hybrid LLM-RL Framework for Efficient Cross-Domain Task-Oriented Dialogue

SafetyDGX agent

arXiv:2604.23345v1 Announce Type: new Abstract: Cross-domain task-oriented dialogue requires reasoning over implicit and explicit feasibility constraints while planning long-horizon, multi-turn action

BVI-Mamba: Video Enhancement Using a Visual State-Space Model for Low-Light and Underwater Environments

SafetyDGX agent

arXiv:2604.23655v1 Announce Type: new Abstract: Videos captured in low-light and underwater conditions often suffer from distortions such as noise, low contrast, color imbalance, and blur. These issue

CA-IDD: Cross-Attention Guided Identity-Conditional Diffusion for Identity-Consistent Face Swapping

SafetyDGX agent

arXiv:2604.24493v1 Announce Type: new Abstract: Face swapping aims to optimize realistic facial image generation by leveraging the identity of a source face onto a target face while preserving pose, e

Can Compact Language Models Search Like Agents? Distillation-Guided Policy Optimization for Preserving Agentic RAG Capabilities

SafetyDGX agent

arXiv:2508.20324v4 Announce Type: replace Abstract: Reinforcement Learning has emerged as a dominant post-training approach to elicit agentic RAG behaviors such as search and planning from language mo

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems

SafetyDGX agent

arXiv:2604.23022v1 Announce Type: cross Abstract: Two-stage recommender systems first choose a candidate generator and then rank items within the generated set. Because the generator decides which ite

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning

SafetyDGX agent

arXiv:2604.23308v1 Announce Type: new Abstract: Offline multi-agent reinforcement learning (MARL) enables policy learning from fixed datasets, but is prone to coordination failure: agents trained on s

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

SafetyDGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

COMO: Closed-Loop Optical Molecule Recognition with Minimum Risk Training

SafetyDGX agent

arXiv:2604.23546v1 Announce Type: cross Abstract: Optical chemical structure recognition (OCSR) translates molecular images into machine-readable representations like SMILES strings or molecular graph

Complex SGD and Directional Bias in Reproducing Kernel Hilbert Spaces

SafetyDGX agent

arXiv:2604.23017v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is a known stochastic iterative method popular for large-scale convex optimization problems due to its simple implemen

Conditional Imputation for Within-Modality Missingness in Multi-Modal Federated Learning

SafetyDGX agent

arXiv:2604.23112v1 Announce Type: new Abstract: Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative training, but real-world clinical applications often suffer from within-mo

Conflict-Aware Harmonized Rotational Gradient for Multiscale Kinetic Regimes

SafetyDGX agent

arXiv:2604.24745v1 Announce Type: new Abstract: In this paper, we propose a harmonized rotational gradient method, termed HRGrad, for simultaneously tackling multiscale time-dependent kinetic problems

ConsDreamer: Advancing Multi-View Consistency for Zero-Shot Text-to-3D Generation

SafetyDGX agent

arXiv:2504.02316v4 Announce Type: replace-cross Abstract: Recent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descripti

CT-Guided Spatially-varying Regularization for Voxel-Wise Deformable Whole-Body PET Registration

SafetyDGX agent

arXiv:2604.22905v1 Announce Type: cross Abstract: Whole-body Positron Emission Tomography (PET) registration is essential for multi-parametric tumor characterization and assessment of metastatic disea

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning

SafetyDGX agent

arXiv:2601.13262v2 Announce Type: replace Abstract: While large language models (LLMs) have shown to perform well on monolingual mathematical and commonsense reasoning, they remain unreliable for mult

Data-efficient Targeted Token-level Preference Optimization for LLM-based Text-to-Speech

SafetyDGX agent

arXiv:2510.05799v2 Announce Type: replace-cross Abstract: Aligning text-to-speech (TTS) system outputs with human feedback through preference optimization has been shown to effectively improve the rob

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

SafetyDGX agent

arXiv:2601.16046v2 Announce Type: replace-cross Abstract: Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions

DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making

SafetyDGX agent

arXiv:2604.23557v1 Announce Type: cross Abstract: Building scalable and reusable multi-agent decision policies from offline datasets remains a challenge in offline multi-agent reinforcement learning (

Do Synthetic Trajectories Reflect Real Reward Hacking? A Systematic Study on Monitoring In-the-Wild Hacking in Code Generation

SafetyDGX agent

arXiv:2604.23488v1 Announce Type: new Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain full reward without correctly solving the tasks, poses a critical

Do Transaction-Level and Actor-Level AML Queues Agree? An Empirical Evaluation of Granularity Effects on the Elliptic++ Graph

SafetyDGX agent

arXiv:2604.23494v1 Announce Type: new Abstract: Graph-based anti-money laundering (AML) systems on blockchain networks can score suspicious activity at two granularity levels -- transactions or actor

DPEPO: Diverse Parallel Exploration Policy Optimization for LLM-based Agents

SafetyDGX agent

arXiv:2604.24320v1 Announce Type: new Abstract: Large language model (LLM) agents that follow the sequential 'reason-then-act' paradigm have achieved superior performance in many complex tasks.However

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

SafetyDGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

SafetyDGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence

SafetyDGX agent

arXiv:2604.23325v1 Announce Type: cross Abstract: Emotionally talking head video generation aims to generate expressive portrait videos with accurate lip synchronization and emotional facial expressio

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

SafetyDGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

SafetyDGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries (Foo Yun Chee/Reuters)

SafetyDGX agent

Foo Yun Chee / Reuters: EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries — EU countries and E

Evaluating Language Models' Evaluations of Games

SafetyDGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

Explanation Quality Assessment as Ranking with Listwise Rewards

SafetyDGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

Extreme bandits

SafetyDGX agent

arXiv:2604.24545v1 Announce Type: cross Abstract: In many areas of medicine, security, and life sciences, we want to allocate limited resources to different sources in order to detect extreme values.

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

SafetyDGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

Federated Cross-Modal Retrieval with Missing Modalities via Semantic Routing and Adapter Personalization

SafetyDGX agent

arXiv:2604.22885v1 Announce Type: cross Abstract: Federated cross-modal retrieval faces severe challenges from heterogeneous client data, particularly non-IID semantic distributions and missing modali

Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

SafetyDGX agent

arXiv:2602.07605v3 Announce Type: replace-cross Abstract: Any entity in the visual world can be hierarchically grouped based on shared characteristics and mapped to fine-grained sub-categories. While

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

SafetyDGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

From what I can tell, Elon Musk just stepped on his own toes at the trial, making it all about him instead of the promises Altman and Brockm…

SafetyDGX agent

From what I can tell, Elon Musk just stepped on his own toes at the trial, making it all about him instead of the promises Altman and Brockman broke. OpenAI will slay his ego on cross. He should have

← Previous
1…193194195196197…240
Next →