AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
20 Apr 2026

GroupDPO: Memory efficient Group-wise Direct Preference Optimization

SafetyDGX agent

arXiv:2604.15602v1 Announce Type: new Abstract: Preference optimization is widely used to align Large Language Models (LLMs) with preference feedback. However, most existing methods train on a single

Hierarchical Codec Diffusion for Video-to-Speech Generation

SafetyDGX agent

arXiv:2604.15923v1 Announce Type: cross Abstract: Video-to-Speech (VTS) generation aims to synthesize speech from a silent video without auditory signals. However, existing VTS methods disregard the h

Importantly, the tweet by @itsolelehmann may be misleading. I stand by what I said but have some reason to think that @AmandaAskell’s views …

SafetyDGX agent

Importantly, the tweet by @itsolelehmann may be misleading. I stand by what I said but have some reason to think that @AmandaAskell’s views may have been misrepresented. Seeking clarification and will

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

SafetyDGX agent

arXiv:2604.15701v1 Announce Type: new Abstract: The significant computational demands of large language models have increased interest in distilling reasoning abilities into smaller models via Chain-o

Information-Consistent Language Model Recommendations through Group Relative Policy Optimization

SafetyDGX agent

arXiv:2512.12858v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in business-critical domains such as finance, education, healthcare, and customer suppo

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des compo…

SafetyDGX agent

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des comportements inquiétants des modèles de pointe, et des solutions

Joint-Centric Dual Contrastive Alignment with Structure-Preserving and Information-Balanced Regularization

SafetyDGX agent

arXiv:2604.16247v1 Announce Type: cross Abstract: We propose HILBERT (HIerarchical Long-sequence Balanced Embedding with Reciprocal contrastive Training), a cross-attentive multimodal framework for le

Language, Place, and Social Media: Geographic Dialect Alignment in New Zealand

SafetyDGX agent

arXiv:2604.15744v1 Announce Type: new Abstract: This thesis investigates geographic dialect alignment in place-informed social media communities, focussing on New Zealand-related Reddit communities. B

Large Language Models for Market Research: A Data-augmentation Approach

SafetyDGX agent

arXiv:2412.19363v3 Announce Type: replace Abstract: Large Language Models (LLMs) have transformed artificial intelligence by excelling in complex natural language processing tasks. Their ability to ge

Learning to Look before Learning to Like: Incorporating Human Visual Cognition into Aesthetic Quality Assessment

SafetyDGX agent

arXiv:2604.15853v1 Announce Type: new Abstract: Automated Aesthetic Quality Assessment (AQA) treats images primarily as static pixel vectors, aligning predictions with human-rating scores largely thro

literally my basic model since 1998. crazy that some people still haven’t figured this out.

SafetyDGX agent

literally my basic model since 1998. crazy that some people still haven’t figured this out. My basic model of capabilities: LLMs are good at problems similar to those that appear in their training dat

LLMs are “ good at math style problems, where you tell them A, B and C are true, and then ask them to figure out D [but] extremely bad at an…

SafetyDGX agent

LLMs are “ good at math style problems, where you tell them A, B and C are true, and then ask them to figure out D [but] extremely bad at anything involving what I would call mature scholarship .. [so

M3R: Localized Rainfall Nowcasting with Meteorology-Informed MultiModal Attention

SafetyDGX agent

arXiv:2604.15377v1 Announce Type: cross Abstract: Accurate and timely rainfall nowcasting is crucial for disaster mitigation and water resource management. Despite recent advances in deep learning, pr

MFC-RFNet: A Multi-scale Guided Rectified Flow Network for Radar Sequence Prediction

SafetyDGX agent

arXiv:2601.03633v2 Announce Type: replace-cross Abstract: Accurate and high-resolution precipitation nowcasting from radar echo sequences is crucial for disaster mitigation and economic planning, yet

Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment

SafetyDGX agent

arXiv:2604.15757v1 Announce Type: new Abstract: This research note identifies a previously overlooked distinction between multi-objective reinforcement learning (MORL), and more conventional single-ob

On theCUBE Pod: IBM’s AI strategy, infrastructure bottlenecks and ecosystem partnerships reshape markets

SafetyDGX agent

Artificial intelligence infrastructure is now the deciding force behind enterprise competitiveness. What was once a backend concern has moved directly into the center of business strategy, shaping how

One-Shot Cross-Geometry Skill Transfer through Part Decomposition

SafetyDGX agent

arXiv:2604.15455v1 Announce Type: new Abstract: Given a demonstration, a robot should be able to generalize a skill to any object it encounters-but existing approaches to skill transfer often fail to

OpenAI’s market share is slipping.

SafetyDGX agent

OpenAI’s market share is slipping. ChatGPT is dying. OpenAI lost the crown as their market share fell below 40 percent for the first time ever. It stands at just 38.7 percent now after four straight m

Order: India's CCI sets May hearing on penalties after Apple failed to submit info on alleged app market abuse; Apple said it fears it could be fined up to $38B (Aditya Kalra/Reuters)

SafetyDGX agent

Aditya Kalra / Reuters: Order: India's CCI sets May hearing on penalties after Apple failed to submit info on alleged app market abuse; Apple said it fears it could be fined up to $38B — Apple has not

PAWN: Piece Value Analysis with Neural Networks

SafetyDGX agent

arXiv:2604.15585v1 Announce Type: cross Abstract: Predicting the relative value of any given chess piece in a position remains an open challenge, as a piece's contribution depends on its spatial relat

PLAF: Pixel-wise Language-Aligned Feature Extraction for Efficient 3D Scene Understanding

SafetyDGX agent

arXiv:2604.15770v1 Announce Type: new Abstract: Accurate open-vocabulary 3D scene understanding requires semantic representations that are both language-aligned and spatially precise at the pixel leve

Preregistered Belief Revision Contracts

SafetyDGX agent

arXiv:2604.15558v1 Announce Type: new Abstract: Deliberative multi-agent systems allow agents to exchange messages and revise beliefs over time. While this interaction is meant to improve performance,

Prototype-Grounded Concept Models for Verifiable Concept Alignment

SafetyDGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability

SafetyDGX agent

arXiv:2604.16106v1 Announce Type: cross Abstract: The Project of AI is a world-building endeavor, wherein those who fund and develop AI systems both operate through and seek to sustain networks of pow

Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2510.10959v3 Announce Type: replace-cross Abstract: Reasoning ability has become a defining capability of Large Language Models (LLMs), with Reinforcement Learning with Verifiable Rewards (RLVR)

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

SafetyDGX agent

arXiv:2604.15577v1 Announce Type: cross Abstract: Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vect

Robust Multispectral Semantic Segmentation under Missing or Full Modalities via Structured Latent Projection

SafetyDGX agent

arXiv:2604.15856v1 Announce Type: cross Abstract: Multimodal remote sensing data provide complementary information for semantic segmentation, but in real-world deployments, some modalities may be unav

Robust Synchronisation for Federated Learning in The Face of Correlated Device Failure

SafetyDGX agent

arXiv:2604.16090v1 Announce Type: cross Abstract: Probabilistic Synchronous Parallel (PSP) is a technique in distributed learning systems to reduce synchronization bottlenecks by sampling a subset of

Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model

SafetyDGX agent

arXiv:2604.16111v1 Announce Type: new Abstract: We study the sample complexity of learning an epsilon-optimal policy in the Stochastic Shortest Path (SSP) problem. We first derive sample complexity bo

Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agents

SafetyDGX agent

arXiv:2504.13541v5 Announce Type: replace-cross Abstract: Training resource-constrained autonomous agents on multiple tasks simultaneously is crucial for adapting to diverse real-world environments. R

Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration

SafetyDGX agent

arXiv:2503.05578v4 Announce Type: replace Abstract: Pose estimation-guided unseen object 6-DoF robotic manipulation is a key task in robotics. However, the scalability of current pose estimation metho

Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting

SafetyDGX agent

arXiv:2604.15794v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degra

Seriously. If you don’t understand the below, you really shouldn’t speculate about AI. It’s absolutely fundamental.

SafetyDGX agent

Seriously. If you don’t understand the below, you really shouldn’t speculate about AI. It’s absolutely fundamental. literally my basic model since 1998. crazy that some people still haven’t figured th

SIMMER: Cross-Modal Food Image--Recipe Retrieval via MLLM-Based Embedding

SafetyDGX agent

arXiv:2604.15628v1 Announce Type: cross Abstract: Cross-modal retrieval between food images and recipe texts is an important task with applications in nutritional management, dietary logging, and cook

Skill-RAG: Failure-State-Aware Retrieval Augmentation via Hidden-State Probing and Skill Routing

SafetyDGX agent

arXiv:2604.15771v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a foundational paradigm for grounding large language models in external knowledge. While adaptive re

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

SafetyDGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

'Taking Stock at FAccT': Using Participatory Design to Co-Create a Vision for the Fairness, Accountability and Transparency Community

SafetyDGX agent

arXiv:2604.16224v1 Announce Type: cross Abstract: As a relatively new forum, ACM FAccT has become a key space for activists and scholars to critically examine emerging AI and ML technologies. It bring

Targeted Exploration via Unified Entropy Control for Reinforcement Learning

SafetyDGX agent

arXiv:2604.14646v2 Announce Type: replace Abstract: Recent advances in reinforcement learning (RL) have improved the reasoning capabilities of large language models (LLMs) and vision-language models (

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models

SafetyDGX agent

arXiv:2604.15383v1 Announce Type: cross Abstract: Large audio-language models (LALMs) generalize across speech, sound, and music, but unified decoders can exhibit a temporal smoothing bias: transient

The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback

SafetyDGX agent

arXiv:2604.16087v1 Announce Type: new Abstract: We study the problem of learning in zero-sum matrix games with repeated play and bandit feedback. Specifically, we focus on developing uncoupled algorit

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning

SafetyDGX agent

arXiv:2603.01283v2 Announce Type: replace Abstract: Deployed RL agents operate in closed-loop systems where reliable performance depends on maintaining coherent coupling between observations, actions,

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

Truncated Kernel Stochastic Gradient Descent with General Losses and Spherical Radial Basis Functions

SafetyDGX agent

arXiv:2510.04237v5 Announce Type: replace Abstract: In this paper, we propose a novel kernel stochastic gradient descent (SGD) algorithm for large-scale supervised learning with general losses. Compar

Unsupervised domain adaptation for radioisotope identification in gamma spectroscopy

SafetyDGX agent

arXiv:2603.05719v2 Announce Type: replace Abstract: Training machine learning models for radioisotope identification using gamma spectroscopy remains an elusive challenge for many practical applicatio

UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval

SafetyDGX agent

arXiv:2604.15827v1 Announce Type: cross Abstract: Conventional information retrieval is concerned with identifying the relevance of texts for a given query. Yet, the conventional definition of relevan

VADF: Vision-Adaptive Diffusion Policy Framework for Efficient Robotic Manipulation

SafetyDGX agent

arXiv:2604.15938v1 Announce Type: new Abstract: Diffusion policies are becoming mainstream in robotic manipulation but suffer from hard negative class imbalance due to uniform sampling and lack of sam

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

SafetyDGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

Whose Facts Win? LLM Source Preferences under Knowledge Conflicts

SafetyDGX agent

arXiv:2601.03746v3 Announce Type: replace Abstract: As large language models (LLMs) are more frequently used in retrieval-augmented generation pipelines, it is increasingly relevant to study their beh

Why Colors Make Clustering Harder:Global Integrality Gaps, the Price of Fairness, and Color-Coupled Algorithms in Chromatic Correlation Clustering

SafetyDGX agent

arXiv:2604.15738v1 Announce Type: new Abstract: Chromatic Correlation Clustering (CCC) extends Correlation Clustering by assigning semantic colors to edges and requiring each cluster to receive a sing

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

SafetyDGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

19 Apr 2026

good to know that a trillion dollars’ investment scaling has thoroughly solved the challenges of common sense.

SafetyDGX agent

Gary Marcus critiques the assumption that massive financial investment in AI scaling alone can solve fundamental challenges related to common sense reasoning in artificial intelligence systems. The po

18 Apr 2026

easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P]

SafetyDGX agent

easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization

EVERYONE who heard about @SeismicOrg (Seismic Foundation)'s report that AI 'salience' is low NEEDS to hear about this! New research from dat…

SafetyDGX agent

EVERYONE who heard about @SeismicOrg (Seismic Foundation)'s report that AI 'salience' is low NEEDS to hear about this! New research from data science wiz @davidshor shows AI is now ahead of ABORTION a

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants …

SafetyDGX agent

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants to be the company that didn't invest in AI when everyone els

17 Apr 2026

3D Instruction Ambiguity Detection

Model ReleasesDGX agent

arXiv:2601.05991v2 Announce Type: replace Abstract: In safety-critical domains, linguistic ambiguity can have severe consequences; a vague command like 'Pass me the vial' in a surgical setting could l

A Mechanistic Account of Attention Sinks in GPT-2: One Circuit, Broader Implications for Mitigation

SafetyDGX agent

arXiv:2604.14722v1 Announce Type: new Abstract: Transformers commonly exhibit an attention sink: disproportionately high attention to the first position. We study this behavior in GPT-2-style models w

Abstract Sim2Real through Approximate Information States

SafetyDGX agent

arXiv:2604.15289v1 Announce Type: new Abstract: In recent years, reinforcement learning (RL) has shown remarkable success in robotics when a fast and accurate simulator is available for a given task.

AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation

SafetyDGX agent

arXiv:2510.01433v2 Announce Type: replace Abstract: Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background fe

Beyond Importance Sampling: Rejection-Gated Policy Optimization

SafetyDGX agent

arXiv:2604.14895v1 Announce Type: new Abstract: We propose a new perspective on policy optimization: rather than reweighting all samples by their importance ratios, an optimizer should select which sa

Bias in Surface Electromyography Features across a Demographically Diverse Cohort

SafetyDGX agent

arXiv:2604.14460v1 Announce Type: cross Abstract: Neuromotor decoding from upper-limb electromyography (sEMG) can enhance human-machine interfaces and offer a more natural means of controlling prosthe

← Previous
1…205206207208209…240
Next →