AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
21 Apr 2026

Emergency Stopping for Liquid-manipulating Robots

SafetyDGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

SafetyDGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

SafetyDGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models

SafetyDGX agent

arXiv:2604.16481v1 Announce Type: new Abstract: Large-scale text-to-image (T2I) diffusion models deliver remarkable visual fidelity but pose safety risks due to their capacity to reproduce undesirable

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

SafetyDGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning

SafetyDGX agent

arXiv:2601.10306v2 Announce Type: replace-cross Abstract: While Reinforcement Learning (RL) has advanced LLM reasoning, applying it to long-context scenarios is hindered by sparsity of outcome rewards

Evolutionary Negative Module Pruning for Better LoRA Merging

SafetyDGX agent

arXiv:2604.17753v1 Announce Type: cross Abstract: Merging multiple Low-Rank Adaptation (LoRA) experts into a single backbone is a promising approach for efficient multi-task deployment. While existing

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

SafetyDGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

FairLogue: Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using the All of Us Research Program

SafetyDGX agent

arXiv:2604.16450v1 Announce Type: cross Abstract: Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess de

Fairness Constraints in High-Dimensional Generalized Linear Models

SafetyDGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

FairNVT: Improving Fairness via Noise Injection in Vision Transformers

SafetyDGX agent

arXiv:2604.16780v1 Announce Type: new Abstract: This paper presents FairNVT, a lightweight debiasing framework for pretrained transformer-based encoders that improves both representation and predictio

'Faithful to What?' On the Limits of Fidelity-Based Explanations

SafetyDGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

SafetyDGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

Fisher Decorator: Refining Flow Policy via A Local Transport Map

SafetyDGX agent

arXiv:2604.17919v1 Announce Type: new Abstract: Recent advances in flow-based offline reinforcement learning (RL) have achieved strong performance by parameterizing policies via flow matching. However

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

SafetyDGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

Foundation Model for Cardiac Time Series via Masked Latent Attention

SafetyDGX agent

arXiv:2603.26475v2 Announce Type: replace Abstract: Electrocardiograms (ECGs) are among the most widely available clinical signals and play a central role in cardiovascular diagnosis. While recent fou

Function Words as Statistical Cues for Language Learning

SafetyDGX agent

arXiv:2601.21191v2 Announce Type: replace Abstract: What statistical properties might support learning abstract grammatical knowledge from linear input? We address this question by examining the stati

Generative Semantic Communication via Alternating Dual-Domain Posterior Sampling

SafetyDGX agent

arXiv:2604.16796v1 Announce Type: new Abstract: Generative semantic communication (SemCom) harnesses pretrained generative priors to improve the perceptual quality of wireless image transmission. Exis

Geometric Stability: The Missing Axis of Representations

SafetyDGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

SafetyDGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

GS-STVSR: Ultra-Efficient Continuous Spatio-Temporal Video Super-Resolution via 2D Gaussian Splatting

SafetyDGX agent

arXiv:2604.18047v1 Announce Type: new Abstract: Continuous Spatio-Temporal Video Super-Resolution (C-STVSR) aims to simultaneously enhance the spatial resolution and frame rate of videos by arbitrary

HAVEN: Hierarchical Adversary-aware Visibility-Enabled Navigation with Cover Utilization using Deep Transformer Q-Networks

SafetyDGX agent

arXiv:2512.00592v2 Announce Type: replace Abstract: Autonomous navigation in partially observable environments requires agents to reason beyond immediate sensor input, exploit occlusion, and ensure sa

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment

SafetyDGX agent

arXiv:2604.17928v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has proven effective for training reasoning-oriented large language models, but existing methods la

Heterogeneous Self-Play for Realistic Highway Traffic Simulation

SafetyDGX agent

arXiv:2604.16406v1 Announce Type: cross Abstract: Realistic highway simulation is critical for scalable safety evaluation of autonomous vehicles, particularly for interactions that are too rare to stu

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

SafetyDGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

SafetyDGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

Human-Centered Supervision for Sentiment Analysis in Telugu: A Systematic Inquiry Beyond Accuracy

SafetyDGX agent

arXiv:2508.01486v3 Announce Type: replace Abstract: Sentiment analysis for low-resource languages remains challenging in an era where interpretability, human alignment, and fairness are increasingly n

Hybrid Multi-Dimensional MRI Prostate Cancer Detection via Hadamard Network-Based Bias Correction and Residual Networks

SafetyDGX agent

arXiv:2604.17107v1 Announce Type: new Abstract: Magnetic Resonance Imaging (MRI) is vital for prostate cancer (PCa) diagnosis. While advanced techniques such as Hybrid Multi-dimensional MRI (HM-MRI) h

Hybrid Spectro-Temporal Fusion Framework for Structural Health Monitoring

SafetyDGX agent

arXiv:2604.16589v1 Announce Type: new Abstract: Structural health monitoring plays a critical role in ensuring structural safety by analyzing vibration responses from engineering systems. This paper p

Hyperbolic Enhanced Representation Learning for Incomplete Multi-view Clustering

SafetyDGX agent

arXiv:2604.16959v1 Announce Type: cross Abstract: Incomplete Multi-View Clustering (IMVC) faces the challenge of learning discriminative representations from fragmentary observations while maintaining

I agree with Connor that most folks concerned about AI have been far too coy about extinction risk, and that ControlAI is one of few excepti…

SafetyDGX agent

I agree with Connor that most folks concerned about AI have been far too coy about extinction risk, and that ControlAI is one of few exceptions. The world needs more earnest communication efforts. We

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to res…

SafetyDGX agent

I love the way this is framed. New congressional districts that would disenfranchise almost half of Virginia GOP voters are designed 'to restore fairness in the upcoming elections.' Modern American po

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

SafetyDGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

Identifying Ethical Biases in Action Recognition Models

SafetyDGX agent

arXiv:2604.17971v1 Announce Type: new Abstract: Human Action Recognition (HAR) models are increasingly deployed in high-stakes environments, yet their fairness across different human appearances has n

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI …

SafetyDGX agent

If you are inclined to take @ohabryka on his word on issues like this, you may want to know some context... Oliver accuses me and ControlAI of telling people to be more coy about extinction risks. I h

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

SafetyDGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

SafetyDGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

Inertia in Moral and Value Judgments of Large Language Models

SafetyDGX agent

arXiv:2408.09049v3 Announce Type: replace Abstract: Large Language Models (LLMs) behave non-deterministically, and prompting has become a common method for steering their outputs. A popular strategy i

Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias

SafetyDGX agent

arXiv:2601.16934v2 Announce Type: replace Abstract: To be discoverable in an embedding-based search process, each part of a document should be reflected in its embedding representation. To quantify an

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

SafetyDGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

Instinct vs. Reflection: Unifying Token and Verbalized Confidence in Multimodal Large Models

SafetyDGX agent

arXiv:2604.17274v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in various perception and reasoning tasks. Despite this success, ens

Integrated Wheel Sensor Communication using ESP32 -- A Contribution towards a Digital Twin of the Road System

SafetyDGX agent

arXiv:2509.04061v2 Announce Type: replace Abstract: While current onboard state estimation methods are adequate for most driving and safety-related applications, they do not provide insights into the

Inter-Agent Relative Representations for Multi-Agent Option Discovery

SafetyDGX agent

arXiv:2512.24827v3 Announce Type: replace Abstract: Temporally extended actions improve the ability to explore and plan in single-agent settings. In multi-agent settings, the exponential growth of the

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

SafetyDGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

J-PARSE: Jacobian-based Projection Algorithm for Resolving Singularities Effectively in Inverse Kinematic Control of Serial Manipulators

SafetyDGX agent

arXiv:2505.00306v5 Announce Type: replace Abstract: J-PARSE is an algorithm for smooth first-order inverse kinematic control of a serial manipulator near kinematic singularities. The commanded end-eff

Jailbreaking Large Language Models with Morality Attacks

SafetyDGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

Last Wednesday, our founder and scientific advisor, @Yoshua_Bengio, was officially appointed an Officer of the Order of the British Empire (…

SafetyDGX agent

Last Wednesday, our founder and scientific advisor, @Yoshua_Bengio, was officially appointed an Officer of the Order of the British Empire (OBE). This prestigious distinction recognizes his contributi

LatentMimic: Terrain-Adaptive Locomotion via Latent Space Imitation

SafetyDGX agent

arXiv:2604.16440v1 Announce Type: new Abstract: Developing natural and diverse locomotion controllers for quadruped robots that can adapt to complex terrains while preserving motion style remains a si

Learning-Based Sparsification of Dynamic Graphs in Robotic Exploration Algorithms

SafetyDGX agent

arXiv:2604.16509v1 Announce Type: cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapi

Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification

SafetyDGX agent

arXiv:2601.21244v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced LLM reasoning, but remains constrained by inefficient exploration under lim

LiDAR-based Crowd Navigation with Visible Edge Group Representation

SafetyDGX agent

arXiv:2604.16741v1 Announce Type: new Abstract: Robot navigation in crowded pedestrian environments is a well-known challenge and we explore the practical deployment of group-based representations in

LLM-Extracted Covariates for Clinical Causal Inference: Rethinking Integration Strategies

SafetyDGX agent

arXiv:2604.16763v1 Announce Type: new Abstract: Causal inference from electronic health records (EHR) is fundamentally limited by unmeasured confounding: critical clinical states such as frailty, goal

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams …

SafetyDGX agent

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams actually close the evals feedback loop. Looking forward to I

Mammo-FM: Breast-specific foundational model for Integrated Mammographic Diagnosis, Prognosis, and Reporting

SafetyDGX agent

arXiv:2512.00198v2 Announce Type: replace Abstract: Breast cancer is one of the leading causes of death among women worldwide. We introduce Mammo-FM, the first foundation model specifically for mammog

Mark Zuckerberg and Meta Platforms $META just sent a memo to employees saying Meta Platforms is installing a new tracking software on the co…

SafetyDGX agent

Mark Zuckerberg and Meta Platforms $META just sent a memo to employees saying Meta Platforms is installing a new tracking software on the computers of all employees in the United States 🇺🇸 so it can t

MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning

SafetyDGX agent

arXiv:2602.17550v3 Announce Type: replace Abstract: Existing Reinforcement Learning with Verifiable Rewards (RLVR) algorithms, such as GRPO, rely on rigid, uniform, and symmetric trust region mechanis

MASSIVE: 🇺🇸 The BBC just validated everything we've been saying. A clear pattern of trades right before major Trump announcements. Iran wa…

SafetyDGX agent

MASSIVE: 🇺🇸 The BBC just validated everything we've been saying. A clear pattern of trades right before major Trump announcements. Iran war. Tariff reversals. Policy shifts. We tracked a whale for wee

Mechanisms of Multimodal Synchronization: Insights from Decoder-Based Video-Text-to-Speech Synthesis

SafetyDGX agent

arXiv:2411.17690v3 Announce Type: replace-cross Abstract: Unified decoder-only transformers have shown promise for multimodal generation, yet the mechanisms by which they synchronize modalities with h

MESA: A Training-Free Multi-Exemplar Deep Framework for Restoring Ancient Inscription Textures

SafetyDGX agent

arXiv:2604.17390v1 Announce Type: new Abstract: Ancient inscriptions frequently suffer missing or corrupted regions from fragmentation, erosion, or other damage, hindering reading, and analysis. We re

Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training (Reuters)

SafetyDGX agent

Reuters: Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training — Meta (META.O) is installing new

← Previous
1…188189190191192…212
Next →