AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
6 May 2026

100% stood the test of time: “What Ilya saw” was Sam’s bad behavior, not AGI.

SafetyDGX agent

100% stood the test of time: “What Ilya saw” was Sam’s bad behavior, not AGI. When Altman got fired “What did Ilya see?” became a wildly popular conspiracy meme. I think we can safely say now that wha

2nd episode of The Roman Forum is an interview with AI Safety/Governance expert Connor Leahy @NPCollapse. Connor is a great speaker and is l…

SafetyDGX agent

2nd episode of The Roman Forum is an interview with AI Safety/Governance expert Connor Leahy @NPCollapse. Connor is a great speaker and is lobbying to get government to ban Superintelligence. My first

A Knowledge-Driven LLM-Based Decision-Support System for Explainable Defect Analysis and Mitigation Guidance in Laser Powder Bed Fusion

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.01100v1 Announce Type: new Abstract: This work presents a knowledge-driven decision-support system that integrates structured defect knowledge with LLM-based reasoning to provide explainabl

A Robust Unsupervised Domain Adaptation Framework for Medical Image Classification Using RKHS-MMD

SafetyDGX agent

arXiv:2605.03787v1 Announce Type: new Abstract: Labeling medical images is a major bottleneck in the field of medical imaging, as it requires domain-specific expertise, and it gets further complicated

A Three-Stage Offline SDRE-Based Control Framework for Human Motion Reproduction on a Suspended Bipedal Robot

SafetyDGX agent

arXiv:2506.04680v2 Announce Type: replace Abstract: During the development of wearable exoskeletons, evaluations involving human subjects pose inherent safety risks. Therefore, systematic testing is o

A Universal Reproducing Kernel Hilbert Space from Polynomial Alignment and IMQ Distance

SafetyDGX agent

arXiv:2605.03262v1 Announce Type: new Abstract: We introduce the Yat kernel $k_{b,arepsilon}(mathbf{w},mathbf{x})=frac{(mathbf{w}^opmathbf{x}+b)^2}{|mathbf{x}-mathbf{w}|^2+arepsilon},qquad bge 0, arep

A US appeals court strikes down a 2023 FCC rule banning broadband access discrimination based on income, race, and more; Chair Brendan Carr welcomes the ruling (Jon Brodkin/Ars Technica)

SafetyDGX agent

Jon Brodkin / Ars Technica: A US appeals court strikes down a 2023 FCC rule banning broadband access discrimination based on income, race, and more; Chair Brendan Carr welcomes the ruling — An appeals

A Vision-Based Shared-Control Teleoperation Scheme for Controlling the Robotic Arm of a Four-Legged Robot

SafetyDGX agent

arXiv:2508.14994v3 Announce Type: replace-cross Abstract: In hazardous and remote environments, robotic systems perform critical tasks demanding improved safety and efficiency. Among these, quadruped

Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

SafetyDGX agent

arXiv:2605.00968v1 Announce Type: cross Abstract: Positional encoding plays a pivotal role in determin?ing the extrapolation and generalization performance of wireless foundation models for channel st

ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms

SafetyDGX agent

arXiv:2605.03212v1 Announce Type: cross Abstract: Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic

Agentic AI-Based Joint Computing and Networking via Mixture of Experts and Large Language Models

SafetyDGX agent

arXiv:2605.02911v1 Announce Type: new Abstract: Future sixth-generation (6G) mobile networks are envisioned to be equipped with a diverse set of powerful, yet highly specialized, optimization experts.

Algebraic Semantics of Governed Execution: Monoidal Categories, Effect Algebras, and Coterminous Boundaries

SafetyDGX agent

arXiv:2605.01032v2 Announce Type: new Abstract: We present an algebraic semantics for governed execution in which governance is axiomatized, compositional, and coterminous with expressibility. The fra

Aligning Inductive Bias for Data-Efficient Generalization in State Space Models

SafetyDGX agent

arXiv:2509.20789v4 Announce Type: replace Abstract: The remarkable success of modern AI has been closely tied to scaling laws, yet the finite supply of high-quality data makes data efficiency--learnin

🚨Alphabet $GOOGL trades at 133x free cash flow. For context: its pre-COVID multiple was ~20x. And free cash flow hasn't grown since 2021. G…

SafetyDGX agent

🚨Alphabet GOOGL trades at 133x free cash flow. For context: its pre-COVID multiple was ~20x. And free cash flow hasn't grown since 2021. GQG Partners — one of the world's top institutional investors —

Am I right that hyperscaling compute is the biggest bet in history? Any counter examples? It’s way more expensive than the Manhattan Project…

SafetyDGX agent

Am I right that hyperscaling compute is the biggest bet in history? Any counter examples? It’s way more expensive than the Manhattan Project, the Apollo project, and railways across the US. If it does

Amazing the shit X gave me in November 2023 for saying Sam wasn’t always candid. He wasn’t. Period.

SafetyDGX agent

Amazing the shit X gave me in November 2023 for saying Sam wasn’t always candid. He wasn’t. Period. Murati testified that, by fall 2023, Sam Altman was “not always” candid with her, and undermined her

And how many times have I told you that Sam is no longer the right CEO for OpenAI?

SafetyDGX agent

And how many times have I told you that Sam is no longer the right CEO for OpenAI? Murati: My issues with Sam were very much around management and providing direction to the organization and decision

Anthropic researchers detail 'model spec midtraining', which adds a stage between pretraining and fine-tuning to improve generalization from alignment training (Anthropic)

SafetyDGX agent

Anthropic: Anthropic researchers detail “model spec midtraining”, which adds a stage between pretraining and fine-tuning to improve generalization from alignment training — Sara Price2, Samuel Marks2,

Architectural Obsolescence of Unhardened Agentic-AI Runtimes

SafetyDGX agent

arXiv:2605.01740v1 Announce Type: cross Abstract: An agentic-AI runtime issues tool calls, sends messages, and actuates devices on behalf of an LLM. Catching the four ways an action can diverge from i

Audio-Visual Intelligence in Large Foundation Models

SafetyDGX agent

arXiv:2605.04045v1 Announce Type: new Abstract: Audio-Visual Intelligence (AVI) has emerged as a central frontier in artificial intelligence, bridging auditory and visual modalities to enable machines

Beyond Activation Alignment: The Geometry of Neural Sensitivity

SafetyDGX agent

arXiv:2605.03222v1 Announce Type: new Abstract: Activation-alignment measures such as Representational Similarity Analysis (RSA), Canonical Correlation Analysis (CCA), and Centered Kernel Alignment (C

BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation

SafetyDGX agent

arXiv:2605.03452v1 Announce Type: new Abstract: High-quality data collection is a fundamental cornerstone for training humanoid whole-body visuomotor policies. Current data acquisition paradigms predo

Bootstrapped Mixed Rewards for RL Post-Training: Injecting Canonical Action Order

SafetyDGX agent

arXiv:2512.04277v3 Announce Type: replace Abstract: Post-training with reinforcement learning (RL) typically optimizes a single scalar objective and ignores structure in how solutions are produced. We

Can Semantic Methods Enhance Team Sports Tactics? A Methodology for Football with Broader Applications

SafetyDGX agent

arXiv:2601.00421v2 Announce Type: replace Abstract: This paper explores how semantic-space reasoning, traditionally used in computational linguistics, can be extended to tactical decision-making in te

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.01758v1 Announce Type: new Abstract: Large multimodal model-based Multi-Agent Systems (MASs) enable collaborative complex problem solving through specialized agents. However, MASs are vulne

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents

SafetyDGX agent

arXiv:2604.06132v2 Announce Type: replace Abstract: Large language models are increasingly deployed as autonomous agents for multi-step workflows in real-world software environments. However, existing

C’mon BBC. Zilis was sharp as a tack on the stand, on her role in the OpenAI *nonprofit* board, and how she managed conflicts as they began …

SafetyDGX agent

C’mon BBC. Zilis was sharp as a tack on the stand, on her role in the OpenAI *nonprofit* board, and how she managed conflicts as they began to develop, and this (which is not really even news since it

Coherent Hierarchical Multi-Label Learning to Defer for Medical Imaging

SafetyDGX agent

arXiv:2605.02734v1 Announce Type: new Abstract: Learning to Defer (L2D) enables a model to predict autonomously or defer to an expert, but prior work largely assumes flat label spaces. We study the fi

Correction: the jury is advisory only. It’s the judge who decides; if she sees it as I do, OpenAI loses.

SafetyDGX agent

Gary Marcus clarifies that in the legal proceeding he's discussing, the jury serves an advisory role while the judge retains decision-making authority on the case outcome. Marcus expresses confidence

Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective

SafetyDGX agent

arXiv:2605.02658v2 Announce Type: new Abstract: Shortcut learning causes deep learning models to rely on non-essential features within the data. However, its formation in deep neural network training

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning

SafetyDGX agent

arXiv:2602.20078v3 Announce Type: replace-cross Abstract: Scaling cooperative multi-agent reinforcement learning (MARL) is fundamentally limited by cross-agent noise. When agents share a common reward

DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment

SafetyDGX agent

arXiv:2605.03327v1 Announce Type: new Abstract: Reinforcement learning is crucial for aligning large language models to perform complex reasoning tasks. However, current algorithms such as Group Relat

Discovering Reinforcement Learning Interfaces with Large Language Models

SafetyDGX agent

arXiv:2605.03408v1 Announce Type: new Abstract: Reinforcement learning systems rely on environment interfaces that specify observations and reward functions, yet constructing these interfaces for new

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment

SafetyDGX agent

arXiv:2605.01899v1 Announce Type: new Abstract: The growing capabilities of large language models (LLMs) have driven their widespread deployment across diverse domains, even in potentially high-risk s

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models

SafetyDGX agent

arXiv:2605.03877v1 Announce Type: new Abstract: Dataset distillation enables efficient training by distilling the information of large-scale datasets into significantly smaller synthetic datasets. Dif

@Dr_Gingerballs At least actual ponzi schemes don't light their cash on fire... they just cant meet redemptions at the level of their inflat…

SafetyDGX agent

@Dr_Gingerballs At least actual ponzi schemes don't light their cash on fire... they just cant meet redemptions at the level of their inflated fake earnings. After all of the hyperscalers burn every l

Efficient Temporal Datalog Materialisation for Composite Event Recognition

SafetyDGX agent

arXiv:2605.02488v1 Announce Type: new Abstract: Several applications demand the timely detection of critical situations, such as threats to safety and transparency, over high-velocity streams of symbo

EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models

SafetyDGX agent

arXiv:2605.02921v1 Announce Type: cross Abstract: As LLMs continue to shape real-world applications, automated jailbreak generation becomes essential to reveal safety weaknesses and guide model improv

False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models

SafetyDGX agent

arXiv:2601.07885v2 Announce Type: replace-cross Abstract: Emoticons are widely used in digital communication to convey affective intent, yet their safety implications for Large Language Models (LLMs)

FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias Correction

SafetyDGX agent

arXiv:2605.03425v1 Announce Type: new Abstract: Differentially private (DP) training protects individual examples by adding noise to gradients, but the injected noise interacts nontrivially with adapt

FINER-SQL: Boosting Small Language Models for Text-to-SQL

SafetyDGX agent

arXiv:2605.03465v1 Announce Type: cross Abstract: Large language models have driven major advances in Text-to-SQL generation. However, they suffer from high computational cost, long latency, and data

FORMULA: FORmation MPC with neUral barrier Learning for safety Assurance

SafetyDGX agent

arXiv:2604.04409v2 Announce Type: replace Abstract: Multi-robot systems (MRS) are essential for large-scale applications such as disaster response, material transport, and warehouse logistics, yet ens

From SFT to RL: Demystifying the Post-Training Pipeline for LLM-based Vulnerability Detection

SafetyDGX agent

arXiv:2602.14012v2 Announce Type: replace-cross Abstract: The integration of LLMs into vulnerability detection (VD) has shifted the field toward more interpretable and context-aware analysis. While po

@GaryMarcus I can't believe this is still a thing that 'experts' haven't caught up with. @GaryMarcus has been saying this forever, and those…

SafetyDGX agent

@GaryMarcus I can't believe this is still a thing that 'experts' haven't caught up with. @GaryMarcus has been saying this forever, and those of us who have actually dug into the tech, analyzed it, use

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

SafetyDGX agent

arXiv:2507.07982v2 Announce Type: replace Abstract: Videos inherently represent 2D projections of a dynamic 3D world. However, our analysis suggests that video diffusion models trained solely on raw v

Global and Local Topology-Aware Attention with Persistent Homology and Euler Biases for Time-Series Forecasting

SafetyDGX agent

arXiv:2605.03163v1 Announce Type: new Abstract: Scientific time series often encode predictive geometric structure, including connectivity, cycles, shell-like geometry, directional changes, and nonlin

Google, Microsoft and xAI agree to allow government safety checks of their AI models prior to release

SafetyDGX agent

Google LLC, Microsoft Corp. and xAI have agreed to share unreleased versions of their artificial intelligence models with the U.S. Department of Commerce to ensure the technologies do not pose a threa

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

SafetyDGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

GRAFT: Auditing Graph Neural Networks via Global Feature Attribution

SafetyDGX agent

arXiv:2605.03377v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong performance on node classification tasks but remain difficult to interpret, particularly with respect to whi

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization

SafetyDGX agent

arXiv:2605.01482v1 Announce Type: new Abstract: Multi-Hop Fact Verification (MHFV) necessitates complex reasoning across disparate evidence, posing significant challenges for Large Language Models (LL

GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning

SafetyDGX agent

arXiv:2605.03403v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has recently shown strong performance in post-training large language models and vision-language models. It ra

Healthcare AI GYM for Medical Agents

SafetyDGX agent

arXiv:2605.02943v1 Announce Type: new Abstract: Clinical reasoning demands multi-step interactions -- gathering patient history, ordering tests, interpreting results, and making safe treatment decisio

Height Control and Optimal Torque Planning for Jumping With Wheeled-Bipedal Robots

SafetyDGX agent

arXiv:2605.03302v1 Announce Type: new Abstract: This paper mainly studies the accurate height jumping control of wheeled-bipedal robots based on torque planning and energy consumption optimization. Du

Heterogeneous Graph Importance Scoring and Clustering with Automated LLM-based Interpretation

SafetyDGX agent

arXiv:2605.02919v1 Announce Type: new Abstract: Urban bridge networks are critical infrastructure whose disruption can cascade into severe impacts on transportation, emergency services, and economic a

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2603.00977v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have recently demonstrated strong capabilities in interactive decision-making, yet they remain fundamentally

How Sam ('You parachute him onto a cannibal island, and he comes back five years later as king”) Altman operates:

SafetyDGX agent

How Sam ('You parachute him onto a cannibal island, and he comes back five years later as king”) Altman operates: counsel: 'by fall of 2023 did you percieve altman was not candid with you? truthful? h

Human-in-the-Loop Uncertainty Analysis in Self-Adaptive Robots Using LLMs

SafetyDGX agent

arXiv:2605.02983v1 Announce Type: new Abstract: Self-adaptive robots operate in dynamic, unpredictable environments where unaddressed uncertainties can lead to safety violations and operational failur

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until …

SafetyDGX agent

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until it's too late. But it's a very simple set up. Hyperscalers g

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or…

SafetyDGX agent

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or my own writings since fall 2023 They would have realized ho

Important nuance: the jury at this Musk-OpenAI trial is an *advisory* jury, hence not binding on the judge, and is only looking at liability…

SafetyDGX agent

Important nuance: the jury at this Musk-OpenAI trial is an *advisory* jury, hence not binding on the judge, and is only looking at liability (not damages, if any). Thanks to @bahhradx for correcting a

← Previous
1…161162163164165…212
Next →