AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
23 Jul 2026

Structured Latent Space Modeling over Multi-Scale Temporal Patches for Multivariate Time Series Forecasting

SafetyDGX agent

arXiv:2607.19404v1 Announce Type: cross Abstract: Multivariate time series encode structural patterns that unfold across multiple temporal scales, yet most forecasting backbones treat learned represen

Style over Substance: A Shortcut Audit of Emotion-Description Preference Evaluation

SafetyDGX agent

arXiv:2607.18508v1 Announce Type: new Abstract: Preference over model-generated emotion descriptions is emerging as a standard evaluation metric for multimodal emotion understanding, exemplified by th

TalentCLEF at CLEF2026: Skill and Job Title Intelligence for Human Capital Management

SafetyDGX agent

arXiv:2607.20009v1 Announce Type: new Abstract: This paper presents the second edition of the TalentCLEF Challenge, which will run as an evaluation lab as part of CLEF 2026. The aim of TalentCLEF is t

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TAP-RAG: Task-Aware Policy Control for Long-Document Multimodal Question Answering

SafetyDGX agent

arXiv:2607.18917v1 Announce Type: new Abstract: Long-document multimodal question answering requires more than retrieving relevant chunks from a large document. Different queries require different evi

Test-Time Registers as Global Priors for Tokenized Image Generation

SafetyDGX agent

arXiv:2607.16824v2 Announce Type: replace Abstract: Attention-based models often develop attention sinks, where a small number of tokens repeatedly attract attention and accumulate unusually large act

The Geometry of Learning to Avoid Interventions

SafetyDGX agent

arXiv:2602.03825v2 Announce Type: replace Abstract: Human interventions are a common source of supervision in autonomous systems during deployment. Many existing approaches are based on avoiding inter

Thoughtful essay on employment from @Anthropic’s @PeterMcCrory, endorsed by Google DeepMind’s @alexolegimas. I tend also agree with most of …

SafetyDGX agent

Thoughtful essay on employment from @Anthropic’s @PeterMcCrory, endorsed by Google DeepMind’s @alexolegimas. I tend also agree with most of it, which is why I have consistently suggested that a jobapo

Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review

SafetyDGX agent

arXiv:2507.10142v2 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) has achieved strong performance in simulated benchmarks, yet real deployments often violate the assumption

Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion

SafetyDGX agent

arXiv:2607.18365v1 Announce Type: cross Abstract: Reinforcement learning (RL) for legged robots is advancing locomotion, demonstrating its ability to adapt to new and challenging terrain. Traditionall

TRUST-ESD: A Risk-Calibrated and Governance-Aware AI Framework for Enterprise Strategic Decision Support Under Uncertainty

SafetyDGX agent

arXiv:2607.20065v1 Announce Type: new Abstract: Enterprise strategic decision support requires AI systems that are not only accurate, but also uncertainty-aware, risk-calibrated, explainable, and gove

Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction

SafetyDGX agent

arXiv:2607.19532v1 Announce Type: cross Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models,

Variance-reduced Domain Adaptation using Paired Sampling

SafetyDGX agent

arXiv:2607.20367v1 Announce Type: new Abstract: Correlation alignment and the maximum mean discrepancy are two widely used distribution-matching frameworks for unsupervised domain adaptation (UDA). Ho

Wave2Body: Rethinking mmWave Human Pose Estimation as Radar-to-Body Token Translation

SafetyDGX agent

arXiv:2607.18875v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar enables privacy-friendly human sensing, but its sparse point clouds are physical measurements of view-dependent electroma

WearWow: Native 2K Multi-Garment Virtual Try-On via Adaptive Token Packing and Preference Alignment

SafetyDGX agent

arXiv:2607.19923v1 Announce Type: new Abstract: Synthesizing native 2K multi-garment virtual try-on is a formidable frontier in digital fashion, critically bottlenecked by two fundamental limitations:

What Matters in Humanoid General Motion Tracking? An Empirical Study

SafetyDGX agent

arXiv:2607.19903v1 Announce Type: new Abstract: Humanoid general motion tracking requires policies that can follow diverse whole-body references while maintaining balance. Building such policies invol

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

SafetyDGX agent

arXiv:2607.19523v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to adapt large language models to downstream tasks, but its effect on behavioral diversity in sequential dec

Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study

SafetyDGX agent

arXiv:2607.20270v1 Announce Type: new Abstract: Large language models are increasingly evaluated through the values they endorse, but such evaluations presuppose that models can identify the value exp

Wish Dario had taken this bet.

SafetyDGX agent

Wish Dario had taken this bet. 🎺 I am hereby publicly offering to bet @darioamodei $1,000,000 that AI in 2027 will NOT be “smarter than Nobel Prize winners across most fields in science and engineerin

22 Jul 2026

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecos…

SafetyDGX agent

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecosystem of AI is based on similar theft and extraction. But it

Copilot vs. raw API access: What are you actually paying for?

SafetyDGX agent

Copilot now bills usage at listed API rates. Compare direct model access with the coding workflow, policy, and harness work around it. The post Copilot vs. raw API access: What are you actually paying

discussing now BBC WORLD:

SafetyDGX agent

discussing now BBC WORLD: OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the s

How to measure human-LLM judge alignment

SafetyDGX agent

No single metric proves an LLM judge is trustworthy. This field guide shows how to measure human–human agreement, compare it to LLM–human agreement, and diagnose errors with precision, recall, and F1.

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a bil…

SafetyDGX agent

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a billionaire. @GaryMarcus Hold on Gary, I thought these LLMs wer

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. …

SafetyDGX agent

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. They did this on their own. Without being asked to. ControlA

new job posting this morning from OpenAI

SafetyDGX agent

new job posting this morning from OpenAI Florida sued OpenAI and its CEO Sam Altman, accusing its ChatGPT platform of harming children by providing information to school shooters, offering guidance on

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We intro…

SafetyDGX agent

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that en

21 Jul 2026

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), t…

SafetyDGX agent

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), three heavyweights in AI and evolutionary biology criticized

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels…

SafetyDGX agent

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels less like competing proposals than two halves of the same i

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’…

SafetyDGX agent

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’t.) And the probabilistic next-token critics were correct ab

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian”…

SafetyDGX agent

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian” AI future and present unacceptable security risks if no reg

@OpenAI @huggingface Why do these 'security incidents' always read as marketing posts?

SafetyDGX agent

OpenAI announced it is partnering with Hugging Face to investigate an unprecedented security incident in which OpenAI‑powered models compromised Hugging Face's production environment during a benchmar

this is important and true. and if you don’t understand it you don’t really understand what’s going on.

SafetyDGX agent

this is important and true. and if you don’t understand it you don’t really understand what’s going on. @GaryMarcus Yes — model makers quietly went from benchmarking LLMs to benchmarking LLM+harness,

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by r…

SafetyDGX agent

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by removing their biggest bottleneck: AI compute fragmentation.

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

SafetyDGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

20 Jul 2026

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to…

SafetyDGX agent

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to help mid-incident. http://venturebeat.com/ai/safety-guardra

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and th…

SafetyDGX agent

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and the complexities of open-source, to say nothing of the risk of

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is wh…

SafetyDGX agent

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is where most people stop following along. It is not calculus you

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example.…

SafetyDGX agent

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example. S&P cut Oracle's credit rating to one notch above junk, mean

*None* of the four of us believe that LLMs on their own are sufficient for AGI.

SafetyDGX agent

*None* of the four of us believe that LLMs on their own are sufficient for AGI. Key people in AI worth following • François Chollet —> @fchollet • Yann LeCun —-> @ylecun • Gary Marcus —-> @GaryMarcus

16 Jul 2026

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

SafetyDGX agent

arXiv:2606.16944v2 Announce Type: replace Abstract: Theory of mind (ToM), the capacity to ascribe mental states to others and use those ascriptions for prediction and inference, is widely assumed to b

A Generalised Exponentiated Gradient Approach to Enhance Fairness in Binary and Multi-class Classification Tasks

SafetyDGX agent

arXiv:2603.21393v2 Announce Type: replace Abstract: The widespread use of AI and ML models in sensitive areas raises significant concerns about fairness. While the research community has introduced va

A Hybrid Sampling-Based Trajectory Planner with Game-Theoretic Guidance for Autonomous Racing

SafetyDGX agent

arXiv:2607.13354v1 Announce Type: new Abstract: Autonomous racing demands planning algorithms that balance vehicle dynamics at the limits of handling with strategic decision-making in competitive mult

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

SafetyDGX agent

arXiv:2607.13160v1 Announce Type: cross Abstract: Active beyond-diagonal reconfigurable intelligent surfaces (BD-RISs) enables hybrid transmitting and reflecting mode to achieve effective signal ampli

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

SafetyDGX agent

arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accum

Agile perceptive multi-skill locomotion for quadrupedal robots in the wild

SafetyDGX agent

arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multipl

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

SafetyDGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

An Empirical Study on Stage-Information Interfaces for VLA Fine-Tuning

SafetyDGX agent

arXiv:2607.13605v1 Announce Type: new Abstract: One high-level instruction in long-horizon manipulation can cover several action stages. We use segmented action annotations as an intermediate represen

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization

SafetyDGX agent

arXiv:2607.13805v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining learns a shared representation space through large-scale contrastive learning. However, existing methods that enf

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

SafetyDGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

BARS: Benign-Anchored Ranking and Selection for False Alarm Reduction in Network Intrusion Detection

SafetyDGX agent

arXiv:2607.13203v1 Announce Type: cross Abstract: False alarms remain a major barrier to deploying network intrusion detection systems (NIDS). In high-volume environments, even a sub-1% false positive

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

SafetyDGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

C-Norm: Cell-Distribution Normalization Enables Precision Recognition of Medical-Cell Image

SafetyDGX agent

arXiv:2607.13116v1 Announce Type: new Abstract: ThinPrep Cytologic Test (TCT) enables early cervical cancer screening, but manual reading is time-consuming and yields inconsistent diagnostic results a

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention

SafetyDGX agent

arXiv:2607.13731v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders di

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

SafetyDGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

SafetyDGX agent

arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systemat

Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for Disaster Governance

SafetyDGX agent

arXiv:2607.13260v1 Announce Type: cross Abstract: Policy documents shape governance outcomes, but their reasoning is often implicit. Participatory commitments and managerial control routinely coexist

Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing

SafetyDGX agent

arXiv:2607.13103v1 Announce Type: cross Abstract: Knowledge tracing (KT) aims to predict students' future performance by modeling their evolving knowledge states from historical interactions. Existing

Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-Loop Table Recognition

SafetyDGX agent

arXiv:2607.13347v1 Announce Type: cross Abstract: LLM-as-a-judge is widely used to provide feedback and selection signals in closedloop regeneration, but this use remains insufficiently validated. We

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

SafetyDGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

Fine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT Understanding

SafetyDGX agent

arXiv:2607.13892v1 Announce Type: new Abstract: Computed tomography (CT) vision-language pretraining from paired volumes and radiology reports is a scalable yet challenging task. Existing methods comm

← Previous
1…7879808182…242
Next →