AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
Safety

Trump's emergency orders pushing coal power are 'illegal' as well as dumb

DGX agent

The Trump administration's Department of Energy has invoked Section 202(c) of the Federal Power Act — a provision granting broad emergency authority over the electricity system that had previously...

safetyars-technica
9 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Again, if you care about computer security, read the red team report: https://red.anthropic.com/2026/mythos-preview/

DGX agent

Anthropic's Frontier Red Team report (April 2026) details the cybersecurity capabilities of Claude Mythos Preview, a general-purpose frontier model that performs strongly across the board but is s...

safetyethan-mollick--x
8 Apr 2026
Safety

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, …

DGX agent

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, cascades latency across your system 🛑 Recitation Errors — saf

safetyjerry-liu--x
8 Apr 2026
Safety

Dudes who won’t tag me because they know their arguments are weak sauce.* *for intellectual exercise you can list the flaws and misrepresent…

DGX agent

I was unable to retrieve the specific tweet at that URL — the X (Twitter) page requires JavaScript/login to load, and the tweet ID `2041954164562145434` does not appear in any indexed search result...

safetygary-marcus--x
8 Apr 2026
Safety

In AI, a lot can change in seven years.

DGX agent

The specific tweet (status ID 2041953155651661977) is not accessible — that ID appears to be from a future date and does not correspond to any retrievable post in the search results. The URL provid...

safetygary-marcus--x
8 Apr 2026
Safety

Incredible.

DGX agent

I was unable to retrieve the specific post at the URL provided (status ID `2041955770644713823`). This post ID does not appear in any search results, and X (formerly Twitter) requires JavaScript/lo...

safetygary-marcus--x
8 Apr 2026
Safety

literally fourteen minutes after my last explanation of why this is a false dichotomy 🤦‍♂️

DGX agent

The specific tweet (status ID 2041904683338625283) is not publicly accessible through search results, and the URL provided appears to reference a future or inaccessible post. The tweet ID is also b...

safetygary-marcus--x
8 Apr 2026
Safety

not surprised by any of this, headline or subheading

DGX agent

I was unable to retrieve the specific tweet at that URL (tweet ID 2041912293475414378). The tweet ID is extremely high — well beyond current Twitter/X ID ranges as of today — suggesting it may be a...

safetygary-marcus--x
8 Apr 2026
Safety

Voice ChatGPT can’t start a timer, but AGI is imminent! 🤦‍♂️

DGX agent

AI critic Gary Marcus uses the irony of ChatGPT's Voice mode being unable to perform a basic task — starting a timer — as a pointed illustration of the gap between AI industry hype and real-world c...

safetygary-marcus--x
8 Apr 2026
Safety

You should read the red team report: https://red.anthropic.com/2026/mythos-preview/

DGX agent

Anthropic's Frontier Red Team published a technical report (April 2026) detailing how their unreleased model, Claude Mythos Preview, autonomously identifies and exploits critical security vulnerabi...

safetyethan-mollick--x
7 Apr 2026
Safety

A Browser-Native Digital Test Range for Benchmarking 4D Ocean-Glider Planning Algorithms

DGX agent

arXiv:2608.13511v1 Announce Type: new Abstract: Repeated in-situ evaluation of ocean-glider planners requires scarce vehicles, operators, deployment and recovery resources, and ocean conditions that c

safetyarxiv-cs-ro
14 Aug 2026
Safety

A Compositional Theory of Curvature in Probabilistic Circuits

DGX agent

arXiv:2608.12869v1 Announce Type: cross Abstract: Probabilistic Circuits (PCs) are generative models that support exact inference and, unlike deep neural networks, admit an exact and tractable measure

safetyarxiv-cs-ai
14 Aug 2026
Safety

Active-Trace Complexity Bounds for Moreau--Yosida Unadjusted Langevin Sampling

DGX agent

arXiv:2608.13467v1 Announce Type: new Abstract: We study the Moreau--Yosida unadjusted Langevin algorithm (MYULA) for the nonsmooth composite target [ pi(dx)propto exp{-f(x)-g(x)},dx, qquad xinmathbb

safetyarxiv-cs-lg
14 Aug 2026
Safety

AI-Driven Multiscenario Interest Rate Forecasting: A Proof of Concept for Banking Asset Management

DGX agent

arXiv:2608.12424v1 Announce Type: cross Abstract: This study focuses on developing an AI-supported prototype for multiperspective interest rate forecasting that combines classical econometric models w

safetyarxiv-cs-lg
14 Aug 2026
Safety

ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs

DGX agent

arXiv:2608.12788v1 Announce Type: new Abstract: The rapid advancement of Auto-Research has surfaced a fundamental evaluation challenge: how can we measure the alignment, logical coherence, and evoluti

safetyarxiv-cs-ai
14 Aug 2026
Safety

Attention from Action, for Action: Emergent Visual Bottlenecks for Policy Learning

DGX agent

arXiv:2608.13422v1 Announce Type: new Abstract: Visual bottlenecks that focus policy inputs on regions of interest (ROIs) can improve data-efficient visuomotor learning by separating where to look fro

safetyarxiv-cs-ro
14 Aug 2026
Safety

Automated Design Optimization via Strategic Search with Large Language Models

DGX agent

arXiv:2511.22651v2 Announce Type: replace-cross Abstract: Optimization methods have long advanced many fields, yet they struggle when faced with design problems where the search space and design param

safetyarxiv-cs-ai
14 Aug 2026
Safety

Better Decomposition, Free Aggregation: A Synthesizer-Folding Framework for Multilingual Multi-Hop Question Answering

DGX agent

arXiv:2608.13160v1 Announce Type: cross Abstract: Multilingual retrieval-augmented generation (mRAG) equips large language models with access to globally distributed external knowledge for complex mul

safetyarxiv-cs-ai
14 Aug 2026
Safety

Beyond Outcome Rewards: Step-Level Self-Distilled Policy Optimization for Deep Search Agents

DGX agent

arXiv:2608.12764v1 Announce Type: cross Abstract: Deep search agents operate over trajectories spanning dozens of steps, yet standard reinforcement learning provides only a single outcome reward per t

safetyarxiv-cs-ai
14 Aug 2026
Safety

Bias Mitigation in Face Recognition via Demographic-based Supervised Contrastive Learning

DGX agent

arXiv:2608.12971v1 Announce Type: new Abstract: Face recognition systems have been shown to be biased toward certain demographic groups by exhibiting different error rates across gender, age, or ethni

safetyarxiv-cs-cv
14 Aug 2026
Safety

ContactGuard: Pre-Contact Execution Monitoring with Action-Conditioned Latent World Models

DGX agent

arXiv:2608.13438v1 Announce Type: cross Abstract: Contact-rich manipulation failures are often detected only after the robot has committed to contact. This is especially limiting in wrist-camera setup

safetyarxiv-cs-ai
14 Aug 2026
Safety

Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation

DGX agent

arXiv:2608.13391v1 Announce Type: new Abstract: Interactive autoregressive video generation demands both low-latency rollouts and precise online control. Few-step distillation accelerates generation b

safetyarxiv-cs-cv
14 Aug 2026
Safety

CROP: Task Relevance via Counterfactuals for Selective On-Policy Distillation

DGX agent

arXiv:2608.13387v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises a student language model on trajectories sampled from its current policy, but assigns equal credit to response t

safetyarxiv-cs-cl
14 Aug 2026
Safety

CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers

DGX agent

arXiv:2608.12773v1 Announce Type: new Abstract: Semi-supervised semantic segmentation has long turned on one question, which pseudo-labels to trust, and a generation of selection rules, dynamic thresh

safetyarxiv-cs-cv
14 Aug 2026
Safety

DAPD: Dual-Anchored Policy Distillation

DGX agent

arXiv:2608.01735v2 Announce Type: replace Abstract: On-policy (self) distillation (OPSD) is increasingly adopted for language-model post-training. It strengthens the teacher with privileged informatio

safetyarxiv-cs-ai
14 Aug 2026
Safety

Decoding Task Progress from VLA Representations

DGX agent

arXiv:2608.13474v1 Announce Type: new Abstract: Vision-language-action models (VLAs) are moving rapidly towards deployment as general-purpose manipulation policies, but we currently lack basic tools f

safetyarxiv-cs-ro
14 Aug 2026
Safety

Decoupled Contrastive Decoding via Expert-Aligned Drafting

DGX agent

arXiv:2608.12913v1 Announce Type: new Abstract: Contrastive Decoding (CD) improves generation quality, but its amateur-model pass makes decoding expensive. Accelerating CD with speculative decoding ra

safetyarxiv-cs-cl
14 Aug 2026
Safety

DiffGRM: Diffusion-based Generative Recommendation Model

DGX agent

arXiv:2510.21805v2 Announce Type: replace-cross Abstract: Generative recommendation (GR) is an emerging paradigm that represents each item via a tokenizer as an n-digit semantic ID (SID) and predicts

safetyarxiv-cs-ai
14 Aug 2026
Model Releases

Doctorina MedBench: A Dialogue-Based Benchmark and Evaluation Framework for Agent-Based Medical AI

DGX agent

arXiv:2603.25821v3 Announce Type: replace-cross Abstract: We present Doctorina MedBench, an evaluation framework for agent-based medical AI based on the simulation of physician-patient interactions. U

model-releasesarxiv-cs-ai
14 Aug 2026
Safety

Doubly Robust Estimation of Causal Effect on CVR with Targeted Regularization

DGX agent

arXiv:2608.13461v1 Announce Type: new Abstract: Post-click conversion rate (CVR) is a key metric in various scenarios including e-commerce and advertising, reflecting the efficiency and user experienc

safetyarxiv-cs-lg
14 Aug 2026
Safety

Dual-Stream Cross-Anchor Correction Grounding Long-Form Captions and the Domain Limits of Object-Level Anchors

DGX agent

arXiv:2608.12746v1 Announce Type: cross Abstract: Object hallucination in multimodal large language models arises when language priors and corpus co-occurrence bias outweigh the visual evidence, with

safetyarxiv-cs-cl
14 Aug 2026
Safety

Embedding networks with the random walk first return time distribution

DGX agent

arXiv:2512.02694v3 Announce Type: replace-cross Abstract: We propose the first return time distribution (FRTD) of a random walk as an interpretable and mathematically grounded node embedding. The FRTD

safetyarxiv-cs-lg
14 Aug 2026
Safety

Entropy-Augmented Multi-Objective Policy Optimization in Multiagent Systems

DGX agent

arXiv:2608.12534v1 Announce Type: cross Abstract: Autonomous agent teams deployed in settings such as marine and extraterrestrial outposts must coordinate actions to achieve optimal outcomes across mu

safetyarxiv-cs-ro
14 Aug 2026
Safety

EU-ETS under attack? The impact of carbon price suppression on the decarbonization of the power sector

DGX agent

arXiv:2608.12363v1 Announce Type: cross Abstract: European countries are debating policies to mitigate the increased energy costs caused by renewed geopolitical tensions, while pursuing decarbonizatio

safetyarxiv-cs-ai
14 Aug 2026
Safety

Evaluation Resolution Confounds Learning-Rule Comparisons in Model-Brain RSA of Early Visual Cortex

DGX agent

arXiv:2608.12408v1 Announce Type: cross Abstract: Representational similarity analysis (RSA) is increasingly used to ask which learning rules give convolutional networks brain-like representations. Be

safetyarxiv-cs-lg
14 Aug 2026
Safety

EvoTale: Continual Character Customization for Expanding Story Worlds

DGX agent

arXiv:2603.16285v2 Announce Type: replace Abstract: Character-centric story visualization aims to synthesize coherent image sequences that depict narrative events and interactions while preserving rec

safetyarxiv-cs-cv
14 Aug 2026
Safety

Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control

DGX agent

arXiv:2506.18278v4 Announce Type: replace-cross Abstract: We introduce an original minimax framework for finite-time performance analysis in queueing control and propose a surprisingly simple Lyapunov

safetyarxiv-cs-lg
14 Aug 2026
Safety

FSGR: Mitigating Token Frequency Bias for Fair SID-Based Generative Recommendation

DGX agent

arXiv:2608.12845v1 Announce Type: cross Abstract: Semantic ID (SID)-based generative recommendation has recently achieved remarkable success. However, existing methods suffer from a previously overloo

safetyarxiv-cs-ai
14 Aug 2026
Safety

Humans are Missing from AI Coding Agent Research

DGX agent

arXiv:2608.12355v1 Announce Type: cross Abstract: Recent progress in AI coding agent research has led to rapid improvements in agents' ability to autonomously perform complex software engineering task

safetyarxiv-cs-ai
14 Aug 2026
Safety

HybridSB-MoE: Dual-Domain Schrodinger Bridges with Scene-Adaptive Expert Routing for Speech Enhancement

DGX agent

arXiv:2608.12715v1 Announce Type: cross Abstract: Generative speech enhancement faces three gaps: spectral models capture harmonic structure but often disrupt phase, waveform models preserve phase but

safetyarxiv-cs-ai
14 Aug 2026
Safety

I-SDPO: Instance-Level Adaptive Self-Distillation Policy Optimization

DGX agent

arXiv:2608.12957v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) learns from reward differences within a rollout group, but receives no useful relative signal when every sam

safetyarxiv-cs-cl
14 Aug 2026
Safety

Intern-S2-Preview: Scientific Agentic Foundation Model

DGX agent

arXiv:2608.13505v1 Announce Type: cross Abstract: Scientific discovery increasingly requires AI systems that can reason over scientific evidence of heterogeneous modalities, interact with scientific t

safetyarxiv-cs-cl
14 Aug 2026
Safety

Into the ORBIT for Time Series: Training Regimes for Foundation Models

DGX agent

arXiv:2608.13262v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have advanced primarily through architectural innovation, while training regimes for large-scale heterogeneous c

safetyarxiv-cs-ai
14 Aug 2026
Safety

It's How You Ask: Gender-Associated Linguistic Bias in LLMs

DGX agent

arXiv:2608.13328v1 Announce Type: cross Abstract: Professional communication is increasingly mediated by LLMs - but do these models serve all users equally? We show that when prompts contain linguisti

safetyarxiv-cs-ai
14 Aug 2026
Safety

Large Language Models Persuade Without Planning Theory of Mind

DGX agent

arXiv:2602.17045v2 Announce Type: replace Abstract: A growing body of work attempts to evaluate the theory of mind (ToM) abilities of humans and large language models (LLMs) using static, non-interact

safetyarxiv-cs-cl
14 Aug 2026
Safety

Latent On-Policy Self-Distillation

DGX agent

arXiv:2608.13040v1 Announce Type: cross Abstract: Enabling agents to learn from experience and internalize it into their policy has become a central problem in self-evolving AI. On-policy self-distill

safetyarxiv-cs-cl
14 Aug 2026
Safety

Learning Under Treatment-Induced Label Indeterminacy with Expert Annotations of Counterfactual Outcomes: A Case Study in Neurological Prognostication

DGX agent

arXiv:2608.12477v1 Announce Type: new Abstract: Clinical prediction models are often developed as if the outcome of interest were cleanly observed for every patient. This assumption fails when treatme

safetyarxiv-cs-lg
14 Aug 2026
Local Ai

LLM-Assisted Dynamic Threat Analysis for Attacker-Reachable Software Weaknesses in Autonomous Vehicles

DGX agent

arXiv:2608.13450v1 Announce Type: cross Abstract: Autonomous vehicles depend on large safety-critical software stacks, where weaknesses reachable from adversarial inputs may affect steering, braking,

local-aiarxiv-cs-lg
14 Aug 2026
← Previous
1…6465666768…302
Next →