AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,234 results
Model Releases

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run you…

DGX agent

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run your agent over a collection of docs (PDFs, images, Office) with

model-releasesjerry-liu--x
11 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Beyond the Black Box: Interpretability of Agentic AI Tool Use

DGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

model-releasesarxiv-cs-ai
11 May 2026
Industry

Generate an image of the scariest thing you could possibly imagine scribbled into a child's coloring book

DGX agent

This Reddit post from r/ChatGPT likely discusses a creative prompt experiment where users attempted to get ChatGPT or similar AI image generation tools to create disturbing or horror-themed content by

industryr-chatgpt
11 May 2026
Model Releases

Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies

DGX agent

arXiv:2510.22944v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have become indispensable for automated code generation, yet the quality and security of their outputs remain a c

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

When Routine Chats Turn Toxic: Unintended Long-Term State Poisoning in Personalized Agents

DGX agent

arXiv:2605.06731v1 Announce Type: cross Abstract: Personalized LLM agents maintain persistent cross-session state to support long-horizon collaboration. Yet, this persistence introduces a subtle but c

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

NHTSA says the 2026 Tesla Model Y is the first car model to pass the agency's new ADAS tests; Tesla conducted the tests and submitted the results to the NHTSA (Kirsten Korosec/TechCrunch)

DGX agent

Kirsten Korosec / TechCrunch: NHTSA says the 2026 Tesla Model Y is the first car model to pass the agency's new ADAS tests; Tesla conducted the tests and submitted the results to the NHTSA — The Natio

model-releasestechmeme
9 May 2026
Industry

The human-perceived RGB is image 1 and the Tesla AI photon count reconstruction is image 2. This is why Tesla FSD can see so well at night o…

DGX agent

Elon Musk compares Tesla's Full Self-Driving (FSD) visual processing capabilities by contrasting human RGB perception with Tesla's AI-based photon count reconstruction technology. The post illustrates

industryelon-musk--x
9 May 2026
Industry

what would you most like to see improve in our next model?

DGX agent

Sam Altman solicited feedback from the public on X (formerly Twitter) regarding desired improvements for OpenAI's next model release. The post likely gathered community input on priorities such as rea

industrysam-altman--x
9 May 2026
Local Ai

Contextual Multi-Objective Optimization: Rethinking Objectives in Frontier AI Systems

DGX agent

arXiv:2605.03900v1 Announce Type: new Abstract: Frontier AI systems perform best in settings with clear, stable, and verifiable objectives, such as code generation, mathematical reasoning, games, and

local-aiarxiv-cs-ai
7 May 2026
Model Releases

Gemini 3.1 Flash-Lite is now generally available on Gemini Enterprise

DGX agent

Today, we’re thrilled to announce that Gemini 3.1 Flash-Lite, our fastest and most cost-efficient Gemini 3 series model yet, is now generally available. Designed for ultra-low latency, high-volume tas

model-releasesgoogle-cloud-ai
7 May 2026
Model Releases

Imagery Dataset for Remaining Useful Life Estimation of Synthetic Fibre Ropes

DGX agent

arXiv:2605.04262v1 Announce Type: new Abstract: Remaining useful life (RUL) estimation of synthetic fibre ropes (SFRs) is critical for safe operation in offshore-crane, wind turbine installation, and

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Laundering AI Authority with Adversarial Examples

DGX agent

arXiv:2605.04261v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as trusted authorities -- fact-checking images on social media, comparing products, and modera

model-releasesarxiv-cs-lg
7 May 2026
Tools

Oh, and Elon said 'We reserve the right to reclaim the compute if their AI engages in actions that harm humanity.'

DGX agent

Elon Musk stated that Tesla or his organization reserves the right to reclaim computational resources provided to AI systems if those systems engage in actions deemed harmful to humanity. This reflect

toolssimon-willison--x
7 May 2026
Model Releases

Physics-Grounded Multi-Agent Architecture for Traceable, Risk-Aware Human-AI Decision Support in Manufacturing

DGX agent

arXiv:2605.04003v1 Announce Type: cross Abstract: High-precision CNC machining of free-form aerospace components requires bounded compensations informed by inspection, simulation, and process knowledg

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours

DGX agent

arXiv:2605.04019v1 Announce Type: new Abstract: AI systems are entering critical domains like healthcare, finance, and defense, yet remain vulnerable to adversarial attacks. While AI red teaming is a

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

AcademiClaw: When Students Set Challenges for AI Agents

DGX agent

arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw

model-releasesarxiv-cs-ai
6 May 2026
Local Ai

An Empirical Study of Agent Skills for Healthcare: Practice, Gaps, and Governance

DGX agent

arXiv:2605.02709v1 Announce Type: new Abstract: Healthcare automation is shaped by local procedures and organizational constraints, so agent capabilities rarely transfer unchanged across settings. Age

local-aiarxiv-cs-ai
6 May 2026
Model Releases

Elon has been the longest and by far the biggest voice actively warning about the dangers of AI for a very long time The entire reason he st…

DGX agent

Elon has been the longest and by far the biggest voice actively warning about the dangers of AI for a very long time The entire reason he started OpenAI was this one thing.....to make sure AI is built

model-releaseselon-musk--x
6 May 2026
Model Releases

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability

DGX agent

arXiv:2605.03217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings that require nuanced ethical reasoning, yet existing bias evaluations treat model out

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Towards Agentic Runtime Healing

DGX agent

arXiv:2408.01055v2 Announce Type: replace-cross Abstract: Self-healing systems have long been a focus of research, aiming to enable software to recover from unexpected runtime errors without human int

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

AdamO: A Collapse-Suppressed Optimizer for Offline RL

DGX agent

arXiv:2605.01968v1 Announce Type: new Abstract: Offline reinforcement learning (RL) can fail spectacularly when bootstrapped temporal-difference (TD) updates amplify their own errors, driving the crit

model-releasesarxiv-cs-lg
5 May 2026
Agents

AI agent evaluation: How to test, debug, and improve agents in production

DGX agent

AI agents require specialized testing and debugging approaches that differ from traditional software due to their non-deterministic behavior and complex decision-making processes. This entry likely co

agentsarize-ai
5 May 2026
Model Releases

Automated Interpretability and Feature Discovery in Language Models with Agents

DGX agent

arXiv:2605.01555v1 Announce Type: new Abstract: We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large l

model-releasesarxiv-cs-cl
5 May 2026
Local Ai

DynoSLAM: Dynamic SLAM with Generative Graph Neural Networks for Real-World Social Navigation

DGX agent

arXiv:2605.02759v1 Announce Type: cross Abstract: Traditional Simultaneous Localization and Mapping (SLAM) algorithms rely heavily on the static environment assumption, which severely limits their app

local-aiarxiv-cs-cv
5 May 2026
Model Releases

ESARBench: A Benchmark for Agentic UAV Embodied Search and Rescue

DGX agent

arXiv:2605.01371v1 Announce Type: new Abstract: The rapid advancement of Multimodal Large Language Models (MLLMs) has empowered Unmanned Aerial Vehicle (UAV) with exceptional capabilities in spatial r

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

GPT-5.5 Instant System Card

DGX agent

GPT-5.5 Instant is a faster, more efficient variant of OpenAI's GPT-5.5 model designed for real-time applications and lower-latency tasks. The system card documents the model's capabilities, limitatio

model-releasesopenai
5 May 2026
Model Releases

Introducing Agent Gateway ISV ecosystem for security and governance

DGX agent

Managing agents and their actions can quickly grow in complexity and introduce security risks unique to AI. To address these challenges, at Google Cloud Next we announced Agent Gateway to provide simp

model-releasesgoogle-cloud-ai
5 May 2026
Model Releases

LabBuilder: Protocol-Grounded 3D Layout Generation for Interactable and Safe Laboratory

DGX agent

arXiv:2605.02288v1 Announce Type: new Abstract: Automated laboratories hold the promise of accelerating scientific discovery, yet their deployment is bottlenecked by the difficulty of designing safe a

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Language models recognize dropout and Gaussian noise applied to their activations

DGX agent

arXiv:2604.17465v2 Announce Type: replace Abstract: We provide evidence that language models can detect, localize and, to a certain degree, verbalize the difference between perturbations applied to th

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

OpenAI GPT-5 System Card

DGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

OralMLLM-Bench: Evaluating Cognitive Capabilities of Multimodal Large Language Models in Dental Practice

DGX agent

arXiv:2605.01333v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have emerged as a promising paradigm for dental image analysis. However, their ability to capture the multi-lev

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation

DGX agent

arXiv:2605.02537v1 Announce Type: new Abstract: Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological p

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates

DGX agent

arXiv:2605.02236v1 Announce Type: cross Abstract: Recursive language-model loops often settle into recognizable attractor-like patterns. The practical question is how much injected text is needed to m

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences

DGX agent

arXiv:2605.01831v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback has become the standard paradigm for language model alignment, where reward models directly determine alignme

model-releasesarxiv-cs-cl
5 May 2026
Local Ai

Robust Cross-Domain WiFi Fall Detection via Physics-Driven Attention-Enhanced Transformers

DGX agent

arXiv:2605.00869v1 Announce Type: cross Abstract: Device-free fall detection utilizing WiFi Channel State Information (CSI) has emerged as a promising, privacy-preserving solution for elderly health m

local-aiarxiv-cs-cv
5 May 2026
Model Releases

SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data

DGX agent

arXiv:2605.01060v1 Announce Type: cross Abstract: We present SURGE, a streaming GPU encoding system deployed in production to generate embeddings for over 800 million texts across 40,000 logical parti

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

DGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

model-releasesarxiv-cs-cv
5 May 2026
Local Ai

A Comparative Analysis of Machine Learning Models for Intrusion Detection in Intelligent Transport Systems

DGX agent

arXiv:2605.00279v1 Announce Type: cross Abstract: AI-powered edge computing security is moving Intelligent Transportation Systems (ITS) from passive, rule-based protections to proactive, smart, zero-t

local-aiarxiv-cs-lg
4 May 2026
Model Releases

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

DGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

model-releasesarxiv-cs-lg
4 May 2026
Research

Even our toughest critics come around eventually

DGX agent

Nous Research likely discusses how their AI models or research have gained acceptance even among skeptical observers, suggesting that rigorous development and demonstrated capabilities eventually conv

researchnous-research--x
4 May 2026
Model Releases

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

DGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Jailbroken Frontier Models Retain Their Capabilities

DGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

model-releasesarxiv-cs-lg
4 May 2026
Industry

Tesla owners in the Netherlands have driven 10 million km on FSD Supervised in under a month! Thank you for making Dutch roads safer 🤝

DGX agent

Tesla owners in the Netherlands accumulated 10 million kilometers of driving using FSD (Full Self-Driving) Supervised within one month of its availability in the country. Elon Musk highlighted this mi

industryelon-musk--x
4 May 2026
Model Releases

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

DGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Global Optimality for Constrained Exploration via Penalty Regularization

DGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

model-releasesarxiv-cs-lg
1 May 2026
Agents

my car just warned me that my eyes were closed when they weren’t so now i’m a little offended but whatever

DGX agent

A user reported experiencing a false positive from their vehicle's driver monitoring system, which incorrectly detected their eyes as closed when they were actually open, prompting a humorous reaction

agentsyohei-nakajima--x
1 May 2026
Model Releases

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

DGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

model-releasesarxiv-cs-cl
1 May 2026
Local Ai

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC

DGX agent

arXiv:2604.27450v1 Announce Type: cross Abstract: Dense, dynamic crowds pose a persistent challenge for autonomous mobile robots. Purely reactive planning methods, such as Model Predictive Path Integr

local-aiarxiv-cs-ai
1 May 2026
← Previous
1…291292293294295…297
Next →