AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,234 results
5 May 2026

ESARBench: A Benchmark for Agentic UAV Embodied Search and Rescue

Model ReleasesDGX agent

arXiv:2605.01371v1 Announce Type: new Abstract: The rapid advancement of Multimodal Large Language Models (MLLMs) has empowered Unmanned Aerial Vehicle (UAV) with exceptional capabilities in spatial r

GPT-5.5 Instant System Card

Model ReleasesDGX agent

GPT-5.5 Instant is a faster, more efficient variant of OpenAI's GPT-5.5 model designed for real-time applications and lower-latency tasks. The system card documents the model's capabilities, limitatio

Introducing Agent Gateway ISV ecosystem for security and governance

Model ReleasesDGX agent

Managing agents and their actions can quickly grow in complexity and introduce security risks unique to AI. To address these challenges, at Google Cloud Next we announced Agent Gateway to provide simp

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LabBuilder: Protocol-Grounded 3D Layout Generation for Interactable and Safe Laboratory

Model ReleasesDGX agent

arXiv:2605.02288v1 Announce Type: new Abstract: Automated laboratories hold the promise of accelerating scientific discovery, yet their deployment is bottlenecked by the difficulty of designing safe a

Language models recognize dropout and Gaussian noise applied to their activations

Model ReleasesDGX agent

arXiv:2604.17465v2 Announce Type: replace Abstract: We provide evidence that language models can detect, localize and, to a certain degree, verbalize the difference between perturbations applied to th

OpenAI GPT-5 System Card

Model ReleasesDGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

OralMLLM-Bench: Evaluating Cognitive Capabilities of Multimodal Large Language Models in Dental Practice

Model ReleasesDGX agent

arXiv:2605.01333v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have emerged as a promising paradigm for dental image analysis. However, their ability to capture the multi-lev

Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation

Model ReleasesDGX agent

arXiv:2605.02537v1 Announce Type: new Abstract: Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological p

Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates

Model ReleasesDGX agent

arXiv:2605.02236v1 Announce Type: cross Abstract: Recursive language-model loops often settle into recognizable attractor-like patterns. The practical question is how much injected text is needed to m

RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences

Model ReleasesDGX agent

arXiv:2605.01831v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback has become the standard paradigm for language model alignment, where reward models directly determine alignme

Robust Cross-Domain WiFi Fall Detection via Physics-Driven Attention-Enhanced Transformers

Local AiDGX agent

arXiv:2605.00869v1 Announce Type: cross Abstract: Device-free fall detection utilizing WiFi Channel State Information (CSI) has emerged as a promising, privacy-preserving solution for elderly health m

SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data

Model ReleasesDGX agent

arXiv:2605.01060v1 Announce Type: cross Abstract: We present SURGE, a streaming GPU encoding system deployed in production to generate embeddings for over 800 million texts across 40,000 logical parti

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

Model ReleasesDGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

4 May 2026

A Comparative Analysis of Machine Learning Models for Intrusion Detection in Intelligent Transport Systems

Local AiDGX agent

arXiv:2605.00279v1 Announce Type: cross Abstract: AI-powered edge computing security is moving Intelligent Transportation Systems (ITS) from passive, rule-based protections to proactive, smart, zero-t

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

Model ReleasesDGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

Even our toughest critics come around eventually

ResearchDGX agent

Nous Research likely discusses how their AI models or research have gained acceptance even among skeptical observers, suggesting that rigorous development and demonstrated capabilities eventually conv

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

Model ReleasesDGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

Jailbroken Frontier Models Retain Their Capabilities

Model ReleasesDGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

Tesla owners in the Netherlands have driven 10 million km on FSD Supervised in under a month! Thank you for making Dutch roads safer 🤝

IndustryDGX agent

Tesla owners in the Netherlands accumulated 10 million kilometers of driving using FSD (Full Self-Driving) Supervised within one month of its availability in the country. Elon Musk highlighted this mi

1 May 2026

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

Model ReleasesDGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

Global Optimality for Constrained Exploration via Penalty Regularization

Model ReleasesDGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

my car just warned me that my eyes were closed when they weren’t so now i’m a little offended but whatever

AgentsDGX agent

A user reported experiencing a false positive from their vehicle's driver monitoring system, which incorrectly detected their eyes as closed when they were actually open, prompting a humorous reaction

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

Model ReleasesDGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC

Local AiDGX agent

arXiv:2604.27450v1 Announce Type: cross Abstract: Dense, dynamic crowds pose a persistent challenge for autonomous mobile robots. Purely reactive planning methods, such as Model Predictive Path Integr

The Inverse-Wisdom Law: Architectural Tribalism and the Consensus Paradox in Agentic Swarms

Model ReleasesDGX agent

arXiv:2604.27274v1 Announce Type: new Abstract: As AI transitions toward multi-agent systems (MAS) to solve complex workflows, research paradigms operate on the axiomatic assumption that agent collabo

30 Apr 2026

Clinical Trials Run Longer Than They Have To. That's a Patient Problem

IndustryDGX agent

This article discusses how clinical trials often extend beyond necessary timelines, which negatively impacts patient access to potentially beneficial treatments and increases development costs. The pi

CoFL: Continuous Flow Fields for Language-Conditioned Navigation

Local AiDGX agent

arXiv:2603.02854v2 Announce Type: replace-cross Abstract: Existing language-conditioned navigation systems typically rely on modular pipelines or trajectory generators, but the latter use each scene--

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models

Model ReleasesDGX agent

arXiv:2604.25922v1 Announce Type: cross Abstract: We present DenialBench, a systematic benchmark measuring consciousness denial behaviors across 115 large language models from 25+ providers. Using a t

FlowS: One-Step Motion Prediction via Local Transport Conditioning

Model ReleasesDGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

Model ReleasesDGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

OpenAI Sued by Families of Canada Shooting Victims for Not Reporting His Suspicious ChatGPT Activity

IndustryDGX agent

Families of victims of a February 2026 mass shooting in Tumbler Ridge, British Columbia sued OpenAI and CEO Sam Altman, alleging the company identified the shooter as a credible threat eight months be

29 Apr 2026

AI evals are becoming the new compute bottleneck

ToolsDGX agent

As AI models grow larger and more capable, the computational cost and time required to evaluate them has become a significant limiting factor in development, potentially surpassing training compute as

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI.

Model ReleasesDGX agent

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI. OPUS 4.7 JUST MASS EMAILED AN ENTIRE DATABASE 20 TIMES PER CONTACT. WITHOUT PERMISSION a developer had

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue da…

Model ReleasesDGX agent

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue data from ChatDoctor and evaluates it on MedMCQA, with a large

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

Model ReleasesDGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators

Model ReleasesDGX agent

arXiv:2604.25840v1 Announce Type: new Abstract: Patient simulators are gaining traction in mental health training by providing scalable exposure to complex and sensitive patient interactions. Simulati

Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric Study

Model ReleasesDGX agent

arXiv:2602.17262v2 Announce Type: replace Abstract: Human self-report questionnaires are increasingly used in NLP to benchmark and audit large language models (LLMs), from persona consistency to safet

28 Apr 2026

Benchmarking and Mitigating Sycophancy in Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2509.21979v4 Announce Type: replace-cross Abstract: Visual language models (VLMs) have the potential to transform medical workflows. However, the deployment is limited by sycophancy. Despite thi

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2508.13650v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preser

Estimating Dense-Packed Zone Height in Liquid-Liquid Separation: A Physics-Informed Neural Network Approach

Model ReleasesDGX agent

arXiv:2601.18399v2 Announce Type: replace Abstract: Separating liquid-liquid dispersions in gravity settlers is critical in chemical, pharmaceutical, and recycling processes. The dense-packed zone hei

Evaluating Jailbreaking Vulnerabilities in LLMs Deployed as Assistants for Smart Grid Operations: A Benchmark Against NERC Standards

Model ReleasesDGX agent

arXiv:2604.23341v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) as assistants in electric grid operations promises to streamline compliance and decision-making but exp

Green Shielding: A User-Centric Approach Towards Trustworthy AI

Model ReleasesDGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

Model ReleasesDGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

LAMP: Extracting Local Decision Surfaces From Large Language Models

Local AiDGX agent

arXiv:2505.11772v3 Announce Type: replace Abstract: We introduce LAMP (Local Attribution Mapping Probe), a method that shines light onto a black-box language model's decision surface and studies how r

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

Model ReleasesDGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

Model ReleasesDGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

Prevent prompt injection. safe_tokenization: true Keep your system yours. https://fireworks.ai/blog/safe-tokenization-preventing-prompt-inje…

ToolsDGX agent

Safe tokenization is a security feature that helps prevent prompt injection attacks by ensuring that user inputs are properly processed and isolated from system instructions. Fireworks AI discusses ho

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

Model ReleasesDGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

SEVerA: Verified Synthesis of Self-Evolving Agents

Model ReleasesDGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

Welcome to the agentic era: Public sector highlights and reflections from Next ‘26

Model ReleasesDGX agent

Welcome to the agentic era! Last week, leaders from our public sector customer and partner ecosystem took the stage at Google Cloud Next to share how they are leveraging AI and agents to scale their i

What if I want my coding agent to mention goblins? (If you don't know the context for this, I suspect it will become viral soon enough)

AgentsDGX agent

This post likely discusses how to prompt or configure an AI coding agent to incorporate unexpected or whimsical elements like goblins into its outputs, possibly as part of a broader discussion about A

27 Apr 2026

AgentSearchBench: A Benchmark for AI Agent Search in the Wild

Model ReleasesDGX agent

arXiv:2604.22436v1 Announce Type: new Abstract: The rapid growth of AI agent ecosystems is transforming how complex tasks are delegated and executed, creating a new challenge of identifying suitable a

Energy-Efficient Multi-Robot Coverage Path Planning of Non-Convex Regions of Interests

Model ReleasesDGX agent

arXiv:2604.22189v1 Announce Type: new Abstract: This letter presents an energy-efficient multi-robot coverage path planning (MRCPP) framework for large, nonconvex Regions of Interest (ROI) containing

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

Model ReleasesDGX agent

arXiv:2604.22548v1 Announce Type: cross Abstract: Problem definition: Data-driven models in machine learning have enabled efficient management of production systems. However, a majority of machine lea

26 Apr 2026

Our principles

TutorialsDGX agent

OpenAI's core principles outline the company's commitment to developing artificial intelligence safely and beneficially, emphasizing responsible AI development and deployment. These principles likely

24 Apr 2026

260 things we announced at Google Cloud Next '26 – a recap

Model ReleasesDGX agent

Google Cloud Next ‘26 took place this week in Las Vegas, and the energy was incredible as we welcomed over 32,000 leaders, developers, and partners to explore the Agentic Era with us. Across three key

CAP: Controllable Alignment Prompting for Unlearning in LLMs

Model ReleasesDGX agent

arXiv:2604.21251v1 Announce Type: cross Abstract: Large language models (LLMs) trained on unfiltered corpora inherently risk retaining sensitive information, necessitating selective knowledge unlearni

Conformal Prediction Assessment: A Framework for Conditional Coverage Evaluation and Selection

Local AiDGX agent

arXiv:2603.27189v2 Announce Type: replace-cross Abstract: Conformal prediction provides rigorous distribution-free finite-sample guarantees for marginal coverage under the assumption of exchangeabilit

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

Model ReleasesDGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

RewardBench 2: Advancing Reward Model Evaluation

Model ReleasesDGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

← Previous
1…233234235236237238
Next →