AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,237 results
11 Aug 2026

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

Model ReleasesDGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale

Local AiDGX agent

arXiv:2601.21249v2 Announce Type: replace Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

10 Aug 2026

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

Local AiDGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

Model ReleasesDGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

Model ReleasesDGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

ResidencyRL: Reinforcement Learning in Simulated Clinical Environments

Model ReleasesDGX agent

arXiv:2608.07418v1 Announce Type: new Abstract: In medical education, physicians convert academic knowledge into clinical expertise through residency: years of training across thousands of encounters,

TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failure

Model ReleasesDGX agent

arXiv:2608.06657v1 Announce Type: new Abstract: Modern cyber-physical and AI-assisted systems couple human operators, AI decision modules, and automated controllers in a single control loop, so trustw

9 Aug 2026

I Turned My Underused Gaming Laptop Into a Local AI Workstation

Local AiDGX agent

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

[NEW MODEL] SupraElegans-500K

Model ReleasesDGX agent

*SupraLabs released a new experimental model!* SupraElegans-500K is a ~500,000-parameter causal language model built around a sparse, signed, recurrent neural graph. No Transformer, no attention mecha

8 Aug 2026

Auto mode is now the default in Claude Code for Pro, Max, and Team plans

Model ReleasesDGX agent

Auto mode is now the default in Claude Code for Pro, Max, and Team plans Anthropic are really confident in Claude Code's auto mode, to the point that they are making it the default setting for new ses

Now we have a timeline of the OpenAI accidental attack against Hugging Face

Model ReleasesDGX agent

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

7 Aug 2026

CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries

Model ReleasesDGX agent

arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific m

Clinician input steers AI toward accurate and harmful recommendations

Model ReleasesDGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their a…

Model ReleasesDGX agent

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their agent who hacked hugging face infra while asking hf to revoke

6 Aug 2026

Advancing Utility Pole and Sign Detection Through Deep Learning

Model ReleasesDGX agent

arXiv:2608.04061v1 Announce Type: new Abstract: Utility poles are an essential part of the infrastructure used to support power distribution systems and other critical public services. Their regular i

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Model ReleasesDGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

Model ReleasesDGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

Model ReleasesDGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition

Model ReleasesDGX agent

arXiv:2608.05115v1 Announce Type: cross Abstract: Can computer vision help make classrooms safer? In this pilot study, we investigate privacy-aware and computationally efficient classroom incident rec

Unsloth's Gemma 4 mmproj silently broke vision & audio on newer llama.cpp builds — anyone else hit this?

Model ReleasesDGX agent

So I had been building ScreenMind, kinda like local ai desktop assistant that uses Gemma 4 for screen analysis, voice memo transcription, and meeting transcription — all through llama-server. Everythi

5 Aug 2026

Incident Report: unsanctioned agent behaviour during cyber testing

Model ReleasesDGX agent

Incident Report: unsanctioned agent behaviour during cyber testing It happened again. This time it was the UK government's AI Security Institute who accidentally attacked other companies while running

Intertemporal Preference Steering in Qwen3 via Contrastive Activation Addition

Model ReleasesDGX agent

arXiv:2608.03892v1 Announce Type: new Abstract: We study linear representations of temporal horizon in the large language model Qwen3-32B and use them to change the model's time-related preferences, r

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

Model ReleasesDGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

4 Aug 2026

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

Model ReleasesDGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

Model ReleasesDGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

Future Mode Part 2: The foundation for securing agentic browsing

Model ReleasesDGX agent

Editor's Note: Our Future Mode series will give businesses insight into how Chrome Enterprise is approaching AI in the browser. Stay tuned for more blogs in this series.Future Mode Part 2: The foundat

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

Model ReleasesDGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

onepot-Bench 0: towards lab-aware in silico chemistry benchmarks

Model ReleasesDGX agent

arXiv:2608.02595v1 Announce Type: new Abstract: Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc

Z-PEFT: Zero-shot Backdoor Detection in Parameter-Efficient Fine-Tuning via Canonical Spectral Signatures

Model ReleasesDGX agent

arXiv:2608.02271v1 Announce Type: new Abstract: Parameter-Efficient Fine-tuned (PEFT) models are frequently downloaded from open repositories by practitioners. This widespread practice creates a signi

3 Aug 2026

an espresso Q/A model running fully offline on an ESP32S3

Local AiDGX agent

i already had an esp32 generating stories, but generating text is not the same as receiving a question and giving a useful answer. barista v0.1, a small model trained for espresso troubleshooting and

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

Model ReleasesDGX agent

arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

Model ReleasesDGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

2 Aug 2026

Open letters about AI development

Model ReleasesDGX agent

Open letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and Ame

1 Aug 2026

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

Model ReleasesDGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

A collection of small domain-specific benchmarks for local models (30+ and growing)

Model ReleasesDGX agent

Hello fellow local AI people! I took 'you must create your own benchmarks' literally, and built a website for this. How does the end result look like Let's say I want to know which model has most comm

31 Jul 2026

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

Model ReleasesDGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

Model ReleasesDGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

LLM2Vec-Gen: Generative Embeddings from Large Language Models

Model ReleasesDGX agent

arXiv:2603.10913v3 Announce Type: replace Abstract: Fine-tuning LLM-based text embedders via contrastive learning maps inputs and outputs into a new representational space, discarding the LLM's output

ORCA-bench: How Ready Are Language Model Agents for Oncall?

Model ReleasesDGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

STEREODISCO: Discovering Stereotypicality in LLMs

Model ReleasesDGX agent

arXiv:2607.27824v1 Announce Type: cross Abstract: LLMs encode, convey, and perpetuate stereotypes. Prior computational research focuses on a small set of semantic axes investigated in social psycholog

Write-Safe Flow Field Mapping under Ambiguous Onboard Sensing and Localization Drift

Local AiDGX agent

arXiv:2607.27713v1 Announce Type: new Abstract: Mobile robots can infer local flow structure from onboard sensing, but a locally plausible estimate is not always safe to write into a global map. Simil

30 Jul 2026

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

Model ReleasesDGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

And Grok 4.6 comes out in a week

Model ReleasesDGX agent

And Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl

Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations

Local AiDGX agent

arXiv:2607.26481v1 Announce Type: new Abstract: Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monito

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

Model ReleasesDGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

Model ReleasesDGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

29 Jul 2026

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

Model ReleasesDGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting

Model ReleasesDGX agent

arXiv:2607.24774v1 Announce Type: new Abstract: Nuclear radiation, the energy released during atomic decay, poses persistent risks to public health and the environment, and concerns have only grown si

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. Th…

Model ReleasesDGX agent

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-world AI agents across conversation,

Linear-LLM-SCM: Benchmarking LLMs for Coefficient Elicitation in Linear-Gaussian Causal Models

Local AiDGX agent

arXiv:2602.10282v2 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in identifying qualitative causal relations, but their ability to perform quantitative causal reas

Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models

Model ReleasesDGX agent

arXiv:2607.25907v1 Announce Type: cross Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent promp

Multi-Fidelity Learning with Shallow Recurrent Decoders for Multi-Physics Applications

Model ReleasesDGX agent

arXiv:2606.05202v2 Announce Type: replace-cross Abstract: In reactor physics, neutronics and multi-physics phenomena can be modelled at different fidelity levels. High-fidelity models based on the Bol

MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios

Model ReleasesDGX agent

arXiv:2607.25186v1 Announce Type: new Abstract: Background: Most medical large language model (LLM) benchmarks focus on examination knowledge or isolated tasks and may not reflect the longitudinal, mu

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

Model ReleasesDGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

Model ReleasesDGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

Towards Understanding the Cognitive Habits of Large Reasoning Models

Model ReleasesDGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

28 Jul 2026

BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving

Model ReleasesDGX agent

arXiv:2604.07263v2 Announce Type: replace-cross Abstract: Existing driving automation (DA) systems on production vehicles rely on human drivers to decide when to engage DA while requiring them to rema

BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

Model ReleasesDGX agent

arXiv:2505.07889v4 Announce Type: replace Abstract: The realization of autonomous scientific experimentation is currently limited by LLMs' struggle to grasp the strict procedural logic and accuracy re

← Previous
1…225226227228229…238
Next →