AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,237 results
Model Releases

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

DGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

model-releasesarxiv-cs-cl
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

DGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Future Mode Part 2: The foundation for securing agentic browsing

DGX agent

Editor's Note: Our Future Mode series will give businesses insight into how Chrome Enterprise is approaching AI in the browser. Stay tuned for more blogs in this series.Future Mode Part 2: The foundat

model-releasesgoogle-cloud-ai
4 Aug 2026
Model Releases

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

DGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

onepot-Bench 0: towards lab-aware in silico chemistry benchmarks

DGX agent

arXiv:2608.02595v1 Announce Type: new Abstract: Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Z-PEFT: Zero-shot Backdoor Detection in Parameter-Efficient Fine-Tuning via Canonical Spectral Signatures

DGX agent

arXiv:2608.02271v1 Announce Type: new Abstract: Parameter-Efficient Fine-tuned (PEFT) models are frequently downloaded from open repositories by practitioners. This widespread practice creates a signi

model-releasesarxiv-cs-lg
4 Aug 2026
Local Ai

an espresso Q/A model running fully offline on an ESP32S3

DGX agent

i already had an esp32 generating stories, but generating text is not the same as receiving a question and giving a useful answer. barista v0.1, a small model trained for espresso troubleshooting and

local-air-localllama
3 Aug 2026
Model Releases

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

DGX agent

arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

DGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

DGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Open letters about AI development

DGX agent

Open letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and Ame

model-releasessimon-willison
2 Aug 2026
Model Releases

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

DGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

model-releasesitamar-friedman--x
1 Aug 2026
Model Releases

A collection of small domain-specific benchmarks for local models (30+ and growing)

DGX agent

Hello fellow local AI people! I took 'you must create your own benchmarks' literally, and built a website for this. How does the end result look like Let's say I want to know which model has most comm

model-releasesr-localllama
1 Aug 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM2Vec-Gen: Generative Embeddings from Large Language Models

DGX agent

arXiv:2603.10913v3 Announce Type: replace Abstract: Fine-tuning LLM-based text embedders via contrastive learning maps inputs and outputs into a new representational space, discarding the LLM's output

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

STEREODISCO: Discovering Stereotypicality in LLMs

DGX agent

arXiv:2607.27824v1 Announce Type: cross Abstract: LLMs encode, convey, and perpetuate stereotypes. Prior computational research focuses on a small set of semantic axes investigated in social psycholog

model-releasesarxiv-cs-lg
31 Jul 2026
Local Ai

Write-Safe Flow Field Mapping under Ambiguous Onboard Sensing and Localization Drift

DGX agent

arXiv:2607.27713v1 Announce Type: new Abstract: Mobile robots can infer local flow structure from onboard sensing, but a locally plausible estimate is not always safe to write into a global map. Simil

local-aiarxiv-cs-ro
31 Jul 2026
Model Releases

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

DGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

model-releasesperplexity--x
30 Jul 2026
Model Releases

And Grok 4.6 comes out in a week

DGX agent

And Grok 4.6 comes out in a week BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-worl

model-releaseselon-musk--x
30 Jul 2026
Local Ai

Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations

DGX agent

arXiv:2607.26481v1 Announce Type: new Abstract: Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monito

local-aiarxiv-cs-lg
30 Jul 2026
Model Releases

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

DGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting

DGX agent

arXiv:2607.24774v1 Announce Type: new Abstract: Nuclear radiation, the energy released during atomic decay, poses persistent risks to public health and the environment, and concerns have only grown si

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. Th…

DGX agent

BREAKING: Grok 4.5 ranked #1 on LaurenBench with a score of 56.9%, ahead of Claude Sonnet 5, GLM 5.2, Claude Opus 5, Kimi K3 and GPT-5.6. The benchmark tests real-world AI agents across conversation,

model-releaseselon-musk--x
29 Jul 2026
Local Ai

Linear-LLM-SCM: Benchmarking LLMs for Coefficient Elicitation in Linear-Gaussian Causal Models

DGX agent

arXiv:2602.10282v2 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in identifying qualitative causal relations, but their ability to perform quantitative causal reas

local-aiarxiv-cs-lg
29 Jul 2026
Model Releases

Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models

DGX agent

arXiv:2607.25907v1 Announce Type: cross Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent promp

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Multi-Fidelity Learning with Shallow Recurrent Decoders for Multi-Physics Applications

DGX agent

arXiv:2606.05202v2 Announce Type: replace-cross Abstract: In reactor physics, neutronics and multi-physics phenomena can be modelled at different fidelity levels. High-fidelity models based on the Bol

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios

DGX agent

arXiv:2607.25186v1 Announce Type: new Abstract: Background: Most medical large language model (LLM) benchmarks focus on examination knowledge or isolated tasks and may not reflect the longitudinal, mu

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

DGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

DGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Understanding the Cognitive Habits of Large Reasoning Models

DGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving

DGX agent

arXiv:2604.07263v2 Announce Type: replace-cross Abstract: Existing driving automation (DA) systems on production vehicles rely on human drivers to decide when to engage DA while requiring them to rema

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

DGX agent

arXiv:2505.07889v4 Announce Type: replace Abstract: The realization of autonomous scientific experimentation is currently limited by LLMs' struggle to grasp the strict procedural logic and accuracy re

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

CallBench: A Benchmark for Dual-Goal Coordination in Phone Call Assistants

DGX agent

arXiv:2607.22635v1 Announce Type: new Abstract: Target-oriented dialogue systems have demonstrated strong capabilities in completing user goals through interactive conversations. However, existing stu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Child-Oriented AIGC Video Risk Reviewing: A Benchmark and Knowledge-Supported Iterative Reasoning Framework

DGX agent

arXiv:2607.22715v1 Announce Type: new Abstract: The rapid growth of Artificial Intelligence-generated content (AIGC) is reshaping video production and circulation, exposing children to an increasing v

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Context-Aware Concept Distillation for Trustworthy Flood Prediction

DGX agent

arXiv:2607.23237v1 Announce Type: cross Abstract: Effective flood risk management relies on accurate forecasting, yet the 'black box' nature of stateof-the-art Deep Learning models creates a barrier t

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Detect early and enforce firmly with Google Cloud's enhanced cost controls for AI spend

DGX agent

Generative AI can make cloud costs difficult to predict. A single five-word prompt can run complex operations and generate significant costs. Traditional metrics like requests per second no longer hel

model-releasesgoogle-cloud-ai
28 Jul 2026
Model Releases

Do Language Models Converge to Themselves? Recursive Self-Refinement as Textual Relaxation

DGX agent

arXiv:2607.22653v1 Announce Type: new Abstract: Large language models are increasingly used in recursive refinement workflows, where an initial draft is repeatedly revised by the same model. Despite t

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories

DGX agent

arXiv:2607.23704v1 Announce Type: cross Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irre

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Learning-based Hierarchical Tracheal Anatomy Understanding from Sparse Surgical Demonstration Annotations for Ultrasound Robots

DGX agent

arXiv:2607.22789v1 Announce Type: cross Abstract: Tracheostomy requires precise localization of the tracheal incision site; however, conventional manual palpation is subjective and often unreliable, w

local-aiarxiv-cs-cv
28 Jul 2026
Model Releases

MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents

DGX agent

arXiv:2607.23870v1 Announce Type: cross Abstract: Smart-city airspace is transforming Uncrewed Aerial Vehicles (UAVs) from passive sensing platforms into cyber-physical decision makers that must follo

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents

DGX agent

arXiv:2606.18037v2 Announce Type: replace Abstract: Tool-using LLM agents increasingly use the Model Context Protocol (MCP) to answer from heterogeneous evidence sources, including search, APIs, datab

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Reconstructing Item Characteristic Curves using Fine-Tuned Large Language Models

DGX agent

arXiv:2601.02580v2 Announce Type: replace-cross Abstract: Traditional methods for determining assessment item parameters, such as difficulty and discrimination, rely heavily on expensive field testing

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

DGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…282283284285286…297
Next →