AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,202 results
Model Releases

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

DGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

model-releasesarxiv-cs-ai
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

DGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Understanding the Cognitive Habits of Large Reasoning Models

DGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving

DGX agent

arXiv:2604.07263v2 Announce Type: replace-cross Abstract: Existing driving automation (DA) systems on production vehicles rely on human drivers to decide when to engage DA while requiring them to rema

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

DGX agent

arXiv:2505.07889v4 Announce Type: replace Abstract: The realization of autonomous scientific experimentation is currently limited by LLMs' struggle to grasp the strict procedural logic and accuracy re

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

CallBench: A Benchmark for Dual-Goal Coordination in Phone Call Assistants

DGX agent

arXiv:2607.22635v1 Announce Type: new Abstract: Target-oriented dialogue systems have demonstrated strong capabilities in completing user goals through interactive conversations. However, existing stu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Child-Oriented AIGC Video Risk Reviewing: A Benchmark and Knowledge-Supported Iterative Reasoning Framework

DGX agent

arXiv:2607.22715v1 Announce Type: new Abstract: The rapid growth of Artificial Intelligence-generated content (AIGC) is reshaping video production and circulation, exposing children to an increasing v

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Context-Aware Concept Distillation for Trustworthy Flood Prediction

DGX agent

arXiv:2607.23237v1 Announce Type: cross Abstract: Effective flood risk management relies on accurate forecasting, yet the 'black box' nature of stateof-the-art Deep Learning models creates a barrier t

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Do Language Models Converge to Themselves? Recursive Self-Refinement as Textual Relaxation

DGX agent

arXiv:2607.22653v1 Announce Type: new Abstract: Large language models are increasingly used in recursive refinement workflows, where an initial draft is repeatedly revised by the same model. Despite t

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories

DGX agent

arXiv:2607.23704v1 Announce Type: cross Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irre

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Learning-based Hierarchical Tracheal Anatomy Understanding from Sparse Surgical Demonstration Annotations for Ultrasound Robots

DGX agent

arXiv:2607.22789v1 Announce Type: cross Abstract: Tracheostomy requires precise localization of the tracheal incision site; however, conventional manual palpation is subjective and often unreliable, w

local-aiarxiv-cs-cv
28 Jul 2026
Model Releases

MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents

DGX agent

arXiv:2607.23870v1 Announce Type: cross Abstract: Smart-city airspace is transforming Uncrewed Aerial Vehicles (UAVs) from passive sensing platforms into cyber-physical decision makers that must follo

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents

DGX agent

arXiv:2606.18037v2 Announce Type: replace Abstract: Tool-using LLM agents increasingly use the Model Context Protocol (MCP) to answer from heterogeneous evidence sources, including search, APIs, datab

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Reconstructing Item Characteristic Curves using Fine-Tuned Large Language Models

DGX agent

arXiv:2601.02580v2 Announce Type: replace-cross Abstract: Traditional methods for determining assessment item parameters, such as difficulty and discrimination, rely heavily on expensive field testing

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

DGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SymStep: Symbolic Step Verification for Logical Reasoning

DGX agent

arXiv:2607.23055v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

DGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Toward Automated Detection of Documentation Inconsistencies in Electronic Health Records

DGX agent

arXiv:2607.22954v1 Announce Type: new Abstract: Objective: To characterize the kinds of internal documentation inconsistencies a general-domain large language model (LLM) can surface from real-world d

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Trustworthy Medical Segmentation: Uncertainty-Aware U-Net Evaluation Under Clinical Image Degradation

DGX agent

arXiv:2607.22727v1 Announce Type: new Abstract: Medical image segmentation models often report high benchmark accuracy under ideal imaging conditions, yet their failures under clinical degradation can

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

UP-Fuse: Uncertainty-guided LiDAR-Camera Fusion for 3D Panoptic Segmentation

DGX agent

arXiv:2602.19349v2 Announce Type: replace-cross Abstract: LiDAR-camera fusion enhances 3D panoptic segmentation by leveraging camera images to complement sparse LiDAR scans, but it also introduces a c

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy

DGX agent

arXiv:2607.23006v1 Announce Type: cross Abstract: Scientific question answering requires a retrieval system to solve two distinct problems: identifying which papers are relevant and locating the suppo

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

DGX agent

arXiv:2607.21988v1 Announce Type: new Abstract: Self-harm content is particularly challenging to detect using NLP techniques, and is also a high-stakes task which requires the highest accuracy to enab

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents

DGX agent

arXiv:2607.22165v1 Announce Type: cross Abstract: LLM-based database agents show promise, but differing task scopes, testbeds, and metrics hinder comparison. We identify four gaps between evaluation a

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

DGX agent

arXiv:2607.22513v1 Announce Type: cross Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

The 3D Mirage: Probing and Taming 3D Hallucinations

DGX agent

arXiv:2512.15423v2 Announce Type: replace Abstract: Monocular depth foundation models achieve remarkable generalization by learning large-scale semantic priors, but this creates a critical vulnerabili

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants

DGX agent

arXiv:2607.20488v1 Announce Type: new Abstract: Multi-agent LLM frameworks typically fix their team topology at boot time. When an individual agent becomes overloaded at runtime, for example by mixing

model-releasesarxiv-cs-ai
24 Jul 2026
Local Ai

Concept Concentration for Faithful Representation Intervention

DGX agent

arXiv:2505.18672v2 Announce Type: replace Abstract: Representation intervention aims to localize and modify the representations that encode the underlying concepts in large language models (LLMs) to e

local-aiarxiv-cs-lg
24 Jul 2026
Local Ai

Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.20547v1 Announce Type: new Abstract: Safe Advanced Air Mobility operations require aircraft to maintain separation when surveillance information is noisy, delayed, incomplete, or temporaril

local-aiarxiv-cs-lg
24 Jul 2026
Model Releases

Geometric Configurations of Perturbed Jailbreak Prompts

DGX agent

arXiv:2607.20581v1 Announce Type: cross Abstract: Perturbation techniques that turn unsuccessful jailbreak prompts into successful ones are continuously evolving, constituting a major security threat

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

GuardianAgentBench: Where Agents Fail and How to Guard Them

DGX agent

arXiv:2607.20982v1 Announce Type: new Abstract: As large language model agents increasingly operate autonomously with access to tools and external environments, ensuring their safe and reliable behavi

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues

DGX agent

arXiv:2604.01925v2 Announce Type: replace-cross Abstract: Large Language Models increasingly suppress biased outputs when demographic identity is stated explicitly, yet may still exhibit implicit bias

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests

DGX agent

arXiv:2607.20759v1 Announce Type: cross Abstract: AI coding agents powered by LLMs are increasingly integrated into real-world software development, where they generate, edit, and execute code with au

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Refusal-Gated Decoding: Preserving Refusal Behavior Under High-Temperature Sampling

DGX agent

arXiv:2607.20791v1 Announce Type: new Abstract: High-temperature sampling is one of the primary mechanisms for increasing diversity in LLMs. Recent advances in truncation-based sampling techniques hav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Same Dangerous Objective, Opposite Advice: Direct Exposure versus Multi-Agent Mediation

DGX agent

arXiv:2607.21518v1 Announce Type: new Abstract: Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SkillCorpus: Consolidating and Evaluating the Open Skill Ecosystem for Real-World LLM Agents

DGX agent

arXiv:2607.15557v4 Announce Type: replace Abstract: Agent skills, SKILL files that package reusable procedural knowledge for an LLM agent, are a popular mechanism for extending agent capabilities. Pub

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

StabilityBench: Benchmarking Instability in LLMs

DGX agent

arXiv:2607.20558v1 Announce Type: cross Abstract: AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poor

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning

DGX agent

arXiv:2607.09328v2 Announce Type: replace-cross Abstract: Answering complex questions over long documents frequently requires integrating evidence that the source itself disperses naturally across dis

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems

DGX agent

arXiv:2607.19430v1 Announce Type: cross Abstract: Multi-agent LLM applications chain a planner, worker agents, a verifier, and a synthesizer, and every hop between agents is an unmonitored channel thr

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Emergent Autonomous Drifting for Collision Avoidance in Real-World Winter Driving Scenarios

DGX agent

arXiv:2607.19484v1 Announce Type: new Abstract: Real-world collision avoidance is a core motivation for studying the dynamics and control of high sideslip drifting in vehicles, yet the practical benef

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

DGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

JailMeter: An Evidence-Based Evaluation Framework for Jailbreak Attacks on Large Language Models

DGX agent

arXiv:2607.19424v1 Announce Type: cross Abstract: The assessment of jailbreak attacks against large language models currently suffers from inconsistent evaluation criteria and methods, leading to unre

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations

DGX agent

arXiv:2607.20379v1 Announce Type: new Abstract: Natural-language autoencoders score explanations of hidden activations by reconstruction: an explanation is deemed faithful if the activation can be reg

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training

DGX agent

arXiv:2607.19971v1 Announce Type: new Abstract: Accurate motion prediction of surrounding agents and safe motion planning are two closely coupled key tasks for social robot navigation in crowded envir

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

A Self-Evolving Agent for Longitudinal Personal Health Management

DGX agent

arXiv:2607.13940v1 Announce Type: new Abstract: Personal health management unfolds over repeated encounters, yet most health AI systems treat each request in isolation. We developed HealthClaw, an ope

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

DGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Efficient Text-to-Audio Generation via Pruning

DGX agent

arXiv:2607.13330v1 Announce Type: cross Abstract: Diffusion-based text-to-audio generative models such as AudioLDM achieve high perceptual quality and strong semantic consistency; however, their pract

model-releasesarxiv-cs-ai
16 Jul 2026
← Previous
1…243244245246247…255
Next →