AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,202 results
Model Releases

Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents

DGX agent

arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages update dynamically or hide important structure be

model-releasesarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

REAP: Relation-Aware Elicitation and Parsing for Closed-Book Knowledge Base Construction from LLMs

DGX agent

arXiv:2608.10963v1 Announce Type: new Abstract: We present the REAP system for the AKBC Shared Task 2026 on constructing knowledge bases from language models in a closed-book setting, subject to a bud

model-releasesarxiv-cs-cl
12 Aug 2026
Local Ai

Agentic AI-driven Immersive Simulation: A Knowledge-Aware Virtual Training Platform forHigh Dose Rate (HDR) Brachytherapy

DGX agent

arXiv:2608.08163v1 Announce Type: new Abstract: The convergence of the Metaverse and Large Language Model (LLM)-based AI agent is catalyzing a shift toward autonomous, immersive, and personalized peda

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms

DGX agent

arXiv:2508.16481v3 Announce Type: replace Abstract: Ensuring the safe use of agentic systems requires a thorough understanding of the range of malicious behaviors these systems may exhibit. In this pa

model-releasesarxiv-cs-lg
11 Aug 2026
Local Ai

Beyond Hazard Resemblance: Contrastive Event Adjudication for Training-Free Video Anomaly Detection

DGX agent

arXiv:2608.09908v1 Announce Type: new Abstract: Video anomaly detection (VAD) aims to identify and temporally localize abnormal events in videos. Supervised methods learn anomaly decision boundaries f

local-aiarxiv-cs-cv
11 Aug 2026
Model Releases

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

DGX agent

arXiv:2608.07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primari

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Large Multimodal Agents for Intelligent Transportation Systems: Architectures, Evidence, and Deployment Challenges

DGX agent

arXiv:2608.08184v1 Announce Type: new Abstract: Large multimodal agents (LMAs) are increasingly proposed for intelligent transportation systems (ITS), but existing studies often conflate multimodality

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

DGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Open-World Hierarchical Perception: Taxonomic Abstraction over Class-Agnostic Proposals for the Safe Handling of Out-of-Vocabulary Road Objects

DGX agent

arXiv:2608.07577v1 Announce Type: cross Abstract: A closed-set detector for autonomous driving must assign every object one of a fixed set of labels. On an object outside that set (a horse-drawn carri

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

DGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale

DGX agent

arXiv:2601.21249v2 Announce Type: replace Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

DGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

DGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

local-aiarxiv-cs-ai
10 Aug 2026
Model Releases

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

DGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

ResidencyRL: Reinforcement Learning in Simulated Clinical Environments

DGX agent

arXiv:2608.07418v1 Announce Type: new Abstract: In medical education, physicians convert academic knowledge into clinical expertise through residency: years of training across thousands of encounters,

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failure

DGX agent

arXiv:2608.06657v1 Announce Type: new Abstract: Modern cyber-physical and AI-assisted systems couple human operators, AI decision modules, and automated controllers in a single control loop, so trustw

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries

DGX agent

arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific m

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Advancing Utility Pole and Sign Detection Through Deep Learning

DGX agent

arXiv:2608.04061v1 Announce Type: new Abstract: Utility poles are an essential part of the infrastructure used to support power distribution systems and other critical public services. Their regular i

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

DGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

DGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition

DGX agent

arXiv:2608.05115v1 Announce Type: cross Abstract: Can computer vision help make classrooms safer? In this pilot study, we investigate privacy-aware and computationally efficient classroom incident rec

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Intertemporal Preference Steering in Qwen3 via Contrastive Activation Addition

DGX agent

arXiv:2608.03892v1 Announce Type: new Abstract: We study linear representations of temporal horizon in the large language model Qwen3-32B and use them to change the model's time-related preferences, r

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

DGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

DGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

DGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

DGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

onepot-Bench 0: towards lab-aware in silico chemistry benchmarks

DGX agent

arXiv:2608.02595v1 Announce Type: new Abstract: Language models are playing an increasingly important role in laboratory science, performing tasks such as experiment planning, execution, and post-hoc

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Z-PEFT: Zero-shot Backdoor Detection in Parameter-Efficient Fine-Tuning via Canonical Spectral Signatures

DGX agent

arXiv:2608.02271v1 Announce Type: new Abstract: Parameter-Efficient Fine-tuned (PEFT) models are frequently downloaded from open repositories by practitioners. This widespread practice creates a signi

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

DGX agent

arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

DGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

DGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM2Vec-Gen: Generative Embeddings from Large Language Models

DGX agent

arXiv:2603.10913v3 Announce Type: replace Abstract: Fine-tuning LLM-based text embedders via contrastive learning maps inputs and outputs into a new representational space, discarding the LLM's output

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

STEREODISCO: Discovering Stereotypicality in LLMs

DGX agent

arXiv:2607.27824v1 Announce Type: cross Abstract: LLMs encode, convey, and perpetuate stereotypes. Prior computational research focuses on a small set of semantic axes investigated in social psycholog

model-releasesarxiv-cs-lg
31 Jul 2026
Local Ai

Write-Safe Flow Field Mapping under Ambiguous Onboard Sensing and Localization Drift

DGX agent

arXiv:2607.27713v1 Announce Type: new Abstract: Mobile robots can infer local flow structure from onboard sensing, but a locally plausible estimate is not always safe to write into a global map. Simil

local-aiarxiv-cs-ro
31 Jul 2026
Local Ai

Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations

DGX agent

arXiv:2607.26481v1 Announce Type: new Abstract: Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monito

local-aiarxiv-cs-lg
30 Jul 2026
Model Releases

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

DGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting

DGX agent

arXiv:2607.24774v1 Announce Type: new Abstract: Nuclear radiation, the energy released during atomic decay, poses persistent risks to public health and the environment, and concerns have only grown si

model-releasesarxiv-cs-ai
29 Jul 2026
Local Ai

Linear-LLM-SCM: Benchmarking LLMs for Coefficient Elicitation in Linear-Gaussian Causal Models

DGX agent

arXiv:2602.10282v2 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in identifying qualitative causal relations, but their ability to perform quantitative causal reas

local-aiarxiv-cs-lg
29 Jul 2026
Model Releases

Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models

DGX agent

arXiv:2607.25907v1 Announce Type: cross Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent promp

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Multi-Fidelity Learning with Shallow Recurrent Decoders for Multi-Physics Applications

DGX agent

arXiv:2606.05202v2 Announce Type: replace-cross Abstract: In reactor physics, neutronics and multi-physics phenomena can be modelled at different fidelity levels. High-fidelity models based on the Bol

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios

DGX agent

arXiv:2607.25186v1 Announce Type: new Abstract: Background: Most medical large language model (LLM) benchmarks focus on examination knowledge or isolated tasks and may not reflect the longitudinal, mu

model-releasesarxiv-cs-cl
29 Jul 2026
← Previous
1…242243244245246…255
Next →