AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,234 results
Model Releases

When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

DGX agent

arXiv:2602.08995v2 Announce Type: replace Abstract: Computer-use agents (CUAs) have made tremendous progress in the past year, yet they still frequently produce misaligned actions that deviate from th

model-releasesarxiv-cs-cl
26 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

DGX agent

arXiv:2606.25487v1 Announce Type: new Abstract: Almost every paper on LLM jailbreaks and prompt injection reports an attack-success rate (ASR), and that number is assigned not by people but by an auto

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

DGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

The 4/elta Bound: Designing Predictable LLM-Verifier Systems for Formal Method Guarantee

DGX agent

arXiv:2512.02080v3 Announce Type: replace-cross Abstract: The integration of Formal Verification tools with Large Language Models (LLMs) offers a path to scale software verification beyond manual work

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding

DGX agent

arXiv:2606.25160v1 Announce Type: cross Abstract: The rapid rise of Vision-Language Models (VLMs) in egocentric visual understanding has made low-latency inference in human-robot collaborative (HRC) t

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Models and GUI Grounding Datasets

DGX agent

arXiv:2606.25760v1 Announce Type: cross Abstract: Computer-use agents turn vision-language model (VLM) predictions into executable GUI clicks, so reliable uncertainty estimates are essential for rejec

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

DGX agent

arXiv:2606.25182v1 Announce Type: new Abstract: Jailbreak attacks reveal a persistent weakness in aligned Large Language Models: carefully crafted prompts can elicit policy-violating responses despite

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

WOLF-VLA: Whole-Body Humanoid Optimal Locomotion Framework for Vision-Language-Action Learning

DGX agent

arXiv:2606.25591v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently demonstrated strong generalization in robotic manipulation, yet their applicability to whole-body, con

model-releasesarxiv-cs-ro
25 Jun 2026
Industry

Elon Musk denies Tesla’s Autopilot caused crash that killed grandmother

DGX agent

A Tesla Model 3 crashed into a residential home in Katy, Texas, killing a 76-year-old grandmother, with the driver claiming Autopilot was active at the time. Elon Musk disputed the claim, and Tesla's

industryars-technica
24 Jun 2026
Model Releases

ESBMC-PLC+: A Unified IEC~61131-3 Formal Verification Framework as a PLCverif Successor

DGX agent

arXiv:2606.23870v1 Announce Type: cross Abstract: PLCverif is the most mature open-source platform for PLC formal verification, developed at CERN and in production use since 2019. Yet it has two funda

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

How Much Can We Trust LLM Search Agents? Measuring Endorsement Vulnerability to Web Content Manipulation

DGX agent

arXiv:2606.16821v2 Announce Type: replace Abstract: Large language model (LLM)-based search agents synthesize open-web content into actionable recommendations on behalf of users, creating a risk that

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for Driving VLAs

DGX agent

arXiv:2606.23938v1 Announce Type: new Abstract: Driving VLA models incorporating Chain-of-Thought (CoT) reasoning are attractive because they leverage pretrained VLM representations and expose interme

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Poster: Exploring the Limits of Audio-Based Detection of Turkish Phone Call Scams

DGX agent

arXiv:2606.24523v1 Announce Type: cross Abstract: Scam phone calls exploit vulnerable communities worldwide, yet research on detection has focused almost exclusively on English and other high-resource

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

RASC+: Retrieval-Constrained LLM Adjudication for Clinical Value Set Authoring

DGX agent

arXiv:2606.23992v1 Announce Type: cross Abstract: Clinical value sets define the standardized terminology codes used in quality measurement, phenotyping, cohort construction, and clinical decision sup

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

T2D-Bench: Evidence-Gated Evaluation of LLM Outputs for Type 2 Diabetes Using a Multi-Layer Clinical-Lifestyle Knowledge Graph

DGX agent

arXiv:2606.24145v1 Announce Type: new Abstract: Large language models (LLMs) can produce clinically fluent recommendations for type 2 diabetes while failing to satisfy guideline constraints or explici

model-releasesarxiv-cs-ai
24 Jun 2026
Research

The best way to understand a complex system is via edge cases and failure modes, because they define the contour of the system.

DGX agent

Edge cases and failure modes reveal the fundamental boundaries and constraints of complex systems more effectively than typical operations, making them valuable for understanding system behavior and l

researchfrancois-chollet--x
24 Jun 2026
Model Releases

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

DGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Attacking the Trusted Imagination: Oracle-Level Integrity Attacks on Imagine-then-Act World Models

DGX agent

arXiv:2606.22966v1 Announce Type: new Abstract: Many recent vision-language-action (VLA) policies adopt an imagine-then-act design. A world-action model (WAM) first imagines a short future as a latent

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Hierarchical Adversarial Bandits for Online Configuration Optimization

DGX agent

arXiv:2505.19061v2 Announce Type: replace Abstract: Motivated by Online Configuration Optimization in large, dynamic parameter spaces, this work studies the nonstochastic multi-armed bandit (MAB) prob

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Large Language Model-Assisted Cleaning of Report-Derived Labels in a Large-Scale Chest CT Dataset

DGX agent

arXiv:2606.22382v1 Announce Type: cross Abstract: Purpose: To evaluate whether large language model (LLM)-assisted label cleaning can identify label-report discordance in CT-RATE, a large-scale public

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

Mind the Noise: Sensitivity of Transformer-based Interaction-Aware Trajectory Prediction Models to Noisy Data

DGX agent

arXiv:2606.21344v1 Announce Type: cross Abstract: Trajectory prediction allows autonomous vehicles to anticipate the future behavior of surrounding objects (or agents) and, accordingly, maximize the s

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

Open models, global networks: How AT&T and GSMA are accelerating telecom innovation with Gemma

DGX agent

Telecommunications is an incredibly complex, highly specialized domain. Modern mobile networks are inherently multi-vendor, featuring diverse and often proprietary data structures. While AI has made m

model-releasesgoogle-cloud-ai
23 Jun 2026
Model Releases

R2HandoverSim: A Simulation Framework and Benchmark for Robot-to-Human Object Handovers

DGX agent

arXiv:2606.21011v1 Announce Type: new Abstract: We present R2HandoverSim, a simulation benchmark for robot-to-human (R2H) object handovers. Although R2H handover methods have advanced rapidly, the lac

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

SparseWorld: Enhancing End-to-End Autonomous Driving via World Models with Sparse Scene Representation

DGX agent

arXiv:2605.24354v2 Announce Type: replace Abstract: Recently, world models have made significant progress in enhancing end-to-end driving systems through both future situation forecasting and improved

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Zero-Shot Vision-Language Models for Classroom Engagement Recognition: A Benchmark Study of Prompt Sensitivity and Cross-Dataset Generalization

DGX agent

arXiv:2606.21861v1 Announce Type: new Abstract: Automated classroom engagement recognition holds substantial promise for scalable learning analytics, yet the suitability of modern Vision-Language Mode

model-releasesarxiv-cs-cv
23 Jun 2026
Applications

1,250 hp hybrid Corvette shatters the Pikes Peak production record

DGX agent

Chevrolet's Corvette ZR1X, powered by a hybrid all-wheel drive system with 1,250 horsepower, set a new Pikes Peak production car record at the 104th running of the hill climb. Driven by IndyCar vetera

applicationsars-technica
22 Jun 2026
Model Releases

Daybreak: Tools for securing every organization in the world

DGX agent

Daybreak is an OpenAI initiative focused on developing and distributing security tools designed to protect organizations globally from cyber threats. The program aims to democratize access to advanced

model-releasesopenai
22 Jun 2026
Industry

GM installs robots at flagship EV factory after laying off 1,300 workers

DGX agent

General Motors temporarily laid off 1,300 workers at its Factory Zero EV plant due to sluggish demand for battery-powered vehicles , following the discontinuation of the $7,500 federal EV tax credit .

industryars-technica
22 Jun 2026
Industry

How Anthropic may have talked itself into an AI export ban

DGX agent

The US government ordered Anthropic to cut off foreign access to its Mythos 5 and Claude Fable 5 AI models in June 2026, citing national security concerns related to jailbreak vulnerabilities that cou

industryars-technica
22 Jun 2026
Tools

Red-Teaming after Mythos — Zico Kolter & Matt Fredrikson, Gray Swan

DGX agent

This episode likely discusses red-teaming methodologies and adversarial testing approaches in AI systems, featuring security researchers Zico Kolter and Matt Fredrikson from Carnegie Mellon University

toolslatent-space
22 Jun 2026
Model Releases

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

DGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Making Models Unmergeable via Scaling-Sensitive Loss Landscape

DGX agent

arXiv:2601.21898v2 Announce Type: replace Abstract: The rise of model hubs has made it easier to access reusable model components, making model merging a practical tool for combining capabilities. Yet

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

LLM-Based Code Documentation Generation and Multi-Judge Evaluation

DGX agent

arXiv:2606.09852v1 Announce Type: cross Abstract: High-quality source code documentation is vital yet often neglected, especially in critical domains like healthcare where reliability and maintainabil

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Uncertainty-Aware Motion Planning for Autonomous Driving in Mixed Traffic Environment

DGX agent

arXiv:2606.09958v1 Announce Type: cross Abstract: In mixed-traffic environments where autonomous and human-driven vehicles may co-exist, motion planning for autonomous vehicles requires anticipating t

model-releasesarxiv-cs-ai
10 Jun 2026
Local Ai

AI-Native Closed-Loop Security for 6G-Enabled Cyber-Physical Systems: From Edge Detection to Network-Wide Mitigation

DGX agent

arXiv:2606.08173v1 Announce Type: cross Abstract: In sixth-generation (6G) networks, billions of cyber-physical systems (CPSs) - autonomous vehicles, smart grids, industrial robots, and remote-surgica

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs (Matthias Bastian/The Decoder)

DGX agent

Matthias Bastian / The Decoder: Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs — Key Poin

model-releasestechmeme
9 Jun 2026
Model Releases

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

DGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word Subsets

DGX agent

arXiv:2606.08497v1 Announce Type: new Abstract: As deep language models (DLMs) are increasingly deployed in high-stakes domains such as healthcare, understanding their decision rationale becomes param

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Initial impressions of Claude Fable 5

DGX agent

I didn't have early access to today's Claude Fable 5 release, but I've spent the past ~5.5 hours putting it through its paces. My initial impressions are that this is something of a beast. It's slow,

model-releasessimon-willison
9 Jun 2026
Model Releases

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

DGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

DGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ProbeAct: Probe-Guided Training-Free Failure Recovery in Vision-Language-Action Models

DGX agent

arXiv:2606.09740v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate strong perfor-1 mance on language-conditioned robotic manipulation within their training dis-2 tribution

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

DGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Safe Polytope-in-Polytope Motion Planning and Control with Control Barrier Functions

DGX agent

arXiv:2606.09719v1 Announce Type: new Abstract: Autonomous mobile robots operating in tight environments require motion planning frameworks that account for the physical footprint of the robot. Simpli

local-aiarxiv-cs-ro
9 Jun 2026
Model Releases

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

DGX agent

arXiv:2603.22793v2 Announce Type: replace Abstract: Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Sound and Complete Neurosymbolic Reasoning with LLM-Grounded Interpretations

DGX agent

arXiv:2507.09751v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities in natural language understanding and generation, but exhibit problems with l

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

DGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…286287288289290…297
Next →