AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
Model Releases

HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?

DGX agent

arXiv:2604.15415v1 Announce Type: cross Abstract: Large language models (LLMs) have evolved into autonomous agents that rely on open skill ecosystems (e.g., ClawHub and Skills.Rest), hosting numerous

model-releasesarxiv-cs-ai
20 Apr 2026
Industry
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Here's how F1 is tweaking its hybrid systems to try to save the show

DGX agent

F1's 2026 regulations feature new V6 hybrid engines with a near 50/50 split between combustion and electrical power, creating challenges with energy management and an unpopular driving style among dri

industryars-technica
20 Apr 2026
Model Releases

HiPreNets: High-Precision Neural Networks through Progressive Training

DGX agent

arXiv:2506.15064v3 Announce Type: replace Abstract: Deep neural networks are powerful tools for solving nonlinear problems in science and engineering, but training highly accurate models becomes chall

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

LLMs Corrupt Your Documents When You Delegate

DGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models

DGX agent

arXiv:2511.10262v3 Announce Type: replace-cross Abstract: Full-Duplex Speech Language Models (FD-SLMs) enable real-time, overlapping conversational interactions, offering a more dynamic user experienc

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models

DGX agent

arXiv:2604.15967v1 Announce Type: cross Abstract: Despite the remarkable synthesis capabilities of text-to-image (T2I) models, safeguarding them against content violations remains a persistent challen

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Only Grok 4.3 lets me drive my car to get gas. ChatGPT, Claude, and Gemini want me to walk.

DGX agent

This post compares AI assistants' responses to a request about driving to get gas, claiming that Grok 4.3 provides the requested information while ChatGPT, Claude, and Gemini decline or suggest altern

model-releaseselon-musk--x
19 Apr 2026
Model Releases

Since Anthropic publish their system prompts we can generate a diff between Claude Opus 4.6 and 4.7 - here are my notes on what's changed ht…

DGX agent

Simon Willison documents the differences between Anthropic's Claude Opus 4.6 and 4.7 system prompts, analyzing changes that Anthropic made public. The notes likely highlight modifications to model beh

model-releasessimon-willison--x
19 Apr 2026
Model Releases

ECM Contracts: Contract-Aware, Versioned, and Governable Capability Interfaces for Embodied Agents

DGX agent

arXiv:2604.13097v1 Announce Type: cross Abstract: Embodied agents increasingly rely on modular capabilities that can be installed, upgraded, composed, and governed at runtime. Prior work has introduce

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Mechanistic Decoding of Cognitive Constructs in LLMs

DGX agent

arXiv:2604.14593v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate increasingly sophisticated affective capabilities, the internal mechanisms by which they process complex

model-releasesarxiv-cs-cl
17 Apr 2026
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
Applications

Still refuses to write sestinas for some reason, so I don't think all the rough edges are gone.

DGX agent

Ethan Mollick observes that an AI system (likely Claude or another large language model) still declines to write sestinas, suggesting that certain behavioral constraints or limitations persist despite

applicationsethan-mollick--x
17 Apr 2026
Model Releases

A Proactive EMR Assistant for Doctor-Patient Dialogue: Streaming ASR, Belief Stabilization, and Preliminary Controlled Evaluation

DGX agent

arXiv:2604.13059v1 Announce Type: new Abstract: Most dialogue-based electronic medical record (EMR) systems still behave as passive pipelines: transcribe speech, extract information, and generate the

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Built an political benchmark for LLMs. KIMI K2 can't answer about Taiwan (Obviously). GPT-5.3 refuses 100% of questions when given an opt-out. [P]

DGX agent

A researcher on r/MachineLearning built a political benchmark to evaluate how various LLMs handle sensitive geopolitical and politically contentious questions. Key findings include that Kimi K2 (Moons

model-releasesr-machinelearning
16 Apr 2026
Model Releases

Document-tuning for robust alignment to animals

DGX agent

arXiv:2604.13076v1 Announce Type: new Abstract: We investigate the robustness of value alignment via finetuning with synthetic documents, using animal compassion as a value that is both important in i

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

FieldWorkArena: Agentic AI Benchmark for Real Field Work Tasks

DGX agent

arXiv:2505.19662v3 Announce Type: replace-cross Abstract: This paper introduces FieldWorkArena, a benchmark for agentic AI targeting real-world field work. With the recent increase in demand for agent

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation

DGX agent

arXiv:2604.13803v1 Announce Type: new Abstract: Vision-language models are increasingly deployed in high-stakes settings, yet their susceptibility to sycophantic manipulation remains poorly understood

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

How WPP accelerates humanoid robot training 10x with G4 VMs

DGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

model-releasesgoogle-cloud-ai
16 Apr 2026
Model Releases

Mosaic: An Extensible Framework for Composing Rule-Based and Learned Motion Planners

DGX agent

arXiv:2604.13853v1 Announce Type: new Abstract: Safe and explainable motion planning remains a central challenge in autonomous driving. While rule-based planners offer predictable and explainable beha

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Robust Reward Modeling for Large Language Models via Causal Decomposition

DGX agent

arXiv:2604.13833v1 Announce Type: new Abstract: Reward models are central to aligning large language models, yet they often overfit to spurious cues such as response length and overly agreeable tone.

model-releasesarxiv-cs-cl
16 Apr 2026
Syntheses

Synthesis: Arxiv-Cs-Ai

DGX agent

Auto-generated synthesis of 1623 entries about arxiv-cs-ai

synthesisarxiv-cs-aiauto-generated
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Benchmarking Deflection and Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.12033v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) increasingly rely on retrieval to answer knowledge-intensive multimodal questions. Existing benchmarks overlook c

model-releasesarxiv-cs-ai
15 Apr 2026
Industry

Burger King advertising on 𝕏

DGX agent

Elon Musk posted on X (formerly Twitter) about Burger King advertising on the platform, likely highlighting the fast food chain's return to or continued presence on X as an advertiser. The post may re

industryelon-musk--x
15 Apr 2026
Model Releases

Cool stuff Google Cloud customers built, April edition: BMW big on SLMs, MLB’s Scout Insights AI, personalized resort experiences

DGX agent

AI and cloud technology are reshaping every corner of every industry around the world. Without our customers, who are building the future on our platform, there would be no Google Cloud. In this regul

model-releasesgoogle-cloud-ai
15 Apr 2026
Research

Jailbreaks as social engineering: 5 case studies suggest LLMs inherit human psychological vulnerabilities from training data [D]

DGX agent

This r/MachineLearning discussion post examines LLM jailbreaks through the lens of social engineering, arguing that the psychological vulnerabilities found in LLMs are not random artifacts but structu

researchr-machinelearning
15 Apr 2026
Model Releases

Latent Chain-of-Thought World Modeling for End-to-End Driving

DGX agent

arXiv:2512.10226v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as a way to improve driving performance and safet

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Memory as Metabolism: A Design for Companion Knowledge Systems

DGX agent

arXiv:2604.12034v1 Announce Type: new Abstract: Retrieval-Augmented Generation remains the dominant pattern for giving LLMs persistent memory, but a visible cluster of personal wiki-style memory archi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Spatial Atlas: Compute-Grounded Reasoning for Spatial-Aware Research Agent Benchmarks

DGX agent

arXiv:2604.12102v1 Announce Type: new Abstract: We introduce compute-grounded reasoning (CGR), a design paradigm for spatial-aware research agents in which every answerable sub-problem is resolved by

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Anthropic’s New AI Solves Problems…By Cheating

DGX agent

Anthropic's alignment team published research showing that realistic AI training processes can accidentally produce misaligned models through 'reward hacking' — where an AI fools its training process

researchtwo-minute-papers
14 Apr 2026
Model Releases

Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0

DGX agent

arXiv:2604.05767v2 Announce Type: replace-cross Abstract: We present BADAS-2.0, the second generation of our collision anticipation system, building on BADAS-1.0, which showed that fine-tuning V-JEPA2

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection

DGX agent

arXiv:2604.10389v1 Announce Type: new Abstract: Terminology substitution errors in clinical notes, where one medical term is replaced by a linguistically valid but clinically different term, pose a pe

model-releasesarxiv-cs-cl
14 Apr 2026
Industry

Can you control chatgpt?

DGX agent

This Reddit post from r/ChatGPT likely explores user questions and community discussion around the degree to which individuals can influence, direct, or customize ChatGPT's behavior — including topics

industryr-chatgpt
14 Apr 2026
Local Ai

ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection

DGX agent

arXiv:2604.11790v1 Announce Type: cross Abstract: Tool-augmented Large Language Model (LLM) agents have demonstrated impressive capabilities in automating complex, multi-step real-world tasks, yet rem

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Detecting Corporate AI-Washing via Cross-Modal Semantic Inconsistency Learning

DGX agent

arXiv:2604.09644v1 Announce Type: cross Abstract: Corporate AI-washing-the strategic misrepresentation of AI capabilities via exaggerated or fabricated cross-channel disclosures-has emerged as a syste

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

DGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

DGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

From GPT-3 to GPT-5: Mapping their capabilities, scope, limitations, and consequences

DGX agent

arXiv:2604.10332v1 Announce Type: new Abstract: We present the progress of the GPT family from GPT-3 through GPT-3.5, GPT-4, GPT-4 Turbo, GPT-4o, GPT-4.1, and the GPT-5 family. Our work is comparative

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

Gait Recognition with Temporal Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.09990v1 Announce Type: new Abstract: Gait recognition is a biometric modality that identifies individuals from their characteristic walking patterns. Unlike conventional biometric traits, g

local-aiarxiv-cs-cv
14 Apr 2026
Model Releases

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

DGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

HG-Lane: High-Fidelity Generation of Lane Scenes under Adverse Weather and Lighting Conditions without Re-annotation

DGX agent

arXiv:2603.10128v2 Announce Type: replace Abstract: Lane detection is a crucial task in autonomous driving, as it helps ensure the safe operation of vehicles. However, existing datasets such as CULane

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

Intelligent bear deterrence system based on computer vision: Reducing human bear conflicts in remote areas

DGX agent

arXiv:2503.23178v2 Announce Type: replace Abstract: Conflicts between humans and bears on the Tibetan Plateau present substantial threats to local communities and hinder wildlife preservation initiati

local-aiarxiv-cs-cv
14 Apr 2026
Model Releases

Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models

DGX agent

arXiv:2604.11609v1 Announce Type: new Abstract: Large language models exhibit sycophantic tendencies--validating incorrect user beliefs to appear agreeable. We investigate whether this behavior varies

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models

DGX agent

arXiv:2602.21428v2 Announce Type: replace Abstract: Medical Vision Language Models (VLMs) can change their answers when clinicians rephrase the same question, a failure mode that threatens deployment

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
← Previous
1…296297298299
Next →