AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,951 results
Model Releases

Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs

DGX agent

arXiv:2509.02372v3 Announce Type: replace-cross Abstract: Large Language Models have become critical to modern software development, but their reliance on uncurated web-scale datasets for training int

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

DGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

agentsarxiv-cs-cl
12 May 2026
Agents

Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments

DGX agent

arXiv:2605.09721v1 Announce Type: cross Abstract: Tool-enabled AI agents are increasingly deployed in cloud-hosted environments and offered as services, where they perform side-effecting operations th

agentsarxiv-cs-ai
12 May 2026
Model Releases

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

DGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

model-releasesarxiv-cs-ai
12 May 2026
Applications

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation

DGX agent

arXiv:2605.10079v1 Announce Type: new Abstract: Video generation has advanced rapidly, producing photorealistic videos from text or image prompts. Meanwhile, film production and social robotics increa

applicationsarxiv-cs-cv
12 May 2026
Research

Speech-based Psychological Crisis Assessment using LLMs

DGX agent

arXiv:2605.10027v1 Announce Type: cross Abstract: Psychological support hotlines provide critical support for individuals experiencing mental health emergencies, yet current assessments largely rely o

researcharxiv-cs-ai
12 May 2026
Local Ai

SymTorch: Symbolic Distillation of Neural Networks

DGX agent

arXiv:2602.21307v2 Announce Type: replace Abstract: What mathematical functions do neural network components learn? Symbolic distillation addresses this question by expressing neural network component

local-aiarxiv-cs-lg
12 May 2026
Agents

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

DGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

agentsarxiv-cs-ai
12 May 2026
Safety

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring

DGX agent

arXiv:2605.09225v1 Announce Type: cross Abstract: Jailbreak attacks -- adversarial prompts that bypass LLM alignment through purely linguistic manipulation -- pose a growing operational security threa

safetyarxiv-cs-ai
12 May 2026
Agents

the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc

DGX agent

The Mini Shai-Hulud attack targets emerging AI-assisted development workflows by compromising multiple integration points including continuous integration systems, code editor hooks, and AI agent conf

agentsyohei-nakajima--x
12 May 2026
Safety

Upholding Epistemic Agency: A Brouwerian Assertibility Constraint for Responsible AI

DGX agent

arXiv:2603.03971v2 Announce Type: replace-cross Abstract: Generative AI can convert uncertainty into hypersuasive, authoritative-seeming verdicts, displacing the justificatory work on which democratic

safetyarxiv-cs-ai
12 May 2026
Model Releases

VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation

DGX agent

arXiv:2605.08553v1 Announce Type: cross Abstract: Large language models can generate useful code from natural language, but their outputs come without correctness guarantees. Verifiable code generatio

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VFM-SDM: A vision foundation model-based framework for training-free, marker-free, and calibration-free structural displacement measurement

DGX agent

arXiv:2605.09677v1 Announce Type: new Abstract: Reliable displacement measurement is fundamental for structural health monitoring and digital engineering workflows, as it provides direct structural re

model-releasesarxiv-cs-cv
12 May 2026
Agents

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

DGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

agentsarxiv-cs-ai
12 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Safety

A Large-Scale Dataset for Molecular Structure-Language Description via a Rule-Regularized Method

DGX agent

arXiv:2602.02320v3 Announce Type: replace-cross Abstract: Molecular function is largely determined by structure. Accurately aligning molecular structure with natural language is therefore essential fo

safetyarxiv-cs-ai
11 May 2026
Research

A Linear-Transformer Hybrid for SNP-Based Genotype-to-Phenotype Prediction in Grapevine

DGX agent

arXiv:2605.06762v1 Announce Type: cross Abstract: Robust genotype-to-phenotype (G2P) prediction is essential for accelerating breeding decisions and genetic gain. However, it remains challenging to me

researcharxiv-cs-ai
11 May 2026
Agents

A Self-Healing Framework for Reliable LLM-Based Autonomous Agents

DGX agent

arXiv:2605.06737v1 Announce Type: cross Abstract: Autonomous agents based on Large Language Models (LLMs) are increasingly being utilized in complex software systems. However, reliability remains a si

agentsarxiv-cs-ai
11 May 2026
Research

A Unified Framework for the Detection and Classification of Fatty Pancreas in Ultrasound Images

DGX agent

arXiv:2605.07466v1 Announce Type: new Abstract: Non-alcoholic fatty pancreas disease (NAFPD) is an underdiagnosed condition associated with metabolic syndrome, insulin resistance, and increased risk o

researcharxiv-cs-cv
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Safety

AI existential crisis for software engineers is to go live in the woods and read poetry. AI existential crisis for creatives is to make thin…

DGX agent

AI existential crisis for software engineers is to go live in the woods and read poetry. AI existential crisis for creatives is to make things until 4am every day because you can't stop now. is this w

safetycristobal-valenzuela--x
11 May 2026
Agents

ATHENA: Agentic Team for Hierarchical Evolutionary Numerical Algorithms

DGX agent

arXiv:2512.03476v2 Announce Type: replace-cross Abstract: Bridging the gap between theoretical conceptualization and computational implementation is a major bottleneck in Scientific Computing (SciC) a

agentsarxiv-cs-ai
11 May 2026
Hardware

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

DGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

hardwarearxiv-cs-cl
11 May 2026
Research

Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation

DGX agent

arXiv:2510.07926v2 Announce Type: replace Abstract: Despite demonstrating remarkable performance across a wide range of tasks, large language models (LLMs) have also been found to frequently produce o

researcharxiv-cs-cl
11 May 2026
Agents

Computer use with any model Hermes Agent × @trycua

DGX agent

This post from Nous Research discusses the integration of computer use capabilities with Hermes Agent models in collaboration with Claude (CUA), enabling AI agents to interact with computer interfaces

agentsnous-research--x
11 May 2026
Model Releases

Deeply Dual Supervised learning for melanoma recognition

DGX agent

arXiv:2508.01994v2 Announce Type: replace Abstract: As the application of deep learning in dermatology continues to grow, the recognition of melanoma has garnered significant attention, demonstrating

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution

DGX agent

arXiv:2605.07208v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to brainstorm and evaluate research ideas, yet assessing such judgments is fundamentally difficult be

model-releasesarxiv-cs-lg
11 May 2026
Safety

Future-proof your data strategy: AlloyDB adds PostgreSQL 18 and new Extended Support

DGX agent

As you look out at your 2026 infrastructure roadmap, your goal is to balance the need for rapid innovation with operational stability. You shouldn't have to choose between adopting the latest database

safetygoogle-cloud-ai
11 May 2026
Tools

Here's the full TIL https://til.simonwillison.net/llms/llm-shebang

DGX agent

Simon Willison shares a technique for using Large Language Models directly from the command line using a shebang (#!) syntax, allowing scripts to be executed with LLM processing without explicit comma

toolssimon-willison--x
11 May 2026
Model Releases

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

DGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Narrow Secret Loyalty Dodges Black-Box Audits

DGX agent

arXiv:2605.06846v1 Announce Type: cross Abstract: Recent work identifies secret loyalties as a distinct threat from standard backdoors. A secret loyalty causes a model to covertly advance the interest

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

OpenAI just released its answer to Claude Mythos

DGX agent

OpenAI is launching Daybreak, an AI initiative focused on detecting and patching vulnerabilities before attackers find them. Daybreak uses the Codex Security AI agent that launched in March to create

model-releasesthe-verge-ai
11 May 2026
Safety

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

DGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

safetyarxiv-cs-ai
11 May 2026
Safety

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

DGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

safetyarxiv-cs-cl
11 May 2026
Model Releases

ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs

DGX agent

arXiv:2605.07338v1 Announce Type: new Abstract: The decline of global shellfish biodiversity poses a severe threat to coastal ecosystems. Although artificial intelligence (AI) technologies show potent

model-releasesarxiv-cs-cv
11 May 2026
Research

Statistical Patterns in the Equations of Physics and the Emergence of a Meta-Law of Nature

DGX agent

arXiv:2408.11065v2 Announce Type: replace-cross Abstract: Physics seeks to uncover the laws of Nature and express them through mathematical equations. Despite the vast diversity of natural phenomena,

researcharxiv-cs-cl
11 May 2026
Safety

STDA-Net: Spectrogram-Based Domain Adaptation for cross-dataset Sleep Stage Classification

DGX agent

arXiv:2605.06736v1 Announce Type: cross Abstract: Accurate sleep stage classification across datasets remains challenging due to variability in EEG channel montages, sampling rates, recording environm

safetyarxiv-cs-ai
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

DGX agent

arXiv:2605.07593v1 Announce Type: new Abstract: Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams,

model-releasesarxiv-cs-cv
11 May 2026
Applications

User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums

DGX agent

arXiv:2509.11777v2 Announce Type: replace Abstract: Customer feedback in industrial forums offers rich but underexplored insights into real-world product experience. Yet systematic analysis remains ch

applicationsarxiv-cs-cl
11 May 2026
Agents

VDCook:DIY video data cook your MLLMs

DGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

agentsarxiv-cs-ai
11 May 2026
Agents

VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents

DGX agent

arXiv:2506.21582v5 Announce Type: replace-cross Abstract: Text analytics has traditionally required specialized knowledge in Natural Language Processing (NLP) or text analysis, which presents a barrie

agentsarxiv-cs-ai
11 May 2026
Model Releases

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models

DGX agent

arXiv:2605.07872v1 Announce Type: cross Abstract: Multimodal reward models have advanced substantially in text and image domains, yet progress in video understanding reward modeling remains severely l

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

Your AI Use Is Breaking My Brain

DGX agent

Your AI Use Is Breaking My Brain Excellent, angry piece by Jason Koebler on how AI writing online is becoming impossible to avoid, filtering it is mentally exhausting and it's even starting to distort

tutorialssimon-willison
11 May 2026
Model Releases

Local AI is having its moment! Below is the number of new GGUF models created each month over the past 8 months & insights from our HF inter…

DGX agent

Local AI is having its moment! Below is the number of new GGUF models created each month over the past 8 months & insights from our HF internal agent (May is partial): - 176,000 total public GGUF mode

model-releasesclem-delangue--x
10 May 2026
Agents

MachinaCheck: Building a Multi-Agent CNC Manufacturability System on AMD MI300X

DGX agent

MachinaCheck is a multi-agent AI system designed to assess the manufacturability of parts for CNC (Computer Numerical Control) machining, built on AMD's MI300X GPU architecture. The system likely leve

agentshugging-face
10 May 2026
← Previous
1…8687888990…104
Next →