AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Research

Fuzzy PyTorch: Rapid Numerical Variability Evaluation for Deep Learning Models

DGX agent

arXiv:2605.25991v1 Announce Type: new Abstract: We introduce Fuzzy PyTorch, a framework for rapid evaluation of numerical variability in deep learning (DL) models. As DL is increasingly applied to div

researcharxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

DGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

model-releasesarxiv-cs-ai
26 May 2026
Agents

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

DGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

agentsarxiv-cs-ai
26 May 2026
Applications

Positivity in classical enumerative geometry: a case study in synchronized AI-assisted mathematics

DGX agent

arXiv:2605.25271v1 Announce Type: cross Abstract: We study the symmetric polynomial prod_{alphain A_{n,d}}igl(1+alpha_1 x_1+dots+alpha_n x_nigr) where A_{n,d}:={alphainZ_{ge 0}^n:|alpha|=d}, which is

applicationsarxiv-cs-ai
26 May 2026
Agents

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

DGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

agentsarxiv-cs-ai
26 May 2026
Model Releases

Querying structural and functional niches on spatial transcriptomics data

DGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

model-releasesarxiv-cs-lg
26 May 2026
Agents

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

DGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

agentsarxiv-cs-cl
26 May 2026
Research

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

DGX agent

arXiv:2605.24899v1 Announce Type: new Abstract: Ontologies represent the conceptual knowledge of a domain. At the core of an ontology is the taxonomy of concepts and subconcepts that represent specifi

researcharxiv-cs-ai
26 May 2026
Model Releases

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

DGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Agents

ExpOS: Explainable Open-Surgery Skills Assessment Using 3D Hand Reconstruction

DGX agent

arXiv:2605.23653v1 Announce Type: new Abstract: Timely and transparent feedback is essential for effective surgical training, yet current assessment remains dependent on expert observation, limiting s

agentsarxiv-cs-cv
25 May 2026
Safety

NeuroNL2LTL: A Neurosymbolic Framework for Natural Language Translation of Linear Temporal Logic

DGX agent

arXiv:2605.22874v1 Announce Type: new Abstract: Effectively translating between natural language (NL) and formal logics like Linear Temporal Logic (LTL) requires expertise that limits formal verificat

safetyarxiv-cs-ai
25 May 2026
Research

Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration

DGX agent

arXiv:2512.11587v2 Announce Type: replace Abstract: Even for the gradient descent (GD) method applied to neural network training, understanding its optimization dynamics, including convergence rate, i

researcharxiv-cs-lg
23 May 2026
Research

Turning Trust to Transactions: Tracking Affiliate Marketing and FTC Compliance in YouTube's Influencer Economy

DGX agent

arXiv:2603.04383v2 Announce Type: replace-cross Abstract: YouTube has evolved into a powerful platform where creators monetize their influence through affiliate marketing, raising concerns about trans

researcharxiv-cs-lg
23 May 2026
Model Releases

HealthCraft: A Reinforcement Learning Safety Environment for Emergency Medicine

DGX agent

arXiv:2605.21496v1 Announce Type: cross Abstract: Frontier language models are being deployed into clinical workflows faster than the infrastructure to evaluate them safely. Static medical-QA benchmar

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms

DGX agent

arXiv:2605.20704v1 Announce Type: cross Abstract: Autonomous AI agents that spawn sub-agent swarms create a safety gap: existing credential revocation mechanisms, OAuth~2.0 introspection, OCSP, and W3

local-aiarxiv-cs-ai
22 May 2026
Agents

Quality and Security Signals in AI-Generated Python Refactoring Pull Requests

DGX agent

arXiv:2605.21453v1 Announce Type: cross Abstract: As AI agents increasingly contribute to code development and maintenance, there is still limited empirical evidence on the quality and risk characteri

agentsarxiv-cs-ai
22 May 2026
Research

Component Influence-Driven Fastener Reduction for Robotic Disassemblability-Aware Design Simplification

DGX agent

arXiv:2605.21026v1 Announce Type: new Abstract: To accelerate automated remanufacturing, robotic disassembly must be considered during the product design phase. However, designers currently lack quant

researcharxiv-cs-ro
21 May 2026
Research

Generative AI Practices, Literacy, and Divides: An Empirical Analysis in the Italian Context

DGX agent

arXiv:2512.03671v2 Announce Type: replace Abstract: The rise of generative AI (GenAI) chatbots accessible via conversational interfaces is transforming digital interactions and holds economic promise.

researcharxiv-cs-cl
21 May 2026
Model Releases

MemGym: a Long-Horizon Memory Environment for LLM Agents

DGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

model-releasesarxiv-cs-cl
21 May 2026
Tutorials

WiXus: A Wheeled-Legged Robot with Wire-Driven Environmental Utilizing to Integrate Mobility and Manipulation

DGX agent

arXiv:2605.20932v1 Announce Type: new Abstract: Wheeled-legged robots, which have wheels at their feet and achieve high mobility by coordinating wheel drive and leg drive, have been developed. These r

tutorialsarxiv-cs-ro
21 May 2026
Agents

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

DGX agent

arXiv:2605.18770v1 Announce Type: cross Abstract: We present a collaborative agentic GraphRAG framework for expert analysis of commercial registry data. Public registries are often formally accessible

agentsarxiv-cs-ai
20 May 2026
Research

Can LLMs Estimate Cognitive Complexity of Reading Comprehension Items?

DGX agent

arXiv:2510.25064v2 Announce Type: replace Abstract: Estimating the cognitive complexity of reading comprehension (RC) items is crucial for assessing item difficulty before it is administered to learne

researcharxiv-cs-cl
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Applications

Prognostic Value of Lung Ultrasound Biomarkers for Readmission Risk in Congestive Heart Failure: A Pilot Data-Driven Analysis

DGX agent

arXiv:2605.18878v1 Announce Type: cross Abstract: Hospital readmission within 30 days of discharge is a leading driver of morbidity, mortality, and avoidable healthcare expenditure in congestive heart

applicationsarxiv-cs-cv
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Vision Harnessing Agent for Open Ad-hoc Segmentation

DGX agent

arXiv:2605.19410v1 Announce Type: new Abstract: Segmentation has become easy when the concept is known, requiring retrieval of a learned visual grounding from text. It remains hard for open ad-hoc con

model-releasesarxiv-cs-cv
20 May 2026
Agents

YAC: Bridging Natural Language and Interactive Visual Exploration with Generative AI for Biomedical Data Discovery

DGX agent

arXiv:2509.19182v2 Announce Type: replace-cross Abstract: Incorporating natural language input has the potential to improve the capabilities of biomedical data discovery interfaces. However, user inte

agentsarxiv-cs-ai
20 May 2026
Model Releases

AI for Auto-Research: Roadmap & User Guide

DGX agent

arXiv:2605.18661v1 Announce Type: new Abstract: AI-assisted research is crossing a threshold: fully automated systems can now generate research papers for as little as $15, while long-horizon agents c

model-releasesarxiv-cs-ai
19 May 2026
Safety

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering

DGX agent

arXiv:2605.17352v1 Announce Type: new Abstract: Despite substantial advances in large language models (LLMs), generating factually consistent responses for knowledge-intensive question answering remai

safetyarxiv-cs-cl
19 May 2026
Model Releases

CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?

DGX agent

arXiv:2605.16679v1 Announce Type: cross Abstract: End-to-end automation of realistic healthcare operations stresses three capabilities underrepresented in current benchmarks: policy density, decisions

model-releasesarxiv-cs-ai
19 May 2026
Agents

DeepArrhythmia: Segment-Contextualized ECG Arrhythmia Classification via Selective Evidence Acquisition

DGX agent

arXiv:2605.16441v1 Announce Type: cross Abstract: Beat-level Electrocardiography (ECG) arrhythmia detection aims to assign an arrhythmia class to each beat in a recording, yet many existing systems tr

agentsarxiv-cs-ai
19 May 2026
Model Releases

Detecting Verbatim LLM Copy-Paste in Homework

DGX agent

arXiv:2605.16336v1 Announce Type: cross Abstract: Large language models (LLMs) have made fluent essay writing, code drafting, and quiz answering instantly available to students at every level, from se

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

Evidence of a Cognitive Shift in AI Education: How Students Are Rethinking Human Intelligence?

DGX agent

arXiv:2605.16292v1 Announce Type: cross Abstract: Perceptions of intelligence shape how learners evaluate and rely on artificial intelligence (AI) systems. Despite rapid advances in AI capabilities, t

tutorialsarxiv-cs-ai
19 May 2026
Agents

Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework

DGX agent

arXiv:2605.16821v1 Announce Type: new Abstract: The rapid evolution of Large Language Model (LLM) agents has produced diverse interaction paradigms, yet few production systems integrate multiple parad

agentsarxiv-cs-ai
19 May 2026
Model Releases

OpenJarvis: Personal AI, On Personal Devices

DGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

model-releasesarxiv-cs-ai
19 May 2026
Safety

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

DGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

safetyarxiv-cs-ai
19 May 2026
Model Releases

SEDD: Scalable and Efficient Dataset Deduplication with GPUs

DGX agent

arXiv:2501.01046v4 Announce Type: replace Abstract: Dataset deduplication is widely recognized as a crucial preprocessing step that enhances data quality and improves the performance of large language

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain

DGX agent

arXiv:2605.17946v1 Announce Type: new Abstract: Multimodal large language models are increasingly used as agent backbones that understand multimodal inputs, plan retrieval actions, invoke external too

model-releasesarxiv-cs-ai
19 May 2026
Hardware

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

DGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

hardwarearxiv-cs-lg
19 May 2026
Model Releases

Your SaaS Is an Insurance Product: A Modeling Framework

DGX agent

arXiv:2605.16699v1 Announce Type: new Abstract: Capped-usage SaaS products -- LLM subscriptions such as Claude Code and ChatGPT, cloud platforms such as Vercel and Cloudflare Workers, corporate benefi

model-releasesarxiv-cs-lg
19 May 2026
Agents

Context, Reasoning, and Hierarchy: A Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP

DGX agent

arXiv:2605.16205v1 Announce Type: new Abstract: Deploying compound LLM agents in adversarial, partially observable sequential environments requires navigating several design dimensions: (1) what the a

agentsarxiv-cs-ai
18 May 2026
Research

How Data Augmentation Shapes Neural Representations

DGX agent

arXiv:2605.15306v1 Announce Type: new Abstract: Data augmentation is widely recognized for improving generalization in deep networks, yet its impact on the geometry of learned representations remains

researcharxiv-cs-lg
18 May 2026
Agents

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

DGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

agentsarxiv-cs-lg
18 May 2026
Research

From User Preferences to Base Score Extraction Functions in Gradual Argumentation (with Appendix)

DGX agent

arXiv:2602.14674v4 Announce Type: replace Abstract: Gradual argumentation is a field of symbolic AI which is attracting attention for its ability to support transparent and contestable AI systems. It

researcharxiv-cs-ai
15 May 2026
Research

MicroscopyMatching: Towards a Ready-to-use Framework for Microscopy Image Analysis in Diverse Conditions

DGX agent

arXiv:2605.14980v1 Announce Type: cross Abstract: Analyzing microscopy images to extract biological object properties (e.g., their morphological organization, temporal dynamics, and population density

researcharxiv-cs-ai
15 May 2026
Model Releases

Near-Miss: Latent Policy Failure Detection in Agentic Workflows

DGX agent

arXiv:2603.29665v2 Announce Type: replace Abstract: Agentic systems for business process automation often require compliance with policies governing conditional updates to the system state. Evaluation

model-releasesarxiv-cs-cl
15 May 2026
← Previous
1…4748495051…109
Next →