AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Research

A Category-Theoretic Analysis of Conformal Prediction

DGX agent

arXiv:2507.04441v4 Announce Type: replace-cross Abstract: Conformal prediction (CP) produces prediction regions with finite-sample, distribution free coverage guarantees, but its interpretation as a q

researcharxiv-cs-lg
5 May 2026
Agents

A Language for Describing Agentic LLM Contexts

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.01920v1 Announce Type: cross Abstract: Large language models are increasingly used within larger systems ('LLM agents'). These make a sequence of LLM calls, each call providing the LLM with

agentsarxiv-cs-cl
5 May 2026
Research

Active Reasoning Vision-Language Models via Sequential Experimental Design

DGX agent

arXiv:2605.01345v1 Announce Type: new Abstract: Visual perception in modern Vision-Language Models (VLMs) is constrained by a fundamental perceptual bandwidth bottleneck: a broad field of view inevita

researcharxiv-cs-cv
5 May 2026
Safety

Adaptive Interpolation-Synthesis for Motion In-Betweening on Keyframe-Based Animation

DGX agent

arXiv:2605.02742v1 Announce Type: cross Abstract: Motion in-betweening is one of the most artistically demanding and time consuming stages of 3D animation, where the expressivity and rhythm of motion

safetyarxiv-cs-lg
5 May 2026
Agents

AgentXRay: White-Boxing Agentic Systems via Workflow Reconstruction

DGX agent

arXiv:2602.05353v3 Announce Type: replace-cross Abstract: Large Language Models have shown strong capabilities in complex problem solving, yet many agentic systems remain difficult to interpret and co

agentsarxiv-cs-cl
5 May 2026
Safety

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

DGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

safetyarxiv-cs-ai
5 May 2026
Safety

Analyzing Adversarial Inputs in Deep Reinforcement Learning

DGX agent

arXiv:2402.05284v2 Announce Type: replace Abstract: In recent years, Deep Reinforcement Learning (DRL) has become a popular paradigm in machine learning due to its successful applications to real-worl

safetyarxiv-cs-lg
5 May 2026
Research

Bucketing the Good Apples: A Method for Diagnosing and Improving Causal Abstraction

DGX agent

arXiv:2605.02234v1 Announce Type: cross Abstract: We present a method for diagnosing interpretation in neural networks by identifying an input subspace where a proposed interpretation is highly faithf

researcharxiv-cs-cl
5 May 2026
Model Releases

Constructing Interpretable Features from Compositional Neuron Groups

DGX agent

arXiv:2506.10920v2 Announce Type: replace Abstract: A central goal for mechanistic interpretability has been to identify the right units of analysis in large language models (LLMs) that causally expla

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features

DGX agent

arXiv:2602.10437v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) decompose language model activations into interpretable features, but existing methods reveal only which features a

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Creating and Evaluating Figurative Language Dataset for Sindhi

DGX agent

arXiv:2605.01323v1 Announce Type: new Abstract: In this article, we introduce SiNFluD, a novel benchmark dataset for Sindhi figurative language classification. We first collect raw text from various b

model-releasesarxiv-cs-cl
5 May 2026
Research

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages

DGX agent

arXiv:2605.02608v1 Announce Type: new Abstract: Transformer-based models achieve state-of-the-art dependency parsing for high-resource languages, yet their advantage over simpler architectures in low-

researcharxiv-cs-cl
5 May 2026
Research

DIAGRAMS: A Review Framework for Reasoning-Level Attribution in Diagram QA

DGX agent

arXiv:2605.00905v1 Announce Type: new Abstract: Diagram question answering (Diagram QA) requires reasoning-level attribution that links each question-answer pair to all visual regions needed to derive

researcharxiv-cs-cl
5 May 2026
Model Releases

FeedbackLLM: Metadata driven Multi-Agentic Language Agnostic Test Case Generator with Evolving prompt and Coverage Feedback

DGX agent

arXiv:2605.01264v1 Announce Type: cross Abstract: Traditional approaches to test case generation often involve manual effort and incur significant computational overhead. Additionally, these approache

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

From Where Things Are to What They Are For: Benchmarking Spatial-Functional Intelligence in Multimodal LLMs

DGX agent

arXiv:2605.02130v1 Announce Type: new Abstract: Human-level agentic intelligence extends beyond low-level geometric perception, evolving from recognizing where things are to understanding what they ar

model-releasesarxiv-cs-cv
5 May 2026
Agents

GA-VisAgent: A Multi-Agent application for code generation and visualization in interactive learning

DGX agent

arXiv:2605.01299v1 Announce Type: new Abstract: Geometric Algebra (GA) presents challenges to learners due to its highly abstract mathematical structure and complex operational rules, as translating a

agentsarxiv-cs-lg
5 May 2026
Agents

Hallucinations Undermine Trust; Metacognition is a Way Forward

DGX agent

arXiv:2605.01428v1 Announce Type: new Abstract: Despite significant strides in factual reliability, errors -- often termed hallucinations -- remain a major concern for generative AI, especially as LLM

agentsarxiv-cs-cl
5 May 2026
Research

IMPACT-Scribe: Interactive Temporal Action Segmentation with Boundary Scribbles and Query Planning

DGX agent

arXiv:2605.01668v1 Announce Type: new Abstract: Dense temporal annotation of procedural activity videos is vital for action understanding and embodied intelligence but remains labor-intensive due to r

researcharxiv-cs-cv
5 May 2026
Research

KANs need curvature: penalties for compositional smoothness

DGX agent

arXiv:2605.02190v1 Announce Type: new Abstract: Kolmogorov-Arnold networks (KANs) offer a potent combination of accuracy and interpretability, thanks to their compositions of learnable univariate acti

researcharxiv-cs-lg
5 May 2026
Tutorials

Learning Koopman operators for coupled systems via information on governing equations of subsystems

DGX agent

arXiv:2605.01835v1 Announce Type: new Abstract: Nonlinear coupled systems are ubiquitous in science and engineering. The analysis and modeling of such systems is challenging due to their high dimensio

tutorialsarxiv-cs-lg
5 May 2026
Research

Linear-Readout Floors and Threshold Recovery in Computation in Superposition

DGX agent

arXiv:2605.01192v1 Announce Type: new Abstract: Two recent approaches to computation in superposition reach different recursive capacity regimes: Hanni et al. certify ilde{O}(d^{3/2}) computable featu

researcharxiv-cs-lg
5 May 2026
Research

LITcoder: A General-Purpose Library for Building and Comparing Encoding Models

DGX agent

arXiv:2509.09152v2 Announce Type: replace Abstract: We introduce LITcoder, an open-source library for building and benchmarking neural encoding models. Designed as a flexible backend, LITcoder provide

researcharxiv-cs-cl
5 May 2026
Agents

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

DGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

agentsarxiv-cs-cv
5 May 2026
Model Releases

MolViBench: Evaluating LLMs on Molecular Vibe Coding

DGX agent

arXiv:2605.02351v1 Announce Type: new Abstract: Molecular Vibe Coding, a paradigm where chemists interact with LLMs to generate executable programs for molecular tasks, has emerged as a flexible alter

model-releasesarxiv-cs-cl
5 May 2026
Local Ai

Open-access model for detecting openly dumped dispersed municipal solid waste from crowdsourced UAV imagery in Sub-Saharan Africa

DGX agent

arXiv:2605.02316v1 Announce Type: new Abstract: Managing municipal solid waste in rapidly urbanizing Sub-Saharan Africa remains challenging due to dispersed informal dumping and limited high-resolutio

local-aiarxiv-cs-cv
5 May 2026
Model Releases

OpenAI GPT-5 System Card

DGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

DGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

model-releasesarxiv-cs-lg
5 May 2026
Research

Principles and Guidelines for Randomized Controlled Trials in AI Evaluation

DGX agent

arXiv:2605.02050v1 Announce Type: cross Abstract: This work establishes a foundational framework for standardizing AI evaluation RCTs (sometimes called human uplift studies). Drawing on established ex

researcharxiv-cs-lg
5 May 2026
Research

Random-Effects Algorithm for Random Objects in Metric Spaces

DGX agent

arXiv:2605.02693v1 Announce Type: cross Abstract: Across many scientific disciplines, multiple observations are collected from the same experimental units, and in modern datasets these observations of

researcharxiv-cs-lg
5 May 2026
Research

Reconstructing conformal field theoretical compositions with Transformers

DGX agent

arXiv:2605.01072v1 Announce Type: cross Abstract: We study the use of transformers to reconstruct the compositions of tensor products of two-dimensional rational conformal field theories (RCFTs) based

researcharxiv-cs-lg
5 May 2026
Safety

Reinforcement Learning from Compiler and Language Server Feedback

DGX agent

arXiv:2510.22907v2 Announce Type: replace Abstract: Coding agents fail when text-level guesses outrun program facts: they hallucinate APIs, drift to the wrong symbol, and apply edits without evidence

safetyarxiv-cs-cl
5 May 2026
Model Releases

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning

DGX agent

arXiv:2605.01489v1 Announce Type: cross Abstract: Frontier scientific reasoning is rapidly emerging as a key foundation for advancing AI agents in automated scientific discovery. Deep research agents

model-releasesarxiv-cs-cl
5 May 2026
Agents

Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis

DGX agent

arXiv:2605.00314v1 Announce Type: cross Abstract: An agent skill is a configuration package that equips an LLM-driven agent with a concrete capability, such as reading email, executing shell commands,

agentsarxiv-cs-ai
5 May 2026
Agents

Sound Source Localization for Spatial Mapping of Surgical Actions in Dynamic Scenes

DGX agent

arXiv:2510.24332v3 Announce Type: replace-cross Abstract: Purpose: Surgical scene understanding is key to advancing computer-aided and intelligent surgical systems. Current approaches predominantly re

agentsarxiv-cs-cv
5 May 2026
Model Releases

Spectral Model eXplainer: a chemically-grounded explainability framework for spectral-based machine learning models

DGX agent

arXiv:2605.02684v1 Announce Type: new Abstract: Spectral-based machine learning models have been increasingly deployed in chemometrics and spectroscopy, where predictive accuracy is as important as ex

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer

DGX agent

arXiv:2605.00924v1 Announce Type: new Abstract: AI-generated content (AIGC) detectors are increasingly deployed in high-stakes settings such as academic integrity screening, yet their reliability rest

model-releasesarxiv-cs-lg
5 May 2026
Agents

SUDP: Secret-Use Delegation Protocol for Agentic Systems

DGX agent

arXiv:2604.24920v2 Announce Type: replace-cross Abstract: Agentic systems increasingly act with user secrets for APIs, messaging platforms, and cloud services. Today's bearer-secret interfaces impleme

agentsarxiv-cs-ai
5 May 2026
Safety

SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion

DGX agent

arXiv:2605.01510v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in high-quality image synthesis, sparking interest in image-guided generation tasks such as subject-dr

safetyarxiv-cs-cv
5 May 2026
Model Releases

TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs

DGX agent

arXiv:2510.15545v4 Announce Type: replace Abstract: Accelerating the inference of large language models (LLMs) has been a critical challenge in generative AI. Speculative decoding (SD) substantially i

model-releasesarxiv-cs-cl
5 May 2026
Local Ai

UnGAP: Uncertainty-Guided Affine Prompting for Real-Time Crack Segmentation

DGX agent

arXiv:2605.02380v1 Announce Type: new Abstract: Real-time crack segmentation is vital for structural health monitoring but is plagued by aleatoric uncertainties arising from varying lighting, blur, an

local-aiarxiv-cs-cv
5 May 2026
Research

Universality in Deep Neural Networks: An approach via the Lindeberg exchange principle

DGX agent

arXiv:2605.02771v1 Announce Type: cross Abstract: We consider the infinite-width limit of a fully connected deep neural network with general weights, and we prove quantitative general bounds on the 2-

researcharxiv-cs-lg
5 May 2026
Research

Virtual Scanning for NSCLC Histology: Investigating the Discriminatory Power of Synthetic PET

DGX agent

arXiv:2605.02746v1 Announce Type: new Abstract: Accurate histological differentiation between adenocarcinoma (ADC) and squamous cell carcinoma (SCC) is critical for personalized treatment in non-small

researcharxiv-cs-cv
5 May 2026
Research

A Novel Patch-Based TDA Approach for Computed Tomography Imaging

DGX agent

arXiv:2512.12108v5 Announce Type: replace Abstract: The development of machine learning models based on computed tomography (CT) imaging has been a major focus due to the promise that imaging holds fo

researcharxiv-cs-cv
4 May 2026
Safety

Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines

DGX agent

arXiv:2605.00410v1 Announce Type: new Abstract: A multi-agent pipeline with N agents typically issues N LLM calls per run. Merging agents into fewer calls (compound execution) promises token savings,

safetyarxiv-cs-cl
4 May 2026
Model Releases

Can Coding Agents Reproduce Findings in Computational Materials Science?

DGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

model-releasesarxiv-cs-cl
4 May 2026
Research

Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments

DGX agent

arXiv:2505.09901v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate or automate human behavior in complex sequential decision-making settings. A na

researcharxiv-cs-cl
4 May 2026
Model Releases

Concolic Testing on Individual Fairness of Neural Network Models

DGX agent

arXiv:2509.06864v2 Announce Type: replace Abstract: This paper introduces PyFair, a formal framework for evaluating and verifying individual fairness of Deep Neural Networks (DNNs). By adapting the co

model-releasesarxiv-cs-lg
4 May 2026
Tutorials

Latent Generative Modeling of Random Fields from Limited Training Data

DGX agent

arXiv:2505.13007v2 Announce Type: replace Abstract: The ability to accurately model random fields plays a critical role in science and engineering for problems involving uncertain, spatially-varying q

tutorialsarxiv-cs-lg
4 May 2026
← Previous
1…9394959697…108
Next →