AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Tutorials

Puzzled By ChatGPT? No more! A Jigsaw Puzzle to Promote AI Literacy and Awareness

DGX agent

arXiv:2605.20404v1 Announce Type: new Abstract: The rapid adoption of Generative AI, including LLM-based chatbots like ChatGPT, has highlighted the need for accessible ways to support public understan

tutorialsarxiv-cs-cl
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Reviving Error Correction in Modern Deep Time-Series Forecasting

DGX agent

arXiv:2605.21088v1 Announce Type: new Abstract: Modern deep-learning models have achieved remarkable success in time-series forecasting. Yet, their performance degrades in long-term prediction due to

researcharxiv-cs-lg
21 May 2026
Local Ai

The General Theory of Localization Methods

DGX agent

arXiv:2605.20635v1 Announce Type: new Abstract: This paper proposes a general machine learning framework called the localization method, which is fundamentally built on two core concepts: localization

local-aiarxiv-cs-lg
21 May 2026
Research

Understanding Model Behavior in Monocular Polyp Sizing

DGX agent

arXiv:2605.20461v1 Announce Type: new Abstract: Accurate polyp size stratification guides surveillance decisions, with lesions larger than 5 mm typically requiring closer follow-up. However, monocular

researcharxiv-cs-cv
21 May 2026
Model Releases

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

DGX agent

arXiv:2605.19988v1 Announce Type: cross Abstract: Documentation has long guided computer system tuning by distilling expert knowledge into per-parameter recommendations. Yet such guides capture only w

model-releasesarxiv-cs-ai
20 May 2026
Tutorials

A Multi-Dimensional Clustering Approach for Identifying Inborn Errors of Immunity

DGX agent

arXiv:2605.18880v1 Announce Type: cross Abstract: Rare diseases such as inborn errors of immunity (IEI) require early diagnosis to prevent end organ damage and improve quality of life. Hurdles in acce

tutorialsarxiv-cs-cv
20 May 2026
Agents

A novel YOLO26-MoE optimized by an LLM agent for insulator fault detection considering UAV images

DGX agent

arXiv:2605.19595v1 Announce Type: cross Abstract: The inspection of electrical power line insulators is essential for ensuring grid reliability and preventing failures caused by damaged or degraded in

agentsarxiv-cs-ai
20 May 2026
Research

Accurate, Efficient, and Explainable Deep Learning Approaches for Environmental Science Problems

DGX agent

arXiv:2605.19366v1 Announce Type: new Abstract: Environmental science plays a pivotal role in safeguarding ecosystems, a domain driven by large-scale, heterogeneous data. In the big data era, artifici

researcharxiv-cs-lg
20 May 2026
Tutorials

Bayesian Symbolic Regression for Missing Physics

DGX agent

arXiv:2603.14918v2 Announce Type: replace-cross Abstract: Model-based approaches for (bio)process systems often suffer from incomplete knowledge of the underlying physical, chemical, or biological law

tutorialsarxiv-cs-lg
20 May 2026
Model Releases

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

DGX agent

arXiv:2605.20176v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has a

model-releasesarxiv-cs-cl
20 May 2026
Local Ai

Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds

DGX agent

arXiv:2605.18827v1 Announce Type: cross Abstract: Multiple-choice QA benchmarks usually evaluate small language models (SLMs) as direct answerers, but deployed language-model systems increasingly rely

local-aiarxiv-cs-lg
20 May 2026
Model Releases

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

DGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

model-releasesarxiv-cs-ai
20 May 2026
Safety

Distributional AGI Safety

DGX agent

arXiv:2512.16856v2 Announce Type: replace Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an e

safetyarxiv-cs-ai
20 May 2026
Model Releases

DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model

DGX agent

arXiv:2602.23622v2 Announce Type: replace-cross Abstract: Significant progress has been made in the field of Instruction-based Image Editing Models (IIEMs). However, while these models demonstrate pla

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Enabling Real-Time Colonoscopic Polyp Segmentation on Commodity CPUs via Ultra-Lightweight Architecture

DGX agent

arXiv:2602.04381v2 Announce Type: replace-cross Abstract: Real-time polyp segmentation is essential for early colorectal cancer detection, yet clinical deployment remains blocked by GPU dependency. We

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design

DGX agent

arXiv:2605.19743v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly applied to engineering design tasks, yet existing evaluation frameworks do not adequately address mul

model-releasesarxiv-cs-ai
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Tutorials

Evaluating Memory Condensation Strategies for Coding Agents in Data-Driven Scientific Discovery

DGX agent

arXiv:2605.18854v1 Announce Type: new Abstract: Coding agents accumulate extensive context during long-running tasks, yet fixed context windows force practitioners to choose between truncation and tas

tutorialsarxiv-cs-lg
20 May 2026
Research

Fine-Tuning Without Forgetting via Loss-Adaptive Learning Rates

DGX agent

arXiv:2605.20005v1 Announce Type: new Abstract: Fine-tuning large language models on new data improves task performance but degrades capabilities learned during pretraining, a phenomenon known as cata

researcharxiv-cs-lg
20 May 2026
Model Releases

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

DGX agent

arXiv:2605.20006v1 Announce Type: new Abstract: Geospatial reasoning requires solving image-grounded problems over the complex spatial structure of a scene. However, developing this capability is hind

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Graph Neural Networks for Community Detection in Graph Signal Analysis

DGX agent

arXiv:2605.19733v1 Announce Type: cross Abstract: Community detection is a central problem in graph analysis, with applications ranging from network science to graph signal processing. In recent years

model-releasesarxiv-cs-lg
20 May 2026
Safety

Guiding Neuro-Symbolic Scenario Generation with Spatio-Temporal Logic

DGX agent

arXiv:2605.19038v1 Announce Type: cross Abstract: The rapid advancement of autonomous driving (AD) technologies has outpaced the development of robust safety evaluation methods. Conventional testing r

safetyarxiv-cs-lg
20 May 2026
Model Releases

Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling

DGX agent

arXiv:2605.18838v1 Announce Type: cross Abstract: Scaling laws predict loss from compute but not how capabilities interact. We measure the coupling between reasoning and truthfulness across 63 base mo

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges

DGX agent

arXiv:2605.19723v1 Announce Type: cross Abstract: Mathematical reasoning is essential for problem-solving in education, science, and industry, serving as a crucial benchmark for evaluating artificial

model-releasesarxiv-cs-ai
20 May 2026
Local Ai

Mechanisms of Object Localization in Vision-Language Models

DGX agent

arXiv:2605.19792v1 Announce Type: new Abstract: Visually-grounded language models (VLMs) are highly effective in linking visual and textual information, yet they often struggle with basic classificati

local-aiarxiv-cs-cv
20 May 2026
Model Releases

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

DGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

DGX agent

arXiv:2605.18825v1 Announce Type: new Abstract: Prefix caching is a key optimization in Large Language Model (LLM) serving, reusing attention Key-Value (KV) states across requests with shared prompt p

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

OpenCompass: A Universal Evaluation Platform for Large Language Models

DGX agent

arXiv:2605.19276v1 Announce Type: new Abstract: In recent years, the field of artificial intelligence has undergone a paradigm shift from task-specific small-scale models to general-purpose large lang

model-releasesarxiv-cs-cl
20 May 2026
Research

OpenComputer: Verifiable Software Worlds for Computer-Use Agents

DGX agent

arXiv:2605.19769v1 Announce Type: new Abstract: We present OpenComputer, a verifier-grounded framework for constructing verifiable software worlds for computer-use agents. OpenComputer integrates four

researcharxiv-cs-ai
20 May 2026
Agents

PASC: Pipeline-Aware Conformal Prediction with Joint Coverage Guarantees for Multi-Stage NLP and LLM Pipelines

DGX agent

arXiv:2605.18812v1 Announce Type: cross Abstract: Modern NLP and LLM systems are pipelines: named entity recognition (NER) -> entity disambiguation (NED) -> entity typing, retrieval-augmented generati

agentsarxiv-cs-cl
20 May 2026
Model Releases

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

DGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Robust Basis Spline Decoupling for the Compression of Transformer Models

DGX agent

arXiv:2605.18794v1 Announce Type: cross Abstract: Decoupling is a powerful modeling paradigm for representing multivariate functions as compositions of linear transformations and univariate nonlinear

model-releasesarxiv-cs-ai
20 May 2026
Agents

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

DGX agent

arXiv:2605.18857v1 Announce Type: cross Abstract: For most of the history of information retrieval (IR), search results were designed for human consumers who could scan, filter, and discard irrelevant

agentsarxiv-cs-ai
20 May 2026
Safety

Toward an AI-Powered Computational Testbed for Workforce Policy

DGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

safetyarxiv-cs-ai
20 May 2026
Agents

Towards Discovery of Polymers for Insulin Delivery via Physics-Grounded Agentic Workflows

DGX agent

arXiv:2605.18831v1 Announce Type: cross Abstract: Cold-chain storage limits access to insulin for hundreds of millions of people; a thermally protective patch polymer could help, but the design space

agentsarxiv-cs-lg
20 May 2026
Research

Transformers Linearly Represent Highly Structured World Models

DGX agent

arXiv:2605.18847v1 Announce Type: cross Abstract: Do transformers, when trained on sequential reasoning traces, build internal models of the underlying task? And if so, does the structure of those int

researcharxiv-cs-ai
20 May 2026
Tutorials

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation

DGX agent

arXiv:2602.09023v4 Announce Type: replace Abstract: Despite strong generalization capabilities, Vision-Language-Action (VLA) models remain constrained by the high cost of expert demonstrations and lim

tutorialsarxiv-cs-ro
20 May 2026
Model Releases

ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins

DGX agent

arXiv:2603.06740v2 Announce Type: replace-cross Abstract: Protein language models (pLMs) have shown strong potential for zero-shot prediction of missense variant effects, yet systematic benchmarking o

model-releasesarxiv-cs-ai
20 May 2026
Research

Where Does Authorship Signal Emerge in Encoder-Based Language Models?

DGX agent

arXiv:2605.19908v1 Announce Type: new Abstract: Authorship attribution models fine-tuned with the same pretrained encoder, data, and loss can differ four-fold in performance depending only on their sc

researcharxiv-cs-cl
20 May 2026
Model Releases

XNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception

DGX agent

arXiv:2603.22453v2 Announce Type: replace Abstract: Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance

model-releasesarxiv-cs-cl
20 May 2026
Research

3D Densification for Multi-Map Monocular VSLAM in Endoscopy

DGX agent

arXiv:2503.14346v3 Announce Type: replace Abstract: Multi-map Sparse Monocular visual Simultaneous Localization and Mapping applied to monocular endoscopic sequences has proven efficient to robustly r

researcharxiv-cs-cv
19 May 2026
Research

A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks

DGX agent

arXiv:2512.04329v2 Announce Type: replace Abstract: Reusing existing neural-network components is central to research efficiency, yet discovering, extracting, and validating such modules across thousa

researcharxiv-cs-cv
19 May 2026
Research

A Unified Framework for Structured Flow Modeling: From Continuous Fields to Data-Driven Representations

DGX agent

arXiv:2605.18250v1 Announce Type: cross Abstract: Many dynamical systems can be described in terms of structured flows combining source/sink behavior, cyclic dynamics, and topology-constrained transpo

researcharxiv-cs-lg
19 May 2026
Model Releases

ADR: An Agentic Detection System for Enterprise Agentic AI Security

DGX agent

arXiv:2605.17380v1 Announce Type: new Abstract: We present the Agentic AI Detection and Response (ADR) system, the first large-scale, production-proven enterprise framework for securing AI agents oper

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Adversarial Agent Collaboration for Correctness Improvements of C to Safe Rust Translation

DGX agent

arXiv:2510.03879v3 Announce Type: replace-cross Abstract: Translating C to memory-safe languages, like Rust, prevents critical memory safety vulnerabilities that are prevalent in legacy C software. Ev

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Agentic AI Governance and Lifecycle Management in Healthcare

DGX agent

arXiv:2601.15630v2 Announce Type: replace Abstract: Healthcare organizations are beginning to embed agentic AI into routine workflows, including clinical documentation support and early-warning monito

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

ALIGN: A Vision-Language Framework for High-Accuracy Accident Location Inference through Geo-Spatial Neural Reasoning

DGX agent

arXiv:2511.06316v3 Announce Type: replace Abstract: In low- and middle-income countries, public safety and urban planning initiatives frequently face a critical shortage of accurate, location-specific

local-aiarxiv-cs-ai
19 May 2026
Safety

AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment

DGX agent

arXiv:2605.18529v1 Announce Type: new Abstract: The alignment of Large Language Models (LLMs) for complex reasoning heavily relies on Reinforcement Learning with Verifiable Rewards (RLVR). However, st

safetyarxiv-cs-ai
19 May 2026
← Previous
1…8586878889…109
Next →