AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

DGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

model-releasesarxiv-cs-lg
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

The Attribution Impossibility: No Feature Ranking Is Faithful, Stable, and Complete Under Collinearity

DGX agent

arXiv:2605.21492v1 Announce Type: new Abstract: No feature ranking can be simultaneously faithful, stable, and complete when features are collinear. For collinear pairs, ranking reduces to a coin flip

safetyarxiv-cs-lg
23 May 2026
Safety

Agentic CLEAR: Automating Multi-Level Evaluation of LLM Agents

DGX agent

arXiv:2605.22608v1 Announce Type: new Abstract: Agentic systems are becoming more capable: agents define strategies, take actions, and interact with different environments. This autonomy poses serious

safetyarxiv-cs-cl
22 May 2026
Agents

An Entity Linking Agent for Question Answering

DGX agent

arXiv:2508.03865v4 Announce Type: replace Abstract: Some Question Answering (QA) systems rely on knowledge bases (KBs) to provide accurate answers. Entity Linking (EL) plays a critical role in linking

agentsarxiv-cs-cl
22 May 2026
Model Releases

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

DGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries

DGX agent

arXiv:2605.21712v1 Announce Type: new Abstract: Transportation safety analysis requires integrating crash records, roadway attributes, and geospatial data through GIS-based workflows, but access remai

local-aiarxiv-cs-cl
22 May 2026
Tutorials

Chain-of-thought obfuscation learned from output supervision can generalise to unseen tasks

DGX agent

arXiv:2601.23086v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning provides a significant performance uplift to LLMs by enabling planning, exploration, and deliberation of their acti

tutorialsarxiv-cs-ai
22 May 2026
Safety

Diagnosis Is Not Prescription: Linguistic Co-Adaptation Explains Patching Hazards in LLM Pipelines

DGX agent

arXiv:2605.21958v1 Announce Type: new Abstract: When a multi-module LLM agent fails, the module most responsible for the failure is not necessarily the best place to intervene. We demonstrate this Dia

safetyarxiv-cs-cl
22 May 2026
Safety

Discovering Implicit Large Language Model Alignment Objectives

DGX agent

arXiv:2602.15338v2 Announce Type: replace-cross Abstract: Large language model (LLM) alignment relies on complex reward signals that often obscure the specific behaviors being incentivized, creating c

safetyarxiv-cs-cl
22 May 2026
Model Releases

Dissecting Embodied Abilities in Multimodal Language Models through Skill-level Evaluation and Diagnosis

DGX agent

arXiv:2510.08759v2 Announce Type: replace Abstract: Understanding the capability bottlenecks of embodied multimodal large language models (MLLMs) is crucial for improving embodied agents. However, exi

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

FRED: A Multi-Modal Autonomous Driving Dataset for Flooded Road Environments

DGX agent

arXiv:2605.22018v1 Announce Type: new Abstract: The Flooded Road Environments Dataset (FRED) is, to our knowledge, the first multi-modal autonomous driving dataset specifically targeting the collectio

model-releasesarxiv-cs-cv
22 May 2026
Research

From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification

DGX agent

arXiv:2605.22003v1 Announce Type: new Abstract: Sentiment analysis, also referred to as opinion mining, primarily tries to extract opinion from any text-based data. In the context of movie reviews and

researcharxiv-cs-cl
22 May 2026
Research

Geometry-Adaptive Explainer for Faithful Dictionary-Based Interpretability under Distribution Shift

DGX agent

arXiv:2605.21849v1 Announce Type: cross Abstract: Mechanistic interpretability aims to explain a model's behavior by identifying causally responsible internal structures. Dictionary-based explainers s

researcharxiv-cs-cl
22 May 2026
Safety

Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewriting

DGX agent

arXiv:2605.22258v1 Announce Type: new Abstract: Large language models (LLMs) require robust toxicity evaluation beyond explicit wording. This setting remains underexplored in Chinese, where toxicity m

safetyarxiv-cs-cl
22 May 2026
Model Releases

OSS: Open Suturing Skills Vision-Based Assessment Challenge 2024-2025

DGX agent

arXiv:2605.22200v1 Announce Type: new Abstract: Achieving high levels of surgical skill through effective training is essential for optimal patient outcomes. Automated, data-driven skill assessment ho

model-releasesarxiv-cs-cv
22 May 2026
Applications

PrivacyAkinator: Articulating Key Privacy Design Decisions by Answering LLM-Generated Multiple-choice Questions

DGX agent

arXiv:2605.20206v1 Announce Type: cross Abstract: NIST's Privacy Risk Assessment Methodology (PRAM) provides a structured framework for privacy experts to assess privacy risks. However, its complexity

applicationsarxiv-cs-ai
22 May 2026
Applications

SE3Kit: A Lightweight Python Library for Specialized Geometric Primitives in Robotics

DGX agent

arXiv:2605.22633v1 Announce Type: new Abstract: The Python robotics ecosystem faces a challenge: while many libraries exist for rigid body transformations, few are both lightweight and mathematically

applicationsarxiv-cs-ro
22 May 2026
Model Releases

Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System

DGX agent

arXiv:2605.20368v1 Announce Type: cross Abstract: Organizations that scan documents for sensitive information face a practical problem. Cloud services require data to be sent to external infrastructur

model-releasesarxiv-cs-ai
22 May 2026
Research

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

DGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

researcharxiv-cs-ro
22 May 2026
Applications

Synthetic Data Alone is Enough? Rethinking Data Scarcity in Pediatric Rare Disease Recognition

DGX agent

arXiv:2605.22767v1 Announce Type: new Abstract: Children with rare genetic diseases often exhibit distinctive facial phenotypes, yet developing computer vision systems for early diagnosis remains chal

applicationsarxiv-cs-cv
22 May 2026
Research

Tokenisation via Convex Relaxations

DGX agent

arXiv:2605.22821v1 Announce Type: new Abstract: Tokenisation is an integral part of the current NLP pipeline. Current tokenisation algorithms such as BPE and Unigram are greedy algorithms -- they make

researcharxiv-cs-cl
22 May 2026
Research

A Comprehensive Comparison of Deep Learning Architectures for COVID-19 Classification on CT & X-ray Imagery

DGX agent

arXiv:2605.20445v1 Announce Type: new Abstract: COVID-19 was a significant challenge that led to the loss of numerous lives daily. Not only a certain country was involved in this outbreak, but even th

researcharxiv-cs-cv
21 May 2026
Research

A Dialogue between Causal and Traditional Representation Learning: Toward Mutual Benefits in a Unified Formulation

DGX agent

arXiv:2605.21058v1 Announce Type: new Abstract: Causal representation learning (CRL) and traditional representation learning have largely developed along different trajectories. Traditional representa

researcharxiv-cs-lg
21 May 2026
Safety

AI-Assisted Competency Assessment from Egocentric Video in Simulation-Based Nursing Education

DGX agent

arXiv:2605.20233v1 Announce Type: new Abstract: Assessing learner competency in clinical simulation requires expert observation that is time-intensive, difficult to scale, and subject to inter-rater v

safetyarxiv-cs-cv
21 May 2026
Local Ai

AIGaitor: Privacy-preserving and cloud-free motion analysis for everyone, using edge computing

DGX agent

arXiv:2605.21421v1 Announce Type: new Abstract: Motion capture is the gold standard for measuring human movement, but clinical use remains limited by cost, technical complexity, and privacy concerns.

local-aiarxiv-cs-cv
21 May 2026
Research

AirfoilGen: A valid-by-construction and performance-aware latent diffusion model for airfoil generation

DGX agent

arXiv:2605.20303v1 Announce Type: new Abstract: Airfoil shape design is a fundamental task in aerospace engineering, with a direct impact on flight stability and fuel consumption. Deep learning has re

researcharxiv-cs-lg
21 May 2026
Research

ArPoMeme: An Annotated Arabic Multimodal Dataset for Political Ideology and Polarization

DGX agent

arXiv:2605.20967v1 Announce Type: new Abstract: Memes have become a prominent medium of political communication in the Arab world, reflecting how humor, imagery, and text interact to express ideologic

researcharxiv-cs-cl
21 May 2026
Agents

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents

DGX agent

arXiv:2605.20616v1 Announce Type: new Abstract: Language agents increasingly operate over streams of related tasks, yet existing memory systems struggle to convert accumulated experience into reusable

agentsarxiv-cs-cl
21 May 2026
Research

Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts

DGX agent

arXiv:2605.20610v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are often interpreted by analysing which categories are routed to which experts. However, routing alone does not reveal

researcharxiv-cs-cv
21 May 2026
Safety

Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs

DGX agent

arXiv:2605.21027v1 Announce Type: new Abstract: Enterprise analytics aims to make organizational data accessible for decision-making, yet non-technical users still face barriers when using traditional

safetyarxiv-cs-cl
21 May 2026
Tutorials

Building a Custom Taxonomy of AI Skills and Tasks from the Ground Up with Job Postings

DGX agent

arXiv:2605.21029v1 Announce Type: new Abstract: Utilizing LLMs for automated taxonomy construction presents a clear opportunity for the comprehensive, yet efficient mapping of potentially complex doma

tutorialsarxiv-cs-cl
21 May 2026
Research

Causal Machine Learning Is Not a Panacea: A Roadmap for Observational Causal Inference in Health

DGX agent

arXiv:2605.20782v1 Announce Type: new Abstract: Objective: The growing availability of large-scale observational clinical datasets and challenges in conducting randomized controlled trials have spurre

researcharxiv-cs-lg
21 May 2026
Applications

CoarseSoundNet: Building a reliable model for ecological soundscape analysis

DGX agent

arXiv:2605.21143v1 Announce Type: cross Abstract: A soundscape is composed of three types of sound: biophony (sounds made by animals), geophony (natural abiotic sounds) and anthropophony (sounds made

applicationsarxiv-cs-lg
21 May 2026
Research

Computational-Statistical Trade-off in Kernel Two-Sample Testing with Random Fourier Features

DGX agent

arXiv:2407.08976v2 Announce Type: replace-cross Abstract: Recent years have seen a surge in methods for two-sample testing, among which the Maximum Mean Discrepancy (MMD) test has emerged as an effect

researcharxiv-cs-lg
21 May 2026
Research

CRANE: Correcting Errors in Raw Nanopore Signals Using Hidden Markov Models

DGX agent

arXiv:2603.20420v2 Announce Type: replace-cross Abstract: Nanopore sequencing can read substantially longer sequences of nucleic acid molecules, called reads, than other sequencing methods, which has

researcharxiv-cs-lg
21 May 2026
Tutorials

Efficient numeracy in language models through single-token number embeddings

DGX agent

arXiv:2510.06824v2 Announce Type: replace Abstract: To drive progress in science and engineering, large language models (LLMs) must be able to process large amounts of numerical data and solve long ca

tutorialsarxiv-cs-lg
21 May 2026
Safety

FineVision: Open Data Is All You Need

DGX agent

arXiv:2510.17269v2 Announce Type: replace Abstract: The advancement of vision-language models (VLMs) is hampered by a fragmented landscape of inconsistent and contaminated public datasets. We introduc

safetyarxiv-cs-cv
21 May 2026
Research

FusionCell: Cross-Attentive Fusion of Layout Geometry and Netlist Topology for Standard-Cell Performance Prediction

DGX agent

arXiv:2605.20287v1 Announce Type: cross Abstract: Standard cells form the building blocks of digital circuits, so their delay and power critically influence chip-level performance; yet characterizatio

researcharxiv-cs-cv
21 May 2026
Research

Large Language Models Unpack Complex Political Opinions through Target-Stance Extraction

DGX agent

arXiv:2603.23531v2 Announce Type: replace Abstract: Political polarization emerges from a complex interplay of beliefs about policies, figures, and issues. However, most computational analyses reduce

researcharxiv-cs-cl
21 May 2026
Research

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

DGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

researcharxiv-cs-lg
21 May 2026
Model Releases

Learning-to-Defer with Expert-Conditional Advice

DGX agent

arXiv:2603.14324v3 Announce Type: replace-cross Abstract: Learning-to-Defer routes each input to the expert that minimizes expected cost, but it assumes that the information available to every expert

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

DGX agent

arXiv:2502.12120v3 Announce Type: replace-cross Abstract: Scaling laws guide the development of large language models (LLMs) by offering estimates for the optimal balance of model size, tokens, and co

model-releasesarxiv-cs-cl
21 May 2026
Safety

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

DGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

safetyarxiv-cs-lg
21 May 2026
Agents

Mem-pi: Adaptive Memory through Learning When and What to Generate

DGX agent

arXiv:2605.21463v1 Announce Type: new Abstract: We present Mem-pi, a framework for adaptive memory in large language model (LLM) agents, where useful guidance is generated on demand rather than retrie

agentsarxiv-cs-cl
21 May 2026
Tutorials

PACD-Net: Pseudo-Augmented Contrastive Distillation for Glycemic Control Estimation from SMBG

DGX agent

arXiv:2605.20751v1 Announce Type: new Abstract: Effective diabetes management requires continuous monitoring of glycemic levels. Clinically, glycemic control is assessed using metrics such as Time in

tutorialsarxiv-cs-lg
21 May 2026
Hardware

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR

DGX agent

arXiv:2605.20863v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has recently unlocked strong reasoning capabilities in large language models (LLMs), triggering

hardwarearxiv-cs-lg
21 May 2026
Model Releases

Post-Hoc Understanding of Metaphor Processing in Decoder-Only Language Models via Conditional Scale Entropy

DGX agent

arXiv:2605.21391v1 Announce Type: new Abstract: Metaphor requires a language model to resolve a token whose contextual meaning diverges from its basic literal sense. Understanding how transformer mode

model-releasesarxiv-cs-cl
21 May 2026
Research

Praxium: Diagnosing Cloud Anomalies with AI-based Telemetry and Dependency Analysis

DGX agent

arXiv:2603.23890v2 Announce Type: replace-cross Abstract: As the modern microservice architecture for cloud applications grows in popularity, cloud services are becoming increasingly complex and more

researcharxiv-cs-lg
21 May 2026
← Previous
1…8485868788…109
Next →