AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Research

Conformal C2ST: Turning weak classifiers into strong two-sample tests

DGX agent

arXiv:2507.17026v2 Announce Type: replace-cross Abstract: The two-sample testing problem, a fundamental task in statistics and machine learning, seeks to determine whether two sets of samples, drawn f

researcharxiv-cs-lg
1 Jun 2026
Applications
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Modal Clinical Knowledge Integration for Mammography Report Generation

DGX agent

arXiv:2605.31093v1 Announce Type: new Abstract: Breast cancer is a major global health concern, and mammography screening plays a central role in early detection. The large volume of screening examina

applicationsarxiv-cs-cv
1 Jun 2026
Safety

dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Development and Deployment

DGX agent

arXiv:2605.31360v1 Announce Type: cross Abstract: The Artificial Intelligence (AI) life cycle requires a thorough understanding of the underlying data dynamics for robust, safe and cost-effective AI d

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories

DGX agent

arXiv:2602.10809v2 Announce Type: replace Abstract: Existing multimodal retrieval systems excel at semantic matching but implicitly assume that query-image relevance can be measured in isolation. This

model-releasesarxiv-cs-cv
1 Jun 2026
Safety

DISCO: Mitigating Bias in Deep Learning with Conditional Distance Correlation

DGX agent

arXiv:2506.11653v3 Announce Type: replace-cross Abstract: Dataset bias often leads deep learning models to exploit spurious correlations instead of task-relevant signals. We introduce the Standard Ant

safetyarxiv-cs-ai
1 Jun 2026
Research

Discovering a Zeta Map Algorithm on Dyck Paths via Mechanistic Interpretability

DGX agent

arXiv:2605.30482v1 Announce Type: new Abstract: Machine learning is increasingly used in mathematical discovery, but in mathematics the desired output is often not a prediction itself, but an explicit

researcharxiv-cs-lg
1 Jun 2026
Research

Discovering Differences in Strategic Behavior Between Humans and LLMs

DGX agent

arXiv:2602.10324v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly deployed in social and strategic scenarios, it becomes critical to understand where and why their b

researcharxiv-cs-ai
1 Jun 2026
Model Releases

Do covariates explain why these groups differ? The choice of reference group can reverse conclusions in the Oaxaca-Blinder decomposition

DGX agent

arXiv:2603.29972v2 Announce Type: replace-cross Abstract: Scientists often want to explain why an outcome is different in two groups. For instance, differences in patient mortality rates across two ho

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision

DGX agent

arXiv:2605.31557v1 Announce Type: new Abstract: Continuous episodic memory is a core capability for autonomous agents operating in dynamic, real-world environments, yet current streaming video benchma

model-releasesarxiv-cs-cv
1 Jun 2026
Safety

From Evidence to Design: Developing an AI-Augmented UX Research Point of View for Digital Wellbeing in Emergency and Public Safety Contexts

DGX agent

arXiv:2605.31146v1 Announce Type: cross Abstract: This paper investigates how User Experience Research (UXR) methods can be combined with AI-supported analysis to develop clearer design direction for

safetyarxiv-cs-ai
1 Jun 2026
Tutorials

Function2Scene: 3D Indoor Scene Layout from Functional Specifications

DGX agent

arXiv:2605.30819v1 Announce Type: new Abstract: Most text-driven 3D indoor scene synthesis methods generate rooms from object-centric prompts, asking what furniture should be placed rather than how th

tutorialsarxiv-cs-cv
1 Jun 2026
Applications

GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration

DGX agent

arXiv:2605.31039v1 Announce Type: new Abstract: Real-world image restoration (IR) is bottlenecked by the scarcity of high-quality paired training data. Synthetic datasets are abundant but often fail t

applicationsarxiv-cs-cv
1 Jun 2026
Model Releases

GUI-C^2: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning

DGX agent

arXiv:2605.30884v1 Announce Type: new Abstract: Existing agentic reinforcement learning methods for GUI grounding have limitations at two levels. At the data level, current approaches typically treat

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs

DGX agent

arXiv:2511.18760v2 Announce Type: replace Abstract: Informal mathematics has been central to modern large language model (LLM) reasoning, offering flexibility and efficient construction of arguments.

model-releasesarxiv-cs-ai
1 Jun 2026
Research

How well does Classification Accuracy capture Concept Drift Detection Quality? An overview of Concept Drift Detection evaluation

DGX agent

arXiv:2605.31186v1 Announce Type: new Abstract: Data streams are nowadays among the most frequently analyzed data structures, with the concept drift posing a major challenge encountered by processing

researcharxiv-cs-lg
1 Jun 2026
Research

Is the Last Layer Sufficient for Uncertainty Quantification?

DGX agent

arXiv:2605.30741v1 Announce Type: cross Abstract: Epistemic uncertainty quantification (UQ) for deep neural networks (DNNs) is a requirement for safe adoption of AI in mission-critical settings. Sever

researcharxiv-cs-lg
1 Jun 2026
Local Ai

iVGR: Internalizing Visually Grounded Reasoning for MLLMs with Reinforcement Learning

DGX agent

arXiv:2605.31096v1 Announce Type: new Abstract: While visually grounded Chain-of-Thought (CoT) has emerged as a promising paradigm to enhance fine-grained perception in multimodal large language model

local-aiarxiv-cs-cv
1 Jun 2026
Research

Kalimati Vegetable Price Index Forecasting with a Momentum Corrected Online Stacking Ensemble

DGX agent

arXiv:2605.30720v1 Announce Type: cross Abstract: Forecasting agricultural commodity prices in emerging economies is difficult due to high volatility, frequent supply disruptions, and strong cultural

researcharxiv-cs-ai
1 Jun 2026
Applications

Learning to Solve and Optimize by Evolving Code

DGX agent

arXiv:2605.31049v1 Announce Type: cross Abstract: Combinatorial and optimization problems are fundamental to many industrial AI applications. Solving large-scale real-world instances of such problems

applicationsarxiv-cs-ai
1 Jun 2026
Agents

LH-Bench: Skill-Grounded Evaluation of Long-Horizon Agents on Subjective Enterprise Tasks

DGX agent

arXiv:2603.22744v2 Announce Type: replace Abstract: Large language models excel on objectively verifiable tasks such as math and programming, where evaluation reduces to unit tests or a single correct

agentsarxiv-cs-ai
1 Jun 2026
Local Ai

LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability

DGX agent

arXiv:2605.31167v1 Announce Type: new Abstract: Assessing whether Large Language Models outputs are factually grounded, epistemically calibrated, and methodologically reproducible is a prerequisite fo

local-aiarxiv-cs-ai
1 Jun 2026
Safety

LVSA: Training-Free Sparse Attention for Long Video Diffusion

DGX agent

arXiv:2605.31057v1 Announce Type: new Abstract: Dense self-attention is the compute and quality bottleneck of long-video diffusion inference: cost grows quadratically with the sequence length, and bey

safetyarxiv-cs-cv
1 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Safety

PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation

DGX agent

arXiv:2601.11702v3 Announce Type: replace-cross Abstract: AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates

safetyarxiv-cs-ai
1 Jun 2026
Agents

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

DGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

agentsarxiv-cs-ai
1 Jun 2026
Research

ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving

DGX agent

arXiv:2502.04671v3 Announce Type: replace Abstract: Neural approaches to theorem proving require robust infrastructure for interfacing with interactive theorem provers (ITPs), extracting structured pr

researcharxiv-cs-ai
1 Jun 2026
Model Releases

Re-examining Low Rank adaptation for private LLM fine-tuning

DGX agent

arXiv:2510.01137v3 Announce Type: replace Abstract: Privacy is a central concern when fine-tuning large language models (LLMs) on sensitive data, and differentially private stochastic gradient descent

model-releasesarxiv-cs-lg
1 Jun 2026
Research

Separating Secrets from Placeholders: A Hybrid CNN-CodeBERT Framework for Three-Class Credential Leakage Detection

DGX agent

arXiv:2605.31520v1 Announce Type: cross Abstract: Credential leakage in public source code repositories poses a critical security threat, with over 23.8 million secrets exposed in 2024 alone. Existing

researcharxiv-cs-ai
1 Jun 2026
Research

SPECTRA: Synthetic IR Test Collections with Relevance Oracles and Controlled Distractor Diagnostics

DGX agent

arXiv:2605.31575v1 Announce Type: cross Abstract: Scalable information retrieval testing needs corpora that are large enough to stress index construction, ranking latency, query routing, and evaluatio

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Spectral Reach: Understanding Neural Scaling as Progress into the Spectral Tail

DGX agent

arXiv:2605.31244v1 Announce Type: new Abstract: Neural scaling laws describe predictable power-law relationships between model size, dataset size, compute, and performance. While these laws guide the

tutorialsarxiv-cs-lg
1 Jun 2026
Safety

UXR PoV for Neuroinclusive Emotion Regulation

DGX agent

arXiv:2605.31131v1 Announce Type: cross Abstract: Attention-deficit/hyperactivity disorder (ADHD) is a psychiatric disorder which presents itself in individuals through patterns of developmentally ina

safetyarxiv-cs-ai
1 Jun 2026
Safety

Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems

DGX agent

arXiv:2506.00175v5 Announce Type: replace-cross Abstract: Modern AI systems are typically developed through multiple stages-pretraining, fine-tuning rounds, and subsequent adaptation or alignment, whe

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Adapting Multilingual Embedding Models to Turkish via Cross-Lingual Tokenizer Surgery and Offline Distillation

DGX agent

arXiv:2605.29992v1 Announce Type: new Abstract: Sentence embeddings are a foundational component for semantic search, clustering, classification, and retrieval-augmented generation. This paper present

model-releasesarxiv-cs-cl
29 May 2026
Agents

An Approach for Thyroid Nodule Analysis Using Thermographic Images

DGX agent

arXiv:2605.29221v1 Announce Type: new Abstract: Thyroid cancer is said to be the second most common type of cancer in female individuals and the third in males by 2030, according to projections. In ge

agentsarxiv-cs-cv
29 May 2026
Research

Bridging Functional and Representational Similarity via Usable Information

DGX agent

arXiv:2601.21568v2 Announce Type: replace Abstract: We present a unified framework for quantifying the similarity between representations through the lens of extit{usable} information, offering a rigo

researcharxiv-cs-lg
29 May 2026
Model Releases

CalArena: A Large-Scale Post-Hoc Calibration Benchmark

DGX agent

arXiv:2605.30188v1 Announce Type: cross Abstract: Reliable probability estimates are critical in many machine learning applications, yet modern classifiers are often poorly calibrated. Post-hoc calibr

model-releasesarxiv-cs-ai
29 May 2026
Agents

Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agent

DGX agent

arXiv:2603.01311v2 Announce Type: replace Abstract: The discovery of novel catalysts tailored for particular applications is a major challenge for the twenty-first century. Traditional methods for thi

agentsarxiv-cs-cl
29 May 2026
Agents

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

DGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

agentsarxiv-cs-ai
29 May 2026
Research

Cross-Chirality Generalization by Axial Vectors for Hetero-Chiral Protein-Peptide Interaction Design

DGX agent

arXiv:2602.20176v2 Announce Type: replace-cross Abstract: D-peptide binders targeting L-proteins have promising therapeutic potential. Despite rapid advances in machine learning-based target-condition

researcharxiv-cs-lg
29 May 2026
Model Releases

DiffSpot: Can VLMs Spot Fine-Grained Visual Differences in Web Interfaces?

DGX agent

arXiv:2605.29615v1 Announce Type: cross Abstract: Vision-language models (VLMs) have made strong progress on high-level image-text alignment, yet their ability to perceive subtle visual differences re

model-releasesarxiv-cs-cl
29 May 2026
Research

Digitally enriching a screening population for pancreatic cancer using routine blood-based measures and clinical histories

DGX agent

arXiv:2605.30275v1 Announce Type: new Abstract: Earlier detection of pancreatic cancer is key to enabling wider access to curative treatment and reducing cancer deaths; however, screening is presently

researcharxiv-cs-lg
29 May 2026
Agents

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

DGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

agentsarxiv-cs-ai
29 May 2026
Research

DLT-Corpus: A Large-Scale Text Collection for the Distributed Ledger Technology Domain

DGX agent

arXiv:2602.22045v2 Announce Type: replace Abstract: We introduce DLT-Corpus, the largest domain-specific text collection for Distributed Ledger Technology (DLT) research to date: 2.98 billion tokens f

researcharxiv-cs-cl
29 May 2026
Safety

Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure

DGX agent

arXiv:2602.08783v3 Announce Type: replace Abstract: Latent or continuous chain-of-thought methods replace explicit textual rationales with a number of internal latent steps, but these intermediate com

safetyarxiv-cs-ai
29 May 2026
Applications

Enhancing Reinforcement Learning in 3D Environments through Semantic Segmentation: A Case Study in ViZDoom

DGX agent

arXiv:2511.11703v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in 3D environments with high-dimensional sensory input poses two major challenges: (1) the high memory consumption

applicationsarxiv-cs-ai
29 May 2026
Model Releases

Entity-Collision: A Stratified Protocol for Attributing Retrieval Lift in Agent Memory

DGX agent

arXiv:2605.29630v1 Announce Type: cross Abstract: End-to-end agent-memory benchmarks report a single hit@k per retriever, confounding lexical leakage (uncontrolled query/gold/distractor entity overlap

model-releasesarxiv-cs-ai
29 May 2026
Agents

Estimating the Empowerment of Language Model Agents

DGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

agentsarxiv-cs-ai
29 May 2026
Tutorials

Explaining Concept Shift with Interpretable Feature Attribution

DGX agent

arXiv:2505.20634v2 Announce Type: replace Abstract: Concept shift occurs when the distribution of labels conditioned on the features changes between domains, which can make even a well-tuned ML model

tutorialsarxiv-cs-lg
29 May 2026
← Previous
1…7980818283…109
Next →