AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
23 Apr 2026

Cover meets Robbins while Betting on Bounded Data: ln n Regret and Almost Sure lnln n Regret

ResearchDGX agent

arXiv:2604.20172v1 Announce Type: new Abstract: Consider betting against a sequence of data in [0,1], where one is allowed to make any bet that is fair if the data have a conditional mean m_0 in (0,1)

Coverage, Not Averages: Semantic Stratification for Trustworthy Retrieval Evaluation

SafetyDGX agent

arXiv:2604.20763v1 Announce Type: cross Abstract: Retrieval quality is the primary bottleneck for accuracy and robustness in retrieval-augmented generation (RAG). Current evaluation relies on heuristi

CrackForward: Context-Aware Severity Stage Crack Synthesis for Data Augmentation

ResearchDGX agent

arXiv:2604.19941v1 Announce Type: new Abstract: Reliable crack detection and segmentation are vital for structural health monitoring, yet the scarcity of well-annotated data constitutes a major challe

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CRAFT: Training-Free Cascaded Retrieval for Tabular QA

Model ReleasesDGX agent

arXiv:2505.14984v2 Announce Type: replace Abstract: Open-Domain Table Question Answering (TQA) involves retrieving relevant tables from a large corpus to answer natural language queries. Traditional d

CreativeGame:Toward Mechanic-Aware Creative Game Generation

AgentsDGX agent

arXiv:2604.19926v1 Announce Type: new Abstract: Large language models can generate plausible game code, but turning this capability into iterative creative improvement remains difficult. In practice,

Cross-Modal Taxonomic Generalization in (Vision-) Language Models

ApplicationsDGX agent

arXiv:2603.07474v2 Announce Type: replace-cross Abstract: What is the interplay between semantic representations learned by language models (LM) from surface form alone to those learned from more grou

CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction

SafetyDGX agent

arXiv:2505.04897v2 Announce Type: replace-cross Abstract: Interactive imitation learning makes an agent's control policy robust by stepwise supervisions from an expert. The recent algorithms mostly em

CXR-LanIC: Language-Grounded Interpretable Classifier for Chest X-Ray Diagnosis

ResearchDGX agent

arXiv:2510.21464v2 Announce Type: replace Abstract: Deep learning models have achieved remarkable accuracy in chest X-ray diagnosis, yet their widespread clinical adoption remains limited by the black

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

Model ReleasesDGX agent

arXiv:2604.20389v1 Announce Type: cross Abstract: The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against

DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles

SafetyDGX agent

arXiv:2604.20771v1 Announce Type: cross Abstract: The Internet of Vehicles (IoV) is advancing modern transportation by improving safety, efficiency, and intelligence. However, the reliance on the Cont

Decentralized Machine Learning with Centralized Performance Guarantees via Gibbs Algorithms

SafetyDGX agent

arXiv:2604.20492v1 Announce Type: cross Abstract: In this paper, it is shown, for the first time, that centralized performance is achievable in decentralized learning without sharing the local dataset

Decision-Focused Federated Learning Under Heterogeneous Objectives and Constraints

AgentsDGX agent

arXiv:2604.20031v1 Announce Type: cross Abstract: We consider what we refer to as {Decision-Focused Federated Learning (DFFL)} framework, i.e., a predict-then-optimize approach employed by a collectio

Decoding Text Spans for Efficient and Accurate Named-Entity Recognition

Local AiDGX agent

arXiv:2604.20447v1 Announce Type: new Abstract: Named Entity Recognition (NER) is a key component in industrial information extraction pipelines, where systems must satisfy strict latency and throughp

Deconstructing Superintelligence: Identity, Self-Modification and Differance

ResearchDGX agent

arXiv:2604.19845v1 Announce Type: new Abstract: Self-modification is often taken as constitutive of artificial superintelligence (SI), yet modification is a relative action requiring a supplement outs

Degrees, Levels, and Profiles of Contextuality

ResearchDGX agent

arXiv:2603.26692v3 Announce Type: replace-cross Abstract: We introduce a new notion, that of a contextuality profile of a system of random variables. Rather than characterizing a system's contextualit

Depression Risk Assessment in Social Media via Large Language Models

ResearchDGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

Model ReleasesDGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation

AgentsDGX agent

arXiv:2604.20841v1 Announce Type: new Abstract: Recent advances in video generative models enable the synthesis of realistic human-object interaction videos across a wide range of scenarios and object

Device-Native Autonomous Agents for Privacy-Preserving Negotiations

Local AiDGX agent

arXiv:2601.00911v3 Announce Type: replace-cross Abstract: Automated negotiations in insurance and business-to-business (B2B) commerce encounter substantial challenges. Current systems force a trade-of

Diagnosing CFG Interpretation in LLMs

SafetyDGX agent

arXiv:2604.20811v1 Announce Type: new Abstract: As LLMs are increasingly integrated into agentic systems, they must adhere to dynamically defined, machine-interpretable interfaces. We evaluate LLMs as

Diagnosing Urban Street Vitality via a Visual-Semantic and Spatiotemporal Framework for Street-Level Economics

ResearchDGX agent

arXiv:2604.19798v1 Announce Type: cross Abstract: Micro-scale street-level economic assessment is fundamental for precision spatial resource allocation. While Street View Imagery (SVI) advances urban

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

Model ReleasesDGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

Differentiable Conformal Training for LLM Reasoning Factuality

Model ReleasesDGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

Differentially Private Clustered Federated Learning with Privacy-Preserving Initialization and Normality-Driven Aggregation

ResearchDGX agent

arXiv:2604.20596v1 Announce Type: new Abstract: Federated learning (FL) enables training of a global model while keeping raw data on end-devices. Despite this, FL has shown to leak private user inform

DISCA: A Digital In-memory Stochastic Computing Architecture Using A Compressed Bent-Pyramid Format

ApplicationsDGX agent

arXiv:2511.17265v2 Announce Type: replace-cross Abstract: Nowadays, we are witnessing an Artificial Intelligence revolution that dominates the technology landscape in various application domains, such

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

Model ReleasesDGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

Distributional Inverse Reinforcement Learning

SafetyDGX agent

arXiv:2510.03013v3 Announce Type: replace Abstract: We propose a distributional framework for offline Inverse Reinforcement Learning (IRL) that jointly models uncertainty over reward functions and ful

Distributional Value Estimation Without Target Networks for Robust Quality-Diversity

ResearchDGX agent

arXiv:2604.20381v1 Announce Type: new Abstract: Quality-Diversity (QD) algorithms excel at discovering diverse repertoires of skills, but are hindered by poor sample efficiency and often require tens

Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs

ApplicationsDGX agent

arXiv:2604.19765v1 Announce Type: cross Abstract: Recent work identifies a sparse set of 'hallucination neurons' (H-neurons), less than 0.1% of feed-forward network neurons, that reliably predict when

Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment

Model ReleasesDGX agent

arXiv:2604.19781v1 Announce Type: cross Abstract: Automated scoring of student work at scale requires balancing accuracy against cost and latency. In 'cascade' systems, small language models (LMs) han

Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models

ResearchDGX agent

arXiv:2604.01965v2 Announce Type: replace-cross Abstract: Scientific knowledge discovery increasingly relies on large language models, yet many existing scholarly assistants depend on proprietary syst

DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data

AgentsDGX agent

arXiv:2604.19859v1 Announce Type: cross Abstract: Edge-scale deep research agents based on small language models are attractive for real-world deployment due to their advantages in cost, latency, and

DRIV-EX: Counterfactual Explanations for Driving LLMs

SafetyDGX agent

arXiv:2603.00696v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as reasoning engines in autonomous driving, yet their decision-making remains opaque. We propose

Dual Causal Inference: Integrating Backdoor Adjustment and Instrumental Variable Learning for Medical VQA

Model ReleasesDGX agent

arXiv:2604.20306v1 Announce Type: cross Abstract: Medical Visual Question Answering (MedVQA) aims to generate clinically reliable answers conditioned on complex medical images and questions. However,

Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving

AgentsDGX agent

arXiv:2604.20183v1 Announce Type: new Abstract: Large Language Models (LLMs) often struggle with structural ambiguity in optimization problems, where a single problem admits multiple related but confl

Duluth at SemEval-2026 Task 6: DeBERTa with LLM-Augmented Data for Unmasking Political Question Evasions

Model ReleasesDGX agent

arXiv:2604.20168v1 Announce Type: new Abstract: This paper presents the Duluth approach to SemEval-2026 Task 6 on CLARITY: Unmasking Political Question Evasions. We address Task 1 (clarity-level class

DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion

Local AiDGX agent

arXiv:2604.20470v1 Announce Type: new Abstract: Leveraging the natural spatiotemporal energy decay in video diffusion offers a path to efficiency, yet relying solely on rigid static masks risks losing

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

Model ReleasesDGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

Effects of Cross-lingual Evidence in Multilingual Medical Question Answering

ResearchDGX agent

arXiv:2604.20531v1 Announce Type: new Abstract: This paper investigates Multilingual Medical Question Answering across high-resource (English, Spanish, French, Italian) and low-resource (Basque, Kazak

Efficient INT8 Single-Image Super-Resolution via Deployment-Aware Quantization and Teacher-Guided Training

ApplicationsDGX agent

arXiv:2604.20291v1 Announce Type: new Abstract: Efficient single-image super-resolution (SISR) requires balancing reconstruction fidelity, model compactness, and robustness under low-bit deployment, w

Efficient Multi-Cohort Inference for Long-Term Effects and Lifetime Value in A/B Testing with User Learning

ResearchDGX agent

arXiv:2604.20777v1 Announce Type: new Abstract: In streaming platforms churn is extremely costly, yet A/B tests are typically evaluated using outcomes observed within a limited experimental horizon. E

Efficient Reinforcement Learning using Linear Koopman Dynamics for Nonlinear Robotic Systems

SafetyDGX agent

arXiv:2604.19980v1 Announce Type: new Abstract: This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed appro

Efficient Symbolic Computations for Identifying Causal Effects

TutorialsDGX agent

arXiv:2604.20516v1 Announce Type: cross Abstract: Determining identifiability of causal effects from observational data under latent confounding is a central challenge in causal inference. For linear

Efficient Test-Time Inference via Deterministic Exploration of Truncated Decoding Trees

ResearchDGX agent

arXiv:2604.20500v1 Announce Type: new Abstract: Self-consistency boosts inference-time performance by sampling multiple reasoning traces in parallel and voting. However, in constrained domains like ma

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

Model ReleasesDGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

Efficiently Closing Loops in LiDAR-Based SLAM Using Point Cloud Density Maps

SafetyDGX agent

arXiv:2501.07399v2 Announce Type: replace Abstract: Consistent maps are key for most autonomous mobile robots, and they often use SLAM approaches to build such maps. Loop closures via place recognitio

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

SafetyDGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

Emergence Transformer: Dynamical Temporal Attention Matters

ResearchDGX agent

arXiv:2604.19816v1 Announce Type: new Abstract: The Transformer, a breakthrough architecture in artificial intelligence, owes its success to the attention mechanism, which utilizes long-range interact

Energy-Based Open-Set Active Learning for Object Classification

ApplicationsDGX agent

arXiv:2604.20083v1 Announce Type: cross Abstract: Active learning (AL) has emerged as a crucial methodology for minimizing labeling costs in deep learning by selecting the most valuable samples from a

Enhancing ASR Performance in the Medical Domain for Dravidian Languages

TutorialsDGX agent

arXiv:2604.19797v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) for low-resource Dravidian languages like Telugu and Kannada faces significant challenges in specialized medical do

Enhancing Research Idea Generation through Combinatorial Innovation and Multi-Agent Iterative Search Strategies

AgentsDGX agent

arXiv:2604.20548v1 Announce Type: cross Abstract: Scientific progress depends on the continual generation of innovative re-search ideas. However, the rapid growth of scientific literature has greatly

Enhancing Speaker Verification with Whispered Speech via Post-Processing

ResearchDGX agent

arXiv:2604.20229v1 Announce Type: cross Abstract: Speaker verification is a task of confirming an individual's identity through the analysis of their voice. Whispered speech differs from phonated spee

Environmental Understanding Vision-Language Model for Embodied Agent

SafetyDGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

Epistemic Constitutionalism Or: how to avoid coherence bias

SafetyDGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

Epistemology gives a Future to Complementarity in Human-AI Interactions

SafetyDGX agent

arXiv:2601.09871v2 Announce Type: replace Abstract: Human-AI complementarity is the claim that a human supported by an AI system can outperform either alone in a decision-making process. Since its int

ESGLens: An LLM-Based RAG Framework for Interactive ESG Report Analysis and Score Prediction

ResearchDGX agent

arXiv:2604.19779v1 Announce Type: new Abstract: Environmental, Social, and Governance (ESG) reports are central to investment decision-making, yet their length, heterogeneous content, and lack of stan

ETac: A Lightweight and Efficient Tactile Simulation Framework for Learning Dexterous Manipulation

SafetyDGX agent

arXiv:2604.20295v1 Announce Type: new Abstract: Tactile sensors are increasingly integrated into dexterous robotic manipulators to enhance contact perception. However, learning manipulation policies t

Evaluating Assurance Cases as Text-Attributed Graphs for Structure and Provenance Analysis

SafetyDGX agent

arXiv:2604.20577v1 Announce Type: cross Abstract: An assurance case is a structured argument document that justifies claims about a system's requirements or properties, which are supported by evidence

Evaluating Black-Box Vulnerabilities with Wasserstein-Constrained Data Perturbations

SafetyDGX agent

arXiv:2603.15867v2 Announce Type: replace Abstract: The growing use of Machine Learning (ML) tools comes with critical challenges, such as limited model explainability. We propose a global explainabil

Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models

Model ReleasesDGX agent

arXiv:2508.17761v3 Announce Type: replace Abstract: In safety-critical applications data-driven models must not only be accurate but also provide reliable uncertainty estimates. This property, commonl

← Previous
1…860861862863864…998
Next →