AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
Human
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
20 Apr 2026

Towards Universal Convergence of Backward Error in Linear System Solvers

Model ReleasesDGX agent

arXiv:2604.16075v1 Announce Type: cross Abstract: The quest for an algorithm that solves an nimes n linear system in O(n^2) time complexity, or O(n^2 ext{poly}(1/epsilon)) when solving up to epsilon r

TPA: Next Token Probability Attribution for Detecting Hallucinations in RAG

ResearchDGX agent

arXiv:2512.07515v4 Announce Type: replace-cross Abstract: Detecting hallucinations in Retrieval-Augmented Generation remains a challenge. Prior approaches attribute hallucinations to a binary conflict

Training Flow Matching: The Role of Weighting and Parameterization

ResearchDGX agent

arXiv:2603.06454v2 Announce Type: replace Abstract: We study the training objectives of denoising-based generative models, with a particular focus on loss weighting and output parameterization, includ

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Training Time Prediction for Mixed Precision-based Distributed Training

ResearchDGX agent

arXiv:2604.16145v1 Announce Type: cross Abstract: Accurate prediction of training time in distributed deep learning is crucial for resource allocation, cost estimation, and job scheduling. We observe

Trajectory Planning for Safe Dual Control with Active Exploration

SafetyDGX agent

arXiv:2604.15507v1 Announce Type: new Abstract: Planning safe trajectories under model uncertainty is a fundamental challenge. Robust planning ensures safety by considering worst-case realizations, ye

Transfer Learning from Foundational Optimization Embeddings to Unsupervised SAT Representations

ResearchDGX agent

arXiv:2604.15448v1 Announce Type: cross Abstract: Foundational optimization embeddings have recently emerged as powerful pre-trained representations for mixed-integer programming (MIP) problems. These

Transformer Neural Processes - Kernel Regression

Model ReleasesDGX agent

arXiv:2411.12502v4 Announce Type: replace-cross Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic p

TriagerX: Dual Transformers for Bug Triaging Tasks with Content and Interaction Based Rankings

ApplicationsDGX agent

arXiv:2508.16860v2 Announce Type: replace-cross Abstract: Pretrained Language Models or PLMs are transformer-based architectures that can be used in bug triaging tasks. PLMs can better capture token s

TRIDENT: Enhancing Large Language Model Safety with Tri-Dimensional Diversified Red-Teaming Data Synthesis

Model ReleasesDGX agent

arXiv:2505.24672v2 Announce Type: replace Abstract: Large Language Models (LLMs) excel in various natural language processing tasks but remain vulnerable to generating harmful content or being exploit

Truncated Kernel Stochastic Gradient Descent with General Losses and Spherical Radial Basis Functions

SafetyDGX agent

arXiv:2510.04237v5 Announce Type: replace Abstract: In this paper, we propose a novel kernel stochastic gradient descent (SGD) algorithm for large-scale supervised learning with general losses. Compar

TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models

ResearchDGX agent

arXiv:2604.15756v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP exhibit strong Out-of-distribution (OOD) detection capabilities by aligning visual and textual representation

TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.15950v1 Announce Type: new Abstract: Pancreatic ductal adenocarcinoma (PDAC) segmentation on contrast-enhanced CT is inherently ambiguous: inter-rater disagreement among experts reflects ge

Two-Dimensional Deep ReLU CNN Approximation for Korobov Functions: A Constructive Approach

Model ReleasesDGX agent

arXiv:2503.07976v2 Announce Type: replace-cross Abstract: This paper investigates approximation capabilities of two-dimensional (2D) deep convolutional neural networks (CNNs), with Korobov functions s

Two-Stage Framework for Efficient UAV-Based Wildfire Video Analysis with Adaptive Compression and Fire Source Detection

Local AiDGX agent

arXiv:2508.16739v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) have become increasingly important in disaster emergency response by facilitating aerial video analysis. Due to the

TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2604.15967v1 Announce Type: cross Abstract: Despite the remarkable synthesis capabilities of text-to-image (T2I) models, safeguarding them against content violations remains a persistent challen

UA-Net: Uncertainty-Aware Network for TRISO Image Semantic Segmentation

ResearchDGX agent

arXiv:2604.15542v1 Announce Type: new Abstract: Tristructural isotropic (TRISO)-coated particle fuels undergo dimensional changes and chemical reactions during high-temperature neutron irradiation. Po

Uncertainty, Vagueness, and Ambiguity in Human-Robot Interaction: Why Conceptualization Matters

ResearchDGX agent

arXiv:2604.15339v1 Announce Type: cross Abstract: Uncertainty, vagueness, and ambiguity are closely related and often confused concepts in human-robot interaction (HRI). In earlier studies, these conc

Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation

ResearchDGX agent

arXiv:2511.02626v3 Announce Type: replace Abstract: Prior works have shown that fine-tuning on new knowledge can induce factual hallucinations in large language models (LLMs), leading to incorrect out

UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs

Model ReleasesDGX agent

arXiv:2604.15871v1 Announce Type: cross Abstract: The evaluation of visual editing models remains fragmented across methods and modalities. Existing benchmarks are often tailored to specific paradigms

Univariate Channel Fusion for Multivariate Time Series Classification

ResearchDGX agent

arXiv:2604.16119v1 Announce Type: new Abstract: Multivariate time series classification (MTSC) plays a crucial role in various domains, including biomedical signal analysis and motion monitoring. Howe

Unsupervised domain adaptation for radioisotope identification in gamma spectroscopy

SafetyDGX agent

arXiv:2603.05719v2 Announce Type: replace Abstract: Training machine learning models for radioisotope identification using gamma spectroscopy remains an elusive challenge for many practical applicatio

Unveiling Stochasticity: Universal Multi-modal Probabilistic Modeling for Traffic Forecasting

ApplicationsDGX agent

arXiv:2604.16084v1 Announce Type: cross Abstract: Traffic forecasting is a challenging spatio-temporal modeling task and a critical component of urban transportation management. Current studies mainly

UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval

SafetyDGX agent

arXiv:2604.15827v1 Announce Type: cross Abstract: Conventional information retrieval is concerned with identifying the relevance of texts for a given query. Yet, the conventional definition of relevan

Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing

ApplicationsDGX agent

arXiv:2604.16280v1 Announce Type: new Abstract: Explaining Machine Learning (ML) results in a transparent and user-friendly manner remains a challenging task of Explainable Artificial Intelligence (XA

VADF: Vision-Adaptive Diffusion Policy Framework for Efficient Robotic Manipulation

SafetyDGX agent

arXiv:2604.15938v1 Announce Type: new Abstract: Diffusion policies are becoming mainstream in robotic manipulation but suffer from hard negative class imbalance due to uniform sampling and lack of sam

VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects

Model ReleasesDGX agent

arXiv:2604.16272v1 Announce Type: cross Abstract: As AI-assisted video creation becomes increasingly practical, instruction-guided video editing has become essential for refining generated or captured

VeriCWEty: Embedding enabled Line-Level CWE Detection in Verilog

ResearchDGX agent

arXiv:2604.15375v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown significant improvement in RTL code generation. Despite the advances, the generated code is often riddled with

Verification Modulo Tested Library Contracts

ResearchDGX agent

arXiv:2604.15533v1 Announce Type: cross Abstract: We consider the problem of verification modulo tested library contracts as a step towards automating the verification of client programs that use comp

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

AgentsDGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

AgentsDGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

VeRVE: Versatile Retrieval for Videos via Unified Embeddings

Local AiDGX agent

arXiv:2601.12193v3 Announce Type: replace Abstract: Modern video retrieval systems are expected to handle diverse tasks ranging from corpus-level retrieval, fine-grained moment localization to flexibl

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

ResearchDGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools

ResearchDGX agent

arXiv:2510.08480v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable potential in bridging visual and textual reasoning, yet their reliance on text

vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2603.13966v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are increasingly evaluated across multiple simulation benchmarks, yet adding each benchmark to an evaluation pip

VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models

Model ReleasesDGX agent

arXiv:2512.14554v5 Announce Type: replace-cross Abstract: The rapid advancement of large language models (LLMs) has enabled new possibilities for applying artificial intelligence within the legal doma

VoodooNet: Achieving Analytic Ground States via High-Dimensional Random Projections

Local AiDGX agent

arXiv:2604.15613v1 Announce Type: cross Abstract: We present VoodooNet, a non-iterative neural architecture that replaces the stochastic gradient descent (SGD) paradigm with a closed-form analytic sol

WARBERT: A Hierarchical BERT-based Model for Web API Recommendation

ResearchDGX agent

arXiv:2509.23175v2 Announce Type: replace-cross Abstract: With the rise of Web 2.0 and microservices, the increasing availability of Web APIs has intensified the need for effective recommendation syst

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions

Model ReleasesDGX agent

arXiv:2604.15823v1 Announce Type: new Abstract: Embodied robotic agents often perceive movies through an egocentric screen-view interface rather than native cinematic footage, introducing domain shift

Weak-Link Optimization for Multi-Agent Reasoning and Collaboration

AgentsDGX agent

arXiv:2604.15972v1 Announce Type: new Abstract: LLM-driven multi-agent frameworks address complex reasoning tasks through multi-role collaboration. However, existing approaches often suffer from reaso

Weak-to-Strong Knowledge Distillation Accelerates Visual Learning

ResearchDGX agent

arXiv:2604.15451v1 Announce Type: new Abstract: Large-scale visual learning is increasingly limited by training cost. Existing knowledge distillation methods transfer from a stronger teacher to a weak

(Weighted) Adaptive Radius Near Neighbor Search: Evaluation for WiFi Fingerprint-based Positioning

ResearchDGX agent

arXiv:2604.15940v1 Announce Type: new Abstract: Fixed Radius Near Neighbor (FRNN) search is an alternative to the widely used k Nearest Neighbors (kNN) search. Unlike kNN, FRNN determines a label or a

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

SafetyDGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

When Cultures Meet: Multicultural Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2502.15972v2 Announce Type: replace-cross Abstract: Text-to-image generation models have achieved strong performance in culturally homogeneous settings, yet their ability to generate multicultur

When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth

ResearchDGX agent

arXiv:2604.15764v1 Announce Type: cross Abstract: Early-exit neural networks enable adaptive computation by allowing confident predictions to exit at intermediate layers, achieving 2-8imes inference s

When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models

SafetyDGX agent

arXiv:2510.09689v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been augmented with web search to overcome the limitations of the static knowledge boundary by accessing up-

When Surfaces Lie: Exploiting Wrinkle-Induced Attention Shift to Attack Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.27759v3 Announce Type: replace Abstract: Visual-Language Models (VLMs) have demonstrated exceptional cross-modal understanding across various tasks, including zero-shot classification, imag

When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems

ApplicationsDGX agent

arXiv:2604.15343v1 Announce Type: cross Abstract: We report a detailed autoethnographic case study of a single-subject who deliberately constructed and operated a multi-modal prompt-engineering system

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

Local AiDGX agent

arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent

Where does output diversity collapse in post-training?

ResearchDGX agent

arXiv:2604.16027v1 Announce Type: cross Abstract: Post-trained language models produce less varied outputs than their base counterparts. This output diversity collapse undermines inference-time scalin

Whose Facts Win? LLM Source Preferences under Knowledge Conflicts

SafetyDGX agent

arXiv:2601.03746v3 Announce Type: replace Abstract: As large language models (LLMs) are more frequently used in retrieval-augmented generation pipelines, it is increasingly relevant to study their beh

Why Colors Make Clustering Harder:Global Integrality Gaps, the Price of Fairness, and Color-Coupled Algorithms in Chromatic Correlation Clustering

SafetyDGX agent

arXiv:2604.15738v1 Announce Type: new Abstract: Chromatic Correlation Clustering (CCC) extends Correlation Clustering by assigning semantic colors to edges and requiring each cluster to receive a sing

Why Fine-Tuning Encourages Hallucinations and How to Fix It

Model ReleasesDGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

SafetyDGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

Winner of CVPR2026 NTIRE Challenge on Image Shadow Removal: Semantic and Geometric Guidance for Shadow Removal via Cascaded Refinement

ResearchDGX agent

arXiv:2604.16177v1 Announce Type: new Abstract: We present a three-stage progressive shadow-removal pipeline for the CVPR2026 NTIRE WSRD+ challenge. Built on OmniSR, our method treats deshadowing as i

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting

Model ReleasesDGX agent

arXiv:2510.17210v3 Announce Type: replace Abstract: The increase in computing power and the necessity of AI-assisted decision-making boost the growing application of large language models (LLMs). Alon

WiseMind: a knowledge-guided multi-agent framework for accurate and empathetic psychiatric diagnosis

AgentsDGX agent

arXiv:2502.20689v4 Announce Type: replace Abstract: Large Language Models (LLMs) offer promising opportunities to support mental healthcare workflows, yet they often lack the structured clinical reaso

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

SafetyDGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines

ResearchDGX agent

arXiv:2604.15376v1 Announce Type: cross Abstract: Multi-step zoom-in pipelines are widely used for GUI grounding, yet the intermediate predictions they produce are typically discarded after coordinate

17 Apr 2026

3AM: 3egment Anything with Geometric Consistency in Videos

ResearchDGX agent

arXiv:2601.08831v5 Announce Type: replace Abstract: Video object segmentation methods like SAM2 achieve strong performance through memory-based architectures but struggle under large viewpoint changes

3D Conditional Image Synthesis of Left Atrial LGE MRI from Composite Semantic Masks

ResearchDGX agent

arXiv:2601.04588v2 Announce Type: replace Abstract: Segmentation of the left atrial (LA) wall and endocardium from late gadolinium-enhanced (LGE) MRI is essential for quantifying atrial fibrosis in pa

← Previous
1…912913914915916…998
Next →