AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
Human
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
58,761 results
Research

Towards Linguistically-informed Representations for English as a Second or Foreign Language: Review, Construction and Application

DGX agent

arXiv:2604.09008v1 Announce Type: cross Abstract: The widespread use of English as a Second or Foreign Language (ESFL) has sparked a paradigm shift: ESFL is not seen merely as a deviation from standar

researcharxiv-cs-ai
13 Apr 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models

DGX agent

arXiv:2604.08815v1 Announce Type: new Abstract: Medical vision-language models (VLMs) show strong performance on radiology tasks but often produce fluent yet weakly grounded conclusions due to over-re

safetyarxiv-cs-cv
13 Apr 2026
Research

Tracing the Chain: Deep Learning for Stepping-Stone Intrusion Detection

DGX agent

arXiv:2604.08800v1 Announce Type: cross Abstract: Stepping-stone intrusions (SSIs) are a prevalent network evasion technique in which attackers route sessions through chains of compromised intermediat

researcharxiv-cs-lg
13 Apr 2026
Safety

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

DGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

safetyarxiv-cs-lg
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Safety

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

DGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

safetyarxiv-cs-ai
13 Apr 2026
Research

Transferable FB-GNN-MBE Framework for Potential Energy Surfaces: Data-Adaptive Transfer Learning in Deep Learned Many-Body Expansion Theory

DGX agent

arXiv:2604.09320v1 Announce Type: cross Abstract: Mechanistic understanding and rational design of complex chemical systems depend on fast and accurate predictions of electronic structures beyond indi

researcharxiv-cs-lg
13 Apr 2026
Model Releases

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

DGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

DGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

TurPy: a physics-based and differentiable optical turbulence simulator for algorithmic development and system optimization

DGX agent

arXiv:2604.07248v2 Announce Type: replace-cross Abstract: Developing optical systems for free-space applications requires simulation tools that accurately capture turbulence-induced wavefront distorti

model-releasesarxiv-cs-cv
13 Apr 2026
Hardware

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

DGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

hardwarearxiv-cs-ai
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Research

UHD Low-Light Image Enhancement via Real-Time Enhancement Methods with Clifford Information Fusion

DGX agent

arXiv:2604.09321v1 Announce Type: cross Abstract: Considering efficiency, ultra-high-definition (UHD) low-light image restoration is extremely challenging. Existing methods based on Transformer archit

researcharxiv-cs-cv
13 Apr 2026
Research

UIPress: Bringing Optical Token Compression to UI-to-Code Generation

DGX agent

arXiv:2604.09442v1 Announce Type: new Abstract: UI-to-Code generation requires vision-language models (VLMs) to produce thousands of tokens of structured HTML/CSS from a single screenshot, making visu

researcharxiv-cs-cl
13 Apr 2026
Safety

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

DGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

safetyarxiv-cs-ai
13 Apr 2026
Research

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

DGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

researcharxiv-cs-lg
13 Apr 2026
Model Releases

Uncertainty Estimation for the Open-Set Text Classification systems

DGX agent

arXiv:2604.08560v1 Announce Type: cross Abstract: Accurate uncertainty estimation is essential for building robust and trustworthy recognition systems. In this paper, we consider the open-set text cla

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Unified Multimodal Uncertain Inference

DGX agent

arXiv:2604.08701v1 Announce Type: new Abstract: We introduce Unified Multimodal Uncertain Inference (UMUI), a multimodal inference task spanning text, audio, and video, where models must produce calib

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

DGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

safetyarxiv-cs-cv
13 Apr 2026
Research

Universal Approximation with XL MIMO Systems: OTA Classification via Trainable Analog Combining

DGX agent

arXiv:2504.12758v3 Announce Type: replace-cross Abstract: In this paper, we show that an eXtremely Large (XL) Multiple-Input Multiple-Output (MIMO) wireless system with appropriate analog combining co

researcharxiv-cs-lg
13 Apr 2026
Research

Unmasking Puppeteers: Leveraging Biometric Leakage to Disarm Impersonation in AI-based Videoconferencing

DGX agent

arXiv:2510.03548v3 Announce Type: replace-cross Abstract: AI-based talking-head videoconferencing systems reduce bandwidth by sending a compact pose-expression latent and re-synthesizing RGB at the re

researcharxiv-cs-ai
13 Apr 2026
Research

Using Synthetic Data for Machine Learning-based Childhood Vaccination Prediction in Narok, Kenya

DGX agent

arXiv:2604.08902v1 Announce Type: new Abstract: Background: Limited data utilization in low-resource settings poses a barrier to the vaccine delivery ecosystem, undermining efforts to achieve equitabl

researcharxiv-cs-lg
13 Apr 2026
Agents

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

DGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

agentsarxiv-cs-ro
13 Apr 2026
Safety

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

DGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

VAGNet: Vision-based accident anticipation with global features

DGX agent

arXiv:2604.09305v1 Announce Type: new Abstract: Traffic accidents are a leading cause of fatalities and injuries across the globe. Therefore, the ability to anticipate hazardous situations in advance

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Variational Quantum Physics-Informed Neural Networks for Hydrological PDE-Constrained Learning with Inherent Uncertainty Quantification

DGX agent

arXiv:2604.09374v1 Announce Type: cross Abstract: We propose a Hybrid Quantum-Classical Physics-Informed Neural Network (HQC-PINN) that integrates parameterized variational quantum circuits into the P

researcharxiv-cs-lg
13 Apr 2026
Safety

Verbalizing LLMs' assumptions to explain and control sycophancy

DGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering

DGX agent

arXiv:2604.08549v1 Announce Type: cross Abstract: We introduce VerifAI, an open-source expert system for biomedical question answering that integrates retrieval-augmented generation (RAG) with a novel

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

ViSAGE @ NTIRE 2026 Challenge on Video Saliency Prediction

DGX agent

arXiv:2604.08613v1 Announce Type: new Abstract: In this report, we present our champion solution for the NTIRE 2026 Challenge on Video Saliency Prediction held in conjunction with CVPR 2026. To exploi

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Vision Transformers for Preoperative CT-Based Prediction of Histopathologic Chemotherapy Response Score in High-Grade Serous Ovarian Carcinoma

DGX agent

arXiv:2604.09197v1 Announce Type: cross Abstract: Purpose. High-grade serous ovarian carcinoma (HGSOC) is characterized by pronounced biological and spatial heterogeneity and is frequently diagnosed a

researcharxiv-cs-ai
13 Apr 2026
Research

VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

DGX agent

arXiv:2604.09531v1 Announce Type: cross Abstract: Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr

researcharxiv-cs-ai
13 Apr 2026
Applications

VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel Optimization

DGX agent

arXiv:2508.13792v2 Announce Type: replace Abstract: The intrinsic dynamics of an object governs its physical behavior in the real world, playing a critical role in enabling physically plausible intera

applicationsarxiv-cs-cv
13 Apr 2026
Safety

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

DGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

safetyarxiv-cs-ai
13 Apr 2026
Safety

Visually-Guided Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

safetyarxiv-cs-ai
13 Apr 2026
Research

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

DGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

researcharxiv-cs-ai
13 Apr 2026
Model Releases

VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning

DGX agent

arXiv:2604.08639v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ m

model-releasesarxiv-cs-ai
13 Apr 2026
Research

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

DGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

researcharxiv-cs-ai
13 Apr 2026
Research

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

DGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

DGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Weak Adversarial Neural Pushforward Method for the Wigner Transport Equation

DGX agent

arXiv:2604.08763v1 Announce Type: cross Abstract: We extend the Weak Adversarial Neural Pushforward Method to the Wigner transport equation governing the phase-space dynamics of quantum systems. The c

researcharxiv-cs-lg
13 Apr 2026
Research

Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels

DGX agent

arXiv:2510.06499v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved remarkable success through imitation learning on vast text corpora, but this paradigm creates a tra

researcharxiv-cs-ai
13 Apr 2026
Tutorials

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

DGX agent

arXiv:2604.08716v1 Announce Type: new Abstract: Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try

tutorialsarxiv-cs-cv
13 Apr 2026
Safety

When & How to Write for Personalized Demand-aware Query Rewriting in Video Search

DGX agent

arXiv:2602.17667v2 Announce Type: replace-cross Abstract: In video search systems, user historical behaviors provide rich context for identifying search intent and resolving ambiguity. However, tradit

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning

DGX agent

arXiv:2510.07517v5 Announce Type: replace Abstract: Multi-agent debate (MAD) aims to improve large language model (LLM) reasoning by letting multiple agents exchange answers and then aggregate their o

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Models

DGX agent

arXiv:2602.22918v2 Announce Type: replace Abstract: Vision-language models (VLMs) can read text from images, but where does this optical character recognition (OCR) information enter the language proc

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Which Pieces Does Unigram Tokenization Really Need?

DGX agent

arXiv:2512.12641v2 Announce Type: replace Abstract: The Unigram tokenization algorithm offers a probabilistic alternative to the greedy heuristics of Byte-Pair Encoding. Despite its theoretical elegan

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails

DGX agent

arXiv:2603.03099v5 Announce Type: replace-cross Abstract: Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essential

model-releasesarxiv-cs-ai
13 Apr 2026
Applications

WildDet3D: Scaling Promptable 3D Detection in the Wild

DGX agent

arXiv:2604.08626v1 Announce Type: new Abstract: Understanding objects in 3D from a single image is a cornerstone of spatial intelligence. A key step toward this goal is monocular 3D object detection--

applicationsarxiv-cs-cv
13 Apr 2026
← Previous
1…12021203120412051206…1225
Next →