AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
13 Apr 2026

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

Towards Linguistically-informed Representations for English as a Second or Foreign Language: Review, Construction and Application

ResearchDGX agent

arXiv:2604.09008v1 Announce Type: cross Abstract: The widespread use of English as a Second or Foreign Language (ESFL) has sparked a paradigm shift: ESFL is not seen merely as a deviation from standar

Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models

Safety
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.08815v1 Announce Type: new Abstract: Medical vision-language models (VLMs) show strong performance on radiology tasks but often produce fluent yet weakly grounded conclusions due to over-re

Tracing the Chain: Deep Learning for Stepping-Stone Intrusion Detection

ResearchDGX agent

arXiv:2604.08800v1 Announce Type: cross Abstract: Stepping-stone intrusions (SSIs) are a prevalent network evasion technique in which attackers route sessions through chains of compromised intermediat

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

SafetyDGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

TutorialsDGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

SafetyDGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

Transferable FB-GNN-MBE Framework for Potential Energy Surfaces: Data-Adaptive Transfer Learning in Deep Learned Many-Body Expansion Theory

ResearchDGX agent

arXiv:2604.09320v1 Announce Type: cross Abstract: Mechanistic understanding and rational design of complex chemical systems depend on fast and accurate predictions of electronic structures beyond indi

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

Model ReleasesDGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

SafetyDGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

TurPy: a physics-based and differentiable optical turbulence simulator for algorithmic development and system optimization

Model ReleasesDGX agent

arXiv:2604.07248v2 Announce Type: replace-cross Abstract: Developing optical systems for free-space applications requires simulation tools that accurately capture turbulence-induced wavefront distorti

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

HardwareDGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

UHD Low-Light Image Enhancement via Real-Time Enhancement Methods with Clifford Information Fusion

ResearchDGX agent

arXiv:2604.09321v1 Announce Type: cross Abstract: Considering efficiency, ultra-high-definition (UHD) low-light image restoration is extremely challenging. Existing methods based on Transformer archit

UIPress: Bringing Optical Token Compression to UI-to-Code Generation

ResearchDGX agent

arXiv:2604.09442v1 Announce Type: new Abstract: UI-to-Code generation requires vision-language models (VLMs) to produce thousands of tokens of structured HTML/CSS from a single screenshot, making visu

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

SafetyDGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

ResearchDGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

Uncertainty Estimation for the Open-Set Text Classification systems

Model ReleasesDGX agent

arXiv:2604.08560v1 Announce Type: cross Abstract: Accurate uncertainty estimation is essential for building robust and trustworthy recognition systems. In this paper, we consider the open-set text cla

Unified Multimodal Uncertain Inference

Model ReleasesDGX agent

arXiv:2604.08701v1 Announce Type: new Abstract: We introduce Unified Multimodal Uncertain Inference (UMUI), a multimodal inference task spanning text, audio, and video, where models must produce calib

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

SafetyDGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

Universal Approximation with XL MIMO Systems: OTA Classification via Trainable Analog Combining

ResearchDGX agent

arXiv:2504.12758v3 Announce Type: replace-cross Abstract: In this paper, we show that an eXtremely Large (XL) Multiple-Input Multiple-Output (MIMO) wireless system with appropriate analog combining co

Unmasking Puppeteers: Leveraging Biometric Leakage to Disarm Impersonation in AI-based Videoconferencing

ResearchDGX agent

arXiv:2510.03548v3 Announce Type: replace-cross Abstract: AI-based talking-head videoconferencing systems reduce bandwidth by sending a compact pose-expression latent and re-synthesizing RGB at the re

Using Synthetic Data for Machine Learning-based Childhood Vaccination Prediction in Narok, Kenya

ResearchDGX agent

arXiv:2604.08902v1 Announce Type: new Abstract: Background: Limited data utilization in low-resource settings poses a barrier to the vaccine delivery ecosystem, undermining efforts to achieve equitabl

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

AgentsDGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

SafetyDGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

VAGNet: Vision-based accident anticipation with global features

Model ReleasesDGX agent

arXiv:2604.09305v1 Announce Type: new Abstract: Traffic accidents are a leading cause of fatalities and injuries across the globe. Therefore, the ability to anticipate hazardous situations in advance

Variational Quantum Physics-Informed Neural Networks for Hydrological PDE-Constrained Learning with Inherent Uncertainty Quantification

ResearchDGX agent

arXiv:2604.09374v1 Announce Type: cross Abstract: We propose a Hybrid Quantum-Classical Physics-Informed Neural Network (HQC-PINN) that integrates parameterized variational quantum circuits into the P

Verbalizing LLMs' assumptions to explain and control sycophancy

SafetyDGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2604.08549v1 Announce Type: cross Abstract: We introduce VerifAI, an open-source expert system for biomedical question answering that integrates retrieval-augmented generation (RAG) with a novel

ViSAGE @ NTIRE 2026 Challenge on Video Saliency Prediction

Model ReleasesDGX agent

arXiv:2604.08613v1 Announce Type: new Abstract: In this report, we present our champion solution for the NTIRE 2026 Challenge on Video Saliency Prediction held in conjunction with CVPR 2026. To exploi

Vision Transformers for Preoperative CT-Based Prediction of Histopathologic Chemotherapy Response Score in High-Grade Serous Ovarian Carcinoma

ResearchDGX agent

arXiv:2604.09197v1 Announce Type: cross Abstract: Purpose. High-grade serous ovarian carcinoma (HGSOC) is characterized by pronounced biological and spatial heterogeneity and is frequently diagnosed a

VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

ResearchDGX agent

arXiv:2604.09531v1 Announce Type: cross Abstract: Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr

VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel Optimization

ApplicationsDGX agent

arXiv:2508.13792v2 Announce Type: replace Abstract: The intrinsic dynamics of an object governs its physical behavior in the real world, playing a critical role in enabling physically plausible intera

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

SafetyDGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

Visually-Guided Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

ResearchDGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning

Model ReleasesDGX agent

arXiv:2604.08639v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ m

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

ResearchDGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

ResearchDGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

Model ReleasesDGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

Weak Adversarial Neural Pushforward Method for the Wigner Transport Equation

ResearchDGX agent

arXiv:2604.08763v1 Announce Type: cross Abstract: We extend the Weak Adversarial Neural Pushforward Method to the Wigner transport equation governing the phase-space dynamics of quantum systems. The c

Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels

ResearchDGX agent

arXiv:2510.06499v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved remarkable success through imitation learning on vast text corpora, but this paradigm creates a tra

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

TutorialsDGX agent

arXiv:2604.08716v1 Announce Type: new Abstract: Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try

When & How to Write for Personalized Demand-aware Query Rewriting in Video Search

SafetyDGX agent

arXiv:2602.17667v2 Announce Type: replace-cross Abstract: In video search systems, user historical behaviors provide rich context for identifying search intent and resolving ambiguity. However, tradit

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2510.07517v5 Announce Type: replace Abstract: Multi-agent debate (MAD) aims to improve large language model (LLM) reasoning by letting multiple agents exchange answers and then aggregate their o

Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Models

Model ReleasesDGX agent

arXiv:2602.22918v2 Announce Type: replace Abstract: Vision-language models (VLMs) can read text from images, but where does this optical character recognition (OCR) information enter the language proc

Which Pieces Does Unigram Tokenization Really Need?

Model ReleasesDGX agent

arXiv:2512.12641v2 Announce Type: replace Abstract: The Unigram tokenization algorithm offers a probabilistic alternative to the greedy heuristics of Byte-Pair Encoding. Despite its theoretical elegan

Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails

Model ReleasesDGX agent

arXiv:2603.03099v5 Announce Type: replace-cross Abstract: Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essential

WildDet3D: Scaling Promptable 3D Detection in the Wild

ApplicationsDGX agent

arXiv:2604.08626v1 Announce Type: new Abstract: Understanding objects in 3D from a single image is a cornerstone of spatial intelligence. A key step toward this goal is monocular 3D object detection--

Wireless Communication Enhanced Value Decomposition for Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.08728v1 Announce Type: new Abstract: Cooperation in multi-agent reinforcement learning (MARL) benefits from inter-agent communication, yet most approaches assume idealized channels and exis

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

TutorialsDGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers

Model ReleasesDGX agent

arXiv:2604.09489v1 Announce Type: cross Abstract: Model poisoning attacks pose a significant security threat to Federated Learning (FL). Most existing model poisoning attacks rely on collusion, requir

Yes, But Not Always. Generative AI Needs Nuanced Opt-in

AgentsDGX agent

arXiv:2604.09413v1 Announce Type: cross Abstract: This paper argues that a one-size-fits-all approach to specifying consent for the use of creative works in generative AI is insufficient. Real-world o

You Can't Fight in Here! This is BBS!

TutorialsDGX agent

arXiv:2604.09501v1 Announce Type: new Abstract: Norm, the formal theoretical linguist, and Claudette, the computational language scientist, have a lovely time discussing whether modern language models

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector

SafetyDGX agent

arXiv:2603.15757v2 Announce Type: replace-cross Abstract: What happens when a pretrained generative robot policy is provided a constant initial noise as input, rather than repeatedly sampling it from

Zero-Shot Generative De-identification: Inversion-Free Flow for Privacy-Preserving Skin Image Analysis

ResearchDGX agent

arXiv:2602.00821v2 Announce Type: replace Abstract: The secure analysis of dermatological images in clinical environments is fundamentally restricted by the critical trade-off between patient privacy

10 Apr 2026

3DrawAgent: Teaching LLM to Draw in 3D with Early Contrastive Experience

Model ReleasesDGX agent

arXiv:2604.08042v1 Announce Type: new Abstract: Sketching in 3D space enables expressive reasoning about shape, structure, and spatial relationships, yet generating 3D sketches through natural languag

A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset

Model ReleasesDGX agent

arXiv:2604.06227v1 Announce Type: new Abstract: Accurate short-term forecasting of agricultural commodity prices is critical for food security planning and smallholder income stabilisation in developi

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

SafetyDGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

A comparative analysis of machine learning models in SHAP analysis

TutorialsDGX agent

arXiv:2604.07258v1 Announce Type: new Abstract: In this growing age of data and technology, large black-box models are becoming the norm due to their ability to handle vast amounts of data and learn i

← Previous
1…10071008100910101011…1025
Next →