AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
Human
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
2 Jul 2026

Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark

Model ReleasesDGX agent

arXiv:2607.00710v1 Announce Type: cross Abstract: Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets conta

Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

ResearchDGX agent

arXiv:2607.00553v1 Announce Type: cross Abstract: Lightweight machine learning models are increasingly proposed for intrusion detection in Industrial Internet of Things (IIoT) networks due to their su

Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.00514v1 Announce Type: cross Abstract: Automatic understanding of dynamic 4D point clouds, the 3D-point sequences captured over time by depth sensors and LiDAR, is central to robotics and e

Crystalite: A Lightweight Transformer for Efficient Crystal Modeling

ResearchDGX agent

arXiv:2604.02270v2 Announce Type: replace-cross Abstract: Generative models for crystalline materials often rely on equivariant graph neural networks, which capture geometric structure well but are co

DART: Difficulty-Adaptive Routing for Zero-Shot Video Temporal Grounding

ResearchDGX agent

arXiv:2607.00672v1 Announce Type: new Abstract: Zero-shot video temporal grounding (VTG) localizes events in untrimmed videos from natural language queries without task-specific training. Existing met

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation

ResearchDGX agent

arXiv:2607.01043v1 Announce Type: cross Abstract: Memory-based discrete vision-language navigation (VLN) agents must act under partial observability, yet even strong frozen backbones remain vulnerable

Dataset Biases and Shortcut Learning in Motion-Based AI-Generated Video Detection

SafetyDGX agent

arXiv:2607.00948v1 Announce Type: new Abstract: The visual quality of AI-generated videos has improved drastically in recent years, making it increasingly difficult for humans to distinguish between r

Decentralized Geometric Control for Cable-Suspended Payload Transport with Adaptive Mass Estimation

SafetyDGX agent

arXiv:2607.00024v1 Announce Type: new Abstract: Cooperative aerial transport requires controllers that respect nonlinear manifold geometry, operate without centralized coordination, and respect operat

Decision-Aware Training for Sample-Based Generative Models

ApplicationsDGX agent

arXiv:2607.01171v1 Announce Type: new Abstract: Sample-based generative models are increasingly used for probabilistic forecasting in high-stakes decision settings, yet their training objectives are b

Decision-focused Sparse Tangent Portfolio Optimization

TutorialsDGX agent

arXiv:2607.00581v1 Announce Type: new Abstract: Sparse tangent portfolio optimization aims to learn an interpretable, low-cardinality portfolio in the tangency direction of the mean-variance frontier.

Decompose, Compare, and Decide: Multimodal LLMs are Implicit Few-Shot Learners

ResearchDGX agent

arXiv:2607.00125v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable abilities when analyzing images, yet translating these capabilities to few-shot im

Deconfounded Lifelong Learning for Autonomous Driving via Dynamic Knowledge Spaces

AgentsDGX agent

arXiv:2603.14354v3 Announce Type: replace-cross Abstract: End-to-End autonomous driving (E2E-AD) systems face challenges in lifelong learning, including catastrophic forgetting, difficulty in knowledg

Decoupled Guidance: Disentangling Subject and Context Pathways in Text-to-Image Personalization

ResearchDGX agent

arXiv:2607.00766v1 Announce Type: new Abstract: Text-to-image personalization aims to generate a user-provided subject in novel scenes described by text. However, most existing methods encode subject

Deep Learning-Driven Black-Box Doherty Power Amplifier with Pixelated Output Combiner and Extended Efficiency Range

ResearchDGX agent

arXiv:2603.16565v2 Announce Type: replace-cross Abstract: This article presents a deep learning-driven inverse design methodology for Doherty power amplifiers (PA) with multi-port pixelated output com

Deep learning with missing data

TutorialsDGX agent

arXiv:2504.15388v3 Announce Type: replace-cross Abstract: In the context of multivariate nonparametric regression with missing covariates, we propose Pattern Embedded Neural Networks (PENNs), which ca

Deep Multitask Learning for Mixed-Type Outcomes with Shared Sparsity

ResearchDGX agent

arXiv:2607.00995v1 Announce Type: cross Abstract: Most existing multitask learning approaches are limited by their reliance on task-specific loss functions tailored to the scale and type of each outco

Destination-Labeled Self-Looping Systems with Dwell: Intrinsic Characterization, Realization Cost, and Recognition

ResearchDGX agent

arXiv:2607.00044v1 Announce Type: cross Abstract: We study a finite-state symbolic controller for systems in which the admissible visible transitions are fixed in advance and each visible state carrie

Detecting the Undetectable: Enhancing Unsupervised time series Anomaly Detection via Active Learning

TutorialsDGX agent

arXiv:2607.00720v1 Announce Type: cross Abstract: Despite the increasing sophistication of industrial AI systems, the ability to reliably detect subtle and noisy anomalies in complex time series data

Device Passport: Enabling Spatio-Temporal Pretrained Models to Generalize Across Input Layouts

ResearchDGX agent

arXiv:2607.00249v1 Announce Type: new Abstract: New device layouts pose a challenging modeling problem due to the lack of large datasets for each specific layout. Biosignal foundation models offer a p

DeWorldSG: Depth-Aware 3D Semantic Scene Graph Generation via World-Model Priors

ResearchDGX agent

arXiv:2607.00889v1 Announce Type: cross Abstract: We present DeWorldSG, a novel framework that generates spatio-temporally robust 3D Semantic Scene Graphs from RGB-D sequences. Existing methods often

Diffeomorphic Optimization

TutorialsDGX agent

arXiv:2607.00947v1 Announce Type: new Abstract: Generative models learn data distributions that reside on a low-dimensional manifold within a higher-dimensional ambient space. Optimizing differentiabl

Diffusion-Based Multi-Class Normality for OOD Detection: An Application to CDP Authentication

ResearchDGX agent

arXiv:2607.00609v1 Announce Type: new Abstract: Reconstruction-based generative models offer a natural framework for unsupervised out-of-distribution (OOD) detection, but multi-class normality modelli

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

SafetyDGX agent

arXiv:2607.01170v1 Announce Type: cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they ar

DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoning

Model ReleasesDGX agent

arXiv:2607.00341v1 Announce Type: cross Abstract: Large language models achieve strong performance on many reasoning tasks when allowed to externalize intermediate steps as Chain-of-Thought (CoT). How

Disentangling Speaker and Language Effects in Cross-Lingual Speaker Verification for Iberian Languages

ResearchDGX agent

arXiv:2607.01161v1 Announce Type: cross Abstract: Cross-lingual speaker verification (SV) systems typically exhibit performance degradation when enrollment and test utterances are spoken in different

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation

SafetyDGX agent

arXiv:2607.01208v1 Announce Type: cross Abstract: Language models deployed in high-stakes roles can potentially favor certain entities, brands, or viewpoints, steering user decisions at scale. Such pr

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

SafetyDGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

Distributed Online Bandit Submodular Maximization with Bounded Sampling Violations

ResearchDGX agent

arXiv:2607.00680v1 Announce Type: new Abstract: We study distributed online submodular maximization under partition matroid constraints, in which multiple agents select a limited number of actions fro

Distributionally Robust Linear Regression With Block Lewis Weights

Model ReleasesDGX agent

arXiv:2607.00252v1 Announce Type: new Abstract: We present an algorithm for the group distributionally robust (GDR) least squares problem. Given m groups, a parameter vector in R^d, and stacked design

Do Time Series Foundation Model Benchmarks Hide Regime-Dependent Failures? Evidence from Traffic Speed Forecasting

ResearchDGX agent

arXiv:2606.18367v2 Announce Type: replace Abstract: Standard benchmarks evaluate time series foundation models (TSFMs) using aggregate metrics, but these can mask severe failures in critical operating

Does Your ViT Still Need U-Net for Segmentation?

Model ReleasesDGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts

SafetyDGX agent

arXiv:2607.00666v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models often fail to perform the same learned tasks under environmental shifts, such as changes in camera pose and shifts

'Don't Say It!': Constraints, Compliance, and Communication when Language Models Play Taboo

ResearchDGX agent

arXiv:2607.00601v1 Announce Type: new Abstract: The game of Taboo requires describing a target word without using a set of forbidden words, so that other players can guess it. This deceptively simple

DriftScope: Measuring The Hidden Effects of Diffusion Model Adaptation

TutorialsDGX agent

arXiv:2607.00183v1 Announce Type: new Abstract: Adapting pre-trained text-to-image diffusion models, whether to learn new visual concepts or erase unwanted ones, is routinely evaluated on its intended

DriveVA: Video Action Models are Zero-Shot Drivers

Model ReleasesDGX agent

arXiv:2604.04198v2 Announce Type: replace Abstract: Generalization is a central challenge in autonomous driving, as real-world deployment requires robust performance under unseen scenarios, sensor dom

DriveVer: Lightweight Trajectory Evaluator as Test-Time Verifier for Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.00399v1 Announce Type: new Abstract: End-to-end autonomous driving models often encounter performance bottlenecks, as training-time scaling leads to high computational costs and diminishing

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images

Model ReleasesDGX agent

arXiv:2607.00338v1 Announce Type: new Abstract: Object detection for Unmanned Aerial Vehicles (UAVs) working in open and dynamic environments is a highly challenging task. While Vision-Language Models

DroneIQA-VLE: Multi-Task Drone Image Quality Assessment via Vision-Language Ensemble

ResearchDGX agent

arXiv:2607.00416v1 Announce Type: new Abstract: We present DroneIQA-VLE, our solution to the ICME 2026 Drone-IQA Grand Challenge on Target-aware Image Quality Assessment for Low-altitude UAV Images. T

Dual-Confidence Contrastive Decoding for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2607.00570v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) increasingly requires models to answer questions from multiple retrieved documents, where only some sources are rel

Dual-Informed Vertical Expansion for Multi-Objective Node Selection in Anytime Conflict-Based Search

SafetyDGX agent

arXiv:2607.00156v1 Announce Type: new Abstract: Conflict-Based Search (CBS) is a leading exact algorithm for Multi-Agent Path Finding (MAPF), but its high-level node-selection rule is usually treated

Dynamic Bidirectional Pattern Memory: A Production-Scale Empirical Characterisation of Inference-Time Gating in Clinical NLP

Model ReleasesDGX agent

arXiv:2607.00870v1 Announce Type: new Abstract: We study inference-time pattern-memory gating in a production-scale clinical natural language processing (NLP) pipeline. The pipeline pairs a generator

E-TIDE: Fast, Structure-Preserving Motion Forecasting from Event Sequences

ResearchDGX agent

arXiv:2603.27757v2 Announce Type: replace Abstract: Event-based cameras capture visual information as asynchronous streams of per-pixel brightness changes, generating sparse, temporally precise data.

EchoRisk: A Multicentre Echocardiography Dataset and Benchmark for Cardio-Oncology

Model ReleasesDGX agent

arXiv:2607.01039v1 Announce Type: cross Abstract: Therapy-induced cardiotoxicity is the leading non-oncological cause of treatment interruption in breast cancer patients, yet early, automated risk str

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

SafetyDGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

HardwareDGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

Efficient Multilingual Reasoning Transfer via Progressive Code-Switching

ResearchDGX agent

arXiv:2607.00485v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong reasoning capabilities in English, yet their performance degrades significantly when required to reas

Efficient Neural Controlled Differential Equations via Attentive Kernel Smoothing

ResearchDGX agent

arXiv:2602.02157v2 Announce Type: replace Abstract: Neural Controlled Differential Equations (Neural CDEs) provide a powerful continuous-time framework for sequence modeling, yet the roughness of the

EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection

Local AiDGX agent

arXiv:2607.00867v1 Announce Type: new Abstract: Long-video reasoning is fundamentally constrained by how models acquire and utilize visual evidence. Existing tool-augmented video frameworks often inte

EgoGapBench: Benchmarking Egocentric Action Selection in Multi-Agent Scenes

Model ReleasesDGX agent

arXiv:2607.00547v1 Announce Type: cross Abstract: Existing egocentric benchmarks have primarily constructed the egocentric setting from first-person-view data, which makes it difficult to evaluate ego

EgoSafetyBench: A Diagnostic Egocentric Video Benchmark for Evaluating Embodied VLMs as Runtime Safety Guards

Model ReleasesDGX agent

arXiv:2607.00218v1 Announce Type: cross Abstract: Vision-language models (VLMs) are now proposed as runtime safety guards for embodied agents in homes and factories. A deployable guard must catch genu

EgoSim: Egocentric World Simulator for Embodied Interaction Generation

ApplicationsDGX agent

arXiv:2604.01001v2 Announce Type: replace-cross Abstract: We introduce EgoSim, a closed-loop egocentric world simulator that generates spatially consistent interaction videos and persistently updates

ELMP: Efficient Learning for Motion Planning via Analytical Policy Gradients

SafetyDGX agent

arXiv:2607.00215v1 Announce Type: new Abstract: Neural Motion Planners (NMPs) enable fast reactive motion generation, but adapting them to new environments typically requires recollecting large expert

EmbodimentSemantic: A Spatial Scene-Graph Dataset and Benchmark for Vision-Language Models on Embodied Manipulation Trajectories

Model ReleasesDGX agent

arXiv:2607.00020v1 Announce Type: new Abstract: Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling

Model ReleasesDGX agent

arXiv:2512.15702v2 Announce Type: replace Abstract: Autoregressive video diffusion models hold promise for world simulation but are vulnerable to exposure bias arising from the train-test mismatch. Wh

Energy-Efficient Real-Time 4-Stage Sleep Classification at 10-Second Resolution

ResearchDGX agent

arXiv:2508.11664v2 Announce Type: replace-cross Abstract: Sleep stage classification is critical for diagnosing and managing disorders like sleep apnea and insomnia. However, conventional methods like

Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking

Model ReleasesDGX agent

arXiv:2412.20750v3 Announce Type: replace Abstract: Large-scale Vision-Language Models (VLMs) have achieved notable progress in aligning visual inputs with text. However, their ability to deeply under

Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis

ResearchDGX agent

arXiv:2607.00363v1 Announce Type: cross Abstract: Flow Matching (FM) has emerged as a powerful paradigm for speech generation but remains constrained by high inference latency and timbre leakage. To a

Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms

SafetyDGX agent

arXiv:2407.15283v2 Announce Type: replace-cross Abstract: Industry is moving toward autonomous, network-connected machines that detect and adapt to changing conditions, including hardware faults. Conv

Enhancing Oracle Bone Inscription Recognition via Multi-Scale Layer Attention

ResearchDGX agent

arXiv:2607.00057v1 Announce Type: cross Abstract: Oracle Bone Inscriptions (OBIs) recognition plays a crucial role in understanding ancient Chinese culture. However, accurately recognizing OBIs remain

Enhancing Robustness in Robot-Environment Interactions through Passive Compliant Degrees of Freedom: A Hybrid Position-Force Control Approach with Feedback Linearization

ResearchDGX agent

arXiv:2607.00571v1 Announce Type: new Abstract: Robot-environment interactions in dynamic or unstructured settings are often degraded by impact shocks, vibrations, and uncertainties in contact geometr

← Previous
1…287288289290291…1025
Next →