AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
Human
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
5 May 2026

Trees and Graphs with Non Log-concave Dominating Set Sequence via AI Tools

ResearchDGX agent

arXiv:2605.02193v1 Announce Type: cross Abstract: We give new examples of graphs and trees with dominating set sequences that are not log-concave. These examples were generated by PatternBoost, a tran

TRIM: A Self-Supervised Video Summarization Framework Maximizing Temporal Relative Information and Representativeness

ResearchDGX agent

arXiv:2506.20588v2 Announce Type: replace Abstract: The increasing ubiquity of video content and the corresponding demand for efficient access to meaningful information have elevated video summarizati

TRIMMER: A New Paradigm for Video Summarization through Self-Supervised Reinforcement Learning

ApplicationsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.01659v1 Announce Type: new Abstract: The rapid growth of video content across domains such as surveillance, education, and social media has made efficient content understanding increasingly

TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation

Model ReleasesDGX agent

arXiv:2605.00907v1 Announce Type: new Abstract: Large language models (LLMs) and multimodal large models (MLLMs) are increasingly used for transportation tasks such as regulation question answering, t

Triple Spectral Fusion for Sensor-based Human Activity Recognition

Model ReleasesDGX agent

arXiv:2605.02743v1 Announce Type: cross Abstract: The field of sensor-based human activity recognition (HAR) mainly uses posture, motion and context data of Inertial Measurement Units (IMUs) to identi

Trust, but Verify: Peeling Low-Bit Transformer Networks for Training Monitoring

Local AiDGX agent

arXiv:2605.02853v1 Announce Type: new Abstract: Understanding whether deep neural networks are effectively optimized remains challenging, as training occurs in highly nonconvex landscapes and standard

TT4D: A Pipeline and Dataset for Table Tennis 4D Reconstruction From Monocular Videos

ResearchDGX agent

arXiv:2605.01234v1 Announce Type: new Abstract: We present TT4D, a large-scale, high-fidelity table tennis dataset. It provides 140+ hours of reconstructed singles and doubles gameplay from monocular

TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization

SafetyDGX agent

arXiv:2605.00224v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human preferences is commonly done via reinforcement learning from human feedback (RLHF) with Proximal Policy

Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Environments

Model ReleasesDGX agent

arXiv:2510.04142v2 Announce Type: replace Abstract: This paper identifies a critical yet underexplored challenge in reasoning alignment from multiple multi-modal large language models (MLLMs): In non-

TwistNet-2D: Learning Second-Order Channel Interactions via Spiral Twisting for Texture Recognition

Model ReleasesDGX agent

arXiv:2602.07262v3 Announce Type: replace Abstract: Second-order feature statistics are central to texture recognition, yet existing mechanisms exhibit a structural tension: bilinear pooling and Gram

Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video

Model ReleasesDGX agent

arXiv:2605.01512v1 Announce Type: new Abstract: Grounding traffic accidents in real CCTV footage is a rare-event problem where training on labeled accident video is often prohibited, yet accurate join

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning

TutorialsDGX agent

arXiv:2605.02765v1 Announce Type: cross Abstract: LLMs are increasingly used for end-user task planning, yet their black-box nature limits users' ability to ensure reliability and control. While recen

UCATSC: Uncertainty-Aware Constrained Traffic Signal Control Under Vision-Based Partial Observability

SafetyDGX agent

arXiv:2602.07784v3 Announce Type: replace Abstract: Camera-based adaptive traffic signal control is inherently partially observable: detections can be missed, vehicle speeds and distances can be noisy

Ultrafast On-chip Online Learning via Spline Locality in Kolmogorov-Arnold Networks

Local AiDGX agent

arXiv:2602.02056v2 Announce Type: replace-cross Abstract: Ultrafast online learning is essential for high-frequency systems, such as controls for quantum computing and nuclear fusion, where adaptation

Ultrasound Vision-Language Alignment via Contrastive Learning

SafetyDGX agent

arXiv:2605.02126v1 Announce Type: new Abstract: Ultrasound foundation models have achieved strong performance on structured prediction tasks but remain exclusively vision-based, limiting zero-shot and

Unbox Responsible GeoAI: Navigating Climate Extreme and Disaster Mapping

ResearchDGX agent

arXiv:2605.00315v1 Announce Type: cross Abstract: As climate extreme and disaster events become more frequent and intense, Geospatial Artificial Intelligence (GeoAI) has emerged as a transformative ap

Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis

SafetyDGX agent

arXiv:2208.01899v2 Announce Type: replace Abstract: Imitation learning learns a policy from expert trajectories. While the expert data is believed to be crucial for imitation quality, it was found tha

Understanding Emergent Misalignment via Feature Superposition Geometry

Model ReleasesDGX agent

arXiv:2605.00842v1 Announce Type: cross Abstract: Emergent misalignment, where fine-tuning on narrow, non-harmful tasks induces harmful behaviors, poses a key challenge for AI safety in LLMs. Despite

Understanding the Performance Plateau in Text-to-Video Retrieval: A Comprehensive Empirical and Linguistic Analysis

Model ReleasesDGX agent

arXiv:2605.00826v1 Announce Type: cross Abstract: Text-to-video retrieval enables users to find relevant video content using natural language queries, a task that has grown increasingly important with

UnGAP: Uncertainty-Guided Affine Prompting for Real-Time Crack Segmentation

Local AiDGX agent

arXiv:2605.02380v1 Announce Type: new Abstract: Real-time crack segmentation is vital for structural health monitoring but is plagued by aleatoric uncertainties arising from varying lighting, blur, an

Unified Map Prior Encoder for Mapping and Planning

SafetyDGX agent

arXiv:2605.02762v1 Announce Type: new Abstract: Online mapping and end-to-end (E2E) planning in autonomous driving remain largely sensor-centric, leaving rich map priors, including HD/SD vector maps,

Unifying Deep Stochastic Processes for Image Enhancement

ResearchDGX agent

arXiv:2605.01568v1 Announce Type: new Abstract: Deep stochastic processes have recently become a central paradigm for image enhancement, with many methods explicitly conditioning the stochastic trajec

Universality in Deep Neural Networks: An approach via the Lindeberg exchange principle

ResearchDGX agent

arXiv:2605.02771v1 Announce Type: cross Abstract: We consider the infinite-width limit of a fully connected deep neural network with general weights, and we prove quantitative general bounds on the 2-

Unsupervised full-field Bayesian inference of orthotropic hyperelasticity from a single biaxial test: a myocardial case study

Model ReleasesDGX agent

arXiv:2510.09498v3 Announce Type: replace-cross Abstract: Cardiac muscle tissue exhibits highly non-linear hyperelastic and orthotropic material behavior during passive deformation. Traditional consti

Unsupervised Learning of Robust Spectral Shape Matching

ResearchDGX agent

arXiv:2304.14419v2 Announce Type: replace Abstract: We propose a novel learning-based approach for robust 3D shape matching. Our method builds upon deep functional maps and can be trained in a fully u

Unsupervised Machine Learning for Detecting Structural Anomalies in European Regional Statistics

SafetyDGX agent

arXiv:2605.02884v1 Announce Type: new Abstract: Ensuring the coherence of regional socio-economic statistics is a central task for national statistical institutes. Traditional validation tools, such a

Validation of an AI-based end-to-end model for prostate pathology using long-term archived routine samples

ResearchDGX agent

arXiv:2605.02614v1 Announce Type: new Abstract: Artificial intelligence (AI) is becoming a clinical tool for prostate pathology, but generalization across variations in sample preparation and preserva

Validation of Whole-Slide Foundation Models for Image Retrieval in TCGA Data

ResearchDGX agent

arXiv:2605.00902v1 Announce Type: new Abstract: Foundation models are reshaping computational histopathology, yet their value for whole-slide image retrieval relative to strong patch-based and supervi

Value Functions for Temporal Logic: Optimal Policies and Safety Filters

SafetyDGX agent

arXiv:2605.01051v1 Announce Type: cross Abstract: While Bellman equations for basic reach, avoid, and reach-avoid problems are well studied, the relationship between value optimality and policy optima

Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision

ResearchDGX agent

arXiv:2604.04934v2 Announce Type: replace Abstract: We present Vanast, a unified framework that generates garment-transferred human animation videos directly from a single human image, garment images,

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation

Model ReleasesDGX agent

arXiv:2605.01517v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) animation generation is pivotal for professional design due to their structural editability and resolution independence.

Variational Matrix-Learning Fourier Networks for Parametric Multiphysics Surrogates

Model ReleasesDGX agent

arXiv:2605.02280v1 Announce Type: new Abstract: Multiphysics simulation is critical for system-technology co-optimization (STCO) in chiplet-based design, but repeated finite-element solutions of PDE-g

VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation

ApplicationsDGX agent

arXiv:2602.21054v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) frequently hallucinate, limiting their safe deployment in real-world applications. Existing LLM self-eval

Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning

AgentsDGX agent

arXiv:2605.01399v1 Announce Type: new Abstract: The conventional Retrieval-Augmented Generation (RAG) paradigm of injecting raw retrieved texts into the Large Language Model (LLM)'s context often resu

VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning

Local AiDGX agent

arXiv:2601.20055v2 Announce Type: replace Abstract: Despite the syntactic fluency of Large Language Models (LLMs), ensuring their logical correctness in high-stakes domains remains a fundamental chall

VeRO: An Evaluation Harness for Agents to Optimize Agents

Model ReleasesDGX agent

arXiv:2602.22480v2 Announce Type: replace-cross Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a target agent through edit-execute-eva

Vibe Coding in Product Teams: Reconfiguring AI-Assisted Workflows, Prototyping, and Collaboration

ResearchDGX agent

arXiv:2509.10652v3 Announce Type: replace-cross Abstract: Generative AI is reshaping product design practices through 'vibe coding,' where product team members express intent in natural language and A

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.01662v1 Announce Type: new Abstract: Large vision-language models (VLMs) have advanced multimodal tasks such as video question answering (QA). However, VLMs face the challenge of selecting

Video Generation with Predictive Latents

TutorialsDGX agent

arXiv:2605.02134v1 Announce Type: new Abstract: Video Variational Autoencoder (VAE) enables latent video generative modeling by mapping the visual world into compact spatiotemporal latent spaces, impr

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation

SafetyDGX agent

arXiv:2601.23286v2 Announce Type: replace Abstract: While recent video diffusion models (VDMs) produce visually impressive results, they fundamentally struggle to maintain 3D structural consistency, o

VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition

Model ReleasesDGX agent

arXiv:2605.02834v1 Announce Type: new Abstract: Videos are unique in their ability to capture actions which transcend multiple frames. Accordingly, for many years action recognition was the quintessen

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking

ResearchDGX agent

arXiv:2605.02638v1 Announce Type: new Abstract: Cross-view Referring Multi-Object Tracking (CRMOT) aims to track multiple objects specified by natural language across multiple camera views, with globa

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA

HardwareDGX agent

arXiv:2605.01935v1 Announce Type: cross Abstract: Vision Mamba (ViM) models offer a compelling efficiency advantage over Transformers by leveraging the linear complexity of State Space Models (SSMs),

Virtual Scanning for NSCLC Histology: Investigating the Discriminatory Power of Synthetic PET

ResearchDGX agent

arXiv:2605.02746v1 Announce Type: new Abstract: Accurate histological differentiation between adenocarcinoma (ADC) and squamous cell carcinoma (SCC) is critical for personalized treatment in non-small

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy

SafetyDGX agent

arXiv:2605.01101v1 Announce Type: cross Abstract: This paper develops Virtual Speech Therapist (VST), an intelligent agent-based platform that streamlines stuttering assessment and delivers customized

Visibility-Aware Mobile Grasping in Dynamic Environments

SafetyDGX agent

arXiv:2605.02487v1 Announce Type: new Abstract: This paper addresses the problem of mobile grasping in dynamic, unknown environments where a robot must operate under a limited field-of-view. The funda

VISTA: Video Interaction Spatio-Temporal Analysis Benchmark

Model ReleasesDGX agent

arXiv:2605.01391v1 Announce Type: new Abstract: Existing benchmarks for Vision-Language Models (VLMs) primarily evaluate spatio-temporal understanding on simple single-action videos, closed attribute

Visual Chart Representations for Cryptocurrency Regime Prediction: A Systematic Deep Learning Study

ResearchDGX agent

arXiv:2605.00875v1 Announce Type: new Abstract: Technical traders have long relied on visual analysis of candlestick charts to identify market patterns and predict price movements. While deep learning

Visual Implicit Autoregressive Modeling

Model ReleasesDGX agent

arXiv:2605.01220v1 Announce Type: new Abstract: Visual Autoregressive Modeling (VAR) based on next-scale prediction achieves strong generation quality, but their explicit deep stacks fix the amount of

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

Model ReleasesDGX agent

arXiv:2605.02735v1 Announce Type: new Abstract: Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evide

Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms

Model ReleasesDGX agent

arXiv:2603.14535v2 Announce Type: replace Abstract: Reinforcement learning has proven its power on various occasions. However, its performance is not always guaranteed when system dynamics change. Ins

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model

Model ReleasesDGX agent

arXiv:2605.01194v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities and generalization in embodied manipulation. However, their decision-makin

VOFA: Visual Object Goal Pushing with Force-Adaptive Control for Humanoids

SafetyDGX agent

arXiv:2605.01518v1 Announce Type: new Abstract: The ability to push large objects in a goal-directed manner using onboard egocentric perception is an essential skill for humanoid robots to perform com

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection

ResearchDGX agent

arXiv:2605.01365v1 Announce Type: new Abstract: Open-vocabulary 3D affordance detection requires localizing interaction regions on point clouds given novel affordance descriptions. Recent methods exte

VRGaussianAvatar: Integrating 3D Gaussian Avatars into VR

ResearchDGX agent

arXiv:2602.01674v2 Announce Type: replace Abstract: We present VRGaussianAvatar, an integrated system that enables real-time full-body 3D Gaussian Splatting (3DGS) avatars in virtual reality using onl

Watch Your Step: Information Injection in Diffusion Models via Shadow Timestep Embedding

ResearchDGX agent

arXiv:2605.00935v1 Announce Type: cross Abstract: Diffusion models have become the foundation of modern generative systems, with most research focusing primarily on improving generation efficiency and

Watermarking LLM Agent Trajectories

Model ReleasesDGX agent

arXiv:2602.18700v2 Announce Type: replace-cross Abstract: LLM agents rely heavily on high-quality trajectory data to guide their problem-solving behaviors, yet producing such data requires substantial

Weight Clipping for Robust Conformal Inference under Unbounded Covariate Shifts

ApplicationsDGX agent

arXiv:2605.02072v1 Announce Type: new Abstract: Conformal prediction (CP) provides powerful, distribution-free prediction sets, but its guarantees rely on the exchangeability of training and test data

What Does a Meow Mean? In Search of Intuitively Understandable Communication by a Nonverbal Companion Robot

ResearchDGX agent

arXiv:2605.01251v1 Announce Type: cross Abstract: Older adults living alone have a number of challenges, and robots can help with some of them--by providing reminders, initiating activity, or offering

← Previous
1…787788789790791…998
Next →