AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
Human
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
23 Jun 2026

Hedgementation = Hedgerow Segmentation: A Remote Sensing Benchmark

Model ReleasesDGX agent

arXiv:2606.23615v1 Announce Type: new Abstract: We propose Hedgementation: a new benchmark to evaluate machine learning models for hedgerow mapping from remote sensing data at country scale and 10m^2

HEM: a margin-based loss for visual categorisation tasks

Model ReleasesDGX agent

arXiv:2501.12191v2 Announce Type: replace-cross Abstract: Training deep neural networks (DNNs) on classification tasks can be performed with a number of different losses, but cross-entropy (CE) loss i

HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval

HardwareDGX agent

arXiv:2606.21633v1 Announce Type: new Abstract: Diffusion LLMs (dLLMs) improve GPU utilization over autoregressive decoding by generating multiple tokens per forward pass, but their KV cache still gro

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HERCULES: An Open-Source Simulation Framework for Heterogeneous Multi-Robot SLAM, Collaborative Perception, and Exploration

Model ReleasesDGX agent

arXiv:2606.22756v1 Announce Type: cross Abstract: We present HERCULES, an open-source simulator and data-collection pipeline for heterogeneous multi-robot autonomy. Built upon the Unreal Engine 5 (UE5

HERMAN: Hierarchical Representation Matching for CLIP-based Class-Incremental Learning

ResearchDGX agent

arXiv:2509.22645v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) aims to endow models with the ability to continuously adapt to evolving data streams. Recent advances in pre-traine

HERO: Hypothesis-Driven Evidence Retrieval from Omics for Multi-Task Breast Cancer Analysis

ResearchDGX agent

arXiv:2606.21174v1 Announce Type: new Abstract: Matched multi-omics can improve WSI-based biomarker and prognosis prediction, but most existing pipelines use omics as a paral lel feature stream or tex

Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

SafetyDGX agent

arXiv:2606.20962v1 Announce Type: new Abstract: High-performing human-human teams learn intelligent and efficient communication and coordination strategies to maximize their joint utility. These teams

Hierarchical Adversarial Bandits for Online Configuration Optimization

Model ReleasesDGX agent

arXiv:2505.19061v2 Announce Type: replace Abstract: Motivated by Online Configuration Optimization in large, dynamic parameter spaces, this work studies the nonstochastic multi-armed bandit (MAB) prob

Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation

ResearchDGX agent

arXiv:2602.03448v2 Announce Type: replace Abstract: Multi-subject image generation aims to synthesize images that faithfully preserve the identities of multiple reference subjects while following text

Hierarchical Pooling for Sheaf Neural Networks

ResearchDGX agent

arXiv:2606.20932v1 Announce Type: new Abstract: Sheaf Neural Networks (SNNs) generalize Graph Neural Networks (GNNs) by replacing scalar node signals with stalk-valued signals and by using restriction

Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra

SafetyDGX agent

arXiv:2606.22922v1 Announce Type: new Abstract: Applying machine learning techniques to solving long-standing mathematical conjectures can be particularly challenging due to their extreme reward spars

Hierarchical Sparse Circuit Extraction from Billion-Parameter Language Models through Scalable Attribution Graph Decomposition

Model ReleasesDGX agent

arXiv:2601.12879v2 Announce Type: replace Abstract: Extracting sparse circuits from billion-parameter transformers is constrained by O(2^n) search cost and pervasive feature reuse across co-active pat

High-Dimensional Differentially Private Quantile Regression: Distributed Estimation and Statistical Inference

ResearchDGX agent

arXiv:2508.05212v2 Announce Type: replace-cross Abstract: With the development of big data and machine learning, privacy concerns have become increasingly critical, especially when handling heterogene

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning

SafetyDGX agent

arXiv:2606.22860v1 Announce Type: new Abstract: Recent advancements in generative imitation learning have significantly propelled the field of robotic manipulation. However, the majority of existing m

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

SafetyDGX agent

arXiv:2606.20189v2 Announce Type: replace Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data

HiMatch-AD: DINOv3-driven Hierarchical Matching for Training-free Medical Anomaly Detection

Model ReleasesDGX agent

arXiv:2606.22556v1 Announce Type: new Abstract: Anomaly detection is essential for medical image analysis, where pathological regions often appear as rare deviations from normal anatomical structures.

Holo-World: Unified Camera, Object and Weather Control for Video World Model

Model ReleasesDGX agent

arXiv:2606.20083v2 Announce Type: replace Abstract: Video world models are moving toward preserving an observed world under controllable camera and object motion while allowing its environmental state

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

SafetyDGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

Homographic Navigation: Geometry-Driven Camera Guidance for Deterministic Planar Capture

Local AiDGX agent

arXiv:2606.22834v1 Announce Type: new Abstract: We present homographic navigation, a geometry-centric framework for guiding camera acquisition toward precise capture of planar regions. Rather than tre

Horizon Adaptive Offline Policy Learning via Value Stitching

SafetyDGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

How Should a Robot Configure Its Laser Scanner for Inspection?

Model ReleasesDGX agent

arXiv:2606.21093v1 Announce Type: cross Abstract: Robotic inspection relies on accurate sensing to acquire high-fidelity geometric measurements for defect detection and metrology. While prior work has

How Should a Simulation-to-Reality Transfer Budget Be Spent?

Model ReleasesDGX agent

arXiv:2606.22062v1 Announce Type: cross Abstract: Simulation-to-reality transfer, often called sim-to-real transfer, is a central challenge in robot learning. Yet, the tradeoff between measuring a sys

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding

ResearchDGX agent

arXiv:2606.20726v1 Announce Type: new Abstract: We introduce a compact empirical model that quantifies how answer accuracy degrades as a function of frame budget B and temporal distance D in long vide

How Well Do Self-Supervised Speech Models Encode Age and Gender in Children's Speech? A Layer-Wise Analysis Across Multiple Architectures

Model ReleasesDGX agent

arXiv:2606.22177v1 Announce Type: cross Abstract: Self-supervised learning (SSL) models have become a central component of modern speech processing systems, as they enable the learning of rich acousti

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning

ResearchDGX agent

arXiv:2606.21734v1 Announce Type: new Abstract: Understanding long videos requires fine-grained perception and multi-step, higher-order reasoning over complex, long-range spatio-temporal dynamics. Vis

HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks

Model ReleasesDGX agent

arXiv:2603.19822v2 Announce Type: replace Abstract: Existing UAV vision-language navigation (VLN) benchmarks have enabled language-guided flight, but they largely focus on long, step-wise route descri

Human and AI collaboration for pulmonary nodule segmentation

ResearchDGX agent

arXiv:2606.22486v1 Announce Type: new Abstract: Medical expert annotators are scarce, and blind reliance on artificial intelligence (AI) can be misleading, motivating approaches in which humans, parti

HumanHalo -- Safe and Efficient 3D Navigation Among Humans via Minimally Conservative MPC

SafetyDGX agent

arXiv:2510.17525v3 Announce Type: replace Abstract: Safe and efficient robotic navigation among humans is essential for integrating robots into everyday environments. Most existing approaches focus on

Humanoid-OmniOcc: Stereo-Based Full-View Occupancy Dataset for Embodied AI

AgentsDGX agent

arXiv:2606.22971v1 Announce Type: cross Abstract: Occupancy prediction at voxel-level granularity is essential for safe robotic navigation and interaction in complex environments. Existing occupancy d

Hybrid Compression: Integrating Pruning and Quantization for Optimized Neural Networks

Model ReleasesDGX agent

arXiv:2606.22935v1 Announce Type: new Abstract: Deep neural networks have witnessed remarkable advancements in recent years and have become integral to various applications. However, alongside these d

HyperQuant: A Rate-Distortion-Optimal Quantization Pipeline for Large Language and Diffusion Models

Model ReleasesDGX agent

arXiv:2606.23406v1 Announce Type: new Abstract: We present HyperQuant (Hadamard, optimallY Packing, Entropy Rice-coding), a unified post-training quantization pipeline for the weights and the KV cache

Hypothesis-Disciplined Multi-Agent Automated Formalization of Asymptotic Statistical Theory

AgentsDGX agent

arXiv:2606.20642v1 Announce Type: cross Abstract: Asymptotic statistical theory is a challenging domain for AI-assisted formalization: its central results mix convergence statements, asymptotic expans

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance

ResearchDGX agent

arXiv:2606.22042v1 Announce Type: new Abstract: Diffusion-based video editing has made significant progress; however, achieving precise and temporally consistent object-level control, especially in mu

IMAGIN-4D: Image-Guided Controllable Interaction Generation

Model ReleasesDGX agent

arXiv:2606.23675v1 Announce Type: new Abstract: Generating human-object interactions (HOI) is central to character animation, robotics, AR/VR, and embodied AI. Recent HOI generation methods synthesize

Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation

SafetyDGX agent

arXiv:2509.17125v2 Announce Type: replace Abstract: Relational object rearrangement (ROR) tasks (e.g., insert flower to vase) require a robot to manipulate objects with precise semantic and geometric

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

SafetyDGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

Improved Algorithms for Nash Welfare in Linear Bandits

SafetyDGX agent

arXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training

SafetyDGX agent

arXiv:2606.22158v1 Announce Type: new Abstract: Achieving human-like reasoning in Vision-Language Models (VLMs) remains a long-standing challenge. Recent approaches leverage Chain-of-Thought (CoT) rat

Improving Robotic Imitation Learning via Trajectory Standardization

SafetyDGX agent

arXiv:2606.22907v1 Announce Type: cross Abstract: Imitation learning for robotic manipulation relies on large sets of human demonstration trajectories, which are often noisy and temporally irregular d

Improving Text-to-Music Generation with Human Preference Rewards

Model ReleasesDGX agent

arXiv:2606.21670v1 Announce Type: cross Abstract: We describe our entry to the efficiency track of the Academic Text-to-Music (ATTM) Grand Challenge at ICME 2026. Beyond the challenge protocol's FAD-C

In-Context Molecular Property Prediction with LLMs: A Blinding Study on Memorization and Knowledge Conflicts

Model ReleasesDGX agent

arXiv:2603.25857v2 Announce Type: replace Abstract: The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecula

In LLM Reasoning, there is Irrationality on top of Value Misalignment

Model ReleasesDGX agent

arXiv:2606.20624v1 Announce Type: cross Abstract: Significant progress has been made in aligning LLMs with target value functions. We argue that, even when an LLM has been well aligned in (post-)train

Incremental Learning in Mirror Flows

ResearchDGX agent

arXiv:2606.23198v1 Announce Type: cross Abstract: We study mirror flows generated by a convex quadratic loss and a general convex lower semicontinuous mirror potential. We show that, when initialized

IndicGuard: A Multilingual Safety Guard Model and Dataset for Indic Languages

Model ReleasesDGX agent

arXiv:2606.22841v1 Announce Type: cross Abstract: As Large Language Models (LLMs) achieve widespread integration across diverse linguistic landscapes, ensuring their safety and alignment with regional

Inductive Generalization for Robotic Manipulation

SafetyDGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l

Influencer Cartels

SafetyDGX agent

arXiv:2405.10231v3 Announce Type: replace-cross Abstract: Social media influencers account for a growing share of marketing worldwide. We demonstrate the existence of a novel form of market failure in

Input-schema identifiability limits in physics-informed surrogates for mechanics-governed flow

ResearchDGX agent

arXiv:2606.20655v1 Announce Type: cross Abstract: Physics-informed and data-driven surrogates are increasingly used to approximate mechanics-governed flow fields, but the target quantities assigned to

Integrated cloud-based architecture for robot-robot and human-robot collaboration using ROS 2--MQTT in Mediterranean Greenhouses

ApplicationsDGX agent

arXiv:2606.22682v1 Announce Type: new Abstract: The imperative to develop more sustainable agriculture demands a transition from isolated automation toward the deployment of multi-robot systems (MRS)

Integrated Marketing Attribution: A Bayesian Framework for Privacy-Safe Granular Measurement Anchored in MMM

ResearchDGX agent

arXiv:2606.16878v3 Announce Type: replace Abstract: Retail marketing measurement increasingly requires granular campaign-level insights without relying on user-level tracking. However, the two dominan

Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

SafetyDGX agent

arXiv:2606.21970v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models, such as Moshi, enable natural, low-latency voice conversations. However, they remain limited to the audio modality

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.22913v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving by incorporating reasoning for better interpretability and planni

Intent-Handover: Grounding Language in Human-Usage Regions for Trustworthy Robot-to-Human Handovers

SafetyDGX agent

arXiv:2503.03579v2 Announce Type: replace-cross Abstract: Spoken instructions in robot-to-human handovers may specify either an object ('the cup') or an intended use ('pour water'); in both cases, suc

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars

ResearchDGX agent

arXiv:2606.22905v1 Announce Type: new Abstract: Recent diffusion-based models have enabled realistic audio-driven avatar generation in real-time streaming. However, existing approaches struggle to mai

Interest Entanglement: The Hidden Barrier to Blind Super-Resolution Optimization

ResearchDGX agent

arXiv:2606.22353v1 Announce Type: new Abstract: Fidelity and perceptual quality are two inherently competing and conflicting objectives in the image super-resolution (SR) task. Different loss function

Interleaved Speech Language Models Latently Work In Text

ResearchDGX agent

arXiv:2606.22473v1 Announce Type: cross Abstract: Speech language models (SLMs) have been extensively studied, with the common paradigm incorporating text data and pre-trained text LMs. A leading appr

Interpretable Kolmogorov-Arnold Network with Feature-Isolated Temporal Attention Mechanism for Electricity Load Forecasting

ResearchDGX agent

arXiv:2606.23425v1 Announce Type: new Abstract: Accurate electricity load forecasting is a crucial prerequisite for stable power system operations. While prevalent deep learning models present competi

Interpretable Machine Learning for Predicting Startup Funding, Patenting, and Exits

ApplicationsDGX agent

arXiv:2510.09465v2 Announce Type: replace Abstract: This study develops an interpretable machine learning framework to forecast startup outcomes, including funding, patenting, and exit. A firm-quarter

Interpretable machine learning of halo gas density profiles: a sensitivity analysis of cosmological hydrodynamical simulations

ResearchDGX agent

arXiv:2512.09021v3 Announce Type: replace-cross Abstract: Stellar and AGN-driven feedback processes affect the distribution of gas on a wide range of scales, from within galaxies well into the interga

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

SafetyDGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

Interpretable Uncertainty Routing Separating Emotion Ambiguity from Distribution Shift in Facial Expression Recognition

ResearchDGX agent

arXiv:2606.22725v1 Announce Type: new Abstract: Facial expression recognition (FER) is inherently ambiguous: human annotators frequently disagree, and models deployed in real environments face distrib

← Previous
1…392393394395396…1040
Next →