AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
Human
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
7 Jul 2026

HiFA4: Training-Free 4-bit FlashAttention on Ascend HIF4 NPUs for LLM Inference

Model ReleasesDGX agent

arXiv:2607.04302v1 Announce Type: cross Abstract: We present HiFA4, a post-training operator-level design that executes both QK^T and PV in FlashAttention as 4-bit HIF4 Cube GEMMs for LLM inference on

High-Fidelity One-Step Generative Visuomotor Policy via Recursive Correction, Frequency Consistency, and Contrastive Flow Matching

SafetyDGX agent

arXiv:2607.03865v1 Announce Type: cross Abstract: Generative models such as diffusion and flow matching have advanced robotic visuomotor policies by modeling multimodal action distributions, but their

High-Precision Formation Control for Heterogeneous Multi-Robot Systems via Hierarchical Hybrid Physics-Informed Deep Reinforcement Learning

Safety
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.03512v1 Announce Type: new Abstract: Existing classical control methods commonly require precise models and struggle to cope with model uncertainties and external disturbances, while end-to

HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control

ResearchDGX agent

arXiv:2607.03449v1 Announce Type: cross Abstract: Current Vision-Language-Action (VLA) models excel at robotic manipulation but often struggle with non-Markovian tasks requiring long-term memory and r

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

ApplicationsDGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

Holo-Captioning: Toward the Text Equivalent of 3D Scenes

Model ReleasesDGX agent

arXiv:2607.02908v1 Announce Type: new Abstract: This work introduces holo-captioning, a novel task that strives to seek the text equivalent of 3D scenes. As the initial step, we formulate holo-caption

Homer: Understanding Long-form Videos with Hierarchical Memory and Agentic Reasoning

AgentsDGX agent

arXiv:2607.02588v1 Announce Type: cross Abstract: Multimodal large language models excel on short clips but struggle on hour-long videos in an online setting, where frames are processed incrementally

Hope for the Best, Prepare for the Worst: Occlusion-Aware Contingency Planning for Autonomous Vehicles

SafetyDGX agent

arXiv:2607.03155v1 Announce Type: new Abstract: The deployment of autonomous vehicles in urban environments introduces significant safety challenges, particularly in scenarios with occlusions, where c

How Do Diffusion Classifiers Decide? A Bias-Centric Evaluation

Model ReleasesDGX agent

arXiv:2607.03831v1 Announce Type: cross Abstract: Diffusion models have recently been repurposed for zero-shot classification, giving rise to diffusion classifiers that identify the best-matching text

How Far is Too Far? Defining the Distance Threshold for Verification Siamese Networks

TutorialsDGX agent

arXiv:2607.05329v1 Announce Type: new Abstract: Siamese verification networks are widely used to compare items such as faces, cars, or signatures. In these scenarios, the network is trained to learn a

How Many Initial Points Does Bayesian Optimization Need?

ResearchDGX agent

arXiv:2607.04356v1 Announce Type: new Abstract: Bayesian Optimization (BO) generally begins with an initialization phase: a batch of n_0 uninformed evaluations. The choice of n_0 remains largely heuri

How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation

Model ReleasesDGX agent

arXiv:2605.06605v2 Announce Type: replace Abstract: Evaluating and predicting the performance of large language models (LLMs) in multi-turn conversational settings is critical yet computationally expe

How many labels do you need? A decision framework for cross-habitat marine species recognition

Model ReleasesDGX agent

arXiv:2607.02559v1 Announce Type: new Abstract: Automated image recognition is increasingly used to scale ecological monitoring beyond manual annotation, yet ecologists lack evidence-based guidance on

How Much is Left? LLMs Linearly Encode Their Remaining Output Length

TutorialsDGX agent

arXiv:2607.05316v1 Announce Type: new Abstract: Large language models generate one token at a time, yet their responses show remarkably consistent length structure: step-by-step solutions converge in

How Much of the Routing Gap Is Real? Decomposing the Router-to-Oracle Gap into Reproducible Specialist Advantage and Single-Draw Label Noise

Model ReleasesDGX agent

arXiv:2607.03436v1 Announce Type: new Abstract: Routing among large language models (LLMs) promises better quality at lower cost, motivated by the reported gap between learned routers and a per-instan

How to Avoid Debate: Scalable AI Safety via Doubly-Efficient Interactive Proofs

SafetyDGX agent

arXiv:2607.03561v1 Announce Type: new Abstract: As AI models continue to develop powerful capabilities, it becomes critical that we are able to verify that their output is aligned with our intentions.

How to Build Digital Humans? From Priors to Photorealistic Avatars

TutorialsDGX agent

arXiv:2607.04341v1 Announce Type: cross Abstract: This state-of-the-art report provides an overview of controllable 3D human avatar creation. We describe current 3D avatar systems, which typically con

How Utilitarian Are OpenAI's Models Really? Replicating and Reinterpreting Pfeffer, Krugel, and Uhl (2025)

SafetyDGX agent

arXiv:2603.22730v2 Announce Type: replace Abstract: Pfeffer, Krugel, and Uhl (2025) report that OpenAI's reasoning model o1-mini produces more utilitarian responses to the trolley problem and footbrid

HUGS: Guiding Unified Dexterous Grasp Synthesis Across Modes and Scales via Learned Human Priors

ApplicationsDGX agent

arXiv:2607.04554v1 Announce Type: new Abstract: Dexterous grasping across diverse object scales requires contact modes ranging from two-finger pinches to bimanual grasps. Existing dexterous grasp synt

Human-Centric Reflective Architecture for Human-AI Collaborative Decision-Making

SafetyDGX agent

arXiv:2607.03025v1 Announce Type: new Abstract: The use of Large Language Models (LLMs) across diverse areas of human activity-ranging from everyday tasks to safety-critical applications-aims to enhan

Human-like Object Grouping in Self-supervised Vision Transformers

Model ReleasesDGX agent

arXiv:2603.13994v2 Announce Type: replace-cross Abstract: Vision foundation models trained with self-supervised objectives achieve strong performance across diverse tasks and exhibit emergent object s

Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation

SafetyDGX agent

arXiv:2510.08807v3 Announce Type: replace-cross Abstract: From loco-motion to dextrous manipulation, humanoid robots have made remarkable strides in demonstrating complex full-body capabilities. Howev

HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better

AgentsDGX agent

arXiv:2607.04884v1 Announce Type: new Abstract: We present HunyuanOCR-1.5, a lightweight end-to-end OCR-specialized vision-language model. HunyuanOCR unifies document parsing, text spotting, informati

HVR-Met: A Hypothesis-Verification-Replanning Agentic System for Extreme Weather Diagnosis

Model ReleasesDGX agent

arXiv:2603.01121v2 Announce Type: replace Abstract: While deep learning-based weather forecasting paradigms have made significant strides, addressing extreme weather diagnostics remains a formidable c

Hybrid Deep Learning for Traceability and Classification of Industrial Slate Tiles

ApplicationsDGX agent

arXiv:2607.04811v1 Announce Type: new Abstract: Applying deep learning to instance-aware reidentification of slate tiles and extraction site classification can improve production efficiency and qualit

Hyper-KGGen: A Skill-Driven Knowledge Extractor for High-Quality Knowledge Hypergraph Generation

Model ReleasesDGX agent

arXiv:2602.19543v2 Announce Type: replace Abstract: Knowledge hypergraphs surpass traditional binary knowledge graphs by encapsulating complex n-ary atomic facts, providing a more comprehensive paradi

Hyperparameter Transfer in Graph Neural Networks

ResearchDGX agent

arXiv:2607.05017v1 Announce Type: cross Abstract: The performance of deep learning models crucially depends on the settings of hyperparameters like learning rate, initialization scale, and weight deca

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion

HardwareDGX agent

arXiv:2607.03012v1 Announce Type: cross Abstract: Video Diffusion Transformers (VDiTs) have demonstrated significant capabilities in high-fidelity video generation. However, their ability to produce l

i-IF-Learn: Iterative Feature Selection and Unsupervised Learning for High-Dimensional Complex Data

TutorialsDGX agent

arXiv:2603.24025v2 Announce Type: replace Abstract: Unsupervised learning of high-dimensional data is challenging due to irrelevant or noisy features obscuring underlying structures. It's common that

IBIS: A Hybrid Inception-BiLSTM and SVM Ensemble for Robust Doppler-based Human Activity Recognition

ApplicationsDGX agent

arXiv:2510.24936v3 Announce Type: replace Abstract: Wi-Fi sensing is a leading technology for Human Activity Recognition (HAR), offering a non-intrusive and cost-effective solution for healthcare and

ICME 2026 Grand Challenge on Cross-Scenario Defect Detection and Fine-Grained Severity Grading for High-Precision Manufacturing

Model ReleasesDGX agent

arXiv:2607.04675v1 Announce Type: new Abstract: This paper presents the IEEE International Conference on Multimedia and Expo (ICME) 2026 Grand Challenge on Cross-Scenario Defect Detection and Fine-Gra

ICR-RL: Deep Reinforcement Learning via In-Context Regression

Model ReleasesDGX agent

arXiv:2509.11259v2 Announce Type: replace-cross Abstract: Recent advancements in machine learning have largely been driven by foundation models (FMs) trained on large, diverse datasets, enabling them

IDEAL-Bench: Indoor Dataset and Evaluation suite for Analyzing 3D Layout reasoning

Model ReleasesDGX agent

arXiv:2607.03614v1 Announce Type: new Abstract: Spatial question answering is the dominant paradigm for evaluating spatial intelligence in Vision-Language Models (VLMs), but it leaves a complementary

Identifiability of Relational Queries in Multi-View Pretraining

ApplicationsDGX agent

arXiv:2607.04735v1 Announce Type: cross Abstract: When data sources are integrated through a shared interface, a downstream query may or may not be determined by what the interface exposes: two global

Identifiability Without Gaussianity: Symbolic World Models and Near-Infinite Temporal Consistency

SafetyDGX agent

arXiv:2606.12471v2 Announce Type: replace-cross Abstract: Klindt, LeCun, and Balestriero (arXiv:2605.26379) proved that Joint-Embedding Predictive Architectures (JEPAs) achieve linear identifiability,

iFLYTEK-Embodied-Omni Technical Report

ResearchDGX agent

arXiv:2607.02542v1 Announce Type: new Abstract: General-purpose embodied agents must understand multimodal instructions, anticipate how their environment will evolve, and produce precise control actio

!Imperio, smolVLA: The Implications of Data Poisoning on Open Source Robotics

ApplicationsDGX agent

arXiv:2607.04146v1 Announce Type: cross Abstract: This work establishes that trigger-word data poisoning of vision language action models is practical, while at the same time the open-source robotics

Implicit Bias of SGD in Multivariate ReLU Networks: Effective Width Collapse

SafetyDGX agent

arXiv:2607.03613v1 Announce Type: new Abstract: We study the implicit bias of noisy stochastic gradient descent in training wide two-layer ReLU networks for multivariate regression. In a mean-field re

Improved Finite-Particle Convergence Rates for Stein Variational Gradient Descent

ResearchDGX agent

arXiv:2409.08469v4 Announce Type: replace-cross Abstract: We provide finite-particle convergence rates for the Stein Variational Gradient Descent (SVGD) algorithm in the Kernelized Stein Discrepancy (

Improved generalization bounds for binary linear classification via isoperimetry

ResearchDGX agent

arXiv:2505.16713v3 Announce Type: replace-cross Abstract: We examine the concentration of uniform generalization errors around their expectation in binary linear classification problems via an isoperi

Improving LLMs via Validator-to-Generator Alignment

SafetyDGX agent

arXiv:2607.02668v1 Announce Type: new Abstract: Large language models are inconsistent: varying prompts or including unrelated information can lead to unexpected changes in model outputs. The generato

ImputeECG: Deep Learning Reconstruction of Complete 12-Lead Electrocardiograms from Incomplete Recordings for Cardiac Assessment

ApplicationsDGX agent

arXiv:2607.05009v1 Announce Type: cross Abstract: Complete digital 12-lead electrocardiograms (ECGs) are essential for AI-enabled cardiovascular assessment, yet many clinical ECG records, particularly

In-span learning: adapting reduced-order models using their own predictions

ResearchDGX agent

arXiv:2607.02937v1 Announce Type: new Abstract: Reduced-order models compress high-dimensional dynamics into low-dimensional representations that can be evaluated rapidly, but they lose accuracy when

Incentivizing Vision Language Models to Search for Long Video Question Answering

AgentsDGX agent

arXiv:2607.02959v1 Announce Type: new Abstract: We introduce VSeek, an agentic framework that transforms long-video question answering (LVQA) from a passive, single-pass perception task into a multi-t

Inclusive KL Gradient Flows: Otto-Wasserstein, Fisher-Rao-Gaussian, and Local-Estimator Dynamics

ResearchDGX agent

arXiv:2411.00214v2 Announce Type: replace-cross Abstract: Otto's Wasserstein gradient flow of the inclusive (forward) Kullback--Leibler (KL) divergence offers a principled framework for analyzing stat

Incremental Learning of Sparse Attention Patterns in Transformers

TutorialsDGX agent

arXiv:2602.19143v2 Announce Type: replace Abstract: This paper studies simple transformers trained on a high-order Markov chain, where the model must incorporate information from multiple past positio

Individual Parameters in Weight-Sparse Transformers Appear Interpretable

ResearchDGX agent

arXiv:2607.02964v1 Announce Type: cross Abstract: A central goal of mechanistic interpretability is to understand how neural networks work and what each individual component does. Dominant circuit-fin

Induction Heads Interpolate N-Grams

TutorialsDGX agent

arXiv:2607.02800v1 Announce Type: new Abstract: Induction heads are attention circuits believed to underlie in-context learning in transformers, yet a precise characterization of the estimators they i

Industrial3D: A Water-Treatment TLS Point Cloud Dataset and Cross-Paradigm Benchmark for MEP Scene Understanding

Model ReleasesDGX agent

arXiv:2603.28660v2 Announce Type: replace Abstract: Automated semantic understanding of dense terrestrial laser scanning (TLS) point clouds is a prerequisite for Scan-to-BIM, digital twin maintenance,

IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation

Model ReleasesDGX agent

arXiv:2511.17384v2 Announce Type: replace-cross Abstract: While Visual Large Language Models (VLLMs) show great promise as embodied agents, they continue to face substantial challenges in spatial reas

Inelastic Constitutive Kolmogorov-Arnold Networks: A generalized framework for automated discovery of interpretable inelastic material models

ResearchDGX agent

arXiv:2602.17750v2 Announce Type: replace-cross Abstract: A key problem of solid mechanics is the identification of the constitutive law of a material, that is, the relation between strain history and

InFlux++: Real and Synthetic Data for Estimating Dynamic Camera Intrinsics

Model ReleasesDGX agent

arXiv:2607.05389v1 Announce Type: new Abstract: Camera intrinsics are vital for recovering 3D structure from 2D video. However, most 3D algorithms assume fixed intrinsics throughout a video, an assump

Information-Geometric Superposed Vowel Evaluation: Part 1. Moraic Syllabary (Japanese)

ResearchDGX agent

arXiv:2607.04154v1 Announce Type: cross Abstract: This paper explains the principles and provides examples of a new method for distinguishing between FAKE human speech synthesized by generative AI and

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection

Model ReleasesDGX agent

arXiv:2607.03795v1 Announce Type: new Abstract: Robust object detection under adverse visual conditions remains a long-standing challenge for multi-modal perception systems. Existing fusion-based meth

Inpainting U-Net for seamless pedestrian-level wind prediction across urban morphologies

Local AiDGX agent

arXiv:2607.02560v1 Announce Type: new Abstract: Pedestrian-level wind prediction is essential for urban design and wind-comfort assessment, but high-fidelity simulations such as LES remain computation

Input Pathways Shape Few-Shot, Not Zero-Shot, Binding in Tiny Transformers: A Fully-Enumerable Study

Model ReleasesDGX agent

arXiv:2607.04926v1 Announce Type: cross Abstract: How does the way information reaches a transformer -- as symbolic tokens, a clean per-factor 'oracle' code, or an entangled perceptual vector -- shape

InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360{eg} Image

ResearchDGX agent

arXiv:2607.03990v1 Announce Type: new Abstract: Recent advances in single image-to-3D generation have enabled high-quality asset synthesis, yet extending these capabilities to indoor scene generation

Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas

SafetyDGX agent

arXiv:2607.04710v1 Announce Type: new Abstract: Inducing cooperation among distributed agents is still a difficult problem in the field of multi-agent reinforcement learning (MARL), particularly in so

Integrated Forward-Inverse Network for Lensless Image Reconstruction

ResearchDGX agent

arXiv:2607.04608v1 Announce Type: new Abstract: Lensless imaging enables compact and versatile computational cameras by replacing bulky optics with thin coded elements. However, reconstruction from th

Integrated Graph Search and Model Predictive Control for Smooth and Efficient Path Planning in Autonomous Vehicles

SafetyDGX agent

arXiv:2607.04259v1 Announce Type: new Abstract: Path planning is a fundamental component of autonomous vehicles, where achieving safe, comfortable, and dynamically feasible paths while ensuring comput

← Previous
1…242243244245246…1005
Next →