AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “tutorials”

GridTimelineEvolution
3,356 results
30 Jul 2026

Chaos Is a LADDER: Domain Generalization Beyond Invariance via Reweighting

TutorialsDGX agent

arXiv:2607.26458v1 Announce Type: cross Abstract: Domain generalization (DG) aims to learn from multiple source domains and generalize to unseen target domains. Most DG methods pursue invariance: they

Explicit Kinematic Guidance from Analytic Concepts for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2607.26513v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models rely mainly on 2D inputs, neglecting the rich object structural information and commonsense knowledge inhere

From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for Embodied Intelligence

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.26903v1 Announce Type: cross Abstract: The key bottleneck in embodied AI is not model architecture but data. Although billions of human manipulation videos exist online, robots cannot direc

how to train lora, make datasheet with ai toolkit

TutorialsDGX agent

Hi! I'm trying to train my first LoRA, but I'm not sure how to handle the dataset captions. Should I describe everything in the photo, or just specific details? Also, can I train it using different im

Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick

TutorialsDGX agent

Learn how to build an inference meta-monitoring system for Amazon SageMaker AI endpoints using Amazon Quick. This governance layer sits above production ML inference pipelines to continuously track pr

Ripple: Real-Time Streaming Audio-Video Generation With Cross-Modal Recurrent Memory

TutorialsDGX agent

arXiv:2607.26818v1 Announce Type: new Abstract: Audio-video generative models achieve impressive quality but suffer from high latency, making them unsuitable for real-time applications. Although sever

SCOUT: Per-Context Reset Curricula for Sparse-Reward Reinforcement Learning

TutorialsDGX agent

arXiv:2607.26417v1 Announce Type: new Abstract: Sparse-reward reinforcement learning often fails because rollouts from the unassisted evaluation start rarely reach later task stages. Reset curricula a

Stacked sessions and pull requests in the GitHub Copilot app

TutorialsDGX agent

Learn how I modernized an old codebase of mine using stacked sessions and pull requests in the GitHub Copilot app. The post Stacked sessions and pull requests in the GitHub Copilot app appeared first

StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation

TutorialsDGX agent

arXiv:2607.26754v1 Announce Type: new Abstract: Recent game world models can generate visually realistic and interactive environments conditioned on player actions. However, games are not defined by p

Surrogate assisted diversity estimation in neural ensemble search

TutorialsDGX agent

arXiv:2607.26940v1 Announce Type: new Abstract: Ensembles are a standard way to improve the performance and robustness of deep neural networks, but their effectiveness crucially depends on both the qu

29 Jul 2026

A Causality-aware Infer-diagnose-refine Framework for Test-time Modality Adaptation in VLA Models

TutorialsDGX agent

arXiv:2607.25516v1 Announce Type: new Abstract: Vision-language-action (VLA) models predict sequential actions to execute tasks specified by language instructions, conditioned on visual observations a

AI-Assisted Knowledge Access for Legacy Enterprise Asset Management in Energy Operations: A Practical Retrieval System

TutorialsDGX agent

arXiv:2607.24792v1 Announce Type: cross Abstract: Energy utilities still run engineering work management, engineering procurement, and inventory processes on long-lived enterprise asset management pla

Algorithmic Separation between Constant-Depth and Logarithmic-Depth Neural Networks

TutorialsDGX agent

arXiv:2607.25200v1 Announce Type: new Abstract: Despite the empirical advantages of deep networks over shallow ones, theoretical depth separations largely concern approximation power, while algorithmi

Analysis of the Shortcut Learning and Clever Hans Effect in CNN based ECG Image Classification

TutorialsDGX agent

arXiv:2607.25117v1 Announce Type: cross Abstract: Deep learning models for ECG image classification may achieve high accuracy by exploiting non-physiological visual cues instead of ECG waveform morpho

Anti-Backdoor Coreset Selection via Cumulative Entropy

TutorialsDGX agent

arXiv:2607.25502v1 Announce Type: new Abstract: Recent training-time defenses against neural backdoors isolate a benign subset from poisoned training data, to learn a backdoor-free model from it. In t

Balanced Soft mixture-of-expert model for Glaucoma Detection

TutorialsDGX agent

arXiv:2607.25324v1 Announce Type: cross Abstract: Glaucoma is a group of eye diseases that damage the optic nerve, often caused by elevated intraocular pressure. It is a leading cause of irreversible

Bi-Level Collaborative Learning for Few-Shot Scribble-Supervised Medical Image Segmentation

TutorialsDGX agent

arXiv:2607.25432v1 Announce Type: new Abstract: Scribble annotations offer an efficient alternative to costly pixel-wise labeling for medical image segmentation, yet in real clinical scenarios, scribb

Contrastive Representation Learning of Longitudinal Disease Trajectories on Temporal Graphs

TutorialsDGX agent

arXiv:2607.25609v1 Announce Type: cross Abstract: Understanding disease trajectories from longitudinal clinical data remains challenging due to complex temporal dynamics and heterogeneous patient coho

Deep Label-Wise Attentive Temporal Convolutional Networks Improve Medical Coding

TutorialsDGX agent

arXiv:2607.25129v1 Announce Type: new Abstract: Medical coding is the task of assigning a set of diagnosis and procedure codes for a hospitalization using recorded notes. It requires aggregating infor

Diffusion Disambiguation Models for Partial Label Learning

TutorialsDGX agent

arXiv:2507.00411v2 Announce Type: replace Abstract: Learning from ambiguous labels is a long-standing problem in practical machine learning applications. The purpose of partial label learning (PLL) is

Explicit Layer Modeling for Video Object Insertion and Layer Decomposition

TutorialsDGX agent

arXiv:2607.25802v1 Announce Type: new Abstract: Most video editing systems still lack explicit layered video representations, limiting their ability to perform realistic compositing, object reuse, and

From Idea to Classroom in Days: Using 'Vibe Coding' to Create a Programming Process Visualizer from IDE Activity Logs

TutorialsDGX agent

arXiv:2607.24757v1 Announce Type: cross Abstract: This paper reports on the rapid development and classroom deployment of a Thonny log visualizer built using AI-assisted ``vibe coding'' to make studen

Gaussian Volumetric Representation for Efficient Shear-Warp Visualization

TutorialsDGX agent

arXiv:2607.25377v1 Announce Type: new Abstract: Medical image visualization requires volumetric rendering algorithms that preserve anatomical fidelity while maintaining high rendering speeds. To addre

HVM-GraphRAG: Holistic-View Multimodal Graph Retrieval-Augmented Generation on Complex Document

TutorialsDGX agent

arXiv:2607.24861v1 Announce Type: cross Abstract: Question answering (QA) over complex documents requires models to retrieve and integrate evidence distributed across distant document regions and moda

Jensen Huang was asked about the people technology left behind. He didn't offer a plan. He said the gap already closed. Huang: 'All of a sud…

TutorialsDGX agent

Jensen Huang was asked about the people technology left behind. He didn't offer a plan. He said the gap already closed. Huang: 'All of a sudden artificial intelligence closed that technology divide.'

Normalizing Flows to Reconstruct Pseudo-PDFs

TutorialsDGX agent

arXiv:2607.25282v1 Announce Type: cross Abstract: We investigate a normalizing-flow approach for reconstructing parton distribution functions (PDFs) from synthetic matrix-element data. Our framework c

On the Design and Evaluation of Human-centered Explainable AI Systems: A Systematic Review and Taxonomy

TutorialsDGX agent

arXiv:2510.12201v2 Announce Type: replace Abstract: As AI becomes more common in everyday living, there is an increasing demand for intelligent systems that are both performant and understandable. Exp

Patterns of Learner-AI Interaction and Academic Performance in an Object-Oriented Programming Course

TutorialsDGX agent

arXiv:2607.24755v1 Announce Type: cross Abstract: This full research paper examines how different forms of learner-AI interaction relate to learning outcomes in object-oriented programming (OOP) cours

PILA: Plug-and-Play Insertion for LLM-native Advertising

TutorialsDGX agent

arXiv:2607.25590v1 Announce Type: new Abstract: How to monetize large language models (LLMs) by naturally integrating sponsored content into their responses, known as LLM-native advertising, has recen

Preliminary Guidelines for Using and Evaluating GenAI Tools to Support Systematic Literature Reviews

TutorialsDGX agent

arXiv:2607.24991v1 Announce Type: cross Abstract: Context: Generative AI (GenAI) and Large Language Models (LLMs) are increasingly used for academic tasks in software engineering and beyond, including

Reinformed Dreamer: An Asymmetric World Model Efficiently Trained through Latent Guidance

TutorialsDGX agent

arXiv:2607.26040v1 Announce Type: new Abstract: Much like humans benefit from guidance while learning, reinforcement learning algorithms may benefit from additional supervision beyond rewards. Leverag

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control

TutorialsDGX agent

arXiv:2607.25337v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting in representation space rather than reconstructing pixels, making them

Unifying Active Learning and Semi-Supervised Learning for Medical Image Segmentation

TutorialsDGX agent

arXiv:2607.25014v1 Announce Type: new Abstract: In practical settings, medical image segmentation models are often developed with limited annotated data rather than fully labeled datasets. Training fr

Using Data-Derived Priors to Guide CNN Architecture Design for NIR Chemometrics

TutorialsDGX agent

arXiv:2607.25636v1 Announce Type: new Abstract: Convolutional neural networks (CNN) for near-infrared (NIR) chemometrics are often designed using generic architectural rules, although spectral dataset

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning

TutorialsDGX agent

arXiv:2607.25236v1 Announce Type: new Abstract: Different research lines use the term world model in different ways, yet they share a common aim: to capture how the world evolves under action in a for

28 Jul 2026

AI Strategy: How to Choose What AI Product to Implement

TutorialsDGX agent

arXiv:2607.23733v1 Announce Type: cross Abstract: Firms struggle to choose AI projects that pay off: two projects can look equally promising to smart, motivated stakeholders and yet deserve opposite d

At the core of our mission is working through how to ensure increasingly powerful AI benefits everyone. We believe that, at some point in th…

TutorialsDGX agent

At the core of our mission is working through how to ensure increasingly powerful AI benefits everyone. We believe that, at some point in the future, AI acceleration for frontier model development may

BeyondFusion: Self-Aligned Latent Diffusion for Calibration-Free Infrared Super-Resolution and Infrared-Visible Fusion

TutorialsDGX agent

arXiv:2607.24110v1 Announce Type: new Abstract: Mobile infrared-visible imaging typically pairs a compact infrared sensor with a high-resolution visible camera for complementary perception. While cros

Controlling Embedding Spaces with Text-Conditioned Transformations

TutorialsDGX agent

arXiv:2607.22919v1 Announce Type: cross Abstract: Multimodal embedding spaces in models like CLIP enable powerful capabilities such as semantic similarity retrieval and cross-modal zero-shot classific

Counterfactual Motion Reliability Learning for Robust UAV Tracking

TutorialsDGX agent

arXiv:2607.23209v1 Announce Type: new Abstract: Infrared unmanned aerial vehicle (UAV) tracking is challenging because the target is often small, low-contrast, and easily confused with thermal distrac

CrossSpine: Multi-scale Cross-sequence Attention with Anatomical Priors for Automated Pfirrmann Grading

TutorialsDGX agent

arXiv:2607.22728v1 Announce Type: new Abstract: Automated grading of Lumbar Disc Degeneration is essential for the objective quantification of structural changes associated with low back pain. Observi

Do Coverage and Mutation Scores of LLM-Generated Test Suites Correlate with Their Effectiveness? (Replicability Study)

TutorialsDGX agent

arXiv:2607.22880v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have driven growing interest in using LLMs to automate test generation. Prior work commonly evaluates

Effect of User-Prompted Priors on Semi-Automated Cancer Lesion Segmentation in Whole-Body Computed Tomography

TutorialsDGX agent

arXiv:2607.24210v1 Announce Type: new Abstract: In clinical oncology studies, metastatic cancer is commonly evaluated using 'Response Evaluation Criteria in Solid Tumors' (RECIST), in which the diamet

Extending Desbordante with Probabilistic Functional Dependency Discovery Support

TutorialsDGX agent

arXiv:2607.23636v1 Announce Type: cross Abstract: Data profiling aims to extract complex patterns from data for further analysis and use that data in domains such as data cleaning, data deduplication,

From Machine Learning to Large-Scale EO Products: Best Practices for Making Maps

TutorialsDGX agent

arXiv:2607.24532v1 Announce Type: new Abstract: Recent years have seen a rapid expansion in the production of large-scale geospatial maps derived from Earth observation (EO) data, driven largely by ad

Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention

TutorialsDGX agent

arXiv:2607.23917v1 Announce Type: new Abstract: We introduce a novel learning problem: decoding gaze into natural language descriptions of human goals across diverse visual tasks. Unlike prior work, w

GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs

TutorialsDGX agent

arXiv:2508.09478v2 Announce Type: replace Abstract: In this work, we present GazeLT, a human visual attention integration-disintegration approach for long-tailed disease classification. A radiologist'

Geometry Meets Semantics: Fractional Gradient Stabilization for Semantic-Driven Bounding Box Optimization in Visual Detection Tasks

TutorialsDGX agent

arXiv:2607.23530v1 Announce Type: new Abstract: Bounding boxes are fundamental for object localization in visual detection tasks. Among them, oriented bounding boxes are widely used in visual detectio

Gradient-Free Continual Learning

TutorialsDGX agent

arXiv:2504.01219v2 Announce Type: replace Abstract: Neural networks are notorious for forgetting old skills when taught new ones - a problem known as catastrophic forgetting. Standard continual learni

Hiding in Plain Sight: An Effective Physical Adversarial Patch Attack against Visual-Infrared Fused Face Detection

TutorialsDGX agent

arXiv:2607.23292v1 Announce Type: cross Abstract: Deep learning-based visual-infrared fused face detection models are increasingly deployed across a wide range of applications, yet they remain suscept

How to deal with text only vector search across multimodal embedding space? [D]

TutorialsDGX agent

My data set is a list of images, each equipped with a a couple sentences of text. A user would search primarily with text only. My default approach is using BM25, but how would I facilitate searching

Integrating Factual and Normative Industrial Knowledge via Constraint-Aware Graph Attention for Process Plan Recommendation

TutorialsDGX agent

arXiv:2607.24213v1 Announce Type: new Abstract: Integrating heterogeneous industrial knowledge, including factual relations and decision constraints, remains a core challenge in industrial information

Invariant Discovery for Networked Systems

TutorialsDGX agent

arXiv:2607.22944v1 Announce Type: cross Abstract: Invariants, the relations expected to hold among measured signals of a network, underpin applications from verification to traffic generation, telemet

Kernel-SDF: An Open-Source Library for Real-Time Signed Distance Function Estimation using Kernel Regression

TutorialsDGX agent

arXiv:2603.29227v2 Announce Type: replace Abstract: Accurate and efficient scene representation is crucial for robotic tasks such as motion planning, manipulation, and navigation. Signed distance func

Learned Interventions in Lean 4 grind

TutorialsDGX agent

arXiv:2607.22972v1 Announce Type: new Abstract: Lean~4's grind{} tactic combines congruence closure, ematch{}ing, and case-splitting into a single automated solver, and like any such solver, it relies

Learning Dense 2D-3D Correspondence for X-ray-to-CT Registration of Knee Bones

TutorialsDGX agent

arXiv:2607.22803v1 Announce Type: cross Abstract: Recovering the 6-DoF pose of the knee bones from a plain radiograph, given the patient's segmented pre-operative CT, turns a routine low-dose image in

Learning Distributions from Multiple Data Providers

TutorialsDGX agent

arXiv:2607.24732v1 Announce Type: cross Abstract: Motivated by learning from heterogeneous and overlapping data providers, we study a stylized model of distribution learning from restricted conditiona

Like a Baby: Visually Situated Neural Language Acquisition

TutorialsDGX agent

arXiv:1805.11546v3 Announce Type: replace-cross Abstract: We examine the benefits of visual context in training neural language models to perform next-word prediction. A multi-modal neural architectur

Like a bilingual baby: The advantage of visually grounding a bilingual language model

TutorialsDGX agent

arXiv:2210.05487v3 Announce Type: replace Abstract: Unlike most neural language models, humans learn language in a rich, multi-sensory and, often, multi-lingual environment. Current language models ty

Mamba-CL: Optimizing Selective State Space Model in Null Space for Continual Learning

TutorialsDGX agent

arXiv:2411.15469v3 Announce Type: replace Abstract: Continual Learning (CL) aims to equip AI models with the ability to learn a sequence of tasks over time, without forgetting previously learned knowl

← Previous
1…56789…56
Next →