AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Diffusion Models Observe Only Gradients: A Geometric Perspective on Score Matching Errors

DGX agent

arXiv:2606.06179v2 Announce Type: replace-cross Abstract: Score-based diffusion models are typically trained by minimizing the L^2 score matching error, and standard theoretical analyses rely on this

researcharxiv-cs-lg
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model

DGX agent

arXiv:2606.30292v1 Announce Type: cross Abstract: We present DreamForge-World 0.1 Preview, a preview foundational world model for real-time interactive world simulation. The system adapts the LongLive

hardwarearxiv-cs-cv
30 Jun 2026
Model Releases

DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation

DGX agent

arXiv:2606.29961v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents can solve complex procedural tasks by interacting with environments over multiple turns, but this ability typi

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Em-ergence of the em-dash: a population-level rise in em-dash frequency in medRxiv preprints at the dawn of the large-language-model era

DGX agent

arXiv:2606.29540v1 Announce Type: cross Abstract: Large language models (LLMs) can leave subtle stylistic traces in assisted text; one of the most cited is the em-dash (Unicode U+2014). Yet no one has

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Evolutional Math: Cross-Validated Island-Model Genetic Programming for Interpretable Symbolic Regression on Small, Wide Datasets

DGX agent

arXiv:2606.28381v1 Announce Type: cross Abstract: Symbolic regression via genetic programming routinely fails on small, wide datasets - a regime common in clinical-trial monitoring, biostatistics, and

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

FDM-MFVT: Few-step Sampling Diffusion Model for Mask-Free Virtual Try-On

DGX agent

arXiv:2606.29319v1 Announce Type: new Abstract: Image-based Virtual Try-On (IVTON) has greatly advanced through diffusion models, yet existing methods require many sampling steps and depend on masks w

safetyarxiv-cs-cv
30 Jun 2026
Research

From Word Sequences to Behavioral Sequences: Adapting Modeling and Evaluation Paradigms for Longitudinal NLP

DGX agent

arXiv:2601.07988v2 Announce Type: replace-cross Abstract: While NLP typically treats documents as independent and unordered samples, in longitudinal studies, this assumption rarely holds: documents ar

researcharxiv-cs-ai
30 Jun 2026
Safety

General and Efficient Steering of Diffusion Models

DGX agent

arXiv:2602.11395v2 Announce Type: replace-cross Abstract: Steering diffusion models toward conditions unseen during training typically requires either retraining with conditional inputs or per-step gr

safetyarxiv-cs-ai
30 Jun 2026
Local Ai

HDDPM: Heteroscedastic Denoising Diffusion Probabilistic Model for Quantitative Low-Count Brain PET Recovery

DGX agent

arXiv:2606.28513v1 Announce Type: cross Abstract: Positron emission tomography (PET) seeks to balance diagnostic quality with ra-diation dose. Low-count PET noise is non-Gaussian, non-stationary, and

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations

DGX agent

arXiv:2606.28556v1 Announce Type: new Abstract: Recent advances in large language models and vision-language models have enabled reasoning over multimodal data, offering opportunities for clinical app

model-releasesarxiv-cs-ai
30 Jun 2026
Local Ai

Improvement of Robot's Simultaneous Localization and Mapping Using an Effective Transformation to Achieve Linear Model

DGX agent

arXiv:2606.28475v1 Announce Type: cross Abstract: Nowadays mobile robots have wide engineering applications. Simultaneous localization and mapping (SLAM) is an important task of these robots. The majo

local-aiarxiv-cs-ai
30 Jun 2026
Agents

Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents

DGX agent

arXiv:2606.29459v1 Announce Type: cross Abstract: Inverse design of metal-organic frameworks (MOFs) requires searching a combinatorially vast space where property labels are expensive and most machine

agentsarxiv-cs-ai
30 Jun 2026
Safety

Mechanistically Eliciting Latent Behaviors in Language Models

DGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

safetyarxiv-cs-ai
30 Jun 2026
Research

Momentum Guidance: Plug-and-Play Guidance for Flow Models

DGX agent

arXiv:2602.20360v2 Announce Type: replace-cross Abstract: Flow-based generative methods offer a simple and effective framework for high-fidelity generation, yet pretrained flow models are rarely used

researcharxiv-cs-cv
30 Jun 2026
Local Ai

Multimodal and Multiscale Spatial-Temporal Semantic Search and Recommendation with AI Foundation Models

DGX agent

arXiv:2606.28369v1 Announce Type: cross Abstract: Semantic search and recommendation of similar documents, such as news and reports about unusual environmental events (e.g., a dead whale washed ashore

local-aiarxiv-cs-ai
30 Jun 2026
Safety

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

DGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

safetyarxiv-cs-ai
30 Jun 2026
Research

Predicting Metastatic Risk from Primary Tissue Architecture via Distance-Aware Spatial Modeling

DGX agent

arXiv:2606.28676v1 Announce Type: cross Abstract: Predicting the risk of distant metastasis from primary tumor tissue histology is a critical yet challenging task in computational pathology. Multiple

researcharxiv-cs-ai
30 Jun 2026
Safety

SGD Provably Prioritizes a Shortcut Spurious Feature in the XOR Model

DGX agent

arXiv:2606.30444v1 Announce Type: cross Abstract: Neural networks are known to be susceptible to over-reliance on spurious correlations. However, the precise mechanism by which models exploit shortcut

safetyarxiv-cs-lg
30 Jun 2026
Tutorials

Simplifying Flow Matching Transformations with Low-Rank Mixture Models

DGX agent

arXiv:2606.29724v1 Announce Type: new Abstract: Normalizing flows are powerful generative models that learn an invertible mapping between complex data distributions and simple latent distributions, ty

tutorialsarxiv-cs-lg
30 Jun 2026
Safety

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport

DGX agent

arXiv:2602.23353v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis posits that neural networks trained on different modalities converge toward a shared statistical model

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?

DGX agent

arXiv:2511.06090v3 Announce Type: replace-cross Abstract: Optimizing the performance of large-scale software repositories demands expertise in code reasoning and software engineering (SWE) to reduce r

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

TAP-VLA: Tactile Annotation Prompting for Vision Language Action Models

DGX agent

arXiv:2606.29089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate impressive reasoning over visual, semantic, and spatial task variations by leveraging large-scale vision

safetyarxiv-cs-ro
30 Jun 2026
Research

The Surprising Effectiveness of Video Diffusion Models for Hand Motion Reconstruction

DGX agent

arXiv:2606.30308v1 Announce Type: new Abstract: 4D hand motion reconstruction from egocentric video is bottlenecked by clear limitations of existing methods: image-based pipelines depend on a detector

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics

DGX agent

arXiv:2512.13660v3 Announce Type: replace-cross Abstract: Spatial tracing, as a fundamental embodied interaction ability for robots, is inherently challenging as it requires multi-step metric-grounded

model-releasesarxiv-cs-cv
30 Jun 2026
Research

When Does Synthetic CT Transfer? A Label-Free Donor/Host Diagnostic for Medical Vision-Language Model Routing on Real Lung CT

DGX agent

arXiv:2606.29232v1 Announce Type: new Abstract: A synthetic measurement of model competence is useful only if it survives the move to real data, yet the real labels that would verify it are exactly wh

researcharxiv-cs-cv
30 Jun 2026
Research

A Gaussian Perspective for Distributional Discrepancy in Generative Diffusion Models

DGX agent

arXiv:2601.13602v3 Announce Type: replace-cross Abstract: This paper introduces an analytical approach to quantifying and optimizing the distributional discrepancy in generative diffusion models. For

researcharxiv-cs-lg
29 Jun 2026
Research

An Approach to Enriching Surgical Video Datasets for Fine-Grained Spatial-Temporal Understanding of Vision-Language Models

DGX agent

arXiv:2604.00784v2 Announce Type: replace Abstract: Surgical video understanding is a crucial prerequisite for advancing Computer-Assisted Surgery. While vision-language models (VLMs) have recently be

researcharxiv-cs-cv
29 Jun 2026
Applications

Are Time-Series Foundation Models Ready for E-Nose Data? An Empirical Assessment of Their Embeddings

DGX agent

arXiv:2606.27672v1 Announce Type: new Abstract: Inspired by advances in natural language processing and computer vision, 'time-series foundation models' (TSFMs) have recently been introduced with the

applicationsarxiv-cs-lg
29 Jun 2026
Research

Class-frequency Guided Noise Schedule for Diffusion Models

DGX agent

arXiv:2606.27696v1 Announce Type: cross Abstract: In this paper, we are the first to examine the correlations between class frequency and the multi-scale noise schedule within diffusion models. For sc

researcharxiv-cs-ai
29 Jun 2026
Tutorials

DFM: Difference Feature Modeling with Text-Guided Gated Contrastive Loss for Remote Sensing Image Change Captioning

DGX agent

arXiv:2606.27410v1 Announce Type: cross Abstract: The primary goal of Remote Sensing Image Change Captioning (RSICC) is to automatically generate descriptions of changes between remote sensing images

tutorialsarxiv-cs-lg
29 Jun 2026
Local Ai

DIM-WAM: World-Action Modeling with Diverse Historical Event Memory

DGX agent

arXiv:2606.27677v1 Announce Type: cross Abstract: World-action models have shown promising robot-manipulation performance by jointly predicting future visual states and actions. However, existing meth

local-aiarxiv-cs-cv
29 Jun 2026
Model Releases

End-to-End Dynamic Sparsity for Resource-Adaptive LLM Inference

DGX agent

arXiv:2606.27743v1 Announce Type: cross Abstract: Large Language Models (LLMs) inference is typically deployed under a static resource assumption, where models execute a fixed computational graph rega

model-releasesarxiv-cs-ai
29 Jun 2026
Local Ai

Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models

DGX agent

arXiv:2606.27593v1 Announce Type: new Abstract: We introduce a categorical framework called ODYSSEY for constructing verifiable, local truth-preserving foundation models as compositions of foundries:

local-aiarxiv-cs-ai
29 Jun 2026
Local Ai

Perceptual 3D Simulation With Physical World Modeling

DGX agent

arXiv:2606.27575v1 Announce Type: new Abstract: Predicting how a scene will evolve after a desired 3D transformation from images is a central goal in vision, graphics, and robotics. Yet unlike ideal s

local-aiarxiv-cs-cv
29 Jun 2026
Tutorials

Physics-Informed Neural Network with Transfer Learning for State Estimation in Lithium-Ion Batteries using the Single Particle Model with Electrolyte

DGX agent

arXiv:2606.28220v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful tool for solving nonlinear partial differential equations (PDEs), including battery

tutorialsarxiv-cs-lg
29 Jun 2026
Safety

PRISON: Unmasking the Criminal Potential of Large Language Models

DGX agent

arXiv:2506.16150v4 Announce Type: replace-cross Abstract: As large language models (LLMs) advance, concerns about their misconduct in complex social contexts intensify. Existing research overlooked th

safetyarxiv-cs-ai
29 Jun 2026
Research

Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs

DGX agent

arXiv:2603.20239v2 Announce Type: replace-cross Abstract: 3D Scene Graphs (3DSGs) provide hierarchical, multi-resolution abstractions that encode the geometric and semantic structure of an environment

researcharxiv-cs-cv
29 Jun 2026
Local Ai

SelectAnyTree: A Promptable Instance Segmentation Model for 3D Forest LiDAR Point Clouds

DGX agent

arXiv:2606.27491v1 Announce Type: new Abstract: Automated instance segmentation of forest LiDAR point clouds is increasingly critical as forest monitoring moves toward scalable, detailed, 3D measureme

local-aiarxiv-cs-cv
29 Jun 2026
Safety

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks

DGX agent

arXiv:2606.27807v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a dominant paradigm for embodied intelligence. However, most existing approaches are built on large-scal

safetyarxiv-cs-ro
29 Jun 2026
Research

VGB for Masked Diffusion Model: Efficient Test-time Scaling for Reward Satisfaction and Sample Editing

DGX agent

arXiv:2606.28301v1 Announce Type: new Abstract: Inference-time scaling is a promising paradigm to improve generative models, especially when outputs must satisfy structural constraints or optimize dow

researcharxiv-cs-lg
29 Jun 2026
Local Ai

EndoUFM: Utilizing Foundation Models for Monocular depth estimation of endoscopic images

DGX agent

arXiv:2508.17916v2 Announce Type: replace Abstract: Depth estimation is a foundational component for 3D reconstruction in minimally invasive endoscopic surgeries. However, existing monocular depth est

local-aiarxiv-cs-cv
26 Jun 2026
Research

Generative Models on Analog Hardware with Dynamics

DGX agent

arXiv:2606.27294v1 Announce Type: cross Abstract: Analog hardware platforms such as coupled oscillators and Analog Ising Machines naturally solve differential equations at a fraction of the energy cos

researcharxiv-cs-lg
26 Jun 2026
Safety

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

DGX agent

arXiv:2606.26502v1 Announce Type: new Abstract: Large reasoning models (LRMs) take longer on harder problems, just as humans do. This surface similarity hides an opposite pattern within items. When an

safetyarxiv-cs-ai
26 Jun 2026
Safety

Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision

DGX agent

arXiv:2606.26801v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for generalizable robotic manipulation. During fine-tuning, however, action supervisio

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models

DGX agent

arXiv:2606.26379v1 Announce Type: new Abstract: Visual prompt tuning has emerged as a parameter-efficient fine-tuning approach for adapting large-scale Vision Transformers (ViTs) to downstream tasks.

model-releasesarxiv-cs-cv
26 Jun 2026
Research

LithoDreamer: A Physics-Informed World Model for Multi-Stage Computational Lithography

DGX agent

arXiv:2606.26713v1 Announce Type: new Abstract: As semiconductor technology nodes scale, computational lithography is essential for ensuring yield and performance. However, lithography is a continuous

researcharxiv-cs-ai
26 Jun 2026
Safety

Racing a Wheeled Quadruped: Active Load Transfer Mitigation via Model Predictive Control

DGX agent

arXiv:2606.26313v1 Announce Type: new Abstract: This paper presents a hierarchical control framework using model predictive control (MPC) and reinforcement learning (RL) for active roll control to man

safetyarxiv-cs-ro
26 Jun 2026
Applications

ReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning Models

DGX agent

arXiv:2606.26986v1 Announce Type: cross Abstract: Open Relation Extraction (OpenRE) requires a model to extract unseen relations between head and tail entities from unstructured text for real-world ap

applicationsarxiv-cs-ai
26 Jun 2026
← Previous
1…177178179180181…1030
Next →