AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
1 Jul 2026

SHMoAReg: Spark Deformable Image Registration via Spatial Heterogeneous Mixture of Experts and Attention Heads

ResearchDGX agent

arXiv:2509.20073v2 Announce Type: replace Abstract: Encoder-Decoder architectures are widely used in deep learning-based Deformable Image Registration (DIR), where the encoder extracts multi-scale fea

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

AgentsDGX agent

arXiv:2606.31693v1 Announce Type: cross Abstract: The wave of AI-native applications is moving shopping beyond page- and feed-based browsing toward intent-driven experiences orchestrated by LLM agents

Signed-Permutation Coordinate Transport for RMSNorm Transformers

Model ReleasesDGX agent

arXiv:2606.31963v1 Announce Type: cross Abstract: Modern LLM workflows move coordinate-indexed objects across checkpoints: steering vectors, sparse autoencoders, top-k neuron sets, attribution lists,

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Simple Supervision Is Hard to Beat: A Bitter Lesson from Sparse Target Labels in Domain-Adaptive Object Detection

ResearchDGX agent

arXiv:2606.30795v1 Announce Type: new Abstract: Source-free domain adaptive object detection adapts a source-trained detector to an unlabeled target domain, typically through teacher-student self-trai

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search

Model ReleasesDGX agent

arXiv:2606.31504v1 Announce Type: new Abstract: We present SimpleSearch-VL, an efficient, reliable, and practical framework for multimodal agentic search. Its core idea is to improve the agent's own s

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders

SafetyDGX agent

arXiv:2606.15054v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) detect features via inner product, so a feature's activation scales with both its directional alignment and the input's n

SkillSpotter: Pose-Aware Multi-View Skilled Action Detection and Grading in Ego-Exo Videos

Model ReleasesDGX agent

arXiv:2606.31127v1 Announce Type: cross Abstract: To enable personalized, real-time coaching using Augmented Reality glasses or fixed camera setups in domains such as sports, cooking, or music, a syst

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

SafetyDGX agent

arXiv:2606.31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, volt

Sparsity-Inducing Divergence Losses for Biometric Verification

ResearchDGX agent

arXiv:2606.31664v1 Announce Type: cross Abstract: Performance in face and speaker verification is largely driven by margin-penalty softmax losses such as CosFace and ArcFace. Recently introduced alpha

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation

ResearchDGX agent

arXiv:2606.31285v1 Announce Type: new Abstract: Human reasoning is inherently multimodal: when problems become difficult, we rarely think in words alone. We often externalize our reasoning by sketchin

SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision

SafetyDGX agent

arXiv:2603.24036v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) enables real-time, photorealistic novel view synthesis, making it a highly attractive representation for model-based vi

SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views

ResearchDGX agent

arXiv:2509.17246v2 Announce Type: replace Abstract: We introduce SPFSplatV2, an efficient feed-forward framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth pose

SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE

ResearchDGX agent

arXiv:2606.32033v1 Announce Type: new Abstract: We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical pri

SpikeLogBERT: Energy-Efficient Log Parsing Using Spiking Transformer Networks

Model ReleasesDGX agent

arXiv:2606.31781v1 Announce Type: cross Abstract: Log parsing is a fundamental step in automated log analysis, transforming raw system logs into structured event templates for downstream tasks such as

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

SafetyDGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

Stage-Transition Dense Reward Modeling for Reinforcement Learning

ResearchDGX agent

arXiv:2606.31377v1 Announce Type: cross Abstract: Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping si

Stealthy Multi-Task Adversarial Attacks

SafetyDGX agent

arXiv:2411.17936v2 Announce Type: replace-cross Abstract: Deep neural networks are highly vulnerable to adversarial perturbations, raising serious safety concerns in the real-world systems. While prio

STEB: Style Text Embedding Benchmark

Model ReleasesDGX agent

arXiv:2606.31741v1 Announce Type: cross Abstract: While semantic embeddings are rigorously evaluated on the Massive Text Embedding Benchmark, the evaluation of style embeddings remains fragmented, wit

StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation

ResearchDGX agent

arXiv:2602.23721v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models integrate visual observations and language instructions to predict robot actions, demonstrating promising

Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance

TutorialsDGX agent

arXiv:2506.20995v4 Announce Type: replace Abstract: We propose a step-by-step video-to-audio (V2A) generation method that provides finer control over the generation process and more realistic audio sy

Streaming Gaussian Encoding for 4D Panoptic Occupancy Tracking

ResearchDGX agent

arXiv:2606.30754v1 Announce Type: new Abstract: Camera-based 4D panoptic occupancy tracking (4D-POT) is a promising paradigm for holistic scene understanding from multi-view imagery, enabling joint re

Structural Preservation and the Logical Expressiveness of Graph Neural Networks

SafetyDGX agent

arXiv:2606.17882v2 Announce Type: replace Abstract: Bridges between graph neural networks (GNNs) and logical formalisms have been established by fixing architectural choices, such as the types of aggr

Structure-Regularized Interpretable TCR-Epitope Prediction

Model ReleasesDGX agent

arXiv:2606.30902v1 Announce Type: cross Abstract: T cell receptor (TCR)-epitope binding prediction is essential for understanding adaptive immunity and developing immunotherapies. Existing sequence- a

Structured SIR: Efficient and Expressive Importance-Weighted Inference for High-Dimensional Image Registration

ResearchDGX agent

arXiv:2603.17415v2 Announce Type: replace-cross Abstract: Image registration is an ill-posed dense vision task, where multiple solutions achieve similar loss values, motivating probabilistic inference

Surprise as a Signal for Plasticity and Metacognition

ResearchDGX agent

arXiv:2606.31495v1 Announce Type: new Abstract: We study a single idea across two settings: that a prediction-error signal, computed by a small predictor over the latent space of a frozen encoder, can

Surrogate Fidelity: When Can Open LLMs Explain Closed Ones?

Model ReleasesDGX agent

arXiv:2606.32008v1 Announce Type: new Abstract: Mechanistic interpretability (MI) requires full access to model internals, yet the APIs for most widely deployed language models at best expose log-prob

Surrogate-Gated Generation and Foundation-Model Embeddings for Bayesian Materials Design

Model ReleasesDGX agent

arXiv:2606.28578v1 Announce Type: cross Abstract: Closed-loop materials discovery iterates between proposing candidate structures and evaluating their properties, and property evaluation dominates the

SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation

ResearchDGX agent

arXiv:2606.31259v1 Announce Type: cross Abstract: Diffusion-based text-to-audio (TTA) models achieve impressive synthesis quality but suffer from high inference latency due to iterative multi-step den

Symmetry in language statistics shapes the geometry of model representations

ResearchDGX agent

arXiv:2602.15029v3 Announce Type: replace-cross Abstract: The internal representations learned by language models consistently exhibit striking geometric structure: calendar months organize into a cir

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation

SafetyDGX agent

arXiv:2606.30849v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have significantly advanced audio-driven portrait animation, but their high computational cost leads to substantial infere

T-QPM: Enabling Temporal Out-Of-Distribution Detection and Domain Generalization for Vision-Language Models in Open-World

ResearchDGX agent

arXiv:2603.18481v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection remains a critical challenge in open-world learning, where models must adapt to evolving data distributions. Whi

TabPATE: Differentially Private Tabular In-Context Learning Without Public Data

ResearchDGX agent

arXiv:2606.31474v1 Announce Type: new Abstract: Tabular foundation models enable accurate in-context learning (ICL) from small labeled datasets, but the private records placed in context can leak thro

TactX: Learning Shared Tactile Representations Across Diverse Sensors

SafetyDGX agent

arXiv:2606.31236v1 Announce Type: new Abstract: Tactile sensors provide critical information for contact-rich manipulation, yet tactile representations and policies remain tightly coupled to each spec

TAG-DLM: Diffusion Language Models for Text-Attributed Graph Learning

Local AiDGX agent

arXiv:2606.31166v1 Announce Type: new Abstract: Text-attributed graphs (TAGs), where each node carries a natural language description, require models to jointly reason over text and graph topology. Ex

Tailored minimal reservoir computing: on the bidirectional connection between nonlinearities in the reservoir and in data

Model ReleasesDGX agent

arXiv:2504.17503v2 Announce Type: replace Abstract: We study how the degree of nonlinearity in the input data affects the optimal design of reservoir computers, focusing on how closely the model's non

TAPE: Tether-Aware Path Planning for Autonomous Exploration of Unknown 3D Cavities Using a Tangle-Compatible Tethered Aerial Robot

AgentsDGX agent

arXiv:2606.30817v1 Announce Type: new Abstract: This letter presents the first method for autonomous exploration of unknown cavities in three dimensions (3D) that focuses on minimizing the distance tr

TaxoMIL: Taxonomy-Constrained Learning for Hierarchical Whole Slide Image Analysis

Model ReleasesDGX agent

arXiv:2606.31100v1 Announce Type: new Abstract: Whole slide image (WSI) analysis is central to computational pathology, with multiple instance learning (MIL) emerging as the standard pipeline for slid

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

SafetyDGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

Teaching LLMs String Matching, Backtracking, and Error Recovery to Deduce Bases and Truth Tables for the Combinatorially Exploding Bit Manipulation Puzzles

Model ReleasesDGX agent

arXiv:2606.23672v2 Announce Type: replace Abstract: This paper presents our algorithmic innovations for the NVIDIA Nemotron Model Reasoning Challenge, focusing on Bit Manipulation Puzzles. In this tas

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

Local AiDGX agent

arXiv:2606.31036v1 Announce Type: new Abstract: Specialist epilepsy expertise is scarce in resource-constrained settings, making LLM-based decision support attractive for frontline clinicians managing

Teaching Models to Teach Themselves: Reasoning at the Edge of Learnability

ResearchDGX agent

arXiv:2601.18778v3 Announce Type: replace-cross Abstract: RL methods for scaling large reasoning models stall on datasets with low initial success rates, and thus little training signal. We investigat

Team MKC at CLPsych 2026: Capturing and Characterizing Mental Health Changes through Social Media Timeline Dynamics

ResearchDGX agent

arXiv:2606.31464v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have motivated their adoption across a wide range of domains, including Artificial Intelligence (AI) f

Technical Report of RoboSpatial Challenge at CVPR 2026: Selective Reasoning Activation and Reference-Frame Disambiguation for Embodied Spatial Reasoning

ResearchDGX agent

arXiv:2606.31645v1 Announce Type: new Abstract: Vision-language models achieve strong general perception but often struggle with the spatial reasoning required for embodied tasks. We present RoboSpati

Temperature Field Reconstruction of Tungsten Monoblock Divertor on EAST using Physics-aware Neural Operator Transformer

Model ReleasesDGX agent

arXiv:2606.31574v1 Announce Type: cross Abstract: Accurate modeling of the divertor temperature field is essential for preventing material melting and damage and for extending the service life of fusi

Temporal Preservation over Processing: Diagnosing and Designing Spatiotemporal Single-Stage Video Detectors

ResearchDGX agent

arXiv:2606.31421v1 Announce Type: cross Abstract: Single-stage video object detectors are increasingly deployed in time-critical applications, yet it remains unclear whether these models genuinely rea

Temporal Training Strategies for Left Atrium and Left Atrial Appendage Segmentation in Dynamic Contrast 4DCT

ResearchDGX agent

arXiv:2606.31444v1 Announce Type: new Abstract: Dynamic contrast-enhanced cardiac CT enables time-resolved analysis of contrast filling and washout in the left atrium (LA) and left atrial appendage (L

TerraDiT-Omega: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive

Local AiDGX agent

arXiv:2606.31029v1 Announce Type: new Abstract: Generative models have achieved remarkable progress, yet applying them to satellite imagery remains challenging. Unlike natural imagery, satellite scene

Test-Time Verification for Text-to-SQL via Outcome Reward Models

SafetyDGX agent

arXiv:2606.30851v1 Announce Type: cross Abstract: Improving the reliability of large language models (LLMs) at inference time is a central challenge in structured reasoning tasks such as Text-to-SQL.

The Bidirectional Process Reward Model

Model ReleasesDGX agent

arXiv:2508.01682v3 Announce Type: replace Abstract: Process Reward Models (PRMs), which assign fine-grained scores to intermediate reasoning steps within a solution trajectory, have emerged as a promi

The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

Model ReleasesDGX agent

arXiv:2606.31273v1 Announce Type: new Abstract: AI-assisted research has entered a stage in which the central question is not only whether systems can generate hypotheses, run experiments, or produce

The Consistency Dilemma in LLMs: Generator-Evaluator Agreement and Vulnerability to Mistakes

AgentsDGX agent

arXiv:2606.30653v1 Announce Type: cross Abstract: Large language models are increasingly deployed in agentic pipelines that depend on the model evaluating its own outputs without external verification

The Decomposition Is the Fingerprint: Per-Component Identity for Agent Skills

Model ReleasesDGX agent

arXiv:2606.31272v1 Announce Type: cross Abstract: AI agents increasingly acquire and execute skills at runtime: bundles of prompt instructions, executable code, and tool declarations fetched from mark

The Geometry of Efficient Nonconvex Sampling

ResearchDGX agent

arXiv:2603.25622v2 Announce Type: replace-cross Abstract: We present an efficient algorithm for uniformly sampling from an arbitrary compact body X subset R^n from a warm start under isoperimetry and

The HydroGym Reinforcement Learning Platform for Fluid Dynamics

ResearchDGX agent

arXiv:2512.17534v2 Announce Type: replace-cross Abstract: Modeling and controlling fluids is critical across science and engineering. Effective flow control can increase lift, reduce drag, enhance mix

The Impact of Dimensionality on the Stability of Node Embeddings

ResearchDGX agent

arXiv:2604.08492v2 Announce Type: replace Abstract: Previous work has shown that node embedding methods can produce different representations and downstream predictions across repeated training runs,

The Label Imitation Game: Turing Test Network for Zero-Shot Pseudo-Label Pruning

ResearchDGX agent

arXiv:2606.30875v1 Announce Type: cross Abstract: Foundation model pseudo-labeling - labeling data strictly via zero-shot inference - enables massive scale, but performance is undermined by hallucinat

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory

SafetyDGX agent

arXiv:2606.31121v1 Announce Type: new Abstract: Sequentially evolving LLM memory enables agents to reuse past experience, but existing systems usually deploy each locally generated memory update witho

The Quadruped Soft Tail: Compliant Grasping and Swabbing for Contamination Surveys in Harsh Environments

ResearchDGX agent

arXiv:2606.30900v1 Announce Type: new Abstract: Beryllium contamination surveys in radioactive areas are challenging for robots in environments cluttered with cables and electronics. To address this p

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action

Model ReleasesDGX agent

arXiv:2606.31916v1 Announce Type: new Abstract: Theory of Mind (ToM) benchmarks for Large Language Models (LLMs) typically rely on passive question-answering formats, but the deployment of LLMs in inc

Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

Model ReleasesDGX agent

arXiv:2606.31648v1 Announce Type: new Abstract: We present LuckyStar 111B, a 111B-parameter hybrid reasoning model developed through a collaboration between Cohere and LG CNS for Korean-English enterp

← Previous
1…305306307308309…1025
Next →