AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
Research

Simple Supervision Is Hard to Beat: A Bitter Lesson from Sparse Target Labels in Domain-Adaptive Object Detection

DGX agent

arXiv:2606.30795v1 Announce Type: new Abstract: Source-free domain adaptive object detection adapts a source-trained detector to an unlabeled target domain, typically through teacher-student self-trai

researcharxiv-cs-cv
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search

DGX agent

arXiv:2606.31504v1 Announce Type: new Abstract: We present SimpleSearch-VL, an efficient, reliable, and practical framework for multimodal agentic search. Its core idea is to improve the agent's own s

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders

DGX agent

arXiv:2606.15054v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) detect features via inner product, so a feature's activation scales with both its directional alignment and the input's n

safetyarxiv-cs-lg
1 Jul 2026
Model Releases

SkillSpotter: Pose-Aware Multi-View Skilled Action Detection and Grading in Ego-Exo Videos

DGX agent

arXiv:2606.31127v1 Announce Type: cross Abstract: To enable personalized, real-time coaching using Augmented Reality glasses or fixed camera setups in domains such as sports, cooking, or music, a syst

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

DGX agent

arXiv:2606.31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, volt

safetyarxiv-cs-ai
1 Jul 2026
Research

Sparsity-Inducing Divergence Losses for Biometric Verification

DGX agent

arXiv:2606.31664v1 Announce Type: cross Abstract: Performance in face and speaker verification is largely driven by margin-penalty softmax losses such as CosFace and ArcFace. Recently introduced alpha

researcharxiv-cs-ai
1 Jul 2026
Research

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation

DGX agent

arXiv:2606.31285v1 Announce Type: new Abstract: Human reasoning is inherently multimodal: when problems become difficult, we rarely think in words alone. We often externalize our reasoning by sketchin

researcharxiv-cs-ai
1 Jul 2026
Safety

SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision

DGX agent

arXiv:2603.24036v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) enables real-time, photorealistic novel view synthesis, making it a highly attractive representation for model-based vi

safetyarxiv-cs-cv
1 Jul 2026
Research

SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views

DGX agent

arXiv:2509.17246v2 Announce Type: replace Abstract: We introduce SPFSplatV2, an efficient feed-forward framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth pose

researcharxiv-cs-cv
1 Jul 2026
Research

SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE

DGX agent

arXiv:2606.32033v1 Announce Type: new Abstract: We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical pri

researcharxiv-cs-cv
1 Jul 2026
Model Releases

SpikeLogBERT: Energy-Efficient Log Parsing Using Spiking Transformer Networks

DGX agent

arXiv:2606.31781v1 Announce Type: cross Abstract: Log parsing is a fundamental step in automated log analysis, transforming raw system logs into structured event templates for downstream tasks such as

model-releasesarxiv-cs-cl
1 Jul 2026
Safety

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

DGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

safetyarxiv-cs-ro
1 Jul 2026
Research

Stage-Transition Dense Reward Modeling for Reinforcement Learning

DGX agent

arXiv:2606.31377v1 Announce Type: cross Abstract: Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping si

researcharxiv-cs-ai
1 Jul 2026
Safety

Stealthy Multi-Task Adversarial Attacks

DGX agent

arXiv:2411.17936v2 Announce Type: replace-cross Abstract: Deep neural networks are highly vulnerable to adversarial perturbations, raising serious safety concerns in the real-world systems. While prio

safetyarxiv-cs-cv
1 Jul 2026
Model Releases

STEB: Style Text Embedding Benchmark

DGX agent

arXiv:2606.31741v1 Announce Type: cross Abstract: While semantic embeddings are rigorously evaluated on the Massive Text Embedding Benchmark, the evaluation of style embeddings remains fragmented, wit

model-releasesarxiv-cs-ai
1 Jul 2026
Research

StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation

DGX agent

arXiv:2602.23721v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models integrate visual observations and language instructions to predict robot actions, demonstrating promising

researcharxiv-cs-cv
1 Jul 2026
Tutorials

Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance

DGX agent

arXiv:2506.20995v4 Announce Type: replace Abstract: We propose a step-by-step video-to-audio (V2A) generation method that provides finer control over the generation process and more realistic audio sy

tutorialsarxiv-cs-cv
1 Jul 2026
Research

Streaming Gaussian Encoding for 4D Panoptic Occupancy Tracking

DGX agent

arXiv:2606.30754v1 Announce Type: new Abstract: Camera-based 4D panoptic occupancy tracking (4D-POT) is a promising paradigm for holistic scene understanding from multi-view imagery, enabling joint re

researcharxiv-cs-cv
1 Jul 2026
Safety

Structural Preservation and the Logical Expressiveness of Graph Neural Networks

DGX agent

arXiv:2606.17882v2 Announce Type: replace Abstract: Bridges between graph neural networks (GNNs) and logical formalisms have been established by fixing architectural choices, such as the types of aggr

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

Structure-Regularized Interpretable TCR-Epitope Prediction

DGX agent

arXiv:2606.30902v1 Announce Type: cross Abstract: T cell receptor (TCR)-epitope binding prediction is essential for understanding adaptive immunity and developing immunotherapies. Existing sequence- a

model-releasesarxiv-cs-lg
1 Jul 2026
Research

Structured SIR: Efficient and Expressive Importance-Weighted Inference for High-Dimensional Image Registration

DGX agent

arXiv:2603.17415v2 Announce Type: replace-cross Abstract: Image registration is an ill-posed dense vision task, where multiple solutions achieve similar loss values, motivating probabilistic inference

researcharxiv-cs-cv
1 Jul 2026
Research

Surprise as a Signal for Plasticity and Metacognition

DGX agent

arXiv:2606.31495v1 Announce Type: new Abstract: We study a single idea across two settings: that a prediction-error signal, computed by a small predictor over the latent space of a frozen encoder, can

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Surrogate Fidelity: When Can Open LLMs Explain Closed Ones?

DGX agent

arXiv:2606.32008v1 Announce Type: new Abstract: Mechanistic interpretability (MI) requires full access to model internals, yet the APIs for most widely deployed language models at best expose log-prob

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Surrogate-Gated Generation and Foundation-Model Embeddings for Bayesian Materials Design

DGX agent

arXiv:2606.28578v1 Announce Type: cross Abstract: Closed-loop materials discovery iterates between proposing candidate structures and evaluating their properties, and property evaluation dominates the

model-releasesarxiv-cs-ai
1 Jul 2026
Research

SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation

DGX agent

arXiv:2606.31259v1 Announce Type: cross Abstract: Diffusion-based text-to-audio (TTA) models achieve impressive synthesis quality but suffer from high inference latency due to iterative multi-step den

researcharxiv-cs-ai
1 Jul 2026
Research

Symmetry in language statistics shapes the geometry of model representations

DGX agent

arXiv:2602.15029v3 Announce Type: replace-cross Abstract: The internal representations learned by language models consistently exhibit striking geometric structure: calendar months organize into a cir

researcharxiv-cs-cl
1 Jul 2026
Safety

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation

DGX agent

arXiv:2606.30849v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have significantly advanced audio-driven portrait animation, but their high computational cost leads to substantial infere

safetyarxiv-cs-cv
1 Jul 2026
Research

T-QPM: Enabling Temporal Out-Of-Distribution Detection and Domain Generalization for Vision-Language Models in Open-World

DGX agent

arXiv:2603.18481v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection remains a critical challenge in open-world learning, where models must adapt to evolving data distributions. Whi

researcharxiv-cs-cv
1 Jul 2026
Research

TabPATE: Differentially Private Tabular In-Context Learning Without Public Data

DGX agent

arXiv:2606.31474v1 Announce Type: new Abstract: Tabular foundation models enable accurate in-context learning (ICL) from small labeled datasets, but the private records placed in context can leak thro

researcharxiv-cs-lg
1 Jul 2026
Safety

TactX: Learning Shared Tactile Representations Across Diverse Sensors

DGX agent

arXiv:2606.31236v1 Announce Type: new Abstract: Tactile sensors provide critical information for contact-rich manipulation, yet tactile representations and policies remain tightly coupled to each spec

safetyarxiv-cs-ro
1 Jul 2026
Local Ai

TAG-DLM: Diffusion Language Models for Text-Attributed Graph Learning

DGX agent

arXiv:2606.31166v1 Announce Type: new Abstract: Text-attributed graphs (TAGs), where each node carries a natural language description, require models to jointly reason over text and graph topology. Ex

local-aiarxiv-cs-cl
1 Jul 2026
Model Releases

Tailored minimal reservoir computing: on the bidirectional connection between nonlinearities in the reservoir and in data

DGX agent

arXiv:2504.17503v2 Announce Type: replace Abstract: We study how the degree of nonlinearity in the input data affects the optimal design of reservoir computers, focusing on how closely the model's non

model-releasesarxiv-cs-lg
1 Jul 2026
Agents

TAPE: Tether-Aware Path Planning for Autonomous Exploration of Unknown 3D Cavities Using a Tangle-Compatible Tethered Aerial Robot

DGX agent

arXiv:2606.30817v1 Announce Type: new Abstract: This letter presents the first method for autonomous exploration of unknown cavities in three dimensions (3D) that focuses on minimizing the distance tr

agentsarxiv-cs-ro
1 Jul 2026
Model Releases

TaxoMIL: Taxonomy-Constrained Learning for Hierarchical Whole Slide Image Analysis

DGX agent

arXiv:2606.31100v1 Announce Type: new Abstract: Whole slide image (WSI) analysis is central to computational pathology, with multiple instance learning (MIL) emerging as the standard pipeline for slid

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

DGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

Teaching LLMs String Matching, Backtracking, and Error Recovery to Deduce Bases and Truth Tables for the Combinatorially Exploding Bit Manipulation Puzzles

DGX agent

arXiv:2606.23672v2 Announce Type: replace Abstract: This paper presents our algorithmic innovations for the NVIDIA Nemotron Model Reasoning Challenge, focusing on Bit Manipulation Puzzles. In this tas

model-releasesarxiv-cs-ai
1 Jul 2026
Local Ai

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

DGX agent

arXiv:2606.31036v1 Announce Type: new Abstract: Specialist epilepsy expertise is scarce in resource-constrained settings, making LLM-based decision support attractive for frontline clinicians managing

local-aiarxiv-cs-lg
1 Jul 2026
Research

Teaching Models to Teach Themselves: Reasoning at the Edge of Learnability

DGX agent

arXiv:2601.18778v3 Announce Type: replace-cross Abstract: RL methods for scaling large reasoning models stall on datasets with low initial success rates, and thus little training signal. We investigat

researcharxiv-cs-cl
1 Jul 2026
Research

Team MKC at CLPsych 2026: Capturing and Characterizing Mental Health Changes through Social Media Timeline Dynamics

DGX agent

arXiv:2606.31464v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have motivated their adoption across a wide range of domains, including Artificial Intelligence (AI) f

researcharxiv-cs-ai
1 Jul 2026
Research

Technical Report of RoboSpatial Challenge at CVPR 2026: Selective Reasoning Activation and Reference-Frame Disambiguation for Embodied Spatial Reasoning

DGX agent

arXiv:2606.31645v1 Announce Type: new Abstract: Vision-language models achieve strong general perception but often struggle with the spatial reasoning required for embodied tasks. We present RoboSpati

researcharxiv-cs-cv
1 Jul 2026
Model Releases

Temperature Field Reconstruction of Tungsten Monoblock Divertor on EAST using Physics-aware Neural Operator Transformer

DGX agent

arXiv:2606.31574v1 Announce Type: cross Abstract: Accurate modeling of the divertor temperature field is essential for preventing material melting and damage and for extending the service life of fusi

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Temporal Preservation over Processing: Diagnosing and Designing Spatiotemporal Single-Stage Video Detectors

DGX agent

arXiv:2606.31421v1 Announce Type: cross Abstract: Single-stage video object detectors are increasingly deployed in time-critical applications, yet it remains unclear whether these models genuinely rea

researcharxiv-cs-ai
1 Jul 2026
Research

Temporal Training Strategies for Left Atrium and Left Atrial Appendage Segmentation in Dynamic Contrast 4DCT

DGX agent

arXiv:2606.31444v1 Announce Type: new Abstract: Dynamic contrast-enhanced cardiac CT enables time-resolved analysis of contrast filling and washout in the left atrium (LA) and left atrial appendage (L

researcharxiv-cs-cv
1 Jul 2026
Local Ai

TerraDiT-Omega: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive

DGX agent

arXiv:2606.31029v1 Announce Type: new Abstract: Generative models have achieved remarkable progress, yet applying them to satellite imagery remains challenging. Unlike natural imagery, satellite scene

local-aiarxiv-cs-cv
1 Jul 2026
Safety

Test-Time Verification for Text-to-SQL via Outcome Reward Models

DGX agent

arXiv:2606.30851v1 Announce Type: cross Abstract: Improving the reliability of large language models (LLMs) at inference time is a central challenge in structured reasoning tasks such as Text-to-SQL.

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

The Bidirectional Process Reward Model

DGX agent

arXiv:2508.01682v3 Announce Type: replace Abstract: Process Reward Models (PRMs), which assign fine-grained scores to intermediate reasoning steps within a solution trajectory, have emerged as a promi

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

The Calibration Turn in AI-Assisted Research: A Conceptual and Methodological Framework for Evidence-Licensed Claims

DGX agent

arXiv:2606.31273v1 Announce Type: new Abstract: AI-assisted research has entered a stage in which the central question is not only whether systems can generate hypotheses, run experiments, or produce

model-releasesarxiv-cs-lg
1 Jul 2026
Agents

The Consistency Dilemma in LLMs: Generator-Evaluator Agreement and Vulnerability to Mistakes

DGX agent

arXiv:2606.30653v1 Announce Type: cross Abstract: Large language models are increasingly deployed in agentic pipelines that depend on the model evaluating its own outputs without external verification

agentsarxiv-cs-ai
1 Jul 2026
← Previous
1…400401402403404…1300
Next →