AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
30 Jun 2026

Towards Long-Form Spatio-Temporal Video Grounding

Local AiDGX agent

arXiv:2602.23294v2 Announce Type: replace Abstract: In real scenarios, videos can span several minutes or even hours. However, existing research on spatio-temporal video grounding (STVG), given a text

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

SafetyDGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics

Model ReleasesDGX agent

arXiv:2512.13660v3 Announce Type: replace-cross Abstract: Spatial tracing, as a fundamental embodied interaction ability for robots, is inherently challenging as it requires multi-step metric-grounded

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment

ResearchDGX agent

arXiv:2606.30313v1 Announce Type: new Abstract: Longitudinal glioblastoma response assessment requires comparing subtle tumor changes across MRI time points using structured clinical criteria such as

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

ResearchDGX agent

arXiv:2606.30543v1 Announce Type: cross Abstract: With the proliferation of speech AI agents, understanding emotional entrainment in conversational interaction has become increasingly important. Emoti

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

Model ReleasesDGX agent

arXiv:2606.30560v1 Announce Type: cross Abstract: Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge r

Traffic-CBM: A Structurally Interpretable Multimodal Framework for Encrypted Traffic Classification

TutorialsDGX agent

arXiv:2606.29909v1 Announce Type: new Abstract: Encrypted traffic classification has achieved strong performance, but its decision process remains difficult to interpret. Existing methods usually comb

TrafficAlign: Aligning Large Language Models for Traffic Scenario Generation

SafetyDGX agent

arXiv:2606.29097v1 Announce Type: new Abstract: Recent research has investigated the use of large language models (LLMs) to generate traffic scenarios for autonomous driving. However, pretrained LLMs

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision

Model ReleasesDGX agent

arXiv:2606.30552v1 Announce Type: cross Abstract: Cross-embodiment transfer in vision-language-action (VLA) models remains challenging because low-level state and action spaces differ fundamentally ac

Trajectory Optimization for Collision-Aware Redundant Robotic Multi-Axis Additive Manufacturing by Constrained Gradient Projection

ApplicationsDGX agent

arXiv:2606.29766v1 Announce Type: new Abstract: Redundant robotic multi-axis additive manufacturing (MAAM) enables support-free and conformal fabrication, but trajectory optimization for long-horizon

TrajRS: Towards Certified Robustness in Pedestrian Trajectory Prediction

SafetyDGX agent

arXiv:2606.28716v1 Announce Type: new Abstract: The robustness of trajectory prediction models is crucial for developing safe autonomous driving systems. Adversarial attacks on trajectory prediction c

Transformer Architectures as Complete Bayes Processes: A Formal Proof in the Measure-Theoretic Kernel Framework

ResearchDGX agent

arXiv:2606.30440v1 Announce Type: cross Abstract: We present a complete formal proof that transformer architectures, when their internal update mechanisms satisfy a Bayes joint-distribution condition,

Transformer-Based Active Learning for Data-Efficient Vaccine Epitope Selection in PRRS

SafetyDGX agent

arXiv:2606.28659v1 Announce Type: cross Abstract: High-fidelity molecular docking simulations can produce biologically relevant estimates of epitope-receptor binding affinity but are computationally e

Transition-Aware best-of-N sampling for Longitudinal Chest X-ray Reports

ResearchDGX agent

arXiv:2606.28393v1 Announce Type: new Abstract: In longitudinal clinical practice, every chest X-ray is read in the context of the patients prior exam, and much of what the radiologist communicates is

Translating Natural Language to Strategic Temporal Specifications via LLMs

Model ReleasesDGX agent

arXiv:2606.30441v1 Announce Type: cross Abstract: A rigorous formalization of system requirements is a fundamental prerequisite for the verification of Multi-Agent Systems (MAS). However, writing corr

Translationese as a Rational Response to Translation Task Difficulty

ApplicationsDGX agent

arXiv:2603.12050v2 Announce Type: replace Abstract: Translations systematically diverge from texts originally produced in the target language, a phenomenon widely referred to as translationese. Transl

Transolver-3: Scaling Up Transformer Solvers to Industrial-Scale Geometries

HardwareDGX agent

arXiv:2602.04940v2 Announce Type: replace Abstract: Deep learning has emerged as a transformative tool for the neural surrogate modeling of partial differential equations (PDEs), known as neural PDE s

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

Model ReleasesDGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

Model ReleasesDGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.29892v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become indispensable for pushing Vision-Language-Action Models (VLAs) beyond static imitation learning. However, exist

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

Model ReleasesDGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

TUGS: Physics-based Compact Representation of Underwater Scenes by Tensorized Gaussian

ApplicationsDGX agent

arXiv:2505.08811v3 Announce Type: replace Abstract: Underwater 3D scene reconstruction is crucial for multimedia applications in adverse environments, such as underwater robotic perception and navigat

Tumor-aware augmentation with task-guided attention analysis improves rectal cancer segmentation from magnetic resonance images

Model ReleasesDGX agent

arXiv:2605.05522v2 Announce Type: replace-cross Abstract: Although self-supervised pretraining is expected to learn broadly transferable representations, its effectiveness across imaging modalities su

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

ResearchDGX agent

arXiv:2606.28548v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have become a useful tool for extracting interpretable features in language models. However, standard SAE architectures opera

Two kinds of robustness are not the same: disentangling fault tolerance and low-SNR robustness in multi-domain event detection on real data

Model ReleasesDGX agent

arXiv:2606.29339v1 Announce Type: cross Abstract: Reliable event detection underpins induced-seismicity monitoring for Carbon dioxide Capture and Storage (CCS) and geothermal operations, distributed a

Two-Stage Prompt Optimization for Few-Shot Relation Extraction: From Reasoning-Guided Search to Gradient-Guided Refinement

ResearchDGX agent

arXiv:2606.29639v1 Announce Type: cross Abstract: Automatic prompt optimization is still underexplored for episodic few-shot relation extraction with smaller language models. We propose a two-stage fr

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models

ApplicationsDGX agent

arXiv:2602.22960v2 Announce Type: replace Abstract: World models based on video generation demonstrate remarkable potential for simulating interactive environments yet suffer from persistent difficult

UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation

Local AiDGX agent

arXiv:2606.29502v1 Announce Type: new Abstract: Skill memories can improve agentic reinforcement learning by reusing past experience as textual guidance, but retrieved skills are not oracular: they ma

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention

Local AiDGX agent

arXiv:2510.16325v3 Announce Type: replace Abstract: Ultra-high-resolution text-to-image generation is increasingly vital for applications requiring fine-grained textures and global structural fidelity

Uncertainty-Aware Generation and Decision-Making Under Ambiguity

ApplicationsDGX agent

arXiv:2606.30578v1 Announce Type: new Abstract: With rapidly improving capabilities, Large Language Models (LLMs) are increasingly used in many complex real-world tasks. Beyond requiring in-depth know

Uncertainty Estimation in Pathology Foundation Models via Deep Mutual Learning

ResearchDGX agent

arXiv:2606.30020v1 Announce Type: new Abstract: Pathology foundation models (PFMs) offer generalizable representations for whole-slide image (WSI) analysis, yet their clinical adoption remains limited

Uncovering Salience-Driven Dynamics in Consumer Confidence with Generative Social Simulation

SafetyDGX agent

arXiv:2606.30395v1 Announce Type: cross Abstract: Consumer confidence is typically modeled as a persistent macroeconomic index, yet its movements arise from households that interpret economic informat

Understanding Evaluation Illusion in Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.29228v1 Announce Type: new Abstract: Despite the capability of parallel decoding, diffusion large language models (dLLMs) require many denoising steps to maintain generation quality, motiva

Understanding LLM Intervention Explanations in Multi-Party Human-Robot Interaction

ResearchDGX agent

arXiv:2606.29460v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly embedded in social robots to support natural group interactions, yet their role in complex multi-party set

UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image

AgentsDGX agent

arXiv:2606.30608v1 Announce Type: new Abstract: Articulated 3D objects are essential for interactive environments in embodied AI, robotics, and virtual reality, but reconstructing their structure and

UniCA: Bi-directional Cross-Attention with Positive Similarity Loss for Robust Multi-Modal Retrieval

Model ReleasesDGX agent

arXiv:2606.28350v1 Announce Type: cross Abstract: Multi-modal retrieval has become increasingly critical for handling the growing volume of integrated visual-textual data in real-world applications, b

Unified Complex-valued Neural Network: A Magnitude-Phase Computational Model for Event-Driven Neuromorphic Learning

Local AiDGX agent

arXiv:2606.29099v1 Announce Type: cross Abstract: Artificial neural networks (ANN) provide accurate continuous-valued representation, whereas spiking neural networks (SNN) offer event-driven temporal

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

Model ReleasesDGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

SafetyDGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation

SafetyDGX agent

arXiv:2603.22282v2 Announce Type: replace-cross Abstract: We present UniMotion, to our knowledge the first unified framework for simultaneous understanding and generation of human motion, natural lang

UniPR-3D: Towards Universal Visual Place Recognition with Visual Geometry Grounded Transformer

ResearchDGX agent

arXiv:2512.21078v3 Announce Type: replace Abstract: Visual Place Recognition (VPR) has been traditionally formulated as a single-image retrieval task. Using multiple views offers clear advantages, yet

UniTriSplat: A Unified 3D Gaussian Splatting Framework with Uniform Spherical Rasterization for Universal Cameras

ResearchDGX agent

arXiv:2606.29794v1 Announce Type: new Abstract: Existing 3D Gaussian Splatting (3DGS) frameworks rely on camera-specific rasterization, suffering from inconsistent solid-angle sampling and degraded pe

UniVAD v2: Unified Visual Anomaly Detection via Support-Conditioned Boundary Construction

ResearchDGX agent

arXiv:2606.29714v1 Announce Type: new Abstract: Unified visual anomaly detection seeks to train a single detector that can be deployed across categories, domains, and application scenarios. In the few

Universality of empirical risk minimization

ResearchDGX agent

arXiv:2202.08832v3 Announce Type: replace-cross Abstract: We study a general class of optimization problems with decision variable oldsymbol{Theta} in R^{p imes k} and cost function which is the sum o

Unlocking the Visual Record of Materials Science: A Large-Scale Multimodal Dataset from Scientific Literature

Model ReleasesDGX agent

arXiv:2606.29667v1 Announce Type: cross Abstract: The materials science literature encodes decades of experimental knowledge in figures, yet this visual record remains locked away and inaccessible to

Unveiling Novelty Evolution in the field of Library and Information Science in China

ResearchDGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

UrbanCDNet: Appearance-Robust and Boundary-Aware Bitemporal Change Detection for Korean Urban Building Monitoring

Model ReleasesDGX agent

arXiv:2606.29781v1 Announce Type: new Abstract: Urban building change detection from bi-temporal aerial imagery is important for redevelopment monitoring, infrastructure management, and unauthorized-c

Using Large Language Models as Low-Cost Statistical Estimators for Human-Response Data

SafetyDGX agent

arXiv:2606.30372v1 Announce Type: new Abstract: Quantitative research across the social and behavioral sciences depends on human subject experiments that are expensive, slow, and subject to sampling b

Value-Action Alignment in Large Language Models under Privacy-Prosocial Conflict

SafetyDGX agent

arXiv:2601.03546v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate decision-making tasks involving personal data sharing, where privacy concerns a

Variance Reduction for Stochastic Gradient Generalized Non-reversible Langevin Monte Carlo Algorithms

ResearchDGX agent

arXiv:2606.28808v1 Announce Type: cross Abstract: We study the leading-order fluctuation of stochastic gradient Euler-Maruyama estimators for generalized non-reversible Langevin dynamics. Under struct

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

Model ReleasesDGX agent

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

VCS-SLAM: Geometry-Validated Semantic Evidence Fusion for 3D Gaussian SLAM

ApplicationsDGX agent

arXiv:2606.29494v1 Announce Type: new Abstract: Visual SLAM performance often deteriorates in complex real-world applications. Semantic 3D Gaussian SLAM commonly fuses 2D semantic priors into a persis

VIB-AVSR: Variational Information Bottleneck for Noise-Robust LLM-Based Audio-Visual Speech Recognition

ResearchDGX agent

arXiv:2606.29632v1 Announce Type: cross Abstract: Audio-Visual Speech Recognition takes two input modalities, acoustic and visual streams, where visual information from lip movements aids recognition

VibES: Induced Vibration for Persistent Event-Based Sensing

ApplicationsDGX agent

arXiv:2508.19094v3 Announce Type: replace Abstract: Event cameras are a bio-inspired class of sensors that asynchronously measure per-pixel intensity changes. Under fixed illumination conditions in st

ViewSplat: View-Adaptive 3D Gaussian Splatting for Feed-Forward Synthesis

ResearchDGX agent

arXiv:2603.25265v2 Announce Type: replace Abstract: We present ViewSplat, a view-adaptive 3D Gaussian splatting network for novel view synthesis from unposed images. While recent feed-forward 3D Gauss

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

Model ReleasesDGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

ViPSim: Collaborating Visual and Parameter Spaces for Consistent Long-Horizon Embodied World Models

Model ReleasesDGX agent

arXiv:2606.28804v1 Announce Type: new Abstract: Embodied World Models (EWMs) have emerged as a scalable and risk-free paradigm for advancing embodied intelligence, enabling the safety-critical evaluat

Virtual Ring Try-On

ResearchDGX agent

arXiv:2606.28792v1 Announce Type: new Abstract: This paper presents an innovative approach that enables the users to capture their hand and try the jewel ring on their hand. The user captures the imag

Vision-driven Preference Synthesis for Mitigating Hallucinations in VLMs

SafetyDGX agent

arXiv:2606.28401v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong performance in visual understanding, yet they still suffer from hallucinations, generating content that

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

SafetyDGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

← Previous
1…335336337338339…1032
Next →