AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
Human
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
18 May 2026

Tube Loss: A Novel Approach for Prediction Interval Estimation

Model ReleasesDGX agent

arXiv:2412.06853v4 Announce Type: replace-cross Abstract: This paper proposes a novel loss function, called 'Tube Loss', for simultaneous estimation of bounds of a Prediction Interval (PI) in the regr

Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance

TutorialsDGX agent

arXiv:2605.15533v1 Announce Type: cross Abstract: Video editing poses a significant challenge. While a series of tuning-free methods circumvent the need for extensive data collection and model trainin

TVRN: Invertible Neural Networks for Compression-Aware Temporal Video Rescaling

ApplicationsDGX agent

arXiv:2605.15579v1 Announce Type: cross Abstract: To fit diverse display and bandwidth constraints, high-frame-rate videos are temporally downscaled to low-frame-rate (LFR) and later upscaled, requiri

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Two-Dimensional Quantization for Geometry-Aware Audio Coding

ResearchDGX agent

arXiv:2512.01537v2 Announce Type: replace-cross Abstract: Recent neural audio codecs have achieved impressive reconstruction quality, typically relying on quantization methods such as Residual Vector

U-SEG: Uncertainty in SEGmentation -- A systematic multi-variable exploration

ResearchDGX agent

arXiv:2605.15421v1 Announce Type: new Abstract: In this study, we explore in depth a few under-studied topics at the intersection of uncertainty estimation and segmentation. Prior work has shown that

UAM: A Dual-Stream Perspective on Forgetting in VLA Training

Model ReleasesDGX agent

arXiv:2605.15735v1 Announce Type: cross Abstract: Vision--language--action (VLA) models are typically built by fine-tuning a pretrained vision--language model (VLM) on action data. However, we show th

Uncertainty-Aware Wildfire Smoke Density Classification from Satellite Imagery via CBAM-Augmented EfficientNet with Evidential Deep Learning

ResearchDGX agent

arXiv:2605.15894v1 Announce Type: cross Abstract: Rapid and accurate wildfire smoke severity assessment from satellite images is essential for emergency response, air quality modeling, and human healt

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

SafetyDGX agent

arXiv:2605.15388v1 Announce Type: new Abstract: Stochastic estimators are fundamental to large-scale optimization, where population quantities must be inferred from noisy oracle observations. Although

Unified Simulation of Lagrangian Particle Dynamics via Transformer

ApplicationsDGX agent

arXiv:2605.15305v1 Announce Type: cross Abstract: A unified simulator that can model diverse physical phenomena without solver-specific redesign is a long-standing goal across simulation science. We p

UniShield: An Adaptive Multi-Agent Framework for Unified Forgery Image Detection and Localization

Local AiDGX agent

arXiv:2510.03161v2 Announce Type: replace-cross Abstract: With the rapid advancements in image generation, synthetic images have become increasingly realistic, posing significant societal risks, such

Universal Approximation of Nonlinear Operators and Their Derivatives

ResearchDGX agent

arXiv:2605.15285v1 Announce Type: cross Abstract: Derivative-Informed Operator Learning (DIOL), i.e. learning a (nonlinear) operator and its derivatives, is an open research frontier at the foundation

Universal Magnetic Structure Prediction from Atomic Coordinates with Near-Experimental Accuracy

ResearchDGX agent

arXiv:2605.16230v1 Announce Type: cross Abstract: Magnetic order is a fundamental property of materials, governing collective behavior and enabling a broad range of functionalities. Yet magnetic struc

'Unlimited Realm of Exploration and Experimentation': Methods and Motivations of AI-Generated Sexual Content Creators

ResearchDGX agent

arXiv:2601.21028v2 Announce Type: replace-cross Abstract: AI-generated media is radically changing the way content is both consumed and produced on the internet, and in no place is this potentially mo

Unlocking Dense Metric Depth Estimation in VLMs

Model ReleasesDGX agent

arXiv:2605.15876v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at 2D tasks such as grounding and captioning, yet remain limited in 3D understanding. A key limitation is their text

Unsupervised 3D Human Pose Estimation via Conditional Multi-view Ancestral Sampling

ResearchDGX agent

arXiv:2605.15583v1 Announce Type: new Abstract: We propose a method of estimating a 3D human pose from a single view without 3D supervision. The key to our method is to leverage the 2D diffusion prior

Unsupervised Domain Shift Detection with Interpretable Subspace Attribution

ResearchDGX agent

arXiv:2605.15920v1 Announce Type: cross Abstract: We developed a tool for detecting domain shifts, namely subtle differences in the probability distributions of datasets. We identify these shifts usin

Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents

SafetyDGX agent

arXiv:2505.11708v3 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) agents are increasingly used to simulate sophisticated cyberattacks, but their decision-making processes remain op

VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation

Local AiDGX agent

arXiv:2605.15661v1 Announce Type: cross Abstract: Classifier-free guidance (CFG) is the primary control over how strongly text semantics move a flow-based sampler, yet standard practice holds its scal

Variational Autoregressive Networks with probability priors

ResearchDGX agent

arXiv:2605.16020v1 Announce Type: new Abstract: Monte Carlo methods are essential across diverse scientific fields, yet their efficiency is frequently hampered by critical slowing down-a sharp increas

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

Model ReleasesDGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

Vector-valued self-normalized concentration inequalities beyond sub-Gaussianity

ResearchDGX agent

arXiv:2511.03606v2 Announce Type: replace-cross Abstract: The study of self-normalized processes plays a crucial role in a wide range of applications, from sequential decision-making to econometrics.

Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems

SafetyDGX agent

arXiv:2605.15228v1 Announce Type: new Abstract: Modern cloud and enterprise systems rely on identity-centric authorization, assuming that callers possessing valid credentials are safe to execute comma

Viability of perturbative expansion for quantum field theories on neurons

ResearchDGX agent

arXiv:2508.03810v4 Announce Type: replace-cross Abstract: Neural Network (NN) architectures that break statistical independence of parameters have been proposed as a new approach for simulating local

Video Models Can Reason with Verifiable Rewards

SafetyDGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

VideoGameBench: Can Vision-Language Models complete popular video games?

Model ReleasesDGX agent

arXiv:2505.18134v3 Announce Type: replace Abstract: Vision-language models (VLMs) have achieved strong results on coding and math benchmarks that are challenging for humans, yet their ability to perfo

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation

Model ReleasesDGX agent

arXiv:2605.16079v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown significant progress in video understanding, yet they face substantial challenges in tasks requiring p

VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?

Model ReleasesDGX agent

arXiv:2510.08398v4 Announce Type: replace Abstract: The recent rapid advancement of Text-to-Video (T2V) generation technologies are engaging the trained models with more world model ability, making th

ViewBridge: Curriculum Knowledge Distillation for Activity View-Invariance Under Extreme Viewpoint Changes

ResearchDGX agent

arXiv:2504.05451v2 Announce Type: replace Abstract: Traditional methods for view-invariant learning rely on controlled multi-view training data with minimal scene clutter. However, they struggle with

Visual Compositional Tuning

ResearchDGX agent

arXiv:2504.21850v3 Announce Type: replace Abstract: Visual instruction tuning (VIT) datasets have grown rapidly in scale, yet the informativeness of individual training samples has largely been overlo

VLMs Trace Without Tracking: Diagnosing Failures in Visual Path Following

Local AiDGX agent

arXiv:2605.15672v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong performance on multimodal benchmarks, but may still lack robust control over basic visual operations. We

VSPO: Vector-Steered Policy Optimization for Behavioral Control

SafetyDGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

WeatherOcc3D: VLM-Assisted Adverse Weather Aware 3D Semantic Occupancy Prediction

Model ReleasesDGX agent

arXiv:2605.16127v1 Announce Type: new Abstract: While multi-modal 3D semantic occupancy prediction typically enhances robustness by fusing camera and LiDAR inputs, its effectiveness is fundamentally c

Weight Concentration Regularization for Improving Pruning Robustness Under High Sparsity

Model ReleasesDGX agent

arXiv:2511.14282v2 Announce Type: replace-cross Abstract: Deep neural networks achieve outstanding performance across vision and language tasks, yet their large parameter counts limit deployment in re

What Is Preference Optimization Doing, and Why?

SafetyDGX agent

arXiv:2512.00778v2 Announce Type: replace Abstract: Preference optimization (PO) is indispensable for large language models (LLMs), with methods such as direct preference optimization (DPO) and proxim

When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making

ResearchDGX agent

arXiv:2602.04003v3 Announce Type: replace Abstract: Most adversarial threats in artificial intelligence (AI) target the computational behavior of models rather than the humans who rely on them. Yet mo

When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective

SafetyDGX agent

arXiv:2605.15959v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) are powerful surrogates for differential equations but are notoriously difficult to train due to spectral bia

When Does Sparse MoE Help in Vision? The Role of Backbone Compute Leverage in Sparse Routing

ResearchDGX agent

arXiv:2605.15484v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) networks promise favorable accuracy-compute trade-offs, yet practical vision deployments are hindered by expert collapse and li

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

SafetyDGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

SafetyDGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

Where to Perch in a Tree: Vision-Guidance for Tree-Grasping Drones

AgentsDGX agent

arXiv:2605.15430v1 Announce Type: cross Abstract: This study demonstrates a method to locate an ideal perch location on a tree for vision-guided autonomous tree-perching drones. Various image processi

Who Owns This Agent? Tracing AI Agents Back to Their Owners

AgentsDGX agent

arXiv:2605.16035v1 Announce Type: cross Abstract: AI agents are increasingly deployed to act autonomously in the world, yet there is still no reliable way to trace a harmful agent back to the account

Whole-body motion planning and safety-critical control for aerial manipulation

SafetyDGX agent

arXiv:2511.02342v3 Announce Type: replace Abstract: Aerial manipulation combines the maneuverability of multirotors with the dexterity of robotic arms to perform complex tasks in cluttered spaces. Yet

Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis

ResearchDGX agent

arXiv:2605.15440v1 Announce Type: new Abstract: Surprisal theory posits that the processing difficulty of a word is determined by its predictability in context, offering a potential link between human

Wind-Aware Optimal Trajectory Planning for Efficient Gliding of Fixed-Wing Aerial Systems

ApplicationsDGX agent

arXiv:2605.15619v1 Announce Type: new Abstract: Gliding offers small fixed-wing UAVs extended endurance and silent operation but requires accurate energy management, especially under wind disturbances

WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes

AgentsDGX agent

arXiv:2605.15843v1 Announce Type: new Abstract: Recent 3D world modeling systems based on generative scene synthesis, such as Marble, can create coherent and explorable 3D environments, yet their outp

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation

AgentsDGX agent

arXiv:2605.15964v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires agents to follow natural-language instructions through closed-loop perception and action in 3D enviro

X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention

AgentsDGX agent

arXiv:2605.15505v1 Announce Type: new Abstract: In enterprise operations, the context required for an AI agent task is scattered across systems of record, static information stores, and communication

XSearch: Explainable Code Search via Concept-to-Code Alignment

Model ReleasesDGX agent

arXiv:2605.16046v1 Announce Type: cross Abstract: Semantic code search has been widely adopted in both academia and industry. These approaches embed natural-language queries and code snippets into a s

Zero-Shot Goal Recognition with Large Language Models

Model ReleasesDGX agent

arXiv:2605.15333v1 Announce Type: new Abstract: Large language models have recently reached near-parity with classical planners on well-known planning domains, yet this competence relies on world-know

15 May 2026

3D Skew-Normal Splatting

Model ReleasesDGX agent

arXiv:2605.15010v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading representation for real-time novel view synthesis and been widely adopted in various downstream ap

A Benchmark for Early-stage Parkinson's Disease Detection from Speech

Model ReleasesDGX agent

arXiv:2605.14066v1 Announce Type: cross Abstract: Early-stage Parkinson's disease (EarlyPD) detection from speech is clinically meaningful yet underexplored, and published results are hard to compare

A Block Coordinate Descent Method for Nonsmooth Composite Optimization under Orthogonality Constraints

ResearchDGX agent

arXiv:2304.03641v4 Announce Type: replace-cross Abstract: Nonsmooth composite optimization with orthogonality constraints has a wide range of applications in statistical learning and data science. How

A Calculus-Based Framework for Determining Vocabulary Size in End-to-End ASR

Model ReleasesDGX agent

arXiv:2605.14427v1 Announce Type: new Abstract: In hybrid automatic speech recognition (ASR) systems, the vocabulary size is unambiguous, typically determined by the number of phones, bi-phones, or tr

A cross-species neural foundation model for end-to-end speech decoding

SafetyDGX agent

arXiv:2511.21740v5 Announce Type: replace-cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication for people with paralysis by translating neural activity into text. Most

A CUBS-Compatible Ultrasound Morphology and Uncertainty-Aware Baseline for Carotid Intima-Media Segmentation and Preliminary Risk Prediction

ResearchDGX agent

arXiv:2605.14949v1 Announce Type: new Abstract: Carotid atherosclerosis is a major contributor to ischemic stroke and transient ischemic attack. Conventional ultrasound assessment is commonly based on

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

Model ReleasesDGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

A Formative Study of Brief Affective Text as a Complement to Wearable Sensing for Longitudinal Student Health Monitoring

ResearchDGX agent

arXiv:2605.14360v1 Announce Type: cross Abstract: Wearable devices capture physiological and behavioral data with increasing fidelity, but the psychological context shaping these outcomes is difficult

A Hardware-Aware, Per-Layer Methodology for Post-Training Quantization of Large Language Models

ResearchDGX agent

arXiv:2605.14929v1 Announce Type: new Abstract: Scaled Outer Product (SOP) is a post-training quantization methodology for large language model weights, designed to deliver near-lossless fidelity at 4

A Heterogeneous Temporal Memory Governance Framework for Long-Term LLM Persona Consistency

ResearchDGX agent

arXiv:2605.14802v1 Announce Type: new Abstract: Large language models often suffer from fact loss, timeline confusion, persona drift, and reduced stability during long-range interaction, especially un

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

ResearchDGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

← Previous
1…681682683684685…1025
Next →