AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
1 Jul 2026

Partition-Guided Distance Saliency: Bridging Decision and Objective Spaces in Many-Objective Optimization

Local AiDGX agent

arXiv:2606.30836v1 Announce Type: new Abstract: Explainability in Many-Objective Optimization (MaO) is currently hindered by the escalating complexity of the Pareto front, which renders the relationsh

Patch-PODiff-ViT: Structured Latent Diffusion with Patchwise POD for Super-Resolution and Uncertainty Quantification

ResearchDGX agent

arXiv:2606.31290v1 Announce Type: new Abstract: Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned lat

Patient-Level Elbow Abnormality Detection: Leakage-Aware Evaluation of Learned Preprocessing, Calibration, and Triage-Oriented Operating Points

Research
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.31348v1 Announce Type: new Abstract: In this study, we examine learned preprocessing pipelines in the context of triage-oriented orthopedic abnormality detection task using elbow radiograph

Personalizing Marketplace Policies with Competing Objectives and Constrained Experiments: Evidence from a Job Marketplace

SafetyDGX agent

arXiv:2606.30932v1 Announce Type: new Abstract: Two-sided marketplaces connect distinct user groups whose interests often conflict -- improving outcomes on one side could degrade the other side's expe

PGUDA: Pressure-Guided Unsupervised Domain Adaptation with Cross-Modal Knowledge Distillation for sEMG-Based Gesture Recognition

ResearchDGX agent

arXiv:2606.31349v1 Announce Type: cross Abstract: Surface electromyography (sEMG)-based gesture recognition has emerged as a promising technology for natural human-computer interaction. However, its p

Phantom: A Unified Face-Swap Deepfake Protection Framework with Latent and Spatial Constraints

TutorialsDGX agent

arXiv:2606.31703v1 Announce Type: new Abstract: Face-swapping deepfakes pose an escalating threat to personal privacy by enabling unauthorized identity manipulation. While adversarial approaches have

Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer

ResearchDGX agent

arXiv:2511.19778v2 Announce Type: replace Abstract: Rotary positional embeddings (RoPE) are widely used in diffusion transformers (DiTs) to encode spatial relationships, yet their behavior with mixed-

{Phi}eat: Physically Grounded Material Feature Representation

ResearchDGX agent

arXiv:2511.11270v2 Announce Type: replace Abstract: While foundation models have emerged as general-purpose visual backbones, their representations are primarily optimized for semantics and lack expli

PhotoQuilt: Training-Free Arbitrary-Resolution Photomosaics via Bootstrapped Tiled Denoising

ResearchDGX agent

arXiv:2606.30968v1 Announce Type: new Abstract: Photomosaics are large images whose local regions are seen as independent tiles while their overall arrangement forms a coherent scene. Generating them

Physics-Constrained Fine-Tuning of Flow-Matching Models for Generation and Inverse Problems

Model ReleasesDGX agent

arXiv:2508.09156v3 Announce Type: replace-cross Abstract: We present a framework for fine-tuning flow-matching generative models to enforce physical constraints and solve inverse problems in scientifi

Physics-informed Conditional Normalizing Flows for Angles-only Cislunar Orbit Determination

ResearchDGX agent

arXiv:2606.30936v1 Announce Type: cross Abstract: Generative Astrodynamics is advanced in this work by extending generative modelling to an orbit determination problem in the cislunar environment. The

PiLoT v2: Pixel-to-Orthogonal Map Alignment for Free-view UAV Geo-localization

SafetyDGX agent

arXiv:2606.31098v1 Announce Type: new Abstract: Real-time, drift-free UAV geo-localization is essential for autonomous missions in GNSS-denied environments. The pioneering system, PiLoT, achieves high

Plan Right, Then Plan Tight: Symbolic RL for Efficient Embodied Reasoning

AgentsDGX agent

arXiv:2606.31260v1 Announce Type: new Abstract: Embodied task planning asks an agent to turn a natural-language instruction into an executable sequence of actions in a physical scene, and is a buildin

Planar-SfM: Camera Pose Estimation via Homography Graph Embeddings

Model ReleasesDGX agent

arXiv:2606.31979v1 Announce Type: new Abstract: Structure from Motion (SfM) systems traditionally struggle with planar scenes, where standard epipolar geometry-based methods become degenerate. Rather

PointSplat: Compact Gaussian Splatting via Human-Centric Prediction

TutorialsDGX agent

arXiv:2606.32036v1 Announce Type: new Abstract: Producing 3D human representations from input views on the fly is essential for immersive live streaming systems, where representation compactness is as

Policy Optimization Achieves Data-Dependent Regret Bounds in MDPs with Unknown Transitions

SafetyDGX agent

arXiv:2606.31769v1 Announce Type: new Abstract: We study policy optimization for online episodic tabular Markov decision processes with unknown transition kernels, aiming for best-of-both-worlds guara

PolicyGuard: From Organizational Policies to Neuro-SymbolicCompliance Review Engines

SafetyDGX agent

arXiv:2606.32004v1 Announce Type: new Abstract: Policy-grounded document review requires determining whether a target document complies with organization-specific policies, guidelines, or playbooks. W

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation

Model ReleasesDGX agent

arXiv:2606.30673v1 Announce Type: cross Abstract: Autoregressive Transformers dominate high-quality mesh generation by producing artist-worthy topologies, yet their inherent sequential decoding induce

PoseGravity: Pose Estimation from Points and Lines with Axis Prior

ResearchDGX agent

arXiv:2405.12646v3 Announce Type: replace Abstract: This paper presents a new algorithm to estimate absolute camera pose given an axis of the camera's rotation matrix. Current algorithms solve the pro

Position: Collaborative Agentic AI Needs Interoperability Across Ecosystems

AgentsDGX agent

arXiv:2505.21550v2 Announce Type: replace-cross Abstract: Collaborative agentic AI is projected to transform entire industries by enabling AI-powered agents to autonomously perceive, plan, and act wit

Position: Vision-Language-Action Models Cannot Be Verified to Perform Physical Reasoning

Model ReleasesDGX agent

arXiv:2606.30686v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) systems, built on pretrained vision-language models (VLMs), have shown rapidly improving performance on robot manipulatio

PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint Tasks

Model ReleasesDGX agent

arXiv:2606.31154v1 Announce Type: cross Abstract: Creating and editing slides is a rich, multimodal activity that is ubiquitous in professional and educational settings, making it an ideal testbed for

Practical High-Fidelity Novel-View Synthesis of Mounted Lepidoptera

ResearchDGX agent

arXiv:2606.31679v1 Announce Type: cross Abstract: Mounted butterflies are among the most striking objects in natural history collections. However, their beauty is notoriously hard to digitize in 3D: t

Predictable GRPO: A Closed-Form Model of Training Dynamics

Model ReleasesDGX agent

arXiv:2606.30789v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard tool for improving the reasoning ability of large language models, yet its training dyna

Preserve the Hard, Regenerate the Rest: Uncertainty-Guided Synthetic Training Data Augmentation with Diffusion Models

AgentsDGX agent

arXiv:2606.31603v1 Announce Type: cross Abstract: Semantic segmentation models struggle with data sparsity and rare or visually diverse regions, e.g., dense regions or small objects in aerial or auton

Pretrained Video Models as Differentiable Physics Simulators for Urban Wind Flows

Model ReleasesDGX agent

arXiv:2603.21210v3 Announce Type: replace Abstract: Designing urban spaces that provide pedestrian wind comfort and safety requires time-resolved Computational Fluid Dynamics (CFD) simulations, but th

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.31830v1 Announce Type: new Abstract: Most end-to-end autonomous driving methods rely solely on instantaneous sensor observations, limiting them to reactive behavior without the anticipatory

PrISM-IQA: Image Quality Assessment Made Practical for Smartphone Photography

Model ReleasesDGX agent

arXiv:2606.31626v1 Announce Type: new Abstract: Existing smartphone image quality assessment (IQA) methods commonly reduce perceptual quality to a single score. However, this scalar formulation is poo

PRISM: Latent Composition Consistency for Single-Image Reflection Removal

ResearchDGX agent

arXiv:2606.31513v1 Announce Type: new Abstract: Single-image reflection removal (SIRR) seeks to recover the transmission layer from a mixture corrupted by reflections -- a severely ill-posed problem.

Private Rate-Constrained Optimization with Applications to Fair Learning

Model ReleasesDGX agent

arXiv:2505.22703v2 Announce Type: replace Abstract: Many problems in trustworthy ML can be expressed as constraints on prediction rates across subpopulations, including group fairness constraints (dem

Probabilistic Inversion with Flow Matching

ResearchDGX agent

arXiv:2606.31288v1 Announce Type: new Abstract: We demonstrate the application of Flow Matching, a technique originating from generative Artificial Intelligence, to probabilistic inversion in geophysi

Probe Choice Changes Canary-Memorization Verdicts: Three Post-Hoc Disagreement Case Studies in a Text-Dominant LoRA-Tuned Autoregressive Testbed

ResearchDGX agent

arXiv:2606.31168v1 Announce Type: cross Abstract: We audit a fixed prefix-window mean-NLL memorization probe (K=20) on a Qwen2.5-VL-7B canary testbed and report three post-hoc cases where it disagrees

Probing Memorization of Tabular In-Context Learning

ApplicationsDGX agent

arXiv:2606.31208v1 Announce Type: new Abstract: Large tabular models (LTMs), i.e., tabular foundation models leveraging in-context learning (ICL), achieve state-of-the-art performance on tabular tasks

Probing Stylistic Appropriation using Large Language Models: An Evaluation Framework for Copyright Infringement under EU Law

Model ReleasesDGX agent

arXiv:2606.31250v1 Announce Type: cross Abstract: Large language models (LLM) trained on web-scale corpora generate output that may infringe copyright, yet existing technical safeguards focus narrowly

Prompting Robot Teams with Natural Language

SafetyDGX agent

arXiv:2509.24575v2 Announce Type: replace-cross Abstract: This paper presents a framework to prompt multi-robot teams with high-level tasks using natural language expressions. Our objective is to use

Proxy-GS: Unified Occlusion Priors for Training and Inference in Structured 3D Gaussian Splatting

ResearchDGX agent

arXiv:2509.24421v5 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has emerged as an efficient approach for achieving photorealistic rendering. Recent MLP-based variants further improve

PruneGround: Plug-and-play Spatial Pruning for 3D Visual Grounding

Local AiDGX agent

arXiv:2606.31148v1 Announce Type: cross Abstract: 3D Visual Grounding (3DVG) aims to localize target objects in 3D scenes given natural language descriptions. Existing approaches typically perform rea

PSCT-Net: Geometry-Aware Pediatric Skull CT Reconstruction via Differentiable Back-Projection and Attention-Guided Refinement

ResearchDGX agent

arXiv:2606.19867v2 Announce Type: replace-cross Abstract: Computed Tomography (CT) is essential for diagnosing pediatric craniofacial abnormalities, yet poses radiation risks to developing anatomies.

PSHuman: Photorealistic Single-image 3D Human Reconstruction using Cross-Scale Multiview Diffusion and Explicit Remeshing

Local AiDGX agent

arXiv:2409.10141v3 Announce Type: replace Abstract: Detailed and photorealistic 3D human modeling is essential for various applications and has seen tremendous progress. However, full-body reconstruct

Qualified Educational Capacity Planning under Heterogeneous Student Support Needs: A Synthetic Benchmark and Decision-Support Framework

Model ReleasesDGX agent

arXiv:2606.30650v1 Announce Type: cross Abstract: Educational support services often face a qualified-capacity problem: staff time is scarce, qualifications decay, new support needs can appear before

Quality-Aware Modulation for Diffusion Transformers

ResearchDGX agent

arXiv:2606.30934v1 Announce Type: new Abstract: Modern text-to-image diffusion models, such as diffusion transformers (DiT), rely on timestep or prompt embeddings to modulate the strength of the denoi

Quality-Controlled Active Learning via Gaussian Processes for Robust Structure-Property Learning in Autonomous Microscopy

AgentsDGX agent

arXiv:2603.29135v2 Announce Type: replace Abstract: Autonomous experimental systems are increasingly used in materials research to accelerate scientific discovery, but their performance is often limit

Quantum Bayesian Networks Can Speed up Reinforcement Learning in Partially Observable Environments

ResearchDGX agent

arXiv:2507.18606v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) provides a principled framework for decision-making in partially observable environments, which can be modeled as

Quantum Flow Matching

ResearchDGX agent

arXiv:2508.12413v4 Announce Type: replace-cross Abstract: The flow matching has rapidly become a dominant paradigm in classical generative modeling, offering an efficient way to interpolate between tw

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference

ResearchDGX agent

arXiv:2606.31519v1 Announce Type: cross Abstract: Long-context Large Language Model inference is severely bottlenecked by the massive Key-Value (KV) cache, yet existing sparse attention methods often

Radial Suppression Accelerates Algorithmic Generalization: A Geometric Analysis of Delayed Generalization

Model ReleasesDGX agent

arXiv:2606.32000v1 Announce Type: cross Abstract: Why do neural networks memorize algorithmic training data long before they generalize? We present a geometric case study demonstrating that, on tasks

RAISE: LLM-based Automated Heuristic Design with Robust Adversary Instance Search

ApplicationsDGX agent

arXiv:2606.31801v1 Announce Type: new Abstract: Automated Heuristic Design (AHD) with Large Language Models (LLMs) has shown remarkable progress in discovering high-quality heuristics. However, existi

Random Reshuffling Dominates Stochastic Gradient Descent

ResearchDGX agent

arXiv:2606.32005v1 Announce Type: cross Abstract: Stochastic Gradient Descent (extsf{SGD}) is one of the most classical optimization algorithms with favorable theoretical guarantees, yet the practical

RASR: Retrieval-Augmented Semantic Reasoning for Fake News Video Detection

ResearchDGX agent

arXiv:2604.06687v2 Announce Type: replace Abstract: Multimodal fake news video detection is a crucial research direction for maintaining the credibility of online information. Existing studies primari

RCL-Mamba: A Dual-domain State Space Model for Measurement-oriented Image Restoration in Rotational Sparse-View Scanning Computed Laminography

ApplicationsDGX agent

arXiv:2606.31353v1 Announce Type: new Abstract: Rotational Scanning Computed Laminography (RCL) is widely utilized for the Non-Destructive Testing(NDT) of large planar components. However, to facilita

RCT: A Robot-Collected Touch-Vision-Language Dataset for Tactile Generalization

ResearchDGX agent

arXiv:2606.31694v1 Announce Type: cross Abstract: For robots manipulating open-world objects, tactile representations must generalize to unseen materials. We introduce RCT (Robotic Contact Tactile), a

ReactionAtlas: Ab origine exploration of chemical reaction networks with machine learning

ResearchDGX agent

arXiv:2606.30778v1 Announce Type: new Abstract: Mapping a chemical reaction network, the graph of minima and transition states (TS) and the elementary reactions connecting them, is the natural languag

Real-Time Source-Free Object Detection

AgentsDGX agent

arXiv:2606.31834v1 Announce Type: cross Abstract: Real-world detectors for autonomous driving, surveillance, and robotics must handle domain-shifts under strict latency and memory constraints, yet exi

Reasoning-aware Speculative Decoding for Efficient Vision-Language-Action Models in Autonomous Driving

AgentsDGX agent

arXiv:2606.31160v1 Announce Type: new Abstract: Modern Vision-Language-Action (VLA) planners for autonomous driving emit a chain-of-causation (CoC) reasoning step before producing a trajectory. The re

Reasoning in machine vision by learning fast and slow thinking

Local AiDGX agent

arXiv:2506.22075v2 Announce Type: replace Abstract: Reasoning is a hallmark of human intelligence, enabling adaptive decision-making in complex unfamiliar scenarios. In contrast, machine intelligence

REDI: Corpus Aware Patch Ranking for DINOv3 Token Reduction

ResearchDGX agent

arXiv:2606.31676v1 Announce Type: new Abstract: Most token reduction methods for Vision Transformers seek favorable tradeoffs between accuracy and efficiency by pruning, merging, or pooling patch toke

Reference-Based Prosody and Rhythm Evaluation for Spoken Dialogue Systems

ResearchDGX agent

arXiv:2606.31055v1 Announce Type: new Abstract: Speech-to-speech (S2S) AI agents are advancing rapidly, yet evaluation lacks interpretable speech-native measures for conversational prosody and rhythm.

Reference-Free Image Quality Assessment for Virtual Try-On via Human Feedback

Model ReleasesDGX agent

arXiv:2603.13057v2 Announce Type: replace Abstract: As virtual try-on (VTON) systems become increasingly important in fashion e-commerce, there is a growing need for reliable reference-free evaluation

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation

TutorialsDGX agent

arXiv:2506.23102v2 Announce Type: replace-cross Abstract: Current CT report generation frameworks predominantly rely on global feature representations, often failing to capture region-specific details

← Previous
1…303304305306307…1025
Next →