AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
Human
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
23 Jul 2026

Nonlinear Bias-Compensated Adaptive Filter and Its Application for Time-Series Prediction

SafetyDGX agent

arXiv:2607.19902v1 Announce Type: new Abstract: Most existing nonlinear adaptive filtering algorithms only account for output noise, neglecting the fact that input noise is also prevalent in practice.

Norm or Direction? Decoding Vision Mambas for High-Resolution Vision

SafetyDGX agent

arXiv:2607.18625v1 Announce Type: new Abstract: Vision Mamba models replace quadratic self-attention with linear complexity selective state space models (SSMs), emerging as efficient visual backbones.

Not All Patches are Equal: Sampling Matters for Visible-Infrared Pre-Training

Model ReleasesDGX agent

arXiv:2607.20238v1 Announce Type: new Abstract: Visible-infrared (VIS-IR) alignment is a key pre-training task for robust multi-sensor perception. Most existing methods use uniform patch-wise contrast

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Notes to Self: Can LLMs Benefit from Experiential Abstractions?

ResearchDGX agent

arXiv:2607.20372v1 Announce Type: new Abstract: Humans distill experience into reusable abstractions, e.g., strategies and cautionary reminders, and apply them to gradually solve problems more effecti

Now We Know? A Systematic Comparison of TerraMind and THOR

ResearchDGX agent

arXiv:2607.18504v1 Announce Type: cross Abstract: Benchmarks for Geospatial Foundation Models (GFMs) increasingly rank models by aggregate score, but such rankings obscure why models differ: how much

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

SafetyDGX agent

arXiv:2607.19061v2 Announce Type: replace-cross Abstract: Hateful optical illusions expose a serious gap in current multimodal safety systems. On original-view hateful illusions, previous work shows t

Nuclear Quantum Effects as a Denoising Problem

ResearchDGX agent

arXiv:2607.19680v1 Announce Type: cross Abstract: Nuclear quantum effects are rigorously captured by imaginary-time path integrals, which map the quantum Boltzmann distribution onto a ring polymer of

Occlusion-Aware Panoptic Segmentation with Joint Position Embedding and Occlusion-Level Attention

ResearchDGX agent

arXiv:2607.18112v2 Announce Type: replace Abstract: Panoptic segmentation in complex scenes remains challenging because of occlusions, yet modern approaches often neglect occlusion modelling. In this

OffNadirLoc: Benchmark and Framework for Challenging UAV-to-Satellite Geo-Localization under Large Off-Nadir Views

Model ReleasesDGX agent

arXiv:2607.19951v1 Announce Type: new Abstract: Cross-view geo-localization between UAV and satellite imagery remains a fundamental yet highly challenging task, especially under large off-nadir views

OLEDLM: A Unified Language Model for OLED Molecular Design

Model ReleasesDGX agent

arXiv:2607.20194v1 Announce Type: new Abstract: The development of organic light-emitting diode (OLED) materials faces the compounded challenges of an astronomically large chemical space, stringent qu

OmniReasoner: Thinking with Long Audio-Video via Native Tool Use

AgentsDGX agent

arXiv:2607.19339v1 Announce Type: new Abstract: Long audio-video reasoning is difficult for omnimodal LLMs because the decisive evidence is often sparse, cross-modal, and too expensive to preserve wit

On Optimization Complexity of Second-Order Certified Unlearning

ResearchDGX agent

arXiv:2607.20192v1 Announce Type: new Abstract: We study machine unlearning: the removal of memorized training data from a trained model. Specifically, we investigate the algorithmic complexity of cer

On the Computational Complexity of Structural Generalization

Model ReleasesDGX agent

arXiv:2607.19573v1 Announce Type: new Abstract: Structural generalization has been measured repeatedly by several benchmarks, yet it has never been formally defined. We give a definition that translat

On the Separability of Information in Diffusion Models

ResearchDGX agent

arXiv:2509.23937v5 Announce Type: replace-cross Abstract: Diffusion models transform noise into data by injecting information that was captured in their neural network during the training phase. In th

On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens

ResearchDGX agent

arXiv:2607.20241v1 Announce Type: cross Abstract: Culturally loaded translation poses unique challenges for machine translation (MT), as meanings are deeply embedded in socio-cultural contexts beyond

One4Many-StablePacker: An Efficient Deep Reinforcement Learning Framework for the 3D Bin Packing Problem

SafetyDGX agent

arXiv:2510.10057v2 Announce Type: replace Abstract: The three-dimensional bin packing problem (3D-BPP) is widely applied in logistics and warehousing. Existing learning-based approaches often neglect

Online Optimization of Difference-of-Convex Compositions with Smooth Mappings

ResearchDGX agent

arXiv:2607.19553v1 Announce Type: cross Abstract: We study online optimization for a broad class of structured non-convex non-smooth problems where each loss is a composition of a difference-of-convex

Online Variance Reduction for Domain Adaptation on Streaming Data

SafetyDGX agent

arXiv:2607.20374v1 Announce Type: new Abstract: This paper studies the problem of stochastic variance reduction (SVR) for the maximum mean discrepancy (MMD) and correlation alignment (CORAL) loss func

OPD-IAD: From Language Judgment to Industrial Anomaly Detection via On-Policy Self-Distillation

Local AiDGX agent

arXiv:2607.18850v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have recently shown strong potential for industrial anomaly detection (IAD) by providing image-level anomaly judgme

Open-Vocabulary Gaze Object Prediction: Benchmark and Method

Model ReleasesDGX agent

arXiv:2607.18827v1 Announce Type: new Abstract: Gaze Object Prediction (GOP) aims to localize and recognize the objects humans attend to, a task crucial for understanding human-centric interactions. H

OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks

SafetyDGX agent

arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions throu

OpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party Skills

Model ReleasesDGX agent

arXiv:2607.20121v1 Announce Type: new Abstract: LLM-based agents leverage third-party skills to extend their capabilities in open-world scenarios. However, third-party skills can introduce extra secur

OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization

SafetyDGX agent

arXiv:2607.19806v1 Announce Type: cross Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended

Optimal Placement of Docking Stations and Resident AUVs for Subsea Pipeline Inspection

AgentsDGX agent

arXiv:2607.19944v1 Announce Type: cross Abstract: A two-stage mixed-integer linear programming framework is introduced for subsea pipeline incident response planning, jointly optimizing Subsea Docking

Optimal Recalibration of an Online Predictor

HardwareDGX agent

arXiv:2607.19689v1 Announce Type: cross Abstract: We study the problem of recalibrating an online predictor [KE17, OKS24]: given an arbitrary 'hint' sequence of forecasts, the learner must output new

Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer Accelerators

Model ReleasesDGX agent

arXiv:2607.19421v1 Announce Type: cross Abstract: Silicon-photonic (SiPh) accelerators have emerged as a promising platform for Vision Transformer (ViT) inference by performing matrix multiplications

OrbitAll: A Unified Quantum Mechanical Representation Deep Learning Framework for All Molecular Systems

ResearchDGX agent

arXiv:2507.03853v2 Announce Type: replace Abstract: We introduce OrbitAll, a geometry- and physics-informed deep learning framework that encodes any molecular system with arbitrary charges, spins, and

OSVE: One Step Video Editing with One Step Diffusion Models

ResearchDGX agent

arXiv:2607.19895v1 Announce Type: cross Abstract: Text-guided video editing with diffusion models is impractically slow, hindered by costly multi-step sampling and inversion. We present OSVE, the firs

Overview of FinMMEval 2026 Task 1: Multilingual Financial Multiple-Choice Question Answering

ApplicationsDGX agent

arXiv:2607.19856v1 Announce Type: cross Abstract: FinMMEval 2026 Task 1 evaluates multilingual financial multiple-choice question answering in English, Chinese, Arabic, and Hindi. The task tests wheth

Overview of FinMMEval 2026 Task 2: Multilingual Financial Short-Answer Question Answering

ResearchDGX agent

arXiv:2607.19867v1 Announce Type: cross Abstract: FinMMEval 2026 Task 2 evaluates short-answer financial question answering over multilingual evidence. Each final-test item pairs an English question w

Pain in 3D: Generating Controllable Synthetic Faces for Automated Pain Assessment

ResearchDGX agent

arXiv:2509.16727v5 Announce Type: replace Abstract: Automated pain assessment from facial expressions is crucial for non-communicative patient. Progress has been limited by two challenges: (i) existin

PathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Pathology Image

Model ReleasesDGX agent

arXiv:2607.19261v1 Announce Type: new Abstract: Whole-slide image (WSI) diagnosis requires identifying diagnostically relevant regions, examining them across magnifications, and integrating multi-scal

Pathologist Attention-Aligned Report Generation for Prostate Histopathology

SafetyDGX agent

arXiv:2607.19624v1 Announce Type: new Abstract: The allocation of visual attention by pathologists during cancer diagnosis is a highly selective process that critically shapes the information extracte

PathReportEval: A Systematic Benchmark for Pathology Report Generation

Model ReleasesDGX agent

arXiv:2607.18448v1 Announce Type: cross Abstract: Pathology report generation from whole-slide images (WSIs) is a rapidly growing multimodal learning problem, yet progress is difficult to measure beca

PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration

SafetyDGX agent

arXiv:2607.16602v2 Announce Type: replace Abstract: Action-conditioned world models are a key component of embodied AI, serving as scalable policy evaluators that reduce reliance on expensive real-wor

PC-Seg: Progressive Cross-View Consistency for 3D OCT Segmentation from Sparse 2D Annotations

TutorialsDGX agent

arXiv:2607.17718v2 Announce Type: replace Abstract: Volumetric segmentation of optical coherence tomography (OCT) images is essential for diagnosing ocular diseases but requires labor-intensive voxel-

PercepCap: Video Captioner with Structured Spatio-Temporal Perception

ResearchDGX agent

arXiv:2607.20389v1 Announce Type: new Abstract: Video captioning requires fine-grained spatio-temporal understanding of videos, including spatial perception of where objects are located and temporal p

PerceptDrive: Perception Prior World-Action Modeling with Adaptive Expert Routing for End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2607.20175v1 Announce Type: new Abstract: Frozen perception foundation models encode rich geometric, semantic, and dynamic knowledge. Yet narrow conditioning interfaces may attenuate task-releva

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization

Model ReleasesDGX agent

arXiv:2607.19653v1 Announce Type: cross Abstract: Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature im

Persian Pixel: A large-scale synthetic OCR dataset for Persian language

ResearchDGX agent

arXiv:2607.20385v1 Announce Type: cross Abstract: Optical Character Recognition (OCR) for Persian remains substantially less mature than for Latin-script languages despite Persian being spoken by more

Personalized Recommendation Tool Learning via Autonomous Language Agents

AgentsDGX agent

arXiv:2607.19739v1 Announce Type: cross Abstract: Although large language models (LLMs) have recently gained traction in recommender systems due to their strong reasoning capabilities and extensive wo

PG-KINN: A Physics-Informed Petrov-Galerkin Kolmogorov-Arnold Network for Solving Forward and Inverse PDEs

Model ReleasesDGX agent

arXiv:2607.20378v1 Announce Type: new Abstract: Physics-informed learning of partial differential equations (PDEs) has been dominated by multilayer perceptrons (MLPs), whose spectral bias and dense pa

PGTT: Phase-Guided Terrain Traversal for Perceptive Legged Locomotion

SafetyDGX agent

arXiv:2510.18348v2 Announce Type: replace-cross Abstract: State-of-the-art perceptive Reinforcement Learning controllers for legged robots typically either (i) impose oscillator-or IK-based gait prior

PhaseAware: Interpretable Human-in-the-Loop Rehabilitation Scoring with Boundary Monitoring

AgentsDGX agent

arXiv:2607.20237v1 Announce Type: new Abstract: Rehabilitation scoring systems are most useful when their outputs can be reviewed and interpreted within clinical workflows. This study presents PhaseAw

PhenSPINE: A Standardized Benchmark for Spine Pathology Diagnosis

Model ReleasesDGX agent

arXiv:2607.19696v1 Announce Type: cross Abstract: The accurate diagnosis of spinal pathologies depends heavily on radiological interpretation, yet automated systems are hindered by the lack of diverse

Physics-Aware Complex-Valued State Space Model with Scattering-Prior Feature Modulation for PolSAR Image Classification

Model ReleasesDGX agent

arXiv:2607.19787v1 Announce Type: cross Abstract: Polarimetric synthetic aperture radar (PolSAR) image classification is a representative task for physics-aware GeoAI, where land-cover semantics are c

Physics Closure Matters for Machine Olfaction: A Maxwell--Stefan Graph Solver for Identifiable Dynamic Gas Unmixing

Local AiDGX agent

arXiv:2607.18544v1 Announce Type: new Abstract: Machine olfaction for gas unmixing is an underconstrained inverse problem in which gas compositions must be inferred from low-dimensional, delayed, and

PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling

ResearchDGX agent

arXiv:2607.20230v1 Announce Type: new Abstract: Accurate modeling of environmental systems is fundamental to scientific understanding and decision-making, yet remains challenging because observations

Pixel-Space Diffusion Transformers

ResearchDGX agent

arXiv:2607.17585v2 Announce Type: replace Abstract: Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual

Plausibility-Driven Prioritization of Candidate Biomedical Annotations

TutorialsDGX agent

arXiv:2607.20163v1 Announce Type: cross Abstract: The rapid growth of biomedical knowledge has made the validation of automatically generated biological annotations a major bottleneck in biomedical cu

PN-QNN: Harnessing Physical Noise as a Native Regularizer in Photonic Hybrid Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2607.20045v1 Announce Type: cross Abstract: Physical noise in near-term quantum hardware is usually treated as a nuisance to suppress. We ask whether it can instead act as a hardware-native regu

Point Ladder Tuning: Parameter-Efficient Hierarchical Adaptation for 3D Point Cloud Understanding

Model ReleasesDGX agent

arXiv:2607.19171v1 Announce Type: new Abstract: Fine-tuning pre-trained point-cloud backbones typically updates all parameters, resulting in substantial computation and memory overhead. More important

Point-Selection Fine-Tuning Framework for Robust Point Cloud Classification

Model ReleasesDGX agent

arXiv:2607.19711v1 Announce Type: new Abstract: Noisy and corrupted points can substantially degrade point cloud recognition performance, especially under challenging corruption settings. In particula

Pointing-Based Object Recognition

ResearchDGX agent

arXiv:2603.15403v2 Announce Type: replace Abstract: This paper presents a comprehensive pipeline for recognizing objects targeted by human pointing gestures using RGB images. As human-robot interactio

PoseIDON: 6DoF Pose Estimation with Foundation Model Features for Marine Sediment Burial Mapping

Local AiDGX agent

arXiv:2506.10386v2 Announce Type: replace Abstract: The burial state of anthropogenic objects on the seafloor provides insight into localized sedimentation dynamics and is also critical for assessing

Position: The Inevitable Transition to Machine Learning in Quantum Chemistry

ResearchDGX agent

arXiv:2607.18281v2 Announce Type: replace-cross Abstract: Finding exact solutions to the quantum many-body problem is computationally intractable (QMA-hard). Traditional approximations for electrons i

Post-Training in Time Series Foundation Models: A Unifying Framework

Model ReleasesDGX agent

arXiv:2607.20002v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for

Posterior Samplings are Missing Modalities Generators for Medical Image Translation

ApplicationsDGX agent

arXiv:2607.18763v1 Announce Type: new Abstract: Magnetic resonance imaging comes in various modality contrasts that provide complementary anatomical and pathological information. Complete multimodal a

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

AgentsDGX agent

arXiv:2607.20268v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at many tasks, they frequently struggle with complex reasoning that requires long-horizon planning and iterativ

Pre-Deployment Complexity Estimation for Federated Perception Systems

Local AiDGX agent

arXiv:2603.28282v2 Announce Type: replace-cross Abstract: Edge AI systems increasingly rely on federated learning to train perception models in distributed, privacy-preserving, and resource-constraine

← Previous
1…179180181182183…998
Next →