AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
23 Jun 2026

Beyond Damage Assessment: Recyclable Material Detection in Aerial Disaster Imagery Using a Lightweight Patch-Based Framework

ResearchDGX agent

arXiv:2606.21279v1 Announce Type: new Abstract: Nowadays, more and more disasters of different natures are appearing. Several disaster assessment approaches have been developed in order to identify da

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies

Model ReleasesDGX agent

arXiv:2606.20599v1 Announce Type: cross Abstract: Tree of Thought (ToT) search has become a promising direction for improving the reasoning capabilities of large language models, but deploying these m

Beyond Flat Labels: Level-Restricted Contrastive Learning for Hierarchical Fine-Grained Vision Classification

Research
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.21838v1 Announce Type: new Abstract: Multimodal contrastive learning has enabled zero-shot visual classification by aligning images with textual categories. However, in hierarchically struc

Beyond Importance: Interchange-Sobol Sensitivity Reveals Task-Specific Content Channels in Transformer Components

ResearchDGX agent

arXiv:2606.20678v1 Announce Type: cross Abstract: Mechanistic interpretability methods summarize a transformer component by a single importance score, conflating two distinct roles: a component may ma

Beyond 'One Language, One Script': Quantifying Orthographic Bias in Multilingual VLMs with PuMVR

Model ReleasesDGX agent

arXiv:2606.20770v1 Announce Type: cross Abstract: Current Vision-Language Models (VLMs) are celebrated for their multilingual capabilities, yet they operate under a flawed assumption: that one languag

Beyond ROC-AUC: Operating-Point Performance Reporting for Biometric Verification

ResearchDGX agent

arXiv:2606.20680v1 Announce Type: new Abstract: A biometric verifier is often deployed with a strict false match budget, so only a narrow, low false match rate (FMR) slice of the score range is used.

Beyond Templates: Revisiting Zero-Shot Remote Sensing through Meta-Prompting

TutorialsDGX agent

arXiv:2606.20702v1 Announce Type: new Abstract: Vision-language models (VLMs) have sparked growing interest in zero-shot Earth Observation (EO) downstream tasks, with further gains enabled by remote-s

Beyond the LUMIR challenge: The pathway to foundational registration models

Model ReleasesDGX agent

arXiv:2505.24160v3 Announce Type: replace-cross Abstract: Medical image challenges have played a transformative role in advancing the field, catalyzing innovation and establishing new performance benc

Beyond the Next Step: Variable-Length Latent World Models for Long-Horizon Planning

ResearchDGX agent

arXiv:2606.21775v1 Announce Type: new Abstract: Recently, world models have emerged as a promising paradigm for building intelligent agents by learning predictive models that estimate future environme

Beyond Time Series: Spatial Reasoning for Epidemic Forecasting via Multimodal Learning

ApplicationsDGX agent

arXiv:2606.22171v1 Announce Type: new Abstract: Epidemic forecasting models typically rely on surveillance data reported over administrative regions, treating them as atomic units, thereby obscuring s

B[FM]^2: Brain Foundation Model via Flow Matching with SplitUNet

SafetyDGX agent

arXiv:2606.20812v1 Announce Type: new Abstract: EEG foundation models can learn generalizable representations from large-scale EEG corpora to enable single-backbone transfer across diverse clinical an

BIFE: Better Interaction, Fewer Errors for Minute-Long Video Generation

Model ReleasesDGX agent

arXiv:2511.22973v2 Announce Type: replace Abstract: Long video generation is a critical step toward building realistic world models, requiring both high visual fidelity and long-range interaction cons

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

SafetyDGX agent

arXiv:2606.23531v1 Announce Type: new Abstract: Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular fie

Biological Sex Determination in Cadavers Using Deep Learning Algorithms from Computed Tomography Images of Pelvis and Skull

ResearchDGX agent

arXiv:2606.22515v1 Announce Type: new Abstract: Sexual identification of decomposed cadavers challenges traditional methods dependent on visual anthropological analysis. This study evaluates state-of-

BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language

ResearchDGX agent

arXiv:2606.22138v1 Announce Type: cross Abstract: We present BioMatrix, the first multimodal foundation model that natively integrates sequences, structures, and natural language for both molecules an

BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation

ResearchDGX agent

arXiv:2606.21398v1 Announce Type: new Abstract: Vision-Language Models (VLMs) for embodied navigation rely on selecting a fixed number of frames from a growing trajectory history. As episodes extend,

Black-Box Continual Learning for Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.22999v1 Announce Type: new Abstract: The rapid deployment of Vision-Language Models (VLMs) in dynamic environments necessitates the ability to learn continuously without forgetting. However

BLENDS: Bayesian Learning-Enhanced Deep Smoothing for GNSS-Denied Environments

SafetyDGX agent

arXiv:2606.22456v1 Announce Type: new Abstract: Maintaining accurate navigation during GNSS outages remains a significant challenge for autonomous systems relying on low-cost inertial sensors. While c

Boosting CVaR Policy Optimization with Quantile Gradients

SafetyDGX agent

arXiv:2601.22100v3 Announce Type: replace Abstract: Optimizing Conditional Value-at-risk (CVaR) using policy gradient (a.k.a CVaR-PG) faces significant challenges of sample inefficiency. This ineffici

Boosting Neural Video Codec via Scale-Driven Online Flow Refinement

ResearchDGX agent

arXiv:2606.23023v1 Announce Type: new Abstract: Although state-of-the-art neural video codecs (NVCs) have achieved remarkable performance, they suffer from limited generalization when encountering com

Boundary-by-Mask: Few-Shot Instance Segmentation with Mask-Conditioned Boundary Learning for Texture-Poor Industrial Parts

ResearchDGX agent

arXiv:2606.21594v1 Announce Type: new Abstract: Recent advances in large pre-trained models have led to remarkable progress in instance segmentation on general images. However, industrial scenarios re

BoxCtrl: 3D-Aware Visual Prompting for Geometric Image Editing

TutorialsDGX agent

arXiv:2606.23270v1 Announce Type: new Abstract: As instruction-based editing models and multimodal large language models advance, diverse image editing tasks have become feasible. However, achieving p

Bracing for Impact: Robust Humanoid Push Recovery and Locomotion with Reduced Order Models

ResearchDGX agent

arXiv:2505.11495v2 Announce Type: replace Abstract: Push recovery during locomotion will facilitate the deployment of humanoid robots in human-centered environments. In this paper, we present a unifie

Brain-Adapter: A Dual-Stream Vision-Language MIL Framework for Comprehensive 3D CT Diagnosis of Acute Intracranial Pathologies

ResearchDGX agent

arXiv:2606.23494v1 Announce Type: new Abstract: Automated diagnosis of 3D brain CT scans is essential for critical care, yet it remains challenging due to the heavy reliance on manual annotations and

Brain-Inspired Stochastic Joint Embedding Representation Learning

TutorialsDGX agent

arXiv:2505.11129v2 Announce Type: replace Abstract: Representation learning is one of the key research topics in machine learning, and the framework of self-supervised learning (SSL) has revolutionize

BranchShine: Compact Raw-Audio-to-IPA Transcription with a RoPE E-Branchformer Encoder

Model ReleasesDGX agent

arXiv:2606.22824v1 Announce Type: new Abstract: Speech-to-IPA transcription is useful when the desired output is pronunciation rather than orthographic text, but competitive multilingual systems are o

Breaking chains with trees: Deep learning with O(log N) parallel time complexity

Local AiDGX agent

arXiv:2606.21497v1 Announce Type: new Abstract: Modern deep neural network architectures are trained via backpropagation, which requires errors to be sequentially propagated through all layers before

Bridge the Gaps: Heterogeneous Attributed Graph Clustering via Quaternion Representation Learning

SafetyDGX agent

arXiv:2606.23199v1 Announce Type: new Abstract: Attributed graph clustering partitions nodes by jointly exploiting node attributes and graph topology. It remains challenging due to attribute heterogen

Bridging Multi-Valued Heuristics and Dimensionality Reduction in Multi-Objective Search

ResearchDGX agent

arXiv:2606.20644v1 Announce Type: cross Abstract: Multi-objective shortest-path (MOSP) algorithms traditionally rely on single-valued heuristics (SVHs), which associate each state with a single admiss

Bridging Semantics and Kinematics: A Modular Framework for Zero-Shot Robotic Manipulation

ResearchDGX agent

arXiv:2606.23157v1 Announce Type: new Abstract: This paper presents a modular training-free framework for zero-shot, language-guided robotic manipulation in semi-structured environments. The architect

Bridging Single Distortion Artifacts and Multifactorial Clinical Quality: Few-shot Biparametric MRI Quality Assessment via Distortion-trained Prototypical Networks

ApplicationsDGX agent

arXiv:2606.18872v2 Announce Type: replace Abstract: Clinical prostate multi-parametric MRI relies heavily on high-quality diffusion-weighted imaging (DWI), yet reading DWI is frequently compromised by

Build Once, Monitor Continuously: Persistent Semantic Mapping via Autonomous Exploration and Open-Vocabulary Object Updates

Local AiDGX agent

arXiv:2409.15493v4 Announce Type: replace-cross Abstract: Persistent semantic monitoring of indoor spaces such as warehouses, hospitals, and offices requires a robot to repeatedly monitor an environme

Bypassing Minimization Bias: A Shift-Invariant Variance Estimator for Off-Equilibrium Local Learning Coefficients

SafetyDGX agent

arXiv:2606.22389v1 Announce Type: new Abstract: Singular Learning Theory leverages the Local Learning Coefficient (LLC) to quantify the geometry of neural network loss landscapes. However, mean-energy

C^2GR: Coupled Comprehensive Generative Replay for a Continually Learnable Universal Segmentation Model

ResearchDGX agent

arXiv:2606.23473v1 Announce Type: new Abstract: Universal segmentation models exhibit significant potential for diverse tasks involving different imaging modalities and segmentation objectives. Task-I

Calibrated Sampling-Free Uncertainty Estimation in Bayesian Deep Learning

ResearchDGX agent

arXiv:2606.16214v3 Announce Type: replace Abstract: Modern deep learning models remain notoriously prone to overconfidence, limiting their reliability in high-stakes applications. Bayesian methods aim

Can LLMs Control Readability? A Multi-Dimensional Evaluation Framework for CEFR-Controlled Arabic Generation

SafetyDGX agent

arXiv:2606.21981v1 Announce Type: cross Abstract: While Large Language Models (LLMs) can generate fluent Arabic text, their ability to reliably control readability levels remains unclear. We propose a

Can Reasoning Models Detect Changes to their Chains of Thought?

ResearchDGX agent

arXiv:2606.22085v1 Announce Type: cross Abstract: There are many reasons one may want to edit a model's chain of thought (CoT) -- e.g., to prefill it with reasoning from a stronger model or to remove

Can Single-View Mesh Reconstruction Generalize to Robot Camera Rotation?

ResearchDGX agent

arXiv:2606.22987v1 Announce Type: new Abstract: Single-view mesh reconstruction predicts object meshes and spatial layouts from a single observation, making it attractive for fast robot spatial reason

CAOA -- Completion-Assisted Object-CAD Alignment

Model ReleasesDGX agent

arXiv:2606.18429v2 Announce Type: replace Abstract: Accurately aligning CAD models to their corresponding objects in indoor RGB-D scans is a central challenge in 3D semantic reconstruction. The task r

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales

Model ReleasesDGX agent

arXiv:2606.21949v1 Announce Type: new Abstract: Accurate and comprehensive video captions with consistent subject references are critical for downstream understanding and generation tasks. However, fe

CAT-Translate: Building Compact Open-Source Models for Japanese-English Translation

ApplicationsDGX agent

arXiv:2606.21413v1 Announce Type: cross Abstract: Nowadays, large multilingual translation models demonstrate impressive translation capabilities in the machine translation benchmarks. This raises a p

CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching

ApplicationsDGX agent

arXiv:2410.12261v5 Announce Type: replace Abstract: Anomaly detection in multivariate time series is challenging as heterogeneous subsequence anomalies may occur. Reconstruction-based methods, which f

Catching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

Model ReleasesDGX agent

arXiv:2606.22699v1 Announce Type: new Abstract: Frontier multimodal models can guess whether a person is lying from a testimony video. To do so, they stream that raw face and voice to a third-party mo

Category-Adaptive Cross-Modal Semantic Refinement and Transfer for Open-Vocabulary Multi-Label Recognition

ResearchDGX agent

arXiv:2412.06190v2 Announce Type: replace Abstract: Benefiting from the generalization capability of CLIP, recent vision language pre-training (VLP) models have demonstrated the ability to capture a w

Causal Discovery in the Era of Agents

AgentsDGX agent

arXiv:2606.23608v1 Announce Type: cross Abstract: Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or in

Causal Gaussian Processes for Robust Treatment Effect Evaluation with Unobserved Confounding

SafetyDGX agent

arXiv:2606.21809v1 Announce Type: new Abstract: The presence of confounding bias poses a key challenge in policy evaluation, as the target causal effects of actions are not identifiable (i.e., underde

Causal Reward World Models: Zero-shot Reward Design for Automated Skill Generation

ResearchDGX agent

arXiv:2606.23280v1 Announce Type: new Abstract: Automated Reward Design (ARD) aims to replace manual reward engineering in reinforcement learning with language-driven reward function synthesis. Howeve

Causal Variational Deep Embedding: A Family of Interventional Generators for Confounded Images

ResearchDGX agent

arXiv:2606.21806v1 Announce Type: new Abstract: Deep generative models reproduce the observational distribution of their training data, inheriting any spurious associations it contains. A common sourc

Causally Fair Node Classification on Non-IID Graph Data

SafetyDGX agent

arXiv:2505.01652v2 Announce Type: replace Abstract: Fair machine learning seeks to identify and mitigate biases in predictions against unfavorable populations characterized by demographic attributes,

CDER-SME: A Cross-Device Event-RGB Micro-Expression Dataset under Multi-Level Stress Induction

Model ReleasesDGX agent

arXiv:2606.20715v1 Announce Type: new Abstract: Micro-expression recognition (MER) in realistic scenarios demands high temporal sensitivity and ecological validity, yet existing benchmarks are largely

CELEUS: Certifiable and Efficient LLM Evaluation via E-Processes

ApplicationsDGX agent

arXiv:2606.20820v1 Announce Type: new Abstract: Can we trust evaluation scores to capture an LLM's true real-world performance? Certifiable evaluation answers this question by providing guarantee for

Central limit theorem for the averaged Adam optimizer

ResearchDGX agent

arXiv:2606.21433v1 Announce Type: cross Abstract: In this article, we analyse convergence of the averaged Adam optimizer to an attracting zero of the Adam vector field. We provide a central limit theo

Certified World Models: Predictability Across Configuration, Horizon, and Resolution

ResearchDGX agent

arXiv:2606.13092v2 Announce Type: replace Abstract: Scale buys interpolation; structure buys certifiable transfer. A world model's average error does not say whether a particular rollout can be truste

CFAgentBench: A Reproducible Environment and Benchmark for Autonomous Construction-Finance Agents

Model ReleasesDGX agent

arXiv:2606.22000v1 Announce Type: cross Abstract: We introduce CFAgentBench, a reproducible, self-hostable environment and benchmark for autonomous construction-finance agents: a CFO/controller-class

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL

SafetyDGX agent

arXiv:2602.03389v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning remains challenging for long-horizon tasks. While hierarchical approaches mitigate this issue by dec

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

Model ReleasesDGX agent

arXiv:2606.22862v1 Announce Type: new Abstract: Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe

Changing Modalities: Adapting Remote Sensing Models to New Satellites and Sensors

ApplicationsDGX agent

arXiv:2606.23356v1 Announce Type: new Abstract: Machine learning models for remote sensing are trained and deployed on a static set of modalities. However, as we equip newer satellites with novel sens

Channel Location Constrains the Auditability of Subliminal Learning

Local AiDGX agent

arXiv:2606.22019v1 Announce Type: new Abstract: Subliminal learning lets a student inherit a teacher's hidden trait from distillation data that never names it. We ask when such transfer can be audited

Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition

Model ReleasesDGX agent

arXiv:2606.21657v1 Announce Type: new Abstract: Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basi

← Previous
1…385386387388389…1040
Next →