AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
25 Jun 2026

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

Model ReleasesDGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

From Sparse and Imperfect 2D Anchors to Consistent 3D Gaussian Street Scenes: Support-Aware Appearance

Model ReleasesDGX agent

arXiv:2606.26007v1 Announce Type: new Abstract: Image priors can synthesize target conditions for 3D Gaussian street scenes, but independently edited views do not define a coherent 3D target. Direct f

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2502.02205v4 Announce Type: replace Abstract: The application of deep learning for partial differential equation (PDE)-constrained control is gaining increasing attention. However, existing meth

Fully Differentiable Neural Forced Alignment via Soft Dynamic Programming

SafetyDGX agent

arXiv:2606.25460v1 Announce Type: cross Abstract: Recent advances in sequence modeling have significantly improved ASR systems, bringing them close to human-level recognition accuracy and enhancing ro

FunPiQ: A New Benchmark for Pixel-Level Quality Assessment in Fundus Images

Model ReleasesDGX agent

arXiv:2606.25915v1 Announce Type: new Abstract: Color fundus photography (CFP) is the most common ophthalmic imaging modality for large-scale screening. However, it is highly susceptible to degradatio

FUTO Swipe: Layout-Agnostic Neural Swipe Decoding

ResearchDGX agent

arXiv:2606.25247v1 Announce Type: cross Abstract: Neural swipe decoders are typically tied to the keyboard they were trained on, requiring a new corpus and training run for each layout. In this report

Fuzzy Quantification over OWL Ontologies and Knowledge Graphs

ResearchDGX agent

arXiv:2606.25778v1 Announce Type: new Abstract: This paper presents a versatile framework for evaluating fuzzy quantification queries over both standard and fuzzy ontologies as well as knowledge graph

G2DP: Diffusion Planning with Spatio-Temporal Grid Guidance

SafetyDGX agent

arXiv:2606.26017v1 Announce Type: new Abstract: In autonomous driving, diffusion-based planners have emerged as a promising paradigm for robust motion planning in dense and interactive traffic, as the

Gastroendoscopy View Synthesis: A New Real Dataset and Evaluation

ResearchDGX agent

arXiv:2606.25427v1 Announce Type: new Abstract: Novel view synthesis (NVS) is an active research topic in computer vision, owing to the success of neural radiance field (NeRF) and 3D Gaussian splattin

Gaussian Mean Field Variational Inference can Overestimate Predictive Variance

Model ReleasesDGX agent

arXiv:2606.25745v1 Announce Type: cross Abstract: Mean Field Variational Inference (MFVI) is widely understood to underestimate posterior variance. By analysing conjugate Bayesian Linear Regression (B

GCT-MARL: Graph-Based Contrastive Transfer for Sample-Efficient Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.25073v1 Announce Type: new Abstract: In cooperative multi-agent reinforcement learning (MARL), from a deployment perspective, it is challenging and expensive to train agents from scratch fo

Generalised Medical Phrase Grounding

SafetyDGX agent

arXiv:2512.01085v3 Announce Type: replace-cross Abstract: Medical phrase grounding (MPG) maps textual descriptions of radiological findings to corresponding image regions. These grounded reports are e

Generating Input Distributions for Explaining Portfolio Optimization Pipelines

Model ReleasesDGX agent

arXiv:2606.25808v1 Announce Type: cross Abstract: We propose a predict-optimize-explain framework that uses gradient-based sample generation to interpret various portfolio models by identifying macroe

Generative AI for Safe and Photorealistic Drone Light Shows

SafetyDGX agent

arXiv:2606.25458v1 Announce Type: new Abstract: Drone light shows are redefining aerial entertainment, yet their widespread adoption is bottlenecked by labor-intensive, manual animation. While generat

Geo-Strat-RL: Learning Geological Event Reasoning from Verifiable Tasks

ResearchDGX agent

arXiv:2606.25000v1 Announce Type: new Abstract: To evaluate whether vision-language models can reason about geological histories, it is necessary to construct observations for which the underlying pro

Geometry-Anchored Transport Framework for Exemplar-Free Class-Incremental Learning

SafetyDGX agent

arXiv:2606.25347v1 Announce Type: cross Abstract: Exemplar-free class-incremental learning (EFCIL) requires stable decision boundaries within a shifting feature space. While maintaining class-conditio

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

Model ReleasesDGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

GeoRanker: Distance-Aware Ranking for Worldwide Image Geolocalization

ResearchDGX agent

arXiv:2505.13731v4 Announce Type: replace Abstract: Worldwide image geolocalization-the task of predicting GPS coordinates from images taken anywhere on Earth-poses a fundamental challenge due to the

Gradient-based inverse lithography for EUV masks via the waveguide method and a physics-informed neural operator

ResearchDGX agent

arXiv:2606.25753v1 Announce Type: new Abstract: Gradient-based inverse lithography technology~(ILT) for extreme ultraviolet~(EUV) masks is presented. A novel framework treats the differentiable wavegu

GRAFT: Graph-Based Affordance Transfer via Part Correspondence

ResearchDGX agent

arXiv:2606.25241v1 Announce Type: new Abstract: Generalizing robotic manipulation to unseen objects remains challenging, as learning-based approaches require many demonstrations and fail in few-shot s

Graph-Based Phonetic Error Correction of Noisy ASR

Local AiDGX agent

arXiv:2606.24889v1 Announce Type: new Abstract: Automatic speech recognition (ASR) systems, despite low overall word error rates, produce residual lexical errors that disproportionately affect semanti

Graph it first! Enabling Reasoning on Long-form Egocentric Videos through Scene Graphs

ResearchDGX agent

arXiv:2606.25842v1 Announce Type: new Abstract: Existing multi-modal large language models (MLLMs) face significant challenges in processing long video sequences due to strict input token limitations.

GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data

Model ReleasesDGX agent

arXiv:2603.14609v2 Announce Type: replace Abstract: Precise spatial understanding in Earth Observation is essential for translating raw aerial imagery into actionable insights for critical application

GROVE: Grounded Pedestrian Simulation via Natural Language for Interactive Social Robot Navigation

ResearchDGX agent

arXiv:2606.25504v1 Announce Type: new Abstract: Pedestrian simulation is a critical component for training and deploying social robot navigation approaches, yet it remains a largely rigid system that

GUI agent: Guided Exploration of User-Sensitive Screens

SafetyDGX agent

arXiv:2606.25705v1 Announce Type: new Abstract: LLM agents are increasingly being used to automate tasks for users within an open GUI environment. They inevitably encounter screens containing user-sen

H-Adapter: Pose-Robust Hairstyle Transfer via Attention-Derived, Source-Aligned Hair Masks

TutorialsDGX agent

arXiv:2606.25578v1 Announce Type: new Abstract: Hairstyle transfer has practical applications such as virtual try-on, yet remains challenging when the source and reference exhibit large head-pose disc

HEART: Coordination of Heterogeneous Expert Agents for Physically Grounded Robotic Task Planning

ApplicationsDGX agent

arXiv:2606.25404v1 Announce Type: new Abstract: Large Language Models (LLMs) can reason over complex instructions but often fail to satisfy the physical and spatial constraints required for robotic ta

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study

ApplicationsDGX agent

arXiv:2606.25973v1 Announce Type: cross Abstract: Software vulnerability remediation is a cognitively demanding task that requires specialized security expertise often lacking in general developers. I

Heterogeneous and Adept Snapshot Distillation for 3D Semantic Segmentation

ResearchDGX agent

arXiv:2606.25278v1 Announce Type: new Abstract: Multi-modal fusion and multi-model ensembling are prevalent in enhancing the performance of 3D semantic segmentation. Despite the impressive performance

Heuresis: Search Strategies for Autonomous AI Research Agents Across Quality, Diversity and Novelty

SafetyDGX agent

arXiv:2606.25198v1 Announce Type: new Abstract: Autonomous AI Research promises to accelerate the scientific progress of machine learning. To realise this goal, current Large Language Model (LLM)-base

HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Homework Assessment

Model ReleasesDGX agent

arXiv:2606.25491v1 Announce Type: new Abstract: Automated homework assessment depends not only on recognizing student answers, but also on accurately locating where each answer and each intermediate r

Hierarchical Graph Learning for Calendar Spread Strategies in Commodity Futures Markets

Model ReleasesDGX agent

arXiv:2606.25811v1 Announce Type: cross Abstract: Commodity futures can be represented hierarchically, with underlying assets at the upper level and individual futures contracts at the lower level. En

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

Model ReleasesDGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

HiFiVe: High-Fidelity Vehicle Generation Leveraging Auto-Regressive 2D Generative Priors

ApplicationsDGX agent

arXiv:2606.25300v1 Announce Type: new Abstract: Existing 3D vehicle generation methods often suffer from low geometric fidelity and blurry textures, hindering their downstream applications. While rece

HiT-JEPA: A Hierarchical Self-supervised Trajectory Embedding Framework for Similarity Computation

Local AiDGX agent

arXiv:2507.00028v2 Announce Type: replace-cross Abstract: The representation of urban trajectory data plays a critical role in effectively analyzing spatial movement patterns. Despite considerable pro

Hitting a Moving Target: Test-Time Adaptation for AI Text Detection under Continual Distribution Shift

Model ReleasesDGX agent

arXiv:2606.25152v1 Announce Type: new Abstract: Deployed approaches for AI text detection often rely on training-time access to labeled datasets of both human-written and AI-generated text. This appro

Holographic Memory for Zero-Shot Compositional Reasoning in Knowledge Graphs: A Mechanistic Study of Where and Why It Fails

ResearchDGX agent

arXiv:2606.24948v1 Announce Type: new Abstract: Knowledge graph embedding (KGE) models predict single-hop links well but have no mechanism for zero-shot compositional queries: multi-hop questions whos

Homogeneity Bias in Open-Weight LLMs Is Robust to Decoding Hyperparameters

SafetyDGX agent

arXiv:2501.02211v2 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominan

Homomorphic Encryptions for Privacy Preserving Vision

ApplicationsDGX agent

arXiv:2606.25216v1 Announce Type: cross Abstract: Legal requirements might prevent organizations from sharing sensitive data like medical or financial details of consumers which prevents them from lev

How Complexity Contributes to Learning Opacity in Machine Learning

TutorialsDGX agent

arXiv:2606.24953v1 Announce Type: new Abstract: Machine learning (ML) algorithms are known to be opaque. We do not know the reasons for their predictions. The learning process leading to the predictio

How Does the Pretraining Distribution Shape In-Context Learning? A Fundamental Trade-Off

ResearchDGX agent

arXiv:2510.01163v2 Announce Type: replace Abstract: The factors driving the performance of in-context learning (ICL) in large language models (LLMs) remain poorly understood despite ICL's surprising e

How Large Language Models Source Brand Reputation Across Languages and Markets

ResearchDGX agent

arXiv:2606.25787v1 Announce Type: cross Abstract: When a large language model (LLM) answers a question about a company, it grounds the answer in retrieved web sources, and those sources decide what th

How Modular Is a Frontier Mixture-of-Experts? A Pre-registered Causal Test in Which Apparent Expert Modularity Mostly Dissolves

ResearchDGX agent

arXiv:2606.25092v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models route each token to a few of many experts, inviting the hypothesis that experts form functional modules tied to c

How Pragmatics Shape Articulation: A Computational Case Study in STEM ASL Discourse

ApplicationsDGX agent

arXiv:2510.23842v2 Announce Type: replace Abstract: Most state-of-the-art sign language models are trained on interpreter or isolated vocabulary data, which overlooks the variability that characterize

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

Model ReleasesDGX agent

arXiv:2606.25487v1 Announce Type: new Abstract: Almost every paper on LLM jailbreaks and prompt injection reports an attack-success rate (ASR), and that number is assigned not by people but by an auto

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

Model ReleasesDGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

Model ReleasesDGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

Hybrid deep learning-based phase diversity method for wavefront reconstruction

ResearchDGX agent

arXiv:2606.25855v1 Announce Type: cross Abstract: The efficiency of high-power laser systems is limited by wavefront distortions in the beam, particularly non-common path aberrations, which reduce the

Hybrid-IR: Dual-Path Hybrid Retrieval with Iterative Reasoning for Complex Medical Question Answering

ResearchDGX agent

arXiv:2606.25338v1 Announce Type: new Abstract: Large language models (LLMs) have shown promising performance across a wide range of biomedical applications, including medical question answering (QA),

Hypergraph Normal World Models for Logical Visual Anomaly Detection

Local AiDGX agent

arXiv:2606.25368v1 Announce Type: new Abstract: Visual anomaly detection is often deployed with only normal training images. Most one-class detectors map test patches or features to a normal reference

ILV: Iterative Latent Volumes for Fast and Accurate Sparse-View CT Reconstruction

ResearchDGX agent

arXiv:2603.14915v2 Announce Type: replace Abstract: A long-term goal in CT imaging is to achieve fast and accurate 3D reconstruction from sparse-view projections, thereby reducing radiation exposure,

Improved Large Language Diffusion Models

ResearchDGX agent

arXiv:2606.25331v1 Announce Type: new Abstract: Modern large language models are predominantly trained with autoregressive factorization and causal attention. We present iLLaDA, an 8B masked diffusion

Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation

Model ReleasesDGX agent

arXiv:2411.15490v2 Announce Type: replace Abstract: Acute ischemic stroke (AIS) requires time-critical decision-making, where inaccurate interpretation of neuroimaging findings can lead to irreversibl

Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors

ResearchDGX agent

arXiv:2606.25971v1 Announce Type: new Abstract: Modern neural network training relies on optimizers such as Adam and Muon which act on each weight matrix as a single object. Yet every weight matrix ca

Improving Zero-Shot Offline RL via Behavioral Task Sampling

Model ReleasesDGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

In-context Region-based Drag: Drag Any Region to Any Shape

ResearchDGX agent

arXiv:2606.25907v1 Announce Type: new Abstract: Diffusion models have shown promise in drag-style editing. Previous works mainly focus on point-based drag, which is inherently ambiguous. This paper fo

In-Context World Modeling for Robotic Control

Model ReleasesDGX agent

arXiv:2606.26025v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages

Model ReleasesDGX agent

arXiv:2606.19157v2 Announce Type: replace-cross Abstract: AudioLLMs enable speech recognition conditioned on textual prompts such as domain descriptions or entity lists. However, it remains unclear wh

Internal Data Repetition Destroys Language Models

Model ReleasesDGX agent

arXiv:2606.24998v1 Announce Type: new Abstract: Language models are running out of high-quality training data, and even aggressively deduplicated corpora retain some amount of repetition. Earlier cont

Interpretable Concept-Guided Polynomial Tabular Kolmogorov-Arnold Network for EEG-Based Mild Cognitive Impairment Detection

TutorialsDGX agent

arXiv:2606.25434v1 Announce Type: new Abstract: Early and scalable detection of mild cognitive impairment (MCI) remains an unresolved clinical challenge. Existing EEG-based screening approaches are co

← Previous
1…360361362363364…1034
Next →