AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
13 Apr 2026

Characterizing Lidar Range-Measurement Ambiguity due to Multiple Returns

ResearchDGX agent

arXiv:2604.09282v1 Announce Type: cross Abstract: Reliable position and attitude sensing is critical for highly automated vehicles that operate on conventional roadways. Lidar sensors are increasingly

Constraining Sequential Model Editing with Editing Anchor Compression

Model ReleasesDGX agent

arXiv:2503.00035v2 Announce Type: replace-cross Abstract: Large language models (LLMs) struggle with hallucinations due to false or outdated knowledge. Given the high resource demands of retraining th

Deep Light Pollution Removal in Night Cityscape Photographs

ResearchDGX agent

arXiv:2604.09145v1 Announce Type: new Abstract: Nighttime photography is severely degraded by light pollution induced by pervasive artificial lighting in urban environments. After long-range scatterin

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Degradation-Robust Fusion: An Efficient Degradation-Aware Diffusion Framework for Multimodal Image Fusion in Arbitrary Degradation Scenarios

TutorialsDGX agent

arXiv:2604.08922v1 Announce Type: new Abstract: Complex degradations like noise, blur, and low resolution are typical challenges in real world image fusion tasks, limiting the performance and practica

Dejavu: Towards Experience Feedback Learning for Embodied Intelligence

SafetyDGX agent

arXiv:2510.10181v3 Announce Type: replace-cross Abstract: Embodied agents face a fundamental limitation: once deployed in real-world environments, they cannot easily acquire new knowledge to improve t

Demystifying the Silence of Correctness Bugs in PyTorch Compiler

ResearchDGX agent

arXiv:2604.08720v1 Announce Type: cross Abstract: Performance optimization of AI infrastructure is key to the fast adoption of large language models (LLMs). The PyTorch compiler (torch.compile), a cor

Detecting Diffusion-generated Images via Dynamic Assembly ForestsDetecting Diffusion-generated Images via Dynamic Assembly Forests

ResearchDGX agent

arXiv:2604.09106v1 Announce Type: new Abstract: Diffusion models are known for generating high-quality images, causing serious security concerns. To combat this, most efforts rely on deep neural netwo

Detection of Hate and Threat in Digital Forensics: A Case-Driven Multimodal Approach

ResearchDGX agent

arXiv:2604.08609v1 Announce Type: cross Abstract: Digital forensic investigations increasingly rely on heterogeneous evidence such as images, scanned documents, and contextual reports. These artifacts

Do We Really Need to Approach the Entire Pareto Front in Many-Objective Bayesian Optimisation?

Model ReleasesDGX agent

arXiv:2604.09417v1 Announce Type: new Abstract: Many-objective optimisation, a subset of multi-objective optimisation, involves optimisation problems with more than three objectives. As the number of

Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight

Model ReleasesDGX agent

arXiv:2501.14377v2 Announce Type: replace Abstract: Autonomous drone racing has risen as a challenging robotic benchmark for testing the limits of learning, perception, planning, and control. Expert h

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

ResearchDGX agent

arXiv:2412.12686v3 Announce Type: replace Abstract: Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their En

Exploring Teachers' Perspectives on Using Conversational AI Agents for Group Collaboration

SafetyDGX agent

arXiv:2602.07142v2 Announce Type: replace-cross Abstract: Collaboration is a cornerstone of 21st-century learning, yet teachers continue to face challenges in supporting productive peer interaction. E

FashionStylist: An Expert Knowledge-enhanced Multimodal Dataset for Fashion Understanding

Model ReleasesDGX agent

arXiv:2604.09249v1 Announce Type: new Abstract: Fashion understanding requires both visual perception and expert-level reasoning about style, occasion, compatibility, and outfit rationale. However, ex

Finite-Sample Analysis of Nonlinear Independent Component Analysis:Sample Complexity and Identifiability Bounds

Model ReleasesDGX agent

arXiv:2604.08850v1 Announce Type: new Abstract: Independent Component Analysis (ICA) is a fundamental unsupervised learning technique foruncovering latent structure in data by separating mixed signals

FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

Model ReleasesDGX agent

arXiv:2604.09114v1 Announce Type: new Abstract: Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-lan

Fisher-Geometric Diffusion in Stochastic Gradient Descent: Optimal Rates, Oracle Complexity, and Information-Theoretic Limits

ResearchDGX agent

arXiv:2603.02417v3 Announce Type: replace-cross Abstract: Classical stochastic-approximation analyses treat the covariance of stochastic gradients as an exogenous modeling input. We show that under ex

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

From Selection to Scheduling: Federated Geometry-Aware Correction Makes Exemplar Replay Work Better under Continual Dynamic Heterogeneity

SafetyDGX agent

arXiv:2604.08617v1 Announce Type: cross Abstract: Exemplar replay has become an effective strategy for mitigating catastrophic forgetting in federated continual learning (FCL) by retaining representat

GAN-Enhanced Deep Reinforcement Learning for Semantic-Aware Resource Allocation in 6G Network Slicing

SafetyDGX agent

arXiv:2604.08576v1 Announce Type: cross Abstract: Sixth-generation (6G) wireless networks must support heterogeneous services: enhanced Mobile Broadband (eMBB) requiring 1 Tbps data rates, massive Mac

Generative Simulation for Policy Learning in Physical Human-Robot Interaction

SafetyDGX agent

arXiv:2604.08664v1 Announce Type: new Abstract: Developing autonomous physical human-robot interaction (pHRI) systems is limited by the scarcity of large-scale training data to learn robust robot beha

GeRM: A Generative Rendering Model From Physically Realistic to Photorealistic

AgentsDGX agent

arXiv:2604.09304v1 Announce Type: new Abstract: For decades, Physically-Based Rendering (PBR) is the fundation of synthesizing photorealisitic images, and therefore sometimes roughly referred as Photo

GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback

SafetyDGX agent

arXiv:2604.08553v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong performance on text-attributed graphs (TAGs) due to their superior semantic understanding ability on te

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

How Should Video LLMs Output Time? An Analysis of Efficient Temporal Grounding Paradigms

Model ReleasesDGX agent

arXiv:2604.08966v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have advanced Video Temporal Grounding (VTG), existing methods often couple output paradigms with differe

Hypergraph Neural Networks Accelerate MUS Enumeration

AgentsDGX agent

arXiv:2604.09001v1 Announce Type: new Abstract: Enumerating Minimal Unsatisfiable Subsets (MUSes) is a fundamental task in constraint satisfaction problems (CSPs). Its major challenge is the exponenti

Imitation Learning for Combinatorial Optimisation under Uncertainty

ResearchDGX agent

arXiv:2601.05383v4 Announce Type: replace Abstract: Imitation learning (IL) provides a data-driven framework for approximating policies for large-scale combinatorial optimisation problems formulated a

Implicit Bias in Deep Linear Discriminant Analysis

SafetyDGX agent

arXiv:2603.02622v2 Announce Type: replace Abstract: While the Implicit Bias(or Implicit Regularization) of standard loss functions has been studied, the optimization geometry induced by discriminative

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation

ResearchDGX agent

arXiv:2604.08646v1 Announce Type: new Abstract: Instruction-based video editing is a natural way to control video content with text, but adapting a video generation model into an editor usually appear

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition

AgentsDGX agent

arXiv:2604.09121v1 Announce Type: cross Abstract: Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale train

Intrinsic Concept Extraction Based on Compositional Interpretability

ResearchDGX agent

arXiv:2603.11795v2 Announce Type: replace Abstract: Unsupervised Concept Extraction aims to extract concepts from a single image; however, existing methods suffer from the inability to extract composa

Investigating Multimodal Large Language Models to Support Usability Evaluation

ApplicationsDGX agent

arXiv:2508.16165v2 Announce Type: replace-cross Abstract: Usability evaluation is an essential method to support the design of effective and intuitive user interfaces (UIs). However, it commonly relie

Learning General Representation of 12-Lead Electrocardiogram with a Joint-Embedding Predictive Architecture

TutorialsDGX agent

arXiv:2410.08559v5 Announce Type: replace-cross Abstract: Electrocardiogram (ECG) captures the heart's electrical signals, offering valuable information for diagnosing cardiac conditions. However, the

M-IDoL: Information Decomposition for Modality-Specific and Diverse Representation Learning in Medical Foundation Model

TutorialsDGX agent

arXiv:2604.08936v1 Announce Type: new Abstract: Medical foundation models (MFMs) aim to learn universal representations from multimodal medical images that can generalize effectively to diverse downst

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

Model ReleasesDGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

MSMO-ABSA: Multi-Scale and Multi-Objective Optimization for Cross-Lingual Aspect-Based Sentiment Analysis

SafetyDGX agent

arXiv:2502.13718v2 Announce Type: replace Abstract: Aspect-based sentiment analysis (ABSA) garnered growing research interest in multilingual contexts in the past. However, the majority of the studies

Music Audio-Visual Question Answering Requires Specialized Multimodal Designs

ResearchDGX agent

arXiv:2505.20638v2 Announce Type: replace-cross Abstract: While recent Multimodal Large Language Models exhibit impressive capabilities for general multimodal tasks, specialized domains like music nec

MV3DIS: Multi-View Mask Matching via 3D Guides for Zero-Shot 3D Instance Segmentation

TutorialsDGX agent

arXiv:2604.08916v1 Announce Type: new Abstract: Conventional 3D instance segmentation methods rely on labor-intensive 3D annotations for supervised training, which limits their scalability and general

NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression

Model ReleasesDGX agent

arXiv:2604.08923v1 Announce Type: new Abstract: Dimensional Aspect-Based Sentiment Analysis (DimABSA) extends traditional ABSA from categorical polarity labels to continuous valence-arousal (VA) regre

Nested Radially Monotone Polar Occupancy Estimation: Clinically-Grounded Optic Disc and Cup Segmentation for Glaucoma Screening

ResearchDGX agent

arXiv:2604.09062v1 Announce Type: new Abstract: Valid segmentation of the optic disc (OD) and optic cup (OC) from fundus photographs is essential for glaucoma screening. Unfortunately, existing deep l

No Single Best Model for Diversity: Learning a Router for Sample Diversity

ResearchDGX agent

arXiv:2604.02319v2 Announce Type: replace Abstract: When posed with prompts that permit a large number of valid answers, comprehensively generating them is the first step towards satisfying a wide ran

Off-the-shelf Vision Models Benefit Image Manipulation Localization

Local AiDGX agent

arXiv:2604.09096v1 Announce Type: new Abstract: Image manipulation localization (IML) and general vision tasks are typically treated as two separate research directions due to the fundamental differen

Offline-First LLM Architecture for Adaptive Learning in Low-Connectivity Environments

Local AiDGX agent

arXiv:2603.03339v5 Announce Type: replace-cross Abstract: Artificial intelligence (AI) and large language models (LLMs) are transforming educational technology by enabling conversational tutoring, per

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

Model ReleasesDGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

On the Representational Limits of Quantum-Inspired 1024-D Document Embeddings: An Experimental Evaluation Framework

SafetyDGX agent

arXiv:2604.09430v1 Announce Type: cross Abstract: Text embeddings are central to modern information retrieval and Retrieval-Augmented Generation (RAG). While dense models derived from Large Language M

On the Terminology and Geometric Aspects of Redundant Parallel Manipulators

ResearchDGX agent

arXiv:2604.09156v1 Announce Type: new Abstract: Parallel kinematics machines (PKM) can exhibit kinematic as well as actuation redundancy. While the meaning of kinematic redundancy has been clarified a

Out-of-the-box: Black-box Causal Attacks on Object Detectors

ResearchDGX agent

arXiv:2512.03730v2 Announce Type: replace-cross Abstract: Adversarial perturbations are a useful way to expose vulnerabilities in object detectors. Existing perturbation methods are frequently white-b

Physics-Informed Reinforcement Learning of Spatial Density Velocity Potentials for Map-Free Racing

Local AiDGX agent

arXiv:2604.09499v1 Announce Type: new Abstract: Autonomous racing without prebuilt maps is a grand challenge for embedded robotics that requires kinodynamic planning from instantaneous sensor data at

Provable Post-Training Quantization: Theoretical Analysis of OPTQ and Qronos

Model ReleasesDGX agent

arXiv:2508.04853v2 Announce Type: replace-cross Abstract: Post-training quantization (PTQ) has become a crucial tool for reducing the memory and compute costs of modern deep neural networks, including

PS-TTS: Phonetic Synchronization in Text-to-Speech for Achieving Natural Automated Dubbing

ResearchDGX agent

arXiv:2604.09111v1 Announce Type: cross Abstract: Recently, artificial intelligence-based dubbing technology has advanced, enabling automated dubbing (AD) to convert the source speech of a video into

R3PM-Net: Real-time, Robust, Real-world Point Matching Network

ApplicationsDGX agent

arXiv:2604.05060v2 Announce Type: replace Abstract: Accurate Point Cloud Registration (PCR) is an important task in 3D data processing, involving the estimation of a rigid transformation between two p

RansomTrack: A Hybrid Behavioral Analysis Framework for Ransomware Detection

Model ReleasesDGX agent

arXiv:2604.08739v1 Announce Type: cross Abstract: Ransomware poses a serious and fast-acting threat to critical systems, often encrypting files within seconds of execution. Research indicates that ran

Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces

AgentsDGX agent

arXiv:2603.21692v2 Announce Type: replace Abstract: As AI agents transition from human-supervised copilots to autonomous platform infrastructure, the ability to analyze their reasoning behavior across

Retrieval Augmented Classification for Confidential Documents

Model ReleasesDGX agent

arXiv:2604.08628v1 Announce Type: cross Abstract: Unauthorized disclosure of confidential documents demands robust, low-leakage classification. In real work environments, there is a lot of inflow and

Revisiting Anisotropy in Language Transformers: The Geometry of Learning Dynamics

ResearchDGX agent

arXiv:2604.08764v1 Announce Type: new Abstract: Since their introduction, Transformer architectures have dominated Natural Language Processing (NLP). However, recent research has highlighted an inhere

Sam Altman's Molotov attack suspect listed the names of other AI CEOs and investors in a 'last warning' note, the feds said

IndustryDGX agent

A Texas man, 20-year-old Daniel Moreno-Gama, was charged with attempted murder after allegedly throwing a Molotov cocktail at OpenAI CEO Sam Altman's San Francisco home on April 10, 2026, with surveil

Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence

SafetyDGX agent

arXiv:2604.09104v1 Announce Type: cross Abstract: Scheming, the covert pursuit of misaligned goals by AI systems, represents a potentially catastrophic risk, yet scheming research suffers from signifi

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

Model ReleasesDGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

{sf TriDeliver}: Cooperative Air-Ground Instant Delivery with UAVs, Couriers, and Crowdsourced Ground Vehicles

ApplicationsDGX agent

arXiv:2604.09049v1 Announce Type: new Abstract: Instant delivery, shipping items before critical deadlines, is essential in daily life. While multiple delivery agents, such as couriers, Unmanned Aeria

SPP-SBL: Space-Power Prior Sparse Bayesian Learning for Block Sparse Recovery

Model ReleasesDGX agent

arXiv:2505.08518v2 Announce Type: replace-cross Abstract: The recovery of block-sparse signals with unknown structural patterns remains a fundamental challenge in structured sparse signal reconstructi

← Previous
1…196197198199200201
Next →