AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
15 Apr 2026

DiffusionPrint: Learning Generative Fingerprints for Diffusion-Based Inpainting Localization

Local AiDGX agent

arXiv:2604.12443v1 Announce Type: new Abstract: Modern diffusion-based inpainting models pose significant challenges for image forgery localization (IFL), as their full regeneration pipelines reconstr

DINO-Explorer: Active Underwater Discovery via Ego-Motion Compensated Semantic Predictive Coding

AgentsDGX agent

arXiv:2604.12933v1 Announce Type: cross Abstract: Marine ecosystem degradation necessitates continuous, scientifically selective underwater monitoring. However, most autonomous underwater vehicles (AU

Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection

TutorialsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12941v1 Announce Type: new Abstract: Continual face forgery detection (CFFD) requires detectors to learn emerging forgery paradigms without forgetting previously seen manipulations. Existin

Disposition Distillation at Small Scale: A Three-Arc Negative Result

Model ReleasesDGX agent

arXiv:2604.11867v1 Announce Type: cross Abstract: We set out to train behavioral dispositions (self-verification, uncertainty acknowledgment, feedback integration) into small language models (0.6B to

Distinct mechanisms underlying in-context learning in transformers

ResearchDGX agent

arXiv:2604.12151v1 Announce Type: new Abstract: Modern distributed networks, notably transformers, acquire a remarkable ability (termed `in-context learning') to adapt their computation to input stati

Distorted or Fabricated? A Survey on Hallucination in Video LLMs

ResearchDGX agent

arXiv:2604.12944v1 Announce Type: cross Abstract: Despite significant progress in video-language modeling, hallucinations remain a persistent challenge in Video Large Language Models (Vid-LLMs), refer

Do Transformers Use their Depth Adaptively? Evidence from a Relational Reasoning Task

ResearchDGX agent

arXiv:2604.12426v1 Announce Type: cross Abstract: We investigate whether transformers use their depth adaptively across tasks of increasing difficulty. Using a controlled multi-hop relational reasonin

Do VLMs Truly 'Read' Candlesticks? A Multi-Scale Benchmark for Visual Stock Price Forecasting

Model ReleasesDGX agent

arXiv:2604.12659v1 Announce Type: cross Abstract: Vision-language models(VLMs) are increasingly applied to visual stock price forecasting, yet existing benchmarks inadequately evaluate their understan

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

SafetyDGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models

SafetyDGX agent

arXiv:2511.00710v4 Announce Type: replace Abstract: Recent studies posit that Reinforcement Learning with Verifiable Rewards (RLVR) primarily amplifies behaviors inherent to the pre-training distribut

Does Visual Token Pruning Improve Calibration? An Empirical Study on Confidence in MLLMs

ResearchDGX agent

arXiv:2604.12035v1 Announce Type: new Abstract: Visual token pruning is a widely used strategy for efficient inference in multimodal large language models (MLLMs), but existing work mainly evaluates i

Domain-Specific Latent Representations Improve the Fidelity of Diffusion-Based Medical Image Super-Resolution

Local AiDGX agent

arXiv:2604.12152v1 Announce Type: cross Abstract: Latent diffusion models for medical image super-resolution universally inherit variational autoencoders designed for natural photographs. We show that

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

Model ReleasesDGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

DoseRAD2026 Challenge dataset: AI accelerated photon and proton dose calculation for radiotherapy

Model ReleasesDGX agent

arXiv:2604.12778v1 Announce Type: cross Abstract: Purpose: Accurate dose calculation is essential in radiotherapy for precise tumor irradiation while sparing healthy tissue. With the growing adoption

DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment

Model ReleasesDGX agent

arXiv:2604.12813v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown promising performance on video quality assessment (VQA) tasks. However, adapting them to new

Drawing on Memory: Dual-Trace Encoding Improves Cross-Session Recall in LLM Agents

Model ReleasesDGX agent

arXiv:2604.12948v1 Announce Type: new Abstract: LLM agents with persistent memory store information as flat factual records, providing little context for temporal reasoning, change tracking, or cross-

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction

Model ReleasesDGX agent

arXiv:2603.07083v2 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) agents operating in high-dimensional observation spaces, such as Dreamer, rely on learning abstract repres

DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos

HardwareDGX agent

arXiv:2604.12270v1 Announce Type: new Abstract: Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency,

Dress-ED: Instruction-Guided Editing for Virtual Try-On and Try-Off

Model ReleasesDGX agent

arXiv:2603.22607v2 Announce Type: replace Abstract: Recent advances in Virtual Try-On (VTON) and Virtual Try-Off (VTOFF) have greatly improved photo-realistic fashion synthesis and garment reconstruct

DRPG (Decompose, Retrieve, Plan, Generate): An Agentic Framework for Academic Rebuttal

AgentsDGX agent

arXiv:2601.18081v2 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in scientific research workflows, automated support for academic rebuttal, a crucial st

Dual-Modality Anchor-Guided Filtering for Test-time Prompt Tuning

Model ReleasesDGX agent

arXiv:2604.12403v1 Announce Type: new Abstract: Test-Time Prompt Tuning (TPT) adapts vision-language models using augmented views, but its effectiveness is hindered by the challenge of determining whi

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

SafetyDGX agent

arXiv:2509.19695v3 Announce Type: replace-cross Abstract: Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient

Dynamic Modeling and Robust Gait Optimization of a Compliant Worm Robot

ResearchDGX agent

arXiv:2604.12031v1 Announce Type: new Abstract: Worm-inspired robots provide an effective locomotion strategy for constrained environments by combining cyclic body deformation with alternating anchori

Dynamic Multi-Robot Task Allocation under Uncertainty and Communication Constraints: A Game-Theoretic Approach

SafetyDGX agent

arXiv:2604.11954v1 Announce Type: cross Abstract: We study dynamic multi-robot task allocation under uncertain task completion, time-window constraints, and incomplete information. Tasks arrive online

E2E-Fly: An Integrated Training-to-Deployment System for End-to-End Quadrotor Autonomy

SafetyDGX agent

arXiv:2604.12916v1 Announce Type: new Abstract: Training and transferring learning-based policies for quadrotors from simulation to reality remains challenging due to inefficient visual rendering, phy

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning

ResearchDGX agent

arXiv:2409.06679v3 Announce Type: replace Abstract: Processing long contexts is increasingly important for Large Language Models (LLMs) in tasks like multi-turn dialogues, code generation, and documen

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

Model ReleasesDGX agent

arXiv:2604.06063v2 Announce Type: replace Abstract: The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation

EEG-Based Multimodal Learning via Hyperbolic Mixture-of-Curvature Experts

Model ReleasesDGX agent

arXiv:2604.12579v1 Announce Type: new Abstract: Electroencephalography (EEG)-based multimodal learning integrates brain signals with complementary modalities to improve mental state assessment, provid

Efficiency of Proportional Mechanisms in Online Auto-Bidding Advertising

ResearchDGX agent

arXiv:2604.12799v1 Announce Type: cross Abstract: The rise of automated bidding strategies in online advertising presents new challenges in designing and analyzing efficient auction mechanisms. In thi

Efficient Adversarial Training via Criticality-Aware Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.12780v1 Announce Type: cross Abstract: Vision Transformer (ViT) models have achieved remarkable performance across various vision tasks, with scalability being a key advantage when applied

Efficient and Scalable Granular-ball Graph Coarsening Method for Large-scale Graph Node Classification

ResearchDGX agent

arXiv:2603.29148v2 Announce Type: replace-cross Abstract: Graph Convolutional Network (GCN) is a model that can effectively handle graph data tasks and has been successfully applied. However, for larg

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

ResearchDGX agent

arXiv:2604.05546v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) enable sophisticated reasoning over images and videos, yet their inference is hindered by a systemic efficiency

Efficient Semantic Image Communication for Traffic Monitoring at the Edge

ResearchDGX agent

arXiv:2604.12622v1 Announce Type: cross Abstract: Many visual monitoring systems operate under strict communication constraints, where transmitting full-resolution images is impractical and often unne

EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports

Model ReleasesDGX agent

arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i

EigenCoin: sassanid coins classification based on Bhattacharyya distance

ResearchDGX agent

arXiv:2604.11932v1 Announce Type: new Abstract: Solving pattern recognition problems using imbalanced databases is a hot topic, which entices researchers to bring it into focus. Therefore, we consider

El Agente Quntur: A research collaborator agent for quantum chemistry

AgentsDGX agent

arXiv:2602.04850v2 Announce Type: replace-cross Abstract: Quantum chemistry is a foundational enabling tool for the fields of chemistry, materials science, computational biology and others. Despite of

Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning

ResearchDGX agent

arXiv:2604.12483v1 Announce Type: cross Abstract: In this article, we propose the optimization of the resolution of time-frequency atoms and the regularization of fitting models to obtain better repre

ELoG-GS: Dual-Branch Gaussian Splatting with Luminance-Guided Enhancement for Extreme Low-light 3D Reconstruction

Model ReleasesDGX agent

arXiv:2604.12592v1 Announce Type: new Abstract: This paper presents our approach to the NTIRE 2026 3D Restoration and Reconstruction Challenge (Track 1), which focuses on reconstructing high-quality 3

EMBER: Autonomous Cognitive Behaviour from Learned Spiking Neural Network Dynamics in a Hybrid LLM Architecture

AgentsDGX agent

arXiv:2604.12167v1 Announce Type: new Abstract: We present (Experience-Modulated Biologically-inspired Emergent Reasoning), a hybrid cognitive architecture that reorganises the relationship between la

Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG

Model ReleasesDGX agent

arXiv:2604.12047v1 Announce Type: new Abstract: PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, table

Enabling Ultra-Fast Cardiovascular Imaging Across Heterogeneous Clinical Environments with A Generalist Foundation Model and Multimodal Database

ResearchDGX agent

arXiv:2512.21652v2 Announce Type: replace-cross Abstract: Multimodal cardiovascular magnetic resonance (CMR) imaging provides comprehensive and non-invasive insights into cardiovascular disease (CVD)

Enhance-then-Balance Modality Collaboration for Robust Multimodal Sentiment Analysis

ResearchDGX agent

arXiv:2604.12518v1 Announce Type: new Abstract: Multimodal sentiment analysis (MSA) integrates heterogeneous text, audio, and visual signals to infer human emotions. While recent approaches leverage c

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

SafetyDGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

Enhancing Clustering: An Explainable Approach via Filtered Patterns

ApplicationsDGX agent

arXiv:2604.12460v1 Announce Type: new Abstract: Machine learning has become a central research area, with increasing attention devoted to explainable clustering, also known as conceptual clustering, w

Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning

ResearchDGX agent

arXiv:2505.19261v2 Announce Type: replace-cross Abstract: Current text-to-image diffusion generation typically employs complete-text conditioning. Due to the intricate syntax, diffusion transformers (

Euler-inspired Decoupling Neural Operator for Efficient Pansharpening

SafetyDGX agent

arXiv:2604.12463v1 Announce Type: cross Abstract: Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spe

Evaluating Differential Privacy Against Membership Inference in Federated Learning: Insights from the NIST Genomics Red Team Challenge

Model ReleasesDGX agent

arXiv:2604.12737v1 Announce Type: cross Abstract: While Federated Learning (FL) mitigates direct data exposure, the resulting trained models remain susceptible to membership inference attacks (MIAs).

Evaluating Language Models for Harmful Manipulation

Local AiDGX agent

arXiv:2603.25326v4 Announce Type: replace Abstract: Interest in the concept of AI-driven harmful manipulation is growing, yet current approaches to evaluating it are limited. This paper introduces a f

Evaluating LLM-Generated ACSL Annotations for Formal Verification

Model ReleasesDGX agent

arXiv:2602.13851v3 Announce Type: replace-cross Abstract: Formal specifications are crucial for building verifiable and dependable software systems, yet generating accurate and verifiable specificatio

Evaluating Relational Reasoning in LLMs with REL

Model ReleasesDGX agent

arXiv:2604.12176v1 Announce Type: new Abstract: Relational reasoning is the ability to infer relations that jointly bind multiple entities, attributes, or variables. This ability is central to scienti

Evaluating Robustness of Large Language Models Against Multilingual Typographical Errors

ApplicationsDGX agent

arXiv:2510.09536v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in multilingual, real-world applications with user inputs -- naturally introducing typographi

Evaluating the Limitations of Protein Sequence Representations for Parkinson's Disease Classification

SafetyDGX agent

arXiv:2604.11852v1 Announce Type: cross Abstract: The identification of reliable molecular biomarkers for Parkinson's disease remains challenging due to its multifactorial nature. Although protein seq

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

Evolution-Inspired Sample Competition for Deep Neural Network Optimization

SafetyDGX agent

arXiv:2604.12568v1 Announce Type: new Abstract: Conventional deep network training generally optimizes all samples under a largely uniform learning paradigm, without explicitly modeling the heterogene

Evolution of Optimization Methods: Algorithms, Scenarios, and Evaluations

ResearchDGX agent

arXiv:2604.12968v1 Announce Type: cross Abstract: Balancing convergence speed, generalization capability, and computational efficiency remains a core challenge in deep learning optimization. First-ord

Evolving the Complete Muscle: Efficient Morphology-Control Co-design for Musculoskeletal Locomotion

ResearchDGX agent

arXiv:2604.12855v1 Announce Type: new Abstract: Musculoskeletal robots offer intrinsic compliance and flexibility, providing a promising paradigm for versatile locomotion. However, existing research t

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

SafetyDGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

Exploring Concept Subspace for Self-explainable Text-Attributed Graph Learning

ResearchDGX agent

arXiv:2604.11986v1 Announce Type: new Abstract: We introduce Graph Concept Bottleneck (GCB) as a new paradigm for self-explainable text-attributed graph learning. GCB maps graphs into a subspace, conc

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

Model ReleasesDGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

FaCT: Faithful Concept Traces for Explaining Neural Network Decisions

Model ReleasesDGX agent

arXiv:2510.25512v2 Announce Type: replace-cross Abstract: Deep networks have shown remarkable performance across a wide range of tasks, yet getting a global concept-level understanding of how they fun

← Previous
1…926927928929930…989
Next →