AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

Hybrid-Vector Retrieval for Visually Rich Documents: Combining Single-Vector Efficiency and Multi-Vector Accuracy

Model ReleasesDGX agent

arXiv:2510.22215v2 Announce Type: replace-cross Abstract: Retrieval over visually rich documents is essential for tasks such as legal discovery, scientific search, and enterprise knowledge management.

HyKey: Hyperspectral Keypoint Detection and Matching in Minimally Invasive Surgery

ResearchDGX agent

arXiv:2604.17446v1 Announce Type: new Abstract: Purpose: 3D reconstruction in minimally invasive surgery (MIS) enables enhanced surgical guidance through improved visualisation, tool tracking, and aug

Hyperbolic Enhanced Representation Learning for Incomplete Multi-view Clustering

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.16959v1 Announce Type: cross Abstract: Incomplete Multi-View Clustering (IMVC) faces the challenge of learning discriminative representations from fragmentary observations while maintaining

Hyperspectral Unmixing Hierarchies

ResearchDGX agent

arXiv:2604.16969v1 Announce Type: new Abstract: Unmixing reveals the spatial distribution and spectral details of different constituents, called endmembers, in a hyperspectral image. Because unmixing

ICAT: Incident-Case-Grounded Adaptive Testing for Physical-Risk Prediction in Embodied World Models

Model ReleasesDGX agent

arXiv:2604.16405v1 Announce Type: cross Abstract: Video-generative world models are increasingly used as neural simulators for embodied planning and policy learning, yet their ability to predict physi

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

SafetyDGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

ICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection

ResearchDGX agent

arXiv:2604.16749v1 Announce Type: cross Abstract: Audio deepfakes pose a significant security threat, yet current state-of-the-art (SOTA) detection systems do not generalize well to realistic in-the-w

Identifying Ethical Biases in Action Recognition Models

SafetyDGX agent

arXiv:2604.17971v1 Announce Type: new Abstract: Human Action Recognition (HAR) models are increasingly deployed in high-stakes environments, yet their fairness across different human appearances has n

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

Model ReleasesDGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

iDocV2: Leveraging Self-Supervision and Open-Set Detection for Improving Pattern Spotting in Historical Documents

Model ReleasesDGX agent

arXiv:2604.16726v1 Announce Type: new Abstract: Considering the imminent massification of digital books, it has become critical to facilitate searching collections through graphical patterns. Current

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

ResearchDGX agent

arXiv:2604.17028v1 Announce Type: new Abstract: Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based crite

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

SafetyDGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

ResearchDGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

SafetyDGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification

Model ReleasesDGX agent

arXiv:2604.17010v1 Announce Type: new Abstract: We introduce a self-play framework for semantic equivalence in Haskell, utilizing formal verification to guide adversarial training between a generator

Improving Radio Interferometry Imaging by Explicitly Modeling Cross-Domain Consistency in Reconstruction

ResearchDGX agent

arXiv:2604.16794v1 Announce Type: new Abstract: Radio astronomy plays a crucial role in understanding the universe, particularly within the realm of non-thermal astrophysics. Images of celestial objec

Improving reproducibility by controlling random seed stability in machine learning based estimation via bagging

ResearchDGX agent

arXiv:2604.17694v1 Announce Type: cross Abstract: Predictions from machine learning algorithms can vary across random seeds, inducing instability in downstream debiased machine learning estimators. We

Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context

ResearchDGX agent

arXiv:2506.10779v2 Announce Type: replace Abstract: Classroom speech and lectures often contain named entities (NEs) such as names of people and special terminology. While automatic speech recognition

In-Context Learning Under Regime Change

ApplicationsDGX agent

arXiv:2604.16988v1 Announce Type: new Abstract: Non-stationary sequences arise naturally in control, forecasting, and decision-making. The data-generating process shifts at unknown times, and models m

In-Context Symbolic Regression for Robustness-Improved Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2603.15250v2 Announce Type: replace Abstract: Symbolic regression aims to replace black-box predictors with concise analytical expressions that can be inspected and validated in scientific machi

In Search of Lost DNA Sequence Pretraining

ResearchDGX agent

arXiv:2604.16570v1 Announce Type: new Abstract: DNA sequence encoding is fundamental to gene function prediction, protein synthesis, and diverse downstream biological tasks. Despite the substantial pr

In Situ Training of Implicit Neural Compressors for Scientific Simulations via Sketch-Based Regularization

ResearchDGX agent

arXiv:2511.02659v3 Announce Type: replace Abstract: Focusing on implicit neural representations, we present a novel in situ training protocol that employs limited memory buffers of full and sketched d

Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation

ResearchDGX agent

arXiv:2604.16881v1 Announce Type: new Abstract: Cross-cultural entity translation remains challenging for large language models (LLMs) as literal or phonetic renderings are usually yielded instead of

IncepDeHazeGAN: Novel Satellite Image Dehazing

ResearchDGX agent

arXiv:2604.16609v1 Announce Type: new Abstract: Dehazing is a technique in computer vision for enhancing the visual quality of images captured in cloudy or foggy conditions. Dehazing helps to recover

Incoherent Deformation, Not Capacity: Diagnosing and Mitigating Overfitting in Dynamic Gaussian Splatting

Model ReleasesDGX agent

arXiv:2604.16747v1 Announce Type: new Abstract: Dynamic 3D Gaussian Splatting methods achieve strong training-view PSNR on monocular video but generalize poorly: on the D-NeRF benchmark we measure an

IncreFA: Breaking the Static Wall of Generative Model Attribution

Model ReleasesDGX agent

arXiv:2604.17736v1 Announce Type: new Abstract: As AI generative models evolve at unprecedented speed, image attribution has become a moving target. New diffusion, adversarial and autoregressive gener

Incremental learning for audio classification with Hebbian Deep Neural Networks

TutorialsDGX agent

arXiv:2604.18270v1 Announce Type: cross Abstract: The ability of humans for lifelong learning is an inspiration for deep learning methods and in particular for continual learning. In this work, we app

Inductive Convolution Nuclear Norm Minimization for Tensor Completion with Arbitrary Sampling

ResearchDGX agent

arXiv:2604.17001v1 Announce Type: new Abstract: The recently established Convolution Nuclear Norm Minimization (CNNM) addresses the problem of extit{tensor completion with arbitrary sampling} (TCAS),

Inertia in Moral and Value Judgments of Large Language Models

SafetyDGX agent

arXiv:2408.09049v3 Announce Type: replace Abstract: Large Language Models (LLMs) behave non-deterministically, and prompting has become a common method for steering their outputs. A popular strategy i

Inference-Time Temporal Probability Smoothing for Stable Video Segmentation with SAM2 under Weak Prompts

ResearchDGX agent

arXiv:2604.17115v1 Announce Type: new Abstract: Interactive video segmentation models such as SAM2 have demonstrated strong generalization across diverse visual domains. However, under weak user super

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation

Model ReleasesDGX agent

arXiv:2510.09275v2 Announce Type: replace Abstract: Medical diagnostics is a high-stakes and complex domain that is critical to patient care. However, current evaluations of large language models (LLM

Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias

SafetyDGX agent

arXiv:2601.16934v2 Announce Type: replace Abstract: To be discoverable in an embedding-based search process, each part of a document should be reflected in its embedding representation. To quantify an

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

SafetyDGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG

Model ReleasesDGX agent

arXiv:2604.16422v1 Announce Type: new Abstract: The injection of domain-specific knowledge is crucial for adapting language models (LMs) to specialized fields such as biomedicine. While most current a

Instant Colorization of Gaussian Splats

ResearchDGX agent

arXiv:2604.17155v1 Announce Type: new Abstract: Gaussian Splatting has recently become one of the most popular frameworks for photorealistic 3D scene reconstruction and rendering. While current raster

Instinct vs. Reflection: Unifying Token and Verbalized Confidence in Multimodal Large Models

SafetyDGX agent

arXiv:2604.17274v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in various perception and reasoning tasks. Despite this success, ens

Instruction-as-State: Environment-Guided and State-Conditioned Semantic Understanding for Embodied Navigation

AgentsDGX agent

arXiv:2604.18223v1 Announce Type: new Abstract: Vision-and-Language Navigation requires agents to follow natural-language instructions in visually changing environments. A central challenge is the dyn

Integrated Wheel Sensor Communication using ESP32 -- A Contribution towards a Digital Twin of the Road System

SafetyDGX agent

arXiv:2509.04061v2 Announce Type: replace Abstract: While current onboard state estimation methods are adequate for most driving and safety-related applications, they do not provide insights into the

Integrating Feature Selection and Machine Learning for Nitrogen Assessment in Grapevine Leaves using In-Field Hyperspectral Imaging

ApplicationsDGX agent

arXiv:2507.17869v3 Announce Type: replace-cross Abstract: Nitrogen (N) is one of the most critical nutrients in winegrape production, influencing vine vigor, fruit composition, and wine quality. Becau

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval

Model ReleasesDGX agent

arXiv:2604.18051v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a challenging image retrieval paradigm that enables to retrieve target images based on multimodal queries consisting o

Inter-Agent Relative Representations for Multi-Agent Option Discovery

SafetyDGX agent

arXiv:2512.24827v3 Announce Type: replace Abstract: Temporally extended actions improve the ability to explore and plan in single-agent settings. In multi-agent settings, the exponential growth of the

Interdisciplinary Workshop on Mechanical Intelligence: Summary Report

ResearchDGX agent

arXiv:2604.16381v1 Announce Type: new Abstract: This report provides a summary of the outcomes of the Interdisciplinary Workshop on Mechanical Intelligence held in 2024. Mechanical Intelligence (MI) r

InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts

Model ReleasesDGX agent

arXiv:2509.10813v3 Announce Type: replace Abstract: The advancement of Embodied AI heavily relies on large-scale, simulatable 3D scene datasets characterized by scene diversity and realistic layouts.

Interpolating Discrete Diffusion Models with Controllable Resampling

Model ReleasesDGX agent

arXiv:2604.17310v1 Announce Type: new Abstract: Discrete diffusion models form a powerful class of generative models across diverse domains, including text and graphs. However, existing approaches fac

Introducing the O-Value: A Universal Standardization for Confusion-Matrix-Based Classification Performance Metrics

ApplicationsDGX agent

arXiv:2505.07033v2 Announce Type: replace-cross Abstract: Many classification performance metrics exist, each suited to a specific application. However, these metrics often differ in scale and can exh

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

ResearchDGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

Is Agentic RAG worth it? An experimental comparison of RAG approaches

AgentsDGX agent

arXiv:2601.07711v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are usually defined by the combination of a generator and a retrieval component that extracts textual c

Is SAM3 ready for pathology segmentation?

ResearchDGX agent

arXiv:2604.18225v1 Announce Type: new Abstract: Is Segment Anything Model 3 (SAM3) capable in segmenting Any Pathology Images? Digital pathology segmentation spans tissue-level and nuclei-level scales

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

SafetyDGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

J-PARSE: Jacobian-based Projection Algorithm for Resolving Singularities Effectively in Inverse Kinematic Control of Serial Manipulators

SafetyDGX agent

arXiv:2505.00306v5 Announce Type: replace Abstract: J-PARSE is an algorithm for smooth first-order inverse kinematic control of a serial manipulator near kinematic singularities. The commanded end-eff

Jailbreaking Large Language Models with Morality Attacks

SafetyDGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

Joint Distillation for Fast Likelihood Evaluation and Sampling in Flow-based Models

ResearchDGX agent

arXiv:2512.02636v3 Announce Type: replace-cross Abstract: Log-likelihood evaluation enables important capabilities in generative models, including model comparison, certain fine-tuning objectives, and

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

Model ReleasesDGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew

Model ReleasesDGX agent

arXiv:2604.18041v1 Announce Type: new Abstract: Despite significant advances in large language models, personalizing them for individual decision-makers remains an open problem. Here, we introduce a s

Jupiter-N Technical Report

Model ReleasesDGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

KaLDeX: Kalman Filter based Linear Deformable Cross Attention for Retina Vessel Segmentation

ResearchDGX agent

arXiv:2410.21160v2 Announce Type: replace-cross Abstract: Background and Objective: In the realm of ophthalmic imaging, accurate vascular segmentation is paramount for diagnosing and managing various

KIRA: Knowledge-Intensive Image Retrieval and Reasoning Architecture for Specialized Visual Domains

Model ReleasesDGX agent

arXiv:2604.16915v1 Announce Type: new Abstract: Retrieval augmented generation (RAG) has transformed text based question answering, yet its extension to visual domains remains hindered by fundamental

Knowing When to Quit: A Principled Framework for Dynamic Abstention in LLM Reasoning

Model ReleasesDGX agent

arXiv:2604.18419v1 Announce Type: cross Abstract: Large language models (LLMs) using chain-of-thought reasoning often waste substantial compute by producing long, incorrect responses. Abstention can m

Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact

Model ReleasesDGX agent

arXiv:2603.00883v2 Announce Type: replace Abstract: LLMs increasingly excel on AI benchmarks, but doing so does not guarantee validity for downstream tasks. This study contrasts LLM alignment on bench

L1 Regularization Paths in Linear Models by Parametric Gaussian Message Passing

ResearchDGX agent

arXiv:2604.16949v1 Announce Type: new Abstract: The paper considers the computation of L1 regularization paths in a state space setting, which includes L1 regularized Kalman smoothing, linear SVM, LAS

← Previous
1…890891892893894…998
Next →