Ratio divergence learning using target energy in restricted Boltzmann machines: Beyond Kullback-Leibler divergence learning

0Citations
Citations of this article
5Readers
Mendeley users who have this article in their library.
Get full text

Abstract

We propose ratio divergence (RD) learning for discrete energy-based models, a method that utilizes both training data and a tractable target energy function. We apply RD learning to restricted Boltzmann machines (RBMs), which are a minimal model that satisfies the universal approximation theorem for discrete distributions. RD learning combines the strength of both forward and reverse Kullback-Leibler divergence (KLD) learning, effectively addressing the "notorious"issues of underfitting with the forward KLD and mode collapse with the reverse KLD. Since the summation of forward and reverse KLD seems to be sufficient to combine the strength of both approaches, we include this learning method as a direct baseline in numerical experiments to evaluate its effectiveness. Numerical experiments demonstrate that RD learning outperforms other learning methods in terms of energy function fitting, mode-covering, and learning stability across various discrete energy-based models. Moreover, the performance gaps between RD learning and the other learning methods become more pronounced as the dimensions of target models increase.

Cite

CITATION STYLE

APA

Ishida, Y., Ichikawa, Y., Dote, A., Miyazawa, T., & Hukushima, K. (2025). Ratio divergence learning using target energy in restricted Boltzmann machines: Beyond Kullback-Leibler divergence learning. Physical Review E, 112(4). https://doi.org/10.1103/fxnm-y5pd

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free