Noise Contrastive Estimation

本文转载自查看原文 2016-10-06 19:59 3694

Notes from Notes on Noise Contrastive Estimation and Negative Sampling
one sample:

\[x_i \to [y_i^0,\cdots,y_{i}^{k}] \]

where \(y_i^0\) are true labeled words , and \(y_i^1,\cdots,y_i^{k}\) are noise samples word index, which is generated by unigram distribution \(q(w)\) of the dataset.
the probability of true data:

\[p(y_i^0=1|x_i,\theta)=\frac{\exp(y_i^0,h_\theta)}{\exp(y_i^0 h_\theta) + k*q(y_i^0)} \]

the noise sample probability:

\[p(y_i^t=0|x_i,\theta)=\frac{k*q(y_i^t)}{\exp(y_i^t h_\theta) + k*q(y_i^t)},t=1,\cdots,k \]

the cost function of this sample:

\[l_{nce}=\log p(y_i^0|x_i,\theta)+\sum_{t=1}^k{\log p(y_i^t|x_i,\theta)} \]

the overall cost function of the dataset:

\[\mathcal{L}_{nce}=\frac{1}{N}\sum_i^N{\left\{\log p(y_i^0|x_i,\theta)+\sum_{t=1}^k{\log p(y_i^t|x_i,\theta)}\right\}} \]

[Noise-Contrastive Estimation of Unnormalized Statistical Models with Applications to Natural Image Statistics]

[Word2vec Parameter Learning Explained]

[Efficient Estimation of Word Representation in Vector Space]

[Distributed Representations of Words and Phrases and their Compositionality]

[Notes on Noise Contrastive Estimation and Negative Sampling]

免责声明！

本站转载的文章为个人学习借鉴使用，本站对版权不负任何法律责任。如果侵犯了您的隐私权益，请联系本站邮箱yoyou2525@163.com删除。

猜您在找 Noise Contrastive Estimation --- 从 NCE 到 InfoNCE Chapter 9:Noise-Estimation Algorithms NCE损失(Noise-Constrastive Estimation Loss) 图像噪声水平估计——An Efficient Statistical Method for Image Noise Level Estimation contrastive loss GraphicsLab Project 之 Curl Noise Contrastive Predictive Coding(CPC) Contrastive Loss (对比损失) Fluid Motion by Curl Noise IMU Noise Model

Noise Contrastive Estimation

Related Paper

免责声明！