arrow
返回

Boosted negative sampling by quadratically constrained entropy maximization

delete2019-07-01
delete2
PRE
AI
T
Taygun Kekeç *
D
David Mimno
D
David M. J. Tax
DOI:10.1016/j.patrec.2019.04.027delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Learning probability densities for natural language representations is a difficult problem because language is inherently sparse and high-dimensional. Negative sampling is a popular and effective way to avoid intractable maximum likelihood problems, but it requires correct specification of the sampling distribution. Previous state of the art methods rely on heuristic distributions that appear to do well in practice. In this work, we define conditions for optimal sampling distributions and demonstrate how to approximate them using Quadratically Constrained Entropy Maximization (QCEM). Our analysis shows that state of the art heuristics are restrictive approximations to our proposed framework. To demonstrate the merits of our formulation, we apply QCEM to matching synthetic exponential family distributions and to finding high dimensional word embedding vectors for English. We are able to achieve faster inference on synthetic experiments and improve the correlation on semantic similarity evaluations on the Rare Words dataset by 4.8%. (C) 2019 Elsevier B.V. All rights reserved.
Keyword:
Word embeddings
Contrastive learning
Negative sampling
Entropy maximization
Semantic similarity
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
7.9K
被引数:
1.6W

机构

D
Delft University of Technology
学者数:
2.6W
论文数: 2.5W
被引数: 3.8W
C
Cornell University
学者数:
6.3W
论文数: 5.4W
被引数: 10.9W
引用论文

引用论文

Representation learning for very short texts using weighted word embedding aggregation
err2016-09-01
err125
errOAAI
errDe Boom, Cedric; Van Canneyt, Steven; Demeester, Thomas; Dhoedt, Bart
err分享
err收藏