Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Apr 21, 2021

Vít Novotný, Michal Štefánik, Eniafe Festus Ayetiran, Petr Sojka

Figure 1 for When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Figure 2 for When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Figure 3 for When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Figure 4 for When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Share this with someone who'll enjoy it:

Abstract:Since the seminal work of Mikolov et al. (2013a) and Bojanowski et al. (2017), word representations of shallow log-bilinear language models have found their way into many NLP applications. Mikolov et al. (2018) introduced a positional log-bilinear language model, which has characteristics of an attention-based language model and which has reached state-of-the-art performance on the intrinsic word analogy task. However, the positional model has never been evaluated on qualitative criteria or extrinsic tasks and its speed is impractical. We outline the similarities between the attention mechanism and the positional model, and we propose a constrained positional model, which adapts the sparse attention mechanism of Dai et al. (2018). We evaluate the positional and constrained positional models on three novel qualitative criteria and on the extrinsic language modeling task of Botha and Blunsom (2014). We show that the positional and constrained positional models contain interpretable information about word order and outperform the subword model of Bojanowski et al. (2017) on language modeling. We also show that the constrained positional model outperforms the positional model on language modeling and is twice as fast.

View paper on

Share this with someone who'll enjoy it:

Title:When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting

Paper and Code