Tacotron-2: Location Sensitive Attention vs BahdanauMonotonic Attention

Created on 24 Feb 2019  路  6Comments  路  Source: Rayhane-mamah/Tacotron-2

Location Sensitive Attention(step 8500)
test-step-000008500-locationsensitive-attention

BahdanauMonotonic Attention(step 8500)
test-step-000008500-bahdanau-attention

These plots show that BahdanauMonotonic Attention is better.

What are the advantages of Location Sensitive Attention?

Most helpful comment

As for Chinese mandarin, the evaluation results sound good for me. Here is my T2 fork and https://github.com/Rayhane-mamah/Tacotron-2/issues/292#issuecomment-444823633 is my samples.

All 6 comments

Show us the samples please? By the way, you had better change the mel loss function into MAE and watch the alignment again.

The results made by location sensitive attention are a little worse(in my opinion).

  1. Do not the alignment plot and the results(audio quality) match?
  2. What are the advantages of Location Sensitive Attention?

As for Chinese mandarin, the evaluation results sound good for me. Here is my T2 fork and https://github.com/Rayhane-mamah/Tacotron-2/issues/292#issuecomment-444823633 is my samples.

Did the synthesized results have any misreading or word missing based on Bahdanau Monotonic Attention compared with Location Sensitive Attention?

I can not be sure because I am under training.

The following plot with Location Sensitive Attention shows missing words.
test-step-000024000-align001

As for Chinese mandarin, the evaluation results sound good for me. Here is my T2 fork and #292 (comment) is my samples.

which is better?location sensitive attention?

Was this page helpful?
0 / 5 - 0 ratings

Related issues

a8568730 picture a8568730  路  6Comments

begeekmyfriend picture begeekmyfriend  路  6Comments

zhf459 picture zhf459  路  12Comments

kwibjo picture kwibjo  路  4Comments

pandaGst picture pandaGst  路  4Comments