Hi. I was looking at create_embeddings.py to see how you derived char embeddings directly from word embeddings.
It looks like you equate a char embedding with the average of all word vectors that contain that char, counting each char multiple times if it occurs more than once in a word. Is that correct?
Did you decide to do this because you got good results or was there some other reason for this?
Thanks!
FA
Hi. I was looking at
create_embeddings.pyto see how you derived char embeddings directly from word embeddings.It looks like you equate a char embedding with the average of all word vectors that contain that char, counting each char multiple times if it occurs more than once in a word. Is that correct?
Did you decide to do this because you got good results or was there some other reason for this?
Thanks!
FA