BPE, various NORM in deep learning

#15 · open · 1 comments

View on GitHub ↗

Albert-Ma

Neural Machine Translation of Rare Words with Subword Units dropout batch normalization layer norm

Comments

Albert-Ma

subword tokenization: https://zhuanlan.zhihu.com/p/38546218