BPEmb
Vector Database ManagementVector Databases

BPEmb

Pre-trained subword embeddings in 275 languages, based on Byte-Pair Encoding (BPE)

Open Source

About

BPEmb is a collection of pre-trained subword embeddings in 275 languages, based on Byte-Pair Encoding (BPE) and trained on Wikipedia. Its intended use is as input for neural models in natural language processing.

Open Source Health

Not enough history
Stars
1,224
Forks
100
License
MIT
Last commit
2 years ago
Python

Related Categories