Bi Att Flow
Bi-directional Attention Flow (BiDAF) network is a multi-stage hierarchical process that represents context at different levels of granularity and uses a bi-directional attention flow mechanism to achieve a query-aware context representation without early summarization.
About
The model has ~2.5M parameters. The model was trained with NVidia Titan X (Pascal Architecture, 2016). The model requires at least 12GB of GPU RAM. If your GPU RAM is smaller than 12GB, you can either decrease batch size (performance might degrade), or you can use multi GPU (see below). The training converges at ~18k steps, and it took ~4s per step (i.e. ~20 hours).
Open Source Health
- Stars
- 1,546
- Forks
- 667
- License
- Apache-2.0
- Last commit
- 3 years ago
Related Categories
Vendor
Allenai
Publisher of Dont Stop Pretraining, Objaverse Xl and Procthor
Quick Links
Open Source
More by Allenai
Related Products
Dont Stop Pretraining
Code associated with the Don't Stop Pretraining ACL 2020 paper
More from this vendor
Textrank
TextRank implementation for Python 3.
Top category match
KoBERT
Korean BERT pre-trained cased (KoBERT)
Top category match
TAADpapers
Must-read Papers on Textual Adversarial Attack and Defense
Top category match
AutoPhrase
AutoPhrase: Automated Phrase Mining from Massive Text Corpora
Top category match