
Eda NLP
Data augmentation for NLP, presented at EMNLP 2019
About
We present EDA: easy data augmentation techniques for boosting performance on text classification tasks. EDA consists of four simple but powerful operations: synonym replacement, random insertion, random swap, and random deletion. On five text classification tasks, we show that EDA improves performance for both convolutional and recurrent neural networks. EDA demonstrates particularly strong results for smaller datasets; on average, across five datasets, training with EDA while using only 50% of the available training set achieved the same accuracy as normal training with all available data. We also performed extensive ablation studies and suggest parameters for practical use.
Open Source Health
- Stars
- 1,653
- Forks
- 311
- License
- Not stated
- Last commit
- 4 years ago
Related Categories
Vendor
Jason Wei
Some dude
Quick Links
Open Source
Related Products
Turbovec
A vector index built on TurboQuant, written in Rust with Python bindings
Top category match
NGT
Nearest Neighbor Search with Neighborhood Graph and Tree for High-dimensional Data
Top category match
Chromem Go
Embeddable vector database for Go with Chroma-like interface and zero third-party dependencies. In-memory with optional persistence.
Top category match
Vecs
Postgres/pgvector Python Client
Top category match
Keras-TextClassification
Macropodus擅长自然语言处理,深度学习与tensorflow,LLM,等方面的知识,Macropodus关注机器翻译,神经网络,知识图谱,语音识别,排序算法,推荐算法,tensorflow,语言模型,nlp,机器学习,人工智能,pytorch,自然语言处理,数据挖掘,分类领域.
Top category match