Thirteenth Workshop on Innovative Use of NLP for Building Educational Applications
Journal or Book Title
Proceedings of the Thirteenth Workshopon Innovative Use of NLP for Building Educational Applications
June 5, 2018
New Orleans, LA
This paper describes the collection and compilation of the OneStopEnglish corpus of texts written at three reading levels, and demonstrates its usefulness for through two applications - automatic readability assessment and automatic text simplification. The corpus consists of 189 texts, each in three versions (567 in total). The corpus is now freely available under a CC by-SA 4.0 license1 and we hope that it would foster further research on the topics of readability assessment and text simplification.
Creative Commons License
This work is licensed under a Creative Commons Attribution-Share Alike 4.0 International License.
Association for Computational Linguistics
Vajjala, Sowmya and Lucic, Ivana, "OneStopEnglish corpus: A new corpus for automatic readability assessment and text simplification" (2018). English Conference Papers, Posters and Proceedings. 10.