| Title | Large Scale Arabic Error Annotation: Guidelines and Framework | 
  
  | Authors | Wajdi Zaghouani, Behrang Mohit, Nizar Habash, Ossama Obeid, Nadi Tomeh, Alla Rozovskaya, Noura Farra, Sarah Alkuhlani and Kemal Oflazer | 
  
  | Abstract | We present annotation guidelines and a web-based annotation framework developed as part of an effort to create a manually annotated Arabic corpus of errors and corrections for various text types. Such a corpus will be invaluable for developing Arabic error correction tools, both for training models and as a gold standard for evaluating error correction algorithms. We summarize the guidelines we created. We also describe issues encountered during the training of the annotators, as well as problems that are specific to the Arabic language that arose during the annotation process. Finally, we present the annotation tool that was developed as part of this project, the annotation pipeline, and the quality of the resulting annotations. | 
  
  | Topics | Grammar and Syntax, Tools, Systems, Applications | 
  
  | Full paper  | Large Scale Arabic Error Annotation: Guidelines and Framework | 
  
  | Bibtex | @InProceedings{ZAGHOUANI14.956, author =  {Wajdi Zaghouani and Behrang Mohit and Nizar Habash and Ossama Obeid and Nadi Tomeh and Alla Rozovskaya and Noura Farra and Sarah Alkuhlani and Kemal Oflazer},
 title =  {Large Scale Arabic Error Annotation: Guidelines and Framework},
 booktitle =  {Proceedings of the Ninth International Conference on Language Resources and Evaluation (LREC'14)},
 year =  {2014},
 month =  {may},
 date =  {26-31},
 address =  {Reykjavik, Iceland},
 editor =  {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Thierry Declerck and Hrafn Loftsson and Bente Maegaard and Joseph Mariani and Asuncion Moreno and Jan Odijk and Stelios Piperidis},
 publisher =  {European Language Resources Association (ELRA)},
 isbn =  {978-2-9517408-8-4},
 language =  {english}
 }
 |