arXiv · 1410.4863
Arabic Language Text Classification Using Dependency Syntax-Based Feature Selection
Abstract
We study the performance of Arabic text classification combining various techniques: (a) tfidf vs. dependency syntax, for feature selection and weighting; (b) class association rules vs. support vector machines, for classification. The Arabic text is used in two forms: rootified and lightly stemmed. The results we obtain show that lightly stemmed text leads to better performance than rootified text; that class association rules are better suited for small feature sets obtained by dependency syntax constraints; and, finally, that support vector machines are better suited for large feature sets based on morphological feature selection criteria.
Explore related subjects
Keep this discovery
Yannis Haralambous, Yassir Elidrissi, Philippe Lenca. 2014-10-17. Arabic Language Text Classification Using Dependency Syntax-Based Feature Selection. https://arxiv.org/abs/1410.4863
Cite the original work for its findings. Save a collection to share your selection of sources.