![]() Mentioning that the languages understood by computers and humans are quite different, yet Machine, due to its capacity to communicate with humans through languages. With no disputes, languages should be recognized as the most amazing artifacts everĭeveloped by mankind to enable communication. ![]() The effect of stopwords is generally found to be quite low in short documents compared with their long counterparts across the four Indian languages.Ĭommunication is fundamental to the evolution and development of all kinds of livingīeings. We also study the effect of stopwords on retrieval performance over document length. For each language, different lengths of the stopword list are explored and evaluated that lead to suggesting its optimal length. Is there any impact of non-corpus-based stopword removal on chosen Indian languages (if yes, to what extent)? Can we recommend, based on experiment, a number of stopwords for chosen Indian languages that are good enough from retrieval point of view? Is there any relationship of stopwords with average document length from retrieval perspective? It is observed that the stopword removal generally improves mean average precision (MAP) significantly compared with the case when it is not done. ![]() The issue was investigated from three viewpoints. We explore and evaluate the effect of stopwords in retrieval performance of different Indian languages such as Marathi, Bengali, Gujarati and Sanskrit.
0 Comments
Leave a Reply. |
AuthorWrite something about yourself. No need to be fancy, just an overview. ArchivesCategories |