Skip to content
akaturk Academic measurement

Article detail · 2023 · article

Machine Learning-Based Text Classification Comparison: Turkish Language Context

Journal Applied Sciences
OpenAlex Open access · gold SJR Q3 Top 10%
Year2023
Citations23OpenAlex
Percentile%92.8
FWCI3.031.00 = world average
Scopus (SJR)Q3

Data source split

  • YÖKSİS venueApplied Sciences
  • OpenAlexOpenAlex enrichment (abstract, citations, topics)

Abstract

OpenAlex English

The growth in textual data associated with the increased usage of online services and the simplicity of having access to these data has resulted in a rise in the number of text classification research papers. Text classification has a significant influence on several domains such as news categorization, the detection of spam content, and sentiment analysis. The classification of Turkish text is the focus of this work since only a few studies have been conducted in this context. We utilize data obtained from customers’ inquiries that come to an institution to evaluate the proposed techniques. Classes are assigned to such inquiries specified in the institution’s internal procedures. The Support Vector Machine, Naïve Bayes, Long Term-Short Memory, Random Forest, and Logistic Regression algorithms were used to classify the data. The performance of the various techniques was then analyzed after and before data preparation, and the results were compared. The Long Term-Short Memory technique demonstrated superior effectiveness in terms of accuracy, achieving an 84% accuracy rate, surpassing the best accuracy record of traditional techniques, which was 78% accuracy for the Support Vector Machine technique. The techniques performed better once the number of categories in the dataset was reduced. Moreover, the findings show that data preparation and coherence between the classes’ number and the number of training sets are significant variables influencing the techniques’ performance. The findings of this study and the text classification technique utilized may be applied to data in dialects other than Turkish.

Topics

Citations

OpenAlex cited_by_count. Not a WoS or Scopus citation count; those sources have no separate column here.

23citationsOpenAlex · cited_by_count (cache / database)

10 publications in the local catalog that cite this work (OpenAlex reference match; not the full global list).

  1. 2024 Research trends in deep learning and machine learning for cloud computing securityCitations 67 · OpenAlex
  2. 2024 Machine Learning-Based Analysis and Prediction of Liver CirrhosisCitations 16 · OpenAlex
  3. 2024 Remote Sensing Image Segmentation for Aircraft Recognition Using U-Net as Deep Learning ArchitectureCitations 13 · OpenAlex
  4. 2024 Enhancing Document Image Retrieval in Education: Leveraging Ensemble-Based Document Image Retrieval Systems for Improved PrecisionCitations 4 · OpenAlex
  5. 2024 LSRM: A New Method for Turkish Text ClassificationCitations 4 · OpenAlex
  6. 2024 Anticipate Movie Theme from Subtitle: A Deep Learning ApproachCitations 2 · OpenAlex
  7. 2025 TÜRKÇE DOĞAL DİL İŞLEME TEMELLİ ÇALIŞMALARIN TEORİK DEĞERLENDİRMESİ: YÖNTEMSEL ZORLUKLAR VE GELECEK PERSPEKTİFLERİCitations 1 · OpenAlex
  8. 2025 Text classification by machine learning algorithms using a new text feature extraction method based on image processingCitations 1 · OpenAlex
  9. 2025 Text classification by machine learning algorithms using a new text feature extraction method based on image processingCitations 1 · OpenAlex
  10. 2024 Sentiment Analysis based on Text with Universal Sentence Encoder and CNN-LSTM ModelsCitations 1 · OpenAlex

Authors

1
  1. AHMET ERCAN TOPCU ANKARA MEDİPOL ÜNİVERSİTESİ 1