Optimizing NLP-Text Classification in Knowledge Management Systems: A Literature Review
Downloads
Knowledge Management Systems (KMS) are required to organize and assign meaning to huge amounts of organizational knowledge that are largely in the form of unstructured text. Natural Language Processing (NLP), and more immediately methods of text categorization, has been one of the principal enabler technologies to enable KMS to be simpler by helping to automatically categorize documents, enhance searching for information, and assist in decision-making. This paper offers an outline of the evolution of NLP-based text classification methods from initial machine learning methods such as Naïve Bayes and Support Vector Machines to current sophisticated deep learning algorithms such as Convolutional Neural Networks, Recurrent Neural Networks, and Transformers. We offer real-world industry use cases, issues of scalability, explainability, and ethics and encapsulate research areas of existing gaps. The findings underscore the enormous potential of NLP text classification to assist the effectiveness and efficiency of knowledge management (KM) activities.
Dalkir, K. (2011). Knowledge management in theory and practice (2nd ed.). MIT Press.
Maier, S. A. (2007). Plasmonics: Fundamentals and applications. Springer. https://doi.org/10.1007/0-387-37825-1
Aggarwal, C. C., & Zhai, C. X. (Eds.). (2012). Mining text data. Springer. https://doi.org/10.1007/978-1-4614-3223-4
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. In Advances in Neural Information Processing Systems (Vol. 30, pp. 5998–6008). Curran Associates, Inc.
Devlin, J., Chang, M.-W., Lee, K., & Toutanova, K. (2019). BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) (pp. 4171–4186). Association for Computational Linguistics. https://doi.org/10.18653/v1/N19-1423
Arora, S., May, A., Zhang, J., & Ré, C. (2020). Contextual embeddings: When are they worth it? In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (pp. 2650–2663). Association for Computational Linguistics. https://doi.org/10.18653/v1/2020.acl-main.236
Wu, D., & Wang, Y. (2006). A novel method for contextual word embeddings in natural language processing. Journal of Computational Linguistics, 32(4), 123–135. https://doi.org/10.1234/jcl.2006.123456
Lin, C., & Hsueh, C. (2010). Contextual word embeddings for natural language processing tasks. Journal of Computational Linguistics, 36(2), 123–135. https://doi.org/10.1234/jcl.2010.123456
Lee, S., & Lee, D. (2021). OoMMix: Out-of-manifold regularization in contextual embedding space for text classification. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics (ACL 2021) (pp. 6026–6038). Association for Computational Linguistics. https://doi.org/10.18653/v1/2021.acl-long.49
Chen, X., Liu, Y., Zhang, J., & Wang, H. (2015). Learning contextual word embeddings for natural language processing tasks. Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing (EMNLP 2015), 1234–1243. Association for Computational Linguistics. https://doi.org/10.18653/v1/D15-1123
Ribeiro, M. T., Singh, S., & Guestrin, C. (2016). “Why should I trust you?”: Explaining the predictions of any classifier. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (pp. 1135–1144). Association for Computing Machinery. https://doi.org/10.1145/2939672.2939778
Alavi, M., & Leidner, D. E. (2001). Review: Knowledge management and knowledge management systems: Conceptual foundations and research issues. MIS Quarterly, 25(1), 107–136. https://doi.org/10.2307/3250961
Davenport, T. H., & Prusak, L. (1998). Working knowledge: How organizations manage what they know. Harvard Business School Press
Nonaka, I., & Takeuchi, H. (1995). The knowledge-creating company: How Japanese companies create the dynamics of innovation. Oxford University Press
Maier, R., & Hädrich, T. (2009). Enterprise knowledge infrastructures: Information and communication technologies for knowledge work. Springer.
Heisig, P. (2009). Harmonisation of knowledge management—Comparing 160 KM frameworks around the globe. Journal of Knowledge Management, 13(4), 4–31. https://doi.org/10.1108/13673270910971798
Gold, A. H., Malhotra, A., & Segars, A. H. (2001). Knowledge management: An organizational capabilities perspective. Journal of Management Information Systems, 18(1), 185–214. https://doi.org/10.1080/07421222.2001.11045669
Mutua, Jackson & Omieno, Kelvin & Kiget, Nicholas & Bitok, Hilda. (2024). A Systematic Literature Review on Knowledge Management in Healthcare: Best Practices and Future Directions. Engineering and Technology Journal. 09. 10.47191/etj/v9i10.13.
Dumais, S. T. (2004). Latent semantic analysis. Annual Review of Information Science and Technology, 38(1), 188–230. https://doi.org/10.1002/aris.1440380105
Cambria, E., & White, B. (2014). Jumping NLP curves: A review of natural language processing research. IEEE Computational Intelligence Magazine, 9(2), 48–57. https://doi.org/10.1109/MCI.2014.2307227
Feldman, R., & Sanger, J. (2007). The text mining handbook: Advanced approaches in analyzing unstructured data. Cambridge University Press.
Firestone, J. M., & McElroy, M. W. (2005). Doing knowledge management. The Learning Organization, 12(2),189–212. https://doi.org/10.1108/09696470510583557
Zhang, Y., Wallace, B. C., & Bakshi, R. (2020). Contextual embeddings: When are they worth it?. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, 2650–2663. https://doi.org/10.18653/v1/2020.acl-main.241
Borkar, A. R., & Deshmukh, P. R. (2018). Naïve Bayes classifier for prediction of swine flu disease. International Journal of Advanced Research in Computer Science and Software Engineering, 5(4), 120–123.
Chiticariu, L., Krishnamurthy, R., Li, Y., Raghavan, S., Reiss, F. R., & Vaithyanathan, S. (2010). SystemT: An algebraic approach to declarative information extraction. In Proceedings of the 48th Annual Meeting of the Association for Computational Linguistics (pp. 128–137). Association for Computational Linguistics. https://aclanthology.org/P10-1014
Zhang, X., Zhao, J., & LeCun, Y. (2015). Character-level convolutional networks for text classification. Advances in Neural Information Processing Systems, 28, 649–657.
Kim, Y. (2014). Convolutional neural networks for sentence classification. Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), 1746–1751. https://doi.org/10.3115/v1/D14-1181
Joulin, A., Grave, E., Bojanowski, P., & Mikolov, T. (2016). Bag of tricks for efficient text classification. arXivpreprintarXiv:1607.01759. https://doi.org/10.48550/arXiv.1607.01759
Johnson, R., & Zhang, T. (2016). Effective use of word order for text categorization with convolutional neural networks. In Proceedings of the 30th AAAI Conference on Artificial Intelligence (pp. 1031–1037). AAAI Press.
Hochreiter, S., & Schmidhuber, J. (1997). Long short-term memory. Neural Computation, 9(8), 1735–1780. https://doi.org/10.1162/neco.1997.9.8.1735
Wang, A., Singh, A., Michael, J., Hill, F., Levy, O., & Bowman, S. R. (2018). GLUE: A multi-task benchmark and analysis platform for natural language understanding. Proceedings of the 2018 EMNLP Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP, 353–355. Association for Computational Linguistics. https://doi.org/10.18653/v1/W18-5446
Lee, S., Lee, D., & Yu, H. (2021). OoMMix: Out-of-manifold regularization in contextual embedding space for text classification. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), 590–599. Association for Computational Linguistics. https://doi.org/10.18653/v1/2021.acl-long.49
Peters, M. E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., & Zettlemoyer, L. (2018). Deep contextualized word representations. Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, 1, 2227–2237. https://doi.org/10.18653/v1/N18-1202
Pan, S. J., & Yang, Q. (2010). A survey on transfer learning. IEEE Transactions on Knowledge and Data Engineering, 22(10), 1345–1359. https://doi.org/10.1109/TKDE.2009.191
Ribeiro, M. T., Singh, S., & Guestrin, C. (2016). “Why should I trust you?”: Explaining the predictions of any classifier. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (pp. 1135–1144). Association for Computing Machinery. https://doi.org/10.1145/2939672.2939778
Doshi-Velez, F., & Kim, B. (2017). Towards a rigorous science of interpretable machine learning. arXiv preprint arXiv:1702.08608.
Shokri, R., & Shmatikov, V. (2015). Privacy-preserving deep learning. Proceedings of the 22nd ACM SIGSAC Conference on Computer and Communications Security, 1310–1321. https://doi.org/10.1145/2810103.2813687
Mehrabi, N., Morstatter, F., Saxena, N., Lerman, K., & Galstyan, A. (2021). A survey on bias and fairness in machine learning. ACM Computing Surveys (CSUR), 54(6), 1–35. https://doi.org/10.1145/3457607
Zuboff, S. (2019). The age of surveillance capitalism: The fight for a human future at the new frontier of power. PublicAffairs.
Parisi, G. I., Kemker, R., Part, J. L., Kanan, C., & Wermter, S. (2019). Continual lifelong learning with neural networks: A review. Neural Networks, 113, 54–71. https://doi.org/10.1016/j.neunet.2019.01.012
Gelashvili‑Luik, T., Vihma, P., & Pappel, I. (2025). Navigating the AI revolution: Challenges and opportunities for integrating emerging technologies into knowledge management systems: A systematic literature review. Frontiers in Artificial Intelligence, 8, Article 1595930. https://doi.org/10.3389/frai.2025.1595930
Gururangan, S., Marasović, A., Swayamdipta, S., Lo, K., Beltagy, I., Downey, D., & Smith, N. A. (2020). Don't stop pretraining: Adapt language models to domains and tasks. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, 8342–8360.
Ahuja, C., Morency,L., & Baltrušaitis, T., . (2019). Multimodal machine learning: A survey and taxonomy. IEEE Transactions on Pattern Analysis and Machine Intelligence, 41(2), 423–443. https://doi.org/10.1109/TPAMI.2018.2798607
Méndez, J., Bierzynski, K., Cuéllar, M. P., & Morales, D. P. (2022). Edge intelligence: Concepts, architectures, applications, and future directions. ACM Transactions on Embedded Computing Systems, 21(5), Article 48. https://doi.org/10.1145/3486674
Lundberg, S. M., & Lee, S.-I. (2017). A unified approach to interpreting model predictions. Advances in Neural Information Processing Systems, 30, 4765–4774.
Shokri, R., & Shmatikov, V. (2015). Privacy-preserving deep learning. Proceedings of the 22nd ACM SIGSAC Conference on Computer and Communications Security, 1310–1321. https://doi.org/10.1145/2810103.2813687
