MARATTO

article · Engineering and Technology Journal

Optimizing NLP-Text Classification in Knowledge Management Systems: A Literature Review

In plain language

Knowledge management systems rely on organising vast amounts of unstructured text to extract meaning. Natural language processing, particularly automated text classification, plays a central role in simplifying these systems by categorising documents automatically, improving information retrieval, and supporting organisational decision-making. The field has evolved from early machine learning approaches, including Naïve Bayes and Support Vector Machines, to advanced deep learning architectures such as Convolutional Neural Networks, Recurrent Neural Networks, and Transformers. Alongside these technical advancements, real-world industry use cases reveal critical operational challenges. These include maintaining system scalability, achieving explainability in automated choices, and managing ethical considerations. Addressing these existing research gaps remains essential, as effective natural language processing tools offer substantial capability to increase both the efficiency and overall effectiveness of knowledge management tasks across diverse enterprise environments.

Key takeaways

  • Natural language processing enables knowledge management systems to automate document categorisation, improve information retrieval, and assist decision-making.
  • Text classification has progressed from traditional machine learning models like Naïve Bayes and Support Vector Machines to advanced deep learning architectures such as Transformers.
  • Implementing text classification in industry highlights persistent challenges related to scalability, model explainability, and ethics.
  • Targeted improvements in text classification methods can boost the overall efficiency and effectiveness of knowledge management activities.

Why it matters

Modern organisations accumulate massive volumes of unstructured text that are difficult to navigate and interpret. Automated text classification methods help transform this unstructured information into searchable, useful assets. By accelerating data discovery and supporting informed decision-making, these technologies can enhance daily workplace productivity, provided that practitioners successfully navigate associated issues around system scale, algorithmic transparency, and ethical use.

Commercialisation angle

This review synthesises text classification methods for enterprise knowledge management, pointing directly to applications in automated document indexing and workplace search tools. Intended users include enterprise software developers and organisations handling large document archives. Because the work reviews existing industry implementations alongside mature machine learning and deep learning models, the underlying technologies appear applied and tested, though prospective adopters must still navigate documented hurdles in scalability and system explainability.

AI-generated from the published abstract. Always read the original work before citing.

Abstract

Knowledge Management Systems (KMS) are required to organize and assign meaning to huge amounts of organizational knowledge that are largely in the form of unstructured text. Natural Language Processing (NLP), and more immediately methods of text categorization, has been one of the principal enabler technologies to enable KMS to be simpler by helping to automatically categorize documents, enhance searching for information, and assist in decision-making. This paper offers an outline of the evolution of NLP-based text classification methods from initial machine learning methods such as Naïve Bayes and Support Vector Machines to current sophisticated deep learning algorithms such as Convolutional Neural Networks, Recurrent Neural Networks, and Transformers. We offer real-world industry use cases, issues of scalability, explainability, and ethics and encapsulate research areas of existing gaps. The findings underscore the enormous potential of NLP text classification to assist the effectiveness and efficiency of knowledge management (KM) activities.

Research topics

  • Text and Document Classification Technologies
  • Organizational and Employee Performance
  • Internet of Things and AI

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.5281/zenodo.22156226

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.