Text classification system based on LLM

Yang Yu, Jinqi Li, Hongpeng Liu, Fan Yang, Xuanyao Yu · Journal of Artificial Intelligence Practice · 2024

Text classification based on long text and multi label text classification have always been a challenge, and the combination of these two problems brings great difficulties to text classification. This study focuses on the problem of long text multi label classification. The GLM large model was used to extract abstracts from text, which effectively reduced the length of the text and retained the main content of the text. Furthermore, the Bert-BiLSTM model was used to improve the accuracy of long sequence text. This model performs particularly well in multi label classification, and can accurately classify all category labels for multi label news classification, with much higher accuracy than conventional models.

Read the paper · More papers on PaperTik