Leveraging seq2seq language generation for multi-level product issue identification

Yang Liu; Varnith Uttam Chordia; Hua Li; Siavash Fazeli; Yifei Sun; Vincent Gao; Na Zhang

Publication

Leveraging seq2seq language generation for multi-level product issue identification

By Yang Liu, Varnith Uttam Chordia, Hua Li, Siavash Fazeli, Yifei Sun, Vincent Gao, Na Zhang

2022

Download Copy BibTeX

Share

Download

Copy BibTeX

Share

In a leading e-commerce business, we receive hundreds of millions of customer feedback from different text communication channels such as product reviews. The feedback can contain rich information regarding customers’ dissatisfaction in the quality of goods and services. To harness such information to better serve customers, in this paper, we created a machine learning approach to automatically identify product issues and uncover root causes from the customer feedback text. We identify issues at two levels: coarse grained (LCoarse) and fine grained (L-Granular). We formulate this multi-level product issue identification problem as a seq2seq language generation problem. Specifically, we utilize transformer-based seq2seq models due to their versatility and strong transfer-learning capability. We demonstrate that our approach is label efficient and outperforms the traditional approach such as multi-class multi-label classification formulation. Based on human evaluation, our fine-tuned model achieves 82.1% and 95.4% human-level performance for L-Coarse and LGranular issue identification, respectively. Furthermore, our experiments illustrate that the model can generalize to identify unseen LGranular issues.

Leveraging seq2seq language generation for multi-level product issue identification

Latest news

Work with us