00
Days
00
Hrs
00
Min
00
Sec
Submit Your Paper

Data Science Talent Demand in China: A Large-Scale Job Posting Analysis and Implications for Curriculum Alignment

Authors

Yang Shiwei

Faculty of Computing and Metal-Technology,Sultan Idris Education University (MY)

Ashardi Abas

Faculty of Computing and Metal-Technology,Sultan Idris Education University (MY)

Article Information

DOI: 10.51583/IJLTEMAS.2026.150300052

Subject Category: Data Science

Volume/Issue: 15/3 | Page No: 630-659

Publication Timeline

Submitted: 2026-04-10

Published: 2026-04-10

Abstract

The rapid expansion of the digital economy has intensified demand for data science talent in China, yet higher education curricula often lag behind evolving industry requirements. While the skills gap is widely acknowledged, few studies offer large-scale empirical evidence linking labor-market signals to curriculum design in the Chinese context. This study analyzes 12,436 data science–related job postings from major Chinese recruitment platforms between 2022 and 2024 to map employer demands and their educational implications. Using text preprocessing, natural language processing, skill extraction, clustering, and regression techniques, we identify key patterns in geographic distribution, required competencies, and salary drivers. Results show that job demand is heavily concentrated in Tier-1 and Tier-2 cities. The most frequently required skills include Python, SQL, machine learning, big data tools (e.g., Spark and Hadoop), statistical analysis, and communication abilities. Salaries are most strongly influenced by city tier, company size, educational qualifications, and proficiency in specialized technical areas such as cloud platforms and deep learning. A notable mismatch persists between university training and market expectations—particularly in applied technical skills and interdisciplinary problem-solving. These findings provide an evidence-based foundation for curriculum redesign, stronger industry–academia collaboration, and more responsive educational planning. Future research should extend to longitudinal forecasting and cross-country comparisons.

Keywords

data science talent demand; job posting analytics; curriculum alignment; China

Downloads

References

1. W. S. Cleveland, “Data science: An action plan for expanding the technical areas of the field of statistics,” International Statistical Review, vol. 69, no. 1, pp. 21–26, 2001. [Google Scholar] [Crossref]

2. V. Dhar, “Data science and prediction,” Communications of the ACM, vol. 56, no. 12, pp. 64–73, 2013. [Google Scholar] [Crossref]

3. W. van der Aalst, “Data science in action,” in Process Mining, Berlin, Germany: Springer, 2016, pp. 3–23. [Google Scholar] [Crossref]

4. L. Cao, “Data science: A comprehensive overview,” ACM Computing Surveys, vol. 50, no. 3, pp. 1–42, 2017. [Google Scholar] [Crossref]

5. I. Y. Song and Y. Zhu, “Big data and data science: What should we teach?,” Expert Systems, vol. 33, no. 4, pp. 364–373, 2016. [Google Scholar] [Crossref]

6. D. W. Xia and Z. L. Zhang, “On training mode of big-data talents in the age of data technology,” Journal of Southwest China Normal University (Natural Science Edition), vol. 41, no. 9, pp. 191–196, 2016. [Google Scholar] [Crossref]

7. Y. M. Chen, “On teaching environment and teaching mode based on ‘Internet +’,” Journal of Southwest China Normal University (Natural Science Edition), vol. 41, no. 3, pp. 228–232, 2016. [Google Scholar] [Crossref]

8. Y. Y. Zhu and Y. Xiong, “Training data scientists in the era of big data,” Big Data Research, vol. 2, no. 3, pp. 106–112, 2016. [Google Scholar] [Crossref]

9. L. B. Wu, “Cultivating big data talent by combining various disciplines and utilizing multiple resources,” Big Data Research, vol. 2, no. 5, pp. 89–94, 2016. [Google Scholar] [Crossref]

10. K. C. C. Chan and T. T. He, “Data science: The demand and development of talents,” Big Data Research, vol. 2, no. 5, pp. 95–106, 2016. [Google Scholar] [Crossref]

11. J. Chen, “Research and application of association rules algorithm in online recruitment system,” M.S. article, Xi’an University of Science and Technology, Xi’an, China, 2009. [Google Scholar] [Crossref]

12. T. Cao, “Research on the application of data mining in employee online recruitment,” M.S. article, Jinan University, Guangzhou, China, 2009. [Google Scholar] [Crossref]

13. J. Wu, “Analysis of occupation technique and ability in network engineering using recruitment information,” in Proc. 9th Int. Conf. Fuzzy Systems and Knowledge Discovery (FSKD), 2012, pp. 1396–1400. [Google Scholar] [Crossref]

14. Y. Zhang, W. Zhao, H. Bao, Y. Li, and K. Zhou, “Data mining of online recruitment information based on K-means and correlation analysis,” Software Engineering, vol. 20, no. 5, pp. 10–14, 2017. [Google Scholar] [Crossref]

15. C. Liu, “Research on recruitment demand information for data jobs,” M.S. article, Lanzhou University of Finance and Economics, Lanzhou, China, 2019. [Google Scholar] [Crossref]

16. H. Tang, “Research on talent demand and talent training approach of educational technology subject based on text mining,” M.S. article, Beijing University of Posts and Telecommunications, Beijing, China, 2019. [Google Scholar] [Crossref]

17. C. Priyadarshini, S. Sreejesh, and M. R. Anusree, “Effect of information quality of employment website on attitude toward the website,” International Journal of Manpower, vol. 38, no. 1, pp. 104–124, 2017. [Google Scholar] [Crossref]

18. R. S. Baker and P. S. Inventado, “Educational data mining and learning analytics,” in Learning Analytics, New York, NY, USA: Springer, 2014, pp. 61–75. [Google Scholar] [Crossref]

19. A. P. Ayala, “Educational data mining: A survey and a data mining-based analysis of recent works,” Expert Systems with Applications, vol. 41, no. 4, pp. 1432–1462, 2014. [Google Scholar] [Crossref]

20. D. J. Hand and N. M. Adams, “Data mining,” Wiley StatsRef: Statistics Reference Online, pp. 1–7, 2014. [Google Scholar] [Crossref]

21. H. Qing, N. Li, W. Luo, and Z. Shi, “Overview of machine learning algorithms under big data,” Pattern Recognition and Artificial Intelligence, vol. 27, no. 4, pp. 327–336, 2014. [Google Scholar] [Crossref]

22. S. Zhou, Z. Xu, and X. Tang, “Method for determining the optimal number of clusters in K-means algorithm,” Computer Applications, vol. 30, no. 8, pp. 1995–1998, 2010. [Google Scholar] [Crossref]

23. Y. Wu, “Overview of clustering algorithms,” Computer Science, vol. 42, no. 6A, pp. 491–499, 2015. [Google Scholar] [Crossref]

24. F. Å. Nielsen, Data Mining Using Python. 2017. [Google Scholar] [Crossref]

25. V. G. Nair, Getting Started with Beautiful Soup. Birmingham, U.K.: Packt Publishing, 2014. [Google Scholar] [Crossref]

26. G. Zaccone, Python Parallel Programming Cookbook. Birmingham, U.K.: Packt Publishing, 2015. [Google Scholar] [Crossref]

27. The State Council of the People’s Republic of China, New Generation Artificial Intelligence Development Plan. Beijing, China, 2017. [Google Scholar] [Crossref]

28. The State Council of the People’s Republic of China, The 14th Five-Year Plan for Digital Economy Development. Beijing, China, 2021. [Google Scholar] [Crossref]

29. National Development and Reform Commission of the People’s Republic of China, Implementation Plan for the Eastern Data and Western Computing Project. Beijing, China, 2022. [Google Scholar] [Crossref]

30. Ministry of Education of the People’s Republic of China, Statistical Bulletin on National Education Development 2023. Beijing, China, 2023. [Google Scholar] [Crossref]

31. Ministry of Education of the People’s Republic of China, Guidelines on Big Data and Artificial Intelligence Teaching in Universities. Beijing, China, 2023. [Google Scholar] [Crossref]

32. China Big Data Industry Alliance, China Big Data Industry Development Report 2023. Beijing, China, 2023. [Google Scholar] [Crossref]

Metrics

Views & Downloads

Similar Articles

© 2026 IJLTEMAS · RSIS International. All rights reserved. ISSN 2278-2540.