Investigating Job Mismatch in Software Industry through News Big Data
DOI:
https://doi.org/10.32890/jict2023.22.1.2Keywords:
News big data, network analysis, software manpower, human resource developmentAbstract
The purpose of this study is to identify issues related to software manpower, which became more important in the era of the Fourth
Industrial Revolution in Korea. The results of this study can provide guidelines for those who establish software manpower training policies for solving the software industry’s human resource paradox. As for the research method, the quantitative text network and qualitative analyses from industry experts were used to interpret the results. A total of 14,752 news data mentioning software manpower were extracted, and data pre-processing for the synonyms and negative words were performed. The network was non-directional and consisted of 14,074 words (nodes) and 1,542,383 word combinations (edges). In addition, the network was clustered based on Modularity, and the degree of connection and eigenvector centrality were used to determine the importance of nodes. The analysis of the results showed that the government’s efforts through the Korean Ministry of Science and ICT were vital in creating jobs that fueled software innovation growth, and that software education was actively promoted to develop software talent. This study had the following implications. It was confirmed that software is making a high contribution to the expansion of business opportunities and job creation in the fields of new technology and software convergence technology. To resolve the software manpower supply-demand mismatch, it is necessary to cultivate high-quality software talent and provide mid- to long-term activities to attract competent human resources. In addition, it is necessary to develop and expand programs that link education and recruitment in terms of public-private cooperation along with government-led investment to strengthen national software competitiveness.
References
Ali, R. M., Sa’adillah, M. D., & Teddy, M. (2020). Indonesian news classification using convolutional neural network. Indonesian Journal of ICT, 22, No. 1 (January) 2023, pp: 31– Journal of Electrical Engineering and Computer Science, 19(2), 1000–1009. http://doi.org/10.11591/ijeecs.v19.i2.pp1000-1009
Blondel, V. D., Guillaume, J. L., & Lambiotte, R. (2008). Fast unfolding of communities in large networks. Journal of Statistical Mechanics: Theory and Experiment, 2008(10). https://doi.org /10.1088%2F17425468%2F2008%2F10%2Fp10008
Bonacich, P. (1972). Factoring and weighting approaches to status scores and clique identification. Journal of Mathematical Sociology, 2(1), 113–120. https://doi.org/10.1080/002225 0X.1972.9989806
Bonacich, P. (1987). Power and centrality: A family of measures. American Journal of Sociology, 92(5), 1170–1182. https://doi.org/10.1086/228631
Brandes, U., Delling, D., Gaertler, M., Gorke, R., Hoefer, M., Nikoloski, Z., & Wagner, D. (2007). On modularity clustering. IEEE Transactions on Knowledge and Data Engineering, 20(2), 172–188. https://doi.org/10.1109/TKDE.2007.190689
Cha, D. (2014). Let’s build a Korean-style software manpower ecosystem! Technology and Management, (February, 2014). 18–21. http://203.234.181.180/webzine/201402.pdf
Choi, J., Lee, H., & Jin, E. (2019). A topic modeling analysis of the news topic on the 4th industrial revolution in Korea: Focusing on the difference by media type and each major period. Journal of Cybercommunication Academic Society, 36(2), 173–219. https://doi.org/120.36494/JCAS.2019.06.36.2.173
Driskell, J. E., & Mullen, B. (2004). Social network analysis. In Handbook of human factors and ergonomics methods (pp. 565–571). CRC Press. https://doi.org/10.1002/9781118131350
Jiapei, L., Shin, S., & Lee, H. (2017). Text mining and visualization of papers reviews using R language. Journal of Information and Communication Convergence Engineering, 15(3), 170–174. https://doi.org/10.6109/JICCE.2017.15.3.170
Jeong, G. W. (2011). Domestic software industry development strategy. Information and Communications Magazine, 29(1), 17–22.
Kim, S. (2020). Exploring the trends and challenges of artificial intelligence education through the analysis of newspapers in Korea, 1991–2020: A topic-modeling approach. Journal of Information and Communication Convergence Engineering, 18(4), 216–221. https://doi.org/10.6109/jicce.2020.18.4.216 Journal of ICT, 22, No. 1 (January) 2023, pp: 31–
Kim, Y., Kim, M., & Hwang B. (2022). Maximum node interconnection by a given sum of Euclidean edge lengths in a cluster node distribution. Journal of Information and Communication Convergence Engineering, 20(2), 90–95. https://doi.org/10.6109/jicce.2022.20.2.90
Lee, D., Huh, J., & Kim, J. (2018, April 23). Labor market forecast of promising SW areas. SPRI (Software Policy & Research Institute) Issue Report 2018-001. https://spri.kr/posts/ view/22049?code=data_all&study_type=issue_reports
Lee, H., & Kwon, J. (2019). A study on the issues and perceptions of 5G industry in digital transformation era: Focused on news network analysis. The Journal of Korean Institute of Information Technology, 17(11), 9–16. https://doi.org/10.14801/ jkiit.2019.17.11.9
Lee, H., & Song, J. (2021). A study on human-resource fostering methods in big data industry: Focusing on news network analysis. Journal of Knowledge Information Technology and Systems, 16(6), 1283–1294. https://doi.org/10.34163/ jkits.2021.16.6.016
Lee, J., Wu, G., & Jung, H. (2021). Deep learning document analysis system based on keyword frequency and section centrality analysis. Journal of Information and Communication Convergence Engineering, 19(1), 48–53. https://doi.org/10.6109/jicce.2021.19.1.
Lee, S. (2010). A preliminary study on the co-author network analysis of Korean library & information science research community. Journal of Korean Library and Information Science Society, 41(2), 297–315. https://doi.org/10.16981/kliss.41.2.201006.297
Li, H. (2018, March). Centrality analysis of online social network big data. In 2018 IEEE 3rd International Conference on Big Data Analysis (ICBDA) (pp. 38–42). IEEE. https://doi.org/10.1109/ ICBDA.2018.8367648.
Maylawati, D. S., & Saptawati, G. A. P. (2017). Set of frequent word item sets as feature representation for text with Indonesian slang. Journal of Physics Conference Series, 801(1). https://doi.org/10.1088/1742-6596/801/1/012066
Newman, M. E. J. (2004). Analysis of weighted networks. Physical Review E, 70(5), 056131. https://doi.org/10.1103/ PhysRevE.70.056131
Newman, M. E. J., & Girvan, M. (2004). Finding and evaluating community structure in networks. Physical Review E, 69(2), 026113. https://doi.org/10.1103/PhysRevE.69.026113 Journal of ICT, 22, No. 1 (January) 2023, pp: 31–
Park, E., Kim, Y., & Park, C. (2017). Comparative analysis of domestic and foreign hospice nursing research topics using text network analysis. Journal of Korean Academy of Nursing, 47(5), 600–612. http://doi.org/10.4040/jkan.2017.47.5.600
Scott, J. (1991). Social network analysis (2nd ed.). London: Sage. https://dx.doi.org/10.4135/9781412985864
Schwab, K. (2016, January 14). The fourth industrial revolution: What it means, how to respond. World Economic Forum. https://www. weforum.org/agenda/2016/01/the-fourth-industrial-revolution-what-it-means-and-how-to-respond/
Software Policy & Research Institute (SPRI). (2019, August). 2018 White Pater of Korea Software Industry. Software Policy & Research Institute. https://spri.kr/posts/view/22745?code=data_ all&study_type=annual_reports
Software Policy & Research Institute (SPRI). (2022). 2018 Software Industry Survey 2021. Software Policy & Research Institute. https://stat.spri.kr/posts/view/23462?code=stat_sw_reports
Tissera, M., & Weerasinghe, R. (2022). Grammatical structure oriented automated approach for surface knowledge extraction from open domain unstructured text. Journal of Information and Communication Convergence Engineering, 20(2), 113–124. https://doi.org/10.6109/jicce.2022.20.2.113
Published
Issue
Section
License
Copyright (c) 2023 Journal of Information and Communication Technology

This work is licensed under a Creative Commons Attribution 4.0 International License.
How to Cite
Research impact
Harvested 2026-09-06Counts differ between services because each indexes a different body of literature. None of them is the whole picture.
2002 - 2020






















