Influential tweeters in relation to highly cited articles in altmetric big data
The relationship between influential tweeters and highly cited articles in the field of information sciences was analysed using Twitter data gathered by Altmetric.com from July 2011 through February 2017. The dataset consists of more than 10,000 tweets, and these mentions, retweets and followers were used to generate a connected, undirected graph. This graph reveals the most influential tweeters by identifying the largest drop in the eigenvalue of adjacency or affinity matrix of a graph when certain nodes are removed; those which, when deleted, cause the greatest drop in the eigenvalue of the graph are considered to be the most influential. The machine-learning model applied in this work utilizes a feature vector containing the accumulated sum of the rank scores of those influential users who tweet a given article, along with known altmetric features such as the user type and post counts for various social media. Finally, the supervised-learning model was trained using Random Forest and Support Vector Machine classifiers with 11 features, including the sum of the ranks of influential users who tweet a given article in our dataset. The results were analysed using Receiver Operating Characteristic (ROC) curves and Precision Recall (PR) curves, which give the commendable outcomes compared to the baseline model. We found that, for the classification of highly cited articles, Twitter users’ score for influence is the most important feature. Finally, we show that our model—which was trained by taking the score for influence into consideration—outperforms the baseline, at 79% for ROC and 90% for PR with the Random Forest Model, effectively identifying the highly cited articles.
KeywordsAltmetrics Influential users Twitter Highly cited articles
The research work has been supported by the NRPU grant no. 6857/Punjab/NRPU/R&D/HEC/2016 funded by the Higher Education Commission of Pakistan.
- Alonso, O., Carson, C., Gerster, D., Ji, X., & Nabar, U. S. (2010). Detecting uninteresting content in text streams. In Proceedings of the SIGIR 2010 workshop on crowdsourcing for search evaluation (CSE 2010) (pp. 39–42).Google Scholar
- Anger, I., & Kittl, C. (2011). Measuring influence on Twitter. In Proceedings of the 11th international conference on knowledge management and knowledge technologies, 31(1), pp. 1–31.Google Scholar
- Chung, F. R. (1997). Spectral graph theory (vol. 92, Regional Conference Series in Mathematics). Rhode Island: American Mathematical Society/Conference Board of the Mathematical Sciences. ISBN: 978-0-8218-0315-8.Google Scholar
- Escamilla, I., Torres-Ruiz, M., Moreno-Ibarra, M., Quintero, R., Guzmán, G., & Luna-Soto, V. (2016). Geocoding tweets approach based on conceptual representations in the context of the knowledge society. International Journal on Semantic Web and Information Systems (IJSWIS), 12(1), 44–61.CrossRefGoogle Scholar
- Haustein, S., Bowman, T. D., & Costas, R. (2015). Interpreting ‘altmetrics’: Viewing acts on social media through the lens of citation and social theories. arXiv preprint arXiv:1502.05701.
- Hussain, A. R., Hameed, M. A., & Sayeedunnissa, S. F. (2012). Measuring influence in social networks using a network amplification score-an analysis using cloud computing. In 2012 12th International conference on hybrid intelligent systems (HIS).Google Scholar
- Kemp, S. (2017). Digital in 2017: Global overview. Retrieved from ‘We are social. https://wearesocial.com/blog/2017/01/digital-in-2017-global-overview. Accessed 10 June 2018.
- Lotan, G., Ananny, M., Gaffney, D., & Pearce, I. (2011). The Arab Spring/The revolutions were tweeted: Information flows during the 2011 Tunisian and Egyptian revolutions. International Journal of Communication, 5, 1375–1405.Google Scholar
- Priem, J., Piwowar, H., & Hemminger, B. (2011). Altmetrics in the wild: An exploratory study of impact metrics based on social media. In Metrics 2011: Symposium on informetric and scientometric research, New Orleans, USA.Google Scholar
- Priem, J., Taraborelli, D., Groth, P., & Neylon, C. (2010). Altmetrics: A manifesto. Available online at http://altmetrics.org/manifesto/. Accessed 10 June 2018.
- Quercia, D., Ellis, J., Capra, L., & Crowcroft, J. (2011). In the mood for being influential on Twitter. In Privacy, security, risk and trust (PASSAT) and 2011 IEEE 3rd international conference on social computing (SocialCom) (pp. 307–314). IEEE.Google Scholar
- Tariq, J., Ahmad, M., Khan, I., & Shabbir, M. (2017). Scalable approximation algorithm for network immunization. arXiv preprint arXiv:1711.00784.
- Tsou, A., Bowman, T.D., Ghazinejad, A., & Sugimoto, C.R. (2015). Who tweets about science? In Proceedings of the 2015 international society for scientometrics and informetrics (pp. 95–100), Istanbul, Turkey.Google Scholar
- Yang, M.-C., Lee, J.-T., Lee, S.-W., & Rim, A. H.-C. (2012). Finding interesting posts in Twitter based on retweet graph analysis. In 35th International ACM SIGIR conference on research and development in information retrieval (pp. 1073–1074), August, Portland, OR.Google Scholar