J. Rudman, The state of authorship attribution studies: some problems and solutions. Comput. Hum. 31(4), 351–365 (1997)
Google Scholar
H. Baayen, H. Van Halteren, F. Tweedie, Outside the cave of shadows: using syntactic annotation to enhance authorship attribution. Liter. Linguist. Comput. 11(3), 121–132 (1996)
Google Scholar
J.F. Burrows, Word-patterns and story-shapes: the statistical analysis of narrative style. Liter. Linguist. Comput. 2(2), 61–70 (1987)
Google Scholar
O. De Vel, A. Anderson, M. Corney, G. Mohay, Mining e-mail content for author identification forensics. ACM SIGMOD Rec. 30(4), 55–64 (2001)
CrossRef
Google Scholar
R. Zheng, Y. Qin, Z. Huang, H. Chen, Authorship analysis in cybercrime investigation, in International Conference on Intelligence and Security Informatics (2003), pp. 59–73
Google Scholar
A. Abbasi, H. Chen, Writeprints: a stylometric approach to identity-level identification and similarity detection in cyberspace. ACM Trans. Inf. Syst. 26(2), 7 (2008)
CrossRef
Google Scholar
S. Argamon, M. Koppel, G. Avneri, Routing documents according to style, in First International Workshop on Innovative Information Systems (1998), pp. 85–92
Google Scholar
D.I. Holmes, The evolution of stylometry in humanities scholarship. Liter. Linguist. Comput. 13(3), 111–117 (1998)
Google Scholar
G.U. Yule, On sentence-length as a statistical characteristic of style in prose: with application to two cases of disputed authorship. Biometrika 30(3/4), 363–390 (1939)
Google Scholar
W.W. Greg, The statistical study of literary vocabulary. JSTOR 291–293 (1944)
Google Scholar
R. Zheng, J. Li, H. Chen, Z. Huang, A framework for authorship identification of online messages: writing-style features and classification techniques. J. Am. Soc. Inf. Sci. Technol. 57(3), 378–393 (2006)
CrossRef
Google Scholar
F. Mosteller, D. Wallace, Inference and disputed authorship: the federalist (1964)
Google Scholar
E. Stamatatos, N. Fakotakis, G. Kokkinakis, Automatic text categorization in terms of genre and author. Comput. Linguist. 26(4), 471–495 (2000)
Google Scholar
M. Corney, O. De Vel, A. Anderson, G. Mohay, Gender-preferential text mining of e-mail discourse, in 18th Annual Proceedings Computer Security Applications Conference, 2002 (2002), pp. 282–289
Google Scholar
M. Gamon, Linguistic correlates of style: authorship classification with deep linguistic analysis features, in Proceedings of the 20th International Conference on Computational Linguistics (2004), p. 611
Google Scholar
E. Frank, S. Kramer, Ensembles of nested dichotomies for multi-class problems, in Proceedings of the Twenty-First International Conference on Machine Learning (2004), p. 39
Google Scholar
E. Frank, M.A. Hall, I.H. Witten, The WEKA Workbench. Online Appendix for “Data Mining: Practical Machine Learning Tools and Techniques” (Morgan Kaufmann, 2016)
Google Scholar
Discovering Email Header Forensic Analysis! (2017). [Online]. http://www.xploreforensics.com/blog/email-header-forensic-analysis.html. Accessed 5 May 2020
J. Diederich, J. Kindermann, E. Leopold, G. Paass, Authorship attribution with support vector machines. Appl. Intell. 19(1–2), 109–123 (2003)
MATH
Google Scholar
O. De Vel, Mining e-mail authorship, in Proc. Workshop on Text Mining, ACM International Conference on Knowledge Discovery and Data Mining (KDD’2000) (2000)
Google Scholar
F. Mosteller, D.L. Wallace, Applied Bayesian and Classical Inference: The Case of the Federalist Papers (Springer Science & Business Media, 2012)
Google Scholar
G. Salton, M.J. McGill, Introduction to modern information retrieval (1986)
Google Scholar
J.R. Quinlan, Induction of decision trees. Mach. Learn. 1(1), 81–106 (1986)
Google Scholar
L.M. Manevitz, M. Yousef, One-class SVMs for document classification. J. Mach. Learn. Res. 2, 139–154 (2001)
MATH
Google Scholar
T. Joachims, Text categorization with support vector machines: learning with many relevant features, in European Conference on Machine Learning (1998), pp. 137–142
Google Scholar
G.-F. Teng, M.-S. Lai, J.-B. Ma, Y. Li, E-mail authorship mining based on SVM for computer forensic, in Proceedings of 2004 International Conference on Machine Learning and Cybernetics, vol. 2 (2004), pp. 1204–1207
Google Scholar
H. Li, D. Shen, B. Zhang, Z. Chen, Q. Yang, Adding semantics to email clustering, in Sixth International Conference on Data Mining, 2006. ICDM’06 (2006), pp. 938–942
Google Scholar
H. Van Halteren, Author verification by linguistic profiling: an exploration of the parameter space. ACM Trans. Speech Lang. Process. 4(1), 1 (2007)
Google Scholar
T. Kucukyilmaz, B.B. Cambazoglu, C. Aykanat, F. Can, Chat mining: predicting user and message attributes in computer-mediated communication. Inf. Process. Manag. 44(4), 1448–1466 (2008)
CrossRef
Google Scholar
Y. Zhao, J. Zobel, Effective and scalable authorship attribution using function words, in Asia Information Retrieval Symposium (2005), pp. 174–189
Google Scholar
O. De Vel, A.M. Anderson, M.W. Corney, G.M. Mohay, Multi-topic e-mail authorship attribution forensics (2001)
Google Scholar
M. Koppel, J. Schler, S. Argamon, Computational methods in authorship attribution. J. Am. Soc. Inf. Sci. Technol. 60(1), 9–26 (2009)
Google Scholar
M. Koppel, S. Argamon, A.R. Shimoni, Automatically categorizing written texts by author gender. Liter. Linguist. Comput. 17(4), 401–412 (2002)
Google Scholar
S. Argamon, M. Šarić, S.S. Stein, Style mining of electronic messages for multiple authorship discrimination: first results, in Proceedings of the Ninth ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (2003), pp. 475–480
Google Scholar
A. Abbasi, H. Chen, J.F. Nunamaker, Stylometric identification in electronic markets: scalability and robustness. J. Manag. Inf. Syst. 25(1), 49–78 (2008)
Google Scholar
O.Y. de Vel, M.W. Corney, A.M. Anderson, G.M. Mohay, Language and gender author cohort analysis of e-mail for computer forensics (2002)
Google Scholar
J. Novak, P. Raghavan, A. Tomkins, Anti-aliasing on the web, in Proceedings of the 13th International Conference on World Wide Web (2004), pp. 30–39
Google Scholar
S.E. Robertson, K.S. Jones, Relevance weighting of search terms. J. Am. Soc. Inf. Sci. 27(3), 129–146 (1976)
Google Scholar
J.R. Quinlan et al., Learning with continuous classes, in 5th Australian Joint Conference on Artificial Intelligence, vol. 92 (1992), pp. 343–348
Google Scholar
S.L. Salzberg, C4. 5: programs for machine learning by j. ross quinlan. morgan kaufmann publishers, inc., 1993. Mach. Learn. 16(3), 235–240 (1994)
Google Scholar
N. Cristianini, J. Shawe-Taylor, An Introduction to Support Vector Machines (Cambridge University Press, Cambridge, 2000)
MATH
Google Scholar
M. Van Uden, Rocchio: relevance feedback in learning classification algorithms, in Proceedings of the ACM SIGIR Conference (1998)
Google Scholar
F. Sebastiani, Machine learning in automated text categorization. ACM Comput. Surv. 34(1), 1–47 (2002)
Google Scholar
R. Agrawal, T. Imieliński, A. Swami, Mining association rules between sets of items in large databases. ACM SIGMOD Rec 22(2), 207–216 (1993)
CrossRef
Google Scholar