References
Atzori, L., Iera, A., & Morabito, G. (2010). The internet of things:
A survey. Computer Networks, 54(15), 2787–2805. https://doi.org/10.1016/j.comnet.2010.05.010
Baltrušaitis, T., Ahuja, C., & Morency, L.-P. (2019). Multimodal
machine learning: A survey and taxonomy. IEEE Transactions on
Pattern Analysis and Machine Intelligence, 41(2), 423–443.
https://doi.org/10.1109/TPAMI.2018.2798607
Box, G. E. P., & Jenkins, G. M. (1970). Time series analysis:
Forecasting and control. Holden-Day.
Breiman, L. (2001). Random forests. Machine Learning,
45(1), 5–32. https://doi.org/10.1023/A:1010933404324
Breiman, L., Friedman, J. H., Olshen, R. A., & Stone, C. J. (1984).
Classification and regression trees. Chapman; Hall/CRC.
Broman, K. W., & Woo, K. H. (2018). Data organization in
spreadsheets. The American Statistician, 72(1), 2–10.
https://doi.org/10.1080/00031305.2017.1375989
Cortes, C., & Vapnik, V. (1995). Support-vector networks.
Machine Learning, 20, 273–297. https://doi.org/10.1007/BF00994018
Cox, D. R. (1958). The regression analysis of binary sequences.
Journal of the Royal Statistical Society, Series B,
20(2), 215–242. https://doi.org/10.1111/j.2517-6161.1958.tb00292.x
Donoho, D. (2017). 50 years of data science. Journal of
Computational and Graphical Statistics, 26(4), 745–766. https://doi.org/10.1080/10618600.2017.1384734
Dunn, J. C. (1973). A fuzzy relative of the ISODATA process and its use
in detecting compact well-separated clusters. Journal of
Cybernetics, 3(3), 32–57. https://doi.org/10.1080/01969727308546046
Fisher, R. A. (1936). The use of multiple measurements in taxonomic
problems. Annals of Eugenics, 7(2), 179–188. https://doi.org/10.1111/j.1469-1809.1936.tb02137.x
Freund, Y., & Schapire, R. E. (1997). A decision-theoretic
generalization of on-line learning and an application to boosting.
Journal of Computer and System Sciences, 55(1),
119–139. https://doi.org/10.1006/jcss.1997.1504
Friedman, J. H. (2001). Greedy function approximation: A gradient
boosting machine. The Annals of Statistics, 29(5),
1189–1232. https://doi.org/10.1214/aos/1013203451
Friedman, J., Hastie, T., & Tibshirani, R. (2010). Regularization
paths for generalized linear models via coordinate descent. Journal
of Statistical Software, 33(1), 1–22. https://doi.org/10.18637/jss.v033.i01
Hochreiter, S., & Schmidhuber, J. (1997). Long short-term memory.
Neural Computation, 9(8), 1735–1780. https://doi.org/10.1162/neco.1997.9.8.1735
Hoerl, A. E., & Kennard, R. W. (1970). Ridge regression: Biased
estimation for nonorthogonal problems. Technometrics,
12(1), 55–67. https://doi.org/10.1080/00401706.1970.10488634
Jaccard, P. (1912). The distribution of the flora in the alpine zone.
New Phytologist, 11(2), 37–50. https://doi.org/10.1111/j.1469-8137.1912.tb05611.x
James, G., Witten, D., Hastie, T., & Tibshirani, R. (2013). An
introduction to statistical learning: With applications in
R (Vol. 103). Springer. https://doi.org/10.1007/978-1-4614-7138-7
Krogh, A., & Hertz, J. A. (1991). A simple weight decay can improve
generalization. Advances in Neural Information Processing
Systems, 4, 950–957.
LeCun, Y., Bottou, L., Bengio, Y., & Haffner, P. (1998).
Gradient-based learning applied to document recognition. Proceedings
of the IEEE, 86(11), 2278–2324. https://doi.org/10.1109/5.726791
Liakos, K. G., Busato, P., Moshou, D., Pearson, S., & Bochtis, D.
(2018). Machine learning in agriculture: A review. Sensors,
18(8), 2674. https://doi.org/10.3390/s18082674
MacQueen, J. (1967). Some methods for classification and analysis of
multivariate observations. Proceedings of the Fifth Berkeley
Symposium on Mathematical Statistics and Probability, 1,
281–297.
Panko, R. R. (1998). What we know about spreadsheet errors. Journal
of End User Computing, 10(2), 15–21. https://doi.org/10.4018/joeuc.1998040102
Pearson, K. (1901). On lines and planes of closest fit to systems of
points in space. The London, Edinburgh, and Dublin Philosophical
Magazine and Journal of Science, 2(11), 559–572. https://doi.org/10.1080/14786440109462720
R Core Team. (2024). R: A language and environment for statistical
computing. R Foundation for Statistical Computing. https://www.R-project.org/
Rabiner, L. R. (1989). A tutorial on hidden markov models and selected
applications in speech recognition. Proceedings of the IEEE,
77(2), 257–286. https://doi.org/10.1109/5.18626
Redmon, J., Divvala, S., Girshick, R., & Farhadi, A. (2016). You
only look once: Unified, real-time object detection. Proceedings of
the IEEE Conference on Computer Vision and Pattern Recognition,
779–788.
Ruder, S. (2016). An overview of gradient descent optimization
algorithms. arXiv Preprint arXiv:1609.04747.
Rumelhart, D. E., Hinton, G. E., & Williams, R. J. (1986). Learning
representations by back-propagating errors. Nature,
323(6088), 533–536. https://doi.org/10.1038/323533a0
Snedecor, G. W., & Cochran, W. G. (1989). Statistical
methods (8th ed.). Iowa State University Press.
Spearman, C. (1904). General intelligence, objectively determined and
measured. The American Journal of Psychology, 15(2),
201–292. https://doi.org/10.2307/1412107
Stone, M. (1974). Cross-validatory choice and assessment of statistical
predictions. Journal of the Royal Statistical Society, Series
B, 36(2), 111–147. https://doi.org/10.1111/j.2517-6161.1974.tb00994.x
Student. (1908). The probable error of a mean. Biometrika,
6(1), 1–25. https://doi.org/10.2307/2331554
Sutton, R. S., & Barto, A. G. (2018). Reinforcement learning: An
introduction (2nd ed.). MIT Press.
Tibshirani, R. (1996). Regression shrinkage and selection via the lasso.
Journal of the Royal Statistical Society, Series B,
58(1), 267–288. https://doi.org/10.1111/j.2517-6161.1996.tb02080.x
Tukey, J. W. (1977). Exploratory data analysis. Addison-Wesley.
Ward, J. H. (1963). Hierarchical grouping to optimize an objective
function. Journal of the American Statistical Association,
58(301), 236–244. https://doi.org/10.1080/01621459.1963.10500845
Wickham, H. (2014). Tidy data. Journal of Statistical Software,
59(10), 1–23. https://doi.org/10.18637/jss.v059.i10
Wickham, H. (2016). ggplot2: Elegant graphics for data analysis
(2nd ed.). Springer-Verlag. https://doi.org/10.1007/978-3-319-24277-4
Wirth, R., & Hipp, J. (2000). CRISP-DM: Towards a
standard process model for data mining. Proceedings of the 4th
International Conference on the Practical Applications of Knowledge
Discovery and Data Mining, 29–40.
Ziemann, M., Eren, Y., & El-Osta, A. (2016). Gene name errors are
widespread in the scientific literature. Genome Biology,
17, 177. https://doi.org/10.1186/s13059-016-1044-7