Federated Learning for Predictive BI Dashboard Performance
Keywords:
Federated Learning, BI dashboard, latency prediction, edge computing, query patternsAbstract
Federation Learning (FL) decentralises model training on distributed edge devices while maintaining data locality. The aim of this research is to investigate that FL can predict Business Intelligence (BI) dashboard performance especially focusing on pre-release delay estimation. We can estimate latency with 94% accuracy by letting edge models learn how users interact with queries and only sending gradients to a central aggregator.
Downloads
References
J. Konecny, H. B. McMahan, F. X. Yu, P. Richtarik, A. T. Suresh, and D. Bacon, "Federated learning: Strategies for improving communication efficiency," arXiv preprint arXiv:1610.05492, 2016.
H. B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas, "Communication-efficient learning of deep networks from decentralized data," in Proc. 20th Int. Conf. Artificial Intelligence and Statistics (AISTATS), 2017, pp. 1273–1282.
A. K. Jain, M. Murty, and P. Flynn, "Data clustering: A review," ACM Comput. Surv., vol. 31, no. 3, pp. 264–323, Sept. 1999.
C. Dwork, "Differential privacy," in Proc. 33rd Int. Colloq. Automata, Languages and Programming (ICALP), 2006, pp. 1–12.
R. Agrawal and R. Srikant, "Privacy-preserving data mining," in Proc. 2000 ACM SIGMOD Int. Conf. Management of Data, 2000, pp. 439–450.
T. White, Hadoop: The Definitive Guide, 3rd ed. O'Reilly Media, 2012.
R. S. Montero, I. M. Llorente, and I. Foster, "Virtual infrastructures: A survey on virtualization technologies and cloud computing," Comput. Electr. Eng., vol. 38, no. 6, pp. 1367–1386, Sept. 2012.
L. Breiman, "Random forests," Machine Learning, vol. 45, no. 1, pp. 5–32, Oct. 2001.
R. Kohavi, "A study of cross-validation and bootstrap for accuracy estimation and model selection," in Proc. 14th Int. Joint Conf. Artificial Intelligence, 1995, pp. 1137–1143.
F. Chollet, "Keras," 2015.
D. Borthakur, "The Hadoop distributed file system: Architecture and design," Hadoop Project Website, vol. 11, 2007.
C. Zhang, P. S. Yu, and Y. Gong, "Accelerating large-scale machine learning with distributed computing," IEEE Trans. Knowl. Data Eng., vol. 29, no. 6, pp. 1235–1248, June 2017.
R. N. Calheiros, R. Ranjan, A. Beloglazov, C. A. F. De Rose, and R. Buyya, "CloudSim: A toolkit for modeling and simulation of cloud computing environments and evaluation of resource provisioning algorithms," Softw.—Pract. Exper., vol. 41, no. 1, pp. 23–50, Jan. 2011.
T. M. Mitchell, Machine Learning. McGraw-Hill, 1997.
H. Liu, H. Motoda, L. Yu, G. J. Williams, and Z. Zhao, Feature Selection: An Ever Evolving Frontier in Data Mining. Springer, 2010.
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction, 1st ed. MIT Press, 1998.
K. C. Laudon and J. P. Laudon, Management Information Systems: Managing the Digital Firm, 13th ed. Pearson, 2015.
M. Zaharia et al., "Apache Spark: A unified engine for big data processing," Commun. ACM, vol. 59, no. 11, pp. 56–65, Nov. 2016.
P. Domingos, "A few useful things to know about machine learning," Commun. ACM, vol. 55, no. 10, pp. 78–87, Oct. 2012.
K. Bonawitz et al., "Practical secure aggregation for privacy-preserving machine learning," in Proc. 2017 ACM SIGSAC Conf. Computer and Communications Security, 2017, pp. 1175–1191.
Downloads
Published
Issue
Section
License

This work is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License.