OPTIMIZING TERADATA, HIVE SQL, AND PYSPARK FOR ENTERPRISE-SCALE FINANCIAL WORKLOADS WITH DISTRIBUTED AND PARALLEL COMPUTING
Keywords:
Distributed Computing, Parallel Processing, Teradata, Hive SQL, PySpark, Financial, Workloads, Big Data Optimization, Enterprise Analytics]Abstract
Financial organizations deal with large amounts of information on transactions, markets and risksthat must be sorted through rapidly and correctly. This research aims to discover how Teradata,Hive SQL and PySpark
References
• Akhund, S. (n.d.). Computing infrastructure and data pipeline for enterprise-scale data preparation.
• Chang, B. R., Tsai, H. F., & Lee, Y. D. (2018). Integrated high-performance platform for fast query response in big data with Hive, Impala, and SparkSQL: A performance evaluation. Applied Sciences, 8(9), 1514. https://doi.org/10.3390/app8091514


