开始使用免费开始使用

Module 2 Quiz: Design and transformations — Question 4

You are building a new pipeline using Serverless for Apache Spark. The source data is highly structured, with well-defined columns like userid and purchaseamount. The primary task involves complex filtering and calculations on these columns. According to modern Spark best practices, which core API should you use to represent and manipulate this data?

本练习是课程的一部分

Build Batch Data Pipelines on Google Cloud

查看课程

动手互动练习

通过我们的互动练习之一,将理论转化为实践

开始练习