Managed Service for Apache Spark documentation
Managed Service for Apache Spark on clusters lets you take advantage of open source data tools for batch processing, querying, streaming, and machine learning. Managed Service for Apache Spark automation helps you create clusters quickly, manage them easily, and save money by turning clusters off when you don't need them. With less time and money spent on administration, you can focus on your jobs and your data.
Use Managed Service for Apache Spark serverless to run Spark batch workloads without provisioning and managing your own cluster. Specify workload parameters, and then submit the workload to the Managed Service for Apache Spark service. The service will run the workload on a managed compute infrastructure, autoscaling resources as needed. Managed Service for Apache Spark charges apply only to the time when the workload is executing.
Go to the Managed Service for Apache Spark product page for more.
立获300美元免费赠金
Google Cloud新用户首次注册即可获得300美元赠金。此外,无论新老用户,均可免费使用20多款产品,积累动手实践经验。
立获300美元免费赠金
Google Cloud新用户首次注册即可获得300美元赠金。此外,无论新老用户,均可免费使用20多款产品,积累动手实践经验。
Documentation resources
Related resources
Run a Spark job on Google Kubernetes Engine
Submit Spark jobs to a running Google Kubernetes Engine cluster from the Dataproc Jobs API.
Introduction to Cloud Dataproc: Hadoop and Spark on Google Cloud
This course features a combination of lectures, demos, and hands-on labs to create a Dataproc cluster, submit a Spark job, and then shut down the cluster.
Machine Learning with Spark on Dataproc
This course features a combination of lectures, demos, and hands-on labs to implement logistic regression using a machine learning library for Apache Spark running on a Dataproc cluster to develop a model for data from a multivariable dataset.
Migrate HDFS Data from On-Premises to Google Cloud
How to move data from on-premises Hadoop Distributed File System (HDFS) to Google Cloud.
Manage Java and Scala dependencies for Apache Spark
Recommended approaches to including dependencies when you submit a Spark job to a Managed Service for Apache Spark cluster.