Managed Service for Apache Spark on GKE מאפשר לכם להפעיל אפליקציות של Big Data באמצעות jobs API של Managed Service for Apache Spark באשכולות GKE.
משתמשים ב Google Cloud מסוף, ב-Google Cloud CLI או ב-Managed Service for Apache Spark API (בקשת HTTP או ספריות לקוח ב-Cloud) כדי ליצור אשכול וירטואלי של Managed Service for Apache Spark ב-GKE, ואז שולחים עבודת Spark, PySpark, SparkR או Spark-SQL לשירות Managed Service for Apache Spark.
Managed Service for Apache Spark on GKE תומך בגרסאות Spark 3.5.
איך פועל Managed Service for Apache Spark ב-GKE
Managed Service for Apache Spark on GKE פורס אשכולות וירטואליים של Managed Service for Apache Spark באשכול GKE. בניגוד ל-Managed Service for Apache Spark באשכולות של Compute Engine, ב-Managed Service for Apache Spark באשכולות וירטואליים של GKE אין מכונות וירטואליות נפרדות של master ו-worker. במקום זאת, כשיוצרים אשכול וירטואלי של Managed Service for Apache Spark ב-GKE, שירות Managed Service for Apache Spark ב-GKE יוצר מאגרי צמתים באשכול GKE. משימות של Managed Service for Apache Spark ב-GKE מופעלות כ-Pods במאגרי הצמתים האלה. מאגרי הצמתים ותזמון הפודים במאגרי הצמתים מנוהלים על ידי GKE.