סקירה כללית של Managed Service for Apache Spark on GKE

‫Managed Service for Apache Spark on GKE מאפשר לכם להפעיל אפליקציות של Big Data באמצעות jobs API של Managed Service for Apache Spark באשכולות GKE. משתמשים ב Google Cloud מסוף, ב-Google Cloud CLI או ב-Managed Service for Apache Spark API (בקשת HTTP או ספריות לקוח ב-Cloud) כדי ליצור אשכול וירטואלי של Managed Service for Apache Spark ב-GKE, ואז שולחים עבודת Spark,‏ PySpark,‏ SparkR או Spark-SQL לשירות Managed Service for Apache Spark.

‫Managed Service for Apache Spark on GKE תומך בגרסאות Spark 3.5.

איך פועל Managed Service for Apache Spark ב-GKE

‫Managed Service for Apache Spark on GKE פורס אשכולות וירטואליים של Managed Service for Apache Spark באשכול GKE. בניגוד ל-Managed Service for Apache Spark באשכולות של Compute Engine, ב-Managed Service for Apache Spark באשכולות וירטואליים של GKE אין מכונות וירטואליות נפרדות של צומת ראשי וצומת עובד. במקום זאת, כשיוצרים אשכול וירטואלי של Managed Service for Apache Spark ב-GKE, שירות Managed Service for Apache Spark ב-GKE יוצר מאגרי צמתים באשכול GKE. משימות של Managed Service for Apache Spark ב-GKE מופעלות כ-Pods במאגרי הצמתים האלה. מאגרי הצמתים ותזמון הפודים במאגרי הצמתים מנוהלים על ידי GKE.