Como executar jobs do Apache Spark no Cloud Dataproc avaliações

46105 avaliações

This code didn't work Py4JavaError on last line: from pyspark.sql import SparkSession, SQLContext, Row gcs_bucket='378837530308' spark = SparkSession.builder.appName("kdd").getOrCreate() sc = spark.sparkContext data_file = "gs://"+gcs_bucket+"//kddcup.data_10_percent.gz" raw_rdd = sc.textFile(data_file).cache() raw_rdd.take(5)

Louise O. · Revisado há about 3 years

Zaklina P. · Revisado há about 3 years

Stergios M. · Revisado há about 3 years

Zainab A. · Revisado há about 3 years

tushar s. · Revisado há about 3 years

Rhaydrick S. · Revisado há about 3 years

Rongzhi Z. · Revisado há about 3 years

Eduardo G. · Revisado há about 3 years

Provides an average understanding. Some of the button names have been updated on the tool/ dataproc compared to the instructions but was able to figure it out

Mario D. · Revisado há about 3 years

Yaozhi L. · Revisado há about 3 years

Aleksandr A. · Revisado há about 3 years

Daniel V. · Revisado há about 3 years

Fatih S. · Revisado há about 3 years

Guillaume P. · Revisado há about 3 years

complex

Maximiliano A. · Revisado há about 3 years

Ignacio V. · Revisado há about 3 years

Paweł M. · Revisado há about 3 years

Olesya S. · Revisado há about 3 years

Ajay T. · Revisado há about 3 years

Hajer D. · Revisado há about 3 years

Barış A. · Revisado há about 3 years

Gastón E. · Revisado há about 3 years

Dejan J. · Revisado há about 3 years

Yannic H. · Revisado há about 3 years

Learnt how to make use of Cloud Storage instead of HDFS to stage data

Ali H. · Revisado há about 3 years

Não garantimos que as avaliações publicadas sejam de consumidores que compraram ou usaram os produtos. As avaliações não são verificadas pelo Google.