{"id":8956,"date":"2026-09-21T06:09:23","date_gmt":"2026-09-21T06:09:23","guid":{"rendered":"https:\/\/prwatech.in\/blog\/?p=8956"},"modified":"2026-09-22T16:09:18","modified_gmt":"2026-09-22T16:09:18","slug":"working-with-dataproc-in-console","status":"publish","type":"post","link":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/","title":{"rendered":"Working with dataproc in console"},"content":{"rendered":"\r\n<h1>Submitting a PySpark Job Through the Dataproc Console<\/h1>\r\n<p>Once a script works, you have two ways to actually run it against a Dataproc cluster: from the command line with spark-submit, or by submitting it as a job through the console. This covers the console route, writing a small PySpark script that counts movie ratings, uploading it to Cloud Storage, then submitting it as a job against an existing cluster and watching it run.<\/p>\r\n<p>This assumes you already have a Dataproc cluster running and the sample data already loaded into HDFS on that cluster, covered in the Working with Dataproc guide in this series.<\/p>\r\n<h2>Before You Start: Getting the Sample Data Onto Your Cluster<\/h2>\r\n<p>If you have not already done this, SSH into your cluster and run the following to create a folder and load the sample ratings file into it:<\/p>\r\n<p>hadoop fs -mkdir \/user\/YOUR_USER_ID\/sparkdata<br \/>hadoop fs -put u.data sparkdata<br \/>hadoop fs -ls sparkdata<\/p>\r\n<p>The last command confirms the file actually landed where the script below expects to find it.<\/p>\r\n<h2>A complete Step by Step Process of Submitting a PySpark Job Through the Dataproc Console<\/h2>\r\n<h3>Step One: Write the Script<\/h3>\r\n<p>In a plain text editor on your own machine, paste the following:<\/p>\r\n<p>from pyspark import SparkConf, SparkContext<br \/>import collections<br \/><br \/>conf = SparkConf().setMaster(&#8220;local&#8221;).setAppName(&#8220;Ratings&#8221;)<br \/>sc = SparkContext(conf = conf)<br \/><br \/>lines = sc.textFile(&#8220;\/user\/YOUR_USER_ID\/sparkdata\/u.data&#8221;)<br \/>ratings = lines.map(lambda x: x.split()[2])<br \/>result = ratings.countByValue()<br \/><br \/>sortedResults = collections.OrderedDict(sorted(result.items()))<br \/>for key, value in sortedResults.items():<br \/>\u00a0\u00a0\u00a0\u00a0print(&#8220;%s %i&#8221; % (key, value))<\/p>\r\n<p>This counts how many ratings fall into each score, from the sample data you loaded into HDFS.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"628\" height=\"319\" class=\"wp-image-8957\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png 628w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320-300x152.png 300w\" sizes=\"auto, (max-width: 628px) 100vw, 628px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Two: Find Your User ID<\/h3>\r\n<p>Open Cloud Shell and run:<\/p>\r\n<p>whoami<\/p>\r\n<p>This gives you your username directly. Running pwd instead also reveals it, since Cloud Shell&#8217;s home directory path includes your username, but whoami is the more direct way to get it.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"546\" height=\"90\" class=\"wp-image-8959\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-322.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-322.png 546w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-322-300x49.png 300w\" sizes=\"auto, (max-width: 546px) 100vw, 546px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Three: Save the File<\/h3>\r\n<p>Save it as ratingscounter.py, replacing YOUR_USER_ID in the script with the actual ID from the previous step.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"628\" height=\"274\" class=\"wp-image-8958\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-321.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-321.png 628w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-321-300x131.png 300w\" sizes=\"auto, (max-width: 628px) 100vw, 628px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Four: Upload the File to a Bucket<\/h3>\r\n<p>Open the console, then<\/p>\r\n\r\n\r\n\r\n<p>Open Cloud Storage &gt; Browser<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"379\" height=\"261\" class=\"wp-image-8960\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-323.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-323.png 379w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-323-300x207.png 300w\" sizes=\"auto, (max-width: 379px) 100vw, 379px\" \/><\/figure>\r\n\r\n\r\n\r\n<p>Upload the script into a bucket, then click it.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"489\" height=\"293\" class=\"wp-image-8961\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-324.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-324.png 489w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-324-300x180.png 300w\" sizes=\"auto, (max-width: 489px) 100vw, 489px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Five: Copy the File&#8217;s URI<\/h3>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"628\" height=\"391\" class=\"wp-image-8962\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-325.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-325.png 628w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-325-300x187.png 300w\" sizes=\"auto, (max-width: 628px) 100vw, 628px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Six: Open Dataproc Jobs<\/h3>\r\n<p>Open the menu, then Dataproc, then Jobs.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"451\" height=\"350\" class=\"wp-image-8963\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-326.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-326.png 451w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-326-300x233.png 300w\" sizes=\"auto, (max-width: 451px) 100vw, 451px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Seven: Submit a Job<\/h3>\r\n<p>Click Submit Job.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"338\" height=\"71\" class=\"wp-image-8964\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-327.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-327.png 338w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-327-300x63.png 300w\" sizes=\"auto, (max-width: 338px) 100vw, 338px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Eight: Configure the Job<\/h3>\r\n<p>Give the job an ID. The region fills in automatically, and you choose which cluster to run it against.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"532\" height=\"241\" class=\"wp-image-8965\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-328.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-328.png 532w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-328-300x136.png 300w\" sizes=\"auto, (max-width: 532px) 100vw, 532px\" \/><\/figure>\r\n\r\n\r\n\r\n<p>Set the job type to PySpark.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"547\" height=\"283\" class=\"wp-image-8966\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-329.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-329.png 547w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-329-300x155.png 300w\" sizes=\"auto, (max-width: 547px) 100vw, 547px\" \/><\/figure>\r\n\r\n\r\n\r\n<p>Paste the script&#8217;s URI into the main Python file field.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"539\" height=\"387\" class=\"wp-image-8967\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-330.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-330.png 539w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-330-300x215.png 300w\" sizes=\"auto, (max-width: 539px) 100vw, 539px\" \/><\/figure>\r\n\r\n\r\n\r\n<h3>Step Nine: Submit and Review<\/h3>\r\n<p>Click Submit.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"185\" height=\"322\" class=\"wp-image-8968\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-331.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-331.png 185w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-331-172x300.png 172w\" sizes=\"auto, (max-width: 185px) 100vw, 185px\" \/><\/figure>\r\n\r\n\r\n\r\n<p>The job runs and returns its result.<\/p>\r\n\r\n\r\n\r\n<figure class=\"wp-block-image size-large\"><img loading=\"lazy\" decoding=\"async\" width=\"628\" height=\"408\" class=\"wp-image-8969\" src=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-332.png\" alt=\"\" srcset=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-332.png 628w, https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-332-300x195.png 300w\" sizes=\"auto, (max-width: 628px) 100vw, 628px\" \/><\/figure>\r\n\r\n\r\n\r\n<h2>Consider Cloud Storage Instead of HDFS for New Work<\/h2>\r\n<p>This example reads from HDFS, which lives on the cluster itself, and disappears the moment you delete that cluster, the same tradeoff covered in the Dataproc Metastore guide in this series. For anything beyond a quick exercise, reading source data directly from a Cloud Storage path instead, using a gs colon slash slash URI in place of an HDFS path, means the data survives independently of any specific cluster, and multiple clusters can read the same source without copying it around first.<\/p>\r\n<h2>Common Mistakes to Avoid<\/h2>\r\n<ul>\r\n<li>Submitting the job before loading the sample data into HDFS. The script has nothing to read and will fail immediately.<\/li>\r\n<li>Leaving the placeholder user ID in the script&#8217;s file path. It needs to match your actual username exactly, or the path will not resolve.<\/li>\r\n<li>Relying on pwd out of habit to find your username. It works here because of how Cloud Shell&#8217;s home directory happens to be structured, but whoami is the direct, reliable way to get it.<\/li>\r\n<li>Storing important source data only in HDFS. It disappears with the cluster, while a Cloud Storage path survives independently.<\/li>\r\n<\/ul>\r\n<p>That covers submitting a PySpark job through the Dataproc console, including the HDFS setup step easy to miss on the way here. To go further, explore <a href=\"https:\/\/prwatech.in\/gcp-training-institutes-in-bangalore\/\"><strong><b>Prwatech&#8217;s Google Cloud training<\/b><\/strong>\u00a0program<\/a>, which includes placement assistance.<\/p>\r\n","protected":false},"excerpt":{"rendered":"<p>Submitting a PySpark Job Through the Dataproc Console Once a script works, you have two ways to actually run it against a Dataproc cluster: from the command line with spark-submit, or by submitting it as a job through the console. This covers the console route, writing a small PySpark script that counts movie ratings, uploading [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[1634,1],"tags":[1415,1412,1413,1414,1411,605,699,700,984,617,683,684,685,611,1400,692],"class_list":["post-8956","post","type-post","status-publish","format-standard","hentry","category-dataproc","category-google-cloud-platform","tag-dataproc","tag-dataproc-cluster","tag-dataproc-cluster-creation","tag-dataproc-cluster-properties","tag-dataproc-in-gcp","tag-gcp","tag-gcp-certification","tag-gcp-cloud-console","tag-gcp-course","tag-google-cloud","tag-google-cloud-certification","tag-google-cloud-console","tag-google-cloud-courses","tag-google-cloud-platform","tag-google-cloud-platform-tutorial","tag-google-cloud-training"],"acf":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v25.7 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>Submitting a PySpark Job Through Dataproc Console - Prwatech<\/title>\n<meta name=\"description\" content=\"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Submitting a PySpark Job Through Dataproc Console - Prwatech\" \/>\n<meta property=\"og:description\" content=\"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process\" \/>\n<meta property=\"og:url\" content=\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/\" \/>\n<meta property=\"og:site_name\" content=\"Prwatech\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/prwatech.in\/\" \/>\n<meta property=\"article:published_time\" content=\"2026-09-21T06:09:23+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-09-22T16:09:18+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png\" \/>\n\t<meta property=\"og:image:width\" content=\"628\" \/>\n\t<meta property=\"og:image:height\" content=\"319\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"author\" content=\"Prwatech\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@Eduprwatech\" \/>\n<meta name=\"twitter:site\" content=\"@Eduprwatech\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Prwatech\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"6 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":\"WebPage\",\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/\",\"url\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/\",\"name\":\"Submitting a PySpark Job Through Dataproc Console - Prwatech\",\"isPartOf\":{\"@id\":\"https:\/\/prwatech.in\/blog\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage\"},\"image\":{\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage\"},\"thumbnailUrl\":\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png\",\"datePublished\":\"2026-09-21T06:09:23+00:00\",\"dateModified\":\"2026-09-22T16:09:18+00:00\",\"author\":{\"@id\":\"https:\/\/prwatech.in\/blog\/#\/schema\/person\/db90baff7744090b2288bbc98fea87f3\"},\"description\":\"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process\",\"breadcrumb\":{\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage\",\"url\":\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png\",\"contentUrl\":\"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png\",\"width\":628,\"height\":319},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/prwatech.in\/blog\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Working with dataproc in console\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/prwatech.in\/blog\/#website\",\"url\":\"https:\/\/prwatech.in\/blog\/\",\"name\":\"Prwatech\",\"description\":\"Share Ideas, Start Something Good.\",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/prwatech.in\/blog\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Person\",\"@id\":\"https:\/\/prwatech.in\/blog\/#\/schema\/person\/db90baff7744090b2288bbc98fea87f3\",\"name\":\"Prwatech\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/prwatech.in\/blog\/#\/schema\/person\/image\/\",\"url\":\"https:\/\/secure.gravatar.com\/avatar\/c00bafc1b04045f31eda917de39891456c44fa47c092b9bb6be0f860a3a30a2f?s=96&d=mm&r=g\",\"contentUrl\":\"https:\/\/secure.gravatar.com\/avatar\/c00bafc1b04045f31eda917de39891456c44fa47c092b9bb6be0f860a3a30a2f?s=96&d=mm&r=g\",\"caption\":\"Prwatech\"},\"url\":\"https:\/\/prwatech.in\/blog\/author\/prwatech123\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"Submitting a PySpark Job Through Dataproc Console - Prwatech","description":"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/","og_locale":"en_US","og_type":"article","og_title":"Submitting a PySpark Job Through Dataproc Console - Prwatech","og_description":"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process","og_url":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/","og_site_name":"Prwatech","article_publisher":"https:\/\/www.facebook.com\/prwatech.in\/","article_published_time":"2026-09-21T06:09:23+00:00","article_modified_time":"2026-09-22T16:09:18+00:00","og_image":[{"width":628,"height":319,"url":"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png","type":"image\/png"}],"author":"Prwatech","twitter_card":"summary_large_image","twitter_creator":"@Eduprwatech","twitter_site":"@Eduprwatech","twitter_misc":{"Written by":"Prwatech","Est. reading time":"6 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/","url":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/","name":"Submitting a PySpark Job Through Dataproc Console - Prwatech","isPartOf":{"@id":"https:\/\/prwatech.in\/blog\/#website"},"primaryImageOfPage":{"@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage"},"image":{"@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage"},"thumbnailUrl":"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png","datePublished":"2026-09-21T06:09:23+00:00","dateModified":"2026-09-22T16:09:18+00:00","author":{"@id":"https:\/\/prwatech.in\/blog\/#\/schema\/person\/db90baff7744090b2288bbc98fea87f3"},"description":"Submit a PySpark job to a Dataproc cluster through the console, including the HDFS setup step the original page assumes already happened. Step by step process","breadcrumb":{"@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#primaryimage","url":"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png","contentUrl":"https:\/\/prwatech.in\/blog\/wp-content\/uploads\/2021\/05\/image-320.png","width":628,"height":319},{"@type":"BreadcrumbList","@id":"https:\/\/prwatech.in\/blog\/google-cloud-platform\/dataproc\/working-with-dataproc-in-console\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/prwatech.in\/blog\/"},{"@type":"ListItem","position":2,"name":"Working with dataproc in console"}]},{"@type":"WebSite","@id":"https:\/\/prwatech.in\/blog\/#website","url":"https:\/\/prwatech.in\/blog\/","name":"Prwatech","description":"Share Ideas, Start Something Good.","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/prwatech.in\/blog\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Person","@id":"https:\/\/prwatech.in\/blog\/#\/schema\/person\/db90baff7744090b2288bbc98fea87f3","name":"Prwatech","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/prwatech.in\/blog\/#\/schema\/person\/image\/","url":"https:\/\/secure.gravatar.com\/avatar\/c00bafc1b04045f31eda917de39891456c44fa47c092b9bb6be0f860a3a30a2f?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/c00bafc1b04045f31eda917de39891456c44fa47c092b9bb6be0f860a3a30a2f?s=96&d=mm&r=g","caption":"Prwatech"},"url":"https:\/\/prwatech.in\/blog\/author\/prwatech123\/"}]}},"_links":{"self":[{"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/posts\/8956","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/comments?post=8956"}],"version-history":[{"count":6,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/posts\/8956\/revisions"}],"predecessor-version":[{"id":11843,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/posts\/8956\/revisions\/11843"}],"wp:attachment":[{"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/media?parent=8956"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/categories?post=8956"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/prwatech.in\/blog\/wp-json\/wp\/v2\/tags?post=8956"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}