Dataproc

GCPAnalytics

Managed Spark and Hadoop clusters for big data processing with per-second billing, serverless Spark (Dataproc Serverless), Presto, Flink, and ephemeral or long-running clusters

Jurisdictional exposure

Provider HQ
USMountain View, USA

Subject to CLOUD Act, FISA-702, DPF

Region locations
APACCNEEAEUUKUSOther44 regions across 7 jurisdictions
Sovereign option
Yes — 2 sovereign-flagged regions available

Attributes

Auto Scaling
Yes
Spark Support
Yes
Hadoop Support
Yes

Sub-services (3)

Dataproc Clusters

Managed Spark and Hadoop cluster provisioning

Dataproc Serverless

Serverless Spark for batch workloads

Dataproc Metastore

Managed Hive Metastore for metadata management

Compliance & Certifications

This service is attested for the following frameworks. Always verify with the provider before relying on a specific compliance posture.

Where this runs

44 regions
28 countries
2sovereign
Sovereign regions (2)
  • T-Systems Sovereign Cloud · FrankfurtT-Systems Sovereign Cloud powered by Google Cloud
  • S3NS Sovereign Cloud · ParisS3NS — Google Cloud + Thales joint venture
Commercial regions (42)

Europe (13)

  • Belgium
  • Finland
  • Paris
  • Berlin
  • Frankfurt
  • Milan
  • Turin
  • Netherlands
  • Warsaw
  • Madrid
  • Stockholm
  • Zurich
  • London

North America (12)

  • Montréal
  • Toronto
  • Querétaro
  • Northern Virginia
  • Columbus
  • Iowa
  • Dallas
  • Las Vegas
  • Los Angeles
  • South Carolina
  • Salt Lake City
  • Oregon

South America (2)

  • São Paulo
  • Santiago

Asia (9)

  • Hong Kong
  • Delhi
  • Mumbai
  • Jakarta
  • Osaka
  • Tokyo
  • Singapore
  • Seoul
  • Taiwan

Oceania (2)

  • Melbourne
  • Sydney

Middle East (3)

  • Tel Aviv
  • Doha
  • Dammam

Africa (1)

  • Johannesburg

Tags

Equivalent services on other platforms

Amazon EMRAWS

Managed big-data platform for running Apache Spark, Hive, Presto, Flink, Trino, and HBase across EC2, EKS, and fully serverless deployments with up to 5x faster Spark runtime and Graviton price-performance

Azure Synapse AnalyticsAzure

Unified analytics service combining data warehousing and big data processing with dedicated and serverless SQL pools, Apache Spark, Data Integration pipelines, and Power BI embedded

Azure HDInsightAzure

Managed open-source analytics clusters for Hadoop, Apache Spark, Apache Hive LLAP, Apache Kafka, and Apache HBase with enterprise security via Enterprise Security Package, autoscale, and integration with ADLS Gen2 — used mainly for migrating existing OSS big-data estates into Azure

Huawei MapReduce ServiceHuawei

Fully managed big-data platform running Apache Spark, Hive, HBase, Flink, Hadoop, and Kudu clusters with autoscaling, Kerberos security, and integration with OBS for compute-storage separation

OVHcloud Data PlatformOVHcloud

Managed data-lakehouse stack combining Apache Spark for batch and stream processing, Iceberg table format, and integration with object-storage data lakes, targeted at sovereign-EU analytics workloads

OVHcloud Cloud AnalyticsOVHcloud

Managed analytics platform combining ingestion, transformation, and visualisation capabilities for European-resident data analytics workloads needing EU-jurisdictional sovereignty

Tencent Elastic MapReduceTencent

Managed big-data platform running Apache Spark, Hadoop, Hive, HBase, Flink, Presto, and ClickHouse with autoscaling, Kerberos authentication, integration with COS for compute-storage separation, and Jupyter notebooks for interactive analysis

Pricing

Pricing model:pay-as-you-go