Jurisdictional exposure
Attributes
- Auto Scaling
- Yes
- Apache Beam
- Yes
- Exactly Once Processing
- Yes
Sub-services (2)
Batch Pipelines
Large-scale batch data processing jobs
Streaming Pipelines
Real-time stream processing with auto-scaling
Compliance & Certifications
This service is attested for the following frameworks. Always verify with the provider before relying on a specific compliance posture.
Where this runs
Sovereign regions (2)
- T-Systems Sovereign Cloud · FrankfurtT-Systems Sovereign Cloud powered by Google Cloud
- S3NS Sovereign Cloud · ParisS3NS — Google Cloud + Thales joint venture
Commercial regions (42)
Europe (13)
- Belgium
- Finland
- Paris
- Berlin
- Frankfurt
- Milan
- Turin
- Netherlands
- Warsaw
- Madrid
- Stockholm
- Zurich
- London
North America (12)
- Montréal
- Toronto
- Querétaro
- Northern Virginia
- Columbus
- Iowa
- Dallas
- Las Vegas
- Los Angeles
- South Carolina
- Salt Lake City
- Oregon
South America (2)
- São Paulo
- Santiago
Asia (9)
- Hong Kong
- Delhi
- Mumbai
- Jakarta
- Osaka
- Tokyo
- Singapore
- Seoul
- Taiwan
Oceania (2)
- Melbourne
- Sydney
Middle East (3)
- Tel Aviv
- Doha
- Dammam
Africa (1)
- Johannesburg
Tags
Equivalent services on other platforms
End-to-end data development platform over MaxCompute, EMR, and Hologres with visual and SQL-based task authoring, scheduled pipelines, data quality rules, data lineage, and a built-in business-glossary catalog
Real-time streaming data collection, processing, and analytics platform with Data Streams, Data Firehose, Video Streams, and Kinesis Data Analytics sub-services
Serverless data integration platform with visual ETL authoring via Glue Studio, a Hive-compatible Data Catalog, automatic crawlers, DataBrew visual prep, and Glue Data Quality for declarative rule-based validation across Spark and Python jobs
Managed streaming ETL that captures data from Kinesis / Kafka / DirectPUT clients, buffers + optionally transforms it via Lambda, and writes to S3 / Redshift / OpenSearch / third-party sinks — no operators to manage. Formerly branded Amazon Kinesis Data Firehose
Managed Apache Flink runtime for stateful stream-processing workloads — Java, Python, Scala, and SQL applications with auto-scaling, savepoint management, and Studio notebooks for interactive development
Managed ETL and ELT service for data integration at scale with 100+ connectors, visual pipeline designer, mapping data flows, and triggers for event-driven orchestration
Real-time data stream processing and analytics engine with SQL-based queries, windowing functions, and native inputs from Event Hubs, IoT Hub, and Blob Storage
Ingest streaming event data over HTTP or from Worker bindings, transform with SQL, and write the result to R2 as Apache Iceberg tables or partitioned Parquet/JSON files for downstream warehouse and analytics tools. Open-beta as of 2026.
Fully managed orchestration service for data engineering and ML pipelines with DAG-based scheduling, dependency tracking, alerting, and multi-task jobs across notebooks, JARs, and Python scripts
Declarative ETL framework for streaming and batch pipelines on the lakehouse — define tables in SQL or Python, DLT handles dependency graph, retries, data quality, and observability
Managed ETL / ELT service for building visual pipelines that move and transform data between OCI Object Storage, Autonomous Database, ADW, and external sources — with Spark-based execution, incremental loads, and DataOps CI/CD via version-controlled workspaces
Continuous data ingestion into Snowflake from cloud storage (file-based Snowpipe) or directly from row-level streams (Snowpipe Streaming) with sub-10s latency and serverless compute