Our Community is getting an upgrade! To get everything ready for the relaunch, we’ll be placing the site in read-only mode starting September 21st.
We really appreciate your understanding while we get things set up behind the scenes. Catch up on all the exciting details about the move here.
Need help or have questions? Drop us a line at [email protected]

Archives of Support Questions (Read Only)

This is an archived board for historical reference. Information and links may no longer be available or relevant
Announcements
This board is archived and read-only for historical reference. To ask a new question, please post a new topic on the appropriate active board.

adding additional volume to data nodes.

avatar
New Member

We are running hadoop HA cluster using AWS EC2 instances with 17 Data ndoes (All instances are M4.4xlarge including name nodes). All the DN's are configured with 16TB (EBS st1) volumes for hdfs.

Now we are running out of HDFS storage and looking to extend the storage. Since 16TB is max limit for st1 EBS we cannot extend the existing volume.

Trying to add additional 16TB volumes to few data nodes and update "DataNode directories" in ambari with this new volume path.

Will this approach impact any performance issue with cluster ? Any other things need be considered in this approach ?

1 ACCEPTED SOLUTION

avatar
Contributor

@Sajesh PP - As you increase the size of a data node you can run into performance problems such as to much read/write activity on a single data node. If this occurs it is better to add new data nodes with additional storage.

View solution in original post

1 REPLY 1

avatar
Contributor

@Sajesh PP - As you increase the size of a data node you can run into performance problems such as to much read/write activity on a single data node. If this occurs it is better to add new data nodes with additional storage.