NetBackup™ for Hadoop Administrator's Guide

Last Published:
Product(s): NetBackup & Alta Data Protection (10.3)
  1. Introduction
    1.  
      Protecting Hadoop data using NetBackup
    2.  
      Backing up Hadoop data
    3.  
      Restoring Hadoop data
    4.  
      NetBackup for Hadoop terms
    5.  
      Limitations
  2. Prerequisites and best practices for the Hadoop plug-in for NetBackup
    1.  
      About deploying the Active Directory plug-in
    2. Prerequisites for the Hadoop plug-in
      1.  
        Operating system and platform compatibility
      2.  
        License for Hadoop plug-in for NetBackup
    3.  
      Preparing the Hadoop cluster
    4.  
      Best practices for deploying the Hadoop plug-in
  3. Configuring NetBackup for Hadoop
    1.  
      About configuring NetBackup for Hadoop
    2. Managing backup hosts
      1.  
        Including a NetBackup client on NetBackup primary server allowed list
      2.  
        Configure a NetBackup Appliance as a backup host
    3.  
      Adding Hadoop credentials in NetBackup
    4. Configuring the Hadoop plug-in using the Hadoop configuration file
      1.  
        Configuring NetBackup for a highly-available Hadoop cluster
      2.  
        Configuring a custom port for the Hadoop cluster
      3.  
        Configuring number of threads for backup hosts
      4.  
        Configuring number of streams for backup hosts
      5.  
        Configuring distribution algorithm and golden ratio for backup hosts
      6. Configuring communication between NetBackup and Hadoop clusters that are SSL-enabled (HTTPS)
        1.  
          ECA_TRUST_STORE_PATH for NetBackup servers and clients
        2.  
          ECA_CRL_PATH for NetBackup servers and clients
        3.  
          HADOOP_SECURE_CONNECT_ENABLED for servers and clients
        4.  
          HADOOP_CRL_CHECK for NetBackup servers and clients
        5.  
          Example values for the parameters in the bp.conf file
    5.  
      Configuration for a Hadoop cluster that uses Kerberos
    6.  
      Hadoop.conf configuration for parallel restore
    7.  
      Create a BigData policy for Hadoop clusters
    8.  
      Disaster recovery of a Hadoop cluster
  4. Performing backups and restores of Hadoop
    1. About backing up a Hadoop cluster
      1.  
        Prerequisites for running backup and restore operations for a Hadoop cluster with Kerberos authentication
      2.  
        Best practices for backing up a Hadoop cluster
      3.  
        Backing up a Hadoop cluster
    2. About restoring a Hadoop cluster
      1.  
        Best practices for restoring a Hadoop cluster
      2. Restoring Hadoop data on the same Hadoop cluster
        1.  
          Restore Hadoop data on the same Hadoop cluster
      3.  
        Restoring Hadoop data on an alternate Hadoop cluster
    3.  
      Best practice for improving performance during backup and restore
  5. Troubleshooting
    1.  
      About troubleshooting NetBackup for Hadoop issues
    2.  
      About NetBackup for Hadoop debug logging
    3. Troubleshooting backup issues for Hadoop data
      1.  
        Backup operation fails with error 6609
      2.  
        Backup operation failed with error 6618
      3.  
        Backup operation fails with error 6647
      4.  
        Extended attributes (xattrs) and Access Control Lists (ACLs) are not backed up or restored for Hadoop
      5.  
        Backup operation fails with error 6654
      6.  
        Backup operation fails with bpbrm error 8857
      7.  
        Backup operation fails with error 6617
      8.  
        Backup operation fails with error 6616
      9.  
        Backup operation fails with error 84
      10.  
        NetBackup configuration and certificate files do not persist after the container-based NetBackup appliance restarts
      11.  
        Unable to see incremental backup images during restore even though the images are seen in the backup image selection
      12.  
        One of the child backup jobs goes in a queued state
    4. Troubleshooting restore issues for Hadoop data
      1.  
        Restore fails with error code 2850
      2.  
        NetBackup restore job for Hadoop completes partially
      3.  
        Extended attributes (xattrs) and Access Control Lists (ACLs) are not backed up or restored for Hadoop
      4.  
        Restore operation fails when Hadoop plug-in files are missing on the backup host
      5.  
        Restore fails with bpbrm error 54932
      6.  
        Restore operation fails with bpbrm error 21296
      7.  
        Hadoop with Kerberos restore job fails with error 2850
      8.  
        Configuration file is not recovered after a disaster recovery
  6.  
    Index

Best practice for improving performance during backup and restore

Performance issues such as slow throughput and high CPU usage are observed during the backup and recovery of Hadoop using the SSL environment (HTTPS). The issue is caused if the internal communications in Hadoop are not encrypted. The HDFS configurations must be tuned correctly in the HDFS cluster to improve the internal communication and performance in Hadoop, which can also improve the backup and recovery performance.

  • For a better backup and restore performance, NetBackup recommended to follow the Hadoop configuration recommendations from Apache or Hadoop distributions in use.

  • If you have Hadoop encryption turned on within the cluster, follow the recommendations from Apache or Hadoop distributions in use to select the right cipher and bit length for data transfer within Hadoop cluster.

  • NetBackup performs better during backup and recovery when AES 128 is used for data encryption during the block data transfer.

  • You can also increase the number of backup hosts in case of backup to get a better performance; when you have more than one folder to be backed up in the Hadoop cluster. You can have maximum one backup host per folder in the Hadoop cluster to get the maximum benefit.

  • You can also increase the number of threads per backup host that are used to fetch data from the Hadoop cluster by NetBackup during backup operation. If you have files with the size in the range of tens of GBs, then you can increase the number of threads for better performance. The default number for threads is 4.

  • You can also increase the number of streams per backup host that are used for parallel streaming.

  • You can choose any one of the data distribution algorithms best suited for your deployment:

    • For small number of large files in your data set, use distribution algorithm 1.

    • For large number of small sized files in your data set, use distribution algorithm 2.

    • For a mix of small number of very large sized files and large number of small sized files in your data set, use the appropriate combination of distribution algorithm and golden ratio. See the example below:

Table: Example for large number of small files and small number of large file case

Data size

Number of backup hosts

Number of threads

Number of streams

Distribution algorithm

Golden ratio

Upto 1 TB

4

16

5

4

80

Upto 50TB

5

32

5

4

80

>50TB

6

32

5

4

80

For more details, refer Apache Hadoop documentation for secure mode.

Additionally for optimal performance, ensure the following:

  • Primary server is not used as a backup host.

  • In case of multiple policies scheduled to be triggered in parallel:

    • Avoid using the same discovery host in all policies.

  • The last Backup_Host entry is different for these policies.

    Note:

    Discovery host is the last entry in the Backup_Host list.