Updated on 2026-09-24 GMT+08:00

ALM-16004 Hive Service Unavailable

Alarm Description

The system checks the Hive service status every 60 seconds. This alarm is generated when the Hive service is unavailable.

This alarm is cleared when the Hive service recovers.

Alarm Attributes

Alarm ID

Alarm Severity

Auto Cleared

16004

Critical

Yes

Alarm Parameters

Parameter

Description

Source

Specifies the cluster for which the alarm was generated.

ServiceName

Specifies the service for which the alarm was generated.

RoleName

Specifies the role for which the alarm was generated.

HostName

Specifies the host for which the alarm was generated.

Impact on the System

The system cannot provide data loading, query, and extraction services.

Possible Causes

  • Hive service unavailability may be related to the faults of the Hive process as well as basic services, such as ZooKeeper, Hadoop distributed file system (HDFS), YARN, and DBService.
    • The ZooKeeper service is abnormal.
    • The HDFS service is abnormal.
    • The YARN service is abnormal.
    • The DBService service is abnormal.
    • The Hive service process is faulty. If the alarm is caused by the Hive process fault, the alarm reporting may be delayed for about 5 minutes.
  • The network communication between the Hive service and basic services is interrupted.
  • The permission on the HDFS temporary directory of Hive is abnormal.
  • The local disk space of the Hive node is insufficient.

Handling Procedure

Check the HiveServer/MetaStore process status.

  1. On Manager, click Cluster > Services > Hive > Instances. On the displayed page, check whether there are HiveServer or MetaStore instances in the Unknown state.

    For details about how to log in to FusionInsight Manager, see Accessing MRS Manager.

  2. Above the Hive instance list, choose More > Restart Instance to restart the HiveServer/MetaStore process.

    The HiveServer/MetaStore instance that is being restarted cannot provide services for external systems, and the SQL tasks being executed on the instance may fail.

  3. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 4.

Check the ZooKeeper service status.

  1. In the alarm list on Manager, check whether there are Process Fault alarms.

  2. Check whether the Process Fault alarm is generated for ZooKeeper in the alarm location information.

  3. Rectify the fault by following the steps provided in ALM-12007 Process Fault.
  4. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 8.

Check the HDFS service status.

  1. In the alarm list, check whether there is an HDFS Service Unavailable alarm.

  2. Rectify the fault by following the steps provided in ALM-14000 HDFS Service Unavailable.
  3. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 11.

Check the YARN service status.

  1. In the alarm list, check whether there is a Yarn Service Unavailable alarm.

  2. Rectify the fault by following the steps provided in ALM-18000 Yarn Service Unavailable.
  3. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 14.

Check the DBService service status.

  1. In the alarm list, check whether there is a DBService Service Unavailable alarm.

  2. Rectify the fault. For details, see ALM-27001 DBService Service Unavailable.
  3. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 17.

Check the network connection between the Hive and ZooKeeper, HDFS, Yarn, and DBService.

  1. On the Manager homepage, choose Cluster > Services > Hive.
  2. Click Instances to go to the Hive instance list page.
  3. Click the host name in the row where HiveServer is located to go to the HiveServer host status page.
  4. Record the IP address in the Basic Information area.
  5. Log in to the host where HiveServer resides as user omm using the IP address obtained in Step 20.
  1. Run the ping command to check whether the network connection between the host that runs HiveServer and the hosts that run the ZooKeeper, HDFS, Yarn, and DBService services is normal. (Obtain the IP addresses of the hosts that run the ZooKeeper, HDFS, Yarn, and DBService services in the same way as that for obtaining the IP address of the HiveServer.)

  2. Contact the network administrator to restore the network.
  3. In the alarm list, check whether Hive Service Unavailable is cleared.

    • If yes, no further action is required.
    • If no, go to Step 25.

Check the permission on the HDFS temporary directory.

  1. Log in to the node where the HDFS client is located and run the following command to go to the HDFS client installation directory:

    Go to the client installation directory.

    cd Client installation directory

    Configure environment variables.

    source bigdata_env

    If Kerberos authentication is enabled for the cluster (in security mode), run the following command to authenticate the user:

    kinit User with the permissions of the supergroup group

  2. Run the following command to check whether the permission on the data warehouse directory is 770:

    hdfs dfs -ls /tmp | grep hive-scratch

  3. Run the following command to restore the default data warehouse permission:

    hdfs dfs -chmod 770 /tmp/hive-scratch

  4. Wait for several minutes and check whether the "Hive Service Unavailable" alarm is cleared.

    • If yes, no further action is required.
    • If no, go to Step 29.

Check whether the local disk space is normal.

  1. Run the following command to check whether the disk usage of the /srv, /var, and /opt directories exceeds 95% in the root directory.

    df -h

  2. Clear unnecessary information in the corresponding directory to ensure that the available disk space is greater than 80%. Wait for several minutes and check whether the "Hive Service Unavailable" alarm is cleared.

    • If yes, no further action is required.
    • If no, go to Step 31.

Collect fault information.

  1. On Manager, choose O&M. In the navigation pane on the left, choose Log > Download.
  2. In the Service area, select the following nodes of the desired cluster.

    • ZooKeeper
    • HDFS
    • Yarn
    • DBService
    • Hive

  3. Click in the upper right corner, and set Start Date and End Date for log collection to 10 minutes before and after the alarm generation time, respectively. Then, click Download.
  4. Send the collected fault logs to O&M personnel for help.

Alarm Clearance

This alarm is automatically cleared after the fault is rectified.

Related Information

None