hive 2.1

一问题

最近有一个场景，要向一个表的多个分区写数据，为了缩短执行时间，采用并发的方式，多个sql同时执行，分别写不同的分区，同时开启动态分区：

set hive.exec.dynamic.partition=true

insert overwrite table test_table partition(dt) select * from test_table_another where dt = 1;

结果发现只有1个sql运行，其他sql都会卡住；
查看hive thrift server线程堆栈发现请求都卡在DbTxnManager上，hive关键配置如下：

hive.support.concurrency=true
hive.txn.manager=org.apache.hadoop.hive.ql.lockmgr.DbTxnManager

配置对应的默认值及注释：

org.apache.hadoop.hive.conf.HiveConf

    HIVE_SUPPORT_CONCURRENCY("hive.support.concurrency", false,

        "Whether Hive supports concurrency control or not. \n" +

        "A ZooKeeper instance must be up and running when using zookeeper Hive lock manager "),

    HIVE_TXN_MANAGER("hive.txn.manager",

        "org.apache.hadoop.hive.ql.lockmgr.DummyTxnManager",

        "Set to org.apache.hadoop.hive.ql.lockmgr.DbTxnManager as part of turning on Hive\n" +

        "transactions, which also requires appropriate settings for hive.compactor.initiator.on,\n" +

        "hive.compactor.worker.threads, hive.support.concurrency (true), hive.enforce.bucketing\n" +

        "(true), and hive.exec.dynamic.partition.mode (nonstrict).\n" +

        "The default DummyTxnManager replicates pre-Hive-0.13 behavior and provides\n" +

        "no transactions."),

二代码分析

hive执行sql的详细过程详见：https://www.cnblogs.com/barneywill/p/10185168.html

hive中执行sql最终都会调用到Driver.run，run会调用runInternal，下面直接看runInternal代码：

org.apache.hadoop.hive.ql.Driver

  private CommandProcessorResponse runInternal(String command, boolean alreadyCompiled)

      throws CommandNeedRetryException {

...

      if (requiresLock()) {

        // a checkpoint to see if the thread is interrupted or not before an expensive operation

        if (isInterrupted()) {

          ret = handleInterruption("at acquiring the lock.");

        } else {

          ret = acquireLocksAndOpenTxn(startTxnImplicitly);

        }

...

  private boolean requiresLock() {

    if (!checkConcurrency()) {

      return false;

    }

    // Lock operations themselves don't require the lock.

    if (isExplicitLockOperation()){

      return false;

    }

    if (!HiveConf.getBoolVar(conf, ConfVars.HIVE_LOCK_MAPRED_ONLY)) {

      return true;

    }

    Queue<Task<? extends Serializable>> taskQueue = new LinkedList<Task<? extends Serializable>>();

    taskQueue.addAll(plan.getRootTasks());

    while (taskQueue.peek() != null) {

      Task<? extends Serializable> tsk = taskQueue.remove();

      if (tsk.requireLock()) {

        return true;

      }

...

  private boolean checkConcurrency() {

    boolean supportConcurrency = conf.getBoolVar(HiveConf.ConfVars.HIVE_SUPPORT_CONCURRENCY);

    if (!supportConcurrency) {

      LOG.info("Concurrency mode is disabled, not creating a lock manager");

      return false;

    }

    return true;

  }

  private int acquireLocksAndOpenTxn(boolean startTxnImplicitly) {

...

      txnMgr.acquireLocks(plan, ctx, userFromUGI);

...

runInternal会调用requiresLock判断是否需要lock，requiresLock有两个判断：

调用checkConcurrency，checkConcurrency会检查hive.support.concurrency=true才需要lock;
调用Task.requireLock，只有部分task才需要lock；

如果判断需要lock，会调用acquireLocksAndOpenTxn，acquireLocksAndOpenTxn会调用HiveTxnManager.acquireLocks来获取lock；

1）先看那些task需要lock：

org.apache.hadoop.hive.ql.parse.DDLSemanticAnalyzer

  private void analyzeAlterTablePartMergeFiles(ASTNode ast,

      String tableName, HashMap<String, String> partSpec)

      throws SemanticException {

...

      DDLWork ddlWork = new DDLWork(getInputs(), getOutputs(), mergeDesc);

      ddlWork.setNeedLock(true);

...

可见DDL操作需要；

2）再看怎样获取lock：

org.apache.hadoop.hive.ql.lockmgr.DbTxnManager

  public void acquireLocks(QueryPlan plan, Context ctx, String username) throws LockException {

    try {

      acquireLocksWithHeartbeatDelay(plan, ctx, username, 0);

...

  void acquireLocksWithHeartbeatDelay(QueryPlan plan, Context ctx, String username, long delay) throws LockException {

    LockState ls = acquireLocks(plan, ctx, username, true);

...

  LockState acquireLocks(QueryPlan plan, Context ctx, String username, boolean isBlocking) throws LockException {

...

      switch (output.getType()) {

        case DATABASE:

          compBuilder.setDbName(output.getDatabase().getName());

          break;

        case TABLE:

        case DUMMYPARTITION:   // in case of dynamic partitioning lock the table

          t = output.getTable();

          compBuilder.setDbName(t.getDbName());

          compBuilder.setTableName(t.getTableName());

          break;

        case PARTITION:

          compBuilder.setPartitionName(output.getPartition().getName());

          t = output.getPartition().getTable();

          compBuilder.setDbName(t.getDbName());

          compBuilder.setTableName(t.getTableName());

          break;

        default:

          // This is a file or something we don't hold locks for.

          continue;

      }

...

    LockState lockState = lockMgr.lock(rqstBuilder.build(), queryId, isBlocking, locks);

    ctx.setHiveLocks(locks);

    return lockState;

  }

可见当开启动态分区时，锁的粒度是DbName+TableName，这样就会导致多个sql只有1个sql可以拿到lock，其他sql只能等待；

三总结

解决问题的方式有几种：

关闭动态分区：set hive.exec.dynamic.partition=false
关闭并发：set hive.support.concurrency=false
关闭事务：set hive.txn.manager=org.apache.hadoop.hive.ql.lockmgr.DummyTxnManager

三者任选其一，推荐第1种，因为在刚才的场景下，不需要动态分区；

【原创】大叔问题定位分享（22）hive同时执行多个insert overwrite table只有1个可以执行的更多相关文章

【原创】大叔问题定位分享（21）spark执行insert overwrite非常慢，比hive还要慢
最近把一些sql执行从hive改到spark,发现执行更慢,sql主要是一些insert overwrite操作,从执行计划看到,用到InsertIntoHiveTable spark-sql> ...
【原创】大叔问题定位分享（15）spark写parquet数据报错ParquetEncodingException: empty fields are illegal, the field should be ommited completely instead
spark 2.1.1 spark里执行sql报错 insert overwrite table test_parquet_table select * from dummy 报错如下: org.ap ...
hive INSERT OVERWRITE table could not be cleaned up.
create table maats.account_channel ROW FORMAT DELIMITED FIELDS TERMINATED BY '^' STORED AS TEXTFILE ...
【原创】大叔问题定位分享（18）beeline连接spark thrift有时会卡住
spark 2.1.1 beeline连接spark thrift之后,执行use database有时会卡住,而use database 在server端对应的是 setCurrentDatabas ...
【原创】大叔问题定位分享（16）spark写数据到hive外部表报错ClassCastException: org.apache.hadoop.hive.hbase.HiveHBaseTableOutputFormat cannot be cast to org.apache.hadoop.hive.ql.io.HiveOutputFormat
spark 2.1.1 spark在写数据到hive外部表(底层数据在hbase中)时会报错 Caused by: java.lang.ClassCastException: org.apache.h ...
【原创】大叔问题定位分享（31）hive metastore报错
hive metastore在建表时报错 [pool-5-thread-2]: MetaException(message:Got exception: java.net.ConnectExcepti ...
【原创】大叔问题定位分享（13）HBase Region频繁下线
问题现象:hive执行sql报错 select count(*) from test_hive_table; 报错 Error: java.io.IOException: org.apache.had ...
【原创】大叔问题定位分享（30）mesos agent启动失败：Failed to perform recovery: Incompatible agent info detected
mesos agent启动失败,报错如下: Feb 15 22:03:18 server1.bj mesos-slave[1190]: E0215 22:03:18.622994 1192 slave ...
【原创】大叔问题定位分享（28）openssh升级到7.4之后ssh跳转异常
服务器集群之间忽然ssh跳转不通 # ssh 192.168.0.1The authenticity of host '192.168.0.1 (192.168.0.1)' can't be esta ...

随机推荐

Java 开发笔记2
Java获取参数名称 https://blog.csdn.net/z69183787/article/details/81117525 DefaultParameterNameDiscoverer() ...
jeecg开发环境搭建
Maven安装步骤见:https://www.cnblogs.com/dyh004/p/8523260.html 修改Maven仓库 1.修改maven仓库存放位置修改maven仓库存放位置:找到 ...
fisher线性判别
fisher 判决方式是监督学习,在新样本加入之前,已经有了原样本. 原样本是训练集,训练的目的是要分类,也就是要找到分类线.一刀砍成两半! 当样本集确定的时候,分类的关键就在于如何砍下这一刀! 若以 ...
Redhat6.4安装Oracle 11gr2 64位注意事项
安装步骤略, 安装步骤参考:https://www.cnblogs.com/jhlong/p/5442459.html 注意的是,会出现找不到一些依赖库,我根据光盘已有的库安装了所有64位的依赖库,强 ...
控制结构(3): 状态机（state machine）
// 上一篇:卫语句(guard clause) // 下一篇:局部化(localization) 基于语言提供的基本控制结构,更好地组织和表达程序,需要良好的控制结构. 前情回顾上次分析了guar ...
第六章· Redis高可用sentinel
sentinel介绍 sentinel实战及配置讲解 sentinel介绍什么是sentinel? Redis-Sentinel是Redis官方推荐的高可用性(HA)解决方案,当用Redis做Mas ...
sgu438-The_Glorious_Karlutka_River
Description SGU似乎死了... 题目搬到了Codeforces... Problem - 99999438 - Codeforces Solution 动态最大流. 考虑如果不求时间, ...
python提取浏览器Cookie
在用浏览器进行网页访问时,会向网页所在的服务器发送http协议的GET或者POST等请求,在请求中除了指定所请求的方法以及URI之外,后面还跟随着一段Request Header.Request He ...
电脑装windows和ubuntu，如何卸载ubuntu系统
电脑装windows和ubuntu,如何卸载ubuntu系统 2018年01月17日 16:28:29 职业炮灰阅读数:684 版权声明:本文为博主原创文章,未经博主允许不得转载. https ...
BZOJ4671异或图
题目描述定义两个结点数相同的图 G1 与图 G2 的异或为一个新的图 G, 其中如果 (u, v) 在 G1 与 G2 中的出现次数之和为 1, 那么边 (u, v) 在 G 中, 否则这条边不在 ...

【原创】大叔问题定位分享（22）hive同时执行多个insert overwrite table只有1个可以执行

一 问题

二 代码分析

三 总结

【原创】大叔问题定位分享（22）hive同时执行多个insert overwrite table只有1个可以执行的更多相关文章

随机推荐

热门专题

一问题

二代码分析

三总结