背景

前段时间我选用了 Airflow 对 wms 进行数据归档,在运行一段时间后,经常发现会报以下错误:

[-- ::,: WARNING/ForkPoolWorker-] Failed operation _store_result.  Retrying  more times.
Traceback (most recent call last):
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in _execute_context
self.dialect.do_execute(
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/default.py", line , in do_execute
cursor.execute(statement, parameters)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in execute
self.errorhandler(self, exc, value)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/connections.py", line , in defaulterrorhandler
raise errorvalue
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in execute
res = self._query(query)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in _query
db.query(q)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/connections.py", line , in query
_mysql.connection.query(self, query)
_mysql_exceptions.OperationalError: (, 'MySQL server has gone away') The above exception was the direct cause of the following exception: Traceback (most recent call last):
File "/usr/local/python38/lib/python3.8/site-packages/celery/backends/database/__init__.py", line , in _inner
return fun(*args, **kwargs)
File "/usr/local/python38/lib/python3.8/site-packages/celery/backends/database/__init__.py", line , in _store_result
task = list(session.query(Task).filter(Task.task_id == task_id))
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/orm/query.py", line , in __iter__
return self._execute_and_instances(context)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/orm/query.py", line , in _execute_and_instances
result = conn.execute(querycontext.statement, self._params)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in execute
return meth(self, multiparams, params)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/sql/elements.py", line , in _execute_on_connection
return connection._execute_clauseelement(self, multiparams, params)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in _execute_clauseelement
ret = self._execute_context(
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in _execute_context
self._handle_dbapi_exception(
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in _handle_dbapi_exception
util.raise_from_cause(sqlalchemy_exception, exc_info)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/util/compat.py", line , in raise_from_cause
reraise(type(exception), exception, tb=exc_tb, cause=cause)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/util/compat.py", line , in reraise
raise value.with_traceback(tb)
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/base.py", line , in _execute_context
self.dialect.do_execute(
File "/usr/local/python38/lib/python3.8/site-packages/sqlalchemy/engine/default.py", line , in do_execute
cursor.execute(statement, parameters)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in execute
self.errorhandler(self, exc, value)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/connections.py", line , in defaulterrorhandler
raise errorvalue
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in execute
res = self._query(query)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/cursors.py", line , in _query
db.query(q)
File "/usr/local/python38/lib/python3.8/site-packages/MySQLdb/connections.py", line , in query
_mysql.connection.query(self, query)
sqlalchemy.exc.OperationalError: (_mysql_exceptions.OperationalError) (, 'MySQL server has gone away')
[SQL: SELECT celery_taskmeta.id AS celery_taskmeta_id, celery_taskmeta.task_id AS celery_taskmeta_task_id, celery_taskmeta.status AS celery_taskmeta_status, celery_tas
kmeta.result AS celery_taskmeta_result, celery_taskmeta.date_done AS celery_taskmeta_date_done, celery_taskmeta.traceback AS celery_taskmeta_traceback
FROM celery_taskmeta
WHERE celery_taskmeta.task_id = %s]
[parameters: ('e909b916-4284-47c4-bc5b-321bc32eb9f9',)]
(Background on this error at: http://sqlalche.me/e/e3q8)

解决过程

查了下资料一般情况下数据库服务器断开连接后,被连接池未收回将会导致以下错误:

MySQL server has gone away

所以看了下 sqlalchemy 的配置:

sql_alchemy_pool_enabled = True

# The SqlAlchemy pool size is the maximum number of database connections
# in the pool. indicates no limit.
sql_alchemy_pool_size = # The maximum overflow size of the pool.
# When the number of checked-out connections reaches the size set in pool_size,
# additional connections will be returned up to this limit.
# When those additional connections are returned to the pool, they are disconnected and discarded.
# It follows then that the total number of simultaneous connections the pool will allow is pool_size + max_overflow,
# and the total number of "sleeping" connections the pool will allow is pool_size.
# max_overflow can be set to - to indicate no overflow limit;
# no limit will be placed on the total number of concurrent connections. Defaults to .
sql_alchemy_max_overflow = # The SqlAlchemy pool recycle is the number of seconds a connection
# can be idle in the pool before it is invalidated. This config does
# not apply to sqlite. If the number of DB connections is ever exceeded,
# a lower config value will allow the system to recover faster.
sql_alchemy_pool_recycle = # Check connection at the start of each connection pool checkout.
# Typically, this is a simple statement like “SELECT ”.
# More information here: https://docs.sqlalchemy.org/en/13/core/pooling.html#disconnect-handling-pessimistic
sql_alchemy_pool_pre_ping = True sql_alchemy_pool_size = # The maximum overflow size of the pool.
# When the number of checked-out connections reaches the size set in pool_size,
# additional connections will be returned up to this limit.
# When those additional connections are returned to the pool, they are disconnected and discarded.
# It follows then that the total number of simultaneous connections the pool will allow is pool_size + max_overflow,
# and the total number of "sleeping" connections the pool will allow is pool_size.
# max_overflow can be set to - to indicate no overflow limit;
# no limit will be placed on the total number of concurrent connections. Defaults to .
sql_alchemy_max_overflow = # The SqlAlchemy pool recycle is the number of seconds a connection
# can be idle in the pool before it is invalidated. This config does
# not apply to sqlite. If the number of DB connections is ever exceeded,
# a lower config value will allow the system to recover faster.
sql_alchemy_pool_recycle = # Check connection at the start of each connection pool checkout.
# Typically, this is a simple statement like “SELECT ”.
# More information here: https://docs.sqlalchemy.org/en/13/core/pooling.html#disconnect-handling-pessimistic
sql_alchemy_pool_pre_ping = True

该配的都配置上了,因为我们的任务是一天跑一次,查了下数据库变量 waits_timeout 是 28800 ,所以直接改成25个小时。

到了第二天发现还是报这个错,很奇怪该配的都配上了,到底是哪里的问题?

仔细翻下报错日志:

File "/usr/local/python38/lib/python3.8/site-packages/celery/backends/database/__init__.py", line , in _store_result
task = list(session.query(Task).filter(Task.task_id == task_id))

难道 Airflow 的 sqlalchemy 配置对 celery 不生效?

翻阅下源码发现果然 Airflow 配置的 sqlalchemy 只对 Airflow 生效

app = Celery(
conf.get('celery', 'CELERY_APP_NAME'),
config_source=celery_configuration)

在继续翻阅 Celery 文档看有没有办法配置

database_short_lived_sessions Default: Disabled by default.

Short lived sessions are disabled by default. If enabled they can drastically reduce performance, especially on systems processing lots of tasks. This option is useful on low-traffic workers that experience errors as a result of cached database connections going stale through inactivity. For example, intermittent errors like (OperationalError) (2006, ‘MySQL server has gone away’) can be fixed by enabling short lived sessions. This option only affects the database backend.

文档告知通过database_short_lived_sessions 参数就可以避免这个问题,但是新的问题又来了,如何在 Airflow 中配置额外的 Celery 配置呢?

解决方案

找到以下文件拷贝到 DAGS 目录下,重新命名为 my_celery_config 随便起

Python/Python37/site-packages/airflow/config_templates/default_celery.py
修改 Airflow.cfg 配置 找到 celery_config_options 将配置改为 刚才起的名字
celery_config_options = my_celery_config.DEFAULT_CELERY_CONFIG
在 my_celery_config 文件中的 DEFAULT_CELERY_CONFIG dict 中就可以随便加自己需要的 Celery 配置

Airflow 使用 Celery 时,如何添加 Celery 配置的更多相关文章

  1. Python3安装Celery模块后执行Celery命令报错

    1 Python3安装Celery模块后执行Celery命令报错 pip3 install celery # 安装正常,但是执行celery 命令的时候提示没有_ssl模块什么的 手动在Python解 ...

  2. celery 分布式异步任务框架(celery简单使用、celery多任务结构、celery定时任务、celery计划任务、celery在Django项目中使用Python脚本调用Django环境)

    一.celery简介: Celery 是一个强大的 分布式任务队列 的 异步处理框架,它可以让任务的执行完全脱离主程序,甚至可以被分配到其他主机上运行.我们通常使用它来实现异步任务(async tas ...

  3. [mark] 使用Sublime Text 2时如何将Tab配置为4个空格

    在Mac OS X系统下,Sublime Text是一款比较赞的编辑器. 作为空格党的自觉,今天mark一下使用Sublime Text 2时如何将Tab配置为4个空格: 方法来自以下两个链接: ht ...

  4. 问题.NET--win7 IIS唯一密钥属性“VALUE”设置为“DEFAULT.ASPX”时,无法添加类型为“add”的重复集合

    问题现象:.NET--win7 IIS唯一密钥属性“VALUE”设置为“DEFAULT.ASPX”时,无法添加类型为“add”的重复集合 问题处理: 内容摘要:    HTTP 错误 500.19 - ...

  5. mybatis JdbcTypeInterceptor - 运行时自动添加 jdbcType 属性

    上代码: package tk.mybatis.plugin; import org.apache.ibatis.executor.ErrorContext; import org.apache.ib ...

  6. SpringBoot运行时动态添加数据源

    此方案适用于解决springboot项目运行时动态添加数据源,非静态切换多数据源!!! 一.多数据源应用场景: 1.配置文件配置多数据源,如默认数据源:master,数据源1:salve1...,运行 ...

  7. IDEA 创建类是自动添加注释和创建方法时快速添加注释

    1.创建类是自动添加注释 /*** @Author: chiyl* @DateTime: ${DATE} ${TIME}* @Description: TODO*/2. 创建方法时快速添加注释2.1 ...

  8. eclipse启动时虚拟机初始内存配置

    eclipse启动时虚拟机初始内存配置: -Xms256M -Xmx512M -XX:PermSize=256m -XX:MaxPermSize=512m

  9. 如何设置SVN提交时强制添加注释

    windows版本: 1.新建一个名为pre-commit.bat的文件并将该文件放在创建的库文件的hooks文件夹中 2.pre-commit.bat文件的内容如下: @echo off set S ...

  10. iOS 10 (X8)上CoreData的使用(包含创建工程时未添加CoreData)

    1.在创建工程时未添加CoreData,后期想要使用CoreData则要在工程Appdelegate.h文件中添加CoreData库和CoreData中的通道类(用来管理类实例和CoreData之间的 ...

随机推荐

  1. uWSGI, Thread, time.sleep 使用问题

    下面的问题,在flask程序独立运行中,都没有问题,但是部署在 uwsgi 上表现异常: 1. 在http请求处理过程中,产出异步线程,放在线程池中,线程的启动时间有比较明显的延迟. 2. 在异步线程 ...

  2. memcached 和 redis 性能测试比对

    网上很多关于memcached 和 redis 区别的介绍,大部分都是说redis比memcached支持的数据类型多的话题,而性能比对确很少,我专门针对两者进行了性能测试比对. 测试内容如下: 两者 ...

  3. Eclipse设置自动提示代码(不用alt+/了)

    在preferences找到如图的相关位置.在输入框里把26个字母加进去,qwer...........

  4. POJ 3669 Meteor Shower BFS求最小时间

    Meteor Shower Time Limit: 1000MS   Memory Limit: 65536K Total Submissions: 31358   Accepted: 8064 De ...

  5. django-redis和redis连接

    redis连接 简单连接 import redis r = redis.Redis(host=) r.set('foo', 'Bar') print r.get('foo') 连接池 import r ...

  6. CentOS 7安装/卸载Redis,配置service服务管理

    Redis简介 Redis功能简介 Redis 是一个开源(BSD许可)的,内存中的数据结构存储系统,它可以用作数据库.缓存和消息中间件. 相比于传统的关系型数据库,Redis的存储方式是key-va ...

  7. 洛谷 P2375 [NOI2014]动物园

    题目传送门 解题思路: 其实对于一个sum[i],其值就等于sum[next[i]] + sum[next[next[i]]] + ... + 1,然后我们可以记忆化,然后题目里又有一个限制,就是前后 ...

  8. generator 和 yield

    yield 的使用 generator 生成器 yield 可以使生成器返回多次 我习惯于从表象推测,不喜欢官方文档,写的字都认识,结果变成句子之后,就一句都看不懂 所以先举一个例子来看一下这个东西怎 ...

  9. java#类的实例化顺序

    关于类的实例化,不用弄的那么细致,这里只说单一类,没有其他父类(排除Obejct)的情况.要实例化一个类,需要加载class文件到jvm并且验证通过了是安全的字节码文件. 初始化大致上是按照如下步骤: ...

  10. [Java] Eclipse 设置相同变量背景色高亮显示

    在Eclipse中,鼠标选中或者光标移动到java类的变量名时,相同变量会被标识显示(设置背景色高亮), 并且侧边滚动条会标出变量的位置, 查找变量十分方便. 1.相同变量标识高亮显示:Window ...