理解这个问题,就是pods在Kubernetes中怎么进行failover

在Kubernetes的work node上有kubelet,会负责监控该work node上的pods,如果有container挂掉了,它会负责重启

但是如果进程没有挂掉,只是hang住,或是死循环,或是死锁了,这个怎么判断

所以还需要引入,liveness probes,用于主动探测Pods是否正常

liveness probe

- An HTTP GET probe performs an HTTP GET request on the container’s IP address, a port and path you specify. If the probe receives a response, and the response code doesn’t represent an error (in other words, if the HTTP response code is 2xx or 3xx), the probe is considered successful. If the server returns an error response code or if it doesn’t respond at all, the probe is considered a failure and the container will be restarted as a result.

- A TCP Socket probe tries to open a TCP connection to the specified port of the container. If the connection is established successfully, the probe is successful. Otherwise, the container is restarted.

- An Exec probe executes an arbitrary command inside the container and checks the command’s exit status code.
If the status code is 0, the probe is successful. All other codes are considered failures.

probe分为三种,Http,Tcp,Exec

创建一个Http Probe,

这样后面,kubelet会定期主动通过定义的probe进行探测,如果probe失败就重启container

通过下面的命令看出当前pod的状态,

$ kubectl get po kubia-liveness

NAME READY STATUS RESTARTS AGE

kubia-liveness 1/1 Running 1 2m

看上一次重启的原因,

When you want to figure out why the previous container terminated, you’ll want to

see those logs instead of the current container’s logs. This can be done by using

the --previous option:

$ kubectl logs mypod --previous

也可以查看pod的详细信息,

$ kubectl describe po kubia-liveness

这里可以看到更详细的restart信息,和具体的liveness probe的命令

Liveness: http-get http://:8080/ delay=0s timeout=1s period=10s #success=1
➥ #failure=3

参数本身也比较好理解,#failure=3,判断failure要重试3次,

其中delay是个比较关键的参数,默认是0,如果你的服务需要初始化时间,很容易造成第一次probe失败,

所以可以设大些,

liveness probe,让我们更有效的探测pod或container的fail,这样kubelet可以更加有效的重启和恢复服务

但是kubelet是在work node上面,如果一个node挂了,怎么办?

这就需要ReplicationController,RC会保障他管理的pod在node间failover

ReplicationController

上面是RC的工作流程图,RC不会去迁移Pod,只是根据数目的对比决定是删除Pod,还是创建新的Pod

从图中,知道有几个要素,

我们如何知道RC管理哪些Pod?通过label selector来筛选(更改RC的label selector或是其中pod的label,都可以改变RC管理的Pod范围)

一般一个RC管理的是同一种的Pod,所以Pod的数目就是副本数,通过replica count来定义 (扩缩容)

最后需要有一个Pod模板,用于创新新的Pod (影响新创建的Pod,不会影响已经在运行的Pod)

Kubernetes是采用declarative approach的方式去管理集群,即你不要下达具体的操作指令,而只需要规定需要达到的状态,比如维护2种RC,每个并发度是10;然后Kubernetes会根据当前的实际状态做具体的操作去满足你所需要的状态

创建RC,

$ kubectl create -f kubia-rc.yaml

replicationcontroller "kubia" created

RC的效果,被删除的Pod,会被重新创建

看RC的状态,

$ kubectl get rc

NAME DESIRED CURRENT READY AGE

kubia 3 3 2 3m

$ kubectl describe rc kubia

看下,实际修改一个RC中的Pod的label,会发生什么?

$ kubectl label pod kubia-dmdck app=foo --overwrite

pod "kubia-dmdck" labeled

该pod会脱离RC的管理,RC会创建一个新的Pod来替代该Pod

可以对RC进行扩缩容,可以通过命令,也可以直接修改rc的配置文件

$ kubectl scale rc kubia --replicas=10

$ kubectl edit rc kubia

删除RC,可以选择保留Pods或不保留

$ kubectl delete rc kubia --cascade=false

replicationcontroller "kubia" deleted

ReplicaSet

新版本的Kubernetes会用replicaSet替换当前的replicationcontroller,

不同在于,首先版本不同,replicaSet属于v1beta2

主要是,selector更为灵活,虽然这里使用matchLabels和原来差不多

但可以使用,matchExpressios

DaemonSet

DaemonSet是种特殊形式,如下图,比如kubelet,就是一种典型的DaemonSet

A DaemonSet makes sure it creates as many pods as there are nodes and deploys each one on its own node

DaemonSet也可以选择部分node,通过nodeSelector

Job

可完成的,就是batch任务

对于Job有两个参数,比较关键

completions,job pod需要被执行几次

parallelism,同时有几个pod被执行

可以通过,activeDeadlineSeconds,来设定最大执行时间,超时会被关闭,算fail

CronJob

定期执行的job

关键是对,schedule的理解,

五项,代表

 Minute

 Hour

 Day of month

 Month

 Day of week.

"0,15,30,45 * * * *", which means at the 0, 15, 30 and 45 minutes mark of every hour
(first asterisk), of every day of the month (second asterisk), of every month (third
asterisk) and on every day of the week (fourth asterisk)

"0,30 * 1 * *", you wanted it to run every 30 minutes, but only on the first day of the month

"0 3 * * 0",if you want it to run at 3AM every Sunday

对于cronJob,都不可能完全精确时间点执行的,可能因为前面的任务拖延或其他问题导致,我们要设个期限,超出这次cronjob就不执行了,否则会堆积大量的cronjob

kubernetes in action - Replication Controller的更多相关文章

  1. kubernetes concepts -- Replication Controller

    Edit This Page ReplicationController NOTE: A Deployment that configures a ReplicaSet is now the reco ...

  2. kubernetes进阶之五:Replication Controller&Replica Sets&Deployments

    一:Replication Controller RC是kubernetes的核心概念之一.它定义了一个期望的场景即声明某种Pod的副本数量在任意时候都要符合某个预期值. 它由以下几个部分组成: 1. ...

  3. kubernetes 1.3管中窥豹- RS(Replica Sets):the next-generation Replication Controller

    前言 kubernates 1.3出了几个新的概念,其中包括deployments,Replica Sets,并且官网称之为是the next-generation Replication Contr ...

  4. Replication Controller、Replica Set

    假如我们现在有一个Pod正在提供线上的服务,我们来想想一下我们可能会遇到的一些场景: 某次运营活动非常成功,网站访问量突然暴增 运行当前Pod的节点发生故障了,Pod不能正常提供服务了 第一种情况,可 ...

  5. MVC路由规则以及前后台获取Action、Controller、ID名方法

    1.前后台获取Action.Controller.ID名方法 前台页面:ViewContext.RouteData.Values["Action"].ToString(); Vie ...

  6. MVC前后台获取Action、Controller、ID名方法 以及 路由规则

    前后台获取Action.Controller.ID名方法 前台页面:ViewContext.RouteData.Values["Action"].ToString();//获取Ac ...

  7. C# -- 等待异步操作执行完成的方式 C# -- 使用委托 delegate 执行异步操作 JavaScript -- 原型:prototype的使用 DBHelper类连接数据库 MVC View中获取action、controller、area名称、参数

    C# -- 等待异步操作执行完成的方式 C# -- 等待异步操作执行完成的方式 1. 等待异步操作的完成,代码实现: class Program { static void Main(string[] ...

  8. kubernetes垃圾回收器GarbageCollector Controller源码分析(二)

    kubernetes版本:1.13.2 接上一节:kubernetes垃圾回收器GarbageCollector Controller源码分析(一) 主要步骤 GarbageCollector Con ...

  9. Replication Controller 和 Replica Set

    使用Replication Controller . Replica Set管理Pod Replication Controller (RC) 简写为RC,可以使用rc作为kubectl工具的快速管理 ...

随机推荐

  1. 「luogu1417」烹调方案

    题目链接 :https://www.luogu.org/problemnew/show/P1417 直接背包 ->  30' 考虑直接背包的问题:在DP时第i种食材比第j种食材更优,但由于j&l ...

  2. 【转】【Linux】Swap与Memory

    背景介绍 Memory指机器物理内存,读写速度低于CPU一个量级,但是高于磁盘不止一个量级.所以,程序和数据如果在内存的话,会有非常快的读写速度.但是,内存的造价是要高于磁盘的,且内存的断电丢失数据也 ...

  3. Java并发编程的4个同步辅助类

    Java并发编程的4个同步辅助类(CountDownLatch.CyclicBarrier.Semphore.Phaser) @https://www.cnblogs.com/lizhangyong/ ...

  4. greenplum加密

    --如下为greenplum5.0数据库加解密--加密函数select encrypt('123456','aa','aes');--加解密函数select convert_from(decrypt( ...

  5. Vue项目中使用webpack配置了别名,引入的时候报错

    chainWebpack(config) { config.resolve.alias .set('@', resolve('src')) .set('assets', resolve('src/as ...

  6. 进入js

    JavaScript概述 ECMAScript和JavaScript的关系 1996年11月,JavaScript的创造者--Netscape公司,决定将JavaScript提交给国际标准化组织ECM ...

  7. python之配置日志的三种方式

    以下3种方式来配置logging: 1)使用Python代码显式的创建loggers, handlers和formatters并分别调用它们的配置函数: 2)创建一个日志配置文件,然后使用fileCo ...

  8. 2018秋季C语言学习总结

    2018秋季开始学习c语言 1.printf格式化输出函数 2.基本数据类型,int整型,float浮点型,double双精度浮点型,char字符型 3.算数运算符 +加法,-减法,*乘法,/除法,% ...

  9. ansible的logging模块用来写日志

    [root@node-1 library]# cat dolog.py #!/bin/env python ANSIBLE_METADATA = { 'metadata_version': 'alph ...

  10. ansible的lookup

    lookup路径: /usr/lib/python2.7/site-packages/ansible/plugins/lookup 所有的lookup插件列表cartesian.py dnstxt.p ...