第 11 章 python线程与多线程

一、什么是线程

在传统操作系统中，每个进程有一个地址空间，而且默认就有一个控制线程。

进程只是用来把资源集中到一起（进程只是一个资源单位，或者说资源集合），而线程才是cpu上的执行单位。

多线程（即多个控制线程）的概念是，在一个进程中存在多个控制线程，多个控制线程共享该进程的地址空间，相当于一个车间内有多条流水线，都共用一个车间的资源。

二、线程的创建开销小

创建进程的开销要远大于线程。

进程之间是竞争关系，线程之间是协作关系。

不同的进程直接是竞争关系，是不同的程序员写的程序运行的。

同一进程的线程之间是合作关系，是同一个程序写的程序内开启动。

三、线程与进程的区别

1、线程共享创建它的进程的地址空间，进程有自己的地址空间。

2、线程可以直接访问其进程的数据段，进程有它们自己的父进程数据段的副本。

3、线程可以直接与进程的其他线程通信，进程必须使用进程间通信来与其他进程通信。

4、新线程很容易创建，新进程需要复制父进程。

5、线程可以对同一进程的线程进行相当大的控制，进程只能对子进程进行执行控制。

6、对主线程的更改（取消，优先级更改等）可能会影响该进程的其他线程的行为，对父进程的更改不会影响子进程。

四、为何要用多线程

多线程指的是，在一个进程中开启多个线程，简单的讲：如果多个任务共用一块地址空间，那么必须咋一个进程内开启多个线程。详细的分为4点：

1、多线程共享一个进程的地址空间

2、线程比进程更轻量级，线程比进程更容易创建可撤销，在许多操作系统中，创建一个线程比创建一个进程要快10-100倍，在有大量线程需要动态和快速修改时，这一特性很有用。

3、若多个线程都是cpu密集型的，那么并不能获得性能上的增强，但是如果存在大量的计算和大量的I/o处理，拥有多个线程这些活动彼此重叠运行，从而会加快程序执行的速度。

4、在多核cpu系统中，为了最大限度的利用多核，可以开启多个线程，比开进程开销要小的多。（这一条并不适用于python）

五、多线程的应用举例

六、经典的线程模型

多个线程共享同一个进程的地址空间中的资源，是对一台计算机上多个进程的模拟，有时也称线程为轻量级的进程，而对一台计算机上多个进程，则共享物理内存，磁盘，打印机等其他物理资源。多线程的运行和多进程的运行类似，是cpu在多个线程之间的快速切换。

线程通常是有益的，但是带来了不小程序设计难度，线程的问题：

1、父进程有多个线程，那么开启的子进程是否需要同样多的线程，如果是，那么附近中某个线程被阻塞，那么copy到子进程后，copy版的线程也要被阻塞。nginx的多线程模式接收用户连接。

2、在同一个进程中，如果一个线程关闭了问题，而另外一个线程正准备往该文件内写内容呢？如果一个线程注意到没有内存了，并开始分配更多的内存，在工作一半时，发生线程切换，新的线程也发现内存不够用了，又开始分配更多的内存，这样内存就被分配了多次，这些问题都是多线程编程的典型问题，需要仔细思考和设计。

七、threading模块介绍

multiprocessing模块的完全模仿了threading模块的接口，二者在使用层面，有很大的相似性，因而不再详细介绍

https://docs.python.org/3/library/threading.html?highlight=threading#

八、开启线程的两种方式

 #方式一

 from threading import Thread

 import time

 def sayhi(name):

     time.sleep(2)

     print('%s say hello' %name)

 if __name__ == '__main__':

     t=Thread(target=sayhi,args=('egon',))

     t.start()

     print('主线程')

 方式一

方式一

 #方式二

 from threading import Thread

 import time

 class Sayhi(Thread):

     def __init__(self,name):

         super().__init__()

         self.name=name

     def run(self):

         time.sleep(2)

         print('%s say hello' % self.name)

 if __name__ == '__main__':

     t = Sayhi('egon')

     t.start()

     print('主线程')

 方式二

方式二

九、在一个进程下开启多个线程与在一个进程下开启多个子进程的区别

 from threading import Thread

 from multiprocessing import Process

 import os

 def work():

     print('hello')

 if __name__ == '__main__':

     #在主进程下开启线程

     t=Thread(target=work)

     t.start()

     print('主线程/主进程')

     '''

     打印结果:

     hello

     主线程/主进程

     '''

     #在主进程下开启子进程

     t=Process(target=work)

     t.start()

     print('主线程/主进程')

     '''

     打印结果:

     主线程/主进程

     hello

     '''

 谁的开启速度快

谁的开启速度快

 from threading import Thread

 from multiprocessing import Process

 import os

 def work():

     print('hello',os.getpid())

 if __name__ == '__main__':

     #part1:在主进程下开启多个线程,每个线程都跟主进程的pid一样

     t1=Thread(target=work)

     t2=Thread(target=work)

     t1.start()

     t2.start()

     print('主线程/主进程pid',os.getpid())

     #part2:开多个进程,每个进程都有不同的pid

     p1=Process(target=work)

     p2=Process(target=work)

     p1.start()

     p2.start()

     print('主线程/主进程pid',os.getpid())

 瞅一瞅pid

瞅一瞅pid

 from  threading import Thread

 from multiprocessing import Process

 import os

 def work():

     global n

     n=0

 if __name__ == '__main__':

     # n=100

     # p=Process(target=work)

     # p.start()

     # p.join()

     # print('主',n) #毫无疑问子进程p已经将自己的全局的n改成了0,但改的仅仅是它自己的,查看父进程的n仍然为100

     n=1

     t=Thread(target=work)

     t.start()

     t.join()

     print('主',n) #查看结果为0,因为同一进程内的线程之间共享进程内的数据

 同一进程内的线程共享该进程的数据？

同一进程内的线程共享该进程的数据？

十、练习

练习一：

 from threading import Thread

 import socket

 s=socket.socket(socket.AF_INET,socket.SOCK_STREAM)

 s.bind(('127.0.0.1',8080))

 s.listen(5)

 def action(conn):

     while True:

         data=conn.recv(1024)

         print(data)

         conn.send(data.upper())

 if __name__ == '__main__':

     while True:

         conn,addr=s.accept()

         p=Thread(target=action,args=(conn,))

         p.start()

多线程并发的socket服务端

 #/usr/bin/python

 #-*- coding:utf-8 -*-

 import socket

 s=socket.socket(socket.AF_INET,socket.SOCK_STREAM)

 s.connect(('127.0.0.1',8080))

 while True:

     msg=input('>>: ').strip()

     if not msg:continue

     s.send(msg.encode('utf-8'))

     data=s.recv(1024)

     print(data)

客户端

练习二：

三个任务，一个接收用户输入，一个将用户输入的内容格式化成大写，一个将格式后的结果存入文件

 from threading import Thread

 msg_l=[]

 format_l=[]

 def talk():

     while True:

         msg=input('>>: ').strip()

         if not msg:continue

         msg_l.append(msg)

 def format_msg():

     while True:

         if msg_l:

             res=msg_l.pop()

             format_l.append(res.upper())

 def save():

     while True:

         if format_l:

             with open('db.txt','a',encoding='utf-8') as f:

                 res=format_l.pop()

                 f.write('%s\n' %res)

 if __name__ == '__main__':

     t1=Thread(target=talk)

     t2=Thread(target=format_msg)

     t3=Thread(target=save)

     t1.start()

     t2.start()

     t3.start()

十一、线程相关的其他方法

Thread实例对象的方法

  # isAlive(): 返回线程是否活动的。

  # getName(): 返回线程名。

  # setName(): 设置线程名。

threading模块提供的一些方法：

  # threading.currentThread(): 返回当前的线程变量。

  # threading.enumerate(): 返回一个包含正在运行的线程的list。正在运行指线程启动后、结束前，不包括启动前和终止后的线程。

  # threading.activeCount(): 返回正在运行的线程数量，与len(threading.enumerate())有相同的结果。

 from threading import Thread

 import threading

 from multiprocessing import Process

 import os

 def work():

     import time

     time.sleep(3)

     print(threading.current_thread().getName())

 if __name__ == '__main__':

     #在主进程下开启线程

     t=Thread(target=work)

     t.start()

     print(threading.current_thread().getName())

     print(threading.current_thread()) #主线程

     print(threading.enumerate()) #连同主线程在内有两个运行的线程

     print(threading.active_count())

     print('主线程/主进程')

     '''

     打印结果:

     MainThread

     <_MainThread(MainThread, started 140735268892672)>

     [<_MainThread(MainThread, started 140735268892672)>, <Thread(Thread-1, started 123145307557888)>]

     主线程/主进程

     Thread-1

     '''

主线程等待子线程结束

from threading import Thread

import time

def sayhi(name):

    time.sleep(2)

    print('%s say hello' %name)

if __name__ == '__main__':

    t=Thread(target=sayhi,args=('egon',))

    t.start()

    t.join()

    print('主线程')

    print(t.is_alive())

    '''

    egon say hello

    主线程

    False

    '''

十二、守护线程

无论是进程还是线程，都遵循：守护xxx会等待主xxx运行完毕后被销毁，需要强调的是：运行完毕并非终止运行

#1.对主进程来说，运行完毕指的是主进程代码运行完毕

#2.对主线程来说，运行完毕指的是主线程所在的进程内所有非守护线程统统运行完毕，主线程才算运行完毕

详细解释：

#1 主进程在其代码结束后就已经算运行完毕了（守护进程在此时就被回收）,然后主进程会一直等非守护的子进程都运行完毕后回收子进程的资源(否则会产生僵尸进程)，才会结束，

#2 主线程在其他非守护线程运行完毕后才算运行完毕（守护线程在此时就被回收）。因为主线程的结束意味着进程的结束，进程整体的资源都将被回收，而进程必须保证非守护线程都运行完毕后才能结束。

from threading import Thread

import time

def sayhi(name):

    time.sleep(2)

    print('%s say hello' %name)

if __name__ == '__main__':

    t=Thread(target=sayhi,args=('egon',))

    t.setDaemon(True) #必须在t.start()之前设置

    t.start()

    print('主线程')

    print(t.is_alive())

    '''

    主线程

    True

    '''

 from threading import Thread

 import time

 def foo():

     print(123)

     time.sleep(1)

     print("end123")

 def bar():

     print(456)

     time.sleep(3)

     print("end456")

 t1=Thread(target=foo)

 t2=Thread(target=bar)

 t1.daemon=True

 t1.start()

 t2.start()

 print("main-------")

 迷惑人的例子

迷惑人的例子

十三、python GIL（Glbal Interpreter Lock）

链接：http://www.cnblogs.com/linhaifeng/articles/7449853.html

十四、同步锁

三个需要注意的点：

#1.线程抢的是GIL锁，GIL锁相当于执行权限，拿到执行权限后才能拿到互斥锁Lock，其他线程也可以抢到GIL，但如果发现Lock仍然没有被释放则阻塞，即便是拿到执行权限GIL也要立刻交出来

#2.join是等待所有，即整体串行，而锁只是锁住修改共享数据的部分，即部分串行，要想保证数据安全的根本原理在于让并发变成串行，join与互斥锁都可以实现，毫无疑问，互斥锁的部分串行效率要更高

#3. 一定要看本小节最后的GIL与互斥锁的经典分析

GIL VS Lock

GIL保证同一时间只能有一个线程来执行。

锁的目的是为了保护共享的数据，同一时间只能有一个线程来修改共享的数据

结论：保护不同的数据就应该加不同的锁。

GIL与Lock是两把锁，保护的数据不一样，前者是解释器级别的（当然保护的就是解释器级别的数据，比如垃圾回收的数据），后者是保护用户自己开发的应用程序的数据，很明显GIL不负责这件事，只能用户自定义加锁处理，即Lock。

过程分析：所有线程抢的是GIL锁，或者说所有线程抢的是执行权限

线程1抢到GIL锁，拿到执行权限，开始执行，然后加了一把Lock，还没有执行完毕，即线程1还未释放Lock，有可能线程2抢到GIL锁，开始执行，执行过程中Lock还没有被线程1释放，于是线程2进入阻塞，被夺走执行权限，有可能线程1拿到GIL，然后正常到释放Lock....这就导致了串行运行的效果。

既然是串行，那我们执行

t1.start()

t1.join

t2.start()

t2.join()

这也是串行执行，为何还要加lock，需知join是等待t1所有的代码执行完，相当于锁住了t1的所有代码，而lock只是锁住一部分操作共享数据的代码。

因为Python解释器帮你自动定期进行内存回收，你可以理解为python解释器里有一个独立的线程，每过一段时间它起wake up做一次全局轮询看看哪些内存数据是可以被清空的，此时你自己的程序 里的线程和 py解释器自己的线程是并发运行的，假设你的线程删除了一个变量，py解释器的垃圾回收线程在清空这个变量的过程中的clearing时刻，可能一个其它线程正好又重新给这个还没来及得清空的内存空间赋值了，结果就有可能新赋值的数据被删除了，为了解决类似的问题，python解释器简单粗暴的加了锁，即当一个线程运行时，其它人都不能动，这样就解决了上述的问题，  这可以说是Python早期版本的遗留问题。

详细

from threading import Thread

import os,time

def work():

    global n

    temp=n

    time.sleep(0.1)

    n=temp-1

if __name__ == '__main__':

    n=100

    l=[]

    for i in range(100):

        p=Thread(target=work)

        l.append(p)

        p.start()

    for p in l:

        p.join()

    print(n) #结果可能为99

锁通常被用来实现对共享资源的同步访问，为每一个共享资源创建一个lock对象，当你需要访问该资源时，调用acquire方法来获取锁对象（如果其他线程已经获得了该锁，则当前线程需要等待其被释放），等待资源访问完后，再调用release方法释放锁：

 import threading

 R=threading.Lock()

 R.acquire()

 '''

 对公共数据的操作

 '''

 R.release()

 from threading import Thread,Lock

 import os,time

 def work():

     global n

     lock.acquire()

     temp=n

     time.sleep(0.1)

     n=temp-1

     lock.release()

 if __name__ == '__main__':

     lock=Lock()

     n=100

     l=[]

     for i in range(100):

         p=Thread(target=work)

         l.append(p)

         p.start()

     for p in l:

         p.join()

     print(n) #结果肯定为0，由原来的并发执行变成串行，牺牲了执行效率保证了数据安全

 分析：

 　　#1.100个线程去抢GIL锁，即抢执行权限

      #2. 肯定有一个线程先抢到GIL（暂且称为线程1），然后开始执行，一旦执行就会拿到lock.acquire()

      #3. 极有可能线程1还未运行完毕，就有另外一个线程2抢到GIL，然后开始运行，但线程2发现互斥锁lock还未被线程1释放，于是阻塞，被迫交出执行权限，即释放GIL

     #4.直到线程1重新抢到GIL，开始从上次暂停的位置继续执行，直到正常释放互斥锁lock，然后其他的线程再重复2 3 4的过程

 GIL锁与互斥锁综合分析（重点！！！）

GIL锁与互斥锁综合分析（重点！！！）

 #不加锁:并发执行,速度快,数据不安全

 from threading import current_thread,Thread,Lock

 import os,time

 def task():

     global n

     print('%s is running' %current_thread().getName())

     temp=n

     time.sleep(0.5)

     n=temp-1

 if __name__ == '__main__':

     n=100

     lock=Lock()

     threads=[]

     start_time=time.time()

     for i in range(100):

         t=Thread(target=task)

         threads.append(t)

         t.start()

     for t in threads:

         t.join()

     stop_time=time.time()

     print('主:%s n:%s' %(stop_time-start_time,n))

 '''

 Thread-1 is running

 Thread-2 is running

 ......

 Thread-100 is running

 主:0.5216062068939209 n:99

 '''

 #不加锁:未加锁部分并发执行,加锁部分串行执行,速度慢,数据安全

 from threading import current_thread,Thread,Lock

 import os,time

 def task():

     #未加锁的代码并发运行

     time.sleep(3)

     print('%s start to run' %current_thread().getName())

     global n

     #加锁的代码串行运行

     lock.acquire()

     temp=n

     time.sleep(0.5)

     n=temp-1

     lock.release()

 if __name__ == '__main__':

     n=100

     lock=Lock()

     threads=[]

     start_time=time.time()

     for i in range(100):

         t=Thread(target=task)

         threads.append(t)

         t.start()

     for t in threads:

         t.join()

     stop_time=time.time()

     print('主:%s n:%s' %(stop_time-start_time,n))

 '''

 Thread-1 is running

 Thread-2 is running

 ......

 Thread-100 is running

 主:53.294203758239746 n:0

 '''

 #有的同学可能有疑问:既然加锁会让运行变成串行,那么我在start之后立即使用join,就不用加锁了啊,也是串行的效果啊

 #没错:在start之后立刻使用jion,肯定会将100个任务的执行变成串行,毫无疑问,最终n的结果也肯定是0,是安全的,但问题是

 #start后立即join:任务内的所有代码都是串行执行的,而加锁,只是加锁的部分即修改共享数据的部分是串行的

 #单从保证数据安全方面,二者都可以实现,但很明显是加锁的效率更高.

 from threading import current_thread,Thread,Lock

 import os,time

 def task():

     time.sleep(3)

     print('%s start to run' %current_thread().getName())

     global n

     temp=n

     time.sleep(0.5)

     n=temp-1

 if __name__ == '__main__':

     n=100

     lock=Lock()

     start_time=time.time()

     for i in range(100):

         t=Thread(target=task)

         t.start()

         t.join()

     stop_time=time.time()

     print('主:%s n:%s' %(stop_time-start_time,n))

 '''

 Thread-1 start to run

 Thread-2 start to run

 ......

 Thread-100 start to run

 主:350.6937336921692 n:0 #耗时是多么的恐怖

 '''

 互斥锁与join的区别（重点！！！）

互斥锁与join的区别（重点！！！）

十五、死锁现象与递归锁

进程也有死锁与递归锁。

所谓死锁：是指两个或两个以上的进程或线程在执行过程中，因争夺资源而造成的一种互相等待的现象，若无外力作用，它们都将无法推进下去，此时称系统处于死锁状态或系统产生了死锁，这些永远在互相等待的进程称为死锁进程，如下就是死锁

 from threading import Thread,Lock

 import time

 mutexA=Lock()

 mutexB=Lock()

 class MyThread(Thread):

     def run(self):

         self.func1()

         self.func2()

     def func1(self):

         mutexA.acquire()

         print('\033[41m%s 拿到A锁\033[0m' %self.name)

         mutexB.acquire()

         print('\033[42m%s 拿到B锁\033[0m' %self.name)

         mutexB.release()

         mutexA.release()

     def func2(self):

         mutexB.acquire()

         print('\033[43m%s 拿到B锁\033[0m' %self.name)

         time.sleep(2)

         mutexA.acquire()

         print('\033[44m%s 拿到A锁\033[0m' %self.name)

         mutexA.release()

         mutexB.release()

 if __name__ == '__main__':

     for i in range(10):

         t=MyThread()

         t.start()

 '''

 Thread-1 拿到A锁

 Thread-1 拿到B锁

 Thread-1 拿到B锁

 Thread-2 拿到A锁

 然后就卡住，死锁了

 '''

解决方法，递归锁，在python中为了支持在同一线程中多次请求同一资源，python提供了可重入锁RLock。

这个RLock内部维护着一个Lock和一个counter变量，counter记录了acquire的次数，从而使得资源可以被多次require，直到一个线程所有的acquire都被release，其他的线程才能获得资源，上面的例子如果使用RLock代替lock，则不会发生死锁：

mutexA=mutexB=threading.RLock() #一个线程拿到锁，counter加1,该线程内又碰到加锁的情况，则counter继续加1，这期间所有其他线程都只能等待，等待该线程释放所有锁，即counter递减到0为止

 #递归锁

 from threading import Thread,Lock,RLock

 import time

 mutex=RLock()

 class Mythread(Thread):

     def run(self):

         self.f1()

         self.f2()

     def f1(self):

         mutex.acquire()

         print('\033[45m%s 抢到A锁\033[0m' %self.name)

         mutex.acquire()

         print('\033[44m%s 抢到B锁\033[0m' %self.name)

         mutex.release()

         mutex.release()

     def f2(self):

         mutex.acquire()

         print('\033[44m%s 抢到B锁\033[0m' %self.name)

         time.sleep(1)

         mutex.acquire()

         print('\033[45m%s 抢到A锁\033[0m' %self.name)

         mutex.release()

         mutex.release()

 if __name__ == '__main__':

     for i in range(20):

         t=Mythread()

         t.start()

递归锁

十六、信号量Semaphore

同进程的一样

Semaphore管理一个内置的计数器

每当调用acquire()时内置计数器-1；调用release()时内置计数器+1；计数器不能小于0；当计数器为0时，acquire()将阻塞线程直到其他线程调用release（）。

实例：（同时只有5个线程可以获得semaphore，即可以限制最大连接数为5）：

 from threading import Thread,Semaphore

 import threading

 import time

 # def func():

 #     if sm.acquire():

 #         print (threading.currentThread().getName() + ' get semaphore')

 #         time.sleep(2)

 #         sm.release()

 def func():

     sm.acquire()

     print('%s get sm' %threading.current_thread().getName())

     time.sleep(3)

     sm.release()

 if __name__ == '__main__':

     sm=Semaphore(5)

     for i in range(23):

         t=Thread(target=func)

         t.start()

 from threading import Thread,current_thread,Semaphore

 import time,random

 sm=Semaphore(5)

 def work():

     sm.acquire()

     print('%s 上厕所' %current_thread().getName())

     time.sleep(random.randint(1,3))

     sm.release()

 if __name__ == '__main__':

     for i in range(20):

         t=Thread(target=work)

         t.start()

与进程池是完全不同的概念，进程池Pool（4），最大只能产生4个进程，而且从头到尾都只是这四个进程，不会产生新的，而信号量是产生一堆线程/进程。（信号量semaphore默认开启线程是cpu个数的*5）

十七、事件 Event

同进程的一样

线程的一个关键特性是每个线程都是独立运行且状态不可预测，如果程序中的其他线程需要通过判断某个线程的状态来确定自己下一步的操作，这时线程同步问题就会变得非常棘手，为了解决这些问题，我们需要使用threading库中的Event对象，对象包含一个可由线程设置的信号标志，它允许线程等待某些事件的发生，在初始情况下，Event对象中的信号标志被设置为假，如果有线程等待一个Event对象，而这个Event对象的标志为假，那么这个线程将会被一直阻塞直至该标志为真。一个线程如果将一个Event对象的信号标志设置为真，它将唤醒所有等待这个Event对象的线程，如果一个线程等待一个已经被设置为真的Event对象，那么它将忽略这个事件，继续执行

 event.isSet()：返回event的状态值；

 event.wait()：如果 event.isSet()==False将阻塞线程；

 event.set()： 设置event的状态值为True，所有阻塞池的线程激活进入就绪状态， 等待操作系统调度；

 event.clear()：恢复event的状态值为False。

 from threading import Thread,Event

 import threading

 import time,random

 def conn_mysql():

     count=1

     while not event.is_set():

         if count > 3:

             raise TimeoutError('链接超时')

         print('<%s>第%s次尝试链接' % (threading.current_thread().getName(), count))

         event.wait(0.5)

         count+=1

     print('<%s>链接成功' %threading.current_thread().getName())

 def check_mysql():

     print('\033[45m[%s]正在检查mysql\033[0m' % threading.current_thread().getName())

     time.sleep(random.randint(2,4))

     event.set()

 if __name__ == '__main__':

     event=Event()

     conn1=Thread(target=conn_mysql)

     conn2=Thread(target=conn_mysql)

     check=Thread(target=check_mysql)

     conn1.start()

     conn2.start()

     check.start()

十八、定时器

定时器，指定n秒后执行某操作

from threading import Timer

def hello():

    print("hello, world")

t = Timer(1, hello)

t.start()  # after 1 seconds, "hello, world" will be printed

十九、线程queue队列

queue队列：使用import queue，用法与进程的Queue一样

queue is especially useful in threaded programming when information must be exchanged safely between multiple threads.

class queue.Queue（maxsize=0）#队列：先进先出

 import queue

 q=queue.Queue(3) #队列：先进先出

 q.put(1)

 q.put(2)

 q.put(3)

 print(q.get())

 print(q.get())

 print(q.get())

 '''

 结果（先进先出）

 1

 2

 3

 '''

队列：先进先出

class queue.LifoQueue（maxsize=0）#last in fisrt out

 import queue

 q=queue.LifoQueue(3) #堆栈：后进先出

 q.put(1)

 q.put(2)

 q.put(3)

 print(q.get())

 print(q.get())

 print(q.get())

 '''

 结果（后进先出）

 3

 2

 1

 '''

堆栈：后进先出

class queue.PriorityQueue（maxsize=0）#存储数据时可设置优先级的队列

 import queue

 q=queue.PriorityQueue()

 #put进入一个元组,元组的第一个元素是优先级(通常是数字,也可以是非数字之间的比较),数字越小优先级越高

 q.put((20,'a'))

 q.put((10,'b'))

 q.put((30,'c'))

 print(q.get())

 print(q.get())

 print(q.get())

 '''

 结果(数字越小优先级越高,优先级高的优先出队):

 (10, 'b')

 (20, 'a')

 (30, 'c')

 '''

其他

 Constructor for a priority queue. maxsize is an integer that sets the upperbound limit on the number of items that can be placed in the queue. Insertion will block once this size has been reached, until queue items are consumed. If maxsize is less than or equal to zero, the queue size is infinite.

 The lowest valued entries are retrieved first (the lowest valued entry is the one returned by sorted(list(entries))[0]). A typical pattern for entries is a tuple in the form: (priority_number, data).

 exception queue.Empty

 Exception raised when non-blocking get() (or get_nowait()) is called on a Queue object which is empty.

 exception queue.Full

 Exception raised when non-blocking put() (or put_nowait()) is called on a Queue object which is full.

 Queue.qsize()

 Queue.empty() #return True if empty

 Queue.full() # return True if full

 Queue.put(item, block=True, timeout=None)

 Put item into the queue. If optional args block is true and timeout is None (the default), block if necessary until a free slot is available. If timeout is a positive number, it blocks at most timeout seconds and raises the Full exception if no free slot was available within that time. Otherwise (block is false), put an item on the queue if a free slot is immediately available, else raise the Full exception (timeout is ignored in that case).

 Queue.put_nowait(item)

 Equivalent to put(item, False).

 Queue.get(block=True, timeout=None)

 Remove and return an item from the queue. If optional args block is true and timeout is None (the default), block if necessary until an item is available. If timeout is a positive number, it blocks at most timeout seconds and raises the Empty exception if no item was available within that time. Otherwise (block is false), return an item if one is immediately available, else raise the Empty exception (timeout is ignored in that case).

 Queue.get_nowait()

 Equivalent to get(False).

 Two methods are offered to support tracking whether enqueued tasks have been fully processed by daemon consumer threads.

 Queue.task_done()

 Indicate that a formerly enqueued task is complete. Used by queue consumer threads. For each get() used to fetch a task, a subsequent call to task_done() tells the queue that the processing on the task is complete.

 If a join() is currently blocking, it will resume when all items have been processed (meaning that a task_done() call was received for every item that had been put() into the queue).

 Raises a ValueError if called more times than there were items placed in the queue.

 Queue.join() block直到queue被消费完毕

二十、python标准模块--concurrent.futures（开启进程池和线程池）

https://docs.python.org/dev/library/concurrent.futures.html

 from concurrent.futures import ProcessPoolExecutor,ThreadPoolExecutor

 from threading import current_thread

 import os,time,random

 def work(n):

     print('%s is running' %current_thread().getName())

     time.sleep(random.randint(1,3))

     return n**2

 if __name__ == '__main__':

     p=ThreadPoolExecutor()

     objs=[]

     for i in range(21):

         obj=p.submit(work,i)

         objs.append(obj)

     p.shutdown()

     for obj in objs:

         print(obj.result())

开启线程池

 #进程池

 import requests #pip3 install requests

 import os,time

 from multiprocessing import Pool

 from concurrent.futures import ProcessPoolExecutor,ThreadPoolExecutor

 def get_page(url):

     print('<%s> get :%s' %(os.getpid(),url))

     respone = requests.get(url)

     if respone.status_code == 200:

         return {'url':url,'text':respone.text}

 def parse_page(obj):

     dic=obj.result()

     print('<%s> parse :%s' %(os.getpid(),dic['url']))

     time.sleep(0.5)

     res='url:%s size:%s\n' %(dic['url'],len(dic['text'])) #模拟解析网页内容

     with open('db.txt','a') as f:

         f.write(res)

 if __name__ == '__main__':

     # p=Pool(4)

     p=ProcessPoolExecutor()

     urls = [

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

     ]

     for url in urls:

         # p.apply_async(get_page,args=(url,),callback=parse_page)

         p.submit(get_page,url).add_done_callback(parse_page)

     p.shutdown()

     print('主进程pid:',os.getpid())

#进程池应用

 #线程池

 import requests #pip3 install requests

 import os,time,threading

 from multiprocessing import Pool

 from concurrent.futures import ProcessPoolExecutor,ThreadPoolExecutor

 def get_page(url):

     print('<%s> get :%s' %(threading.current_thread().getName(),url))

     respone = requests.get(url)

     if respone.status_code == 200:

         return {'url':url,'text':respone.text}

 def parse_page(obj):

     dic=obj.result()

     print('<%s> parse :%s' %(threading.current_thread().getName(),dic['url']))

     time.sleep(0.5)

     res='url:%s size:%s\n' %(dic['url'],len(dic['text'])) #模拟解析网页内容

     with open('db.txt','a') as f:

         f.write(res)

 if __name__ == '__main__':

     # p=Pool(4)

     p=ThreadPoolExecutor(3)

     urls = [

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

         'http://www.baidu.com',

     ]

     for url in urls:

         # p.apply_async(get_page,args=(url,),callback=parse_page)

         p.submit(get_page,url).add_done_callback(parse_page)

     p.shutdown()

     print('主进程pid:',os.getpid())

#线程池应用

 from concurrent.futures import ProcessPoolExecutor,ThreadPoolExecutor

 import os,time,random

 def work(n):

     print('%s is running' %os.getpid())

     time.sleep(random.randint(1,3))

     return n**2

 if __name__ == '__main__':

     p=ProcessPoolExecutor()

     objs=[]

     for i in range(10):

         obj=p.submit(work,i)

         objs.append(obj)

     p.shutdown()

     for obj in objs:

         print(obj.result())

 #新方法

 if __name__ == '__main__':

     p = ProcessPoolExecutor()

     obj=p.map(work,range(10))

     p.shutdown()

     print(list(obj))

map(func, *iterables, timeout=None, chunksize=1)

二十一、python并发编程之协程

详细地址：http://www.cnblogs.com/fanglingen/articles/7486020.html

二十二、python并发编程之io模型

详细地址：http://www.cnblogs.com/fanglingen/articles/7490901.html

第 11 章 python线程与多线程的更多相关文章

第11章 Windows线程池（1）_传统的Windows线程池
第11章 Windows线程池 11.1 传统的Windows线程池及API (1)线程池中的几种底层线程 ①可变数量的长任务线程:WT_EXECUTELONGFUNCTION ②Timer线程:调用 ...
Windows核心编程：第11章 Windows线程池
Github https://github.com/gongluck/Windows-Core-Program.git //第11章 Windows线程池.cpp: 定义应用程序的入口点. // #i ...
python 线程、多线程
复习进程知识: python:主进程,至少有一个主线程启动一个新的子进程:Process,pool 给每一个进程设定一下执行的任务:传一个函数+函数的参数如果是进程池:map函数:传入一个任务函数 ...
《python解释器源码剖析》第11章--python虚拟机中的控制流
11.0 序在上一章中,我们剖析了python虚拟机中的一般表达式的实现.在剖析一遍表达式是我们的流程都是从上往下顺序执行的,在执行的过程中没有任何变化.但是显然这是不够的,因为怎么能没有流程控制呢 ...
python——线程与多线程进阶
之前我们已经学会如何在代码块中创建新的线程去执行我们要同步执行的多个任务,但是线程的世界远不止如此.接下来,我们要介绍的是整个threading模块.threading基于Java的线程模型设计.锁( ...
python——线程与多线程基础
我们之前已经初步了解了进程.线程与协程的概念,现在就来看看python的线程.下面说的都是一个进程里的故事了,暂时忘记进程和协程,先来看一个进程中的线程和多线程.这篇博客将要讲一些单线程与多线程的基础 ...
第11章 Windows线程池（3）_私有的线程池
11.3 私有的线程池 11.3.1 创建和销毁私有的线程池 (1)进程默认线程池当调用CreateThreadpoolwork.CreateThreadpoolTimer.CreateThread ...
第11章 Windows线程池（2）_Win2008及以上的新线程池
11.2 Win2008以上的新线程池 (1)传统线程池的优缺点: ①传统Windows线程池调用简单,使用方便(有时只需调用一个API即可) ②这种简单也带来负面问题,如接口过于简单,无法更多去控制 ...
python线程互斥锁Lock（29）
在前一篇文章 python线程创建和传参中我们介绍了关于python线程的一些简单函数使用和线程的参数传递,使用多线程可以同时执行多个任务,提高开发效率,但是在实际开发中往往我们会碰到线程同步问题, ...

随机推荐

Linux-SSH免密登陆原理
HTMLTestRunner_PY3脚本代码
HTMLTestRunner_PY3.py文件代码如下: # -*- coding: utf-8 -*- """ A TestRunner for use with th ...
灵活的理解JavaScript中的this指向（一）
this是JavaScript中的关键字之一,在编写程序的时候经常会用到,正确的理解和使用关键字this尤为重要.首先必须要说的是,this的指向在函数定义的时候是确定不了的,只有函数执行的时候才能确 ...
ORACLE之字符集修改（10g）
当从oracle服务器将数据导出成dmp文件后,再导入到本地的oracle数据库时,出现: IMP-00019: 由于 ORACLE 错误 12899 而拒绝行 IMP-00003: 遇到 ORACL ...
【知识强化】第二章数据的表示和运算 2.4 算术逻辑单元ALU
从本节开始我们就进入到本章的最后一节内容了,也就是我们算术逻辑单元的它的实现.这部分呢是数字电路的一些知识,所以呢,如果你没有学过数字电路的话,也不要慌张,我会从基础开始给大家补起.那么在计算机当中, ...
linux精简开机启动服务
1.可以使用 setup-system services 里面调整,这样调整起来效率低 2.或者 ntsysv 调出来 3.使用脚本一件关闭 #LANG=en chkconfig --list #停止 ...
C# 模拟页面登录
using System; using System.Collections; using System.Collections.Generic; using System.IO; using Sys ...
docker安装es
下载镜像 docker pull docker.elastic.co/elasticsearch/elasticsearch:6.8.1 创建容器并映射docker run -e ES_JAVA_OP ...
02scikit-learn模型训练
模型训练 In [6]: import numpy as np import matplotlib.pyplot as plt from sklearn.linear_model import Lin ...
[题目] 4座塔的Hanoi
题目地址经典递推题. 解出 n (1<=n<=12) 个盘子 \(4\) 座塔的Hanoi(汉诺塔)问题最少需多少步?(1到12每个答案分别占一行) 题解在原Hanoi问题中 \(d[ ...

第 11 章 python线程与多线程

第 11 章 python线程与多线程的更多相关文章

随机推荐

热门专题