使用python模拟登陆百度

#!/usr/bin/python

# -*- coding: utf- -*-

"""

Function:   Used to demostrate how to use Python code to emulate login baidu main page: http://www.baidu.com/

Note:       Before try to understand following code, firstly, please read the related articles:

            ()【整理】关于抓取网页，分析网页内容，模拟登陆网站的逻辑/流程和注意事项

http://www.crifan.com/summary_about_flow_process_of_fetch_webpage_simulate_login_website_and_some_notice/

            () 【教程】手把手教你如何利用工具(IE9的F12)去分析模拟登陆网站(百度首页)的内部逻辑过程

http://www.crifan.com/use_ie9_f12_to_analysis_the_internal_logical_process_of_login_baidu_main_page_website/

            () 【教程】模拟登陆网站 之 Python版

http://www.crifan.com/emulate_login_website_using_python

Version:    --

Author:     Crifan

"""

import re;

import cookielib;

import urllib;

import urllib2;

import optparse;

#------------------------------------------------------------------------------

# check all cookies in cookiesDict is exist in cookieJar or not

def checkAllCookiesExist(cookieNameList, cookieJar) :

    cookiesDict = {};

    for eachCookieName in cookieNameList :

        cookiesDict[eachCookieName] = False;

    allCookieFound = True;

    for cookie in cookieJar :

        if(cookie.name in cookiesDict) :

            cookiesDict[cookie.name] = True;

    for eachCookie in cookiesDict.keys() :

        if(not cookiesDict[eachCookie]) :

            allCookieFound = False;

            break;

    return allCookieFound;

#------------------------------------------------------------------------------

# just for print delimiter

def printDelimiter():

    print '-'*;

#------------------------------------------------------------------------------

# main function to emulate login baidu

def emulateLoginBaidu():

    print "Function: Used to demostrate how to use Python code to emulate login baidu main page: http://www.baidu.com/";

    print "Usage: emulate_login_baidu_python.py -u yourBaiduUsername -p yourBaiduPassword";

    printDelimiter();

    # parse input parameters

    parser = optparse.OptionParser();

    parser.add_option("-u","--username",action="store",type="string",default='',dest="username",help="Your Baidu Username");

    parser.add_option("-p","--password",action="store",type="string",default='',dest="password",help="Your Baidu password");

    (options, args) = parser.parse_args();

    # export all options variables, then later variables can be used

    for i in dir(options):

        exec(i + " = options." + i);

    printDelimiter();

    print "[preparation] using cookieJar & HTTPCookieProcessor to automatically handle cookies";

    cj = cookielib.CookieJar();

    opener = urllib2.build_opener(urllib2.HTTPCookieProcessor(cj));

    urllib2.install_opener(opener);

    printDelimiter();

    print "[step1] to get cookie BAIDUID";

    baiduMainUrl = "http://www.baidu.com/";

    resp = urllib2.urlopen(baiduMainUrl);

    #respInfo = resp.info();

    #print "respInfo=",respInfo;

    for index, cookie in enumerate(cj):

        print '[',index, ']',cookie;

    printDelimiter();

    print "[step2] to get token value";

    getapiUrl = "https://passport.baidu.com/v2/api/?getapi&class=login&tpl=mn&tangram=true";

    getapiResp = urllib2.urlopen(getapiUrl);

    #print "getapiResp=",getapiResp;

    getapiRespHtml = getapiResp.read();

    #print "getapiRespHtml=",getapiRespHtml;

    #bdPass.api.params.login_token='5ab690978812b0e7fbbe1bfc267b90b3';

    foundTokenVal = re.search("bdPass\.api\.params\.login_token='(?P<tokenVal>\w+)';", getapiRespHtml);

    if(foundTokenVal):

        tokenVal = foundTokenVal.group("tokenVal");

        print "tokenVal=",tokenVal;

        printDelimiter();

        print "[step3] emulate login baidu";

        staticpage = "http://www.baidu.com/cache/user/html/jump.html";

        baiduMainLoginUrl = "https://passport.baidu.com/v2/api/?login";

        postDict = {

            #'ppui_logintime': "",

            'charset'       : "utf-8",

            #'codestring'    : "",

            'token'         : tokenVal, #de3dbf1e8596642fa2ddf2921cd6257f

            'isPhone'       : "false",

            'index'         : "",

            #'u'             : "",

            #'safeflg'       : "",

            'staticpage'    : staticpage, #http%3A%2F%2Fwww.baidu.com%2Fcache%2Fuser%2Fhtml%2Fjump.html

            'loginType'     : "",

            'tpl'           : "mn",

            'callback'      : "parent.bdPass.api.login._postCallback",

            'username'      : username,

            'password'      : password,

            #'verifycode'    : "",

            'mem_pass'      : "on",

        };

        postData = urllib.urlencode(postDict);

        # here will automatically encode values of parameters

        # such as:

        # encode http://www.baidu.com/cache/user/html/jump.html into http%3A%2F%2Fwww.baidu.com%2Fcache%2Fuser%2Fhtml%2Fjump.html

        #print "postData=",postData;

        req = urllib2.Request(baiduMainLoginUrl, postData);

        # in most case, for do POST request, the content-type, is application/x-www-form-urlencoded

        req.add_header('Content-Type', "application/x-www-form-urlencoded");

        resp = urllib2.urlopen(req);

        #for index, cookie in enumerate(cj):

        #    print '[',index, ']',cookie;

        cookiesToCheck = ['BDUSS', 'PTOKEN', 'STOKEN', 'SAVEUSERID'];

        loginBaiduOK = checkAllCookiesExist(cookiesToCheck, cj);

        if(loginBaiduOK):

            print "+++ Emulate login baidu is OK, ^_^";

        else:

            print "--- Failed to emulate login baidu !"

    else:

        print "Fail to extract token value from html=",getapiRespHtml;

if __name__=="__main__":

    emulateLoginBaidu();

使用python模拟登陆百度的更多相关文章

【教程】模拟登陆百度之Java代码版
[背景] 之前已经写了教程,分析模拟登陆百度的逻辑: [教程]手把手教你如何利用工具(IE9的F12)去分析模拟登陆网站(百度首页)的内部逻辑过程然后又去用不同的语言: Python的: [教程]模 ...
模拟登陆百度以及Selenium 的基本用法
模拟登陆百度,需要依赖于selenium 模块,调用浏览器,执行python命令先来说一下这个selenium模块啦...... 本文参考内容来自 Selenium官网 SeleniumPython ...
Python模拟登陆新浪微博
上篇介绍了新浪微博的登陆过程,这节使用Python编写一个模拟登陆的程序.讲解与程序如下: 1.主函数(WeiboMain.py): import urllib2 import cookielib i ...
Python模拟登陆万能法-微博|知乎
Python模拟登陆让不少人伤透脑筋,今天奉上一种万能登陆方法.你无须精通HTML,甚至也无须精通Python,但却能让你成功的进行模拟登陆.本文讲的是登陆所有网站的一种方法,并不局限于微博与知乎,仅 ...
Python模拟登陆TAPD
因为在wiki中未找到需要的数据,查询也很迷,打算用python登录tapd抓取所需项目下的wiki数据,方便查找. 2018-9-30 19:12:44 几步走模拟登录tapd 抓取wiki页左侧 ...
Python模拟登陆淘宝并统计淘宝消费情况的代码实例分享
Python模拟登陆淘宝并统计淘宝消费情况的代码实例分享支付宝十年账单上的数字有点吓人,但它统计的项目太多,只是想看看到底单纯在淘宝上支出了多少,于是写了段脚本,统计任意时间段淘宝订单的消费情况,看 ...
Selenium模拟登陆百度贴吧
Selenium模拟登陆百度贴吧 from selenium import webdriver from time import sleep from selenium.webdriver.commo ...
python 模拟登陆，请求包含cookie信息
需求: 1.通过GET方法,访问URL地址一,传入cookie参数 2.根据地址一返回的uuid,通过POST方法,传入cooki参数实现思路: 1.理解http的GET和POST差别 (网上有很多 ...
python模拟登陆之下载
好长时间没有更新博客了,哈哈. 今天公司给了这么一个需求,现在我们需要去淘宝获取上一天的订单号,然后再根据订单号去另一个接口去获取订单详情,然后再给我展示到web! 中间涉及到的技术点有: 模拟登陆 ...

随机推荐

css一些不为人所熟知的知识点
1.设置a标签内字体水平居中:text-algin:center 2.设置a标签内字体水平居中:line-height:height 3.如何设置td宽度固定<table style=" ...
spark学习之IDEA配置spark并wordcount提交集群
这篇文章包括以下内容 (1)IDEA中scala的安装 (2)hdfs简单的使用,没有写它的部署 (3) 使用scala编写简单的wordcount,输入文件和输出文件使用参数传递 (4)IDEA打包 ...
org.hibernate.id.IdentifierGenerationException: ids for this class must be manually assigned before calling save()
org.hibernate.id.IdentifierGenerationException: ids for this class must be manually assigned before ...
3 手写Java HashMap核心源码
手写Java HashMap核心源码上一章手写LinkedList核心源码,本章我们来手写Java HashMap的核心源码. 我们来先了解一下HashMap的原理.HashMap 字面意思 has ...
洛谷 - P1020 - 导弹拦截 - 最长上升子序列
https://www.luogu.org/problemnew/show/P1020 终于搞明白了.根据某定理,最少需要的防御系统的数量就是最长上升子序列的数量. 呵呵手写二分果然功能很多,想清楚自 ...
lightoj 1422【区间DP·分类区间首元素的情况】
题意: 给你n天分别要穿的衣服种类,可以套着穿, 一旦脱下来就不能再穿,求n天至少要几件. 思路: 区间DP dp[i][j]代表i到j需要至少几件衣服第i天的衣服在第i天穿上了,dp[i][j]= ...
ubuntu 14 安装XML::Simple 模块
最近需要用到perl 来解析xml 文件,从网上搜索了一下,大部分都建议使用XML::Simple 模块来解析,这里记录一下安装过程方法一: 直接使用CPAN 来安装模块 $ perl -MCPAN ...
Hackintosh
条件:Mac环境(也可在Windows电脑上用虚拟机建立).两只(一只亦可)16G及以上优盘.一块64G以上SSD固态(机械)硬盘.一台待折腾的Windows电脑 1.创建安装盘: ·app stor ...
Java的12个语法糖【转】
本文转载自公众号 Hollis 原创: 会反编译的 Hollis 侵权删本文从 Java 编译原理角度,深入字节码及 class 文件,抽丝剥茧,了解 Java 中的语法糖原理及用法,帮助大家在学 ...
MyBatis源码解析（一）
 <bean id="sq ...

使用python模拟登陆百度

使用python模拟登陆百度的更多相关文章

随机推荐

热门专题