python3爬虫之如何使用浏览器cookie

发布时间：2020-12-01 10:52:39 阅读：201 作者：小新栏目：编程语言

Python开发者专用服务器限时活动，0元免费领，库存有限，领完即止！点击查看>>

小编给大家分享一下python3爬虫之如何使用浏览器cookie，希望大家阅读完这篇文章后大所收获，下面让我们一起去探讨吧！

以网页提取标题为例

>>> import re

>>> get_title = lambda html: re.findall('<title>(.*?)</title>', html, flags=re.DOTALL)[0].strip()

未登录情况下下载得到的标题：

>>> import urllib2

>>> url = 'https://bitbucket.org/'

>>> public_html = urllib2.urlopen(url).read()

>>> get_title(public_html)

'Git and Mercurial code management for teams'

使用第三方库browsercookie，获取cookie再下载：

>>> import urllib.request

>>> public_html = urllib.request.urlopen(url).read()

>>> opener = urllib.request.build_opener(urllib.request.HTTPCookieProcessor(cj))

看完了这篇文章，相信你对python3爬虫之如何使用浏览器cookie有了一定的了解，想了解更多相关知识，欢迎关注亿速云行业资讯频道，感谢各位的阅读！

亿速云「云服务器」，即开即用、新一代英特尔至强铂金CPU、三副本存储NVMe SSD云盘，价格低至29元/月。点击查看>>

向AI问一下细节

python3爬虫之如何使用浏览器cookie

猜你喜欢

最新资讯

相关推荐

开发者交流群：

相关标签