天天看點

python淘寶爬蟲_python爬蟲爬取淘寶商品資訊

本文執行個體為大家分享了python爬取淘寶商品的具體代碼,供大家參考,具體内容如下

import requests as req

import re

def getHTMLText(url):

try:

r = req.get(url,timeout=30)

r.raise_for_status()

r.encoding = r.apparent_encoding

return r.text

except:

return ""

def parasePage(ilt,html):

try:

plt = re.findall(r'\"view_price\"\:\"[\d\.]*\"',html)

tlt = re.findall(r'\"raw_title\"\:\".*?\"',html)

for i in range(len(plt)):

price = eval(plt[i].split(':')[1])

title = eval(tlt[i].split(':')[1])

ilt.append([price,title])

except:

print("")

def printGoodsList(ilt):

tplt = "{:4}\t{:8}\t{:16}"

print(tplt.format("序列号","價格","商品名稱"))

count = 0

for j in ilt:

count = count + 1

print(tplt.format(count,j[0],j[1]))

def main():

goods = "python爬蟲"

depth = 3

start_url = 'https://s.taobao.com/search?q=' + goods

infoList = []

for i in range(depth):

try:

url = start_url + '&s=' + str(44*i)

html = getHTMLText(url)

parasePage(infoList,html)

except:

continue

printGoodsList(infoList)

main()

效果圖:

python淘寶爬蟲_python爬蟲爬取淘寶商品資訊

更多内容請參考專題《python爬取功能彙總》進行學習。

以上就是本文的全部内容,希望對大家的學習有所幫助,也希望大家多多支援程式設計小技巧。