亚洲精品久久久中文字幕-亚洲精品久久片久久-亚洲精品久久青草-亚洲精品久久婷婷爱久久婷婷-亚洲精品久久午夜香蕉

您的位置:首頁技術文章
文章詳情頁

python - seleium 爬網頁數據,只能怕當前頁,如果我輸入兩頁的話,會出現初始頁數據下載兩次的情況

瀏覽:99日期:2022-07-16 15:35:58

問題描述

import requestsfrom lxml import html,etreefrom selenium import webdriverimport time, json#how many page do you want to scanpage_numnotint = input('how many page do you want to scan')page_num = int(page_numnotint)file_name = ’jd_goods_data.json’url = ’https://list.jd.com/list.html?cat=1713,3264,3414&page=1&delivery=1&sort=sort_totalsales15_desc&trans=1&JL=4_10_0#J_main ’driver = webdriver.Chrome()driver.get(url)base_html = driver.page_sourceselctor = etree.HTML(base_html)date_info = []name_data, price_data = [], []jd_goods_data = {}for q in range(page_num): i = int(1) while True:name_string = ’//*[@id='plist']/ul/li[%d]/p/p[3]/a/em/text()’ %(i)price_string = ’//*[@id='plist']/ul/li[%d]/p/p[2]/strong[1]/i/text()’ %(i)if i == 60: breakelse: i += 1name = selctor.xpath(name_string)[0]name_data.append(name)price = selctor.xpath(price_string)[0]price_data.append(price)jd_goods_data[name] = priceprint(name_data)with open(file_name, ’w’) as f: json.dump(jd_goods_data, f) time.sleep(2) driver.find_element_by_xpath(’//*[@id='J_bottomPage']/span[1]/a[10]’).click() time.sleep(2)# for k, v in jd_goods_data.items(): # print(k,v) # with open(file_name, ’w’) as f: # json.dump(jd_goods_data, f)

問題解答

回答1:

import requestsfrom lxml import html,etreefrom selenium import webdriverimport time, json#how many page do you want to scanpage_numnotint = input('how many page do you want to scan')page_num = int(page_numnotint)file_name = ’jd_goods_data.json’driver = webdriver.Chrome()date_info = []name_data, price_data = [], []jd_goods_data = {}for q in range(page_num): url = ’https://list.jd.com/list.html?cat=1713,3264,3414&page={page}&delivery=1&sort=sort_totalsales15_desc&trans=1&JL=4_10_0#J_main’.format(page=q) driver.get(url) base_html = driver.page_source selctor = etree.HTML(base_html) i = 1 while True:name_string = ’//*[@id='plist']/ul/li[%d]/p/p[3]/a/em/text()’ %(i)price_string = ’//*[@id='plist']/ul/li[%d]/p/p[2]/strong[1]/i/text()’ %(i)if i == 60: breakelse: i += 1name = selctor.xpath(name_string)[0]name_data.append(name)price = selctor.xpath(price_string)[0]price_data.append(price)jd_goods_data[name] = priceprint(name_data)with open(file_name, ’w’) as f: json.dump(jd_goods_data, f)driver.quit()

標簽: Python 編程
主站蜘蛛池模板: 天天在线天天综合网色 | 国产吧在线 | 五月桃花网婷婷亚洲综合 | 国产婷婷一区二区三区 | 久久精品视频在线 | 在线第一福利视频观看 | 国产精品日韩欧美一区二区 | 成人亚洲欧美日韩在线观看 | 久久婷婷激情综合色综合也去 | 国产色婷婷精品免费视频 | 国产精品视频国产永久视频 | 手机看片日韩国产一区二区 | 美女被免费网站视频九色 | 日本一级特黄刺激爽大片 | 性满足久久久久久久久 | 看黄视频在线观看 | 亚洲欧美日韩中文字幕网址 | 久久国产视频一区 | www色婷婷| 做爰成人五级在线视频 | 国产亚洲精品视频中文字幕 | 岛国毛片在线观看 | 成人精品视频一区二区三区尤物 | 精品一区二区三区在线观看l | 国产20岁美女一级毛片 | 日本亚洲精品色婷婷在线影院 | 亚洲欧美日韩成人一区在线 | 国内在线播放 | japanesexxxx护士 | 国产高清视频在线免费观看 | 亚洲精品久一区 | 久草看片 | 最新国产精品亚洲 | 成人国产精品一区二区网站 | 国产亚洲一区二区精品张柏芝 | 国产一区二区三区四区在线 | 欧美日韩在线网站 | 亚洲色图欧洲色图 | 香蕉视频黄色在线观看 | 99亚洲乱人伦精品 | 国产精品视频国产永久视频 |