6、通过xpath获取网页数据

时间：2018-03-21 17:29:09 阅读：334 评论：0 收藏：0 [点我收藏+]

标签：UI ike index post chrome 文件 pre html xpath

from urllib import request
from lxml import etree
# 请求的url
url = "http://www.dfenqi.cn/Product/Index"
# 请求的头文件
headers = {
    "User-Agent": "Mozilla/5.0 (Windows NT 10.0; WOW64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/64.0.3282.186 Safari/537.36"
}
# 创建请求对象
req = request.Request(url,headers = headers)
# 创建处理器对象
httpHandler = request.HTTPHandler()
# 创建opener
opener = request.build_opener(httpHandler)
# 发送请求
response = opener.open(req)
# 读取源文件
html = response.read().decode(‘utf-8‘)
# 创建xpath关系
xpath = "//div[@class=‘liebiao‘]/ul/li/p/text()"
# 获取属性值列表
# xpath = "//div[@class=‘liebiao‘]/ul/li/p/@class"
# 将html转换成可解析对象
selector = etree.HTML(html)
# 返回xpath查询列表
goodsList = selector.xpath(xpath)
# 显示商品标题
for goods in goodsList:
    print(goods)

6、通过xpath获取网页数据

标签：UI ike index post chrome 文件 pre html xpath

原文地址：https://www.cnblogs.com/toloy/p/8618007.html

踩

(0)

评论一句话评论（0）

分享档案

更多>

2021年07月29日 (22)
2021年07月28日 (40)
2021年07月27日 (32)
2021年07月26日 (79)
2021年07月23日 (29)
2021年07月22日 (30)
2021年07月21日 (42)
2021年07月20日 (16)
2021年07月19日 (90)
2021年07月16日 (35)

周排行