xpath的使用：定位，獲取文本和屬性值

myPage = '''
TITLE

*
****

Hello,\nworld!
-- by Adam

放在尾部的其他一些說(shuō)明

'''

站在用戶的角度思考問(wèn)題，與客戶深入溝通，找到興山網(wǎng)站設(shè)計(jì)與興山網(wǎng)站推廣的解決方案，憑借多年的經(jīng)驗(yàn)，讓設(shè)計(jì)與互聯(lián)網(wǎng)技術(shù)結(jié)合，創(chuàng)造個(gè)性化、用戶體驗(yàn)好的作品，建站類型包括：成都網(wǎng)站建設(shè)、成都網(wǎng)站制作、企業(yè)官網(wǎng)、英文網(wǎng)站、手機(jī)端網(wǎng)站、網(wǎng)站推廣、申請(qǐng)域名、網(wǎng)頁(yè)空間、企業(yè)郵箱。業(yè)務(wù)覆蓋興山地區(qū)。

html = etree.fromstring(myPage)

#一、定位
divs1 = html.xpath('//div')
divs2 = html.xpath('//div[@id]')
divs3 = html.xpath('//div[@class="foot"]')
divs4 = html.xpath('//div[@]')
divs5 = html.xpath('//div[1]')
divs6 = html.xpath('//div[last()-1]')
divs7 = html.xpath('//div[position()<3]')
divs8 = html.xpath('//div|//h2')
divs9 = html.xpath('//div[not(@)]')

二、取文本 text() 區(qū)別 html.xpath('string()')

text1 = html.xpath('//div/text()')
text2 = html.xpath('//div[@id]/text()')
text3 = html.xpath('//div[@class="foot"]/text()')
text4 = html.xpath('//div[@*]/text()')
text5 = html.xpath('//div[1]/text()')
text6 = html.xpath('//div[last()-1]/text()')
text7 = html.xpath('//div[position()<3]/text()')
text8 = html.xpath('//div/text()|//h2/text()')

#三、取屬性 @
value1 = html.xpath('//a/@href')
value2 = html.xpath('//img/@src')
value3 = html.xpath('//div[2]/span/@id')

#四、定位（進(jìn)階）
#1.文檔(DOM)元素(Element)的find，findall方法
divs = html.xpath('//div[position()<3]')
for div in divs:
ass = div.findall('a') # 這里只能找到:div->a, 找不到:div->p->a
for a in ass:
if a is not None:
#print(dir(a))
print(a.text, a.attrib.get('href')) #文檔(DOM)元素(Element)的屬性：text, attrib

2.與1等價(jià)

a_href = html.xpath('//div[position()<3]/a/@href')
print(a_href)

#3.注意與1、2的區(qū)別
a_href = html.xpath('//div[position()<3]//a/@href')
print(a_href)

參考：https://www.cnblogs.com/hhh6460/p/5079465.html

當(dāng)前文章：xpath的使用：定位，獲取文本和屬性值
本文鏈接：http://m.jiaotiyi.com/article/ppiice.html

網(wǎng)站建設(shè)知識(shí)

xpath的使用：定位，獲取文本和屬性值

二、取文本 text() 區(qū)別 html.xpath('string()')

2.與1等價(jià)

其他資訊

網(wǎng)站建設(shè)知識(shí)

xpath的使用：定位，獲取文本和屬性值

二、取文本 text() 區(qū)別 html.xpath('string()')

2.與1等價(jià)

其他資訊

xpath的使用：定位，獲取文本和屬性值

二、取文本 text() 區(qū)別 html.xpath('string()')