python 3利用BeautifulSoup抓取div标签的方法示例

时间：2021-05-22

前言

本文主要介绍的是关于python 3用BeautifulSoup抓取div标签的方法示例，分享出来供大家参考学习，下面来看看详细的介绍：

示例代码：

# -*- coding:utf-8 -*-#python 2.7#XiaoDeng#http://tieba.baidu.com/p/2460150866#标签操作from bs4 import BeautifulSoupimport urllib.requestimport re#如果是网址，可以用这个办法来读取网页#html_doc = "http://tieba.baidu.com/p/2460150866"#req = urllib.request.Request(html_doc) #webpage = urllib.request.urlopen(req) #html = webpage.read()html="""<html><head><title>The Dormouse's story</title></head><body>The Dormouse's storyOnce upon a time there were three little sisters; and their names were<a href="http://example.com/elsie" rel="external nofollow" class="sister" id="xiaodeng"></a>,<a href="http://example.com/lacie" rel="external nofollow" rel="external nofollow" class="sister" id="link2">Lacie</a> and<a href="http://example.com/tillie" rel="external nofollow" class="sister" id="link3">Tillie</a>;<a href="http://example.com/lacie" rel="external nofollow" rel="external nofollow" class="sister" id="xiaodeng">Lacie</a>and they lived at the bottom of a well.<div class="ntopbar_loading"><img src="http://simg.sinajs.cn/blog7style/images/common/loading.gif">加载中…</div><div class="SG_connHead"> 个人资料 <div class="info_list"> <ul class="info_list1"> <li>博客等级：<img src="http://simg.sinajs.cn/blog7style/images/common/sg_trans.gif" real_src="http://simg.sinajs.cn/blog7style/images/common/number/9.gif" /></li> <li>博客积分：0</li> </ul> <ul class="info_list2"> <li>博客访问：3,971</li> <li>关注人气：0</li> <li>获赠金笔：0支</li> <li>赠出金笔：0支</li> <li class="lisp" id="comp_901_badge">荣誉徽章：</li> </ul> </div><div class="atcTit_more"><a href="http://blog.sina.com.cn/" rel="external nofollow" rel="external nofollow" target="_blank">更多>></a></div> ..."""soup = BeautifulSoup(html, 'html.parser') #文档对象# 类名为xxx而且文本内容为hahaha的divfor k in soup.find_all('div',class_='atcTit_more'):#,string='更多' print(k) #<div class="atcTit_more"><a href="http://blog.sina.com.cn/" rel="external nofollow" rel="external nofollow" target="_blank">更多>></a></div>

总结

以上就是这篇文章的全部内容了，希望本文的内容对大家的学习或者工作能带来一定的帮助，如果有疑问大家可以留言交流，谢谢大家的支持。

python 3利用BeautifulSoup抓取div标签的方法示例

相关文章

Python实现抓取城市的PM2.5浓度和排名

Python爬虫包BeautifulSoup学习实例（五）

python爬虫开发之Beautiful Soup模块从安装到详细使用方法与实例

Python中BeautifuSoup库的用法使用详解

python基于BeautifulSoup实现抓取网页指定内容的方法