脚本之家,脚本语言编程技术及教程分享平台!
分类导航

Python|VBS|Ruby|Lua|perl|VBA|Golang|PowerShell|Erlang|autoit|Dos|bat|

服务器之家 - 脚本之家 - Python - python爬取酷狗音乐Top500榜单

python爬取酷狗音乐Top500榜单

2022-09-09 10:46Ding Jiaxiong Python

大家好,本篇文章主要讲的是python爬取酷狗音乐Top500榜单,感兴趣的同学赶快来看一看吧,对你有帮助的话记得收藏一下

网页情况

python爬取酷狗音乐Top500榜单

爬取数据包含

歌曲排名、歌手、歌曲名、歌曲时长

 

python 代码

import requests #请求网页获取网页数据
  from bs4 import BeautifulSoup #解析网页数据
  import time #时间库
  #user-Agent,伪装成浏览器,便于爬虫的稳定性
  headers = {
      "User-Agent":
      "Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/97.0.4692.71 Safari/537.36"
  }
  def get_info(url):
      web_data = requests.get(url,headers= headers)
      soup = BeautifulSoup(web_data.text,"lxml")
      ranks = soup.select("span.pc_temp_num")
      titles = soup.select("div.pc_temp_songlist > ul > li > a")
      times = soup.select("span.pc_temp_tips_r > span")
      for rank,title,time in zip(ranks,titles,times):
          data = {
              "rank":rank.get_text().strip(),
              "singer":title.get_text().replace("
","").replace("	","").split("-")[1],
              "song":title.get_text().replace("
","").replace("	","").split("-")[0],
              "time":time.get_text().strip()
          }
          print(data)
  if __name__ == "__main__":
      urls = ["https://www.kugou.com/yy/rank/home/{}-8888.html".format(str(i)) for i in range(1,24)]
      for url in urls:
          get_info(url)
          time.sleep(1)

 

运行效果

python爬取酷狗音乐Top500榜单

python爬取酷狗音乐Top500榜单

总结

到此这篇关于python爬取酷狗音乐Top500榜单的文章就介绍到这了,更多相关python爬取酷狗音乐榜单内容请搜索服务器之家以前的文章或继续浏览下面的相关文章希望大家以后多多支持服务器之家!

原文链接:https://blog.csdn.net/weixin_44226181/article/details/122800556

延伸 · 阅读

精彩推荐