美文网首页python_spider
python爬取链家租房之获取房屋页面的详细信息(房名,地址,房

python爬取链家租房之获取房屋页面的详细信息(房名,地址,房

作者: 宁静消失何如 | 来源:发表于2017-06-20 17:53 被阅读30次
    __author__ = 'Lee'
    from bs4 import BeautifulSoup
    import requests
    import time
    url = 'https://bj.lianjia.com/zufang/101101613377.html'
    headers = {
        'User-Agent':'Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/58.0.3029.110 Safari/537.36',
    }
    web_data = requests.get(url,headers=headers)
    soup = BeautifulSoup(web_data.text,'lxml')
    title = soup.title.text #房名
    address = soup.select('div.zf-room > p > a')[0].text  #地址
    price = soup.select(' div.price > span.total')[0].text + '元'
    area = (soup.select('div.zf-room > p ')[0].text).split(':')[-1]
    home_url = url
    print({'title':title ,
           'address':address,
           'price':price,
           'area':area,
           'home_url':home_url
           })
    

    相关文章

      网友评论

        本文标题:python爬取链家租房之获取房屋页面的详细信息(房名,地址,房

        本文链接:https://www.haomeiwen.com/subject/lslaqxtx.html