无法使用美丽汤打印仅文本

时间:2020-04-18 22:34:48

标签: python html python-3.x text beautifulsoup

我正在努力在python3上创建我的第一个项目之一。当我使用以下代码时:

def scrape_offers():
    r = requests.get("https://www.olx.bg/elektronika/kompyutrni-aksesoari-chasti/aksesoari-chasti/q-1070/?search%5Border%5D=filter_float_price%3Aasc", cookies=all_cookies)
    soup = BeautifulSoup(r.text,"html.parser")
    offers = soup.find_all("div",{'class':'offer-wrapper'})

    for offer in offers:
        offer_name = offer.findChildren("a", {'class':'marginright5 link linkWithHash detailsLink'})
        print(offer_name.text.strip())

我收到以下错误:

Traceback (most recent call last):
  File "scrape_products.py", line 45, in <module>
    scrape_offers()
  File "scrape_products.py", line 40, in scrape_offers
    print(offer_name.text.strip())
  File "/usr/local/lib/python3.7/site-packages/bs4/element.py", line 2128, in __getattr__
    "ResultSet object has no attribute '%s'. You're probably treating a list of elements like a single element. Did you call find_all() when you meant to call find()?" % key
AttributeError: ResultSet object has no attribute 'text'. You're probably treating a list of elements like a single element. Did you call find_all() when you meant to call find()?

我已经在StackOverFlow上阅读了许多类似的案例,但我仍然无法自救。如果有人有任何想法,请帮助:)

P.S .:如果我在没有.text的情况下运行代码,则会显示整个<a class=...> ... </a>

1 个答案:

答案 0 :(得分:1)

findchildren返回一个列表。有时您会得到一个空列表,有时会得到一个包含一个元素的列表。

您应该添加一条if语句来检查返回列表的长度是否大于1,然后打印文本。

import requests
from bs4 import BeautifulSoup
def scrape_offers():
    r = requests.get("https://www.olx.bg/elektronika/kompyutrni-aksesoari-chasti/aksesoari-chasti/q-1070/?search%5Border%5D=filter_float_price%3Aasc")
    soup = BeautifulSoup(r.text,"html.parser")
    offers = soup.find_all("div",{'class':'offer-wrapper'})

    for offer in offers:
        offer_name = offer.findChildren("a", {'class':'marginright5 link linkWithHash detailsLink'})
        if (len(offer_name) >= 1):
            print(offer_name[0].text.strip())

scrape_offers()