Как получить центральный текст c помощью html2text

У меня есть переменная html, которая хранит в себе разметку всей страницы. Мне нужно получить только центральный(главный текст страницы). Например, у меня есть html статьи CNN(https://edition.cnn.com/2020/06/16/africa/africa-coronavirus-cases-prevention-intl/index.html), мне нужно получить только текст самой статьи, тоесть:

Johannesburg (CNN)On January 28, at around one in the morning, Dr. John Nkengasong's cellphone rang in Addis Ababa. Nigerian officials told Nkengasong, the Director of the Africa Centres for Disease Control and Prevention (CDC), that a recently arrived Italian businessman had tested positive for Covid-19. He later recovered. But the force of infection, mostly coming from Europe, seeded the virus in countries throughout the continent, say health officials. As imported cases increased, and community transmission began, the World Health Organization began sounding the alarm in press conferences and statements about an unfolding crisis on the continent. They said Covid-19 patients could quickly overwhelm the weak health infrastructure...

С помощью данного кода я получаю весь текст страницы, включая весь ненужный мусор от которого нужно избавиться. Есть какие-нибудь идеи на этот счёт?

html = get_html()
text_marker = html2text.HTML2Text()
text = text_marker.handle(html)

Ответы (0 шт):