Parse the given HTML and returns token objects (words with attached tags). This parses only the content of a page; anything in the head is ignored, and the and elements are themselves optional. The content is then parsed by lxml, which ensures the validity of the
(html, include_hrefs=True)
source not stored for this graph (policy: none)
no test coverage detected