Process the XHTML DOM. This should extract the useful context from the page. This should normally be overridden by a subclass to process a specific type of page, i.e. Wikipedia entry.
(Node node, URL url, Network network)
source not stored for this graph (policy: none)
no test coverage detected