MCPcopy Create free account
hub / github.com/Gyyyn/OpenWebTTS / extract_readable_content

Function extract_readable_content

functions/webpage.py:4–152  ·  view source on GitHub ↗

Analyzes an HTML document to extract the main readable content, similar to a "reader mode". This is useful for passing clean text to a Text-to-Speech (TTS) engine. Args: html_content (str): The HTML content of the page as a string. Returns: BeautifulSoup.Tag or Non

(html_content)

Source from the content-addressed store, hash-verified

source not stored for this graph (policy: none)

Callers 1

read_websiteFunction · 0.90

Calls 1

get_element_scoreFunction · 0.85

Tested by

no test coverage detected