/opt/alt/python-internal/lib64/python3.11/html/__pycache__
Edit: /opt/alt/python-internal/lib64/python3.11/html/__pycache__/parser.cpython-311.pyc (19396B)
§
qJ[¾§˜ô�ã óÞ — d Z ddlZddlZddlmZ dgZ ej d¦ « Z ej d¦ « Z ej d¦ « Z ej d¦ « Z
ej d ¦ « Z ej d
¦ « Z ej d¦ « Z
ej d¦ « Z ej d
¦ « Z ej dej ¦ « Z ej d
¦ « Z ej d¦ « Z G d„ dej ¦ « ZdS )zA parser for HTML and XHTML.é N)ÚunescapeÚ
HTMLParserz[&<]z
&[a-zA-Z#]z%&([a-zA-Z][-.a-zA-Z0-9]*)[^a-zA-Z0-9]z)(?:[0-9]+|[xX][0-9a-fA-F]+)[^0-9a-fA-F]z <[a-zA-Z]ú>z--\s*>z+([a-zA-Z][^\t\n\r\f />\x00]*)(?:\s|/(?!>))*z]((?<=[\'"\s/])[^\s/>][^\s/=>]*)(\s*=+\s*(\'[^\']*\'|"[^"]*"|(?![\'"])[^>\s]*))?(?:\s|/(?!>))*aF
<[a-zA-Z][^\t\n\r\f />\x00]* # tag name
(?:[\s/]* # optional whitespace before attribute name
(?:(?<=['"\s/])[^\s/>][^\s/=>]* # attribute name
(?:\s*=+\s* # value indicator
(?:'[^']*' # LITA-enclosed value
|"[^"]*" # LIT-enclosed value
|(?!['"])[^>\s]* # bare value
)
\s* # possibly followed by a space
)?(?:\s|/(?!>))*
)*
)?
\s* # trailing whitespace
z#\s*([a-zA-Z][-.a-zA-Z0-9:_]*)\s*>c ó² — e Zd ZdZdZddœd„Zd„ Zd„ Zd„ Zd Z d
„ Z
d„ Zd„ Zd
„ Z
d„ Zdd„Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd„ Zd S ) r aE Find tags and other markup and call handler functions.
Usage:
p = HTMLParser()
p.feed(data)
...
p.close()
Start tags are handled by calling self.handle_starttag() or
self.handle_startendtag(); end tags by self.handle_endtag(). The
data between tags is passed from the parser to the derived class
by calling self.handle_data() with the data as argument (the data
may be split up in arbitrary chunks). If convert_charrefs is
True the character references are converted automatically to the
corresponding Unicode character (and self.handle_data() is no
longer split in chunks), otherwise they are passed by calling
self.handle_entityref() or self.handle_charref() with the string
containing respectively the named or numeric reference as the
argument.
)ÚscriptÚstyleT)Úconvert_charrefsc ó<