4 ms·
Exact same thing. Odd unsolved memory leak, using lxml for HTML parsing. I think I'll stay away from it in the future. Unfortunate, since it's the fastest HTML
by rudolf0 10y ago
Exact same thing. Odd unsolved memory leak, using lxml for HTML parsing.
I think I'll stay away from it in the future. Unfortunate, since it's the fastest HTML parser in Python that I know of.
- maxerickson 10y agoHave you compared it to the C implementation of ElementTree? With stanzas like for event, n in ElementTree.iterparse(xmlfile, events=("end",)): if n.tag == "blar": # do stuff n.clear() it doesn't even take much memory to parse files of arbitrary size.