What HTML parsing libraries do you recommend in Java

html, html-content-extraction, java, parsing

Solution

NekoHTML, TagSoup, and JTidy will allow you to parse HTML and then process with XML tools, like XPath.

Problem

I want to parse some HTML in order to find the values of some attributes/tags etc. What HTML parsers do you recommend? Any pros and cons?

Original source