01091nas a2200193 4500000000100000000000100001008004100002260003200043100001400075700001900089700002100108700002200129700001400151700001300165245002200178490000800200520066700208020002200875 2012 d bSpringeraBerlin Heidelberg1 aP. Pawlas1 aAdam Domański1 aJoanna Domańska1 aAndrzej Kwiecień1 aPiotr Gaj1 aP. Stera00aComputer Networks0 v2913 aThis article describes the universal web pages content parser-cross-platform application enhancing the process of data extraction from the web pages. In this implementation user friendly interface, possibility of significant automation and reusability of already created patterns had been the key elements. Moreover, the original approach to the issue of parsing the not well-formed HTML, stating the application`s core, is precisely presented. Universal web pages content parser shows that the simplified web scrapping utility may be available to masses and not well-formed HTML sources may feed useful tree-like data structures as well as the well-formed ones. a978-3-642-31216-8