Package0.14.8XML
tagsoup
Parsing and extracting information from (possibly malformed) HTML/XML documents
- Version0.14.8
- CategoryXML
- LicenceBSD-3-Clause
- AuthorNeil Mitchell <ndmitchell@gmail.com>
- MaintainerNeil Mitchell <ndmitchell@gmail.com>
- Homepagegithub.com/ndmitchell/tagsoup#readme
- Pinned byhackage tagsoup 0.14.8
- Sourcehackage.haskell.org/package/tagsoup-0.14.8
Modules
5 modules- Text.HTML.TagSoup34This module is for working with HTML/XML. It deals with both well-formed XML and
- Text.HTML.TagSoup.Entity6This module converts between HTML/XML entities (i.e. &) and
- Text.HTML.TagSoup.Match17Combinators to match tags. Some people prefer to use (~==) from
- Text.HTML.TagSoup.Tree11NOTE: This module is preliminary and may change at a future date. This module is intended to help converting a list of tags into a
- Text.StringLike3WARNING: This module is not intended for use outside the TagSoup library. This module provides an abstraction for String's as used inside…
Description
TagSoup is a library for parsing HTML/XML. It supports the HTML 5 specification, and can be used to parse either well-formed XML, or unstructured and malformed HTML from the web. The library also provides useful functions to extract information from an HTML document, making it ideal for screen-scraping.
Users should start from the Text.HTML.TagSoup module.
Depends on
4 packages- base-4.20.2.0with GHC
- bytestring-0.12.2.0with GHC
- containers-0.7with GHC
- text-2.1.3with GHC