Skip to content
Sections
>> Trisquel >> Packages >> aramo >> libs >> libhtmlcxx3v5
etiona  ] [  nabia  ] [  aramo  ]
[ Source: htmlcxx  ]

Package: libhtmlcxx3v5 (0.87-2)

simple HTML parser library for C++

htmlcxx is a simple non-validating CSS1 and HTML parser for C++. Although there are several other html parsers available, htmlcxx has some characteristics that make it unique:

 * STL like navigation of DOM tree, using excellent tree.hh library from
   Kasper Peeters
 * It is possible to reproduce exactly, character by character, the original
   document from the parse tree
 * Bundled CSS parser
 * Optional parsing of attributes
 * C++ code that looks like C++ (not so true anymore)
 * Offsets of tags/elements in the original document are stored in the nodes
   of the DOM tree

The parsing politics of htmlcxx were created trying to mimic Mozilla Firefox (http://www.mozilla.org) behavior. So you should expect parse trees similar to those create by Firefox. However, differently from Firefox, htmlcxx does not insert non-existent stuff in your html. Therefore, serializing the DOM tree gives exactly the same bytes contained in the original HTML document.

Other Packages Related to libhtmlcxx3v5

  • depends
  • recommends
  • suggests
  • dep: libc6 (>= 2.14) [amd64]
    GNU C Library: Shared libraries
    also a virtual package provided by libc6-udeb
    dep: libc6 (>= 2.17) [arm64, ppc64el]
    dep: libc6 (>= 2.4) [armhf]
  • dep: libgcc-s1 (>= 3.3.1) [not armhf]
    GCC support library
    dep: libgcc-s1 (>= 3.5) [armhf]
  • dep: libstdc++6 (>= 5.2)
    GNU Standard C++ Library v3

Download libhtmlcxx3v5

Download for all available architectures
Architecture Package Size Installed Size Files
amd64 34.2 kB111 kB [list of files]
arm64 32.5 kB103 kB [list of files]
armhf 29.8 kB74 kB [list of files]
ppc64el 41.1 kB151 kB [list of files]