File: PKG-INFO

package info (click to toggle)
python-tidylib 0.3.2~dfsg-9
  • links: PTS, VCS
  • area: main
  • in suites: forky, sid, trixie
  • size: 160 kB
  • sloc: python: 411; makefile: 16; sh: 7
file content (68 lines) | stat: -rw-r--r-- 2,901 bytes parent folder | download | duplicates (5)
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
Metadata-Version: 1.1
Name: pytidylib
Version: 0.3.2
Summary: Python wrapper for HTML Tidy (tidylib) on Python 2 and 3
Home-page: http://countergram.com/open-source/pytidylib/
Author: Jason Stitt
Author-email: js@jasonstitt.com
License: UNKNOWN
Description: `PyTidyLib`_ is a Python package that wraps the `HTML Tidy`_ library. This
        allows you, from Python code, to "fix" invalid (X)HTML markup. Some of the
        library's many capabilities include:
        
        * Clean up unclosed tags and unescaped characters such as ampersands
        * Output HTML 4 or XHTML, strict or transitional, and add missing doctypes
        * Convert named entities to numeric entities, which can then be used in XML
          documents without an HTML doctype.
        * Clean up HTML from programs such as Word (to an extent)
        * Indent the output, including proper (i.e. no) indenting for ``pre`` elements,
          which some (X)HTML indenting code overlooks.
        
        Changes
        =======
        
        * 0.3.2: Initialization bug fix
        
        * 0.3.1: find_library support while still allowing a list of library names
        
        * 0.3.0: Refactored to use Tidy and PersistentTidy classes while keeping the
        functional interface (which will lazily create a global Tidy() object) for
        backward compatibility. You can now pass a list of library names and base
        options when instantiating Tidy. The keep_doc argument is now deprecated
        and does nothing; use PersistentTidy.
        
        * 0.2.4: Bugfix for a strange memory allocation corner case in Tidy.
        
        * 0.2.3: Python 3 support (2 + 3 cross compatible) with passing Tox tests.
        
        Small example of use
        ====================
        
        The following code cleans up an invalid HTML document and sets an option::
        
            from tidylib import tidy_document
            document, errors = tidy_document('''<p>f&otilde;o <img src="bar.jpg">''',
              options={'numeric-entities':1})
            print document
            print errors
        
        Docs
        ====
        
        Documentation is shipped with the source distribution and is available at
        the `PyTidyLib`_ web page.
        
        .. _`HTML Tidy`: http://tidy.sourceforge.net/
        .. _`PyTidyLib`: http://countergram.com/open-source/pytidylib/
        
Platform: UNKNOWN
Classifier: Development Status :: 5 - Production/Stable
Classifier: Environment :: Other Environment
Classifier: Intended Audience :: Developers
Classifier: License :: OSI Approved :: MIT License
Classifier: Programming Language :: Python
Classifier: Programming Language :: Python :: 3
Classifier: Natural Language :: English
Classifier: Topic :: Utilities
Classifier: Topic :: Text Processing :: Markup :: HTML
Classifier: Topic :: Text Processing :: Markup :: XML