chardet
Universal encoding detector
Universal character encoding detector ------------------------------------- Detects - ASCII, UTF-8, UTF-16 (2 variants), UTF-32 (4 variants) - Big5, GB2312, EUC-TW, HZ-GB-2312, ISO-2022-CN (Traditional and Simplified Chinese) - EUC-JP, SHIFT_JIS, ISO-2022-JP (Japanese) - EUC-KR, ISO-2022-KR (Korean) - KOI8-R, MacCyrillic, IBM855, IBM866, ISO-8859-5, windows-1251 (Cyrillic) - ISO-8859-2, windows-1250 (Hungarian) - ISO-8859-5, windows-1251 (Bulgarian) - windows-1252 (English) - ISO-8859-7, windows-1253 (Greek) - ISO-8859-8, windows-1255 (Visual and Logical Hebrew) - TIS-620 (Thai) Requires Python 2.1 or laterFilename | Platform | Type | Version | Uploaded On | Size |
---|---|---|---|---|---|
chardet-2.0.1.tar.gz | Windows | Source | 2.0.1 | 2014-11-12 11:05:03 | 148.5 KB |
- License: LGPL
- Author: Mark Pilgrim
- Home Page: http://chardet.feedparser.org/
- Download URL: http://chardet.feedparser.org/download/python2-chardet-2.0.1.tgz
- Classifiers:
- Development Status :: 4 - Beta
- Intended Audience :: Developers
- Operating System :: OS Independent
- Programming Language :: Python
- Topic :: Software Development :: Libraries :: Python Modules
- License :: OSI Approved :: GNU Library or Lesser General Public License (LGPL)
- Programming Language :: Python :: 3
- Environment :: Other Environment
- Topic :: Text Processing :: Linguistic