Universal Encoding Detector https://github.com/chardet/chardet
Go to file
2024-01-06 10:57:06 +01:00
python-chardet.spec rebuilt with -py36 support [release 3.0.4-2mamba;Fri Mar 15 2019] 2024-01-06 10:57:06 +01:00
README.md automatic version update by autodist [release 2.2.1-1mamba;Fri Jan 17 2014] 2024-01-06 10:57:06 +01:00

python-chardet

Character encoding auto-detection in Python 2 and 3. As smart as your browser. Open source.

Detects:

  • ASCII, UTF-8, UTF-16 (2 variants), UTF-32 (4 variants)
  • Big5, GB2312, EUC-TW, HZ-GB-2312, ISO-2022-CN (Traditional and Simplified Chinese)
  • EUC-JP, SHIFT_JIS, ISO-2022-JP (Japanese)
  • EUC-KR, ISO-2022-KR (Korean)
  • KOI8-R, MacCyrillic, IBM855, IBM866, ISO-8859-5, windows-1251 (Cyrillic)
  • ISO-8859-2, windows-1250 (Hungarian)
  • ISO-8859-5, windows-1251 (Bulgarian)
  • windows-1252 (English)
  • ISO-8859-7, windows-1253 (Greek)
  • ISO-8859-8, windows-1255 (Visual and Logical Hebrew)
  • TIS-620 (Thai)