KOI8-F

KOI8-F or KOI8 Unified is an 8-bit character set.[1] It was designed by Peter Cassetta[2] of Fingertip Software (now defunct) as an attempt to support all the encoded letters from both KOI8-E (ISO-IR-111) and KOI8-RU (and hence also, KOI8-U and KOI8-R), along with some of the pseudographics from KOI8-R,[3][2] with some additional punctuation in the remaining space, sourced partly from Windows-1251.[2] This encoding was only used in the software of that company.

KOI8 Unified
Alias(es)KOI8-F
Language(s)Belarusian, Ukrainian, Russian, Bulgarian, Serbian Cyrillic, Macedonian
Created byPeter Cassetta (Fingertip Software)
Classification8-bit KOI, extended ASCII
ExtendsKOI8-B
Based onKOI8-RU, KOI8-E
Other related encoding(s)KOI8-R, KOI8-U

Character set

The following table shows the KOI8-F encoding. Each character is shown with its equivalent Unicode code point. Differences from ISO-IR-111 are boxed; other relevant encodings which are matched, if any, are noted in footnotes.

KOI8-F[4]
0 1 2 3 4 5 6 7 8 9 A B C D E F
0x
1x
2x  SP  ! " # $ % & ' ( ) * + , - . /
3x 0 1 2 3 4 5 6 7 8 9 : ; < = > ?
4x @ A B C D E F G H I J K L M N O
5x P Q R S T U V W X Y Z [ \ ] ^ _
6x ` a b c d e f g h i j k l m n o
7x p q r s t u v w x y z { | } ~
8x [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1] [lower-alpha 1]
9x [lower-alpha 1] [lower-alpha 2] [lower-alpha 2] [lower-alpha 2] [lower-alpha 2] title="U+2219 BULLET OPERATOR or U+2022 BULLET" style="padding:0px;background:#FFD"}}|∙/•[lower-alpha 3] [lower-alpha 2] [lower-alpha 2] © [lower-alpha 2] NBSP[lower-alpha 4] » ® « ·[lower-alpha 1] ¤
Ax NBSP[lower-alpha 4] ђ ѓ ё є ѕ і ї ј љ њ ћ ќ ґ[lower-alpha 5] ў џ
Bx Ђ Ѓ Ё Є Ѕ І Ї Ј Љ Њ Ћ Ќ Ґ[lower-alpha 5] Ў Џ
Cx ю а б ц д е ф г х и й к л м н о
Dx п я р с т у ж в ь ы з ш э щ ч ъ
Ex Ю А Б Ц Д Е Ф Г Х И Й К Л М Н О
Fx П Я Р С Т У Ж В Ь Ы З Ш Э Щ Ч Ъ
  Differences from ISO-IR-111
  1. Matching KOI8-R, KOI8-U, KOI8-RU.
  2. Matching Windows-1251 and Windows-1252.
  3. May be U+2219, which matches RFC 1489 (KOI8-R),[4] or U+2022, which matches Windows-1251 and Windows-1252.
  4. The non-breaking space is encoded twice: first at 0x9A matching KOI8-R, and then at 0xA0 matching KOI8-E (the latter of which also happens to be its location in Windows-1251 and Windows-1252).
  5. Matching KOI8-U and KOI8-RU.

See also

References

  1. Nechayev, Valentin (2013) [2001]. "Review of 8-bit Cyrillic encodings universe". Archived from the original on 2016-12-05. Retrieved 2016-12-05.
  2. Czyborra, Roman (1998-11-30) [1998-05-25]. "The Cyrillic Charset Soup". Archived from the original on 2016-12-03. Retrieved 2016-12-03.
  3. "KOI8 Unified". Fingertip Software. Archived from the original on 1998-01-09. Retrieved 2020-02-11.
  4. Leisher, Mark (2008) [1998-03-05]. "KOI8 Unified Cyrillic to Unicode 2.1 mapping table". Department of Mathematical Sciences, New Mexico State University. Retrieved 2020-05-02.
This article is issued from Wikipedia. The text is licensed under Creative Commons - Attribution - Sharealike. Additional terms may apply for the media files.