UTF-7
Character encoding
UTF-7 (7-bit Unicode Transformation Format) is an obsolete variable-length character encoding for representing Unicode text using a stream of ASCII characters. It was originally intended to provide a means of encoding Unicode text for use in Internet E-mail messages that was more efficient than the combination of UTF-8 with quoted-printable.
Nº Q276826 ★
Common · Literature
UTF-7
Character encoding
UTF-7 (7-bit Unicode Transformation Format) is an obsolete variable-length character encoding for representing Unicode text using a stream of ASCII characters. It was originally intended to provide a means of encoding Unicode text for use in Internet E-mail messages that was more efficient than the combination of UTF-8 with quoted-printable.
From Wikipedia
UTF-7 (7-bit Unicode Transformation Format) is an obsolete variable-length character encoding for representing Unicode text using a stream of ASCII characters. It was originally intended to provide a means of encoding Unicode text for use in Internet E-mail messages that was more efficient than the combination of UTF-8 with quoted-printable. UTF-7 (according to its RFC) isn't a "Unicode Transformation Format", as the definition can only encode code points in the BMP (the first 65536 Unicode code points, which does not include emojis and many other characters). However if a UTF-7 translator is to/from UTF-16 then it can (and probably does) encode each surrogate half as though it was a 16-bit code point, and thus can encode all code points. It is unclear if other UTF-7 software (such as translators to UTF-32 or UTF-8) support this. UTF-7 has never been an official standard of the Unicode Consortium. It is known to have security issues, which is why software has been changed to disable its use. It is prohibited in HTML 5.
Text: Wikipédia, CC BY-SA 4.0. ·
Related cards
-
U
UTF-8
Variable-width encoding (into one to four bytes) and transformation format of code points for the universal character set defined by ISO/IEC 10646 and The Unicode® Standard, compatible with ASCII
Nº Q193537 ★★★★
Not listed
-
Base64
Group of binary-to-text encoding schemes using 64 symbols (plus padding)
Nº Q726780 ★★★★
Not listed
-
U
UTF-16
Format for transforming texts encoded with one or two 16-bit code units per character of the universal character set defined by the ISO/IEC 10646 standard and the Unicode Standard
Nº Q740701 ★★★
Not listed
-
U
UTF-32
Format for transforming texts encoded with 4 bytes per code point in the universal character set defined by ISO/IEC 10646 and the Unicode Standard
Nº Q736068 ★
Not listed
-
Unicode
Computing industry standard for the consistent encoding, representation and handling of text expressed in most of the world's writing systems
Nº Q8819 ★★★★
Not listed
-
Byte
Unit of digital information equal to 8 bits
Nº Q8799 ★★★★
Not listed