请输入您要查询的百科知识:

 

词条 Unicode subscripts and superscripts
释义

  1. Uses

  2. Superscripts and subscripts block

  3. Other superscript and subscript characters

  4. Latin and Greek tables

  5. Composite characters

  6. References

{{SpecialChars}}

Unicode has subscripted and superscripted versions of a number of characters including a full set of Arabic numerals.[1] These characters allow any polynomial, chemical and certain other equations to be represented in plain text without using any form of markup like HTML or TeX.

The World Wide Web Consortium and the Unicode Consortium have made recommendations on the choice between using markup and using superscript and subscript characters: "When used in mathematical context (MathML) it is recommended to consistently use style markup for superscripts and subscripts.... However, when super and sub-scripts are to reflect semantic distinctions, it is easier to work with these meanings encoded in text rather than markup, for example, in phonetic or phonemic transcription."[2]

Uses

The intended use[2] when these characters were added to Unicode was to allow chemical and algebra formulas and phonetics to be written without markup, but produce true superscripts and subscripts. Thus "H₂O" (using a subscript character) is supposed to be identical to "H2O" (with subscript markup).

In reality most fonts that include these characters ignore the Unicode definition, and design the digits for mathematical numerator and denominator glyphs, which are smaller than normal characters but are aligned with the cap line and the baseline, respectively. When used with the solidus, these glyphs are useful for making arbitrary diagonal fractions (similar to the ½ glyph). Trying to make fractions using existing software super/subscripts look messier (example: 1/2), so font designers provided this alternative. This also makes the superscript letters useful for ordinal indicators, more closely matching the ª and º characters. However it makes them incorrect for normal super and subscripts, and generally formulas look better using markup than these characters.

Unicode intended to produce diagonal fractions through a different mechanism but it is very poorly supported. The fraction slash U+2044 is visually similar to the solidus, but when used with the ordinary digits (not the superscripts and subscripts) is intended to tell a layout system that a fraction such as ¾ should be rendered[3] using automatic glyph substitution[4] for the digits. Some browsers support this[5] but not in all fonts, a selection of fonts is shown in the below table.

Characters Font Result
bd|Vulgar Fraction One Half}}Default}} ½
b9|Superscript one}}, {{unichar|2f|Solidus}}, {{unichar|2082|Subscript two}} ¹/₂
b9|Superscript one}}, {{unichar|2044|Fraction slash}}, {{unichar|2082|Subscript two}} ¹⁄₂
{{unichar|31|Digit one}}, {{unichar|2044|Fraction slash}}, {{unichar|32|Digit two}} 1⁄2
Arial1⁄2
Cambria1⁄2
Consolas1⁄2
Times New Roman1⁄2

Superscripts and subscripts block

{{main|Superscripts and Subscripts|l1=Superscripts and Subscripts (Unicode block)}}

The most common superscript digits (1, 2, and 3) were in ISO-8859-1 and were therefore carried over into those positions in the Latin-1 range of Unicode. The rest were placed in a dedicated section of Unicode at {{U+|2070}} to U+209F. The two tables below show these characters. Each superscript or subscript character is preceded by a normal x to show the subscripting/superscripting. The table on the left contains the actual Unicode characters; the one on the right contains the equivalents using HTML markup for the subscript or superscript.

>
Unicode characters
}}}}}}}}}}}}}}}}}}}}}}}}}}
x⁰ xⁱ x⁴ x⁵ x⁶ x⁷ x⁸ x⁹ x⁺ x⁻ x⁼ x⁽ x⁾ xⁿ
x₀ x₁ x₂ x₃ x₄ x₅ x₆ x₇ x₈ x₉ x₊ x₋ x₌ x₍ x₎
xₐ xₑ xₒ xₓ xₔ xₕ xₖ xₗ xₘ xₙ xₚ xₛ xₜ
Simulated using <sup> or <sub> tags
}}}} x2 x3 }}}}}}}}}} x1 }}}}}}}}}}}}
x0 xi x4 x5 x6 x7 x8 x9 x+ x x= x( x) xn
x0 x1 x2 x3 x4 x5 x6 x7 x8 x9 x+ x x= x( x)
xa xe xo xx xə xh xk xl xm xn xp xs xt
{{legend|silver|outline=#aaa|Reserved for future use.}}{{legend|#ececec|outline=#aaa|Other characters from Latin-1 not related to super- or sub-scripts.}}

Other superscript and subscript characters

Unicode version 11.0 also includes subscript and superscript characters that are intended for semantic usage, in the following blocks:[1][6]

  • The Latin-1 Supplement block contains the feminine and masculine ordinal indicators ª and º.
  • The Latin Extended-C block contains one additional superscript, ⱽ, and one additional subscript ⱼ.
  • The Latin Extended-D block contains three superscripts: ꝰ ꟸ ꟹ.
  • The Latin Extended-E block contains four superscripts: ꭜ ꭝ ꭞ ꭟ.
  • The Combining Diacritical Marks block contains medieval superscript letter diacritics. These letters are written directly above other letters appearing in medieval Germanic manuscripts, and so these glyphs do not include spacing, for example uͤ. They are shown here over the dotted circle placeholder ◌: ◌ͣ ◌ͤ ◌ͥ ◌ͦ ◌ͧ ◌ͨ ◌ͩ ◌ͪ ◌ͫ ◌ͬ ◌ͭ ◌ͮ ◌ͯ.
  • The Combining Diacritical Marks Supplement block contains additional medieval superscript letter diacritics, enough to complete the basic lowercase Latin alphabet except for j, q and y, a few small capitals and ligatures (ae, ao, av), and additional letters: ◌ᷓ ◌ᷔ ◌ᷕ ◌ᷖ ◌ᷗ ◌ᷘ ◌ᷙ ◌ᷚ ◌ᷛ ◌ᷜ ◌ᷝ ◌ᷞ ◌ᷟ ◌ᷠ ◌ᷡ ◌ᷢ ◌ᷣ ◌ᷤ ◌ᷥ ◌ᷦ ◌ᷧ ◌ᷨ ◌ᷩ ◌ᷪ ◌ᷫ ◌ᷬ ◌ᷭ ◌ᷮ ◌ᷯ ◌ᷰ ◌ᷱ ◌ᷲ ◌ᷳ ◌ᷴ. There is also a combining subscript: ◌᷊.
  • The Spacing Modifier Letters block has superscripted letters and symbols used for phonetic transcription: ʰ ʱ ʲ ʳ ʴ ʵ ʶ ʷ ʸ ˀ ˁ ˠ ˡ ˢ ˣ ˤ.
  • The Phonetic Extensions block has several sub- and super-scripted letters and symbols: Latin/IPA ᴬ ᴭ ᴮ ᴯ ᴰ ᴱ ᴲ ᴳ ᴴ ᴵ ᴶ ᴷ ᴸ ᴹ ᴺ ᴻ ᴼ ᴽ ᴾ ᴿ ᵀ ᵁ ᵂ ᵃ ᵄ ᵅ ᵆ ᵇ ᵈ ᵉ ᵊ ᵋ ᵌ ᵍ ᵏ ᵐ ᵑ ᵒ ᵓ ᵖ ᵗ ᵘ ᵚ ᵛ ᵢ ᵣ ᵤ ᵥ, Greek ᵝ ᵞ ᵟ ᵠ ᵡ ᵦ ᵧ ᵨ ᵩ ᵪ, Cyrillic ᵸ, other ᵎ ᵔ ᵕ ᵙ ᵜ. These are intended to indicate secondary articulation.
  • The Phonetic Extensions Supplement block has several more: Latin/IPA ᶛ ᶜ ᶝ ᶞ ᶟ ᶠ ᶡ ᶢ ᶣ ᶤ ᶥ ᶦ ᶧ ᶨ ᶩ ᶪ ᶫ ᶬ ᶭ ᶮ ᶯ ᶰ ᶱ ᶲ ᶳ ᶴ ᶵ ᶶ ᶷ ᶸ ᶹ ᶺ ᶻ ᶼ ᶽ ᶾ, Greek ᶿ.
  • The Cyrillic Extended-B block contains two Cyrillic superscripts: ꚜ ꚝ.
  • The Cyrillic Extended-A and -B blocks contains multiple medieval superscript letter diacritics, enough to complete the basic lowercase Cyrillic alphabet used in Church Slavonic texts, also includes an additional ligature (ст): ◌ⷠ ◌ⷡ ◌ⷢ ◌ⷣ ◌ⷤ ◌ⷥ ◌ⷦ ◌ⷧ ◌ⷨ ◌ⷩ ◌ⷪ ◌ⷫ ◌ⷬ ◌ⷭ ◌ⷮ ◌ⷯ ◌ⷰ ◌ⷱ ◌ⷲ ◌ⷳ ◌ⷴ ◌ⷵ ◌ⷶ ◌ⷷ ◌ⷸ ◌ⷹ ◌ⷺ ◌ⷻ ◌ⷼ ◌ⷽ ◌ⷾ ◌ⷿ ◌ꙴ ◌ꙵ ◌ꙶ ◌ꙷ ◌ꙸ ◌ꙹ ◌ꙺ ◌ꙻ ◌ꚞ ◌ꚟ.
  • The Georgian block contains one superscripted Mkhedruli letter: ჼ.
  • The Kanbun block has superscripted annotation characters used in Japanese copies of Classical Chinese texts: ㆒ ㆓ ㆔ ㆕ ㆖ ㆗ ㆘ ㆙ ㆚ ㆛ ㆜ ㆝ ㆞ ㆟.
  • The Tifinagh block has one superscript letter : ⵯ.

Latin and Greek tables

Consolidated, the Unicode standard contains superscript and subscript versions of a subset of Latin and Greek letters. Here they are arranged in order for comparison (or for copy and paste convenience). Since these characters come from different ranges, they may not be of the same size and position, depending on the typeface:

Latin superscript and subscript letters
ABCDEFGHIJKLMNOPQRSTUVWXYZ
Superscript capitalᴿ
Superscript small cap
Superscript minusculeʰʲˡʳˢʷˣʸ
Subscript minuscule
Greek superscript and subscript letters
ΑΒΓΔΕΖΗΘΙΚΛΜΝΞΟΠΡΣΤΥΦΧΨΩ
Superscript minusculeᶿ
Subscript minuscule
other IPA superscript letters
ɐ ɑ ɒ ɔ ɕ ð ə ɜ ɟ ɡ ɦ ɥ ɨ ʝ ɭ ɱ ɯ ɰ ŋ ɲ ɳ ɵ œ ɹ ɻ ʁ ʂ ʃ ƫ ʉ ʊ ʋ ʌ ɣ ʐ ʑ ʒ ɸ ʔ ʕ
ʱ ʴ ʵ ʶ ˠ ˀ ˁ,ˤ

See also small caps in Unicode.

Composite characters

Primarily for compatibility with earlier character sets, Unicode contains a number of characters that compose super- and subscripts with other symbols.[1] In most fonts these render much better than attempts to construct these symbols from the above characters or by using markup.

  • The Latin-1 Supplement block contains the precomposed fractions ½, ¼, and ¾. The copyright © and registered trademark signs ® are also in this block.
  • The General Punctuation block contains the permille sign ‰ and the per-ten-thousand sign ‱, and Basic Latin has the percent sign %.
  • The Number Forms block contains several precomposed fractions: ⅐ ⅑ ⅒ ⅓ ⅔ ⅕ ⅖ ⅗ ⅘ ⅙ ⅚ ⅛ ⅜ ⅝ ⅞ ⅟ ↉.
  • The Letterlike Symbols block contains a few symbols composed of subscript and superscript characters: ℀ ℁ ℅ ℆ № ℠ ™ ⅍.
  • The Enclosed Alphanumeric Supplement block contains three superscript abbreviations 🅪 🅫 🅬: MC for {{lang|fr|marque de commerce}} (trademark), MD for {{lang|fr|marque déposée}} (registered trademark), both used in Canada; MR for marca registrada (registered trademark) in Spanish and Portuguese speaking countries[7]
  • The Miscellaneous Technical block has one additional subscript, a subscript 10 (⏨), for the purpose of scientific notation.

References

{{Portal|Writing}}
1. ^{{cite web|url=https://www.unicode.org/Public/UCD/latest/ucd/UnicodeData.txt|title=UCD: UnicodeData.txt|work=The Unicode Standard|accessdate=2016-05-14}}
2. ^{{cite web |url=http://www.w3.org/TR/unicode-xml/#Superscripts |title=Unicode in XML and other Markup Languages |author=Martin Dürst, Asmus Freytag |date=16 May 2007 |publisher=W3C |accessdate=13 September 2010}}
3. ^{{cite web |url=http://www.w3.org/TR/unicode-xml/#Fraction |title=Fraction Slash |author=Martin Dürst, Asmus Freytag |date=16 May 2007 |publisher=W3C |accessdate=13 September 2010}}
4. ^For a general overview and technical information on glyph substitution (though not specifically for fractions): [https://www.microsoft.com/typography/otspec/gsub.htm GSUB — Glyph Substitution Table] in the [https://www.microsoft.com/typography/otspec/default.htm OpenType specification] on the Microsoft Typography site.
5. ^Such as [https://www.google.com/chrome/index.html Chrome] on Windows, [https://www.mozilla.org/en-US/firefox/new/ Firefox]
6. ^{{cite web|url=https://www.unicode.org/Public/UCD/latest/ucd/Scripts.txt|title=UCD: Scripts.txt|work=The Unicode Standard|accessdate=2017-06-20}}
7. ^{{Cite web|url=http://www.unicode.org/L2/L2017/17066r-marca-registrada.pdf|title=L2/17-066R: Proposal to encode the Marca Registrada sign|date=2017-03-01|first=Eduardo Marín|last=Silva}}
{{Unicode navigation}}{{DEFAULTSORT:Subscripts and superscripts, Unicode}}Unicodeblock Hoch- und tiefgestellte Zeichen

1 : Unicode

随便看

 

开放百科全书收录14589846条英语、德语、日语等多语种百科知识,基本涵盖了大多数领域的百科知识,是一部内容自由、开放的电子版国际百科全书。

 

Copyright © 2023 OENC.NET All Rights Reserved
京ICP备2021023879号 更新时间:2024/11/16 1:01:21