Unicode Utilities: Character Properties

help | character | properties | confusables | unicode-set | compare-sets | regex | bnf-regex | breaks | transform | bidi | bidi-c | idna | languageid


 ‐ 
2010
HYPHEN
Dash Punctuation
confuse: , , , , , , , , , ,
Normative, Informative, Contributory, and (Provisional) UCD properties for U+2010
AgeV1_1
AlphabeticNo
ASCII_Hex_DigitNo
Bidi_ClassOther_Neutral
Bidi_ControlNo
Bidi_MirroredNo
Bidi_Mirroring_Glyphnull
Bidi_Paired_Bracketnull
Bidi_Paired_Bracket_TypeNone
BlockGeneral_Punctuation
Canonical_Combining_ClassNot_Reordered
Case_Folding‐ <U+2010>
Case_IgnorableNo
CasedNo
Changes_When_CasefoldedNo
Changes_When_CasemappedNo
Changes_When_LowercasedNo
Changes_When_NFKC_CasefoldedNo
Changes_When_TitlecasedNo
Changes_When_UppercasedNo
Composition_ExclusionNo
DashYes
Decomposition_Mapping‐ <U+2010>
Decomposition_TypeNone
Default_Ignorable_Code_PointNo
DeprecatedNo
DiacriticNo
East_Asian_WidthAmbiguous
EmojiNo
Emoji_ComponentNo
Emoji_ModifierNo
Emoji_Modifier_BaseNo
Emoji_PresentationNo
Equivalent_Unified_Ideographnull
Expands_On_NFCNo
Expands_On_NFDNo
Expands_On_NFKCNo
Expands_On_NFKDNo
Extended_PictographicNo
ExtenderNo
FC_NFKC_Closure‐ <U+2010>
Full_Composition_ExclusionNo
General_CategoryDash_Punctuation
Grapheme_BaseYes
Grapheme_Cluster_BreakOther
Grapheme_ExtendNo
Grapheme_LinkNo
Hangul_Syllable_TypeNot_Applicable
Hex_DigitNo
HyphenYes
ID_Compat_Math_ContinueNo
ID_Compat_Math_StartNo
ID_ContinueNo
ID_StartNo
IdeographicNo
IDS_Binary_OperatorNo
IDS_Trinary_OperatorNo
IDS_Unary_OperatorNo
Indic_Conjunct_BreakNone
Indic_Positional_CategoryNot_Applicable
Indic_Syllabic_CategoryConsonant_Placeholder
ISO_Commentnull
Jamo_Short_Namenull
Join_ControlNo
Joining_GroupNo_Joining_Group
Joining_TypeNon_Joining
Line_BreakUnambiguous_Hyphen
Logical_Order_ExceptionNo
LowercaseNo
Lowercase_Mapping‐ <U+2010>
MathNo
Modifier_Combining_MarkNo
NameHYPHEN
Name_Aliasnull
NFC_Quick_CheckYes
NFD_Quick_CheckYes
NFKC_Casefold‐ <U+2010>
NFKC_Quick_CheckYes
NFKC_Simple_Casefold‐ <U+2010>
NFKD_Quick_CheckYes
Noncharacter_Code_PointNo
Numeric_TypeNone
Numeric_ValueNaN
Other_AlphabeticNo
Other_Default_Ignorable_Code_PointNo
Other_Grapheme_ExtendNo
Other_ID_ContinueNo
Other_ID_StartNo
Other_LowercaseNo
Other_MathNo
Other_UppercaseNo
Pattern_SyntaxYes
Pattern_White_SpaceNo
Prepended_Concatenation_MarkNo
Quotation_MarkNo
RadicalNo
Regional_IndicatorNo
ScriptCommon
Script_ExtensionsCommon
Sentence_BreakOther
Sentence_TerminalNo
Simple_Case_Folding‐ <U+2010>
Simple_Lowercase_Mapping‐ <U+2010>
Simple_Titlecase_Mapping‐ <U+2010>
Simple_Uppercase_Mapping‐ <U+2010>
Soft_DottedNo
Terminal_PunctuationNo
Titlecase_Mapping‐ <U+2010>
Unicode_1_Namenull
Unified_IdeographNo
UppercaseNo
Uppercase_Mapping‐ <U+2010>
Variation_SelectorNo
Vertical_OrientationRotated
White_SpaceNo
Word_BreakOther
XID_ContinueNo
XID_StartNo
Non-UCD properties for U+2010
Basic_EmojiNo
Identifier_StatusAllowed
Identifier_TypeInclusion
IDNA2008_CategoryDisallowed
Link_Bracketnull
Link_EmailNo
Link_TermInclude
Math_ClassPunctuation
RGI_EmojiNo
RGI_Emoji_Flag_SequenceNo
RGI_Emoji_Keycap_SequenceNo
RGI_Emoji_Modifier_SequenceNo
RGI_Emoji_QualificationNone
RGI_Emoji_Tag_SequenceNo
RGI_Emoji_Zwj_SequenceNo
Other UCD data for U+2010
Arabic_Shaping_Schematic_Namenull
CJK_Radicalnull
Do_Not_Emit_Dispreferrednull
Do_Not_Emit_Dispreferred_TypeNone
Do_Not_Emit_Preferrednull
Do_Not_Emit_TypeNone
Emoji_DCMnull
Emoji_KDDInull
Emoji_SBnull
Emoji_Variation_BaseNo
emoji_variation_sequencenull
Name_Alias_Abbreviationnull
Name_Alias_Alternatenull
Name_Alias_Controlnull
Name_Alias_Correctionnull
Name_Alias_Figmentnull
Named_Sequencesnull
Names_List_Aliasnull
Names_List_Block_HeaderGeneral Punctuation
Names_List_Commentnull
Names_List_Cross_Ref- <U+002D>|⁠­ <U+00AD>
Names_List_Formal_Aliasnull
Names_List_NameHYPHEN
Names_List_SubheaderDashes
Names_List_Subheader_Noticenull
Non_Unihan_Numeric_ValueNaN
normalization_correction_correctednull
normalization_correction_originalnull
normalization_correction_versionnull
Other_Joining_TypeDeduce_From_General_Category
Pretty_BlockGeneral Punctuation
Standardized_Variantnull
Standardized_Variation_BaseNo
Other information on U+2010
ANYYes
ASCIINo
bmpYes
Confusable_MA- <U+002D>
exemplar
exemplar_aux
exemplar_punctaf|⁠agq|⁠ak|⁠am|⁠ar|⁠as|⁠asa|⁠ast|⁠az|⁠az-Cyrl|⁠ba|⁠bas|⁠bem|⁠bez|⁠bg|⁠bho|⁠blo|⁠bm|⁠bn|⁠bo|⁠br|⁠brx|⁠bs|⁠bs-Cyrl|⁠bua|⁠ca|⁠ccp|⁠ce|⁠cgg|⁠chr|⁠ckb|⁠cs|⁠csw|⁠cv|⁠cy|⁠da|⁠dav|⁠de|⁠dje|⁠dsb|⁠dua|⁠dyo|⁠dz|⁠ebu|⁠ee|⁠el|⁠en|⁠eo|⁠es|⁠eu|⁠ewo|⁠fa|⁠ff|⁠fi|⁠fil|⁠fo|⁠fr|⁠fur|⁠fy|⁠ga|⁠gaa|⁠gd|⁠gl|⁠gsw|⁠gu|⁠guz|⁠gv|⁠haw|⁠he|⁠hi|⁠hi-Latn|⁠hr|⁠hsb|⁠ia|⁠id|⁠ie|⁠ii|⁠is|⁠ja|⁠jmc|⁠jv|⁠ka|⁠kab|⁠kam|⁠kde|⁠kea|⁠kgp|⁠khq|⁠ki|⁠kk|⁠kk-Arab|⁠kl|⁠kln|⁠kn|⁠ko|⁠kok|⁠kok-Latn|⁠ks|⁠ksb|⁠ksf|⁠ksh|⁠ku|⁠kw|⁠kxv|⁠ky|⁠lag|⁠lb|⁠lg|⁠lij|⁠lkt|⁠ln|⁠lo|⁠lrc|⁠lt|⁠lu|⁠luo|⁠luy|⁠lv|⁠mas|⁠mer|⁠mfe|⁠mg|⁠mgh|⁠mi|⁠mk|⁠ml|⁠mn|⁠mni|⁠mr|⁠ms|⁠mua|⁠my|⁠mzn|⁠naq|⁠nd|⁠nds|⁠nl|⁠nmg|⁠nso|⁠nus|⁠nyn|⁠oc|⁠om|⁠or|⁠os|⁠pa|⁠pa-Arab|⁠pcm|⁠pl|⁠prg|⁠pt|⁠qu|⁠rm|⁠rn|⁠ro|⁠rof|⁠ru|⁠rw|⁠rwk|⁠saq|⁠sat|⁠sbp|⁠sc|⁠scn|⁠sd-Deva|⁠se|⁠seh|⁠ses|⁠sg|⁠shi|⁠shn|⁠si|⁠sk|⁠smn|⁠sn|⁠so|⁠sq|⁠sr|⁠sr-Latn|⁠st|⁠su|⁠sv|⁠syr|⁠szl|⁠ta|⁠teo|⁠tg|⁠th|⁠ti|⁠tn|⁠to|⁠tr|⁠tt|⁠twq|⁠tyv|⁠tzm|⁠uz|⁠uz-Arab|⁠uz-Cyrl|⁠vai|⁠vai-Latn|⁠vec|⁠vi|⁠vmw|⁠vun|⁠wae|⁠wo|⁠xh|⁠xog|⁠yav|⁠yi|⁠yo|⁠yrl|⁠yue|⁠yue-Hans|⁠za|⁠zgh|⁠zh-Hant|⁠zu
HanTypena
Idn_2008NV8
Idn_Mapping‐ <U+2010>
Idn_Statusvalid
idna2003valid
idna2008cdisallowed
isNFCYes
isNFDYes
isNFKCYes
isNFKDYes
isNFMYes
Math_Class_ExPunctuation
Math_Descriptive_CommentsISOPUB
Math_Entity_Name‐
Math_Entity_Sethyphen
toIdna2003null
toNFC‐ <U+2010>
toNFD‐ <U+2010>
toNFKC‐ <U+2010>
toNFKD‐ <U+2010>
toNFM‐ <U+2010>
toUts46nnull
toUts46tnull
uca0514
uca205
uca2.501
uca305

The list includes both Unicode Character Properties and some additions (like idna2003 or subhead)