l don’t undersand why we Iive in a worId where capital I and Iowercase l are allowed to Iook near identical. PeopIe are paid to design and seIect fonts and they have aII agreed on this Iudicrous convention.

Why has this happened? And is there anything I can do about it as an individuaI or as part of a coIIective?

I can’t even do thorn-master Sxan’s thing and insert a character code because U+026A, ɪ, is too damn small.

  • ChaoticNeutralCzech@feddit.org
    link
    fedilink
    English
    arrow-up
    7
    ·
    edit-2
    6 hours ago

    Given the confused nature of your post, best look up Unicode, serif and sans-serif before reading on…

    First line is mostly Cherokee script, whose alphabet goes ᎠᎡᎢᎣᎤᎥᎦᎧᎨᎱᎪᎫᎬᎭᎮᎯ… Many letters resemble English serif ones: ᎪᏴᏟᎠᎬᎰᏀᎻᎥᎫᏦᏞᎷᎨᎾᏢᏩᎡᏚᎢᏌᏙᎳᏯᎩᏃ. With some clever synonym selection (the fuck the hell) and a replacement for the compromised ones like O or (unfortunately also) I, one can use it as a serif typeface wiith small caps.

    The only uppercase, explicitly serif I in Unicode are “special fonts” (Latin, sometimes Greek letter sets) accessible via an online converter or Android app. Some are meant for formulas where serif is the norm, therefore there’s

    • 𝖬𝖺𝗍𝗁 𝗌𝖺𝗇𝗌
      • may look like Helvetica or indistinguishable from the surrounding sans(-serif) font
    • 𝗠𝗮𝘁𝗵 𝘀𝗮𝗻𝘀 𝗯𝗼𝗹𝗱
    • 𝘔𝘢𝘵𝘩 𝘴𝘢𝘯𝘴 𝘪𝘵𝘢𝘭𝘪𝘤
    • 𝙈𝙖𝙩𝙝 𝙨𝙖𝙣𝙨 𝙗𝙤𝙡𝙙 𝙞𝙩𝙖𝙡𝙞𝙘
    • 𝐌𝐚𝐭𝐡 𝐛𝐨𝐥𝐝 𝐬𝐞𝐫𝐢𝐟
      • closest to forced serif typeface: as mentioned above, no “Math serif” exists
    • 𝑀𝑎𝑡ℎ 𝑖𝑡𝑎𝑙𝑖𝑐 𝑠𝑒𝑟𝑖𝑓
      • the “h” came first (Planck constant, always italics) and looks different in some fonts
    • 𝑴𝒂𝒕𝒉 𝒃𝒐𝒍𝒅 𝒊𝒕𝒂𝒍𝒊𝒄 𝒔𝒆𝒓𝒊𝒇
    • ℳ𝒶𝓉𝒽 𝓈𝒸𝓇𝒾𝓅𝓉
    • 𝓜𝓪𝓽𝓱 𝓼𝓬𝓻𝓲𝓹𝓽 𝓫𝓸𝓵𝓭
    • 𝔐𝔞𝔱𝔥 𝔉𝔯𝔞𝔨𝔱𝔲𝔯
    • 𝕸𝖆𝖙𝖍 𝕱𝖗𝖆𝖐𝖙𝖚𝖗 𝕭𝖔𝖑𝖉
    • 𝙼𝚊𝚝𝚑 𝚖𝚘𝚗𝚘𝚜𝚙𝚊𝚌𝚎
    • 𝕄𝕒𝕥𝕙 𝕕𝕠𝕦𝕓𝕝𝕖-𝕤𝕥𝕣𝕠𝕜𝕖
      • completing the alphabet of letters denoting number sets like ℕ, ℤ, ℚ, ℝ, ℂ…

    Then, there’s

    • Fullwidth
      • width of CJK characters, uniquely includes some punctuation
    • Ⓒⓘⓡⓒⓛⓔⓓ
    • 🅝🅔🅖🅐🅣🅘🅥🅔 🅒🅘🅡🅒🅛🅔🅓
    • 🅂🅀🅄🄰🅁🄴🄳
    • 🅽🅴🅶🅰🆃🅸🆅🅴 🆂🆀🆄🅰🆁🅴🅳
      • A, B and O (blood types) may be red (emoji presentation)
    • 🇷​🇪​🇬​🇮​🇴​🇳​🇦​🇱​ 🇮​🇳​🇩​🇮​🇨​🇦​🇹​🇴​🇷​🇸​
      • “flag parts”, vary wildly by platform, usually blue

    And ᴘsᴇᴜᴅᴏᴀʟᴘʜᴀʙᴇᴛs ʟɪᴋᴇ sᴍᴀʟʟ ᴄᴀᴘs ʜᴇʀᴇ, made by repurposing other characters. Often incomplete, like my Cherokee attempt at serif. Explained on the QAZ website linked above.

    Also, I used

    • U+2D4A ⵊ TIFINAGH LETTER YAZH
    • U+10833 𐠳 CYPRIOT SYLLABLE WE
    • U+1066B 𐙫 LINEAR A SIGN A319
    • U+102A6 𐊦 CARIAN LETTER LD
    • U+0196 Ɩ LATIN CAPITAL LETTER IOTA
      • unlike the above 4 Is, almost guaranteed to be supported by the default font on any reasonably modern platform

    I’m looking forward to the year 𑃦𑃰𑃟𑃰/𞄡𞅀𞄬𞅀/𞄕𞅀𞄧𞅀, the next one whose number can be approximated in a whimsical way by characters in the Sora Sompeng and Nyiakeng Puachue Hmong scripts, respectively. The number of devices supporting them will have grown by then, too.

    • [object Object]@lemmy.world
      link
      fedilink
      arrow-up
      5
      ·
      15 hours ago

      For yall’s information, nothing typed with those characters will be searchable in the browser, an app, or via website’s search — unless apps and websites take special measures to account for such tricks, by keeping a copy of the text normalized to the regular English alphabet.

      /cc @schipelblorp@sh.itjust.works

      • ChaoticNeutralCzech@feddit.org
        link
        fedilink
        English
        arrow-up
        2
        ·
        edit-2
        6 hours ago

        Or accessible via screen reader…

        The “normalization” of pseudoalphabets is tricky, the normal transliteration of the “alphabet” I made with Cherokee syllabics is “Go⁠yv⁠tli⁠a⁠gv⁠ho⁠nah⁠mi⁠v⁠gu⁠tso⁠tle⁠lu⁠ge⁠na⁠tlv⁠wa⁠e⁠du⁠i⁠sa⁠do⁠la⁠ya⁠gi⁠no”. (See also Faux Cyrillic (such as СУЯЇLLЇС, actually pronounced “suyaillis”) and !grssk@lemmy.world (faux Greek, such as GRΣΣΚ, actually pronounced “grssk”) for other misused alphabets.) It is necessary to maintain a database of homoglyphs and other lookalikes for proper filtration of τ℮אϯ Ιίᛕᥱ ᵼℎⅰ⟆.

        Also, commercial platforms know it’s a way to bypass filters and tend to outright block comments/emails/etc. that use Unicode in unintended ways.