#unicode

10 posts · Last used 23d

Back to Timeline
W3C Developers @w3cdevs@w3c.social · Jul 20, 2026
Unicode strings that look the same aren't always the same. The @w3c@w3c.social's "Character Model for String Matching" document explains how case mapping, case folding, #emoji and other #Unicode details affect #interoperability and why specifications should define matching rules carefully #timetogiveinput #FPWD ▶️ https://www.w3.org/TR/charmod-norm/ The goal of this document is to provide a common reference for specification authors, software #developers, and content creators. Feedback wlc: https://github.com/w3c/charmod-norm/issues
6
0
10
Martin Edwards @medwds@infosec.exchange · Jul 17, 2026
17 July is the one day of the year when the calendar #emoji is correct! The word emoji combines the Japanese e (絵, picture) + moji (文字, character) and dates back to the late 80s. Its similarity to the English "emotion" is just a coincidence. Nowadays, emoji are standardised by the Unicode Consortium, a Californian non-profit that has its roots in facilitating a universal digital representation of the world's writing systems. #Unicode doesn't specify exactly how emoji should look – so they're different on Android and Apple devices, for example – but it lists their names. The first major character set used in computing was #ASCII, the American Standard Code for Information Interchange, published in 1963. Using 7 bits, it could represent 128 characters — more than enough for a to z, A to Z, 0 to 9 and common punctuation marks, but certainly indicative of a North American worldview. Later "extended" versions of ASCII doubled this to 256 possibilities, allowing for dozens of accented letters, characters like æ and ß, and symbols like © and ¾. And double-byte character sets like Shift JIS meant that Chinese, Korean and Japanese could be represented in full. But multiple standards existed, and it became just a bit too common to see ����. The answer was Unicode: a single, multi-byte character encoding standard for everyone. Plus emoji! #language #i18n
0
0
0
Leeloo @leeloo@c.im · Jul 16, 2026
Replying to @0xabad1dea@infosec.exchange
@0xabad1dea@infosec.exchange Yesh, that's definitely what you do if you find a random snake in a hotel. Stuff it in your bra and try to sneak it through airport security/customs. Hey, #Unicode geeks, where's the facepalm emoji?
0
1
0
James Widman @JamesWidman@mastodon.social · Jul 15, 2026
a nice unexpected benefit of godbolt.org is that it's making me aware of more system languages. today, i learned about the C3 language via the godbolt language pull-down menu. for prototyping purposes (not for production, which requires memory safety), C3 seemed to check a lot of boxes. but then i got to the part where it milkshake-ducks itself: https://c3-lang.org/faq/rejected-ideas/#unicode-identifiers they might as well have asked, "why would anyone think in their native language, even in their personal experimental code?"
1
5
0
Jim DeLaHunt @jdlh@mstdn.ca · Apr 27, 2026
Replying to @fediforum@mastodon.social
@Profpatsch@mastodon.xyz You "added #Unicode username support to #flohmarkt" last month — wonderful to hear! I will look for your demo at #FediForum. I will propose a session on "#Globally-inclusive Fediverse handles". If you were present, it would improve the session. #UniversalAcceptance #Multilingual
1
0
6
Replying to @liilliil@flipboard.social
Оу, у нас же в #unicode есть четырёхстрочный/восьмиточечный #брайль! Можно его заюзать 2026 → ⠶⠘
0
0
0

You've seen all posts