← All Tags

#bidirectional-text

7 episodes

#5488: How Emoji Went From 176 Characters to 3,600

From a 176-character set built for Japanese pagers to a 3,600-character Unicode standard — and why your emoji breaks in production.

unicodebidirectional-textprofessional-communication

#5487: Why Your Spreadsheet Mangles José's Name

A deep dive into character sets, from ASCII to UTF-8, and why José becomes "José" in your spreadsheet.

unicodebidirectional-textdata-integrity

#5486: Why Hebrew URLs Turn Into Percent-Sign Gibberish

Hebrew renders fine in the domain but explodes into percent-hex in the path. Two different standards explain the split.

unicodebidirectional-textinternet-security

#5466: Hebrew Words Hidden in English Text

Daniel wants a classifier that spots Hebrew written in Latin letters — and it turns out nobody's built one.

linguisticsbidirectional-textautomatic-speech-recognition

#5461: TTS Can't Pronounce Hebrew Inside English

Your TTS reads Hebrew words with English phonetics. Here's why — and why the obvious fix doesn't work yet.

text-to-speechbidirectional-textlarge-language-models

#2593: The Politics of Unicode: Paleo-Hebrew, Han Unification, and Who Decides What a Character Is

What it takes to build a custom keyboard for an ancient biblical script, from Unicode politics to font design.

unicodekeyboard-layoutsbidirectional-text

#775: When Your Cursor Has a Mind of Its Own

Stop fighting your cursor! Discover why mixing RTL and LTR languages breaks your layout and how to fix it using Unicode and CSS.

software-developmentusabilitylinguisticsbidirectional-textunicode