#bidirectional-text
7 episodes
#5488: How Emoji Went From 176 Characters to 3,600
From a 176-character set built for Japanese pagers to a 3,600-character Unicode standard — and why your emoji breaks in production.
#5487: Why Your Spreadsheet Mangles José's Name
A deep dive into character sets, from ASCII to UTF-8, and why José becomes "José" in your spreadsheet.
#5486: Why Hebrew URLs Turn Into Percent-Sign Gibberish
Hebrew renders fine in the domain but explodes into percent-hex in the path. Two different standards explain the split.
#5466: Hebrew Words Hidden in English Text
Daniel wants a classifier that spots Hebrew written in Latin letters — and it turns out nobody's built one.
#5461: TTS Can't Pronounce Hebrew Inside English
Your TTS reads Hebrew words with English phonetics. Here's why — and why the obvious fix doesn't work yet.
#2593: The Politics of Unicode: Paleo-Hebrew, Han Unification, and Who Decides What a Character Is
What it takes to build a custom keyboard for an ancient biblical script, from Unicode politics to font design.
#775: When Your Cursor Has a Mind of Its Own
Stop fighting your cursor! Discover why mixing RTL and LTR languages breaks your layout and how to fix it using Unicode and CSS.