←back to thread

296 points todsacerdoti | 1 comments | | HN request time: 0.206s | source
Show context
rryan ◴[] No.44373939[source]
Don't make me tap the sign: There is no such thing as "bytes". There are only encodings. UTF-8 is the encoding most people are using when they talk about modeling "raw bytes" of text. UTF-8 is just a shitty (biased) human-designed tokenizer of the unicode codepoints.
replies(2): >>44377004 #>>44377091 #
hiddencost ◴[] No.44377091[source]
Well akshually...

I assume you started programming some time this millennia? That's the only way I can explain this "take".

replies(2): >>44377568 #>>44385622 #
1. vaxman ◴[] No.44385622[source]
Roger, who spoke only Chinglish and never paused between words, was working on a VAX FORTRAN program that exchanged tapes with IBM mainframes and a memory mapped section, inventing a new word in the process that still has me rolling decades later: ebsah-dicky-asky-codah