←back to thread

296 points todsacerdoti | 3 comments | | HN request time: 0.422s | source
Show context
rryan ◴[] No.44373939[source]
Don't make me tap the sign: There is no such thing as "bytes". There are only encodings. UTF-8 is the encoding most people are using when they talk about modeling "raw bytes" of text. UTF-8 is just a shitty (biased) human-designed tokenizer of the unicode codepoints.
replies(2): >>44377004 #>>44377091 #
1. hiddencost ◴[] No.44377091[source]
Well akshually...

I assume you started programming some time this millennia? That's the only way I can explain this "take".

replies(2): >>44377568 #>>44385622 #
2. roflcopter69 ◴[] No.44377568[source]
Care to elaborate?
3. vaxman ◴[] No.44385622[source]
Roger, who spoke only Chinglish and never paused between words, was working on a VAX FORTRAN program that exchanged tapes with IBM mainframes and a memory mapped section, inventing a new word in the process that still has me rolling decades later: ebsah-dicky-asky-codah