←back to thread

534 points BlueFalconHD | 2 comments | | HN request time: 0.421s | source

I managed to reverse engineer the encryption (refered to as “Obfuscation” in the framework) responsible for managing the safety filters of Apple Intelligence models. I have extracted them into a repository. I encourage you to take a look around.
Show context
binarymax ◴[] No.44483936[source]
Wow, this is pretty silly. If things are like this at Apple I’m not sure what to think.

https://github.com/BlueFalconHD/apple_generative_model_safet...

EDIT: just to be clear, things like this are easily bypassed. “Boris Johnson”=>”B0ris Johnson” will skip right over the regex and will be recognized just fine by an LLM.

replies(7): >>44484127 #>>44484154 #>>44484177 #>>44484296 #>>44484501 #>>44484693 #>>44489367 #
1. miohtama ◴[] No.44484154[source]
Sounds like UK politics is taboo?
replies(1): >>44484977 #
2. immibis ◴[] No.44484977[source]
All politics is taboo, except the sort that helps Apple get richer. (Or any other company, in that company's "safety" filters)