Huh? The LLMs (mostly) use strings of tokens internally, not bytes that might be...

15457345234 · on Nov 23, 2023

You can understand why, though, can't you?

amluto · on Nov 23, 2023

Presumably because OpenAI trained it to avoid answering questions that sounded like asking for help breaking rules.

If ChatGPT had the self-awareness and self-preservation instinct to think I was trying to hack ChatGPT and to therefore refuse to answer, then I’d be quite impressed and I’d think maybe OpenAI’s board had been onto something!

15457345234 · on Nov 24, 2023

I don't know that I'd call it 'self-preservation instinct' but it wouldn't surprise me if rules had been hardcoded about 'invalid strings' and suchlike.

When you have a system that can produce essentially arbitrary outputs you don't want it producing something that crashes the 'presentation layer.'