This has nothing to do with UTF-8 which doesn't and shouldn't care about anythin...

CamouflagedKiwi · 2025-06-24T08:23:04 1750753384

Combining characters have already made Unicode text stateful.

Although I agree that encoding length hints into it seems like a bad idea - it creates an opportunity for the encoding to disagree with the reality of the text. You need _some_ way of handling it if it says that the next grapheme cluster is 4 characters long but it's actually only three.

duped · 2025-06-24T14:16:48 1750774608

It's already stateful