in the (human-observable) limit, these language models will all be speaking chinglish:
- most highly-educated digital workers speak english or chinese
- peak token efficiency <> expressiveness
- compact translations of those four-word chinese idioms (成语) or continued language compaction with gen-z slang and beyond
- common grammatical structure amplified between the two languages when expressing concepts in corporate environments (this++ if/when there's also overlap with hindi)
or maybe more simply: we're purely in a digital data numbers game (against a perfect ai translation amplifier) from here on out