@AndrewCritchPhD There's a fair case to be made for all anthropomorphic self-references of contemporary chatbots to be deceptive, but we might need a new kind of grammar for the speech of products and services.
@AndrewCritchPhD 'Person' is a much more capacious term than 'human', as we already have artificial persons (i.e. corporations). At the policy level, the question is going to be about the relationship between corporate persons and robots pretending to be people but deployed by corporations.
@Miles_Brundage Seems like the original sin of the AGI definition is that it's based on a comparison with a human as a unit of cognition. What would you say are better categories/concepts to use? I'd welcome recommended reading on this.
This is a great paper. It points out:
1. Humans do not even approximately behave according to rational choice theory
2. There is no reason to think advanced AI will "inevitably" maximize some utility function
3a. Human preferences are derivative / constructed, so aligning AI by matching its behavior to our stated preferences is wrongheaded
3b. We can align AIs directly to some normative ideal of a "good assistant / programmer / driver / etc." instead
4. Aggregating preferences across people is fraught with philosophical and mathematical difficulties. We should not aim to align AI to the "collective will of humanity."
I had a great time at the 6th Annual Symposium on Applications of Contextual Integrity!
It's a great community that's always growing intellectually, with a sincere and friendly vibe.
Thanks to everyone who organized it.