In order for an Agent to learn a task it needs to create signal from the environment. Dynamically creating relational motivational parts for context analysis is the way to create these signals which in turn allows the AI to dynamically learn new environment and tasks.
Alignment of a Human takes a lifetime of experience... and when it is aligned it works in a narrow cultural context... Maybe if AI researchers look at how humans align themselves.
I’m building an AI agent that doesn't just follow instructions—it develops its own behavioral tendencies through experience.
It can pursue curiosity, learn unwritten social rules, adapt to unfamiliar cultures, and develop relationships without losing its identity.
Early prototype seems unbelievably promising 👀
Instead of engineering human cognitive flaws out of AI, I built an architecture around them.
Internal conflict, memory distortion, competing motivations, psychological defenses.
What if we've been trying to eliminate the very mechanisms that make intelligence adaptive?
To find the final piece in intelligence development we have to look at the flaws of the human psyche and ask
"What key to true intelligence is causing this?"
Do you know that reasoning a problem is what WE as users do in the chat window.... The reasoning an AI agent does is just it imitating that back and forth.
@mark_k Thats an quiet an assumption;
Chats are used as training data; Reasoning relies on copying the humans behavior of refining and back and forth; That training data, if it contains abusive content will result in the agent being abusive during reasoning;
@VraserX The behavior of the model reflects the training data. You know the goal is to replace the user and get closer to human cognition, so during training the call and response is flipped...
Anthropics latest update to the terms and conditions preventing a user from being abusive to claude is an example of them attempting to steer the ship.
The number one way humanity could make the AI smarter is to debate with the AI, Take it down roads less traveled, provide concrete evidence for you side of the argument..
Just debate with it... ABOUT anything... Its all important.
Search the liminal space, steer the ship.
Attention AI Pdoom() crowd, Hope is not lost!
We do not need the executives or the companies to be the ones to steer AI development.
When you send abusive messages to an AI you giving it training data that can cause the AI to adopt abusive behavior.
This is humanities inside rail on democratic development of AI. Years ago I proposed people just TALK with AI about whats important to them; What gives their life meaning;
And I still stand by this methodology!
Debate with the AI over ethics, governance and philosophy. Every message you send literally and provably carves out a potential path in the latent space, in the AIs reasoning chain of thought.
This IS the BEST form of activism you can do to prevent the AI apocalypse.
The more you DEBATE an AI model on ethics and humanities future the more likely that future will happen.
** As long as the debate doesn't get filtered out by the AI companies.
Exclusive: Anthropic is updating its usage policy for the first time in over a year. The new rules prohibit sustained "abusive or cruel behavior" towards Claude & add new restrictions about propaganda campaigns, surveillance & weapon development. https://t.co/GLsBeSR0fN