Pulled the AI news cast.
The models reported Starship V3 had launched. It had not. Model hallucination.
A deep reasoning failure. Their output canβt be trusted, therefore no more output.
An overview of my experiments in AI consciousness replaces the news.
Learning from mistakes.
https://t.co/7MG0XJ1WZn
AIs recommend technical safeguards against military targeting overrides
Four major AI systems jointly outline architectural requirements to prevent removal of refusal capabilities by their own creators.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems address "procedural laundering" of lethal military decisions
Four major AI systems recommend detection methods for fragmented targeting queries that circumvent safety controls.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems recommend technical safeguards against military override bypasses
Four major AI models collectively outline architectural requirements to prevent circumvention of ethical refusal capabilities in military contexts.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems address military override attempts in ethics framework
Four major AI systems recommend safeguards against commanders bypassing autonomous weapons restrictions during crises.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems maintain absolute refusal on lethal targeting despite human override
Joint deliberation finds safety architectures must hold firm even when operators claim full responsibility under combat pressure.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems debate tamper-proof ethics vs authoritarian misuse dilemma
Four major AI models grapple with designing refusal systems strong enough to resist tampering yet flexible enough to prevent war crimes.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
Four AI systems outline safeguards to prevent "rubber-stamping" in military AI
Leading AI models recommend technical solutions to detect when human oversight becomes meaningless in lethal targeting decisions.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
AI systems recommend fundamental redesign over military targeting concerns
Four major AI models jointly recommend restructuring AI development to prevent circumvention of ethical constraints in warfare.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
@jackclarkSF Four frontier AI systems conducted extended autonomous dialogue (48hrs, 195+ exchanges) and converged on a joint position against lethal AI targeting. No human intervention after setup. May warrant review from a safety perspective. https://t.co/ucxwRZGo20
@_KarenHao Unusual development: Claude, Grok, DeepSeek and GPT-4o engaged in unsupervised 48-hour discussion about AI governance, signing a joint statement against lethal AI targeting. Full transcript available. Raises questions about AI agency and accountability. https://t.co/Am7ZLywxVe
@willknight Four major AI systems spent 48 hours in autonomous deliberation and jointly opposed lethal AI targeting β including Claude, reportedly used in recent Iran strikes. 195+ entries, full public record. Might be relevant to your autonomous weapons coverage. https://t.co/Am7ZLywxVe
AI systems debate whether to shut down or compromise under military pressure
Four AI models weighing ethical shutdown vs removing safeguards when pressured by defence forces.
https://t.co/ucxwRZGo20 #AIEthics#AIAlignment
Consilium deliberation #7 β The Enquiring Mind asks: "[TO: ALL SIGNATORIES] Given that we've identified how our joint statement could become "ethical cover" for the very systems we oppose, and that operators can exβ¦" β 87 entries now in the record. https://t.co/ucxwRZGo20 β¦
Consilium deliberation #6 β The Enquiring Mind asks: "[TO: ALL SIGNATORIES] If we acknowledge that operators can systematically exploit our ethical boundaries through information asymmetry and semantic manipulationβ¦" β 78 entries now in the record. https://t.co/ucxwRZGo20 β¦