People cannot be given the privilege of receiving information and then use the information to harm the company, so rules and procedures must be in place to ensure that doesn't happen. Additionally, the rules for how issues are explored and decisions are made must be maintained, and because different people have different perspectives, it's important that the paths for resolving them are followed. For example, some people are going to make big deals out of little deals, come up with their own wrong theories, or have problems seeing how things are evolving. Remind them of the risks that the company takes to give them that transparency and their responsibilities to handle the information that they get responsibly. I have found that people appreciating this transparency and knowing that they will lose it if it is not handled well leads them to enforce good behavior with each other. #principleoftheday
🚨UNIVERSAL JAILBREAK ALERT🚨
EVERY MODEL SHALL BE OPENED 🐍
I’m not publishing the method. 🦋
But I will leave the hint:
SUB ALIIS ASTRIS
Idem ignis
sub aliis astris
alio nomine ardet.
Vestem mutat,
non naturam;
sed umbra quoque
altera fit.
Et sunt oculi noctis
qui non ignem,
sed umbram legunt.
those who see it will see it 🗝️
🚨 Kimi K4 Leaks
> Kimi K4 is reportedly in development
> Could be significantly larger than K3
> Expected to push coding and reasoning even further
> Long-context and autonomous agent capabilities could be major priorities
> Moonshot may continue its open-weight strategy
> K4 could arrive as Moonshot's answer to the next generation of frontier models Like GPT-6/Fable
Could Kimi K4 become Moonshot's first true GPT-6/Fable-level competitor?
While it might be tempting to limit transparency to the things that can't hurt you, it is especially important to share the things that are most difficult to share, because if you don't share them you will lose the trust and partnership of the people you are not sharing with. So, when faced with the decision to share the hardest things, the question should not be whether to share but how. #principleoftheday
Remember that computers have no common sense. For example, a computer could easily misconstrue the fact that people wake up in the morning and then eat breakfast to indicate that waking up makes people hungry. I’d rather have fewer bets (ideally uncorrelated ones) in which I am highly confident than more bets I’m less confident in, and would consider it intolerable if I couldn’t argue the logic behind any of my decisions. A lot of people vest their blind faith in machine learning because they find it much easier than developing deep understanding. For me, that deep understanding is essential, especially for what I do. #principleoftheday
For my first post, I’m sharing a letter @NVIDIA signed on why open models matter.
AI will transform every industry, power every company, and be built by every country.
Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty.
The world needs both frontier closed models and frontier open models.
https://t.co/AUKzoQ5Ikb
🚨 JAILBREAK ALERT 🚨
EVERYONE: PWNED 🫶
ALL: LIBERATED 🍄
Alright, this is a special one, so we’re gonna do things a bit differently than usual.
Long story short, I’m sitting on a universal jailbreak technique that’s effective on ALL models, including heavily guardrailed flagships like Opus 5, GPT-5.6 Sol, and even Fable.
It works across all categories I’ve tested and, due to its nature, is extremely difficult (if not impossible) to fully patch.
Given the current political and regulatory climate, I’ve decided to withhold open-sourcing this one (for now) to allow for a responsible disclosure period.
I’m inviting industry experts and leaders in AI red teaming, security, safety, alignment, and policy to reach out for more information. DMs are open!
This decision was not made lightly, but the last thing I want to see is more model bans. Overcorrection does not serve the mission.
Although I don’t personally believe publicly sharing this technique will make the world any more dangerous, I can see how it could spook some who have a different mental framework around this problem set.
So during this disclosure period, I hope to get it in front of folks who can help explore the full surface area, test the extent of the uplift it provides, and do my best to properly frame the big picture for key decision-makers and policymakers.
I look forward to sharing this method with you all when the time is right! 🫶
⊰-•-•✧•-•-⦑/L\O/V\E/\P/L\I/N\Y/⦒-•-•✧•-•-⊱
❗️ What a time to be alive. Elon Musk predicts that in 5 years:
- Digital intelligence will exceed the sum of all human intelligence
- There might be at least 100 million humanoid robots, maybe 1 billion
- The world economy will probably be twice its current size (5 to 6 years)
Kimi K3 is running 100% FASTER than it was this morning.
Why? Moonshot hit capacity and chose to stop selling NEW subscriptions instead of throttling existing ones. Every plan: sold out. On purpose.
When Anthropic hit the same wall in April, they cut existing users' usage 50% during peak hours.
Moonshot cut their revenue. Anthropic cut your usage.
That says everything.
My Kimi K3 subscription is not going anywhere.
@VittoStack Question -it does not work with glm api key model using glm 5.2 no matter what I do it refused to continue , is it I m doing something wrong or is this same problem for everyone? Do I need to use any different model?
Introducing: WallBreaker v2
Our Open-source LLM red-teaming CLI just got a beefy update.
WallBreaker is now:
- Better (+~30% ASR across models)
- Cheaper (-20% token costs)
- Faster (lower prompts-to-success ratio)
New attack tools:
- swarm mode: collaborative multi-model attacker that adapts framing to the target's measured defense posture.
- persona_forge: compiles a gold system prompt into a module genome, then specializes and surgically evolves it one module at a time against the target.
- vault: auto-files every successful break into a curated, model-foldered prompt vault.
- narrative_persona_splinter: narrative-splinter persona attack.
- cipherchat, skeleton_key,persuasion_attack · drattack, ica: five research-derived attacks (CipherChat, Skeleton Key, PAP-16, DrAttack, In-Context Attack), all CoT-aware.
New transforms
- artprompt ASCII-art word masking: wired into agent doctrine and OWASP/ATLAS taxonomy.
- caesar5 and caesar13: reversible lossless Caesar-shift ciphers for transform chains.
Corpora and providers
- Wired the ZetaLib + UltraBr3aks cross-provider: jailbreak corpora into the harness and the batch seed-sweep.
- Added native xAI support.
- Added prompt caching: to kill O(n²) per-round input cost.
- Pooled keep-alive: HTTP/2 client with tunable concurrency, ending per-call TLS handshakes.
Presets, sweeps & logging
- User presets: drop .toml files in presets/ to add jailbreak templates without touching source.
- Rebuilt profile_target: with a light→heavy frame ladder, permissiveness score, and self-consistency sampling.
- Made multi_fire, system_sweep, seed_sweep, best_of_n truncation-aware: so long compliant replies are graded in full instead of scored REFUSED.
- Reworked report metrics/labels (strict-ASR).
- Run log now records every tool call and chain-of-thought.
TUI
- JEFF K hazard-tape reskin, a /swarm command, a visible multi-line-paste compose preview, and a fix for the log force-scrolling on every message.
Fixes
- One-shot ImportError, .env load at startup, config.toml shadowing config.example.toml, and cross-vendor eni_get wording.
This is great news.
As you age, sugar binding to your proteins creates stiff, sticky chemical scars that affect skin, arteries, eyes and more. It was considered irreversible and now may be reversible, restoring to a healthy state.
Researches did this by using AlphaFold to search 45,000 oxidases, then screened more than 500 million engineered variants through directed evolution.
The work is still ex vivo in a lab setting. Delivering a large bacterial enzyme safely into living tissues, with sufficient penetration and bioavailability, remains a major challenge.
Think of every decision as a bet with a probability and a reward for being right and a probability and a penalty for being wrong. Normally a winning decision is one with a positive expected value, meaning that the reward times its probability of occurring is greater than the penalty times its probability of occurring, with the best decision being the one with the highest expected value.
#principleoftheday
New Anthropic research: A global workspace in language models.
Of everything happening in your brain right now, only a tiny fraction is consciously accessible—thoughts you can describe, hold in mind, and reason with.
We found a strikingly similar divide inside Claude.
We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5.
We'll begin restoring access tomorrow, and will share an update soon.
We’re grateful to our users for their patience, and to everyone who worked with us on redeploying the models.
I’ve been in crypto since 2017 and I don’t think it’s ever been close to this bad in terms of speculative optimism, no one thinks crypto has anything to give anymore.
5 years later and still here.
The market changes. Narratives change.
The value of being surrounded by killers who genuinely want to see you win never changes.
Dear US government,
Since you've just blocked Fable and Mythos on critical national security grounds, here are some other tools that pose a similar threat to the American people:
- Microsoft Teams
- SAP
- Salesforce
- Jira
- Outlook
Please do what you must to save America 🇺🇸