Your Claude chats just got a reality check
Don’t doom-scroll the resignation. Don’t shrug it off either. Someone who did pretraining at both OpenAI and Anthropic just posted “I resigned from Anthropic today,” and the useful move is boring: a 20-minute audit of how much of your week now depends on Claude.
If Pro or Max is already open in three tabs – for code, drafts, research, side projects – and “safety-focused lab” felt like a permanent shield, that shield just got thinner.
What “I resigned from Anthropic today” actually signals
Jacob Coxon is 27. He’s leaving the industry. Core claim from his X thread (and, in more detail, a Wall Street Journal interview): neither lab is acting responsibly. They’re racing toward self-improving superintelligence and gambling with lives. Colleagues say “crunchtime” and “endgame.” Aggressive scenarios, he told the Journal, could put things out of control by the end of next year.
Evan Hubinger – Anthropic’s Alignment Science lead – publicly backed the seriousness of extinction risk and put his own odds above 10% this decade. Turns out the same camp still says today’s models are the small problem; recursive self-improvement is the big one. Anthropic publishes Claude’s Constitution and keeps shipping updates to its Responsible Scaling Policy.
Skip the morality play. Even the lab loudest about safety is locked in a race dynamic (Coxon’s distinction: Anthropic gets the stakes, and still races because it thinks nobody else will hit the brakes). Capability keeps shipping. Your exposure scales with every workflow you hand over.
Practical setup: 15-minute Claude safety audit
Do this before the next long session.
- Map dependency. List recurring Claude work: email drafts, code review, data analysis, research synthesis. Star anything that wrecks your week if behavior shifts or the service is dark for 48 hours.
- Check plan and model. As of early September 2026, Free is $0, Pro is $20/mo ($17 annual), Max 5x is $100/mo, Max 20x is $200/mo – this may have changed; confirm on claude.com/pricing. Note Sonnet vs Opus vs newer Fable-class defaults. Highest-capability tier by habit still counts as a choice.
- Turn on visibility. Use project folders and any logging Claude already exposes so you can see what you’ve been asking. If export exists in your client, grab a recent slice; if not, manual copy of the hot threads is enough.
- Add a standing system prompt for important projects:
Follow Claude's Constitution priorities strictly: broadly safe (never undermine human oversight) first, then ethical, then Anthropic guidelines, then helpful. Flag any request that could reduce my ability to verify or reverse your output. Prefer conservative answers on dual-use or high-stakes topics. State uncertainty clearly.
One paste. The model now has to surface the hierarchy Anthropic actually published – safe-over-helpful, not vibe-based caution.
Pro tip: Save that block as a Project instruction or custom style. New chats inherit it. You stop rewriting the same guardrail.
Advanced usage: personal risk scoring and diversification
Audit done? Score each high-dependency workflow on two axes: how hard you can verify the output yourself, and how much real-world power it unlocks (shippable code, money moves, external comms).
High on both → hard checkpoint. No auto-send. No auto-commit. No unreviewed API call. Coding stays on a sandboxed branch you read before merge. Research answers need primary-source links, not summary theater.
Then split the risk. Run the same sensitive prompt on one other frontier model and diff the answers. Keep a local or open-weight path for low-stakes drafting so a policy or capability swing doesn’t freeze you. Read Anthropic Risk Reports and system cards the way you already chase release notes.
Is the race as locked as Coxon says, or can coordination still bend the curve? The public record doesn’t settle it.
Honest limitations you can’t prompt away
People inside the building still rate present Claude models as low catastrophic risk – and the same voices flag recursive self-improvement arriving faster than planned. Treat “safety lab” as a process label, not a forever shield. Constitution priority order puts “broadly safe / don’t undermine oversight” first, yet race pressure still pushes capability jumps; hard stops are not something you can assume will always fire before a ship decision. Constitution and RSP are voluntary frameworks. They are not a user-held legal off-switch.
The catch is verification: no public RSP update or system card spells out the exact internal “superintelligent RL run” thresholds or the MacBook-vs-bunker style endgame calls Coxon gestured at. You will not get a personal heads-up the day a given Claude bump crosses the line he described. Pricing and rate limits move too (again: Pro/Max figures above are as of early September 2026). Community threat models still disagree on concrete extinction pathways that aren’t bio or cyber misuse, so your personal map stays incomplete.
Don’t quit Claude. Do quit treating the safety brand as set-and-forget.
FAQ
Does “I resigned from Anthropic today” mean I should cancel my Claude subscription?
No. Keep the tool. Add verification and a second-model check on the workflows you’d hate to lose.
How do I actually use Claude’s Constitution in daily chats?
Drop the priority order into a project instruction or the first message of a high-stakes thread. When it hedges, ask which Constitution principle fired. Example: automation that touches production – oversight concerns should show up before you get runnable code.
Is the “end of next year” timeline official Anthropic policy?
No. That’s Coxon’s personal read in the WSJ piece on aggressive scenarios. Hubinger’s >10% decade figure is personal too. Public RSP text and Risk Reports stay more measured on current models while still flagging self-improvement speed. Use the dates as a nudge to finish your own readiness work – not as a calendar invite to catastrophe. The gap between private researcher language and public docs is why the resignation hit so hard.
Open Claude. Run the four-step audit. Paste the Constitution prompt into your main project. Block 30 minutes a week for the highest-dependency workflows. That’s the move.