Claude's System Prompts Go Public
Anthropic published the core system prompts shaping Claude’s web and mobile behavior, including instructions for tone, formatting, vision, and safety. The release offers developers an unusual look at how model behavior is steered after training.
Publishing the prompts makes Claude more transparent, but also shows how fragile some behavioral controls can be when expressed as plain-language instructions.
- –Developers can study Anthropic’s approach to response style, refusals, image handling, and conversational behavior
- –The prompts apply to Claude’s consumer apps, not the Claude API, so API users still control their own system instructions
- –Hacker News commenters noted that Claude frequently ignores or contradicts several published directives
- –The release highlights the difference between learned behavior, system-level steering, and user prompts
- –Public prompts improve accountability while giving competitors and jailbreak researchers more material to analyze
DISCOVERED
46d ago
2026-08-16
PUBLISHED
46d ago
2026-08-16
RELEVANCE
AUTHOR
tosh