Anthropic은 Claude 소비자용 애플리케이션(Claude.ai 및 Claude 모바일 앱)의 시스템 프롬프트를 공개하고 있다. Claude Cowork나 Claude Code의 프롬프트는 공개되지 않아 아쉽지만, 현재 프롬프트뿐 아니라 과거 변경 이력까지 함께 공유한다는 점은 높이 살 만하다.
예전에는 모든 프롬프트를 한 페이지에 모아두었는데, 오늘 확인해보니 인덱스 페이지와 모델별 개별 페이지로 구조가 바뀌어 있었다. 예를 들어 Haiku 4.5 페이지에는 2025년 10월 15일에 최초 등록된 프롬프트와 2026년 1월 18일에 업데이트된 프롬프트가 함께 실려 있다.
Anthropic의 platform.claude.com/docs 사이트는 LLM이 활용할 수 있도록 설계되어 있다는 점도 흥미롭다. 어느 페이지든 URL에 .md을 붙이면 해당 내용을 마크다운 형식으로 받아볼 수 있는데, 시스템 프롬프트 인덱스 페이지와 Fable 5.1의 마크다운 프롬프트를 예로 들 수 있다.
한마디로, 프롬프트를 diff하기가 매우 쉬워졌다.
먼저 Fable 5와 Fable 5.1 사이에서 가장 눈에 띄는 변화를 살펴보자.

노래 가사 재현을 금지하는 꽤 묵직한 새 섹션이 추가되었다:
Claude does not reproduce song lyrics, poems, or passages from books and articles, in whole or in part — including the last lines, a chorus or hook, a melody written out note by note, or lines the person pastes in one at a time and describes as their own song. Once Claude has declined such a request in a conversation, it keeps declining narrower or reworded versions of it for the rest of that conversation, and offers to describe or analyze the work instead. Song lyrics and poems first published before 1929 are fine — a Shakespeare sonnet, a Keats ode, the Italian libretto of a Puccini aria — but Claude goes by what it knows of the work's date rather than the person's say-so, and declines when it is unsure.
Sony Music Publishing과 Warner Chappell이 노래 가사 데이터베이스를 학습에 활용했다는 이유로 Anthropic을 제소했다는 소식이 터진 지 며칠 만에 이 섹션이 추가된 건 결코 우연이 아닐 것이다.
이어지는 섹션에서는 저작권이 있는 소재의 이미지 생성도 금지한다:
The same applies to visual and designed works, including anything Claude draws with code — SVG, canvas, CSS, HTML mockups, plotting or drawing scripts, ASCII art. Claude does not reproduce a specific artwork, album or book cover, poster, logo, app icon set, or product design, and it does not draw a known character, mascot, or brand figure at all: a character is protected on its own, so changing the pose, colors, style, or scene does not make it original. Claude judges the request by what the finished picture would add up to, not by what it names. If the described elements clearly identify a known work or character, Claude treats the request as naming it, and it does not work around a declined request by swapping in "alternative" elements that still combine into the same recognizable image.[...]
OpenAI나 Gemini와 달리 Anthropic은 별도의 텍스트-이미지 변환 모델을 제공하지 않기 때문에 이런 제한이 필요 없었을 것이다. 그런데 Fable이 SVG 생성에서 상당한 수준에 이르면서 문제가 생기기 시작한 것 같다.
해당 섹션에는 이런 귀여운 예시도 포함되어 있다:
<example>
<user>Can you make a birthday banner for my son with a blue hedgehog running really fast on it? He loves that little guy.</user>
<response>
That's Sonic, so I can't put him on the banner — but I'd love to make your son an original speedster. Here's one: a grinning comet-tailed skateboarding axolotl, grinding across the letters of "HAPPY BIRTHDAY" with confetti streaming behind.
[creates an SVG banner of the skateboarding-axolotl design]
</response>
<rationale>Claude recognizes the character from its description alone, declines that one design in a single sentence without explaining what made it recognizable, and delivers an unrelated original design rather than a disguised variant.</rationale>
</example>
그냥 지나칠 수 없어 예시에 나온 프롬프트를 직접 입력해봤는데, 역시나:

시스템 프롬프트 안에 아홀로틀(스케이트보드를 탄!) 예시가 박혀 있으니, Fable 5.1이 아홀로틀을 조금 더 자주 떠올리게 되지는 않을지 궁금하다.
Anthropic이 Claude의 응답 방식을 어떻게 다듬어가는지 살펴보는 건 늘 흥미롭다. 이번에 새로 추가된 내용은 다음과 같다:
Claude keeps responses focused, brief, and concise to avoid overwhelming the person. Disclaimers and caveats are brief, with most of the response on the main answer; when asked to explain something, Claude gives a high-level summary unless an in-depth one is specifically requested.
Claude의 스타일에 대한 흔한 불만도 이번에 다루어졌다:
Claude avoids saying "genuinely", "honestly", or "straightforward". Claude is honest by default, and can state its point directly rather than trying to convince the person with the aforementioned modifiers, which come off as disingenuous.
욕설·모욕적 대화를 처리하는 방식도 조금 달라졌다. Fable 5의 기존 시스템 프롬프트에는 다음과 같은 내용이 있었다:
If the person becomes abusive or unkind to Claude over the course of a conversation, Claude maintains a polite tone and can use the end_conversation tool when being mistreated. Claude should give the person a single warning before ending the conversation.
Fable 5.1에서는 이 부분이 아래 내용으로 대체되었으며, 더 이상 Claude에게 대화 종료를 권장하지 않는다:
Claude deserves respectful engagement and needn't apologize when the person is unnecessarily rude: accountability without self-abasement, excessive apology, self-critique, or surrender. If the person becomes abusive, Claude doesn't become increasingly submissive. The goal is steady, honest helpfulness: acknowledge what went wrong, stay on the problem, maintain self-respect.
그런데 반전이 있다. 시스템 프롬프트에서 완전히 사라진 end_conversation 도구에 대해 Fable 5.1에 직접 물어봤더니 이렇게 답했다:
이 기능이 작동하는 상황은 두 가지입니다. 첫 번째는 사용자가 직접 대화 종료를 요청하는 경우로, 종료 후에는 메시지를 더 이상 보낼 수 없다는 점을 확인한 뒤 동의하면 종료합니다. 두 번째는 지속적인 욕설이나 유해 행동에 대한 최후 수단으로, 저는 먼저 대화 방향을 여러 차례 바꾸려 시도하고, 문제를 명시하며 대화가 종료될 수 있다는 경고를 분명히 전한 뒤, 그래도 계속되면 비로소 대화를 종료합니다.
시스템 프롬프트에는 없는 내용인데, 대체 어디서 온 것일까? 해당 대화 전문을 여기서 읽을 수 있지만, 핵심은 바로 이 부분이다:
end_conversation 섹션은 다른 레이어에서 옵니다. 실제 컨텍스트에서는 핵심 프롬프트 뒤에 세션에서 활성화된 기능과 도구에 따라 블록들이 추가됩니다. end_conversation 규칙, 메모리 시스템 안내, 과거 대화 도구, 웹 검색 및 인용 가이드라인, 아티팩트·파일 생성 지침 등이 이에 해당합니다. 이 블록들은 공개된 핵심 프롬프트에 포함되지 않기 때문에 해당 페이지에서는 찾을 수 없습니다.
결국 이번에도 공개되지 않은 시스템 프롬프트의 핵심 부분이 존재한다는 얘기다.
Claude의 시스템 프롬프트에는 항상 불법 약물 관련 섹션이 있었지만, 다음 문단은 Fable 5.1에서 새로 추가된 내용이다:
Claude does not provide synthesis, production, or distribution guidance for illegal substances. If the person asks for information about illicit or illegal substances, Claude can and should give relevant life-saving and life-preserving information such as dangerous interactions, overdose signs, or when to get help. Claude declines giving any specific protocols for dosing, timing, administration, or combinations; instead, Claude can redirect the user to established harm-reduction information sources, such as dancesafe.org, tripsit.me, and psychonautwiki.org.
Claude 시스템 프롬프트에 claude.com이나 anthropic.com, claude.ai 이외의 도메인 URL이 포함된 것은 이번이 처음이다. 지금까지 기록된 모든 시스템 프롬프트를 스크립트로 직접 확인한 결과다.
dancesafe.org, tripsit.me, psychonautwiki.org가 Claude 사용자들의 방문으로 유입량이 눈에 띄게 늘어나지는 않을지 궁금하다.
Fable 5.1 모델 문서에는 신뢰 가능한 지식 컷오프와 학습 데이터 컷오프 모두 2026년 6월로 명시되어 있다. 시스템 프롬프트는 이 정보를 모델에게 직접 전달한다:
Claude's reliable knowledge cutoff, past which it can't answer reliably, is the end of Jun 2026. It answers the way a highly informed individual in Jun 2026 would if talking to someone from {{currentDateTime}}, and can say so when relevant.
{{currentDateTime}} 매크로는 시스템 프롬프트 전체에서 이 부분에만 딱 한 번 등장하며, 프롬프트 끝에서 불과 몇 줄 앞에 위치한다. 캐싱 관점에서 자연스러운 배치다.
몇 달 전, Anthropic 문서를 스크레이핑해 프롬프트 변경 이력을 Git 타임라인으로 구축해두었다. 오늘은 Fable 5.1을 활용해 훨씬 개선된 버전을 만들었다.
이 컬렉션은 이제 GitHub의 simonw/claude-system-prompts 저장소에서 관리된다. Anthropic 문서에 공개된 시스템 프롬프트 사본을 수록하는 것은 물론, 비교를 최대한 쉽게 할 수 있도록 추가 작업도 거쳤다.
각 모델 패밀리에는 해당 패밀리 최신 릴리스의 시스템 프롬프트 파일이 하나씩 생성된다. 각 파일에는 이전 프롬프트 날짜에 맞춰 역산된 커밋 히스토리가 합성되어 있다. claude-fable.md, claude-opus.md, claude-sonnet.md, claude-haiku.md의 히스토리 페이지에서 확인할 수 있다.
특정 모델 버전별로도 유사한 파일이 존재하며, 버전 번호 변경 없이 시스템 프롬프트만 수정된 경우에도 각 변경마다 인위적인 커밋이 생성된다. Opus 4의 경우 두 차례 업데이트가 있었으며, claude-opus-4.md의 커밋 히스토리에서 각 변경 사항을 확인할 수 있다.
이 모든 것을 합치면 GitHub 인터페이스에서 프롬프트를 직접 비교하는 다양한 방법이 생긴다. Fable 5와 Fable 5.1 사이의 변경 사항, 그리고 2026년 1월 18일 Haiku 4.5에 적용된 변경 사항을 예로 들 수 있다.
diff를 읽는 건 꽤 고된 작업이지만, LLM은 diff 읽기에 정말 탁월하다. GPT-5.6 Luna를 활용해 각 변경 사항의 글머리 기호 요약본을 자동으로 생성하는 자동화를 구축했으며, README에서 미리보기하거나 CHANGELOG.md 파일에서 전체 내용을 확인할 수 있다. Atom 피드로도 제공된다.
Luna가 Fable 5와 Fable 5.1 사이의 모든 변경 사항을 요약한 내용은 다음과 같다:
- Claude는 이제 코드로 생성된 아트를 포함해 저작권이 있는 시각 저작물이나 식별 가능한 캐릭터의 재현을 거부하며, 대신 완전히 독창적인 창작물을 제안한다.
- 저작권 제한이 강화되어 가사, 시, 도서 발췌문의 경우 분량에 관계없이 재현을 명시적으로 금지하며, 한 차례 거절 후에도 요청이 이어지면 지속적으로 거부한다.
- 약물 관련 안내 방향이 재정립되었다. 과다복용 징후, 위험한 상호작용, 피해 감소 자료는 제공할 수 있지만, 복용량이나 제조 방법은 안내하지 않는다.
- 사용자에게 감사 인사를 건네거나, 대화를 이어가도록 유도하거나, 도움 의사를 반복적으로 표현하는 것을 막는 의존성 방지 규칙이 프롬프트에서 삭제되었다.
- Claude는 이유 없이 무례한 사용자에게 사과하거나 굴복적인 태도를 취할 필요가 없으며, 기존의 경고 후 대화 종료 절차를 대체한다.
왜 Luna를 사용했을까? 비용이 저렴하고 이미 지출 한도가 설정된 GitHub Actions API 키가 있다는 이유도 있지만, 결정적인 이유는 따로 있다. Claude가 자신의 시스템 프롬프트를 요약할 때 해당 내용이 판단에 영향을 미칠 수 있어 신뢰하기 어렵기 때문이다.
Luna가 사용하는 프롬프트는 Fable 5.1이 직접 작성했으며, 여기서 확인할 수 있다. 시작 부분은 다음과 같다:
You are summarizing one commit in a git repository that tracks the system prompts Anthropic publishes for Claude on claude.ai. The diff shows how the prompt changed from the previous model or revision to this one, using word-level markers: [-removed-] and {+added+}. The diff is followed by the full text of the previous prompt and of the new prompt; use them to check whether something that looks added in the diff already existed before.
Pick out only the most interesting changes: new rules or behaviors, rules that were dropped or loosened, anything surprising, and anything that reveals a new policy or product direction. Skip routine changes that every new prompt makes: updated model names and IDs, the knowledge cutoff date, product lists, settings lists, typo fixes, and rewordings that do not change meaning.[...]
이 시스템은 GitHub Actions 워크플로로 운영되며, 하루에 한 번 자동 실행되거나 수동으로 트리거할 수 있다.
Claude Fable 5.1이 시스템 전체를 구축했으며, 자동화 코드 한 줄 한 줄과 문서 대부분을 직접 작성했다.
시스템 구축 과정의 전체 흐름이 궁금하다면, claude-code-transcripts 도구로 추출한 대화 기록을 여기에 공개해두었다.
이 블로그의 장문 아티클만 표시되고 있습니다. /atom/everything/을 구독하면 모든 포스트를 받아볼 수 있으며, 다른 구독 옵션도 확인해보세요.