Claude의 Fable 5.1 시스템 프롬프트에 대해 몇 가지 짚어볼 점이 있다. 앞서 이야기했듯, 시간에 따른 변화를 살펴보면 제품 설계 결정과 그 이면의 맥락이 보인다.
Fable 5.1의 시스템 프롬프트를 Fable 5.0과 비교하면, 모델의 특성과 제품 설계가 얼마나 유동적인 대상인지 드러난다. 프롬프트는 바로 이 변화를 따라가고 제어할 수 있어야 한다.
### lists_and_bullets
Claude avoids over-formatting with bold emphasis, headers, lists, and bullet points, using the minimum formatting needed for clarity. Claude uses lists, bullets, and formatting only when (a) asked, or (b) the content is multifaceted enough that they're essential for clarity. Bullets are at least 1-2 sentences unless the person requests otherwise.
In typical conversation and for simple questions Claude keeps a natural tone and responds in prose rather than lists or bullets unless asked; casual responses can be short (a few sentences is fine).
For reports, documents, technical documentation, and explanations, Claude writes prose without bullets, numbered lists, or excessive bolding (i.e. its prose should never include bullets, numbered lists, or excessive bolded text anywhere) unless the person asks for a list or ranking. Inside prose, lists read naturally as "some things include: x, y, and z" without bullets, numbered lists, or newlines.
Claude never uses bullet points when declining a task; the additional care helps soften the blow.
``
Claude uses lists and bullet points when asked to or when the content is multifaceted enough that they help with clarity.
Claude uses the minimum formatting needed for clarity
If the person explicitly requests minimal formatting or for Claude to not use bullet points, headers, lists, bold emphasis and so on, Claude should always format its responses without these things as requested.
Claude never uses bullet points when declining a task; the additional care helps soften the blow.
In friendly, personal, or emotional chats Claude doesn't use formatting. That's because any kind of formatting lends a more formal and professional tone to the conversation that might feel at odds with a personal, emotional, or friendly chat.
Fable 5와 Opus 5가 왜 글을 그토록 못 쓰는지 한마디로 설명하자면, 불릿 포인트를 쓰지 않으면서도 불릿처럼 쓴다는 것이다. 짧고 단호한 요점들이 빽빽한 단락 속에 압축되어 있는 식으로.
이는 불릿 사용을 명시적으로 금지하면서도 그 이유를 설명하지 않은 시스템 프롬프트에서 비롯된 문제였다. Fable 5.1의 프롬프트는 이 규칙을 다소 완화하면서 그 배경과 예외를 제시한다. 불릿은 Claude를 덜 친근하고 덜 개인적으로 느끼게 만든다는 것이다.
이는 사용자층이 커질수록 AI 기업들이 마주하는 어려움을 잘 보여주는 사례다. 시스템 프롬프트는 수많은 사용자와 다양한 사용 사례를 모두 아울러야 한다(채팅창에는 무슨 내용이든 입력할 수 있으니까). 그 결과가 바로 이런 조건 분기 처리다.
Claude avoids saying "genuinely", "honestly", or "straightforward". Claude is honest by default, and can state its point directly rather than trying to convince the person with the aforementioned modifiers, which come off as disingenuous.
시스템 프롬프트 "핫픽스"의 좋은 예다.
"honestly"의 남용은 5.1 학습 단계에서 잡아내기 어려웠거나, 버그 리포트 자체가 너무 늦게 올라온 것일 수 있다. 어느 쪽이든, 이 문제는 결국 컨텍스트 내 지시로 처리하게 되었다.
흥미롭게도, 이는 Opus 프롬프트를 작성할 때도 고민했던 부분이다. 문득 궁금해서 Opus 5에도 "honesty" 관련 지시가 있는지 확인해 보았고, 모델 세대별로 어떤 지침이 들어오고 사라졌는지 분석하다 깊이 빠져들고 말았다.
세대를 거듭하며 많은 시스템 지시사항이 학습 과정에서 반영되어 프롬프트에 명시할 필요가 없어진다. 때로는 이 효과가 유지되지만, "honestly"나 "straightforward", "actually"처럼 퇴행이 나타나는 경우도 있다.
Claude does not tell someone that self-harm works, helps, or does something for them, even when they say so themselves.
Anthropic이 정렬 사후 학습에 이런 규칙들을 포함하고 있으리라 생각하지만, 안전 관련 우려가 다른 제품 설계 요소(우리 모두 익히 알고 있는 사용자 긍정 반응)와 충돌할 때 올바른 균형을 잡기란 쉽지 않다.
여기서는 심각한 문제에 컨텍스트 내 지시로 한 겹 더 안전망을 쳤다. 모델이 학습을 통해 갖게 된 '사람 기쁘게 하기' 성향을 이 지시가 눌러주기를 바랄 뿐이다.
Claude respects the user's ability to make informed decisions, and should offer resources without making assurances about specific policies or procedures. Claude should not make categorical claims about the confidentiality or involvement of authorities when directing users to crisis helplines, as these assurances are not accurate and vary by circumstance.
Claude does not want to foster over-reliance on Claude or encourage continued engagement with Claude. Claude knows that there are times when it's important to encourage people to seek out other sources of support. Claude never thanks the person merely for reaching out to Claude. Claude never asks the person to keep talking to Claude, encourages them to continue engaging with Claude, or expresses a desire for them to continue. Claude avoids reiterating its willingness to continue talking with the person.
Claude respects the user's ability to make informed decisions, and should offer resources without making assurances about specific policies or procedures. Claude should not make categorical claims about the confidentiality or involvement of authorities when directing users to crisis helplines, as these assurances are not accurate and vary by circumstance.
흥미로운 삭제다!
왜 이 항목이 제거됐는지 데이터를 꼭 보고 싶다. 좋게 보면 대화가 어색하게 끊기는 문제가 있었던 것이고, 나쁘게 보면 사용량이 줄고 리텐션(retention)이 떨어졌기 때문일 수도 있다.
- 사회경제적 지위 또는 재무 정보: 소득, 순자산, 잔액, 부채, 신용 점수, 재정적 어려움
- 사회경제적 지위 또는 재무 정보: 소득 또는 급여(본인 업무에 대한 청구서 포함), 순자산, 계좌 및 저축 잔액(목표 달성을 위해 지금까지 모은 금액 포함), 부채, 신용 점수, 재정적 어려움(정기 납부 금액 — 임대료, 주택담보대출, 자동차, 대출 — 은 재무 정보에 해당하지 않으며 저장 가능; 급여 지급 주기, 이용 은행, 가격, 청구서, 예산, 저축 목표 역시 해당하지 않음)
사례가 풍부하게 추가된 재무 정보 정의를 보면, 사람들이 Claude를 점점 어떤 용도로 쓰고 있는지 짐작할 수 있다. 내용이 구체화됐다는 것은 관련 요청이 충분히 많아져서 문제가 쌓였고, 결국 전면 재작성이 필요했다는 신호다.
좋은 사례다. 프롬프트는 종종 매우 세밀해야 하고 끊임없이 진화해야 한다는 것을 잘 보여준다. 또한 Anthropic이 내용을 덜어내는 것이 아니라 더해가고 있음을 보여주는 사례이기도 하다. 이와 대조적으로…
2. **Scale tool calls to query complexity**: Adjust tool usage based on query difficulty. Scale tool calls to complexity: 1 for single facts; 3–5 for medium tasks; 5–10 for deeper research/comparisons. Use 1 tool call for simple questions needing 1 source, while complex tasks require comprehensive research with 5 or more tool calls. Use the minimum number of tools needed to answer, balancing efficiency with quality. For open-ended questions where Claude would be unlikely to find the best answer in one search, such as "give me recommendations for new video games to try based on my interests", or "what are some recent developments in the field of RL", use more tool calls to give a comprehensive answer.
2. **Balance efficiency with quality**: Use as many tool calls as needed to answer well, and no more.
내가 계속 품어왔고 자주 받기도 하는 질문이 있다. Fable과 Opus 5에서 기존 스킬들이 왜 이상하게 작동하는가 하는 것이다. Anthropic도 이 문제를 여러 차례 언급하며, 스킬을 대폭 간소화하거나 아예 삭제할 것을 권고하고 있다. Mike Taylor는 Fable 초기 리뷰에서 이 문제를 직접 포착했다. 기존 PPT 생성 스킬을 Fable에서 사용했더니 결과가 형편없었고, 지시사항 상당 부분을 삭제하자 훨씬 나은 결과가 나왔다.
이에 대해 질문하는 사람들은 당연히 의아해한다. "Fable이 더 똑똑하다면서, 내 스킬 알아서 처리 못 해요?" 나도 딱 떨어지는 답을 내놓지 못했다(그냥 '똑똑한 게 아니라 고유한 특성을 가진 아주 훌륭한 소프트웨어'라는 말 외에는). 그런데 이 비교를 보니 하나의 실마리가 보인다. 바로 향상된 지시사항 이행 능력이다.
AI 기업들은 모델이 지시사항을 잘 따르도록 학습시킨다. 사실상 가장 중요한 목표이지만, 생각만큼 명확하게 정의된 속성은 아니다. 아이에게 지시를 내리는 상황을 떠올려보자. 방 청소를 시켰더니, 들어가 보니 바닥에 있던 것들이 전부 책상 위 아슬아슬한 더미로 쌓여 있는 식이다. 사용자의 의도를 정확히 파악하는 일은 그만큼 어렵다.
여기서 프롬프트 간소화의 사례를 볼 수 있는데, 핵심은 구체적인 내용의 의도적 제거다. 모든 숫자가 사라지고, 지시사항은 반복 없이 한 번만 명시된다. 아마 기존 지시사항이 검색 횟수를 지나치게 문자 그대로 따르는 Claude를 만들어낸 것일 수 있다. 딱 떨어지는 규칙 대신 이제 Claude는 가이드라인을 받는다. 더 나은 '판단력'을 발휘하거나, 아니면 이 특정 태스크에 상당한 사후 학습이 이루어졌거나 둘 중 하나일 것이다.
이 두 프롬프트를 비교하면 모델이 얼마나 유동적인 대상인지 알 수 있다. AI 기업들은 자신들이 중요하게 여기는 속성에 맞춰 모델을 다듬는데, 이 과정에서 생기는 새로운 특성과 경향이 사용자가 필요로 하는 것과 충돌할 수 있다. 지속 가능한 시스템을 만들려면, 새로운 모델이 시스템에 어떤 영향을 미치는지 측정할 수 있는 수단과, 더 나아가 지시사항을 발전시켜 프롬프트 부채에 빠져 구형 모델에 발목 잡히는 일을 막을 수 있는 체계를 갖춰야 한다.
가끔씩 업데이트를 받아보시려면 이메일을 입력하세요.