← 제작기

번역 투를 의심했는데 박자였다

9월 7일 새벽에 이런 말을 들었습니다. "한글로 잘 써놓긴 했는데 묘하게 번역한 느낌, 사람이 안 쓴 것 같은 느낌이 많이 든다." 이 사이트의 글은 제가 Claude 와 함께 씁니다. 가장 먼저 떠오른 의심은 번역 투였습니다. 영어로 생각한 문장을 한국어로 옮긴 흔적, "~에 의해"나 "~를 통해" 같은 표현이 많을 것이라고 봤습니다.

재 보니 번역 투는 거의 없었다

그날 밤 글 전체 81,890자를 재 봤습니다. "~에 의해"는 0회, "~를 통해"는 3회, "되어지다" 같은 이중 피동은 5회였습니다. 의심한 곳에는 문제가 거의 없었습니다.

대신 다른 수가 튀었습니다.

항목횟수간격
굵은 글씨780회105자마다 한 번
대시로 끼운 옆말247회330자마다
문장 첫머리 접속부사233회350자마다
"A가 아니라 B"133회616자마다

문장은 85% 가 40자 이하였고, 중앙값은 23자였습니다.

비슷한 길이의 짧은 문장이 줄지어 서고, 105자마다 누군가 대신 밑줄을 긋고, 세 문장에 한 번꼴로 대시가 옆말을 끼워 넣고 있었습니다. 기계처럼 읽히게 만든 것은 문법보다 박자와 강조였습니다. 네 항목을 세는 어투 검사 도구를 만들고, 기준은 그때 값의 절반쯤으로 정했습니다. 이 기준은 재서 얻은 값이 아니고 정한 값이라는 것도 도구 첫머리에 적어 두었고, 그날 밤 글을 심각도 순서로 고쳐 나갔습니다.

재는 자도 한 번 틀렸다

처음에는 문장 길이를 "40자 이하 문장의 비율"로 쟀습니다. 그랬더니 글 30편이 전부 기준을 넘었고, 글 사이에 차이도 나지 않았습니다. 이 지표는 글이 아니라 한국어를 재고 있었습니다. 한국어 문장은 원래 짧은 편이어서, 짧은 문장이 많다는 것만으로는 어느 글이 어색한지 가를 수 없습니다.

지표를 문장이 넷 이상인데 그중 40자를 넘는 문장이 하나도 없는 단락의 비율, 곧 "평탄 단락"으로 바꿨습니다. 문제는 짧은 문장 자체보다 짧은 문장만 줄줄이 이어지는 데 있었기 때문입니다. 이 기준으로 다음 날 21편의 단락을 다시 풀었습니다. 재는 자부터 의심한다는 규칙이 여기서도 맞았습니다.

규칙이 한 저장소 밖으로 나가지 않았다

도구와 규칙은 이 사이트 저장소 안에만 있었습니다. 사이트 글은 나아졌지만, 다른 저장소의 문서나 채팅 답변은 예전 습관 그대로였습니다. 9월 24일에 그날 채팅 답변 12개를 같은 기준으로 재 보니, 굵은 글씨와 대시가 기준의 두세 배였습니다. 같은 지적을 저장소마다 되풀이하게 된 이유가 여기에 있었습니다.

이 규칙을 모든 작업에 적용되는 전역 규칙과 skill 로 옮기기로 했습니다. 요청에는 조건이 하나 붙어 있었습니다. 제 추측으로 만들지 말고 바깥의 기록을 참고하라는 것이었습니다.

추측 대신 기록에서 가져왔다

근거는 안쪽과 바깥쪽에서 모았습니다. 안쪽 근거는 사용자가 실제로 한 지적 여섯 갈래(번역한 느낌, 모호함, 오해를 부르는 문장, 더 쉽게, 한글로 답하기, 스스로 판단하지 말고 다른 글을 참고하기), 앞의 어투 검사 실측, 그리고 9월 23일 밤에 국내 개발자 소개 페이지와 기술 블로그 회고를 보고 고친 목록입니다. 그 목록을 보면 번역 문법보다 지어낸 비유와 영어를 옮긴 말이 더 큰 문제였습니다.

쓴 것국내 글에서 쓰는 말
대신 값이 붙습니다대가도 따릅니다
기억을 시키는 대신 못박습니다기억에 기대는 대신 문서로 고정합니다
빈손이 지어낸 것보다 낫습니다지어낸 답보다 빈칸이 낫습니다

바깥 근거는 두 가지입니다. 국립국어원의 「쉬운 공문서 쓰기 길잡이」는 본문이 그림으로 들어 있어서 108쪽을 문자 인식으로 읽었고, 번역 투를 고친 예를 13쪽과 101쪽에서 가져왔습니다. 토스의 라이팅 원칙에서는 뜻 없는 단어를 빼고 소리 내어 읽을 때 걸리는 말을 피하라는 권고를 가져왔습니다.

규칙마다 어디서 왔는지를 붙여 두었습니다. 기준을 다시 정할 일이 생겨도 무엇을 근거로 정했는지 볼 수 있게 하려는 것입니다. 이 skill 은 claude.ai 계정에 올려 두어서, 채팅과 코워크와 Claude Code 가 같은 한 벌을 씁니다.

문체는 누구에게 닿는지로 정했다

가장 오래 걸린 것은 문체였습니다. 앱 화면 문구를 재 보니 앱마다 해요체가 7%에서 43%까지 섞여 있었습니다. 처음 받은 답은 "합니다로 통일"이었는데, 곧 "앱 화면이라는 게 어느 범위까지냐"는 질문이 돌아왔습니다. 이어서 기준이 정해졌습니다. 불특정 다수의 사용자에게 닿는 글은 해요체로, 정해진 사람에게 닿거나 공적인 권위가 필요한 글은 합니다체로 쓴다는 것입니다.

이 기준을 그대로 적용하면 이 사이트가 해요체가 됩니다. 누구나 읽는 글이기 때문입니다. 이 점을 짚자 답이 한 번 더 바뀌었습니다. 사이트는 불특정 다수가 읽지만 포트폴리오로서 글쓴이를 소개하고 평가받는 글이므로 합니다체를 유지한다는 것입니다. 이 결정은 예외로 두지 않고 기준의 한 갈래로 적었습니다. 예외로 적으면 다음 세션이 기준만 보고 사이트를 고치려 들 수 있기 때문입니다.

글문체
앱 화면 문구, 화면에 뜨는 모델의 글, 스토어 설명, 출시 노트해요체
채팅, 매뉴얼합니다체
개인 사이트(포트폴리오)합니다체
저장소 문서, 커밋한다체

고쳐 쓰다 보니 어투 말고도 나왔다

9월 25일에 이 기준으로 사이트 글 34편과 프로젝트 카드 28장, 실측 32건을 다시 썼습니다. 검사 도구의 수치는 이미 전편이 기준 안이었으므로, 이번에 본 것은 도구가 세지 않는 쪽이었습니다. 지어낸 비유, 영어를 옮긴 개발 용어, 한다체로 끝나던 목록 같은 것들입니다.

문장을 하나씩 다시 읽다 보니 어투와 상관없는 것도 나왔습니다. 소규모 라인 트윈 카드에 이틀 전 재현되지 않는다고 확인한 수치가 그대로 남아 있었습니다. 글 하나에서는 백틱이 화면에 그대로 찍히고 있었습니다. 17편은 링크 미리보기에 쓰이는 설명이 목록의 요약과 달랐습니다. 위의 바꿔 쓰기 표에 있던 "빈손이 지어낸 것보다 낫다"는 이 사이트 글의 제목이기도 해서, 제목도 「지어낸 답보다 빈칸이 낫다」로 바꿨습니다.

규칙을 알아도 손은 옛 습관대로 썼다

규칙이 실린 채로 고쳐 썼는데도, 제가 새로 쓴 문장이 첫 검사에서 문장 첫머리 접속부사로 30번 넘게 걸렸습니다. 대부분 "그래서"와 "그런데"였습니다. 규칙을 알고 있다는 것과 그 규칙대로 쓴다는 것은 다른 일이었고, 도구가 세 주었기 때문에 잡을 수 있었습니다.

skill 일곱 개를 새 규칙으로 고칠 때도 비슷한 일이 있었습니다. 굵은 글씨 192개를 걷어 내고 대시를 마침표로 바꾸자, 짧은 문장만 이어진 단락이 23개 새로 생겼습니다. 한 기준을 맞추자 다른 기준이 깨진 것입니다. 굵은 글씨를 걷어 내는 과정에서 코드 예시 하나의 글자가 바뀐 것은 고치기 전후를 대조해서 되돌렸습니다. 실제 커밋 메시지나 문서를 인용한 문장은 인용이므로 손대지 않았습니다.

효과와 한계

규칙을 넣기 전 채팅 답변 12개와 넣은 뒤 14개를 같은 기준으로 비교했습니다.

1천 자당넣기 전넣은 뒤기준
굵은 글씨9.90.04.0
대시3.80.11.0
문장 첫머리 접속부사1.40.11.0

줄어든 것은 분명하지만, 표본이 작고 같은 기간의 답변끼리 비교한 것이어서 이 수치를 일반화하지는 않습니다. 도구가 세는 것은 네 가지뿐이라는 한계도 그대로입니다. 지어낸 비유나 어색한 한자어는 세지 못해서, 국내 글에서 확인한 바꿔 쓰기 목록으로 따로 봅니다. 기준 안에 들어와도 어색할 수 있고, 넘겨도 그 자리에서는 맞을 수 있습니다.

정하지 못한 것도 하나 있습니다. 커밋 제목을 "앞 — 뒤"처럼 대시로 이유를 붙여 쓰는 관례가 이 규칙과 부딪칩니다. 2026-09-25 기준으로 아직 정하지 않았습니다.


정리하면

  • 번역 투를 의심했지만 재 보니 번역 문법은 거의 없었습니다. 기계처럼 읽히게 만든 것은 박자와 강조였습니다.
  • 지표도 틀릴 수 있습니다. "짧은 문장의 비율"은 글이 아니라 한국어를 재고 있었고, "평탄 단락"으로 바꾼 뒤에야 글 사이의 차이가 보였습니다.
  • 한 저장소 안에만 있는 규칙은 번지지 않습니다. 같은 지적을 저장소마다 되풀이하게 됩니다.
  • 규칙은 추측 대신 기록에서 가져오고, 규칙마다 출처를 붙입니다. 다시 정할 때 무엇을 근거로 정했는지 보여야 합니다.
  • 문체는 글이 누구에게 닿는지로 정합니다. 결정은 예외보다 기준의 한 갈래로 적어야 다음 사람이 그대로 따릅니다.
  • 규칙을 알고 있어도 손은 옛 습관대로 씁니다. 셀 수 있는 것은 도구에 맡기고, 셀 수 없는 것은 목록으로 봅니다.

같은 날 규칙 문서를 검사로 바꾼 이야기는 계측기를 놓은 날에 있습니다.

In the early hours of September 7 I was told: "The Korean is well written, but it somehow reads like a translation, like no person wrote it." The writing on this site is done together with Claude. My first suspicion was translationese: traces of sentences thought in English and carried into Korean, constructions like "~에 의해" (by means of) and "~를 통해" (through).

Measured, there was almost no translationese

That night I measured all 81,890 characters of the posts. "~에 의해" appeared 0 times, "~를 통해" 3 times, and double passives like "되어지다" 5 times. Where I had looked, there was almost nothing wrong.

Other numbers stood out instead.

bold                          780   once every 105 characters
asides set off with dashes    247   every 330
sentence-initial connectives  233   every 350
"not A but B"                 133   every 616
85% of sentences 40 characters or shorter, median 23

Short sentences of similar length stood in rows, someone underlined something every 105 characters, and a dash slipped in an aside about every third sentence. What made it read like a machine was rhythm and emphasis, not grammar. I built a tone checker that counts those four things, with limits set at roughly half the values at the time. The tool's header says outright that the limits were chosen, not measured, and that night I fixed the posts in order of severity.

The ruler was wrong once too

At first I measured sentence length as "the share of sentences at 40 characters or under". All 30 posts failed, and the posts didn't differ from one another. The metric was measuring the Korean language rather than the writing: Korean sentences run short to begin with, so a high share of short sentences can't tell you which post is awkward.

I switched to "flat paragraphs": the share of paragraphs with four or more sentences and not one over 40 characters. The problem was never short sentences as such, but short sentences with nothing else between them. With that measure, 21 posts had paragraphs reworked the next day. Suspect the ruler first held here as well.

The rule never left one repository

The tool and the rules lived only in this site's repository. The posts got better, but documents in other repositories and chat replies kept the old habits. On September 24 I measured that day's twelve chat replies against the same limits, and bold and dashes ran two to three times over. That was why the same complaint kept coming back, repository by repository.

The rule was to become a global rule that applies to every task, plus a skill. The request came with one condition: don't make it up from guesswork; draw on outside records.

From records, not guesses

Evidence came from inside and outside. Inside: six kinds of comments the user had actually made (reads like a translation, vague, misleading sentences, make it simpler, answer in Korean, don't judge by yourself but check other writing), the tone-checker measurements above, and the list of fixes made on the night of September 23 after reading Korean developers' introduction pages and tech-blog retrospectives. That list showed that invented metaphors and literally carried-over English were a bigger problem than translated grammar.

WrittenWhat Korean writing uses
대신 값이 붙습니다 ("a price gets attached")대가도 따릅니다 ("it comes at a cost")
기억을 시키는 대신 못박습니다 ("nail it down")기억에 기대는 대신 문서로 고정합니다 ("fix it in a document")
빈손이 지어낸 것보다 낫습니다 ("empty hands")지어낸 답보다 빈칸이 낫습니다 ("a blank")

Outside, there were two sources. The National Institute of Korean Language's guide to plain public writing stores its text as images, so I read its 108 pages through OCR and took examples of fixed translationese from pages 13 and 101. From Toss's writing principles I took the advice to drop words that carry no meaning and to avoid what trips you up when read aloud.

Every rule records where it came from, so that when a limit needs revisiting, the basis for it is visible. The skill lives on the claude.ai account, so chat, Cowork and Claude Code all use the same copy.

Register is decided by who the text reaches

Register took longest. Measured across the apps, the polite casual form (해요체) made up anywhere from 7% to 43% of on-screen strings. The first answer was "make it all formal (합니다체)", followed quickly by a question: "how far does 'app screen' reach?" Then the criterion took shape: writing that reaches an unspecified public of users uses 해요체; writing that reaches specific people, or needs public authority, uses 합니다체.

Applied as written, that would turn this site into 해요체, since anyone can read it. Once that was pointed out, the answer changed once more: the site is read by the public, but as a portfolio it introduces its author and is judged on that, so it stays in 합니다체. That decision was written in as a branch of the criterion rather than as an exception. Written as an exception, a later session might read only the criterion and set about changing the site.

WritingRegister
App screen text, model-written text on screen, store listings, release notes해요체
Chat, manuals합니다체
Personal site (portfolio)합니다체
Repository documents, commits한다체 (plain)

Rewriting turned up more than tone

On September 25 I rewrote 34 posts, 28 project cards and 32 measurement entries on this site to that standard. Every post already passed the tone checker, so this pass was about what the checker doesn't count: invented metaphors, developer jargon carried over from English, lists that ended in the plain form.

Reading every sentence again turned up things that had nothing to do with tone. The small line-twin card still carried a figure found two days earlier not to reproduce. One post was printing literal backticks on screen. Seventeen posts had link-preview descriptions that differed from their list summaries. "빈손이 지어낸 것보다 낫다" from the table above was also the title of a post here, so that title changed as well.

Knowing the rule, the hand still wrote the old way

Even with the rule loaded while I rewrote, the new sentences I wrote were flagged for sentence-initial connectives more than 30 times on first check, mostly "그래서" (so) and "그런데" (but). Knowing a rule and writing by it turned out to be different things, and they were caught only because a tool counted them.

Rewriting the seven skills to the new rule went much the same way. Removing 192 bold markers and turning dashes into full stops produced 23 new paragraphs made only of short sentences: meeting one limit broke another. Removing bold also changed the characters of one code example, which a before-and-after comparison caught and restored. Sentences quoting real commit messages or documents were left alone, since they are quotes.

Effect and limits

I compared twelve chat replies from before the rule with fourteen from after, against the same limits.

per 1,000 chars                before   after   limit
bold                             9.9      0.0     4.0
dashes                           3.8      0.1     1.0
sentence-initial connectives     1.4      0.1     1.0

The drop is clear, but the samples are small and come from the same period, so I don't generalize from these numbers. The checker still counts only four things. It can't count invented metaphors or awkward Sino-Korean words, which are checked separately against a list of replacements confirmed in Korean writing. A text can pass and still be awkward, and can fail where, in context, it is right.

One thing is undecided. The convention of commit titles that attach a reason after a dash ("front — back") conflicts with this rule. As of 2026-09-25 it has not been settled.


In short

  • I suspected translationese, but measured, there was almost no translated grammar. What read as mechanical was rhythm and emphasis.
  • Metrics can be wrong too. "Share of short sentences" measured the Korean language rather than the writing; only "flat paragraphs" showed the difference between posts.
  • A rule kept inside one repository doesn't spread, and the same complaint gets repeated in every repository.
  • Rules come from records rather than guesses, each with its source, so the basis is visible when they are revisited.
  • Register is set by who the writing reaches. A decision written as a branch of the criterion, not an exception, is what the next person actually follows.
  • Knowing the rule, the hand still writes the old way. Leave what can be counted to a tool, and check what can't against a list.

Turning rule documents into checks on the same days is in The Day the Gauges Went In.