메뉴
HN
Hacker News • 29일 전

클로드 쿼터 10분 만에 소진되어 원인 분석 도구 'Tare' 제작

IMP
6/10
핵심 요약

Hacker News에 공개된 'Tare'는 Claude Code의 사용량 로그를 Claude Code 스스로 읽게 해서, 자연어로 질문만 하면 토큰이 어디에 소비됐는지 원인을 찾아주는 도구입니다. 대시보드나 별도 명령어 없이 한 줄 설치(npx)로 사용할 수 있으며, 반복 전송되는 컨텍스트('컨테이너' 무게)를 분리해 진짜 원인을 보여줍니다.

번역된 본문

tare — Claude Code에게 사용량이 어디로 갔는지 물어보세요.

사용량 한도에 걸렸는데 이유를 모르겠거나, 쿼터가 예전보다 빨리 소진되거나, 뭔가 뒤에서 토큰을 잡아먹고 있다는 의심이 들 때가 있습니다. 이 모든 질문에 답하는 기록은 이미 여러분의 컴퓨터에 있습니다 — Claude Code는 모든 요청을 로그로 남깁니다. tare는 Claude Code가 자기 자신의 로그를 읽도록 해서, 그냥 물어보기만 하면 되게 만듭니다. 대시보드도 없고, 배울 명령어도 없습니다. 한 번 설치한 뒤 평범한 말로 질문하면 됩니다.

Tare(포장 무게): 내용물을 알기 위해 빼는 컨테이너의 무게입니다. 세션 비용의 대부분은 컨테이너 — 반복해서 다시 전송되는 컨텍스트 — 이며, 이 도구들이 정확히 그 부분을 차감해 보여줍니다.

설치 아무 터미널에서 명령 한 줄:

npx skills add kelviq/tare -g -y --copy --agent claude-code

끝입니다. 계정도, 패키지도, 설정도 필요 없습니다. 새 Claude Code 세션을 시작하면 바로 작동합니다 — / 를 입력해 tare가 나타나는지 확인하세요. 나타나지 않는다면 문제 해결 참고사항을 보세요 — Claude Code가 스킬을 한 번도 설치한 적 없는 머신에서는 설치 프로그램이 이를 놓칠 수 있습니다. (수동 설치, 팀 전체 설치, Claude Code 플러그인으로 설치하는 방법은 INSTALL.md에 있습니다.)

그냥 질문하기 Claude Code를 열고 사람에게 물어보듯 질문하면 됩니다. 아래 질문들이 모두 작동합니다 — 정확한 단어를 쓸 필요 없이, 한도에 대한 불평만 해도 충분합니다:

한도에 걸렸을 때

  • 어제 왜 사용량 한도에 걸렸지?
  • 저녁 세션 시작 10분 만에 잠겼는데 어떻게 가능하지?
  • 5시간 한도인가요, 주간 한도인가요?
  • 이번 주에 쿼터가 왜 이렇게 빨리 소모되지?

토큰이 어디로 가는지

  • 이번 주에 내 토큰이 실제로 어디로 갔지?
  • 어떤 프로젝트가 내 쿼터를 잡아먹고 있지?
  • 어느 모델이 가장 비용이 많이 나오지?
  • Claude가 계속 다시 읽는 파일 중 가장 비싼 것은?
  • 어제 그 거대한 세션은 실제로 얼마나 들었지?
  • MCP 서버들이 내 컨텍스트를 많이 차지하나?
  • 서브에이전트와 스킬이 오버헤드를 얼마나 추가하나?

큰 작업 시작 전

  • 지금 5시간 윈도우가 얼마나 찼지?
  • 큰 리팩터링을 지금 시작해도 안전할까, 아니면 윈도우가 비울 때까지 기다려야 할까?

잊고 있던 것 확인

  • 뭔가 백그라운드에서 Claude Code를 실행 중인가?
  • 내가 자는 동안 Claude Code가 활동했나?
  • 지난달에 자동화를 설정했는데 얼마나 비용이 들고 있지?

설정 튜닝

  • 작업 사이에 /clear를 쓰기 시작했는데 실제로 효과가 있었나? 이번 주와 지난주를 비교해줘.
  • 최신 Claude Code 업데이트가 사용량을 바꿨나?
  • API를 직접 결제하면 사용량이 얼마나 나올까?
  • 어떤 변경 하나가 가장 많이 아껴줄까?

리포트와 공유

  • 브라우저에서 열 수 있는 사용량 리포트를 만들어줘.
  • 개인정보 없이 공개적으로 게시할 수 있는 요약을 줘.
  • 사용량을 스프레드시트로 내보내줘.

명령어 입력을 선호한다면? /tare는 전체 진단을 실행하고, 자주 묻는 질문의 변형도 받습니다 — 물론 평범한 말로 물어도 항상 작동합니다:

  • /tare — 전체 진단: 토큰이 어디로 갔고 왜 그랬는지
  • /tare usage — 한눈에 보는 사용량 패널 — /usage와 비슷하지만 원인 귀속 포함
  • /tare window — 5시간 윈도우가 얼마나 찼는지 — 시작해도 안전한가?
  • /tare report [일수] — HTML 리포트 생성 후 열기
  • /tare tools [일수] — 무엇이 컨텍스트를 채우는지
  • /tare week — 이번 주와 지난주 비교
  • /tare share [일수] — 공개 게시 가능한 개인정보 제거 요약
  • /tare 어제 왜 한도에 걸렸지 — 어떤 질문이든 인자로 작동

돌려받는 것 숫자의 벽이 아니라 — 원인입니다. "어제 왜 한도에 걸렸지?"에 대한 답은 이렇습니다:

어제 사용량의 99%는 당신이 아니라 당신이 실행 중인 어떤 도구에서 나왔습니다. 웹사이트 프로젝트에서 어떤 것이 1,553개의 짧은 Claude Code 세션을 생성했습니다 — 9,022개의 요청, 한 번에 최대 51개 세션 동시 실행. 당신이 직접 작업한 그날의 요청은 93개였습니다. 새 세션은 매번 컨텍스트를 처음부터 다시 만드는데, 이것이 토큰을 쓰는 가장 비싼 방식입니다...

...그 뒤에 증거, 확인할 것, 바꿀 것이 따라옵니다. 그리고 모든 게 실제로 정상일 때는 그렇게 말해줍니다: 사용량 비례 적정, 이상 없음, 당신에게 정상인 수치는 이것입니다.

실제 출력 결과는 examples/에 있습니다 — /tare share로 생성한 공유 가능한 요약 등이 있습니다.

원문 보기
원문 보기 (영어)
tare Ask Claude Code where your usage went. You hit a usage limit and don't know why. Your quota drains faster than it used to. You suspect something is eating tokens in the background. The records that answer all of this are already on your computer — Claude Code keeps a log of every request it makes. tare teaches Claude Code to read its own logs, so you can just ask. No dashboards, no commands to learn. Install once, then ask in plain English. Tare: the weight of the container, subtracted to find what's inside. Most of what a session costs is the container — context re-sent again and again — and that's exactly what these tools subtract. Install One command, in any terminal: npx skills add kelviq/tare -g -y --copy --agent claude-code That's it. Nothing else to set up — no accounts, no packages, no configuration. Start a new Claude Code session and it's live — type / and check that tare appears. If it doesn't, see the troubleshooting note — on machines where Claude Code has never installed a skill before, the installer can miss it. (Other ways to install — by hand, for a whole team, or as a Claude Code plugin — are in INSTALL.md .) Then just ask Open Claude Code and ask your question the way you'd ask a person. These all work — the words don't have to match, complaining about your limits is enough: When you hit a limit Why did I hit my usage limit yesterday? I got locked out ten minutes into my evening session — how is that possible? Did I hit the 5-hour limit or the weekly cap? Why am I burning through my quota so much faster this week? Where your tokens go Where did my tokens actually go this week? Which of my projects is eating my quota? Which model is costing me the most? What's the most expensive file Claude keeps re-reading? How much did that giant session yesterday actually cost me? Are my MCP servers adding a lot to my context? How much overhead do subagents and skills add? Before you start something big How full is my 5-hour window right now? Is it safe to start a big refactor now, or should I wait for my window to clear? Checking for things you forgot Is something running Claude Code in the background? Was Claude Code active while I was asleep? I set up an automation last month — what is it costing me? Tuning your setup I started using /clear between tasks — did it actually help? Compare this week to last. Did the latest Claude Code update change my usage? What would my usage cost if I were paying for the API directly? What one change would save me the most? Reports and sharing Make me a usage report I can open in my browser. Give me a summary I can post publicly — with nothing private in it. Export my usage to a spreadsheet. Prefer typing commands? /tare runs the full diagnosis, and takes variants for the common asks — asking in plain words always works too: /tare full diagnosis — where tokens went and why /tare usage at-a-glance panel — like /usage , with attribution /tare window how full is the 5-hour window — safe to start? /tare report [days] build the HTML report and open it /tare tools [days] what is filling my context /tare week compare this week with last /tare share [days] redacted summary safe to post publicly /tare why did I hit the limit yesterday any question works as the argument What you get back Not a wall of numbers — a cause. The answer to "why did I hit my limit yesterday?" looks like this: 99% of yesterday's usage came from a tool you're running, not from you. Something spawned 1,553 short Claude Code sessions in your website project — 9,022 requests, up to 51 sessions running at once. Your own hands-on work that day was 93 requests. Each fresh session rebuilds its context from scratch, which is the most expensive way to spend tokens... ...followed by the evidence, what to check, and what to change. And when everything is actually fine, it says that: usage proportionate, no anomaly, here's what's normal for you. Real output lives in examples/ — a shareable summary produced by /tare share , and the HTML report. What it knows that a raw token count doesn't Correct totals. Claude Code's log format repeats each API response several times over; naive counting inflates totals — by 86% on the data this was built against. tare deduplicates properly. The real cost of context. A file read early in a long session gets re-sent with every later message. tare charges tools for what they caused , not just what they returned — which is how one big file read early can quietly dominate a week. The rolling window. Limits don't reset when you walk away; work from four hours ago still counts. tare can tell you how full your window was at the exact moment you were locked out. The shape of automation. Hundreds of short parallel sessions is a script, not a person. tare recognises the signature and says so. Private by design Everything runs on your machine and nothing leaves it. The scripts make no network connections at all. When you ask for a shareable summary, it contains totals, dates and tool names only — no prompts, no file paths or contents, no commands, no session or account identifiers — so you can post it publicly or send it to a colleague and ask "what am I missing?" Don't take that on faith: SECURITY.md states exactly what each file reads, writes and sends, and shows how to verify every claim yourself with one grep. Requirements Claude Code on macOS or Linux Python 3.9+ — already present on every Mac; no packages to install Currently reads Claude Code's logs only, not other coding agents' Demo tare-screen-recording.mp4 For developers Everything the skill does, you can also do by hand: three dependency-free Python scripts with recipes for scripting, cron, CSV export, live per-request telemetry and more — see CLI.md . Issues and PRs welcome. Useful directions: a live TUI, Windows paths, aggregating anonymised summaries across users to spot patterns no single person can see, and better token estimation for tool results. MIT licensed.