§03.06

How to Track and Cut Claude Code Token Usage

Run /usage to see session tokens, cost and plan limits — /cost is now just an alias. Then cut waste with /clear, subagents and the right model.

published 21 Aug 2026 updated 06 Sept 2026 checked against docs 06 Sept 2026 3 min in Claude Code commands Markdown

Step 6 of 6 · Make Claude Code yours

On this page6 sections
  1. 1. Open /usage
  2. 2. Read the attribution
  3. 3. Clear rather than compact
  4. 4. Match the model to the job
  5. 5. Push verbose work out of your window
  6. Verify it worked

One command tells you where the tokens went. Note that /cost is now an alias for /usage — same screen, so use whichever you can remember.

1. Open /usage

/usage

What it does: shows the Session block — total cost, API duration, and per-model input, output and cache token counts. The dollar figure is computed locally at list price, so treat it as an estimate rather than an invoice.

On Pro, Max, Team and Enterprise plans the same screen adds your plan usage bars. Press d or w to switch between the last 24 hours and the last 7 days.

2. Read the attribution

The plan breakdown attributes recent usage to skills, subagents, plugins and individual MCP servers, and flags any behavior — long context, cache misses — accounting for 10% or more. It’s the fastest way to discover that one forgotten MCP server is eating a fifth of your week.

3. Clear rather than compact

/clear starts a fresh context and costs nothing. /compact has to read everything it summarizes, so it is itself a large request. Session totals reset when /clear starts a new session. See managing context for when each one wins.

4. Match the model to the job

Sonnet handles most coding work and costs less than Opus; reserve Opus for architecture and multi-step reasoning. Switch mid-session with /model. For mechanical subagent tasks, set model: haiku in the agent’s frontmatter.

Extended thinking is on by default and its tokens bill as output. Lower it with /effort for simple work.

5. Push verbose work out of your window

Use a subagent to run the full test suite and report only the failures.

What it does: keeps thousands of lines of output in the subagent’s context and returns a summary to yours. A PreToolUse hook that greps a log before Claude reads it does the same job even more cheaply.

Verify it worked

Run /usage at the start of a task and again at the end. If the total climbed while you barely typed, it’s long context — every turn re-sends the whole conversation, priced in tokens. Your first message after an hour-long break also misses the prompt cache and reprocesses everything.

Source: Manage costs effectively.

← All Claude Code commands plates · Search all guides

↑↓ move↵ openalt+↵ copy first command

Keyboard

⌘/ctrl+K or /
Search all guides
alt+↵
In search: copy the guide's first command
j / k
Move through a list of guides
c
On a guide: copy its first command
t
Toggle light / dark
?
This list