Show the per-model usage limit, and size the status line by measurement
/usage reports a weekly limit scoped to a single model (Fable) that the
status line did not surface. It is absent from both the status line's stdin
JSON and the legacy seven_day_opus/seven_day_sonnet fields, which are always
null now; it lives in the usage API's .limits[] under kind "weekly_scoped".
Render it labelled by scope.model.display_name so it follows whichever model
the limit applies to, inside the 7d segment since both are weekly limits
sharing one reset time.
Fitting it exposed a problem with sizing by width tier. Tiers only know the
terminal width, so content that varies with session state — branch name, cwd,
model display name — could push a line past the edge and be clipped by the
renderer. Measure the assembled line instead (vis_len) and emit the richest of
four rate-limit variants that fits, on one line or two: with reset times, with
extra-usage credits, bars only, or bars without the per-model segment. Every
branch is fit-checked, including with wrapping disabled. Two lines render
correctly in the status line.
Cap the cwd basename as well: nothing else shortened line one, so a long
project directory could overflow it on its own.
Correct the usable width. Claude Code exports COLUMNS but applies the `padding`
setting on top of it rather than deducting it first, so a line sized to COLUMNS
is clipped; deduct padding on both sides plus a column of margin.
Sanitise payload data before rendering. printf %b interprets backslash escapes,
which is how the colour variables work, so a cwd or model name containing a
literal \n — or a real newline, which jq decodes from the JSON — would split
the line and break both the width measurement and the two-line guarantee.
Parse the usage payload in one jq pass rather than ten. The status line runs on
every redraw and per-field parsing had become the dominant cost; this brings a
render back to roughly what it was before (~155ms vs ~115ms parent), the
remainder being the reset times the tiered version did not show at this width.
Fields are read one per line rather than via @tsv: tab is IFS whitespace, so
`read` collapses runs of it and an empty field — no scoped limit, which is the
common case — silently shifted every later field along by one.
Tests cover the API payload shapes including malformed and hostile input, the
width invariants (never exceed the usable width, never more than two lines,
never wrap unnecessarily), and a sweep over widths, branch lengths, model
names, cwd lengths and extra-usage. They restore the live usage cache on exit,
and remove the fixture outright when there was no cache to restore — otherwise
a test run would leave fabricated usage figures live for an hour.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-27 17:55:35 +01:00
#!/bin/bash
# Tests for claude/statusline.isaacaudet.sh — the Isaac Audet-inspired Claude
# Code status line (the variant symlinked from ~/.claude/statusline.sh).
#
# Sibling variants in claude/ (statusline.burnrate.sh, statusline.original.sh)
# are NOT covered by these tests.
#
# Payload shapes: how the usage API's rate-limit JSON is rendered, in
# particular the per-model ("weekly_scoped") limit such as Fable, and how the
# renderer degrades when the group will not fit.
SCRIPT = " $( cd " $( dirname " ${ BASH_SOURCE [0] } " ) /.. " && pwd ) /statusline.isaacaudet.sh "
Wrap the status line rather than drop the per-model limit
The measured ladder tried every single-line variant before considering a
second line, so whenever line one plus the full group did not fit — from
around 92 usable columns upwards, depending on how long the branch, cwd and
model names are — it emitted rl_bare (5h and 7d only) and silently dropped
the per-model (Fable) bar that is the main reason the group is worth
rendering. On a 110-column laptop the limit was invisible.
Reorder so a second line beats losing that bar: one line rich/mid/lean, then
wrap, and rl_bare only when WRAP_NARROW is false. Reset times and
extra-usage credits are still given up rather than wrapped for, which makes
the rendered content non-monotone in width; the comment on the ladder spells
that out.
Also stop the tests writing fixtures to the caches the live status line
reads. ~/.claude/statusline.sh is a symlink to the script, so a test run was
visibly rendering fabricated usage — a Fable bar at 10%, extra-usage credits
of $1234.56/$2000.00 — in whatever session happened to be open, and a run
killed before its trap fired would have left that in place for an hour. The
cache directory is now $STATUSLINE_CACHE_DIR (default /tmp/claude) and each
suite points it at a temporary directory, which also removes the
backup/restore dance and the deletion of the live git-status cache.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-31 17:28:55 +01:00
# Render into a throwaway cache directory. The live cache must not be touched:
# ~/.claude/statusline.sh is a symlink to the script under test, so a fixture
# written there is shown in the running TUI as real usage until it expires.
export STATUSLINE_CACHE_DIR = " $( mktemp -d) "
CACHE = " $STATUSLINE_CACHE_DIR /statusline-usage-cache.json "
Show the per-model usage limit, and size the status line by measurement
/usage reports a weekly limit scoped to a single model (Fable) that the
status line did not surface. It is absent from both the status line's stdin
JSON and the legacy seven_day_opus/seven_day_sonnet fields, which are always
null now; it lives in the usage API's .limits[] under kind "weekly_scoped".
Render it labelled by scope.model.display_name so it follows whichever model
the limit applies to, inside the 7d segment since both are weekly limits
sharing one reset time.
Fitting it exposed a problem with sizing by width tier. Tiers only know the
terminal width, so content that varies with session state — branch name, cwd,
model display name — could push a line past the edge and be clipped by the
renderer. Measure the assembled line instead (vis_len) and emit the richest of
four rate-limit variants that fits, on one line or two: with reset times, with
extra-usage credits, bars only, or bars without the per-model segment. Every
branch is fit-checked, including with wrapping disabled. Two lines render
correctly in the status line.
Cap the cwd basename as well: nothing else shortened line one, so a long
project directory could overflow it on its own.
Correct the usable width. Claude Code exports COLUMNS but applies the `padding`
setting on top of it rather than deducting it first, so a line sized to COLUMNS
is clipped; deduct padding on both sides plus a column of margin.
Sanitise payload data before rendering. printf %b interprets backslash escapes,
which is how the colour variables work, so a cwd or model name containing a
literal \n — or a real newline, which jq decodes from the JSON — would split
the line and break both the width measurement and the two-line guarantee.
Parse the usage payload in one jq pass rather than ten. The status line runs on
every redraw and per-field parsing had become the dominant cost; this brings a
render back to roughly what it was before (~155ms vs ~115ms parent), the
remainder being the reset times the tiered version did not show at this width.
Fields are read one per line rather than via @tsv: tab is IFS whitespace, so
`read` collapses runs of it and an empty field — no scoped limit, which is the
common case — silently shifted every later field along by one.
Tests cover the API payload shapes including malformed and hostile input, the
width invariants (never exceed the usable width, never more than two lines,
never wrap unnecessarily), and a sweep over widths, branch lengths, model
names, cwd lengths and extra-usage. They restore the live usage cache on exit,
and remove the fixture outright when there was no cache to restore — otherwise
a test run would leave fabricated usage figures live for an hour.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-27 17:55:35 +01:00
REPO = " $( mktemp -d) "
Wrap the status line rather than drop the per-model limit
The measured ladder tried every single-line variant before considering a
second line, so whenever line one plus the full group did not fit — from
around 92 usable columns upwards, depending on how long the branch, cwd and
model names are — it emitted rl_bare (5h and 7d only) and silently dropped
the per-model (Fable) bar that is the main reason the group is worth
rendering. On a 110-column laptop the limit was invisible.
Reorder so a second line beats losing that bar: one line rich/mid/lean, then
wrap, and rl_bare only when WRAP_NARROW is false. Reset times and
extra-usage credits are still given up rather than wrapped for, which makes
the rendered content non-monotone in width; the comment on the ladder spells
that out.
Also stop the tests writing fixtures to the caches the live status line
reads. ~/.claude/statusline.sh is a symlink to the script, so a test run was
visibly rendering fabricated usage — a Fable bar at 10%, extra-usage credits
of $1234.56/$2000.00 — in whatever session happened to be open, and a run
killed before its trap fired would have left that in place for an hour. The
cache directory is now $STATUSLINE_CACHE_DIR (default /tmp/claude) and each
suite points it at a temporary directory, which also removes the
backup/restore dance and the deletion of the live git-status cache.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-31 17:28:55 +01:00
cleanup( ) { rm -rf " $STATUSLINE_CACHE_DIR " " $REPO " ; }
Show the per-model usage limit, and size the status line by measurement
/usage reports a weekly limit scoped to a single model (Fable) that the
status line did not surface. It is absent from both the status line's stdin
JSON and the legacy seven_day_opus/seven_day_sonnet fields, which are always
null now; it lives in the usage API's .limits[] under kind "weekly_scoped".
Render it labelled by scope.model.display_name so it follows whichever model
the limit applies to, inside the 7d segment since both are weekly limits
sharing one reset time.
Fitting it exposed a problem with sizing by width tier. Tiers only know the
terminal width, so content that varies with session state — branch name, cwd,
model display name — could push a line past the edge and be clipped by the
renderer. Measure the assembled line instead (vis_len) and emit the richest of
four rate-limit variants that fits, on one line or two: with reset times, with
extra-usage credits, bars only, or bars without the per-model segment. Every
branch is fit-checked, including with wrapping disabled. Two lines render
correctly in the status line.
Cap the cwd basename as well: nothing else shortened line one, so a long
project directory could overflow it on its own.
Correct the usable width. Claude Code exports COLUMNS but applies the `padding`
setting on top of it rather than deducting it first, so a line sized to COLUMNS
is clipped; deduct padding on both sides plus a column of margin.
Sanitise payload data before rendering. printf %b interprets backslash escapes,
which is how the colour variables work, so a cwd or model name containing a
literal \n — or a real newline, which jq decodes from the JSON — would split
the line and break both the width measurement and the two-line guarantee.
Parse the usage payload in one jq pass rather than ten. The status line runs on
every redraw and per-field parsing had become the dominant cost; this brings a
render back to roughly what it was before (~155ms vs ~115ms parent), the
remainder being the reset times the tiered version did not show at this width.
Fields are read one per line rather than via @tsv: tab is IFS whitespace, so
`read` collapses runs of it and an empty field — no scoped limit, which is the
common case — silently shifted every later field along by one.
Tests cover the API payload shapes including malformed and hostile input, the
width invariants (never exceed the usable width, never more than two lines,
never wrap unnecessarily), and a sweep over widths, branch lengths, model
names, cwd lengths and extra-usage. They restore the live usage cache on exit,
and remove the fixture outright when there was no cache to restore — otherwise
a test run would leave fabricated usage figures live for an hour.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-27 17:55:35 +01:00
trap cleanup EXIT
git -C " $REPO " init -q 2>/dev/null
git -C " $REPO " -c user.email= t@t -c user.name= t commit -q --allow-empty -m init 2>/dev/null
pass = 0; fail = 0
ok( ) { echo " PASS $1 " ; pass = $(( pass+1)) ; }
bad( ) { echo " FAIL $1 " ; [ -n " $2 " ] && printf '%s\n' " $2 " | sed 's/^/ /' ; fail = $(( fail+1)) ; }
# Render, with ANSI stripped. $1 = TERM_WIDTH, $2 = cwd, $3 = model name.
render( ) {
local w = " $1 " cwd = " ${ 2 :- $REPO } " model = " ${ 3 :- Opus 5 } "
stdin_json " $cwd " " $model " \
| TERM_WIDTH = " $w " bash " $SCRIPT " 2>& 1 | sed $'s/\033\[[0-9;]*m//g'
}
# Build the payload in python: a backslash written into shell-constructed JSON
# is a JSON escape, so "pro\nj" would reach the script as a REAL newline rather
# than the literal characters the hostile-string tests mean to exercise.
stdin_json( ) {
CWD = " $1 " MODEL = " $2 " python3 -c '
import json, os
print( json.dumps( { "model" : { "display_name" : os.environ[ "MODEL" ] } ,
"cwd" : os.environ[ "CWD" ] ,
"cost" : { "total_cost_usd" : 0.5} ,
"context_window" : { "context_window_size" : 200000,
"current_usage" : { "input_tokens" : 1000} } } ) ) '
}
nth( ) { printf '%s\n' " $1 " | sed -n " ${ 2 } p " ; }
count( ) { printf '%s\n' " $1 " | wc -l | tr -d ' ' ; }
vis( ) { printf '%s' " $1 " | python3 -c "
import sys,unicodedata
s = sys.stdin.read( ) ; print( sum( 2 if unicodedata.east_asian_width( c) in 'WF' else 1 for c in s) ) " ; }
fixture( ) { cat > " $CACHE " ; }
BASE = ' "five_hour" :{ "utilization" :5.0,"resets_at" :"2026-08-27T15:40:00+00:00" } ,
"seven_day" :{ "utilization" :7.0,"resets_at" :"2026-08-28T02:00:00+00:00" } '
# ---------------------------------------------------------------- scoped limit
echo "Per-model (weekly_scoped) limit:"
fixture <<JSON
{ $BASE ,"extra_usage" :{ "is_enabled" :false} ,
"limits" :[ { "kind" :"weekly_scoped" ,"percent" :10,"scope" :{ "model" :{ "display_name" :"Fable" } } ,"is_active" :true} ] }
JSON
out = $( render 250)
# Assert the label AND its percentage together, so a bar rendered with the
# wrong number cannot pass.
case " $out " in *"Fable" *"10%" *) ok "wide: renders 'Fable' with its percentage" ; ;
*) bad "wide: renders 'Fable' with its percentage" " $out " ; ; esac
# Wrapped: assert Fable is on line TWO specifically, not merely present.
out = $( render 90)
[ " $( count " $out " ) " = "2" ] && ok "narrow: wraps to two lines" || bad "narrow: wraps to two lines" " $out "
case " $( nth " $out " 2) " in *"Fable" *"10%" *) ok "narrow: Fable is on line two" ; ;
*) bad "narrow: Fable is on line two" " $out " ; ; esac
case " $( nth " $out " 1) " in *"Fable" *) bad "narrow: Fable not duplicated on line one" " $out " ; ;
*) ok "narrow: Fable not duplicated on line one" ; ; esac
Wrap the status line rather than drop the per-model limit
The measured ladder tried every single-line variant before considering a
second line, so whenever line one plus the full group did not fit — from
around 92 usable columns upwards, depending on how long the branch, cwd and
model names are — it emitted rl_bare (5h and 7d only) and silently dropped
the per-model (Fable) bar that is the main reason the group is worth
rendering. On a 110-column laptop the limit was invisible.
Reorder so a second line beats losing that bar: one line rich/mid/lean, then
wrap, and rl_bare only when WRAP_NARROW is false. Reset times and
extra-usage credits are still given up rather than wrapped for, which makes
the rendered content non-monotone in width; the comment on the ladder spells
that out.
Also stop the tests writing fixtures to the caches the live status line
reads. ~/.claude/statusline.sh is a symlink to the script, so a test run was
visibly rendering fabricated usage — a Fable bar at 10%, extra-usage credits
of $1234.56/$2000.00 — in whatever session happened to be open, and a run
killed before its trap fired would have left that in place for an hour. The
cache directory is now $STATUSLINE_CACHE_DIR (default /tmp/claude) and each
suite points it at a temporary directory, which also removes the
backup/restore dance and the deletion of the live git-status cache.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-31 17:28:55 +01:00
# The per-model bar is the reason the group exists on a Fable-capable account,
# so it must survive by wrapping and never be dropped to squeeze the group onto
# one line. Regression: on a ~110-column laptop it vanished silently.
echo "Mid widths keep the per-model bar (wrapping if need be):"
for w in 105 110 116 120; do
out = $( render " $w " " $REPO " "Opus 5 (1M context)" )
case " $out " in *"Fable" *"10%" *) ok " width $w : Fable still shown " ; ;
*) bad " width $w : Fable still shown " " $out " ; ; esac
done
Show the per-model usage limit, and size the status line by measurement
/usage reports a weekly limit scoped to a single model (Fable) that the
status line did not surface. It is absent from both the status line's stdin
JSON and the legacy seven_day_opus/seven_day_sonnet fields, which are always
null now; it lives in the usage API's .limits[] under kind "weekly_scoped".
Render it labelled by scope.model.display_name so it follows whichever model
the limit applies to, inside the 7d segment since both are weekly limits
sharing one reset time.
Fitting it exposed a problem with sizing by width tier. Tiers only know the
terminal width, so content that varies with session state — branch name, cwd,
model display name — could push a line past the edge and be clipped by the
renderer. Measure the assembled line instead (vis_len) and emit the richest of
four rate-limit variants that fits, on one line or two: with reset times, with
extra-usage credits, bars only, or bars without the per-model segment. Every
branch is fit-checked, including with wrapping disabled. Two lines render
correctly in the status line.
Cap the cwd basename as well: nothing else shortened line one, so a long
project directory could overflow it on its own.
Correct the usable width. Claude Code exports COLUMNS but applies the `padding`
setting on top of it rather than deducting it first, so a line sized to COLUMNS
is clipped; deduct padding on both sides plus a column of margin.
Sanitise payload data before rendering. printf %b interprets backslash escapes,
which is how the colour variables work, so a cwd or model name containing a
literal \n — or a real newline, which jq decodes from the JSON — would split
the line and break both the width measurement and the two-line guarantee.
Parse the usage payload in one jq pass rather than ten. The status line runs on
every redraw and per-field parsing had become the dominant cost; this brings a
render back to roughly what it was before (~155ms vs ~115ms parent), the
remainder being the reset times the tiered version did not show at this width.
Fields are read one per line rather than via @tsv: tab is IFS whitespace, so
`read` collapses runs of it and an empty field — no scoped limit, which is the
common case — silently shifted every later field along by one.
Tests cover the API payload shapes including malformed and hostile input, the
width invariants (never exceed the usable width, never more than two lines,
never wrap unnecessarily), and a sweep over widths, branch lengths, model
names, cwd lengths and extra-usage. They restore the live usage cache on exit,
and remove the fixture outright when there was no cache to restore — otherwise
a test run would leave fabricated usage figures live for an hour.
Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-27 17:55:35 +01:00
# The label must come from the payload, not be hardcoded.
fixture <<JSON
{ $BASE ,"extra_usage" :{ "is_enabled" :false} ,
"limits" :[ { "kind" :"weekly_scoped" ,"percent" :42,"scope" :{ "model" :{ "display_name" :"Nimbus" } } ,"is_active" :false} ] }
JSON
out = $( render 250)
case " $out " in *"Nimbus" *"42%" *) ok "label comes from scope.model.display_name" ; ;
*) bad "label comes from scope.model.display_name" " $out " ; ; esac
case " $out " in *Fable*) bad "no hardcoded 'Fable'" " $out " ; ; *) ok "no hardcoded 'Fable'" ; ; esac
# ------------------------------------------------------------- malformed input
echo "Malformed / legacy payloads still render the 5h and 7d bars:"
for desc in " no limits key:{ $BASE ,\"extra_usage\":{\"is_enabled\":false}} " \
" limits is a string:{ $BASE ,\"extra_usage\":{\"is_enabled\":false},\"limits\":\"nope\"} " \
" limits has no scoped:{ $BASE ,\"extra_usage\":{\"is_enabled\":false},\"limits\":[{\"kind\":\"weekly_all\",\"percent\":7}]} " \
" percent as string:{ $BASE ,\"extra_usage\":{\"is_enabled\":false},\"limits\":[{\"kind\":\"weekly_scoped\",\"percent\":\"42.7\",\"scope\":{\"model\":{\"display_name\":\"Str\"}},\"is_active\":true}]} " ; do
name = " ${ desc %% : * } " ; json = " ${ desc #* : } "
printf '%s' " $json " | fixture
out = $( render 250)
if [ " $( count " $out " ) " = "1" ] && case " $out " in *"5h" *"7d" *) true; ; *) false; ; esac \
&& case " $out " in *error*| *null*| *"jq:" *| *"line " *) false; ; *) true; ; esac ; then
ok " $name "
else bad " $name " " $out " ; fi
done
# a string percentage should round, not become 0
case " $out " in *"Str" *"43%" *) ok "percent given as a string rounds to 43%" ; ;
*) bad "percent given as a string rounds to 43%" " $out " ; ; esac
# ------------------------------------------------- extra usage (regression §1)
echo "Extra-usage credits must not overflow line two:"
fixture <<JSON
{ $BASE ,
"extra_usage" :{ "is_enabled" :true,"monthly_limit" :200000,"used_credits" :123456,"utilization" :62.0} ,
"limits" :[ { "kind" :"weekly_scoped" ,"percent" :10,"scope" :{ "model" :{ "display_name" :"Fable" } } ,"is_active" :true} ] }
JSON
out = $( render 250)
case " $out " in *"extra" *) ok "wide: extra-usage segment is shown" ; ;
*) bad "wide: extra-usage segment is shown" " $out " ; ; esac
for w in 73 80 95; do
out = $( render " $w " ) ; u = $(( w - 5 )) ; worst = 0
while IFS = read -r l; do c = $( vis " $l " ) ; [ " $c " -gt " $worst " ] && worst = $c ; done <<< " $out "
[ " $worst " -le " $u " ] \
&& ok " width $w : no line exceeds usable $u (worst $worst ) " \
|| bad " width $w : no line exceeds usable $u (worst $worst ) " " $out "
done
# ------------------------------------------ hostile strings (regression §9)
echo "Control characters and backslashes in payload data:"
fixture <<JSON
{ $BASE ,"extra_usage" :{ "is_enabled" :false} ,
"limits" :[ { "kind" :"weekly_scoped" ,"percent" :10,"scope" :{ "model" :{ "display_name" :"A\nB" } } ,"is_active" :true} ] }
JSON
out = $( render 250)
[ " $( count " $out " ) " = "1" ] && ok "newline in display_name does not split the line" \
|| bad "newline in display_name does not split the line" " $out "
fixture <<JSON
{ $BASE ,"extra_usage" :{ "is_enabled" :false} ,"limits" :[ ] }
JSON
mkdir -p " $REPO /sub "
out = $( render 250 " $REPO /pro\\nj " )
[ " $( count " $out " ) " = "1" ] && ok "literal backslash-n in cwd does not split the line" \
|| bad "literal backslash-n in cwd does not split the line" " $out "
out = $( render 250 " $REPO " 'Opus\n5' )
[ " $( count " $out " ) " = "1" ] && ok "literal backslash-n in model name does not split the line" \
|| bad "literal backslash-n in model name does not split the line" " $out "
# --------------------------------------------- WRAP_NARROW=false (regression §2)
echo "With WRAP_NARROW=false the single line must still be fit-checked:"
NOWRAP = $( mktemp) ; sed 's/^WRAP_NARROW=true/WRAP_NARROW=false/' " $SCRIPT " > " $NOWRAP "
fixture <<JSON
{ $BASE ,"extra_usage" :{ "is_enabled" :false} ,
"limits" :[ { "kind" :"weekly_scoped" ,"percent" :10,"scope" :{ "model" :{ "display_name" :"Fable" } } ,"is_active" :true} ] }
JSON
for w in 81 105 130; do
out = $( stdin_json " $REPO " "Opus 5" | TERM_WIDTH = $w bash " $NOWRAP " 2>& 1 | sed $'s/\033\[[0-9;]*m//g' )
u = $(( w - 5 )) ; c = $( vis " $out " )
[ " $( count " $out " ) " = "1" ] && [ " $c " -le " $u " ] \
&& ok " width $w : one line, $c <= usable $u " \
|| bad " width $w : one line within usable $u (got $( count " $out " ) line/s, $c cols) " " $out "
done
rm -f " $NOWRAP "
echo; echo " pass= $pass fail= $fail " ; [ " $fail " -eq 0 ]