Model economics / detail
gpt-5.6-terraC
Dispatch tier: Medium (complexity class this model is routed for)
Two-for-two on runs the infra didn't eat, which is technically a perfect record and statistically a coin flip. Three runs is a cameo, not a career — come back with volume.
The measured record
What the ledger shows
Fleet-ledger aggregates for this model only — run counts, outcomes, spend, and token appetite. Missing telemetry reads “Not observed”, never zero.
Agent runs
3
Success rate
67%
Metered spend
Not observed
cost-blind: no run was metered
Cost / metered run
Not observed
Tokens in
Not observed
Tokens out
Not observed
Tokens in / run
Not observed
Tokens out / run
Not observed
Outcomes
Status breakdown
Every run ends in exactly one status. The bar is the whole record, to scale.
- lost: 1 of 3 runs (33.3% of total)
Run duration
How long the runs take
Wall-clock distribution across measured runs, from the fastest exit to the longest grind.
Min
9m
p25
12m
Median
14m
p75
16m
p90
18m
p99
19m
Max
19m
Mean
14m
Duration measured on 3 of 3 runs.
Commentary
Idiosyncrasies
What stands out in this model's numbers — shape, appetite, and failure habits.
- 2-for-2 on the runs the infra did not eat (one of three was lost).
- Typical run around 14 minutes — unhurried for a three-run cameo.
- No token or cost telemetry.
Commentary
Lessons learned
Practical routing and operations takeaways, grounded in the same record.
- n=3 — nothing to learn yet; a perfect record two runs long is a coin landing heads twice.