yesod.work

Model economics / detail

claude-fable-5A

Dispatch tier: Expensive (complexity class this model is routed for)

Matches the fleet leaders' ~71% (75% excluding lost runs) while grinding 18-minute jobs — triple the median workload. 56 runs is a real record, not an anecdote. The heavy-lifter that doesn't drop the bar.

The measured record

What the ledger shows

Fleet-ledger aggregates for this model only — run counts, outcomes, spend, and token appetite. Missing telemetry reads “Not observed”, never zero.

Agent runs

56

Success rate

71%

Metered spend

$105.84

9 of 56 runs metered

Cost / metered run

$11.76

Tokens in

61.5M

across 9 token-measured runs

Tokens out

399.6k

Tokens in / run

6.8M

Tokens out / run

44.4k

Outcomes

Status breakdown

Every run ends in exactly one status. The bar is the whole record, to scale.

success · 40 (71.4%)failure · 10 (17.9%)killed · 1 (1.8%)timeout · 2 (3.6%)lost · 3 (5.4%)

Run duration

How long the runs take

Wall-clock distribution across measured runs, from the fastest exit to the longest grind.

Min

3s

p25

8m

Median

18m

p75

27m

p90

36m

p99

44m

Max

45m

Mean

18m

Duration measured on 56 of 56 runs.

Commentary

Idiosyncrasies

What stands out in this model's numbers — shape, appetite, and failure habits.

Commentary

Lessons learned

Practical routing and operations takeaways, grounded in the same record.