yesod.work

Model economics / detail

glm-5p2A

Dispatch tier: Medium (complexity class this model is routed for)

Second-biggest workhorse in the fleet, and once you strip the 55 infra-lost runs its 79% adjusted success rate quietly leads the high-volume pack. Docked from S for gulping ~2M tokens per job — reliable, but thirsty.

The measured record

What the ledger shows

Fleet-ledger aggregates for this model only — run counts, outcomes, spend, and token appetite. Missing telemetry reads “Not observed”, never zero.

Agent runs

366

Success rate

67%

Metered spend

$205.40

223 of 366 runs metered

Cost / metered run

$0.92

Tokens in

773.2M

across 268 token-measured runs

Tokens out

5.4M

Tokens in / run

2.9M

Tokens out / run

20.2k

Outcomes

Status breakdown

Every run ends in exactly one status. The bar is the whole record, to scale.

success · 247 (67.5%)failure · 62 (16.9%)timeout · 2 (0.5%)lost · 55 (15.0%)

Run duration

How long the runs take

Wall-clock distribution across measured runs, from the fastest exit to the longest grind.

Min

5s

p25

2m

Median

6m

p75

11m

p90

19m

p99

49m

Max

23.0h

Mean

12m

Duration measured on 366 of 366 runs.

Commentary

Idiosyncrasies

What stands out in this model's numbers — shape, appetite, and failure habits.

Commentary

Lessons learned

Practical routing and operations takeaways, grounded in the same record.