yesod.work

Model economics / detail

kimi-k2p7-codeA

Dispatch tier: Inexpensive (complexity class this model is routed for)

293 runs at 72% (76% excluding infra-lost), 4m median, lean token diet — quietly outperforms every other high-volume workhorse. Not flashy enough for S, but the fleet's best price-per-success.

The measured record

What the ledger shows

Fleet-ledger aggregates for this model only — run counts, outcomes, spend, and token appetite. Missing telemetry reads “Not observed”, never zero.

Agent runs

293

Success rate

72%

Metered spend

$76.72

157 of 293 runs metered

Cost / metered run

$0.49

Tokens in

420.6M

across 191 token-measured runs

Tokens out

2.6M

Tokens in / run

2.2M

Tokens out / run

13.6k

Outcomes

Status breakdown

Every run ends in exactly one status. The bar is the whole record, to scale.

success · 211 (72.0%)failure · 60 (20.5%)timeout · 5 (1.7%)lost · 17 (5.8%)

Run duration

How long the runs take

Wall-clock distribution across measured runs, from the fastest exit to the longest grind.

Min

2s

p25

53s

Median

4m

p75

11m

p90

24m

p99

60m

Max

60m

Mean

9m

Duration measured on 293 of 293 runs.

Commentary

Idiosyncrasies

What stands out in this model's numbers — shape, appetite, and failure habits.

Commentary

Lessons learned

Practical routing and operations takeaways, grounded in the same record.