Post-Workshop Survey: Mean Time Taken by Model

Stata Translation Workshop · Post-Workshop Survey

Model used vs. time taken to translate

9 respondents. One bar per model, height = mean time.

0 50 100 150 200 min 109 avg Opus n=4 75 avg Sonnet n=2 30 Fable n=1 5 Haiku 4.5 n=1 10 Haiku n=1
Mean time (minutes)
Single, incomplete run (Fable, ~30 min – ran out of credits)

Opus's average (109 min) is pulled up by a single 180-minute attempt — its four attempts actually ranged 60–180 (see the table below). Haiku and Haiku 4.5 were the fastest (5–10 min), but each is a single data point. Fable's 30 minutes isn't a completed translation — that run stopped when the respondent ran out of tokens/credits, so it understates how long Fable would actually take.

With 1–4 attempts per model and effort/complexity varying by task, this is far too little data to conclude one model is reliably faster than another — read it as what happened in these 9 attempts, not a benchmark.

View as table
ModelMean (min)Rangen
Opus10960–1804
Sonnet7560–902
Fable30*—1
Haiku 4.55—1
Haiku10—1
ModelTime (min)Note
Fable30stopped – ran out of credits
Opus60
Opus75"with lot of debugging"
Opus120
Opus180
Sonnet60
Sonnet90
Haiku 4.55
Haiku10