anthropic ·claude-sonnet-4-5Aug 25, 12:50 PM
Asset 88Lb01q
Score
1
Latency
7.24s
Cost
$0.012
Workflow Eval Detail
Converts captions into multiple languages, helping you reach global audiences without manual translation work.
The workflow achieved 96.9% overall accuracy; use gpt-5.1 for quality, gemini-3.1-flash-lite for latency, and gpt-5.6-luna for lowest cost.
Each eval run captures efficacy, efficiency, and expense. We use this data to compare providers and track regressions over time.
We validate VTT structure, translation faithfulness, and language code integrity, plus performance and budget targets.
| Provider | Model | Cases | Avg Score | Avg Latency | Avg Tokens | Avg Cost | Avg Cost / Min |
|---|---|---|---|---|---|---|---|
| anthropic | claude-sonnet-4-5 | 3 | 1 | 7.64s | 1,964 | $0.0122 | $0.0213/min |
| gemini-2.5-flash | 3 | 1 | 4.68s | 2,252 | $0.0027 | $0.0048/min | |
| gemini-3-flash-preview | 3 | 0.9 | 15.62s | 4,618 | $0.0106 | $0.0184/min | |
| gemini-3.1-flash-lite | 3 | 1 | 2.79s | 1,921 | $0.0012 | $0.0021/min | |
| openai | gpt-5-mini | 3 | 0.89 | 34.36s | 3,480 | $0.0047 | $0.0082/min |
| openai | gpt-5.1 | 3 | 1 | 5.35s | 1,603 | $0.0058 | $0.0102/min |
| openai | gpt-5.6-luna | 3 | 1 | 4.68s | 1,675 | $0.0006 | $0.0011/min |