Trinity Large Thinking
arcee-ai/trinity-large-thinking
Compare with Virtuoso Large, GLM 5.2, DeepSeek V4 Flash 0731
CompareTrinity Large Thinking
- Arcee AI100%Healthy24h 99.88%
- p50
- 327ms
- p99
- 1.39s
- Throughput
- 163 tps
- Context
- 262K
- $/Mtok
- $0.250 / $0.800
arcee-ai - Parasail100%Healthy24h 99.56%
- p50
- 155ms
- p99
- 433ms
- Throughput
- 72 tps
- Context
- 262K
- $/Mtok
- $0.220 / $0.850
parasail/fp4fp4
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window, averaged over the hour.
Throughput
Median tokens per second.