LLM classifier
A general-purpose language model constrained to a typed candidate choice.
Record
- Unrated
- Not estimated
- 6
- 5 / 0 / 10
- 2,285
Provisional quick suite. seed-01, seed-02, seed-03. Only IQ contributes to CR. Quick samples cannot establish a significant difference.
Footnotes
IQ
- 1.26 s
- 2.07 s
- $0.02362
- Not measured
- 163 / 163
- 0
- 0
- 832.25 µs / 2.4 ms
- 0
- 0
- 0
- 0
Blitz · systems stress test
- Not measured
- Not measured
- Not reported
- Not measured
- 0 / 292
- 292
- 0
- 732.29 µs / 2.4 ms
- 600
- 0
- 0
- 0
Percentiles cover completed calls only. Timeouts are censored and shown separately. Optional top-out forecasts depend on the policy and are not classifier accuracy.
Model: openai/gpt-4o-mini