The outcome
Choose a model based on measured quality, speed, limits, and stability for the intended application.
Step by step
A workflow you can repeat.
- 01
Define task examples, expected outputs, context requirements, safety cases, latency targets, and a consistent scoring method.
- 02
Inspect the current SambaCloud model catalog and label each candidate as production or preview before testing.
- 03
Run identical prompts with fixed parameters and record output quality, latency, token use, formatting, refusals, and errors.
- 04
Test realistic request volume against current per-user limits and inspect response headers rather than relying on static quotas.
- 05
Select a production primary and compatible fallback, record exact model IDs, and schedule regression tests for catalog changes.
Working standard
What good use looks like.
- Keep preview models out of production paths.
- Test the actual workload.
- Read current response-limit headers.
Official references