The outcome

Choose a model based on measured quality, speed, limits, and stability for the intended application.

Step by step

A workflow you can repeat.

  1. 01

    Define task examples, expected outputs, context requirements, safety cases, latency targets, and a consistent scoring method.

  2. 02

    Inspect the current SambaCloud model catalog and label each candidate as production or preview before testing.

  3. 03

    Run identical prompts with fixed parameters and record output quality, latency, token use, formatting, refusals, and errors.

  4. 04

    Test realistic request volume against current per-user limits and inspect response headers rather than relying on static quotas.

  5. 05

    Select a production primary and compatible fallback, record exact model IDs, and schedule regression tests for catalog changes.

Working standard

What good use looks like.

  • Keep preview models out of production paths.
  • Test the actual workload.
  • Read current response-limit headers.

Official references

Check the current product documentation.