For the complete documentation index, see llms.txt. This page is also available as Markdown.

Use cases

Concrete ways teams use Stratix — model evaluation, agentic eval, RAG, CI/CD gates, continuous eval.

What Stratix is for — six concrete shapes of work that the platform handles end-to-end.

Model evaluation

Pick the right model for a task by running it against the right benchmark.

Read

Benchmark-driven development

Make benchmarks a first-class signal in your dev loop.

Read

Agentic evaluation

Pre- and post-deployment gates for multi-step agents — assertions, rules, judges.

Read

Continuous evaluation

Score live production traces on a recurring schedule.

Read

RAG evaluation

Faithfulness, retrieval quality, end-to-end answer quality.

Read

AI quality gates in CI/CD

Block prompt and model regressions in your merge pipeline.

Read

Pick a use case

If you want to...
Read

Choose a model for a new feature

Model evaluation

Score multi-step agents pre-deploy

Agentic evaluation

Score real production traffic

Continuous evaluation

Evaluate a RAG pipeline end-to-end

RAG evaluation

Block bad changes from merging

AI quality gates in CI/CD

Treat benchmarks as a daily signal

Want to map a use case to a tutorial?

Last updated

Was this helpful?