Skip to main content

How-to guides

Practical guides, each answering one concrete "how do I X". They assume the ground the tutorial covers (a project, a device, a model, a run) and go deeper on one problem at a time, in any order. Entries marked as coming are planned and land here as they are written.

Benchmarks and results

  • Read benchmark results: the metrics a run returns, the per-sample evidence behind a score, and how to pull both through the API and the CLI.
  • Compare two runs (coming): pick a baseline, read a regression, and share the comparison.

Definitions

Automation

  • Set up MCP for Claude: give Claude Code or Claude Desktop the platform as a set of tools, over the deployment's own endpoint or the CLI's stdio server.
  • Run benchmarks from CI (coming): an API key, clika-rt apply, and clika-rt benchmarks watch as a gate.

Administration

For the vocabulary these guides use, see Concepts. For the commands, see the CLI reference.