|
Continuum C++ API
Unified runtime for token + tensor execution
|
Inputs the kit drives the backend with. More...
#include <conformance.hpp>
Public Attributes | |
| std::string | model_id = "conformance/model" |
| std::string | prompt = "You are a careful assistant. Answer briefly. Question: what is a prefix cache?" |
| std::int32_t | prefix_len = -1 |
| std::int32_t | max_tokens = 16 |
| bool | expect_deterministic = true |
| Require identical output for identical cold inputs at temperature 0. | |
| std::string | tensor_op = "identity" |
| Tensor-path op; the backend must return its single input unchanged. | |
Inputs the kit drives the backend with.
| bool continuum::backend::ConformanceOptions::expect_deterministic = true |
Require identical output for identical cold inputs at temperature 0.
| std::int32_t continuum::backend::ConformanceOptions::max_tokens = 16 |
| std::string continuum::backend::ConformanceOptions::model_id = "conformance/model" |
| std::int32_t continuum::backend::ConformanceOptions::prefix_len = -1 |
| std::string continuum::backend::ConformanceOptions::prompt = "You are a careful assistant. Answer briefly. Question: what is a prefix cache?" |
Token-path prompt. Its first prefix_len bytes act as the cached prefix on the warm run (default: half the prompt).
| std::string continuum::backend::ConformanceOptions::tensor_op = "identity" |
Tensor-path op; the backend must return its single input unchanged.