BenchOptions
The knobs of ModelGraph.bench(); every field has a default. warmup untimed runs precede iterations timed ones; fill picks the synthesized input values; device places the inputs (None: the device of the graph's first bound input, else the CPU); reuse_output_buffers writes every run after the first into preallocated outputs, so the timed runs exclude output allocation; batch_axis overrides the guessed axis items are counted along; profile_dir, when set, receives a profile of the measured runs.
batch_axis (property)
The axis items are counted along, or None for the guessed one.
device (property)
The device the inputs are placed on (a Device or its string form), or None for the default.
fill (property)
How synthesized inputs are filled (a FeedFill).
iterations (property)
Measured runs; every one lands in the report.
profile_dir (property)
Where a profile of the measured runs is written, or None.
reuse_output_buffers (property)
Write every run after the first into preallocated outputs.
warmup (property)
Untimed runs before the measured ones.
__init__
__init____init__(self, *, warmup: int = 3, iterations: int = 20, fill: clika_runtime._core.io.FeedFill = FeedFill.Zeros, device: object | None = None, reuse_output_buffers: bool = True, batch_axis: object | None = None, profile_dir: object | None = None) -> None
init(self, *, warmup: int = 3, iterations: int = 20, fill: clika_runtime._core.io.FeedFill = FeedFill.Zeros, device: object | None = None, reuse_output_buffers: bool = True, batch_axis: object | None = None, profile_dir: object | None = None) -> None
BenchOptions(*, warmup=3, iterations=20, fill=FeedFill.Zeros, device=None, reuse_output_buffers=True, batch_axis=None, profile_dir=None)
The knobs of one bench() call; device is a Device or its string form ('cpu', 'cuda:1').