//clika-runtime/io.clika.runtime/Ops/qkLayerNorm
qkLayerNorm
[common]
fun qkLayerNorm(query: Tensor, key: Tensor? = null, value: Tensor? = null, queryWeight: Tensor? = null, queryBias: Tensor? = null, keyWeight: Tensor? = null, keyBias: Tensor? = null, headDim: Long = 0, eps: Double? = null): List<Tensor>
qkLayerNorm(query: Tensor, key: Tensor? = null, value: Tensor? = null, queryWeight: Tensor? = null, queryBias: Tensor? = null, keyWeight: Tensor? = null, keyBias: Tensor? = null, headDim: Long = 0L, eps: Double? = null): the qk_layer_norm operator. Per-head LAYER norm over packed attention projections, the mean-subtracting sibling of qk_rms_norm. Each contiguous head_dim run of query (and key, when present) is normalized as (x - mean) / sqrt(var + eps) * w + b; value, when given, is normalized the same way with no weight and no bias (it has no affine slots).