For the complete documentation index, see llms.txt. Markdown versions of all pages are available by appending .md to any URL (e.g. /max/get-started.md).
Mojo function
build_block_scaled_configs
def build_block_scaled_configs[a_type: DType, b_type: DType, c_type: DType, sfa_dtype: DType, sfb_dtype: DType, N: Int, K: Int, transpose_b: Bool = True]() -> Set[BlockScaledMatmulConfig[a_type, b_type, c_type, sfa_dtype, sfb_dtype, transpose_b]]
Build a set of BlockScaledMatmulConfig instances by sweeping M from 8 to 8192.
Parameters:
- βa_type (
DType):DTypeof the A (left) operand elements. - βb_type (
DType):DTypeof the B (right) operand elements. - βc_type (
DType):DTypeof the output matrix elements. - βsfa_dtype (
DType):DTypeof the A operand block scaling factors. - βsfb_dtype (
DType):DTypeof the B operand block scaling factors. - βN (
Int): The N dimension (output columns) of the matmul. - βK (
Int): The K dimension (contraction axis) of the matmul. - βtranspose_b (
Bool): Whether the B operand is stored transposed (defaults toTrue).
Returns:
Set[BlockScaledMatmulConfig[a_type, b_type, c_type, sfa_dtype, sfb_dtype, transpose_b]]: A set of unique BlockScaledMatmulConfig instances covering the swept M range.
Was this page helpful?
Thank you! We'll create more content like this.
Thank you for helping us improve!