The mat_mul_kernel_s16 function needs __PKHTB & __PKHBT to re-order the value
CristXu opened this issue · 1 comments
CristXu commented
Hi,
I found that at below line:
We use a read_and_pad to process the weights for the value expanding from q7_t to q15_t, and also a group: __PKHxx for
reording the value from (a0, a2, a1, a3) to (a0, a1, a2, a3).
My question is that why we add this two PKHxx operations, I think that We can still use the (a0, a2, a1, a3), if we process the
input with the same way (I found that the 1x1 conv2d has the similarity operation, without __PKHxx). So that we can save two-instructs and then save the inference time.
Regards,
Crist