Model Optimizer
/
Model Optimizer
/
modelopt API
/
torch
/
kernels
/
quantization
/
linear_attention
/
int8
int8
#
Signed narrow-range INT8 QDQ for recurrent-state tiles.
Back to top
Previous
decode
Next
sparsity