Arm backend: Enable test_llama_tosa_BI and related fixes (#10681)
First problem solved by adding quantize of scalar_tensor: The where.self
operator got a scalar_tensor input which was not quantized. This
happened since the where.self quantization annotator uses the parent
specs, which in this case where non-existing. Adding the quantization of
scalar_tensor sorts this out.
Secondly when quantizing scalar_tensor the following assert triggers:
expecting kwargs for aten op IR to be empty
Hence setting scalar_tensor kwargs to {}.
Finally trying to quantize -inf fails for scalar_tensor nodes fails Fix
it by adding the pass from qnn backend to replace -inf/inf. Hence adding
new pass ReplaceInfValues.
Co-authored-by: Per Åstrand <per.astrand@arm.com> M
Måns Nilsson committed
51befeecf7c94138c297960df8a202e683f41737
Parent: 6da46fb
Committed by GitHub <noreply@github.com>
on 5/5/2025, 2:31:04 PM