tutorial to investigate the xnnpack to save initial time and memory consumptions.
Open
- Dominant language
- C
- Stars
- 2.5k
- Forks
- 560
- Avg merge
- 1d 6h
- Merged PRs (30d)
- 163
Description
To whom it may concern,
Xnnpack delegate are successfully created on my tflite model deployment and accelerate the inference. However, the initial time and the memory consumption increased (200ms-->700ms, 18MB --> 70MB) after turn on the xnnpack delegate.
Is such increase normal? Could you please provide the tutorial to investigate how to save the initial time and memory consumption?
I'd appreciate any help I can get.
Kind regards,
Li
Contributor guide
Assessment
This issue has not been assessed yet.