INT 8 quantization is not implemented when selected to export TensorFlow Lite #748

franciscocostela · 2024-06-28T20:17:06Z

Search before asking

I have searched the HUB issues and found no similar bug report.

HUB Component

Export

Bug

I trained an Object Detection Model using both YoloV5n and YoloV8n. In the Deploy tab, I selected TensorFlow Lite - advanced and selected 'Int 8 Quantization'. Then I clicked Export and downloaded the model when the button Download became available.

However, when I inspect the file, it looks like the quantization is never applied. This happens for both YoloV5n and YoloV8n

Is there anything that I am not doing correctly or is this really a bug?

Environment

No response

Minimal Reproducible Example

No response

Additional

No response

github-actions · 2024-06-28T20:17:33Z

👋 Hello @franciscocostela, thank you for raising an issue about Ultralytics HUB 🚀! Please visit our HUB Docs to learn more:

Quickstart. Start training and deploying YOLO models with HUB in seconds.
Datasets: Preparing and Uploading. Learn how to prepare and upload your datasets to HUB in YOLO format.
Projects: Creating and Managing. Group your models into projects for improved organization.
Models: Training and Exporting. Train YOLOv5 and YOLOv8 models on your custom datasets and export them to various formats for deployment.
Integrations. Explore different integration options for your trained models, such as TensorFlow, ONNX, OpenVINO, CoreML, and PaddlePaddle.
Ultralytics HUB App. Learn about the Ultralytics App for iOS and Android, which allows you to run models directly on your mobile device.
- iOS. Learn about YOLO CoreML models accelerated on Apple's Neural Engine on iPhones and iPads.
- Android. Explore TFLite acceleration on mobile devices.
Inference API. Understand how to use the Inference API for running your trained models in the cloud to generate predictions.

If this is a 🐛 Bug Report, please provide screenshots and steps to reproduce your problem to help us get started working on a fix.

If this is a ❓ Question, please provide as much information as possible, including dataset, model, environment details etc. so that we might provide the most helpful response.

We try to respond to all issues as promptly as possible. Thank you for your patience!

sergiuwaxmann · 2024-07-01T09:07:35Z

@franciscocostela Hello!
Can you check if the exported model size is 4x smaller than the fp32 one?

franciscocostela · 2024-07-01T16:53:03Z

Hi Sergiu,

Yes - It is about 4x smaller. These are the sizes of the files:
Original FP32 - 11.6MB
Fp16 - 5.8MB
Int8 - 3.0MB

I am trying to run the TFLite file through a conversion pipeline to deploy it into a camera but it fails with an error message about the file not being quantized. When I inspect it with Netron, I see that the quantization bias is FLOAT32. INT8 is used in some of the convolution layers but not all of them (see screenshot). This seems to trigger the error message using the conversion pipeline.

sergiuwaxmann · 2024-07-02T08:57:27Z

@franciscocostela Based on the file size, the quantization is applied.

github-actions · 2024-08-02T00:20:59Z

👋 Hello there! We wanted to give you a friendly reminder that this issue has not had any recent activity and may be closed soon, but don't worry - you can always reopen it if needed. If you still have any questions or concerns, please feel free to let us know how we can help.

For additional resources and information, please see the links below:

Docs: https://docs.ultralytics.com
HUB: https://hub.ultralytics.com
Community: https://community.ultralytics.com

Feel free to inform us of any other issues you discover or feature requests that come to mind in the future. Pull Requests (PRs) are also always welcomed!

Thank you for your contributions to YOLO 🚀 and Vision AI ⭐

franciscocostela added the bug Something isn't working label Jun 28, 2024

ultralytics deleted a comment from pderrenger Jul 1, 2024

ultralytics deleted a comment from pderrenger Jul 2, 2024

sergiuwaxmann self-assigned this Jul 2, 2024

github-actions bot added the Stale Stale and schedule for closing soon label Aug 2, 2024

github-actions bot closed this as not planned Won't fix, can't repro, duplicate, stale Aug 12, 2024

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

INT 8 quantization is not implemented when selected to export TensorFlow Lite #748

INT 8 quantization is not implemented when selected to export TensorFlow Lite #748

franciscocostela commented Jun 28, 2024

github-actions bot commented Jun 28, 2024

sergiuwaxmann commented Jul 1, 2024

franciscocostela commented Jul 1, 2024

sergiuwaxmann commented Jul 2, 2024

github-actions bot commented Aug 2, 2024

INT 8 quantization is not implemented when selected to export TensorFlow Lite #748

INT 8 quantization is not implemented when selected to export TensorFlow Lite #748

Comments

franciscocostela commented Jun 28, 2024

Search before asking

HUB Component

Bug

Environment

Minimal Reproducible Example

Additional

github-actions bot commented Jun 28, 2024

sergiuwaxmann commented Jul 1, 2024

franciscocostela commented Jul 1, 2024

sergiuwaxmann commented Jul 2, 2024

github-actions bot commented Aug 2, 2024