tensorflow / tensorflow/model-optimization

Input and resource quantization

Open
#1,003 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

feature request
Dominant language
Python
Stars
1.6k
Forks
349
Avg merge
3d 2h
Merged PRs (30d)
1

Description

System information

  • TensorFlow version (you are using): tensorflow 2.9.1
  • Are you willing to contribute it (Yes/No): No

Motivation
The input and resource need interface for customer layer quantization,
Normalization is a basic structure in the DNN, but we cannot quantize it easily.

  1. No input quantization interface in customer configuration, it will make quantization mul-add structure in float32 instead of int8 in tflite file.
  2. No resource quantization interface in customer configuration, it will make batchnorm - (running mean, running variance) wo/ QAT.

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The issue names no files, tests, or entry points. Start by locating customer quantization configuration and the TensorFlow Lite conversion and QAT paths; compare how input tensors and batch-normalization resources are currently handled. Done should include agreed interfaces and coverage showing that the relevant operations can be quantized as intended.

Written by the indexing model from the issue text.

Assessment

Tech stack
keras, python, tensorflow
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.