International Journal of Innovative Computing and Applications (Print), v. 7, p. 147, 2016.

Reconstruct and quantify the RNA molecules in a cell at a given moment is an important problem in molecular biology that allows one to know which genes are being expressed and at which intensity level. Such problem is known as Transcriptome Reconstruction and Quantification Problem (TRQP). Although several approaches were already designed that solve the TRQP, none of them model it as a combinatorial optimization problem. In order to narrow this gap, we present here a new combinatorial optimization problem called Maximum Similarity Partitioning Problem (MSPP) that models the TRQP. In addition, we prove that the MSPP is NP-complete in the strong sense and present a greedy heuristic for it.