PT - JOURNAL ARTICLE AU - Kabir, Md Humayun AU - Djordjevic, Djordje AU - O’Connor, Michael D. AU - Ho, Joshua W. K. TI - C3: An R package for cross-species compendium-based cell-type identification AID - 10.1101/267880 DP - 2018 Jan 01 TA - bioRxiv PG - 267880 4099 - http://biorxiv.org/content/early/2018/02/19/267880.short 4100 - http://biorxiv.org/content/early/2018/02/19/267880.full AB - Cell type identification from an unknown sample can often be done by comparing its gene expression profile against a gene expression database containing profiles of a large number of cell-types. This type of compendium-based cell-type identification strategy is particularly successful for human and mouse samples because a large volume of data exists for these organisms. However, such rich data repositories often do not exist for most non-model organisms. This makes transcriptome-based sample classification in these species challenging. We propose to overcome this challenge by performing a cross-species compendium comparison. The key is to utilise a recently published cross-species gene set analysis (XGSA) framework to correct for biases that may arise due to potentially complex homologous gene mapping between two species. The framework is implemented as an open source R package called C3. We have evaluated the performance of C3 using a variety of public data in NCBI Gene Expression Omnibus. We also compared the functionality and performance of C3 against some similar gene expression profile matching tools. Our evaluation shows that C3 is a simple and effective method for cell type identification. C3 is available at https://github.com/VCCRI/C3.