IDEAS home Printed from https://ideas.repec.org/a/bpj/sagmbi/v12y2013i6p743-755n6.html
   My bibliography  Save this article

Estimation of weighted log partial area under the ROC curve and its application to MicroRNA expression data

Author

Listed:
  • Hossain Ahmed

    (Clinical Epidemiology and Biostatistics, McMaster University, 1280 Main Street West, Hamilton, Ontario L8S4K1, Canada)

  • Beyene Joseph

    (Clinical Epidemiology and Biostatistics, McMaster University, 1280 Main Street West, Hamilton, Ontario L8S4K1, Canada)

Abstract

MicroRNAs (miRNAs) are short non-coding RNAs that play critical roles in numerous cellular processes through post-transcriptional functions. The aberrant role of miRNAs has been reported in a number of diseases. A robust computational method is vital to discover novel miRNAs where level of noise varies dramatically across the different miRNAs. In this paper, we propose a flexible rank-based procedure for estimating a weighted log partial area under the receiver operating characteristic (ROC) curve statistic for selecting differentially expressed miRNAs. The statistic combines results taking partial area under the curve (pAUC) and their corresponding variance. The proposed method does not involve complicated formulas and does not require advanced programming skills. Two real datasets are analyzed to illustrate the method and a simulation study is carried out to assess the performance of different miRNA ranking statistics. We conclude that the proposed method offers robust results with large samples for miRNA expression data, and the method can be used as an alternative analytical tool for identifying a list of target miRNAs for further biological and clinical investigation.

Suggested Citation

  • Hossain Ahmed & Beyene Joseph, 2013. "Estimation of weighted log partial area under the ROC curve and its application to MicroRNA expression data," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 12(6), pages 743-755, December.
  • Handle: RePEc:bpj:sagmbi:v:12:y:2013:i:6:p:743-755:n:6
    DOI: 10.1515/sagmb-2013-0035
    as

    Download full text from publisher

    File URL: https://doi.org/10.1515/sagmb-2013-0035
    Download Restriction: For access to full text, subscription to the journal or payment for the individual article is required.

    File URL: https://libkey.io/10.1515/sagmb-2013-0035?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Margaret Sullivan Pepe & Gary Longton & Garnet L. Anderson & Michel Schummer, 2003. "Selecting Differentially Expressed Genes from Microarray Experiments," Biometrics, The International Biometric Society, vol. 59(1), pages 133-142, March.
    2. Efron B. & Tibshirani R. & Storey J.D. & Tusher V., 2001. "Empirical Bayes Analysis of a Microarray Experiment," Journal of the American Statistical Association, American Statistical Association, vol. 96, pages 1151-1160, December.
    3. Victor Ambros, 2004. "The functions of animal microRNAs," Nature, Nature, vol. 431(7006), pages 350-355, September.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Ahmed Hossain & Hafiz T.A. Khan, 2016. "Identification of genomic markers correlated with sensitivity in solid tumors to Dasatinib using sparse principal components," Journal of Applied Statistics, Taylor & Francis Journals, vol. 43(14), pages 2538-2549, October.

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Ahmed Hossain & Hafiz T.A. Khan, 2016. "Identification of genomic markers correlated with sensitivity in solid tumors to Dasatinib using sparse principal components," Journal of Applied Statistics, Taylor & Francis Journals, vol. 43(14), pages 2538-2549, October.
    2. Ahmed Hossain & Joseph Beyene, 2015. "Application of skew-normal distribution for detecting differential expression to microRNA data," Journal of Applied Statistics, Taylor & Francis Journals, vol. 42(3), pages 477-491, March.
    3. Debashis Ghosh & Arul Chinnaiyan, 2004. "Covariate adjustment in the analysis of microarray data from clinical studies," The University of Michigan Department of Biostatistics Working Paper Series 1030, Berkeley Electronic Press.
    4. Bickel David R., 2008. "Correcting the Estimated Level of Differential Expression for Gene Selection Bias: Application to a Microarray Study," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 7(1), pages 1-27, March.
    5. Youngchao Ge & Sandrine Dudoit & Terence Speed, 2003. "Resampling-based multiple testing for microarray data analysis," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 12(1), pages 1-77, June.
    6. Pounds Stanley B. & Gao Cuilan L. & Zhang Hui, 2012. "Empirical Bayesian Selection of Hypothesis Testing Procedures for Analysis of Sequence Count Expression Data," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 11(5), pages 1-32, October.
    7. Niels Lundtorp Olsen & Alessia Pini & Simone Vantini, 2021. "False discovery rate for functional data," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 30(3), pages 784-809, September.
    8. Li-Xuan Qin & Steven G. Self, 2006. "The Clustering of Regression Models Method with Applications in Gene Expression Data," Biometrics, The International Biometric Society, vol. 62(2), pages 526-533, June.
    9. Wen Shi & Xi Chen & Jennifer Shang, 2019. "An Efficient Morris Method-Based Framework for Simulation Factor Screening," INFORMS Journal on Computing, INFORMS, vol. 31(4), pages 745-770, October.
    10. Hossain, Ahmed & Beyene, Joseph & Willan, Andrew R. & Hu, Pingzhao, 2009. "A flexible approximate likelihood ratio test for detecting differential expression in microarray data," Computational Statistics & Data Analysis, Elsevier, vol. 53(10), pages 3685-3695, August.
    11. Dørum Guro & Snipen Lars & Solheim Margrete & Saebo Solve, 2011. "Smoothing Gene Expression Data with Network Information Improves Consistency of Regulated Genes," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 10(1), pages 1-26, August.
    12. José María Galván-Román & Ángel Lancho-Sánchez & Sergio Luquero-Bueno & Lorena Vega-Piris & Jose Curbelo & Marcos Manzaneque-Pradales & Manuel Gómez & Hortensia de la Fuente & Mara Ortega-Gómez & Javi, 2020. "Usefulness of circulating microRNAs miR-146a and miR-16-5p as prognostic biomarkers in community-acquired pneumonia," PLOS ONE, Public Library of Science, vol. 15(10), pages 1-13, October.
    13. Daniel Yekutieli, 2015. "Bayesian tests for composite alternative hypotheses in cross-tabulated data," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 24(2), pages 287-301, June.
    14. Gong Chen & Qing Zhou, 2010. "Heterogeneity in DNA Multiple Alignments: Modeling, Inference, and Applications in Motif Finding," Biometrics, The International Biometric Society, vol. 66(3), pages 694-704, September.
    15. Kshitij Srivastava & Anvesha Srivastava, 2012. "Comprehensive Review of Genetic Association Studies and Meta-Analyses on miRNA Polymorphisms and Cancer Risk," PLOS ONE, Public Library of Science, vol. 7(11), pages 1-1, November.
    16. Ghosh Debashis, 2012. "Incorporating the Empirical Null Hypothesis into the Benjamini-Hochberg Procedure," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 11(4), pages 1-21, July.
    17. Ruth Heller & Saharon Rosset, 2021. "Optimal control of false discovery criteria in the two‐group model," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 83(1), pages 133-155, February.
    18. Yu Lianbo & Gulati Parul & Fernandez Soledad & Pennell Michael & Kirschner Lawrence & Jarjoura David, 2011. "Fully Moderated T-statistic for Small Sample Size Gene Expression Arrays," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 10(1), pages 1-22, September.
    19. Werft, W. & Benner, A. & Kopp-Schneider, A., 2012. "On the identification of predictive biomarkers: Detecting treatment-by-gene interaction in high-dimensional data," Computational Statistics & Data Analysis, Elsevier, vol. 56(5), pages 1275-1286.
    20. repec:dau:papers:123456789/13437 is not listed on IDEAS
    21. Xing Chen & Jun Yin & Jia Qu & Li Huang, 2018. "MDHGI: Matrix Decomposition and Heterogeneous Graph Inference for miRNA-disease association prediction," PLOS Computational Biology, Public Library of Science, vol. 14(8), pages 1-24, August.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:bpj:sagmbi:v:12:y:2013:i:6:p:743-755:n:6. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Peter Golla (email available below). General contact details of provider: https://www.degruyter.com .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.