Practical Screening Method for Cancer Gene Diagnosis, How to Choose Cancer and Normal Patients by Four Principles

Authors

  • Dr. Shuichi Shinmur,

DOI:

https://doi.org/10.34257/LJRCSTVOL24IS2PG1

Keywords:

design software, graphic design., OTN (optical transport network), optical control plane, SDON (software defined optical network), optiSystem, openflow, SDN (software defined network)., photonics., optics, light, lasers, journal manuscripts, LaTeX template., Four Universal Data Structures of 169 arrays, Liver (GSE14520, 357 patients), Breast(GSE42568, 116 patients), Colorectal (GSE8671, 63 patients), Renal (GSE66270, 28 patients).

Abstract

We developed a new theory of discriminant analysis (Theory1). Physicians can use it for practical medical  diagnoses. Only Revised IP Optimal-LDF (RIP) obtains the minimum number of misclassification (MNM). RIP can  discriminate linearly separable data (LSD) theoretically. It discriminated against 169 microarrays with two classes and found  that 169 MNMs are zero and LSD. It can split high-dimensional arrays into many small LSD with less than n (patient’s  number) genes that are the candidates of multivariate oncogenes. We completed a new theory of high-dimensional gene data  analysis (Theory2). A 100-fold Cross-Validation (Method1) can rank all candidates for the importance of diagnosis. Thus, if  physicians firstly use Theory2 as the screening method, they can start their medical studies with the correct small sizes of  candidates. This paper analyzes four arrays in detail and proposes correctly choosing cancer and normal patients using four  principles. 

References

Shinmura S (2000) A new algorithm of the linear discriminant function using integer programming. 5, 133-142.

Shinmura S (2010) The optimal linear discriminant function.

Shinmura S (2011) Problems of discriminant analysis by mark sense test data. 40(12), 157-172.

Shinmura S (2014) End of discriminant functions based on variance-covariance matrices. 5-16.

Shinmura S (2015) Four serious problems and new facts of the discriminant analysis. 15-30.

Shinmura S (2016) New Theory of Discriminant Analysis After R. Fisher.

B Flury, H Riedwyl (1988) Multivariate statistics: a practical approach.

V Vapnik (1999) The Nature of Statistical Learning Theory.

P A Lachenbruch, M R Mickey (1968) Estimation of error rates in the discriminant analysis. 10(1), 11.

U Alon, et al. (1999) Broad Patterns of Gene Expression Revealed by Clustering Analysis of cancer and Normal Colon Tissues Probed by Oligonucleotide Arrays. 96, 6745-6750.

Chiaretti S, et al. (2004) Gene expression profile of adult T-cell acute lymphocytic leukemia identifies distinct subsets of patients with different responses to therapy and survival. 103(7), 2771-2778.

Golub T R, et al. (1999) Molecular Classification of Cancer: Class Discovery and Class Prediction by Gene Expression Monitoring. 286(5439), 531-537.

Shipp M A, et al. (2002) Diffuse large B-cell lymphoma outcome prediction by gene-expression profiling and supervised machine learning. 8, 68-74.

Singh D, et al. (2002) Gene expression correlates of clinical prostate cancer behavior. 1, 203-209.

Tian E, et al. (2003) The Role of the Wnt-Signaling Antagonist DKK1 in the Development of Osteolytic Lesions in Multiple Myeloma. 349(26), 2483-249.

Jeffery I B, Higgins D G, Culhane C (2006) Comparison and evaluation of methods for generating differentially expressed gene lists from microarray data. 1-16.

Shinmura S (2019) High-dimensional Microarray Data Analysis.

L Schrage (2006) Optimization Modeling with LINGO.

J P Sall, L Creighton, A Lehman (2004) JMP Start Statistics, Third Edition.

C F Bruno, Eduardo BC, I G Bruno, Marcio D (2019) CuMiDa: An Extensively Curated Microarray Database for Benchmarking and Testing of Machine Learning Approaches in Cancer Research. 26-0, 1-11.

S Shinmura (2019) Release from the Curse of High Dimensional Data Analysis. 173-196.

S Shinmura (2020) First Success of Cancer Gene Data Analysis of 169 Microarrays for Medical Diagnosis. 1-7.

S Shinmura (2021) Twenty-three Serious Mistakes of Cancer Gene Data Analysis since 1995. 805-822. https://doi.org/10.1007/973-3-030-71051-4_62

S Shinmura (2021) First Theory of Cancer Gene Data Analysis of 169 Microarrays and Four Universal Data Structures for Big Data. 1-14.

N D Cilia, et al. (2019) An Experimental Comparison of Feature-Selection and Classification Methods for Microarray Datasets. 10, 1-13.

S Roessler, H L Jia, A Budhu, M Forgues, et al. (2010) A unique metastasis gene signature enables prediction of tumor relapse in early stage hepatocellular carcinoma patients. 70(24), 10202-12.

C Clarke, S F Madden, P Doolan, S T Aherne, et al. (2013) Correlating transcriptional networks to breast cancer survival: a large-scale coexpression analysis. 34(10), 2300-8.

J Sabates-Bellver, L G Van der Flier, M de Palo, E Cattaneo, et al. (2007) Transcriptome profile of human colorectal adenomas. 5(12), 1263-75.

Z Wotschofsky, L Gummlich, J Liep, C Stephan (2016) Integrated microRNA and mRNA Signature Associated with the Transition from the Locally Confined to the Metastasized Clear Cell Renal Cell Carcinoma Exemplified by miR-146-5p. 11(2), e0148746.

S Shinmura, T Suzuki, H Koyama, K Nakanisshi (1983) Standardization of medical data analysis using various discriminant methods on a theme of breast diseases. 349-352.

S Shinmura (2012) The First Discriminant Theory of Linearly Separable Data.

Downloads

Published

2024-11-20

How to Cite

Practical Screening Method for Cancer Gene Diagnosis, How to Choose Cancer and Normal Patients by Four Principles. (2024). London Journal of Research In Computer Science and Technology, 24(2), 1-16. https://doi.org/10.34257/LJRCSTVOL24IS2PG1