EP2046997A2 - Konventionelle genexpressionssignatur in dilatierter kardiomyopathie - Google Patents
Konventionelle genexpressionssignatur in dilatierter kardiomyopathieInfo
- Publication number
- EP2046997A2 EP2046997A2 EP07866600A EP07866600A EP2046997A2 EP 2046997 A2 EP2046997 A2 EP 2046997A2 EP 07866600 A EP07866600 A EP 07866600A EP 07866600 A EP07866600 A EP 07866600A EP 2046997 A2 EP2046997 A2 EP 2046997A2
- Authority
- EP
- European Patent Office
- Prior art keywords
- gene
- microarray
- nucleic acid
- genes
- cardiomyopathy
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Withdrawn
Links
Classifications
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12Q—MEASURING OR TESTING PROCESSES INVOLVING ENZYMES, NUCLEIC ACIDS OR MICROORGANISMS; COMPOSITIONS OR TEST PAPERS THEREFOR; PROCESSES OF PREPARING SUCH COMPOSITIONS; CONDITION-RESPONSIVE CONTROL IN MICROBIOLOGICAL OR ENZYMOLOGICAL PROCESSES
- C12Q1/00—Measuring or testing processes involving enzymes, nucleic acids or microorganisms; Compositions therefor; Processes of preparing such compositions
- C12Q1/68—Measuring or testing processes involving enzymes, nucleic acids or microorganisms; Compositions therefor; Processes of preparing such compositions involving nucleic acids
- C12Q1/6876—Nucleic acid products used in the analysis of nucleic acids, e.g. primers or probes
- C12Q1/6883—Nucleic acid products used in the analysis of nucleic acids, e.g. primers or probes for diseases caused by alterations of genetic material
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12Q—MEASURING OR TESTING PROCESSES INVOLVING ENZYMES, NUCLEIC ACIDS OR MICROORGANISMS; COMPOSITIONS OR TEST PAPERS THEREFOR; PROCESSES OF PREPARING SUCH COMPOSITIONS; CONDITION-RESPONSIVE CONTROL IN MICROBIOLOGICAL OR ENZYMOLOGICAL PROCESSES
- C12Q2600/00—Oligonucleotides characterized by their use
- C12Q2600/158—Expression markers
Definitions
- the present invention is drawn to a convenient and highly effective microarray tool and method for diagnosing a cardiomyopathy state, especially dilated cardiomyopathy, in a subject, with a high degree of accuracy.
- DCM Dilated cardiomyopathy
- natriuretic peptides are the typical markers for diagnosing and managing heart failure. However, there is significant heterogeneity in expression levels of these natriuretic molecular disease markers that is not explained by left ventricular function alone.
- Transcriptional signature analysis is a powerful technique for identifying potential molecular targets that could ultimately become important for diagnosis and therapy of heart failure. Yet, differences in platform technologies, experimental design, and the biological heterogeneity associated with the use of human tissue samples are obstacles to the successful comparison and integration of results obtained by different microarray studies in heart failure. In addition, substantial regional variation in gene expression exists in mammalian myocardium including atrium, ventricle and septum or left and right side of the heart. See Barth et al., "Functional profiling of human atrial and ventricular gene expression," Pfl ⁇ gers Archiv.
- the present invention takes into account the various variables that can otherwise confound a cardiomyopathy diagnosis, by integrating independent microarray studies from large numbers of failing and non-failing hearts. Accordingly, the present invention provides a convenient and highly effective tool and method for diagnosing a cardiomyopathic state with accuracy.
- a general aspect of the present invention is the detection of particular nucleic acid or protein expression levels in a biological sample, which is useful for preparing an expression profile of that sample, wherein the profile is indicative of a healthy or diseased condition associated with that sample.
- the invention provides expression profiles in various conditions associated or symptomatic of cardiomyopathy. Hence, the invention produces expression profiles that are useful for distinguishing, for instance, between cardiomyopathic tissue and healthy tissue, failing and non-failing heart conditions, and ischemic and non-ischemic cardiomyopathy.
- an aspect of the invention entails comparing the expression profile from a biological sample from an individual, such as a human patient, with the expression profile of a healthy individual or the expression profile of an individual who does not have the cardiomyopathic condition that the tested individual is believed or suspected of having.
- nucleic acids each one of which shares sequence identity with either the sense or antisense strand sequence of a particular target gene or target nucleic acid implicated in a particular cardiomyopathic condition, that is useful for determining the expression levels of those target genes in any given sample.
- each nucleic acid comprises a sequence that is completely or partially complementary to a sequence of the target gene or target nucleic acid.
- a nucleic acid sequence of the collection is completely or partially complementary to 4, 5, 6, 7, 8, 9, 10, 1 1, 12, 13, 14, 15, 16, 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29, 30, 31 , 32, 33, 34, 35, 36, 37, 38, 39, 40, 41 , 42, 43, 44, 45, 46, 47, 48, 49, 50, 51, 52, 53, 54, 55, 56, 57, 58, 59, 60, 61, 62, 63, 64, 65, 66, 67, 68, 69, 70, 71, 72, 73, 74, 75, 76, 77, 78, 79, 80, 81, 82, 83, 84, 85, 86, 87, 88, 89, 90, 91, 92, 93, 94, 95, 96, 97, 98, 99, 100, 110, 120, 130, 140, 150, 160, 170, 180, 190, 200, 250, 300, 350, 400, 450,
- a nucleic acid sequence that is completely complementary is one that comprises a sequence that is identical in its complement to the equivalent sequence of the target. That is, the nucleic acid of the collection comprises a sequence that has no nucleotide mismatches compared to the corresponding target sequence.
- an array such as a micro- or nanoarrays, that comprises a collection of nucleic acid molecules, wherein each nucleic acid molecule comprises a nucleotide sequence that is complementary to a target sequence of one of the following genes listed below from SEQ ID NOs: 1-27.
- the collection of nucleic acids that are affixed or immobilized on the surface of the array may include nucleic acids that share sequence identity with some or all of these twenty-seven sequences.
- the array may comprise multiple nucleic acids each of which
- Synuclein alpha non A4 component of amyloid precursor
- SNCA non A4 component of amyloid precursor
- Asporin (SEQ ID NO: 2);
- Pleckstrin homology-like domain, family A, member 1 (PHLDAl) (SEQ ID NO: 4); (5) Frizzled-related protein (“FRZB”) (SEQ ID NO: 5);
- MYH6 Myosin heavy chain 6, cardiac muscle, alpha (cardiomyopathy, hypertrophic 1) ("MYH6”) (SEQ ID NO: 6);
- CCL2 Chemokine (C-C motif) ligand 2
- CCL2 Chemokine (C-C motif) ligand 2
- ODCl Ornithine decarboxylase 1
- Retinoic acid receptor responder tazarotene induced 1 (“RARRESl ”) (SEQ ID NO: 9);
- MYHlO Myosin heavy chain 10, non-muscle (MYHlO) (SEQ ID NO: 12);
- FCN3 Ficolin (collagen/ fibrinogen domain containing) 3 (Hakata antigen) (FCN3”) (SEQ ID NO: 13);
- S100A8 SEQ ID NO: 14
- CORIN Corin serine peptidase
- NPPA N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoeth
- PCOLCE2 Procollagen C-endopeptidase enhancer 2
- NPPB N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl)-2-aminoethyl-N-(2-aminoethyl) (SEQ ID NO: 18);
- ATF3 Activating transcription factor 3
- CTGF Connective tissue growth factor
- G0/G lswitch 2 (SEQ ID NO: 23);
- KLHL3 Kelch-like 3 (Drosophila)
- ZBTB16 Zinc finger and BTB domain containing 16
- AEBPl AE binding protein 1
- ETS variant gene 5 (ets-related molecule) (ETV5) (SEQ ID NO: 27).
- each of the nucleic acid molecules in the collection on the microarray comprises a different sequence as compared to each of the other nucleic acid molecules in the collection.
- each of the nucleic acid molecules in the collection comprises a complementary sequence to a target gene that is different from the complementary target gene sequence in each of the other nucleic acid molecules in the collection.
- the nucleic acid molecule is an oligonucleotide or a probe.
- the oligonucleotide or probe is labeled with a moiety to promote signal detection after the oligonucleotide or probe hybridizes to its corresponding target sequence.
- the collection comprises nucleic acid molecules that comprise complementary sequences to at least the following sequences: a myosin heavy polypeptide 10 (non-muscle) gene, a synuclein gene, an alpha putative lymphocyte G0/G1 switch gene gene, an ets variant gene 5, an AE binding protein 1 gene, a kelch-like 3 gene, a zinc finger and BTB domain containing 16 gene, and a procollagen C-endopeptidase enhancer 2 gene.
- Another aspect of the present invention is directed to a method for diagnosing a cardiomyopathic state in a subject comprising: (i) exposing the microarray of claim 1 to an isolated biological sample (ii) determining which nucleic acid molecules of the collection hybridize to its corresponding target gene sequence, and (iii) comparing that hybridization pattern to the hybridization pattern of a control for the same genes, wherein a difference between the two patterns is indicative that the individual has a cardiomyopathic disease.
- the biological sample is blood or a tissue biopsy.
- the tissue biopsy is a heart muscle biopsy.
- the cardiomyopathic disease is dilated cardiomyopathy or idiopathic dilated cardiomyopathy.
- Another aspect of the present invention is directed to an assay for diagnosing a cardiomyopathic state in a subject, comprising: (i) determining the gene expression levels of one or more of the genes or nucleic acids selected from the group consisting of:
- the gene or nucleic acid is in a biological sample isolated from the subject, and (ii) comparing the expression levels to a biological sample from a healthy individual, wherein a difference in the expression levels indicates that the subject has a cardiomyopathy disease.
- the cardiomyopathy disease is dilated cardiomyopathy or idiopathic dilated cardiomyopathy.
- the expression levels that are determined are the levels of either the respective gene transcripts or proteins.
- the expression levels of one or more, at least about two, at least about three, at least about four, at least about five, at least about six, at least about seven, at least about eight, at least about nine, at least about ten, at least about eleven, at least about twelve, at least about thirteen, at least about fourteen, at least about fifteen, at least about sixteen, at least about seventeen, at least about eighteen, at least about nineteen, at least about twenty, at least about twenty one, at least about twenty two, at least about twenty three, at least about twenty-four, at least about twenty five, at least about twenty six, or about twenty seven genes are determined in the subject's biological sample.
- the expression levels of all 27 genes are determined in the subject's biological sample and compared to the expression levels of the same 27 genes in the healthy subject's biological sample.
- the biological sample is blood or a tissue biopsy.
- the tissue biopsy is cardiac muscle.
- the step of determining gene expression levels is performed by at least one of a method selected from the group consisting of microarray, PCR, RT-PCR, TaqMan RT-PCR, Northern Blot, Western Blot, antibody detection, ELISA, or any combination thereof.
- the present invention also encompasses the diagnosis of other cardiomyopathies using the compositions and methods described herein.
- the present invention encompasses the diagnosis or prognosis of ischemic cardiomyopathy (ICM) using the genes, detection methods, and assays disclosed herein.
- ICM ischemic cardiomyopathy
- Figure 1 Shows a functional analysis based on Gene Ontology for selected gene classes comparing up-regulated genes in DCM (yellow bars) and down- regulated (green bars) genes in DCM of Dataset A (open bars) and B (striped bars). Differences in gene classes marked by an asterisk were statistically significant (Fisher's exact test, p ⁇ 0.05) according to "FatiGO" (20).
- FIG. 2 Shows a Prediction analysis for microarrays (“PAM”) classification.
- PAM microarrays
- the first step PAM classification was applied to all four datasets separately. Very low misclassification rates were found in datasets A, B and D for the classification of non-failing (NF) vs. DCM samples, whereas the classification algorithm did not show any power in Dataset C.
- the second step the procedure was repeated with the smallest gene signature obtained from Dataset B, now achieving more than 90% accuracy for classifying DCM and NF samples across all studies, including Dataset C.
- Figure 3 Shows mean expression ⁇ S. E. M of pro-BNP in NF (black bars) and DCM samples (red bars) in datasets A-D. Statistical comparison was carried out by Student's t-test.
- the present invention provides methods and compositions for identifying genetic biomarkers that are useful for classifying and diagnosing cardiomyopathy disease states.
- the present invention identifies genes and biological processes that are either known to be involved with, or are implicated in, dilated cardiomyopathy (DCM).
- DCM dilated cardiomyopathy
- NF non-failing
- protein is understood to include the terms “polypeptide” and “peptide” (which, at times, may be used interchangeably herein) within its meaning.
- Recombinant proteins or polypeptides refer to proteins or polypeptides produced by recombinant DNA techniques, i.e., produced from cells, microbial or mammalian, transformed by an exogenous recombinant DNA expression construct encoding the desired protein or polypeptide. Proteins or polypeptides expressed in most bacterial cultures will typically be free of glycan. Proteins or polypeptides expressed in yeast may have a glycosylation pattern different from that expressed in mammalian cells.
- a DNA or polynucleotide “coding sequence” is a DNA or polynucleotide sequence that is transcribed into mRNA and translated into a polypeptide in a host cell when placed under the control of appropriate regulatory sequences. The boundaries of the coding sequence are the start codon at the 5' N-terminus and the translation stop codon at the 3' C-terminus.
- a coding sequence can include prokaryotic sequences, cDNA from eukaryotic mRNA, genomic DNA sequences from eukaryotic DNA, and synthetic DNA sequences. A transcription termination sequence will usually be located 3' to the coding sequence.
- DNA or polynucleotide sequence is a heteropolymer of deoxyribonucleotides (bases adenine, guanine, thymine, cytosine). DNA or polynucleotide sequences can be assembled from synthetic cDNA-derived DNA fragments and short oligonucleotide linkers.
- analogs when referring to the nucleic acids of this invention mean analogs, fragments, derivatives, and variants of such nucleotides having, for example, at least about 60% sequence identity, at least about 70% sequence identity, at least about 80% sequence identity, at least about 90% sequence identity, at least about 91% sequence identity, at least about 92% sequence identity, at least about 93% sequence identity, at least about 94% sequence identity, at least about 95% sequence identity, at least about 96% sequence identity, at least about 97% sequence identity, at least about 98% sequence identity, at least about 99% sequence identity, or at least about 100% sequence identity to the native or naturally occurring nucleic acid, as described herein.
- Similarity between two polynucleotides is determined by comparing the amino acid sequence corresponding to each polynucleotide to the amino acid sequence corresponding to the second polynucleotide.
- An amino acid of one amino acid sequene is similar to the corresponding amino acid of a second amino acid sequence if it is identical or a conservative amino acid substitution.
- Conservative substitutions include those described in Dayhoff, M.O., ed., The Atlas of Protein Sequence and Structure 5, National Biomedical Research Foundation, Washington, D. C. (1978), and in Argos, P. (1989) EMBOJ. 8:779-785.
- amino acids belonging to one of the following groups represent conservative changes or substitutions:
- -Ala Pro, GIy, GIn, Asn, Ser, Thr: -Cys, Ser, Tyr, Thr; -VaI, lie, Leu, Met, Ala, Phe; -Lys, Arg, His; -Phe, Tyr, Trp, His; and
- “Mammal” includes humans and domesticated animals, such as cats, dogs, swine, cattle, sheep, goats, horses, rabbits, and the like.
- One aspect of the present invention is directed to a collection of 27 genes that represent useful targets for successfully distinguishing non-failing heart samples from dilated cardiomyopathy heart samples with over 90% accuracy. That is, identifying the presence or expression level of one or more or all of these 27 genes can be indicative of a cardiomyopathy disease phenotype.
- expression levels of any combination of the 27 genes can be measured, such as the expression levels of one or more, at least about two, at least about three, at least about four, at least about five, at least about six, at least about seven, at least about eight, at least about nine, at least about ten, at least about eleven, at least about twelve, at least about thirteen, at least about fourteen, at least about fifteen, at least about sixteen, at least about seventeen, at least about eighteen, at least about nineteen, at least about twenty, at least about twenty one, at least about twenty two, at least about twenty three, at least about twenty-four, at least about twenty five, at least about twenty six, or about twenty seven genes can be determined in the subject's biological sample.
- the expression levels of all 27 genes are determined in the subject's biological sample and compared to the expression levels of the same 27 genes in the healthy subject's biological sample.
- the group of 27 genes includes the following, which are categorized based on existing knowledge of their individual involvement in DCM.
- BNP natriuretic peptide precursor B
- NPPA natriuretic peptide precursor A
- CORRIN cardiomyopathy corin
- C-C motif chemokine (C-C motif) ligand 2 (CCL2) myosin, heavy polypeptide 6, cardiac, alpha (MYH6)
- ATF3 activating transcription factor 3
- CGF connective tissue growth factor
- FRZB frizzled-related protein
- PLDAl pleckstrin homology-like domain, family Al
- ODC l retinoic acid receptor responder 1
- AXT2L1 complement factor H-related 3
- SPOCK osteonectin proteoglycan
- AEBPl AE binding protein 1
- KLHL3 kelch-like 3
- ZBTB 16 BTB domain containing 16
- procollagen C-endopeptidase enhancer 2 PCOLCE2
- the present invention provides an arrangement of markers for these 27 genes, and the markers can collectively be used to determine whether a particular biological sample is indicative of cardiomyopathic heart disease.
- the abbreviations are used in Tables elsewhere in this application.
- the present invention encompasses the use of markers to all 27 genes, as well as the use of markers to subsets of the 27 genes. For instance, markers to one or more of the Group 4 genes may be arranged or combined alongside one or more markers of the Group 1 genes. Hence, the present invention contemplates various combinations of gene markers based on the four groups outlined above, such as one or more markers of each Group according to the following exemplary combinations:
- Microarrays are useful for identifying cancer-specific genes and inflammatory-specific genes. See DeRisi et a , Nat. Genet. 14(4):457-60 (1996); and Heller et al. , Proc. NaU. Acad. Sci. USA 94(6):2150-55 ( 1997).
- a microarray may typically be composed of a number of unique, single- stranded nucleic acid sequences, usually either synthetic antisense oligonucleotides or fragments of cDNAs, fixed to a solid support.
- microarray connotes an array of polynucleotides or oligonucleotides that are placed, arranged, or otherwise affixed on to a substrate, such as paper, nylon or other type of membrane, filter, chip, glass slide, or any other such suitable solid support.
- a substrate such as paper, nylon or other type of membrane, filter, chip, glass slide, or any other such suitable solid support.
- one microarray may comprise one or more combinations of nucleic acid sequences that share sequence identity with, or share sequence identity with the complement of, one or more of the 27 genes denoted above characterized into the four groups.
- One embodiment of the invention uses solid support-based oligonucleotide hybridization methods to detect gene expression.
- Solid support-based methods suitable for practicing the present invention are widely known and are described, for example, in PCT application WO 95/ 1 1755; Huber et al., Anal. Biochem. 299: 24 (2001); Meiyanto et al, Biotechniques. 31 : 406 (2001); Relogio et al, Nucleic Adds Res. 30:e51 (2002).
- Any solid surface to which oligonucleotides can be bound, covalently or non-covalently can be used.
- Such solid supports include, but are not limited to, filters, polyvinyl chloride dishes, silicon or glass based chips.
- the nucleic acid molecule can be directly bound to the solid support or bound through a linker arm, which is typically positioned between the nucleic acid sequence and the solid support.
- a linker arm that increases the distance between the nucleic acid molecule and the substrate can increase hybridization efficiency.
- the solid support is coated with a polymeric layer that provides linker arms with a lot of reactive ends/sites.
- a common example of this type is glass slides coated with polylysine (see, U.S. Patent No. 5667976), which are commercially available.
- the linker arm may be synthesized as part of or conjugated to the nucleic acid molecule, and then this complex is bonded to the solid support.
- the streptavidin-biotinylated reaction is stable enough to withstand stringent washing conditions and is sufficiently stable that it is not cleaved by laser pulses used in some detection systems, such as matrix-assisted laser desorption/ ionization time of flight (MALDI-TOF) mass spectrometry. Therefore, streptavidin may be covalently attached to a solid support, and the nucleic acid molecule is labeled with a biotin group (or vice versa).
- MALDI-TOF matrix-assisted laser desorption/ ionization time of flight
- biotinylated nucleic acid molecule effectively sticks wherever it is placed on the streptavidin-covered support surface.
- an amino- coated silicon wafer is reacted with the n-hydroxysuccinimido-ester of biotin and complexed with streptavidin.
- Biotinylated oligonucleotides are bound to the surface at a concentration of about 20 fmol DNA per mm 2 .
- the support is coated with hydraztde groups, then treated with carbodiimide.
- Carboxy-modified nucleic acid molecules are then coupled to the treated support.
- Epoxide-based chemistries are also being employed with amine modified oligonucleotides.
- Other chemistries for coupling nucleic acid molecules to solid substrates are known to those of skill in the art.
- the nucleic acid molecules are typically delivered to the substrate material. Because of the miniaturization of the arrays, delivery techniques should be capable of positioning very small amounts of liquids (e.g., less than 1 nanoliter) in very small regions (e.g., 100 m diameter dots), very close to one another (e.g., 250 m separation) and amenable to automation. Several techniques and apparatus are available to achieve such delivery. Among these are mechanical mechanisms [e.g., arrayers from GeneticMicroSystems, MA, USA) and ink-jet technology. Very fine pipettes may also be used. Other formats are also suitable within the context of this invention. For example, a 96-well format with fixation of the nucleic acids to a nitrocellulose or nylon membrane may also be employed.
- the probes After the nucleic acid molecules have been bound to the solid support, it is often useful to block reactive sites on the solid support that are not consumed in binding to the nucleic acid molecule. Otherwise, the probes will, to some extent, bind directly to the solid support itself, giving rise to so-called non-specific binding. Non-specific binding can sometimes hinder the ability to detect low levels of specific binding.
- a variety of effective blocking agents e.g., milk powder, serum albumin or other proteins with free amine groups, polyvinylpyrrolidine
- the choice depends at least in part upon the binding chemistry.
- An oligonucleotide may preferably be about 6 to about 60 nucleotides in length, or any length in between these two parameters. In other embodiments of the invention, an oligonucleotide may be about 15 to about 30 nucleotides in length, or about 20 to about 25 nucleotides in length. For a certain type of microarray, it may be preferable to use oligonucleotides which are about 7 to about 10 nucleotides in length.
- the microarray may comprise oligonucleotides which cover the known 5', or 3', sequence, sequential oligonucleotides which cover the full length sequence; or unique oligonucleotides selected from particular areas along the length of the sequence.
- Polynucleotides used in the microarray may be oligonucleotides that are specific to a gene or genes of interest in which at least a fragment of the sequence is known or that are specific to one or more unidentified cDNAs which are common to a particular cell type, development or disease state.
- oligonucleotide arrays i.e. microarrays, to simultaneously observe the expression of a number of genes or gene products.
- Oligonucleotide arrays comprise two or more oligonucleotide probes provided on a solid support, wherein each probe occupies a unique location on the support.
- the location of each probe may be predetermined, such that detection of a detectable signal at a given location is indicative of hybridization to an oligonucleotide probe of a known identity.
- Each predetermined location can contain more than one molecule of a probe, but each molecule within the predetermined location has an identical sequence. Such predetermined locations are termed features.
- each oligonucleotide is located at a unique position on an array at least 2, at least 3, at least 4, at least 5, at least 6, or at least 10 times.
- Oligonucleotide probe arrays for detecting gene expression can be made and used according to conventional techniques described, for example, in Lockhart et al, Natl Biotech. 14: 1675 (1996), McGaIl et al, Proc. Natl Acad. Sd. USA 93: 13555 (1996), and Hughes et al, Nature Biotechnol. 19:342 (2001).
- a variety of oligonucleotide array designs is suitable for the practice of this invention.
- the one or more oligonucleotides include a plurality of oligonucleotides that each hybridize to a different gene expressed in a particular tissue type.
- oligonucleotides of the present invention hybridize to nucleic acid sequences of any of the following genes: Synuclein alpha (non A4 component of amyloid precursor) ("SNCA"), Asporin ("ASPN"), Secreted frizzled- related protein 4 ("SFRP4"), Pleckstrin homology-like domain, family A, member 1 ("PHLDAl”), Frizzled- related protein (“FRZB”), Myosin heavy chain 6, cardiac muscle, alpha (cardiomyopathy, hypertrophic 1) ("MYH6"), Chemokine (C-C motif) ligand 2 (“CCL2”), Ornithine decarboxylase 1 (“ODCl”), Retinoic acid receptor responder (tazarotene induced) 1 (“RARRESl”), Complement factor H
- a detectable molecule also referred to herein as a label
- a label will be incorporated or added to an array's nucleic acid sequences.
- Many types of molecules can be used within the context of this invention. Such molecules include, but are not limited to, fluorochromes, chemiluminescent molecules, chromogenic molecul es, radioactive molecules, mass spectometry tags, proteins, and the like. Other labels will be readily apparent to one skilled in the art. Indirect detection can also be used within the context of this invention. Proteins and other molecules are available that will bind to double-stranded DNA but not to single-stranded DNA. Thus, hybridization can be measured.
- a nucleic acid sample obtained from an individual can be amplified and, optionally labeled with a detectable label.
- Any method of nucleic acid amplification and any detectable label suitable for such purpose can be used.
- amplification reactions can be performed using, e.g. Ambion's MessageAmp, which creates "antisense” RNA or "aRNA" (complementary in nucleic acid sequence to the RNA extracted from the sample tissue).
- the RNA can optionally be labeled using CyDye fluorescent labels.
- CyDye fluorescent labels are coupled to the aaUTPs in a non-enzymatic reaction.
- labeled amplified antisense RNAs are precipitated and washed with appropriate buffer, and then assayed for purity.
- purity can be assay using a NanoDrop spectrophotometer.
- the nucleic acid sample is then contacted with an oligonucleotide array having, attached to a solid substrate (a "microarray slide"), oligonucleotide sample probes capable of hybridizing to nucleic acids of interest which may be present in the sample.
- the step of contacting is performed under conditions where hybridization can occur between the nucleic acids of interest and the oligonucleotide probes present on the array.
- the array is then washed to remove non-specifically bound nucleic acids and the signals from the labeled molecules that remain hybridized to oligonucleotide probes on the solid substrate are detected.
- the step of detection can be accomplished using any method appropriate to the type of label used.
- the step of detecting can accomplished using a laser scanner and detector.
- on can use and Axon scanner which optionally uses GenePix Pro software to analyze the position of the signal on the microarray slide. Data from one or more microarray slides can analyzed by any appropriate method known in the art.
- Oligonucleotide probes used in the methods of the present invention can be generated using PCR.
- PCR primers used in generating the probes are chosen, for example, based on the sequences of SEQ ID NOs: 1-27.
- oligonucleotide control probes also are used.
- Exemplary control probes can fall into at least one of three categories referred to herein as (1) normalization controls, (2) expression level controls and (3) negative controls.
- one or more of these control probes may be provided on the array with the inventive cell cycle gene-related oligonucleotides.
- Normalization controls correct for dye biases, tissue biases, dust, slide irregularities, malformed slide spots, etc.
- Normalization controls are oligonucleotide or other nucleic acid probes that are complementary to labeled reference oligonucleotides or other nucleic acid sequences that are added to the nucleic acid sample to be screened.
- the signals obtained from the normalization controls, after hybridization provide a control for variations in hybridization conditions, label intensity, reading efficiency and other factors that can cause the signal of a perfect hybridization to vary between arrays.
- signals ⁇ e.g., fluorescence intensity or radioactivity) read from all other probes used in the method are divided by the signal from the control probes, thereby normalizing the measurements.
- Virtually any probe can serve as a normalization control. Hybridization efficiency varies, however, with base composition and probe length. Preferred normalization probes are selected to reflect the average length of the other probes being used, but they also can be selected to cover a range of lengths. Further, the normalization control(s) can be selected to reflect the average base composition of the other probes being used. In one embodiment, only one or a few normalization probes are used, and they are selected such that they hybridize well ⁇ i.e., without forming secondary structures) and do not match any test probes. In one embodiment, the normalization controls are mammalian genes.
- Expression level controls probes hybridize specifically with constitutively expressed genes present in the biological sample. Virtually any constitutively expressed gene provides a suitable target for expression level control probes. Typically, expression level control probes have sequences complementary to subsequences of constitutively expressed "housekeeping genes" including, but not limited to certain photosynthesis genes.
- Negative control probes are not complementary to any of the test oligonucleotides [i.e., the inventive cell cycle gene-related oligonucleotides), normalization controls, or expression controls.
- the negative control is a mammalian gene which is not complementary to any other sequence in the sample.
- background and background signal intensity refer to hybridization signals resulting from non-specific binding or other interactions between the labeled target nucleic acids (i.e., mRNA present in the biological sample) and components of the oligonucleotide array. Background signals also can be produced by intrinsic fluorescence of the array components themselves. A single background signal can be calculated for the entire array, or a different background signal can be calculated for each target nucleic acid. In a one embodiment, background is calculated as the average hybridization signal intensity for the lowest 5 to 10 percent of the oligonucleotide probes being used, or, where a different background signal is calculated for each target gene, for the lowest 5 to 10 percent of the probes for each gene.
- background can be calculated as the average hybridization signal intensity produced by hybridization to probes that are not complementary to any sequence found in the sample [e.g., probes directed to nucleic acids of the opposite sense or to genes not found in the sample).
- background can be calculated as the average signal intensity produced by regions of the array that lack any oligonucleotides probes at all.
- the nucleic acid molecules are directly or indirectly coupled to an enzyme.
- a chromogenic substrate is applied and the colored product is detected by a camera, such as a charge-coupled camera.
- enzymes include alkaline phosphatase, horseradish peroxidase and the like.
- the invention also provides methods of labeling nucleic acid molecules with cleavable mass spectrometry tags (CMST) (see for example, U.S. Patent No: 60279890). After an assay is complete, and the uniquely CMST-labeled probes are distributed across the array, a laser beam is sequentially directed to each member of the array.
- CMST cleavable mass spectrometry tags
- the light from the laser beam both cleaves the unique tag from the tag-nucleic acid molecule conjugate and volatilizes it.
- the volatilized tag is directed into a mass spectrometer. Based on the mass spectrum of the tag and knowledge of how the tagged nucleotides were prepared, one can unambiguously identify the nucleic acid molecules to which the tag was attached (see, e.g., WO9905319).
- the nucleic acids can be labeled readily by any of a variety of techniques.
- the nucleic acids can be labeled during the reaction by incorporation of a labeled dNTP or use of labeled amplification primer.
- the amplification primers include a promoter for an RNA polymerase, a post-reaction labeling can be achieved by synthesizing RNA in the presence of labeled NTPs.
- Amplified fragments that were unlabeled during amplification or unamplified nucleic acid molecules can be labeled by one of a number of end labeling techniques or by a transcription method, such as nick- translation, random-primed DNA synthesis.
- PCR-based methods are used to detect gene expression. These methods include reverse-transcriptase-mediated polymerase chain reaction (RT-PCR) including real-time and endpoint quantitative reverse- transcriptase-mediated polymerase chain reaction (Q-RTPCR). These methods are well known in the art. For example, methods of quantitative PCR can be carried out using kits and methods that are commercially available from, for example, Applied BioSystems and Stratagene®. See also Kochanowski, QUANTITATIVE PCR PROTOCOLS (Humana Press, 1999); Innis et al., supra.; Vandesompele et al., Genome Biol. 3: RESEARCH0034 (2002); Stein, CeH MoI. Life Sd.
- RT-PCR reverse-transcriptase-mediated polymerase chain reaction
- Q-RTPCR quantitative reverse-transcriptase-mediated polymerase chain reaction
- Q-RTPCR relies on detection of a fluorescent signal produced proportionally during amplification of a PCR product. See Innis et al, supra.
- this technique employs PCR oligonucleotide primers, typically 15-30 bases long, that hybridize to opposite strands and regions flanking the DNA region of interest.
- a probe ⁇ e.g., TaqMan®, Applied Biosystems is designed to hybridize to the target sequence between the forward and reverse primers traditionally used in the PCR technique.
- the probe is labeled at the 5' end with a reporter fluorophore, such as 6-carboxyfluorescein (6-FAM) and a quencher fluorophore like 6-carboxy-tetramethyl-rhodamine (TAMRA).
- a reporter fluorophore such as 6-carboxyfluorescein (6-FAM) and a quencher fluorophore like 6-carboxy-tetramethyl-rhodamine (TAMRA).
- 6-FAM 6-carboxyfluorescein
- TAMRA 6-carboxy-tetramethyl-rhodamine
- the forward and reverse amplification primers and internal hybridization probe is designed to hybridize specifically and uniquely with one nucleotide derived from the transcript of a target gene.
- the selection criteria for primer and probe sequences incorporates constraints regarding nucleotide content and size to accommodate TaqMan® requirements.
- SYBR Green® can be used as a probe-less Q-RTPCR alternative to the Taqman®-type assay, discussed above.
- a device measures changes in fluorescence emission intensity during PCR amplification. The measurement is done in "real time," that is, as the amplification product accumulates in the reaction. Other methods can be used to measure changes in fluorescence resulting from probe digestion. For example, fluorescence polarization can distinguish between large and small molecules based on molecular tumbling (see U.S. patent No. 5,593,867).
- stringent hybridization and washing conditions are useful for nucleic acid molecules over about 500 bp.
- Stringent hybridization conditions include a solution comprising about 1 M Na+ at 25° to 30 0 C below the Tm; e.g., 5 x SSPE, 0.5% SDS, at 65DC; see, Ausubel, et al, Current Protocols in Molecular Biology, Greene Publishing, 1995; Sambrook et al.. Molecular Cloning: A Laboratory Manual, Cold Spring Harbor Press, 1989).
- Tm is dependent on both the G+C content and the concentration of Na+.
- Tm 81.5 + 0.41(%(G+C)) - log 10 [Na+]. Washing conditions are generally performed at least at equivalent stringency conditions as the hybridization. If the background levels are high, washing may be performed at higher stringency, such as around 15°C below the Tm.
- Low stringency hybridizations are performed at conditions approximately 40 0 C below Tm, and are used for short fragments, e.g., less than about 500 bp. For fragments between about 100 and 500 bp, the Tm decreases about 1.5°C for every fewer 50 bp than 500. For very small fragments, e.g., less than about 50 bp, a formula for calculating Tm is 2°C for each AT pair and 4°C for each GC pair. Very high stringency hybridizations are performed at conditions approximately 10°C below Tm.
- Hybridization conditions are tailored to the length and GC content of the oligonucleotide. Suitable hybridization conditions may be found in Sambrook et al., supra, Ausubel et al., supra, and furthermore hybridization solutions may contain additives such as tetramethylammonium chloride or other chaotropic reagents or hybotropic reagents to increase specificity of hybridization (see for example, PCT/US97/ 17413).
- Hybridization may be detected in a variety of ways and with a variety of equipment.
- the methods may be categorized as those that rely upon detectable molecules incorporated into the diversity panels and those that rely upon measurable properties of double-stranded nucleic acids (i.e., hybridized nucleic acids) that distinguish them from single-stranded nucleic acids (i.e., unhybridized nucleic acids).
- the latter category of methods includes intercalation of dyes, such as ethidium bromide, into double-stranded nucleic acids, differential absorbance properties of double and single stranded nucleic acids, binding of proteins that preferentially bind double-stranded nucleic acids, and the like.
- a radioactive label is used, autoradiography or storage phosphor screens (Phosphorlmager) are common methods of detection.
- An alternative detection system that can be used with radioactive, fluorescent or chemiluminescent labels is a CCD integrated silicon wafer.
- a charge-coupled device designed to detect high energy beta particles or photons, is placed in direct contact with a silicon support for an array. Upon binding of the sample to the immobilized nucleic acids, a radioisotope decay product or photon is generated. Electron-hole pairs are generated in the silicon and then electrons are collected by the CCD.
- An alternative detection system for fluorescent molecules is a lens based camera detecting one or more fluorescent labels.
- these cameras include epifluorescent microscopes, confocal microscopes, and charge-coupled cameras.
- a laser excites a fluorescent label, the emitted light is collected through a bandpass filter, and the signal is detected by a photomultiplier tube that has electronics for counting photons.
- labels are also amenable to use with either a lens-based camera or a CCD.
- chemiluminescent labels or chromogenic substrates can be detected with a lens-based charge-coupled camera.
- the label is a cleavable mass-spectrometry tag.
- Such labels are then detected using a mass-spectrometer.
- Many detection systems are commercially available (e.g., Affymetrix, Santa Clara, CA).
- Affymetrix e.g., Affymetrix, Santa Clara, CA.
- One skilled in the art is able to choose an appropriate detection means and equipment for the label used.
- Patterns of hybridization can be expressed as presence or absence of hybridization, the degree of hybridization, or some combination of these.
- the simplest analysis is performed by determining the presence or absence of hybridization.
- the complexity of the genome of the organism to be genotyped is greater than the complexity of the genome(s) represented on the array, the absence of hybridization conclusively signifies a polymorphism.
- the complexity is less than on the array, the absence of hybridization can signify either a polymorphism or a lack of representation of those sequences in the probing diversity panel.
- the presence of hybridization does not necessarily signify the absence of a polymorphism under either scenario.
- the pattern of hybridization is informative.
- each addressable area is queried for hybridization using a method appropriate to the label. For example, when fluorescent labels are used, such as Cy3 and Cy5, both green and red signals are assayed.
- a method appropriate to the label For example, when fluorescent labels are used, such as Cy3 and Cy5, both green and red signals are assayed.
- positive and negative controls are included on the array, signals are compared to the controls and each addressable area is assigned a value, e.g., 1 for detectable hybridization and 0 for no detectable hybridization. In general, a value of 1 is assigned for detection over a threshold level and 0 assigned for detection under a threshold level. It will be appreciated by those skilled in the art that detection of polymorphisms is based primarily on finding a binary distribution of signal values for any particular array feature when hybridized with multiple diversity panels.
- the panels are the same as those used to create the diversity array (see Example 5).
- a diversity panel is generated from a heterozygote for a polymorphism, one will then detect a trimodal distribution.
- two threshold values are calculated, the: first threshold separates the "0" cluster (lack of hybridization) from the "0/ 1° cluster (heterozygote) and the second threshold separates the "0/ 1" cluster from the "1" cluster (hybridization present).
- Conventional statistical methods may be used to determine the threshold levels.
- the genotype of the organism may then be expressed as a value for each addressable area.
- the addressable array is a 96-spot format (a grid of 8 rows (A-G) x 12 columns ( 1- 12))
- the value for hybridization is 1 and no detectable hybridization is 0, then it is possible to visualize the individual's expression profile on a two-dimensional grid reflecting those 1/0 detections.
- relative values are assigned to each addressable location. The relative values will generally be normalized to controls. All data can be collected into database formats to facilitate comparisons as well as perform further analyses, such as construction of genotype trees.
- oligonucleotides of some or all of those 27 genes disclosed herein may be used as components of a microarray.
- the present invention is not limited to markers, such as oligonucleotides and nucleic acid probes, that specifically target the denoted 27 genes.
- Nucleic acid probes or oligonucleotides also may be designed to target isoforms and homologs of any one of the 27 genes.
- Proteins also can be observed by any means known in the art, including immunological methods, enzyme assays and protein array/ proteomics techniques, for determining the expression profile of a sample instead of, or in addition to, determining expression levels by detecting nucleic acid transcripts. Measurement of the translational state of proteins can be performed according to several protein methods. For example, whole genome monitoring of protein — the "proteome” — can be carried out by constructing a microarray in which binding sites comprise immobilized, preferably monoclonal, antibodies specific to a plurality of proteins having an amino acid sequence of any of SEQ ID NOs: 236-470 and 718-737 or proteins encoded by the genes of SEQ ID NOs: 1-235 and 698-717 or conservative variants thereof.
- proteins can be separated by two-dimensional gel electrophoresis systems.
- Two-dimensional gel electrophoresis is well-known in the art and typically involves isoelectric focusing along a first dimension followed by SDS-PAGE electrophoresis along a second dimension. See, e.g., Hames et al, , GEL ELECTROPHORESIS OF PROTEINS: A PRACTICAL APPROACH (IRL Press, 1990).
- the resulting electropherograms can be analyzed by numerous techniques, including mass spectrometric techniques, western blotting and immunoblot analysis using polyclonal and monoclonal antibodies, and internal and N-terminal micro- sequencing.
- a nucleic acid marker to, say, the corin gene can be a nucleic acid sequence, such as an oligonucleotide or probe, that is complementary to a corin- specific gene sequence; so that, when it is affixed to a substrate surface, the probe will anneal to or hybridize to a corin gene nucleic acid transcript, be it genomic DNA, cDNA, or RNA.
- an oligonucleotide may be designed to any and all of the 27 genes denoted above.
- the present invention contemplates the presence of multiple oligonucleotides on a particular substrate that is designed to anneal to or hybridize to the same gene.
- the "pairs" will be identical, except for one nucleotide which preferably is located in the center of the sequence.
- the second oligonucleotide in the pair serves as a control.
- the number of oligonucleotide pairs may range from two to one million.
- the oligomers are synthesized at designated areas on a substrate using a light-directed chemical process.
- the substrate may be paper, nylon or other type of membrane, filter, chip, glass slide or any other suitable solid support.
- the gene of interest may be examined using a computer algorithm which starts at the 5' or more preferably at the 3' end of the nucleotide sequence.
- the algorithm identifies oligomers of defined length that are preferably unique to the gene, have a GC content within a range suitable for hybridization, and lack predicted secondary structure that may interfere with hybridization.
- an oligonucleotide may be synthesized on the surface of the substrate by using a chemical coupling procedure and an ink jet application apparatus, as described in PCT application WO95/2511 16 (Baldeschweiler et al.) which is incorporated herein in its entirety by reference.
- a "gridded" array analogous to a dot or slot blot may be used to arrange and link cDNA fragments or oligonucleotides to the surface of a substrate using a vacuum system, thermal, UV, mechanical or chemical bonding procedures.
- An array such as those described above, may be produced by hand or by using available devices (slot blot or dot blot apparatus), materials (any suitable solid support), and machines (including robotic instruments).
- oligonucleotides, probes, or pieces of target-specific nucleic acids may be labeled in such fashion that a detectable signal is generated from their annealing or hybridizing to the target nucleic acid.
- a microarray of the present invention is placed into contact with a biological sample, such as blood, urine, saliva, phlegm, gastric juices, cultured cells, tissue biopsies, or other tissue preparations; and a detection system may then be used to measure the absence, presence, and/or amount of hybridization for all of the distinct sequences simultaneously.
- a biological sample such as blood, urine, saliva, phlegm, gastric juices, cultured cells, tissue biopsies, or other tissue preparations.
- the present invention contemplates a microarray of at least about one, at least about two, or any number in between about two and up to 27, or all of SEQ ID NOs: 1-27 genes denoted above in Groups 1-4, which can be used as described herein to evaluate a biological sample from a subject to determine whether that subject is symptomatic or is afflicted with a cardiomyopathic disease, such as dilated cardiomyopathy.
- a cardiomyopathic disease such as dilated cardiomyopathy.
- the present invention is not limited to the use of a microarray assay, however, for diagnosing a cardiomyopathic disease via gene expression level analysis.
- the present invention also contemplates the use of techniques such as polymerase chain reaction (PCR), quantitative-PCR and Real Time PCR, Quantitative Competitive Reverse Transcription-PCR and Real Time Detection 5'-Nuclease-PCR.
- PCR polymerase chain reaction
- quantitative-PCR quantitative-PCR
- Real Time PCR Quantitative Competitive Reverse Transcription-PCR
- Real Time Detection 5'-Nuclease-PCR Real Time Detection 5'-Nuclease-PCR.
- TaqMan RT-PCR also is known as TaqMan RT-PCR.
- TaqMan RT-PCR is useful for correlating the concentration of a protein in a sample tissue to its mRNA expression. See Hirayama et al., "Concentrations of Thrombopoietin in Bone Marrow in Normal Subjects and in Patients with Idiopathic Thrombocytopenic Purpura, Aplastic Anemia, and Essential Thrombocythemia Correlate With Its mRNA Expression of Bone Marrow Stromal Cells," Blood, 92(1): 46 52, 1998.
- This quantitative replicative method relies on the presence of a 5'-nuclease assay in the RT-PCR reactions, wherein a probe specific for the target protein, contains a fluorescent moiety such as 6-carboxyfluorescein (FAM) on its 5'-end, and a phosphate-capped quencher fluor moiety such as 6-carboxytetramethylfluorescein (TAMRA) on its 3'-end.
- FAM 6-carboxyfluorescein
- TAMRA 6-carboxytetramethylfluorescein
- TaqMan RT-PCR also may be coupled with an ABI Prism 7700 Sequence Detection System, or Competitive PCR for quantification of DNA. See Desjardin et al., "Comparison of the ABI 7700 System (TaqMan) and Competitive PCR for Quantification of IS61 10 DNA in Sputum During Treatment of Tuberculosis," J. Clin. Microbiol., 36(7): 1964- 1976, 1998.
- Another assay contemplated by the present invention is an ELISA assay to detect protein products of one or more of the denoted 27 genes from a subject's biological sample.
- ELISA assays and associated plate-reader apparatus are well known to those skilled in the art.
- PAM Prediction analysis for microarrays
- the PAM method Based on the smallest gene set for classification from Dataset B, the PAM method classified all four human heart failure datasets with low misclassification rates. Notably, despite the large variation of gene expression values for single genes, the classifier as a whole is highly valuable to distinguish DCM and NF samples. These results support the usefulness of this molecular approach for diagnostic applications. In addition, the classificator gene set based on DCM and NF hearts also achieved a similarly high accuracy of classification in ICM like in DCM samples, suggesting that this gene set could be representative of molecular changes of heart failure in general.
- the classificator based on Dataset B performed as well in Datasets A and D as if one used classificators generated from these two datasets alone.
- the gene signature from Dataset B was also able to accurately discriminate NF and DCM samples in Dataset C.
- differences in gene expression were greater between left ventricular assist-device (LVAD) and non- LVAD hearts in the DCM group than between DCM and NF samples (13). This peculiarity might impede the PAM approach for identifying a useful classifier between DCM and NF within this Dataset itself.
- the classifier gene signature can be grouped into different functional sets with respect to the pathogenesis of DCM.
- Up-regulation of the cardiomyopathy markers pro-ANP and pro-BNP is well established in heart failure (23) and mediated by neurohormonal dysregulation (27).
- Activation of pro-f ⁇ brotic stress hormone pathways lead to prominent structural remodeling in DCM, exemplified by deregulation of genes coding for sarcomer structure and extracellular matrix proteins like myosin 6 and 10, asporin, procollagen C-endopeptidase enhancer 2 (PCOLCE2), kelch-like 3 (KLHL3) and AE binding protein 1 (AEBPl).
- PCOLCE2 procollagen C-endopeptidase enhancer 2
- KLHL3 kelch-like 3
- AEBPl AE binding protein 1
- transcripts of this set of classifier genes including the transcription factor ZBTB 16 (28), the connective tissue growth factor CTGF (29) and the chemokine CCL2 (30), characterize important targets of the renin-angiotensin system in failing myocardium, as they all have been shown to be induced by angiotensin-II.
- transcripts of this classifier gene set belong to anti-apoptotic (PHLDAl, SNCA, CCL2) and cell growth processes (FRZB, SFRP4, SPOCK, CTGF).
- FRZB frizzled-related protein
- SFRP4 secreted frizzled related protein 4
- FRZB frizzled-related protein
- SFRP4 secreted frizzled related protein 4
- CFHL3 complement factor H-related 3
- FCN3 ficolin 3
- CCL2 chemokine ligand 2
- S100A8 calgranulin A
- CCL2 is a prominent member of the broader functional group of immune and inflammatory processes and was found to be down- regulated in both datasets A and B.
- This chemokine capable of interacting with TNF-alpha and IL-6-related pathways, has been localized to the cardiomyocyte compartment by immunohistochemistry (24). It promotes attraction and invasion of activated leukocytes into the failing myocardium, but is also involved in shaping the extracellular matrix by modulating the activity of matrix metalloproteinases and collagen turnover (35) as well as cell proliferation and induction of apoptosis (24). Down-regulation of CCL2 transcripts in end-stage heart failure may therefore represent an adaptive mechanism to promote cell survival.
- additional chemokines like CCLl 1 and CCL 18 were also found to be down- regulated in Affymetrix and Unigene arrays, respectively.
- Dataset A cDNA microarray study with 28 septal myocardial samples obtained from 13 DCM hearts at the time of transplantation and 15 NF donor hearts which were not transplanted because of palpable coronary calcifications. The latter patient group was not known to have any history of overt cardiovascular disease. Detailed patient characteristics are listed in Table 3.
- Dataset B oligonucleotide microarray study with twelve independent subendocardial left ventricular samples were collected from seven DCM patients and five NF donors. Detailed patient characteristics are listed in Table 3.
- RNA isolation, sample preparation, labeling, hybridization to RZPD LJnigene 3.1 cDNA (37.5K) and to Affymetrix U133A (22.2K) arrays was carried out as described previously (6, 11 , 12).
- Dataset C six NF, 21 DCM and 10 ICM samples hybridized to Affymetrix HG-U 133 A arrays (13). Normalized gene expression data were downloaded from Gene Expression Omnibus (accession number GSE1869).
- Dataset D available online through a program for genomic application funded by the National Heart, Lung, and Blood Institute and consisted of 14 NF, 27 DCM and 32 ICM samples hybridized to Affymetrix HG-U 133 2.0 plus arrays (http: / / www.cardiogenomics.org) (14).
- Prediction analysis for microarrays was used for classification.
- the ability to correctly classify the status of DCM and NF samples was assessed by complete cross-validation implemented in the Bioconductor package "MCRestimate" (22).
- MCRestimate Bioconductor package "MCRestimate"
- the samples in every study were randomly divided into equally sized subsets. In each following step, one subset was left aside and the classifier (filtering and PAM) was built on the remaining samples (training set).
- the status (NF vs. DCM) of the left-out samples was predicted and compared with the clinically diagnosed status. Optimization of the PAM parameter and of the number of genes remaining after variance filtering was achieved through a second cross-validation within each training set.
- To estimate the variability of the cross-validation result based on different sample compositions of the training set the procedure was repeated 50 times. A sample was called "misclassified” if it was incorrectly classified in more than half of all cross-validations.
- Dataset A 1353 transcripts were up-regulated and 384 were down -regulated in DCM.
- Dataset B 399 transcripts were up-regulated and 75 transcripts were down- regulated in DCM.
- up-regulation was about four- to five-times more common than down -regulation, indicating a net transcriptional activation in heart failure.
- 76 transcripts were found to be consistently deregulated in both studies, representing an approximate 16% overlap at the single gene level between both microarray studies.
- NPPB pro-brain natriuretic peptide
- CCL2 chemokine ligand 2
- differentially expressed genes were related to their respective GO classes. Thereby, it was possible to identify specific biological processes which were consistently enriched in up- or down-regulated transcripts of both studies. For example, both studies showed a marked up-regulation of transcripts involved in protein biosynthesis in DCM.
- extracellular matrix protein 2 (a member of the small leucine rich proteoglycans (SLRP), important for collagen fibrillogenesis), asporin and most other members of the SLRP family were found to be up-regulated as well (decorin, lumican, biglycan, fibromodulin, osteoglycin, and osteomodulin), highlighting their importance in extracellular remodeling.
- Z-disc genes coding for Z-disc components were noted, including caldesmon 1, sarcospan, sarcoglycan epsilon, utrophin, spectrin, titin, vinculin, sarcoglycan D and G, aJpha-actinin, LIM-domain binding 3, and alpha-2-capping protein.
- the Z-disc is thought to act as a sensor, linking biomechanical forces to the activation of stress pathways (25).
- pro- and anti-apoptotic programs may determine if relevant loss of myocytes occurs.
- up-regulation of anti- (FGFl, DSIPI, CCL2) and pro-apoptotic transcripts (BCLAFl , FOXO3A) were noted in these studies.
- the second goal of the experiments of the present invention was to identify a specific set of transcripts which could reliably classify DCM and NF samples.
- the classification method "PAM" was performed on four independent microarray studies. Very low misclassification rates were found in two studies (Datasets A and B) and in Dataset D for the classification of NF versus DCM samples ( Figure 2). Specifically, one out of twelve samples was misclassified in Dataset B. Likewise, Datasets A and D showed similar results, with one out of 28 and three out of 41 misclassified samples, respectively. In contrast, the classification algorithm did not show any predictive power in Dataset C. This was unexpected as the expression levels of established molecular cardiomyopathy markers, including pro-BNA or pro- ANP, suggested a clear separation into NF and failing ventricular samples ( Figure 3).
- the 27-gene signature included known marker genes of heart failure: pro- BNP, pro-ANP, corin (converts pro- ANP to biologically active ANP), transcripts encoding for sarcomer structure proteins (MYH6, MYHlO), anti-apoptotic processes (CCL2, PHLDAl , SNCA), cell growth (FRZB, SFRP4, SPOCK 1 CTGF) and cell cycle control (G0S2, ETV5, RARRES l).
- pro- BNP pro- ANP
- pro-ANP corin (converts pro- ANP to biologically active ANP)
- CCL2, PHLDAl , SNCA anti-apoptotic processes
- FRZB SFRP4, SPOCK 1 CTGF
- G0S2, ETV5, RARRES l cell cycle control
- RNA samples used for microarray hybridization were amplified once. Linear amplification was performed using the MessageAmpTM aRNA Kit (Ambion, Huntingdon, United Kingdom) according to the manufactures instructions.
- RNA amplified RNA
- Agilent 2100 bioanalyzer Agilent Technologies GmbH, Waldbronn, Germany.
- all aRNA samples showed a length distribution of 50 - 6000 nucleotides with maximum peaks at about 900 - 1000 nucleotide.
- a slight length shortening of the aRNA samples was observed.
- Utility of T7 RNA polymerase based linear amplification has been shown previously. See Sultmann et al., "Gene expression in kidney cancer is associated with cytogenetic abnormalities, metastasis formation, and patient survival," Clin Cancer Res. 2005; l l :646-655.
- Cy3- and Cy5-labeled probes were purified with Microcon YM-30 columns (Milipore, Bedford, MA, USA), combined and resuspended in 50 ⁇ l Ix DIG-Easy hybridization buffer (Roche Diagnostics, Mannheim, Germany), containing 10x Denhardt's solution and 2 ng/ ⁇ l Cotl-DNA (Invitrogen). Hybridizations were carried out in duplicate on Unigene 3.1 microarrays.
- the hybridized arrays were scanned with the GenePix 4000B microarray scanner (Axon Instruments Inc., Union City, CA, USA), and analyzed using GenePix Pro 4.1 software (Axon Instruments).
- HG-U 133A chip (Affymetrix, Santa Clara, CA, USA) representing 22.283 probe sets was used for each human heart sample.
- the sequences were derived from GenBank, dbEST and RefSeq. Sequence clusters were created from Build 133 of UniGene (April 20, 2001). Further information about the Gene Chip System can be obtained at www.affymetrix.com. See Liu et al., "NetAffx: Affymetrix probesets and annotations," Nucleic Acids Res., 2003;31 :82-6. mRNA-Preparation and Hybridization to Affymetrix HG-U 133A Microarravs
- Double-stranded cDNA was synthesized from lO ⁇ g total RNA by using the Superscript double-stranded cDNA synthesis kit (Invitrogen, Düsseldorf, Germany) with an HPLC-purified oligo(dT) primer containing a T7 RNA polymerase promoter (GENSET, La Jolla, CA, USA) following the manufacturer's protocol.
- Biotinylated cRNA probes were synthesized by in vitro transcription using ENZO BioArray RNA transcript labeling kit (ENZO Diagnostics, Farmingdale, NY, USA). Fragmentation of lO ⁇ g biotinylated cRNA as well as subsequent steps of hybridization, washing and staining followed instructions provided by Affymetrix (Affymetrix, Santa Clara, CA, USA).
- the hybridized arrays were scanned with the GeneChip Scanner 2500 (Affymetrix, Santa Clara, CA, USA) and preprocessed using Microarray Suite 5 software (Affymetrix, Santa Clara, CA, USA).
- Gene-specific primers and probes were designed using Primer 3 software (Applied Biosystems, Foster City, CA, USA) to amplify fragments of 70-150 base pairs in length close to the 3'-end of the transcript. Real-time PCR was performed in triplicate for each sample with 10 ⁇ l aliquot of diluted cDNA (1 :3).
- a 2x Universal PCR Master-Mix from Perkin Elmer (containing AmpliTaq GoldTM DNA-Polymerase, AmpErase UNG, dNTPs with dUTPs, passive reference dye and optimized buffer including MgCh), 900 nM primer and 200 nM probe were used.
- PCR-amplification of cDNA started with a "hot start -activation of SureStart Taq polymerase at 95°C for 10 minutes, followed by 40 cycles of 15s denaturation at 95°C, annealing for 60s at 58°C, and 10s elongation at 72"C. All experimental results for the samples with a coefficient of variation >10% were retested.
- ⁇ Ct relative quantification method based on the REST-program developed by Pfaffl was used. See Pfaffl et al., "Relative expression software tool (REST) for group-wise comparison and statistical analysis of relative expression results in real-time PCR," Nucleic Acids Res., 2002;30(9):e36.
- Taqman validation of twelve candidate genes served as housekeeping gene.
- GAPDH served as housekeeping gene.
- q-value based on "Significance Analysis of Microarrays" (SAM) is given, whereas Taqman data was analyzed by Student's t-test.
- SAM Signal Analysis of Microarrays
- BCL2-associated transcription factor 1 basic helix-loop-helix domain containing, class B, 3, chromosome 16 open reading frame 45, caldesmon 1, CDC 14 cell division cycle 14 homolog B (S. cerevisiae), carbohydrate (N- acetylglucosamine 6-O) sulfotransferase 5, collagen, type V, alpha 1, collagen, type VlII, alpha 1 , coatomer protein complex, s ⁇ bunit zeta 2, cofactor required for SpI transcriptional activation, subunit 6, 77kDa, connective tissue growth factor, discs, large homolog 1 (Drosophila), dynein, cytoplasmic, light polypeptide 1, dedicator of cytokinesis 9, delta sleep inducing peptide, immunoreactor, extracellular matrix protein 2, female organ and adipocyte specific, exostoses (multiple) 1 , Fibroblast growth factor 1 (acidic), hypothetical protein FLJ22662, forkhead box O3A
- septin 2 sarcoglycan, epsilon, solute carrier family 25 (mitochondrial carrier), member 5, solute carrier family 30 (zinc transporter), member 1, sparc/osteonectin, cwcv and kazal-like domains proteoglycan (testican), sprouty-related, EVH l domain containing 2, sprouty homolog 1, antagonist of FGF signaling (Drosophila), sarcospan, t-complex-associated-testis-expressed 1-like 1 , transmembrane protein 43, thioredoxin domain containing 7, and exportin 1 (CRM l homolog, yeast).
- transcripts classifying DCM and NF samples (generated by PAM classification from dataset B and listed in alphabetical order). In addition to expression values for DCM and NF samples of dataset B, the ranking of the single genes for classification of datasets A-D based on PAM-parameters is given. When transcripts were represented by two probe sets, the ranking of both is indicated. Genes signed with a line in dataset A were not resent in the cDNA arra dataset.
- Kittleson MM Minhas KM, Irizarry RA, et al. Gene expression analysis of ischemic and nonischemic cardiomyopathy: shared and distinct genes in the development of heart failure. Physiol Genomics. 2005;21:299-307.
Landscapes
- Chemical & Material Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Proteomics, Peptides & Aminoacids (AREA)
- Health & Medical Sciences (AREA)
- Organic Chemistry (AREA)
- Wood Science & Technology (AREA)
- Analytical Chemistry (AREA)
- Zoology (AREA)
- Genetics & Genomics (AREA)
- Engineering & Computer Science (AREA)
- Pathology (AREA)
- Immunology (AREA)
- Microbiology (AREA)
- Molecular Biology (AREA)
- Biotechnology (AREA)
- Biophysics (AREA)
- Physics & Mathematics (AREA)
- Biochemistry (AREA)
- Bioinformatics & Cheminformatics (AREA)
- General Engineering & Computer Science (AREA)
- General Health & Medical Sciences (AREA)
- Measuring Or Testing Involving Enzymes Or Micro-Organisms (AREA)
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US83295906P | 2006-07-25 | 2006-07-25 | |
| PCT/IB2007/004191 WO2008053358A2 (en) | 2006-07-25 | 2007-07-24 | A common gene expression signature in dilated cardiomyopathy |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| EP2046997A2 true EP2046997A2 (de) | 2009-04-15 |
Family
ID=39344668
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP07866600A Withdrawn EP2046997A2 (de) | 2006-07-25 | 2007-07-24 | Konventionelle genexpressionssignatur in dilatierter kardiomyopathie |
Country Status (3)
| Country | Link |
|---|---|
| EP (1) | EP2046997A2 (de) |
| JP (1) | JP2009544306A (de) |
| WO (1) | WO2008053358A2 (de) |
Families Citing this family (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| ES2526928T3 (es) * | 2008-04-30 | 2015-01-16 | The Governing Council Of The University Of Toronto | Uso de SFRP-3 en la evaluación de la insuficiencia cardíaca |
| CA3051839A1 (en) | 2017-02-17 | 2018-08-23 | Bristol-Myers Squibb Company | Antibodies to alpha-synuclein and uses thereof |
| CN110809718A (zh) * | 2017-06-21 | 2020-02-18 | 韩国生命工学研究院 | 利用血液生物标志物诊断肌肉衰弱相关疾病的方法和试剂盒 |
| CN113355332B (zh) * | 2021-07-22 | 2022-09-06 | 青岛市妇女儿童医院 | Heg1基因突变体及其应用 |
Family Cites Families (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US6610480B1 (en) * | 1997-11-10 | 2003-08-26 | Genentech, Inc. | Treatment and diagnosis of cardiac hypertrophy |
| WO2003006687A2 (en) * | 2001-07-10 | 2003-01-23 | Medigene Ag | Novel target genes for diseases of the heart |
| WO2003040407A2 (en) * | 2001-11-09 | 2003-05-15 | Max-Planck-Gesellschaft | Novel markers for cardiopathies |
| JP2007514441A (ja) * | 2003-12-16 | 2007-06-07 | ジョシュア エム ヘア | 虚血性と非虚血性の心筋症を区別する遺伝子発現プロファイルの同定 |
-
2007
- 2007-07-24 EP EP07866600A patent/EP2046997A2/de not_active Withdrawn
- 2007-07-24 JP JP2009521384A patent/JP2009544306A/ja active Pending
- 2007-07-24 WO PCT/IB2007/004191 patent/WO2008053358A2/en not_active Ceased
Non-Patent Citations (1)
| Title |
|---|
| See references of WO2008053358A3 * |
Also Published As
| Publication number | Publication date |
|---|---|
| WO2008053358A2 (en) | 2008-05-08 |
| WO2008053358A3 (en) | 2008-11-20 |
| JP2009544306A (ja) | 2009-12-17 |
| WO2008053358A8 (en) | 2009-08-27 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| EP2504451B1 (de) | Verfahren zur vorhersage des klinischen verlaufs von krebs | |
| US10131948B2 (en) | Transcriptomic biomarkers for individual risk assessment in new onset heart failure | |
| US20180094323A1 (en) | Test Kits and Methods for Their Use to Detect Genetic Markers for Transitional Cell Carcinoma of the Bladder | |
| US20180251843A1 (en) | Diagnostic transcriptomic biomarkers in inflammatory cardiomyopathies | |
| WO2011044927A1 (en) | A method for the diagnosis or prognosis of an advanced heart failure | |
| US20210302437A1 (en) | Transcriptomic biomarker of myocarditis | |
| EP3152327B1 (de) | Biomarker von pulmonaler hypertonie | |
| US20100304987A1 (en) | Methods and kits for diagnosis and/or prognosis of the tolerant state in liver transplantation | |
| US20120004127A1 (en) | Gene expression markers for colorectal cancer prognosis | |
| US10584383B2 (en) | Mitochondrial non-coding RNAs for predicting disease progression in heart failure and myocardial infarction patients | |
| CA2728688A1 (en) | In vitro diagnosis/prognosis method and kit for assessment of tolerance in liver transplantation | |
| WO2008053358A2 (en) | A common gene expression signature in dilated cardiomyopathy | |
| CA2549712A1 (en) | Identification of a gene expression profile that differentiates ischemic and nonischemic cardiomyopathy | |
| US20140171371A1 (en) | Compositions And Methods For The Diagnosis of Schizophrenia | |
| CA2525179A1 (en) | A gene equation to diagnose rheumatoid arthritis | |
| WO2006132983A2 (en) | Differential expression of molecules associated with vascular disease risk | |
| AU2014259525B2 (en) | A transcriptomic biomarker of myocarditis | |
| JP2010261920A (ja) | 2型糖尿病診断剤 | |
| US20180142297A1 (en) | Systems and methods for characterizing granulomatous diseases | |
| HK1175820B (en) | Methods to predict clinical outcome of cancer | |
| HK1175820A (en) | Methods to predict clinical outcome of cancer | |
| HK1233687B (en) | Mitochondrial non-coding rnas for predicting disease progression in heart failure and myocardial infarction patients | |
| HK1233687A1 (en) | Mitochondrial non-coding rnas for predicting disease progression in heart failure and myocardial infarction patients | |
| EP2313519A1 (de) | In-vitro-diagnose/prognose-verfahren und kit zur beurteilung der toleranz bei lebertransplantation |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION HAS BEEN WITHDRAWN |
|
| 17P | Request for examination filed |
Effective date: 20090217 |
|
| AK | Designated contracting states |
Kind code of ref document: A2 Designated state(s): AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IS IT LI LT LU LV MC MT NL PL PT RO SE SI SK TR |
|
| AX | Request for extension of the european patent |
Extension state: AL BA HR MK RS |
|
| 18W | Application withdrawn |
Effective date: 20090325 |