Matches in SemOpenAlex for { <https://semopenalex.org/work/W1448572673> ?p ?o ?g. }
- W1448572673 endingPage "e1634" @default.
- W1448572673 startingPage "e1634" @default.
- W1448572673 abstract "Estimating and comparing microbial diversity are statistically challenging due to limited sampling and possible sequencing errors for low-frequency counts, producing spurious singletons. The inflated singleton count seriously affects statistical analysis and inferences about microbial diversity. Previous statistical approaches to tackle the sequencing errors generally require different parametric assumptions about the sampling model or about the functional form of frequency counts. Different parametric assumptions may lead to drastically different diversity estimates. We focus on nonparametric methods which are universally valid for all parametric assumptions and can be used to compare diversity across communities. We develop here a nonparametric estimator of the true singleton count to replace the spurious singleton count in all methods/approaches. Our estimator of the true singleton count is in terms of the frequency counts of doubletons, tripletons and quadrupletons, provided these three frequency counts are reliable. To quantify microbial alpha diversity for an individual community, we adopt the measure of Hill numbers (effective number of taxa) under a nonparametric framework. Hill numbers, parameterized by an order q that determines the measures’ emphasis on rare or common species, include taxa richness ( q = 0), Shannon diversity ( q = 1, the exponential of Shannon entropy), and Simpson diversity ( q = 2, the inverse of Simpson index). A diversity profile which depicts the Hill number as a function of order q conveys all information contained in a taxa abundance distribution. Based on the estimated singleton count and the original non-singleton frequency counts, two statistical approaches (non-asymptotic and asymptotic) are developed to compare microbial diversity for multiple communities. (1) A non-asymptotic approach refers to the comparison of estimated diversities of standardized samples with a common finite sample size or sample completeness. This approach aims to compare diversity estimates for equally-large or equally-complete samples; it is based on the seamless rarefaction and extrapolation sampling curves of Hill numbers, specifically for q = 0, 1 and 2. (2) An asymptotic approach refers to the comparison of the estimated asymptotic diversity profiles. That is, this approach compares the estimated profiles for complete samples or samples whose size tends to be sufficiently large. It is based on statistical estimation of the true Hill number of any order q ≥ 0. In the two approaches, replacing the spurious singleton count by our estimated count, we can greatly remove the positive biases associated with diversity estimates due to spurious singletons and also make fair comparisons across microbial communities, as illustrated in our simulation results and in applying our method to analyze sequencing data from viral metagenomes." @default.
- W1448572673 created "2016-06-24" @default.
- W1448572673 creator A5003229507 @default.
- W1448572673 creator A5049404782 @default.
- W1448572673 date "2016-02-01" @default.
- W1448572673 modified "2023-10-12" @default.
- W1448572673 title "Estimating and comparing microbial diversity in the presence of sequencing errors" @default.
- W1448572673 cites W1901435509 @default.
- W1448572673 cites W1903024071 @default.
- W1448572673 cites W190681677 @default.
- W1448572673 cites W1969295076 @default.
- W1448572673 cites W1980179247 @default.
- W1448572673 cites W1990746682 @default.
- W1448572673 cites W1993602546 @default.
- W1448572673 cites W2000072710 @default.
- W1448572673 cites W2008627263 @default.
- W1448572673 cites W2020302957 @default.
- W1448572673 cites W2022825068 @default.
- W1448572673 cites W2034833171 @default.
- W1448572673 cites W2039516620 @default.
- W1448572673 cites W2041457600 @default.
- W1448572673 cites W2044560677 @default.
- W1448572673 cites W2053925957 @default.
- W1448572673 cites W2078194008 @default.
- W1448572673 cites W2082092506 @default.
- W1448572673 cites W2082221931 @default.
- W1448572673 cites W2087671769 @default.
- W1448572673 cites W2093891969 @default.
- W1448572673 cites W2095306947 @default.
- W1448572673 cites W2096207020 @default.
- W1448572673 cites W2098436674 @default.
- W1448572673 cites W2103336148 @default.
- W1448572673 cites W2106235393 @default.
- W1448572673 cites W2106308757 @default.
- W1448572673 cites W2106651885 @default.
- W1448572673 cites W2106747537 @default.
- W1448572673 cites W2107974372 @default.
- W1448572673 cites W2108718991 @default.
- W1448572673 cites W2110048017 @default.
- W1448572673 cites W2115186425 @default.
- W1448572673 cites W2116601594 @default.
- W1448572673 cites W2120108422 @default.
- W1448572673 cites W2120474334 @default.
- W1448572673 cites W2129297365 @default.
- W1448572673 cites W2129994723 @default.
- W1448572673 cites W2137690603 @default.
- W1448572673 cites W2141034720 @default.
- W1448572673 cites W2141673200 @default.
- W1448572673 cites W2149573313 @default.
- W1448572673 cites W2150914285 @default.
- W1448572673 cites W2151141101 @default.
- W1448572673 cites W2156232632 @default.
- W1448572673 cites W2160113159 @default.
- W1448572673 cites W2162315106 @default.
- W1448572673 cites W2162720870 @default.
- W1448572673 cites W2163236435 @default.
- W1448572673 cites W2166655850 @default.
- W1448572673 cites W2170024099 @default.
- W1448572673 cites W2171783890 @default.
- W1448572673 cites W4238811048 @default.
- W1448572673 doi "https://doi.org/10.7717/peerj.1634" @default.
- W1448572673 hasPubMedCentralId "https://www.ncbi.nlm.nih.gov/pmc/articles/4741086" @default.
- W1448572673 hasPubMedId "https://pubmed.ncbi.nlm.nih.gov/26855872" @default.
- W1448572673 hasPublicationYear "2016" @default.
- W1448572673 type Work @default.
- W1448572673 sameAs 1448572673 @default.
- W1448572673 citedByCount "71" @default.
- W1448572673 countsByYear W14485726732016 @default.
- W1448572673 countsByYear W14485726732017 @default.
- W1448572673 countsByYear W14485726732018 @default.
- W1448572673 countsByYear W14485726732019 @default.
- W1448572673 countsByYear W14485726732020 @default.
- W1448572673 countsByYear W14485726732021 @default.
- W1448572673 countsByYear W14485726732022 @default.
- W1448572673 countsByYear W14485726732023 @default.
- W1448572673 crossrefType "journal-article" @default.
- W1448572673 hasAuthorship W1448572673A5003229507 @default.
- W1448572673 hasAuthorship W1448572673A5049404782 @default.
- W1448572673 hasBestOaLocation W14485726731 @default.
- W1448572673 hasConcept C102366305 @default.
- W1448572673 hasConcept C105795698 @default.
- W1448572673 hasConcept C117251300 @default.
- W1448572673 hasConcept C117354338 @default.
- W1448572673 hasConcept C185429906 @default.
- W1448572673 hasConcept C18903297 @default.
- W1448572673 hasConcept C2779234561 @default.
- W1448572673 hasConcept C30968088 @default.
- W1448572673 hasConcept C33923547 @default.
- W1448572673 hasConcept C53565203 @default.
- W1448572673 hasConcept C54355233 @default.
- W1448572673 hasConcept C86803240 @default.
- W1448572673 hasConcept C97256817 @default.
- W1448572673 hasConceptScore W1448572673C102366305 @default.
- W1448572673 hasConceptScore W1448572673C105795698 @default.
- W1448572673 hasConceptScore W1448572673C117251300 @default.
- W1448572673 hasConceptScore W1448572673C117354338 @default.
- W1448572673 hasConceptScore W1448572673C185429906 @default.
- W1448572673 hasConceptScore W1448572673C18903297 @default.