Transcription of Estimating species richness - UVM
1 OUP CORRECTED PROOF FINAL, 18/10/2010, SPi CHAPTER 4. Estimating species richness Nicholas J. Gotelli and Robert K. Colwell Introduction studies continue to ignore some of the fundamental sampling and measurement problems that can com- Measuring species richness is an essential objec- promise the accurate estimation of species richness tive for many community ecologists and conserva- (Gotelli & Colwell 2001). tion biologists. The number of species in a local In this chapter we review the basic statisti- assemblage is an intuitive and natural index of cal issues involved with species richness estima- community structure, and patterns of species rich- tion. Although a complete review of the subject is ness have been measured at both small ( Blake beyond the scope of this chapter, we highlight sam- & Loiselle 2000) and large ( Rahbek & Graves pling models for species richness that account for 2001) spatial scales.
2 Many classic models in commu- undersampling bias by adjusting or controlling for nity ecology, such as the MacArthur Wilson equi- differences in the number of individuals and the librium model (MacArthur & Wilson 1967) and number of samples collected (rarefaction) as well as the intermediate disturbance hypothesis (Connell models that use abundance or incidence distribu- 1978), as well as more recent models of neutral tions to estimate the number of undetected species theory (Hubbell 2001), metacommunity structure (estimators of asymptotic richness ). (Holyoak et al. 2005), and biogeography (Gotelli et al. 2009) generate quantitative predictions of the number of coexisting species . To make progress in State of the field modelling species richness , these predictions need Sampling models for biodiversity data to be compared with empirical data.
3 In applied ecology and conservation biology, the number of Although the methods of Estimating species rich- species that remain in a community represents the ness that we discuss can be applied to assemblages ultimate scorecard' in the fight to preserve and of organisms that have been identified by genotype restore perturbed communities ( Brook et al. ( Hughes et al. 2000), to species , or to some 2003). higher taxonomic rank, such as genus or family ( Yet, in spite of our familiarity with species rich- Bush & Bambach 2004), we will write species ' to ness, it is a surprisingly difficult variable to mea- keep it simple. Because we are discussing estima- sure. Almost without exception, species richness tion of species richness , we assume that one or more can be neither accurately measured nor directly samples have been taken, by collection or observa- estimated by observation because the observed tion, from one or more assemblages for some speci- number of species is a downward-biased estimator fied group or groups of organisms.
4 We distinguish for the complete (total) species richness of a local two kinds of data used in richness studies: (1) inci- assemblage. Hundreds of papers describe statistical dence data, in which each species detected in a sam- methods for correcting this bias in the estimation ple from an assemblage is simply noted as being of species richness (see also Chapter 3), and spe- present, and (2) abundance data, in which the abun- cial protocols and methods have been developed dance of each species is tallied within each sample. for Estimating species richness for particular taxa Of course, abundance data can always be converted ( Agosti et al. 2000). Nevertheless, many recent to incidence data, but not the reverse. 39. OUP CORRECTED PROOF FINAL, 18/10/2010, SPi 40 B I O L O G I C A L DI V E R S I T Y.
5 Box Observed and estimated richness Sobs is the total number of species observed in a sample, or ACE (for abundance data). in a set of samples. Sest is the estimated number of species in the . 10. Srare = fk is the number of rare species in a sample (each assemblage represented by the sample, or by the set of k =1. samples, where est is replaced by the name of an estimator. with 10 or fewer individuals). Abundance data. Let fk be the number of species each S . obs Sabund = fk is the number of abundant species in a represented by exactly k individuals in a single sample. k =11. Thus, f0 is the number of undetected species ( species sample (each with more than 10 individuals). present in the assemblage but not included in the sample), . 10. f1 is the number of singleton species , f2 is the number of nrare = k fk is the total number of individuals in the k =1.
6 Doubleton species , etc. The total number of individuals in rare species . S . obs The sample coverage estimate is C AC E = 1 nrfar1 e , the the sample is n = fk . k =1. proportion of all individuals in rare species that are not Replicated incidence data. Let qk be the number of singletons. Then the ACE estimator of species richness is f1. species present in exactly k samples in a set of replicate SACE = Sabund + CSrACar eE + C AC 2 , where 2 ACE is the E ACE. incidence samples. Thus, q0 is the number of undetected coefficient of variation, species ( species present in the assemblage but not included . 10 . in the set of samples), q1 is the number of unique species , k(k 1)fk S . rare k=1 . q2 is the number of duplicate species , etc. The total number 2 ACE = max 1, 0.
7 S CACE (nrare ) (nrare 1) . obs of samples is m = qk . k =1. The formula for ACE is undefined when all rare species Chao 1 (for abundance data) are singletons (f1 = nrare , yielding CACE = 0). In this case, compute the bias-corrected form of Chao1 instead. f2. SChao1 = Sobs + 2 1f2 is the classic form, but is not defined when f2 = 0 (no doubletons). ICE (for incidence data). 1). SChao1 = Sobs + f2(1 ( ff21+1) is a bias-corrected form, always . 10. Sinfr = qk is the number of infrequent species in a obtainable. k =1. 2 3 4. 1 f1 f1 1 f1. var(SChao1 ) = f2 2 f2. + f2. + 4 f2. for sample (each found in 10 or fewer samples). S . f1 > 0 and f2 > 0 (see Colwell 2009, Appendix B of Sfreq =. obs qk is the number of frequent species in a EstimateS User's Guide for other cases and for asymmetrical k =11.)
8 Confidence interval computation). sample (each found in more than 10 samples).. 10. ninfr = kqk is the total number of incidences in the Chao 2 (for replicated incidence data) k =1. infrequent species . q12. SChao2 = Sobs + is the classic form, but is not defined The sample coverage estimate is CICE = 1 niqnf1 r , the 2q2. when q2 = 0 (no duplicates). q1 (q1 1) proportion of all incidences of infrequent species that are SChao2 = Sobs + m 1 m 2(q2 +1). is a bias-corrected form, not uniques. Then the ICE estimator of species richness is S nf r always obtainable.. 2 3 4 CICE = Sfreq + Ci ICE + CqICE. 1. 2 ICE , where 2 ICE is the coefficient var(SChao2 ) = q2 12 qq12 + qq12 + 14 qq12 for of variation, . q1 > 0 and q2 > 0 (see Colwell 2009, Appendix B of 10.
9 S k(k 1)qk . EstimateS User's Guide for other cases and for asymmetrical infr minfr k=1 . 2 ICE = max 1, 0 . confidence interval computation). C ICE (minfr 1) (ninfr ). 2 . OUP CORRECTED PROOF FINAL, 18/10/2010, SPi Estimating species richness 41. The formula for ICE is undefined when all infrequent Jackknife estimators (for incidence data). species are uniques (q1 = ninfr , yielding CICE = 0). In this case, compute the bias-corrected form of Chao2 The first-order jackknife richness estimator is instead.. m 1. Sjackknife1 = Sobs + q1. m Jackknife estimators (for abundance data). The second-order jackknife richness estimator is The first-order jackknife richness estimator is . Sjackknife1 = Sobs + f1 q (2m 3) q2 (m 2)2. Sjackknife2 = Sobs + 1 . The second-order jackknife richness estimator is m m (m 1).
10 Sjackknife2 = Sobs + 2f1 f2. By their nature, sampling data document only inferences about the number of colours ( species ) in the verified presence of species in samples. The the entire jar. This process of statistical inference absence of a particular species in a sample may depends critically on the biological assumption that represent either a true absence (the species is not the community is closed,' with an unchanging total present in the assemblage) or a false absence (the number of species and a steady species abundance species is present, but was not detected in the distribution. Jellybeans may be added or removed sample; see Chapter 3). Although the term pres- from the jar, but the proportional representation of ence/absence data' is often used as a synonym for colours is assumed to remain the same.