# PCoA biplot with Jaccard distance matrix

**URL:** https://forum.qiime2.org/t/pcoa-biplot-with-jaccard-distance-matrix/23550
**Category:** General Discussion
**Tags:** pcoa, biplot
**Created:** [July 4, 2022, 3:02pm UTC](https://forum.qiime2.org/t/pcoa-biplot-with-jaccard-distance-matrix/23550 "2022-07-04T15:02:32Z")
**Posts on this page:** 3
**Page:** 1

<div class="post-metadata">

### Author: ![fracand](https://forum.qiime2.org/user_avatar/forum.qiime2.org/fracand/32/5996_2.png) [@fracand](https://forum.qiime2.org/u/fracand)
#### Post date: [July 4, 2022, 3:02pm UTC](https://forum.qiime2.org/t/pcoa-biplot-with-jaccard-distance-matrix/23550/1 "2022-07-04T15:02:32Z")

</div>

Hello everyone,  
I have a question regarding the correct utilization of the relative frequency table in a PCoA biplot based on Jaccard distance matrix.  
Usually, when I calculate PCoA biplot, I proced with the calculation of the distance matrix from my feature table, then its PCoA, convert the same feature table utilized earlier to a relative frequency table,  
and then produce the biplot.  
In the last dataset I analyzed, I noticed that a feature present in every sample was among the most important ones, but since I'm using Jaccard I found this result strange.  
I read in this post [Questions about PCoA biplots](https://forum.qiime2.org/t/questions-about-pcoa-biplots/18152) that projection of the feature on the PCoA space is due to the calculation of a covariance matrix between the PCoA matrix and the feature table, so the projection depends also by the feature abundance, and not only by its presence/absence, if I understood correctly.

So to neutralize the problem I utilized a new relative frequency table where every feature, if present, has the same frequency. In this way, the contribution of the omnipresent feature was very close to 0 (10^-18 on PCo1 and 2).

My question is: is it correct to utilize a relative frequency table where every present feature has the same weight when calculating PCoA biplot of a qualitative metric like Jaccard?

Thank you for your attention, I hope I've been clear with my explanation!

---

<div class="post-metadata">

### Author: ![colinbrislawn](https://forum.qiime2.org/user_avatar/forum.qiime2.org/colinbrislawn/32/6221_2.png) [@colinbrislawn](https://forum.qiime2.org/u/colinbrislawn)
#### Post date: [August 1, 2022, 3:46pm UTC](https://forum.qiime2.org/t/pcoa-biplot-with-jaccard-distance-matrix/23550/2 "2022-08-01T15:46:08Z")

</div>

> [@fracand](#):
>
> My question is: is it correct to utilize a relative frequency table where every present feature has the same weight when calculating PCoA biplot of a qualitative metric like Jaccard?

Sure, I would be OK with this method if I found it in a paper and you explained how you rescaled your feature table to make all frequencies either a constant or zero. This makes your distances calculations and also your bi-plot vectors be binary, which is fine.

I guess the other option would be to use weighted distances like Bray-Curtis or weighted Jaccard (Ružička / Ruzicka index), then use relative abundance values for the biplot. But that's not what you want.

I'm not sure how reviewer 3 would feel 🙃

---

<div class="post-metadata">

### Author: ![gregcaporaso](https://forum.qiime2.org/user_avatar/forum.qiime2.org/gregcaporaso/32/17769_2.png) [@gregcaporaso](https://forum.qiime2.org/u/gregcaporaso)
#### Post date: [August 1, 2022, 5:47pm UTC](https://forum.qiime2.org/t/pcoa-biplot-with-jaccard-distance-matrix/23550/3 "2022-08-01T17:47:53Z")

</div>

> [@colinbrislawn](#):
>
> I guess the other option would be to use weighted distances

I would lean toward this option suggested by @colinbrislawn for biplots. I did a quick search for "qualitative biplot" and I'm not turning up anything very informative. Based on how the loadings are calculated though (essentially a correlation between abundances and PCoA axis values) I'm not sure exactly what the loadings would be telling you in this case.
