# Specific formula for F score, precision and recall NER

**URL:** <https://support.prodi.gy/t/specific-formula-for-f-score-precision-and-recall-ner/4422>\
**Category:** Uncategorized\
**Tags:** usage, spacy, training\
**Created:** [July 9, 2021, 7:45am UTC](https://support.prodi.gy/t/specific-formula-for-f-score-precision-and-recall-ner/4422 "2021-07-09T07:45:39Z")\
**Posts on this page:** 1\
**Showing post:** 2

<div class="post-metadata">

**Author:** ![ines](https://sea2.discourse-cdn.com/flex020/user_avatar/support.prodi.gy/ines/32/3_2.png) [@ines](https://support.prodi.gy/u/ines)\
**Post date:** [July 10, 2021, 4:55am UTC](https://support.prodi.gy/t/specific-formula-for-f-score-precision-and-recall-ner/4422/2 "2021-07-10T04:55:46Z")

</div>

Hi! If you're training from manually created annotation, the evaluation all happens within spaCy and doesn't depend on Prodigy. spaCy uses a very standard NER evaluation. If you're working with spaCy v2.x, you can view the code here:

> <https://github.com/explosion/spaCy/blob/v2.x/spacy/scorer.py>

For spaCy v3.x, it's here:

> <https://github.com/explosion/spaCy/blob/master/spacy/scorer.py>

> [@Aziz](#):
>
> I'm trying to compare prodigy ner results with Bert ner ?

If you want to do a comparative evaluation, you can also just run both your models over your evaluation data and then calculate the accuracy however you want to, and consistently for both evaluations.

Some thing to keep in mind here: if you're using a non-spaCy model with a tokenizer that doesn't preserve the original text, this may impact your evaluation. It probably also makes sense to train with spaCy v3 directly (you can use `prodigy data-to-spacy` and `spacy convert` to convert your annotations), so you can train a transformer-based that's more directly comparable to another model initialised with transformer weights. Otherwise, your evaluation might not be very meaningful.

---

_[View the full topic](https://support.prodi.gy/t/specific-formula-for-f-score-precision-and-recall-ner/4422)._
