# Do the outputted models using textcat.batch-train make use of word vectors?

**URL:** <https://support.prodi.gy/t/do-the-outputted-models-using-textcat-batch-train-make-use-of-word-vectors/1351>\
**Category:** Uncategorized\
**Tags:** usage, textcat, spacy\
**Created:** [March 28, 2019, 10:03am UTC](https://support.prodi.gy/t/do-the-outputted-models-using-textcat-batch-train-make-use-of-word-vectors/1351 "2019-03-28T10:03:38Z")\
**Posts on this page:** 3\
**Page:** 1

<div class="post-metadata">

**Author:** ![bigbeaker](https://avatars.discourse-cdn.com/v4/letter/b/7feea3/32.png) [@bigbeaker](https://support.prodi.gy/u/bigbeaker)\
**Post date:** [March 28, 2019, 10:03am UTC](https://support.prodi.gy/t/do-the-outputted-models-using-textcat-batch-train-make-use-of-word-vectors/1351/1 "2019-03-28T10:03:38Z")

</div>

Hi guys

Quick question on the models outputted by prodigy - do they use word vectors?  
The model I trained and outputted seems like it doesn’t have any knowledge of word vectors and the model is quite small in size ~10mb

If not, how can I use the word-vectors I have for pretraining?

Thanks!

---

<div class="post-metadata">

**Author:** ![ines](https://sea2.discourse-cdn.com/flex020/user_avatar/support.prodi.gy/ines/32/3_2.png) [@ines](https://support.prodi.gy/u/ines)\
**Post date:** [March 28, 2019, 10:29am UTC](https://support.prodi.gy/t/do-the-outputted-models-using-textcat-batch-train-make-use-of-word-vectors/1351/2 "2019-03-28T10:29:52Z")

</div>

If word vectors are present in the model you're updating then yes, spaCy will use those representations during training. This can sometimes give you a nice boost in accuracy.

If your model is only 10mb, it's likely that you started off with an `sm` model that doesn't have word vectors. So instead, try using a model like `en_core_web_md` or `en_core_web_lg`.

> [@bigbeaker](#):
>
> If not, how can I use the word-vectors I have for pretraining?

Actual _pre-training_ is different – that's something we just introduced in [spaCy v2.1](https://spacy.io/usage/v2-1). Here, you're pre-training weights using lots of raw unlabelled text and word vectors. Also [see this blog post](https://explosion.ai/blog/spacy-v2-1#pretraining-example1) for examples. Once you have that artifact, you can pass it in when you train your model. In the next version of Prodigy, which will introduce support for spaCy v2.1, you'll also be able to pass in those pre-trained weights files in the `textcat.teach` and `ner.teach` recipes.

---

<div class="post-metadata">

**Author:** ![bigbeaker](https://avatars.discourse-cdn.com/v4/letter/b/7feea3/32.png) [@bigbeaker](https://support.prodi.gy/u/bigbeaker)\
**Post date:** [March 28, 2019, 12:28pm UTC](https://support.prodi.gy/t/do-the-outputted-models-using-textcat-batch-train-make-use-of-word-vectors/1351/3 "2019-03-28T12:28:10Z")

</div>

ah got it! makes sense

Thanks ines
