There is a newer version of the record available.

Published November 18, 2021 | Version v1.1

BioVAE: a pre-trained latent variable language model for biomedical text mining

Description

We release BioVAE, the first large-scale pre-trained latent variable language model for the biomedical domain, which uses the OPTIMUS framework to train on large volumes of biomedical text.

This version contains the encoder parts of the pre-trained models for text mining tasks such as named entity recognition or relation extraction.

Explanation of each file:

  • pm-full-lt32-beta00: latent_size = 32, beta=0.0
  • pm-full-lt32-beta05: latent_size = 32, beta=0.5
  • pm-full-lt768-beta00: latent_size = 768, beta=0.0
  • pm-full-lt768-beta05: latent_size = 768, beta=0.5

Files

Files (1.6 GB)

Name Size
md5:72a862847b94cfca44176e960a5ce64a
408.2 MB Download
md5:3a553eb9131242104100e58fa7ac6737
408.2 MB Download
md5:d4e2febce9c9d8a75aad9a5dfaefc48f
412.4 MB Download
md5:40c2f83b6ebe04ede1efd234625981ae
412.5 MB Download

Additional details