I have field within a pandas dataframe with a text field for which I want to generate BioBERT embeddings. Is there a simple way with which I can generate the vector embeddings? I want to use them within another model.
here is a hypothetical sample of the data frame
| Visit Code | Problem Assessment |
|---|---|
| 1234 | ge reflux working diagnosis well |
| 4567 | medication refill order working diagnosis note called in brand benicar 5mg qd 30 prn refill |
I have tried this package, but receive an error upon installation https://pypi.org/project/biobert-embedding
Error:
Collecting biobert-embedding
Using cached biobert-embedding-0.1.2.tar.gz (4.8 kB)
ERROR: Could not find a version that satisfies the requirement torch==1.2.0 (from biobert-embedding) (from versions: 0.1.2, 0.1.2.post1, 0.1.2.post2, 1.7.1)
ERROR: No matching distribution found for torch==1.2.0 (from biobert-embedding)
Any help is GREATLY appreciated!