microsoft/deberta-v3-base
mitDeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing DeBERTa( improves the BERT and Ro...
google-bert/bert-base-cased
apache-2.0BERT base model (cased) Pretrained model on English language using a masked language modeling (MLM) objective. It was introduced in this pap...
microsoft/BiomedNLP-BiomedBERT-base-uncased-abstract-fulltext
mitMSR BiomedBERT (abstracts + full text) This model was previously named "PubMedBERT (abstracts + full text)". You can either adopt the new mo...
medicalai/ClinicalBERT
ClinicalBERT This model card describes the ClinicalBERT model, which was trained on a large multicenter dataset with a large corpus of 1.2B ...
nlpaueb/legal-bert-base-uncased
cc-by-sa-4.0LEGAL-BERT: The Muppets straight out of Law School LEGAL-BERT is a family of BERT models for the legal domain, intended to assist legal NLP ...
ctheodoris/Geneformer
apache-2.0Geneformer Geneformer is a foundational transformer model pretrained on a large-scale corpus of human single cell transcriptomes to enable c...
FacebookAI/roberta-large
mit--- language: en tags: - exbert license: mit datasets: - bookcorpus - wikipedia --- RoBERTa large model Pretrained model on English language...
microsoft/deberta-v3-large
mitDeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing DeBERTa( improves the BERT and Ro...
esm2_t33_650M_UR50D
mitOfficial fill-mask model by facebook.
hfl/chinese-roberta-wwm-ext-large
apache-2.0Please use 'Bert' related functions to load this model! Chinese BERT with Whole Word Masking For further accelerating Chinese natural langua...