Is there a way to correctly load a pre-trained transformers model without the configuration file?

milesc · August 12, 2021, 8:58pm

The configuration file I have been provided is showing “vocab_size”: 30522, same as original BERT. I will try to obtain the correct config file.

{
“attention_probs_dropout_prob”: 0.1,
“hidden_act”: “gelu”,
“hidden_dropout_prob”: 0.1,
“hidden_size”: 1024,
“initializer_range”: 0.02,
“intermediate_size”: 4096,
“max_position_embeddings”: 512,
“num_attention_heads”: 16,
“num_hidden_layers”: 24,
“type_vocab_size”: 2,
“vocab_size”: 30522
}

If vocab size is different, do you think I might also need the trained tokenizer?

Topic		Replies	Views
Load weight from local ckpt file Beginners	9	9010	February 23, 2021
Load EncoderDecoderModel from a checkpoint Models	0	305	March 9, 2023
Differences between Config.from_pretrained and Model.from_pretrained 🤗Transformers	1	1177	July 20, 2021
Loading a checkpoint from training GPT2LMHeadModel 🤗Transformers	0	473	May 23, 2023
Loading pytorch_pretrained_bert models with transformers Beginners	2	1923	April 29, 2021

Is there a way to correctly load a pre-trained transformers model without the configuration file?

Related topics