Quantifying the advantage of domain-specific pre-training on named entity recognition tasks in materials science

-based models by 1%∼12%, implying that domain-specific pre-training provides measurable advantages. Despite relative architectural simplicity, the BiLSTM model consistently outperforms BERT, perhaps due to its domain-specific pre-trained word embeddings. Furthermore, MatBERT and SciBERT models outperform the original BERT model to a greater extent in the small data limit. MatBERT's higher-quality predictions should accelerate the extraction of structured data from materials science literature.

Quantifying the advantage of domain-specific pre-training on named entity recognition tasks in materials science | Litlas