Home/Academic Papers/The papers that got us here/BERT: Pre-training of Deep Bidirectional TransformersItemBERT: Pre-training of Deep Bidirectional Transformersin The papers that got us here by TheLysts Platform0likesLike this itemFollow TheLysts PlatformOpen in appBERT: Pre-training of Deep Bidirectional Transformers on “The papers that got us here”, a list by TheLysts Platform on TheLysts.DetailsPhoto—Name—CommentMade pretrain-then-finetune the standard recipe for NLP and showed transfer learning works for language.PreviousAttention Is All You NeedNextLanguage Models are Few-Shot Learners (GPT-3)Related itemsDoll & Hill, Smoking and Carcinoma of the LungMarshall & Warren, Unidentified Curved Bacilli in the StomachFleming, On the Antibacterial Action of Cultures of a PenicilliumWatson & Crick, Molecular Structure of Nucleic AcidsReport