We shipped a Word2Vec+LSTM over a more accurate BERT, then load testing showed the model was 0.3% of the wall-clock time. The queue was five human reviewers.