Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Benchmark Model Rank Results
natural-language-inference-on-multinliMT-DNN-ensemble#15Matched: 87.9Mismatched: 87.4
sentiment-analysis-on-sst-2-binary-classificationMT-DNN-ensemble#13Accuracy: 96.5