Prompting for explanations improves Adversarial NLI. Is this true? {Yes} it is {true} because {it weakens superficial cues}

Benchmark Model Rank Results
natural-language-inference-on-anli-testT0-11B (explanation prompting)A1: 75.6A2: 60.6A3: 59.9
natural-language-inference-on-anli-testT5-3B (explanation prompting)A1: 81.8A2: 72.5A3: 74.8