Rethinking with Retrieval: Faithful Large Language Model Inference

Benchmark Model Rank Results
question-answering-on-strategyqaRethinking with retrieval (GPT-3)#2Accuracy: 77.73