INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair
Open paper
Benchmark
Model
Rank
Results
code-generation-on-mbpp
GPT-3.5 Turbo + INTERVENOR
#21
Accuracy: 69.8
code-generation-on-mbpp
GPT-3.5 Turbo (few-shot)
#64
Accuracy: 45.4
code-generation-on-mbpp
GPT-3.5 Turbo (0-shot)
#72
Accuracy: 39.8
Rank counts only results with a code link.