MultiPL-E: Go

The Go subset of MultiPL-E: HumanEval and MBPP function-completion problems translated into Go and scored with pass@1.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorycoding
Subcategorymultilingual code generation
Page statusactive
Metricpass@1
Directionhigher_is_better
Unit%
Dataset licenceMIT
PublisherNortheastern University Programming Research Lab (nuprl)

What it measures

This subset translates the HumanEval and MBPP prompts into Go by rewriting each problem's function signature, docstring and tests with Go syntax and typing, then asks the model to complete the function body in Go. The underlying algorithmic problem is unchanged from the Python original; only the surface language differs.

Task format

Function completion in Go: given a translated signature, docstring and (for HumanEval-derived items) doctests, the model generates a function body, which is compiled or interpreted with a real Go toolchain inside a container and checked against translated unit tests.

Models reporting this benchmark

These figures come from the model cards, which carry one collection date per card and no per-score attribution. They are shown as reported, not as verified evidence.
ModelProviderScoreCard as of
Claude Opus 4Anthropic85.52026-04
Claude Opus 4.6Anthropic85.52026-04
Claude Sonnet 4Anthropic82.82026-04
Claude Sonnet 4.5Anthropic82.82026-04
Claude Sonnet 4.5 (latest)Anthropic82.82026-04
GPT-4.1OpenAI82.12026-04
Gemini 2.5 ProGoogle DeepMind81.22026-04
DeepSeek R1DeepSeek80.52026-04
DeepSeek R1 0528DeepSeek80.52026-04
DeepSeek R1 0528 NVFP4 v2NVIDIA80.52026-04
DeepSeek ReasonerDeepSeek80.52026-04
GPT-4oOpenAI79.12026-04
GPT-4o (2024-05-13)OpenAI79.12026-04
GPT-4o (2024-08-06)OpenAI79.12026-04
GPT-4o (2024-11-20)OpenAI79.12026-04
GPT-4o miniOpenAI79.12026-04
Qwen2.5 Coder 32B InstructAlibaba / Qwen Team76.52026-04
Qwen2.5 Coder 32B Instruct AWQAlibaba / Qwen Team76.52026-04
Gemma 4 31BGoogle DeepMind73.22026-04
gemma 4 31B itGoogle DeepMind73.22026-04
gemma 4 31B it GGUFUnsloth73.22026-04
Gemma 4 31B IT NVFP4NVIDIA73.22026-04
Codestral (latest)Mistral AI72.52026-04
Mistral Large (latest)Mistral AI72.12026-04
Mistral Large 2.1Mistral AI72.12026-04
Mistral Large 3Mistral AI72.12026-04
Gemma 4 26BGoogle DeepMind71.52026-04
Qwen2.5 Coder 14B InstructAlibaba / Qwen Team70.12026-04
Llama 3.3 70B Instruct NVFP4NVIDIA68.52026-04
Llama-3.3-70B-InstructMeta68.52026-04
phi 4Microsoft67.52026-04
Phi 4 mini instructMicrosoft67.52026-04
Llama 3.1 70BMeta66.82026-04
Llama 3.1 70B InstructMeta66.82026-04
Qwen2.5 Coder 7B InstructAlibaba / Qwen Team63.82026-04
Qwen2.5 Coder 7B Instruct GPTQ Int4Alibaba / Qwen Team63.82026-04
CodeLlama 34B Instruct hfMeta55.22026-04

Data

This page as JSON · Edit on GitHub