MultiPL-E: Rust

The Rust subset of MultiPL-E: HumanEval and MBPP function-completion problems translated into Rust and scored with pass@1.

unassessed

This page is a discovery lead. Nobody has yet assessed it against the catalogue contract, so it carries no disposition. Absence of evidence here is not evidence of staleness.
Categorycoding
Subcategorymultilingual code generation
Page statusactive
Metricpass@1
Directionhigher_is_better
Unit%
Dataset licenceMIT
PublisherNortheastern University Programming Research Lab (nuprl)

What it measures

This subset translates the HumanEval and MBPP prompts into Rust by rewriting each problem's function signature, docstring and tests with Rust syntax and typing, then asks the model to complete the function body in Rust. The underlying algorithmic problem is unchanged from the Python original; only the surface language differs.

Task format

Function completion in Rust: given a translated signature, docstring and (for HumanEval-derived items) doctests, the model generates a function body, which is compiled or interpreted with a real Rust toolchain inside a container and checked against translated unit tests.

Models reporting this benchmark

These figures come from the model cards, which carry one collection date per card and no per-score attribution. They are shown as reported, not as verified evidence.
ModelProviderScoreCard as of
Claude Opus 4Anthropic82.12026-04
Claude Opus 4.6Anthropic82.12026-04
Claude Sonnet 4Anthropic78.52026-04
Claude Sonnet 4.5Anthropic78.52026-04
Claude Sonnet 4.5 (latest)Anthropic78.52026-04
GPT-4.1OpenAI76.22026-04
Gemini 2.5 ProGoogle DeepMind75.52026-04
DeepSeek R1DeepSeek75.12026-04
DeepSeek R1 0528DeepSeek75.12026-04
DeepSeek R1 0528 NVFP4 v2NVIDIA75.12026-04
DeepSeek ReasonerDeepSeek75.12026-04
GPT-4oOpenAI72.82026-04
GPT-4o (2024-05-13)OpenAI72.82026-04
GPT-4o (2024-08-06)OpenAI72.82026-04
GPT-4o (2024-11-20)OpenAI72.82026-04
GPT-4o miniOpenAI72.82026-04
Qwen2.5 Coder 32B InstructAlibaba / Qwen Team70.22026-04
Qwen2.5 Coder 32B Instruct AWQAlibaba / Qwen Team70.22026-04
Gemma 4 31BGoogle DeepMind67.52026-04
gemma 4 31B itGoogle DeepMind67.52026-04
gemma 4 31B it GGUFUnsloth67.52026-04
Gemma 4 31B IT NVFP4NVIDIA67.52026-04
Codestral (latest)Mistral AI65.82026-04
Gemma 4 26BGoogle DeepMind65.22026-04
Mistral Large (latest)Mistral AI63.52026-04
Mistral Large 2.1Mistral AI63.52026-04
Mistral Large 3Mistral AI63.52026-04
Qwen2.5 Coder 14B InstructAlibaba / Qwen Team62.52026-04
phi 4Microsoft60.82026-04
Phi 4 mini instructMicrosoft60.82026-04
Llama 3.3 70B Instruct NVFP4NVIDIA60.52026-04
Llama-3.3-70B-InstructMeta60.52026-04
Llama 3.1 70BMeta58.22026-04
Llama 3.1 70B InstructMeta58.22026-04
Qwen2.5 Coder 7B InstructAlibaba / Qwen Team55.22026-04
Qwen2.5 Coder 7B Instruct GPTQ Int4Alibaba / Qwen Team55.22026-04
CodeLlama 34B Instruct hfMeta42.12026-04

Data

This page as JSON · Edit on GitHub