HumanEvalPlus_multilingual is a multilingual version of the benchmark HumanEval+, covering six languages: English, French, German, Spanish, Chinese, and Swahili. Each sample is a Python code generation problem from HumanEval+ (an extended version of OpenAI's HumanEval with 80x more tests), with the natural-language prompt translated into the five target languages.