Best open-source LLM for multilingual translation

Find a Open-source translation LLM being performant without relying on a proprietary API has become a concrete priority for teams handling large volumes of multilingual content (contracts, customer support, technical documentation). Unlike a cloud API billed by the token, a locally hosted open-weight model provides complete control over translated data and a fixed cost independent of volume. This article reviews the general-purpose models in the catalog that are best suited to translation, the required hardware configuration, and the role of specialized models such as Aya 32B or Madlad-400 in this landscape.

Why choose an open-weights model for translation

Translating sensitive documents (legal, HR, medical) raises a confidentiality issue when the text passes through a third-party API. A model running locally via llama.cpp, vLLM or Ollama eliminates this network transit: the files remain on the company's infrastructure. That's the reasoning detailed in our guide on document translation with a local LLM and in the one dedicated to multilingual customer support.

The second advantage is the license. Most recent multilingual models in the catalog are published under Apache 2.0 or MIT, allowing commercial use without royalties—a point you should systematically verify before deploying to production.

Multilingual general-purpose models to prioritize

The large mixture-of-experts models released in 2025-2026 generally natively support dozens of languages, thanks to massively multilingual training corpora. Options in the catalog include:

For long text (contracts, technical manuals), prioritizing a large context matters as much as model size: beyond 100,000 tokens, the entire document fits in a single request without manual splitting, preserving terminology consistency.

Specialized translation models: Aya 32B and Madlad-400

Outside the catalog's general-purpose models, two families are specifically designed for multilingual translation and frequently appear in comparisons: Aya 32B (Cohere's Aya Expanse family, focused on covering more than 20 languages) and Madlad-400 (trained on a corpus covering more than 400 languages). These models are not included in BestLLMfor's main catalog, and their exact specifications (VRAM per quantization, tokens/sec) must be confirmed on a case-by-case basis depending on the variant tested.

To compare Aya 32B with comparable general-purpose models, see our comparison Aya Expanse 32B vs Qwen 2.5 32B as well as the lighter version Aya Expanse 8B vs Llama 3 8B. In practice, a specialized model like Aya excels at languages with limited training data, while a large general-purpose model like Qwen 3.5 or Mistral Large 3 remains more versatile for mixed content (translation + rewriting + summarization in the same pipeline).

Hardware configuration and quantization

The quantization choice directly determines the required VRAM:

For a workstation with a single 24 GB or 32 GB GPU, only compact models such as Mistral Small 4 in Q4 quantization remain accessible; models above 200B require a cluster or a Mac with a large amount of unified memory. Our European multilingual guide and our selection of French-language models detail the VRAM/quality tradeoffs for more modest configurations.

Tokens/sec throughput varies significantly with the GPU, context length, and concurrent load; without a measurement on the target hardware, this figure should be confirmed rather than estimated in advance.

Concrete use cases

FAQ

Q: What is the best open-source LLM for high-volume translation?

For a large workload with hardware available, Qwen 3.5 397B-A17B or Mistral Large 3 675B offer good multilingual coverage under an Apache 2.0 license. For more modest infrastructure, Mistral Small 4 (119B, Q4 VRAM ~72 GB) remains the most accessible option in the catalog.

Q: Is Aya 32B better than the general-purpose models in the catalog?

For languages that are underrepresented in general-purpose corpora, Aya 32B may outperform them. For common language pairs (English–French, English–Chinese), general-purpose models such as Qwen 3.5 or GLM 5.3 Flash are often competitive, with a longer context. See the comparison Aya Expanse 32B vs Qwen 2.5 32B.

Q: Is Madlad-400 suitable for enterprise deployment?

Madlad-400 targets very broad language coverage (more than 400 languages), which is useful for niche use cases. Its exact VRAM requirements by quantization must be confirmed for the variant used; it is not included in this comparison's main catalog.

Q: Do you need a dedicated GPU to run a translation LLM locally?

It depends on the size of the model you choose. A compact model such as Mistral Small 4 in Q4 fits on a 24–32 GB GPU or a Mac with sufficient unified memory. Models over 300B require multiple GPUs or a machine with large unified memory.

Q: Which license should you choose for commercial use?

Prefer Apache 2.0 or MIT, which impose neither royalties nor commercial-use restrictions—this applies to Qwen 3.5, Mistral Large 3, GLM 5.3 Flash, and DeepSeek V3.2 in this comparison. Always verify the exact license on the model page before deployment.

Q: How can I measure translation quality before choosing a model?

Test on a representative sample from your domain (contracts, support tickets, documentation) rather than relying solely on generic benchmarks such as MMLU. Also consult theOpen LLM Leaderboard for an initial indication of general ability before job-specific testing.

Conclusion

Choose one Open-source translation LLM depends primarily on the volume to process and the available hardware: general-purpose models such as Qwen 3.5, Mistral Large 3, or GLM 5.3 Flash cover most use cases under permissive licenses, while specialized models such as Aya 32B or Madlad-400 remain relevant for low-resource languages. To refine this choice for your hardware configuration, use the configurator or explore the entire catalog.

The hardware for running an LLM locally

To run these models comfortably locally, a RTX 5070 Ti offers an excellent price/performance ratio. Compare prices:

Amazon GMKtec EVO-X2 64GB / 1TB (Ryzen AI Max+ 395) →

Affiliate links — BestLLMfor may earn a commission on purchases, at no extra cost to you. As an Amazon Associate, BestLLMfor earns from qualifying purchases.

Article published and updated on by Mohamed Meguedmi · Data source: /api/models.json · Content license: CC BY 4.0.

An error or update to report? Contribute.