North Small Translate - Cohere Documentation
Multilingual Reasoning Image Inputs Safety Modes Citations Tool Use Structured Outputs For both trial keys and production keys, North Small Translate is free until rate limits are reached. Learn more about rate limits for different models and key types here .

- Learn more about rate limits for different models and key types here .
- Suggested Hardware : Two H100s or one B200 (W4A16) North Small Translate is a mixture-of-experts (MoE) model purpose-built for machine translation.
- It has 218 billion total parameters, with 25 billion active parameters, and a 16K context length.
Multilingual Reasoning Image Inputs Safety Modes Citations Tool Use Structured Outputs For both trial keys and production keys, North Small Translate is free until rate limits are reached. Learn more about rate limits for different models and key types here . To learn more about using Cohere models in production, reach out to sales.cohere.com . Suggested Hardware : Two H100s or one B200 (W4A16) North Small Translate is a mixture-of-experts (MoE) model purpose-built for machine translation. It has 218 billion total parameters, with 25 billion active parameters, and a 16K context length. The model provides flexible deployment options for research and enterprise use. Its weights are openly available in W4A16, FP8, and BF16 under the Creative Commons Attribution-NonCommercial 4.0 license What can North Small Translate be used for? North Small Translate is well-suited for: Knowledge management : Translate internal documentation, wikis, standard operating procedures, technical manuals, historical reports, and institutional knowledge. Safety and operations : Translate safety manuals, emergency procedures, maintenance instructions, manufacturing procedures, incident reports, field-service documentation, and operational alerts. Internal communication : Translate company announcements, leadership communications, HR policies, employee handbooks, benefits information, onboarding materials, newsletters, and intranet content. Localization and customer support : Translate product content and customer communications while retaining control over where the model and data are hosted. North Small Translate supports English and more than 50 language and locale variants. Its tier-one languages are Modern Standard Arabic, German, French, Japanese, Korean, Russian, and Ukrainian. All supported languages and locale variants North Small Translate is available on the free tier through the Chat V2 API . Ask the model to translate the supplied content into a target language: co = ClientV2 ( api_key = " <YOUR_API_KEY> " ) content = " Enterprises need accurate translations of business-critical documents. " " content " : f "Translate everything that follows into { target_language } : \n\n { content } " , print ( response . message . content [ 0 ]. text ) Cohere deployment (free tier - Cohere API) : Try the model on the free tier through the Chat V2 API. Usage is subject to limitations set out in the applicable commercial agreements with customers, including non-commercial use under the CC BY-NC 4.0 license. Hardware : Suggested hardware by quantization format: Machine translation uses software to convert text from one language into another. A purpose-built generative model can act as the translation engine in an application, workflow, or translation platform. A model gives developers control over integration, prompting, deployment, data residency, and customization. Self-hosted and private deployments can also support sovereignty requirements and reduce dependence on a fixed When is a translation platform a better fit? A translation platform wraps a model with features such as collaboration, content management, project management, and formatting. Use a platform when those workflow features are more important than direct control over the model.
Sources
Related stories

Introducing CARE-X: Towards Clinically Useful Radiology VLMs with Auxiliary Supervision, Reward-Aligned Learning, and Tool-Augmented Measurement
Research Note: CARE-X is a research model and not a Microsoft product offering or medical device.

Red Hat Shrinks Nemotron 3.5 Lightning's 30B Agent Model by Half With FP8
Takeaways − Red Hat AI released an FP8 quantized build of NVIDIA Nemotron 3.5 Lightning 30B A3B. Cuts GPU memory and disk by roughly 50% versus the BF16 reference weights.

Kling 4.0 Debuts 30s AI Video Model - Briefs Finance
Warning : Undefined variable $stocks in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : Undefined variable $funds in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 472 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d...

Claude’s New addTools() Can Reuse 98.7% of Your Next Request. Editing tools[] Reuses None.
Author(s): Chew Loong Nian – AI ENGINEER Originally published on Towards AI.