Back to North
Open WeightsCodingtextUpdated June 9, 2026

North Mini Code: 30B Sparse MoE Coding Model

Model Overview

North Mini Code (1.0) is a specialized 30B total parameter, sparse Mixture-of-Experts (MoE) coding model with ~3B active parameters per token developed by Cohere and Cohere Labs.

Designed for agentic software engineering and developer workflows, it achieves near-30B-scale reasoning and multi-file code editing performance with the low compute overhead of a 3B parameter model.


Key Features

  • Sparse MoE Architecture: 30B total parameters across 128 experts with 8 active per token (~3B active parameters).
  • Hybrid Attention Design: Interleaves sliding-window attention (with RoPE) and global attention in a 3:1 ratio.
  • Extended Context Capability: Features a 256K token context window with up to 64K output token generation.
  • Agentic Optimization: Post-trained via two-stage SFT and RLVR for multi-file repo changes, terminal execution, and scaffolds like OpenCode and SWE-Agent.
  • Open Weights & Deployability: Released under the permissive Apache 2.0 license for single-GPU local execution.

Verified Project Links


Performance & Benchmarks

  • SWE-bench Verified: 67.6% (pass@1) / 80.2% (pass@10).
  • HumanEval: 50.0%.
  • Artificial Analysis Coding Index: 33.4.

Key Features

Sparse MoE Architecture: 30B total parameters across 128 experts with 8 active per token (~3B active parameters) for low latency

Feature 01

Hybrid Attention Design: Interleaves sliding-window attention and global attention in a 3:1 ratio

Feature 02

Extended Context Capability: Features a 256K token context window with up to 64K output token generation capability

Feature 03

Agentic Optimization: Post-trained via two-stage SFT and RLVR for multi-file repo changes, terminal execution, and OpenCode/SWE-Agent scaffolds

Feature 04

Open Weights & Deployability: Released under Apache 2.0 license, enabling local execution and FP8 quantized single-GPU deployment

Feature 05

You might also want to compare

Verified Sources

Tags

moecoding-agentlow-latencycohereapache-2.0

Model Specs

open-weights

Parameters

30B (3B active)

Context Window

256K tokens

License

Apache-2.0

Deployment

self-hostable

Resources & Links

Lineage

Model Family

Part of the North family

Only release in this line currently tracked.

Curator Notes

Verified release from Cohere and Cohere Labs. Released under Apache 2.0 license.

Compare Specs

Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.

Compare Model