Side-by-side comparison of IBM Granite 4.1 8B (IBM · USA) and Step-3.7-Flash (StepFun) for self-hosted deployment of the open-weight model. IBM Granite 4.1 8B is rated EU-ready; Step-3.7-Flash is conditional. They part ways on training data: IBM Granite 4.1 8B is "Disclosed", Step-3.7-Flash is "Undisclosed".
| Field | ||
|---|---|---|
| Summary | ||
| Verdict | EU-ready Per the published model card, Granite 4.1 8B is an Apache 2.0 9B-parameter dense decoder with a 131k-token context, sourced from publicly-available datasets, internal synthetic data and human-curated material. IBM continues the unusual-for-the-industry training-data transparency that anchored the Granite 3 family, and offers IP indemnification when the model is consumed via watsonx — a strong default for regulated enterprise pilots that need a defensible weights-available alternative to hyperscaler frontier models. | Conditional Per the published Apache 2.0 LICENSE, the Step-3.7-Flash weights ship without commercial restriction — a 198B / 11B-active vision-language MoE with a 256k-token context aimed at tool-heavy and agentic workflows. The remaining EU-readiness gaps are the entirely undisclosed training corpus and StepFun's Shanghai-based vendor jurisdiction; deploy on self-managed EU infrastructure and document the Art. 50 transparency story for any synthetic or biometric output. |
| Last reviewed | 2026-05-03 | 2026-05-30 |
| Open-weight | ||
| Licence | Apache 2.0 | Apache 2.0 |
| Commercial use | Unrestricted | Unrestricted |
| Training data | Disclosed | Undisclosed |
| Origin | USA | China (Shanghai) |
| Performance & pricing? | ||
| Quality index | 12/100 | — |
| Speed | 91 tok/s | — |
| Blended price | $0.06/M | — |
| Context window | — | — |
| Evidence | ||
| Sources | ||
No overlapping sources between the two entries.