Models to evaluate
| Model | Architecture / size | License | Hardware class | Provider |
|---|---|---|---|---|
| Ministral 3 8B Instruct 2512Mistral AI | 8B-class · vision-language edge model | Apache 2.0 | ≤8B · consumer/local | 🇫🇷 France |
| Devstral Small 2 24B Instruct 2512Mistral AI | 24B · coding agent model | Apache 2.0 | 17–32B · high-memory workstation | 🇫🇷 France |
| Mistral Small 4 119B A6BMistral AI | 119B / 6.5B active · multimodal MoE | Apache 2.0 | Model-specific · large / specialized | 🇫🇷 France |
| Granite 4.2 8BIBM | 8B · 128K enterprise reasoning | Apache 2.0 | ≤8B · consumer/local | 🇺🇸 United States |
| Qwen3 14BQwen / Alibaba | 14B · dense general-purpose model | Apache 2.0 | 9–16B · workstation/local | 🇨🇳 China |
| gpt-oss-20bOpenAI | 20B · compact reasoning | Apache 2.0 | 17–32B · high-memory workstation | 🇺🇸 United States |
| Gemma 3 12B ITGoogle DeepMind | 12B · multimodal | Gemma Terms | 9–16B · workstation/local | 🇺🇸 United States |
| Phi-4Microsoft | 14B · dense | MIT | 9–16B · workstation/local | 🇺🇸 United States |
Decision criteria
Choose an EU/EEA infrastructure boundary before selecting external services.
Keep provider origin separate from where prompts and documents are processed.
Review the exact checkpoint license and software stack.
Document legal role, processors, retention and international transfers separately.
Deployment reality
Validate the exact checkpoint, precision or quantization, runtime, context length and concurrency target. Weight memory alone does not capture KV cache, runtime workspaces, multimodal encoders or distributed-serving overhead.
OWM keeps license, provider origin and data residency separate. A provider-country label is provenance metadata; the deployer determines where inference and connected services run.