MoE model optimized for German/English with 78B total, 3B active parameters—on-premise deployable with abstention training for regulated sectors, available on Hugging Face under Apache 2.0.
Summary
Enables sovereign deployment in compliance-heavy domains (public admin, aerospace, manufacturing) without exfiltrating data to third-party inference. On-device inference reduces latency and regulatory friction for teams building mission-critical EU/German workloads.
Why it matters
Enables sovereign deployment in compliance-heavy domains (public admin, aerospace, manufacturing) without exfiltrating data to third-party inference. On-device inference reduces latency and regulatory friction for teams building mission-critical EU/German workloads.
Implementation verdict
Replaces closed-model inference for regulated German-language tasks. Requires on-premise GPU capacity and familiarity with MoE routing. Worth evaluating now if you target public sector or automotive verticals—benchmarks show parity with 4x larger active-parameter models (Nemotron 3 Super 120B-A12B). Production-ready; full model weights and tech report available.
Sources
Dev Signal
Get briefs like this in your inbox — free, every weekday.
100+ sources compressed into one 4-minute read. Ranked, cited, implementation-ready.