
Soofi S achieves dense-model parity with 3B active parameters per token
Soofi S 30B-A3B represents a strategic shift toward efficient, regionally optimized foundation models. This open-source Mixture-of-Experts hybrid activates only 3B parameters per token while maintaining constant inference cache during long-context inference, delivering throughput gains over dense competitors in high-concurrency settings. Pretrained on 27 trillion tokens with German language weighting, it matches 14-27B dense models on English and German benchmarks while achieving top code performance across both languages among open baselines. The release signals growing momentum in European sovereign AI infrastructure and challenges the assumption that scale alone determines competitive positioning.62























