The MiMo-V2.5 series introduces two open-source Mixture-of-Experts (MoE) language models MiMo-V2.5-Pro with 1.02T parameters (42B activated) and MiMo-V2.5 with 310B parameters (15B activated) both featuring a context window of one million tokens. It would be great to run MiMo V2.5 with llama.cpp!
The MiMo-V2.5 series introduces two open-source Mixture-of-Experts (MoE) language models MiMo-V2.5-Pro with 1.02T parameters (42B activated) and MiMo-V2.5 with 310B parameters (15B activated) both featuring a context window of one million tokens. It would be great to run MiMo V2.5 with llama.cpp!