Jamba-1.5 models perform well in multiple languages, even though we include only a very small fraction of non-english data in the post-training phase. Therefore, we speculate the models are able to use the learned multilingual capabilities from the pre-training phase. 6/7
Jamba-1.5 Models Demonstrate Strong Multilingual Performance Despite Limited Post-Training Data
By
–
