600M parameter language model fine-tuned on 2.47M instruction examples using ChatML format with response-only loss masking.
Base model trained from scratch on 39B tokens (70% EN, 30% FR) on Google Cloud TPUs with JAX.
Base Model | Model Card | Base Demo
<|im_start|>
<|im_end|>
Note: This is a 600M parameter model. Responses may contain factual errors or inconsistencies. The model performs better in English than French due to training data distribution.
Trained with Google TPU Research Cloud (TRC) program