Instructions to use nvidia/Nemotron-4-Mini-Hindi-4B-Base with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- NeMo
How to use nvidia/Nemotron-4-Mini-Hindi-4B-Base with NeMo:
# tag did not correspond to a valid NeMo domain.
- Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -13,7 +13,7 @@ library_name: nemo
|
|
| 13 |
# Model Overview
|
| 14 |
|
| 15 |
Nemotron-4-Mini-Hindi-4B-Base is a base model pre-trained on Hindi and English corpus. The Nemotron-Mini-4B-Base (Minitron-4B) is subject to continuous pre-training using Hindi and English data (400B tokens) exclusively to create a strong base model for Hindi, English, and Hinglish. We make extensive use of synthetic data during the continuous pre-training stage. The base small language model (SLM) is optimized through distillation, pruning, and quantization for speed and on-device deployment.
|
| 16 |
-
Please refer to our [arXiv paper](https://arxiv.org/abs/2410.14815) for more details.
|
| 17 |
|
| 18 |
This model is for research and development only.
|
| 19 |
|
|
|
|
| 13 |
# Model Overview
|
| 14 |
|
| 15 |
Nemotron-4-Mini-Hindi-4B-Base is a base model pre-trained on Hindi and English corpus. The Nemotron-Mini-4B-Base (Minitron-4B) is subject to continuous pre-training using Hindi and English data (400B tokens) exclusively to create a strong base model for Hindi, English, and Hinglish. We make extensive use of synthetic data during the continuous pre-training stage. The base small language model (SLM) is optimized through distillation, pruning, and quantization for speed and on-device deployment.
|
| 16 |
+
Please refer to our [arXiv paper](https://arxiv.org/abs/2410.14815) for more details. The intruction tuned version of this model is [nvidia/Nemotron-4-Mini-Hindi-4B-Instruct](https://huggingface.co/nvidia/Nemotron-4-Mini-Hindi-4B-Instruct).
|
| 17 |
|
| 18 |
This model is for research and development only.
|
| 19 |
|