Remove Calibration Dataset section
Browse files
README.md
CHANGED
|
@@ -77,22 +77,6 @@ This model is NVFP4 quantized with nvidia-modelopt **v0.44.0**
|
|
| 77 |
|
| 78 |
## Training and Evaluation Datasets:
|
| 79 |
|
| 80 |
-
## Calibration Dataset:
|
| 81 |
-
**Links:** Various datasets from [NVIDIA's Nemotron Post-Training V3 Collection](https://huggingface.co/collections/nvidia/nemotron-post-training-v3) are used to train this model. Both the prompts and synthetically-generated responses in those datasets are used as-is. The following datasets are used:
|
| 82 |
-
- [nvidia/Nemotron-Science-v1](https://huggingface.co/datasets/nvidia/Nemotron-Science-v1)
|
| 83 |
-
- [nvidia/Nemotron-Instruction-Following-Chat-v1](https://huggingface.co/datasets/nvidia/Nemotron-Instruction-Following-Chat-v1)
|
| 84 |
-
- [nvidia/Nemotron-Competitive-Programming-v1](https://huggingface.co/datasets/nvidia/Nemotron-Competitive-Programming-v1)
|
| 85 |
-
- [nvidia/Nemotron-Math-v2](https://huggingface.co/datasets/nvidia/Nemotron-Math-v2)
|
| 86 |
-
- [nvidia/Nemotron-SFT-SWE-v2](https://huggingface.co/datasets/nvidia/Nemotron-SFT-SWE-v2)
|
| 87 |
-
- [nvidia/Nemotron-SFT-Multilingual-v1](https://huggingface.co/datasets/nvidia/Nemotron-SFT-Multilingual-v1)
|
| 88 |
-
- [nvidia/Nemotron-SFT-Agentic-v2](https://huggingface.co/datasets/nvidia/Nemotron-SFT-Agentic-v2)
|
| 89 |
-
- [nvidia/Nemotron-SFT-Instruction-Following-Chat-v2](https://huggingface.co/datasets/nvidia/Nemotron-SFT-Instruction-Following-Chat-v2)
|
| 90 |
-
|
| 91 |
-
**Data Modality:** Text<br>
|
| 92 |
-
**Data Collection Method by Dataset:** Hybrid: Synthetic, Human, Automated<br>
|
| 93 |
-
**Labeling Method by Dataset:** Hybrid: Synthetic, Human, Automated<br>
|
| 94 |
-
**Properties:** The Nemotron Post-Training V3 Collection datasets are post-training datasets curated by NVIDIA containing multi-turn conversations across diverse topics. Total of ~2.9M samples, majority synthetic, others sourced from commercially-friendly datasets.
|
| 95 |
-
|
| 96 |
### Training Dataset
|
| 97 |
**Data Modality:** Text, Image, Video<br>
|
| 98 |
**Image Training Data Size:** Undisclosed<br>
|
|
|
|
| 77 |
|
| 78 |
## Training and Evaluation Datasets:
|
| 79 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 80 |
### Training Dataset
|
| 81 |
**Data Modality:** Text, Image, Video<br>
|
| 82 |
**Image Training Data Size:** Undisclosed<br>
|