llama-duo/synth_summarize_dataset
Viewer • Updated • 903k • 238 • 5
How to use llama-duo/gemma7b-summarize-gemini1.5flash-80k with PEFT:
from peft import PeftModel
from transformers import AutoModelForCausalLM
base_model = AutoModelForCausalLM.from_pretrained("google/gemma-7b")
model = PeftModel.from_pretrained(base_model, "llama-duo/gemma7b-summarize-gemini1.5flash-80k")This model is a fine-tuned version of google/gemma-7b on the llama-duo/synth_summarize_dataset dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.8742 | 0.9982 | 280 | 2.1938 |
| 0.7213 | 2.0 | 561 | 2.1462 |
| 0.675 | 2.9982 | 841 | 2.1484 |
| 0.6439 | 4.0 | 1122 | 2.2149 |
| 0.569 | 4.9982 | 1402 | 2.3224 |
| 0.5317 | 6.0 | 1683 | 2.4839 |
| 0.472 | 6.9982 | 1963 | 2.6540 |
| 0.4306 | 8.0 | 2244 | 2.8791 |
| 0.4106 | 8.9982 | 2524 | 3.0011 |
| 0.4021 | 9.9822 | 2800 | 3.0229 |
Base model
google/gemma-7b