Model save

Browse files

Files changed (4) hide show

README.md +61 -0
adapter_config.json +2 -2
adapter_model.safetensors +1 -1
training_args.bin +1 -1

README.md ADDED Viewed

	@@ -0,0 +1,61 @@

+---
+library_name: peft
+license: llama3.2
+base_model: meta-llama/Llama-3.2-11B-Vision-Instruct
+tags:
+- trl
+- sft
+- generated_from_trainer
+model-index:
+- name: fine-tuned-visionllama_5
+  results: []
+---
+<!-- This model card has been generated automatically according to the information the Trainer had access to. You
+should probably proofread and complete it, then remove this comment. -->
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/august-gislerud-rolfsen-relu/Dummy/runs/9mfryeev)
+# fine-tuned-visionllama_5
+This model is a fine-tuned version of [meta-llama/Llama-3.2-11B-Vision-Instruct](https://huggingface.co/meta-llama/Llama-3.2-11B-Vision-Instruct) on an unknown dataset.
+It achieves the following results on the evaluation set:
+- eval_loss: 2.0710
+- eval_runtime: 8.0583
+- eval_samples_per_second: 0.62
+- eval_steps_per_second: 0.124
+- epoch: 0.0489
+- step: 425
+## Model description
+More information needed
+## Intended uses & limitations
+More information needed
+## Training and evaluation data
+More information needed
+## Training procedure
+### Training hyperparameters
+The following hyperparameters were used during training:
+- learning_rate: 0.0002
+- train_batch_size: 1
+- eval_batch_size: 8
+- seed: 42
+- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
+- lr_scheduler_type: constant
+- lr_scheduler_warmup_ratio: 0.03
+- num_epochs: 2
+### Framework versions
+- PEFT 0.13.0
+- Transformers 4.45.1
+- Pytorch 2.2.2+cu121
+- Datasets 3.0.1
+- Tokenizers 0.20.3

adapter_config.json CHANGED Viewed

@@ -20,8 +20,8 @@
   "rank_pattern": {},
   "revision": null,
   "target_modules": [
-    "q_proj",
-    "v_proj"
   ],
   "task_type": "CAUSAL_LM",
   "use_dora": false,

   "rank_pattern": {},
   "revision": null,
   "target_modules": [
+    "v_proj",
+    "q_proj"
   ],
   "task_type": "CAUSAL_LM",
   "use_dora": false,

adapter_model.safetensors CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:141b7dc1a732738e873a9bb6e26889fe9b39efc1e08ea9a4ca4ceadefd393aed
 size 23641256

 version https://git-lfs.github.com/spec/v1
+oid sha256:78584636ede9a36a020bf4fc93240f002220daa3dbc2779147eb7a39f187edc8
 size 23641256

training_args.bin CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:05f741cc6ae801095ee7b66243e5dba82cf5e8ae4c0e6b612bcc4b98a093ce79
 size 5496

 version https://git-lfs.github.com/spec/v1
+oid sha256:e540f97775b099d85753089c65f68a2c4a692045a369e17ac2a4d5edccf54f79
 size 5496