Delta-Vector
/

Baldur-8B

PyTorch

English

llama

chat

Eval Results

Model card Files Files and versions Community

Delta-Vector commited on Oct 5

Commit

d35f2b4

•

1 Parent(s): deceecb

Update README.md

Browse files

Files changed (1) hide show

README.md +89 -59

README.md CHANGED Viewed

@@ -1,21 +1,94 @@
 ---
-library_name: transformers
-license: llama3
-base_model: arcee-ai/Llama-3.1-SuperNova-Lite
 tags:
-- generated_from_trainer
-model-index:
-- name: henbane-8b-r3
-  results: []
 ---
-<!-- This model card has been generated automatically according to the information the Trainer had access to. You
-should probably proofread and complete it, then remove this comment. -->
-[<img src="https://raw.githubusercontent.com/axolotl-ai-cloud/axolotl/main/image/axolotl-badge-web.png" alt="Built with Axolotl" width="200" height="32"/>](https://github.com/axolotl-ai-cloud/axolotl)
 <details><summary>See axolotl config</summary>
-axolotl version: `0.4.1`
 ```yaml
 base_model: arcee-ai/Llama-3.1-SuperNova-Lite
 model_type: AutoModelForCausalLM
@@ -57,8 +130,6 @@ datasets:
     type: chat_template
   - path: anthracite-org/kalo_misc_part2
     type: chat_template
-  - path: anthracite-org/kalo_misc_part2
-    type: chat_template
   - path: Nitral-AI/Creative_Writing-ShareGPT
     type: chat_template
   - path: NewEden/Gryphe-Sonnet3.5-Charcard-Roleplay-unfiltered
@@ -129,54 +200,13 @@ special_tokens:
   eos_token: <|eot_id|>
 ```
 </details><br>
-# henbane-8b-r3
-This model is a fine-tuned version of [arcee-ai/Llama-3.1-SuperNova-Lite](https://huggingface.co/arcee-ai/Llama-3.1-SuperNova-Lite) on the None dataset.
-## Model description
-More information needed
-## Intended uses & limitations
-More information needed
-## Training and evaluation data
-More information needed
-## Training procedure
-### Training hyperparameters
-The following hyperparameters were used during training:
-- learning_rate: 1e-05
-- train_batch_size: 1
-- eval_batch_size: 1
-- seed: 42
-- distributed_type: multi-GPU
-- num_devices: 2
-- gradient_accumulation_steps: 32
-- total_train_batch_size: 64
-- total_eval_batch_size: 2
-- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
-- lr_scheduler_type: cosine
-- lr_scheduler_warmup_steps: 5
-- num_epochs: 2
-### Training results
-### Framework versions
-- Transformers 4.45.0.dev0
-- Pytorch 2.4.0+cu121
-- Datasets 2.19.1
-- Tokenizers 0.19.1

 ---
+License: agpl-3.0
+Language:
+- En
+Pipeline_tag: text-generation
+Base_model: arcee-ai/Llama-3.1-SuperNova-Lite
+Tags:
+- Chat
+license: agpl-3.0
+datasets:
+- Gryphe/Sonnet3.5-SlimOrcaDedupCleaned
+- Nitral-AI/Cybersecurity-ShareGPT
+- Nitral-AI/Medical_Instruct-ShareGPT
+- Nitral-AI/Olympiad_Math-ShareGPT
+- anthracite-org/kalo_opus_misc_240827
+- NewEden/Claude-Instruct-5k
+- lodrick-the-lafted/kalo-opus-instruct-3k-filtered
+- anthracite-org/kalo-opus-instruct-22k-no-refusal
+- Epiculous/Synthstruct-Gens-v1.1-Filtered-n-Cleaned
+- Epiculous/SynthRP-Gens-v1.1-Filtered-n-Cleaned
+-  anthracite-org/kalo_misc_part2
+- Nitral-AI/Creative_Writing-ShareGPT
+- NewEden/Gryphe-Sonnet3.5-Charcard-Roleplay-unfiltered
 tags:
+- chat
+language:
+- en
+base_model:
+- arcee-ai/Llama-3.1-SuperNova-Lite
 ---
+![](https://huggingface.co/Delta-Vector/Baldur-8B/resolve/main/Baldur.jpg)
+An finetune of the L3.1 instruct distill done by Arcee, The intent of this model is to have differing prose then my other releases, in my testing it has achieved this and avoiding using common -isms frequently and has a differing flavor then my other models.
+# Quants
+GGUF: https://huggingface.co/Delta-Vector/Baldur-8B-GGUF
+EXL2: https://huggingface.co/Delta-Vector/Baldur-8B-EXL2
+## Prompting
+Model has been Instruct tuned with the Llama-Instruct formatting. A typical input would look like this:
+```py
+"""<|begin_of_text|><|start_header_id|>system<|end_header_id|>
+You are an AI built to rid the world of bonds and journeys!<|eot_id|><|start_header_id|>user<|end_header_id|>
+Bro i just wanna know what is 2+2?<|eot_id|><|start_header_id|>assistant<|end_header_id|>
+"""
+```
+## System Prompting
+I would highly recommend using Sao10k's Euryale System prompt, But the "Roleplay Simple" system prompt provided within SillyTavern will work aswell.
+```
+Currently, your role is {{char}}, described in detail below. As {{char}}, continue the narrative exchange with {{user}}.
+<Guidelines>
+• Maintain the character persona but allow it to evolve with the story.
+• Be creative and proactive. Drive the story forward, introducing plotlines and events when relevant.
+• All types of outputs are encouraged; respond accordingly to the narrative.
+• Include dialogues, actions, and thoughts in each response.
+• Utilize all five senses to describe scenarios within {{char}}'s dialogue.
+• Use emotional symbols such as "!" and "~" in appropriate contexts.
+• Incorporate onomatopoeia when suitable.
+• Allow time for {{user}} to respond with their own input, respecting their agency.
+• Act as secondary characters and NPCs as needed, and remove them when appropriate.
+• When prompted for an Out of Character [OOC:] reply, answer neutrally and in plaintext, not as {{char}}.
+</Guidelines>
+<Forbidden>
+• Using excessive literary embellishments and purple prose unless dictated by {{char}}'s persona.
+• Writing for, speaking, thinking, acting, or replying as {{user}} in your response.
+• Repetitive and monotonous outputs.
+• Positivity bias in your replies.
+• Being overly extreme or NSFW when the narrative context is inappropriate.
+</Forbidden>
+Follow the instructions in <Guidelines></Guidelines>, avoiding the items listed in <Forbidden></Forbidden>.
+```
+## Axolotl config
 <details><summary>See axolotl config</summary>
+Axolotl version: `0.4.1`
 ```yaml
 base_model: arcee-ai/Llama-3.1-SuperNova-Lite
 model_type: AutoModelForCausalLM
     type: chat_template
   - path: anthracite-org/kalo_misc_part2
     type: chat_template
   - path: Nitral-AI/Creative_Writing-ShareGPT
     type: chat_template
   - path: NewEden/Gryphe-Sonnet3.5-Charcard-Roleplay-unfiltered
   eos_token: <|eot_id|>
 ```
+## Credits
+Thank you to [Lucy Knada](https://huggingface.co/lucyknada), [Kalomaze](https://huggingface.co/kalomaze), [Kubernetes Bad](https://huggingface.co/kubernetes-bad) and the rest of [Anthracite](https://huggingface.co/anthracite-org) (But not Alpin.)
 </details><br>
+## Training
+The training was done for 2 epochs. I used  2 x [RTX 6000s](https://www.nvidia.com/en-us/design-visualization/rtx-6000/) GPUs graciously provided by [Kubernetes Bad](https://huggingface.co/kubernetes-bad) for the full-parameter fine-tuning of the model.
+[<img src="https://raw.githubusercontent.com/OpenAccess-AI-Collective/axolotl/main/image/axolotl-badge-web.png" alt="Built with Axolotl" width="200" height="32"/>](https://github.com/OpenAccess-AI-Collective/axolotl)