--- base_model: - LatitudeGames/Wayfarer-Large-70B-Llama-3.3 - marcelbinz/Llama-3.1-Centaur-70B - flammenai/Mahou-1.5-llama3.1-70B - EVA-UNIT-01/EVA-LLaMA-3.33-70B-v0.0 - zerofata/L3.3-GeneticLemonade-Unleashed-v3-70B - deepcogito/cogito-v2-preview-llama-70B - tdrussell/Llama-3-70B-Instruct-Storywriter - Ppoyaa/MythoNemo-L3.1-70B-v1.0 - Blackroot/Mirai-3.0-70B - Sao10K/70B-L3.3-mhnnn-x1 - TheDrummer/Fallen-Llama-3.3-70B-v1 - nvidia/Llama-3.1-Nemotron-70B-Instruct-HF - TheDrummer/Anubis-70B-v1.1 - Doctor-Shotgun/L3.3-70B-Magnum-Diamond - Black-Ink-Guild/Pernicious_Prophecy_70B - watt-ai/watt-tool-70B - ReadyArt/Forgotten-Safeword-70B-v5.0 - nbeerbower/Llama3.1-Gutenberg-Doppel-70B library_name: transformers tags: - mergekit - merge --- # GoldDiamondGold-70b ![image/png](https://cdn-uploads.huggingface.co/production/uploads/633e85093a17ab61de8d9073/8BCqqt8KoN6DUrjHJXpi0.png) This is a merge of pre-trained language models created using [mergekit](https://github.com/cg123/mergekit). ## Motivation Inspired from the sapphire model by BruhzWater, the base for this is cogito. Rest of the models is what I think would be good for a SCE attempt. Basically combining all the best bits in each model and see what I get out of this. There's 3 interesting models that I highlighted in the previous merge that went into this too. ## Vibes Seems OK. I think it's better than the previous model. That felt super sloppy. ## Merge Details ### Merge Method This model was merged using the [SCE](https://arxiv.org/abs/2408.07990) merge method using [deepcogito/cogito-v2-preview-llama-70B](https://huggingface.co/deepcogito/cogito-v2-preview-llama-70B) as a base. ### Models Merged The following models were included in the merge: * [LatitudeGames/Wayfarer-Large-70B-Llama-3.3](https://huggingface.co/LatitudeGames/Wayfarer-Large-70B-Llama-3.3) * [marcelbinz/Llama-3.1-Centaur-70B](https://huggingface.co/marcelbinz/Llama-3.1-Centaur-70B) * [flammenai/Mahou-1.5-llama3.1-70B](https://huggingface.co/flammenai/Mahou-1.5-llama3.1-70B) * [EVA-UNIT-01/EVA-LLaMA-3.33-70B-v0.0](https://huggingface.co/EVA-UNIT-01/EVA-LLaMA-3.33-70B-v0.0) * [zerofata/L3.3-GeneticLemonade-Unleashed-v3-70B](https://huggingface.co/zerofata/L3.3-GeneticLemonade-Unleashed-v3-70B) * [tdrussell/Llama-3-70B-Instruct-Storywriter](https://huggingface.co/tdrussell/Llama-3-70B-Instruct-Storywriter) * [Ppoyaa/MythoNemo-L3.1-70B-v1.0](https://huggingface.co/Ppoyaa/MythoNemo-L3.1-70B-v1.0) * [Blackroot/Mirai-3.0-70B](https://huggingface.co/Blackroot/Mirai-3.0-70B) * [Sao10K/70B-L3.3-mhnnn-x1](https://huggingface.co/Sao10K/70B-L3.3-mhnnn-x1) * [TheDrummer/Fallen-Llama-3.3-70B-v1](https://huggingface.co/TheDrummer/Fallen-Llama-3.3-70B-v1) * [nvidia/Llama-3.1-Nemotron-70B-Instruct-HF](https://huggingface.co/nvidia/Llama-3.1-Nemotron-70B-Instruct-HF) * [TheDrummer/Anubis-70B-v1.1](https://huggingface.co/TheDrummer/Anubis-70B-v1.1) * [Doctor-Shotgun/L3.3-70B-Magnum-Diamond](https://huggingface.co/Doctor-Shotgun/L3.3-70B-Magnum-Diamond) * [Black-Ink-Guild/Pernicious_Prophecy_70B](https://huggingface.co/Black-Ink-Guild/Pernicious_Prophecy_70B) * [watt-ai/watt-tool-70B](https://huggingface.co/watt-ai/watt-tool-70B) * [ReadyArt/Forgotten-Safeword-70B-v5.0](https://huggingface.co/ReadyArt/Forgotten-Safeword-70B-v5.0) * [nbeerbower/Llama3.1-Gutenberg-Doppel-70B](https://huggingface.co/nbeerbower/Llama3.1-Gutenberg-Doppel-70B) ### Configuration The following YAML configuration was used to produce this model: ```yaml models: # Mirai is Mirai. - model: Blackroot/Mirai-3.0-70B # Narration - model: EVA-UNIT-01/EVA-LLaMA-3.33-70B-v0.0 # Claude 3 Sonnet/Opus prose style and quality - model: Doctor-Shotgun/L3.3-70B-Magnum-Diamond # "For the *Action*"? - model: marcelbinz/Llama-3.1-Centaur-70B # Better writing style, "creativity" shift. (fiction books) - model: tdrussell/Llama-3-70B-Instruct-Storywriter # Roleplaying and Story Writing - model: Ppoyaa/MythoNemo-L3.1-70B-v1.0 # Sao10K - model: Sao10K/70B-L3.3-mhnnn-x1 # Medical - model: Black-Ink-Guild/Pernicious_Prophecy_70B # Dialogue reinforcement - model: LatitudeGames/Wayfarer-Large-70B-Llama-3.3 # Extra details - model: TheDrummer/Anubis-70B-v1.1 # "Meanness" - model: TheDrummer/Fallen-Llama-3.3-70B-v1 # Antique history - model: nbeerbower/Llama3.1-Gutenberg-Doppel-70B # Normalization? - model: nvidia/Llama-3.1-Nemotron-70B-Instruct-HF # Doomer tilt (From negative llama) + "Taboo"(?) - model: ReadyArt/Forgotten-Safeword-70B-v5.0 # ERP/RP enhancement + Anime tilt - model: zerofata/L3.3-GeneticLemonade-Unleashed-v3-70B # Tool Calling - model: watt-ai/watt-tool-70B # Short, casual dialogue (Anime tilt) - model: flammenai/Mahou-1.5-llama3.1-70B merge_method: sce base_model: deepcogito/cogito-v2-preview-llama-70B select_topk: 0.33 parameters: normalize: true dtype: bfloat16 ```