Twilight-Large

This is a merge of pre-trained language models created using mergekit by @softwareweaver. Use the prompt format that Mistral Large uses.

GGUF Quants

https://huggingface.co/mradermacher/Twilight-Large-123B-GGUF

https://huggingface.co/mradermacher/Twilight-Large-123B-i1-GGUF

Use --chat-template llama2 when using llama.cpp

Control Vectors

You can use Control Vectors for Mistral Large https://huggingface.co/jukofyork/creative-writing-control-vectors-v3.0/tree/main/Mistral-Large-Instruct-2407

Control vectors allow fine-tuned control over LLMs, enabling more precise/targeted text generation. More info https://huggingface.co/jukofyork/creative-writing-control-vectors-v3.0

Sample Generations

Some sample generations https://huggingface.co/softwareweaver/Twilight-Large-123B/discussions

Please add your own generations to the community tab. This allows others to evaluate the model outputs before downloading it.

Merge Details

Merge Method

This model was merged using the della_linear merge method using mistralai/Mistral-Large-Instruct-2407 as a base.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

models:
  - model: TheDrummer/Behemoth-123B-v1
    parameters:
      weight: 0.25
      density: 0.9
  - model: schnapper79/lumikabra-123B_v0.4
    parameters:
      weight: 0.3
      density: 0.9      
merge_method: della_linear
base_model: mistralai/Mistral-Large-Instruct-2407
parameters:
  epsilon: 0.05
  lambda: 1
  int8_mask: true
dtype: bfloat16

Kotokin
/

softwareweaver_Twilight-Large-123B-exl2-4bpw