Mixtral-8x7B-Instruct-v0.1

View as Markdown

Mixtral-8x7B-Instruct-v0.1 has a checked-in NeMo AutoModel recipe for text generation. The Hugging Face configuration declares the MixtralForCausalLM architecture.

Fine-Tune Mixtral-8x7B-Instruct-v0.1

Follow the installation instructions, then run the recipe from the repository root:

automodel examples/llm_finetune/mistral/mixtral-8x7b-v0-1_squad_peft.yaml --nproc-per-node 8

Use the Slurm launcher guide for the multi-node run.

Choose a Workflow

GoalStart Here
Run the primary recipe for this modelUse the recipe configuration.
Try another checked-in recipe for this modelUse the alternate recipe.

Configuration

SettingConfiguration
Hardware-
StrategyFSDP2; tp_size=4, pp_size=1, cp_size=1
Nodes2
FeaturesCheckpointing (enabled=true), LoRA
Advancedbf16

More Recipes for This Model

Model Reference

Model Architecture

PropertyValue
TaskText generation
ArchitectureMixtralForCausalLM

Available Models