/ 5 models

    MIST

    Our flagship large language model family. Built with advanced merging (DARE-TIES, frankenmerging) and GRPO reasoning tuning on the Llama 3.1 lineage, spanning 8B, 70B, and 140B parameters, with 4-bit quantised releases for accessible deployment.

    MISTLABS

    MIST-Mini-8B

    General-purpose 8B merged LLM built on the Llama 3.1 lineage with advanced model merging.

    725 4
    View
    MISTLABS

    MIST-1-70B

    General-purpose 70B merged LLM built on the Llama 3.1 lineage with advanced model merging.

    547 1
    View
    MISTLABS

    MIST-1-140B

    General-purpose 140B frankenmerge LLM built on the Llama 3.1 lineage with advanced model merging.

    207 0
    View
    MISTLABS

    MIST-1-140B-4bit

    4-bit NF4 quantised release of MIST-1-140B: 140B-scale weights on accessible hardware.

    178 0
    View
    MISTLABS

    MIST-Mini-8B-Thinking

    Reasoning-tuned MIST model trained with GRPO to think step-by-step before answering.

    343 3
    View

    Explore other families

    More of the stack, organised the same way.