Models / Mistral AI
LLM
Chat

Mistral

7.3B model surpassing Llama 2 13B, nearing CodeLlama 7B on code, with GQA for speed and SWA for efficient long-sequence handling.

This model is not available on Together’s Serverless API.

Pick a supported alternative from the Model Library.

Related models
  • Model provider
    Mistral AI
  • Type
    LLM
    Chat
  • Main use cases
    Chat
    Small & Fast
  • Parameters
    7B
  • Context length
    8192
  • Quantization level
    FP16
  • Category
    Chat