Models / Qwen
Chat
Reasoning

Qwen3 235B A22B FP8 Throughput

Hybrid instruct + reasoning model (232Bx22B MoE) optimized for high-throughput, cost-efficient inference and distillation.

This model is not available on Together’s Serverless API.

Related models
  • Model provider
    Qwen
  • Type
    Chat
    Reasoning
  • Main use cases
    Chat
    Reasoning
    Medium General Purpose
    Function Calling
  • Features
    Function Calling
  • Parameters
    235.1B
  • Context length
    40k