🚀 DeepSeek V4 Pro 0813 vs. GPT-5.6 Sol on DeepSWE →

🤝 Together AI & Y Combinator announce partnership to deliver the first dedicated YC GPU cluster →

⚡ On-demand B200s now available on Together GPU Clusters →

🚀 Now serving MiniMax-M3 for efficient inference →

  • Inference

    • Serverless Inference

      High-performance inference as APIs

    • Batch Inference

      Inference for batch workloads

    • Provisioned Throughput

      Token-based capacity with SLAs

    • Dedicated Model Inference

      Inference on custom hardware

    • Dedicated Container Inference

      Inference for custom models

    MiniMax M3
    Gemma 4 31B
    DeepSeek V4 Pro
    GLM-5.2
    White circular shape with uneven edges and three extended finger-like projections on a black background.
    kimi K2.7 Code
    OpenAI logo with a symmetrical abstract geometric knot design.
    gpt-oss-120B

    Model library

    Explore the top open-source models

  • Compute

    Accelerated Compute

    • GPU Clusters

      Reliable GPU clusters at scale

    • AI Factory

      Custom infrastructure at frontier scale

    Developer Environments

    • Sandbox

      Build development environments for AI

    Storage

    • Managed Storage

      Store model weights & data securely

    • GB300

    • GB200

    • B200

    • H200

    • H100

  • Model Shaping

    • Custom Training

      From reinforcement learning to full control

    • Fine-Tuning

      Shape models with your data

    • Evaluations

      Measure model quality

    kimi K2.7 Code
    Gemma 4 31B-it FP8
    GLM 5.1 FP4
    OpenAI logo with a symmetrical abstract geometric knot design.
    gpt-oss-120b
    Qwen3.5 397B A17b
    Llama 4 Maverick

    Model library

    Fine-tune top open-source models

  • Research

    • Research

      Systems research for production AI

    • Research blog

      All our research publications

    Featured publications

    • FlashAttention

    • ATLAS

    • Kernel Collection

    • ThunderKittens

    • DSGym

    Show all
  • Developers

    • Documentation

      Technical docs for Together AI

    • Demos

      Our open-source demo apps

    • Cookbooks

      Practical implementation guides

    • Voice Agents

      Build voice agents for production

    • Open-source AI

      Build better with open models

    • Model Library

    • Playground

    • Together Chat

    • Which LLM to use

    • Open-source ROI calculator

  • Company

    Resources

    • Customer stories

      Testimonials from AI Natives

    • Startup accelerator

      Build and scale your startup

    • Customer support

      Find answers to your questions

    • Blog

      Our latest news & blog posts

    • Events

      Explore our events calendar

    Company

    • About

      Get to know us

    • Careers

      Join our mission

    • Press

      Together in the news

  • Pricing

    • Serverless Inference

      High-performance inference as APIs

    • Batch Inference

      Inference for batch workloads

    • Provisioned Throughput

      Token-based capacity with SLAs

    • Dedicated Model Inference

      Inference on custom hardware

    • Dedicated Container Inference

      Inference for custom models

    MiniMax M3
    Gemma 4 31B
    DeepSeek V4 Pro
    GLM-5.2
    White circular shape with uneven edges and three extended finger-like projections on a black background.
    kimi K2.7 Code
    OpenAI logo with a symmetrical abstract geometric knot design.
    gpt-oss-120B

    Model library

    Explore the top open-source models

  • Accelerated Compute

    • GPU Clusters

      Reliable GPU clusters at scale

    • AI Factory

      Custom infrastructure at frontier scale

    Developer Environments

    • Sandbox

      Build development environments for AI

    Storage

    • Managed Storage

      Store model weights & data securely

    • GB300

    • GB200

    • B200

    • H200

    • H100

    • Custom Training

      From reinforcement learning to full control

    • Fine-Tuning

      Shape models with your data

    • Evaluations

      Measure model quality

    kimi K2.7 Code
    Gemma 4 31B-it FP8
    GLM 5.1 FP4
    OpenAI logo with a symmetrical abstract geometric knot design.
    gpt-oss-120b
    Qwen3.5 397B A17b
    Llama 4 Maverick

    Model library

    Fine-tune top open-source models

    • Research

      Systems research for production AI

    • Research blog

      All our research publications

    Featured publications

    • FlashAttention

    • ATLAS

    • Kernel Collection

    • ThunderKittens

    • DSGym

    Show all
    • Documentation

      Technical docs for Together AI

    • Demos

      Our open-source demo apps

    • Cookbooks

      Practical implementation guides

    • Voice Agents

      Build voice agents for production

    • Open-source AI

      Build better with open models

    • Model Library

    • Playground

    • Together Chat

    • Which LLM to use

    • Open-source ROI calculator

  • Resources

    • Customer stories

      Testimonials from AI Natives

    • Startup accelerator

      Build and scale your startup

    • Customer support

      Find answers to your questions

    • Blog

      Our latest news & blog posts

    • Events

      Explore our events calendar

    Company

    • About

      Get to know us

    • Careers

      Join our mission

    • Press

      Together in the news

Contact sales
Contact sales
Sign in
Explore Research

Research blog

All
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
Inference
AdapTive-LeArning Speculator System (ATLAS): A New Paradigm in LLM Inference via Runtime-Learning Accelerators

ATLAS delivers up to 4x faster LLM inference, powered by Together Turbo’s latest research.

Junxiong Wang, Shirley Wu, Zelei Shao, Vikranth Srivatsa, Jue Wang, Roy Yuan, Qingyang Wu, Alpay Ariyak, Rupert Wu, Wai Tong Chung, Chenfeng Xu, Yonatan Oren, Pragaash Ponnusamy, Yineng Zhang, Avner May, Leon Song, Tri Dao, Percy Liang, Ce Zhang, Ben Athiwaratkun

Blue circles with letters A T L A S connected with arrows on a globe grid background.
Agents
How Together AI Uses AI Agents to Automate Complex Engineering Tasks: Lessons from Developing Efficient LLM Inference Systems

Shang Zhu, Federico Bianchi, Wai Tong Chung, Zain Hasan, Rupert Wu, Ce Zhang, James Zou, Ben Athiwaratkun

Website banner: How Together AI uses AI agents to automate complex engineering tasks with a Learn More button.
Agents
Back to The Future: Evaluating AI Agents on Predicting Future Events

Federico Bianchi, Junlin Wang, Zain Hasan, Shang Zhu, Roy Yuan, Clémentine Fourrier, James Zou

Retro futuristic scene with a DeLorean car, neon grid floor, mountains, and a striped glowing sun.
Inference
DeepSWE: Training a Fully Open-sourced, State-of-the-Art Coding Agent by Scaling RL

Michael Luo*, Naman Jain*, Jaskirat Singh*, Sijun Tan*, Ameen Patel*, Qingyang Wu*, Alpay Ariyak*, Colin Cai*, Tarun Venkat, Shang Zhu, Ben Athiwaratkun, Manan Roongta, Ce Zhang, Li Erran Li, Raluca Ada Popa, Koushik Sen, Ion Stoica

Chart showing SWE-Bench performance vs model size for various models with DeepSWE-Preview + TTS leading at 59%.
Previous
Load more
7 / 20

No search result

Try expanding your search or changing the filters.

Be at the forefront of AI innovation

From optimized training and model shaping to large-scale production inference

See open roles
  • Products

    • Accelerated Compute

    • Serverless Inference

    • Provisioned Throughput

    • Dedicated Inference

    • Fine-Tuning

    • Sandbox

    • Evaluations

  • Models

    See all models

    DeepSeek

    Meta

    Qwen

    Google

    OpenAI

    Mistral AI

    Custom models

  • Developers

    • Research

    • Docs

    • Open-source AI

    • OSS ROI calculator

    Pricing

    • Pricing overview

    • Inference

    • Fine-Tuning

    • GPU Clusters

  • Resources

    • Blog

    • About us

    • Careers

    • Customer Stories

    • Support

  • Privacy Policy

  • Terms of service

  • Cookie Policy

  • Consent Preferences

© 2026 Together AI. All Rights Reserved.