Providers / GPU cloud
Together AI
A platform for running open models through an inference API, with fine-tuning and dedicated GPU clusters for larger jobs.
- Type
- GPU cloud
- Tends to suit
- Teams that want to call open models by API first, and move to dedicated GPUs or fine-tuning later.
- How it is priced
- Priced through a network token or credits
- Getting access
- Self-serve API signup; sales contact for reserved capacity and large clusters.
What we checked
- Serverless inference is priced per million tokens, as shown on the pricing page. (source)
- The home page also lists GPU clusters, from self-serve instant clusters to larger deployments, and fine-tuning. (source)
Current prices
Prices change often, so we do not copy them. Go to the source.
See the provider's own pricing page
Written by ComputeFinder from the provider's own public pages on 2 October 2026. Not yet reviewed line by line by a second person. If something is wrong, tell us.