Together AI vs Anthropic Pricing (2026)
How do these two stack up on price? Here's what each one costs, what you get, and where the value sits.
| Together AI | Anthropic | |
|---|---|---|
| Starts at | Custom | $20/mo |
| Number of plans | 7 | 3 |
| Free plan | — | |
| Free trial | — | — |
| Pricing model | usage-based | flat-rate |
Serverless Inference
- Chat models (price per 1M tokens)
- Vision models (price per 1M tokens)
- Image generation models (price per image or per mp)
- Audio / TTS models (price per 1M characters)
- Video generation models (price per video)
- Transcription models (price per audio minute)
- Embeddings models (price per 1M tokens)
- Moderation models (price per 1M tokens)
- Batch API pricing available
Provisioned Throughput
- Reserved dedicated capacity in throughput units (PTUs)
- Fixed capacity per PTU with model-dependent tokens-per-minute
- MiniMax M3 support
- GLM-5.2 support
Dedicated Inference
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
- NVIDIA HGX H100 on-demand ($5.49/GPU/hr)
- NVIDIA HGX B200 on-demand ($8.99/GPU/hr)
- NVIDIA HGX H200 (contact us)
- NVIDIA HGX B300 (contact us)
- NVIDIA GB200 NVL72 (contact us)
- NVIDIA GB300 NVL72 (contact us)
- Reserved capacity available
GPU Clusters
- On-demand pay-as-you-go GPU capacity (hourly)
- Reserved capacity (7-30 days, 31-90 days, 91-180 days, 181+ days)
- NVIDIA HGX H100 on-demand ($3.99/GPU/hr)
- NVIDIA HGX H200 on-demand ($5.99/GPU/hr)
- NVIDIA HGX B200 on-demand ($8.19/GPU/hr)
- NVIDIA GB200 NVL72 (contact us)
- NVIDIA GB300 NVL72 (contact us)
Sandbox
- Code Sandbox: VM sandboxes for large development environments ($0.0446/vCPU/hr, $0.0149/GiB RAM/hr)
- Code Interpreter: Execute LLM-generated code securely via API ($0.03/session)
Storage
- Shared Filesystem ($0.16/GiB/month)
- High-bandwidth, parallel filesystem colocated with compute
Fine-Tuning
- Supervised Fine-Tuning (LoRA)
- Supervised Fine-Tuning (Full Fine-Tuning)
- Direct Preference Optimization (LoRA)
- Direct Preference Optimization (Full Fine-Tuning)
- Standard pricing for models up to 100B parameters
- Specialized pricing for DeepSeek, GLM, Kimi, Llama 4, Qwen3, and other large models
- Minimum charge per job: $4.00 (standard)
Free
- Chat on web, iOS, Android, and on your desktop
- Generate code and visualize data
- Write, edit, and create content
- Ability to search the web
- Memory across conversations
- Create files and execute code
- Unlock more from Claude with desktop extensions
- Connect Slack and Google Workspace services
- Integrate any context or tool through connectors with remote MCP
- Extended thinking for complex work
Pro
- More usage*
- Includes Claude Code
- Includes Claude Cowork
- Includes Claude Design
- Includes Claude Science
- Access to unlimited projects to organize chats and documents
- Access to Research
- Ability to use more Claude models
- Claude for Microsoft 365
Max
- Choose 5x or 20x more usage than Pro*
- Higher output limits for all tasks
- Early access to advanced Claude features
- Priority access at high traffic times
Together AI vs Anthropic FAQ
- Which one is cheaper?
- One or both use custom pricing, so it depends on your specific needs.
- Can I use either one for free?
- Anthropic has a free plan. Together AI doesn't — you'll need to pay from day one.
- How do they charge?
- Different approach here. Together AI uses usage-based pricing, while Anthropic goes with flat-rate. That changes the math depending on your team size and usage.
- Which one is a better deal?
- Depends on what you need. Together AI: Together AI is pitching itself as the serious infrastructure layer for AI builders — not the cheapest (Replicate or self-hosting can undercut on specific models), but more flexible and enterprise-ready than most managed inference APIs. At $5.49/GPU/hr for H100 dedicated, they're competitive with CoreWeave and Lambda Labs while bundling more managed services around it. Anthropic: Anthropic is matching OpenAI's ChatGPT Plus almost dollar-for-dollar at the Pro tier — this is a deliberate 'same price, better model' play targeting GPT defectors. The $100 Max tier signals they're comfortable competing with Copilot and Gemini Advanced for enterprise-adjacent power users without discounting to win.
Keep tabs on both.
We'll monitor pricing changes for Together AI and Anthropic and let you know when something moves.