Anthropic vs Together AI Pricing (2026)

How do these two stack up on price? Here's what each one costs, what you get, and where the value sits.

Anthropic Together AI
Starts at $20/mo Custom
Number of plans 3 7
Free plan
Free trial
Pricing model flat-rate usage-based

Free

$0/mo
  • Chat on web, iOS, Android, and on your desktop
  • Generate code and visualize data
  • Write, edit, and create content
  • Ability to search the web
  • Memory across conversations
  • Create files and execute code
  • Unlock more from Claude with desktop extensions
  • Connect Slack and Google Workspace services
  • Integrate any context or tool through connectors with remote MCP
  • Extended thinking for complex work

Pro

$20/mo
  • More usage*
  • Includes Claude Code
  • Includes Claude Cowork
  • Includes Claude Design
  • Includes Claude Science
  • Access to unlimited projects to organize chats and documents
  • Access to Research
  • Ability to use more Claude models
  • Claude for Microsoft 365

Max

$100/mo
  • Choose 5x or 20x more usage than Pro*
  • Higher output limits for all tasks
  • Early access to advanced Claude features
  • Priority access at high traffic times

Serverless Inference

Custom
  • Chat models (price per 1M tokens)
  • Vision models (price per 1M tokens)
  • Image generation models (price per image or per mp)
  • Audio / TTS models (price per 1M characters)
  • Video generation models (price per video)
  • Transcription models (price per audio minute)
  • Embeddings models (price per 1M tokens)
  • Moderation models (price per 1M tokens)
  • Batch API pricing available

Provisioned Throughput

Custom
  • Reserved dedicated capacity in throughput units (PTUs)
  • Fixed capacity per PTU with model-dependent tokens-per-minute
  • MiniMax M3 support
  • GLM-5.2 support

Dedicated Inference

Custom
  • Guaranteed performance (no sharing)
  • Support for custom models
  • Autoscaling & traffic spike handling
  • NVIDIA HGX H100 on-demand ($5.49/GPU/hr)
  • NVIDIA HGX B200 on-demand ($8.99/GPU/hr)
  • NVIDIA HGX H200 (contact us)
  • NVIDIA HGX B300 (contact us)
  • NVIDIA GB200 NVL72 (contact us)
  • NVIDIA GB300 NVL72 (contact us)
  • Reserved capacity available

GPU Clusters

Custom
  • On-demand pay-as-you-go GPU capacity (hourly)
  • Reserved capacity (7-30 days, 31-90 days, 91-180 days, 181+ days)
  • NVIDIA HGX H100 on-demand ($3.99/GPU/hr)
  • NVIDIA HGX H200 on-demand ($5.99/GPU/hr)
  • NVIDIA HGX B200 on-demand ($8.19/GPU/hr)
  • NVIDIA GB200 NVL72 (contact us)
  • NVIDIA GB300 NVL72 (contact us)

Sandbox

Custom
  • Code Sandbox: VM sandboxes for large development environments ($0.0446/vCPU/hr, $0.0149/GiB RAM/hr)
  • Code Interpreter: Execute LLM-generated code securely via API ($0.03/session)

Storage

Custom
  • Shared Filesystem ($0.16/GiB/month)
  • High-bandwidth, parallel filesystem colocated with compute

Fine-Tuning

Custom
  • Supervised Fine-Tuning (LoRA)
  • Supervised Fine-Tuning (Full Fine-Tuning)
  • Direct Preference Optimization (LoRA)
  • Direct Preference Optimization (Full Fine-Tuning)
  • Standard pricing for models up to 100B parameters
  • Specialized pricing for DeepSeek, GLM, Kimi, Llama 4, Qwen3, and other large models
  • Minimum charge per job: $4.00 (standard)

Anthropic vs Together AI FAQ

Which one is cheaper?
One or both use custom pricing, so it depends on your specific needs.
Can I use either one for free?
Anthropic has a free plan. Together AI doesn't — you'll need to pay from day one.
How do they charge?
Different approach here. Anthropic uses flat-rate pricing, while Together AI goes with usage-based. That changes the math depending on your team size and usage.
Which one is a better deal?
Depends on what you need. Anthropic: Anthropic is matching OpenAI's ChatGPT Plus almost dollar-for-dollar at the Pro tier — this is a deliberate 'same price, better model' play targeting GPT defectors. The $100 Max tier signals they're comfortable competing with Copilot and Gemini Advanced for enterprise-adjacent power users without discounting to win. Together AI: Together AI is pitching itself as the serious infrastructure layer for AI builders — not the cheapest (Replicate or self-hosting can undercut on specific models), but more flexible and enterprise-ready than most managed inference APIs. At $5.49/GPU/hr for H100 dedicated, they're competitive with CoreWeave and Lambda Labs while bundling more managed services around it.

Keep tabs on both.

We'll monitor pricing changes for Anthropic and Together AI and let you know when something moves.

Start tracking free »