Together AI vs Avoma Pricing (2026)
How do these two stack up on price? Here's what each one costs, what you get, and where the value sits.
| Together AI | Avoma | |
|---|---|---|
| Starts at | Custom | $25/mo |
| Number of plans | 7 | 6 |
| Free plan | — | — |
| Free trial | — | |
| Pricing model | usage-based | per-seat |
Serverless Inference
- Chat models (price per 1M tokens)
- Vision models (price per 1M tokens)
- Image generation models (price per image or per mp)
- Audio / TTS models (price per 1M characters)
- Video generation models (price per video)
- Transcription models (price per audio minute)
- Embeddings models (price per 1M tokens)
- Moderation models (price per 1M tokens)
- Batch API pricing available
Provisioned Throughput
- Reserved dedicated capacity in throughput units (PTUs)
- Fixed capacity per PTU with model-dependent tokens-per-minute
- MiniMax M3 support
- GLM-5.2 support
Dedicated Inference
- Guaranteed performance (no sharing)
- Support for custom models
- Autoscaling & traffic spike handling
- NVIDIA HGX H100 on-demand ($5.49/GPU/hr)
- NVIDIA HGX B200 on-demand ($8.99/GPU/hr)
- NVIDIA HGX H200 (contact us)
- NVIDIA HGX B300 (contact us)
- NVIDIA GB200 NVL72 (contact us)
- NVIDIA GB300 NVL72 (contact us)
- Reserved capacity available
GPU Clusters
- On-demand pay-as-you-go GPU capacity (hourly)
- Reserved capacity (7-30 days, 31-90 days, 91-180 days, 181+ days)
- NVIDIA HGX H100 on-demand ($3.99/GPU/hr)
- NVIDIA HGX H200 on-demand ($5.99/GPU/hr)
- NVIDIA HGX B200 on-demand ($8.19/GPU/hr)
- NVIDIA GB200 NVL72 (contact us)
- NVIDIA GB300 NVL72 (contact us)
Sandbox
- Code Sandbox: VM sandboxes for large development environments ($0.0446/vCPU/hr, $0.0149/GiB RAM/hr)
- Code Interpreter: Execute LLM-generated code securely via API ($0.03/session)
Storage
- Shared Filesystem ($0.16/GiB/month)
- High-bandwidth, parallel filesystem colocated with compute
Fine-Tuning
- Supervised Fine-Tuning (LoRA)
- Supervised Fine-Tuning (Full Fine-Tuning)
- Direct Preference Optimization (LoRA)
- Direct Preference Optimization (Full Fine-Tuning)
- Standard pricing for models up to 100B parameters
- Specialized pricing for DeepSeek, GLM, Kimi, Llama 4, Qwen3, and other large models
- Minimum charge per job: $4.00 (standard)
Startup
- Up to 25 Paid Seats
- Unlimited Free View-only Seats
- Unlimited 1:1 Scheduling
- Automatic Video Recording
- Unlimited Real-time Transcription
- Unlimited AI Summary Notes
- "Ask Avoma" for a Meeting
- AI Email Follow-ups
- Auto Save Notes to CRM Records
- Dialer Integration
Organization
Popular- Up to 100 Paid Seats
- Custom AI Topics & Templates
- Group and Round-robin Scheduling
- Limited Conversation Intelligence
- Smart Playlist & AI Automations
- Translate Transcription & Notes
- Custom AI Email Templates
- Enforce Org-level Policies
- API Integration & Webhooks
- Customer Success Manager
Enterprise
- Minimum 10 Paid Seats
- Global "Ask Avoma"
- Designated Success Manager
- Concierge Onboarding & Training
- Quarterly Business Reviews
- Unlimited Usage Intelligence
- Team-Specific Access Controls
- Mutually Signed DPAs
- Single Sign-On (SAML/OIDC)
- Data Retention Policies
- HIPAA Compliance
Conversation Intelligence
- AI Coaching Recommendations
- Auto Call Scoring with AI
- Custom AI Scorecards
- SPICED, Sandler, MEDDICC, etc. AI Scorecards
- Smart Playlists – Rule Based Auto Curation
- Real-time Answer Assistant
- Global "Ask Avoma" across all your conversations
- Member & Scorecard Performance Trend Dashboard
- Interaction & Talk-Pattern Skills
- Standard Topic Intelligence
- Custom Keyword Trackers
- Custom Smart Topic Trackers
Revenue Intelligence
- "Ask Avoma" across Accounts
- "Ask Avoma" across Deals
- Dealboard for Pipeline Review
- 2-way CRM Field Updates
- AI-powered Deal Risk Alerts
- AI Deal Health Score
- AI Sales Methodology Tracker
- Roll-up & Aggregated Forecasting
- AI Win-Loss Deal Analysis
- Pipeline and Forecasting Reports
- Native CRM Integration
- Dialer Integration
Lead Router
- 1:1 Scheduling Links
- Customize Event Titles
- Embed Scheduler to your Website
- Save contacts to CRM
- Group Scheduling
- Round Robin Scheduling
- Real-time Lead Qualification and Routing
- Round Robin Scheduling based on Custom Rules and CRM fields
- Lead Routing based CRM Contact and Account Owner
- Lead Routing based on CRM Custom fields
- Form Field Based Router
- SDR to AE Handoff Router
Together AI vs Avoma FAQ
- Which one is cheaper?
- One or both use custom pricing, so it depends on your specific needs.
- Can I use either one for free?
- Neither has a free plan. But Avoma offers a free trial.
- How do they charge?
- Different approach here. Together AI uses usage-based pricing, while Avoma goes with per-seat. That changes the math depending on your team size and usage.
- Which one is a better deal?
- Depends on what you need. Together AI: Together AI is pitching itself as the serious infrastructure layer for AI builders — not the cheapest (Replicate or self-hosting can undercut on specific models), but more flexible and enterprise-ready than most managed inference APIs. At $5.49/GPU/hr for H100 dedicated, they're competitive with CoreWeave and Lambda Labs while bundling more managed services around it. Avoma: Avoma sits in the mid-market sweet spot — cheaper than Gong or Chorus for full revenue intelligence, but more capable than basic transcription tools like Otter or Fireflies. They're clearly targeting mid-size revenue teams that want Gong-lite functionality without the Gong-sized contract, and the modular add-on structure lets them compete on entry price while upselling aggressively.
Keep tabs on both.
We'll monitor pricing changes for Together AI and Avoma and let you know when something moves.