Together AI vs Avoma Pricing (2026)

How do these two stack up on price? Here's what each one costs, what you get, and where the value sits.

Together AI Avoma
Starts at Custom $25/mo
Number of plans 7 6
Free plan
Free trial
Pricing model usage-based per-seat

Serverless Inference

Custom
  • Chat models (price per 1M tokens)
  • Vision models (price per 1M tokens)
  • Image generation models (price per image or per mp)
  • Audio / TTS models (price per 1M characters)
  • Video generation models (price per video)
  • Transcription models (price per audio minute)
  • Embeddings models (price per 1M tokens)
  • Moderation models (price per 1M tokens)
  • Batch API pricing available

Provisioned Throughput

Custom
  • Reserved dedicated capacity in throughput units (PTUs)
  • Fixed capacity per PTU with model-dependent tokens-per-minute
  • MiniMax M3 support
  • GLM-5.2 support

Dedicated Inference

Custom
  • Guaranteed performance (no sharing)
  • Support for custom models
  • Autoscaling & traffic spike handling
  • NVIDIA HGX H100 on-demand ($5.49/GPU/hr)
  • NVIDIA HGX B200 on-demand ($8.99/GPU/hr)
  • NVIDIA HGX H200 (contact us)
  • NVIDIA HGX B300 (contact us)
  • NVIDIA GB200 NVL72 (contact us)
  • NVIDIA GB300 NVL72 (contact us)
  • Reserved capacity available

GPU Clusters

Custom
  • On-demand pay-as-you-go GPU capacity (hourly)
  • Reserved capacity (7-30 days, 31-90 days, 91-180 days, 181+ days)
  • NVIDIA HGX H100 on-demand ($3.99/GPU/hr)
  • NVIDIA HGX H200 on-demand ($5.99/GPU/hr)
  • NVIDIA HGX B200 on-demand ($8.19/GPU/hr)
  • NVIDIA GB200 NVL72 (contact us)
  • NVIDIA GB300 NVL72 (contact us)

Sandbox

Custom
  • Code Sandbox: VM sandboxes for large development environments ($0.0446/vCPU/hr, $0.0149/GiB RAM/hr)
  • Code Interpreter: Execute LLM-generated code securely via API ($0.03/session)

Storage

Custom
  • Shared Filesystem ($0.16/GiB/month)
  • High-bandwidth, parallel filesystem colocated with compute

Fine-Tuning

Custom
  • Supervised Fine-Tuning (LoRA)
  • Supervised Fine-Tuning (Full Fine-Tuning)
  • Direct Preference Optimization (LoRA)
  • Direct Preference Optimization (Full Fine-Tuning)
  • Standard pricing for models up to 100B parameters
  • Specialized pricing for DeepSeek, GLM, Kimi, Llama 4, Qwen3, and other large models
  • Minimum charge per job: $4.00 (standard)

Startup

$29/mo
  • Up to 25 Paid Seats
  • Unlimited Free View-only Seats
  • Unlimited 1:1 Scheduling
  • Automatic Video Recording
  • Unlimited Real-time Transcription
  • Unlimited AI Summary Notes
  • "Ask Avoma" for a Meeting
  • AI Email Follow-ups
  • Auto Save Notes to CRM Records
  • Dialer Integration

Organization

Popular
$39/mo
  • Up to 100 Paid Seats
  • Custom AI Topics & Templates
  • Group and Round-robin Scheduling
  • Limited Conversation Intelligence
  • Smart Playlist & AI Automations
  • Translate Transcription & Notes
  • Custom AI Email Templates
  • Enforce Org-level Policies
  • API Integration & Webhooks
  • Customer Success Manager

Enterprise

$39/mo
  • Minimum 10 Paid Seats
  • Global "Ask Avoma"
  • Designated Success Manager
  • Concierge Onboarding & Training
  • Quarterly Business Reviews
  • Unlimited Usage Intelligence
  • Team-Specific Access Controls
  • Mutually Signed DPAs
  • Single Sign-On (SAML/OIDC)
  • Data Retention Policies
  • HIPAA Compliance

Conversation Intelligence

$35/mo
  • AI Coaching Recommendations
  • Auto Call Scoring with AI
  • Custom AI Scorecards
  • SPICED, Sandler, MEDDICC, etc. AI Scorecards
  • Smart Playlists – Rule Based Auto Curation
  • Real-time Answer Assistant
  • Global "Ask Avoma" across all your conversations
  • Member & Scorecard Performance Trend Dashboard
  • Interaction & Talk-Pattern Skills
  • Standard Topic Intelligence
  • Custom Keyword Trackers
  • Custom Smart Topic Trackers

Revenue Intelligence

$35/mo
  • "Ask Avoma" across Accounts
  • "Ask Avoma" across Deals
  • Dealboard for Pipeline Review
  • 2-way CRM Field Updates
  • AI-powered Deal Risk Alerts
  • AI Deal Health Score
  • AI Sales Methodology Tracker
  • Roll-up & Aggregated Forecasting
  • AI Win-Loss Deal Analysis
  • Pipeline and Forecasting Reports
  • Native CRM Integration
  • Dialer Integration

Lead Router

$25/mo
  • 1:1 Scheduling Links
  • Customize Event Titles
  • Embed Scheduler to your Website
  • Save contacts to CRM
  • Group Scheduling
  • Round Robin Scheduling
  • Real-time Lead Qualification and Routing
  • Round Robin Scheduling based on Custom Rules and CRM fields
  • Lead Routing based CRM Contact and Account Owner
  • Lead Routing based on CRM Custom fields
  • Form Field Based Router
  • SDR to AE Handoff Router

Together AI vs Avoma FAQ

Which one is cheaper?
One or both use custom pricing, so it depends on your specific needs.
Can I use either one for free?
Neither has a free plan. But Avoma offers a free trial.
How do they charge?
Different approach here. Together AI uses usage-based pricing, while Avoma goes with per-seat. That changes the math depending on your team size and usage.
Which one is a better deal?
Depends on what you need. Together AI: Together AI is pitching itself as the serious infrastructure layer for AI builders — not the cheapest (Replicate or self-hosting can undercut on specific models), but more flexible and enterprise-ready than most managed inference APIs. At $5.49/GPU/hr for H100 dedicated, they're competitive with CoreWeave and Lambda Labs while bundling more managed services around it. Avoma: Avoma sits in the mid-market sweet spot — cheaper than Gong or Chorus for full revenue intelligence, but more capable than basic transcription tools like Otter or Fireflies. They're clearly targeting mid-size revenue teams that want Gong-lite functionality without the Gong-sized contract, and the modular add-on structure lets them compete on entry price while upselling aggressively.

Keep tabs on both.

We'll monitor pricing changes for Together AI and Avoma and let you know when something moves.

Start tracking free »