Per-Message Pricing: Examples & Companies

18 companies in the corpus Updated partial analysis
Definition

Per-Message Pricing is a billing unit where each individual message or reply in a conversation is metered, common in AI chat and voice platforms.

Also known as: Per-Reply PricingMessage-Based Billing

What is it

Per-message pricing is a billing unit where each individual message or reply in a conversation is metered, common in AI chat and voice platforms.

The message is the most intuitive unit in conversational AI. It mirrors how users already think — “I sent a message, I got a reply” — and maps onto the single LLM call that drives inference cost. Where token pricing is precise but opaque to non-technical buyers, message pricing trades granularity for a meter that is human-readable before the buyer opens the docs.

The unit shows up in two commercial contexts. Consumer AI platforms — Poe, Grok, Janitor AI, and the now-sunset Cognosys — use message caps to differentiate subscription tiers rather than billing per message, so the message rate is the invisible floor even when no dollar-per-message line item ever appears. Voice and chat API platforms take the opposite tack: Retell AI and Vapi publish explicit per-message rates for their chat-agent products — $0.002 and $0.005 respectively — treating the message as a pure usage meter alongside per-minute voice billing.

Edge cases reveal how elastic “message” can be. Intercom Fin carries messages through a Proactive Support Plus add-on, separate from its headline per-resolution model, and Pi rate-limits on messages purely to control cost — its effective rate is $0, and the meter never converts to a charge. A message meter in the frontmatter does not imply a per-message bill. For how to pick a unit like this, see choosing the right usage metric.

Same 5,000 messages · the model sets the bill
The message count is fixed — the model moves the bill 6× TIER CAP hidden Grok · Poe rate inside plan FLAT PER MSG $0.005 Vapi · SMS / chat model-agnostic LIGHT MODEL ~$10 Retell · $0.002/msg 5,000 messages FRONTIER MODEL ~$65 Retell · $0.013/msg same 5,000 msgs ← LEGIBLE, CHEAP MODEL DRIVES SPEND →

How it works

A message meter counts conversational turns and multiplies by a rate. The rate can be an explicit dollar figure (API platforms), a points cost that varies by model (Poe), or an implicit ceiling encoded as a tier cap (consumer apps). What differs across vendors is what actually gets counted and how the price is exposed.

DimensionWhat it controlsExample on this page
Rate exposureWhether the message price is a public dollar figure or hidden inside a subscriptionRetell AI publishes “from $0.002/AI msg”; Grok hides it behind flat tiers
Unit of countingUser turn, AI reply, or bothRetell counts each AI message; Poe charges per message to a bot
Rate variabilityFlat per message vs. model-dependentVapi is a flat $0.005/msg; Poe’s points cost swings ~10–20 pts (light text) to thousands (frontier/video)
PackagingMetered directly vs. bundled into a capRetell/Vapi meter directly; Cognosys capped Free at 100 msg/mo, Pro (~$15/mo) at ~1,000

The clearest direct-metered example is Retell AI. Its chat agents are billed per AI message from $0.002, and the effective rate depends on the model — a GPT 5.1 chat message computes to about $0.013. There is no platform fee, no contract, and new accounts get $10 in free credits.

Unit math (Retell chat agent): 5,000 GPT 5.1 messages × $0.013/msg ≈ $65. The same volume at the $0.002 floor rate (a lighter model) is $10 — the model choice is the dominant cost lever, not the message count.

Poe shows the points-abstraction variant: one subscription buys a monthly (or daily) compute-points allowance, and every message spends points at a published, model-dependent rate. A lightweight text model costs ~10–20 points per message, so a $19.99 plan stretches to tens of thousands of cheap messages — but a single high-end video generation can cost more points than thousands of text messages. The message is the meter; the model mix is the real cost driver.


Companies using this

Eight companies in the corpus list messages as a billing unit, split between voice/chat API platforms that meter it directly (Retell AI, Vapi) and consumer apps that use message caps to gate subscription tiers (Poe, Grok, Janitor AI, Cognosys).


Patterns observed

  • Direct metering lives in the API layer, caps live in the app layer. The only companies quoting a public per-message dollar rate — Retell AI and Vapi — are developer platforms selling chat agents. Every consumer product expresses the unit as a tier cap instead: Cognosys at 100 messages/month, Janitor AI at ~50/day, Grok at ~10 prompts per 2 hours.

  • The message is almost never the only meter. On every company here, messages rides alongside another unit — minutes for Retell AI and Vapi, points for Poe, tokens for Grok and Janitor AI, resolutions for Intercom Fin. Message pricing is a legibility layer over an underlying token or compute cost, not a standalone model.

  • Model choice, not message volume, drives the bill. Where the rate is model-dependent — Poe’s points, Retell AI’s per-model chat rates — the same message count can cost 5–10× on the model selected. Buyers budgeting on message count alone consistently underestimate frontier-model spend.


Counterexamples & variants

The clearest counterexample is Pi. It lists messages as a billing unit, but the app has never been monetized — Inflection AI rations capacity with rate limits and cooldowns rather than dollars, so the effective rate is $0 and the meter never converts to a bill. It is a pure cost-control meter, not a pricing model.

Cognosys is a lifecycle variant: it ran a clean message-metered freemium ladder (Pro ~1,000 msg/mo at $15, Ultimate unlimited at $59) but is now sunset after the Cohere acquisition, so its rate card is historical rather than purchasable. Intercom Fin inverts the emphasis: its economics run on $0.99-per-resolution outcome pricing, with messages surfacing only in a peripheral Proactive Support add-on (500/mo at $99). The message unit is real but commercially minor.


What this means for buyers vs vendors

For buyers

Confirm exactly what a “message” is before you model spend — Retell AI counts each AI reply, Poe charges per message to a bot, and consumer caps count prompts. Then model on your worst-case model mix, not message volume, since a model-dependent rate is what actually moves the bill. Use the pricing calculator to stress-test a high-volume month and read usage invoicing and billing cycles to understand how overages and caps are reconciled.

For vendors

Per-message billing fits when your unit cost is predictable per exchange and your buyer is non-technical enough that tokens would confuse them — the message is a legibility win that Retell AI turned into a self-serve growth lever. Pair it with a second meter that tracks real cost (minutes, points, or tokens) so a frontier-model exchange doesn’t erode margin, and start with the fundamentals in introduction to usage-based pricing. Expect to publish caps or free-tier limits, since every consumer app on this page gates messages rather than charging for them directly.


Company Product Pricing modelBilling unitsFree tier Verified
ActiveCampaignMarketing automation, email marketing, and sales CRM platform priced by contact countNo2026-07-12
Bland AIAI phone call automation platform — inbound and outbound voice agents at scaleYes2026-07-21
BrazeEnterprise customer-engagement platform for cross-channel messaging (email, push, SMS, in-app), with BrazeAI (Sage AI) bundled.No2026-07-12
CognosysAutonomous AI agents (rebranded Ottogrid, acquired by Cohere)Yes2026-07-30
Constant ContactEmail and digital marketing platform for small businesses, with a separate Lead Gen & CRM product (ex-SharpSpring) and a bundled AI content assistantNo2026-08-06
Customer.ioCustomer engagement platform combining Journeys (behavioral messaging), an AI Agent, and Data Pipelines (CDP)No2026-07-12
GrokxAI's consumer and business AI assistantYes2026-06-16
Intercom FinFin AI Agent for customer serviceNo2026-08-04
IterableCross-channel marketing automation and AI customer engagement platform (email, SMS, push, in-app, web, OTT)No2026-07-12
Janitor AIConsumer AI character chat / roleplay platformYes2026-06-16
KeapAll-in-one CRM, sales, and marketing-automation platform for small businessesNo2026-07-06
KlaviyoB2C marketing CRM for email, SMS, and push, priced on active profiles plus channel volume, with bundled predictive analytics and Klaviyo AI.Yes2026-07-12
Microsoft Dynamics 365Microsoft's enterprise CRM + ERP suite — Sales, Customer Service, Field Service, Business Central, Finance and Supply Chain, with Copilot woven inNo2026-07-06
PiPi — personal, emotionally intelligent AI assistant (consumer app)Yes2026-06-16
PoeMulti-model AI chat subscription (by Quora)Yes2026-07-28
Retell AIConversational voice-agent API platformNo2026-07-22
Snowflake CortexAI functions and model APIs on SnowflakeYes2026-07-06
VapiVoice AI infrastructure for developersNo2026-06-09

Explore this theme in the knowledge graph

FAQ

What is per-message pricing in AI products?

Per-message pricing charges each individual message or reply in a conversation as the billable unit. Consumer AI apps like Poe and Grok use message caps to gate subscription tiers, while voice AI platforms like Retell AI and Vapi charge fractional cents per AI message on their chat agents.

How does per-message pricing differ from per-token pricing?

Tokens measure the raw units the underlying model processes; a message is a complete conversational turn — a prompt plus its reply. One message can consume a few hundred to several thousand tokens depending on context length, so per-message rates are higher in absolute terms but far more legible to non-technical buyers.

Which AI companies charge per message?

Retell AI charges from $0.002 per AI message on chat agents; Vapi charges $0.005 per SMS/chat message; Poe uses compute points where each message to a model costs a published point rate. Consumer apps like Poe, Grok, Janitor AI, and the sunset Cognosys bundle messages into subscription tiers rather than billing per message directly.

How much does per-message chat cost on voice AI platforms?

Retell AI bills chat agents from $0.002 per AI message (GPT 5.1 works out to about $0.013 per message), and Vapi charges a flat $0.005 per SMS or chat message on top of its $0.05-per-minute voice hosting. Both give new accounts $10 in free credits to start.

What are the risks of per-message pricing for buyers?

The main risks are unpredictable costs when conversation volume spikes and bill shock when a single session with a frontier model burns through a points budget. Buyers should confirm whether a message counts each user turn, each AI reply, or both — Poe, Retell, and Cognosys all define it differently.

Related billing units

Back to companies