What is it
Per-generation pricing is a billing unit where each AI-generated creative asset — image, video, or design — is counted as a “generation” and metered accordingly.
The generation is the natural billing atom for creative AI because users think in outputs, not compute cycles. When a marketer asks for “100 product images,” they are counting generations. FLORA, Frase, and Typeface all meter generations because it aligns with how customers measure the work done — the campaign, the batch, the content run. Choosing this unit is a value-metric decision as much as a billing one.
Platforms wire the same atom up very differently. FLORA rejects abstract credits: every generation bills at the underlying model’s published API rate in real dollars, so users see the exact cost per generation and agencies can bill clients directly from FLORA’s usage history. Frase caps “AI generations” as one of seven per-plan meters, bundling a fixed count into each tier.
The unit clusters in creative and marketing AI because outputs are countable discrete artifacts — images and short videos — unlike text, where volume is a continuous token dimension. OctoAI made the boundary explicit before its 2024 shutdown: it billed text per 1M tokens but metered its image (“Media Gen”) endpoints per image generated and per GPU compute-second.
How it works
Each generation is a completed AI creative output. The platform increments a counter once the output is delivered, then either debits an included allowance or bills the generation at a per-unit or per-model rate. The four companies in the corpus package that same atom in four distinct ways:
| Company | Generation unit | How it’s priced |
|---|---|---|
| FLORA | One image or one video | Billed at the published model API rate in real dollars (e.g. Nano Banana 2 at $0.151/image); drawn from a per-seat usage pool ($12 / $50 / $100/seat on Starter / Pro / Max) |
| Frase | One “AI generation” of content | Capped per plan — 25 / 100 / 350 generations on Starter $49 / Professional $129 / Scale $299 per month; overage opt-in pay-as-you-go |
| OctoAI (discontinued) | One image output (SDXL, SD 1.5, SVD) | Usage-metered per image generated and/or per GPU compute-second; $10 free signup credit |
| Typeface | One brand content asset | Blended into enterprise quotes: per-seat (reported ~$49–$99/user/mo) plus generation/credit usage and brand workspaces |
The core unit math for a plan-capped model like Frase:
Monthly cost = tier fee + (max(0, generations_used − included_allowance) × overage_rate)
FLORA’s dollar-denominated variant replaces the abstract allowance with a real-dollar pool:
Usage spend = Σ (generations × published_model_rate), drawn down against seats × included pool; anything above the pool is billed pay-as-you-go at the same API rates.
Worked example on FLORA: a 4-seat Pro team ($50/seat/mo = $200 platform) includes a pooled usage budget of 4 × $50 = $200/month. At the Nano Banana 2 rate of $0.151/image, that pool covers roughly 1,300 image generations before overage — after which each additional image still bills at exactly $0.151, with no credit-conversion markup. That transparency is the whole pitch, and the same reasoning behind broader usage-based pricing models: meter the thing the customer already counts.
Companies using this
Four corpus companies meter AI creative outputs as generations: FLORA with dollar-transparent, model-rate generations across images and video; Frase capping AI generations inside its content-platform tiers; OctoAI (now discontinued) metering image generation per output and compute-second; and Typeface folding generation/credit usage into sales-led enterprise quotes.
| Company | Product | Pricing model | Billing units | Free tier | Verified |
|---|---|---|---|---|---|
| FLORA | AI-powered creative canvas and workflow platform | Yes | 2026-07-23 | ||
| Frase | Agentic SEO and GEO platform that researches, writes, optimizes, and tracks AI-search visibility for content teams. | No | 2026-08-04 | ||
| OctoAI | Generative AI inference platform (acquired by NVIDIA, sunset Oct 2024) | No | 2026-07-30 | ||
| Typeface | Arc enterprise marketing AI platform | No | 2026-06-16 |
Explore this theme in the knowledge graph
FAQ
What is per-generation pricing in AI creative platforms?
Per-generation pricing meters each AI-generated output — an image, a video, or a design variant — as a discrete unit. FLORA bills every generation at its underlying model's published API rate (e.g. Nano Banana 2 at $0.151/image) so users see the exact dollar cost per generation; Frase caps 'AI generations' per plan (25 / 100 / 350 on Starter / Professional / Scale); Typeface blends generation/credit usage into enterprise quotes. The generation unit abstracts over model complexity and output format so the platform can change infrastructure without repricing the customer-facing unit.
How does per-generation pricing compare to token-based pricing for AI creative tools?
Token-based pricing is the standard for language models but maps poorly to visual content, where the 'size' of an output is pixels or frames, not tokens. Per-generation pricing uses the finished asset as the unit — one image is one generation regardless of resolution. OctoAI showed the split directly: it billed text per 1M tokens (Llama 3 8B at $0.15/1M) but metered its Media Gen endpoints (SDXL, SD 1.5) per image generated and per GPU compute-second, because token counting doesn't describe a picture.
How much does per-generation AI pricing cost?
It varies by how the generation is packaged. FLORA's paid plans start at $18/seat/mo (Starter) up to $200/seat/mo (Max), each including a dollar-denominated usage pool ($12 / $50 / $100 per seat) spent at published model rates. Frase bundles a fixed number of AI generations per plan (25 / 100 / 350) inside its $49 / $129 / $299 per-month tiers. Typeface is sales-led with no public rate card — third parties report indicative seat tiers around $49 and $99 per user/month plus generation/credit usage.
Is per-generation pricing the same as credit-based pricing?
They overlap but differ. In pure credit systems, credits are an abstract currency and a generation may cost a variable number of credits by model. Per-generation pricing keeps the generation itself as the visible unit. FLORA deliberately rejects abstract credits — it prices every generation in real dollars at published API rates so agencies can bill clients from FLORA's usage history — while Typeface and OctoAI blend generations with credits and compute.
Related billing units
- Credit-Based BillingA billing unit where customers pre-purchase or are allocated a pool of credits that deplete as they use the product, often at variable rates per feature.
- Token-Based PricingA billing unit common in LLM and AI products, where customers are charged per input and output token processed.
- Per-Seat PricingA billing unit where the vendor charges a fixed fee per named user, regardless of how much each user consumes.
- Per-Resolution PricingA billing unit unique to AI customer-support products, where the vendor charges only when an AI agent resolves a customer issue without escalation.
- Bandwidth-Based PricingA billing unit where customers are charged per gigabyte of data transferred out of the platform.
- Per-Function-Invocation PricingA billing unit where customers are charged per serverless function invocation, often combined with a separate compute-time charge.
- CPU-Hour PricingA billing unit where customers are charged for the CPU time their workloads consume, typically measured in vCPU-seconds or vCPU-hours.
- GB-Hour PricingA billing unit where customers are charged for the memory their workloads consume over time, measured in gigabyte-hours.
- GPU-Hour PricingA billing unit where customers are charged for GPU time consumed, typically measured per-second or per-hour by GPU type.
- Per-API-Call PricingA billing unit where customers are charged per API request, regardless of payload size or processing time.
- Per-GB Storage PricingA billing unit where customers are charged per gigabyte of data stored on the platform per month.
- Media-Minute PricingA billing unit where customers are charged per minute of audio or video processed — used by speech, voice, and video AI vendors.
- Per-Request PricingA billing unit where customers are charged per request served — the generic meter for inference endpoints, search, scraping, and browser infrastructure.
- Per-Event PricingA billing unit where customers are charged per event ingested — the native meter of observability and billing-infrastructure platforms.
- Vector Storage PricingA billing unit where customers are charged for vectors stored or indexed — the storage dimension of vector database pricing.
- Per-Character PricingA billing unit where customers are charged per character of text processed — the standard meter for text-to-speech and translation.
- Per-Document PricingA billing unit where customers are charged per document processed or generated — common in AI writing, SEO, and document-intelligence tools.
- Per-Page PricingA billing unit where customers are charged per page crawled, parsed, or rendered — the meter for web scraping and document parsing.
- Per-Transaction PricingA billing unit where customers are charged per financial or billing transaction processed — the meter of billing and accounting platforms.
- Active-User PricingA billing unit where customers are charged per monthly or daily active user rather than per provisioned seat.
- Per-Task PricingA billing unit where customers are charged per task an automation or agent executes — Zapier's historical unit, now spreading to AI agents.
- Per-Unit PricingA billing unit used by robotics, hardware AI, and some SaaS companies where the metered object is a physical or abstract 'unit' — a robot deployed, a device sold, or a defined deliverable.
- Workflow Execution PricingA billing unit where each end-to-end workflow or automation run is metered and billed, regardless of the compute steps it contains.
- Per-Message PricingA billing unit where each individual message or reply in a conversation is metered, common in AI chat and voice platforms.
- Per-Invoice PricingA billing unit used by billing infrastructure platforms where each invoice generated or processed is metered as the primary cost driver.
- Per-Action PricingA billing unit where each discrete action taken by an AI agent or automation is metered — common in browser automation and agentic workflow tools.
- Per-Image PricingA billing unit where each AI-generated image is metered, common in image generation APIs and multimodal AI platforms.
- Per-Conversation PricingA billing unit where each complete customer conversation — from first message to resolution — is metered as a single chargeable event.
- Per-Record PricingA billing unit where each data record processed, labeled, or extracted is metered — common in data platforms and web scraping services.
- Per-Word PricingA billing unit common in translation and localization platforms where the metered object is the word count of content processed.
- Per-Video PricingA billing unit where each AI-generated video is metered, common in video generation and synthetic media platforms.
- Milestone-Based PricingA billing unit used in drug discovery and biotech AI where payment is tied to achieving defined research milestones rather than time or compute consumed.
- Per-Outcome PricingA billing unit where payment is triggered by verified outcomes delivered — distinct from outcome-based pricing models, this refers specifically to 'outcomes' as a countable billing unit.
- Per-Datapoint PricingA billing unit where each individual data measurement or signal ingested is metered — common in cloud cost intelligence and ML evaluation platforms.
- Per-Interaction PricingA billing unit where each patient-agent or user-agent interaction is metered, common in healthcare AI and customer engagement platforms.
- Data Licensing PricingA pricing structure where access to proprietary datasets or data assets is licensed separately from the software or services, common in AI training data and clinical data platforms.
- Robot-Hour PricingA billing unit where each hour a robot or autonomous system operates is metered — the robotics equivalent of a GPU-hour.
- Per-Contact PricingA billing unit where each contact or lead in the database is metered, common in AI sales development and outbound automation platforms.
- Per-Mailbox PricingA billing unit where each connected email mailbox or sending account is metered, common in AI outbound sales and email automation platforms.
- Browser-Hour PricingA billing unit where each hour of headless browser compute time is metered, common in web scraping and browser automation platforms.
- Per-Ticket PricingA billing unit where each customer support ticket handled by an AI agent is metered — common in AI customer service platforms.
- Per-Log PricingA billing unit where each LLM request log ingested or stored is metered — common in AI observability and evaluation platforms.
- Per-Trace PricingA billing unit where each distributed trace — a complete record of an LLM request chain — is metered, common in AI observability platforms.
- Per-IP PricingA billing unit where each IP address or proxy endpoint allocated is metered — used by web scraping proxy providers.
- Per-Device PricingA billing unit where each hardware device or endpoint connected to the AI platform is metered.
- Per-Case PricingA billing unit used in legal AI platforms where each case or matter processed by the AI is metered.
- Per-Report PricingA billing unit where each AI-generated report or analysis document is metered as a discrete output.