What is it
Per-Page Pricing is a billing unit where customers are charged per page crawled, parsed, or rendered — the meter for web scraping and document parsing. The page is where web-data infrastructure and document AI converge on the same noun: Firecrawl meters pages scraped into clean markdown, LlamaIndex’s LlamaParse meters PDF pages parsed into structured output, Mistral AI’s OCR bills $2 per 1,000 pages read, Unstructured charges a flat $0.03 per page of document ETL, and Exa and You.com both sell full-page content retrieval at $1 per 1,000 pages beside their per-call search APIs.
The unit’s appeal is the same as the document’s: a buyer with a URL list or a PDF folder can count their job before running it. Its weakness is that “a page” abstracts over enormous compute variance — a static blog post and a scanned 600-dpi financial table are the same unit and very different work — which is why most of this cohort prices modes on top of the page count rather than one blended rate.
The SEO and docs tools carry the unit in subscription form: Frase and Scalenut bundle audit-page quotas into monthly tiers, and Mintlify treats the documentation page itself as the rendered unit inside its docs platform — pages as an allowance or a hosted deliverable rather than a metered API line.
How it works
The metered cluster prices pages × mode rate, usually denominated in credits:
| Job | Vendor & rate | Mechanics |
|---|---|---|
| Plain content fetch | Exa Contents $1/1k pages; You.com Livecrawl $1/1k | Known URLs in, text/markdown out |
| Scrape / crawl | Firecrawl 1 credit/page (Scrape, Crawl, Map, Monitor) | Hobby $16/5k credits → Standard $83/100k → Scale $599/1M |
| OCR | Mistral $2/1k pages ($3 with annotations) | Flat rate, no modes |
| Document ETL | Unstructured $0.03/page flat, any file type & pipeline | 15,000 free no-expiration pages; Fast/Hi-Res/VLM all one price |
| Structured parse | LlamaParse 1 / 3 / 10 / 45 credits per page (1k credits = $1.25) | Fast → Cost-effective → Agentic → Agentic Plus |
| Docs pages (platform) | Mintlify Starter $0, AI credits 5,000 included then $0.01/credit | Documentation pages hosted; AI features metered in credits |
| Page audits (quota) | Frase 50–1,000 pages/mo by tier; Scalenut similar | Allowance inside the subscription, no overage rate |
Worked example — the mode is the price. The four bills above are the same 100,000 pages: Firecrawl scraping lands on the Standard tier’s 100k-credit allotment ($83), Mistral OCR runs $200, and Unstructured’s flat $0.03/page ETL comes to $3,000 whether the pages are clean text or VLM-parsed images. LlamaParse spreads widest — $125 on Fast (1 credit/page) to $5,625 on Agentic Plus (45 credits/page), a 45x swing on an identical page count set entirely by how much reasoning the parse applies.
Worked example — endpoint drift. Firecrawl’s “1 credit = 1 page” holds for Scrape, Crawl, Map, and Monitor — but Search bills 2 credits per 10 results and Interact bills 2 credits per browser-minute. A workload that mixes endpoints can’t be estimated from page count alone; the usage-metric guide calls this unit drift, and it’s the first thing to audit on any credit-denominated rate card.
Companies using this
9 in-corpus companies meter pages: the API cluster (Firecrawl, LlamaIndex, Mistral AI, Unstructured, Exa, You.com) billing pages scraped, parsed, OCR’d, ETL’d, or fetched, plus the content pair (Frase, Scalenut) bundling audit-page quotas into subscription tiers and Mintlify metering documentation pages inside its docs platform.
Patterns observed
The commodity floor is real and public: plain page retrieval settled at $1 per 1,000 pages at both Exa and You.com — undifferentiated page-touching is priced like a utility, and margin lives only in the intelligence layered above it.
The category shows a clear repricing gravity toward simpler meters over time. Unstructured walked from compute-hour billing (~$12.93/1,000 pages) to strategy-tiered per-page (Fast $1, Hi-Res $10 per 1,000) to a single flat $0.03/page, and LlamaParse’s v2 collapsed a 1–90-credit maze into four flat tiers — each step trading a “fairer” ladder for a legible one.
Credits are the standard wrapper. Firecrawl and LlamaParse both denominate pages in credits with tier-sized monthly pools, which buys repricing flexibility without touching the dollar price of a credit. Mintlify applies the wrapper to a docs platform: documentation pages are the hosted unit, but AI features draw on 5,000 included credits at $0.01/credit overage — the page and the credit live side by side, with a hard cap so AI spend can’t surprise the buyer.
Counterexamples & variants
Frase and Scalenut show the unit defanged: pages appear as audit quotas (50/mo on Frase Starter, 1,000 at the top) inside flat subscriptions, with no per-page overage rate published — the page count gates the tier you need rather than metering a bill, which makes them packaging, not pricing.
Unstructured is the sharpest counterexample to the mode ladder: one flat rate for every strategy deliberately removes the “if I turn on the good parser, what happens to my bill?” objection — the opposite bet from a per-mode credit spread. And Mintlify is the semantic variant: its “page” is a published documentation page, not a crawled or parsed one, so the meter is closer to a hosting deliverable than a per-request API line — a reminder that “pages-rendered” spans both the input a scraper consumes and the output a docs platform serves.
What this means for buyers vs vendors
For buyers
Price the mode, not the page: get a representative sample of your documents and run them through the cheapest tier that passes quality, because the spread between modes (1–45 credits on LlamaParse, versus one flat rate on Unstructured) dwarfs the spread between vendors at any single mode. On crawl workloads, ask what bills on failure — redirects, soft-404s, and bot walls are a real fraction of any large URL list — and watch reprocessing, since re-running a corpus after a chunking or embedding change pays full per-page rates again. For bundled audit quotas (Frase, Scalenut) or docs-page platforms (Mintlify), translate the tier into an effective per-page price at your actual monthly volume before comparing against the metered APIs.
For vendors
Pick a lane and commit to it: publish the mode ladder with each rung’s quality difference demonstrable, or flatten it entirely as Unstructured did — the category’s repricing history shows buyers reward legible page pricing and punish mode mazes. Keep the commodity rung at the market floor as the funnel, take margin on intelligence above it, and if you wrap pages in credits, hold the page-to-credit exchange rate stable per endpoint — every footnote on “1 credit = 1 page” is forecast error you’re exporting to the buyer, and the prepaid-credits guide covers why that trust erosion compounds. Buyers estimating token-and-page costs alongside the parse bill can sanity-check model spend with the Mistral pricing calculator.
| Company | Product | Pricing model | Billing units | Free tier | Verified |
|---|---|---|---|---|---|
| Exa | AI web search API for agents — search, contents, deep research, and monitoring endpoints billed per request | Yes | 2026-07-14 | ||
| Firecrawl | Web-scraping and data-extraction API for AI agents — scrape, crawl, map, search, and extract pages into clean markdown/JSON | Yes | 2026-06-30 | ||
| Frase | Agentic SEO and GEO platform that researches, writes, optimizes, and tracks AI-search visibility for content teams. | No | 2026-06-24 | ||
| LlamaIndex | RAG/agent orchestration framework + LlamaCloud document parsing | Yes | 2026-07-23 | ||
| Mintlify | AI-native developer documentation | Yes | 2026-06-15 | ||
| Mistral AI | Open and commercial LLM APIs | Yes | 2026-07-06 | ||
| Scalenut | AI search visibility (GEO) and SEO content platform — tracks brand presence in AI answers and generates ready-to-rank content | No | 2026-06-07 | ||
| Snowflake Cortex | AI functions and model APIs on Snowflake | Yes | 2026-07-06 | ||
| Unstructured | Document ingestion / ETL API | Yes | 2026-07-14 | ||
| You.com | Web search, contents, research, and finance-research APIs for AI systems | Yes | 2026-07-22 |
Explore this theme in the knowledge graph
FAQ
What is per-page pricing?
Per-page pricing is a billing unit where customers are charged per page crawled, parsed, or rendered — the meter for web scraping (Firecrawl), document parsing (LlamaIndex's LlamaParse), OCR (Mistral), document ETL (Unstructured), and content-fetch APIs (Exa, You.com). One page is one unit, regardless of what it took to process.
How much does it cost to scrape or parse 1,000 pages?
It depends on the job: plain content fetch runs $1 per 1,000 pages (Exa Contents, You.com Livecrawl), OCR runs $2 per 1,000 (Mistral, $3 with annotations), scraping is about $0.60–$3.20 per 1,000 via Firecrawl's credit tiers, document ETL is a flat $30 per 1,000 on Unstructured ($0.03/page), and structured parsing spans $1.25 to $56 per 1,000 on LlamaParse depending on mode.
Which companies use per-page pricing?
Nine in this corpus: Exa, Firecrawl, Frase, LlamaIndex, Mintlify, Mistral, Scalenut, Unstructured, and You.com. The API cluster meters pages directly or through credits; Frase and Scalenut bundle page audits as monthly quotas; Mintlify meters documentation pages inside a docs platform.
Why do parse tiers cost up to 45x more per page?
Because the page is an abstraction over wildly different compute: a static HTML page and a scanned table-dense PDF cost very different amounts to process. LlamaParse prices the mode (Fast 1 credit, Cost-effective 3, Agentic 10, Agentic Plus 45 per page) so buyers choose the effort level per document rather than paying a blended rate — while Unstructured takes the opposite bet with one flat $0.03/page for every strategy.
Do failed or empty pages bill?
Policies differ by vendor and endpoint, and it materially affects crawl economics — large crawls hit redirects, soft-404s, and bot walls constantly. Check whether the meter counts attempts or successes, and whether reprocessing re-bills (on Unstructured every re-run of a corpus pays full rate again), before estimating from a URL list.
Related billing units
- Credit-Based BillingA billing unit where customers pre-purchase or are allocated a pool of credits that deplete as they use the product, often at variable rates per feature.
- Token-Based PricingA billing unit common in LLM and AI products, where customers are charged per input and output token processed.
- Per-Seat PricingA billing unit where the vendor charges a fixed fee per named user, regardless of how much each user consumes.
- Per-Resolution PricingA billing unit unique to AI customer-support products, where the vendor charges only when an AI agent resolves a customer issue without escalation.
- Bandwidth-Based PricingA billing unit where customers are charged per gigabyte of data transferred out of the platform.
- Per-Function-Invocation PricingA billing unit where customers are charged per serverless function invocation, often combined with a separate compute-time charge.
- CPU-Hour PricingA billing unit where customers are charged for the CPU time their workloads consume, typically measured in vCPU-seconds or vCPU-hours.
- GB-Hour PricingA billing unit where customers are charged for the memory their workloads consume over time, measured in gigabyte-hours.
- GPU-Hour PricingA billing unit where customers are charged for GPU time consumed, typically measured per-second or per-hour by GPU type.
- Per-API-Call PricingA billing unit where customers are charged per API request, regardless of payload size or processing time.
- Per-GB Storage PricingA billing unit where customers are charged per gigabyte of data stored on the platform per month.
- Media-Minute PricingA billing unit where customers are charged per minute of audio or video processed — used by speech, voice, and video AI vendors.
- Per-Request PricingA billing unit where customers are charged per request served — the generic meter for inference endpoints, search, scraping, and browser infrastructure.
- Per-Event PricingA billing unit where customers are charged per event ingested — the native meter of observability and billing-infrastructure platforms.
- Vector Storage PricingA billing unit where customers are charged for vectors stored or indexed — the storage dimension of vector database pricing.
- Per-Character PricingA billing unit where customers are charged per character of text processed — the standard meter for text-to-speech and translation.
- Per-Document PricingA billing unit where customers are charged per document processed or generated — common in AI writing, SEO, and document-intelligence tools.
- Per-Transaction PricingA billing unit where customers are charged per financial or billing transaction processed — the meter of billing and accounting platforms.
- Active-User PricingA billing unit where customers are charged per monthly or daily active user rather than per provisioned seat.
- Per-Task PricingA billing unit where customers are charged per task an automation or agent executes — Zapier's historical unit, now spreading to AI agents.
- Per-Unit PricingA billing unit used by robotics, hardware AI, and some SaaS companies where the metered object is a physical or abstract 'unit' — a robot deployed, a device sold, or a defined deliverable.
- Workflow Execution PricingA billing unit where each end-to-end workflow or automation run is metered and billed, regardless of the compute steps it contains.
- Per-Message PricingA billing unit where each individual message or reply in a conversation is metered, common in AI chat and voice platforms.
- Per-Invoice PricingA billing unit used by billing infrastructure platforms where each invoice generated or processed is metered as the primary cost driver.
- Per-Action PricingA billing unit where each discrete action taken by an AI agent or automation is metered — common in browser automation and agentic workflow tools.
- Per-Image PricingA billing unit where each AI-generated image is metered, common in image generation APIs and multimodal AI platforms.
- Per-Conversation PricingA billing unit where each complete customer conversation — from first message to resolution — is metered as a single chargeable event.
- Per-Record PricingA billing unit where each data record processed, labeled, or extracted is metered — common in data platforms and web scraping services.
- Per-Word PricingA billing unit common in translation and localization platforms where the metered object is the word count of content processed.
- Per-Video PricingA billing unit where each AI-generated video is metered, common in video generation and synthetic media platforms.
- Milestone-Based PricingA billing unit used in drug discovery and biotech AI where payment is tied to achieving defined research milestones rather than time or compute consumed.
- Per-Outcome PricingA billing unit where payment is triggered by verified outcomes delivered — distinct from outcome-based pricing models, this refers specifically to 'outcomes' as a countable billing unit.
- Per-Datapoint PricingA billing unit where each individual data measurement or signal ingested is metered — common in cloud cost intelligence and ML evaluation platforms.
- Per-Interaction PricingA billing unit where each patient-agent or user-agent interaction is metered, common in healthcare AI and customer engagement platforms.
- Data Licensing PricingA pricing structure where access to proprietary datasets or data assets is licensed separately from the software or services, common in AI training data and clinical data platforms.
- Robot-Hour PricingA billing unit where each hour a robot or autonomous system operates is metered — the robotics equivalent of a GPU-hour.
- Per-Contact PricingA billing unit where each contact or lead in the database is metered, common in AI sales development and outbound automation platforms.
- Per-Mailbox PricingA billing unit where each connected email mailbox or sending account is metered, common in AI outbound sales and email automation platforms.
- Browser-Hour PricingA billing unit where each hour of headless browser compute time is metered, common in web scraping and browser automation platforms.
- Per-Generation PricingA billing unit where each AI-generated creative asset — image, video, or design — is counted as a 'generation' and metered accordingly.
- Per-Ticket PricingA billing unit where each customer support ticket handled by an AI agent is metered — common in AI customer service platforms.
- Per-Log PricingA billing unit where each LLM request log ingested or stored is metered — common in AI observability and evaluation platforms.
- Per-Trace PricingA billing unit where each distributed trace — a complete record of an LLM request chain — is metered, common in AI observability platforms.
- Per-IP PricingA billing unit where each IP address or proxy endpoint allocated is metered — used by web scraping proxy providers.
- Per-Device PricingA billing unit where each hardware device or endpoint connected to the AI platform is metered.
- Per-Case PricingA billing unit used in legal AI platforms where each case or matter processed by the AI is metered.
- Per-Report PricingA billing unit where each AI-generated report or analysis document is metered as a discrete output.