Pricing

Priced on what actually costs us money.

Ingestion is billed per page, not per document. Hosted models are metered only when troveGEN supplies them — on BYO Vector DB, every model you bring is metered at zero. One global price in USD, and a hard stop instead of a surprise invoice.

Start here

One free budget, either line.

Every account begins on the trial. It grants hosted vectors AND bring-your-own, so you can evaluate on whichever line you actually intend to buy rather than discovering the difference after paying. It is a one-time budget, not a monthly allowance — a real corpus to judge retrieval quality on, but nothing you could operate a product on.

Free Trial

$0

Evaluate on your own documents.

  • 500 pages — one-time, not monthly
  • 2,000 searches · 200 answers
  • 1 project · 1 GB · 100 MB per document
  • 100 OCR pages
  • Hosted vectors or your own store
  • Full product, no feature locks
Choose plan

BYO Vector DB · pipeline + search

Your vector database. Everything else from us.

Parse, OCR, chunk and embed into your own store, then hybrid-search it with reranking, filters and ACL enforcement. Ingestion and search are one plan, not two: troveGEN can only search an index it populated, so selling retrieval separately would have been selling something that cannot stand alone. You bring the storage — and every model you bring is metered at zero.

BYO Starter

$49/mo

Everything the two old lines gave, at one price.

  • 10,000 pages / month
  • 50,000 searches · 5,000 answers
  • $0.0049 per page
  • 3 projects · 2,000 OCR pages
  • Hybrid search + reranking
  • Premium connectors
Choose plan

BYO Pro

$199/mo

Production volume with governance.

  • 60,000 pages / month
  • 400,000 searches · 50,000 answers
  • 15 projects · 12,000 OCR pages
  • Governance: ACL, PII vault, versioning
  • Agentic retrieval · LLM reranking
  • Priority support
Choose plan

troveGEN-Z · fully managed

Nothing to run, nothing to connect.

troveGEN hosts the vectors, the embedding model, the reranker and the LLM. Paid tiers get a dedicated database schema and vector namespace, not a shared table. This is the line for teams who want retrieval to be someone else's operational problem — if you would rather own the stack, that is BYO Vector DB above.

troveGEN-Z Launch

$25/mo

A side project or an early-stage team.

  • 1,500 pages / month
  • 8,000 searches · 1,500 answers
  • 2 projects · 2 GB
  • 300 OCR pages
  • Hosted vectors included
Choose plan

troveGEN-Z Growth

$349/mo

Scaling products with governance needs.

  • 40,000 pages / month
  • 200,000 searches · 25,000 answers
  • 20 projects · 75 GB
  • 8,000 OCR pages
  • Governance: ACL, PII vault, versioning
  • Agentic retrieval · LLM reranking
Choose plan

troveGEN-Z Scale

$999/mo

High-volume production workloads.

  • 150,000 pages / month
  • 800,000 searches · 100,000 answers
  • 60 projects · 250 GB
  • 30,000 OCR pages
  • Everything in Growth
  • Priority support
Choose plan

Sovereign · self-hosted

Your infrastructure, your rules.

troveGEN deployed inside your network or private cloud, air-gapped if required.

Sovereign

Quoted

Regulated, air-gapped or data-residency-bound deployments.

  • Unlimited pages, searches and projects
  • Runs fully air-gapped
  • Custom models · data residency
  • Negotiated SLA
  • Priced on deployment scope
Talk to us

Hosted model usage

Only when we supply the model.

These are pass-through rates for third-party model spend troveGEN actually pays for. They apply in full on troveGEN-Z, which is fully managed by design. On BYO Vector DB you connect your own embedder, reranker and LLM — and every one of them is metered at zero, which is the concrete dollar value of owning your stack.

Metered unitIncludedOverageOn BYO Vector DB
Embedding tokens2M – 400M / month by plan$0.30 – $0.60 per 1MNot metered
Reranking calls500 – 300,000 / month by plan$4.00 – $7.00 per 1,000Not metered
Generation tokens1M – 430M / month by plan$3.00 – $6.00 per 1MNot metered
OCR pages100 – 30,000 / month by plan$6 – $10 per 1,000Always ours

How it compares

The category has a floor. We sit under it.

Enterprise managed RAG carries a $50–$700/month infrastructure floor before a single query. Published rates, checked August 2026.

ServicePublished rateNote
troveGEN BYO Starter$0.0049 / pageSearch included, and vectors stay in your own database
Ragie (Starter)$0.010 / pagePlus $0.002 per stored page per month
LlamaParse$0.00125 – $0.056 / pageParsing only — no retrieval
Azure Document Intelligence$0.0015 – $0.030 / pageParsing only
AWS Bedrock Data Automation$0.010 / pageParsing only
Pinecone (Standard)$50 / mo minimumStorage + reads, no ingestion pipeline
Bedrock KB + OpenSearch Serverless$345 – $700 / mo floorBilled before a single query
Azure AI Search$73 – $1,014 / moAlways-on capacity

Billing questions

How the meter works.

Why bill per page instead of per document?

Because a document is unbounded. A one-kilobyte note and a 400-page scanned manual are both "one document", but the manual costs roughly 400 times as much to parse, OCR, embed and store. Per-document pricing lets one customer subsidise another by orders of magnitude and invites gaming by concatenation. A page is the parser's real page count where the format has one, otherwise about 3,000 characters of extracted text — measured on extracted text, so you never pay for markup or whitespace.

What counts as hosted-model usage?

Embedding tokens and reranking calls, but only when troveGEN supplies the model. On BYO Vector DB you connect your own OpenAI or Cohere key and those calls are metered at zero here — you pay your provider directly. troveGEN-Z is fully managed, so every model is ours and these rates always apply. Published rates are $0.30–$0.60 per million embedding tokens, $4.00–$7.00 per thousand reranks, and $3.00–$6.00 per million generation tokens. Reranking is priced per search query rather than per token, and is the single largest model cost we carry — the published floor sits above what a cross-encoder provider charges us, so no tier is sold below cost.

Which line should I pick?

It comes down to one question: do you want to own the vector database? If yes, take BYO Vector DB — you run the store, connect your own embedder, reranker and LLM if you want them, and pay nothing here for models you supply. If you would rather not operate any of that, take troveGEN-Z, where we host the vectors, the embedding model, the reranker and the LLM. A workspace holds one subscription, so the lines never stack, and the free trial grants both paths so you can try the one you intend to buy.

Is OCR charged separately?

Yes, because it is the most expensive step in the pipeline by an order of magnitude. Each plan includes an OCR page allowance, with overage at $6–$10 per 1,000 pages. You are billed for pages actually recognised, not the document's page count — a capped run reads fewer, and charging for pages we never looked at would be wrong. You can also set OCR to off, in which case a scanned file is rejected with a clear reason rather than silently indexed as an empty shell.

What happens when I hit a limit?

The request is refused with a clear message naming the metric and the number. There is no silent metered overage — we do not have mid-cycle invoicing, and quietly running up a bill we cannot invoice would be worse than an honest stop.

Is pricing the same everywhere?

Yes. One global list price in USD, wherever you are. Data residency is a deployment choice on the Sovereign tier, not a pricing tier.

Can I change plans later?

Yes, and upgrading from the free trial carries your existing corpus across into your new dedicated schema in a single transaction — your documents are never in two places or neither. Annual billing is ten times the monthly price, so two months are free.

Start on the free trial.

500 pages and 2,000 searches, one time. Enough to load a real corpus and judge the retrieval quality on documents you actually care about.