Initial ingestion
One-time cost
Budget estimate for a phased retrieval-augmented search system. All amounts in euros, excluding VAT.
One-time cost
Per month, including cloud rental
Per month, including hybrid search and cloud rental
First 12 months · search only
Server rental for the self-hosted hybrid search service, including vector and keyword index storage and backups.
Enter the document count, pages and an embedding model in Phase 1 to estimate the index size.
One-time total = documents × pages/document × tokens/page ÷ (1 − overlap/100) × embedding rate ÷ 1,000,000 + additional compute
Used only for the index sizing hint.
Affects the index size hint, not the embedding cost.
Monthly usage = searches/month × (query embedding cost/search + reranking cost/search) + document update embedding cost/month
Searches made directly, excluding assistant conversations, which are counted in Phase 3.
Each one is fully re-embedded with the Phase 1 assumptions.
Estimated searches per month: —
Monthly usage = Phase 2 usage + assistant turns/month × [retrieval cost/turn + (input tokens/turn × blended input rate + output tokens/turn × output rate) ÷ 1,000,000]
Every turn runs one hybrid retrieval plus one generation.
Estimated assistant turns per month: —
Complete model input: system prompt, retrieved passages and conversation history.
Include billable reasoning tokens.
0 % is the conservative default. A stable system prompt and long conversations can reach 50 % or more; cache write premiums are ignored.
| Model | Input / request rate | Cached input | Output | Unit |
|---|