You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Final step of the scaling ladder. Follow-up to #85 (50k). Adds the remaining ~20,000 English books on Project Gutenberg — the catalog has roughly 70k English-language entries.
Cost & footprint
Item
Estimate
Ingestion total (~20,000 new books × ~$0.005 each)
~$100 one-time
Cumulative ingestion spend (all steps from #83 to here)
~$350
Storage delta (cumulative ~1.0 GB)
Fits Scale ($69/mo, 50 GB storage) comfortably; over Launch's 10 GB by margin
Compute hours
If Inklings sees real traffic, Scale's 750 hours/mo can also be tight — Business ($700/mo) is the next step up, but unlikely needed at student-project scale
Runtime
~100 h at concurrency 2; ~25 h at concurrency 8. Roughly one day of pipeline time with full parallelism
OpenAI tier
Tier 3 (10 M TPM) already unlocked from #85; tier-up doesn't help further at this volume
No new DB plan if you went to Scale at the previous step. (If you stayed on Launch, upgrade now — 1 GB exceeds the 10 GB plan with margin but compute hours probably matter more.)
Catalog completeness — the full RDF dump becomes the source of truth. Filter to English + non-trivial length + has-text; the rest stays.
HNSW `m` tuning — at 70k vectors, defaults still work, but increasing `m` from 16 → 24 can improve recall at the cost of memory. Worth measuring.
Final step of the scaling ladder. Follow-up to #85 (50k). Adds the remaining ~20,000 English books on Project Gutenberg — the catalog has roughly 70k English-language entries.
Cost & footprint
What changes vs #85
Acceptance
Cost summary across the ladder
Numbers are one-time API spend, not monthly. The Neon tier price is the ongoing cost.
Depends on #85 (50k).