A production-grade Retrieval-Augmented Generation pipeline with hybrid BM25 + pgvector retrieval, RAGAS evaluation, LangSmith observability, and full GCP deployment via Terraform + GitHub Actions.
Ingestion, retrieval, and evaluation pipelines running on GCP Cloud Run backed by PostgreSQL + pgvector and Redis caching.
Connected to the production API. Upload a document, query it with three retrieval strategies, and evaluate quality with RAGAS.
Drop a file or browse
PDF or TXT · up to 50 MB
Results appear here after a query
Run the same query across all three retrieval strategies simultaneously and compare ranking differences.
Evaluate retrieval quality with four RAGAS metrics. Add multiple samples for aggregate scoring.
Full OpenAPI docs available at the live Swagger UI.