Skip to content
All projects

AI & LLM Agents

Production-Grade RAG System

Retrieval-augmented generation with hybrid search, semantic caching and RAGAS-based evaluation.

Role
Developer
Production-Grade RAG System cover

Overview

A RAG pipeline designed like a production service rather than a demo:

  • Hybrid search (keyword + vector) for better recall
  • Semantic caching to cut latency and LLM cost on repeated questions
  • Systematic quality measurement with RAGAS (faithfulness, answer relevance, context precision)

Repository link will be added when the code is published.