Skip to content
#

ragas-evaluation

Here are 137 public repositories matching this topic...

An enterprise-grade, full-stack AI travel planner which provides data-driven itineraries for Lucknow, India and showcases production-ready architecture, combining a FastAPI backend with a Streamlit frontend. It leverages an advanced agentic RAG system, context-aware responses by integrating a local knowledge base with live, external APIs.

  • Updated Jul 9, 2026
  • Python

Universal Agent Evaluation Framework (UAEF) is a framework-agnostic evaluation system for AI agents. Invoke any agent and score it on multiple metrics spanning tool calling, response quality, safety, performance, and reasoning. Track experiments against baselines, catch regressions automatically, and get LLM-generated insights.

  • Updated Sep 21, 2026
  • Python

A high-performance Retrieval-Augmented Generation pipeline for technical Q&A workloads. Combines hybrid retrieval (dense + BM25), query expansion, Reciprocal Rank Fusion (RRF), and cross-encoder re-ranking to improve retrieval precision and answer grounding. Evaluated with Ragas, showing measurable gains in context recall and faithfulness.

  • Updated Apr 30, 2026
  • Jupyter Notebook

Add this topic to your repo

To associate your repository with the ragas-evaluation topic, visit your repo's landing page and select "manage topics."

Learn more