Skip to content

Evaluations

  • Context Relevance: Is the retrieved context relevant/useful.
  • Groundedness: Did the LLM generate the answer based on retrieved context or hallucinate.
  • Answer Relevance: Did the generated response answer the user query.