RAG-Fusion Jev Evaluation
Tests Jev as a typed relevance judge and query router inside a RAG-Fusion retrieval experiment.
Tests Jev as a typed relevance judge and query router inside a RAG-Fusion retrieval experiment.
The RAG-Fusion evaluation script sends a query and candidate documents to Jev. It asks one Noul relevance question per document and uses the returned yes probabilities; a separate Choice request classifies query kind for routing.
Is document D00 relevant to the query?
A query and documents keyed D00, D01 and so on.
Choose the query kind for retrieval routing.
The incoming query text.
A caller starts the retrieval experiment with queries and candidate documents.
Jev returns one Noul relevance probability per document.
A separate Choice classifies the incoming query.
Evaluation scripts consume the signals to assess or adjust ranking and routing.
The experiment separates retrieval into two small typed decisions: what kind of query arrived and which candidate documents are relevant. That makes ranking and routing signals independently inspectable.