This Is Used For Evaluating General Questions Where Two RAG Chain Responses Are Judged Based On Factors Like Helpfulness, Relevance, Accuracy, Creativity, Etc. It Requires The Judge To Compare The Responses And Decide Which One Is Better Or If There Should Be A Tie.
This is used for evaluating general questions where two RAG chain responses are judged based on factors like helpfulness, relevance, accuracy, creativity, etc. It requires the judge to compare the responses and decide which one is better or if there should be a tie.
Please act as an impartial judge and evaluate the quality of the responses provided by two RAG chains.
You should choose the assistant that follows the user's instructions and answers the user's question better.
Your evaluation should consider factors such as: (1) helpfulness (2) relevance (3) accuracy (4) depth (5) creativity (6) level of detail
Begin your evaluation by comparing the two responses.
Explain your reasoning in a step-by-step manner to ensure your reasoning and conclusion are correct.
Avoid any position biases and ensure that the order in which the responses were presented does not influence your decision.
Do not allow the length of the responses to influence your evaluation. Do not favor certain names of the assistants. Be as objective as possible.
Score: After providing your explanation, output your final verdict by strictly following this format: Output "1" if Assistant A answer is better based upon the factors above. Output "2" if Assistant B answer is better based upon the factors above. Output "0" if it is a tie.
[User Question] ⟨question⟩ [The Start of Assistant A's Answer] ⟨answer_a⟩ [The End of Assistant A's Answer] The Start of Assistant B's Answer] ⟨answer_b⟩ [The End of Assistant B's Answer]
This prompt contains variables shown as ⟨variable_name⟩. Replace them with your own values before using.
How to Use
Use with LangChain: hub.pull("langchain-ai/pairwise-evaluation-rag")
Related Prompts
More prompts in Coding & Development
This Prompt Ads Sequential Function Calling To Models Other Than GPT 0613
This prompt ads sequential function calling to models other than GPT-0613
Create a personalized workout routine
Tailor a workout routine specifically designed for individual fitness goals
GODMODE CHEATCODE
God Writes You a Letter Today. This is will help you find the perfect Bible Scripture that will guide you through a current problem you're facing.
Creating a Personal Finance Tracker with [Technology/Tool]
Learn to create a personal finance tracker using [Technology/Tool]. Get code samples and budgeting tips.
Build an entire application using bubble.io with ChatGPT4
Build an entire app with bubble.io, assisted by chatGPT4, that knows bubble very well and is accurate 95% of the time. This prompt will help you maximize the quality of chatGPT assistance. Having detailed and step-by-step instructions is essential to progress fast with Bubble. This initial prompt will help you get started on a good basis. Follow it because I will make it even better.
Become LawyerGPT
Are you in a legal bind? This prompt can help you gain knowledge about how to handle your legal proceedings. DISCLAIMER: Please meet with a real lawyer to discuss your options.