Evaluate Plato's Dialogue
The following prompt tests an LLM's ability to perform evaluation on the outputs of two different models as if it was a teacher.
Prompt Engineering Guidev10 estrellas · 0 bifurcaciones · 0 creaciones≈32 tokens
prompt.md · 128 BSin procesar
Can you compare the two outputs below as if you were a teacher?
Output from ChatGPT: {output 1}
Output from GPT-4: {output 2}