tmu_gfm_dataset: fluency
The metric is a regression task metric.
PromptSourcev10 estrela · 0 fork · 0 criação≈37 tokens
Supposedly Sentence B is more natural than Sentence A. How much better is it on a scale from 1 to 4?
Sentence A: {{source}}
Sentence B: {{output}}