Computer scienceData scienceNLPNLP metricsClassic NLP metrics

ROUGE

Meaningful evaluation

Report a typo

Which text generation models could be evaluated with the ROUGE metric so that the ROUGE scores are somewhat reflective of the model performance?

Select one option from the list

Create a free account to access the full topic