ROUGE scores summarisation by measuring how much of a human reference summary appears in the generated one.
BLEU emphasises precision and was built for translation. ROUGE emphasises recall and was built for summarisation, where missing a point matters more.
ROUGE scores generated summaries by measuring how much of a human reference summary appears in the output. It emphasises recall, meaning coverage of the reference.
That emphasis is the difference from BLEU. In summarisation, missing an important point matters more than adding a few extra words, so the metric rewards covering the reference.
Think of it like this. Think of marking a summary against a checklist of points that should be mentioned. You are scoring what was covered, not penalising every extra sentence.