It fetches both pages, extracts their visible text, and compares overlapping five-word phrases between them to produce a similarity percentage — the same shingling technique used by most near-duplicate detection tools. A handful of example overlapping phrases are shown as evidence.
Why this matters for SEO
Search engines generally pick one version of near-duplicate content to show in results and treat the other as redundant. If that's not the outcome you want — two pages competing for the same ranking instead of each earning its own — either differentiate the content or point one at the other with a canonical tag.