Where this breaks

This tool measures wording similarity, not truth. The clearest way to show that is to let it fail in front of you: paste a statement and its exact opposite, and the model still calls them similar — often more similar than a genuine paraphrase. Edit either box below; everything recomputes from real, live, on-device inference.

Loading the model…

Contradiction

Genuine paraphrase

Edit either box above — the numbers recompute from real, live, on-device inference (no server call).