Faithful Summary: test results
Tested 2026-09-30, skill version 1.0.0 at the time of loading this page. We run every skill on a strong and a weak model before it is listed, and publish both verdicts, including where the weak one fails.
Verdicts
- Date
- 2026-09-30
- Strong · claude-sonnet-5-5 (Claude Code alias "sonnet")
- Same language on en/bg/de/es/ru, exact numbers and names, no opinions. Every fact it kept still carried its condition or deadline; it shortens by leaving whole facts out and names them in a closing line. 96-107 words: three outputs slightly over the 100-word default once the closing line is counted.
- Weak · claude-haiku-4-5-20251001 (Claude Code alias "haiku")
- Passes all six cases (en/bg/de/es/ru + contradictions), 86-102 words, and keeps the named conditions. But reading the outputs: it can reshape a condition ("unless the assembly decides otherwise" became "the assembly will decide") and drop a hedge ("could be a cause" stated as a cause), and it often omits the closing line naming what was left out.
Note
Tested on six fictional texts of 241-345 words: default at most 100 words, most important facts kept whole with their conditions, a closing line naming what was left out. Machine checks catch only the conditions they name, not reshaped wording; one run per model and case. Details: test results files.
What was not measured
- Models other than the two named above were not run.
- Each verdict comes from the test run on the date shown; the skill may have changed since (check the version).
- Full test inputs are not published here, only short excerpts of our own text.
- Results on your own texts, languages and domains can differ.