Detectors do not reliably label "this was Claude" versus "this was ChatGPT". They estimate how generated the prose looks. Claude output can score high, low or mixed depending on length, prompting and how much a person edited afterward.
If you need model-aware intuition, look at sentence-level signals and at comparisons that discuss coverage honestly. Do not expect a detector to name the model that produced a paragraph.

