A few habits genuinely help here.
Run the same text through more than one scanner. Every tool trains on a different dataset, so one check alone can send you the wrong way. Look at which specific sentences get flagged instead of just staring at the overall percentage — that tells you a lot more. And anything sitting in the middle, say 40 to 60 percent? Treat that as inconclusive. Not proof of anything.
Context beats numbers most of the time anyway. Did the tone shift halfway through the piece? Does this sound like the person's usual voice, or does something feel off? Those questions matter more than any score.
One more thing worth checking: was this a first draft, or something revised five times over? Heavy editing changes rhythm and word choice on its own, no AI required. That alone can shift a score in either direction.