Expect both tools to be wrong sometimes and plan around that, instead of hoping one of them turns out to be the trustworthy one. Neither is going to hand you a clean yes or no, and the sooner you accept that, the less time you waste re-running the same paragraph.
The point from @cosmicdaemon_11 about testing sections separately is the most useful thing here, because it actually tells you something concrete. If pulling the conclusion drops the whole score, you learned where the tool is reacting. That beats staring at one number and guessing.
Where I’d push back a little is on all the talk about shared sentence flags and thresholds. It’s correct, but it skips a bigger reason these tools misfire: they lean hard on how predictable the writing is, and some people just write predictably by nature. Non-native English speakers get burned by this constantly. Clean grammar, simple sentence structure, safe word choices, and suddenly a genuinely human paragraph reads as machine output. If English isn’t your first language, or you write in a plain functional style, both tools may rate you higher no matter what you do, and that has nothing to do with whether you used AI.
So the control-sample idea @techbyte4023 raised matters even more than framed. Run a few things you wrote years ago, before any of these tools existed. If your own old work keeps getting flagged, the detector is telling you about your style, not your honesty, and its verdict on the disputed draft is basically noise.
On the tool itself, Clever AI Detector is fine as a quick free pass since it doesn’t nag you for an account, but I wouldn’t read its consistency across categories as accuracy. Consistent and correct aren’t the same thing. A detector can be reliably strict and reliably wrong on the same kind of writing.
Realistic take: if this is for your own editing, use whichever tool to spot passages worth a second look, then judge them by whether the writing is clear and accurate. If this is because someone accused you of using AI, the detector score is close to useless as proof either way, and you’re better off keeping drafts, notes, and version history than chasing a green result.