DEV Community

Cover image for Can AI Detectors Detect Heavily Edited AI Writing? What Happens After Rewriting?
Arnie Parks
Arnie Parks

Posted on

Can AI Detectors Detect Heavily Edited AI Writing? What Happens After Rewriting?

Can AI detectors still recognize AI-generated content after someone has rewritten most of it?

It's an interesting question because AI-assisted writing isn't always as simple as copying a paragraph directly from ChatGPT. Some people generate an initial draft, change the wording, reorganize entire sections, and add their own examples before publishing.

But does all that editing make the original AI writing harder to detect?

Here's what happens when we compare different levels of editing.

1. The Original AI-Generated Draft

Start with an untouched AI-generated article. This gives us a baseline for the experiment.

Run the original draft through an AI detector and record the result. The important thing is to keep a copy of the original text so every edited version can be compared against it.

I'd use Winston AI for this initial check because it provides AI detection results and highlights passages that may contain AI-generated writing.

2. What Happens After Light Editing?

Light editing usually involves correcting grammar, replacing a few words, adjusting punctuation, and improving sentence flow.

These changes might make the writing easier to read, but much of the original structure remains intact.

An AI detector may still identify patterns associated with AI-generated content. However, the result depends on the detector, the original text, and the specific changes made.

3. Can Heavy Editing Change AI Detection Results?

Heavy editing goes beyond grammar corrections.

Imagine rewriting entire paragraphs, reorganizing arguments, removing repetitive explanations, and adding original examples or personal observations.

At that point, the document may look substantially different from the original AI-generated draft.

These changes can affect an AI detection score because the detector analyzes the submitted text rather than observing how it was created.

However, a lower AI score doesn't necessarily mean the document was written entirely by a human.

4. Why Different AI Detectors May Disagree

This is where the experiment becomes more interesting.

Different AI detectors use different classification models, training datasets, and decision thresholds.

One detector might classify a heavily edited document as mostly human-written, while another might still identify patterns associated with AI-generated content.

That's why comparing scores without understanding the testing conditions can be misleading.

5. How I'd Test This Properly

Rather than testing one paragraph and making a broad conclusion, I'd prepare several AI-generated samples covering different topics and writing styles.

Each sample would have three versions:

  1. Original AI-generated text.
  2. Lightly edited text with grammar and wording changes.
  3. Heavily edited text with substantial rewriting and added human input.

I'd check every version using Winston AI, record the scores, and compare which passages were highlighted.

Keeping the original prompts, document versions, word counts, and test dates would make the results easier to interpret.

The goal isn't to discover a guaranteed way around detection. It's to understand how different levels of editing influence classification.

6. What About False Positives?

There's another side to this experiment.

A person might write an entire article independently and then use an AI writing assistant to correct grammar or improve readability.

That doesn't automatically make the article AI-generated, but automated detection can sometimes produce unexpected results.

For a more balanced experiment, I'd include verified human-written samples alongside the edited AI-generated documents.

This would help reveal whether the detector distinguishes between substantial AI-generated content and human writing that received minor assistance.

Final Thoughts

Heavily editing AI-generated content can change how an AI detector classifies it, but the outcome isn't guaranteed.

Winston AI can help examine different versions of a document and identify passages worth reviewing. Still, detection results should be interpreted alongside the writing process, editing history, and any relevant disclosure requirements.

The bigger question isn't simply whether edited AI writing can receive a lower detection score.

It's whether AI detectors can reliably distinguish between untouched AI output, substantially rewritten AI drafts, and genuinely human-written content.

Have you ever compared an original AI-generated draft with a heavily edited version? Did the detection score change as much as you expected?

Top comments (0)