For AI watermarking text (like anthropic), how does it work in the case of partial things?
Example 1: I have a transcript of a talk. I want to use AI to clean it up to be readable. Remove clutter, filler words, stutters. And re-organize a bit to move all the all the audience questions to the end. Then I notice the speaker rambled a bit and said the same thing 3 times, so I consolidate it down to 1 sentence and fix the minor transitions needed for that. But basically trying to be as faithful as possible (even in word choice) to the speaker/transcript, just make is a pleasant read.
Example 2: What if I write a post by hand. It's almost perfect. But I have 3 examples in the post and I want to put them in a different order. So I ask AI to put example #
3# first. I might need to remove some terminology explanation from the original #
1# and put in the new #
1# so the reader gets all terms explained when they first encounter them. And a few minor transitions might need editing. But mostly just moving things around.
Might some parts get watermarked, like the transitions or consolidation bits if there's a few different wordings that'd be fine?
How much of the piece needs to be watermarked for it to identify the general post as watermarked? Is it dependent on how confident the model is in the tokens it chose?