Probably no way to do something like that and be imperceptible to humans. Could probably find ways to mangle variable names adversarially, but that would also make it harder for humans. The best thing would probably be trying to do reliable bot detection, then serve the bots adversarially constructed text that *looks" plausible, but is nonsense/broken. I've heard about "tar pits" and "markov tarpits;" I don't know if they try to make the content look legitimate and/or adverserially construct the content to try to maximize damage during training.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: