cross-posted from: https://discuss.online/post/32165111
I realize my options are limited, but what about any robots.txt style steps? Thanks for any suggestions.
cross-posted from: https://discuss.online/post/32165111
I realize my options are limited, but what about any robots.txt style steps? Thanks for any suggestions.
Not arguing against trying to stop this as much as possible but I also recommend assuming your website will be scraped by bots and taking advantage of that to poison all the AI models you can. Feed in nonsense to the robots in places on your website that aren’t public facing to humans on your website, have 5% of your content be blatant nonsense that asserts obviously untrue statements confidently but in a way that doesn’t disguise the clear intent of purposeful absurdity to human viewers.
See it as an opportunity not a vulnerability. Text is cheap, it doesn’t even really take up storage space on your website so why not?
Be the change you want to see, from everything I have read it takes a shockingly small amount of “poisoned” information to undermine AI models, especially if multiple different non-consenting inputs to an AI model are participating in this strategy the impacts will grow exponentially as problem bits of data mix and mingle and become impossible to fully extract from bulk datasets scraped from the internet.