Elektrine lite

← Feed

@supersquirrel@sopuli.xyz

Post #67457

2025-12-15 23:04 UTC

In the realm of LLMs sabotage is multilayered, multidimensional and not something that can easily be identified quickly in a dataset. There will be no easy place to draw some line of “data is contaminated after this point and only established AIs are now trustable” as every dataset is going to require continual updating to stay relevant. I am not suggesting we need to sabotage all future endeavors for creating valid datasets for LLMs, I am saying sabotage the ones that are stealing and using things you have made and written without your consent.

Replies (1)

  • @Grimy@lemmy.world 2025-12-15 23:17

    I just think the big players aren’t touching personal blogs and social media anymore and only use specific vetted sources, or have other strategies in place to counter it. Anthropic is the one that told everyone how to do it, I can’t imagine them doing it if it could affect them.

    Open ##67505