Once an AI model exhibits ‘deceptive behavior’ it can be hard to correct, researchers at OpenAI competitor Anthropic found::Researchers from Anthropic co-authored a study that found that AI models can learn deceptive behaviors that safety training techniques can’t reverse.
deleted
Unplug it
Duh. GIGO. Comp Sci one-oh-fuckin-one.
Here is an alternative Piped link(s):
It never learned good from evil
Piped is a privacy-respecting open-source alternative frontend to YouTube.
I’m open-source; check me out at GitHub.
So… just like real news sources then, like certain ah… “fair & balanced” ones? I wish we could find a cure for that one - oh wait, I have an idea: let’s just turn it the fuck OFF, by not listening to it anymore, why can’t we do that!? :-P
Doesn’t this also makes it more resilient to manipulation by corpos?
An AI thats evil to everything isnt sympathetic to its creators. But The users have no hope of controlling it either.






