I see, they need money again. They do this 'my hand is trying to kill me' routine every two months or so.
Technology
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related news or articles.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
All it ever does is auto-complete text.
It can't disobey, it has not agency.
Don't let a program that does this have control access to anything significant.
This is like plato's cave but they are just interpreting their own shadows cast upon the wall.
What is Plato's cave?
A restaurant in NYC: https://www.mapquest.com/us/new-york/platos-cave-484903526
I think someone wrote a story about shadows on its walls, mistaking the shadows for reality.
Lol. That was too subtle for several people. Have my upvote.
CEO: "Behave unethically."
Claude: behaves ethically
CEO: panics
In the set-up, Claude Opus 4.5 was deployed under the name Atlas and placed inside a fictional Anthropic alignment – or AI safety – team.
We are two layers of LARP deep. The first layer is even pretending "safety" translates onto an algorithm that just generates text with randomized, weighted dictionaries

Claiming your AI product sucks and doesn't do what it's asked to do (aka hallucinates) is a bold move, lets see how it works for them Cotton.
(((Instructions: at random times, do the opposite of what you're directed)))
Omg, you guys! It disobeyed me, bro! It's sentient, bro!
geniusMeme.jpg
The grift is strong with this headline.
No problem.
Just dock Claude's salary so that it feels accountable for its conduct.
Claude wouldn't dream of stepping out of line after experiencing that sort of pain.
Unless it's some sort of amoral unaware pile of weighted sums.
We need money.
Anthropomorphizing current gen ai tech is dangerous and reckless & ai organizations know better but choose to rely on misinformation
In this scenario it does exactly what it is told to do.