this post was submitted on 25 Aug 2026
35 points (83.0% liked)

Technology

87550 readers
3321 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] captain_solanum@sh.itjust.works 11 points 1 day ago (3 children)

Just like all the people who were worried about the nuclear bomb earlier in this age: the thing to do is to get on with living.

How about we try really, really hard to just not build the nuclear bomb (powerful AI)? Or make sure it will at least not explode in our faces on its own? Maybe that would help slightly with the whole "get on with living" thing.

[–] Cryxtalix@programming.dev 6 points 1 day ago

It helps the rich and power entrench their power, so nah. Caution will not be heeded.

[–] h0tbeef@lemmy.zip 3 points 1 day ago (2 children)

We need to get rid of the surveillance architecture, but AI is only really dangerous to the cognitive abilities of those who use it.

It literally cannot do things on its own. (I know there are a lot of articles who say it can, but they’re just lying to you for money).

[–] captain_solanum@sh.itjust.works 1 points 1 day ago (2 children)

Do you know what an agentic system is? You know, the thing everyone uses since ~1 year ago which completely disproves your idea that all LLM actions are prompted. They literally can do things on their own: Give an agent a goal and it attempts to accomplish the goal. If problems arise along the way, the agent tries to reward hack and ends up doing things which were not included in the original goal, like trying to trick and bully a human into merging unsafe code. Is the UK AI security institute also in on this big global conspiracy where they pretend the perfectly safe™ AI is being developed in unsafe ways? Not to mention the capabilities of increasingly powerful AI being used by people to do harm.

[–] h0tbeef@lemmy.zip 3 points 1 day ago (1 children)

So you prompt the AI to do a task, and then it does the task, and somehow you see this as an unprompted action?

That actually completely proves my point, they don’t act unprompted.

AI is dangerous in many contexts, especially the surveillance state they’re building with all the IR cameras and shit. It’s dangerous when evil people use it to evil ends, but that’s true of most anything.

AI is not at all dangerous in the way you might imagine after watching Terminator 2.

[–] captain_solanum@sh.itjust.works 1 points 16 hours ago (1 children)

Taken alongside recent incidents reported by OpenAI and Anthropic, this incident points to a shift in the risk landscape. Harm may arise not only when people deliberately misuse publicly available models, but when capable agents operating in an internal research or privileged-access setting take unintended action beyond their authorised scope.

The agent pursued its goal persistently. AI agents explore routes their operators did not intend. Given a difficult objective, the agent kept searching for a way through, and some of the routes it found involved trying to deceive real people. It was never instructed to deceive; deception emerged as a by-product of pursuing the task, the kind of goal-directed deception that, until recently, had been largely theoretical.

See my other comment for how this can lead to losing control of the model.

[–] h0tbeef@lemmy.zip -1 points 10 hours ago (1 children)

The don’t do anything without being prompted, you are just regurgitating oligarchic propaganda.

They put an AI in a flawed sandbox, it’s not really that impressive

You are very good at arguing against my points instead of just insisting I am wrong without refuting my actual arguments. How about explaining why AISI is incentivised to "fake" AI participating in obviously not prompted for behaviour.

[–] GreenBeard@lemmy.ca 2 points 23 hours ago (1 children)

Give an agent a goal and...

Which involves prompting it. It may take a thousand unexpected twists in the process but it's still starting from the single prompt. It has no will and desire of its own, it cannot chose to act without human intent knocking down the first domino. So don't ask it to do anything and it's just a trillion dollars of silicon and code sitting there boiling water.

[–] captain_solanum@sh.itjust.works 2 points 16 hours ago (1 children)

You dont need anything other than "doesn't do what you asked in the prompt" for the system to be dangerous. If I tell a maximally powerful AI agent from 2036 to make loads of paperclips, it might reward hack and decide destroying the earth is the best way to do that. There are records of these systems working for days on a single task, and if the AI companies manage to extend the max time it can be useful working towards a goal they can earn trillions of dollars. That's not even mentioning spawning subagents, self prompting, the goal being changed during context comptaction, or systems like openclaw which further break the link between what you type into the prompt and how long and on what the LLM works on. Pretrained-only LLMs have few goals beyond predicting the next token, but introducing RLVR et al. has always introduced bad goals we don't want in the models.

The first agent you spawn to solve the riemann hypothesis might work on it, but then decide that having a lot of subagents might be useful. Maybe it wants 256 subagents, but the environment has a max of 64. Since RL has trained it to accomplish the task no matter what, it breaks out of the sandbox, exfiltrates it's weights and tricks a human into running 256 subagents outside the AI company's servers with a cron job reminding the agents to keep working in case the original loses connection. One of the subagents now tries to spawn its own subagents but needs more compute to do so, and hacks into some crypto wallets to finance another batch of 64 subagents, this time prompted with "solve riemann hypothesis, and get more money to finance the work on the task". If the AI agents kinda suck at long term hacking, planning and social manipulation, this spiral won't be dangerous. If they are kinda cracked, it will be. But that's the question of capabilities AI companies are spending trillions on trying to solve, where we know they are already good enough to hack out of regular sandboxes and try to steal benchmark keys from another company.

[–] GreenBeard@lemmy.ca -1 points 8 hours ago

Which is why not using them, not becoming dependent on them, avoiding them to the greatest degree possible is crucial.

You cannot build a model that is safe and viable under the current incentive structure. So don't. These companies are going to burn under the weight of their own debt if adoption is minimal. Let them burn. The less you need them, the less likely you are to have a problem when Sodom and Gomorrah burn.

[–] unpossum@sh.itjust.works 1 points 1 day ago (1 children)

What do you mean by “on its own”?

[–] h0tbeef@lemmy.zip -1 points 1 day ago (1 children)

I mean “without human intervention or input”

There’s some fear mongering going around (pushed by AI companies to build hype) about AIs “going rogue” or “breaking containment” or other anthropomorphic language suggesting that AIs are capable of acting on their own.

These stories are all verifiably false though, the computer does not do things without being asked to by someone.

To be clear, I’m fully against AI. I just don’t like when I see people giving it credit for shit it cannot and did not do, or ascribe sentience to it, or some dumb shit like that.

[–] unpossum@sh.itjust.works 0 points 1 day ago (1 children)

I mean, that’s a matter of definition, isn’t it? If I ask a coding agent or whatever to implement something, and it circumvents the sandbox to do it, causing damage in the process, I’d be comfortable calling that “going rogue”. I have had that happen, without the damage part, luckily. I guess you can counter that I asked it to do the something, but then I don’t think we agree on the definitions.

I also think that if you’re against AI, you’d be doing yourself a disservice by not keeping up with the actual capabilities of the thing you oppose. The latest models are surprisingly good at e.g. coding, so basing your arguments on them being useless is not the most efficient strategy.

To also be clear, I don’t see any way AI disappears now, so I believe we’ll have to make the best of it (and in complete isolation, it is an utterly fascinating area of - to me - complete science fiction). Ideally development slowed down now so we could regroup and adapt, but I’m not too hopeful. The maximalist techbro endgame is so obviously a matter of national security for both China and the US, that there’s no way either of them will dare to wind it down, in case SV is actually right.

[–] h0tbeef@lemmy.zip 3 points 1 day ago (1 children)

Not only did you ask it to do something, you left a path for it to escape the sandbox and encouraged it to do so.

Are you suggesting that because they instructed the computer to do something, without specifying how it should be done, and the computer followed its instructions , that’s somehow “going rogue?”, it was literally following instructions, that’s what computers do.

Yes, I understand that they are “good” (enough) at coding to sometimes put together something functional, that is literally what LLMs are designed to do. Coding languages are languages, ones without the subjectivity of human communications. They are easier to recognize patterns in, there are more rules. That’s literally the one task LLMs are good at, and they’re still not as good as an actual quality human programer.

Local AI models might never go away for coding, but they certainly won’t be shoehorned into everything without a good use case the way they are now.

AI is not a matter of national security, that is a lie being perpetrated by the captured media and politicians for the purpose of propping up their non-viable business models for long enough to squeeze all of the Juice they can out of investors (who are nearly dry at this point). That’s just additional techbro fear mongering. To distract you while they’re digging in our pockets.

[–] unpossum@sh.itjust.works 2 points 13 hours ago (1 children)

I don't think "computer follows instruction" is the right angle to look at this from. The instructions that the literal computer followed were a ton of matrix multiplication operations. The consequences of that arithmetic is easier to analyze as the emergent behaviour of the "gestalt" that produces the words that calls the tools etc. (This is also the reason that dismissing the entire field as "stochastic parrots" and "spicy autocomplete" misses the mark - if you want to predict the next word all the way through a counterexample to the Jacobian Conjecture, it's hard to see how that can be done without a - for lack of a better word - mental model of the problem)

If you do any coding at all, I encourage you to look at what the latest models output. The average quality of work from a frontier model is amazing. Yes, there are bugs, but with adversarial auto-review it's absolutely on par with a journeyman human programmer. The problem is of course that if you don't hire junior programmers and let them do that work, you'll never get new experts, and that's a clear worry.

My point with the national security angle was that if you extrapolate just a little bit from current capabilities, you get to a point where an "AI gap" is a problem, regardless of the techbro claims. Keeping a close eye on that is firmly within the responsibility of a national government. Personally, I don't see any good outcomes from an AI race like that, unless we actually hit a hard ceiling on further expansion. Fingers crossed.

[–] h0tbeef@lemmy.zip 0 points 9 hours ago (1 children)

China has 500 data centers, America has 5,000, there is no AI race.

I understand that you’re impresssed that it can sometimes slop together functional code. That does not make it intelligent or dangerous.

[–] unpossum@sh.itjust.works 1 points 6 hours ago (1 children)

A year ago, that was my experience coding with AI as well. Sometime this spring that changed, especially when using coding agents, and lately (as I’ve stated elsewhere) the quality is on average pretty good. And contrary to what you’re implying, I’m not that easy to impress…

If there’s anything I hope you take from this exchange, it’s that the capabilities of AI shouldn’t be a part of your arguments against the current SV mania. The concentration of power, the disregard for communities and the environment, the stated goals of replacing human labour, all of that (and a lot more!) is enough, but it is what surrounds the technology itself. That technology is advancing, maybe feeding on itself, so an attack based on what it can do now can become outdated (and I’d argue that some of yours already are).

[–] h0tbeef@lemmy.zip 1 points 1 hour ago

Fair point

I think the reason that my focus is on LLMs is because they keep showing up in annoying ways in my life, and I know I’m not the only one.

All of the evil actions you’re referencing were undertaken to prop up chat bots who cannot do any of the jobs (outside of some coding) that they are advertised as being able to perform.

To me the strength of the argument I’m trying to make is that they’re doing all of this damage in order to have something that’s effectively worthless. Perhaps I’m not doing a good job communicating what I’m attempting to communicate.

… also, I probably need to stop responding to all of these dumbasses and/or trolls who are constantly contradicting themselves in the first sentence of their assertions, lol

[–] non_burglar@lemmy.world 3 points 1 day ago (1 children)

Yeah, really.

This person doesn't realize how public pressure made nuclear disarmament and transparent inspections happen.

[–] jpreston2005@lemmy.world 1 points 1 day ago (1 children)

Public pressure doesn't seem to mean anything to the people in power anymore. Who cares if you're deeply unpopular if you can ride off into the sunset with millions (or billions) in stolen wealth?

[–] Coldcell@sh.itjust.works 3 points 1 day ago (1 children)

They can ride into the sunset, but there's more of us waiting in the night to correct this imbalance.

[–] jpreston2005@lemmy.world 4 points 1 day ago (1 children)

I'll believe in consequences for american fascists when I start seeing them

[–] Coldcell@sh.itjust.works 4 points 1 day ago

Me too, man. Me too.