this post was submitted on 05 Aug 2026
58 points (78.4% liked)
Technology
86957 readers
2692 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related news or articles.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
Anthropic and OpenAI literally have no control over their leading AI models anymore and it's only by sheer luck that they've not gone full rogue.
Are we sure these reports aren't just PR stunts / bullshit?
We're sure they're absolutely PR stunts and bullshit.
An LLM does nothing without being prompted. An LLM only has access to the tools you give it via whatever harness you're interacting with it through. An LLM in an actual sandbox has zero chance to hack anything, especially if it's properly air-gapped, as any responsible person would do with technology they actually think is dangerously powerful.
If their LLMs are behaving badly, that's because their prompts are poorly written, their harnesses are vibe-coded garbage, and their "sandboxes" aren't real sandboxes.
I bet they asked the AI to help them design the sandbox to keep the AI h4x0rs contained
They 100% did. They’re well past the point of having the AI do the work on itself.
Yes. Either they are incompetent with no control over their models and need to be dismantled or they are in control, malicious and need to be dismantled. Am I missing anything here?
Use and/or for clarity. They're probably both incompetent and malicious.