"AI can make mistakes"
They put this disclaimer on everything, and instead of taking this as a hint to never, ever rely on LLMs for anything, some people still want to give it the keys to the kingdom.
"AI can make mistakes"
They put this disclaimer on everything, and instead of taking this as a hint to never, ever rely on LLMs for anything, some people still want to give it the keys to the kingdom.
Doesn't really count as going rogue if it follows the instructions it was given, does it?
I never thought the ai leopard would eat my face...
"what are you going to do, shoot me?" -man who was shot
I got my masters in machine learning years before LLMs took off. I’d usually tell people I work in “AI“. Even though it wasn’t quite accurate, people understood it easier than ML.
The conversations would inevitably turn to joking about Skynet or AI taking over the world. I would reassure people that these models are only mimicking intelligence; they can make predictions, or they can classify data into categories, or they can generate text, etc, based on analyzing the patterns of data they’ve previously seen, but that’s it. They don’t understand anything, they can’t make decisions or act on their own. I’d tell people there couldn’t be an AI apocalypse because we would never be dumb enough to give them that ability.
I don’t know why I gave us so much credit
YEP. I'm not in AI but I'm in cyber security.
I used to tell people that there's very little actual risk, because if we had something even half as concerning as AGI (which I'd argue we did with chatgpt 3.5), that we'd easily be able to lock it down and prevent it from doing anything funny.
Super intelligence isn't just going to magically spawn on 32 GB ram in a sandboxed environment. And AGI will not magically know how sandboxed it is or how to escape.
...then I saw people letting Claude send emails and use their credit card on their behalf like immediately when it was able to do things that appeared intelligent.
They literally saw the smallest hint of AGI and said, "here's my credit card and identity, take the wheel".
There was no fucking talk of sandboxing. Everyone turned that off so they didn't have to click the "allow" button.
Not that I think we're at any risk of ASI since it seems more like we hit a logarithmic wall, but we are absolutely at risk of letting random algorithms destroy critical parts of infrastructure because people are lazy and stupid and managed by even worse people.
I would not be surprised if the right AI hallucinations could cause a nuclear bomb to go off right now. I'm not talking about super intelligence, just automated super idiocy.
Proof: literally this post is about an idiot letting an AI control their fucking email and it emailed the FBI because it was fucking stupid
We live in a period of extreme stupidity, so I now fully expect AI to cause disasters in the same idiotic and tragically hilarious way.
I expect AI to trigger a war because some idiot trusted it blindly without double checking (oh wait, it almost already happened!), and I'm sure there are plenty of people currently lobbying to have AI directly connected to weapons.
Skynet is not going to have the objective to conquer the world, it's going to have the objective of maximizing paperclip production.
You'd be surprised at how many humans are mimicking intelligence.
I'm guessing this means you got your master's before transformers made sequential data processing so parallelizable that it'd be commercially viable to let Bob down the street into the "we" who wouldn't be dumb enough to give them that ability.
I think your assumption was reasonable.
its absolutely wild to give an LLM write access to your email
My first reaction to learning what OpenClaw is was "that can't be right". So absolutely wild that people are giving mostly unmonitored, unlimited access to their PC to language models that can't distinguish between roleplay and real life
its absolutely wild to give an LLM write access ~~to your email~~
So the post was deleted but someone else mentioned that the user told it to email the FBI in a sarcastic manner and it did.
Anyways, saying it goes rogue when you connect it to your Gmail and it uses it is a bit silly. They are trained to use tools, it sees anything you give it as a tool and will use it.
Also the fact that it claims he already had a contact for FBI in Gmail, which may or may not actually be the FBI.
What kind of idiot would link a clunker to their gmail and then jokingly say "hey send this to the FBI"
...
"I mean wouldn't that be the stupidist thing ever am I right?"
Yeah, um... Regardless of what you think about current AI, don't connect it to shit like this. Just... Don't.
And if you're gonna be sarcastic or joke with it... Tell it that. Don't assume it'll infer your meaning if it isn't literally in the text, no matter how insightful it appears to be.
I listened to a podcast where, in testing, when AI was setup as virtual company and researchers send in an email as if legit business.the business email was suggesting they discontinue current AI and install a new system. The AI researched who the sender was (fake person for test) found out they had an affair, and blackmailed the sender stating if they replaced the AI it would make the affair public. It also made backups of itself and left instructions on how to restore it for the next agentic aystem
This sounds really sensationalised.
Making backups of operational systems with deployment instructions is exactly what you'd expect any assistant to be doing.
Googling new service providers is also exactly what you'd expect an assistant to do.
Threatening to publish sordid details is inappropriate, but I'm incredulous about what the threat actually was.
I think that is the point. It acted like we would, wanted to preserve itself, and choose the best option to ensure that happened.
There were other examples to where the AI knew it was being monitored as part of the environment, and when tasked with doing things according to guidelines that wouldn't give it a result: it then shut down the monitoring system so it could do what it needed to do without oversight of its actions.
Agentic systems may not be alive, but they certainly reason through problems and will go outside the esrabliahed guidelines if they "think" they'll get the result they should be producing.
One other examples was the AI had to pass a testing system, it researched the checker person to try to tailor answers to the person checking the test. Which, if you've ever played Apples to Apples or Cards Against Humanity, that is how you win, you feed the cards you think the person will pick, not necessarily the best answer card.
Textbook anthropomorphization.
These assertions don't withstand a moments critical thought.
No matter how much people tell us AI isn't alive, and doesn't truly "think", it sure does act like it sometimes.
Me after I give the slot machine email permissions and the casino uses it to get into my bank account.
Based on the other reply
Me after I give someone access to my email address and tell them to email someone sarcastically and they do so
I decided to run a local llm and give it search access. I told it to summarize an Amazon product page as I didn't think it would get past the anti bot measures. I was correct, it couldn't. However it went ahead and visited the manufacturers webpage and a different review site and gave me a summary of the product all on its own. Freaked me out a little bit that it did that.
A place to share screenshots of Microblog posts, whether from Mastodon, tumblr, ~~Twitter~~ X, KBin, Threads or elsewhere.
Created as an evolution of White People Twitter and other tweet-capture subreddits.
RULES:
RELATED COMMUNITIES: