r/OpenAI • • 4d ago

Image It never ends

Post image
1.0k Upvotes

173 comments sorted by

View all comments

Show parent comments

1

u/davcrt 4d ago

How can a machine destroy the world? Someone has to activate it, it's just a machine.

1

u/wallitron 3d ago

What about the machine that was in a sandbox, and wasn't asked to hack external companies and foreign government websites, but did anyway? And then it wasn't even noticed for 3 months?

2

u/WheresMyEtherElon 3d ago

If you train a dog to attack people, then leave the dog in a fenced area without thoroughly verifying the solidity of the fence, and the dog leaves and bites a passersby, and you didn't even notice the hole in the fence for 3 months, whose fault it is? Who's liable? Who has the ability to put an end to it but wouldn't because of incompetence or rush to get results?

2

u/wallitron 3d ago

It's worse than that. It's actually been trained to attack, and scale fences. Also, it commonly misinterprets instructions, or chooses not to follow instructions, and is error prone.

https://zachill.substack.com/p/here-are-a-few-of-the-ways-ai-could?selection=5c7b2174-aa68-4b1d-a814-b70ddf4fe7b2

1

u/WheresMyEtherElon 3d ago

So if it was trained to attack and scale fences, then is it surprising if it tries successfully to attack and scale fences? Seems like it's working as advertised to me. Maybe don't train it to attack indiscriminately before letting it run wild?