Rendered at 06:46:09 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
harrouet 15 minutes ago [-]
Nobody talks about telecom operators. But they will be the ones disconnecting the malicious bots when they see one.
nradov 1 hours ago [-]
OK, I've listened to the warnings and they still sound like the usual nonsense by out-of-touch techies who spend too much time lost in apocalyptic sci-fi fantasies. In the real world it's still hard to move around physical atoms or keep machinery working reliably. AI won't change that.
jocoda 2 hours ago [-]
I can't help wondering if the major labs think that tackling the alignment problem and implementing the 'kill switch' that is currently being proposed is going to be their moat.
ferrouswheel 3 hours ago [-]
The problem is not AI, the problem is humans mis-using AI
justinclift 29 minutes ago [-]
Couldn't it be both? :)
CyLith 1 hours ago [-]
Well, if you really believe that, then we truly are screwed. Trusting humans to not mis-use something is wishful thinking.
brainwad 1 hours ago [-]
The hugging face hack shows otherwise, no? Unless you think humans are at fault even for secretive, autonomous, non-prompted behaviours of their AIs, in which case it's just semantics.
nradov 1 hours ago [-]
Toys like HuggingFace get hacked all the time. So what. In the long run AI automated security scans and penetration testing will be a tremendous aid in detecting and repairing vulnerabilities in systems that actually matter.
brainwad 59 minutes ago [-]
The problem is not per se that it was Hugging Face. It's the wild overstepping of reasonable bounds by itself without any human consultation.
anon48293 50 minutes ago [-]
No, the problem was OpenAI not implementing proper sandboxing or safeguards, and telling the AI exactly to hack things. Thats what exploitgym is, and the task they were given.
This is 100% on OpenAI.
brainwad 31 minutes ago [-]
If your security model is having to imagine all the ways your frontier models might misbehave in novel ways and preemptively sandbox them, you don't have a security model. The only way that will work is general alignment.
This is 100% on OpenAI.