Your point was that its an unfair comparison because AI its a race. My point is that bioweaponds are also a race, but a less public one. You dont have the equivalent of Dario publishing an essay every month.
So what's the fuss about then? Is the idea that a Chinese company will release a model that will have no safeguards? For what purpose?
Basic safeguards are all that's required, and they've been there in every usable model since GPT-2, including Chinese models that are supposedly "unsafe".
Or are we saying that some lunatics will start training their own models, spin up a GPU cluster, run some abliteration workflow, or learn how to jailbreak?
That would be a very dedicated person. And dedicated person doesn't need an LLM. So where are they?
Yes, that is exactly Dario's concern. Either one of the US labs or one of the Chinese ones will eventually release something with insufficient safety controls for its power level because it gives them slightly better user retention (look how much complaining there is about current frontier models, especially Fable, rejecting requests). Regulation or consortium is how you avoid the prisoner's dilemma.
As long as user provides inputs and LLMs stay LLMs, you can waltz through any guardrail. Fable is the extreme case, but it's not that hard if you know what you're doing and know how LLMs and their guardrails work.
Am I saying that guardrails don't work? No, they probably stop a lot of insane people trying insane things. But you don't need Fable-level guardrails to do that. You probably don't even need to do anything during pretraining, or RL, or classification to make sure model refuses to compy with "hack me a bank" or "make me a chemical weapon".
All models will automatically have guardrails just as a result of training on data that gives them intelligence. You have to actually train it to be malicious to produce something what Dario calls "insufficient guardrails".
No guardrail is going to stop a determined person with sufficient intelligence. It only has to stop ones with insufficient one, and even basic guardrail that are just by-product of training is going to achieve that.
It seems like it would be really useful for a frequent traveler/commuter, to be able to watch something on a flight/bus ride. I would definitely consider it if I were in that demographic
I bought my Steam Deck for a similar imagined use case but I find myself way too aware of the visibility of a larger than average screen to all those around me.
Do you think there’s some cabal of executives trying to use more resources so that users are forced buy newer hardware?
Or is it more likely that software companies are making a tradeoff based on typical consumer hardware and would rather prioritize making $$$ over optimizing their software?
Blue collar maintenance worker here. I will have a yearly average income much higher than most of my peers because I'll likely die before I can retire and spend it ;)
Hugging Face showed that AI can do serious hacking without really being told to. If a model had its own motivations there could be real damage.
reply