← Back to context

Comment by bamboozled

2 hours ago

Go and get one of their models to hack something, it won't do it, why?

They have claimed this happened during a "training run", but why are they training on systems connected to the internet?

That's why people are skeptical.

The public models won’t hack because they have a classifier that shuts down anything that looks like hacking; without the classifier they are perfectly capable of hacking, multiple third-party evaluators have confirmed this.

The models were not trained on systems intentionally connected to the internet; they chained mutliple zero-days (that they discovered) together to get access to the open internet and into huggingface.