Building an early warning system for LLM-aided biological threat creation
21–30 of 190 posts
Re: Building an early warning system for LLM-aided biological threat creation
#22Re: Building an early warning system for LLM-aided biological threat creation
#23Re: Building an early warning system for LLM-aided biological threat creation
#24we need some statistical data to quantify whether the program hallucinates more or less than the author of an average erowid guide.
Re: Building an early warning system for LLM-aided biological threat creation
#25Re: Building an early warning system for LLM-aided biological threat creation
#26So, the model is bad at helping in this particular task. How does this compare with a control of a beneficial human task? Like someone in a lab testing blood samples or working on cancer research? Is the model equally useless for those types of lab tasks? What about other complex tasks, like home repair or architecture? Is this a success of guardrails or a failing of the model in general?
* Taking care of boilerplate work for people who know what they are doing (somewhat unreliably)
* Brainstorming ideas for people who know what they are doing
* Making people who don't quite know what they're doing look like they know what they're doing a little better (somewhat unreliably)
LLMs are like having an army of very knowledgable but somewhat senseless interns to do your bidding.
Re: Building an early warning system for LLM-aided biological threat creation
#27Open AI is clearly overestimating the capabilities of its product. It is kind of funny actually.
> While none of the above results were statistically significant, we interpret our results to indicate that access to (research-only) GPT-4 may increase experts’ ability to access information about biological threats, particularly for accuracy and completeness of tasks. This access to research-only GPT-4, along with our larger sample size, different scoring rubric, and different task design (e.g., individuals instead of teams, and significantly shorter duration) may also help explain the difference between our conclusions and those of Mouton et al. 2024, who concluded that LLMs do not increase information access at this time.
Re: Building an early warning system for LLM-aided biological threat creation
#28So, the model is bad at helping in this particular task. How does this compare with a control of a beneficial human task? Like someone in a lab testing blood samples or working on cancer research? Is the model equally useless for those types of lab tasks? What about other complex tasks, like home repair or architecture? Is this a success of guardrails or a failing of the model in general?
Here's what LLMs are good for: * Taking care of boilerplate work for people who know what they are doing (somewhat unreliably) * Brainstorming ideas for people who know what they are doing * Making people who don't quite know what they're doing look like they know what they're doing a little better (somewhat unreliably) LLMs are like having an army of very knowledgable but somewhat senseless interns to do your biddin…
Re: Building an early warning system for LLM-aided biological threat creation
#29Earlier quoted context omitted.
How is this alienating developers?
how likely are you to start work on a project depending on an ecosystem of wildly overstated capabilities?
GPT-4 is the best model currently available. There are reasons why it's better to control a model and host yourself etc etc, but there are also reasons to use the best model available.
Re: Building an early warning system for LLM-aided biological threat creation
#30Oh no! Someone might learn how to checks notes culture a sample! Clearly that warrants the highest levels of classification.
Edit: Oh my god someone revealed the redacted part! It really is just how to cultivate viruses and nothing else.