Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
That is against their ToS though if you use your new LLM commercially.
How to Finetune GPT-Like Large Language Models on a Custom Dataset
21–30 of 126 posts
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#22Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
Is "ca" "can" or "can't"?
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#23Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#24Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#25Can someone explain why I'd want to use fine-tuning instead of a vector database (or some other way of storing data/context)?
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#26Can someone explain why I'd want to use fine-tuning instead of a vector database (or some other way of storing data/context)?
My next step is to play with fine tunning, but I have no results to report yet.
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#27Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#28Can someone explain why I'd want to use fine-tuning instead of a vector database (or some other way of storing data/context)?
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#29Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
Yes, almost all improved LLama models are tuned exactly that way (trained on examples of questions and answers from say GPT 4). If OpenAI stole copyrighted works to train their models it is morally fair game to do the same to them regardless of their TOS. It's not like they can prove it anyway.
Plus there's the other point where they also say that everything generated by their models is public domain, so which one is it eh?
Re: How to Finetune GPT-Like Large Language Models on a Custom Dataset
#30Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
> I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? Yes, almost all improved LLama models are tuned exactly that way (trained on examples of questions and answers from say GPT 4). If OpenAI stole copyrighted works to train their models it is morally fair game to do the same to them regardless of their TOS. It's not like they can prove it anyway. Plus there's the other po…
The current attempts to spur on regulation by OpenAI is moat building