Earlier quoted context omitted.
>Could someone explain in simple terms exactly what fine-tuning does? Fine-tuning shows the model examples of sequences it should produce. The model is updated to become more likely to produce sequences like those examples. What precisely 'like those examples' means for brand new prompts unlike those in the training distribution is the black magic of generalization. >Does it show the model how to answer questions, or…
Thank you for this answer! > I generally recommend retrieval Yes that's what everyone's saying and it's also what we're working on. I was wondering what fine-tuning may be used for. Are there use cases where fine-tuning might be worth it (esp; given all the hard work it entails)? > akin to a student taking a test with open notes they can refer to, rather than trying to remember a textbook they read a week ago Excelle…
GPT-3.5 Turbo fine-tuning and API updates
241–244 of 244 posts
Re: GPT-3.5 Turbo fine-tuning and API updates
#242Earlier quoted context omitted.
I'm curious about this. Can you point me to, e.g. some example code for setting up an inference endpoint with a base llama2 model on modal.com?
Here's one if their tutorials using vLLM, and they have a few other guides and example repos as well. https://modal.com/docs/guide/ex/vllm_inference https://github.com/modal-labs Alternatively, Runpod is fairly cheap and easy to get stuff running in a few minutes and can be point/click only using their templates. https://www.runpod.io/console/gpu-secure-cloud?template=f1pf... ("serverless" example) https://github.com…
Re: GPT-3.5 Turbo fine-tuning and API updates
#243Earlier quoted context omitted.
Can’t you get by with ChatGPT-4 for these personal assistant type questions? That’s what I do and my 20 a month goes a long way. I’d be interested to see if I am missing out on anything using GPT to is way in contrast to the API.
I use it with a tool that is wired into my terminal that changes my files for me [1]. That alone makes me several times more productive compared to copy pasting back and forth between the chat window. If the chat window makes me twice as productive the command line tool probably makes me 5x as productive. At that kind of output on a developer salary the $70-200 a month is absolute peanuts compared to what you get in…
Re: GPT-3.5 Turbo fine-tuning and API updates
#244Earlier quoted context omitted.
This one seems to be a deal-breaker, if you already know what types of language you want, why would you want openai moderating your parameter tuning set.
Why do you care at all, let alone "dealbreaker". You need a model specifically fine tuned towards something dangerous?
why do you care about free speech? i have nothing to say