Live data from Hacker News

Hello Dolly: Democratizing the magic of ChatGPT with open models

databricks.com

161–170 of 194 posts

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#161
post #101

Earlier quoted context omitted.

data transfer might actually be the problem there not something like trying to hide the model

bittorrent, come on

context `come on`, what is the point in sharing the weights ?

Someone lay out the reason they should package the weights with this, when they're allowing you to apply your own ?

This repo isn't what you think it is.

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#163

Earlier quoted context omitted.

bittorrent, come on

context `come on`, what is the point in sharing the weights ? Someone lay out the reason they should package the weights with this, when they're allowing you to apply your own ? This repo isn't what you think it is.

‘come on’ meaning that there is an obvious 23 year old solution to data transfer constraints called bittorrent

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#164
post #71

Earlier quoted context omitted.

It's quite clearly a reference to WALL-E the environmentally conscious robot, which is pronounced as you'd expect. I like to think of it as DALL-E the surrealist robot painter.

I totally failed to make that connection! Was that the intended reference? What's the link to WALL-E?

WALL-E is a robot.

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#165
It's a great example of how LLMs can be fine-tuned for specific behaviors. The gain in output quality relative to the expense of tweaking an older, relatively small model is quite impressive. Dolly may not be the Bugatti of LLM chatbots. But it illustrates that high performance is increasingly within reach for the rest of us.

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#167

Earlier quoted context omitted.

Thought the same until I read this: > Contact us at hello-dolly@databricks.com if you would like to get access to the trained weights.

https://github.com/databrickslabs/dolly it’s now available on GitHub

That’s the repo with the code to train the model to get the weights, not the trained weights.

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#168
post #107

Earlier quoted context omitted.

The 3 hours is the instruction fine-tuning. The base foundational model is GPT-J which was already provided by Eleuther-AI and has been around for a couple of years. Note: I work at Databricks and am familiar with this project but didn't work on it.

Do you know why GPT-J is being used instead of NeoX or any of the other larger open source models?

If fine tuning a small model, which can be run on consumer hardware once trained, provides quality results, why use a larger base model?

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#169
post #4

This is really great news and something I felt was missing from the market so far. It seems everyone wants to create `moats` or walled-gardens with some aspect of their models etc. Nice job DataBricks, nice numbers too. Looking forward to more improvements.

Thought the same until I read this: > Contact us at hello-dolly@databricks.com if you would like to get access to the trained weights.

See above, there are simply legal uncertainties about commercial use for users, so want to make sure anyone getting them knows this clearly. That said, you can recreate these weights for like $30.

Re: Hello Dolly: Democratizing the magic of ChatGPT with open models

#170

Earlier quoted context omitted.

https://github.com/databrickslabs/dolly it’s now available on GitHub

That’s the repo with the code to train the model to get the weights, not the trained weights.

Did you try emailing?
Post reply on HN