Falcon 40B LLM (which beats Llama) now Apache 2.0
81–90 of 143 posts
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#82For people who directly want to check the benchmark - https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderb...
Very cool but how does this compare to GPT-4 (before it was nerfed)? I feel like the best benchmark atm is the orig gpt-4 version.
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#8340B is pretty large, right? I expect it would take 70GB or so of RAM to run it. That's some expensive hardware ($10,000 or more).
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#8440B is pretty large, right? I expect it would take 70GB or so of RAM to run it. That's some expensive hardware ($10,000 or more).
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#8540B is pretty large, right? I expect it would take 70GB or so of RAM to run it. That's some expensive hardware ($10,000 or more).
Here somebody quantized it down to 29929.56MB .
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#86Chatting with Falcon 40B feels like chatting with GPT4, very capable model.
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#87Amy implementation for this akin to llama.cpp?
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#8840B is pretty large, right? I expect it would take 70GB or so of RAM to run it. That's some expensive hardware ($10,000 or more).
Sure, it might be a lot slower, but that's a lot better than "I give up, go buy $20K worth of hardware"
Re: Falcon 40B LLM (which beats Llama) now Apache 2.0
#8940B is pretty large, right? I expect it would take 70GB or so of RAM to run it. That's some expensive hardware ($10,000 or more).
Some people use second hand P40 GPUs, which go for around 200-300$. Combine 3 of them with SLI and you've got 72GB of VRAM for less then $1000