it doesn't seem downloadable to run locally, a shame.
Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
31–40 of 70 posts
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#32Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#33I look forward to trying this out, mostly because I’m very frustrated with censored models. I am experimenting with summarizing and navigating documents for forensic psychiatry work, much of which involves subjects that instantly hit the guard rails of LLMs. So far, I have had zero luck getting help from OpenAI/Anthropic or vendors of their models to request an exception for uncensored models. I need powerful models…
Mistrial-Nemo should be able to do this.
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#34Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#35Isn't 63% => 54% regression on MMLU-Pro a huge issue? They said that it excels at advanced reasoning but that seems like a big drawback there.
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#36Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#37I look forward to trying this out, mostly because I’m very frustrated with censored models. I am experimenting with summarizing and navigating documents for forensic psychiatry work, much of which involves subjects that instantly hit the guard rails of LLMs. So far, I have had zero luck getting help from OpenAI/Anthropic or vendors of their models to request an exception for uncensored models. I need powerful models…
How good are these models at summarization anyways? I tried uploading obscure books I've already read, to GPT4 and Claude 3 and asked them to summarize the plot and particular details, as well as asking how many times does a particular thing happen in the book, and the results have been hit and miss. I certainly would not trust these models to create comprehensive and correct summaries of highly sensitive records.
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#38Earlier quoted context omitted.
How good are these models at summarization anyways? I tried uploading obscure books I've already read, to GPT4 and Claude 3 and asked them to summarize the plot and particular details, as well as asking how many times does a particular thing happen in the book, and the results have been hit and miss. I certainly would not trust these models to create comprehensive and correct summaries of highly sensitive records.
Asking "how many times does a particular thing happen in the book" is always going to be hard, because LLMs are notoriously bad at counting.
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#39PAYMENT TANGENT for my fellow entrepreneurs here that take Visa/Mastercard payments: I tried to sign up to Lambda Labs just now to check out Hermes 3. Created an account, verified my email address, entered my billing info... ... but then it says they only accept CREDIT cards, NOT DEBIT cards. I had never heard of this, so I tried it anyway. I entered my business Mastercard (from mercury.com FWIW), that's never been r…
> Anyone know why a business would choose to only accept credit not debit cards? Maybe they want to place a temporary charge to verify the card's valid? I don't believe you can do so with a debit card.
Re: Hermes 3: The First Fine-Tuned Llama 3.1 405B Model
#40I look forward to trying this out, mostly because I’m very frustrated with censored models. I am experimenting with summarizing and navigating documents for forensic psychiatry work, much of which involves subjects that instantly hit the guard rails of LLMs. So far, I have had zero luck getting help from OpenAI/Anthropic or vendors of their models to request an exception for uncensored models. I need powerful models…
try google's gemini models, safety filtering can be completely disabled via cloud studio or api