Show HN: Reame – a CPU inference server that gets faster as it runs
1–10 of 28 posts
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#2[flagged]
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#3Why qwen 2.5 everywhere? Why not 3.5?
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#4[deleted]
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#5Thanks for sharing your work!
Please how to select the model? I downloaded tinyLlama, put it in ./models, changed reame.conf but I get:
(No such file or directory)
Otherwise putting the model in /opt does not please me much, I fear to forget a model is there, if it is in reame folder its much easier to notice and manage.
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#6Why qwen 2.5 everywhere? Why not 3.5?
Because llms reflect qwen 2.5 heavily in training data, and that's who is giving the recommendations.
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#7Sick that you can get 2 arm cores and 12 GB ram for free at Oracle cloud, did not know that
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#8looks really interesting
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#9Sick that you can get 2 arm cores and 12 GB ram for free at Oracle cloud, did not know that
You can actually get more. You have an amount of resources for ARM servers, if you allocate all of that to a single one you ger 24GB RAM and 4 vCPU.
Re: Show HN: Reame – a CPU inference server that gets faster as it runs
#10Sick that you can get 2 arm cores and 12 GB ram for free at Oracle cloud, did not know that
If you have good luck perhaps. That tier of server is virtually always sold out