Viewing profile — danielhanchen
danielhanchen
HN member- Joined
- Wed, Sep 08, 2021, 12:01 PM UTC
- HN karma
- 3,534
- Public activity
- 666 items
- HN profile
- View on Hacker News ↗
About danielhanchen
1. Used to work at NVIDIA RAPIDS cuML
2. Discord: https://discord.gg/unsloth
3. Github: https://github.com/danielhanchen
4. Twitter / X: x.com/danielhanchen
5. Email: my handle @ gmail.com
6. Bug fixes for Gemma: https://news.ycombinator.com/item?id=39671146
7. Bug fixes for Gradient Accumulation: https://x.com/danielhanchen/status/1846235913443262891?lang=en
Recent public activity
-
comment
Comment #48931117
Oh thanks for sharing! The llama.cpp PRs should generally be fine for now - I'm fixing a few small edge cases as well!
-
comment
Comment #48626978
Very cool write-up and GitHub repo!
-
comment
Comment #48096146
Thank you appreciate the support! It's all thanks to you guys and the community!
-
comment
Comment #48047664
Update - Just got rid of the spiced up intro
-
comment
Comment #48047659
Thank you!
-
comment
Comment #48047651
Oh thanks :) We're also going to add MTP support soon for Qwen3.6! 95% of it is fully human done - the maths, algos, code snippets, screenshots & benchmarks are done / conducted by…
-
comment
Comment #47990698
[dead]
- story
-
comment
Comment #47958823
Sorry on the delay - so it installs https://github.com/Blaizzy/mlx-vlm and other components and sets up the commands - you don't need to use it but we thought it might be easier fo…
-
comment
Comment #47958818
Sorry on the delay - oh haha that would be cool :) We did release 2bit dynamic ones, but unsure if they'll be helpful
-
comment
Comment #47958814
Yes we do! Sorry on the delay
-
comment
Comment #47958476
We use Duck Duck Go - sorry on the delayed response as well
-
comment
Comment #47958472
Thank you and appreciate it! Sorry on the delayed reply as well
-
comment
Comment #47958471
Oh yes LM Link is cool!
-
comment
Comment #47958470
Hey sorry on the delay - we just added API support, so you can access a remote server - it includes optional python, tool call, bash and web search support if you enable them. For …
-
comment
Comment #47958453
Hey! Sorry for not replying sooner - yes we'll keep publishing more KLD - sadly some are saying we are "optimizing" for KLD now since we posted so many haha - but the whole purpose…
-
comment
Comment #47958401
Hey so sorry didn't reply sooner - yes the docker used to be I think 4-8GB ish since CUDA sadly itself is 4GB I think, and PyTorch takes the rest. So unfortunately the Unsloth Dock…
-
comment
Comment #47958386
Apologies as well didn't reply sooner - Studio supports AMD out of the box now! We worked with AMD to make it work! One thing that is still missing is pre-compiled AMD ROCM binarie…
-
comment
Comment #47958366
Oh my apologies I didn't respond - if only HN had a notifier haha Oh yes we added a custom folder button which can pull .gguf files for now from any folder - it supports LM Studio …
-
comment
Comment #47865679
We made Unsloth Studio which should help :) 1. Auto best official parameters set for all models 2. Auto determines the largest quant that can fit on your PC / Mac etc 3. Auto deter…
-
comment
Comment #47865374
Haha :)
-
comment
Comment #47865365
Haha :) We had some issues with Kimi-2.6 since it was int4 and we were investigating how to handle it :)
-
comment
Comment #47865351
We also made some dynamic MLX ones if they help - it might be faster for Macs, but llama-server definitely is improving at a fast pace. https://huggingface.co/unsloth/Qwen3.6-27B-U…
-
comment
Comment #47814720
Yes sadly CUDA 13.2 is broken - NVIDIA will push a fix in CUDA 13.3
-
comment
Comment #47803754
Yes we have started doing diffusion GGUFs but it's in it's infancy :) But yes we do generate images to test quants out!