Live data from Hacker News

Viewing profile — danielhanchen

danielhanchen

HN member
Joined
Wed, Sep 08, 2021, 12:01 PM UTC
HN karma
3,534
Public activity
666 items

About danielhanchen

Unsloth github.com/unslothai/unsloth - finetune Llama 2x faster + use 70% less VRAM

1. Used to work at NVIDIA RAPIDS cuML

2. Discord: https://discord.gg/unsloth

3. Github: https://github.com/danielhanchen

4. Twitter / X: x.com/danielhanchen

5. Email: my handle @ gmail.com

6. Bug fixes for Gemma: https://news.ycombinator.com/item?id=39671146

7. Bug fixes for Gradient Accumulation: https://x.com/danielhanchen/status/1846235913443262891?lang=en

Recent public activity

  1. comment
    Comment #48931117

    Oh thanks for sharing! The llama.cpp PRs should generally be fine for now - I'm fixing a few small edge cases as well!

  2. comment
    Comment #48626978

    Very cool write-up and GitHub repo!

  3. comment
    Comment #48096146

    Thank you appreciate the support! It's all thanks to you guys and the community!

  4. comment
    Comment #48047664

    Update - Just got rid of the spiced up intro

  5. comment
    Comment #48047659

    Thank you!

  6. comment
    Comment #48047651

    Oh thanks :) We're also going to add MTP support soon for Qwen3.6! 95% of it is fully human done - the maths, algos, code snippets, screenshots & benchmarks are done / conducted by…

  7. comment
  8. story
  9. comment
    Comment #47958823

    Sorry on the delay - so it installs https://github.com/Blaizzy/mlx-vlm and other components and sets up the commands - you don't need to use it but we thought it might be easier fo…

  10. comment
    Comment #47958818

    Sorry on the delay - oh haha that would be cool :) We did release 2bit dynamic ones, but unsure if they'll be helpful

  11. comment
    Comment #47958814

    Yes we do! Sorry on the delay

  12. comment
    Comment #47958476

    We use Duck Duck Go - sorry on the delayed response as well

  13. comment
    Comment #47958472

    Thank you and appreciate it! Sorry on the delayed reply as well

  14. comment
    Comment #47958471

    Oh yes LM Link is cool!

  15. comment
    Comment #47958470

    Hey sorry on the delay - we just added API support, so you can access a remote server - it includes optional python, tool call, bash and web search support if you enable them. For …

  16. comment
    Comment #47958453

    Hey! Sorry for not replying sooner - yes we'll keep publishing more KLD - sadly some are saying we are "optimizing" for KLD now since we posted so many haha - but the whole purpose…

  17. comment
    Comment #47958401

    Hey so sorry didn't reply sooner - yes the docker used to be I think 4-8GB ish since CUDA sadly itself is 4GB I think, and PyTorch takes the rest. So unfortunately the Unsloth Dock…

  18. comment
    Comment #47958386

    Apologies as well didn't reply sooner - Studio supports AMD out of the box now! We worked with AMD to make it work! One thing that is still missing is pre-compiled AMD ROCM binarie…

  19. comment
    Comment #47958366

    Oh my apologies I didn't respond - if only HN had a notifier haha Oh yes we added a custom folder button which can pull .gguf files for now from any folder - it supports LM Studio …

  20. comment
    Comment #47865679

    We made Unsloth Studio which should help :) 1. Auto best official parameters set for all models 2. Auto determines the largest quant that can fit on your PC / Mac etc 3. Auto deter…

  21. comment
  22. comment
    Comment #47865365

    Haha :) We had some issues with Kimi-2.6 since it was int4 and we were investigating how to handle it :)

  23. comment
    Comment #47865351

    We also made some dynamic MLX ones if they help - it might be faster for Macs, but llama-server definitely is improving at a fast pace. https://huggingface.co/unsloth/Qwen3.6-27B-U…

  24. comment
    Comment #47814720

    Yes sadly CUDA 13.2 is broken - NVIDIA will push a fix in CUDA 13.3

  25. comment
    Comment #47803754

    Yes we have started doing diffusion GGUFs but it's in it's infancy :) But yes we do generate images to test quants out!