Viewing profile — sangwulee
sangwulee
HN member- Joined
- Wed, Jul 02, 2025, 8:12 AM UTC
- HN karma
- 53
- Public activity
- 13 items
- HN profile
- View on Hacker News ↗
About sangwulee
No profile information was provided.
Recent public activity
-
comment
Comment #48661917
A lot of coffee for sure. Regarding the training cost, it's hard to give a good estimate because we used a shared kubernetes cluster with inference + research workloads.
-
comment
Comment #44751668
The highest quality finetuning data was hand curated internally. I would say our post training pipeline is quite similar to SeedDream 2.0 ~ 3.0 series from ByteDance. Similar to th…
-
comment
Comment #44751624
I actually tried a few experiments in early exploration stages! I trained a small classifier to judge AI vs non-AI images. Use it as a reward model to do small RL / post training e…
-
comment
Comment #44751325
The architecture is the same so we found that some LoRAs work out-of-the box, but some LoRAs don't. In those cases, I would expect people to re-run their LoRA finetuning with the t…
-
comment
Comment #44750498
We used two types of datasets for post-training. Supervised finetuning data and preference data used for RLHF stage. You can actually use less than < 1M samples to significantly bo…
-
comment
Comment #44750448
We have not added a separate RTX accelerated version for FLUX.1 Krea, but the model is fully compatible with existing FLUX.1 dev codebase. I don't think we made a separate onnx exp…
-
comment
Comment #44750403
FLUX.1 is one of the most popular open weights text-to-image models. We distilled Krea-1 to FLUX.1 [dev] model so that the community can adopt it seamlessly into existing ecosystem…
-
comment
Comment #44750079
Quick napkin math assuming bfloat16 format : 1B * 16 bits = 16B bits = 2GB. Since it's a 12B parameter model, you get around ~24GB. Downcasting to bfloat16 from float32 comes with …
-
comment
Comment #44749918
I love owls. Photorealism was one of the focus areas for training because "AI look" (e.g. plastic skin) was biggest complaint for FLUX.1 model series. Photorealism was achieved wit…
-
comment
Comment #44749849
Thank you! Glad you find it helpful. The model is focused on photorealism so it should be able to generate most realistic scenes. Although, I think using 3D engines would be more s…
-
comment
Comment #44748228
Hi there, I'm Sangwu Lee, one of the researchers behind this model. I'm happy to answer any questions here. --- I also commented in this other submission: https://news.ycombinator.…
-
comment
Comment #44748201
Hello HackerNews. My name is Sangwu Lee . I work for Krea and I led the research efforts around the post-training for this model. I'll try to answer any questions you may have, but…
-
comment
Comment #44746421
Hi! I'm lead researcher on Krea-1. FLUX.1 Krea is a 12B rectified flow model distilled from Krea-1, designed to be compatible with FLUX architecture. Happy to answer any technical …