Taming randomness in ML models with hypothesis testing and marimo
1–10 of 10 posts
Re: Taming randomness in ML models with hypothesis testing and marimo
#2Re: Taming randomness in ML models with hypothesis testing and marimo
#3Re: Taming randomness in ML models with hypothesis testing and marimo
#4[flagged]
Re: Taming randomness in ML models with hypothesis testing and marimo
#5[flagged]
- upset about “AI slop” (the image is clearly not AI)
- mentions tech buzzwords that annoy you
- claiming the article is not as rigorous as an academic paper
Perhaps I’m just old school. But I miss the HN where the best way to get upvotes was to be insightful and not to send low-effort snarky replies
Re: Taming randomness in ML models with hypothesis testing and marimo
#6[flagged]
Re: Taming randomness in ML models with hypothesis testing and marimo
#7Re: Taming randomness in ML models with hypothesis testing and marimo
#8Re: Taming randomness in ML models with hypothesis testing and marimo
#9Good post. I’ve been thinking about doing offline testing of LLM tasks a bit these days and have come to the conclusion that old school testing is the best until more mature features can be developed. Specifically, I mean running a power analysis to determine sample size, random sampling based on that and then running tests like a z test to see if there is a difference and between what bounds. Tests are expensive and…
Re: Taming randomness in ML models with hypothesis testing and marimo
#10Good post. I’ve been thinking about doing offline testing of LLM tasks a bit these days and have come to the conclusion that old school testing is the best until more mature features can be developed. Specifically, I mean running a power analysis to determine sample size, random sampling based on that and then running tests like a z test to see if there is a difference and between what bounds. Tests are expensive and…
Have you seen LLM testing tools like promptfoo?