Without benchmarking LLMs, you're likely overpaying
karllorey.com
Without benchmarking LLMs, you're likely overpaying
1–10 of 100 posts
Re: Without benchmarking LLMs, you're likely overpaying
#2It sounds like he's building some kind of ai support chat bot.
I despise these things.
Re: Without benchmarking LLMs, you're likely overpaying
#3Since building a custom agent setup to replace copilot, adopting/adjusting Claude Code prompts, and giving it basic tools, gemini-3-flash is my go-to model unless I know it's a big and involved task. The model is really good at 1/10 the cost of pro, super fast by comparison, and some basic a/b testing shows little to no difference in output on the majority of tasks I used
Cut all my subs, spend less money, don't get rate limited
Re: Without benchmarking LLMs, you're likely overpaying
#4> He's a non-technical founder building an AI-powered business. It sounds like he's building some kind of ai support chat bot. I despise these things.
Re: Without benchmarking LLMs, you're likely overpaying
#5I'd second this wholeheartedly Since building a custom agent setup to replace copilot, adopting/adjusting Claude Code prompts, and giving it basic tools, gemini-3-flash is my go-to model unless I know it's a big and involved task. The model is really good at 1/10 the cost of pro, super fast by comparison, and some basic a/b testing shows little to no difference in output on the majority of tasks I used Cut all my sub…
Re: Without benchmarking LLMs, you're likely overpaying
#6I'd second this wholeheartedly Since building a custom agent setup to replace copilot, adopting/adjusting Claude Code prompts, and giving it basic tools, gemini-3-flash is my go-to model unless I know it's a big and involved task. The model is really good at 1/10 the cost of pro, super fast by comparison, and some basic a/b testing shows little to no difference in output on the majority of tasks I used Cut all my sub…
Re: Without benchmarking LLMs, you're likely overpaying
#7> He's a non-technical founder building an AI-powered business. It sounds like he's building some kind of ai support chat bot. I despise these things.
Re: Without benchmarking LLMs, you're likely overpaying
#8> He's a non-technical founder building an AI-powered business. It sounds like he's building some kind of ai support chat bot. I despise these things.
[flagged]
Re: Without benchmarking LLMs, you're likely overpaying
#9I'd second this wholeheartedly Since building a custom agent setup to replace copilot, adopting/adjusting Claude Code prompts, and giving it basic tools, gemini-3-flash is my go-to model unless I know it's a big and involved task. The model is really good at 1/10 the cost of pro, super fast by comparison, and some basic a/b testing shows little to no difference in output on the majority of tasks I used Cut all my sub…
I've been using the smaller models ever since. Nano/mini, flash, etc.