I tried the Fugu models with some real world tales in C# and unity using mcp and open code. I exhausted the $20 plan 5 hour window in one prompt to review my theme system and plan some color changes. So I upgraded to the $100 to see the implementation and result. Well the result was worse than Opus, incredibly slow, and I ended up exhausting the new 5 hour window and have used 35% of the weekly now and it hardly crea…
We provide a similar service for Godot instead of Unity, and 20$ plan being exhausted in one prompt on a top model like Opus sounds about right. That's the life when you pay API prices and can't afford 10x subsidies.
Asian AI startups launch Mythos-like models
201–210 of 211 posts
Re: Asian AI startups launch Mythos-like models
#202I tried the Fugu models with some real world tales in C# and unity using mcp and open code. I exhausted the $20 plan 5 hour window in one prompt to review my theme system and plan some color changes. So I upgraded to the $100 to see the implementation and result. Well the result was worse than Opus, incredibly slow, and I ended up exhausting the new 5 hour window and have used 35% of the weekly now and it hardly crea…
Which unity mcp do you use? I've been playing around with the official one, but was wondering what other folks use. Ran into a package conflict issue with the popular coplay one
Re: Asian AI startups launch Mythos-like models
#203Earlier quoted context omitted.
Yeah this is about the worst way you could imagine to evaluate an AI model. If you’d given it a real task you’d have been impressed. I was floored by the day I spent with Fable. Got weeks of work done.
Same. It was one shotting unbelievably well compared to 4.8.
Re: Asian AI startups launch Mythos-like models
#204Earlier quoted context omitted.
I don't even look at benchmarks anymore. I just try different models as they're released on our large, proprietary, systems software codebases in real, shipping products or projects that will ship eventually. It's pretty clear which models help me do my job better or faster. I'm fortunate enough to have the token budget to use basically as much as I need, for now. No need for benchmarks, evals, marketing, system card…
I am not against AI but I do wonder how you guys handle the fact this leaks all your code and is stored forever on servers belonging to God knows who? I “trust” OpenAI and Anthropic (somewhat) but to be honest I still feel only safe using it on code without any secret sauce whatsoever. Luckily that’s a lot of code, but still. I wonder how others are looking at this? (FYI I feel the same about Github and we also don’t…
I don't.
My company doesn't host any code on GitHub, we have our own Git servers.
Regarding the privacy issues with OpenAI, Anthropic or anyone else, we engineers are just using the tools we've been authorized to use by the company. If there's a security issue, that would need to be worked out between the company and the model providers.
Would I use external model providers through an API for personal projects? For everything I'm doing now, probably. Could I see a day where I'm working on something too sensitive to be willing to give any data to these companies? Possibly.
Re: Asian AI startups launch Mythos-like models
#205Earlier quoted context omitted.
More specifically, political lobotomy shouldn’t affect coding ability.
You’d be quite surprised, I think. Fine tuning a model on one axis can have drastic impacts on another that as a human we would expect to be completely unrelated.
The practical reality is that the Chinese and American models might have very different politics. But the most relevant factor in model performance is the quality and volume of training data, not ideology of the base model. Unless you are suggesting something very particular about the way Grok was neutered.
Re: Asian AI startups launch Mythos-like models
#206Earlier quoted context omitted.
If training a good model requires talent then that’s the answer to the question this thread is trying to answer: is training a good model actually that hard ?
Talent to do.. what? This could mean a lot of things. Navigating astronomically huge fundamentally not so hard but still really tangly and hairy projects requiring both excellent short- and long-term vision in an overheated domain with angry people and lots of money is a skill all of its own.
Re: Asian AI startups launch Mythos-like models
#207Earlier quoted context omitted.
How would that work in practice?
All American services won't be allowed to provide the models. Huggingface for example. The same way BYDs are illegal in the US.
Re: Asian AI startups launch Mythos-like models
#208Earlier quoted context omitted.
We provide a similar service for Godot instead of Unity, and 20$ plan being exhausted in one prompt on a top model like Opus sounds about right. That's the life when you pay API prices and can't afford 10x subsidies.
Not sure if you meant Fable/Mythos instead of Opus, but I can comfortable work several hours a day using Opus on Max-Ultracode on the CC 10x.
Re: Asian AI startups launch Mythos-like models
#209Earlier quoted context omitted.
Anthropic always publishes 3p benchmarks every time they announce a new model
No, stop right there. Anything published by Anthropic implicitly is not third party. For it to be third party, the third party has to be the one publishing it.
Re: Asian AI startups launch Mythos-like models
#210Earlier quoted context omitted.
Not sure if you meant Fable/Mythos instead of Opus, but I can comfortable work several hours a day using Opus on Max-Ultracode on the CC 10x.
Yes because you don't pay API costs with the Claude Code plans.